Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,932 papers · 148 categories

Trend · papers per month

222444666888 · Jun 202019922001200920172026
48 results for Permutation Invariant Training

Prob-PIT improves speech separation by considering output-label permutations as random variables.

problem Overconfident output-label assignment in PIT leads to unreliable speech separation.
method Prob-PIT treats output-label permutations as a discrete latent random variable with a uniform prior distribution and maximizes the log-likelihood function.
result Prob-PIT significantly outperforms PIT in terms of Signal to Distortion Ratio and Signal to Interference Ratio.

Improves full conformal prediction for stochastic non-conformity measures.

problem Inability of existing conditions to guarantee full conformal prediction validity under stochastic settings.
method Introduces a new sufficient condition: Conditional Independence & Permutation Invariance in Distribution.
result Corrects the insufficient condition and provides a new sufficient condition for full conformal prediction validity.

Generative model for set-valued data using permutation invariant flows.

problem Modeling set-valued data with conditional generative models.
method Conditional generative probabilistic model using continuous normalizing flows with permutation equivariant dynamics.
result Significantly outperforms non-permutation invariant baselines in log likelihood and domain-specific metrics.

Variational inference struggles with weight symmetries in neural networks, leading to biased posteriors.

problem Weight space symmetries in neural networks cause multimodal posteriors, challenging variational inference.
method Developed a symmetrization mechanism to create permutation invariant variational posteriors.
result Symmetrized variational posteriors have a better fit to the true posterior and improved predictive performance.

Neural networks learn from ensemble forecasts without considering their order.

problem Improving reliability of probabilistic weather forecasts.
method Permutation-invariant neural networks for postprocessing ensemble forecasts.
result Models achieve state-of-the-art prediction quality in surface temperature and wind gust forecasts.

Representations of sets are challenging to learn because operations on sets should be permutation-invariant. To this end, we propose a Permutation-Optimisation module that learns how to permute a set end-to-end. The permuted set can be further processed to learn a permutation-invariant representation of that set, avoid…

2018-12-10abs ↗pdf ↗

Paper develops deep neural networks for wireless tasks with reduced complexity.

problem Reducing training complexity for deep neural networks in wireless systems.
method Develops permutation invariant DNNs (PINNs) leveraging wireless task properties.
result Demonstrates dramatic reduction in training complexity for PINNs.

A new method computes Teichmüller polynomials from integer permutations.

problem Computing Teichmüller polynomials for fibered 3-manifolds.
method Using integer permutations to characterize pseudo-Anosov homeomorphisms and train tracks.
result Direct implementation of McMullen's algorithm for Teichmüller polynomials.

Generates valid Euclidean distance matrices for molecular structures.

problem Generating point clouds in arbitrary rotations and translations is challenging.
method Developed a neural network architecture that produces valid Euclidean distance matrices invariant to rotations and translations.
result The architecture can generate molecular structures in a one-shot fashion by producing Euclidean distance matrices with a three-dimensional embedding.

Improves efficiency and scalability in multi-agent reinforcement learning.

problem Non-stationarity and inefficiency in critic networks due to agent permutations.
method Proposes a permutation invariant critic (PIC) to avoid changes in critic output due to agent permutations.
result Achieves improvements of test episode reward between 15% to 50% on challenging multi-agent particle environment (MPE).

SetGAN improves GANs by making them more stable and diverse.

problem Training instability and mode collapse in GANs.
method Adversarial architecture that processes sets of generated and real samples, discriminating between their origins in a flexible, permutation invariant manner.
result SetGAN produces more accurate models of the input data with less sensitivity to hyperparameters.

Enhances GNNs by capturing node relationships, outperforming 2-WL test.

problem Inability of conventional GNNs to fully capture node relationships due to permutation invariance.
method Develops permutation-sensitive aggregation mechanism using permutation groups.
result Proves superior expressivity compared to 2-WL test and not less than 3-WL test.

SPAN learns functions over sets invariant to permutations, outperforming existing methods.

problem Learning functions over sets invariant to permutations.
method SPAN architecture that combines neural networks with adversarial permutations.
result SPAN achieves nearly permutation-invariant functions while maintaining accuracy.

The paper models financial correlation matrices using permutation invariant Gaussian models and predicts market anomalies.

problem Modeling and predicting financial correlation matrices from high-frequency data.
method Constructing permutation invariant Gaussian matrix models with 4 parameters, using graph theory and polynomial functions.
result The permutation invariant Gaussian matrix model predicts the expectation values of cubic and quartic polynomials with strong evidence of fit.

New graph foundation models respect symmetries for broader applicability.

problem Tailored graph machine learning architectures limit broader applicability.
method Investigates symmetries for label and feature permutations, proving network universal approximator.
result Universal approximator on multisets respecting node and feature permutations.

This paper improves model fusion by training-time neuron alignment, reducing barriers in multi-model fusion.

problem Diverse neuron permutations across different settings hinder model fusion performances.
method Training-time neuron alignment using fixed neuron anchors to reduce training-time permutations.
result Training-time neuron alignment improves fusion of pretrained models and federated learning performances.

Permutation-equivariant neural networks improve auction mechanisms by reducing regret and sample complexity.

problem Designing optimal auction mechanisms that balance revenue and bidders' regret.
method Introduced permutation-equivariant neural networks to auction mechanisms.
result Permutation-equivariant neural networks decrease expected ex-post regret and improve model generalizability.

New method improves transfer and robustness of supervised contrastive learning.

problem Class collapse in supervised contrastive learning leads to poor representation quality.
method Adding a weighted class-conditional InfoNCE loss and a class-conditional autoencoder.
result Improves transfer and robustness on 5 standard datasets and 3 worst-group robustness datasets.

We study the problem of designing models for machine learning tasks defined on \emph{sets}. In contrast to traditional approach of operating on fixed dimensional vectors, we consider objective functions defined on sets that are invariant to permutations. Such problems are widespread, ranging from estimation of populati…

2017-03-10abs ↗pdf ↗

A new DNN structure reduces complexity for wireless tasks.

problem Reducing complexity in training deep neural networks for wireless tasks.
method Proposes a DNN with special structure using permutation invariant a priori information.
result The proposed DNN structure reduces training complexity and model parameters.

This work refines claims about neural network connectivity, showing that simultaneous linear connectivity is possible under certain conditions.

problem Neural networks' loss landscapes are non-convex due to permutation symmetries, leading to high loss barriers between permuted networks.
method The authors introduce and analyze three claims of increasing strength regarding the connectivity of neural networks, focusing on permutations that align networks.
result The authors provide evidence that strong linear connectivity may be possible under certain conditions, specifically when interpolating among three networks of increasing width.

We introduce a simple permutation equivariant layer for deep learning with set structure.This type of layer, obtained by parameter-sharing, has a simple implementation and linear-time complexity in the size of each set. We use deep permutation-invariant networks to perform point-could classification and MNIST-digit sum…

2016-11-14abs ↗pdf ↗

New model preserves symmetry in multivariate time series, improving performance.

problem Implicit ordering in MTS models violates inherent exchangeability.
method Permutation-equivariant 2D state space model with canonical architecture.
result Eliminates sequential dependency chains and simplifies stability analysis.

A new nonparametric test measures dependence between variables using decision trees.

problem Measuring statistical dependence between two variables robustly and efficiently.
method An ensemble of decision trees discriminates between observed and permuted samples without generating the latter.
result The method effectively detects complex relationships from noisy data.

A new method optimizes slicing directions for SW distances to improve high-dimensional probability measure comparison.

problem Challenging identification of informative slicing directions for SW distances.
method Constrained learning approach to optimize slicing directions, using continuous relaxations and gradient-based primal-dual approach.
result Demonstrated efficacy in learning more informative slicing directions on various high-dimensional data.

Drinfel'd used associators to construct families of universal representations of braid groups. We consider semi-associators (i.e., we drop the pentagonal axiom and impose a normalization in degree one). We show that the process may be reversed, to obtain semi-associators from universal representations of 3-braids. We v…

2007-08-04abs ↗pdf ↗