Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

1234 · Jun 202019922001200920172026
48 results for entropy-minimizing

Entropy-minimal measure calculated for a stochastic volatility model.

problem Calculating the entropy-minimal equivalent martingale measure in a stochastic volatility model.
method Revised related theory, calculated entropy-minimal measure.
result Entropy-minimal measure for the exponential Ornstein-Uhlenbeck model.

Proposes MEDM to balance entropy minimization and diversity maximization for better domain adaptation.

problem Trivial solutions in entropy minimization for unsupervised domain adaptation.
method Introduces diversity maximization to balance with entropy minimization, controlled by deep embedded validation.
result MEDM outperforms state-of-the-art methods on four domain adaptation datasets.

The paper constructs diffeomorphisms on a G2G_2-manifold achieving entropy bounds.

problem Achieving Yomdin's homological lower bound for topological entropy on G2G_2-manifolds.
method Constructs diffeomorphisms mimicking Farb-Looijenga's for K3 surfaces and acts freely on Teichmüller space.
result The homotopy moduli space of G2G_2 structures on the manifold has an infinite fundamental group.

This paper improves test-time adaptation for distribution shifts using confidence maximization and input transformation.

problem Improving deep networks' performance on data shifted from the training distribution.
method Proposes a novel loss function combining confidence maximization and batch-wise entropy maximization with an input transformation module.
result Significantly improves robustness of pretrained networks to corruptions on benchmarks like ImageNet-C.

A new method for generative modeling of discrete data using geometric latent subspaces.

problem Learning generative models for discrete data with statistical dependencies.
method Geometric latent-subspace framework in exponential parameter space of product manifolds of categorical distributions.
result Low-dimensional latent space encodes statistical dependencies and accurately models high-dimensional discrete data.

This work improves deep learning from noisy crowdsourced labels.

problem Learning label correction and neural classifier from noisy crowdsourced data.
method Coupled Cross-Entropy Minimization (CCEM) with identifiability and regularization.
result The CCEM criterion correctly identifies annotators' confusion and neural classifier under realistic conditions.

We explain SSL objectives as log-likelihoods in a data curation model.

problem Lack of understanding of SSL objectives as log-likelihoods.
method Formulate SSL objectives as a log-likelihood in a generative model of data curation.
result SSL methods can be understood as lower-bounds on a principled log-likelihood.

Study on entropy stability in product spaces of negatively curved symmetric spaces.

problem Stability of minimal entropy rigidity in product spaces of negatively curved symmetric spaces.
method Analysis of minimal entropy sequences and proof of intrinsic uniqueness of spherical Plateau solutions.
result Entropy-minimizing sequences converge to the model space after removing subsets whose n-volume converges to zero.

We introduce Negative Sampling in Semi-Supervised Learning (NS3L), a simple, fast, easy to tune algorithm for semi-supervised learning (SSL). NS3L is motivated by the success of negative sampling/contrastive estimation. We demonstrate that adding the NS3L loss to state-of-the-art SSL algorithms, such as the Virtual Adv…

2019-11-12abs ↗pdf ↗

Paper proposes a new hierarchical attention mechanism for multi-scale data.

problem Challenges in applying neural attention to multi-scale, multi-modal data.
method Developed a mathematical framework for multi-modal, multi-scale data and derived optimal neural attention mechanics.
result Proposed hierarchical attention mechanism improves transformer performance in multi-scale, multi-modal settings.

New theorems on compactness and finiteness for specific types of self-shrinkers.

problem Characterizing rotationally symmetric self-shrinkers with constraints.
method Compactness and finiteness theorems for self-shrinkers with specific symmetries and constraints.
result Existence of entropy minimizing self-shrinkers diffeomorphic to S1imesSn1S^1 imes S^{n-1} for each n2n \geq 2.

Study examines uncertainty in adversarially trained models and proposes improved AT methods.

problem Uncertainty quantification in adversarially trained models for safety-critical applications.
method Investigates conformal prediction (CP) for adversarial attacks, proposes Beta-weighting loss with entropy minimization for improved prediction set size (PSS).
result Proposed AT-UR method improves CP efficiency and prediction set size.

In this paper, we propose a general framework to learn a robust large-margin binary classifier when corrupt measurements, called anomalies, caused by sensor failure might be present in the training set. The goal is to minimize the generalization error of the classifier on non-corrupted measurements while controlling th…

2016-10-21abs ↗pdf ↗

In this paper, we propose a general framework to learn a robust large-margin binary classifier when corrupt measurements, called anomalies, caused by sensor failure might be present in the training set. The goal is to minimize the generalization error of the classifier on non-corrupted measurements while controlling th…

2015-07-16abs ↗pdf ↗

We propose a general framework for neural network compression that is motivated by the Minimum Description Length (MDL) principle. For that we first derive an expression for the entropy of a neural network, which measures its complexity explicitly in terms of its bit-size. Then, we formalize the problem of neural netwo…

2018-12-18abs ↗pdf ↗

In 2004, Taubes introduced the space of minimal hyperbolic germs with elements consisting of the first and second fundamental form of an equivariant immersed minimal disk in hyperbolic 3-space. Herein, we initiate a further study of this space by studying the behavior of a dynamically defined function which records the…

2014-04-03abs ↗pdf ↗

The paper improves semi-supervised learning using ff-divergences and αα-Rényi divergences.

problem Improving semi-supervised learning with noisy pseudo-labels.
method Inspired by ff-divergences and αα-Rényi divergences, the paper develops new empirical risk functions and regularization techniques.
result The new methods show better performance than traditional self-training methods, especially in noisy pseudo-label scenarios.

The paper proves asymptotic normality for multinomial logistic regression on null covariates.

problem Classical asymptotic normality results fail in high-dimensional multinomial logistic models.
method Developed asymptotic normality and chi-square results for multinomial logistic MLE on null covariates.
result Validated new methodology to test feature significance in high-dimensional classification problems.

Adaptive importance sampling for estimating point process statistics.

problem Estimating the expected value of a statistic of a locally stable point process.
method Adaptive importance sampling with Poisson point processes and cross-entropy minimization.
result The proposed estimator converges to the target value almost surely and is asymptotically normal.

Timely detection of abrupt anomalies is crucial for real-time monitoring and security of modern systems producing high-dimensional data. With this goal, we propose effective and scalable algorithms. Proposed algorithms are nonparametric as both the nominal and anomalous multivariate data distributions are assumed unkno…

2018-09-14abs ↗pdf ↗

A new approach for test-time adaptation detects and reacts to distribution shifts.

problem Improving test-time accuracy under distribution shifts.
method Online self-training with a detection tool based on entropy values and betting martingales.
result The classifier's entropy values match those of the source domain, building invariance to distribution shifts.

Gradient descent biases linear models in next-token prediction towards data entropy.

problem Optimization bias in next-token prediction models.
method Analysis of gradient descent on linear models with sparse conditional distributions.
result Gradient descent selects parameters that equate token logits differences to log-odds in the data subspace.

GOIMDA selects inputs to maximize expected influence on a goal functional, reducing data acquisition needs.

problem Challenges in active data acquisition for learning and optimization tasks in deep neural networks.
method GOIMDA uses inverse curvature and goal gradient to select inputs maximizing expected influence on a specified goal functional.
result GOIMDA achieves target performance with fewer labeled samples or function evaluations compared to baselines.

The need to reason about uncertainty in large, complex, and multi-modal datasets has become increasingly common across modern scientific environments. The ability to transform samples from one distribution PP to another distribution QQ enables the solution to many problems in machine learning (e.g. Bayesian inference…

2018-01-25abs ↗pdf ↗