Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

54108162216 · May 202619922001200920172026
48 results for length principle

Study h-principles for non-integrable distributions on manifolds.

problem Existence and classification of maximally non-integrable distributions of derived length one.
method Introduced formal structures and used h-principles to discuss existence and classification.
result Discussed existence and classification of maximally non-integrable distributions of derived length one.

The paper proves a nonholonomic version of Maupertuis-Jacobi principle and shows that nonholonomic trajectories minimize length.

problem Nonholonomic dynamics and their length minimization.
method Contact bundle formulation and geometric equivalence between problems.
result Regular solutions of nonholonomic mechanical problems are reparametrizations of geodesics with minimized Riemannian length.

Random walks on hyperbolic spaces follow predictable large deviation principles.

problem Understanding the behavior of random walks on hyperbolic spaces.
method Large deviation principles for displacement and translation distances.
result Translation and displacement distances satisfy large deviation principles with the same rate function.

We analyze differences between two information-theoretically motivated approaches to statistical inference and model selection: the Minimum Description Length (MDL) principle, and the Minimum Message Length (MML) principle. Based on this analysis, we present two revised versions of MML: a pointwise estimator which give…

2013-01-30abs ↗pdf ↗

The paper proves rigidity of bordered polyhedral surfaces using variational principles.

problem Determining the rigidity of bordered polyhedral surfaces.
method Using the variational principle, the paper shows that bordered polyhedral surfaces are determined by boundary values and discrete curvatures on interior edges.
result The paper re-proves the classical result that two Euclidean or hyperbolic cyclic polygons are congruent if their side lengths are equal.

The Minimum Description Length (MDL) principle states that the optimal model for a given data set is that which compresses it best. Due to practial limitations the model can be restricted to a class such as linear regression models, which we address in this study. As in other formulations such as the LASSO and forward …

2009-10-21abs ↗pdf ↗

Study shows LLC correlates with neural network compressibility.

problem Evaluating limits of neural network compression.
method Extended minimum description length principle using singular learning theory.
result Complexity estimates based on LLC are linearly correlated with compressibility.

A new method avoids overfitting in network reconstruction by using the minimum description length principle.

problem Determining the optimal model complexity in network reconstruction to prevent overfitting.
method Hierarchical Bayesian inference and weight quantization based on the minimum description length principle.
result The method yields increased accuracy in reconstructing both artificial and empirical networks.

Bayesian methods detect clusters in noisy data more reliably.

problem Noisy data distorts traditional clustering methods, leading to unreliable results.
method Bayesian community detection using Minimum Description Length principle.
result Bayesian methods identify more robust clusters in noisy data.

New approach to learning kernels from data using AIT principles.

problem Learning kernels from data in machine learning.
method Sparse Kernel Flows method based on AIT principles.
result Sparse Kernel Flows aligns with MDL principle and offers a robust theoretical foundation.

New architectures improve KANs, making them more interpretable and accurate.

problem Improving Kolmogorov-Arnold networks while maintaining interpretability.
method Overprovisioned architectures combined with sparsification, deep supervision, and depth selection, optimized with a minimum description length objective.
result Combining sparsification with depth selection achieves competitive or superior accuracy while discovering smaller models.

Robust low-rank matrix estimation is a topic of increasing interest, with promising applications in a variety of fields, from computer vision to data mining and recommender systems. Recent theoretical results establish the ability of such data models to recover the true underlying low-rank matrix when a large portion o…

2011-09-28abs ↗pdf ↗

Let S=Γ\HS=Γ\backslash \mathbb{H} be a hyperbolic surface of finite topological type, such that the Fuchsian group ΓPSL2(R)Γ\le \operatorname{PSL}_2(\mathbb{R}) is non-elementary, and consider any generating set S\mathfrak S of ΓΓ. When sampling by an nn-step random walk in π1(S)Γπ_1(S) \cong Γ with each step given by an element…

2018-07-10abs ↗pdf ↗

Paper proves method for calculating NML code length works for continuous models.

problem Uncertainty in calculating NML code length for continuous models.
method Introduced a novel decomposition approach based on the coarea formula to prove correctness for continuous cases.
result Method accurately calculates NML code length for continuous models.

During the past few years Boolean matrix factorization (BMF) has become an important direction in data analysis. The minimum description length principle (MDL) was successfully adapted in BMF for the model order selection. Nevertheless, a BMF algorithm performing good results from the standpoint of standard measures in…

2019-01-28abs ↗pdf ↗

This paper extends the work in [Suzuki, 1996] and presents an efficient depth-first branch-and-bound algorithm for learning Bayesian network structures, based on the minimum description length (MDL) principle, for a given (consistent) variable ordering. The algorithm exhaustively searches through all network structures…

2013-01-16abs ↗pdf ↗

This paper studies the combinatorial Yamabe flow on hyperbolic surfaces with boundary. It is proved by applying a variational principle that the length of boundary components is uniquely determined by the combinatorial conformal factor. The combinatorial Yamabe flow is a gradient flow of a concave function. The long ti…

2010-03-20abs ↗pdf ↗

Paper proposes a compression principle for neural networks using Bayesian optimization.

problem Finding methods for making generalizable predictions in machine learning.
method Compression principle and Bayesian optimization approach.
result Optimal predictive models minimize total compressed message length of data and model definition.

A central problem in analyzing networks is partitioning them into modules or communities. One of the best tools for this is the stochastic block model, which clusters vertices into blocks with statistically homogeneous pattern of links. Despite its flexibility and popularity, there has been a lack of principled statist…

2016-05-23abs ↗pdf ↗

Given a system of equations in a "random" finitely generated subgroup of the braid group, we show how to find a small ordered list of elements in the subgroup, which contains a solution to the equations with a significant probability. Moreover, with a significant probability, the solution will be the first in the list.…

2004-04-05abs ↗pdf ↗

Minkowski's second theorem can be stated as an inequality for nn-dimensional flat Finsler tori relating the volume and the minimal product of the lengths of closed geodesics which form a homology basis. In this paper we show how this fundamental result can be promoted to a principle holding for a larger class of Finsl…

2018-10-18abs ↗pdf ↗

Kernel networks' stability edge linked to Fisher Information singularity.

problem Understanding the stability edge in high-capacity kernel Hopfield networks.
method Statistical manifold analysis and Riemannian geometry.
result The Ridge of Optimization corresponds to the Edge of Stability, revealing a dual equilibrium.

New model explains how concepts grow based on experience.

problem Existing models assume fixed representation; new model allows for growth.
method Geometric framework with MDL criterion for basis extension.
result Conceptual growth is selective and conservative, exposing or amplifying residual error.

DL/FBF improves GPSR solutions by selecting compact, generalising expressions.

problem Overfitting and structural bloat in symbolic regression with genetic programming.
method Description length (DL) and fractional Bayes factor (FBF) criteria for selecting compact, generalising expressions.
result DL/FBF post-selection improves test performance compared to AIC/BIC baseline.

New method improves bivariate causal discovery by accurately estimating cause variable complexity.

problem Improper estimation of cause variable complexity in current MDL-based methods.
method Rate-distortion MDL (RDMDL) using information dimension for cause variable complexity estimation.
result RDMDL achieves competitive performance on Tübingen dataset.

Optimizes data splitting for shorter conformal prediction intervals.

problem Minimizing prediction interval length while maintaining coverage.
method Theoretical framework for optimal data splitting in split conformal prediction.
result Analytical characterizations of length-optimal split ratios in various settings.

Study on lengths and curvatures of harmonic functions on smooth and singular surfaces.

problem Investigate logarithmic convexity and isoperimetric inequalities of harmonic functions on surfaces.
method Analyzes geodesic curvature, uses Laplace-type equations, and studies growth estimates.
result Generalizes results on logarithmic convexity and isoperimetric inequalities for harmonic functions.

Study improves neural network performance in sequential learning for image classification.

problem Improving neural network performance in sequential learning for image classification.
method Evaluation of approaches for computing prequential description lengths, proposing forward-calibration and replay-streams.
result Improved description lengths for image classification datasets, outperforming previous results.