Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

16314762 · Jun 202019922001200920182026
48 results for Equiangular tight frame

Sparse coding in learned dictionaries has been established as a successful approach for signal denoising, source separation and solving inverse problems in general. A dictionary learning method adapts an initial dictionary to a particular signal class by iteratively computing an approximate factorization of a training …

2012-05-28abs ↗pdf ↗

The paper uncovers symmetries in large language models through layer-peeled optimization.

problem Understanding geometric structure in large language model weights and context embeddings.
method Constrained layer-peeled optimization program to analyze symmetries in next-token distributions.
result Symmetries in target next-token distributions are transferred to optimal model weights and context embeddings.

An equiangular hyperbolic Coxeter polyhedron is a hyperbolic polyhedron where all dihedral angles are equal to π/n for some fixed integer n at least 2. It is a consequence of Andreev's theorem that either n=3 and the polyhedron has all ideal vertices or that n=2. Volume estimates are given for all equiangular hyperboli…

2008-04-16abs ↗pdf ↗

Deep nets exhibit 'Neural Collapse' during training's final phase, simplifying decision-making.

problem Understanding and optimizing deep learning training phases.
method Direct measurements on three deepnet architectures across seven datasets.
result Deep nets exhibit 'Neural Collapse' during training's final phase, simplifying decision-making.

Frames for Rn\R^n can be thought of as redundant or linearly dependent coordinate systems, and have important applications in such areas as signal processing, data compression, and sampling theory. The word "frame" has a different meaning in the context of differential geometry and topology. A moving frame for the tang…

2012-09-25abs ↗pdf ↗

Language models allocate information storage, not collapsing into uniform representations.

problem Incomplete neural collapse in language model representations.
method Analyzing variance and information sharing across 14 models, proving an information floor.
result Within-class variance is allocated information storage, not collapsed into uniform representations.

Redundancy helps speed up slow nodes in distributed learning.

problem Slow nodes (stragglers) bottleneck distributed optimization and learning performance.
method Encode data with redundancy, dynamically exclude stragglers, and compensate losses.
result Optimization algorithms converge to solutions even with straggling nodes.

This paper studies how to compress neural networks while maintaining accuracy.

problem Compressing a two-layer neural network with fewer nodes without losing accuracy.
method Using tools from high-dimensional probability, the authors minimize the L_2 loss between the target and compressed networks.
result The error rate of the approximation is shown as a function of input dimension and network size in the mean-field limit.

This work justifies neural collapse under MSE loss and analyzes the optimization landscape.

problem Understanding neural collapse in deep neural networks under MSE loss.
method Global landscape analysis of vanilla nonconvex MSE loss.
result The only global minimizers are neural collapse solutions.

Graphs with maximum degree Δ have at most O(1) equiangular lines for λ < 3/sqrt(2).

problem Finding the maximum number of equiangular lines in graphs with a given maximum degree.
method Using eigenfunctions and nodal domains to estimate the multiplicity of eigenvalues.
result The maximum multiplicity of λ as the second largest eigenvalue is O(1) for graphs with maximum degree Δ and cyclomatic number.

We study the configuration space of equilateral and equiangular spatial hexagons for any bond angle by giving explicit expressions of all the possible shapes. We show that the chair configuration is isolated, whereas the boat configuration allows one-dimensional deformations which form a circle in the configuration spa…

2011-05-25abs ↗pdf ↗

New model explains neural collapse and limits on minority classes in imbalanced datasets.

problem Understanding and predicting performance limits of deep learning models on imbalanced datasets.
method Layer-Peeled Model, a nonconvex optimization program isolating top layers and applying constraints.
result Reveals a new phenomenon called Minority Collapse that limits deep learning models on minority classes.

Deep linear networks exhibit collapsing features and classifiers across datasets.

problem Understanding the collapse of features and classifiers in deep linear networks.
method Theoretical and empirical analysis of deep linear networks with MSE and CE losses.
result Deep linear networks exhibit NC properties, collapsing features and classifiers to orthogonal vectors.

New method denoises graph signals using wavelets, scalable for large graphs.

problem Denoising graph signals with overcomplete tight frames and correlated noise.
method Data-driven wavelet tight frame, Stein's unbiased risk estimate, Chebyshev-Jackson polynomial approximations, Monte-Carlo strategy.
result Method scales to large graphs and finds applications in differential privacy.

This paper extends neural collapse to class-imbalanced datasets using an unconstrained ReLU feature model.

problem Understanding neural collapse in class-imbalanced datasets with cross-entropy loss.
method Generalized neural collapse to class-imbalanced settings using an unconstrained ReLU feature model.
result Class-means converge to orthogonal vectors with different lengths, and classifier weights align to these vectors.

Quantifies how geodesic planes isolate in hyperbolic 3-manifolds.

problem Understanding isolation properties of geodesic planes in hyperbolic 3-manifolds.
method Quantitative estimates of geodesic planes in frame bundles, using tight areas and densities.
result Polynomial estimates of isolation properties with degree given by modified critical exponents.

Pipeline combines ETF preprocessing with tabular model for cross-modal inference.

problem Transferability of tabular models across different modalities.
method Fixed comparison object, ETF preprocessing, in-context inference.
result Pipeline is broadly competitive, runs faster, and produces well-calibrated probabilities.

We show that an oriented elliptic 3-manifold admits a universally tight positive contact structure iff the corresponding group of deck transformations on S3S^3 preserves a standard contact structure pointwise. We also relate univerally tight contact structures on 3-manifolds covered by S3S^3 to the exceptional isomorph…

2001-12-24abs ↗pdf ↗

We analyze neural collapse in neural networks, showing that features collapse to vertices of a Simplex ETF.

problem Understanding and optimizing the features learned in the last layer of neural networks during training.
method Simplified unconstrained feature model, studying the global optimization landscape of cross-entropy loss with weight decay.
result The global minimizers of the loss are Simplex ETFs, and other critical points are strict saddles with negative curvature.

Deep nets trained with MSE loss exhibit Neural Collapse, collapsing features and classifiers to class means.

problem Understanding Neural Collapse in MSE-trained deep nets.
method Developed a new MSE loss decomposition and introduced the central path concept.
result Exact dynamics of Neural Collapse along the central path can be predicted.

We study pairs of curves with Poncelet's porism properties and compute their vertex curves.

problem Understanding pairs of curves with Poncelet's porism properties.
method Developed formulas to compute vertex curves for given envelope curves and vice versa, for all sufficiently regular pairs of Poncelet curves.
result Formulas produce all possible sufficiently regular pairs of Poncelet curves, including sets of curves analogous to pencils of conic sections.

New wavelet frames constructed from reproducing kernels for continuous and discrete domains.

problem Generating wavelet frames on non-Euclidean structures.
method Spectral filtering of integral operators associated with reproducing kernels.
result Discrete frames as Monte Carlo estimates of continuous frames, with finite-sample rates derived.

This paper extends neural collapse to imbalanced data under cross-entropy loss.

problem Analyzing neural collapse in deep networks with imbalanced data.
method Using the unconstrained feature model and cross-entropy loss, the paper studies neural collapse in imbalanced datasets.
result Feature vectors within the same class collapse to a single mean vector, but angles between them depend on sample size.

It is well-known that a knot in a contact manifold (M,C)(M,C) transverse to a trivialized contact structure possesses the natural framing given by the first of the trivialization vectors along the knot. If the Euler class eCH2(M)e_C\in H^2(M) of CC is nonzero, then CC is nontrvivializable and the natural framing of transvers…

2001-03-31abs ↗pdf ↗

Six quaternionic lines with optimal angles found in 2D quaternion space.

problem Finding optimal configurations of quaternionic lines in 2D space.
method Simple presentation of lines as orbit of a reflection group, finding other optimal designs.
result Optimal spherical designs of 10, 15, and 20 lines in quaternion space.

This work explains neural collapse in shallow neural networks and its impact on generalization.

problem Understanding neural collapse in shallow neural networks and its effect on generalization.
method Analysis of two and three-layer ReLU neural networks, focusing on data dimension, sample size, and signal-to-noise ratio.
result Neural collapse occurs in shallow ReLU networks under certain conditions related to data properties and network architecture.

Parseval networks improve deep nets' robustness to adversarial examples.

problem Improving deep neural networks' robustness to adversarial attacks.
method Constraining the Lipschitz constant and maintaining Parseval tight frames in weight matrices.
result Parseval networks maintain accuracy and robustness to adversarial examples compared to vanilla networks.

New deep learning methods improve CT image quality from few projections.

problem Sparse-view CT images suffer from streaking artifacts due to limited projections.
method Inspired by deep convolutional framelets, propose new U-Net variants that satisfy the frame condition.
result New U-Net variants provide better reconstruction performance for sparse-view CT.

Paper tackles image reconstruction from limited data using polyhedral norms and convex regularizers.

problem Learning convex regularizers for image reconstruction from limited data.
method Imposes amplitude-equivariance, approximates functionals with polyhedral norms, identifies synthesis and analysis forms, proposes a trainable tight frame architecture.
result Proposed framework outperforms sparsity-based methods in denoising and biomedical image reconstruction.

We construct a simple topological invariant of certain 3-manifolds, including quotients of the 3-sphere by finite groups, based on the fact that the tangent bundle of an orientable 3-manifold is trivialisable. This invariant is strong enough to yield the classification of lens spaces of odd, prime order. We also use pr…

2001-03-27abs ↗pdf ↗

A standard convexity condition on the boundary of a symplectic manifold involves an induced positive contact form (and contact structure) on the boundary; the corresponding concavity condition involves an induced negative contact form. We present two methods of symplectically attaching 2-handles to convex boundaries of…

1999-12-17abs ↗pdf ↗

A twisted curve in Euclidean 3-space E^3 can be considered as a curve whose position vector can be written as linear combination of its Frenet vectors. In the present study we study the twisted curves of constant ratio in E^3 and characterize such curves in terms of their curvature functions. Further, we obtain some re…

2014-10-21abs ↗pdf ↗

SSTQ improves privacy-preserving vector quantization with low communication cost.

problem Achieving local differential privacy in distributed optimization with low communication cost.
method Combines overcomplete equal-norm tight frames, coordinate subsampling, and privacy-aware one-dimensional quantization.
result Achieves optimal mean squared error scaling with only log2N+b\lceil \log_2 N \rceil + b bits per client.

A theory of feature geometry using spectral analysis of weight matrices.

problem Current methods decompose neural network activations into sparse linear features, losing geometric structure.
method Develops a theory by analyzing the spectra of weight-derived matrices, introducing the frame operator.
result Features collapse onto single eigenspaces, organizing into tight frames, and admit discrete classification.

New set class preserves Fourier series terms for planar ovals, leading to isoperimetric inequalities.

problem Investigate geometric properties of kkth Order Preserving Sets and ovals.
method Introduce and analyze kkth Order Preserving Sets and Midpoint Sets; study geometric properties and isoperimetric inequalities.
result Established an isoperimetric-type inequality relating perimeter and area of ovals and their associated sets.