Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

24477194 · May 202619922001200920172026
48 results for Wise conjecture

Wise's Quasiconvex Hierarchy Theorem classifying hyperbolic virtually compact special groups in terms of quasiconvex hierarchies played an essential role in Agol's proof of the Virtual Haken Conjecture. Answering a question of Wise, we construct a new virtual quasiconvex hierarchy for relatively hyperbolic virtually co…

2019-03-28abs ↗pdf ↗

In this paper we extend Witten-Helffer-Sjöstrand theory from selfadjoint Laplacians based on fiber wise Hermitian structures, to non-selfadjoint Laplacians based on fiber wise non-degenerate symmetric bilinear forms. As an application we verify, up to sign, the conjecture about the comparison of the Milnor-Turaev torsi…

2006-10-28abs ↗pdf ↗

This paper uncovers the low-rank structure of neural network Hessians.

problem Understanding the structure of Hessians in neural networks.
method Proposes a decoupling conjecture to decompose layer-wise Hessians into Kronecker products of smaller matrices.
result Proves the structure of top eigenspaces in 2-layer networks and shows high overlap in top eigenvectors across different models.

Given a Kähler fiber space p:XYp:X\to Y whose generic fiber is of general type, we prove that the fiberwise singular Kähler-Einstein metric induces a semipositively curved metric on the relative canonical bundle KX/YK_{X/Y} of pp. We also propose a conjectural generalization of this result for relative twisted Kähler-Eins…

2017-10-04abs ↗pdf ↗

The paper proves Sard's theorem for polynomial maps in infinite dimensions.

problem The validity of Sard's theorem for polynomial maps in infinite-dimensional Banach manifolds.
method Sharp quantitative criteria for the validity of Sard's theorem.
result The paper provides criteria for the validity of Sard's theorem in infinite-dimensional Banach manifolds.

Given a sequence of regular finite coverings of complete Riemannian manifolds, we consider the covering solenoid associated with the sequence. We study the leaf-wise Laplacian on the covering solenoid. The main result is that the spectrum of the Laplacian on the covering solenoid equals the closure of the union of the …

2019-07-15abs ↗pdf ↗

We prove that cubulated hyperbolic groups are virtually special. The proof relies on results of Haglund and Wise which also imply that they are linear groups, and quasi-convex subgroups are separable. A consequence is that closed hyperbolic 3-manifolds have finite-sheeted Haken covers, which resolves the virtual Haken …

2012-04-12abs ↗pdf ↗

Many real datasets contain values missing not at random (MNAR). In this scenario, investigators often perform list-wise deletion, or delete samples with any missing values, before applying causal discovery algorithms. List-wise deletion is a sound and general strategy when paired with algorithms such as FCI and RFCI, b…

2017-05-25abs ↗pdf ↗

xDeepInt learns both vector-wise and bit-wise feature interactions.

problem Learning feature interactions for CTR prediction and recommendation.
method Polynomial Interaction Network (PIN) architecture with subspace-crossing mechanism.
result xDeepInt outperforms state-of-the-art models in CTR prediction and recommendation.

We prove a freeness theorem for low-rank subgroups of one-relator groups. Let FF be a free group, and let wFw\in F be a non-primitive element. The primitivity rank of ww, π(w)π(w), is the smallest rank of a subgroup of FF containing ww as an imprimitive element. Then any subgroup of the one-relator group $G=F/\langle…

2018-03-07abs ↗pdf ↗

Proposes SROF for row-wise fusion in federated learning for multivariate responses.

problem Heterogeneous client models with shared variable-level structure.
method Sparse Row-wise Fusion (SROF) regularizer and RowFed algorithm.
result Empirically shows consistent error reduction and stronger variable-level cluster recovery.

Improves early stopping in deep networks by adjusting stepsizes.

problem Epoch-wise double descent in deep networks.
method Analytical and empirical study of bias-variance tradeoffs in different network layers.
result Eliminating epoch-wise double descent through adjusting stepsizes of different layers improves early stopping performance.

Layer-wise preconditioning methods improve neural network optimization and feature learning.

problem Suboptimal feature learning in standard optimization algorithms.
method Layer-wise preconditioning methods that introduce preconditioners per axis of each layer's weight tensors.
result Layer-wise preconditioning is necessary for provable feature learning in linear and single-index models.

Proposes QEP to mitigate quantization error propagation in layer-wise post-training quantization.

problem Growth of quantization errors across layers degrades performance, especially in low-bit regimes.
method Quantization Error Propagation (QEP) framework that explicitly propagates and compensates for quantization errors.
result QEP-enhanced layer-wise PTQ achieves substantially higher accuracy, especially in low-bit regimes.

The study proves Gaussian universality of deep random features learning.

problem Understanding the test error in deep random features learning.
method Proving Gaussian universality of test error in ridge regression and arbitrary convex losses.
result Sharp asymptotic formula for test error in ridge regression setting.

The paper proves Gorenstein contractions for multiscale differentials on nodal curves.

problem Proving Gorenstein contractions for multiscale differentials on nodal curves.
method Addressing the conjecture by Ranganathan and Wise, showing contractions level by level.
result Multiscale differentials can be contracted to Gorenstein singularities, level by level, from the top down.

We consider billiard ball motion in a convex domain of the Euclidean plane bounded by a piece-wise smooth curve influenced by the constant magnetic field. We show that if there exists a polynomial in velocities integral of the magnetic billiard flow then every smooth piece γγ of the boundary must be algebraic and eith…

2016-05-11abs ↗pdf ↗

A new conformal prediction framework for two-stage models identifies stage-wise uncertainty.

problem Limited coverage guarantees and lack of modular structure understanding in existing conformal prediction methods.
method Decomposes prediction residuals into stage-specific components, calibrates parameters using FWER control, and adapts to non-stationary settings.
result Improves coverage and identifies stage-wise error contributions compared to standard conformal methods.

We develop and analyze efficient "coordinate-wise" methods for finding the leading eigenvector, where each step involves only a vector-vector product. We establish global convergence with overall runtime guarantees that are at least as good as Lanczos's method and dominate it for slowly decaying spectrum. Our methods a…

2017-02-25abs ↗pdf ↗

Random layer-wise pruning profiles are as effective as metric-based ones for various datasets.

problem Reduction of model size and computational resources in neural networks.
method Conducted baseline experiments, developed RL-based search algorithm for finding transferable layer-wise pruning profiles.
result RL-based layer-wise pruning profiles are as good or better than best profiles found on the original dataset via exhaustive search.

This paper explains double descent in linear neural networks, identifying new factors.

problem Understanding double descent in linear neural networks.
method Gradient flow derivation and necessary conditions for double descent.
result Singular values of input-output covariance matrix are important for double descent in two-layer models.

A Finsler space (M,F)(M,F) is called flag-wise positively curved, if for any xMx\in M and any tangent plane PTxM\mathbf{P}\subset T_xM, we can find a nonzero vector yPy\in \mathbf{P}, such that the flag curvature KF(x,y,P)>0K^F(x,y, \mathbf{P})>0. Though compact positively curved spaces are very rare in both Riemannian and Finsler g…

2016-06-06abs ↗pdf ↗

Layer-wise networks have a closed-form solution and a stopping criterion.

problem Training networks one layer at a time without backpropagation.
method Proved the closed-form solution using the kernel Mean Embedding and Neural Indicator Kernel.
result Layer-wise networks have a closed-form solution and a stopping criterion.

New method quantifies uncertainty at class level for better decision-making.

problem Improving cost-sensitive decision-making in classification tasks.
method Label-wise decomposition of uncertainty measures based on non-categorical metrics.
result Proposed measures adhere to desirable properties and improve uncertainty quantification.

New methods estimate point-wise dependency from neural MI models.

problem Estimating point-wise dependency between different events.
method Developed two methods: Probabilistic Classifier and Density-Ratio Fitting.
result Demonstrated effectiveness in MI estimation, self-supervised representation learning, and cross-modal retrieval.

AdaTrans adapts to feature and sample transfer in high-dimensional regression.

problem High-dimensional linear regression with more features than samples.
method F-AdaTrans and S-AdaTrans methods using fused-penalties and adaptive weights.
result AdaTrans achieves convergence rates close to oracle estimators and near-minimax optimal rates.

This paper considers the matrix completion problem. We show that it is not necessary to assume joint incoherence, which is a standard but unintuitive and restrictive condition that is imposed by previous studies. This leads to a sample complexity bound that is order-wise optimal with respect to the incoherence paramete…

2013-10-01abs ↗pdf ↗

Few-shot learning aims to train efficient predictive models with a few examples. The lack of training data leads to poor models that perform high-variance or low-confidence predictions. In this paper, we propose to meta-learn the ensemble of epoch-wise empirical Bayes models (E3BM) to achieve robust predictions. "Epoch…

2019-04-17abs ↗pdf ↗

Linear NDCG is used for measuring the performance of the Web content quality assessment in ECML/PKDD Discovery Challenge 2010. In this paper, we will prove that the DCG error equals a new pair-wise loss.

2013-03-11abs ↗pdf ↗

BlockEcho method improves imputation of block-wise missing data.

problem Block-wise missing data reduces interpolation capability and predictive power.
method Integrates Matrix Factorization (MF) within Generative Adversarial Networks (GAN) to retain long-distance inter-element relationships.
result Superior performance on public datasets across three domains, especially at higher missing rates.

StrTransformer recovers sources without labels by optimizing latent matrices and enforcing structural constraints.

problem Unsupervised blind source recovery in signal processing.
method Source-wise structured Transformer framework with latent source matrix optimization, structural regularization, and branch-specific weights.
result StrTransformer learns distinct temporal-scale structures and recovers source-aligned latent trajectories.

Continuum-wise hyperbolicity is exactly the pseudo-Anosov dynamics with spine singularities.

problem Classification of continuum-wise hyperbolic surface homeomorphisms
method Proving a complete structural classification
result Every cwF_F-hyperbolic homeomorphism is pseudo-Anosov with spine singularities

In the classical best arm identification (Best-11-Arm) problem, we are given nn stochastic bandit arms, each associated with a reward distribution with an unknown mean. We would like to identify the arm with the largest mean with probability at least 1δ1-δ, using as few samples as possible. Understanding the sample c…

2016-08-22abs ↗pdf ↗