Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

6.3%12.5%18.8%25.0% · Jul 199319922001200920182026
48 results for Self-Normalized Deviation Inequalities

Study improves self-normalized bounds for vector-valued processes beyond sub-Gaussianity.

problem Limited understanding of self-normalized concentration for vector-valued processes outside sub-Gaussian frameworks.
method Developed concentration inequalities for self-normalized processes with light tails (e.g., Bennett, Bernstein bounds) for vector-valued data.
result Provided new insights and bounds for self-normalized processes with non-sub-Gaussian distributions.

Calculation of the log-normalizer is a major computational obstacle in applications of log-linear models with large output spaces. The problem of fast normalizer computation has therefore attracted significant attention in the theoretical and applied machine learning literature. In this paper, we analyze a recently pro…

2015-06-12abs ↗pdf ↗

New bounds on self-normalized martingales improve online linear regression performance.

problem Improving regret bounds in online linear regression.
method Characterizing scale-invariant bounds on self-normalized martingales.
result For d=1d=1, O(logT)O(\log T) doubly-uniform regret is possible; for d>1d>1, sublinear doubly-uniform regret is impossible.

We improve bounds for stochastic processes, especially those with heavy tails.

problem Bounding the concentration of sub-ψψ processes with heavy tails.
method Variational approach to concentration, focusing on sub-Gaussian and other tail conditions.
result First dimension-free self-normalized empirical Bernstein inequality.

A new method for evaluating and selecting policies in contextual bandits improves confidence intervals and policy quality.

problem Evaluating and selecting policies in contextual bandits with logged data.
method Self-normalized Importance Weighting (SN) estimator with Efron-Stein tail inequality and multiplicative bias control.
result The method provides tighter confidence intervals and better policy selection compared to competitors.

Unified stopping rules ensure accurate policies in contextual learning.

problem Stopping data collection to ensure accurate policies in personalized decision problems.
method Developed unified stopping rules based on GLR statistics for pairwise action comparisons.
result Unified stopping rules achieve target precision with fewer samples than benchmarks.

Deviation inequalities for stochastic approximation methods.

problem Establishing bounds on the deviation of stochastic approximation methods.
method Martingale approximation method for separately Lipschitz functions.
result Established various deviation inequalities for stochastic approximation by averaging and minimization.

Deviation inequalities and limit laws for random walks on metric spaces.

problem Understanding random walks on metric spaces with contracting isometries.
method Adapting Gouëzel's pivotal time construction to establish deviation inequalities.
result Exponential bounds and limit laws for random walks on mapping class groups and CAT(0) spaces.

New inequalities for matrix supermartingales converge under various conditions.

problem Convergence and maximal inequalities of supermartingales in positive semidefinite matrices.
method Developed new concentration inequalities for matrix supermartingales.
result New inequalities for matrix supermartingales under different tail conditions.

New inequality criterion for a mean field equation on spheres.

problem Finding uniqueness in a mean field equation on spheres.
method Established a new Moser-Trudinger-Onofri inequality with a constraint on moments deviation.
result A threshold for deviation is a uniqueness criterion for the mean field equation.

The paper improves inequalities for nearly spherical sets using quermassintegrals.

problem Improving inequalities for nearly spherical sets.
method Establishing quantitative Alexandrov-Fenchel inequalities for quermassintegrals.
result Lower bounds on the (k,m)(k,m)-isoperimetric deficit found using spherical deviation and asymmetry.

Novel concentration inequalities are obtained for the missing mass, i.e. the total probability mass of the outcomes not observed in the sample. We derive distribution-free deviation bounds with sublinear exponents in deviation size for missing mass and improve the results of Berend and Kontorovich (2013) and Yari Saeed…

2015-03-20abs ↗pdf ↗

The paper provides a finite-sample deviation bound for stable autoregressive processes.

problem Deviation bounds for least squares estimators in Gaussian AR(n) processes.
method Utilizes martingale concentration inequalities and tail-bound for χ² distributed variables.
result Problem-dependent finite-time bound on the deviation probability of AR(n) process parameters.

In this paper, we study the risk bounds for samples independently drawn from an infinitely divisible (ID) distribution. In particular, based on a martingale method, we develop two deviation inequalities for a sequence of random variables of an ID distribution with zero Gaussian component. By applying the deviation ineq…

2012-02-14abs ↗pdf ↗

New algorithm reduces regret for logistic bandits without κκ dependency.

problem Logistic bandits have poor frequentist regret guarantees due to large κκ.
method Optimistic algorithm based on self-normalized martingale tail-inequality.
result Achieves ildeO(T) ilde{\mathcal{O}}(\sqrt{T}) regret with no κκ dependency.

We are concerned with obtaining novel concentration inequalities for the missing mass, i.e. the total probability mass of the outcomes not observed in the sample. We not only derive - for the first time - distribution-free Bernstein-like deviation bounds with sublinear exponents in deviation size for missing mass, but …

2015-03-10abs ↗pdf ↗

The paper develops a method for self-normalized inference in adaptive experiments.

problem Adaptive experiments require a fixed horizon for ATE estimation, but propensities can change.
method The method uses self-normalized martingale limit theory to estimate ATE.
result The Studentized statistic is asymptotically N(0,1) at the prespecified horizon.

Sharp concentration results for sums of heavy-tailed random variables.

problem Analyzing sums of independent heavy-tailed random variables.
method Using concentration inequalities and large deviation principles for distributions satisfying specific tail bounds.
result Sharp concentration inequalities and large deviation results for sums of heavy-tailed random variables.

This paper studies semiparametric contextual bandits, a generalization of the linear stochastic bandit problem where the reward for an action is modeled as a linear function of known action features confounded by an non-linear action-independent term. We design new algorithms that achieve O~(dT)\tilde{O}(d\sqrt{T}) regret …

2018-03-12abs ↗pdf ↗

The paper provides bounds for high-dimensional U-statistics with novel order-explicit inequalities.

problem Bounding the deviation of high-dimensional U-statistics from their Hájek projections.
method Develops novel order-explicit moment inequalities for higher-order Hoeffding components.
result The maximum deviation of a high-dimensional U-statistic from its Hájek projection is of order Op(φbn1log2(dn))O_p(φb n^{-1}\log^2(dn)).

Study on discrepancy principle for learning algorithms in nonparametric regression.

problem Determining optimal iteration number in nonparametric regression with unknown optimal iteration.
method Investigates discrepancy principle and modified principles for kernelized spectral filters, using deviation inequalities and change-of-norm arguments.
result Classical discrepancy principle is adaptive for slow rates, while modified principles are adaptive for faster rates.

A new linear contextual bandit algorithm with improved regret bound.

problem Efficiently solving linear contextual bandit problems with reduced regret.
method Proposes a novel estimator embedded with exploration and a self-normalized bound.
result Regret bound matches lower bound of Ω(dT)Ω(\sqrt{dT}) up to logarithmic factors.

Self Normalizing Flows improve normalizing flows by reducing computational complexity.

problem Efficient gradient computation in normalizing flows, especially in Jacobian determinant terms.
method Introducing Self Normalizing Flows that replace expensive terms with learned approximate inverses.
result Models can be trained more quickly and perform better than functionally constrained counterparts.

Improved concentration inequalities for sub-Weibull variables enhance statistical and machine learning applications.

problem Improving concentration inequalities for sub-Weibull random variables.
method Developed new concentration inequalities for sums of independent sub-Weibull random variables, including a new sub-Weibull parameter.
result New concentration inequalities with sharper constants and a mixture of sub-Gaussian and sub-Weibull tails.

The paper improves generalization bounds for classifier chains with interdependent labels.

problem Improving generalization for classifier chains with multiple interdependent labels.
method Using large deviation inequalities for weakly dependent sequences, the paper derives a new generalization error bound.
result The derived bound explicitly shows dependencies between class labels and provides insights into the chain's order.

Paper develops a new inequality for non-causal machine learning.

problem Current concentration inequalities cannot be applied to non-causal machine learning.
method Develops a framework for non-causal random fields and proves a Hoeffding-type inequality.
result Obtains a Hoeffding-type concentration inequality for non-causal random fields.

New activation function SERLU improves neural network performance.

problem Improving neural network performance and avoiding overfitting.
method Introducing a new activation function (SERLU) that breaks monotonicity while preserving self-normalizing property and developing shift-dropout for regularization.
result SERLU-based neural networks provide consistently promising results compared to other activation functions.

Sharp upper bounds derived for Alexandrov-Fenchel deficit using weighted Minkowski integral formulas.

problem Deriving upper bounds for the Alexandrov-Fenchel deficit.
method Using weighted Minkowski integral formulas and an integral formula for the deficit in Jensen's inequality.
result Quantitative estimates under weaker convexity assumptions, including a distance term.

This article provides a new toolbox to derive sparse recovery guarantees from small deviations on extreme singular values or extreme eigenvalues obtained in Random Matrix Theory. This work is based on Restricted Isometry Constants (RICs) which are a pivotal notion in Compressed Sensing and High-Dimensional Statistics a…

2016-04-05abs ↗pdf ↗