Paper proves robust estimators' generalization guarantees without dimensionality issues.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Exact minibatch MH method improves scalability for large datasets.
Exact generalization guarantees for robust models using Wasserstein distance are established.
DPA autoencoders learn data distribution and intrinsic dimensionality with guarantees.
This work provides a guaranteed tensor recovery method by combining low-rankness and smoothness priors.
Exact bounds derived for neural network outputs with noisy inputs.
The paper proposes a method to learn the structure of continuous-action games with non-parametric utilities using a limited number of samples.
This paper solves the multiple reference model problem in RLHF with exact solutions and sample complexity guarantees.
This paper provides an algorithm for simulating improper (or noncircular) complex-valued stationary Gaussian processes. The technique utilizes recently developed methods for multivariate Gaussian processes from the circulant embedding literature. The method can be performed in operations, where…
New variational flows improve Monte Carlo and normalization tasks.
For optimization on large-scale data, exactly calculating its solution may be computationally difficulty because of the large size of the data. In this paper we consider subsampled optimization for fast approximating the exact solution. In this approach, one gets a surrogate dataset by sampling from the full data, and …
Exact causal network discovery is polynomial for sparse networks.
In this paper, we propose a convergent parallel best-response algorithm with the exact line search for the nondifferentiable nonconvex sparsity-regularized rank minimization problem. On the one hand, it exhibits a faster convergence than subgradient algorithms and block coordinate descent algorithms. On the other hand,…
The paper explores partial identifiability in nonnegative matrix factorization under specific conditions.
We consider the Orthogonal Least-Squares (OLS) algorithm for the recovery of a -dimensional -sparse signal from a low number of noisy linear measurements. The Exact Recovery Condition (ERC) in bounded noisy scenario is established for OLS under certain condition on nonzero elements of the signal. The new result a…
New method for community detection in sparse directed SBMs with exact recovery guarantees.
We address some theoretical guarantees for Schatten- quasi-norm minimization () in recovering low-rank matrices from compressed linear measurements. Firstly, using null space properties of the measurement operator, we provide a sufficient condition for exact recovery of low-rank matrices. This condition…
Nonconvex matrix recovery is known to contain no spurious local minima under a restricted isometry property (RIP) with a sufficiently small RIP constant . If is too large, however, then counterexamples containing spurious local minima are known to exist. In this paper, we introduce a proof technique that is capa…
Several useful variance-reduced stochastic gradient algorithms, such as SVRG, SAGA, Finito, and SAG, have been proposed to minimize empirical risks with linear convergence properties to the exact minimizer. The existing convergence results assume uniform data sampling with replacement. However, it has been observed in …
Improved causal discovery methods for large graphs without strict assumptions.
Exact learning improves naive Bayes classifier performance for small samples.
We suggest using the max-norm as a convex surrogate constraint for clustering. We show how this yields a better exact cluster recovery guarantee than previously suggested nuclear-norm relaxation, and study the effectiveness of our method, and other related convex relaxations, compared to other clustering approaches.
New algorithm for computing Wasserstein barycenters with guarantees.
The paper certifies AI reliability via sampling and calibration, providing exact guarantees.
The topology of a power grid affects its dynamic operation and settlement in the electricity market. Real-time topology identification can enable faster control action following an emergency scenario like failure of a line. This article discusses a graphical model framework for topology estimation in bulk power grids (…
Proposes a partitioned least squares model for feature grouping.
This paper considers compressed sensing and affine rank minimization in both noiseless and noisy cases and establishes sharp restricted isometry conditions for sparse signal and low-rank matrix recovery. The analysis relies on a key technical tool which represents points in a polytope by convex combinations of sparse v…
New method solves group synchronization with cycle-edge message passing.
We reformulate data-dependent constraints to ensure they are always met with high probability.
New algorithms improve GP inference without approximations, achieving better results.
This study improves audit sampling by using sequential procedures with statistical guarantees.
Efficiently calculates privacy guarantees for 2020 Census data.
We consider the exact recovery problem in the hypergraph stochastic block model (HSBM) with blocks of equal size. More precisely, we consider a random -uniform hypergraph with vertices partitioned into clusters of size . Hyperedges are added independently with probability if is…
SyncRank recovers global ranking from noisy comparisons with theoretical guarantees.
Efficiently tunes hyperparameters with dynamic accuracy method.
When the linear measurements of an instance of low-rank matrix recovery satisfy a restricted isometry property (RIP)---i.e. they are approximately norm-preserving---the problem is known to contain no spurious local minima, so exact recovery is guaranteed. In this paper, we show that moderate RIP is not enough to elimin…
New method for valid and exact statistical inference of multi-dimensional change-points.
Exact inference method for Wasserstein distance with finite-sample coverage.
LPF provides formal guarantees for aggregating multi-evidence in probabilistic tasks.
Wasserstein distance plays increasingly important roles in machine learning, stochastic programming and image processing. Major efforts have been under way to address its high computational complexity, some leading to approximate or regularized variations such as Sinkhorn distance. However, as we will demonstrate, regu…
T-SCI improves Cox-MLP's guaranteed coverage for censored data.
This paper investigates gradient recovery schemes for data defined on discretized manifolds. The proposed method, parametric polynomial preserving recovery (PPPR), does not require the tangent spaces of the exact manifolds, and they have been assumed for some significant gradient recovery methods in the literature. Ano…
Recently theoretical guarantees have been obtained for matrix completion in the non-uniform sampling regime. In particular, if the sampling distribution aligns with the underlying matrix's leverage scores, then with high probability nuclear norm minimization will exactly recover the low rank matrix. In this article, we…
Low-rank factorization is a standard way to make structured optimization problems in machine learning more tractable by replacing matrix variables with compact factors. For positive semidefinite (PSD) variables, the symmetric Burer--Monteiro factorization (sBMF) writes with a single low-rank factor . A r…
I analyse the frequentist regret of the famous Gittins index strategy for multi-armed bandits with Gaussian noise and a finite horizon. Remarkably it turns out that this approach leads to finite-time regret guarantees comparable to those available for the popular UCB algorithm. Along the way I derive finite-time bounds…
In this note we compare two recently proposed semidefinite relaxations for the sparse linear regression problem by Pilanci, Wainwright and El Ghaoui (Sparse learning via boolean relaxations, 2015) and Dong, Chen and Linderoth (Relaxation vs. Regularization A conic optimization perspective of statistical variable select…
More and more AI services are provided through APIs on cloud where predictive models are hidden behind APIs. To build trust with users and reduce potential application risk, it is important to interpret how such predictive models hidden behind APIs make their decisions. The biggest challenge of interpreting such predic…
FedSARSA converges with heterogeneous agents, achieving linear speed-up.