Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,341 papers · 148 categories

Trend · papers per month

2.6%5.1%7.7%10.2% · Jun 201219922001200920182026
48 results for sparse fit

The paper tackles identifying a small segment of a population with sparse linear regression.

problem Identifying a small segment of a population with a sparse linear regression fit.
method Algorithms for joint identification of a significant segment of a population with a sparse linear regression fit under the sup norm, using k-DNF conditions and s-sparse regression fits.
result Preliminary algorithms and challenges for future work in non-sparse regression and expected error.

Kernel Multigrid accelerates Back-fitting for additive Gaussian Processes.

problem Slow convergence of Back-fitting in training additive Gaussian Processes.
method Kernel Packets (KP) and Sparse Gaussian Process Regression (GPR) to enhance Back-fitting.
result Kernel Multigrid reduces the required iterations to O(logn)\mathcal{O}(\log n).

Efficiently finds sparse solutions to max-plus equations for convex regression.

problem Finding sparse solutions to max-plus equations for convex multivariate regression.
method Polynomial-time algorithm for sparse approximate solutions.
result Optimal piecewise-linear fitting with minimum number of regions.

A new distributed algorithm for fitting sparse additive models with feature division and decorrelation.

problem Fitting high-dimensional sparse additive models efficiently and accurately.
method Divide, decorrelate, and conquer approach.
result Effective and efficient recovery of sparsity patterns and statistical inference for each component.

SAEs struggle with curved activation manifolds, revealing layer-dependent scaling laws.

problem Sparse autoencoders' reconstruction error varies across layers, not fitting existing scaling laws.
method Cross-layer study of 844 SAE checkpoints, fitting and regressing on manifold geometry.
result Manifold geometry predicts layer-dependent width exponents in SAEs, with transferable coefficients.

The paper tackles sparse model fitting in distributed machine learning with graph-structured data.

problem Sparse model fitting across a distributed collection of heterogeneous data sets.
method Basis Pursuit Denoising with a total variation penalty, using ADMM for distributed methods.
result Recovery is successful with fewer samples than solving problems independently, or using methods with large overlap in signal supports.

New estimators improve sparse semiparametric additive modeling.

problem Sparse semiparametric additive modeling with structured sparsity.
method Combines group subset selection with shrinkage for nonconvex optimization.
result New estimators outperform alternatives in synthetic and real-world data.

A new framework DECO for distributed sparse regression reduces model dimensionality and improves accuracy.

problem Sparse regression challenges in high-dimensional datasets.
method DECO framework for feature space partitioning, decorrelating features, and distributed computation.
result DECO achieves consistent variable selection and parameter estimation with nearly optimal convergence rate.

We introduce a new algorithm, called adaptive sparse backfitting algorithm, for solving high dimensional Sparse Additive Model (SpAM) utilizing symmetric, non-negative definite smoothers. Unlike the previous sparse backfitting algorithm, our method is essentially a block coordinate descent algorithm that guarantees to …

2014-09-08abs ↗pdf ↗

Sparse feature selection improves batch RL efficiency.

problem High-dimensional batch RL with many features.
method Sparse linear function approximation, Lasso, group Lasso, fitted Q-evaluation, fitted Q-iteration.
result Sparse feature selection makes batch RL more sample efficient.

The power of sparse signal modeling with learned over-complete dictionaries has been demonstrated in a variety of applications and fields, from signal processing to statistical inference and machine learning. However, the statistical properties of these models, such as under-fitting or over-fitting given sets of data, …

2011-10-11abs ↗pdf ↗

A new estimator learns sparse linear models with context-dependent coefficients.

problem Sparse linear models lack flexibility compared to deep neural networks for handling feature groups.
method Contextual lasso estimator using a deep neural network with lasso regularization.
result Learned models can be sparser than standard lasso without sacrificing predictive power.

Gaussian Processes improve data interpolation from diverse experiments.

problem Interpolation of sparse and inconsistent datasets from various experiments.
method Used Gaussian Processes (GP) for data interpolation, including uncertainty quantification.
result GPs successfully interpolate data and quantify uncertainties, demonstrating consistency across different sources.

There has been a lot of work fitting Ising models to multivariate binary data in order to understand the conditional dependency relationships between the variables. However, additional covariates are frequently recorded together with the binary data, and may influence the dependence relationships. Motivated by such a d…

2012-09-27abs ↗pdf ↗

Regularized MLE for MoE models tackles high-dimensional heterogeneous data.

problem Fitting and feature selection in Mixtures-of-Experts models for high-dimensional data.
method Proposes a regularized maximum likelihood estimation approach with hybrid EM/MM algorithms.
result Automatic recovery of sparse solutions without thresholding and matrix inversion.

New method estimates multivariate Gaussian fields using sparse precision matrix.

problem Estimating covariance matrices for large multivariate Gaussian fields.
method Sparse Precision Matrix Selection (SPS) algorithm for multivariate GRFs.
result Theoretical rates of convergence for estimated covariance and parameters validated.

Sparse matrices simplify computation of GP variances and likelihoods.

problem Efficient computation of posterior variance and log-likelihood for additive Matérn GPs.
method Represented posterior mean, variance, log-likelihood, and gradient using sparse matrices.
result Efficient computation of posterior mean, variance, log-likelihood, and gradient in O(nlogn)O(n \log n) time.

Sparse-input neural networks handle high-dimensional data with fewer features.

problem Neural networks struggle with high-dimensional data where input features exceed observations.
method Sparse group lasso penalty on first-layer input weights.
result Sparse-input neural networks achieve better performance than existing methods in high-dimensional data with complex interactions.

The paper offers streamlined algorithms for fitting complex linear mixed models.

problem Linear mixed models with crossed random effects in large dimensions.
method Mean field variational Bayes algorithms with various relaxations and storage strategies.
result Different inference strategies have varying trade-offs between accuracy and computational demands.

The study reveals decision trees' limitations in fitting data from additive models, proving a generalization lower bound.

problem Understanding the generalization performance of decision trees on additive models.
method Analyzing decision tree algorithms with sparse additive models, proving generalization lower bounds.
result Generalization lower bounds for decision trees on sparse additive models are much worse than minimax rates.

New method uses reinforcement learning to sample from complex data structures efficiently.

problem Constructing reliable samples from high-dimensional polytopes for goodness-of-fit tests.
method Markov decision process and reinforcement learning for sampling.
result Demonstrated scalable tools from linear algebra for theoretical guarantees in non-linear algebra context.

New methods improve deep learning models for sparse, high-dimensional data.

problem Underfitting in inference networks for sparse, high-dimensional data.
method Iterative optimization inspired by stochastic variational inference and improvements in sparse data representation.
result State-of-the-art results on text-count dataset and excellent recommendation results.

New algorithms find conditions and linear rules with high probability and loss.

problem Finding conditions and rules with high probability and loss in conditional sparse regression.
method Efficient algorithms for identifying conditions and rules with optimal probability and loss.
result Achieved algorithms that nearly match the probability of the ideal condition and improve the approximation to the target loss.

We study the problem of variable selection in convex nonparametric regression. Under the assumption that the true regression function is convex and sparse, we develop a screening procedure to select a subset of variables that contains the relevant variables. Our approach is a two-stage quadratic programming method that…

2014-11-07abs ↗pdf ↗

The paper tackles reward-relevance in offline RL with sparse decision dynamics.

problem Offline reinforcement learning with sparse decision dynamics and estimation sparsity.
method Reward-filtered least-squares policy evaluation using thresholded lasso.
result The method provides theoretical guarantees with sample complexity dependent on sparse component size.

New method detects nonlinear causality in multivariate time series data.

problem Detecting nonlinear causal relationships in multidimensional time series.
method Sparse additive models (SpAMs) with B-spline bases and group-lasso optimization.
result The method can accurately estimate nonlinear causal relationships in β-mixing time series.

New method discovers causal relationships in sparse linear data.

problem Discovering cause-effect relationships in sparse linear data.
method Uses structural matrix to reconstruct data and identify causal structures without independence tests.
result Outperforms existing methods in sparse causal structure recovery.

Gaussian graphical models (GGM) have been widely used in many high-dimensional applications ranging from biological and financial data to recommender systems. Sparsity in GGM plays a central role both statistically and computationally. Unfortunately, real-world data often does not fit well to sparse graphical models. I…

2014-06-10abs ↗pdf ↗

SRF learns sparse rule models by screening out features efficiently.

problem Learning optimal sparse rule models is computationally intractable due to the large number of possible rules.
method SRF uses meta safe screening (mSS) to efficiently screen out multiple features, improving the learning of sparse rule models.
result SRF provides a general framework for fitting sparse rule models and can handle group regularization.

Proposes a method to emulate sparse priors using L1 regularization without complex transformations.

problem Sparse priors in under-determined estimation problems.
method Parameter transform to emulate sparse priors under L2 regularization.
result L1 regularization can be achieved with a remapping of parameters under normal priors.

Jointly learns feature and sample relevancies for robust sparse recovery.

problem Sparse recovery sensitivity to data contaminants like outliers or misspecified noise.
method Jointly learns feature and sample relevancies via marginal likelihood optimization.
result Consistent sparse and robust prediction models across diverse tasks.