Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

131261392522 · Jun 202019922001200920172026
48 results for smoothness parameters

The paper establishes bounds on the smoothness parameter in Gaussian process interpolation.

problem Estimating the smoothness parameter in Gaussian process models.
method Approximation theory in Sobolev spaces and general theorems on parameter estimation.
result Maximum likelihood estimation recovers the true smoothness for certain classes of functions.

Smoothness analysis of adversarial training reveals LL_\infty constraints cause more non-smoothness.

problem Non-smoothness of adversarial training loss function.
method Analyzed the smoothness of adversarial training loss function using optimal attacks for model parameters.
result The LL_\infty constraint causes more non-smoothness than L2L_2 constraint.

A flow defined by a nonsingular smooth vector field XX on a closed manifold MM is said to be parameter rigid if given any real valued smooth function ff on MM, there are a smooth funcion gg and a constant cc such that f=X(g)+cf=X(g)+c holds. We show that the parameter rigid flows on closed orientable 3-manifolds are sm…

2010-02-01abs ↗pdf ↗

NSGD-M optimizes machine learning models without hyperparameter tuning, even under relaxed smoothness.

problem Training machine learning models with optimal complexity under relaxed smoothness assumptions.
method Normalized Stochastic Gradient Descent with Momentum (NSGD-M) without stepsize tuning.
result NSGD-M achieves nearly optimal complexity without prior knowledge of problem parameters.

Deep neural networks with specific parameter sets can approximate smooth functions efficiently.

problem Approximating smooth functions with deep neural networks.
method Deep neural networks with ReLU activation and specific parameter sets {0,±12,±1,2}\{0,\pm \frac{1}{2}, \pm 1, 2\} are used to approximate CβC_β-smooth functions.
result The constructed networks can approximate CβC_β-smooth functions with parameters {0,±12,±1,2}\{0,\pm \frac{1}{2}, \pm 1, 2\} efficiently, achieving the same convergence rate as sparse networks with parameters in [1,1][-1,1].

A new method for efficient BNC parameter estimation outperforms HDP smoothing.

problem Efficiently estimating parameters for Bayesian network classifiers to match or exceed random forest performance.
method Uses log-linear regression to approximate hierarchical Dirichlet process (HDP) smoothing, making the approach simpler and faster.
result Our method outperforms HDP smoothing while being orders of magnitude faster and competitive with random forests.

A non-Bayesian, regression-based or generalized least squares (GLS)-based approach is formally proposed to estimate a class of time-varying AR parameter models. This approach has partly been used by Ito et al. (2014, 2016a,b), and is proven to be efficient because, unlike conventional methods, it does not require Kalma…

2017-07-21abs ↗pdf ↗

PF-LaCG removes the need for knowing smoothness and strong convexity parameters for locally accelerated CG.

problem Locally accelerated CG requires knowledge of smoothness and strong convexity parameters.
method Parameter-Free Locally Accelerated CG (PF-LaCG) algorithm.
result PF-LaCG achieves local acceleration without requiring knowledge of smoothness and strong convexity parameters.

For a smooth family of exact forms on a smooth manifold, an algorithm for computing a primitive family smoothly dependent on parameters is given. The algorithm is presented in the context of a diagram chasing argument in the Čech-de Rham complex. In addition, explicit formulas for such primitive family are presented.

2019-03-19abs ↗pdf ↗

Efficient estimators for smooth Hilbert-valued parameters with theoretical guarantees.

problem Estimating smooth Hilbert-valued parameters with theoretical guarantees.
method Pathwise differentiable Hilbert-valued parameters, efficient influence functions, regularized one-step estimators.
result Theoretical guarantees for efficient estimators even when nuisance functions are arbitrary.

New matching estimators correct bias in multivariate settings without smoothing parameters.

problem Bias in nearest-neighbor and matching estimators in multiple dimensions.
method Polynomial least squares fits on Voronoi tessellations.
result Novel estimators converge at n\sqrt{n} rate under mild smoothness assumptions.

T-LoHo model detects structured sparsity and smoothness on graph data.

problem Detecting structured sparsity and smoothness in graph-structured data.
method Tree-based Low-rank Horseshoe (T-LoHo) prior for multivariate parameters.
result Improves anomaly detection on road networks compared to other methods.

A comprehensive methodology is provided for smoothing noisy, irregularly sampled data with non-Gaussian noise using smoothing splines. We demonstrate how the spline order and tension parameter can be chosen a priori from physical reasoning. We also show how to allow for non-Gaussian noise and outliers which are typical…

2019-04-26abs ↗pdf ↗

Improved regret bounds for structured linear contextual bandits with Gaussian noise.

problem Optimizing bandit learning algorithms for structured contexts with Gaussian perturbations.
method Proposed simple greedy algorithms for structured linear contextual bandits with Gaussian noise.
result Unified regret analysis for structured parameters with geometric quantities as bounds.

Averaged SGD optimizes a smoothed objective, leading to better generalization.

problem Improving generalization performance in machine learning models.
method Analyzed the smoothed objective function of SGD and proved that averaged SGD can optimize this smoothed function efficiently.
result Averaged SGD can efficiently optimize a smoothed objective, leading to better generalization.

Novel method for SDE calibration from sparse data using neural flows.

problem Calibrating SDEs from sparse, noisy observations.
method Characterization of posterior SDE using neural networks trained to solve a PDE with multiplicative updates.
result Significant improvement in scalability and accuracy compared to classical methods.

Here, we study different update rules in stochastic gradient descent (SGD) for online forecasting problems. The selection of the learning rate parameter is critical in SGD. However, it may not be feasible to tune this parameter in online learning. Therefore, it is necessary to have an update rule that is not sensitive …

2019-05-21abs ↗pdf ↗

We say that a topologically embedded 3-sphere in a smoothing of Euclidean 4-space is a barrier provided, roughly, no diffeomorphism of the 4-manifold moves the 3-sphere off itself. In this paper we construct infinitely many one parameter families of distinct smoothings of 4-space with barrier 3-spheres. \par The existe…

1998-07-26abs ↗pdf ↗

A new federated learning algorithm improves on existing methods by exploiting data smoothness.

problem Federated learning optimization with smooth loss functions.
method Federated Low Rank Gradient Descent (FedLRGD) algorithm.
result FedLRGD outperforms Federated Averaging (FedAve) in federated oracle complexity under certain conditions.

This work introduces two strategies for training network classifiers with heterogeneous agents. One strategy promotes global smoothing over the graph and a second strategy promotes local smoothing over neighbourhoods. It is assumed that the feature sizes can vary from one agent to another, with some agents observing in…

2019-10-30abs ↗pdf ↗

Untuned SGD converges but with an exponential dependence on smoothness, adaptive methods prevent this.

problem The exponential dependence on smoothness in untuned SGD's convergence rate.
method Untuned SGD with arbitrary stepsize η, adaptive methods like NSGD, AMSGrad, and AdaGrad.
result Adaptive methods prevent the exponential dependence on smoothness in SGD.

The paper proves weaker conditions for global smoothings of special Lagrangian submanifolds with conical singularities.

problem Conditions for global smoothings of special Lagrangian submanifolds with isolated conical singularities.
method Proof of weaker conditions for global smoothings.
result Global smoothings are possible under weaker hypotheses than previously known.

If the fundamental group of the complement of a smooth embedding f: S^2 \subset R^4 is a cyclic group, the map can be deformed to the standard embedding by a generic one-parameter family with at most cusp singularities. If two smooth embeddings are connected by such a deformation, they will be called cusp equivalent. W…

1999-11-20abs ↗pdf ↗

Multiple generalized additive models (GAMs) are a type of distributional regression wherein parameters of probability distributions depend on predictors through smooth functions, with selection of the degree of smoothness via L2L_2 regularization. Multiple GAMs allow finer statistical inference by incorporating explana…

2018-09-25abs ↗pdf ↗

We perform a geometric study of the equilibrium locus of the flow that models the diffusion process over a circular network of cells. We prove that when considering the set of all possible values of the parameters, the equilibrium locus is a smooth manifold with corners, while for a given value of the parameters, it is…

2015-09-25abs ↗pdf ↗

A neural network with a single hidden layer can't represent certain multivariable functions.

problem Representing certain multivariable functions with a neural network having only one hidden layer.
method Developed a continuum version of a one-hidden-layer neural network with ReLU activation, and proved constraints on its parameters and second derivative.
result Existence of a smooth binary function that cannot be precisely represented by any such neural network.

Dropout improves regularization in flexible models for rare features.

problem Understanding theoretical properties of dropout in generalized linear models.
method Theoretical analysis and application to adaptive smoothing with B-splines.
result Dropout prefers rare features in mean and dispersion parameters.

We consider the non-parametric regression problem under Huber's εε-contamination model, in which an εε fraction of observations are subject to arbitrary adversarial noise. We first show that a simple local binning median step can effectively remove the adversary noise and this median estimator is minimax optimal up t…

2018-05-26abs ↗pdf ↗

The study reveals a persistent bias in the distribution of holonomy on compact hyperbolic 3-manifolds.

problem The distribution of holonomy on compact hyperbolic 3-manifolds is not uniformly distributed.
method An asymptotic count of closed geodesics by their length and holonomy, and analysis of spectral parameters.
result A normalized, smoothed bias count of holonomy is distributed according to a probability distribution, controlled by the number of zero spectral parameters.

This paper improves neural network approximation for analytic functions with adjustable depth and width.

problem Approximating analytic functions using neural networks with depth and width parameters.
method Characterizes approximation rates as a joint function of width (N) and depth (L) for ReLU networks.
result Establishes upper bounds for analytic function approximation rates of O(N^(-CL^τ)) with τ influenced by N and L.

Proposes a new method for estimating non-pathwise differentiable functional parameters.

problem Estimating dose-response curves for continuous exposure.
method Targeted Highly Adaptive Lasso (HAL) for non-pathwise differentiable functional parameters.
result The Targeted HAL-MLE achieves dimension-free rates up to log(n) factors and outperforms other methods in simulations.

Black-box variational inference tries to approximate a complex target distribution though a gradient-based optimization of the parameters of a simpler distribution. Provable convergence guarantees require structural properties of the objective. This paper shows that for location-scale family approximations, if the targ…

2019-01-24abs ↗pdf ↗

We present a 1-parameter family of finite action solutions to the S0(2,1)S0(2,1) Hitchin's equations and explore some of its basic properties. For a fixed value of the parameter, the solution is smooth. We conclude by showing a multi-particle generalization of our basic solutions.

2000-11-13abs ↗pdf ↗