Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,042 papers · 148 categories

Trend · papers per month

65130195260 · Jun 202019922001200920172026
48 results for procedure names

Given a finite family of functions, the goal of model selection aggregation is to construct a procedure that mimics the function from this family that is the closest to an unknown regression function. More precisely, we consider a general regression model with fixed design and measure the distance between functions by …

2012-03-12abs ↗pdf ↗

Researchers create a method to join hyperboloidal data sets without violating the shear-free condition.

problem Creating consistent initial data sets for simulations of spacetime.
method Developed a new gluing procedure that maintains the shear-free condition using special Hölder spaces and elliptic operators.
result Successfully constructed hyperboloidal initial data sets that preserve the shear-free condition.

Random forests are powerful non-parametric regression method but are severely limited in their usage in the presence of randomly censored observations, and naively applied can exhibit poor predictive performance due to the incurred biases. Based on a local adaptive representation of random forests, we develop its regre…

2019-02-08abs ↗pdf ↗

Framework mitigates risk non-monotonicity in high-dimensional predictions.

problem Risk non-monotonicity in high-dimensional predictions.
method Model-agnostic framework using cross-validation and data-driven methodologies (zero- and one-step).
result Modified prediction procedures achieve monotonic asymptotic risk behavior.

We solve the integration problem for generalized complex manifolds, obtaining as the natural integrating object a weakly holomorphic symplectic groupoid, which is a real symplectic groupoid with a compatible complex structure defined only on the associated stack, i.e., only up to Morita equivalence. We explain how such…

2016-11-11abs ↗pdf ↗

The paper proves deep ReLU networks can die and proposes a new initialization method to prevent it.

problem Dying ReLU neurons in deep neural networks.
method The paper rigorously proves the dying ReLU problem and proposes a new randomized asymmetric initialization method.
result The new initialization method effectively prevents the dying ReLU problem.

A method to uniformly sample graph-encoded surfaces of fixed size.

problem Sampling uniformly from combinatorial isomorphism types of balanced triangulations of surfaces.
method Relies on connections between graph-encoded surfaces and permutations, and basic properties of the symmetric group.
result Empirical mean genus of the sample is very close to a specific formula as nn increases.

Non-convex optimization is ubiquitous in machine learning. Majorization-Minimization (MM) is a powerful iterative procedure for optimizing non-convex functions that works by optimizing a sequence of bounds on the function. In MM, the bound at each iteration is required to \emph{touch} the objective function at the opti…

2015-06-25abs ↗pdf ↗

This paper deals with finding an nn-dimensional solution xx to a system of quadratic equations of the form yi=ai,x2y_i=|\langle{a}_i,x\rangle|^2 for 1im1\le i \le m, which is also known as phase retrieval and is NP-hard in general. We put forth a novel procedure for minimizing the amplitude-based least-squares empirical los…

2017-05-29abs ↗pdf ↗

We present a methodology to extract the backbone of complex networks based on the weight and direction of links, as well as on nontopological properties of nodes. We show how the methodology can be applied in general to networks in which mass or energy is flowing along the links. In particular, the procedure enables us…

2009-02-05abs ↗pdf ↗

We study adaptive data-dependent dimensionality reduction in the context of supervised learning in general metric spaces. Our main statistical contribution is a generalization bound for Lipschitz functions in metric spaces that are doubling, or nearly doubling. On the algorithmic front, we describe an analogue of PCA f…

2013-02-12abs ↗pdf ↗

ISLET efficiently estimates low-rank tensors with optimal performance and speed.

problem Efficient estimation of low-rank tensors with optimal performance and speed.
method Importance sketching for low-rank tensor estimation.
result ISLET achieves sharp minimax optimality in mean-squared error under low-rank Tucker assumptions.

We study a transformation of metric measure spaces introduced by Gigli and Mantegazza consisting in replacing the original distance with the length distance induced by the transport distance between heat kernel measures. We study the smoothing effect of this procedure in two important examples. Firstly, we show that in…

2016-03-01abs ↗pdf ↗

Wave propagator constructed on globally hyperbolic spacetimes.

problem Wave propagation on globally hyperbolic spacetimes.
method Reinterpretation of wave propagator construction on ultrastatic Lorentzian manifolds, reduction to static backgrounds, generalization to globally hyperbolic spacetimes.
result Global wave propagator constructed on globally hyperbolic spacetimes.

We study a maturity randomization technique for approximating optimal control problems. The algorithm is based on a sequence of control problems with random terminal horizon which converges to the original one. This is a generalization of the so-called Canadization procedure suggested by Carr [Review of Financial Studi…

2006-02-21abs ↗pdf ↗

Efficiently estimates shrinkage coefficient for RTME using LOOCV approximation.

problem Estimating optimal shrinkage coefficient for Regularized Tyler's M-estimator.
method Proposes an approximate LOOCV method to estimate αα efficiently.
result Significant speedup and accuracy improvement over existing methods.

Paper fine-tunes LLMs using user edits, unifying preference, supervision, and reward feedback.

problem Adapting LLMs to user preferences and feedback types.
method Derives bounds for learning algorithms from user edits, proposes an ensembling procedure.
result Ensembling procedure outperforms individual feedback methods and robustly adapts to different user-edit distributions.

DNA-SE uses deep learning to solve semiparametric problems efficiently.

problem Solving semiparametric integral equations in high dimensions.
method Formulates semiparametric estimation as a bi-level optimization problem and uses DNN to approximate solutions.
result Demonstrates numerical and statistical advantages over traditional methods.

The completion of tensors, or high-order arrays, attracts significant attention in recent research. Current literature on tensor completion primarily focuses on recovery from a set of uniformly randomly measured entries, and the required number of measurements to achieve recovery is not guaranteed to be optimal. In add…

2016-11-03abs ↗pdf ↗

There has been a recent surge of interest in studying permutation-based models for ranking from pairwise comparison data. Despite being structurally richer and more robust than parametric ranking models, permutation-based models are less well understood statistically and generally lack efficient learning algorithms. In…

2017-10-28abs ↗pdf ↗

The equivalence problem of curves with values in a Riemannian manifold, is solved. The domain of validity of Frenet's theorem is shown to be the spaces of constant curvature. For a general Riemannian manifold new invariants must thus be added. There are two important generic classes of curves; namely, Frenet curves and…

2012-07-19abs ↗pdf ↗

We show that the herding procedure of Welling (2009) takes exactly the form of a standard convex optimization algorithm--namely a conditional gradient algorithm minimizing a quadratic moment discrepancy. This link enables us to invoke convergence results from convex optimization and to consider faster alternatives for …

2012-03-20abs ↗pdf ↗

Novel method for learning Gaussian graphical models from paired data.

problem Learning Gaussian graphical models for dependent groups.
method Introducing twin order to explore the search space more efficiently.
result The twin order makes the model space a distributive lattice, leading to more efficient model exploration.

In this paper, we consider the problem of "hyper-sparse aggregation". Namely, given a dictionary F={f1,...,fM}F = \{f_1, ..., f_M \} of functions, we look for an optimal aggregation algorithm that writes f~=j=1Mθjfj\tilde f = \sum_{j=1}^M θ_j f_j with as many zero coefficients θjθ_j as possible. This problem is of particular interest when…

2009-12-08abs ↗pdf ↗

Probabilistic graphical models are graphical representations of probability distributions. Graphical models have applications in many fields including biology, social sciences, linguistic, neuroscience. In this paper, we propose directed acyclic graphs (DAGs) learning via bootstrap aggregating. The proposed procedure i…

2014-06-09abs ↗pdf ↗

Quadratic discriminant analysis (QDA) is a standard tool for classification due to its simplicity and flexibility. Because the number of its parameters scales quadratically with the number of the variables, QDA is not practical, however, when the dimensionality is relatively large. To address this, we propose a novel p…

2015-10-01abs ↗pdf ↗

The paper analyzes two ISGD modes for statistical inference, deriving error bounds and confidence intervals.

problem Statistical inference with implicit SGD for smooth convex functions.
method Proximal Robbins-Monro (proxRM) and proximal Polyak-Ruppert (proxPR) procedures for ISGD.
result Derives non-asymptotic error bounds and confidence interval estimators for model parameters.

Enhances count process modelling with Markov-modulated non-homogeneous Poisson process.

problem Count data modelling challenges, especially in complex scenarios.
method Introduces a flexible frequency perturbation measure into Markov-modulated Poisson process framework.
result Natural incorporation of observed event arrivals and latent factors.