Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

62123185246 · May 202619922001200920172026
48 results for Deviation Control

MPC framework reduces execution costs and schedule deviations in trading.

problem Executing large orders in markets under time and liquidity constraints.
method Model Predictive Control (MPC) framework balancing order completion, market impact, and opportunity cost.
result Significant reductions in slippage and schedule shortfall compared to benchmarks.

SAM improves generalization in overparameterized models, but its behavior in tensorized models is less understood.

problem Understanding the implicit regularization of SAM in tensorized models.
method Scale-invariance analysis and gradient flow analysis to derive Norm Deviation as a measure of core norm imbalance, and propose Deviation-Aware Scaling (DAS).
result DAS achieves competitive or improved performance over SAM, while offering reduced computational overhead.

This study improves audit sampling by using sequential procedures with statistical guarantees.

problem Improving audit efficiency and reliability with statistical methods.
method Formulated as a sequential testing problem, defining null and alternative hypotheses, stopping and decision rules, and exact boundary conditions.
result Exact design yields ex ante control of decision error probabilities, and simulation-based implementation approximates this design.

Proposes a two-stage method for testing variable interactions with FDR control.

problem Testing pairwise interactions in high-dimensional data with dependence.
method Two-stage testing procedure with FDR control using Cramér type moderate deviation technique.
result The proposed method controls FDR and has comparable or improved statistical power.

The aim of this paper is to generalize the PAC-Bayesian theorems proved by Catoni in the classification setting to more general problems of statistical inference. We show how to control the deviations of the risk of randomized estimators. A particular attention is paid to randomized estimators drawn in a small neighbor…

2007-12-11abs ↗pdf ↗

We provide a direct proof of Cramér's theorem for geodesic random walks in a complete Riemannian manifold (M,g)(M,g). We show how to exploit the vector space structure of the tangent spaces to study large deviation properties of geodesic random walks in MM. Furthermore, we reveal the geometric obstructions one runs into …

2018-11-23abs ↗pdf ↗

We consider a long-term optimal investment problem where an investor tries to minimize the probability of falling below a target growth rate. From a mathematical viewpoint, this is a large deviation control problem. This problem will be shown to relate to a risk-sensitive stochastic control problem for a sufficiently l…

2010-01-13abs ↗pdf ↗

Three training methods for language models are shown to be variations of one another.

problem Training language models to reason effectively using different methods.
method Three training methods: GRPO, Dr. GRPO, and DAPO.
result All three methods adjust a single number: standard deviation, measuring disagreement in answers.

Interpolating models can have heavy-tailed risk, leading to rare but severe errors.

problem Interpolating models' tail risk is poorly understood, affecting rare but impactful errors.
method Large-deviation methods to study the fragility of high-dimensional linear interpolators.
result Ridgeless regression exhibits heavy-tailed risk, while ridge-regularized estimators have better tail behavior.

Control charts have traditionally been used in industrial statistics, but are constantly seeing new areas of application, especially in the age of Industry 4.0. This paper introduces a new method, which is suitable for applications in the healthcare sector, especially for monitoring a health-characteristic of a patient…

2019-02-14abs ↗pdf ↗

Distribution grids are currently challenged by frequent voltage excursions induced by intermittent solar generation. Smart inverters have been advocated as a fast-responding means to regulate voltage and minimize ohmic losses. Since optimal inverter coordination may be computationally challenging and preset local contr…

2018-07-10abs ↗pdf ↗

Study develops a data-based model for in-cylinder pressure and cyclic variations in RCCI engines.

problem Lack of models capturing cyclic variations in combustion concepts like RCCI.
method Combines Principle Component Decomposition and Gaussian Process Regression.
result Model predicts combustion measures with high accuracy, especially peak-pressure rise-rate.

Risk control and optimal diversification constitute a major focus in the finance and insurance industries as well as, more or less consciously, in our everyday life. We present a discussion of the characterization of risks and of the optimization of portfolios that starts from a simple illustrative model and ends by a …

1998-02-05abs ↗pdf ↗

New framework embeds generalization in learning dynamics using large deviation theory.

problem Improving generalization and robustness in learning problems.
method Gradient methods from continuous-time perspective with Freidlin-Wentzell theory of large deviations.
result Asymptotic probability estimate for rare events in learning dynamics.

DQNs can approximate optimal Q-functions with high accuracy on compact sets.

problem Approximating optimal Q-functions in continuous-time Markov Decision Processes.
method Stochastic control, FBSDEs, residual network approximation theorems, large deviation bounds, viscosity solutions.
result DQNs can approximate optimal Q-functions on compact sets with arbitrary accuracy and high probability.

Vanishing gradients hinder reinforcement finetuning of language models.

problem Vanishing gradients impede the optimization of language models using reinforcement finetuning.
method The study identifies vanishing gradients as a fundamental optimization obstacle in reinforcement finetuning and proposes an initial supervised finetuning phase to mitigate this issue.
result An initial supervised finetuning phase is crucial for successful reinforcement finetuning of language models, as it helps prevent vanishing gradients and maximizes rewards.

Suppose kk centers are fit to mm points by heuristically minimizing the kk-means cost; what is the corresponding fit over the source distribution? This question is resolved here for distributions with p4p\geq 4 bounded moments; in particular, the difference between the sample cost and distribution cost decays with $…

2013-11-08abs ↗pdf ↗

Paper relaxes symmetry conditions for universal feature selection in noisy data.

problem Feature selection in noisy data with weak symmetry.
method Developed a universal feature selection framework using singular value decomposition of canonical dependence matrix.
result Selected features achieve asymptotically optimal error exponents up to a residual term.

Unified framework for FDR control in knockoffs, validating Gaussian knockoffs.

problem Asymptotic FDR control in knockoffs with user-specified distributions.
method Unified theoretical framework, three conditions on approximate knockoff statistics, Gaussian knockoffs generator based on moments matching.
result Gaussian knockoffs generator achieves asymptotic FDR control.

This survey reviews portfolio selection problem for long-term horizon. We consider two objectives: (i) maximize the probability for outperforming a target growth rate of wealth process (ii) minimize the probability of falling below a target growth rate. We study the asymptotic behavior of these criteria formulated as l…

2014-08-27abs ↗pdf ↗

For a stochastic factor model we maximize the long-term growth rate of robust expected power utility with parameter λ(0,1)λ\in(0,1). Using duality methods the problem is reformulated as an infinite time horizon, risk-sensitive control problem. Our results characterize the optimal growth rate, an optimal long-term trading s…

2012-03-06abs ↗pdf ↗

Paper tackles offline CMDP problems with near-optimal algorithm and sample complexity bound.

problem Offline CMDP problems with only offline data available.
method DPDL algorithm using single-policy concentrability coefficient CC^* and deviation control mechanism.
result DPDL algorithm matches sample complexity lower bound with ildeO((1γ)1) ilde{\mathcal{O}}((1-γ)^{-1}) factor.

Paper introduces a new project control method using Monte Carlo and statistical learning.

problem Project control under uncertainty.
method Integrates Earned Value Methodology with Monte Carlo simulation and statistical learning.
result Estimates probabilities of project success and duration.

A simplified model for fixed income portfolio optimisation.

problem Modeling interest rates and credit risk in fixed income portfolios.
method Proposes a two-factor model for the time evolution of the efficient frontier.
result The efficient frontier is mainly controlled by linear constraints, with standard deviation less important.

MADE improves exploration in RL by maximizing deviation from explored regions.

problem Efficient exploration in high-dimensional RL tasks with sparse rewards.
method Proposes a new exploration approach via maximizing the deviation of the occupancy of the next policy from explored regions, adding it as an adaptive regularizer to the RL objective.
result Significantly improves sample efficiency in navigation and locomotion tasks.