Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

60120179239 · May 202619922001200920172026
48 results for stopping control

The paper tackles optimal stopping problems using reinforcement learning and singular control.

problem Continuous-time and state-space optimal stopping problems.
method Formulated as a singular control problem with randomized stopping times and penalized cumulative residual entropy.
result Identified unique optimal exploratory strategy through dynamic programming.

A framework for robust exploration in reinforcement learning under ambiguity.

problem Optimal stopping under ambiguity in reinforcement learning.
method Continuous-time robust reinforcement learning framework using gg-expectation and backward stochastic differential equations.
result Constructs a robust exploratory stopping time approximating the optimal stopping time under ambiguity.

The paper defines and solves time-inconsistent stopping control problems in multi-dimensional diffusion models.

problem Time-inconsistent problems in control and stopping strategies.
method Formal definition of weak equilibria, extended HJB system, and verification methodology.
result Explicit equilibrium solutions and existence of non-constant equilibria.

Paper optimizes aquaculture feeding and harvesting strategies for profit maximization.

problem Maximizing farm profit through optimal feeding and harvesting decisions under stochastic price dynamics.
method Developed a simplified aquaculture model and two numerical solution approaches: finite difference scheme and PINN-based method combined with DeepOS algorithm.
result PINN-based method achieves comparable accuracy to finite differences but is more scalable.

We use martingale and stochastic analysis techniques to study a continuous-time optimal stopping problem, in which the decision maker uses a dynamic convex risk measure to evaluate future rewards. We also find a saddle point for an equivalent zero-sum game of control and stopping, between an agent (the "stopper") who c…

2009-09-27abs ↗pdf ↗

Study optimal stopping for diffusion processes with unknown primitives, applying RL and martingale methods.

problem Optimal stopping for diffusion processes with unknown model primitives.
method Continuous-time reinforcement learning framework, variational inequality formulation, stochastic optimal control, entropy regularizer, semi-analytical optimal Bernoulli distribution, policy improvement theorem, policy iterations.
result Demonstrated high accuracy in learning value functions and characterizing free boundaries for various optimal stopping problems.

New algorithms control FDX while achieving more power in online multiple testing.

problem Problems with previous online multiple testing methods, including high FDX and low power.
method Developed new dynamic algorithms that adjust testing levels based on accumulated wealth.
result SupLORD algorithm achieves higher power and FDR control in synthetic experiments.

Unified approach to stochastic control, filtering, and stopping using rough paths.

problem Addressing gaps in classical problems of stochastic control, filtering, and stopping.
method Combining rough path theory with controlled rough paths to provide a pathwise deterministic framework.
result Established rigorous connection between candidate solutions and Hamilton-Jacobi-Bellman equation.

This paper extends stock trading results to include stop-loss orders.

problem Generalizing stock trading results with stop-loss orders.
method Geometric Brownian motion model, affine feedback controller, closed-form expression for cumulative distribution function.
result Affine feedback controller with stop-loss order generalizes results without stop-loss orders.

CITE algorithm provides anytime-valid certification of model outputs.

problem Challenges in controlling error levels in LLM self-consistency.
method Certification by Intersection-union Testing with E-processes (CITE) algorithm.
result Provable control of false certification at any prescribed level under arbitrary stopping rules.

Backdoor attacks on DRL-based traffic controllers cause stop-and-go waves or crashes.

problem Vulnerability of DRL-based traffic controllers to machine learning attacks.
method Developed a trigger design methodology based on traffic physics principles.
result Backdoored models can cause stop-and-go traffic waves or AV crashes when triggered.

Study optimal stopping in random exploration, deriving HJB and designing a reinforcement learning algorithm.

problem Optimal stopping problem in continuous time with random exploration.
method Transformed optimal stopping to optimal control problem, derived HJB equation, designed reinforcement learning algorithm.
result Convergence rate of policy iteration and comparison to classical optimal stopping.

The paper analyzes optimal retirement timing considering age-dependent mortality risk.

problem Optimal retirement timing under age-dependent mortality risk.
method Formulated as a stochastic control and optimal stopping problem, transformed into a finite time horizon, three-dimensional degenerate optimal stopping problem.
result Existence of an optimal retirement boundary, characterized as a unique solution to a nonlinear integral equation.

In this paper, we present a family of a control-stopping games which arise naturally in equilibrium-based models of market microstructure, as well as in other models with strategic buyers and sellers. A distinctive feature of this family of games is the fact that the agents do not have any exogenously given fundamental…

2017-08-01abs ↗pdf ↗

In this paper, we investigate dynamic optimization problems featuring both stochastic control and optimal stopping in a finite time horizon. The paper aims to develop new methodologies, which are significantly different from those of mixed dynamic optimal control and stopping problems in the existing literature, to stu…

2014-06-26abs ↗pdf ↗

Study on games with degenerate diffusion matrices, proving value existence and convergence.

problem Zero-sum games between singular controller and stopper with degenerate diffusion.
method Probabilistic approach using parameterized approximations, convergence analysis.
result Existence of value and optimal stopping times for the game with degenerate dynamics.

Develops a numerical algorithm for stochastic impulse control using regression surrogates.

problem Optimal impulse control in stochastic processes.
method Generates statistical surrogates for continuation and intervention functions, recursively trained over simulated state trajectories.
result Demonstrates flexibility and extensibility of the numerical scheme through case studies.

Optimal retirement timing and consumption under shortfall risk management

problem Optimal portfolio, consumption, and endogenous early retirement problem
method Maximizing expected lifetime consumption utility while managing the maximum wealth shortfall relative to a benchmark
result Geometric structure of the stopping set and feedback-form optimal retirement boundary

We solve the problem of optimal stopping of a Brownian motion subject to the constraint that the stopping time's distribution is a given measure consisting of finitely-many atoms. In particular, we show that this problem can be converted to a finite sequence of state-constrained optimal control problems with additional…

2016-04-11abs ↗pdf ↗

Solves inventory control with unknown demand trend using singular control.

problem Optimally managing inventory with an unknown demand trend.
method Formulates as a stochastic control problem under partial observation, solves equivalent separated problem using transition between formulations, and applies viscosity theory.
result Constructs an optimal control rule and shows bounded Lipschitz continuity of free boundaries.

Derives a new formula for optimal stopping problems with exploding derivatives.

problem Optimal stopping problems with complex boundary conditions.
method Develops a change of variable formula for functions with exploding derivatives near a surface.
result Derives a formula similar to Itô's but with less restrictive conditions.

Inspired by recent work of P.-L. Lions on conditional optimal control, we introduce a problem of optimal stopping under bounded rationality: the objective is the expected payoff at the time of stopping, conditioned on another event. For instance, an agent may care only about states where she is still alive at the time …

2019-01-17abs ↗pdf ↗

Using a bondholder who seeks to determine when to sell his bond as our motivating example, we revisit one of Larry Shepp's classical theorems on optimal stopping. We offer a novel proof of Theorem 1 from from \cite{Shepp}. Our approach is that of guessing the optimal control function and proving its optimality with mar…

2016-05-03abs ↗pdf ↗

We study a robust optimal stopping problem with respect to a set $\cP$ of mutually singular probabilities. This can be interpreted as a zero-sum controller-stopper game in which the stopper is trying to maximize its pay-off while an adverse player wants to minimize this payoff by choosing an evaluation criteria from $\…

2013-01-01abs ↗pdf ↗

Study speculative trading using RL with exploratory framework.

problem Sequential optimal stopping problem over entry and exit times with general utility function and price process.
method Formulated as a sequential optimal stopping problem, solved using Cox processes driven by bounded, non-randomized intensity controls. Characterized randomized control via probability measure over jump intensities and regularized objective function by Shannon's entropy. Established error estimates and convergence of RL objective to value function.
result Closed-form solutions for optimal policy and value function are derived.

We show that deliberately introducing a nested simulation stage can lead to significant variance reductions when comparing two stopping times by Monte Carlo. We derive the optimal number of nested simulations and prove that the algorithm is remarkably robust to misspecifications of this number. The method is applied to…

2014-02-02abs ↗pdf ↗

We study the existence of optimal actions in a zero-sum game infτsupPEP[Xτ]\inf_τ\sup_PE^P[X_τ] between a stopper and a controller choosing a probability measure. This includes the optimal stopping problem infτE(Xτ)\inf_τ\mathcal{E}(X_τ) for a class of sublinear expectations E()\mathcal{E}(\cdot) such as the GG-expectation. We show that …

2012-12-10abs ↗pdf ↗

The paper introduces a limit version of multiple stopping options such that the holder selects dynamically a weight function that control the distribution of the payments (benefits) over time. In applications for commodities and energy trading, a control process can represent the quantity that can be purchased by a fix…

2010-12-07abs ↗pdf ↗

Neural networks solve variational inequalities for optimal stopping problems.

problem Solving variational inequalities for optimal stopping problems in finance.
method Proposed neural network approach using loss functions directly incorporating variational inequality on whole domain.
result Existence and convergence of neural networks whose losses converge to zero.

Paper studies early-stopped mirror descent for noisy sparse phase retrieval.

problem Recovering a sparse signal from noisy quadratic measurements.
method Early-stopped mirror descent with hyperbolic entropy mirror map.
result Achieves nearly minimax-optimal rate of convergence for kk-sparse signals.

New findings control FDR for online testing methods under positive dependence.

problem Maintaining FDR control for online testing methods under positive dependence.
method Developed new methods to control FDR for online testing procedures under positive dependence.
result SAFFRON and LORD control FDR under positive dependence, not just conditional superuniformity.

From the Hamilton-Jacobi-Bellman equation for the value function we derive a non-linear partial differential equation for the optimal portfolio strategy (the dynamic control). The equation is general in the sense that it does not depend on the terminal utility and provides additional analytical insight for some optimal…

2013-11-11abs ↗pdf ↗