Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

2525037551,006 · Jun 202019922001200920172026
48 results for control problems

A new method solves complex control problems with random coefficients.

problem Solving LQ McKean-Vlasov control problems with random coefficients.
method Decomposes the problem into two decoupled stochastic optimal control problems.
result The sum of optimal controls of auxiliary problems equals the original problem's optimal control.

Motion planning and control are key problems in a collection of robotic applications including the design of autonomous agile vehicles and of minimalist manipulators. These problems can be accurately formalized within the language of affine connections and of geometric control theory. In this paper we overview recent r…

2002-09-17abs ↗pdf ↗

New control theory for self-path-dependent problems solves unique constraints.

problem Optimal control with self-path-dependent constraints in stochastic systems.
method Introduces new HJB equations for variational inequalities with historical maximum controls.
result Value functions are viscosity solutions to HJB equations under Lipschitz conditions.

We geometrically describe optimal control problems in terms of Morse families in the Hamiltonian framework. These geometric structures allow us to recover the classical first order necessary conditions for optimality and the starting point to run an integrability algorithm. Moreover the integrability algorithm is adapt…

2012-11-19abs ↗pdf ↗

We provide bounds on control learning error in stochastic systems.

problem Learning optimal controls in stochastic environments with uncontrolled parts.
method Dynamic programming and mean-field interpretation of neural networks.
result Non-asymptotic bounds on generalization error for stable overparametrised settings.

We explore a new method for discrete-time control problems using randomization and entropy.

problem Discrete-time linear-exponential quadratic Gaussian (LEQG) control problem.
method Introduce exploration through randomization and apply duality between free energy and relative entropy.
result Reduced LEQG problem to equivalent risk-neutral LQG control problem with entropy regularization.

New framework for policy gradient methods in continuous time reinforcement learning.

problem Addressing policy gradient methods for continuous time reinforcement learning.
method Control randomisation technique to derive policy gradient representation for various Markovian control problems.
result Demonstrated application to optimal switching problems in the energy sector.

Solves inventory control with unknown demand trend using singular control.

problem Optimally managing inventory with an unknown demand trend.
method Formulates as a stochastic control problem under partial observation, solves equivalent separated problem using transition between formulations, and applies viscosity theory.
result Constructs an optimal control rule and shows bounded Lipschitz continuity of free boundaries.

Abstract: Surveying connections between ML and Control Theory.

problem Addressing the intersection of Machine Learning and Control Theory.
method Develops connections through reinforcement learning, supervised learning, deep learning, and stochastic gradient descent.
result Machine Learning and Control Theory are interconnected, with ML solving large control problems and Control Theory providing tools for ML.

Optimal controls for conformal Laplacian obstacle problems on spheres and manifolds.

problem Optimal control of conformal metrics with constant scalar curvature.
method Analysis of optimal control problem on Riemannian manifolds with positive Yamabe invariant.
result Existence of smooth optimal controls inducing metrics with constant scalar curvature.

New method for handling multi-dimensional singular controls with jump costs in mean-field problems.

problem Handling jump costs in multi-dimensional singular controls.
method Introducing two-layer parametrisations to interpolate jumps on both distributional and pathwise levels.
result Derivation of a DPP and characterisation of the value function as a minimal super-solution to a quasi-variational inequality.

Langevin algorithms enhance training of deep neural networks for stochastic control problems.

problem Training acceleration for deep neural networks in stochastic control problems.
method Application of Langevin algorithms to minimize the loss of deep neural networks in stochastic control problems.
result Langevin algorithms improve training on various stochastic control problems.

The Noether theorem is extended to stochastic control problems using contact symmetries.

problem Stochastic optimal control problems.
method Exploiting jet bundles and contact geometry, the authors prove the existence of conserved quantities.
result Optimal control problems admit infinitely many conserved quantities in the form of local martingales.

The paper defines and solves time-inconsistent stopping control problems in multi-dimensional diffusion models.

problem Time-inconsistent problems in control and stopping strategies.
method Formal definition of weak equilibria, extended HJB system, and verification methodology.
result Explicit equilibrium solutions and existence of non-constant equilibria.

Method solves learning problem with hierarchical control objectives.

problem Learning high-dimensional nonlinear functions with model validation accuracy.
method Successive approximation method in functional spaces for hierarchical optimal control.
result Nested algorithm for solving optimal control problem.

Study of multidimensional control problems with reflection controls.

problem Solving control problems with reflection controls in multidimensional settings.
method Gradient descent algorithm for polytope approximations, data-driven domain estimator, episodic learning algorithm.
result Data-driven solutions for unknown diffusion dynamics with sublinear regret.

Investigates optimal strategies for behavioral control problems with finite variation controls.

problem Behavioral singular stochastic control problems with finite variation controls.
method Abstract framework, applied to storage management and portfolio investment problems, using CPT preferences and Skorokhod representation theorem.
result Existence of optimal strategies for various goal functionals, including CPT preferences.

Many real world stochastic control problems suffer from the "curse of dimensionality". To overcome this difficulty, we develop a deep learning approach that directly solves high-dimensional stochastic control problems based on Monte-Carlo sampling. We approximate the time-dependent controls as feedforward neural networ…

2016-11-02abs ↗pdf ↗

Study optimizes trading in multiple assets with cross-effects.

problem Optimizing trade execution in multiple assets with cross-impact effects.
method Formulated as a stochastic control problem, extended to progressively measurable controls, solved using linear-quadratic control theory.
result Cross-hedging effects can be optimal, e.g., trading in an asset without an initial position.

Study uses DRL with Lagrangian relaxation to solve temporal control tasks with STL constraints.

problem Optimal control problems with temporal logic constraints.
method Extended CMDP formulation, Lagrangian relaxation, two-phase constrained DRL algorithm.
result Demonstrated learning performance of the proposed algorithm through simulations.

We consider the exploration-exploitation tradeoff in linear quadratic (LQ) control problems, where the state dynamics is linear and the cost function is quadratic in states and controls. We analyze the regret of Thompson sampling (TS) (a.k.a. posterior-sampling for reinforcement learning) in the frequentist setting, i.…

2017-03-27abs ↗pdf ↗

Study uses viscosity solutions to solve control problems involving measure-valued martingales.

problem Stochastic control problems with measure-valued martingale state processes.
method Viscosity solution approach exploiting structural properties of MVM processes.
result Value function is the unique viscosity solution to the HJB equation.

In this paper we propose a new methodology for solving an uncertain stochastic Markovian control problem in discrete time. We call the proposed methodology the adaptive robust control. We demonstrate that the uncertain control problem under consideration can be solved in terms of associated adaptive robust Bellman equa…

2017-06-07abs ↗pdf ↗

Nonparametric adaptive robust control tackles model uncertainty in stochastic processes.

problem Model uncertainty in stochastic processes.
method Adaptive robust control methodology using online learning and uncertainty reduction, empirical distribution, and Lagrangian duality.
result Nonparametric adaptive robust control approach is preferable to traditional robust frameworks.

Efficient deep policy gradient method for continuous-time control problems.

problem Optimal control in continuous time with fine time discretization.
method Multi-scale deep policy gradient method with varying time discretization.
result Targeted efficiency in computational resources achieved through multi-scale approach.

In this paper, we adapt stochastic Perron's method to analyze a stochastic target problem with unbounded controls in a jump diffusion set-up. With this method, we construct a viscosity sub-solution and super-solution to the associated Hamiltonian-Jacobi-Bellman (HJB) equations. Under comparison principles, uniqueness o…

2016-04-13abs ↗pdf ↗

The paper tackles robust control with uncertain dependence using data-driven methods.

problem Nonparametric robust control under dependence uncertainty in multi-period stochastic systems.
method Nonparametric adaptive robust control framework using stochastic gradient descent ascent algorithm.
result The controller benefits from knowing more about the uncertain model.

The paper tackles optimal stopping problems using reinforcement learning and singular control.

problem Continuous-time and state-space optimal stopping problems.
method Formulated as a singular control problem with randomized stopping times and penalized cumulative residual entropy.
result Identified unique optimal exploratory strategy through dynamic programming.