Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

82165247329 · Jun 202019922001200920172026
48 results for Hamilton-Jacobi-Bellman system

Optimizes control of infectious disease spread using stochastic methods.

problem Optimizing control of highly infectious diseases like COVID-19.
method Reformulated Hamilton-Jacobi-Bellman equation as stochastic minimum principle, leading to forward-backward stochastic differential equations.
result Numerous numerical solutions presented under various scenarios.

We solve a continuous-time game-theoretic problem for Kihlstrom-Mirman preferences.

problem Dynamic inconsistency in preferences due to multiattribute utility theory.
method Formalized an equilibrium control theory for continuous-time Markov processes.
result Equilibrium strategy and value function as solution to extended HJB system.

The paper tackles data-driven optimal control of unknown nonlinear systems using RKHS.

problem Unknown nonlinear dynamics and stage cost functions.
method Embed state densities into RKHS, learn Markov operators, solve Hamilton-Jacobi-Bellman recursions.
result Solves a wide range of nonlinear control problems, including depth regulation.

A new option pricing model handles non-constant risk aversion and transaction costs.

problem Deriving a pricing model for options with varying risk aversion.
method Developed a transformation method to solve the penalized nonlinear PDE and used finite difference discretization.
result Derived bounds on option prices and proposed a numerical scheme.

The paper analyzes convergence of neural SDEs as sample size increases.

problem Understanding the limiting behavior of neural SDEs as sample size grows.
method Analyzes Hamilton-Jacobi-Bellman equation and uses stochastic maximum principle.
result Convergence of minima and optimal parameters of neural SDEs as sample size increases.

Study optimal futures trading strategies for assets with multiscale central tendency price model.

problem Optimal dynamic trading of futures with multiscale central tendency price model.
method Derive no-arbitrage futures prices, solve HJB equations for optimal strategies.
result Optimal trading strategies depend on asset parameters and futures risk premia.

New method learns policies from offline data using operator models.

problem Limited understanding of approximation errors in offline reinforcement learning.
method Linking reinforcement learning to Hamilton-Jacobi-Bellman equation, proposing operator-theoretic algorithm.
result Global convergence of the value function and finite-sample guarantees derived.

The paper develops RL methods for optimal switching between multiple states.

problem Optimal switching between multiple states in continuous time.
method Entropy-regularized exploration, HJB equations, policy improvement, value function convergence.
result The RL algorithm converges to optimal policies as temperature parameter vanishes.

Study solves optimal portfolio selection using HJB equation.

problem Optimal portfolio selection problem.
method Maximal monotone operator method, Banach fixed-point theorem, Fourier transform, monotone operators technique.
result Existence and uniqueness of solution to HJB equation.

We study an optimal investment/consumption problem in a model capturing market and credit risk dependencies. Stochastic factors drive both the default intensity and the volatility of the stocks in the portfolio. We use the martingale approach and analyze the recursive system of nonlinear Hamilton-Jacobi-Bellman equatio…

2018-06-19abs ↗pdf ↗

Deep neural nets approximate high-dimensional HJB equations efficiently.

problem Approximating solutions to high-dimensional HJB equations.
method Deep neural networks for approximating solutions.
result Deep neural networks can approximate solutions without the curse of dimensionality.

We study the mean field games equations, consisting of the coupled Kolmogorov-Fokker-Planck and Hamilton-Jacobi-Bellman equations. The equations are complemented by initial and terminal conditions. It is shown that with some specific choice of data, this problem can be reduced to solving a quadratically nonlinear syste…

2019-11-21abs ↗pdf ↗

Market makers optimize trading with a new implicit scheme for complex inequalities.

problem Optimizing trading in a limit order book with stochastic and impulse control.
method Implicit numerical scheme coupled with policy iteration algorithm.
result Convergence to the unique viscosity solution of the HJBQVI.

Deep-MacroFin uses neural networks to solve complex economic models efficiently.

problem Solving high-dimensional partial differential equations in continuous time economics.
method Leverages deep learning, specifically Multi-Layer Perceptrons and Kolmogorov-Arnold Networks, optimized with HJB equations.
result Offers a more efficient solution (5imes imes less memory, 40imes imes fewer FLOPs) for 50D economic models.

Deep learning for HJB PDEs using synthetic data and residual minimization.

problem Solving Hamilton-Jacobi-Bellman PDEs for optimal control problems.
method Gradient-augmented synthetic dataset for supervised learning, residual minimization.
result Improves accuracy and efficiency of deep learning for HJB PDEs.

We solve continuous-time reinforcement learning using distributional Hamilton-Jacobi-Bellman equations.

problem Predicting the distribution of returns in continuous-time, stochastic environments.
method We derive a distributional Hamilton-Jacobi-Bellman equation for Itô diffusions and Feller-Dynkin processes, and propose an algorithm based on a JKO scheme.
result We propose an online control algorithm that can be used to approximately solve the distributional HJB equation.

Develops a reinforcement learning algorithm for learning deterministic equilibrium policies in time-inconsistent control problems.

problem Learning equilibrium policies in time-inconsistent control problems.
method Continuous-time model-free reinforcement learning algorithm using deterministic policy gradient approach.
result Learned equilibrium policies in general time-inconsistent control problems.

This paper optimizes reinsurance contracts with belief differences between insurer and reinsurer.

problem Dynamic reinsurance design with heterogeneous beliefs under mean-variance framework.
method Modeling surplus process, applying partitioned domain optimization, solving HJB system.
result Optimal reinsurance contracts with belief heterogeneity are more complex than standard contracts.

A neural network approach solves optimal decumulation problems for pension plans.

problem Optimal asset allocation and withdrawal strategies for DC pension holders.
method Data-driven neural network optimization with customized activation functions.
result The neural network approach learns near-optimal solutions comparable to HJB PDE methods.

Investor optimizes portfolio under dynamic risk preferences.

problem Optimizing investment under uncertain future risk attitudes.
method Developed a general equilibrium framework and solved for subgame-perfect equilibrium policies.
result Equilibrium policies include a novel hedging component to counteract anticipated risk aversion changes.

Investigates time-inconsistent portfolio selection under MMV preferences.

problem Time-inconsistent optimal strategies for MMV preferences.
method Nash equilibrium controls for MMV and MV preferences, solving FBSDE and HJB equations.
result MMV optimal strategies lead to higher investment amounts than MV strategies, narrowing over time.

Optimal contracts are found for agents with quadratic effort costs.

problem Finding optimal contracts in principal-agent problems with quadratic effort costs.
method Modeling the problem using Hamilton-Jacobi-Bellman (HJB) equations and proving the existence of classical solutions.
result Existence of optimal contracts for agents with quadratic effort costs is proven.

A model optimizes carbon emission reduction and allowance purchasing for companies.

problem Optimizing carbon emissions and allowance purchasing for companies.
method Established an optimal control model involving two stochastic processes with two control variables, converted into an HJB equation, proved existence and uniqueness of solution.
result Proved the existence and uniqueness of the solution to the HJB equation.

The Noether theorem is extended to stochastic control problems using contact symmetries.

problem Stochastic optimal control problems.
method Exploiting jet bundles and contact geometry, the authors prove the existence of conserved quantities.
result Optimal control problems admit infinitely many conserved quantities in the form of local martingales.

New method uses neural networks to solve complex PDEs from optimal control theory.

problem Solving high-dimensional Hamilton-Jacobi-Bellman PDEs.
method Iterative diffusion optimization techniques, focusing on path measures and divergences.
result Favourable properties of log-variance divergence for Monte Carlo estimators.

In this paper, we consider equilibrium strategies under Volterra processes and time-inconsistent preferences embracing mean-variance portfolio selection (MVP). Using a functional Itô calculus approach, we overcome the non-Markovian and non-semimartingale difficulty in Volterra processes. The equilibrium strategy is the…

2019-07-26abs ↗pdf ↗

The paper solves a complex financial optimization problem using a novel mathematical technique.

problem Optimizing portfolio selection in financial markets.
method Maximal monotone operator method and Riccati transformation.
result Existence and uniqueness of a solution to the transformed parabolic equation in a Sobolev space.

We study the pricing and the hedging of claim ψ which depends on the default times of two firms A and B. In fact, we assume that, in the market, we can not buy or sell any defaultable bond of the firm B but we can only trade defaultable bond of the firm A. Our aim is then to find the best price and hedging of ψ using o…

2012-09-26abs ↗pdf ↗

Paper solves a complex stopping problem using regularization and HJB equations.

problem Time-inconsistent mean-variance optimal stopping problem
method Vanishing regularization method to derive HJB equations and prove existence of solutions
result Formally recovers variational inequalities for original problem