Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

0111 · Jun 201019922001200920172026
9 results for QVIs

Safe reinforcement learning tackles safety constraints with linear approximations.

problem Ensuring safety in reinforcement learning without violating constraints.
method Modeling safety as a linear cost function, developing SLUCB-QVI and RSLUCB-QVI algorithms for MDPs with linear function approximation.
result Achieved a nearly optimal regret bound for safe reinforcement learning, matching state-of-the-art unsafe algorithms.

Develops a new method for pricing GMWBs with jumps and stochastic interest rates.

problem Pricing guaranteed minimum withdrawal benefits (GMWBs) with jumps and stochastic interest rates.
method Combines semi-Lagrangian method with Fourier pricing and Green's function.
result Mathematically demonstrates convergence to the viscosity solution of the HJB-QVI.

In this paper, we study a risk process modeled by a Brownian motion with drift (the diffusion approximation model). The insurance entity can purchase reinsurance to lower its risk and receive cash injections at discrete times to avoid ruin. Proportional reinsurance and excess-of-loss reinsurance are considered. The obj…

2011-12-17abs ↗pdf ↗

The present paper is devoted to the study of a bank salvage model with finite time horizon and subjected to stochastic impulse controls. In our model, the bank's default time is a completely inaccessible random quantity generating its own filtration, then reflecting the unpredictability of the event itself. In this fra…

2019-10-07abs ↗pdf ↗

This paper deals with numerical solutions to an impulse control problem arising from optimal portfolio liquidation with bid-ask spread and market price impact penalizing speedy execution trades. The corresponding dynamic programming (DP) equation is a quasi-variational inequality (QVI) with solvency constraint satisfie…

2010-06-04abs ↗pdf ↗

Study risk-sensitive reinforcement learning with entropic risk measures and generative models.

problem Risk-sensitive reinforcement learning in discounted MDPs with recursive entropic risk measures.
method Introduced Model-Based ERM QQ-Value Iteration (MB-RS-QVI) and derived PAC bounds on sample complexity for value and policy learning.
result PAC bounds show exponential dependence on β/(1γ)|β|/(1-γ), with tight bounds in SS and AA.

RAmmStein optimizes liquidity management in AMMs by learning to rebalance efficiently.

problem Optimal control of concentrated liquidity in decentralized exchanges.
method Formulates as an optimal control problem, uses Deep Reinforcement Learning with HJB-QVI.
result Achieves highest net ROI (1.60%) compared to greedy strategies, reduces rebalancing frequency by 85%.