Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,878 papers · 148 categories

Trend · papers per month

3.6%7.1%10.7%14.3% · May 201619922001200920172026
48 results for discounted stochastic games

Paper solves discounted stochastic games with near-optimal time and sample complexity.

problem Solving discounted stochastic two-player games with optimal complexity.
method Generalizes Q-learning to two-player strategy computation, overcoming limitations of existing methods.
result Near-optimal εε-strategy computation with polylogarithmic factors in 1γ1 - γ and ε2ε^{-2}.

The existence of stationary Markov perfect equilibria in stochastic games is shown under a general condition called "(decomposable) coarser transition kernels". This result covers various earlier existence results on correlated equilibria, noisy stochastic games, stochastic games with finite actions and state-independe…

2013-11-07abs ↗pdf ↗

Study on mean field games with singular controls and their applications.

problem Optimal productivity expansion in dynamic oligopolies.
method Existence and uniqueness of mean field equilibria through nonlinear equations, Abelian limit for discounted and ergodic games.
result Valid connection between discounted and ergodic games, approximation of Nash equilibria.

Paper analyzes robust strategies in a pension plan game with ambiguous financial markets.

problem Analyzing robust strategies in a defined benefit pension plan game with ambiguous financial markets.
method Formulated and solved two robust non-zero-sum games using stochastic dynamic programming.
result Explicit forms and optimality of the solutions are shown for the firm and union.

We develop an option pricing model based on a tug-of-war game. This two-player zero-sum stochastic differential game is formulated in the context of a multi-dimensional financial market. The issuer and the holder try to manipulate asset price processes in order to minimize and maximize the expected discounted reward. W…

2014-10-07abs ↗pdf ↗

We propose a simple model of the banking system incorporating a game feature where the evolution of monetary reserve is modeled as a system of coupled Feller diffusions. The Markov Nash equilibrium generated through minimizing the linear quadratic cost subject to Cox-Ingersoll-Ross type processes creates liquidity and …

2016-11-21abs ↗pdf ↗

Consider a two-player zero-sum stochastic game where the transition function can be embedded in a given feature space. We propose a two-player Q-learning algorithm for approximating the Nash equilibrium strategy via sampling. The algorithm is shown to find an εε-optimal strategy using sample size linear to the number …

2019-06-02abs ↗pdf ↗

Study on investment strategy for agents with periodic preferences and discounting.

problem Investment decisions by agents with periodic S-shaped preferences and present bias.
method Infinite-horizon, continuous-time portfolio selection problem with quasi-hyperbolic discounting.
result Time-consistent planning strategy can be formulated as an equilibrium to a static mean field game.

The paper analyzes Q-learning in 2-player Markov games and provides gap-dependent logarithmic regret bounds.

problem Analyzing the cumulative regret of Nash Q-learning in 2-player turn-based stochastic Markov games.
method Proposed gap-dependent logarithmic upper bounds for cumulative regret in episodic tabular setting and discounted game setting.
result The proposed bounds match theoretical lower bounds up to a logarithmic term.

The paper analyzes a game where players must balance short-term and long-term interests, leading to cooperative or competitive outcomes.

problem Analyzing time inconsistency in inter-personal decision-making under non-exponential discounting.
method Iterative procedures and Zorn's lemma to find Nash equilibria between players' intra-personal equilibria.
result Inter-personal equilibria exist and depend on the impatience levels of the players.

Study time-inconsistent portfolio optimization for competitive agents with relative performance criteria.

problem Time-inconsistent mean field and n-agent games under relative performance criteria.
method Construct open-loop equilibrium strategies for n-agent games and mean field games.
result Explicit solutions for n-agent games and mean field games, unique in a special class of equilibria.

Investor and firm optimize sustainable investment and emission reduction through a dynamic game.

problem Optimal sustainable investment and emission reduction in a dynamic game setting.
method Formulated as a nonzero-sum dynamic game, solved via variational inequalities and verified in a diffusive setup.
result Nash equilibria show moving boundaries increasing with emission abatement, triggered by both investor and firm actions.

Study optimal portfolio strategies with time-varying discount rates.

problem Optimizing portfolio decisions with a non-constant discount rate.
method Introduced subgame perfect strategies to handle time inconsistency, using fixed point iteration to find the utility-weighted discount rate.
result Subgame perfect strategies are equivalent to optimal strategies under certain utility function assumptions.

Asset prices contain information about the probability distribution of future states and the stochastic discounting of those states as used by investors. To better understand the challenge in distinguishing investors' beliefs from risk-adjusted discounting, we use Perron-Frobenius Theory to isolate a positive martingal…

2014-11-28abs ↗pdf ↗

Investment decisions shift earlier as patience decreases, with implications for pasting conditions.

problem Investment timing under decreasing impatience.
method Game-theoretic framework with continuous-time capacity expansion problem.
result Decreasing impatience leads to earlier investment decisions, but can violate smooth pasting conditions.

Study dynamic asset allocation in incomplete markets using game theory and nonlocal BSDEs.

problem Dynamic mean-variance asset allocation in general incomplete markets with non-exponential discounting.
method Game-theoretic approach, decomposition into myopic and hedging strategies, nonlocal BSDEs, fixed-point theorem.
result Well-posedness of solutions to BSDEs, existence of equilibrium control policy.

The valuation process that economic agents undergo for investments with uncertain payoff typically depends on their statistical views on possible future outcomes, their attitudes toward risk, and, of course, the payoff structure itself. Yields vary across different investment opportunities and their interrelations are …

2010-01-08abs ↗pdf ↗

Paper tackles time inconsistency in portfolio management with stochastic volatility and power utility.

problem Time inconsistency in portfolio management with stochastic volatility and power utility.
method Extended Hamilton Jacobi Bellman (HJB) equation, fixed point iteration, and linear parabolic PDE.
result Subgame perfect strategies are characterized and solved through numerical experiments.

Model shows how discount rates affect intergenerational equity in climate mitigation.

problem Intergenerational equity in climate mitigation decisions.
method Extended DICE model with stochastic discount rates and financing extensions.
result Discount-rate uncertainty amplifies intergenerational inequality in climate mitigation.

Stochastic dividend discount models (Hurley and Johnson, 1994 and 1998, Yao, 1997) present expressions for the expected value of stock prices when future dividends evolve according to some random scheme. In this paper we try to offer a more precise view on this issue proposing a closed-form formula for the variance of …

2013-11-01abs ↗pdf ↗

Novel approach to Nash equilibrium in mean-field stochastic games with operator resolvents.

problem Finding Nash equilibrium in mean-field stochastic games with mean-field interaction.
method Proposed a novel approach to derive Nash equilibrium semi-explicitly using operator resolvents and stochastic Fredholm equations.
result Equilibrium of the NN-player game converges to mean-field equilibrium, and ε\varepsilon-Nash equilibrium derived as a by-product.

Approximates discounted moments for financial products using polynomial expansions.

problem Approximating discounted moments of stochastic processes for financial applications.
method High-order power series expansion of the infinitesimal generator.
result Error decreases to around 10 to 100 times machine precision for higher orders.

We propose an analytically tractable variation of the minority game in which rational agents use probabilistic strategies. In our model, NN agents choose between two alternatives repeatedly, and those who are in the minority get a pay-off 1, others zero. The agents optimize the expectation value of their discounted fu…

2012-12-29abs ↗pdf ↗

We determine the optimal strategy for investing in a Black-Scholes market in order to maximize the probability that wealth at death meets a bequest goal bb, a type of goal-seeking problem, as pioneered by Dubins and Savage (1965, 1976). The individual consumes at a constant rate cc, so the level of wealth required fo…

2015-03-03abs ↗pdf ↗

The paper solves investment problems with uncertain factors using game theory.

problem Optimal forward investment in an incomplete market with model uncertainty.
method Combining stochastic differential games and ergodic BSDE approach.
result Representation of robust forward performance processes in factor form.

We study the problem of super-replication for game options under proportional transaction costs. We consider a multidimensional continuous time model, in which the discounted stock price process satisfies the conditional full support property. We show that the super-replication price is the cheapest cost of a trivial s…

2011-03-06abs ↗pdf ↗

The paper proposes a new SDF scaled by time-varying volatility from S&P 500 options.

problem Estimating the SDF from option prices and predicting the equity premium.
method Utilizes S&P 500 options data to recover a stable, non-monotonic SDF.
result The SDF exhibits a hump on the put side, which transitions into a W-shape with maturity.

The paper reviews historical and modern approaches to asset pricing probability measures.

problem Constructing or selecting probability measures for asset pricing.
method Historical review of various approaches including state price theory, martingale measures, and modern data-driven methods.
result Modern asset pricing involves constructing, transforming, or selecting probability measures to represent market prices.

Deep fictitious play converges to Nash equilibrium in stochastic differential games.

problem Finding Nash equilibrium in large stochastic differential games.
method Decouples the game into sub-optimization problems and solves each player's optimal strategy with deep BSDE method.
result Deep fictitious play converges to the true Nash equilibrium.

In this paper we propose and analyze a class of NN-player stochastic games that include finite fuel stochastic games as a special case. We first derive sufficient conditions for the Nash equilibrium (NE) in the form of a verification theorem. The associated Quasi-Variational-Inequalities include an essential game comp…

2018-09-10abs ↗pdf ↗

A simple strategy optimizes broker-client trading, reducing price discounts for informed traders.

problem Optimizing broker-client trading to balance client flow and informed trader losses.
method Modelled as a stochastic control problem, derived optimal strategy in closed form, introduced algorithm.
result Optimal strategy reduces price discounts for informed traders, balancing client flow and informed trader losses.

New method handles large reward variations in reinforcement learning.

problem Optimal policy not achievable with existing methods for non-deterministic processes.
method Introduces conjugated distributional operator for handling real returns.
result Guaranteed theoretical convergence for a wide class of transformations.

Study time-inconsistent consumption-investment in incomplete markets with general discount functions.

problem Time-inconsistent consumption-investment problems in incomplete markets.
method Coupled forward-backward stochastic differential equation approach.
result Uniqueness of open-loop equilibrium pair proved.

Algorithm converges to Nash equilibria in competitive games.

problem Finding Nash equilibria in decentralized, competitive Markov games.
method Decentralized Optimistic Gradient Descent/Ascent with a critic.
result Converges to the set of Nash equilibria under self-play.

Algorithm learns Nash equilibria in stochastic games using entropy-regularized policies.

problem Learning Nash equilibria in zero-sum stochastic games is computationally expensive.
method Entropy-regularized soft policies for Q-function updates.
result Algorithm converges to Nash equilibrium under certain conditions.

Paper proves existence and uniqueness of solutions to nonlocal systems, generalizing stochastic game theory.

problem Time inconsistency in stochastic differential games.
method Proves existence and uniqueness of solutions to nonlocal fully-nonlinear parabolic systems.
result Generalizes stochastic game theory to include time-inconsistent preferences.

This paper considers the problem of consumption and investment in a financial market within a continuous time stochastic economy. The investor exhibits a change in the discount rate. The investment opportunities are a stock and a riskless account. The market coefficients and discount factor switch according to a finite…

2013-03-06abs ↗pdf ↗

Graphon game model simplifies stochastic interactions among agents.

problem Complex interactions among heterogeneous agents in stochastic games.
method Introduced a discrete-time graphon game formulation with a representative player.
result Existence and uniqueness of graphon equilibrium proven with mild assumptions.