Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

3867721,1581,544 · Jun 202019922001200920172026
48 results for Nash Learning

Machine learning detects NASH patients from medical claims data.

problem Detecting undiagnosed NASH patients for screening and management.
method Gradient-boosted decision trees trained on administrative medical claims data.
result Model precision for NASH detection is significantly higher than NASH incidence.

Algorithm learns Nash equilibria in stochastic games using entropy-regularized policies.

problem Learning Nash equilibria in zero-sum stochastic games is computationally expensive.
method Entropy-regularized soft policies for Q-function updates.
result Algorithm converges to Nash equilibrium under certain conditions.

The paper examines Nash equilibrium in GANs for stationary Gaussian processes.

problem Existence and uniqueness of Nash equilibrium in GANs for stationary Gaussian processes.
method Analyzes the existence of Nash equilibrium in GANs for stationary Gaussian processes, considering different discriminator families.
result The existence of Nash equilibrium depends on the discriminator family and symmetry properties of the generator family.

This research uses reinforcement learning to find optimal emission offsets in greenhouse gas markets.

problem Finding optimal emission offsets in greenhouse gas markets to control excess emissions.
method Utilized reinforcement learning, specifically Nash-DQN, to estimate market Nash equilibria.
result Emitting firms can achieve significant financial savings by abiding by the Nash equilibria found in the market.

A RL approach finds Nash equilibrium for turn-based zero-sum games.

problem Finding Nash equilibrium in two-player turn-based zero-sum games.
method EIS method combining exploration, policy improvement, and supervised learning.
result EIS method finds an ε-approximate value function of Nash equilibrium in O(ε^(-(d+4))) steps.

A new algorithm reduces memory and computational needs for reinforcement learning.

problem Memory and computational inefficiency in model-free reinforcement learning.
method Memory-Efficient Nash Q-Learning (ME-Nash-QL) for two-player zero-sum games.
result Proves ME-Nash-QL reduces space and sample complexity for tabular and long-horizon cases.

Let MM and NN be Nash manifolds, and ff and gg Nash maps from MM to NN. If MM and NN are compact and if ff and gg are analytically R-L equivalent, then they are Nash R-L equivalent. In the local case, CinftyC^infty R-L equivalence of two Nash map germs implies Nash R-L equivalence. This shows a difference of Nash…

2010-04-23abs ↗pdf ↗

Paper optimizes reinforcement learning in self-play games with reduced steps.

problem Optimizing reinforcement learning algorithms for self-play in two-player zero-sum games.
method Proposes optimistic Nash Q-learning and Nash V-learning algorithms with improved sample complexity.
result Achieves sample complexity of O(SAB)O(SAB) for Nash Q-learning and O(S(A+B))O(S(A+B)) for Nash V-learning, closing the gap with lower bounds.

Proves minimax sample complexity for turn-based stochastic games.

problem Proving theoretical guarantees for reinforcement learning in turn-based stochastic games.
method Developing absorbing TBSG and reward perturbation techniques to handle statistical dependence.
result Empirical Nash equilibrium strategy approximates true Nash equilibrium in turn-based stochastic games.

We introduce a new loss function for evaluating forecasts and estimate models using it.

problem Lack of a decision-theoretic foundation for evaluating forecasts using the Nash-Sutcliffe efficiency.
method We introduce and analyze the Nash-Sutcliffe loss function and its application in estimating models.
result Nash-Sutcliffe loss provides a decision-theoretic foundation for evaluating and estimating models.

Nash integrates covariate-specific side info into sparse regression via neural networks.

problem Sparse linear regression struggles with covariates exhibiting structure or coming from heterogeneous sources.
method Neural Adaptive Shrinkage (Nash) framework that integrates side information into sparse regression via neural networks. Uses split variational empirical Bayes algorithm.
result Nash improves accuracy and adaptability over existing methods in real data experiments.

Agents trained with reinforcement learning deviate from Nash equilibrium in optimal execution game.

problem Deviation of reinforcement learning strategies from Nash equilibrium in optimal execution game.
method Two-player optimal execution game with reinforcement learning algorithms (Double Deep Q-Learning).
result Strategies learned by agents deviate significantly from Nash equilibrium, exhibiting supra-competitive solutions.

Deep fictitious play converges to Nash equilibrium in stochastic differential games.

problem Finding Nash equilibrium in large stochastic differential games.
method Decouples the game into sub-optimization problems and solves each player's optimal strategy with deep BSDE method.
result Deep fictitious play converges to the true Nash equilibrium.

Policy mirror ascent achieves Nash equilibrium in mean field games without a population generative model.

problem Achieving Nash equilibrium in mean field games without a population generative model.
method Policy mirror ascent, contractive operator, single-path TD learning.
result Policy mirror ascent converges to Nash equilibrium within O~(ε2)\widetilde{\mathcal{O}}(\varepsilon^{-2}) samples.

Model-free learning for multi-agent stochastic games is an active area of research. Existing reinforcement learning algorithms, however, are often restricted to zero-sum games, and are applicable only in small state-action spaces or other simplified settings. Here, we develop a new data efficient Deep-Q-learning method…

2019-04-23abs ↗pdf ↗

Algorithm converges to Nash equilibria in competitive games.

problem Finding Nash equilibria in decentralized, competitive Markov games.
method Decentralized Optimistic Gradient Descent/Ascent with a critic.
result Converges to the set of Nash equilibria under self-play.

Algorithm finds Nash equilibria in complex games with function approximation.

problem Learning Nash equilibria in two-player zero-sum Markov Games with nonlinear function approximation.
method Online learning algorithm using upper and lower confidence bounds derived from optimism in the face of uncertainty.
result Achieves O(T)O(\sqrt{T}) regret with polynomial complexity, under mild assumptions.

Method identifies mixed Nash equilibria in high dimensions for training mixtures of GANs.

problem Finding Nash equilibria in two-player zero-sum continuous games, especially in high dimensions.
method Parametrizing mixed strategies as mixtures of particles, updating their positions and weights using gradient descent-ascent.
result Global convergence to an approximate equilibrium for the related Langevin gradient-ascent dynamic.

Paper shows equivalence between two alignment methods and introduces a new algorithm.

problem Ensuring human alignment of large language models for useful, safe, and pleasant user experience.
method Introduces IPO-MD algorithm, showing equivalence between IPO and Nash-MD methods.
result Equivalence between IPO and Nash-MD methods proven when considering online version of IPO.

New algorithms converge faster to Nash equilibrium in zero-sum games with bandit feedback.

problem Learning in zero-sum games with bandit feedback without communication.
method Developed two uncoupled algorithms achieving optimal rate of Ω(T1/4)Ω(T^{-1/4}).
result Achieved optimal rate of Ω(T1/4)Ω(T^{-1/4}) for convergence of policy profiles to Nash equilibrium.

A Nash game theory approach allocates capital requirements among financial institutions.

problem Allocating systemic risk measures among financial institutions.
method Proposes a Nash allocation rule inspired by game theory.
result Provides sufficient conditions for the existence and uniqueness of Nash allocation rules.

In this paper we review our earlier work on quantum computing and the Nash Equilibrium, in particular, tracing the history of the discovery of new Nash Equilibria and then reviewing the ways in which quantum computing may be expected to generate new classes of Nash equilibria. We then extend this work through a substan…

2007-07-03abs ↗pdf ↗

New algorithm reduces online learning regret in uninformed Markov games.

problem Achieving no external regret in uninformed Markov games is impossible.
method Empirical Nash-value regret, parameter-free algorithm, adaptive restart.
result Achieves O(min{K+(CK)1/3,LK})O(\min \{\sqrt{K} + (CK)^{1/3},\sqrt{LK}\}) regret bound.

This paper simplifies the Nash Bargaining Solution for use in intellectual property cases.

problem Limited application of Nash Bargaining Solution in assigning intellectual property damages.
method Normalizes the Nash Bargaining Solution and provides a methodology for determining bargaining weight.
result Clarifies the application of Nash Bargaining Solution to specific case facts.

Proposes a new criterion for selecting Nash equilibria considering both utility and inequality.

problem Finding a fair Nash equilibrium in group decision-making.
method Introduces entropy-norm space for geometric selection of strict Nash equilibria.
result The closest entropy-norm pair to the largest entropy-norm pair in rescaled space is the most suitable equilibrium.

Transformers learn to play games in-context, proving Nash equilibrium.

problem Understanding in-context game-playing capabilities of pre-trained transformers.
method Theoretical guarantees and constructional results for transformer architecture in multi-agent games.
result Pre-trained transformers can learn Nash equilibrium in-context for two-player zero-sum games.

Develops a game-theoretic approach to solve SGEP efficiently.

problem Efficiently solving the symmetric generalized eigenvalue problem for large datasets.
method Formulates SGEP as a Nash equilibrium in a game-theoretic context and develops a parallelizable algorithm.
result Achieves O(dk)O(dk) runtime complexity, making it feasible for large-scale problems.

Pessimistic model-based algorithm finds Nash equilibria in zero-sum Markov games from offline data.

problem Learning Nash equilibria in two-player zero-sum Markov games from limited data.
method Pessimistic model-based algorithm with Bernstein-style lower confidence bounds (VI-LCB-Game).
result Proves sample complexity no larger than CclippedS(A+B)(1γ)3ε2\frac{C_{\mathsf{clipped}}^\star S(A+B)}{(1-γ)^3 \varepsilon^2}, achieving minimax optimality.

We formulate a general framework for competitive gradient-based learning that encompasses a wide breadth of multi-agent learning algorithms, and analyze the limiting behavior of competitive gradient-based learning algorithms using dynamical systems theory. For both general-sum and potential games, we characterize a non…

2018-04-16abs ↗pdf ↗

Ancient Ricci flows with bounded Nash entropy have uniform Sobolev inequalities.

problem Bounding Nash entropy in ancient Ricci flows.
method Uniformly bounded Nash entropy implies uniform bounds on the ν-functional, leading to uniform logarithmic and Sobolev inequalities.
result Uniform logarithmic and Sobolev inequalities on ancient Ricci flows with bounded Nash entropy.