Machine learning detects NASH patients from medical claims data.
problem Detecting undiagnosed NASH patients for screening and management.
method Gradient-boosted decision trees trained on administrative medical claims data.
result Model precision for NASH detection is significantly higher than NASH incidence.
Algorithm learns Nash equilibria in stochastic games using entropy-regularized policies.
problem Learning Nash equilibria in zero-sum stochastic games is computationally expensive.
method Entropy-regularized soft policies for Q-function updates.
result Algorithm converges to Nash equilibrium under certain conditions.
A new method for RLHF using proximal point Nash learning.
problem Capturing real human preferences in RLHF.
method Proximal point Nash learning, embedding self-play updates into a proximal point framework.
result High-probability last-iterate convergence for the combined method.
The paper examines Nash equilibrium in GANs for stationary Gaussian processes.
problem Existence and uniqueness of Nash equilibrium in GANs for stationary Gaussian processes.
method Analyzes the existence of Nash equilibrium in GANs for stationary Gaussian processes, considering different discriminator families.
result The existence of Nash equilibrium depends on the discriminator family and symmetry properties of the generator family.
This research uses reinforcement learning to find optimal emission offsets in greenhouse gas markets.
problem Finding optimal emission offsets in greenhouse gas markets to control excess emissions.
method Utilized reinforcement learning, specifically Nash-DQN, to estimate market Nash equilibria.
result Emitting firms can achieve significant financial savings by abiding by the Nash equilibria found in the market.
A RL approach finds Nash equilibrium for turn-based zero-sum games.
problem Finding Nash equilibrium in two-player turn-based zero-sum games.
method EIS method combining exploration, policy improvement, and supervised learning.
result EIS method finds an ε-approximate value function of Nash equilibrium in O(ε^(-(d+4))) steps.
A new algorithm reduces memory and computational needs for reinforcement learning.
problem Memory and computational inefficiency in model-free reinforcement learning.
method Memory-Efficient Nash Q-Learning (ME-Nash-QL) for two-player zero-sum games.
result Proves ME-Nash-QL reduces space and sample complexity for tabular and long-horizon cases.
Let M M M and N N N be Nash manifolds, and f f f and g g g Nash maps from M M M to N N N . If M M M and N N N are compact and if f f f and g g g are analytically R-L equivalent, then they are Nash R-L equivalent. In the local case, C i n f t y C^infty C i n f t y R-L equivalence of two Nash map germs implies Nash R-L equivalence. This shows a difference of Nash…
In this paper, we study the problem of learning the set of pure strategy Nash equilibria and the exact structure of a continuous-action graphical game with quadratic payoffs by observing a small set of perturbed equilibria. A continuous-action graphical game can possibly have an uncountable set of Nash euqilibria. We p…
Nash's theorem proved with Günther's trick
problem Proving Nash's smooth embedding theorem
method Using Günther's trick
result Nash's theorem proved
Paper optimizes reinforcement learning in self-play games with reduced steps.
problem Optimizing reinforcement learning algorithms for self-play in two-player zero-sum games.
method Proposes optimistic Nash Q-learning and Nash V-learning algorithms with improved sample complexity.
result Achieves sample complexity of O ( S A B ) O(SAB) O ( S A B ) for Nash Q-learning and O ( S ( A + B ) ) O(S(A+B)) O ( S ( A + B )) for Nash V-learning, closing the gap with lower bounds. Proves minimax sample complexity for turn-based stochastic games.
problem Proving theoretical guarantees for reinforcement learning in turn-based stochastic games.
method Developing absorbing TBSG and reward perturbation techniques to handle statistical dependence.
result Empirical Nash equilibrium strategy approximates true Nash equilibrium in turn-based stochastic games.
GANs may not have Nash equilibria, but proximal training can find solutions.
problem Existence of Nash equilibria in GANs optimization.
method Proximal training approach to find solutions.
result Proximal training finds solutions to GAN problems.
We introduce a new loss function for evaluating forecasts and estimate models using it.
problem Lack of a decision-theoretic foundation for evaluating forecasts using the Nash-Sutcliffe efficiency.
method We introduce and analyze the Nash-Sutcliffe loss function and its application in estimating models.
result Nash-Sutcliffe loss provides a decision-theoretic foundation for evaluating and estimating models.
Nash integrates covariate-specific side info into sparse regression via neural networks.
problem Sparse linear regression struggles with covariates exhibiting structure or coming from heterogeneous sources.
method Neural Adaptive Shrinkage (Nash) framework that integrates side information into sparse regression via neural networks. Uses split variational empirical Bayes algorithm.
result Nash improves accuracy and adaptability over existing methods in real data experiments.
Agents trained with reinforcement learning deviate from Nash equilibrium in optimal execution game.
problem Deviation of reinforcement learning strategies from Nash equilibrium in optimal execution game.
method Two-player optimal execution game with reinforcement learning algorithms (Double Deep Q-Learning).
result Strategies learned by agents deviate significantly from Nash equilibrium, exhibiting supra-competitive solutions.
A new approach to fine-tuning LLMs with human feedback.
problem Inability of current reward models to fully represent human preferences.
method Introducing NLHF, a new pipeline for LLM fine-tuning using pairwise human feedback.
result NLHF produces a sequence of policies converging to the regularized Nash equilibrium.
We propose local symplectic surgery, a two-timescale procedure for finding local Nash equilibria in two-player zero-sum games. We first show that previous gradient-based algorithms cannot guarantee convergence to local Nash equilibria due to the existence of non-Nash stationary points. By taking advantage of the differ…
Distributed strategic learning has been getting attention in recent years. As systems become distributed finding Nash equilibria in a distributed fashion is becoming more important for various applications. In this paper, we develop a distributed strategic learning framework for seeking Nash equilibria under stochastic…
Deep fictitious play converges to Nash equilibrium in stochastic differential games.
problem Finding Nash equilibrium in large stochastic differential games.
method Decouples the game into sub-optimization problems and solves each player's optimal strategy with deep BSDE method.
result Deep fictitious play converges to the true Nash equilibrium.
Policy mirror ascent achieves Nash equilibrium in mean field games without a population generative model.
problem Achieving Nash equilibrium in mean field games without a population generative model.
method Policy mirror ascent, contractive operator, single-path TD learning.
result Policy mirror ascent converges to Nash equilibrium within O ~ ( ε − 2 ) \widetilde{\mathcal{O}}(\varepsilon^{-2}) O ( ε − 2 ) samples. Optimal geometric estimates for Kähler manifolds with bounded Nash entropy
problem Optimal geometric estimates for compact Kähler manifolds
method Proving Sobolev-type inequality and local volume noncollapsing with optimal exponents
result Uniformly bounded q q q -Nash entropy Algorithm learns fair division from noisy feedback in uncertain markets.
problem Learning fair division in uncertain markets with noisy feedback.
method Wrapper algorithms using dual averaging to learn item and agent values from bandit feedback.
result Asymptotically achieves optimal Nash social welfare in linear Fisher markets.
Model-free learning for multi-agent stochastic games is an active area of research. Existing reinforcement learning algorithms, however, are often restricted to zero-sum games, and are applicable only in small state-action spaces or other simplified settings. Here, we develop a new data efficient Deep-Q-learning method…
New method finds all Nash equilibria via vector optimization.
problem Finding all Nash equilibria in games.
method Formulate vector optimization problem to find Pareto optimal solutions.
result Characterize set of all Nash equilibria as Pareto optimal solutions.
The h-cobordism theorem is a noted theorem in differential and PL topology. A generalization of the h-cobordism theorem for possibly non simply connected manifolds is the so called s-cobordism theorem. In this paper, we prove semialgebraic and Nash versions of these theorems. That is, starting with semialgebraic or Nas…
Algorithm converges to Nash equilibria in competitive games.
problem Finding Nash equilibria in decentralized, competitive Markov games.
method Decentralized Optimistic Gradient Descent/Ascent with a critic.
result Converges to the set of Nash equilibria under self-play.
Algorithm finds Nash equilibria in complex games with function approximation.
problem Learning Nash equilibria in two-player zero-sum Markov Games with nonlinear function approximation.
method Online learning algorithm using upper and lower confidence bounds derived from optimism in the face of uncertainty.
result Achieves O ( T ) O(\sqrt{T}) O ( T ) regret with polynomial complexity, under mild assumptions. Method identifies mixed Nash equilibria in high dimensions for training mixtures of GANs.
problem Finding Nash equilibria in two-player zero-sum continuous games, especially in high dimensions.
method Parametrizing mixed strategies as mixtures of particles, updating their positions and weights using gradient descent-ascent.
result Global convergence to an approximate equilibrium for the related Langevin gradient-ascent dynamic.
Paper shows equivalence between two alignment methods and introduces a new algorithm.
problem Ensuring human alignment of large language models for useful, safe, and pleasant user experience.
method Introduces IPO-MD algorithm, showing equivalence between IPO and Nash-MD methods.
result Equivalence between IPO and Nash-MD methods proven when considering online version of IPO.
We show by counterexample that policy-gradient algorithms have no guarantees of even local convergence to Nash equilibria in continuous action and state space multi-agent settings. To do so, we analyze gradient-play in N-player general-sum linear quadratic games, a classic game setting which is recently emerging as a b…
New algorithms converge faster to Nash equilibrium in zero-sum games with bandit feedback.
problem Learning in zero-sum games with bandit feedback without communication.
method Developed two uncoupled algorithms achieving optimal rate of Ω ( T − 1 / 4 ) Ω(T^{-1/4}) Ω ( T − 1/4 ) . result Achieved optimal rate of Ω ( T − 1 / 4 ) Ω(T^{-1/4}) Ω ( T − 1/4 ) for convergence of policy profiles to Nash equilibrium. The study characterizes Nash maps between semialgebraic sets and their properties.
problem Existence of surjective Nash maps between semialgebraic sets.
method Characterization of semialgebraic subsets and their images under Nash maps.
result Characterization of semialgebraic sets that are Nash images of the unit ball.
A method is provided to resolve Lie algebroids with singularities.
problem Resolving Lie algebroids with singular points.
method Nash-type blow-up construction for Lie algebroids.
result A short exact sequence is established linking the blow-up to Lie algebroids.
A Nash game theory approach allocates capital requirements among financial institutions.
problem Allocating systemic risk measures among financial institutions.
method Proposes a Nash allocation rule inspired by game theory.
result Provides sufficient conditions for the existence and uniqueness of Nash allocation rules.
In this paper we review our earlier work on quantum computing and the Nash Equilibrium, in particular, tracing the history of the discovery of new Nash Equilibria and then reviewing the ways in which quantum computing may be expected to generate new classes of Nash equilibria. We then extend this work through a substan…
We study discrete-time mean-field Markov games with infinite numbers of agents where each agent aims to minimize its ergodic cost. We consider the setting where the agents have identical linear state transitions and quadratic cost functions, while the aggregated effect of the agents is captured by the population mean o…
New algorithm reduces online learning regret in uninformed Markov games.
problem Achieving no external regret in uninformed Markov games is impossible.
method Empirical Nash-value regret, parameter-free algorithm, adaptive restart.
result Achieves O ( min { K + ( C K ) 1 / 3 , L K } ) O(\min \{\sqrt{K} + (CK)^{1/3},\sqrt{LK}\}) O ( min { K + ( C K ) 1/3 , L K }) regret bound. This paper simplifies the Nash Bargaining Solution for use in intellectual property cases.
problem Limited application of Nash Bargaining Solution in assigning intellectual property damages.
method Normalizes the Nash Bargaining Solution and provides a methodology for determining bargaining weight.
result Clarifies the application of Nash Bargaining Solution to specific case facts.
Proposes a new criterion for selecting Nash equilibria considering both utility and inequality.
problem Finding a fair Nash equilibrium in group decision-making.
method Introduces entropy-norm space for geometric selection of strict Nash equilibria.
result The closest entropy-norm pair to the largest entropy-norm pair in rescaled space is the most suitable equilibrium.
Transformers learn to play games in-context, proving Nash equilibrium.
problem Understanding in-context game-playing capabilities of pre-trained transformers.
method Theoretical guarantees and constructional results for transformer architecture in multi-agent games.
result Pre-trained transformers can learn Nash equilibrium in-context for two-player zero-sum games.
Paper refines royalty determination using Bayesian methods.
problem Determining a reasonable royalty with risk and uncertainty.
method Bayesian Cost approach to refine Nash Bargaining Solution.
result Nash Bargaining Solution emerges as more reliable.
Develops a game-theoretic approach to solve SGEP efficiently.
problem Efficiently solving the symmetric generalized eigenvalue problem for large datasets.
method Formulates SGEP as a Nash equilibrium in a game-theoretic context and develops a parallelizable algorithm.
result Achieves O ( d k ) O(dk) O ( d k ) runtime complexity, making it feasible for large-scale problems. Pessimistic model-based algorithm finds Nash equilibria in zero-sum Markov games from offline data.
problem Learning Nash equilibria in two-player zero-sum Markov games from limited data.
method Pessimistic model-based algorithm with Bernstein-style lower confidence bounds (VI-LCB-Game).
result Proves sample complexity no larger than C c l i p p e d ⋆ S ( A + B ) ( 1 − γ ) 3 ε 2 \frac{C_{\mathsf{clipped}}^\star S(A+B)}{(1-γ)^3 \varepsilon^2} ( 1 − γ ) 3 ε 2 C clipped ⋆ S ( A + B ) , achieving minimax optimality. We formulate a general framework for competitive gradient-based learning that encompasses a wide breadth of multi-agent learning algorithms, and analyze the limiting behavior of competitive gradient-based learning algorithms using dynamical systems theory. For both general-sum and potential games, we characterize a non…
An complete exposition of Matthias Gunther's elementary proof of Nash's isometric embedding theorem.
Ancient Ricci flows with bounded Nash entropy have uniform Sobolev inequalities.
problem Bounding Nash entropy in ancient Ricci flows.
method Uniformly bounded Nash entropy implies uniform bounds on the ν-functional, leading to uniform logarithmic and Sobolev inequalities.
result Uniform logarithmic and Sobolev inequalities on ancient Ricci flows with bounded Nash entropy.
We prove that the infinite family of homotopy 4-spheres constructed by Daniel Nash are all diffeomorphic to 4-sphere.