Study Nash equilibrium in non-zero-sum game with Bermudan strategies.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We consider two-player non-zero-sum stopping games in discrete time. Unlike Dynkin games, in our games the payoff of each player is revealed after both players stop. Moreover, each player can adjust her own stopping strategy according to the other player's action. In the first part of the paper, we consider the game wh…
Paper analyzes robust strategies in a pension plan game with ambiguous financial markets.
Paper studies competitive networks where teams aim to minimize their own objectives, adapting to each other's strategies.
The paper analyzes strategic interactions in a multi-agent reinsurance chain using game theory.
We first study an optimal stopping problem in which a player (an agent) uses a discrete stopping time in order to stop optimally a payoff process whose risk is evaluated by a (non-linear) -expectation. We then consider a non-zero-sum game on discrete stopping times with two agents who aim at minimizing their respect…
New approach improves robustness of deep neural networks without overfitting.
We study optimal behavior of energy producers under a CO_2 emission abatement program. We focus on a two-player discrete-time model where each producer is sequentially optimizing her emission and production schedules. The game-theoretic aspect is captured through a reduced-form price-impact model for the CO_2 allowance…
This paper investigates a hybrid stochastic differential reinsurance and investment game between one reinsurer and two insurers, including a stochastic Stackelberg differential subgame and a non-zero-sum stochastic differential subgame. The reinsurer, as the leader of the Stackelberg game, can price reinsurance premium…
New method refines model predictions as design evolves.
Game contingent claims (GCCs) generalize American contingent claims by allowing the writer to recall the option as long as it is not exercised, at the price of paying some penalty. In incomplete markets, an appealing approach is to analyze GCCs like their European and American counterparts by solving option holder's an…
The paper analyzes reinsurance strategies in a competitive multi-agent system.
Modeling dynamic groundwater markets with price formation and trading strategies.
Study strategic competition in commodity markets using impulse-switching controls.
In this paper, we apply the idea of fictitious play to design deep neural networks (DNNs), and develop deep learning theory and algorithms for computing the Nash equilibrium of asymmetric -player non-zero-sum stochastic differential games, for which we refer as \emph{deep fictitious play}, a multi-stage learning pro…
Investors' strategic trading affects asset prices, modeled as a game.
New research shows no-regret learning is impossible in Markov games under certain assumptions.
Image recognition systems have demonstrated tremendous progress over the past few decades thanks, in part, to our ability of learning compact and robust representations of images. As we witness the wide spread adoption of these systems, it is imperative to consider the problem of unintended leakage of information from …
Improved GAN training stability through tunable classification losses.
In recent years, constrained optimization has become increasingly relevant to the machine learning community, with applications including Neyman-Pearson classification, robust optimization, and fair machine learning. A natural approach to constrained optimization is to optimize the Lagrangian, but this is not guarantee…
The usage of machine learning models has grown substantially and is spreading into several application domains. A common need in using machine learning models is collecting the data required to train these models. In some cases, labeling a massive dataset can be a crippling bottleneck, so there is need to develop model…
Dual-objective GANs reduce training instabilities with tunable α-loss parameters.
We show that many machine learning goals, such as improved fairness metrics, can be expressed as constraints on the model's predictions, which we call rate constraints. We study the problem of training non-convex models subject to these rate constraints (or any non-convex and non-differentiable constraints). In the non…
Paper presents content-based models for game recommendation in cold start scenarios.
In this work, we ask the following question: Can visual analogies, learned in an unsupervised way, be used in order to transfer knowledge between pairs of games and even play one game using an agent trained for another game? We attempt to answer this research question by creating visual analogies between a pair of game…
Potential games, originally introduced in the early 1990's by Lloyd Shapley, the 2012 Nobel Laureate in Economics, and his colleague Dov Monderer, are a very important class of models in game theory. They have special properties such as the existence of Nash equilibria in pure strategies. This note introduces graphical…
IGGP learns game rules from varying quality game play, finding no overall trend.
We introduce a topological combinatorial game called the Region Smoothing Swap Game. The game is played on a game board derived from the connected shadow of a link diagram on a (possibly non-orientable) surface by smoothing at crossings. Moves in the game are performed on regions of the diagram and can switch the direc…
We present a new general board game (GBG) playing and learning framework. GBG defines the common interfaces for board games, game states and their AI agents. It allows one to run competitions of different agents on different games. It standardizes those parts of board game playing and learning that otherwise would be t…
Just as war is sometimes fallaciously represented as a zero sum game -- when in fact war is a negative sum game - stock market trading, a positive sum game over time, is often erroneously represented as a zero sum game. This is called the "zero sum fallacy" -- the erroneous belief that one trader in a stock market exch…
The existence of stationary Markov perfect equilibria in stochastic games is shown under a general condition called "(decomposable) coarser transition kernels". This result covers various earlier existence results on correlated equilibria, noisy stochastic games, stochastic games with finite actions and state-independe…
Combinatorial two-player games have recently been applied to knot theory. Examples of this include the Knotting-Unknotting Game and the Region Unknotting Game, both of which are played on knot shadows. These are turn-based games played by two players, where each player has a separate goal to achieve in order to win the…
We start briefly surveying research on optimal stopping games since their introduction by E.B.Dynkin more than 40 years ago. Recent renewed interest to dynkin's games is due, in particular, to the study of Israeli (game) options introduced in 2000. We discuss the work on these options and related derivative securities …
Educational game on crypto investment helps students grasp macroeconomics.
Simplified NFT games discussed with methods for extracting value.
We introduce TextWorld, a sandbox learning environment for the training and evaluation of RL agents on text-based games. TextWorld is a Python library that handles interactive play-through of text games, as well as backend functions like state tracking and reward assignment. It comes with a curated list of games whose …
Paper tackles hidden game problem in AI alignment and language games.
AEC Games model represents software MARL environments better than POSGs.
Gradient Descent Ascent converges to von-Neumann solution in hidden zero-sum games.
Generalizes region select game to -colored knot diagrams.
We introduce CSE for MLSF games and devise online learning algorithms for achieving no-external Stackelberg-regret.
Deep Reinforcement Learning automates match-3 game testing.
Unified framework for Bayesian and Frequentist statistics.
EBIs/ESOs substantially change the traditional production/service function because ESOs/EBIs can have different psychological effects(motivation or de-motivation), and can create intangible capital and different economic payoffs. Although Game Theory is flawed, it can be helpful in describing interactions in ESO/EBIs t…
New game approximates mean curvature flow evolution.
In this paper we investigate the Follow the Regularized Leader dynamics in sequential imperfect information games (IIG). We generalize existing results of Poincaré recurrence from normal-form games to zero-sum two-player imperfect information games and other sequential game settings. We then investigate how adapting th…
Study on mean field games with singular controls and their applications.
Gradient methods converge exponentially in concave network games.