AlphaZero assesses new chess variants for balance and dynamics.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
The AlphaZero algorithm for the learning of strategy games via self-play, which has produced superhuman ability in the games of Go, chess, and shogi, uses a quantitative reward function for game outcomes, requiring the users of the algorithm to explicitly balance different components of the reward against each other, s…
Paper solves learning imperfect-information games with fewer episodes.
LAFF algorithm balances adaptability and non-exploitability in repeated games.
Paper explains adversarial training's robust overfitting through a minimax game perspective.
A new game-theoretic approach balances downside risk with expected reward.
New ranking system balances fairness and user utility.
We introduce CSE for MLSF games and devise online learning algorithms for achieving no-external Stackelberg-regret.
Game-theoretic analysis of mining gaps in blockchain systems.
Study learns optimal strategies in imperfect information games with self-play.
Interpersonal relations are fickle, with close friendships often dissolving into enmity. In this work, we explore linguistic cues that presage such transitions by studying dyadic interactions in an online strategy game where players form alliances and break those alliances through betrayal. We characterize friendships …
In this paper I give a brief introduction to a family of simple but non-trivial models designed to increase our understanding of collective processes in markets, the so-called Minority Games, and their non-equilibrium statistical mathematical analysis. Since the most commonly studied members of this family define disor…
Algorithm learns to play against unknown opponents in sequential games.
RL optimizes resource allocation in MG by balancing experience and exploration.
New model considers wealth and time affecting risk aversion in portfolio selection.
Optimal penalties for RECs balance environmental and revenue impacts.
Discrete Green's functions are the inverses or pseudo-inverses of combinatorial Laplacians. We present compact formulas for discrete Green's functions, in terms of the eigensystems of corresponding Laplacians, for products of regular graphs with or without boundary. Explicit formulas are derived for the cycle, torus, a…
MAXMINLCB optimizes unknown target functions with preference feedback using a Stackelberg game approach.
The paper analyzes a game where players must balance short-term and long-term interests, leading to cooperative or competitive outcomes.
Study uses MFG approach to model equilibrium pricing with market clearing condition.
With a large number of sensors and control units in networked systems, distributed support vector machines (DSVMs) play a fundamental role in scalable and efficient multi-sensor classification and prediction tasks. However, DSVMs are vulnerable to adversaries who can modify and generate data to deceive the system to mi…
Adversarial training methods typically align distributions by solving two-player games. However, in most current formulations, even if the generator aligns perfectly with data, a sub-optimal discriminator can still drive the two apart. Absent additional regularization, the instability can manifest itself as a never-end…
We derive a class of macroscopic differential equations that describe collective adaptation, starting from a discrete-time stochastic microscopic model. The behavior of each agent is a dynamic balance between adaptation that locally achieves the best action and memory loss that leads to randomized behavior. We show tha…
The 1/3 Financial Rule helps prevent household bankruptcy through balanced spending, savings, and debt repayment.
Paper addresses inefficiency in converting EFGs to NFGs for learning.
Distributed Support Vector Machines (DSVM) have been developed to solve large-scale classification problems in networked systems with a large number of sensors and control units. However, the systems become more vulnerable as detection and defense are increasingly difficult and expensive. This work aims to develop secu…
Paper assesses how features influence classification of COVID-19 patients.
Study multi-agent RL in OTC markets, learning from agents' interactions.
Bayesian rating system for large competitions improves prediction and efficiency.
This paper introduces a novel framework for designing fair and sustainable unemployment benefits, grounded in cooperative game theory and real-time fiscal policy. The labor market is modeled as a coalitional game, where a random subset of participants is employed, generating stochastic economic output. To ensure fairne…
Game theory models storage investment to balance market competition and profits.
PropFair algorithm ensures fair performance in federated learning.
Insurance contracts for autonomous AI agents must be actuarially sound and resistant to gaming.
We consider the problem of prediction by a machine learning algorithm, called learner, within an adversarial learning setting. The learner's task is to correctly predict the class of data passed to it as a query. However, along with queries containing clean data, the learner could also receive malicious or adversarial …
One of the key issues for imitation learning lies in making policy learned from limited samples to generalize well in the whole state-action space. This problem is much more severe in high-dimensional state environments, such as game playing with raw pixel inputs. Under this situation, even state-of-the-art adversary-b…
We introduce a novel framework for adversarial training where the target distribution is annealed between the uniform distribution and the data distribution. We posited a conjecture that learning under continuous annealing in the nonparametric regime is stable irrespective of the divergence measures in the objective fu…
New approach to optimal dividend control with mean-variance criterion.
Non-recurring traffic congestion is caused by temporary disruptions, such as accidents, sports games, adverse weather, etc. We use data related to real-time traffic speed, jam factors (a traffic congestion indicator), and events collected over a year from Nashville, TN to train a multi-layered deep neural network. The …
Esports have become major international sports with hundreds of millions of spectators. Esports games generate massive amounts of telemetry data. Using these to predict the outcome of esports matches has received considerable attention, but micro-predictions, which seek to predict events inside a match, is as yet unkno…
A new algorithm balances exploration and exploitation in online decision-making.
NAMEx merges experts using Nash bargaining for improved performance.
Paper presents content-based models for game recommendation in cold start scenarios.
In this work, we ask the following question: Can visual analogies, learned in an unsupervised way, be used in order to transfer knowledge between pairs of games and even play one game using an agent trained for another game? We attempt to answer this research question by creating visual analogies between a pair of game…
Paper examines financial engineering problems and introduces AlphaZero for better replication strategies.
Potential games, originally introduced in the early 1990's by Lloyd Shapley, the 2012 Nobel Laureate in Economics, and his colleague Dov Monderer, are a very important class of models in game theory. They have special properties such as the existence of Nash equilibria in pure strategies. This note introduces graphical…
IGGP learns game rules from varying quality game play, finding no overall trend.
We introduce a topological combinatorial game called the Region Smoothing Swap Game. The game is played on a game board derived from the connected shadow of a link diagram on a (possibly non-orientable) surface by smoothing at crossings. Moves in the game are performed on regions of the diagram and can switch the direc…
We present a new general board game (GBG) playing and learning framework. GBG defines the common interfaces for board games, game states and their AI agents. It allows one to run competitions of different agents on different games. It standardizes those parts of board game playing and learning that otherwise would be t…