Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

14274154 · Feb 202019922001200920172026
48 results for Snake game

Framework for multi-agent RL with human feedback in a Snake game.

problem Improving multi-agent reinforcement learning with human feedback.
method Developed a simulated game environment for offline model training and online competitions. Introduced HILL methods and reward manipulation heuristics.
result Agents with HILL methods outperform those without in online competitions.

This paper presents a novel approach to the technical analysis of wireheading in intelligent agents. Inspired by the natural analogues of wireheading and their prevalent manifestations, we propose the modeling of such phenomenon in Reinforcement Learning (RL) agents as psychological disorders. In a preliminary step tow…

2018-11-14abs ↗pdf ↗

Paper studies geometric and combinatorial properties of circular snakes.

problem Exploring geometric and combinatorial properties of circular snakes.
method Definition and investigation of outer Lipschitz geometry, decomposition of Valette link, construction of combinatorial objects, weakly outer Lipschitz classification.
result Existence of canonical decomposition and necessary/sufficient criteria for removing segments or Hölder triangles.

The snake charmer algorithm permits us to deform a piecewise smooth curve starting from the origin in R^d, so that its end follows a given path. When this path is a loop, a holonomy phenomenon occurs. We prove that the holonomy orbits are closed manifolds diffeomorphic to real Stiefel manifolds. A survey of the snake c…

2006-03-27abs ↗pdf ↗

The purpose of this paper is to give a simpler proof to the problem of controllability of a Hilbert snake \cite{PeSa}. Using the action of the Möbius group of the unit sphere on the configuration space, in the context of a separable Hilbert space. We give a generalization of the Theorem of accessibility contained in \c…

2014-12-21abs ↗pdf ↗

A new snake model improves segmentation of SEM images.

problem Efficiently segmenting overlapping electronic structures in SEM images.
method Geodesic tracking on projective line bundle with a geometric criterion for switching between fast spatial snakes and minimizing geodesics.
result Improved robust and automatic segmentation of overlapping electronic structures in SEM images.

We present a new and very concrete connection between cluster algebras and knot theory. This connection is being made via continued fractions and snake graphs. It is known that the class of 2-bridge knots and links is parametrized by continued fractions, and it has recently been shown that one can associate to each con…

2017-10-23abs ↗pdf ↗

Link Floer homology is split into snake complexes and local systems.

problem Classifying link Floer complexes over specific rings.
method Classifying isomorphism and chain homotopy equivalence classes of free chain complexes over a specific ring, then applying these results to link Floer complexes.
result Link Floer complexes split uniquely into snake complexes and local systems.

We show that the Snake on a square SC(S1)SC(S^1) is homotopy equivalent to the space AC(S1)AC(S^1) which was investigated in the previous work by Eda, Karimov and Repov\vs. We also introduce related constructions CSC()CSC(-) and CAC()CAC(-) and investigate homotopical differences between these four constructions. Finally, we explici…

2013-05-27abs ↗pdf ↗

New q-deformed integers help compute Jones polynomials efficiently.

problem Computing Jones polynomials of rational links efficiently.
method Defining q-deformed integers from pairs of coprime integers and using them to compute Jones polynomials.
result Efficient algorithm for computing Jones polynomials of rational links.

Under appropriate assumptions, we generalize the concept of linear almost Poisson struc- tures, almost Lie algebroids, almost differentials in the framework of Banach anchored bundles and the relation between these objects. We then obtain an adapted formalism for mechanical systems which is illustrated by the evolution…

2011-11-25abs ↗pdf ↗

Recently, it has been shown that the Jones polynomial, in [LS19], and the Alexander polynomial, in [NT18], of rational knots can be obtained by specializing FF-polynomials of cluster variables. At the core of both results are continued fractions, which parameterize rational knots and are used to obtain cluster variabl…

2019-10-22abs ↗pdf ↗

We generalise surface cluster algebras to the case of infinite surfaces where the surface contains finitely many accumulation points of boundary marked points. To connect different triangulations of an infinite surface, we consider infinite mutation sequences. We show transitivity of infinite mutation sequences on tria…

2017-04-06abs ↗pdf ↗

This paper was motivated by work of Arnold where he explains how to count "snakes", i.e. Morse functions on the real axis with prescribed behavior at infinity. This leads immediately to a count of excellent Morse functions on the circle, where following Thom's terminology, excellent means that no two critical points li…

2005-12-21abs ↗pdf ↗

Paper presents content-based models for game recommendation in cold start scenarios.

problem Cold start problem in game recommendation where new games and players have no historical data.
method Uses survey data to develop content-based interaction models that generalize to new games, players, and both.
result Content models outperform collaborative filtering in predicting new interactions.

Potential games, originally introduced in the early 1990's by Lloyd Shapley, the 2012 Nobel Laureate in Economics, and his colleague Dov Monderer, are a very important class of models in game theory. They have special properties such as the existence of Nash equilibria in pure strategies. This note introduces graphical…

2015-05-06abs ↗pdf ↗

TAMIS improves MIA on synthetic data, reducing cost and requiring less knowledge.

problem Empirical assessment of privacy in machine learning algorithms.
method Improves MAMA-MIA by recovering graphical model from synthetic data and introducing a more accurate attack score.
result TAMIS achieves better or similar performance to MAMA-MIA on synthetic data challenges.

IGGP learns game rules from varying quality game play, finding no overall trend.

problem Learn game rules from varying quality game play.
method Used Sancho's intelligent game traces and ILP systems (Metagol, Aleph, ILASP) to induce game rules from traces of varying quality and volume.
result No overall trend in accuracy of learned game rules from varying quality and volume of training data.

We introduce a topological combinatorial game called the Region Smoothing Swap Game. The game is played on a game board derived from the connected shadow of a link diagram on a (possibly non-orientable) surface by smoothing at crossings. Moves in the game are performed on regions of the diagram and can switch the direc…

2019-09-26abs ↗pdf ↗

The existence of stationary Markov perfect equilibria in stochastic games is shown under a general condition called "(decomposable) coarser transition kernels". This result covers various earlier existence results on correlated equilibria, noisy stochastic games, stochastic games with finite actions and state-independe…

2013-11-07abs ↗pdf ↗

Game theory helps analyze ESOs/EBIs in production and service sectors.

problem Economic incentives affect traditional production/service functions and create intangible capital.
method Uses game theory to analyze interactions in ESO/EBI transactions.
result No perfect Nash Equilibria for two-stage games involving many participants.

Combinatorial two-player games have recently been applied to knot theory. Examples of this include the Knotting-Unknotting Game and the Region Unknotting Game, both of which are played on knot shadows. These are turn-based games played by two players, where each player has a separate goal to achieve in order to win the…

2018-07-29abs ↗pdf ↗

We start briefly surveying research on optimal stopping games since their introduction by E.B.Dynkin more than 40 years ago. Recent renewed interest to dynkin's games is due, in particular, to the study of Israeli (game) options introduced in 2000. We discuss the work on these options and related derivative securities …

2012-09-09abs ↗pdf ↗

The paper explores how regularization can lead to convergence in imperfect information games.

problem Finding equilibrium in imperfect information games with imperfect information.
method Investigates Follow the Regularized Leader dynamics and how adding a regularization term can lead to strong convergence guarantees.
result The approach leads to algorithms that converge exactly to the Nash equilibrium in imperfect information games.

Educational game on crypto investment helps students grasp macroeconomics.

problem Weak connections between microeconomic decision-making and macroeconomic concepts in classroom games.
method Design and study of an educational game on cryptocurrency investment.
result Engages students in understanding macroeconomics through incentivized individual investment decisions.

We introduce TextWorld, a sandbox learning environment for the training and evaluation of RL agents on text-based games. TextWorld is a Python library that handles interactive play-through of text games, as well as backend functions like state tracking and reward assignment. It comes with a curated list of games whose …

2018-06-29abs ↗pdf ↗

Gradient Descent Ascent converges to von-Neumann solution in hidden zero-sum games.

problem Understanding dynamics of zero-sum games with hidden structure.
method Gradient Descent Ascent applied to hidden zero-sum games with specific convex-concave structure.
result Gradient Descent Ascent converges to von-Neumann solution in strictly convex-concave hidden games.

We introduce CSE for MLSF games and devise online learning algorithms for achieving no-external Stackelberg-regret.

problem Learning equilibrium in leader-follower games with noisy bandit feedback.
method Proposed Correlated Stackelberg Equilibrium (CSE) and online learning algorithms balancing exploration and exploitation.
result Achieves no-external Stackelberg-regret, converging to approximate CSE.

Study on mean field games with singular controls and their applications.

problem Optimal productivity expansion in dynamic oligopolies.
method Existence and uniqueness of mean field equilibria through nonlinear equations, Abelian limit for discounted and ergodic games.
result Valid connection between discounted and ergodic games, approximation of Nash equilibria.

Federated learning linked to mean-field games for large-scale learning.

problem Large-scale distributed and privacy-preserving learning algorithms.
method Established a connection between federated learning and mean-field games, presenting federated learning as a differential game.
result Properties of the equilibrium of the federated learning game were discussed.