Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

14274154 · Feb 202019922001200920182026
48 results for Pong game

Study ping-pong dynamics in hyperbolic-like groups with non-simple points.

problem Investigate the ping-pong dynamics of hyperbolic-like groups.
method Explicitly provide a proper ping-pong partition for any pair of non-cyclic point stabilizers.
result Existence of a proper ping-pong partition for any pair of non-cyclic point stabilizers.

While deep reinforcement learning has successfully solved many challenging control tasks, its real-world applicability has been limited by the inability to ensure the safety of learned policies. We propose an approach to verifiable reinforcement learning by training decision tree policies, which can represent complex p…

2018-05-22abs ↗pdf ↗

New groups discovered with unique properties in a specific space.

problem Finding new discrete subgroups with special properties in a mathematical space.
method Proved by showing groups play ping-pong on cones, related to crooked surfaces.
result Infinite family of discrete subgroups with remarkable properties in Sp4(R){Sp}_4(\mathbb{R}).

Paper introduces a technique to simplify RNN policies for better understanding and analysis.

problem Difficulty in explaining and analyzing RNN policies due to continuous-valued memory vectors and observation features.
method Quantized Bottleneck Insertion technique to learn finite representations of RNN vectors and features.
result Finite representations of RNN policies can be as small as 3 discrete memory states and 10 observations, improving interpretability.

We prove that all atoroidal automorphisms of Out(FN)Out(F_N) act on the space of projectivized geodesic currents with generalized north-south dynamics. As an application, we produce new examples of non virtually cyclic, free and purely atoroidal subgroups of Out(FN)Out(F_N) such that the corresponding free group extension is hyp…

2017-11-21abs ↗pdf ↗

In this paper, we prove a quantitative version of the Tits alternative for negatively pinched manifolds XX. Precisely, we prove that a nonelementary discrete isometry subgroup of Isom(X)\mathrm{Isom}(X) generated by two non-elliptic isometries gg, ff contains a free subgroup of rank 22 generated by isometries fN,hf^N , h

2018-06-19abs ↗pdf ↗

We prove that if φ,ψOut(FN)φ,ψ\in Out(F_N) are hyperbolic iwips (irreducible with irreducible powers) such that <φ,ψ>Out(FN)<φ,ψ>\le Out(F_N) is not virtually cyclic then some high powers of φφ and ψψ generate a free subgroup of rank two, all of whose nontrivial elements are again hyperbolic iwips. Being a hyperbolic iwip element of $…

2009-02-24abs ↗pdf ↗

Local-to-global principle for Morse actions on symmetric spaces.

problem Recognizing Morse actions on symmetric spaces.
method Equivariant Morse quasiisometric embeddings of trees into symmetric spaces.
result Algorithmic recognizability of Morse actions and construction of Morse Schottky subgroups.

We prove that every acylindrically hyperbolic group that has no non-trivial finite normal subgroup satisfies a strong ping pong property, the PnaiveP_{naive} property: for any finite collection of elements h1,,hkh_1, \dots, h_k, there exists another element γ1γ\neq 1 such that for all ii, $\langle h_i, γ\rangle = \langle h_i …

2016-10-13abs ↗pdf ↗

Researchers discover all affinely homogeneous models for surfaces in 4D space.

problem Identifying all affinely homogeneous models for surfaces in 4D space.
method Improved power series method of equivalence, capturing invariants at the origin, creating branches, and infinitesimalizing calculations.
result Find several inequivalent terminal branches yielding each to some nonempty moduli space of homogeneous models.

Defines Wodzicki residue using groupoids and fibered distributions.

problem Defining and understanding the Wodzicki residue in noncommutative geometry.
method Using groupoid language and filtered manifolds, defining the residue and showing its properties.
result The groupoidal residue is a trace on pseudodifferential operators and matches the usual residue in certain cases.

Visual analogies help transfer knowledge between Atari games.

problem Can visual analogies transfer knowledge between Atari games?
method Created visual analogies between pairs of Atari games and used them to train policies for one game using data from another.
result Visual analogies can be used to transfer knowledge between Atari games.

Paper presents content-based models for game recommendation in cold start scenarios.

problem Cold start problem in game recommendation where new games and players have no historical data.
method Uses survey data to develop content-based interaction models that generalize to new games, players, and both.
result Content models outperform collaborative filtering in predicting new interactions.

Potential games, originally introduced in the early 1990's by Lloyd Shapley, the 2012 Nobel Laureate in Economics, and his colleague Dov Monderer, are a very important class of models in game theory. They have special properties such as the existence of Nash equilibria in pure strategies. This note introduces graphical…

2015-05-06abs ↗pdf ↗

IGGP learns game rules from varying quality game play, finding no overall trend.

problem Learn game rules from varying quality game play.
method Used Sancho's intelligent game traces and ILP systems (Metagol, Aleph, ILASP) to induce game rules from traces of varying quality and volume.
result No overall trend in accuracy of learned game rules from varying quality and volume of training data.

This note removes technical assumptions and characterizes relatively dominated representations.

problem Geometrically finiteness and Anosov conditions in higher-rank settings.
method Characterization using eigenvalue gaps and limit maps.
result Relatively dominated representations are characterized using eigenvalue gaps and limit maps.

The paper introduces MDP homomorphic networks for faster reinforcement learning.

problem Current reinforcement learning approaches do not exploit symmetries in the joint state-action space.
method Equivariant neural networks with group-structured symmetries (reflections, rotations).
result MDP homomorphic networks converge faster than unstructured baselines on various tasks.

The existence of stationary Markov perfect equilibria in stochastic games is shown under a general condition called "(decomposable) coarser transition kernels". This result covers various earlier existence results on correlated equilibria, noisy stochastic games, stochastic games with finite actions and state-independe…

2013-11-07abs ↗pdf ↗

TextWorld is a Python library for RL agents in text-based games.

problem Training RL agents on text-based games with varying challenges and sparse rewards.
method Developed a Python library with backend functions for state tracking and reward assignment. Enables users to create new games with precise control over difficulty and scope.
result Demonstrated the effectiveness of TextWorld in training RL agents on a curated list of games and generated sets of games.

Game theory helps analyze ESOs/EBIs in production and service sectors.

problem Economic incentives affect traditional production/service functions and create intangible capital.
method Uses game theory to analyze interactions in ESO/EBI transactions.
result No perfect Nash Equilibria for two-stage games involving many participants.

We start briefly surveying research on optimal stopping games since their introduction by E.B.Dynkin more than 40 years ago. Recent renewed interest to dynkin's games is due, in particular, to the study of Israeli (game) options introduced in 2000. We discuss the work on these options and related derivative securities …

2012-09-09abs ↗pdf ↗

The paper explores how regularization can lead to convergence in imperfect information games.

problem Finding equilibrium in imperfect information games with imperfect information.
method Investigates Follow the Regularized Leader dynamics and how adding a regularization term can lead to strong convergence guarantees.
result The approach leads to algorithms that converge exactly to the Nash equilibrium in imperfect information games.

Educational game on crypto investment helps students grasp macroeconomics.

problem Weak connections between microeconomic decision-making and macroeconomic concepts in classroom games.
method Design and study of an educational game on cryptocurrency investment.
result Engages students in understanding macroeconomics through incentivized individual investment decisions.

The paper proposes a method to learn continuous-action graphical games from perturbed equilibria.

problem Learning the exact structure of continuous-action graphical games from limited data.
method A 12\ell_{12}- block regularized method to recover the graphical game structure.
result The method recovers the exact structure of the graphical game under certain conditions.

Gradient Descent Ascent converges to von-Neumann solution in hidden zero-sum games.

problem Understanding dynamics of zero-sum games with hidden structure.
method Gradient Descent Ascent applied to hidden zero-sum games with specific convex-concave structure.
result Gradient Descent Ascent converges to von-Neumann solution in strictly convex-concave hidden games.

We introduce CSE for MLSF games and devise online learning algorithms for achieving no-external Stackelberg-regret.

problem Learning equilibrium in leader-follower games with noisy bandit feedback.
method Proposed Correlated Stackelberg Equilibrium (CSE) and online learning algorithms balancing exploration and exploitation.
result Achieves no-external Stackelberg-regret, converging to approximate CSE.