Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

22446587 · May 202619922001200920172026
48 results for arcade formalism

Complex environments and tasks pose a difficult problem for holistic end-to-end learning approaches. Decomposition of an environment into interacting controllable and non-controllable objects allows supervised learning for non-controllable objects and universal value function approximator learning for controllable obje…

2019-01-27abs ↗pdf ↗

Efficient exploration in complex environments remains a major challenge for reinforcement learning. We propose bootstrapped DQN, a simple algorithm that explores in a computationally and statistically efficient manner through use of randomized value functions. Unlike dithering strategies such as epsilon-greedy explorat…

2016-02-15abs ↗pdf ↗

The distributional perspective on reinforcement learning (RL) has given rise to a series of successful Q-learning algorithms, resulting in state-of-the-art performance in arcade game environments. However, it has not yet been analyzed how these findings from a discrete setting translate to complex practical application…

2019-10-01abs ↗pdf ↗

In this work we present a technique to use natural language to help reinforcement learning generalize to unseen environments. This technique uses neural machine translation, specifically the use of encoder-decoder networks, to learn associations between natural language behavior descriptions and state-action informatio…

2017-07-26abs ↗pdf ↗

Deep Reinforcement Learning (DRL) has shown impressive performance on domains with visual inputs, in particular various games. However, the agent is usually trained on a fixed environment, e.g. a fixed number of levels. A growing mass of evidence suggests that these trained models fail to generalize to even slight vari…

2020-01-27abs ↗pdf ↗

We introduce a novel Deep Reinforcement Learning (DRL) algorithm called Deep Quality-Value (DQV) Learning. DQV uses temporal-difference learning to train a Value neural network and uses this network for training a second Quality-value network that learns to estimate state-action values. We first test DQV's update rules…

2018-09-30abs ↗pdf ↗

Study finds reinforcement learning performance plateaus due to environmental interference.

problem Catastrophic interference hinders sample efficiency in reinforcement learning.
method Empirical study in ALE, controlled experiments, analysis of prediction errors.
result Interference causes performance plateaus and degrades policies used to reach them.

The General Video Game AI (GVGAI) competition and its associated software framework provides a way of benchmarking AI algorithms on a large number of games written in a domain-specific description language. While the competition has seen plenty of interest, it has so far focused on online planning, providing a forward …

2018-06-06abs ↗pdf ↗

Deep RL policies share adversarial features across different MDPs.

problem Understanding decision boundaries and loss landscapes in neural policies.
method Investigating similarities in high sensitivity directions across MDPs using Arcade Learning Environment.
result High sensitivity directions for neural policies are correlated across MDPs, suggesting shared non-robust features.

New algorithms improve reinforcement learning stability and performance.

problem Stability issues in TD learning algorithms with function approximation and off-policy sampling.
method Developed and adapted emphatic temporal difference (ETD(λλ)) algorithms for deep reinforcement learning.
result Demonstrated improved performance in Atari games and small problems.

In this paper we argue for the fundamental importance of the value distribution: the distribution of the random return received by a reinforcement learning agent. This is in contrast to the common approach to reinforcement learning which models the expectation of this return, or value. Although there is an established …

2017-07-21abs ↗pdf ↗

This paper investigates whether learning contingency-awareness and controllable aspects of an environment can lead to better exploration in reinforcement learning. To investigate this question, we consider an instantiation of this hypothesis evaluated on the Arcade Learning Element (ALE). In this study, we develop an a…

2018-11-05abs ↗pdf ↗

Detects adversarial directions to make reinforcement learning policies more robust.

problem Adversarial attacks exploit non-robust directions in reinforcement learning policies, leading to instability.
method Local quadratic approximation of deep neural policy loss to identify non-robust directions.
result Provides a theoretical basis for distinguishing safe from adversarial observations.

This work improves understanding of reinforcement learning state representations.

problem Lack of precise characterization of how and when state representations generalize.
method Developed a bound on the generalization error based on effective dimension.
result Bound quantifies the tension between generalization and approximation.

Explores local structure of morphisms and formal submanifolds in formal manifolds theory.

problem Understanding the local structure of morphisms and formal submanifolds in formal manifolds.
method Study of formal manifolds, including local structure of constant rank morphisms and formal submanifolds.
result Developed the local structure of constant rank morphisms and formal submanifolds.

Study non-formal pseudo-differential operators over formal ones.

problem Understanding structure of non-formal pseudo-differential operators.
method Diffeological principal bundles, smoothing connections.
result Structure of diffeological bundle of non-formal pseudo-differential operators over formal ones.

The isotropy action on certain symmetric spaces is shown to be equivariantly formal.

problem Understanding the equivariant formality of isotropy actions on symmetric spaces.
method Developed a new approach to prove equivariant formality for (Z2Z2)(\mathbb{Z}_2\oplus \mathbb{Z}_2)-symmetric spaces.
result Symmetric spaces with (Z2Z2)(\mathbb{Z}_2\oplus \mathbb{Z}_2)-symmetry are equivariantly formal and formal in the Sullivan sense.

Strong formal properties for toric and homogeneous Kähler manifolds.

problem Understanding formal properties of Kähler manifolds.
method Analyzing rationally and strongly formal properties of toric and homogeneous Kähler manifolds.
result Toric and homogeneous Kähler manifolds are both rationally and strongly formal.

A metric is formal if all products of harmonic forms are again harmonic. The existence of a formal metric implies Sullivan formality of the manifold, and hence formal metrics can exist only in presence of a very restricted topology. We show that a warped product metric is formal if and only if the warping function is c…

2010-01-13abs ↗pdf ↗

The study shows strong formality in certain complex manifolds.

problem Investigating strong formality in complex manifolds.
method Adapting ss-strong formality from Fernandez and Muñoz to the pluripotential setting.
result Compact Kähler manifolds and generalized complete intersections are strongly formal.

Abstract MDPs enable strategic exploration and fast reward transfer in complex environments.

problem Challenging to learn accurate MDPs for high-dimensional states.
method Learn an abstract MDP over low-dimensional coarse states, using an abstraction function.
result Achieves superhuman performance on Pitfall! and higher reward with fewer samples.

New hierarchies derived from KP hierarchy using non-formal operators and Yang-Mills action.

problem Formal solutions of KP hierarchy and their non-formal counterparts.
method Developed new hierarchies of non-linear equations on non-formal pseudo-differential operators.
result Expressed one hierarchy as Yang-Mills action minimization.

We define the notion of a formal connection for a smooth family of star products with fixed underlying symplectic structure. Such a formal connection allows one to relate star products at different points in the family. This generalizes the formal Hitchin connection introduced by the first author. We establish a necess…

2014-10-07abs ↗pdf ↗

Defines formal exponentials for graded manifolds and linearizes QP-manifolds.

problem Formal exponentials and linearizations of QP-manifolds.
method Definition of formal exponential maps, Grothendieck connections, and connections on tangent bundles.
result Linearizes QP-manifolds at points, giving formal tangent spaces LL_\infty-algebra structures.

Study on geometrically formal metrics on complex manifolds.

problem Existence and properties of geometrically formal metrics on complex manifolds.
method Topological and cohomological obstructions, detailed analysis for specific manifolds, and metric constructions.
result Existence and non-existence conditions for geometrically formal metrics on various complex manifolds.