Soft Actor-Critic is a state-of-the-art reinforcement learning algorithm for continuous action settings that is not applicable to discrete action settings. Many important settings involve discrete actions, however, and so here we derive an alternative version of the Soft Actor-Critic algorithm that is applicable to dis…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
New reinforcement learning framework for adapting to new actions.
Simplifies large action space bandits by selecting representative actions.
The Markov decision process (MDP) formulation used to model many real-world sequential decision making problems does not efficiently capture the setting where the set of available decisions (actions) at each time step is stochastic. Recently, the stochastic action set Markov decision process (SAS-MDP) formulation has b…
We present some methods to construct smooth circle actions on symplectic manifolds with non-symplectic fixed point sets or non-symplectic cyclic isotropy point sets. All such actions are not compatible with any symplectic form.
Study fixed-point sets of -actions on quaternionic manifolds.
Intelligent agents can learn to represent the action spaces of other agents simply by observing them act. Such representations help agents quickly learn to predict the effects of their own actions on the environment and to plan complex action sequences. In this work, we address the problem of learning an agent's action…
In many real-world sequential decision making problems, the number of available actions (decisions) can vary over time. While problems like catastrophic forgetting, changing transition dynamics, changing rewards functions, etc. have been well-studied in the lifelong learning literature, the setting where the action set…
We present and study a partial-information model of online learning, where a decision maker repeatedly chooses from a finite set of actions, and observes some subset of the associated losses. This naturally models several situations where the losses of different actions are related, and knowing the loss of one action p…
We establish a necessary and sufficient condition for pairs of integers to arise as the weights at the fixed points of an effective circle action on a compact almost complex 4-manifold with a discrete fixed point set. As an application, we provide a necessary and sufficient condition for a pair of integers to arise as …
A new RL approach learns near-equivalent actions for healthcare decisions.
We enhance conformal prediction for risk-averse decisions with action-conditional guarantees.
We extend Forester's rigidity theorem so as to give a complete characterization of rigid group actions on trees (an action is rigid if it is the only reduced action in its deformation space, in particular it is invariant under automorphisms preserving the set of elliptic subgroups).
New findings on group actions and stabilizers of infinite sets.
We study actions of finite groups on moduli spaces of stable holomorphic vector bundles and relate the fixed-point sets of those actions to representation varieties of certain orbifold fundamental groups.
We extend the equivariant holomorphic Morse inequalities of circle actions to cases with torus and non-Abelian group actions on holomorphic vector bundles over Kahler manifolds and show the necessity of the Kahler condition. For torus actions, there is a set of inequalities for each choice of action chambers specifying…
Algorithm optimizes bandit decisions with changing action sets using Gaussian processes.
New method for linear bandits with unknown sparsity, improving sparse regret bounds.
We examine free orientation-reversing group actions on orientable handlebodies, and free actions on nonorientable handlebodies. A classification theorem is obtained, giving the equivalence classes and weak equivalence classes of free actions in terms of algebraic invariants that involve Nielsen equivalence. This is app…
New STDP rule for spiking neurons solves discrete action reinforcement learning tasks.
Study quotients of curve complex actions by mapping class group.
New action poisoning attacks improve LinUCB's performance by changing action signals.
Over 50 years of work on group actions on -manifolds, from the 1960's to the present, from knotted fixed point sets to Seiberg-Witten invariants, is surveyed. Locally linear actions are emphasized, but differentiable and purely topological actions are also discussed. The presentation is organized around some of the …
Galois action on manifold structures of complex varieties is abelian.
Study homeomorphisms on fine curve graph of surfaces, revealing new types of dynamics.
Study circle actions on unitary manifolds with discrete fixed points.
We present a simple approach to questions of topological orbit equivalence for actions of countable groups on topological and smooth manifolds. For example, for any action of a countable group on a topological manifold where the fixed sets for any element are contained in codimension two submanifolds, every orbit e…
This paper investigates the adversarial Bandits with Knapsack (BwK) online learning problem, where a player repeatedly chooses to perform an action, pays the corresponding cost, and receives a reward associated with the action. The player is constrained by the maximum budget that can be spent to perform actions, an…
We give a systematic treatment of the stability theory for action of a real reductive Lie group G on a topological space. More precisely, we introduce an abstract setting for actions of non-compact real reductive Lie groups on topological spaces that admit functions similar to the Kempf-Ness function. The point of this…
Proper actions on bornological spaces are characterized with compatible coarse structures.
The use of Reinforcement Learning in real-world scenarios is strongly limited by issues of scale. Most RL learning algorithms are unable to deal with problems composed of hundreds or sometimes even dozens of possible actions, and therefore cannot be applied to many real-world problems. We consider the RL problem in the…
Extends Cheeger's method to Lie groupoid actions on manifolds.
Topologically and geometrically engaging actions have proved to be useful to obtain rigidity results for semisimple Lie group actions. We show that the action of a simple noncompact Lie group on a compact manifold preserving a unimodular rigid geometric structure of algebraic type (e.g. a connection together with a vol…
We use the equivariant Yang-Mills moduli space to investigate the relation between the singular set, isotropy representations at fixed points, and permutation modules realized by the induced action on homology for smooth group actions on certain 4-manifolds.
Most model-free reinforcement learning methods leverage state representations (embeddings) for generalization, but either ignore structure in the space of actions or assume the structure is provided a priori. We show how a policy can be decomposed into a component that acts in a low-dimensional space of action represen…
This paper examines a proposal for gauging non-linear sigma models with respect to a Lie algebroid action. The general conditions for gauging a non-linear sigma model with a set of involutive vector fields are given. We show that it is always possible to find a set of vector fields which will (locally) admit a Lie alge…
The study examines how perturbations of lattice actions on group boundaries behave.
Given a group action, known by its infinitesimal generators, we exhibit a complete set of syzygies on a generating set of differential invariants. For that we elaborate on the reinterpretation of Cartan's moving frame by Fels and Olver (1999). This provides constructive tools for exploring algebras of differential inva…
We consider an adversarial online learning setting where a decision maker can choose an action in every stage of the game. In addition to observing the reward of the chosen action, the decision maker gets side observations on the reward he would have obtained had he chosen some of the other actions. The observation str…
Let be a symplectic manifold, equipped with a semifree symplectic circle action with a finite, nonempty fixed point set. We show that the circle action must be Hamiltonian, and must have the equivariant cohomology and Chern classes of .
Abstract: Proves generic torus diffeomorphisms act parabolically and non-properly on fine curve graph and have generalized rotation sets.
New RL method handles large state-action spaces with complex models.
Classifies timelike translating solitons in Minkowski space.
Sengupta's lower bound for the Yang-Mills action on smooth connections on a bundle over a Riemann surface generalizes to the space of connections whose action is finite. In this larger space the inequality can always be saturated. The Yang-Mills critical sets correspond to critical sets of the energy action on a space …
A finite nonabelian simple group does not admit a free action on a homology sphere, and the only finite simple group which acts on a homology sphere with at most 0-dimensional fixed point sets ("pseudofree action") is the alternating group A_5 acting on the 2-sphere. Our first main theorem is the finiteness result that…
Study of entropy-regularized LQG MFGs with exploratory actions.
We prove that certain volume preserving actions of Lie groups and their lattices do not preserve rigid geometric structures in the sense of Gromov. The actions considered are the "exotic" examples obtained by Katok and Lewis and the first author, by blowing up closed orbits in the well known actions on homogeneous spac…
New algorithm reduces sleeping bandits' regret to O(sqrt(T)).