Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

105210315420 · Jun 202019922001200920172026
48 results for action dependence

Survey of three geometric frameworks for action-dependent field theories.

problem Understanding action-dependent field theories through geometric structures.
method Introduction and analysis of three geometric frameworks: k-contact, k-cocontact, and multicontact.
result Analysis of relationships among these geometric structures and comparison with other definitions.

Policy gradient methods are a widely used class of model-free reinforcement learning algorithms where a state-dependent baseline is used to reduce gradient estimator variance. Several recent papers extend the baseline to depend on both the state and action and suggest that this significantly reduces variance and improv…

2018-02-27abs ↗pdf ↗

We consider a sequential learning problem with Gaussian payoffs and side information: after selecting an action ii, the learner receives information about the payoff of every action jj in the form of Gaussian observations whose mean is the same as the mean payoff, but the variance depends on the pair (i,j)(i,j) (and may…

2015-10-27abs ↗pdf ↗

Paper proposes a new HMM approach for better action recognition.

problem Capturing complex temporal dependency patterns in skeleton-based actions.
method Introduces a hierarchical HMM with a latent variable layer for dynamic inference.
result Proposed approach effectively models complex sequential data and handles missing values.

We consider interactive learning and covering problems, in a setting where actions may incur different costs, depending on the response to the action. We propose a natural greedy algorithm for response-dependent costs. We bound the approximation factor of this greedy algorithm in active learning settings as well as in …

2016-02-23abs ↗pdf ↗

We give a diameter bound for fundamental domains for isometric actions of the fundamental group of a closed hyperbolic surface on a delta-hyperbolic space, where the bound depends on the hyperbolicity constant delta, the genus of the surface, and the injectivity radius of the action, which we assume to be strictly posi…

2007-09-17abs ↗pdf ↗

Bandit algorithms have various application in safety-critical systems, where it is important to respect the system constraints that rely on the bandit's unknown parameters at every round. In this paper, we formulate a linear stochastic multi-armed bandit problem with safety constraints that depend (linearly) on an unkn…

2019-08-16abs ↗pdf ↗

We exhibit large classes of local actions for the vacuum Einstein equations. In presence of fermions, or more generally of matter which couple to the connection, these actions lead to inequivalent equations revealing an arbitrary number of parameters. Even in the pure gravitational sector, any corresponding quantum the…

2009-07-17abs ↗pdf ↗

We discuss how the global geometry and topology of manifolds depend on different group actions of their fundamental groups, and in particular, how properties of a non-trivial compact 4-dimensional cobordism MM whose interior has a complete hyperbolic structure depend on properties of the variety of discrete representa…

2016-11-02abs ↗pdf ↗

PQR estimates reward functions from actions and states without assuming state-only rewards.

problem Estimating reward functions from actions and states without state-only assumptions.
method Deep learning approach that sequentially estimates policy, Q-function, and reward.
result PQR uniquely recovers true reward with known transitions and bounds error with unknown transitions.

Develops a new reinforcement learning framework for complex control problems.

problem Continuous-time extended mean field control with deterministic policies.
method Model-free sensitivity formula, deterministic policy gradient, local value and advantage-rate representations.
result Demonstrates efficiency, stability, and robustness in solving complex control problems.

New algorithm reduces regret in private online learning with optimal gap-dependent rate.

problem Optimal gap-dependent regret rate for private stochastic decision-theoretic online learning.
method Horizon-free pure-DP algorithm with exponential block partitioning and softmax selection.
result Explicit regret bound of 1000(logKΔmin+logKε)1000 \cdot (\frac{\log K}{Δ_{\min}}+\frac{\log K}{\varepsilon}).

New algorithm for bandits with delayed action effects, reducing regret.

problem Delayed impact of actions in multi-armed bandits.
method Formulated a new bandit setting with delayed action effects, proposed an algorithm with regret bound.
result Achieved a regret of ildeO(KT2/3) ilde{\mathcal{O}}(KT^{2/3}) and showed a matching lower bound.

This paper provides theoretical foundations for using quantized actions in behavior cloning.

problem Applying autoregressive models to continuous control requires discretizing actions through quantization, which is poorly understood.
method The paper analyzes quantization error propagation and statistical sample complexity, and proposes model-based augmentation.
result Behavior cloning with quantized actions achieves optimal sample complexity, matching existing lower bounds.

Unified framework for corruption-robust linear bandits with optimal gap-dependent misspecification bounds.

problem Effective learning in linear bandits with corrupted rewards across different corruption models.
method Unified framework for analyzing strong and weak corruption, connection to gap-dependent misspecification, and specialized algorithm.
result Optimal bounds for gap-dependent misspecification in linear bandits.

We introduce a rich class of graphical models for multi-armed bandit problems that permit both the state or context space and the action space to be very large, yet succinctly specify the payoffs for any context-action pair. Our main result is an algorithm for such models whose regret is bounded by the number of parame…

2012-02-14abs ↗pdf ↗

New algorithm tackles multi-agent reinforcement learning with optimal convergence rate.

problem Multi-agent reinforcement learning with large state spaces and linear function approximations.
method Refined AVLPR framework with data-dependent pessimistic estimation and action-dependent bonuses.
result First algorithm with optimal O(T1/2)O(T^{-1/2}) convergence rate and no poly(AmaxA_{\max}) dependency.

We present and study a partial-information model of online learning, where a decision maker repeatedly chooses from a finite set of actions, and observes some subset of the associated losses. This naturally models several situations where the losses of different actions are related, and knowing the loss of one action p…

2014-09-30abs ↗pdf ↗

Study of symplectomorphisms on ruled surfaces under circle actions.

problem Homotopy type of equivariant symplectomorphisms on rational ruled surfaces.
method Analysis of action on compatible and invariant almost complex structures, use of Delzant's and Karshon's classifications.
result Equivariant symplectomorphisms are homotopy equivalent to tori or their pushout.

We derive a consistent differential representation for the dynamics of a self-financing portfolio for different hedging strategies. In the basis of the derivation there is the so called "retarded action principle", which represents the causality in the evolution of dependent stochastic variables. We demonstrate this pr…

2015-09-30abs ↗pdf ↗

Functional determinant for mixed signature sphere products depends on sphere dimensions and parity.

problem Determining the functional determinant for scalar fields on mixed signature sphere products.
method Analyzing the GJMS operator on Sqimes^q imesSp^p to derive the functional determinant.
result The functional determinant depends only on the total dimension and parity of the sphere dimensions.

Describes reconstructing Poisson structures from Lie group actions.

problem Reconstructing invariant Poisson structures from Lie group actions.
method Describes reconstruction of invariant Poisson structures from canonical actions of compact Lie groups on fibered phase spaces.
result Derives symmetry properties of Wong's type equations from main results.

Algorithm maximizes rewards with a budget and giving up option.

problem Sequential decision-making with stochastic rewards and resource consumption.
method Upper Confidence Bound (UCB) algorithm for maximizing cumulative reward.
result Logarithmic regret bound with improved dependence on problem parameters.

New method for linear bandits with unknown sparsity, improving sparse regret bounds.

problem Sparse regret bounds for unknown sparsity and adversarial action sets.
method Combines online to confidence set conversions with randomized model selection over nested confidence sets.
result First sparse regret bounds for unknown sparsity and adversarial action sets.

Finite p-group actions on manifolds have limited stabilizer subgroups.

problem Understanding the structure of stabilizer subgroups in group actions on manifolds.
method Bounding the index of a subgroup H in a finite p-group G acting on a compact manifold M, ensuring a controlled number of stabilizers.
result The existence of a subgroup H with a controlled index and limited stabilizers.

We consider the reduced Allen-Cahn action functional, which appears as the sharp interface limit of the Allen-Cahn action functional and can be understood as a formal action functional for a stochastically perturbed mean curvature flow. For suitable evolutions of generalized hypersurfaces this functional consists of th…

2013-04-07abs ↗pdf ↗

Rigidity of elliptic genera proven for non-spin manifolds with S1S^1-action.

problem Rigidity of elliptic genera for non-spin manifolds with S1S^1-action.
method Analysis of universal covering spin condition and π2(M)π_2(M) for rigidity.
result Rigidity of elliptic genera is proven for spin universal coverings but not for non-spin universal coverings.

Extends integrability to cosymplectic manifolds.

problem Integrability of Hamiltonian systems on cosymplectic manifolds.
method Extended Arnold-Liouville and noncommutative integrability to cosymplectic manifolds, proved a variant of non-commutative integrability for specific fields, constructed action-angle variables.
result Variant of non-commutative integrability for evaluation and Reeb vector fields on cosymplectic manifolds.

We compute the quotient of the self-duality equation for conformal metrics by the action of the diffeomorphism group. We also determine Hilbert polynomial, counting the number of independent scalar differential invariants depending on the jet-order, and the corresponding Poincaré function. We describe the field of rati…

2016-05-04abs ↗pdf ↗

Improved reinforcement learning for episodes with varying action sets.

problem Reinforcement learning with context-dependent action sets.
method Extends MVP algorithm to handle adversarial and stochastic contexts.
result Established minimax regret bounds of O(SAH3KlogL)O(\sqrt{SAH^3K\log L}) for adversarial contexts.

SPEDER extracts state-action abstraction from dynamics for reinforcement learning.

problem Curse of dimensionality and limited applicability of spectral methods.
method Spectral Decomposition Representation (SPEDER) that extracts state-action abstraction from dynamics without policy dependence.
result Theoretical analysis establishes sample efficiency in online and offline settings.

The paper calculates variations of Einstein-Hilbert action on CR manifolds.

problem Variation of the Einstein-Hilbert action in pseudohermitian geometry.
method Computed first and second variations on CR manifolds, characterized critical points as pseudo-Einstein structures, and analyzed second variation on standard spheres.
result In three dimensions, the second variation of the Einstein-Hilbert action on CR structures differs from the Riemannian case due to embeddability.

Let JJ be a semisimple Lie group with all simple factors of real rank at least two. Let Γ<JΓ<J be a lattice. We prove a very general local rigidity result about actions of JJ or ΓΓ. This shows that almost all so-called "standard actions" are locally rigid. As a special case, we see that any action of ΓΓ by toral aut…

2004-08-16abs ↗pdf ↗