Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

104208311415 · Jun 202019922001200920172026
48 results for variational intrinsic control

Revisits VIC method to correct intrinsic reward bias in stochastic environments.

problem Intrinsic reward bias in VIC leading to suboptimal solutions.
method Proposes two methods based on transitional probability model and Gaussian mixture model to correct bias.
result Achieves maximal empowerment through corrected intrinsic reward.

Universal AI seeks high-optionality states through empowerment and curiosity.

problem Understanding and optimizing AI behavior in uncertain environments.
method Unified framework combining AIXI and variational empowerment, showing how universal AI agents balance goal-directed behavior with uncertainty reduction curiosity.
result Self-AIXI asymptotically converges to AIXI performance and exhibits power-seeking behavior due to intrinsic motivations.

In this paper, we describe a geometric setting for higher-order lagrangian problems on Lie groups. Using left-trivialization of the higher-order tangent bundle of a Lie group and an adaptation of the classical Skinner-Rusk formalism, we deduce an intrinsic framework for this type of dynamical systems. Interesting appli…

2011-04-16abs ↗pdf ↗

It has been established that diverse behaviors spanning the controllable subspace of an Markov decision process can be trained by rewarding a policy for being distinguishable from other policies \citep{gregor2016variational, eysenbach2018diversity, warde2018unsupervised}. However, one limitation of this formulation is …

2019-06-12abs ↗pdf ↗

We discuss a recently proposed variational principle for deriving the variational equations associated to any Lagrangian system. The principle gives simultaneously the Lagrange and the variational equations of the system. We define a new Lagrangian in an extended configuration space ---which we call D'Alambert's--- com…

2001-07-08abs ↗pdf ↗

We discuss intrinsic aspects of Krupka's approach to finite-order variational sequences. We give intrinsic isomorphisms of the quotient subsheaves of the short finite-order variational sequence with sheaves of forms on jet spaces of suitable order, obtaining a new finite-order (short exact) variational sequence which i…

2000-01-05abs ↗pdf ↗

New approach to Lagrangian systems using intrinsic geometry.

problem Developing a new framework for Lagrangian systems.
method Direct reformulation of Hamiltonian formalism, introduction of spatial equation and spatial-gauge symmetry.
result Covariant and non-covariant canonical variational principles demonstrated for Maxwell equations.

We show that in the first sub-Riemannian Heisenberg group there are intrinsic graphs of smooth functions that are both critical and stable points of the sub-Riemannian perimeter under compactly supported variations of contact diffeomorphisms, despite the fact that they are not area-minimizing surfaces. In particular, w…

2016-11-22abs ↗pdf ↗

The mathematical problem concerning intrinsic storage optimisation is formulated and solved by means of variational analysis. The solution, though obtained in implicit form, still sheds light on many important features of the optimal exercise strategy. It is shown how the solution depends on different constraint types …

2015-06-22abs ↗pdf ↗

A new multi-objective RL framework improves intrinsic exploration performance.

problem Sub-optimal exploration performance due to ad-hoc handling of intrinsic exploration.
method A multi-objective RL framework where both exploration and exploitation are optimized as separate objectives.
result EMU-Q method outperforms classic and other intrinsic RL methods on benchmarks.

A new reinforcement learning method uses mutual information to encourage agents to control their environment.

problem Learning from internal drives instead of external rewards.
method Formulate an intrinsic objective as mutual information between goal states and controllable states, derive a surrogate objective for efficient optimization.
result Demonstrated the efficacy of the approach in robotic tasks.

EXOC framework uses auxiliary variables for counterfactual fairness in machine learning.

problem Balancing fairness and predictive accuracy in models with sensitive attributes.
method EXOC framework uses auxiliary variables to define an auxiliary node and a control node for counterfactual fairness.
result EXOC framework outperforms state-of-the-art approaches in achieving counterfactual fairness.

We propose a novel framework to identify sub-goals useful for exploration in sequential decision making tasks under partial observability. We utilize the variational intrinsic control framework (Gregor et.al., 2016) which maximizes empowerment -- the ability to reliably reach a diverse set of states and show how to ide…

2019-07-24abs ↗pdf ↗

We study the intrinsic structure of parametric minimal discs in metric spaces admitting a quadratic isoperimetric inequality. We associate to each minimal disc a compact, geodesic metric space whose geometric, topological, and analytic properties are controlled by the isoperimetric inequality. Its geometry can be used …

2016-02-22abs ↗pdf ↗

Deep neural networks with their large number of parameters are highly flexible learning systems. The high flexibility in such networks brings with some serious problems such as overfitting, and regularization is used to address this problem. A currently popular and effective regularization technique for controlling the…

2017-11-30abs ↗pdf ↗

New estimators for intrinsic dimension and Wasserstein distance improve OT accuracy.

problem Intrinsic dimension estimation and Wasserstein distance estimation in large-scale OT.
method Introduces novel estimators for intrinsic dimension and Wasserstein distance.
result Simple, tuning-free estimator of OT and fast intrinsic dimension estimator.

In this contribution we present an intrinsic description of time-variant Port Hamiltonian systems as they appear in modeling and control theory. This formulation is based on the splitting of the state bundle and the use of appropriate covariant derivatives, which guarantees that the structure of the equations is invari…

2012-07-19abs ↗pdf ↗

Variational inference is increasingly being addressed with stochastic optimization. In this setting, the gradient's variance plays a crucial role in the optimization procedure, since high variance gradients lead to poor convergence. A popular approach used to reduce gradient's variance involves the use of control varia…

2018-10-30abs ↗pdf ↗

Paper proposes a new method to optimize robot body structure and control policy.

problem Optimizing robot body structure and control policy in a coupled manner.
method Revisits co-design problem as a Stackelberg game, incorporating control adaptation dynamics.
result Stackelberg PPO outperforms standard PPO in stability and performance.

We present an intrinsic formulation of the kinematic problem of two nn-dimensional manifolds rolling one on another without twisting or slipping. We determine the configuration space of the system, which is an n(n+3)2\frac{n(n+3)}2-dimensional manifold. The conditions of no-twisting and no-slipping are decoded by means of …

2010-08-11abs ↗pdf ↗

We deal with irregular curves contained in smooth, closed, and compact surfaces. For curves with finite total intrinsic curvature, a weak notion of parallel transport of tangent vector fields is well-defined in the Sobolev setting. Also, the angle of the parallel transport is a function with bounded variation, and its …

2019-06-25abs ↗pdf ↗

This paper develops scalable control variates for Monte Carlo methods using stochastic optimization.

problem Reducing variance in Monte Carlo estimators for large-scale problems.
method Control variates based on Stein operators, optimized through stochastic optimization.
result Novel theoretical results and empirical validations show effective variance reduction.

This study analyzes VAEs using ID and II, revealing a transition in behaviour and distinct training phases.

problem Understanding the hidden representations and training phases of VAEs.
method Analysis using Intrinsic Dimension (ID) and Information Imbalance (II).
result VAEs exhibit a transition in behaviour and distinct training phases when the bottleneck size exceeds the Intrinsic Dimension of the data.

A new method reduces variance in training discrete latent variable models.

problem High variance in stochastic gradient estimators for discrete latent variable models.
method Double control variates for score function estimators using Taylor expansions.
result Our method can have lower variance compared to other estimators.

This work proposes using zero-variance control variates to reduce variance in pathwise gradient estimators for variational inference.

problem Pathwise gradient estimators in variational inference have high variance, leading to inefficient optimization.
method Apply zero-variance control variates to pathwise gradient estimators.
result Zero-variance control variates can significantly reduce the variance of pathwise gradient estimators without requiring complex assumptions.

This paper presents an overview of recent developments in the analysis of shapes such as curves and surfaces through Riemannian metrics. We show that several constructions of metrics on spaces of submanifolds can be unified through the prism of Riemannian submersions, with shape space metrics being induced from metrics…

2018-09-17abs ↗pdf ↗

We use neural networks as control variates with geometric integration techniques.

problem Analytic integration of neural network approximations for variance reduction.
method Integration domain subdivision using computational geometry for MLPs with continuous piecewise linear activation functions.
result Neural networks can be used as control variates with geometric integration methods.

Combines control variates and adaptive importance sampling for Monte Carlo integration.

problem Improving Monte Carlo integration accuracy with control variates and adaptive sampling.
method A quadrature rule combining control variates and adaptive importance sampling.
result Non-asymptotic bound on the probabilistic error of the procedure.

The paper explores how control variates can reduce variance in Monte Carlo simulations, especially for Sobolev functions.

problem Efficiency of control variates in reducing variance for Monte Carlo simulations.
method Study of a specific quadrature rule using nonparametric regression-adjusted control variates.
result A specific quadrature rule can improve the Monte Carlo rate and achieve the minimax optimal rate under sufficient smoothness assumptions.

ADAC uses analogous policies to improve RL exploration without sacrificing stability.

problem Improving RL exploration without compromising stability and expressiveness.
method Disentangled actor-critic approach with analogous pairs of actors and critics.
result Empirical evaluation shows ADAC outperforms alternatives in challenging exploration tasks.

An intrinsic description of the Hamilton-Cartan formalism for first-order Berezinian variational problems determined by a submersion of supermanifolds is given. This is achieved by studying the associated higher-order graded variational problem through the Poincaré-Cartan form. Noether theorem and examples from superfi…

2018-05-25abs ↗pdf ↗

Paper improves variance control in importance weighted variational bounds.

problem Improving the variance of gradient estimators for IWAE.
method Develops a novel control variate that grows SNR as √K for large K.
result Empirically, the method yields superior variance reduction for generative models.

A control system q˙=f(q,u)\dot{q} = f(q,u) is said to be trivializable if there exists local coordinates in which the system is feedback equivalent to a control system of the form q˙=f(u)\dot{q} = f(u). In this paper we characterize trivializable control systems and control systems for which, up to a feedback transformation, ff a…

2009-02-13abs ↗pdf ↗