Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,738 papers · 148 categories

Trend · papers per month

69139208277 · May 202619922001200920172026
48 results for intrinsic control

Revisits VIC method to correct intrinsic reward bias in stochastic environments.

problem Intrinsic reward bias in VIC leading to suboptimal solutions.
method Proposes two methods based on transitional probability model and Gaussian mixture model to correct bias.
result Achieves maximal empowerment through corrected intrinsic reward.

A new multi-objective RL framework improves intrinsic exploration performance.

problem Sub-optimal exploration performance due to ad-hoc handling of intrinsic exploration.
method A multi-objective RL framework where both exploration and exploitation are optimized as separate objectives.
result EMU-Q method outperforms classic and other intrinsic RL methods on benchmarks.

A new reinforcement learning method uses mutual information to encourage agents to control their environment.

problem Learning from internal drives instead of external rewards.
method Formulate an intrinsic objective as mutual information between goal states and controllable states, derive a surrogate objective for efficient optimization.
result Demonstrated the efficacy of the approach in robotic tasks.

We study the intrinsic structure of parametric minimal discs in metric spaces admitting a quadratic isoperimetric inequality. We associate to each minimal disc a compact, geodesic metric space whose geometric, topological, and analytic properties are controlled by the isoperimetric inequality. Its geometry can be used …

2016-02-22abs ↗pdf ↗

New estimators for intrinsic dimension and Wasserstein distance improve OT accuracy.

problem Intrinsic dimension estimation and Wasserstein distance estimation in large-scale OT.
method Introduces novel estimators for intrinsic dimension and Wasserstein distance.
result Simple, tuning-free estimator of OT and fast intrinsic dimension estimator.

In this contribution we present an intrinsic description of time-variant Port Hamiltonian systems as they appear in modeling and control theory. This formulation is based on the splitting of the state bundle and the use of appropriate covariant derivatives, which guarantees that the structure of the equations is invari…

2012-07-19abs ↗pdf ↗

Paper proposes a new method to optimize robot body structure and control policy.

problem Optimizing robot body structure and control policy in a coupled manner.
method Revisits co-design problem as a Stackelberg game, incorporating control adaptation dynamics.
result Stackelberg PPO outperforms standard PPO in stability and performance.

We present an intrinsic formulation of the kinematic problem of two nn-dimensional manifolds rolling one on another without twisting or slipping. We determine the configuration space of the system, which is an n(n+3)2\frac{n(n+3)}2-dimensional manifold. The conditions of no-twisting and no-slipping are decoded by means of …

2010-08-11abs ↗pdf ↗

A control system q˙=f(q,u)\dot{q} = f(q,u) is said to be trivializable if there exists local coordinates in which the system is feedback equivalent to a control system of the form q˙=f(u)\dot{q} = f(u). In this paper we characterize trivializable control systems and control systems for which, up to a feedback transformation, ff a…

2009-02-13abs ↗pdf ↗

New RL approach uses future state and action visitation measures for better exploration.

problem Improving exploration in reinforcement learning.
method Intrinsic reward based on future state and action visitation measures, using contraction operators.
result Policies achieve good state-action space coverage and high performance.

This work discusses a closed-loop control strategy for complex systems utilizing scarce and streaming data. A discrete embedding space is first built using hash functions applied to the sensor measurements from which a Markov process model is derived, approximating the complex system's dynamics. A control strategy is t…

2016-04-11abs ↗pdf ↗

The paper introduces a new intrinsic reward method for exploration in reinforcement learning.

problem Improving exploration in reinforcement learning agents.
method Intrinsic rewards proportional to the entropy of future state-action features.
result The new objective leads to improved visitation of features within individual trajectories.

Universal AI seeks high-optionality states through empowerment and curiosity.

problem Understanding and optimizing AI behavior in uncertain environments.
method Unified framework combining AIXI and variational empowerment, showing how universal AI agents balance goal-directed behavior with uncertainty reduction curiosity.
result Self-AIXI asymptotically converges to AIXI performance and exhibits power-seeking behavior due to intrinsic motivations.

This paper improves robot grasping by integrating meta-control and latent-space imagination.

problem Dual-system approaches fail to consider the reliability of the learned model when making multiple-step predictions.
method A meta-controller arbitrates between model-based and model-free decisions based on local reliability, encouraging actions that improve the model and generating imagined experiences for additional training.
result Our approach learns near-optimal grasping policies in dense- and sparse-reward environments, outperforming baseline and state-of-the-art methods.

Study on rolling Stiefel manifolds with specific metrics.

problem Intrinsic and extrinsic rolling of Stiefel manifolds with αα-metrics.
method Investigation of intrinsic rolling of normal naturally reductive homogeneous spaces, derivation of ODEs for rolling, and explicit solutions.
result Explicit solutions for intrinsic and extrinsic rolling of Stiefel manifolds.

A well known question in differential geometry is to control the constant in isoperimetric inequality by intrinsic curvature conditions. In dimension 2, the constant can be controlled by the integral of the positive part of the Gaussian curvature. In this paper, we showed that on simply connected conformal flat manifol…

2013-06-07abs ↗pdf ↗

Low-dimensional structure in images helps deep learning models generalize better.

problem Understanding the intrinsic dimensionality of images for better model performance.
method Applied dimension estimation tools to popular image datasets and used GANs to manipulate intrinsic dimensionality.
result Natural image datasets have very low intrinsic dimensionality, which aids neural networks in learning and generalizing.

We prove that for the mean curvature flow of two-convex hypersurfaces the intrinsic diameter stays uniformly controlled as one approaches the first singular time. We also derive sharp Ln1L^{n-1}-estimates for the regularity scale of the level set flow with two-convex initial data. Our proof relies on a detailed analysis…

2017-10-27abs ↗pdf ↗

In this paper, we describe a geometric setting for higher-order lagrangian problems on Lie groups. Using left-trivialization of the higher-order tangent bundle of a Lie group and an adaptation of the classical Skinner-Rusk formalism, we deduce an intrinsic framework for this type of dynamical systems. Interesting appli…

2011-04-16abs ↗pdf ↗

Study of Brown--York mass for four-dimensional asymptotically flat manifolds.

problem Calculating mass for hypersurfaces in four-dimensional asymptotically flat manifolds.
method Intrinsic definition of mean curvature, expansion analysis for large uniformly convex hypersurfaces.
result Shape-dependent correction to ADM mass for nearly round surfaces vanishes under certain conditions.

Novelty search in low-dimensional space improves sample efficiency in exploration tasks.

problem Efficient exploration in complex environments with sparse rewards.
method Combines model-based and model-free objectives to learn a low-dimensional representation. Uses intrinsic novelty rewards based on nearest neighbor distances in this space.
result Our approach achieves more sample-efficient exploration compared to strong baselines on various tasks.

In [7] Klainerman introduced the hyperboloidal method to prove the global existence results for nonlinear Klein-Gordon equations by using commuting vector fields. In this paper, we extend the hyperboloidal method from Minkowski space to Lorentzian spacetimes. This approach is developed in [14] for proving, under the ma…

2016-07-06abs ↗pdf ↗

A new method monitors unstructured 3D shapes without registration.

problem Error-prone registration and mesh reconstruction steps in PCD monitoring.
method Intrinsic geometric properties of shapes, using Laplacian and geodesic distances.
result Effective monitoring of defects without registration and mesh reconstruction.

It has been established that diverse behaviors spanning the controllable subspace of an Markov decision process can be trained by rewarding a policy for being distinguishable from other policies \citep{gregor2016variational, eysenbach2018diversity, warde2018unsupervised}. However, one limitation of this formulation is …

2019-06-12abs ↗pdf ↗

EXOC framework uses auxiliary variables for counterfactual fairness in machine learning.

problem Balancing fairness and predictive accuracy in models with sensitive attributes.
method EXOC framework uses auxiliary variables to define an auxiliary node and a control node for counterfactual fairness.
result EXOC framework outperforms state-of-the-art approaches in achieving counterfactual fairness.

Deep neural networks solve stochastic control problems with delay.

problem Challenges in stochastic control problems with delay due to path-dependence and high dimensions.
method Employing recurrent neural networks (RNNs) to parameterize policies and optimize objectives.
result RNNs, especially LSTMs, efficiently capture path-dependence and outperform feedforward networks in training and performance.

Survey explores geometric aspects of policy optimization in control systems.

problem Understanding the geometric relationships between control design and optimization.
method Geometric perspective on policy optimization, focusing on parameterization and topology.
result Implications of policy geometry on stability and performance of local search algorithms.

Federated framework learns causal states to predict counterfactuals without centralizing data.

problem Decentralized counterfactual reasoning in coupled industrial systems with private data.
method Federated causal representation learning in state-space systems.
result Proves convergence to centralized oracle and provides privacy guarantees.

We study deformations of Lie groupoids by means of the cohomology which controls them. This cohomology turns out to provide an intrinsic model for the cohomology of a Lie groupoid with values in its adjoint representation. We prove several fundamental properties of the deformation cohomology including Morita invariance…

2015-10-08abs ↗pdf ↗

The famous Nash embedding theorem published in 1956 was aiming for the opportunity to use extrinsic help in the study of (intrinsic) Riemannian geometry, if Riemannian manifolds could be regarded as Riemannian submanifolds. However, this hope had not been materialized yet according to \cite{G}. The main reason for this…

2013-07-07abs ↗pdf ↗

Study on NNs for forecasting time series with novel control variable combinations.

problem Forecast future time series with novel combinations of control variables.
method Modular NN architecture with inductive bias for independence of control variables.
result Modular NN architecture improves forecasting of dependent variables up to large horizons.

New budget quantifies drift in closed-loop learning, improving reproducibility.

problem Characterizing statistical learning under distributional drift in closed-loop settings.
method Introduces an intrinsic drift budget CTC_T quantifying cumulative information-geometric motion of the data distribution.
result Proves a drift-feedback bound of order T1/2+CT/TT^{-1/2}+C_T/T for prequential reproducibility, up to controlled second-order remainder terms.

Robot learns multiple tasks hierarchically by transferring knowledge.

problem Learning multiple complex tasks in open-ended environments.
method Task-oriented procedures, goal-babbling, imitation learning, active learning, intrinsic motivation.
result Robots can learn complex tasks more efficiently by transferring knowledge from simpler ones.

In the present paper we give a historical account -ranging from classical to modern results- of the problem of rolling two Riemannian manifolds one on the other, with the restrictions that they cannot instantaneously slip or spin one with respect to the other. On the way we show how this problem has profited from the d…

2013-01-27abs ↗pdf ↗

Introduces intrinsic Hopf-Lax semigroup linking to intrinsic slope.

problem Understanding intrinsic Hopf-Lax semigroup and its relation to intrinsic slope.
method Introduces and proves the link between intrinsic Hopf-Lax semigroup and intrinsic slope.
result Intrinsic Hopf-Lax semigroup is a subsolution of Hamilton-Jacobi type equality.