Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

8.3%16.7%25.0%33.3% · Jan 199319922001200920182026
48 results for open loop control

Novel algorithm for optimal control of nonlinear systems.

problem Optimal control of nonlinear stochastic dynamical systems with unknown dynamics.
method Decoupled data-based approach combining open-loop and closed-loop control.
result Performance of D2C algorithm is approximately optimal and significantly reduces training time.

Paper uses deep reinforcement learning for better control of rocket engines during start-up phases.

problem Lack of optimal control during transient phases of liquid rocket engines.
method Deep reinforcement learning approach for optimal control of a gas-generator engine's continuous start-up phase.
result Deep reinforcement learning controller achieves highest performance and minimal computational effort.

Paper tackles stochastic control with mean and higher-order moments, finding Nash equilibria.

problem Time-inconsistent stochastic control problems with mean and higher-order moments.
method Developed closed-loop and open-loop Nash equilibrium controls using PDEs and maximum principles.
result Identical closed-loop and open-loop Nash equilibria controls, independent of state value and random path.

Method improves volatility targeting for index construction.

problem High turnover, leverage spikes, and sensitivity to estimation error in existing volatility-targeting strategies.
method Proportional-control approach for setting index weights that corrects tracking error through feedback.
result The proportional-control approach achieves the target volatility more effectively than open-loop alternatives.

Investigates time-inconsistent portfolio selection under MMV preferences.

problem Time-inconsistent optimal strategies for MMV preferences.
method Nash equilibrium controls for MMV and MV preferences, solving FBSDE and HJB equations.
result MMV optimal strategies lead to higher investment amounts than MV strategies, narrowing over time.

This paper uses NLDT to find interpretable control rules from complex DRL policies.

problem Complex, non-interpretable policies from black-box AI methods.
method Evolutionary optimization of NLDT for hierarchical control rules.
result Interpretable control rules with similar performance to black-box DRL.

New algorithm avoids re-planning in tree-search algorithms, reducing suboptimal actions.

problem Avoiding re-planning in tree-search algorithms to reduce suboptimal actions.
method A new algorithm that decides at each step whether to re-plan or use a sub-tree, based on sub-tree statistics.
result The probability of selecting a suboptimal action converges to zero and decays logarithmically.

Study on estimating unstable open-loop matrices from state trajectories.

problem System identification for stochastic continuous-time dynamics.
method Employing randomized control inputs to estimate unstable open-loop matrix.
result Estimation error decays with trajectory length, signal-to-noise ratio, and excitability.

Control Contraction Metrics (CCMs) provide a nonlinear controller design involving an offline search for a Riemannian metric and an online search for a shortest path between the current and desired trajectories. In this paper, we generalize CCMs to Finsler geometry, allowing the use of non-Riemannian metrics. We provid…

2018-03-02abs ↗pdf ↗

Combining causality, control, and reinforcement learning for system control.

problem Learning to control dynamical systems using causal, control, and reinforcement learning approaches.
method Combining causal identification, control strategies, and reinforcement learning to control dynamical systems.
result Combining different learning paradigms for effective system control.

Framework for robust control in cooperative systems with uncertain common noise.

problem Optimizing collective behavior of agents in the presence of uncertain common noise.
method Proposes a robust mean-field control framework and proves existence of optimal controls.
result Existence of optimal open-loop controls linked to a lifted robust Markov decision problem.

We show LLMs can be locally linear, enabling better control of activations.

problem Suboptimal control of LLM activations during generation.
method Model LLM inference as a linear dynamical system, compute feedback controllers using Jacobians, and adapt classical control theory.
result Robust, fine-grained control of LLM activations across models and tasks.

The paper analyzes strategic irreversible investments with novel dynamic strategies.

problem Tradeoff between preemption incentives and option value of waiting in oligopolistic markets.
method Developed novel Markov perfect equilibrium to handle singular control of optimal investment.
result Simpler strategies lead to a 'preemption trap' with zero net present values.

Efficient learning-based MPC for unknown nonlinear systems with state constraints.

problem Control of discrete-time nonlinear systems with unknown dynamics and state constraints.
method Receding horizon reinforcement learning (r-LPC) using Koopman operator-based prediction model.
result Proven closed-loop recursive feasibility, robustness, and asymptotic stability under function approximation errors.

Reinforcement Learning optimizes low-thrust interplanetary trajectories under disturbances.

problem Designing robust interplanetary trajectories in the presence of disturbances.
method Reformulated as a Markov Decision Process, RL algorithm Proximal Policy Optimization trained on a deep neural network.
result Deep neural network provides robust nominal trajectory and guidance law.

DeepRacing uses neural networks to predict trajectories for autonomous racing in video games.

problem Training algorithms for high-speed autonomous racing in realistic environments.
method Developed a virtual testbed using F1 video games, trained neural networks to predict trajectories and control commands.
result Trajectory prediction outperforms end-to-end control methods in autonomous racing simulations.

AdaptOn achieves logarithmic regret in adaptive control of unknown partially observable linear systems.

problem Adaptive control in partially observable linear dynamical systems.
method AdaptOn algorithm that estimates system dynamics through online learning and gradient descent.
result AdaptOn achieves a logarithmic regret bound of polylog(T) after T steps.

Improved algorithm for online planning with tighter bounds.

problem Online planning in Markov Decision Processes with open-loop policies and budget constraints.
method Proposed KLOLOP algorithm with tighter upper-confidence bounds and efficient implementation.
result KLOLOP leads to better practical performances with sample complexity bound retained.

Framework learns robust control policies from expert demonstrations.

problem Adversarial robustness and closed-loop generalization in feedback control policies.
method Lipschitz-constrained loss minimization for certified robustness and generalization.
result Finite sample bound on policy learning error and robust closed-loop stability.

Paper proposes a method to monitor industrial processes under closed-loop control.

problem Difficulty distinguishing between real process faults and normal operating conditions changes.
method Develops a distributed monitoring system by capturing static and dynamic characteristics of large-scale closed-loop industrial processes.
result The method effectively distinguishes between real process faults and normal operating conditions changes.

Develops a method to plan exploration that learns strong policies with fewer samples.

problem Lack of efficient exploration in reinforcement learning for real-world tasks.
method Plans an action sequence that maximizes information gain about the optimal trajectory.
result 2x fewer samples than exploration baselines and 200x fewer than model-free methods.

In this paper, we formulate a general time-inconsistent stochastic linear--quadratic (LQ) control problem. The time-inconsistency arises from the presence of a quadratic term of the expected state as well as a state-dependent term in the objective functional. We define an equilibrium, instead of optimal, solution withi…

2011-11-03abs ↗pdf ↗

A new approach optimizes weights in DLP for better risk-adjusted performance.

problem Optimizing time-varying weights in Double Linear Policy (DLP) for better risk-adjusted performance.
method Stochastic Model Predictive Control (SMPC) framework to maximize risk-adjusted returns while enforcing constraints.
result Empirical results show improved risk-adjusted performance and drawdown control.

Formula adjusts steady-state models for control confounding.

problem Learning steady-state models from operational data can be flawed due to control confounding.
method Derives a formula to adjust for control confounding using structural dynamical causal models.
result Estimates a causal steady-state model from closed-loop operational data.

Geodesic loops escape from balls at a sublinear rate imply virtually abelian fundamental group.

problem Understanding fundamental groups of open manifolds with nonnegative Ricci curvature.
method Generalizing the Cheeger-Gromoll splitting theorem to sublinear escape rates.
result Fundamental groups of open manifolds with nonnegative Ricci curvature are virtually abelian if geodesic loops escape sublinearly.

Paper certifies neural network control policies against persistent adversarial perturbations.

problem Neural networks' fragility to adversarial perturbations in control systems.
method Combining neural network certification tools with robust control theory.
result Certifies neural network policies in a control loop under l-infinity norm bounded adversarial perturbations.

We propose a model of inter-bank lending and borrowing which takes into account clearing debt obligations. The evolution of log-monetary reserves of NN banks is described by coupled diffusions driven by controls with delay in their drifts. Banks are minimizing their finite-horizon objective functions which take into a…

2016-07-21abs ↗pdf ↗

Action chunking and data exploration improve behavior cloning in robotics.

problem Exponential errors in learning from demonstrations for continuous control tasks.
method Action chunking and exploratory data collection.
result Control-theoretic stability is key to improving imitation learning.

This work discusses a closed-loop control strategy for complex systems utilizing scarce and streaming data. A discrete embedding space is first built using hash functions applied to the sensor measurements from which a Markov process model is derived, approximating the complex system's dynamics. A control strategy is t…

2016-04-11abs ↗pdf ↗

Advocates a local feedback approach for RL in unknown systems.

problem Finding optimal feedback laws in unknown nonlinear dynamical systems.
method Searches over a local feedback representation consisting of an open-loop sequence and an optimal linear feedback law.
result Results in highly efficient training and superior performance compared to global methods.

The paper studies the n-loop Kontsevich invariant of knots with same Alexander polynomial.

problem Understanding the n-loop Kontsevich invariant for knots with identical Alexander polynomials.
method Analyzes the subspace generated by the n-loop Kontsevich invariant of knots with genus ≤ g and same Alexander polynomial.
result For n ≥ 2, the subspace is finite-dimensional.

Deep learning models outperform classical methods in forecasting neural activity.

problem Improving forecasting of neural activity using deep learning models.
method Systematic evaluation of eight probabilistic deep learning models against classical statistical models and baseline methods.
result Several deep learning models consistently outperform classical approaches in forecasting neural activity.

We show that if the monodromy of an open book decomposition has sufficiently high displacement distance, acting on the loop and arc complex for a page, then it is the unique minimal Euler characteristic open book for the manifold. In particular, we show that such an open book induces the unique (up to isotopy) minimal …

2011-10-10abs ↗pdf ↗

Paper characterizes equilibrium strategies for stochastic control with higher-order moments.

problem Stochastic control problems with higher-order moments.
method Novel characterization of time-consistent control problems, deriving equilibrium conditions via BSDEs.
result Derives sufficient and necessary conditions for an open-loop Nash equilibrium control (ONEC) in a novel way.

Let (M,d) be a metric space. For 0<r<R, and p in M let G(p,r,R) be the group obtained by considering all loops based at p whose image is contained in the closed ball of radius r and identifying two loops if there is a homotopy betweeen them that is contained in the open ball of radius R. In this paper we study the asym…

2005-10-07abs ↗pdf ↗

Geometrically characterizes virtual nonlinear nonholonomic constraints using symplectic methods.

problem Characterizing virtual nonlinear nonholonomic constraints geometrically.
method Geometric characterization using symplectic structures and Chetaev equations.
result A unique control law exists to satisfy virtual constraints, and closed-loop dynamics are projections of uncontrolled dynamics.

New method allows sheets to morph into multiple shapes via spatially varying stimuli.

problem Limitation of current shape-programmed sheets to achieve only one target geometry.
method Patterning the stimulus itself for spatiotemporal control over local deformation magnitudes.
result A single physical sample can be induced to traverse a continuous family of target geometries.

Study time-inconsistent consumption-investment in incomplete markets with general discount functions.

problem Time-inconsistent consumption-investment problems in incomplete markets.
method Coupled forward-backward stochastic differential equation approach.
result Uniqueness of open-loop equilibrium pair proved.