The paper tackles data-driven optimal control of unknown nonlinear systems using RKHS.
problem Unknown nonlinear dynamics and stage cost functions.
method Embed state densities into RKHS, learn Markov operators, solve Hamilton-Jacobi-Bellman recursions.
result Solves a wide range of nonlinear control problems, including depth regulation.
The paper tackles robust control with uncertain dependence using data-driven methods.
problem Nonparametric robust control under dependence uncertainty in multi-period stochastic systems.
method Nonparametric adaptive robust control framework using stochastic gradient descent ascent algorithm.
result The controller benefits from knowing more about the uncertain model.
Random features enhance control of complex systems.
problem Flexible nonlinear models for control-affine systems.
method Random features approximations for control-affine structure.
result Methods formalized and shown to relate to ADP and AD kernels.
Unified approach for data-driven control of stochastic processes.
problem Developing practical strategies for stochastic control problems with unknown dynamics.
method Reduction to rate-optimal estimators of invariant distribution risk.
result Data-driven strategies can achieve better performance than known methods.
This paper optimizes autonomous vehicle controllers using data-driven methods.
problem Designing robust controllers for autonomous vehicles that handle external and internal disturbances.
method Data-driven approach using principal component analysis and time delay neural networks.
result Improved controller performance through a feed-forward compensator.
End-to-end algorithm for controlling bilinear systems with probabilistic noise.
problem Controlling bilinear systems with noisy data.
method Proposes an end-to-end algorithm using statistical learning theory and robust controller design.
result Derived finite sample identification error bounds and structurally suitable for control.
Data-driven control of robotic systems using Koopman operators with error bounds.
problem Real-time control of nonlinear robotic systems with unknown dynamics.
method Constructing a Koopman operator-based linear representation using higher-order derivatives of nonlinear dynamics, with error bounds derived from Taylor series accuracy analysis.
result The Koopman model provides marginally better performance than competing nonlinear modeling methods and can be efficiently controlled using linear control design tools.
This paper compares model-based and model-free control methods using neural networks.
problem Comparing model-based and model-free control methods for unknown nonlinear systems.
method Utilizes Deep Koopman Representation (DKRC) and Deep Deterministic Policy Gradient (DDPG) for control.
result DKRC outperforms DDPG in terms of control strategies and accuracy for unknown dynamics.
Data-driven method approximates Koopman generator for system identification and control.
problem Approximating Koopman generator for system identification and control.
method gEDMD (extended dynamic mode decomposition) for deterministic and stochastic systems.
result Data-driven approximation of Koopman generator for system identification and control.
Study of multidimensional control problems with reflection controls.
problem Solving control problems with reflection controls in multidimensional settings.
method Gradient descent algorithm for polytope approximations, data-driven domain estimator, episodic learning algorithm.
result Data-driven solutions for unknown diffusion dynamics with sublinear regret.
Paper develops a novel approach for optimal control using kernel methods.
problem Optimal control of nonlinear stochastic systems.
method Infinitesimal generator approach in reproducing kernel Hilbert spaces.
result Data-driven solution to optimal control problems.
Paper uses DeePC to improve urban traffic lights, reducing congestion and emissions.
problem Urban traffic congestion in expanding cities.
method Behavioral system theory, data-driven control, DeePC algorithm.
result DeePC outperforms existing traffic control methods in travel time and CO2 emissions.
A new deep reinforcement learning model for urban traffic control.
problem Complex traffic dynamics in urban intersections.
method Combines deep learning tricks to solve multiple intersections control problems efficiently.
result Outperforms traditional rule-based approaches in simulations.
Survey of deep RL in intelligent transportation systems.
problem Optimizing traffic signals and autonomous driving using deep RL.
method Comprehensive review of deep RL applications in traffic control and autonomous driving.
result Summarizes existing works in deep RL-based transportation applications.
The paper proposes a method to identify causal structure in complex dynamical systems.
problem Spurious correlations in data-driven models limit the performance of control systems.
method The method leverages controllability concepts to compute input trajectories and uses causal inference techniques.
result The method reliably identifies the true causal structure of control systems from real-world data.
The abstract discusses open data resources for studying and controlling the spread of COVID-19.
problem Understanding and controlling the spread of COVID-19.
method Identification and description of open data resources and data-driven methodologies.
result Identification of variables and open data resources for analyzing COVID-19.
Paper adds a restart mechanism to a drawdown control policy for better trading performance.
problem Missed profitable opportunities when drawdown limit is close to reality.
method Integrates a data-driven restart mechanism into the drawdown modulation trading system.
result The restart mechanism improves trading performance even with transaction costs.
OptCS optimizes model selection after conformal inference, controlling FDR and power loss.
problem Challenges in model selection for conformal inference, especially when limited labeled data and many model choices are available.
method OptCS framework that allows valid statistical testing after flexible data-driven model optimization, using novel multiple testing procedures.
result Valid conformal p-values constructed despite substantial data reuse, maintaining FDR control.
A neural network approach solves optimal decumulation problems for pension plans.
problem Optimal asset allocation and withdrawal strategies for DC pension holders.
method Data-driven neural network optimization with customized activation functions.
result The neural network approach learns near-optimal solutions comparable to HJB PDE methods.
New RL algorithms improve control tasks with data reuse.
problem Real-world control requires performance guarantees and data efficiency.
method Generalized Policy Improvement combining on-policy guarantees and sample reuse.
result Extensive experimental analysis shows benefits of new algorithms.
Safe control of systems with unknown dynamics using persistent excitation.
problem Tension between safety and exploration in data-driven control.
method System identification through persistent excitation, robust constraint satisfaction, and synthesis of feedback controllers.
result Non-asymptotic guarantees on estimation and controller performance.
A framework for data-driven decision-making in infectious disease control.
problem Optimizing trade-offs between public health and economic impacts.
method Multi-objective model-based reinforcement learning.
result Pareto-optimal policies minimizing long-term costs.
This paper examines how data affects risk measures in uncertain distributions.
problem How does distributional ambiguity affect risk measures?
method Formulated and derived simpler dual problems for infinite and finite dimensional robust moment problems.
result Developed theory and conducted experiments in inventory control and portfolio management.
Data-driven method for error estimation without needing class complexity.
problem Constructing confidence intervals for a class of estimates.
method Data-driven approach to derive high-probability upper bounds on maximum error.
result Method naturally adapts to unknown correlation structures and works for finite and infinite classes.
End-to-end framework optimizes constrained trajectories using data-driven methods.
problem Optimizing trajectories under constraints with limited dynamics knowledge.
method Data-driven approach decomposes trajectories into function basis, uses maximum a posteriori for optimization, and incorporates linear constraints.
result Commanding results in aeronautics and sailing route optimization.
This work evaluates and benchmarks calibration metrics for data-driven regression models.
problem Conflicting results from different calibration metrics make it hard to compare and interpret model performance.
method Systematically extracted and benchmarked 14 regression calibration metrics across various data types and recalibration methods.
result Many metrics disagree on the same recalibration result, highlighting the need for careful metric selection.
New control strategy minimizes infected individuals in SIS epidemics.
problem Developing effective control strategies for SIS epidemics.
method Stochastic optimal control of SDEs with jumps, using treatment intensities.
result Control strategy consistently outperforms alternatives in synthetic data.
Automates bias control in reinforcement learning algorithms.
problem Overestimation bias in reinforcement learning algorithms.
method Data-driven approach for automatic selection of bias control hyperparameters.
result Significant reduction in the number of interactions while maintaining performance.
The paper proposes a data-driven method for optimal power flow and voltage regulation in distribution grids.
problem Optimal power flow and voltage regulation in decentralized power grids.
method The approach uses a network model, historic data, and regression to find functions approximating optimal reactive power injections for inverters.
result The method achieves near-optimal results in voltage- and capacity-constrained loss minimization and voltage flattening.
Paper examines vulnerabilities in data-driven pricing schemes.
problem Vulnerability of clustering-oriented pricing schemes to malicious user behavior.
method Defined a notion of disguising to identify strategic behaviors of malicious users, characterized sensitivity zones to evaluate malicious user percentages, conducted cost benefit analysis.
result Concluded with a vulnerability analysis of data-driven pricing schemes.
We provide bounds on control learning error in stochastic systems.
problem Learning optimal controls in stochastic environments with uncontrolled parts.
method Dynamic programming and mean-field interpretation of neural networks.
result Non-asymptotic bounds on generalization error for stable overparametrised settings.
Policy certifies inventory levels meeting service requirements.
problem Maintaining stock levels meeting service requirements despite unknown demand.
method Data-driven order policy using online learning and integral action.
result Valid inference method for finite samples.
Paper proposes MA-BERT for efficient data-driven ATM models.
problem Long training time and need for large datasets in data-driven ATM models.
method Multi-Agent Bidirectional Encoder Representations from Transformers (MA-BERT) and transfer learning framework.
result MA-BERT saves training time and achieves high performance with little data.
New method shows data-driven causal studies can be misleading.
problem Misattribution of causality in data-driven earth science studies.
method Subsample-based ensemble approach for robust causality analysis.
result Transfer entropy-based causal graphs can be spurious.
Improved MCMC sampling for expensive, irregular likelihoods.
problem Bayesian inference challenges with irregular, expensive likelihoods.
method Adapt subset samplers, introduce data-driven proxies, adaptive controller.
result Improved HINTS algorithm achieves best sampling error in fixed budget.
Study optimal pricing and inventory control in dynamic settings with censored demand.
problem Optimal pricing and inventory control in dynamic settings with censored demand.
method Approximate optimal policy via high-order MDP, propose novel algorithms for solving Bellman equations.
result Established finite-sample regret bounds and demonstrated efficacy through numerical experiments.
Develops a new model for controllable and realistic traffic simulation.
problem Lack of models that offer both controllability and realism in traffic simulation.
method Guided Conditional Diffusion (CTG) model using diffusion modeling and differentiable logic.
result Improves controllability-realism tradeoff over strong baselines.
Paper derives uniform error bounds for Gaussian process regression for safer control applications.
problem Quantifying model error in Gaussian process regression for safety-critical applications.
method Employing Gaussian process distribution and continuity arguments, derive uniform error bounds under weaker assumptions.
result Derives novel uniform error bounds for Gaussian process regression under weaker assumptions.
New framework calibrates decision robustness using inverse conformal risk control.
problem Inadequate robustness levels in decision-making due to ad hoc choices.
method Constructs valid estimators to trace miscoverage-regret Pareto frontier.
result Provides distribution-free, finite-sample guarantees on robustness levels.
Perfect tracking control for real-world Euler-Lagrange systems is challenging due to uncertainties in the system model and external disturbances. The magnitude of the tracking error can be reduced either by increasing the feedback gains or improving the model of the system. The latter is clearly preferable as it allows…
The design of a reward function often poses a major practical challenge to real-world applications of reinforcement learning. Approaches such as inverse reinforcement learning attempt to overcome this challenge, but require expert demonstrations, which can be difficult or expensive to obtain in practice. We propose var…
Paper introduces a method to control early classification accuracy gaps.
problem Maintaining accuracy in early classification without full input processing.
method Statistical framework for a calibrated stopping rule.
result Reduces up to 94% of timesteps while controlling accuracy gaps.
Paper proposes method for optimal control of unknown systems with latent states.
problem Jointly estimating dynamics and latent states in systems with unmeasurable states.
method Combination of particle Markov chain Monte Carlo methods and scenario theory.
result Probabilistic performance guarantees for optimal input trajectories.
Paper uses VAEs to control IVS features for financial modeling.
problem Generating realistic IVSs with desired characteristics.
method Variational autoencoder architecture with controllable latent variables.
result Controlled generation of IVSs with specified features.
We introduce a novel data-driven order reduction method for nonlinear control systems, drawing on recent progress in machine learning and statistical dimensionality reduction. The method rests on the assumption that the nonlinear system behaves linearly when lifted into a high (or infinite) dimensional feature space wh…
This paper uses RL to optimize bid-ask spreads for illiquid corporate bonds.
problem Optimizing bid-ask spreads for illiquid corporate bonds.
method Data-driven approach using Reinforcement Learning.
result Trained RL agent's behavior shows reasonable optimal bid-ask spreads.
MAC Net improves natural language question answering with data-driven reasoning.
problem Natural Language Question Answering requires complex reasoning.
method MAC Net architecture separates memory and control for iterative reasoning.
result MAC Net achieves high efficiency and interpretability in NLP tasks.
Paper uses RL to optimize ICU load during COVID-19.
problem Optimizing ICU load during a pandemic.
method Combines epidemic model, Bayesian inference, and RL for adaptive intervention levels.
result RL policies reduce ICU burden compared to historical interventions.