Proposes a deep reinforcement learning model for efficient variable speed limits control.
problem Improving traffic flow, safety, and emissions on freeways with varying speed limits.
method Uses a novel actor-critic architecture for deep reinforcement learning to manage dynamic speed limits.
result The proposed method enhances efficiency, safety, and emissions compared to traditional control methods.
Study develops a machine learning-based ramp metering model to improve freeway efficiency.
problem Improving ramp metering to maintain freeway efficiency under various traffic conditions.
method Machine learning approach using historical data to predict and manage traffic flow.
result The novel model outperforms a baseline traffic-responsive ramp metering algorithm.
Dynamic linear models improve travel time prediction for congested freeways.
problem Accurate travel time prediction for congested freeways.
method Dynamic linear models (DLMs) with time-varying parameters.
result Significant improvements in travel time prediction accuracy, especially for short-term predictions.
SECRM-2D improves RL-based autonomous driving with safety guarantees.
problem Safety and efficiency trade-offs in RL-based autonomous driving.
method RL-based controller with safety constraints for efficient and comfortable driving.
result SECRM-2D avoids crashes and improves efficiency and comfort compared to baselines.
New PRGP model improves traffic flow estimation.
problem Lack of models combining physics and ML for traffic flow.
method Physics regularized Gaussian process (PRGP) with discrete formulations.
result PRGP model outperforms calibrated physics models and ML methods.
CitySim dataset captures vehicle trajectories for safety research.
problem Lack of fine-grain vehicle trajectories for safety-oriented research.
method Five-step procedure: video stabilization, object filtering, stitching, detection, and error filtering.
result CitySim dataset improves safety evaluations and facilitates digital-twin research.
Exploration is a fundamental aspect of Reinforcement Learning, typically implemented using stochastic action-selection. Exploration, however, can be more efficient if directed toward gaining new world knowledge. Visit-counters have been proven useful both in practice and in theory for directed exploration. However, a m…
Deep learning classifies land use from high-resolution aerial imagery.
problem Variations in land features in aerial imagery due to sensor settings and context.
method Used deep convolutional neural networks to classify land use from VHR orthophoto mosaics.
result Deep learning can accurately classify land use from high-resolution visible band multispectral imagery.
DistPre predicts traffic speeds efficiently for large networks.
problem Fine-grained, accurate speed prediction for large-scale transportation networks.
method Customizes LSTM models on a cluster, sharing trained models between detectors.
result Efficient and accurate fine-grained traffic-speed prediction.
New model improves traffic flow predictions with physics and machine learning.
problem Inaccurate predictions in traffic flow modeling with small or noisy datasets.
method Developed a physics regularized Gaussian process (PRGP) model to encode physical models into ML architecture and regularize the training process.
result The PRGP model outperforms previous methods in estimation precision and input robustness.
Predicts morning traffic congestion using social media data from the previous evening.
problem Challenges in predicting early morning traffic dynamics.
method Mining Twitter messages to understand evening/midnight work and rest patterns.
result People's tweeting patterns before the morning commute are associated with traffic congestion.
Neural PID controllers improve control system performance and are more interpretable.
problem Lack of interpretability in neural PID controllers limits their use in control engineering.
method Extensive study using four benchmark systems with and without noise and disturbances, applying GDNN to PID controllers.
result Neural PID controllers outperform standard PID and model-based control in most tasks.
Boosting improves control of complex systems.
problem Improving performance of controllers for dynamical systems.
method Proposes a boosting framework for online control of dynamical systems.
result An efficient boosting algorithm that combines weak controllers into a more accurate one.
Derives optimal control conditions using calculus of variations.
problem Optimizing Markov control in stochastic control problems.
method Calculus of variations approach to derive necessary conditions.
result Solves the Merton portfolio optimization problem.
Framework simplifies vision-based control and goal discovery.
problem Learning proportional control from visual data.
method Introduces NewtonianVAE for proportional control and goal discovery.
result Dramatic simplification and acceleration of vision-based controllers.
The paper introduces new tests for global controllability in hybrid systems.
problem Global controllability in hybrid systems with discrete events.
method Geometric formulation of hybrid systems and analysis of jump points.
result Hybrid systems can be globally controllable even if continuous systems are not.
This work tackles force control for contact-rich manipulation tasks with rigid robots using RL.
problem Challenges in working with real robotic hardware, especially position-controlled robots.
method Combines RL with traditional force control techniques, implementing parallel position/force control and admittance control.
result Validated methods on both simulation and real robot (UR3 e-series) for force control.
Paper studies constrained control games with a novel approximation method.
problem Games with constrained control directions.
method Approximation procedure based on L1-stability estimates and almost sure convergence. result Existence of game's value and optimal strategy for the stopper.
RL applied to TCLs for power consumption control.
problem Optimizing power consumption using TCLs with RL.
method Modelica-based reinforcement learning (Q-learning) for stochastic TCLs.
result Q-learning parameters affect controller performance.
A framework integrates machine learning with robust control for safer, more reliable systems.
problem Combining machine learning with robust control for systems with stringent safety and reliability requirements.
method Integrates Gaussian Process Regression and state-of-the-art robust controller synthesis within a framework that provides rigorous guarantees.
result Demonstrated improved performance with more data while maintaining rigorous guarantees.
Paper optimizes energy-based controller for swinging up a pendulum using entropy search.
problem Finding optimal parameters for energy-based controllers is hard.
method Bayesian optimization (Entropy Search) applied to energy-based controller design.
result Optimal controller improves performance of a swinging pendulum.
Paper uses deep reinforcement learning for better control of rocket engines during start-up phases.
problem Lack of optimal control during transient phases of liquid rocket engines.
method Deep reinforcement learning approach for optimal control of a gas-generator engine's continuous start-up phase.
result Deep reinforcement learning controller achieves highest performance and minimal computational effort.
Neural ODEs control graph dynamics with low energy feedback.
problem Controlling complex dynamical systems on graphs.
method Neural Ordinary Differential Equation Control (NODEC) framework.
result NODEC learns low-energy control signals for graph dynamical systems.
Just as an explicit parameterisation of system dynamics by state, i.e., a choice of coordinates, can impede the identification of general structure, so it is too with an explicit parameterisation of system dynamics by control. However, such explicit and fixed parameterisation by control is commonplace in control theory…
Unified control theory and machine learning for safety in uncertain systems.
problem Safety guarantees for systems with measurement model uncertainty.
method Measurement-Robust Control Barrier Functions (MR-CBFs) for control synthesis.
result MR-CBFs ensure safety in perception systems with measurement model uncertainty.
Paper studies optimal control for a specific geometric problem.
problem Optimal control problem associated with the Paneitz obstacle problem.
method Existence and regularity results for optimal controls.
result Existence of optimal controls and their properties.
Optimizes dividend policies in a Brownian model with controlled rates.
problem Realistic optimal dividend policies in a stochastic control problem.
method Delayed linear control strategies for refracted diffusion processes.
result Optimality of delayed linear control strategies for dividend payments.
New method uses neural nets to control systems safely with disturbances.
problem Designing safe control laws for systems with disturbances.
method Imitation learning to train neural network controllers that satisfy CBF constraints.
result Demonstrated on a unicycle model with external disturbances.
Paper proposes a new method to optimize robot body structure and control policy.
problem Optimizing robot body structure and control policy in a coupled manner.
method Revisits co-design problem as a Stackelberg game, incorporating control adaptation dynamics.
result Stackelberg PPO outperforms standard PPO in stability and performance.
Neural network HDP improves virtual inertia control for non-inductive grids.
problem Traditional virtual inertia controllers are not suitable for non-inductive grids.
method Adaptive neural network heuristic dynamic programming (HDP) for optimal control.
result The proposed HDP controller outperforms traditional controllers in virtual inertia control.
Non-bilinear observations make optimal control harder, showing non-convex costs and non-affine optimal controllers.
problem Optimal control from bilinear observations in linear systems is challenging.
method Analytical and numerical methods to study the non-convex cost-to-go and non-affine optimal controllers.
result The Separation Principle does not hold for bilinear observations, leading to non-convex costs and non-affine optimal controllers.
Defense strategy improves controller robustness against adversarial attacks.
problem Adversarial attacks on learning-enabled controllers in CPS.
method Two-stage defense strategy treating controller and environment as black-boxes with unknown dynamics.
result Defense strategy effectively improves controller robustness in realistic control domains.
This paper considers control systems defined on Lie algebroids. After deriving basic controllability tests for general control systems, we specialize our discussion to the class of mechanical control systems on Lie algebroids. This class of systems includes mechanical systems subject to holonomic and nonholonomic const…
Survey of theoretical foundations for policy optimization in control.
problem Understanding the theoretical properties of gradient-based methods in control and reinforcement learning.
method Interdisciplinary review of optimization landscape, convergence, and sample complexity for various control problems.
result Recent theoretical results on stability and robustness in learning-based control.
Neuro-inspired RL solves complex control problems with fewer controllers.
problem Solving nonlinear control problems with unknown dynamics efficiently.
method Hierarchical RL framework combining limb coordination and reinforcement learning.
result Local LQR controllers combined with a reinforcement learner solve global nonlinear problems.
Survey combines FL and control for better adaptability and privacy.
problem Combining FL and control for better adaptability and privacy.
method Combining Federated Learning (FL) and control methods.
result Combining FL and control enhances adaptability, scalability, generalization, and privacy.
New method for handling multi-dimensional singular controls with jump costs in mean-field problems.
problem Handling jump costs in multi-dimensional singular controls.
method Introducing two-layer parametrisations to interpolate jumps on both distributional and pathwise levels.
result Derivation of a DPP and characterisation of the value function as a minimal super-solution to a quasi-variational inequality.
Researchers use Meyer-σ-fields to model information flow in irreversible investment problems.
problem Modeling information flows in stochastic control problems with jumps.
method Using Meyer-σ-fields as a tool to model information flow.
result Different signals on exogenous jumps lead to different optimal controls.
Researchers develop a method to control nonlinear systems with Koopman operator regression.
problem Controlling nonlinear systems with finite action spaces.
method Koopman operator regression for dynamics estimation and model predictive control for control.
result The method yields a linear switching predictive model for control.
A new Q-learning controller improves line follower robot control.
problem Challenges in controlling line follower robots due to unknown mechanical characteristics and uncertainties.
method Simulated annealing based Q learning method to address controller performance issues.
result The proposed controller outperforms conventional P controllers in line follower robots.
Anticipatory model generates music with control over events.
problem Controlling symbolic music generation.
method Interleaving event and control sequences to predict future events.
result Anticipatory model matches autoregressive models in performance and can infill control tasks.
Reinforcement learning controls complex, black-boxed systems like greenhouses.
problem Controlling nonlinear, complex, and black-boxed systems.
method Actor-critic reinforcement learning approach.
result Successfully maintained greenhouse conditions for 20 times longer than other methods.
Meta-learning control algorithm with finite-time guarantees for unknown systems.
problem Online control of unknown linear systems with constraints.
method Provable regret guarantees for an iterative control algorithm.
result Regret bounds of O(T3/4) for controller cost and constraint violation. New algorithm achieves logarithmic regret for adversarial online control.
problem Online linear-quadratic control in systems with adversarial disturbances.
method Characterization of optimal offline control law, reduced to online learning with approximate advantage functions.
result First algorithm with logarithmic regret for arbitrary adversarial disturbance sequences.
Designs adaptive controller for networked control systems with wireless data transmission.
problem Adaptive control in networked systems with unreliable wireless channels.
method Upper Confidence Bounds for Networked Control Systems (UCB-NCS) learning rule.
result Non-asymptotic performance guarantees with a regret bound of O(C√T).
New robust control method for uncertain systems using bootstrapped noise.
problem Designing controllers robust to model uncertainties in finite data.
method Least-squares model estimator, bootstrap resampling, multiplicative noise LQR.
result Significantly outperforms certainty equivalent controllers in numerical tests.
Since governments give stimulus to firms and expect the spillover effect by fiscal policies, it is important to know the effectiveness that they can control the economy. To clarify the controllability of the economy, we investigate a firm production network observed exhaustively in Japan and what firms should be direct…
New control strategy minimizes infected individuals in SIS epidemics.
problem Developing effective control strategies for SIS epidemics.
method Stochastic optimal control of SDEs with jumps, using treatment intensities.
result Control strategy consistently outperforms alternatives in synthetic data.