We provide a dynamic programming principle for stochastic optimal control problems with expectation constraints. A weak formulation, using test functions and a probabilistic relaxation of the constraint, avoids restrictions related to a measurable selection but still implies the Hamilton-Jacobi-Bellman equation in the …
New method optimizes pumped hydroelectric storage with state constraints.
problem Optimal management of pumped hydroelectric production with state constraints.
method Transformed constrained problem into an unconstrained one in augmented spaces with state constraints penalized.
result Solved the problem using dynamic programming.
Counterexample shows state-constrained optimal control problems can have Young measure gaps.
problem Existence of Young measure gaps in state-constrained optimal control problems.
method Provided a counterexample for smooth controllable systems state-constrained to the unit ball.
result Gap occurs in a regular setting with non-convex Lagrangian density.
Global constraints and reranking have not been used in cognates detection research to date. We propose methods for using global constraints by performing rescoring of the score matrices produced by state of the art cognates detection systems. Using global constraints to perform rescoring is complementary to state of th…
Study leader-follower games with terminal state constraints using McKean-Vlasov SDEs.
problem Leader-follower games with terminal state constraints.
method Linear McKean-Vlasov forward-backward SDEs, existence and uniqueness results, convergence results.
result Existence and uniqueness of solutions for leader-follower games with constraints.
The paper solves a consumption-investment problem with state-dependent lower bounds.
problem A life-time consumption-investment problem with a state-dependent lower bound on consumption.
method Transformed the problem into a state-independent control problem to apply standard theory.
result Explicit optimal strategies provided for both homogeneous and non-homogeneous constraints.
This paper proposes a method to safely adjust exploration in RL to satisfy constraints.
problem Unsafe exploration in reinforcement learning violates constraints on controlled object states.
method Automatic adjustment of exploration inputs and variance-covariance matrix for safety.
result The method guarantees satisfaction of joint chance constraints with specified probability.
Formal constraints improve RL safety in complex environments.
problem Safety constraints in reinforcement learning for complex environments.
method Specify constraints in formal languages, instantiate as finite automata, augment MDP states, learn dense cost function.
result Improved safety in training RL algorithms over various constraints.
New method adds user constraints to Markov chains for better data reduction.
problem No systematic framework to impose user-defined constraints on Markov chains.
method Path entropy maximization to derive transition probabilities with user constraints.
result Improved nonlinear dimensionality reduction with user-prescribed constraints.
Surveying nonparametric inference with shape constraints, past and future.
problem Statistical inference under shape constraints.
method Historical overview and future directions.
result Outlook on future research directions.
Study proves steady state space hypersurfaces are hyperplanes under certain curvature constraints.
problem Characterizing complete spacelike hypersurfaces in steady state space.
method Extended Omori-Yau's maximum principle.
result Proves complete spacelike hypersurfaces are hyperplanes under specific curvature conditions.
Solves optimal control with state constraints using probabilistic methods.
problem Optimal control of diffusion processes within state constraints.
method Probabilistic representation and optimal control under mild conditions.
result Explicit formulae for optimally controlled dynamics in examples.
State-augmented algorithm optimizes wireless network resource management.
problem Optimizing resource allocation in multi-user wireless networks.
method Proposes a state-augmented algorithm using dual variables.
result Feasible and near-optimal resource decisions achieved.
Paper tackles SMPC for linear systems with unknown noise distribution.
problem Stochastic MPC for linear systems with chance state constraints and unknown noise distribution.
method Reformulate chance constraints, design robust benchmark SMPC, and develop adaptive SMPC with online noise statistics learning.
result Adaptive SMPC guarantees time-uniform satisfaction of unknown reformulated state constraints with high probability.
Study optimal consumption with relaxed benchmarks and drawdown constraints.
problem Optimal consumption under relaxed benchmark tracking and consumption drawdown constraint.
method Transformed stochastic control problem into regular control problem with state-control constraints, then solved using dual transform and optimal consumption behavior.
result Closed-form solution for optimal investment and consumption in feedback form.
Extends reinforcement learning to continuous state spaces with safety constraints.
problem Safety-critical reinforcement learning in continuous state spaces with unknown dynamics.
method Introduces a novel Budgeted Bellman Optimality operator and applies it to continuous state spaces.
result Validated on spoken dialogue and autonomous driving applications.
Paper derives constraints for Bayesian Knowledge Tracing parameters.
problem Issues with EM algorithm in BKT parameter estimation.
method From first principles, derives constraints on BKT parameter space.
result Novel algorithm respects derived constraints for parameter estimation.
Unified framework for integrating linear constraints in time series forecasting.
problem Challenges in traditional time series forecasting algorithms.
method Unified framework combining linear constraints in time series forecasting.
result Exact minimizer of the constrained empirical risk can be computed efficiently using linear algebra.
Proposes a constraint for deep clustering to handle both simple and complex topologies.
problem Limited prior knowledge for deep clustering methods to perform well on complex topologies.
method Introduces a constraint using symmetric InfoNCE to enhance deep clustering performance.
result The constraint improves deep clustering methods' performance on both simple and complex topologies.
Solves optimal control with constraints for stochastic systems.
problem Optimal control of constrained stochastic linear-quadratic systems.
method State separation theorem and Riccati equations for explicit solution.
result Explicit piecewise affine optimal control policy.
Holistic GLMs add constraints for better model quality.
problem Improving classical linear regression models.
method Sparsity-inducing, sign-coherence, and linear constraints.
result Holistic GLMs reliably solve GLMs for various responses.
The paper extends utility maximization by integrating partial information and robust VaR constraints.
problem Optimal investment under partial information and robust VaR-type constraints.
method Combines partial information and robust regulatory constraints (VaR) to solve the utility maximization problem.
result Optimal wealth is a decreasing function of state price density, and depends on the overall evolution of the estimated market price of risk.
Neural SDEs model suicide risk with compact state space constraints.
problem Modeling suicide risk with irregular, noisy, and partially observed data.
method Developed neural SDEs confined to compact state spaces, addressing domain constraints and numerical stability.
result Improved forecasts and optimization dynamics over standard models on EMA datasets.
We consider n risk-averse agents who compete for liquidity in an Almgren--Chriss market impact model. Mathematically, this situation can be described by a Nash equilibrium for a certain linear-quadratic differential game with state constraints. The state constraints enter the problem as terminal boundary conditions f…
Efficient learning-based MPC for unknown nonlinear systems with state constraints.
problem Control of discrete-time nonlinear systems with unknown dynamics and state constraints.
method Receding horizon reinforcement learning (r-LPC) using Koopman operator-based prediction model.
result Proven closed-loop recursive feasibility, robustness, and asymptotic stability under function approximation errors.
A multi-task GP model tracks time-varying transition probabilities between two states.
problem Tracking time-varying transition probabilities between 'moves' and 'pauses' states.
method Kernel-based multi-task Gaussian Process model with time-variability and constraints.
result Enforces constraints while learning transition probabilities.
The paper shows how Lagrangian duality improves deep learning for constrained problems.
problem Learning optimization problems with complex constraints in science and engineering.
method Lagrangian duality applied to deep learning models.
result Lagrangian duality brings significant benefits for constrained learning tasks.
We study the projected gradient descent method on low-rank matrix problems with a strongly convex objective. We use the Burer-Monteiro factorization approach to implicitly enforce low-rankness; such factorization introduces non-convexity in the objective. We focus on constraint sets that include both positive semi-defi…
Iterative method learns unknown constraints for MPC control.
problem Learning to satisfy unknown polyhedral state constraints in iterative MPC.
method Collects and improves estimates of unknown constraints using collected data, designs an MPC controller to satisfy the estimated constraints.
result Robust and probabilistic guarantees of constraint satisfaction as a function of task iterations.
The paper introduces an adjacency constraint to improve goal-conditioned HRL.
problem Training inefficiency in goal-conditioned HRL due to large action space.
method Restricting the high-level action space to a k-step adjacent region of the current state.
result The adjacency constraint preserves optimal hierarchical policies and improves HRL performance.
Paper finds unique viscosity solution to complex control problems.
problem Complex stochastic control problems with singular terminal state constraints.
method Establishes existence of unique nonnegative continuous viscosity solution using novel comparison principle.
result Unique viscosity solution to HJB equation for linear-quadratic control problems.
Study relaxes boundedness constraints in Ramsey consumption problem.
problem Optimal consumption in the stochastic Ramsey problem without boundedness constraints.
method Non-standard stochastic differential equation, probabilistic arguments, viscosity solutions.
result Value function is the unique classical solution to a nonlinear elliptic equation, leading to optimality of feedback consumption process.
Improves GATs by adding margin-based constraints to prevent over-fitting and over-smoothing.
problem Over-fitting and over-smoothing in GATs.
method Margin-based constraints on attention weights and graph structure.
result Significant improvements over previous GATs on various datasets.
HardCoRe-NAS finds fitting neural networks adhering to hard resource constraints.
problem Finding fitting neural networks that adhere to hard resource constraints.
method Accurate formulation of resource requirement and scalable search method.
result HardCoRe-NAS generates state-of-the-art architectures strictly satisfying hard resource constraints.
Study capacity constraints in continual learning with a simple model.
problem Understanding optimal resource allocation for agents with limited memory and compute resources.
method Analyzes a capacity-constrained linear-quadratic-Gaussian (LQG) sequential prediction problem and demonstrates optimal capacity allocation strategies.
result Derives a solution to the capacity-constrained LQG sequential prediction problem and shows how to optimally allocate capacity across sub-problems in the steady state.
A new ML method teaches constraints directly to models.
problem Addressing safety and fairness in AI systems.
method Directly teaching constraint satisfaction to ML models using a constraint solver.
result Empirically, our approach performs well on fairness and synthetic constraints.
Proposes a new method for GNNs that avoids iterative node state convergence.
problem Iterative computation of node states in GNNs is inefficient and requires many epochs.
method Constrained optimization in the Lagrangian framework to learn transition function and node states simultaneously.
result The proposed method compares favorably with existing models on various benchmarks.
This paper studies the problem of optimal investment with CRRA (constant, relative risk aversion) preferences, subject to dynamic risk constraints on trading strategies. The market model considered is continuous in time and incomplete. the prices of financial assets are modeled by Itô processes. The dynamic risk constr…
Stabilized neural differential equations enforce constraints on dynamical systems.
problem Ensuring dynamical systems preserve known constraints like conservation laws.
method SNDEs with a stabilization term to enforce manifold constraints.
result SNDEs outperform existing methods and broaden constraint types.
Algorithm reduces episode count for CMDPs with constraints.
problem Online decision-making with constraints in episodic CMDPs.
method Optimistic planning using linear programming for PAC guarantee.
result Probably approximately correct (PAC) guarantee on episode count.
Paper proposes a new framework for semi-supervised learning.
problem Limited annotated data in supervised learning.
method Fuzzy domain constraint-based framework for semi-supervised learning.
result Enhances model quality for semi-supervised learning.
The paper introduces MU for NMF with β-divergences and disjoint constraints.
problem Nonnegative matrix factorization with constraints.
method Design multiplicative updates for NMF based on β-divergences with disjoint constraints. result Multiplicative updates satisfy constraints and decrease the objective function.
This work proposes an online learning approach to tighten constraints in stochastic control problems.
problem Solving chance-constrained stochastic optimal control problems is computationally challenging.
method Reformulate chance constraints as a binary regression problem and use a GP model to learn constraint-tightening parameters online.
result The approach tightens constraints more effectively, leading to lower costs in numerical experiments.
Declarative entropy constraints improve semi-supervised learning.
problem Improving semi-supervised learning performance.
method Declarative specification of entropy constraints for semi-supervised learning.
result Consistent improvements on SSL benchmarks, including a new state-of-the-art result.
We solve the problem of optimal stopping of a Brownian motion subject to the constraint that the stopping time's distribution is a given measure consisting of finitely-many atoms. In particular, we show that this problem can be converted to a finite sequence of state-constrained optimal control problems with additional…
Optimizes wireless network resource management with state-augmented policies.
problem Optimizing network-wide utility with user performance constraints.
method State-augmented parameterization of RRM policy, using dual variables.
result Superior trade-off between mean, minimum, and 5th percentile rates.
IPO optimizes reinforcement learning with constraints for better performance.
problem Maximizing long-term reward while satisfying cumulative constraints in decision problems.
method Interior-point Policy Optimization (IPO) using logarithmic barrier functions.
result IPO outperforms state-of-the-art baselines in reward maximization and constraint satisfaction.
New RL algorithm learns good actions from offline data, reducing uncertainty and divergence.
problem Limited applicability of current RL algorithms in real-world settings due to high costs of exploration.
method Proposes an algorithm for batch RL using a fixed offline dataset, with penalties for policy and value constraints.
result Compared favorably to state-of-the-art methods on 32 continuous-action benchmarks.