Paper optimizes industrial refrigeration using adaptive exploration.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
New method solves problem using global Cartan decompositions.
Enhances quantum circuit synthesis using deep learning and geometric methods.
The paper studies time-optimal problems on specific Lie groups, describing orbits and integrals.
In this paper we study the sub-Finsler geometry as a time-optimal control problem. In particular, we consider non-smooth and non-strictly convex sub-Finsler structures associated with the Heisenberg, Grushin, and Martinet distributions. Motivated by problems in geometric group theory, we characterize extremal curves, d…
We consider in this paper the regularity problem for time-optimal trajectories of a single-input control-affine system on a n-dimensional manifold. We prove that, under generic conditions on the drift and the controlled vector field, any control u associated with an optimal trajectory is smooth out of a countable set o…
We consider control-linear left-invariant time-optimal problems on step 2 Carnot groups with strictly convex set of control parameters (in particular, sub-Finsler problems). We describe all linear-in-momenta Casimirs on the dual of the Lie algebra. In the case of rank 3 Lie groups we describe the symplectic foliation o…
This study examines abnormal geodesics in 2D-Zermelo navigation problems, revealing their role in separating time minimal and maximal curves.
We use martingale and stochastic analysis techniques to study a continuous-time optimal stopping problem, in which the decision maker uses a dynamic convex risk measure to evaluate future rewards. We also find a saddle point for an equivalent zero-sum game of control and stopping, between an agent (the "stopper") who c…
The quantum navigation problem of finding the time-optimal control Hamiltonian that transports a given initial state to a target state through quantum wind, that is, under the influence of external fields or potentials, is analysed. By lifting the problem from the state space to the space of unitary gates realising the…
Researchers describe Casimir functions for 3- and 4-step nilpotent Lie groups.
The goal of this paper is to describe Zermelo's navigation problem on Riemannian manifolds as a time-optimal control problem and give an efficient method in order to evaluate its control curvature. We will show that up to change the Riemannian metric on the manifold the control curvature of Zermelo's problem has a simp…
ARTEO algorithm optimizes safety-critical systems with uncertainty.
We consider the Zermelo navigation problem on the ellipsoid of revolution (spheroid) in the presence of a perturbation determined by a mild velocity vector field, , with application of Finsler metric of Randers type in the context of the corresponding optimal control represented by a time-efficient ship's he…
In this paper we study a sub-Finsler geometric problem on the free-nilpotent group of rank 2 and step 3. Such a group is also called Cartan group and has a natural structure of Carnot group, which we metrize considering the norm on its first layer. We adopt the point of view of time-optimal control theory…
New continuous-time optimization algorithms converge in finite time to local minima.
On the ground of origins of the theory of Lie groups and Lie algebras, their (co)adjoint representations, and the Pontryagin maximum principle for the time-optimal problem are given an independent foundation for methods of geodesic vector field to search for normal geodesics of left-invariant (sub-)Finsler metrics on L…
Optimizes trading policies using future price forecasts.
Efficient deep policy gradient method for continuous-time control problems.
Given a sequential learning algorithm and a target model, sequential machine teaching aims to find the shortest training sequence to drive the learning algorithm to the target model. We present the first principled way to find such shortest training sequences. Our key insight is to formulate sequential machine teaching…
Study optimizes trading in multiple assets with cross-effects.
Pronounced variability due to the growth of renewable energy sources, flexible loads, and distributed generation is challenging residential distribution systems. This context, motivates well fast, efficient, and robust reactive power control. Real-time optimal reactive power control is possible in theory by solving a n…
A great deal of academic and theoretical work has been dedicated to optimal liquidation of large orders these last twenty years. The optimal split of an order through time (`optimal trade scheduling') and space (`smart order routing') is of high interest \rred{to} practitioners because of the increasing complexity of t…
The stochastic knapsack has been used as a model in wide ranging applications from dynamic resource allocation to admission control in telecommunication. In recent years, a variation of the model has become a basic tool in studying problems that arise in revenue management and dynamic/flexible pricing; and it is in thi…
New method learns policies from offline data using operator models.
This paper tackles robust control of noisy systems with uncertain distributions.
We develop a theory for solving continuous time optimal stopping problems for non-linear expectations. Our motivation is to consider problems in which the stopper uses risk measures to evaluate future rewards.
Imitation learning is a control design paradigm that seeks to learn a control policy reproducing demonstrations from expert agents. By substituting expert demonstrations for optimal behaviours, the same paradigm leads to the design of control policies closely approximating the optimal state-feedback. This approach requ…
This paper addresses the model-free nonlinear optimal problem with generalized cost functional, and a data-based reinforcement learning technique is developed. It is known that the nonlinear optimal control problem relies on the solution of the Hamilton-Jacobi-Bellman (HJB) equation, which is a nonlinear partial differ…
The paper is devoted to the local classification of generic control-affine systems on an n-dimensional manifold with scalar input for any n>3 or with two inputs for n=4 and n=5, up to state-feedback transformations, preserving the affine structure. First using the Poincare series of moduli numbers we introduce the intr…
In this article we propose a Weighted Stochastic Mesh (WSM) Algorithm for approximating the value of a discrete and continuous time optimal stopping problem. We prove that in the discrete case the WSM algorithm leads to semi-tractability of the corresponding optimal problems in the sense that its complexity is bounded …
AuON is a linear-time optimizer that improves upon Muon's performance without approximate orthogonal matrices.
Paper approximates Kelly betting for wealth growth.
We consider remodeling the planar search patterns, in the presence of the river-type perturbation represented by the weak vector field, basing on the time-optimal paths as Finslerian solutions to the Zermelo navigation problem via Randers metric.
This paper develops a framework for training and evaluating neural networks for MPC.
In this paper, we examine higher order difference problems. Using the "squeezing" argument, we derive both Euler's condition and the transversality condition. In order to derive the two conditions, two needed assumptions are identified. A counterexample, in which the transversality condition is not satisfied without th…
Researchers find optimal paths on a specific geometric group.
Continuous-time optimal stopping solved with deep reinforcement learning
Paper proposes new -regret measure for non-episodic RL.
A new RL method handles uncertainty and constraints in real-time optimization.
Consider power utility maximization of terminal wealth in a 1-dimensional continuous-time exponential Levy model with finite time horizon. We discretize the model by restricting portfolio adjustments to an equidistant discrete time grid. Under minimal assumptions we prove convergence of the optimal discrete-time strate…
The paper studies problem of continuous time optimal portfolio selection for a incom- plete market diffusion model. It is shown that, under some mild conditions, near optimal strategies for investors with different performance criteria can be constructed using a limited number of fixed processes (mutual funds), for a m…
Modeling market dynamics with informed and uninformed traders and fads.
Optimizes portfolio growth rate for a behavioral investor considering terminal relative growth rate.
Researchers found optimal paths on a specific geometric group.
Solves time-optimal navigation on slippery slopes with cross gravitational wind.
Paper proposes efficient image inversion and editing using rectified stochastic differential equations.
Study uses RL to optimize investment with financial constraints, showing exploration benefits.