Article presents QR and LQ decomposition algorithms for various matrix sizes and ranks.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
New theory extends LQ control to non-exponential discount scenarios.
A new method solves complex control problems with random coefficients.
We solve a complex trade execution problem by simplifying it into a known LQ control problem.
RL solves discrete LQ control with Gaussian optimal policy.
Paper solves MV portfolio selection in jump-diffusion models with no-shorting constraint.
Study of LQ MFGs in infinite-dimensional Hilbert spaces.
Developed LQ MFG theory with common noise, proving existence and uniqueness.
We study the global convergence of policy optimization for finding the Nash equilibria (NE) in zero-sum linear quadratic (LQ) games. To this end, we first investigate the landscape of LQ games, viewing it as a nonconvex-nonconcave saddle-point problem in the policy space. Specifically, we show that despite its nonconve…
We consider the exploration-exploitation tradeoff in linear quadratic (LQ) control problems, where the state dynamics is linear and the cost function is quadratic in states and controls. We analyze the regret of Thompson sampling (TS) (a.k.a. posterior-sampling for reinforcement learning) in the frequentist setting, i.…
Neural operators learn to solve LQ MFGs efficiently in infinite dimensions.
In this paper we introduce the notion of cofrontal mappings, as the dual objects to frontal mappings, and study their basic local and global properties. Cofrontals are very special mappings and far from generic nor stable except for the case of submersions. It is observed that any smooth mapping can be -approximat…
Regret analysis is challenging in Multi-Agent Reinforcement Learning (MARL) primarily due to the dynamical environments and the decentralized information among agents. We attempt to solve this challenge in the context of decentralized learning in multi-agent linear-quadratic (LQ) dynamical systems. We begin with a simp…
We study the problem of adaptive control of a high dimensional linear quadratic (LQ) system. Previous work established the asymptotic convergence to an optimal controller for various adaptive control schemes. More recently, for the average cost LQ problem, a regret bound of was shown, apart form logarit…
Choquet regularization improves exploration in RL.
The paper solves a complex control problem with stochastic elements and switching conditions.
This work optimizes RL algorithms using entropy regularisation for continuous-time LQ problems.
Study on PG learning for LQ MFC problems with common noise, proving convergence and sample complexity.
We study the performance of the certainty equivalent controller on Linear Quadratic (LQ) control problems with unknown transition dynamics. We show that for both the fully and partially observed settings, the sub-optimality gap between the cost incurred by playing the certainty equivalent controller on the true system …
We study in this paper a class of constrained linear-quadratic (LQ) optimal control problem formulations for the scalar-state stochastic system with multiplicative noise, which has various applications, especially in the financial risk management. The linear constraint on both the control and state variables considered…
Development systems for deep learning (DL), such as Theano, Torch, TensorFlow, or MXNet, are easy-to-use tools for creating complex neural network models. Since gradient computations are automatically baked in, and execution is mapped to high performance hardware, these models can be trained end-to-end on large amounts…
We discuss two generalizations of the collar lemma. The first is the stable neighborhood theorem which says that a (not necessarily simple) closed geodesic in a hyperbolic surface has a \lq\lq stable neighborhood\rq\rq whose width only depends on the length of the geodesic. As an application, we show that there is a lo…
Model-free approaches for reinforcement learning (RL) and continuous control find policies based only on past states and rewards, without fitting a model of the system dynamics. They are appealing as they are general purpose and easy to implement; however, they also come with fewer theoretical guarantees than model-bas…
Over the moduli space of rank semi-stable lattices is a universal family of tori. Along the fibers, there are natural differential operators and differential equations, particularly, the heat equations and the Fokker-Planck equations in statistical mechanics. In this paper, we explain why, by taking averages over t…
New algorithms improve blind source separation for linear-quadratic mixtures.
This paper develops q-learning methods for mean-field control problems.
Post-training optimizes model performance beyond base model limits.
Motivated by the study of linear quadratic optimal control problems, we consider a dynamical system with a constant, quadratic Hamiltonian, and we characterize the number of conjugate times in terms of the spectrum of the Hamiltonian vector field . We prove the following dichotomy: the number of conjugate time…
Unitons, i.e.\ harmonic spheres in a unitary group, correspond to \lq uniton bundles\rq, i.e.\ holomorphic bundles over the compactified tangent space to the complex line with certain triviality and other properties. In this paper, we use a monad representation similar to Donaldson's representation of instanton bundles…
This paper formulates and studies a stochastic maximum principle for forward-backward stochastic Volterra integral equations (FBSVIEs in short), while the control area is assumed to be convex. Then a linear quadratic (LQ in short) problem for backward stochastic Volterra integral equations (BSVIEs in short) is present …
New formula for 3-manifold invariants using combinatorial methods.
This paper is devoted to study the effects arising from imposing a value-at-risk (VaR) constraint in mean-variance portfolio selection problem for an investor who receives a stochastic cash flow which he/she must then invest in a continuous-time financial market. For simplicity, we assume that there is only one investm…
This paper studies a class of continuous-time scalar-state stochastic Linear-Quadratic (LQ) optimal control problem with the linear control constraints. Applying the state separation theorem induced from its special structure, we develop the explicit solution for this class of problem. The revealed optimal control poli…
This paper examines the problem of learning with a finite and possibly large set of p base kernels. It presents a theoretical and empirical analysis of an approach addressing this problem based on ensembles of kernel predictors. This includes novel theoretical guarantees based on the Rademacher complexity of the corres…
In this paper, we continue our study on a general time-inconsistent stochastic linear--quadratic (LQ) control problem originally formulated in [6]. We derive a necessary and sufficient condition for equilibrium controls via a flow of forward--backward stochastic differential equations. When the state is one dimensional…
This paper is concerned with an optimal reinsurance and investment problem for an insurance firm under the criterion of mean-variance. The driving Brownian motion and the rate in return of the risky asset price dynamic equation cannot be directly observed. And the short-selling of stocks is prohibited. The problem is f…
Study evaluates various regularization methods for electricity price forecasting.
The paper solves TIC LQ control problems using stochastic differential games.
Study of geometric analysis on asymmetric metric spaces, including heat flow and Sobolev spaces.
In this paper, we formulate a general time-inconsistent stochastic linear--quadratic (LQ) control problem. The time-inconsistency arises from the presence of a quadratic term of the expected state as well as a state-dependent term in the objective functional. We define an equilibrium, instead of optimal, solution withi…
Stabilization of linear systems with unknown dynamics is a canonical problem in adaptive control. Since the lack of knowledge of system parameters can cause it to become destabilized, an adaptive stabilization procedure is needed prior to regulation. Therefore, the adaptive stabilization needs to be completed in finite…
Invariants count inflections and vertices in singular plane curves.
Paper proposes a RL approach for ALM with superior performance.
We develop robust Markov Decision Processes with risk measures for uncertain environments.
Researchers describe and compare decompositions of Poincaré duality pairs.
The paper proposes and discusses semiorthogonal decompositions for moduli spaces of vector bundles.
The paper classifies decompositions of 3-sphere and lens spaces with handlebodies.
We combine aspects of the notions of finite decomposition complexity and asymptotic property C into a notion that we call finite APC-decomposition complexity. Any space with finite decomposition complexity has finite APC-decomposition complexity and any space with asymptotic property C has finite APC-decomposition comp…