The paper solves a control problem using reflections to track a benchmark process.
problem Optimal consumption with a benchmark process that grows over time.
method Introduced two auxiliary state processes with reflections to transform the problem into a more tractable form.
result Established the existence of a unique classical solution to the dual PDE.
Study of multidimensional control problems with reflection controls.
problem Solving control problems with reflection controls in multidimensional settings.
method Gradient descent algorithm for polytope approximations, data-driven domain estimator, episodic learning algorithm.
result Data-driven solutions for unknown diffusion dynamics with sublinear regret.
In this paper, we present a family of a control-stopping games which arise naturally in equilibrium-based models of market microstructure, as well as in other models with strategic buyers and sellers. A distinctive feature of this family of games is the fact that the agents do not have any exogenously given fundamental…
Study of billiards in sub-Finsler geometry, including unusual orbits.
problem Exploring billiard dynamics in sub-Finsler spaces.
method Symplectic and variational approaches, control theory.
result Unusual orbits like gliding and creeping orbits exist.
The paper extends Merton's problem by adding benchmark tracking, finding optimal strategies.
problem Maximizing consumption utility with a trade-off against benchmark performance.
method Developed a convex duality theorem and derived optimal strategies for specific cases.
result Found optimal portfolio and consumption strategies for CRRA utility and geometric Brownian motion benchmarks.
We study the optimal dividend problem for a firm's manager who has partial information on the profitability of the firm. The problem is formulated as one of singular stochastic control with partial information on the drift of the underlying process and with absorption. In the Markovian formulation, we have a 2-dimensio…
Study optimal consumption with relaxed benchmarks and drawdown constraints.
problem Optimal consumption under relaxed benchmark tracking and consumption drawdown constraint.
method Transformed stochastic control problem into regular control problem with state-control constraints, then solved using dual transform and optimal consumption behavior.
result Closed-form solution for optimal investment and consumption in feedback form.
We solve explicitly a two-dimensional singular control problem of finite fuel type for infinite time horizon. The problem stems from the optimal liquidation of an asset position in a financial market with multiplicative and transient price impact. Liquidity is stochastic in that the volume effect process, which determi…
This paper optimizes tracking portfolios in incomplete markets using reinforcement learning.
problem Optimizing tracking portfolios in incomplete markets with capital injection.
method Reinforcement learning approach for optimal control in reflected diffusion processes.
result Satisfactory performance of the q-learning algorithm in numerical examples.
We characterise the value function of the optimal dividend problem with a finite time horizon as the unique classical solution of a suitable Hamilton-Jacobi-Bellman equation. The optimal dividend strategy is realised by a Skorokhod reflection of the fund's value at a time-dependent optimal boundary. Our results are obt…
FinReflectKG - EvalBench benchmarks financial KG extraction from SEC 10-K filings.
problem Lack of universal benchmark and evaluation framework for financial KG construction.
method Agentic and holistic evaluation principles, deterministic commit-then-justify judging protocol, binary and ordinal evaluations.
result Reflection-based extraction outperforms single-pass extraction in comprehensiveness, precision, and relevance.
This paper studies the bail-out optimal dividend problem with regime switching under the constraint that the cumulative dividend strategy is absolutely continuous. We confirm the optimality of the regime-modulated refraction-reflection strategy when the underlying risk model follows a general spectrally negative Markov…
We consider controller-stopper problems in which the controlled processes can have jumps. The global filtration is represented by the Brownian filtration, enlarged by the filtration generated by the jump process. We assume that there exists a conditional probability density function for the jump times and marks given t…
Learning to control an environment without hand-crafted rewards or expert data remains challenging and is at the frontier of reinforcement learning research. We present an unsupervised learning algorithm to train agents to achieve perceptually-specified goals using only a stream of observations and actions. Our agent s…
The paper monitors stock market relationships using network analysis and statistical control charts.
problem Detecting abnormal changes in the financial market network structure.
method Network construction using distance methods, hierarchical clustering, and Shewhart control charts.
result Abnormal changes in financial market relationships can be detected using statistical process control.
Optimal retirement timing and consumption under shortfall risk management
problem Optimal portfolio, consumption, and endogenous early retirement problem
method Maximizing expected lifetime consumption utility while managing the maximum wealth shortfall relative to a benchmark
result Geometric structure of the stopping set and feedback-form optimal retirement boundary
The present paper is devoted to the study of a bank salvage model with finite time horizon and subjected to stochastic impulse controls. In our model, the bank's default time is a completely inaccessible random quantity generating its own filtration, then reflecting the unpredictability of the event itself. In this fra…
Equivalences are known between problems of singular stochastic control (SSC) with convex performance criteria and related questions of optimal stopping, see for example Karatzas and Shreve [SIAM J. Control Optim. 22 (1984)]. The aim of this paper is to investigate how far connections of this type generalise to a non co…
Paper tackles risk-sensitive decision-making under uncertainty.
problem Risk-sensitive decision-making problem under uncertainty.
method Formulated as a stochastic control problem, delineated necessary optimality conditions.
result Illustrative examples from optimal betting and inventory management support the theory.
In this work, we address the problem of modifying textual attributes of sentences. Given an input sentence and a set of attribute labels, we attempt to generate sentences that are compatible with the conditioning information. To ensure that the model generates content compatible sentences, we introduce a reconstruction…
Combines control variates and adaptive importance sampling for Monte Carlo integration.
problem Improving Monte Carlo integration accuracy with control variates and adaptive sampling.
method A quadrature rule combining control variates and adaptive importance sampling.
result Non-asymptotic bound on the probabilistic error of the procedure.
New method controls posterior collapse in VAEs without network architecture constraints.
problem Posterior collapse in VAEs reduces diversity of generated samples.
method Introduces Latent Reconstruction (LR) loss to control posterior collapse.
result Controls posterior collapse on various datasets without architectural constraints.
Optimizes Ethena's yield strategy by controlling stETH and ETH futures positions.
problem Maximizes Ethena's yield while managing price impacts.
method Formulates and solves stochastic control problems for Ethena's yield-generating strategy.
result Explicitly determines optimal control rates for stETH and ETH futures.
This paper examines a Markovian model for the optimal irreversible investment problem of a firm aiming at minimizing total expected costs of production. We model market uncertainty and the cost of investment per unit of production capacity as two independent one-dimensional regular diffusions, and we consider a general…
A new method models continuous-time counterfactual outcomes using neural controlled differential equations.
problem Estimating personalized healthcare outcomes over irregularly sampled data.
method Interpreting data as samples from a continuous-time process, modeling latent trajectory using controlled differential equations, and using adversarial training for time-dependent confounding.
result TE-CDE consistently outperforms existing approaches in irregularly sampled scenarios.
A framework for robust exploration in reinforcement learning under ambiguity.
problem Optimal stopping under ambiguity in reinforcement learning.
method Continuous-time robust reinforcement learning framework using g-expectation and backward stochastic differential equations. result Constructs a robust exploratory stopping time approximating the optimal stopping time under ambiguity.
New text-to-image diffusion models improve scene understanding for AI agents.
problem Fine-grained scene understanding for AI agents from text and images.
method Pre-trained text-to-image diffusion models optimized for generating images from text prompts.
result Policies learned with Stable Control Representations outperform state-of-the-art approaches on various control tasks.
Unified approach for data-driven control of stochastic processes.
problem Developing practical strategies for stochastic control problems with unknown dynamics.
method Reduction to rate-optimal estimators of invariant distribution risk.
result Data-driven strategies can achieve better performance than known methods.
GAICF proposes a framework for governing generative AI in banking.
problem Generative AI's impact on financial decision-making and governance.
method SR 26-2-compatible governance framework for generative AI applications.
result GAICF aligns generative AI practices with SR 26-2 supervisory expectations.
GAICF proposes a framework for managing generative AI risks in banking.
problem Generative AI's impact on financial decision-making and governance.
method SR 26-2-compatible governance framework for generative AI.
result GAICF aligns generative AI practices with SR 26-2 supervisory expectations.
Study optimizes portfolio to minimize relative drawdown duration, penalizing unfavorable performance states.
problem Minimizing relative drawdown duration in portfolio optimization relative to a benchmark.
method Introduces a benchmark-relative drawdown-duration criterion penalizing unfavorable performance states. Uses a one-dimensional Markovian representation and Hamilton-Jacobi-Bellman equation.
result Derives explicit projection-based characterization of the optimal feedback control and identifies geometric settings for unique strong solutions.
Optimizes bank capital structure under Basel III constraints, simplifying complex dynamics.
problem Optimizing risky investments, dividends, and capital structure under Basel III constraints.
method Formulated as a stochastic control problem, reducing dynamics to a one-dimensional process in leverage ratio.
result Simple policy: pay dividends at an upper barrier and recapitalize at the distress boundary.
The paper uses conformal prediction to detect railway signals with confidence.
problem Deploying deep learning models in certified systems requires accurate uncertainty estimates.
method The paper uses conformal prediction and risk control to detect railway signals.
result The conformal prediction framework provides reliable and trustworthy uncertainty estimates for model performance.
A discrete subgroup of the group of isometries of the hyperbolic space is called reflective if up to a finite index it is generated by reflections in hyperplanes. The main result of this paper is a complete classification of the reflective (and quasi-reflective) subgroups among the Bianchi groups and their extensions.
Hybrid model learns interpretable meal-level glycemic control.
problem Lack of flexible, interpretable meal-level glycemic control methods.
method Hybrid variational autoencoder grounding latent space to mechanistic differential equation.
result Unsupervised representation discovers separation between individuals based on disease severity.
DART2 enhances multiple testing by leveraging ancillary information robustly.
problem Enhancing multiple testing power with uncertain ancillary information.
method Distance-assisted multiple testing procedure (DART2) that handles both helpful and misleading ancillary information.
result DART2 asymptotically controls FDR and improves power when ancillary information is helpful, maintaining FDR and power otherwise.
New method produces reflections with nonseparating fixed points.
problem Constructing hyperbolic manifolds with reflective symmetries.
method Standard method for constructing closed hyperbolic manifolds.
result Fixed point sets of reflections are nonseparating.
Survey explores interactions between four conformal dynamics branches.
problem Understanding complex dynamics through different mathematical concepts.
method Examples and general results with technical tools.
result Dynamical relations between Schwarz reflection parameter spaces and anti-rational maps/ reflection groups.
One reflection suffices for orthogonal weights, reducing GPU usage.
problem Efficiently computing orthogonal weight matrices without high GPU utilization.
method Use an auxiliary neural network to compute one reflection instead of many.
result One reflection is sufficient for orthogonal weights, improving GPU utilization.
Financial event studies often misestimate causal effects due to misspecified factor models.
problem Misspecification of factor models in financial event studies leads to inconsistent estimates of causal effects.
method Proposed synthetic control methods to construct replicating portfolios from control securities.
result Synthetic control methods provide more accurate estimates of causal effects in event studies.
Role mining tackles the problem of finding a role-based access control (RBAC) configuration, given an access-control matrix assigning users to access permissions as input. Most role mining approaches work by constructing a large set of candidate roles and use a greedy selection strategy to iteratively pick a small subs…
ComiRec framework predicts user interests for personalized recommendations.
problem Predicting user interests from sequential behavior data.
method ComiRec framework captures multiple user interests and balances recommendation accuracy and diversity.
result ComiRec achieves significant improvements over state-of-the-art models in sequential recommendation.
A model used for velocity control during car following was proposed based on deep reinforcement learning (RL). To fulfil the multi-objectives of car following, a reward function reflecting driving safety, efficiency, and comfort was constructed. With the reward function, the RL agent learns to control vehicle speed in …
Minimal surfaces in 3-sphere created by reflections from polygons, with new examples based on pentagons.
problem Constructing minimal surfaces in 3-sphere using reflections.
method Minimal n-gon solves free boundary problem; curvature lines combinatorics investigated. result New examples of minimal reflection surfaces based on pentagons.
Stocks of more resilient firms outperformed during the pandemic, reflecting disaster risk.
problem The impact of social distancing on firms' operations and stock performance.
method Cross-sectional analysis of firms' resilience and stock performance, controlling for risk factors.
result Stocks of more resilient firms are expected to yield significantly lower returns than less resilient ones, reflecting disaster risk.
Study thin hyperbolic reflection groups and their properties.
problem Characterize and enumerate thin hyperbolic reflection groups.
method Analyze Zariski dense subgroups of hyperbolic isometries, apply Vinberg algorithm.
result All thin hyperbolic reflection groups are enumerable.
We propose a new method for the numerical solution of backward stochastic differential equations (BSDEs) which finds its roots in Fourier analysis. The method consists of an Euler time discretization of the BSDE with certain conditional expectations expressed in terms of Fourier transforms and computed using the fast F…
Generative model improves EMG pattern recognition accuracy.
problem Stochastic characteristics of EMG signals not fully considered in existing classification methods.
method Scale mixture-based stochastic generative model with variational Bayesian learning.
result Proposed method outperforms conventional classifiers in EMG pattern recognition.