A new method for risk-sensitive reinforcement learning using Spectral Risk Measures.
problem Incorporating risk sensitivity into reinforcement learning algorithms.
method Proposes a novel framework for optimizing Spectral Risk Measures in both online and offline RL algorithms.
result Demonstrates consistent outperformance over existing risk-sensitive methods in various domains.
Develops an actor-critic algorithm for risk-sensitive Markov decision processes.
problem Risk-sensitive cost criterion in Markov decision processes.
method Actor-critic algorithm with function approximation.
result Asymptotic convergence of the actor-critic algorithm.
Method analyzes complexity of empirical risk landscapes for generalized linear models.
problem Understanding the complexity of empirical risk landscapes in generalized linear models.
method Kac-Rice method and replicated method from theoretical physics.
result Explicit variational formulas for the number of critical points of empirical risk landscapes.
Deployment of emerging technologies and rapid change in industries has created a lot of risk for initiating the new projects. Many techniques and suggestions have been introduced but still lack the gap from various prospective. This paper proposes a reliable project scheduling approach. The objectives of project schedu…
Motivated by optimal investment problems in mathematical finance, we consider a variational problem of Neyman-Pearson type for law-invariant robust utility functionals and convex risk measures. Explicit solutions are found for quantile-based coherent risk measures and related utility functionals. Typically, these solut…
Study on risk measures in reinforcement learning using Monte-Carlo simulations.
problem Lack of satisfactory risk measures in reinforcement learning.
method Generalized approximation scheme based on Monte-Carlo simulations, neural architecture for risk estimation.
result Variance of reward-to-go does not adequately capture risk in reinforcement learning.
Develops a risk score to assist ECMO planning for critically ill patients with viral or unspecified pneumonia.
problem Lack of a risk score to guide ECMO planning for critically ill patients.
method Leverages machine learning to develop the PEER score.
result PEER score predicts mortality and decompensation in patients eligible for ECMO.
Small stocks drive market crashes by suppressing resilience.
problem The single-security price limit exacerbates market panic during crashes.
method Simplified dynamic model on networks of investors and stocks, empirical verification.
result Unexpected linear association between price limit and critical market confidence.
A framework identifies worst-case decision points in safety-critical scenarios, improving risk assessment by 10 hours.
problem Identifying worst-case outcomes in safety-critical decision-making under uncertainty.
method Explicitly estimating distributions of expected return to identify dead-ends, tuning based on risk tolerance.
result Significantly improves risk assessment, providing indications 10 hours earlier and increasing detection by 20%.
CARL safely adapts RL agents for safety-critical tasks.
problem Safety hazards in RL for safety-critical tasks.
method CARL combines model-based RL and cautious adaptation.
result CARL achieves higher rewards with fewer failures in safety-critical tasks.
Recently, along with the emergence of food scandals, food supply chains have to face with ever-increasing pressure from compliance with food quality and safety regulations and standards. This paper aims to explore critical factors of compliance risk in food supply chain with an illustrated case in Vietnamese seafood in…
Study landscape of non-convex empirical risk with degenerate population risk.
problem Degenerate non-convex population risk in machine learning problems.
method Analyze population risk first, then connect to empirical risk landscape.
result Established correspondence between empirical and population risk critical points.
Develops RL for dynamic risk assessment in stochastic optimization.
problem Time-consistent risk assessment in stochastic optimization problems.
method Model-free reinforcement learning with dynamic convex risk measures, time-consistent dynamic programming, policy gradient updates, actor-critic neural network optimization.
result Demonstrates optimal policies for statistical arbitrage, financial hedging, and robot control.
Paper proposes a risk-aware decision-making framework for real-world sequential decisions.
problem Real-world sequential decision-making problems often have critical constraints that learning solutions often neglect.
method Actor multi-critic architecture with risk characterization.
result Our approach consistently satisfies system constraints with minimal performance toll.
Systematic and multifactor risk models are revisited via methods which were already successfully developed in signal processing and in automatic control. The results, which bypass the usual criticisms on those risk modeling, are illustrated by several successful computer experiments.
Proposes a new framework for risk-sensitive RL using deep nets.
problem Risk-sensitive reinforcement learning problems.
method Conditional elicitability, scoring functions, deep neural networks.
result Dynamic spectral risk measures can be approximated by deep nets.
Paper tackles complex risk in deep neural networks.
problem Complex risk in deep neural networks.
method Developed new approach for complex risk statistics.
result Derived dual representation for complex risk.
Paper develops risk statistics for portfolios considering regulator-based risk.
problem Traditional risk statistics fail to describe regulator-based risk.
method Develop dual representation for regulator-based risk statistics.
result Derived dual representation for regulator-based risk statistics.
This work tackles risk-sensitive deep RL by optimizing policies with variance constraints.
problem Risk and aleatoric uncertainty in deep reinforcement learning.
method Lagrangian and Fenchel dualities to transform the problem into an unconstrained saddle-point policy optimization problem, and an actor-critic algorithm to iteratively update policy, Lagrange multiplier, and Fenchel dual variable.
result The proposed actor-critic algorithm finds a globally optimal policy at a sublinear rate.
This paper critiques the Standardized Measurement Approach (SMA) for operational risk and recommends maintaining Advanced Measurement Approach (AMA).
problem Weaknesses and failures of the Standardized Measurement Approach (SMA) in operational risk.
method Critical review and analysis of SMA and AMA approaches.
result SMA is unstable, insensitive to risk, and implicitly related to systemic risk in the banking sector.
We study the feasibility and noise sensitivity of portfolio optimization under some downside risk measures (Value-at-Risk, Expected Shortfall, and semivariance) when they are estimated by fitting a parametric distribution on a finite sample of asset returns. We find that the existence of the optimum is a probabilistic …
A new RL framework for risk-sensitive decision-making using convex scoring functions.
problem Time-inconsistent risk measures in reinforcement learning.
method Convex scoring functions, augmented state space, auxiliary variable, customized Actor-Critic algorithm.
result Theoretical guarantees for approximation and convergence under certain conditions.
Study shows flash crashes in finance are self-organized criticality events.
problem Understanding and predicting anomalous price events in high-frequency finance.
method Investigated volume distributions during flash crashes and linked them to self-organized criticality.
result Volume distributions during flash crashes indicate a diverging second moment, suggesting self-organized criticality.
Improves RL generalization by minimizing adversarial risk.
problem Overfitting to training environments and poor generalization to unseen scenarios.
method Introduces minimax formulation and distributional framework to RL.
result Trained policy shows improved generalization to different environments.
New algorithm for risk-sensitive reinforcement learning with natural policy gradients.
problem Risk-sensitive reinforcement learning with downside risk constraints.
method Introduce a new Bellman equation to estimate the lower partial moment of returns, use natural policy gradients, and extend Reward Constrained Policy Optimization.
result Sample-efficient estimation of partial moments and effective risk-sensitive control.
Reinforcement learning for continuous-time risk-sensitive asset allocation
problem Continuous-time risk-sensitive asset allocation
method Free energy-entropy duality reformulation and q-learning actor-critic method result Optimal policy learning with high accuracy
Paper proposes risk-averse reinforcement learning algorithms.
problem Managing model uncertainty in reinforcement learning.
method Entropic risk constrained policy gradient and actor-critic algorithms.
result Demonstrates usefulness of risk-averse algorithms on various domains.
In many sequential decision-making problems we may want to manage risk by minimizing some measure of variability in rewards in addition to maximizing a standard criterion. Variance related risk measures are among the most common risk-sensitive criteria in finance and operations research. However, optimizing many such c…
The study uses AI to optimize trading in FX markets by considering size-dependent fees and risk-aversion.
problem Optimizing trading in FX markets with size-dependent fees and risk-aversion.
method Fitted Natural Actor-Critic (FNC) Reinforcement Learning algorithm.
result The algorithm effectively trades with variable order sizes, reducing transaction costs and promoting risk-averse behavior.
We study the sensitivity to estimation error of portfolios optimized under various risk measures, including variance, absolute deviation, expected shortfall and maximal loss. We introduce a measure of portfolio sensitivity and test the various risk measures by considering simulated portfolios of varying sizes N and for…
The paper introduces a new risk statistic considering the time value of money.
problem Traditional risk statistics do not fully account for the time value of money.
method Introducing set-valued risk statistics with the time value of money.
result The new risk statistic provides a more accurate quantification of portfolio risk.
Optimal decision-making using prediction sets to minimize risk.
problem Using prediction sets optimally for decision-making in uncertain scenarios.
method Decision-theoretic framework that seeks to minimize expected loss against a worst-case distribution.
result ROCP algorithm reduces critical mistakes compared to baselines, especially in costly out-of-set errors.
The year 2017 saw the rise and fall of the crypto-currency market, followed by high variability in the price of all crypto-currencies. In this work, we study the abrupt transition in crypto-currency residuals, which is associated with the critical transition (the phenomenon of critical slowing down) or the stochastic t…
Deep RL solves dynamic risk pricing for complex financial models.
problem Dynamic risk measures in financial derivatives pricing.
method Deterministic actor-critic deep reinforcement learning (ACRL) for time-consistent expectile risk.
result High-quality hedging policies and prices for complex financial instruments.
Deep Hedging learns optimal strategies for various risk levels.
problem Finding optimal hedging policies for diverse risk aversions.
method Continuous Reinforcement Learning with actor-critic algorithm.
result Demonstrated effectiveness in a stochastic volatility model.
Paper proposes efficient method for estimating risk measures in complex models.
problem Accurately estimating distortion risk measures in computationally expensive models.
method Integrates importance sampling and machine learning for efficient Monte Carlo estimation.
result Demonstrates significant reduction in computational cost for estimating risk measures.
External or internal shocks may lead to the collapse of a system consisting of many agents. If the shock hits only one agent initially and causes it to fail, this can induce a cascade of failures among neighoring agents. Several critical constellations determine whether this cascade remains finite or reaches the size o…
Develops a new method for risk diversification using dynamic risk measures.
problem Dynamic risk diversification in investment portfolios.
method Introduces dynamic risk contributions and a recursive optimization approach for coherent dynamic distortion risk measures.
result Dynamic risk budgeting strategies can be solved using deep learning.
We analyze the landscape of empirical risk minimization for high-dimensional models, predicting phase transitions and critical point properties.
problem Understanding the complexity and structure of high-dimensional empirical risk landscapes.
method Using the Kac-Rice formula, we analyze the expected number of critical points and their spectral properties, providing detailed predictions.
result We derive complete topological phase diagrams for the phase retrieval problem, predicting BBP-type transitions and critical point stability.
An interbank market lets participants pool the risk arising from the combination of illiquid investments and random withdrawals by depositors. But it also creates the potential for one bank's failure to trigger off avalanches of further failures. We simulate a model of interbank lending to study the interplay of these …
Several authors have recently developed risk-sensitive policy gradient methods that augment the standard expected cost minimization problem with a measure of variability in cost. These studies have focused on specific risk-measures, such as the variance or conditional value at risk (CVaR). In this work, we extend the p…
Acute kidney injury (AKI) in critically ill patients is associated with significant morbidity and mortality. Development of novel methods to identify patients with AKI earlier will allow for testing of novel strategies to prevent or reduce the complications of AKI. We developed data-driven prediction models to estimate…
New method certifies risks of LLM outputs, improving accuracy and reliability.
problem Uncertain and incorrect outputs from large language models.
method Information-lift certificates using PAC-Bayes bounds and skeleton design.
result Achieves 77.0% coverage at 2% risk, outperforming baselines.
The practice of valuation by marking-to-market with current trading prices is seriously flawed. Under leverage the problem is particularly dramatic: due to the concave form of market impact, selling always initially causes the expected leverage to increase. There is a critical leverage above which it is impossible to e…
Recurring international financial crises have adverse socioeconomic effects and demand novel regulatory instruments or strategies for risk management and market stabilization. However, the complex web of market interactions often impedes rational decisions that would absolutely minimize the risk. Here we show that, for…
SBCA optimizes portfolios by fusing price data and text sentiment.
problem Insufficient integration of multi-modal information in traditional portfolio optimization models.
method Cross-modal BERT-driven Actor-Critic framework with gated fusion and constraint embedding.
result SBCA outperforms benchmarks in portfolio value, return, Sharpe ratio, and maximum drawdown.
Survey finds many adversarial machine learning threats are not critical for most entities.
problem Adversarial machine learning threats and their impact on model accuracy.
method Literature review and analysis of real-world occurrences of adversarial attacks.
result Many adversarial machine learning threats do not warrant the cost of robust models.
The study analyzes how large language models form and express investor risk profiles.
problem Understanding how large language models (LLMs) form and express investor risk profiles.
method Examined three LLMs (GPT, Gemini, and Llama) and assessed their responses to a standardized risk questionnaire under varying prompts.
result LLMs generally form long-term investment profiles, but they exhibit different risk tolerance levels.