This study shows how trade policy uncertainty affects stock-T bill correlations.
problem The impact of trade policy uncertainty on stock-T bill relationships.
method Extended Dynamic Conditional Correlation (DCC) framework incorporating exogenous variables.
result Trade policy uncertainty significantly alters stock-T bill correlations, especially under specific political conditions.
The study shows how trade uncertainty affects stock-bond correlations over time.
problem Impact of trade policy uncertainty on stock-bond correlations.
method Daily data analysis using GARCH-based models (CCC, STCC, DCC) with TPU and political dummy variables.
result Time-varying correlation models better capture the dynamics of stock-bond correlations than constant models.
Paper optimizes financial trading strategies under uncertain market conditions.
problem Guaranteeing robust positive expected profits in financial systems.
method Transformed semi-infinite constraints into structured policies and proposed a novel graphical approach.
result Demonstrated superior risk-adjusted returns and downside risk compared to conventional strategies.
A new method predicts stock ranking uncertainty to improve trading performance during regime shifts.
problem Ranking models fail during regime shifts, leading to suboptimal performance.
method Adapting DEUP to rankers, predicting rank displacement and uncertainty, and proposing a two-level deployment policy.
result The two-level deployment policy improves risk-adjusted performance and indicates DEUP adds value mainly as a tail-risk guard.
Bayesian framework improves trading robustness against market shifts.
problem Insufficient robustness and overfitting in trading models.
method Bayesian Robust Framework integrating macro-conditioned GAN and adversarial learning.
result Framework outperforms state-of-the-art models in diverse financial instruments.
AdaRL improves robust RL by adaptively adjusting policy complexity.
problem Handling epistemic uncertainty in environment dynamics.
method Bi-level optimization framework with adaptive rank adjustment.
result AdaRL outperforms existing methods on MuJoCo benchmarks.
Trading off exploration and exploitation in an unknown environment is key to maximising expected return during learning. A Bayes-optimal policy, which does so optimally, conditions its actions not only on the environment state but on the agent's uncertainty about the environment. Computing a Bayes-optimal policy is how…
This paper measures financial market resilience in China and identifies key uncertainties.
problem Measuring financial market resilience in China.
method Quantitative analysis of total financial market and sub-markets, Diebold-Yilmaz connectedness approach.
result Financial market resilience in China is event-driven and influenced by geopolitical risks, economic and trade policy uncertainty, and U.S.-China tensions.
This paper combines RL with CPPI and TIPP for better trading strategies.
problem Challenges in quantitative trading due to swift dynamics and uncertainties.
method Fusion of CPPI and TIPP with MADDPG framework for multi-agent reinforcement learning.
result CPPI-MADDPG and TIPP-MADDPG outperform traditional strategies in real-market shares.
Study aims to optimize financial investments by balancing risk and reward efficiently.
problem Balancing risk and reward in dynamic financial investments.
method Proposes a reinforcement learning method to maximize expected quadratic utility, focusing on first and second moments of rewards.
result The proposed method yields MV-efficient policies that maximize expected reward without increasing variance.
MOPO optimizes offline RL by penalizing dynamics uncertainty.
problem Learning policies from offline data with distributional shift.
method Modify model-based RL to avoid distributional shift.
result MOPO outperforms model-free and standard model-based RL.
Model shows how financial markets can decarbonize under climate uncertainty.
problem Decarbonization of financial markets under climate uncertainty.
method Mean-field game approach to model firm decisions and investor interactions.
result Climate uncertainty weakens the impact of green-minded investors on decarbonization.
Develops optimal uncertainty quantification for risk-averse decision makers.
problem Quantifying prediction uncertainty for risk-sensitive domains.
method Decision-theoretic foundations connecting uncertainty quantification with risk-averse decision-making.
result Risk-Averse Calibration (RAC) algorithm provides optimal prediction sets for risk-averse decision makers.
The trade-off between the cost of acquiring and processing data, and uncertainty due to a lack of data is fundamental in machine learning. A basic instance of this trade-off is the problem of deciding when to make noisy and costly observations of a discrete-time Gaussian random walk, so as to minimise the posterior var…
Adversarial attacks can manipulate deep trading policies, compromising their performance.
problem Adversarial attacks can compromise deep reinforcement learning trading policies.
method Developed a threat model and proposed two attack techniques.
result Demonstrated the effectiveness of adversarial attacks against DQN trading agents.
Study examines how economic policy uncertainty impacts stock markets.
problem Dynamic relationship between economic policy uncertainty and stock markets.
method Used symmetric thermal optimal path (TOPS) method.
result Different interaction patterns observed in emerging and developed markets.
RDMM integrates RL and deep learning for better trading performance.
problem Improving financial trading performance with complex market dynamics.
method Reinforced Deep Markov Model (RDMM) integrating RL and deep learning.
result RDMM outperforms benchmarks in optimal execution problems.
Deep learning predicts uncertainty to optimize Eurodollar futures trading.
problem Optimizing investment size in high-frequency Eurodollar futures trading.
method Deep learning models to estimate prediction uncertainty, scaling investment size.
result Clear outperformance with Sharpe ratio metric compared to alternative strategies.
Robust Markov Decision Processes (RMDPs) intend to ensure robustness with respect to changing or adversarial system behavior. In this framework, transitions are modeled as arbitrary elements of a known and properly structured uncertainty set and a robust optimal policy can be derived under the worst-case scenario. In t…
BCPO optimizes offline RL policies by converting uncertainty into conservative bounds.
problem Offline RL's fragility under distribution shifts and model errors.
method Bayesian approach with credible lower bounds and KL regularization.
result BCPO yields an uncertainty-calibrated policy that avoids exploiting model errors.
We consider the core reinforcement-learning problem of on-policy value function approximation from a batch of trajectory data, and focus on various issues of Temporal Difference (TD) learning and Monte Carlo (MC) policy evaluation. The two methods are known to achieve complementary bias-variance trade-off properties, w…
Optimal dynamic allocation of carbon allowances reduces emissions efficiently.
problem Reducing carbon emissions from firms over time with dynamic allocation and trading.
method Variational approach to solve the Stackelberg game between regulator and firms.
result Optimal policies lead to constant abatement effort and allowance price, outperforming static allocations.
This study analyzes economic policy uncertainty indices using visibility graphs.
problem Understanding the role of economic policy uncertainty in global economies.
method Visibility graph algorithm applied to economic policy uncertainty indices.
result The economic policy uncertainty indices exhibit persistent behavior and scale-free networks.
Study examines how uncertainty visualization affects analyst trust in automated classification systems.
problem The impact of uncertainty on analyst trust in automated classification systems.
method Empirical study evaluating different active learning query policies and visualizations.
result Query policy significantly influences analyst trust in automated classification systems.
There is bountiful evidence that political uncertainty stemming from presidential elections or doubt about the direction of future policy make financial markets significantly volatile, especially in proximity to close elections or elections that may prompt radical policy changes. Although several studies have examined …
Model-based reinforcement learning algorithms tend to achieve higher sample efficiency than model-free methods. However, due to the inevitable errors of learned models, model-based methods struggle to achieve the same asymptotic performance as model-free methods. In this paper, We propose a Policy Optimization method w…
Paper proposes a novel policy distillation method for better order execution in noisy markets.
problem Effective order execution in noisy and imperfect market conditions.
method Policy distillation method to guide reinforcement learning towards optimal trading strategies.
result Significant improvements over various baselines in order execution.
RRPI improves offline RL by optimizing policies against worst-case dynamics.
problem Offline RL's performance degrades under distribution shift and transition uncertainty.
method Formulates offline RL as robust policy optimization, treating transition kernel as decision variable.
result RRPI achieves strong average performance on D4RL benchmarks, outperforming recent baselines.
Study finds monetary policy uncertainty negatively impacts Bitcoin returns.
problem Impact of monetary policy and uncertainty on cryptocurrencies market.
method Markov Switching Means VAR (MSM-VAR) method.
result Monetary policy uncertainty leads to a decline in Bitcoin returns.
DNN policies improve stochastic AC OPF for power grid optimization.
problem Optimizing power grid operations under uncertainty.
method Deep neural network (DNN) policies for real-time generator dispatch decisions.
result DNN policies enforce feasibility constraints and produce near optimal solutions.
New framework improves cost-benefit analysis of policies.
problem Limitations of MVPF in welfare analysis.
method Developed an axiomatic framework to create RPV.
result RPV provides better equity-efficiency trade-off quantification.
New methods improve robust decision-making under uncertainty in off-policy evaluation.
problem Statistical uncertainty and causal considerations in off-policy evaluation.
method Marginal Ratio (MR) estimator, Conformal Off-Policy Prediction (COPP), causal bounds.
result Improved robustness and uncertainty quantification in off-policy decision-making.
UTE improves reinforcement learning by measuring action uncertainty, enhancing policy learning efficiency.
problem Degrading performance of action repetition in reinforcement learning, especially with sub-optimal actions.
method UTE uses ensemble methods to measure uncertainty during action extension, allowing strategic exploration or certainty.
result UTE outperforms existing action repetition algorithms, significantly enhancing policy learning efficiency.
New algorithm improves deep learning stability with limited data.
problem Stability and robustness in reinforcement learning with scarce data.
method Uncertainty-aware trust region approach to policy optimization.
result Stable policy updates adapt to uncertainty levels during learning.
ESRL uses uncertainty quantification to learn safe, optimal policies in offline RL.
problem Challenges in interpreting and measuring uncertainty of learned policies in offline RL.
method Expert-Supervised Reinforcement Learning (ESRL) framework that uses hypothesis testing and posterior distributions.
result The framework can learn safe and optimal policies with theoretical guarantees and independent sample efficiency.
Paper adds a restart mechanism to a drawdown control policy for better trading performance.
problem Missed profitable opportunities when drawdown limit is close to reality.
method Integrates a data-driven restart mechanism into the drawdown modulation trading system.
result The restart mechanism improves trading performance even with transaction costs.
There are few papers about the international trade of flowers, so it is believed that this paper, with this topic, could be an important contribution to the international scientific community. It is intended to analyze if the international trade flowers tendencies and policies are adapted to the actual world global con…
We study option pricing and hedging with uncertainty about a Black-Scholes reference model which is dynamically recalibrated to the market price of a liquidly traded vanilla option. For dynamic trading in the underlying asset and this vanilla option, delta-vega hedging is asymptotically optimal in the limit for small u…
Study shows oil prices but not COVID-19 cases affect US economic policy uncertainty.
problem Effect of COVID-19 and crude oil prices on US economic policy uncertainty.
method Used ARDL model with daily data from January 21-March 13, 2020.
result Crude oil price dynamics increase US economic policy uncertainty, while COVID-19 cases have mixed effects.
End-to-end policy learning improves statistical arbitrage trading.
problem Traditional mean reversion trading strategies in statistical arbitrage are limited.
method We use Autoencoder architectures and policy learning to develop trading strategies.
result End-to-end training yields superior gross returns.
Unified pair trading approach using hierarchical reinforcement learning.
problem Decoupling pair selection and trading leads to limited performance.
method Hierarchical reinforcement learning framework for joint pair selection and trading.
result Unified approach outperforms existing methods on real-world stock data.
Study optimizes natural resource harvesting under model uncertainty using risk measures.
problem Optimal harvesting policy selection for natural resources under model uncertainty.
method Investigated using neoclassical growth model dynamics and convex risk measures, specifically Fréchet risk measures.
result Robust harvesting strategies quantifying operational and marginal risk under model uncertainty.
Proposes DGCN with trajectory sampling for data-efficient policy search in MBRL.
problem Improving data efficiency in model-based reinforcement learning.
method Combines trajectory sampling and DGCN for uncertainty propagation in probabilistic world models.
result Improves sample-efficiency over other uncertainty propagation methods and probabilistic models.
SPReD uses uncertainty to decide imitation from demonstrations.
problem Learning from sparse rewards with few demonstrations.
method Ensemble methods to model Q-value distributions, probabilistic and advantage-based uncertainty quantification.
result Significant gains in reinforcement learning across multiple tasks.
Learning a policy using only observational data is challenging because the distribution of states it induces at execution time may differ from the distribution observed during training. We propose to train a policy by unrolling a learned model of the environment dynamics over multiple time steps while explicitly penali…
New trading policies preserve robust gains in presence of transaction costs.
problem Maintaining robust gains in asset trading with transaction costs.
method Proposed double linear trading policies, analyzed with Monte Carlo simulations and historical data.
result Desired robust positive expected gain can be preserved under certain conditions.
Paper proposes risk-averse reinforcement learning algorithms.
problem Managing model uncertainty in reinforcement learning.
method Entropic risk constrained policy gradient and actor-critic algorithms.
result Demonstrates usefulness of risk-averse algorithms on various domains.
Accurately estimating uncertainties in neural network predictions is of great importance in building trusted DNNs-based models, and there is an increasing interest in providing accurate uncertainty estimation on many tasks, such as security cameras and autonomous driving vehicles. In this paper, we focus on the two mai…