Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

50100149199 · Jun 202019922001200920172026
48 results for risk-averse decision

Develops optimal uncertainty quantification for risk-averse decision makers.

problem Quantifying prediction uncertainty for risk-sensitive domains.
method Decision-theoretic foundations connecting uncertainty quantification with risk-averse decision-making.
result Risk-Averse Calibration (RAC) algorithm provides optimal prediction sets for risk-averse decision makers.

We enhance conformal prediction for risk-averse decisions with action-conditional guarantees.

problem Uncertainty quantification and safety guarantees for machine learning decisions.
method Action-conditional conformal prediction, pinball-loss minimization.
result Action-conditional prediction sets optimize risk-averse decision-making.

Risk measures applied to dynamic Markov processes with varying risk aversion.

problem Investigating dynamic risk measures in Markov decision processes with varying risk aversion.
method Distributional viewpoint on law-invariant convex risk measures, applied to Markov decision processes with latent costs and random actions.
result Existence of optimal policies in finite and infinite time horizons under mild assumptions.

New insights into risk aversion for complex decision models.

problem Understanding risk aversion in non-monotone decision models.
method Characterization of probabilistic risk aversion for generalized rank-dependent functions.
result Probabilistic risk aversion is determined by the distortion function, which is convex or scaled quantile-spread mixtures.

The paper proposes a method to learn and leverage contextual preference distributions for better decision-making.

problem Heterogeneous and context-dependent human preferences in decision-making problems.
method A sequential learning-and-optimization pipeline using a bounded-variance score function gradient estimator to train a predictive model mapping contextual features to preference distributions.
result The approach reduces average post-decision surprise by up to 25 times compared to risk-averse baselines in a ridesharing environment.

Proposes new rule for ranking investment prospects over long horizons.

problem Ranking investment prospects over long horizons considering bounded risk aversion.
method Introduces asymptotic fractional-order stochastic dominance with bounded relative risk aversion.
result Establishes equivalent conditions for the new rule under lognormal returns without mean non-negativity constraint.

PDTS improves robustness in sequential decision-making.

problem Robust active task sampling for efficient and reliable decision-making.
method Characterizes robust active task sampling as a Markov decision process, proposes PDTS method.
result Significantly improves zero-shot and few-shot adaptation robustness.

Online learning has traditionally focused on the expected rewards. In this paper, a risk-averse online learning problem under the performance measure of the mean-variance of the rewards is studied. Both the bandit and full information settings are considered. The performance of several existing policies is analyzed, an…

2018-07-24abs ↗pdf ↗

A new method for risk-averse decision-making in Markov processes with improved regret bounds.

problem Risk-averse decision-making in Markov processes.
method Introduces mini-batch measures and multipattern risk-averse problems in a feature-based QQ-learning method.
result Proves a high-probability regret bound of O(H2NHK)\mathcal{O}\big(H^2 N^H \sqrt{ K}\big) for the QQ-learning method.

Investigates how diversification preferences relate to risk attitudes.

problem Connecting diversification preferences to risk attitudes.
method Analyzes diversification preferences for various pairs of risks under different conditions.
result Diversification preferences for certain pairs of risks imply specific levels of risk aversion.

Modeling consumption and investment decisions with reference point and drawdown constraints.

problem Modeling consumption and investment decisions with reference point and drawdown constraints.
method Solving a stochastic control problem to derive value function, optimal consumption plan, and investment strategy in semi-explicit forms.
result Five important thresholds of wealth, all as functions of hh, and significant economic implications.

In decision under risk, the primal moments of mean and variance play a central role to define the local index of absolute risk aversion. In this paper, we show that in canonical non-EU models dual moments have to be used instead of, or on par with, their primal counterparts to obtain an equivalent index of absolute ris…

2016-12-10abs ↗pdf ↗

Motivated by applications in clinical trials and finance, we study the problem of online convex optimization (with bandit feedback) where the decision maker is risk-averse. We provide two algorithms to solve this problem. The first one is a descent-type algorithm which is easy to implement. The second algorithm, which …

2018-10-01abs ↗pdf ↗

Paper develops NPG for risk-averse RL with ECRMs, proving global convergence.

problem Ensuring reliable performance in stochastic RL problems with risk-averse policies.
method Developed natural policy gradient updates for ECRMs-based RL problems, proving global optimality and iteration complexity.
result Global convergence of risk-averse NPG algorithm with ECRMs.

Develops optimal decision-making framework for uncertain counterfactuals.

problem Ensuring reliability of predictions in high-stakes decisions.
method Policy-Coupled Risk-Averse Conformal Prediction (PC-RACP).
result Optimal prediction sets for counterfactual decisions with valid coverage.

Paper explores how risk-averse individuals' willingness to pay for insurance varies with risk probability.

problem Understanding how risk-averse individuals' willingness to pay for insurance varies with risk probability.
method Analyzes willingness to pay (WTP) for partial risk reduction within the dual theory of decision.
result In dual theory, reducing the probability of risk and providing insurance can be complementary if the surplus increases with risk reduction.

Optimal wind farm placement using quantile constraints for better power output.

problem Optimizing wind farm placement to maximize power output considering spatial and temporal wind speed correlations.
method Used a probabilistic neural network with ReLU activation functions to reformulate constraints as linear ones, embedding them into a two-stage stochastic optimization problem.
result The constraint learning approach outperforms classical methods, especially for risk-averse investors.

The paper examines how loss aversion impacts multi-armed bandit decisions over long periods.

problem The impact of loss aversion on multi-armed bandit decisions over long periods.
method A new central limit theorem for measures with history-dependent variances, derived under risk aversion in gains and risk loving in losses.
result Consequences of loss aversion for asymptotic properties are derived in analytical results.

Risk-averse model uncertainty framework for safe reinforcement learning.

problem Safe decision making in uncertain environments.
method Risk-averse perspective towards model uncertainty using coherent distortion risk measures; equivalent to distributionally robust safe reinforcement learning problems; efficient, model-free implementation.
result Demonstrates robust performance and safety across perturbed test environments.

Investors with anxiety about drawdowns may use stop-loss and trailing stops as optimal selling strategies.

problem Investors' anxiety about drawdowns affects optimal selling strategies.
method Mathematical analysis of optimal stopping with random discounting.
result Stop-loss and trailing stops can be optimal selling strategies under anxiety about drawdowns.

The paper tackles risk-averse multi-armed bandit with linear payoffs.

problem Risk-averse contextual multi-armed bandit problem with linear payoffs.
method Apply Thompson Sampling algorithm for disjoint model and provide comprehensive regret analysis.
result Proved an O((1+ρ+1ρ)dlnTlnKδdKT1+2εlnKδ1ε)O((1+ρ+\frac{1}ρ) d\ln T \ln \frac{K}δ\sqrt{d K T^{1+2ε} \ln \frac{K}δ \frac{1}ε}) regret bound for mean-variance criterion.

Optimizes costs in uncertain Markov systems using risk filters.

problem Optimizing costs in systems with model uncertainty and unknown parameters.
method Risk filters and Bellman principle of optimality applied to Bayesian framework.
result Derives the Bellman principle for non-standard risk-averse control problems.

This article focuses on the work of O. Chanel and G. Chichilnisky (2013) on the flaws of expected utility theory while assessing the value of life. Expected utility is a fundamental tool in decision theory. However, it does not fit with the experimental results when it comes to catastrophic outcomes ---see, for example…

2015-08-25abs ↗pdf ↗

Simple policy outperforms complex ones in cloud auto-scaling.

problem Predicting resource scaling for large-scale cloud applications with limited deployment throughput.
method Probabilistic workload forecast for auto-scaling decisions based on risk aversion.
result The proposed policy outperforms sophisticated and simple benchmark policies in real-world and synthetic data.

The paper explores how investors make decisions under disappointment aversion, finding that they prefer not to invest.

problem Continuous-time portfolio selection under generalized disappointment aversion.
method Sufficient and necessary condition for equilibrium strategies via fully nonlinear integral equation.
result Equilibrium strategy under disappointment aversion leads to less investment in the stock market compared to classical utility theory.

Study on optimal fees in hedge funds with first-loss compensation.

problem Determining the best fee structure for hedge funds with first-loss compensation.
method Solved the manager's non-concave utility maximization problem, calculated Pareto optimal first-loss schemes, and maximized a decision criterion on this set.
result Traditional fees are not Pareto optimal, and the preferred first-loss coverage guarantee varies with investor and market factors.

The maximum entropy principle can be used to assign utility values when only partial information is available about the decision maker's preferences. In order to obtain such utility values it is necessary to establish an analogy between probability and utility through the notion of a utility density function. According…

2007-09-05abs ↗pdf ↗

ESRL uses uncertainty quantification to learn safe, optimal policies in offline RL.

problem Challenges in interpreting and measuring uncertainty of learned policies in offline RL.
method Expert-Supervised Reinforcement Learning (ESRL) framework that uses hypothesis testing and posterior distributions.
result The framework can learn safe and optimal policies with theoretical guarantees and independent sample efficiency.

Optimizes financial decisions with illiquid assets using Kelly criterion.

problem Determining optimal betting strategies in games with external capital constraints.
method Dynamic programming and WKB approximation for multi-round games; Kelly criterion for single-round games.
result Rational players adjust their risk-taking based on the proportion of their capital locked away.

TRIBE model uses LLMs to simulate human trading behavior in bond markets.

problem Complexities in decentralized bond market transactions.
method Agent-based model augmented with LLMs to simulate human-like decision-making.
result Slight trade aversion in LLMs can lead to complete market collapse.

Different models of capital exchange among economic agents have been proposed recently trying to explain the emergence of Pareto's wealth power law distribution. One important factor to be considered is the existence of risk aversion. In this paper we study a model where agents posses different levels of risk aversion,…

2003-11-06abs ↗pdf ↗

In real-world decision-making problems, for instance in the fields of finance, robotics or autonomous driving, keeping uncertainty under control is as important as maximizing expected returns. Risk aversion has been addressed in the reinforcement learning literature through risk measures related to the variance of retu…

2019-12-06abs ↗pdf ↗

Robo-advisors estimate clients' risk aversion using interactive questionnaires.

problem Estimating risk aversion of non-expert clients using adaptive questionnaires.
method Model risk aversion with cost functions and spectral risk measures. Use inverse reinforcement learning to design questions maximizing distinguishing power.
result Designing questions by maximizing distinguishing power achieves satisfactory accuracy in learning risk aversion with fewer than 50 questions.

Risk aversion is a key element of utility maximizing hedge strategies; however, it has typically been assigned an arbitrary value in the literature. This paper instead applies a GARCH-in-Mean (GARCH-M) model to estimate a time-varying measure of risk aversion that is based on the observed risk preferences of energy hed…

2011-03-30abs ↗pdf ↗

Study optimal investment decisions for diverse risk-tolerant agents.

problem Optimizing investment choices for agents with varying risk preferences.
method Characterizes optimal behavior using certainty equivalents and lognormal risks.
result Derives optimal decision menus under known and uncertain preference distributions.

Modeling informed trading with risk-averse market makers.

problem Understanding informed trading and its impact on market liquidity and risk premia.
method Connections between optimal transport theory and Kyle's model, including new characterizations of profits and duality.
result Liquidity is lower, assets exhibit short-term reversals, and risk premia depend on market maker inventories, which are mean reverting.