Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

85171256341 · Jun 202019922001200920172026
48 results for Risk Metrics

Proposes resilience metrics for large blackout costs with logarithmic resilience.

problem Large variations in blackout costs make estimating risk impractical.
method Uses mean of log of large blackout costs, tail slope index, and frequency.
result Solves problems of heavy tail and large variations in blackout costs.

AlphaSharpe uses LLMs to improve financial metrics robustness and predictive power.

problem Traditional financial metrics struggle with robustness and generalization in volatile markets.
method Iterative optimization of financial metrics using LLMs, including crossover, mutation, and evaluation.
result AlphaSharpe discovers enhanced risk-return metrics with 3x predictive power and 2x portfolio performance.

Developed a new risk measure, CRI, for evaluating concentrated portfolios.

problem Current risk assessment methods fail to adequately evaluate concentrated portfolios.
method Modified Herfindahl-Hirschman index to create CRI.
result CRI provides a single numeric score for evaluating portfolio risks.

Paper proposes a natural hedging framework with graphical assessment for longevity risk management.

problem Lack of a unified framework for natural hedging and graphical risk assessment.
method Structured natural hedging framework integrated with a graphical risk metric.
result Demonstrates flexibility, interpretability, and practical value for longevity risk management.

We introduce simple cost and risk proxy metrics that can be attached to Treasury issuance strategy to complement analysis of the resulting portfolio weighted-average maturity (WAM). These metrics are based on mapping issuance fractions to their long-term, asymptotic portfolio implications for cost and risk under mechan…

2018-02-09abs ↗pdf ↗

The paper assesses fairness in risk score models, focusing on epistemic value.

problem Fairness of risk score models in communicating uncertainty.
method Identified key fairness desiderata, developed metrics for quantitative assessment, and applied methodology in two case studies.
result Introduced a novel calibration error metric for meaningful comparisons between groups of different sizes.

MDS selects assets by combining daily returns and intraday risk curves, improving portfolio performance.

problem High estimation error in large-scale asset selection.
method Metric Dependence Screening (MDS) incorporating high frequency information as object valued data.
result MDS improves portfolio performance over benchmarks by preserving intraday risk dynamics.

A new framework assesses liquidity risk in perpetual futures exchanges.

problem Measuring and predicting liquidation execution risk in perpetual futures markets.
method Slippage-at-Risk (SaR) framework, comprising three metrics: cross-sectional slippage quantile, expected slippage, and aggregate dollar-denominated tail slippage.
result SaR provides a forward-looking assessment of liquidation execution risk, predictive of systemic stress.

The paper analyzes the generalization of deep neural networks for metric and similarity learning.

problem Lack of rigorous understanding of generalization performance in metric and similarity learning.
method Derive explicit form of true metric, construct structured deep ReLU neural network, establish excess risk bounds.
result Explicit excess risk bounds for metric and similarity learning are derived.

Proposes a method to choose thresholds for LLM evaluation metrics.

problem Ensuring reliable large language models (LLMs) with correct threshold selection.
method Identify risks, stakeholders' risk tolerance, and use ground-truth data to determine thresholds.
result Demonstrates a concrete example with the Faithfulness metric and HaluBench dataset.

We develop a statistical framework to benchmark and select large language models based on their risks.

problem Benchmarking and selecting large language models based on their associated risks.
method A distributional framework using first and second order stochastic dominance, linked to mean-risk models in finance.
result Formalizes a risk-aware approach for model selection, balancing risk and utility.

Investigates model risk and semi-static hedging for martingale constrained models.

problem Model risk distributionally robust sensitivities for functionals on the Wasserstein space.
method Introduces distributionally robust problem with semi-static hedging strategies.
result Explicit characterizations of model risk optimal semi-static hedging strategies.

New metrics improve understanding of predictive system reliability.

problem Evaluating conditional coverage of predictive systems.
method Casting conditional coverage estimation as a classification problem, using excess risk of the target coverage (ERT) metrics.
result Modern classifiers provide higher statistical power for estimating conditional coverage.

New metrics quantify implementation risk in portfolio backtesting, revealing systematic differences in engine implementations.

problem Systematic divergence in backtested portfolio metrics due to differences in engine implementations.
method Formalized implementation risk, proposed four metrics, executed 15 strategies through five engines, analyzed source-code defects.
result Implementation risk introduces measurable ambiguity in performance attribution, but does not alter investment decisions.

Algorithmic risk assessments are increasingly used to help humans make decisions in high-stakes settings, such as medicine, criminal justice and education. In each of these cases, the purpose of the risk assessment tool is to inform actions, such as medical treatments or release conditions, often with the aim of reduci…

2019-08-30abs ↗pdf ↗

The paper analyzes worst-case distortion risk metrics and weighted entropy under partial information.

problem Analyzing worst-case distortion risk metrics and weighted entropy with limited information.
method General distributions, partial information (mean and variance), various entropies and risk measures.
result Provides worst-case results for distortion risk metrics and weighted entropy.

Despite their numerous successes, there are many scenarios where adversarial risk metrics do not provide an appropriate measure of robustness. For example, test-time perturbations may occur in a probabilistic manner rather than being generated by an explicit adversary, while the poor train--test generalization of adver…

2019-12-10abs ↗pdf ↗

Defines computable learning for binary classification over metric spaces.

problem Defines computable PAC learning for binary classification over computable metric spaces.
method Provides sufficient conditions for ERM learners to be computable and bounds the strong Weihrauch degree of an ERM learner.
result Gives a hypothesis class that does not admit any proper computable PAC learner with computable sample function.

We present a framework and analysis of consistent binary classification for complex and non-decomposable performance metrics such as the F-measure and the Jaccard measure. The proposed framework is general, as it applies to both batch and online learning, and to both linear and non-linear models. Our work follows recen…

2016-10-23abs ↗pdf ↗

This paper improves the robustness of risk estimation for financial positions.

problem Ensuring robustness of risk measures in the presence of data noise.
method Proposes a quantitative approach using the Fortet-Mourier metric to quantify the variation of true probability measures.
result Derives explicit error bounds for discrepancies between laws of estimators based on true and perturbed data.

Paper explores generalization of minimax learners, proposing a new metric.

problem Understanding how minimax learners perform on unseen data.
method Proposes a new metric, the primal gap, to study generalization of minimax learners.
result Derives generalization error bounds for the primal gap in nonconvex-concave settings.

New risk metric for RL in finance considers time splits of returns.

problem Optimizing financial decisions with a balance between return and risk.
method Developed a new risk metric for reinforcement learning that allows for flexible target levels of rewards over time.
result Proposed risk metric optimizes for arbitrary time splits of returns, improving upon classical risk measures.

Paper introduces lexical ratio to measure portfolio diversification.

problem Traditional diversification metrics overlook non-numerical relationships.
method Uses textual data to capture diversification dimensions through entropy-based insights.
result Lexical ratio (LR) outperforms traditional metrics in optimizing portfolio returns.

This paper evaluates investment risks in LATAM AI startups using DCF method.

problem Unique challenges and risks faced by LATAM tech startups.
method Total Addressable Market (TAM), Serviceable Available Market (SAM), and Serviceable Obtainable Market (SOM) metrics; Discounted Cash Flow (DCF) method.
result Developed a ranking of emerging powers in Latin America for tech startup investment.

RATE metrics evaluate treatment prioritization rules, subsuming existing methods.

problem Comparing and testing the quality of treatment prioritization rules.
method Rank-weighted average treatment effect (RATE) metrics.
result RATE metrics enable asymptotically exact inference in various study settings.