Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,341 papers · 148 categories

Trend · papers per month

2457 · Jul 202519922001200920182026
48 results for underperformance

Lower bound on portfolio underperformance risk over time.

problem Minimizing risk of a portfolio underperforming a benchmark over long periods.
method Modelled prices of securities as geometric Brownian motions with nonlinear coefficients and economic factor modeled by Ito equation. Obtained a tight lower bound on underperformance probability.
result Lower bound on decay rate of underperformance probability is tight and can be achieved with epsilon-optimal portfolios under certain conditions.

We develop a simple stock selection model to explain why active equity managers tend to underperform a benchmark index. We motivate our model with the empirical observation that the best performing stocks in a broad market index often perform much better than the other stocks in the index. Randomly selecting a subset o…

2015-10-13abs ↗pdf ↗

ETFs with 2x and 3x leverage underperformed the S&P 500 index due to compounding and volatility.

problem ETFs with higher leverage failed to match the performance of the underlying index.
method Analyzed the performance of leveraged ETFs compared to the S&P 500 index, accounting for compounding and volatility.
result Two-thirds of the underperformance was due to compounding and volatility, with the rest due to covariance.

FedDANE adapts DANE for federated learning, but underperforms compared to existing methods.

problem Federated learning's practical constraints and device heterogeneity.
method Adapted DANE for federated learning, providing convergence guarantees for convex and non-convex functions.
result Empirically, FedDANE underperforms compared to FedAvg and FedProx.

A financial market is called "diverse" if no single stock is ever allowed to dominate the entire market in terms of relative capitalization. In the context of the standard Ito-process model initiated by Samuelson (1965) we formulate this property (and the allied, successively weaker notions of "weak diversity" and "asy…

2008-03-20abs ↗pdf ↗

New methods for handling time-varying label noise in time series classification.

problem Temporal label noise in time series classification tasks.
method Proposed methods to estimate temporal label noise function directly from data.
result Our methods lead to state-of-the-art performance under diverse types of temporal label noise.

Leveraged ETFs can outperform their targets in certain market conditions, contrary to the volatility drag hypothesis.

problem The long-term performance decay of leveraged ETFs due to volatility drag.
method Unified framework incorporating AR(1) and AR-GARCH models, continuous-time regime switching, and flexible rebalancing frequencies.
result Return dynamics, including return autocorrelation, volatility clustering, and regime persistence, determine LETF performance.

Study finds rough volatility models underperform in SPX option pricing.

problem Inconsistency of rough volatility models with SPX option prices.
method Empirical study using SPX options data, comparing rough and Markovian models.
result Rough volatility models with H(0,1/2)H \in (0,1/2) are inconsistent with SPX smiles, especially at short maturities.

This study examines the tracking errors of commodity leveraged ETFs, finding many underperform significantly.

problem Tracking errors of commodity leveraged ETFs over longer horizons.
method Constructed a benchmark process accounting for volatility decay and used it to examine ETFs' performance.
result Many commodity leveraged ETFs underperform significantly against a benchmark, quantified via realized effective fee.

ChatGPT struggles in predicting stock movements, underperforming traditional methods.

problem Predicting stock market movements using ChatGPT.
method Zero-shot analysis of ChatGPT's multimodal stock prediction capabilities.
result ChatGPT underperforms traditional methods and state-of-the-art models in predicting stock movements.

BS-NAS broadens and shrinks search space for optimal neural architectures.

problem Suboptimal channel numbers and model averaging effects in One-Shot NAS methods.
method Broadening with spring block for channel search, shrinking with underperforming operations removal, evolutionary algorithm for optimal architecture search.
result BS-NAS achieves state-of-the-art performance on ImageNet.

The paper examines how optimizer comparisons in deep learning are influenced by hyperparameter tuning.

problem The sensitivity of optimizer comparisons to hyperparameter tuning protocols.
method Empirical comparisons of optimizers with and without varying hyperparameter search spaces.
result Inclusion relationships between optimizers matter in practice and can contradict recent empirical comparisons.

Framework selects real estate redevelopment uses by integrating value, risk, complexity, and irreversibility.

problem Persistent underperformance of real estate assets due to structural misalignment.
method Integrates real-options logic and multi-criteria decision analysis.
result Reduces over-complexification and misalignment in strategic use selection.

Quantum model outperforms classical in training but underperforms in real-world metrics.

problem Mismatch between proxy reward signals and true investment objectives in financial domains.
method Hybrid quantum-classical reinforcement learning framework with automated feature engineering.
result Quantum models achieve higher training rewards but underperform in real-world metrics.

Study introduces new financial ratios for better predicting company performance.

problem Lack of progress in predicting company performance and assessing financial risks.
method Developed new financial and macroeconomic ratios, supervised learning models, and Bayesian models.
result New proposed variables improve model accuracy and FNN performs best across multiple tasks.

Study reveals AI skin cancer classifiers underperform for darker skin phototypes, advocating for fairness auditing.

problem AI bias in dermatology, particularly for darker skin phototypes.
method Predictive Representativity (PR) framework, evaluating classifiers on HAM10000 and BOSQUE Test sets.
result Substantial performance disparities by skin phototype, highlighting AI bias.

Calibrating a trading rule using a historical simulation (also called backtest) contributes to backtest overfitting, which in turn leads to underperformance. In this paper we propose a procedure for determining the optimal trading rule (OTR) without running alternative model configurations through a backtest engine. We…

2014-08-06abs ↗pdf ↗

Study finds traditional technical indicators underperform in high-frequency trading, suggesting risk management over prediction.

problem Inadequately explored effectiveness of technical indicators in high-frequency trading, particularly at minute-level frequency.
method Evaluation of random forest models with traditional technical indicators on minute-level SPY data.
result In-sample performance is superior to out-of-sample, with risk-adjusted metrics not outperforming a simple buy-and-hold strategy.

Machine learning portfolios perform well with simple imputation of missing data.

problem Handling missing values in machine learning portfolios constructed from cross-sectional return predictors.
method Simple imputation with cross-sectional means compared to rigorous expectation-maximization methods.
result Simple imputation performs well due to the structure of missing data.

With the increasing size of today's data sets, finding the right parameter configuration in model selection via cross-validation can be an extremely time-consuming task. In this paper we propose an improved cross-validation procedure which uses nonparametric testing coupled with sequential analysis to determine the bes…

2012-06-11abs ↗pdf ↗

Maximizes stock portfolio predictability using machine learning.

problem Improving stock portfolio performance through predictive modeling.
method Optimal constrained weights in the MPP constructed using Elastic Net, Random Forest, and Support Vector Regression models.
result MPP portfolios can outperform or underperform the index based on the time period.

This thesis identifies share buybacks and predicts their impact on stock performance.

problem Recognizing and predicting the impact of share buybacks on stock performance.
method NLP approaches for automated detection of share buybacks, machine learning models for prediction.
result Most companies underperform after a share buyback, but some significantly outperform.

FinFlowRL combines imitation and reinforcement learning for better financial control.

problem Traditional stochastic control methods fail in real-world finance due to changing market conditions.
method FinFlowRL uses imitation learning to pretrain an adaptive meta policy, then finetunes it with reinforcement learning.
result FinFlowRL consistently outperforms individual strategies across various market conditions.

ZeroS improves Transformers by adding negative weights, matching or beating softmax attention.

problem Limited performance of linear attention methods, especially in long context sequences.
method Proposes Zero-Sum Linear Attention (ZeroS) that removes the zero-order term and reweights zero-sum softmax residuals.
result ZeroS matches or exceeds standard softmax attention across various benchmarks, theoretically expanding representable functions.

Study finds financial YouTube channel 3PROTV predicts stock market performance and sentiment changes.

problem Determining the informational value of financial YouTube channels.
method Analyzing 3PROTV's content and its impact on stock market performance and sentiment.
result 3PROTV's content, particularly negative sentiment, predicts stock market performance and sentiment changes.

Bayesian model averaging fails under covariate shift, affecting neural networks' performance.

problem Bayesian model averaging's failure in neural networks under covariate shift.
method Explained the issue and proposed novel priors to improve robustness.
result Bayesian model averaging is problematic under covariate shift, especially with linear feature dependencies.

AlphaZeroBeta uses deep reinforcement learning for market-neutral portfolios, outperforming traditional methods.

problem Traditional portfolio management methods often fail during market regime shifts or when assumptions break down.
method Combines a composite reward function and CNN-GRU policy trained end-to-end via Recurrent PPO.
result Achieves higher Sharpe ratios than baselines while maintaining near-zero benchmark correlations.

Growth rate of real GDP per capita is represented as a sum of two components -- a monotonically decreasing economic trend and fluctuations related to a specific age population change. The economic trend is modeled by an inverse function of real GDP per capita with a numerator potentially constant for the largest develo…

2008-11-06abs ↗pdf ↗

Bayesian rating system for large competitions improves prediction and efficiency.

problem Rating systems for large, competitive events like online programming contests.
method Developed a Bayesian rating system for many participants, proving robustness and runtime.
result The system outperforms existing systems in accuracy and computation speed.

Single tree outperforms random forest in testing accuracy.

problem The challenge of improving single decision tree performance.
method Gradient-based entire tree optimization framework, scaled sigmoid approximation, numerical stability algorithm, subtree polish strategy.
result Optimized single tree outperforms classic random forest by 2.03% on average.

This study improves hyperparameter optimization for categorical and non-normal data.

problem Bayesian hyperparameter optimization struggles with categorical hyperparameters and non-normal data.
method Integrates conformalized quantile regression to address estimation weaknesses and provides robust calibration guarantees.
result Quantile surrogate architectures and acquisition functions yield superior performance compared to existing methods.

BLAE solves batched linear bandits with optimal regret and practical performance.

problem Batched linear bandit problem with limited adaptivity.
method Integrates arm elimination with regularized G-optimal design, achieving minimax optimal regret.
result Achieves minimax optimal regret in both large-KK and small-KK regimes with O(loglogT)O(\log\log T) batches.

Improved Thompson Sampling outperforms existing Bayesian optimization methods.

problem Thompson Sampling's performance in Bayesian optimization is suboptimal compared to other methods.
method Developed Stagger Thompson Sampler (STS), which more precisely samples the optimal arm with less computation.
result STS outperforms TS, PSS, and other acquisition methods in various optimization tasks.