Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

140279419558 · Jun 202019922001200920172026
48 results for Extreme Samples

Paper introduces SPADE method to protect classifiers from OOD and adversarial samples.

problem Protecting classifiers from out-of-distribution and adversarial samples.
method SPADE method based on GEV model in latent space.
result Provable protection against OOD and adversarial samples.

The paper examines extreme value statistics of high-dimensional sample covariances, with applications in finance and image analysis.

problem Statistical validation of normal conditions in high-dimensional time series data.
method Generalizes the maximal deviation of sample autocovariances to high dimensions and applies Gumbel-type extreme value asymptotics.
result Gumbel-type extreme value asymptotics holds true for high-dimensional sample covariances.

We assess cluster stability by trimming extreme points and tracking data range reduction.

problem Assessing stability of one-dimensional clusters.
method Probabilistic method using diameter-shrinkage ratio to track data range reduction.
result Our method achieves higher accuracy than classical tests in small or noisy samples.

A new method reduces uncertainty in predicting rare extreme events without assuming their presence in training data.

problem Predicting rare and extreme events in complex systems with high uncertainty.
method Extreme Event Aware (e2a or η) learning, which enforces extreme event statistics during training.
result Models generate unprecedented extreme events even when training data lacks extremes.

Improved analysis for extreme multi-class CRL with better sample complexity.

problem Theoretical sample complexity of CRL in extreme multi-class settings is poorly understood.
method Improved U-Statistics estimator to capture class concentration, proving O(k)\mathcal{O}(k) sample complexity.
result Sample complexity is O(k)\mathcal{O}(k) for extreme multi-class learning, independent of class distribution.

This paper tackles label-efficient evaluation in extreme class imbalance.

problem Challenges in obtaining a sufficient sample for accurate evaluation in tasks with extreme class imbalance.
method Develops a framework for online evaluation based on adaptive importance sampling.
result Establishes strong consistency and a central limit theorem for performance estimates.

The paper proposes a new variant of a decision tree, called an Extreme Learning Tree. It consists of an extremely random tree with non-linear data transformation, and a linear observer that provides predictions based on the leaf index where the data samples fall. The proposed method outperforms linear models on a bench…

2019-12-19abs ↗pdf ↗

Deep learning models complex multivariate extremes using geometric shapes.

problem Modeling complex extremal dependencies in high-dimensional data.
method Geometric representation and deep learning for flexible semi-parametric models.
result First approach to modeling limit sets using deep learning for high-dimensional data.

New method reduces bias in learning from large action spaces using selective importance sampling.

problem Learning from large-scale recommendation systems with bandit feedback and supervised labels.
method Selective Importance Sampling (sIS) and Policy Optimization for eXtreme Models (POXM) algorithm.
result POXM method significantly outperforms existing methods in learning from bandit feedback on XMC tasks.

Combines GANs and EVT for better modeling of spatial climate extremes.

problem Modeling dependencies between climate extremes, especially in high-dimensional spaces.
method Generative Adversarial Networks (GANs) combined with Extreme Value Theory (EVT).
result evtGAN outperforms classical GANs and statistical approaches in modeling spatial extremes.

The paper provides bounds for the empirical angular measure and applies them to improve statistical learning in extreme regions.

problem Estimating the angular measure in high-dimensional data with different distributions.
method Established bounds for the maximal deviations of the empirical angular measure from the true measure, using rank transformation and analyzing the most extreme observations.
result The bounds provide performance guarantees for statistical learning procedures in extreme regions, such as binary classification and anomaly detection.

Framework reconstructs missing spatio-temporal data for extreme value prediction.

problem Predicting extreme values from incomplete spatio-temporal data.
method Convolutional deep neural networks and autoencoder-like models for conditional sampling.
result Framework produces accurate reconstructions of missing data for extremal values.

Study optimizes sampling to avoid extreme tail risks in unknown heavy-tailed distributions.

problem Identify optimal alternative with minimal extreme tail risk from unknown heavy-tailed distributions.
method Data-driven sequential sampling policies to maximize likelihood of selecting the optimal alternative.
result Proposed methods outperform existing approaches in identifying the optimal alternative.

Extends geometric approach to model non-stationary extremal dependence.

problem Capturing evolving extremal dependence in multivariate data.
method Geometric framework for non-stationary multivariate extreme value modelling.
result Framework can capture various dependence forms and is robust to different model formulations.

Prediction intervals in supervised Machine Learning bound the region where the true outputs of new samples may fall. They are necessary in the task of separating reliable predictions of a trained model from near random guesses, minimizing the rate of False Positives, and other problem-specific tasks in applied Machine …

2019-12-19abs ↗pdf ↗

The hidden tail of empirical distributions is analyzed using extreme value theory.

problem Understanding the bias between in-sample mean and true statistical mean for large nn.
method Extreme value theory applied to empirical distributions and their moments.
result The hidden moment of order 0 for power law distributions follows an exponential distribution with expectation 1/n1/n.

Develops statistical framework for analyzing functional data extremes.

problem Analyzing extremes of functional data in Hilbert spaces.
method Regular variation in Hilbert spaces, Peaks-Over-Threshold framework, functional PCA.
result Proposes a dimension reduction method for functional extreme observations.

New method simulates multivariate extreme events using GANs and Aitchison coordinates.

problem Simulating multivariate extreme events for economic risk assessment.
method Wasserstein-Aitchison GAN approach combining tail dependence and marginal tail modeling.
result Strong performance in capturing tail dependence and generating accurate extreme observations.

Paper proposes efficient GCN learning method for limited data.

problem Learning GCNs from data with extremely limited annotations.
method Adaptive sampling strategy and model compression.
result Cut down annotation requirement by 90% and compress parameters 6x.

Study on consistency of ML methods for moving objects in non-stationary environments.

problem Consistency of machine learning methods for moving objects in non-stationary environments.
method Least squares, ridge regression, and s\ell_s-penalized least squares methods under non-stationary spatial-temporal sampling.
result Consistency and asymptotic normality of the estimates under weak conditions.

New method learns graphical models with latent variables for extreme events.

problem Learning graphical models with latent variables for multivariate extremes.
method Tractable convex program exttt{eglatent} for Hüsler-Reiss models.
result Consistently recovers conditional graph and latent variables.

Novel SVM approach for extreme quantile regression with heavy tailed inputs.

problem Learning from extreme values in quantile regression.
method Support Vector Machine framework for handling high-dimensional and nonlinear settings.
result Established finite-sample learning guarantees under mild regularity assumptions.

Training a classifier over a large number of classes, known as 'extreme classification', has become a topic of major interest with applications in technology, science, and e-commerce. Traditional softmax regression induces a gradient cost proportional to the number of classes CC, which often is prohibitively expensive…

2020-02-15abs ↗pdf ↗

The Extreme Deconvolution method fits a probability density to a dataset where each observation has Gaussian noise added with a known sample-specific covariance, originally intended for use with astronomical datasets. The existing fitting method is batch EM, which would not normally be applied to large datasets such as…

2019-11-26abs ↗pdf ↗

Improved Hawkes model forecasts extreme financial returns more accurately.

problem Forecasting extreme tail events in financial log-returns.
method 2T-POT Hawkes model with multiple exceedance thresholds.
result 2T-POT Hawkes model outperforms GARCH-EVT model in risk forecasting.

Study tail risk in high-frequency finance using L1L_1-regularized regression.

problem Measuring tail risk dynamics in high-frequency financial markets.
method Dynamic extreme value regression model with L1L_1-regularized maximum likelihood estimator.
result Severity of extreme losses well predicted by low price impact in high volatility periods.

This paper develops DRO estimators for EVT statistics using point processes.

problem Scarcity of extreme data leads to model misspecification error in EVT.
method Developed DRO estimators informed by semi-parametric max-stable constraints in the space of point processes.
result Proposed DRO estimators improve out-of-sample performance and are validated on synthetic and real data.

By decomposing asset returns into potential maximum gain (PMG) and potential maximum loss (PML) with price extremes, this study empirically investigated the relationships between PMG and PML. We found significant asymmetry between PMG and PML. PML significantly contributed to forecasting PMG but not vice versa. We furt…

2019-01-07abs ↗pdf ↗