New framework tests mean-variance spanning in high dimensions.
problem Testing mean-variance spanning in high-dimensional asset spaces.
method Robust Student-t statistic based on batch-mean method, combined using Cauchy combination test.
result Advantages of diversification vary by economic conditions and cross-country.
Proposes a modified Morgan-Pitman test for evaluating variances in machine learning models.
problem Limited ability to account for sampling variability in model selection.
method Enhances the classic Morgan-Pitman test for robustness in non-linear models with heavy-tailed distributions or outliers.
result Demonstrates the test's effectiveness and practical utility in model evaluation and selection.
Adversarial training leads to large generalization gap, decomposed into bias and variance.
problem Understanding the large generalization gap in adversarially trained models.
method Bias-Variance decomposition of test risk as a function of adversarial perturbation radius.
result Bias increases monotonically with adversarial perturbation radius and is dominant in test risk.
Develops abstention procedure for nonparametric regression via variance testing.
problem Prediction with selective abstention in error-critical machine learning.
method Nonparametric heteroskedastic regression via testing hypothesis on conditional variance.
result Non-asymptotic risk bounds and convergence regimes for the estimator.
Modern neural networks show no bias-variance tradeoff with increased parameters.
problem The traditional bias-variance tradeoff does not hold in over-parameterized neural networks.
method Empirical measurements and theoretical analysis of bias and variance in modern neural networks.
result Bias and variance can decrease as the number of parameters grows in over-parameterized neural networks.
This work uses ANOVA to understand how different factors contribute to test error in machine learning models.
problem Understanding why overparametrized models generalize well despite potentially fitting noise.
method Analysis of variance (ANOVA) to decompose test error into components of variance.
result The interaction between training samples and initialization can dominate variance, and there are phase transitions in variance behavior.
Reduces quantifier variance with accuracy optimization of base classifier.
problem Minimizing quantifier variance under prior probability shift.
method Optimizes the Brier score of a base classifier for training data.
result Optimizing Brier score on training data reduces quantifier variance on test data.
A new permutation method improves two-sample testing power.
problem Two-sample testing with improved power and validity.
method Structured block-restricted cross-swaps.
result Block-restricted permutations achieve higher power than full permutations.
Study tests GMVP weights in high-dimensional settings, comparing sample and shrinkage estimators.
problem Testing GMVP weights in high-dimensional settings with varying sample size and asset count.
method Developed two tests based on sample and shrinkage estimators of GMVP weights.
result Shrinkage estimator test performs well even for high asset counts.
Adaptive sampling method reduces variance in stochastic optimization.
problem Reducing variance in stochastic optimization with limited gradient computations.
method Adaptive increase in sample size based on inner product test.
result Algorithm converges globally on nonconvex functions and linearly on strongly convex functions.
Introduces TPV to analyze model robustness without labels.
problem Analyzing post-training robustness of machine learning models.
method Parameter perturbations and test prediction variance (TPV) as a unifying framework.
result TPV connects various perturbations under a single lens, providing insights into model stability.
Paper explains why Dropout and BN lead to worse performance when combined and proposes solutions.
problem Worse performance when Dropout and BN are combined.
method Theoretical analysis and experiments on various networks to identify variance shift and propose solutions.
result Dropout shifts variance of a specific neural unit, while BN maintains accumulated variance, leading to unstable predictions.
BYOV combines SSL and Bayesian methods for uncertainty estimation.
problem Model uncertainty in applications.
method Combines Bootstrap Your Own Latent (BYOL) and Bayes by Backprop (BBB).
result BYOV improves model calibration and reliability with various augmentations.
Empirical study finds variance swap rate is affine in spot variance for S&P500 data.
problem Investigating the relationship between variance swap rate and spot variance.
method Empirical analysis using S&P500 data from 2006-2018, testing different models.
result Affine relationship between variance swap rate and spot variance is supported.
Develops new e-processes and confidence sequences for Gaussian means with unknown variance.
problem Constructing valid t-tests and confidence sequences for Gaussian means with unknown variance.
method Explores generalized nonintegrable martingales and extended Ville's inequality, developing two new e-processes and confidence sequences.
result Analyzes the width of resulting confidence sequences with a polynomial dependence on error probability, proving it to be unavoidable and even better than classical fixed-sample t-tests.
Paper proposes a new importance sampling method for reducing variance.
problem Reducing variance in importance sampling when training and testing data come from different distributions.
method A new variant of importance sampling that reduces variance by orders of magnitude.
result The new estimator can improve estimates of treatment effectiveness using limited data.
Improved ANOVA test under differential privacy with higher statistical power.
problem Carrying out ANOVA tests while maintaining privacy.
method Developed a new test statistic \(F_1\) and a method to compute its reference distribution.
result Our test \(F_1\) achieves a significant improvement in statistical power compared to previous methods.
Faster convergence of kernel mean embeddings using variance information.
problem Speeding up the convergence rate of kernel mean embeddings.
method Leveraging variance information in reproducing kernel Hilbert space and estimating variance from data.
result Efficiently estimate variance information from data to achieve distribution-agnostic convergence bounds.
Deep networks generalize well even when they fit training data perfectly, thanks to overparametrization.
problem Understanding generalization in overparametrized deep networks.
method Random features regression, asymptotic analysis, ensemble averaging.
result Bias remains constant beyond the interpolation threshold, while variance components decay with overparametrization.
This paper proposes a new AED framework for multi-metric experiments with fixed budget.
problem Statistical power challenges in testing multiple metrics simultaneously.
method Two-phase structure: adaptive exploration followed by validation. SHRVar algorithm with relative-variance-based sampling.
result Achieves provable error probability that decreases exponentially.
A/B testing improves marketing decisions by selecting effective stratification variables.
problem Improving the sensitivity of A/B testing through stratified sampling.
method Designing an algorithm to select a subset of stratification variables for variance reduction.
result The subset selection method outperforms other variance reduction techniques in A/B testing.
Improves A/B testing for long-term outcomes in dynamic systems.
problem Estimating long-term effects from short-term A/B testing data.
method Develops optimal inference techniques and localized information sharing methods.
result New estimator reduces variance linearly with test arms and matches lower bounds.
We revisit resampling procedures for error estimation in binary classification in terms of U-statistics. In particular, we exploit the fact that the error rate estimator involving all learning-testing splits is a U-statistic. Thus, it has minimal variance among all unbiased estimators and is asymptotically normally dis…
The paper uses the variance-gamma model to price options and explain excess kurtosis.
problem Explaining excess kurtosis in stock price data.
method Random-time subordination, Laplace distribution, Esscher transform.
result The variance-gamma model explains excess kurtosis in log-returns data.
New insights into bias and variance in over-parameterized models.
problem Understanding bias and variance in over-parameterized models.
method Analytic expressions derived from statistical physics for two minimal models.
result Over-parameterized models can overfit even in noiseless conditions.
Study proposes memory-efficient backpropagation for linear layers in neural networks.
problem Significant memory usage in backpropagation through linear layers in neural networks.
method Randomized matrix multiplications to reduce memory usage with a moderate decrease in test accuracy.
result Demonstrated benefits of the proposed method on fine-tuning pre-trained models.
Scaling laws in linear regression explain model performance improvements with size and data.
problem Disagreement between empirical neural scaling laws and conventional wisdom on variance error.
method Infinite dimensional linear regression setup, one-pass SGD, Gaussian prior, power-law spectrum.
result Variance error is dominated by other errors, disappearing from the bound due to SGD's implicit regularization.
A simple method treats heteroscedastic variance variatively, improving model calibration and sample quality.
problem Brittle optimization impacts model likelihoods for mean and variance estimation.
method Proposes a variational approach to heteroscedastic variance, improving predictive mean and variance calibration.
result The proposed method significantly improves parameter calibration and sample quality for regression and VAEs.
USNRT uses tree-structured learning to improve uncertainty quantification of variance networks.
problem Improving uncertainty quantification of variance networks.
method Tree-structured local neural network model that partitions feature space into regions for training region-specific neural networks to predict mean and variance.
result USNRT shows superior performance in estimating uncertainty with variances on UCI datasets compared to recent methods.
This paper analyzes the posterior variance of Gaussian processes and derives a new bound.
problem Lack of suitable analysis of posterior variance for finite and infinite training data.
method Derives a novel bound for posterior variance requiring only local information.
result Proves sufficient conditions for the convergence of posterior variance to zero and demonstrates improved average learning bound.
Proposes incorporating noise sources in machine learning evaluation for more reliable conclusions.
problem Inadequate handling of nondeterminism in machine learning research leads to unreliable results.
method Uses linear mixed effects models (LMEMs) and generalized likelihood ratio tests (GLRT) to analyze performance evaluation scores and assess performance differences.
result Demonstrates how to incorporate various sources of noise and data properties into statistical significance testing and reliability analysis.
New algorithm detects changes in high-dimensional data with mean and variance.
problem Challenges in detecting changes in high-dimensional data with mean and variance.
method Complete graph-based approach to detect changes of mean and variance from low to high-dimensional online data.
result The proposed method outperforms existing methods in terms of detection power.
New method uses machine learning to improve statistical inference.
problem Performing inference on conditional functionals with scarce labeled data.
method Combines localization with prediction-based variance reduction.
result Valid and sharp confidence intervals for conditional functionals.
Detects underspecification in pre-trained models using local ensembles.
problem Underspecification in pre-trained models where many predictors are consistent with training data.
method Uses local second-order information to approximate prediction variance across an ensemble of models.
result Capable of detecting underspecification in pre-trained models on test data.
The paper examines skill estimation and variance under model misspecification in IRT.
problem Underestimation and overestimation of skills when non-compensatory model is misspecified as compensatory.
method Theoretical approach to analyze underestimation and overestimation of skills and variance.
result Overestimation of skills occurs around the origin and asymptotic variance differs under model misspecification.
PPAT uses predictions to improve risk estimation in active testing.
problem Exploiting informative predictions from black-box models for efficient risk estimation.
method Combines LURE estimator with prediction-powered control variate.
result PPAT outperforms existing methods in risk estimation and uncertainty quantification.
Paper tests for time-varying entropy in stock prices, finding periods of inefficiency.
problem Testing for time-varying entropy in stock price dynamics.
method Unbiased approximation of Shannon entropy variance, optimal rolling window selection, hypothesis testing.
result Existence of periods of market inefficiency for meme stocks.
MFVI can overestimate predictive variance compared to the exact posterior
problem MFVI underestimates posterior variance
method Analyzing conjugate Bayesian Linear Regression
result MFVI can overestimate predictive variance compared to the exact posterior
Method selects number of communities in weighted networks.
problem Selecting the number of communities in weighted networks.
method Proposes a novel weighted DCSBM and uses a sequential testing framework with spectral clustering and matrix scaling.
result Method is consistent in estimating the true number of communities under mild conditions.
Normalization effects on deep neural networks impact output variance and test accuracy.
problem The impact of normalization on deep neural networks' statistical behavior and test accuracy.
method Asymptotic expansion analysis of neural network's output for different γi values. result Equal γi values (one) provide the best statistical behavior and test accuracy. Develops a convex surrogate for variance in risk minimization.
problem Balancing approximation and estimation error in optimization.
method Combines distributionally robust optimization and empirical likelihood.
result Shows faster convergence rates than empirical risk minimization.
Kernel-based tests for shape constraints in finance.
problem Enforcing shape relations on latent functions in financial econometrics.
method Kernel-based nonparametric framework for mean-variance optimization.
result Established statistical properties and a joint Wald-type statistic for testing shape constraints.
We find an unbiased estimator for MMD variance.
problem Efficiently estimating the variance of MMD estimators.
method Extending and correcting previous work, we derive an unbiased estimator for MMD variance.
result We provide a truly unbiased estimator for MMD variance with no additional computational cost.
Unified method for MMD variance estimation improves accuracy and computational efficiency.
problem Variance estimation for MMD in nonparametric testing.
method Unified finite-sample characterization of MMD variance through U-statistic and Hoeffding decomposition; exact acceleration method for univariate case.
result Unified estimators improve accuracy and computational efficiency for MMD variance.
New Riemannian optimization improves variance estimation in mixed models.
problem Challenges in estimating variance parameters in linear mixed models due to constraints.
method Formulated as an optimization problem on a Riemannian manifold, using Riemannian gradient and Hessian.
result Yields higher quality variance parameter estimates compared to existing methods.
New method reduces variance in complex probabilistic model optimization.
problem High variance in stochastic optimisation of complex models.
method Use recognition network to approximate optimal control variate for each mini-batch.
result Sub-optimal variance reduction is improved with new approach.
Study shows variance gamma model outperforms Black-Scholes for USD-INR currency options.
problem Complex pricing of currency options with multi-assets.
method Examined USD-INR currency options, tested several models, compared performance.
result Variance gamma model outperforms Black-Scholes model in various volatility regimes.
Method clusters molecular systems based on dynamics or structure similarity.
problem Clustering molecular systems based on dynamics or structure similarity.
method Ward's minimum variance clustering using Jensen-Shannon divergence.
result Method avoids overfitting in supervised learning.