Unified framework for estimating high-dimensional conditional factor models.
problem Estimating high-dimensional conditional latent factor models with practical limitations.
method Constrained nuclear norm regularization and cross-validation for parameter selection.
result Imposing homogeneity improves model predictability, with new method outperforming alternatives.
Sparse APCA identifies sparse factors in financial returns over time.
problem Analyzing co-movements of high-dimensional panel data over time.
method Sparse asymptotic PCA with truncated power method for sparse factors and sequential deflation for multi-factor cases.
result Identification of nine risk factors influencing the S&P 500 stock market.
Method recovers causal diffusion mechanisms from steady-state data without parametric assumptions.
problem Recovering causal diffusion mechanisms from steady-state gene expression data.
method Non-parametric kernel estimator for drift function, cross-validation for hyperparameter tuning.
result Full causal mechanism can be non-parametrically identified under weak non-explosion criterion.
The paper defines cross-section continuity for angular momentum definitions and finds the CWY definition valid.
problem Defining angular momentum at null infinity and ensuring its continuity across different cross-sections.
method Introducing cross-section continuity as a criterion and proving it for specific angular momentum definitions.
result The Chen-Wang-Yau definition of angular momentum satisfies cross-section continuity, while the Compere-Nichols modification does not.
We resolve parts (A) and (B) of Problem 1.100 from Kirby's list by showing that many nontrivial links arise as cross-sections of unknotted holomorphic disks in the four-ball. The techniques can be used to produce unknotted ribbon surfaces with prescribed cross-sections, including unknotted Lagrangian disks with nontriv…
The study identifies flat manifolds with unique cusp cross-sections in arithmetic hyperbolic manifolds.
problem Characterizing flat manifolds that have unique cusp cross-sections in arithmetic hyperbolic manifolds.
method Algebraic characterization of cusp cross-sections in arithmetic hyperbolic manifolds.
result Construction of flat manifolds with unique cusp cross-sections and proof of their existence in all dimensions n≥32. New algorithm improves asset ranking for better cross-sectional portfolios.
problem Sub-optimal ranking of assets in cross-sectional systematic strategies.
method Learning-to-rank algorithms to enhance portfolio construction.
result Modern machine learning ranking algorithms boost Sharpe Ratios by approximately threefold.
Set-Sequence model learns cross-sectional dynamics directly from time series data.
problem Predicting large cross-sections of time series data with latent cross-sectional dynamics.
method A model that learns cross-sectional structure directly, enhancing expressivity and eliminating manual feature engineering.
result Significantly outperforms strong baselines in equity portfolio optimization and loan risk prediction.
Conditions for flat manifolds as cusp cross-sections in arithmetic hyperbolic manifolds.
problem Determining when a flat manifold can be a cusp cross-section in arithmetic hyperbolic manifolds.
method Analyzing rational representations of holonomy groups and quasi-arithmetic manifolds.
result Conditions for a flat manifold to appear as a cusp cross-section in every commensurability class of arithmetic hyperbolic manifolds.
TQA improves prediction intervals for time series data by adjusting quantiles for both cross-sectional and longitudinal coverage.
problem Constructing reliable prediction intervals for cross-sectional time series data.
method Temporal Quantile Adjustment (TQA) method that adjusts the quantile in Conformal Prediction to account for both cross-sectional and longitudinal coverage.
result TQA improves longitudinal coverage while preserving cross-sectional coverage, as validated through extensive experimentation.
The paper evaluates forecast accuracy of realized volatility measures in large cross-sections.
problem Forecast evaluation of realized volatility measures in large cross-sections of financial data.
method Equal predictive accuracy testing procedures, LASSO shrinkage, measurement error correction, cross-sectional jump component measures.
result The augmented HAR model outperforms the standard HAR model in forecasting realized volatility.
CPTD improves prediction intervals in time series regression with cross-sectional data.
problem Constructing valid prediction intervals in time series regression with a cross-section.
method Conformal Prediction with Temporal Dependence (CPTD) for post-hoc, light-weight approach.
result CPTD maintains cross-sectional validity while improving longitudinal coverage.
Unified model learns from both time-series and cross-sectional momentum features.
problem Separate time-series and cross-sectional momentum strategies do not consider concurrent relationships.
method Spatio-Temporal Momentum strategies using neural networks to combine both types of momentum.
result Simple neural network with single fully connected layer generates trading signals for all assets.
Study on stability of surfaces in null cones under area-preserving variations.
problem Investigating stability of spacelike cross sections of null cones.
method Area-preserving variations, Hawking energy analysis, spherical cross sections.
result Only round spheres are stable cross sections of the standard Minkowski lightcone.
Motivated by a question of Hirzebruch on the possible topological types of cusp cross-sections of Hilbert modular varieties, we give a necessary and sufficient condition for a manifold M to be diffeomorphic to a cusp cross-section of a Hilbert modular variety. Specialized to Hilbert modular surfaces, this proves that e…
In this article, we derive concentration inequalities for the cross-validation estimate of the generalization error for empirical risk minimizers. In the general setting, we prove sanity-check bounds in the spirit of \cite{KR99} \textquotedblleft\textit{bounds showing that the worst-case error of this estimate is not m…
Deep learning predicts cross-sectional stock prices for practical investment.
problem Predicting stock prices using cross-sectional factors.
method Deep learning model for daily stock price prediction.
result Profitable investment framework demonstrated in Japanese stock market.
In this article, we derive concentration inequalities for the cross-validation estimate of the generalization error for stable predictors in the context of risk assessment. The notion of stability has been first introduced by \cite{DEWA79} and extended by \cite{KEA95}, \cite{BE01} and \cite{KUNIY02} to characterize cla…
Classifies Nil 3-manifolds as cross-sections of complex hyperbolic surfaces.
problem Identifying Nil 3-manifolds as cross-sections of complex hyperbolic surfaces.
method Comprehensive classification of commensurability classes of cusped, arithmetic, and non-arithmetic complex hyperbolic 2-manifolds.
result Some Nil 3-manifolds are cross-sections in every commensurability class, while others are cross-sections in only one.
In this article, we derive concentration inequalities for the cross-validation estimate of the generalization error for subagged estimators, both for classification and regressor. General loss functions and class of predictors with both finite and infinite VC-dimension are considered. We slightly generalize the formali…
Study predicts risk of true-lumen narrowing after ATAAD surgery using CT data.
problem Early post-surgery risk assessment for aortic dissection patients.
method Retrospective study with CT data, derived cross-sectional shapes, form factor (FF) for morphology assessment, linear discriminant analysis (LDA) for risk classification, LOPO-CV for prediction.
result Machine-learning model accurately predicts risk for all high-risk patients and low-risk patients, potentially reducing hospital visits.
Machine learning portfolios perform well with simple imputation of missing data.
problem Handling missing values in machine learning portfolios constructed from cross-sectional return predictors.
method Simple imputation with cross-sectional means compared to rigorous expectation-maximization methods.
result Simple imputation performs well due to the structure of missing data.
When selecting a classification algorithm to be applied to a particular problem, one has to simultaneously select the best algorithm for that dataset \emph{and} the best set of hyperparameters for the chosen model. The usual approach is to apply a nested cross-validation procedure; hyperparameter selection is performed…
Simple bounds show most cross-sectional predictability findings are likely true.
problem Determining the validity of cross-sectional return predictability findings.
method Developed simple and intuitive bounds on the false discovery rate (FDR).
result Bounds show the FDR is small, indicating most findings are likely true.
A fast bootstrap method estimates cross-validation standard error.
problem Uncertainty quantification in cross-validation estimates.
method Random-effects model to estimate variance component.
result Valid confidence intervals for model performance.
Paper uses HPCA for better stock correlation modeling.
problem Challenges in modeling cross-sectional correlations between thousands of stocks.
method Hierarchical Principal Component Analysis (HPCA) and statistical clustering.
result HPCA provides better cross-sectional correlations than classic PCA.
New cross-validation methods for Gaussian process regression with efficient gradient computation.
problem Estimating parameters of Gaussian process covariance functions.
method Derive new cross-validation criteria and efficient adjoint computation of gradients.
result Efficient method for evaluating cross-validation criteria and their gradients.
PRISM-VQ combines financial priors with vector quantization for better stock prediction.
problem Predicting cross-sectional stock returns is hard due to low signal-to-noise ratios and changing market conditions.
method Integrates expert priors, vector-quantized latent factors, and dynamic factor loadings.
result Consistent improvements in cross-sectional return prediction and portfolio performance.
The present paper is devoted to some results concerning with the complete lifts of an almost complex structure and a connection in a manifold to its (0,q)-tensor bundle along the corresponding cross-section.
Pricing assets has attracted significant attention from the financial technology community. We observe that the existing solutions overlook the cross-sectional effects and not fully leveraged the heterogeneous data sets, leading to sub-optimal performance. To this end, we propose an end-to-end deep learning framework t…
Weyl's tube formula holds for various cross-sections under symmetry conditions.
problem Can the volume of tubes around submanifolds be calculated for non-round cross-sections?
method Investigated the volume of tubes with general cross-sections D under symmetry conditions.
result The volume of tubes around submanifolds can be calculated for general cross-sections under symmetry conditions.
The main purpose of present paper is to study the affine connection induced from the horizontal lift on the cross-section determined by a vector field in Mn with respect to the adapte frame of .
Cross-validation estimates model performance on unseen data, not training data.
problem Understanding how cross-validation estimates prediction error and its limitations.
method Analyzing linear models and popular prediction error estimates, introducing nested cross-validation.
result Cross-validation estimates the average prediction error of models fit on other unseen training sets, not the model at hand.
Modeling how individuals evolve over time is a fundamental problem in the natural and social sciences. However, existing datasets are often cross-sectional with each individual observed only once, making it impossible to apply traditional time-series methods. Motivated by the study of human aging, we present an interpr…
Improves test set performance and reduces out-of-sample disappointment for unstable models.
problem Ensuring strong test set performance via cross-validation for unstable models.
method Nested k-fold cross-validation with hyperparameter selection based on a weighted sum of cross-validation metric and model stability measure.
result Improves out-of-sample MSE for sparse ridge regression and CART by 4% and 2% respectively, compared to k-fold cross-validation.
Pipeline integrates cross-sectional and longitudinal multi-omics data for IBD research.
problem Integrating diverse data types from the same individuals for disease understanding.
method Statistical and deep learning methods for variable selection, feature extraction, and joint integration.
result Identified microbial pathways, metabolites, and genes discriminating IBD status.
The paper develops a cross-validation method for improving signal denoising techniques.
problem Improving signal denoising methods for nonparametric regression.
method Develops a general cross-validation framework for signal denoising and applies it to Trend Filtering and Dyadic CART.
result Cross validated versions of Trend Filtering and Dyadic CART achieve nearly optimal convergence rates.
A new model explains asset returns with a single factor, improving cross-sectional performance.
problem Understanding the cross-section of asset returns with complex models.
method Proposes a non-linear single-factor asset pricing model with a nonparametric link function estimated jointly with sieve-based estimators.
result The model delivers superior cross-sectional performance with a low-dimensional approximation of the link function.
Optimizes Lasso hyperparameters using leave-one-out CV.
problem Finding optimal hyperparameters for Lasso regression.
method Develops an algorithm to compute exact or approximate leave-one-out CV.
result Algorithm finds optimal hyperparameters for Lasso.
Study shows different types of volatility and skewness changes affect stock prices.
problem Different types of volatility and skewness changes affect stock prices.
method Used intraday data for individual stocks to analyze cross-section of asset returns.
result Idiosyncratic transitory and persistent shocks to volatility and skewness are priced differently in stock returns.
LPCI provides valid prediction intervals for longitudinal data.
problem Current conformal prediction methods for time series data lack cross-sectional coverage when applied to longitudinal datasets.
method Modeling residual data as a quantile fixed-effects regression problem, constructing prediction intervals with a trained quantile regressor.
result LPCI achieves valid cross-sectional coverage and outperforms existing benchmarks in terms of longitudinal coverage rates.
Study evaluates cross-validation methods for clinical ECG classification, finding leave-source-out more reliable.
problem Overoptimistic cross-validation estimates for new patient sources.
method Empirical evaluation of K-fold and leave-source-out cross-validation methods.
result Leave-source-out cross-validation provides more reliable performance estimates.
A new method improves super learner validation efficiency.
problem Improving the efficiency of super learner validation.
method Bootstrap Bias Corrected Cross Validation applied to Super Learning.
result Bootstrap Bias Corrected Cross Validation proved efficient and cost-effective.
Cross-validation is the workhorse of modern applied statistics and machine learning, as it provides a principled framework for selecting the model that maximizes generalization performance. In this paper, we show that the cross-validation risk is differentiable with respect to the hyperparameters and training data for …
Cross-validation methods help learn dynamical systems from data.
problem Learning surrogate models for dynamical systems from limited data.
method Variants of cross-validation (Kernel Flows, MMD, Lyapunov exponents).
result Simple approaches for kernel selection in dynamical system emulators.
This text is a survey on cross-validation. We define all classical cross-validation procedures, and we study their properties for two different goals: estimating the risk of a given estimator, and selecting the best estimator among a given family. For the risk estimation problem, we compute the bias (which can also be …
New method tests Granger non-causality in panel data with cross-sectional dependencies.
problem Testing Granger non-causality in panel data with cross-sectional dependencies.
method Proposes a new approach to aggregate p-values from panel members to test Granger non-causality, showing lower FDR.
result Our approach discovers true causal relations in panel data, unlike state-of-the-art methods.
The paper improves confidence intervals for test error using cross-validation.
problem Improving confidence intervals for test error in machine learning.
method Develops central limit theorems and consistent estimators for cross-validation.
result Provides asymptotically-exact confidence intervals and hypothesis tests.