Reduces test set maintenance effort by 80-100%.
problem Lack of proper and up-to-date test sets in real-world scenarios.
method Simple technique to reduce labeling effort.
result Significant reduction in test set maintenance effort (80-100%).
Reduces identity testing of reversible Markov chains to simpler symmetric chain tests.
problem Testing identity of reversible Markov chains from a single trajectory.
method Using lumping-congruent Markov embeddings, the problem is simplified to testing symmetric chains over a larger state space.
result Achieves state-of-the-art sample complexity for identity testing.
We present a new test for studying asphericity and diagrammatic reducibility of group presentations. Our test can be applied to prove diagrammatic reducibility in cases where the classical weight test fails. We use this criterion to generalize results of J. Howie and S.M. Gersten on asphericity of LOTs and of Adian pre…
Introduces directed diagrammatic reducibility with group and topological implications.
problem Diagrammatic reducibility in relative presentations.
method Adapting classical tools for diagrammatic reducibility to directed diagrammatic reducibility.
result Strong group theoretic and topological consequences of directed diagrammatic reducibility.
Active testing reduces label costs for efficient model evaluation.
problem Real-world applications require expensive test labels, disconnecting from existing model evaluation methods.
method Derives acquisition strategies to select test points efficiently, addressing label bias and variance.
result Active testing improves model evaluation efficiency without sacrificing accuracy.
Improves test set performance and reduces out-of-sample disappointment for unstable models.
problem Ensuring strong test set performance via cross-validation for unstable models.
method Nested k-fold cross-validation with hyperparameter selection based on a weighted sum of cross-validation metric and model stability measure.
result Improves out-of-sample MSE for sparse ridge regression and CART by 4% and 2% respectively, compared to k-fold cross-validation.
A new algorithm reduces CI tests for causal graph recovery.
problem Exponential CI tests limit causal discovery algorithms.
method CCPG (Causal Consistent Partition Graph) with polynomial CI tests.
result CCPG efficiently recovers causal graph with polynomial tests.
A new method reduces computational costs for testing RF variable importance measures.
problem Testing variable importance measures from random forests is computationally expensive and challenging.
method Sequential permutation testing and sequential p-value estimation to reduce computational costs.
result Theoretical properties of sequential tests are confirmed, maintaining type-I error and high power.
Reduces deep learning training data for faster testing.
problem Resource-intensive deep learning training with full data sets.
method Evaluated different training set reduction methods.
result Training set reduction is useful in resource-constrained environments.
Study suggests variable selection may not significantly reduce power in multivariate tests.
problem The feasibility of parsimonious variable selection in Hotelling's T2-test.
method Investigation of power loss when selecting small subsets of variables from multivariate data.
result Some evidence suggests no significant power loss over a wide range of alternatives.
E-CIT framework reduces CITs' computational burden and improves causal discovery performance.
problem High computational cost of traditional CITs in causal discovery.
method E-CIT framework using divide-and-aggregate strategy with stable distribution p-value combination.
result Significant reduction in computational burden and competitive performance in causal discovery.
Study uses random matrix test to find significant factors in cryptocurrency forecasts.
problem Determining the optimal number of factors in cryptocurrency forecast models.
method Applied a random matrix test to a forecast model of Reduced Rank Regression (RRR) on cryptocurrencies.
result Consistent results with visual inspection, minimal computational cost compared to cross-validation.
Hypothesis testing in singular models is fundamentally about identifiable vs. non-identifiable parameters.
problem Testing in singular models is inherently problematic due to non-identifiability and degeneracy of Fisher information.
method Formalized the overlap obstruction and showed that hypotheses over non-identifiable parameters are untestable, while those over identifiable parameters reduce to classical testing.
result Hypotheses over non-identifiable parameters are untestable, while those over identifiable parameters reduce to classical testing.
Regularizing for or against class selectivity in DNNs improves test accuracy.
problem The necessity and sufficiency of class selectivity in DNNs.
method Direct regularization of class selectivity in convolutional neural networks.
result Reducing class selectivity improves test accuracy, while increasing it decreases it.
Enhances image quality to improve test-time adaptation accuracy.
problem Reducing accuracy loss due to distribution shift in deep networks.
method Integrates image enhancement with TTA methods to reduce prediction uncertainty.
result TECA method increases accuracy of TTA methods without hyperparameters.
Paper introduces MultiTargeted testing, improving adversarial testing efficiency.
problem Improving efficiency and effectiveness of adversarial testing methods.
method Introduces MultiTargeted testing, using alternative surrogate losses.
result MultiTargeted outperforms other PGD-based methods, requiring fewer iterations.
PPAT uses predictions to improve risk estimation in active testing.
problem Exploiting informative predictions from black-box models for efficient risk estimation.
method Combines LURE estimator with prediction-powered control variate.
result PPAT outperforms existing methods in risk estimation and uncertainty quantification.
Optimizes group testing for COVID-19 to reduce test numbers.
problem Minimizing tests for accurate infection detection.
method Bayesian approach with genetic algorithms and sub-modularity.
result Greedy-adaptive method provides theoretical guarantees.
Reduces quantifier variance with accuracy optimization of base classifier.
problem Minimizing quantifier variance under prior probability shift.
method Optimizes the Brier score of a base classifier for training data.
result Optimizing Brier score on training data reduces quantifier variance on test data.
In quantitative finance, we often fit a parametric semimartingale model to asset prices. To ensure our model is correct, we must then perform goodness-of-fit tests. In this paper, we give a new goodness-of-fit test for volatility-like processes, which is easily applied to a variety of semimartingale models. In each cas…
Research benchmarks LLMs in medical domain to reduce hallucinations.
problem Hallucinations in medical LLMs can lead to incorrect information.
method Developed Med-HALT dataset and testing methods.
result Significant performance differences among LLMs identified.
We introduce online learning algorithms which are independent of feature scales, proving regret bounds dependent on the ratio of scales existent in the data rather than the absolute scale. This has several useful effects: there is no need to pre-normalize data, the test-time and test-space complexity are reduced, and t…
We introduce online learning algorithms which are independent of feature scales, proving regret bounds dependent on the ratio of scales existent in the data rather than the absolute scale. This has several useful effects: there is no need to pre-normalize data, the test-time and test-space complexity are reduced, and t…
This paper discusses how usage patterns and preferences of inhabitants can be learned efficiently to allow smart homes to autonomously achieve energy savings. We propose a frequent sequential pattern mining algorithm suitable for real-life smart home event data. The performance of the proposed algorithm is compared to …
aLTT selects hyperparameters efficiently with statistical guarantees.
problem Statistical validity and efficiency in hyperparameter selection.
method Sequential data-dependent multiple hypothesis testing with early termination.
result Reduces testing rounds while maintaining statistical validity.
We present a novel Metropolis-Hastings method for large datasets that uses small expected-size minibatches of data. Previous work on reducing the cost of Metropolis-Hastings tests yield variable data consumed per sample, with only constant factor reductions versus using the full dataset for each sample. Here we present…
New tests detect asphericity in complex pairs, simplifying previous proofs.
problem Detecting asphericity in complex pairs (L,K) where K is a subcomplex of L. method Developed relative weight tests for injective labeled oriented trees.
result Injective labeled oriented trees are aspherical, strengthening previous results.
Paper proposes a new method for more accurate group testing of infected patients.
problem Identifying infected patients efficiently with reduced tests and corrected errors.
method Adaptive design of pools based on Bayesian posterior prediction using belief propagation algorithm.
result The proposed method results in more accurate identification of infected patients.
TTT improves transformer models for in-context learning.
problem Improving transformer models for efficient in-context learning.
method Gradient-based TTT method for linear transformers, with theoretical and empirical analysis.
result TTT significantly reduces the sample size required for in-context learning.
We first study superpolynomial associated to triply-graded reduced colored HOMFLY-PT homology. We propose conjectures of congruent relations and cyclotomic expansion for it. We prove conjecture of N=1 for torus knot case, through which we obtain the corresponding invariant α(T(m,n))=−(m−1)(n−1)/2. This is closely r…
Given a polarized complex manifold, projection of a torus-equivariant test configuration to holomorphic vector fields was introduced by G. Székelyhidi, as the limit of the associated C∗-actions. We show that there actually holds the moment convergence of the weight distributions. Our analytic approach at th…
Quick and accurate medical diagnosis is crucial for the successful treatment of a disease. Using machine learning algorithms, we have built two models to predict a hematologic disease, based on laboratory blood test results. In one predictive model, we used all available blood test parameters and in the other a reduced…
Neuron-specific dropout reduces overfitting and data needs for neural networks.
problem Overfitting and insufficient training data for deep neural networks.
method Compares training and validation passes of a layer, drops targeted neurons based on feature analysis.
result Achieves similar or better testing accuracy with less data, reducing overfitting.
Geodesic rays of class C^{1,1} are constructed for any test configuration of a positive line bundle L on X using resolution of singularities. The construction reduces to finding a subsolution of the corresponding Monge-Ampere equation. Geometrically, this is accomplished by the use a positive line bundle on the resolut…
A cascaded autoencoder defends machine learning models from adversarial attacks.
problem Adversarial attacks on machine learning models.
method Denoising and dimensionality reduction using cascaded autoencoders.
result Preprocessed data with cascaded autoencoder pipeline improves model accuracy against adversarial perturbations.
It is common practice to decay the learning rate. Here we show one can usually obtain the same learning curve on both training and test sets by instead increasing the batch size during training. This procedure is successful for stochastic gradient descent (SGD), SGD with momentum, Nesterov momentum, and Adam. It reache…
Adaptive sequential testing optimizes epidemic control by learning optimal test strategies.
problem Optimizing test allocation in epidemics with network and temporal dependence.
method Adaptive sequential design with Online Super Learner for optimal test strategies.
result Superior performance in simulated university COVID-19 pandemic.
New framework tests mean-variance spanning in high dimensions.
problem Testing mean-variance spanning in high-dimensional asset spaces.
method Robust Student-t statistic based on batch-mean method, combined using Cauchy combination test.
result Advantages of diversification vary by economic conditions and cross-country.
Tent adapts models during testing by minimizing entropy of predictions.
problem Adapting models to new data during testing with limited information.
method Test entropy minimization (tent) and online channel-wise affine transformations.
result Reduces generalization error on various datasets and benchmarks.
Fractal analysis is carried out on the stock market indices of seven European countries and the US. We find evidence of long range dependence in the log return series of the Mibtel (Italy) and the PX Glob (Czech Republic). Long range dependence implies that predictable patterns in the log returns do not dissipate quick…
We study the problem of structured prediction under test-time budget constraints. We propose a novel approach applicable to a wide range of structured prediction problems in computer vision and natural language processing. Our approach seeks to adaptively generate computationally costly features during test-time in ord…
A new method reduces CI tests for causal structure learning.
problem Exponential CI tests in constraint-based methods.
method Recursive Markov boundary-based approach.
result Significantly reduces CI tests compared to existing methods.
A new kernel test reduces noise in MMD by focusing on leading eigen-directions.
problem Noise in trailing directional components degrades power of standard kernel two-sample tests.
method Truncate MMD spectral decomposition, retaining only leading eigen-directions.
result Our method achieves superior power and robustness, especially in high-dimensional and unbalanced settings.
A method for safe online classification reduces test costs while maintaining low error rates.
problem Sequential testing for binary disease outcomes with unknown logistic model parameters.
method Joint estimation of logistic parameter and feature distribution with a conservative threshold.
result Achieves target error with high probability and requires minimal excess tests.
Tests validity of DML estimators without assumptions.
problem Validating DML estimators without making assumptions.
method Develops tests to falsify assumptions for DML estimators.
result Falsifies assumptions for DML estimators with non-trivial power.
Prototype selection improves DS techniques' accuracy and reduces computational cost.
problem Improving the performance of dynamic selection techniques.
method Prototype selection techniques that edit validation data to remove noise and redundant instances.
result Improves DS techniques' classification accuracy and reduces computational cost.
Sequential Kernel-based Conditional Independence Testing via Adaptive Betting
problem Testing conditional independence
method Testing-by-betting on an adaptively optimized Kernel Conditional Independence statistic
result Significantly reduces Type I error inflation while preserving high power
GaussDetect-LiNGAM eliminates Gaussianity tests for causal discovery.
problem Causal direction identification without Gaussianity assumptions.
method Leverages the equivalence between noise Gaussianity and residual independence in reverse regression.
result Gaussianity tests replaced with robust kernel-based independence tests.