The paper improves high-dimensional linear regression prediction and estimation using auxiliary samples.
problem Estimating and predicting high-dimensional linear regression models with auxiliary samples.
method Proposes Trans-Lasso for data-driven transfer learning, establishing optimality for prediction and estimation.
result Knowledge from auxiliary samples can improve learning performance in target problems.
GEAR uses auxiliary data to estimate optimal decisions in studies with limited primary outcomes.
problem Estimating optimal decisions when primary outcomes are not available in experimental samples.
method GEAR uses augmented inverse propensity weighting to estimate optimal decisions based on auxiliary data.
result GEAR estimators and value estimators have established asymptotic properties and are validated in simulations and a real application.
New sampling algorithms improve efficiency in latent Gaussian models.
problem Efficient sampling from complex target distributions.
method Combines auxiliary variables, Gibbs sampling, and Taylor expansions.
result Marginal samplers are superior in asymptotic variance, but slower in computing time.
Bayesian optimization with faster convergence without auxiliary optimization.
problem Time-consuming and hard to implement Bayesian optimization methods.
method Eliminates auxiliary optimization and delta-cover sampling requirements.
result Achieves exponential convergence rate.
Proposes a Thompson sampling algorithm for multi-objective contextual bandit problems with auxiliary constraints.
problem Real-world applications with multiple competing objectives and auxiliary constraints.
method Thompson sampling algorithm for multi-outcome contextual bandit problems with auxiliary constraints.
result Empirically evaluated and applied to a real-world video transcoding problem.
ScoreFusion fuses multiple diffusion models to enhance generative modeling of a target population.
problem Enhancing generative modeling of a target population with limited data.
method ScoreFusion uses KL barycenters of auxiliary populations and recasts the learning problem as score matching in denoising diffusion.
result ScoreFusion achieves a dimension-free sample complexity bound in total variation distance.
Learn dynamics of a system using auxiliary data from similar systems.
problem Learning dynamics of a linear system with limited data.
method Weighted least squares approach, incorporating auxiliary data.
result Auxiliary data can help reduce intrinsic error due to noise.
The paper introduces methods to solve optimization problems with auxiliary data.
problem Solving multistage optimization problems with uncertain data and auxiliary information.
method Utilizes machine learning techniques like kNN, CART, and RF to develop methods for optimization.
result Demonstrates asymptotic and finite sample optimality of the proposed methods.
New MCMC methods use auxiliary variables to sample from intractable distributions.
problem Sampling from distributions with unknown normalizing constants.
method Unified Markov chain Monte Carlo framework with auxiliary variables.
result New algorithms outperform existing methods on synthetic and real datasets.
Study improves treatment effect estimation using unlabeled covariates.
problem Estimating treatment effects with limited labeled data.
method Developed efficiency bounds and estimators for semi-supervised setting.
result Estimators using unlabeled covariates have lower asymptotic variance.
Meta two-sample testing uses auxiliary data to quickly find powerful tests from limited samples.
problem Challenges in identifying powerful kernels for distinguishing complex distributions with limited data.
method Introduces meta two-sample testing (M2ST) to leverage abundant auxiliary data on related tasks.
result Proposed algorithms improve over baselines and identify powerful tests from scarce observations.
Improved malware detection by adding auxiliary loss terms to a neural network.
problem Malware detection accuracy with a single label.
method Fit deep neural networks to multiple auxiliary prediction targets derived from metadata.
result Significant improvement in detection performance, reducing false negatives by 42.6% at a low false positive rate.
A novel kernel learning framework detects abrupt changes in time series data.
problem Detecting abrupt changes in time series data with fewer assumptions.
method KL-CPD, a novel kernel learning framework that optimizes a lower bound of test power via an auxiliary generative model.
result Significantly outperformed other state-of-the-art methods in benchmark datasets and simulation studies.
BiCoGAN improves cGANs by disentangling latent and auxiliary variables.
problem Improving disentanglement of latent and auxiliary variables in cGANs.
method BiCoGAN uses bidirectional training with extrinsic factor loss and dynamically-tuned importance weight.
result BiCoGAN encodes auxiliary variables more accurately and disentangles latent and auxiliary variables effectively.
Boosts A/B test precision using auxiliary data from historical users.
problem Small sample sizes and imprecise estimates in A/B tests.
method Coupling design-based causal estimation with machine-learning models of historical user data.
result Effect estimates using auxiliary data are roughly equivalent to increasing sample size by 20%, or up to 50-80% in some cases.
New methods improve sampling from complex dynamical models.
problem Sampling from high-dimensional, non-linear latent dynamical models is computationally challenging.
method Introduce auxiliary MCMC and Particle Gibbs samplers with improved performance and parallelisation.
result Enhanced samplers maintain performance in high-dimensional latent spaces and support parallelisation.
We introduce a new family of estimators for unnormalized statistical models. Our family of estimators is parameterized by two nonlinear functions and uses a single sample from an auxiliary distribution, generalizing Maximum Likelihood Monte Carlo estimation of Geyer and Thompson (1992). The family is such that we can e…
Pseudo-marginal Metropolis-Hastings (pmMH) is a powerful method for Bayesian inference in models where the posterior distribution is analytical intractable or computationally costly to evaluate directly. It operates by introducing additional auxiliary variables into the model and form an extended target distribution, w…
AutoSeM automatically selects and balances auxiliary tasks in MTL.
problem Choosing and balancing auxiliary tasks in MTL.
method AutoSeM uses a Beta-Bernoulli multi-armed bandit with Thompson Sampling for task selection and a Gaussian Process for learning the mixing ratio.
result AutoSeM achieves significant performance boosts on GLUE language understanding tasks.
A new Gibbs sampler method speeds up Bayesian inference.
problem Efficient sampling from complex posterior distributions.
method Recycling auxiliary samples within Gibbs estimators.
result Significant improvement in accuracy and computational efficiency.
New algorithm samples neural network posteriors efficiently.
problem Challenges of sampling multimodal Bayesian posteriors for neural networks.
method Greedy Bayes method using log-concave coupling of posterior and auxiliary random variable.
result Log-concave coupling facilitates efficient sampling of neuron weights.
XMixup improves transfer learning accuracy by 1.9% with less training time.
problem Efficiently transfer knowledge from large source datasets to target tasks with small samples.
method Cross-domain Mixup technique that selects auxiliary samples from source datasets and augments training samples via mixup strategy.
result Improves accuracy by 1.9% on average over six real-world transfer learning datasets.
Transforms conditional density estimation into a nonparametric regression problem.
problem Conditional density estimation in high dimensions.
method Introduces auxiliary samples to transform into nonparametric regression.
result Estimator converges to true conditional density in data limit.
TAC-GAN improves image diversity in AC-GAN by minimizing class distribution divergence.
problem Low diversity in AC-GAN's generated samples as class count increases.
method TAC-GAN introduces twin auxiliary classifiers to address class separability issues.
result TAC-GAN effectively minimizes divergence between generated and real data distributions.
New method improves understanding of machine learning model performance.
problem Understanding how well machine learning models generalize from training data to unseen data.
method Auxiliary Distribution Method to derive new generalization error bounds.
result Upper bounds on generalization errors are tighter and more applicable.
Enhances network embedding with auxiliary info using matrix factorization.
problem Lack of flexible incorporation of auxiliary info (content and labels) in network embedding.
method Explicit matrix factorization incorporating structure, content, and label info.
result Unified framework for learning network embedding with structure, content, and label info.
New framework for sequential experiments with unknown data arrival.
problem Sequential decision-making with unknown information arrival.
method Generalized MAB framework for arbitrary arrival processes.
result Upper and lower bounds on minimax complexities.
Bayesian deep learning improves geostatistical mapping with auxiliary data.
problem Traditional geostatistical methods are limited in feature learning and uncertainty estimation.
method Deep neural networks learn complex relationships from auxiliary data for probabilistic mapping.
result Deep learning produces detailed, probabilistic maps with uncertainty estimates.
Deep neural networks improve two-sample testing.
problem Efficiently distinguishing between two unknown distributions.
method Deep learning representations for two-sample testing.
result Significant reduction in type-2 error rate compared to existing methods.
We propose a simulation method for multidimensional Hawkes processes with differing decays.
problem Simulating and calibrating Hawkes processes with various decay rates.
method Superposition theory of point processes, decomposition of inter-arrival times, auxiliary variables, Gibbs samplers, adaptive rejection sampling.
result Significant improvement in algorithm speed and accurate simulation of Hawkes processes.
Trans-Ising combines auxiliary datasets to estimate high-dimensional Ising models.
problem Limited target sample sizes and difficulty in using auxiliary binary datasets of unknown relevance.
method Trans-Ising uses a loss-based source screening rule and a two-stage estimation procedure.
result Trans-Ising achieves lower estimation errors than target-only estimation and naive data pooling.
HydaLearn dynamically adjusts task weights for better MTL performance.
problem Constant loss weights in MTL lead to poor results due to drifting relevance and varying mini-batch composition.
method HydaLearn uses mini-batch gradients to dynamically adjust task weights.
result HydaLearn improves performance on synthetic and real-world data.
Study uses auxiliary data to estimate system dynamics, reducing noise error.
problem Estimating system dynamics from similar but not identical systems.
method Weighted least squares approach with performance guarantees.
result Effective use of auxiliary data reduces estimation error due to process noise.
CCAC calibrates DNN classifiers on OOD datasets by separating mis-classified samples.
problem Calibrating DNN classifiers on out-of-distribution datasets is challenging.
method CCAC introduces an auxiliary class to map DNN output to calibrated confidence, separating mis-classified from correctly classified samples.
result CCAC consistently outperforms prior methods on various DNN models, datasets, and applications.
HiSS sampling overcomes local mode traps in rugged discrete spaces.
problem Sampling multimodal discrete distributions with gradient-based methods.
method Integrates Metropolis-within-Gibbs framework with logistic convolution.
result HiSS outperforms alternatives on various tasks, including Ising models and binary neural networks.
Privileged Information Dropout improves RL performance without distillation.
problem Improving sample efficiency and performance in reinforcement learning.
method Introducing Privileged Information Dropout to directly incorporate privileged information into RL agent inputs.
result Privileged Information Dropout outperforms distillation and auxiliary tasks in a partially-observed environment.
Generative model disentangles dark matter halo properties.
problem Entangling physical factors in generative model latent spaces.
method Auxiliary-variable-guided framework with halo mass and concentration.
result Reveals mass-concentration scaling relation and identifies unusual halo formation.
The paper develops a minimax optimal method for high-dimensional regression using auxiliary data.
problem High-dimensional additive regression with heavy-tailed errors and transfer learning.
method Smooth backfitting estimator with local linear smoothing, followed by a two-stage estimation method.
result The method achieves the minimax optimal rate under certain conditions.
NCDSSM models irregularly sampled time series with improved imputation and forecasting.
problem Accurate modeling of irregularly sampled time series with missing observations.
method Neural Continuous-Discrete State Space Model (NCDSSM) with amortized inference for auxiliary variables and flexible dynamic state parameterizations.
result Improved imputation and forecasting performance on multiple benchmark datasets.
New MCMC method speeds up quantum physics simulations by a factor of 100.
problem Simulating quantum many-body systems with high computational complexity.
method FFT-accelerated MCMC with coupled particle and auxiliary variables.
result Achieves O(NlogN) scaling, significantly faster than traditional O(N3) methods. ATLAS separates invariant and transferable latent factors across diverse environments.
problem Transfer learning and robust prediction in heterogeneous environments.
method ATLAS leverages invariance principle to disentangle latent factors and uses auxiliary labels for robust prediction.
result Near-oracle performance and robust transferable prediction in new environments.
A new method for generating samples in multi-class scenarios using GANs and classifiers.
problem Generating samples for specific classes in multi-class scenarios.
method Versatile Auxiliary Classifier with Generative Adversarial Network (VAC+GAN) method where the generator is conditional and the classification error is backpropagated.
result The method improves sample generation for specific classes in multi-class scenarios.
New method finds balanced clusters in graphs using auxiliary information.
problem Finding balanced clusters in graphs with population-level constraints.
method Proposes individual-level balancing constraint and develops spectral clustering algorithms.
result Establishes first statistical consistency result for constrained spectral clustering.
Random forests and LASSO methods improve small area estimation using auxiliary data.
problem Estimating household consumption in small areas with limited sampled data.
method Model-based small area estimation using random forests and LASSO with auxiliary information.
result Bayesian shrinkage performed best in terms of bias, MSE, and prediction interval coverages.
AuxiLearn combines auxiliary tasks into a single loss function.
problem Improving neural network performance on a main task using auxiliary tasks.
method Implicit differentiation to learn a network that combines auxiliary tasks into a single coherent objective function.
result AuxiLearn consistently outperforms competing methods in various tasks and domains.
New method learns auxiliary labels automatically for improved generalisation.
problem Improving generalisation in supervised learning without additional data.
method Trains two neural networks: label-generation and multi-task networks.
result MAXL outperforms single-task learning on 7 image datasets.
New method uses correlated auxiliary feedback to reduce regret in parameterized bandits.
problem Reducing regret in parameterized bandits with correlated auxiliary feedback.
method Develops a reward estimator using auxiliary feedback with tight confidence bounds.
result Shows significant reduction in regret compared to standard methods.
ADRL improves participant selection in MCS systems.
problem Designing a participant selection algorithm for different MCS systems with multiple goals.
method Auxiliary-task based deep reinforcement learning (ADRL) using transformers and pointer networks.
result ADRL outperforms other baselines in various MCS settings.