We propose an efficient method for estimating covariate effects in doubly-stochastic spatial models.
problem Computational demands and restrictive assumptions in existing doubly-stochastic spatial models.
method Penalized regression method for estimating covariate effects in doubly-stochastic point processes.
result Consistency and asymptotic normality of the covariate effect estimates achieved despite model misspecification.
Doubly SGD improves convergence for intractable objective optimization.
problem Optimizing objectives in sum of intractable expectations.
method Doubly SGD with doubly stochastic gradients and independent minibatching.
result Established convergence of doubly SGD under general conditions, including dependent component gradient estimators.
New method improves off-policy critic evaluation in reinforcement learning.
problem High variance and instability in off-policy policy evaluation.
method Doubly robust estimators applied to actor-critic algorithms.
result Doubly robust estimation significantly improves performance in continuous control tasks.
The paper improves boundary detection and density estimation on noisy data.
problem Detecting boundary points and estimating density on noisy data from compact manifolds.
method Doubly stochastic scaling of the Gaussian heat kernel via Sinkhorn iterations.
result The new estimates of boundary points and density outperform standard methods, especially under noise.
Proposes SDRG to adjust missingness in machine learning models.
problem Systemic missingness in observational data leads to biased parameter estimation.
method Introduces SDRG using two models: weight-corrected gradients and per-covariate control variates.
result Empirically demonstrates convergence in training image classifiers with missing data.
New method reduces variance in complex probabilistic model optimization.
problem High variance in stochastic optimisation of complex models.
method Use recognition network to approximate optimal control variate for each mini-batch.
result Sub-optimal variance reduction is improved with new approach.
Efficiently approximates softmax probabilities for large-scale inference.
problem High cost of computing softmax probabilities for large-scale inference.
method Introduces a lower bound on softmax probabilities as a product of pairwise probabilities, scalable through stochastic optimization and subsampling.
result Demonstrates that the new bound has interesting theoretical properties and can be used in classification problems.
This paper discusses properties of a Doubly Stochastic Poisson Process (DSPP) where the intensity process belongs to a class of affine diffusions. For any intensity process from this class we derive an analytical expression for probability distribution functions of the corresponding DSPP. A specification of our results…
Robustly infers manifold density and geometry under high-dimensional noise.
problem Inaccurate kernel density estimation under high-dimensional noise.
method Doubly stochastic normalization of Gaussian kernel.
result Robust tools for density estimation, noise magnitude estimation, and distance approximation.
Consistency of DSDP proved with exponential convergence.
problem Consistency analysis for Doubly Stochastic Dirichlet Process.
method Proved components consistency with simulation and real-world experiments.
result Exponential convergence of posterior probability.
Doubly-stochastic normalization improves robustness to heteroskedastic noise.
problem Robustness to heteroskedastic noise in affinity matrix construction.
method Doubly-stochastic normalization of the Gaussian kernel.
result Doubly-stochastic normalization converges to clean matrix with rate m − 1 / 2 m^{-1/2} m − 1/2 under heteroskedastic noise. Paper proposes Sinkformers for Transformers with doubly stochastic attention.
problem Improving Transformer models' accuracy in vision and natural language processing.
method Using Sinkhorn's algorithm to make attention matrices doubly stochastic instead of SoftMax normalization.
result Sinkformers enhance model accuracy in vision and natural language processing tasks.
New algorithms learn graph structures privately, matching best results.
problem Private learning of graph structures with multiple blocks.
method Sum-of-squares relaxation and exponential mechanism for score function.
result Matches statistical utility of previous best non-private methods.
New method solves doubly-nonconvex composite optimization problems.
problem Solving composite optimization problems with both functions nonconvex.
method Stochastic gradient descent with quasiconvex penalty function.
result Convergence properties for doubly-nonconvex composite optimization.
Proposes a new simulator for complex arrival processes.
problem Modeling and simulating complex arrival processes with non-stationary and multi-dimensional rates.
method Integrates Monte Carlo and GANs to model a broad class of arrival processes.
result Consistent and efficient estimation of the simulator using Wasserstein distance.
Estimates outcomes under hypothetical scenarios using a flexible framework.
problem Adapting to sudden shifts in treatment patterns.
method Doubly robust estimator using incremental interventions.
result Achieves n \sqrt{n} n -consistency and asymptotic normality. Bayesian Neural Networks built block-by-block with uncertainty estimates.
problem Building interpretable and uncertainty-aware neural networks.
method Bayesian Neural Networks (BNNs) constructed using blocks, with doubly stochastic variational inference for posterior approximation.
result Uncertainty estimates provided for Bayesian Neural Networks.
DSVNP uses global and local latent variables for improved neural process predictions.
problem Limited expressiveness of vanilla neural processes in capturing target-specific local variation.
method Introduces DSVNP combining global and local latent variables for prediction.
result Competitive prediction performance in multi-output regression and uncertainty estimation.
Method uses deep learning to estimate traffic intensity.
problem Estimating stochastic intensity of traffic processes.
method Deep neural networks for nonlinear filtering.
result Deep learning method accurately estimates traffic intensity.
New methods combine machine learning with doubly robust estimators for better treatment effect estimation.
problem Estimating average treatment effects from observational data.
method Doubly robust methods using machine learning techniques.
result Machine learning improves the performance of doubly robust estimators.
The paper analyzes generalization properties of scalable kernel methods.
problem Understanding the generalization of doubly stochastic learning algorithms.
method Theoretical analysis of different variants of doubly stochastic learning algorithms in nonparametric regression.
result Derivation of generalization error convergence results for the algorithms.
A new algorithm speeds up optimal transport for machine learning.
problem Optimal transport for machine learning with additional terms.
method Forward-backward splitting algorithm based on Bregman distances.
result Significant improvement in speed and performance for domain adaptation.
Paper presents a new doubly robust estimator for survival analysis with improved consistency.
problem Consistency of doubly robust estimators in high dimensions with flexible data-adaptive methods.
method Data-adaptive regression estimators, Gaussianization, cross-fitting.
result The estimator converges at n 1 / 2 n^{1/2} n 1/2 rate for a large class of data-adaptive nuisance estimators. Two new estimators improve VAE training for hierarchical and prior parameters.
problem Efficient gradient estimation for VAEs with hierarchical and prior parameters.
method Developed two generalizations of Doubly-Reparameterized Gradient Estimators (DReGs) for VAEs.
result Improved training of conditional and hierarchical VAEs on image modeling tasks.
Graph alignment problem solved with convex relaxations for correlated matrices.
problem Recovering hidden vertex permutations from correlated Gaussian matrices.
method Convex relaxations of the quadratic assignment problem over doubly stochastic matrices.
result The solution of the convex relaxation concentrates around the ground-truth permutation matrix for certain correlation parameters.
FDSKL algorithm trains vertically partitioned data with kernels securely and efficiently.
problem Training vertically partitioned data with kernels while maintaining privacy.
method FDSKL algorithm using random features and doubly stochastic gradients for federated learning.
result FDSKL achieves sublinear convergence and guarantees data security.
Corrects mismatch in consistency of nuisance estimators for doubly robust methods.
problem Mismatch in consistency of nuisance estimators in doubly robust methods.
method Calibrated debiased machine learning (calibrated DML) with isotonic regression adjustment.
result Calibrated DML yields doubly robust asymptotic normality with slower convergence of nuisance estimators.
Paper develops unbiased gradient estimator for continuous-time models.
problem Estimating unbiased gradient of log-likelihood for continuous-time models.
method Doubly randomized scheme with coupled conditional particle filter (CCPF).
result Unbiased gradient estimate facilitates gradient-based algorithms.
Geometric approach for unsupervised word embedding alignment.
problem Learning alignment between word embeddings of source and target languages.
method Formulates alignment as domain adaptation on the manifold of doubly stochastic matrices, employing Riemannian conjugate gradient algorithm.
result Empirically outperforms state-of-the-art methods on bilingual lexicon induction tasks.
New DL algorithm estimates OFDM channels without pilots.
problem Estimating OFDM channels in deep fading conditions.
method Deep learning (DL) for blind channel estimation.
result First theory on MSE performance of DL-based estimator.
The paper tackles batch policy learning in Markov Decision Processes, focusing on average reward maximization.
problem Maximizing long-term average reward in Markov Decision Processes with batch learning.
method Doubly robust estimator for average reward, optimization algorithm for optimal policy, finite-sample regret guarantee.
result The proposed method achieves semiparametric efficiency and provides a finite-sample regret guarantee.
We introduce local expectation gradients which is a general purpose stochastic variational inference algorithm for constructing stochastic gradients through sampling from the variational distribution. This algorithm divides the problem of estimating the stochastic gradients over multiple variational parameters into sma…
New estimators improve causal inference in machine learning studies.
problem Improving causal inference in machine learning models.
method Doubly-robust cross-fit estimators for average causal effect.
result Doubly-robust cross-fit estimators outperform other methods in simulations.
A large number of statistical models are "doubly-intractable": the likelihood normalising term, which is a function of the model parameters, is intractable, as well as the marginal likelihood (model evidence). This means that standard inference techniques to sample from the posterior, such as Markov chain Monte Carlo (…
We study a doubly reflected backward stochastic differential equation (BSDE) with integrable parameters and the related Dynkin game. When the lower obstacle L L L and the upper obstacle U U U of the equation are completely separated, we construct a unique solution of the doubly reflected BSDE by pasting local solutions and…
New method combines strengths of two PCL approaches without density ratio estimation.
problem Estimating causal functions in Proxy Causal Learning with unobserved confounders and proxies.
method Kernel-based doubly robust estimators combining treatment and outcome bridges, density ratio-free.
result Outperforms existing methods on PCL benchmarks, including a prior doubly robust method.
New method for unbiased sampling of doubly-intractable distributions.
problem Hard computation of normalizing constants for complex probability distributions.
method Adapting random series truncation and Markov chain coupling for unbiased estimation of 1/Z.
result Estimators with lower variance and higher positive estimates.
A new algorithm improves regret bounds for contextual bandits.
problem Complexity of missing data in multi-armed bandits.
method Doubly Robust (DR) Thompson Sampling with contexts.
result Improved regret bound with i l d e O ( d T ) ilde{O}(d\sqrt{T}) i l d e O ( d T ) . Natural experiment dataset reveals inconsistent treatment effect estimators.
problem Inconsistent results from over 20 estimators on a new dataset.
method Created a benchmark to evaluate estimator accuracy, derived variance formula, introduced new estimator.
result Doubly robust estimators outperform others by orders of magnitude.
The paper confirms a conjecture about manifolds with positive curvature.
problem Estimating the width of manifolds with positive sectional curvature.
method Establishing an optimal Lipschitz lower bound for functions on manifolds.
result Characterization of doubly warped product metrics with positive constant curvature.
Estimates exponential family distributions using a novel doubly dual embedding technique.
problem Estimating exponential family distributions with smoothness and efficiency.
method Doubly dual embedding for avoiding partition function computation and flexible sampling.
result Improves memory and time efficiency while offering stronger statistical properties.
New estimator for causal effects in large datasets.
problem Unobserved confounding in large-scale data.
method Doubly robust estimator combining imputation, IPW, and cross-fitting.
result Error converges to Gaussian distribution at parametric rate.
We introduce a new method to handle permutations efficiently using variational inference.
problem Efficient probabilistic reasoning about permutations in high-dimensional spaces.
method We reparameterize the Birkhoff polytope to enable variational inference over permutations.
result Our method enables efficient and accurate Bayesian inference over permutations.
This paper investigates robust and efficient DR/RDR estimators for WATEs.
problem Lack of systematic investigation into robustness and efficiency conditions for WATE estimation.
method Proposes three RDR estimators using semiparametric efficient influence function and double/debiased machine learning.
result Demonstrates the practical relevance of the methods in medical and social sciences.
New methods estimate policy value and gradients for deterministic policies from off-policy data.
problem Estimating policy value and gradients for deterministic policies from off-policy data.
method Proposed new doubly robust estimators based on kernelization approaches.
result Demonstrated a rate independent of horizon length for policy value and gradient estimation.
Improves SSL with doubly robust estimation of unlabeled class distribution.
problem Limited labeled data and long-tailed class distributions in unlabeled data.
method Explicitly estimate unlabeled class distribution using doubly robust estimator.
result Improves performance of SSL methods on unlabeled data.
Doubly robust method reduces label cost for noisy crowdsourced data.
problem Efficiently label large datasets with noisy, non-expert labels.
method Doubly robust estimation framework using supervised learners.
result Significantly reduced variance in estimation with adaptive worker/item selection.
New estimator improves ATT estimation efficiency with external controls.
problem Reduced efficiency when incorporating external controls into ATT estimation.
method Proposes a novel doubly robust estimator for ATT that maintains higher efficiency than standard approaches.
result Demonstrates improved efficiency of the new estimator compared to standard approaches, even under model misspecification.