Proposes a method to estimate policy values in reinforcement learning with unmeasured confounders.
problem Estimating policy values in reinforcement learning with unmeasured confounders.
method Develops a two-way deconfounder algorithm using a neural tensor network to learn unmeasured confounders and system dynamics.
result Consistent policy value estimation through model-based estimator.
Clarifies the theory of the deconfounder by Imai and Jiang.
problem Theoretical requirements for the deconfounder algorithm.
method Clarifies the assumption of 'no unobserved single-cause confounders' using empirical studies.
result Imai and Jiang's clarification of the assumption does not hold for counterexamples proposed by Ogburn et al. (2020).
Spectral deconfounding improves machine learning models by reducing hidden confounding effects.
problem Machine learning models can be misled by hidden confounders, leading to unreliable predictions.
method Develops a nonlinear spectral deconfounding framework for gradient boosting that modifies boosting dynamics to slow down in confounding-aligned directions.
result Spectrally deconfounded boosting improves estimation of the target function under hidden confounding and is more scalable.
Selective deconfounding improves ATE estimation with less data.
problem Estimating ATE with unobserved confounders using limited data.
method Combining confounded and deconfounded observational data for ATE estimation.
result Selective deconfounding can significantly reduce the amount of deconfounded data needed.
New method estimates treatment effects over time with unobserved confounders.
problem Estimating treatment effects from observational data with unobserved confounders.
method Sequential Deconfounder using Gaussian process latent variable model.
result Unbiased estimates of individualized treatment responses over time.
Deconfounding scores improve causal effect estimation with weak overlap.
problem Challenges in causal treatment effect estimation due to weak overlap in high-dimensional data.
method Propose deconfounding scores to preserve identification and target estimation while improving overlap.
result Prognostic scores are overlap-optimal under a broad family of generalized linear models with Gaussian features.
Spatial Deconfounder tackles interference and confounding in spatial data.
problem Interference and unmeasured spatial factors confound causal inference in spatial domains.
method Two-stage method using CVAE with spatial prior to reconstruct confounder, then estimate causal effects.
result Nonparametric identification of direct and spillover effects under weak assumptions.
Counterexamples show deconfounder fails to control multi-cause confounding.
problem Deconfounding method fails to handle multi-cause confounding.
method Incorrectly inferring independence and joint independence from conditional independence.
result Two simple counterexamples demonstrate deconfounder's failure.
New method deconfounds deep learning feature representations using counterfactual approach.
problem Improving model stability in deep learning models under dataset shifts.
method Adopting last layer features of DNNs trained with softmax activation for logistic regression, and applying counterfactual deconfounding.
result Counterfactual deconfounding can be applied to DNN feature representations, improving model stability.
Deconfounding scores improve causal effect estimation with weak overlap.
problem Poor overlap in treatment and control groups makes causal effect estimators brittle.
method Introduces feature representations that improve overlap without introducing bias.
result Deconfounding scores satisfy a zero-covariance condition that is identifiable in observed data.
The estimation of treatment effects is a pervasive problem in medicine. Existing methods for estimating treatment effects from longitudinal observational data assume that there are no hidden confounders, an assumption that is not testable in practice and, if it does not hold, leads to biased estimates. In this paper, w…
This commentary has two goals. We first critically review the deconfounder method and point out its advantages and limitations. We then briefly consider three possible ways to address some of the limitations of the deconfounder method.
Causal inference from observational data often assumes "ignorability," that all confounders are observed. This assumption is standard yet untestable. However, many scientific studies involve multiple causes, different variables whose effects are simultaneously of interest. We propose the deconfounder, an algorithm that…
The treatment effects of medications play a key role in guiding medical prescriptions. They are usually assessed with randomized controlled trials (RCTs), which are expensive. Recently, large-scale electronic health records (EHRs) have become available, opening up new opportunities for more cost-effective assessments. …
Deconfounds neural network representation similarity metrics to improve consistency and accuracy.
problem Confounding by population structure in similarity metrics like RSA and CKA.
method Covariate adjustment regression to adjust for confounders.
result Improves detection of semantically similar neural networks and consistency in transfer learning.
Unobserved confounding is a major hurdle for causal inference from observational data. Confounders---the variables that affect both the causes and the outcome---induce spurious non-causal correlations between the two. Wang & Blei (2018) lower this hurdle with "the blessings of multiple causes," where the correlation st…
The empirical practice of using factor models to adjust for shared, unobserved confounders, Z, in observational settings with multiple treatments, A, is widespread in fields including genetics, networks, medicine, and politics. Wang and Blei (2019, WB) formalizes these procedures and develops the …
Stable health predictions need deconfounding test set features.
problem Stability of predictions in health machine learning is compromised by selection biases.
method Deconfounding the test set features improves prediction stability across different environments.
result Improved stability achieved by deconfounding test set features.
One of the most prevalent symptoms among the elderly population, dementia, can be detected by classifiers trained on linguistic features extracted from narrative transcripts. However, these linguistic features are impacted in a similar but different fashion by the normal aging process. Aging is therefore a confounding …
Paper proposes methods to use observational data for reinforcement learning, addressing confounding issues.
problem Using observational data for reinforcement learning can lead to misleading outcomes due to unobserved confounders.
method The paper introduces two deconfounding methods in deep reinforcement learning to adjust for confounders.
result The proposed deconfounding methods improve the accuracy of reinforcement learning models using observational data.
We propose a general formulation for addressing reinforcement learning (RL) problems in settings with observational data. That is, we consider the problem of learning good policies solely from historical data in which unobserved factors (confounders) affect both observed actions and rewards. Our formulation allows us t…
VTD uses deep embeddings to estimate treatment effects from longitudinal data without unconfoundedness assumption.
problem Challenges in estimating individualized treatment effects from longitudinal observational data due to confounding bias.
method Leverages deep variational embeddings and observed proxies to learn hidden confounders.
result Effective in estimating treatment effects when hidden confounding is the leading bias.
The goal of recommendation is to show users items that they will like. Though usually framed as a prediction, the spirit of recommendation is to answer an interventional question---for each user and movie, what would the rating be if we "forced" the user to watch the movie? To this end, we develop a causal approach to …
This study shows ESG ratings reduce equity crash risk during market downturns.
problem Decoupling of alpha from tail risk resilience in traditional models.
method Double Machine Learning for structural deconfounding, state-dependent analysis.
result High ESG ratings reduce crash incidence during systemic drawdowns.
AP-Calculus offers a new framework for causal inference in Bayesian networks.
problem Causal inference in Bayesian networks with complex architectures.
method Introduces Attribution Projection Calculus (AP-Calculus) to determine causal relationships.
result Proves that for each label, exactly one intermediate node acts as a deconfounder.
The aim of this comment (set to appear in a formal discussion in JASA) is to draw out some conclusions from an extended back-and-forth I have had with Wang and Blei regarding the deconfounder method proposed in "The Blessings of Multiple Causes" [arXiv:1805.06826]. I will make three points here. First, in my role as th…
New method uncovers hidden causal connections in multivariate point process networks.
problem Unobserved hidden variables confound causal discovery in high-dimensional point process networks.
method Proposes a deconfounding procedure to estimate causal interactions among observed nodes with unknown unobserved processes.
result The method accurately identifies causal interactions among observed processes, even with hidden variables.
A new estimator reduces bias and improves efficiency for staggered adoption studies.
problem Bias in difference-in-differences estimates for staggered adoption studies.
method Fused Extended Two-Way Fixed Effects (FETWFE) estimator with automatic parameter selection.
result FETWFE identifies correct restrictions with probability tending to one, improving efficiency.
DWTS uses observational data to improve clinical trial efficiency.
problem Lack of definitive conclusions from randomized clinical trials due to insufficient patient cohorts and confounding biases.
method DWTS combines observational data with randomized clinical trials using Doubly Debiased LASSO (DDL) to identify reliable covariates.
result DWTS reduces cumulative regret in clinical trials compared to standard methods.
New method debiases counterfactual distributions using observational data.
problem Estimating counterfactual distributions under interventions without relying on observational data.
method Flow-matching approach to learn counterfactual distributions from observational data.
result Deconfounding flows outperform existing debiased counterfactual distribution estimators.
This paper corrects climate model biases using a factor model approach.
problem Systematic biases in GCM outputs due to unobserved confounders.
method Factor model approach to learn latent confounders from historical data and apply them to enhance bias correction.
result Significant improvements in the accuracy of precipitation outputs.
DecoR estimates causal effects in confounded time series data.
problem Estimating causal effects in time series with unobserved confounders.
method Robust regression in the frequency domain.
result Proves upper bounds for estimation error of DecoR, implying consistency.
A method estimates causal parameters using a latent variable recovery.
problem Estimating causal parameters in contexts with multiple causes and unobserved confounding.
method Substitute adjustment via recovery of latent variables.
result Substitute adjustment estimates adjusted regression parameters under certain conditions.
Learning sparse linear models with two-way interactions is desirable in many application domains such as genomics. l1-regularised linear models are popular to estimate sparse models, yet standard implementations fail to address specifically the quadratic explosion of candidate two-way interactions in high dimensions, a…
Unified techniques improve stability and replicability in changing data.
problem Concept drift in data generating distribution.
method Removing hidden confounding and causal regularization.
result Improves stability, replicability, and robustness in heterogeneous data.
Develops a method to estimate treatment effects using noisy proxies over time.
problem Estimating individualized treatment effects from noisy proxies of confounders.
method Deconfounding Temporal Autoencoder (DTA) combining autoencoder and causal regularization.
result Improves treatment effect estimates by leveraging noisy proxies and learning hidden confounders.
DTS improves robustness of bandit algorithms in nonstationary environments.
problem Brittle behavior of multi-armed bandit algorithms in nonstationary exogenous factors.
method Deconfounded Thompson Sampling (DTS) that projects population-level performance while controlling for context.
result DTS provides resilience to exogenous variation and balances exploration and exploitation.
We extend the Caffarelli-Cordoba estimates to the vector case in two ways, one of which has no scalar counterpart, and we give a few applications for minimal solutions.
We propose a communicationally and computationally efficient algorithm for high-dimensional distributed sparse learning. At each iteration, local machines compute the gradient on local data and the master machine solves one shifted l1 regularized minimization problem. The communication cost is reduced from constant …
This paper studies a new application of deep learning (DL) for optimizing constellations in two-way relaying with physical-layer network coding (PNC), where deep neural network (DNN)-based modulation and demodulation are employed at each terminal and relay node. We train DNNs such that the cross entropy loss is directl…
A new knot move preserves pass-move equivalence and differs in count.
problem Defining and analyzing a new knot move.
method Introducing the 1-2-move and comparing it to pass-move and #-move. result Equivalence under 1-2-move for knots is equivalent to pass-move equivalence. This work is devoted to new constructions of symplectically fat fiber bundles. The latter are constructed in two ways: using the Kirwan map and expressing the fatness condition in terms of the isotropy representation related to the G-structure over some homogeneous spaces.
Mix2FLD improves FL accuracy with FD, reducing convergence time.
problem Uplink-downlink capacity asymmetry in federated learning.
method Two-way mixup of local samples and model parameters, preserving privacy.
result Achieves up to 16.7% higher test accuracy with reduced convergence time.
The paper proposes methods to optimize pAUC for deep learning using DRO.
problem Optimizing partial AUC for deep learning models.
method Proposes gradient-based methods using DRO formulations for pAUC maximization.
result Proves convergence of proposed algorithms for optimizing pAUC.
This paper considers a transmission control problem in network-coded two-way relay channels (NC-TWRC), where the relay buffers random symbol arrivals from two users, and the channels are assumed to be fading. The problem is modeled by a discounted infinite horizon Markov decision process (MDP). The objective is to find…
Regularized variants of Principal Components Analysis, especially Sparse PCA and Functional PCA, are among the most useful tools for the analysis of complex high-dimensional data. Many examples of massive data, have both sparse and functional (smooth) aspects and may benefit from a regularization scheme that can captur…
Proposes TNPM for better node popularity in directed and bipartite networks.
problem Lack of node popularity consideration in community detection of directed and bipartite networks.
method Two-Way Node Popularity Model (TNPM) with Delete-One-Method (DOM) and Two-Stage Divided Cosine Algorithm (TSDC).
result Improved estimation accuracy and computational efficiency demonstrated through real-world applications.
Variables in many massive high-dimensional data sets are structured, arising for example from measurements on a regular grid as in imaging and time series or from spatial-temporal measurements as in climate studies. Classical multivariate techniques ignore these structural relationships often resulting in poor performa…