
The estimation of average treatment effect (ATE) as a causal parameter is carried out in two steps, where in the first step, the treatment and outcome are modeled to incorporate the potential confounders, and in the second step, the predictions are inserted into the ATE estimators such as the augmented inverse probability weighting (AIPW) estimator. Due to the concerns regarding the non-linear or unknown relationships between confounders and the treatment and outcome, there has been interest in applying non-parametric methods such as machine learning (ML) algorithms instead. Some of the literature proposes to use two separate neural networks (NNs) where there is no regularization on the network’s parameters except the stochastic gradient descent (SGD) in the NN’s optimization. Our simulations indicate that the AIPW estimator suffers extensively if no regularization is utilized. We propose the normalization of AIPW (referred to as nAIPW) which can be helpful in some scenarios. nAIPW, provably, has the same properties as AIPW, that is, the double-robustness and orthogonality properties. Further, if the first-step algorithms converge fast enough, under regulatory conditions, nAIPW will be asymptotically normal. We also compare the performance of AIPW and nAIPW in terms of the bias and variance when small to moderate L1 regularization is imposed on the NNs.
FOS: Computer and information sciences, instrumental variables, Science, Physics, QC1-999, Q, Machine Learning (stat.ML), neural networks, Astrophysics, 310, doubly robust estimation, Statistics - Applications, Statistics - Computation, Article, QB460-466, Methodology (stat.ME), Statistics - Machine Learning, Applications (stat.AP), causal inference, semi-parametric theory, Statistics - Methodology, Computation (stat.CO)
FOS: Computer and information sciences, instrumental variables, Science, Physics, QC1-999, Q, Machine Learning (stat.ML), neural networks, Astrophysics, 310, doubly robust estimation, Statistics - Applications, Statistics - Computation, Article, QB460-466, Methodology (stat.ME), Statistics - Machine Learning, Applications (stat.AP), causal inference, semi-parametric theory, Statistics - Methodology, Computation (stat.CO)
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 6 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 10% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 10% |
