使用第一阶段共变数将观测数据和实验数据结合起来 (Combining Observational and Experimental Data Using First-stage Covariates) - 专知论文

会员服务 ·

0

估计/估计量 · 查准率/准确率 · 可约的 · 数据集 · 可辨认的 ·

2022 年 6 月 24 日

Combining Observational and Experimental Data Using First-stage Covariates

翻译：使用第一阶段共变数将观测数据和实验数据结合起来

Randomized controlled trials generate experimental variation that can credibly identify causal effects, but often suffer from limited scale, while observational datasets are large, but often violate desired identification assumptions. To improve estimation efficiency, I propose a method that combines experimental and observational datasets when 1) units from these two datasets are similar and 2) some characteristics of these units are observed. I show that if these characteristics can partially explain treatment assignment in the observational data, they can be used to derive moment restrictions that, in combination with the experimental data, improve estimation efficiency. I outline three estimators (weighting, shrinkage, or GMM) for implementing this strategy, and show that my methods can reduce variance by up to 50% in typical experimental designs; therefore, only half of the experimental sample is required to attain the same statistical precision. If researchers are allowed to design experiments differently, I show that they can further improve the precision by directly leveraging this correlation between characteristics and assignment. I apply my method to a search listing dataset from Expedia that studies the causal effect of search rankings, and show that the method can substantially improve the precision.

翻译：由随机控制的试验会产生实验变异,可以令人信服地确定因果关系,但往往受到有限规模的影响,而观察数据集则庞大,但往往违反预期的识别假设。为了提高估计效率,我提议一种方法,将实验和观察数据集结合起来,只要1个来自这两个数据集的单元相似,2个单元的一些特点被观察到。我表明,如果这些特性可以部分解释观察数据中的治疗任务,它们可以用来产生与实验数据相结合的瞬间限制,提高估计效率。我概述了执行这一战略的3个估计数据(加权、缩水或GMM),并表明在典型的实验设计中,我的方法可以将差异减少高达50%;因此,只有一半的实验样品需要达到同样的统计精确度。如果允许研究人员以不同的方式设计实验,我表明它们可以通过直接利用这些特性和任务之间的关联来进一步提高精确度。我的方法用于搜索Expedia的数据集,以研究搜索等级的因果关系,并表明该方法可以大大改进精确度。

0

相关内容

估计/估计量

估计/估计量

不可错过！UIUC最新《统计强化学习》课程！

专知会员服务

54+阅读 · 2020年9月7日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

社交网络上议题社群的公共焦虑研究，中国人民大学新闻学院塔娜讲师，第八届全国社会媒体处理大会SMP2019

社交网络上议题社群的公共焦虑研究，中国人民大学新闻学院塔娜讲师，第八届全国社会媒体处理大会SMP2019

专知会员服务

15+阅读 · 2019年10月23日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

高血糖激活滑膜AGE-RAGE-PKC轴致骨关节炎易感的机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

Chemerin通过调节p38MAPK通路参与动脉粥样硬化分子机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

稀疏植被覆盖条件下土壤盐渍化高光谱遥感定量反演与动态监测

国家自然科学基金

0+阅读 · 2014年12月31日

从ERS信号通路探讨慢性心理应激影响T2DM大鼠IR及逍遥散的干预机制

国家自然科学基金

0+阅读 · 2013年12月31日

miRNA在补体介导甲型H1N1流感肺部炎症损伤中的作用及调控机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

东北地区海港城市与内陆腹地关系演化模式研究

国家自然科学基金

0+阅读 · 2012年12月31日

小麦小分子RNA TaMIR167和TaMIR1139应答和抵御低磷逆境的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

基于蛋白质组学和代谢组学整合分析的Paraconiothyrium variable GHJ-4降解木质素的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

Reality-based Interaction用户界面模型和评估方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

脂肪因子Chemerin在骨骼肌胰岛素抵抗发生中的作用及其机制

国家自然科学基金

0+阅读 · 2008年12月31日

Modeling racial/ethnic differences in COVID-19 incidence with covariates subject to non-random missingness

Arxiv

0+阅读 · 2022年8月16日

An unified framework for point-level, areal, and mixed spatial data: the Hausdorff-Gaussian Process

Arxiv

0+阅读 · 2022年8月16日

The Correlated Arc Orienteering Problem

Arxiv

0+阅读 · 2022年8月16日

When Does Differentially Private Learning Not Suffer in High Dimensions?

Arxiv

0+阅读 · 2022年8月15日

A Practical Guide to Counterfactual Estimators for Causal Inference with Time-Series Cross-Sectional Data

Arxiv

0+阅读 · 2022年8月14日

Doubly Robust Estimation under Covariate-induced Dependent Left Truncation

Arxiv

0+阅读 · 2022年8月14日

Optimal Recovery for Causal Inference

Arxiv

0+阅读 · 2022年8月13日

Causal Discovery in Probabilistic Networks with an Identifiable Causal Effect

Arxiv

0+阅读 · 2022年8月13日

A Scalable Probabilistic Model for Reward Optimizing Slate Recommendation

Arxiv

0+阅读 · 2022年8月10日

Adversarial and Contrastive Variational Autoencoder for Sequential Recommendation

Arxiv

17+阅读 · 2021年3月19日

VIP会员

文章信息

相关主题

估计/估计量

查准率/准确率

相关VIP内容

不可错过！UIUC最新《统计强化学习》课程！

专知会员服务

54+阅读 · 2020年9月7日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

社交网络上议题社群的公共焦虑研究，中国人民大学新闻学院塔娜讲师，第八届全国社会媒体处理大会SMP2019

社交网络上议题社群的公共焦虑研究，中国人民大学新闻学院塔娜讲师，第八届全国社会媒体处理大会SMP2019

专知会员服务

15+阅读 · 2019年10月23日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【博士论文】在低维和高维空间中分析、建模和转换潜在表征

从无人机到数据：揭示边缘计算作为新作战域

可解释人工智能的基础

大规模视觉模型中的基于提示的适应：综述

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Modeling racial/ethnic differences in COVID-19 incidence with covariates subject to non-random missingness

Arxiv

0+阅读 · 2022年8月16日

An unified framework for point-level, areal, and mixed spatial data: the Hausdorff-Gaussian Process

Arxiv

0+阅读 · 2022年8月16日

The Correlated Arc Orienteering Problem

Arxiv

0+阅读 · 2022年8月16日

When Does Differentially Private Learning Not Suffer in High Dimensions?

Arxiv

0+阅读 · 2022年8月15日

A Practical Guide to Counterfactual Estimators for Causal Inference with Time-Series Cross-Sectional Data

Arxiv

0+阅读 · 2022年8月14日

Doubly Robust Estimation under Covariate-induced Dependent Left Truncation

Arxiv

0+阅读 · 2022年8月14日

Optimal Recovery for Causal Inference

Arxiv

0+阅读 · 2022年8月13日

Causal Discovery in Probabilistic Networks with an Identifiable Causal Effect

Arxiv

0+阅读 · 2022年8月13日

A Scalable Probabilistic Model for Reward Optimizing Slate Recommendation

Arxiv

0+阅读 · 2022年8月10日

Adversarial and Contrastive Variational Autoencoder for Sequential Recommendation

Arxiv

17+阅读 · 2021年3月19日

相关基金

高血糖激活滑膜AGE-RAGE-PKC轴致骨关节炎易感的机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

Chemerin通过调节p38MAPK通路参与动脉粥样硬化分子机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

稀疏植被覆盖条件下土壤盐渍化高光谱遥感定量反演与动态监测

国家自然科学基金

0+阅读 · 2014年12月31日

从ERS信号通路探讨慢性心理应激影响T2DM大鼠IR及逍遥散的干预机制

国家自然科学基金

0+阅读 · 2013年12月31日

miRNA在补体介导甲型H1N1流感肺部炎症损伤中的作用及调控机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

东北地区海港城市与内陆腹地关系演化模式研究

国家自然科学基金

0+阅读 · 2012年12月31日

小麦小分子RNA TaMIR167和TaMIR1139应答和抵御低磷逆境的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

基于蛋白质组学和代谢组学整合分析的Paraconiothyrium variable GHJ-4降解木质素的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

Reality-based Interaction用户界面模型和评估方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

脂肪因子Chemerin在骨骼肌胰岛素抵抗发生中的作用及其机制

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员