弱仪器实用指南

UNSW Business School Research Paper Series Pub Date : 2021-09-27 DOI:10.2139/ssrn.3846841

M. Keane, Timothy Neal

{"title":"弱仪器实用指南","authors":"M. Keane, Timothy Neal","doi":"10.2139/ssrn.3846841","DOIUrl":null,"url":null,"abstract":"We provide a simple survey of the literature on weak instruments, aimed at giving practical advice to applied researchers. It is well-known that 2SLS has poor properties if instruments are exogenous but “weak.” We clarify these properties, explain weak instrument tests, and examine how behavior of 2SLS depends on instrument strength. A common standard for “strong” instruments is a ﬁrst-stage F-statistic of at least 10. But 2SLS has some poor properties in that context: It has low power, and the 2SLS standard error estimate tends to be artiﬁcially small in samples where the 2SLS parameter estimate is most contaminated by the OLS bias. This causes t-tests to give very misleading results. Surprisingly, this problem persists even if the ﬁrst-stage F is in the thousands. Robust tests like Anderson-Rubin greatly alleviate these problems, and should be used in lieu of the t-test even with strong instruments. In many realistic settings a ﬁrst-stage F well above 10 may be necessary to give high conﬁdence that 2SLS will outperform OLS. For example, in the archetypal application of estimating returns to education, we argue one needs F of at least 50.","PeriodicalId":23435,"journal":{"name":"UNSW Business School Research Paper Series","volume":"26 1","pages":""},"PeriodicalIF":0.0000,"publicationDate":"2021-09-27","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"7","resultStr":"{\"title\":\"A Practical Guide to Weak Instruments\",\"authors\":\"M. Keane, Timothy Neal\",\"doi\":\"10.2139/ssrn.3846841\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"We provide a simple survey of the literature on weak instruments, aimed at giving practical advice to applied researchers. It is well-known that 2SLS has poor properties if instruments are exogenous but “weak.” We clarify these properties, explain weak instrument tests, and examine how behavior of 2SLS depends on instrument strength. A common standard for “strong” instruments is a ﬁrst-stage F-statistic of at least 10. But 2SLS has some poor properties in that context: It has low power, and the 2SLS standard error estimate tends to be artiﬁcially small in samples where the 2SLS parameter estimate is most contaminated by the OLS bias. This causes t-tests to give very misleading results. Surprisingly, this problem persists even if the ﬁrst-stage F is in the thousands. Robust tests like Anderson-Rubin greatly alleviate these problems, and should be used in lieu of the t-test even with strong instruments. In many realistic settings a ﬁrst-stage F well above 10 may be necessary to give high conﬁdence that 2SLS will outperform OLS. For example, in the archetypal application of estimating returns to education, we argue one needs F of at least 50.\",\"PeriodicalId\":23435,\"journal\":{\"name\":\"UNSW Business School Research Paper Series\",\"volume\":\"26 1\",\"pages\":\"\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2021-09-27\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"7\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"UNSW Business School Research Paper Series\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.2139/ssrn.3846841\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"UNSW Business School Research Paper Series","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.2139/ssrn.3846841","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 7

摘要

我们提供了一个关于弱仪器的文献的简单调查，旨在给应用研究人员提供实用的建议。众所周知，如果仪器是外生的，那么2SLS的性能很差，但“弱”。我们澄清了这些特性，解释了弱仪器测试，并研究了2SLS的行为如何取决于仪器强度。“强”仪器的通用标准是第一阶段f统计量至少为10。但在这种情况下，2SLS具有一些较差的特性:它具有低功率，并且在2SLS参数估计最受OLS偏差污染的样本中，2SLS标准误差估计往往人为地很小。这导致t检验给出非常具有误导性的结果。令人惊讶的是，即使第一阶段的F是数千，这个问题仍然存在。像Anderson-Rubin这样的稳健测试极大地缓解了这些问题，即使使用强大的工具，也应该使用它来代替t检验。在许多实际情况下，第一阶段F远高于10可能是有必要的，因为2SLS将优于OLS。例如，在评估教育回报的原型应用中，我们认为一个人需要至少50分的F。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

A Practical Guide to Weak Instruments

We provide a simple survey of the literature on weak instruments, aimed at giving practical advice to applied researchers. It is well-known that 2SLS has poor properties if instruments are exogenous but “weak.” We clarify these properties, explain weak instrument tests, and examine how behavior of 2SLS depends on instrument strength. A common standard for “strong” instruments is a ﬁrst-stage F-statistic of at least 10. But 2SLS has some poor properties in that context: It has low power, and the 2SLS standard error estimate tends to be artiﬁcially small in samples where the 2SLS parameter estimate is most contaminated by the OLS bias. This causes t-tests to give very misleading results. Surprisingly, this problem persists even if the ﬁrst-stage F is in the thousands. Robust tests like Anderson-Rubin greatly alleviate these problems, and should be used in lieu of the t-test even with strong instruments. In many realistic settings a ﬁrst-stage F well above 10 may be necessary to give high conﬁdence that 2SLS will outperform OLS. For example, in the archetypal application of estimating returns to education, we argue one needs F of at least 50.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

UNSW Business School Research Paper Series

自引率

0.00%

发文量