Analyses of Public Use Decennial Census Data with Multiply Imputed Industry and Occupation Codes
本文介绍多重插补处理调查无应答的方法,并描述一个将1970年美国普查公共使用样本的行业和职业代码重新校准到1980年标准的项目,通过分析比较使用插补值的大数据集与使用真实值的小数据集的效用,展示仅用一次插补而非多次插补会低估变异性。
"This paper gives a brief introduction to multiple imputation for handling non-response in surveys. We then describe a recently completed project in which multiple imputation was used to recalibrate industry and occupation codes in 1970 U.S. census public use samples to the 1980 standard. Using analyses of data from the project, we examine the utility of analysing a large data set having imputed values compared with analysing a small data set having true values, and we provide examples of the amount by which variability is underestimated by using just one imputation rather than multiple imputations."