Combining Survey and Non-survey Data for Improved Sub-area Prediction Using a Multi-level Model
Author(s) -
Jae Kwang Kim,
Zhonglei Wang,
Zhengyuan Zhu,
Nathan B. Cruze
Publication year - 2018
Publication title -
journal of agricultural biological and environmental statistics
Language(s) - English
Resource type - Journals
SCImago Journal Rank - 0.621
H-Index - 52
eISSN - 1537-2693
pISSN - 1085-7117
DOI - 10.1007/s13253-018-0320-2
Subject(s) - small area estimation , statistics , sampling (signal processing) , survey data collection , computer science , mean squared error , econometrics , data mining , mathematics , estimator , filter (signal processing) , computer vision
Combining information from different sources is an important practical problem in survey sampling. Using a hierarchical area-level model, we establish a framework to integrate auxiliary information to improve state-level area estimates. The best predictors are obtained by the conditional expectations of latent variables given observations, and an estimate of the mean squared prediction error is discussed. Sponsored by the National Agricultural Statistics Service of the US Department of Agriculture, the proposed model is applied to the planted crop acreage estimation problem by combining information from three sources, including the June Area Survey obtained by a probability-based sampling of lands, administrative data about the planted acreage and the cropland data layer, which is a commodity-specific classification product derived from remote sensing data. The proposed model combines the available information at a sub-state level called the agricultural statistics district and aggregates to improve state-level estimates of planted acreages for different crops. Supplementary materials accompanying this paper appear on-line.
Accelerating Research
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom
Address
John Eccles HouseRobert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom