Harmonization Crosswalks and Code for India's ASUSE/UNAE Non-Farm Unincorporated Enterprise Survey Series (2010-11 to 2025)
收藏资源简介:
This dataset provides the derived harmonization infrastructure used to construct a unified establishment-level analysis file from six rounds of India's national non-farm unincorporated enterprise surveys: the Unincorporated Non-Agricultural Enterprises Survey (UNAE, NSS rounds 67 and 73, 2010-11 and 2015-16) and the Annual Survey of Unincorporated Sector Enterprises (ASUSE, 2021-22, 2022-23, 2023-24, and 2025), both conducted by India's Ministry of Statistics and Programme Implementation (MoSPI). The deposit contains four components. First, a concordance mapping the three-digit National Sample Survey (NSS) region code to state and union territory, covering all 88 officially published regions and three empirically identified corrections not present in the published concordance, validated to zero unmatched establishments across all six survey rounds. Second, a classification of all 72 two-digit National Industrial Classification (NIC 2008) divisions observed in the ASUSE 2021-22 sample into contact-intensity tiers (High, Medium, Low), together with a five-digit sub-classification of the 57 retail-trade product codes within Division 47. Third, a validated mapping of the numeric worker-category codes used in each survey round's Employment Particulars block to standardized categories (working owner, formal hired worker, informal hired worker, other worker, and self-help-group member), including identification of subtotal rows requiring exclusion, a source of measurement error not previously documented for this survey series. Fourth, an R script implementing the full harmonization pipeline, including establishment-key construction for each round's distinct naming convention, state derivation, sector classification, employment-quality and registration-status variable construction, and merging with the Oxford COVID-19 Government Response Tracker. This deposit does not include the underlying survey microdata, which is collected and licensed by MoSPI and must be obtained directly from the National Statistical Office. The crosswalks and code provided here are intended to let researchers with independent access to this microdata reconstruct a harmonized, analysis-ready dataset without re-deriving the state concordance, sector classification, or worker-category mapping from scratch, each of which required substantial validation against official published aggregates during the construction of the accompanying research paper.




