遇见数据集

electricsheepasia/asia-ilo-emp-nifl-sex-ocu-est-nb-informal-employment-by-sex-occupation-and-establis

收藏
Hugging Face2026-05-25 更新2026-05-31 收录
官方服务:

资源简介:

该数据集包含亚洲24个国家2000年至2025年期间的非正规就业数据,按性别、职业和企业规模分类,共计32,460条观测记录。数据来源于国际劳工组织(ILO)的ILOSTAT数据库,通过其REST API获取,并经过ILO统一处理,确保符合国际劳工统计学家会议(ICLS)的定义。数据集涵盖一个核心指标:EMP_NIFL_SEX_OCU_EST_NB,即按性别、职业和企业规模划分的非正规就业人数(以千计)。数据结构包括国家代码、国家名称、数据来源、指标代码、性别分类(总计、男性、女性)、职业分类、企业规模分类、观测年份、观测值及其状态标志等列。数据按年度频率提供,并包含数据质量说明,例如ILO会选择最佳来源以处理同一国家同年份的多源数据。该数据集由Electric Sheep Asia重新打包,旨在为亚洲地区提供机器学习就绪的标准化数据,便于研究者和开发者使用HuggingFace的`load_dataset()`功能快速加载和分析。

This dataset contains informal employment data for 24 Asian countries from 2000 to 2025, disaggregated by sex, occupation, and establishment size, with a total of 32,460 observations. The data is sourced from the International Labour Organization (ILO) ILOSTAT database, retrieved via its REST API and harmonized by ILO following International Conference of Labour Statisticians (ICLS) definitions. It covers one core indicator: EMP_NIFL_SEX_OCU_EST_NB, representing informal employment by sex, occupation, and establishment size (in thousands). The dataset schema includes columns such as country code, country name, data source, indicator code, sex classification (total, male, female), occupation classification, establishment size classification, observation year, observed value, and status flags. Data is provided at annual frequency and includes quality caveats, such as ILOs selection of the best source for multiple sources per country-year. Repackaged by Electric Sheep Asia, this dataset aims to provide a standardized, machine-learning-ready data layer for Asia, enabling researchers and developers to quickly load and analyze data using HuggingFaces `load_dataset()` functionality.

提供机构:
electricsheepasia
二维码
社区交流群
二维码
科研交流群
商业服务