electricsheepafrica/africa-ilo-luu-xlux-sex-nb-jobs-gap-thousands
收藏资源简介:
该数据集名为“就业缺口(千)| 非洲(ILOSTAT)”,包含来自ILOSTAT(国际劳工组织统计数据库)的“其他劳动力未充分利用指标”数据。具体涵盖324个观测值,涉及28个非洲国家(如南非、毛里求斯、卢旺达等),时间范围为2008年至2025年。核心指标为“LUU_XLUX_SEX_NB”,即“就业缺口(千)”,用于衡量劳动力市场中的未充分利用情况。数据集以表格形式组织,包括国家代码、国家名称、数据来源、指标代码、指标标签、性别分类(总计、男性、女性)、年份、观测值、观测状态及相关注释等列。数据通过ILOSTAT REST API获取,并经过ILO的标准化处理,确保与国际劳工统计学家会议(ICLS)定义一致。数据集适用于表格分类、回归和时间序列预测等机器学习任务,由Electric Sheep Africa重新打包为Parquet格式,便于使用Hugging Face的`load_dataset()`快速加载。数据集遵循CC-BY-4.0许可证,使用时需引用原始ILO来源及重新打包方。
This dataset is titled Jobs gap (thousands) | Africa (ILOSTAT) and contains data on Other measures of labour underutilization sourced from ILOSTAT, the International Labour Organizations statistics database. It includes 324 observations across 28 African countries (e.g., South Africa, Mauritius, Rwanda) spanning the years 2008 to 2025. The core indicator is LUU_XLUX_SEX_NB, which represents Jobs gap (thousands), measuring labour underutilization in the job market. The data is structured in tabular format with columns such as country code, country name, data source, indicator code, indicator label, sex disaggregation (total, male, female), year, observed value, observation status, and related notes. Data is retrieved via the ILOSTAT REST API and harmonized by the ILO following International Conference of Labour Statisticians (ICLS) definitions. It is suitable for machine learning tasks like tabular classification, regression, and time-series forecasting. Repackaged by Electric Sheep Africa into Parquet format for easy loading with Hugging Faces `load_dataset()`. The dataset is licensed under CC-BY-4.0, requiring citation of both the original ILO source and the repackaging.




