electricsheepeurope/europe-ilo-ees-tees-sex-edu-dsb-nb-employees-by-sex-education-and-disability-status-t
收藏资源简介:
该数据集名为“按性别、教育和残疾状况划分的雇员数(千)| 欧洲(ILOSTAT)”,是一个表格型数据集,专注于时间序列预测和分类回归任务。它包含24,041个观测值,覆盖32个欧洲国家,时间跨度为2002年至2025年,涉及1个核心指标:EES_TEES_SEX_EDU_DSB_NB,即按性别、教育和残疾状况划分的雇员数(单位:千)。数据来源于国际劳工组织(ILO)的ILOSTAT数据库,通过REST API获取并过滤为欧洲国家代码。数据集包括多个列,如国家代码、国家名称、数据源、指标代码、性别分解(总计、男性、女性)、教育分类、残疾状况分类、年份、观测值、观测状态标志以及相关注释。数据质量方面,为年度频率,ILO选择最佳数据源,分解列仅在指标发布时非空。该数据集由Electric Sheep Europe重新打包,旨在为欧洲提供统一的、机器学习就绪的数据层,便于研究人员和开发者使用HuggingFace的load_dataset()快速加载和分析数据。许可证为CC-BY-4.0,使用时需引用原始源和重新打包方。
This dataset, named Employees by sex, education and disability status (thousands) | Europe (ILOSTAT), is a tabular dataset focused on time-series forecasting and classification/regression tasks. It contains 24,041 observations across 32 European countries, spanning from 2002 to 2025, with 1 core indicator: EES_TEES_SEX_EDU_DSB_NB, which represents employees by sex, education, and disability status (in thousands). The data is sourced from the International Labour Organization (ILO)s ILOSTAT database, retrieved via REST API and filtered to European country codes. The dataset includes columns such as country code, country name, data source, indicator code, sex disaggregation (total, male, female), education classification, disability status classification, year, observed value, observation status flags, and related notes. In terms of data quality, it is annual frequency, with ILO selecting the best source for each country-year, and disaggregation columns are non-null only when published. Repackaged by Electric Sheep Europe, it aims to provide a unified, ML-ready data layer for Europe, enabling researchers and developers to quickly load and analyze data using HuggingFaces load_dataset(). Licensed under CC-BY-4.0, users must cite both the original source and the repackaging when using the dataset.




