electricsheepeurope/europe-ilo-luu-xlu4-sex-edu-geo-rt-composite-rate-of-labour-underutilization-lu4-by-s
收藏资源简介:
该数据集包含来自国际劳工组织(ILO)ILOSTAT数据库的“其他劳动力利用不足指标”数据,具体针对欧洲地区。数据集涵盖了34个欧洲国家从1998年至2025年的28,552个观测值,核心指标为“按性别、教育程度和城乡地区划分的劳动力利用不足综合率(LU4)”(百分比)。数据通过ILOSTAT REST API获取,并过滤为欧洲国家代码,经过ILO的统计部门使用国际劳工统计学家会议(ICLS)定义进行标准化处理。数据集包含多个列,如国家代码、国家名称、数据来源、指标代码、性别分类(总计、男性、女性)、教育分类、城乡分类、年份、观测值、观测状态等,支持表格分类、回归和时间序列预测等机器学习任务。数据以年度频率发布,并包含数据质量说明,如使用最佳来源和分类列的非空条件。数据集由Electric Sheep Europe重新打包,以Parquet格式发布,便于使用Hugging Face的`load_dataset()`加载。
This dataset contains Other measures of labour underutilization data from the International Labour Organization (ILO) ILOSTAT database, specifically for Europe. It includes 28,552 observations across 34 European countries from 1998 to 2025, with the key indicator being Composite rate of labour underutilization (LU4) by sex, education and rural / urban areas (%). Data is pulled directly from the ILOSTAT REST API and filtered to Europe ISO3 country codes, harmonized by the ILOs Department of Statistics using International Conference of Labour Statisticians (ICLS) definitions. The dataset features columns such as country code, country name, source, indicator code, sex disaggregation (total, male, female), education classification, rural/urban classification, year, observed value, observation status, etc., supporting machine learning tasks like tabular classification, regression, and time-series forecasting. Data is published at annual frequency and includes quality caveats, such as the use of the best source and non-null conditions for disaggregation columns. Repackaged by Electric Sheep Europe in Parquet format for easy loading with Hugging Faces `load_dataset()`.




