electricsheepasia/asia-ilo-ear-ehrm-sex-nb-median-hourly-earnings-of-employees-by-sex-local-c
收藏资源简介:
该数据集包含亚洲25个国家按性别划分的员工小时收入中位数(本地货币)统计数据,涵盖1996年至2025年期间的663个观测值。数据来源于国际劳工组织(ILO)的ILOSTAT数据库,这是一个全球领先的劳动统计数据源,整合了就业、失业、工资、工作时间等多个领域的指标。数据集通过ILOSTAT REST API获取,并筛选了亚洲国家的ISO3代码,确保数据覆盖印度尼西亚、菲律宾、土耳其、柬埔寨、越南、斯里兰卡、巴基斯坦、泰国、巴勒斯坦、亚美尼亚、印度、约旦、蒙古、东帝汶、孟加拉国等25个国家和地区。数据按性别维度进行细分(包括总计、男性、女性等),并包含国家代码、国家名称、数据来源、指标代码、观测年份、观测值、数据状态等列。数据集适用于表格分类、表格回归和时间序列预测等机器学习任务,可用于分析亚洲国家收入趋势、性别收入差异等研究。数据以Parquet格式发布,便于使用HuggingFace的`load_dataset()`函数快速加载和处理。
This dataset contains median hourly earnings (in local currency) statistics for employees in 25 Asian countries, disaggregated by gender, with 663 observations spanning the period from 1996 to 2025. This dataset is sourced from the ILOSTAT database of the International Labour Organization (ILO), a leading global labor statistics data source that integrates indicators across multiple domains including employment, unemployment, wages, and working hours. The dataset was obtained via the ILOSTAT REST API, and ISO3 codes for Asian countries were filtered to ensure coverage of 25 Asian countries and regions including Indonesia, the Philippines, Turkey, Cambodia, Vietnam, Sri Lanka, Pakistan, Thailand, Palestine, Armenia, India, Jordan, Mongolia, Timor-Leste, and Bangladesh. The data is disaggregated by gender (including total, male, female, etc.), and includes columns such as country code, country name, data source, indicator code, observation year, observed value, and data status. This dataset is suitable for machine learning tasks such as tabular classification, tabular regression, and time series forecasting, and can be used for research analyzing income trends and gender income gaps in Asian countries. The data is released in Parquet format, enabling fast loading and processing using HuggingFace's `load_dataset()` function.




