electricsheepasia/asia-ilo-ees-tees-sex-ocu-nb-employees-by-sex-and-occupation-thousands
收藏资源简介:
该数据集名为按性别和职业划分的员工数量(千)| 亚洲(ILOSTAT),包含21,769个观测值,覆盖40个亚洲国家,时间跨度为1996年至2025年。数据来源于国际劳工组织(ILO)的ILOSTAT数据库,这是一个全球领先的劳动统计数据源,整合了就业、失业、工资、工作时间、童工、非正规经济、社会保护、职业伤害和可持续发展目标(SDG)体面工作目标等指标。数据集的核心指标为EES_TEES_SEX_OCU_NB,表示按性别和职业划分的员工数量(以千计),数据通过ILOSTAT REST API获取,并过滤为亚洲国家代码。数据集采用表格格式,包含列如国家代码、国家名称、数据来源、指标、性别分类、年份、观测值等,适用于表格分类、回归和时间序列预测任务。数据经过ILO的标准化处理,使用国际劳工统计学家会议(ICLS)定义,并标注了数据来源和质量状态(如临时数据、不可靠数据)。数据集由Electric Sheep Asia重新打包,旨在为亚洲提供统一、机器学习就绪的数据层,便于研究人员和开发者使用。许可证为CC-BY-4.0,使用时需引用原始来源和重新打包方。
This dataset is named Employees by sex and occupation (thousands) | Asia (ILOSTAT) and contains 21,769 observations across 40 Asia countries, spanning the years 1996–2025. It is sourced from the International Labour Organization (ILO)s ILOSTAT database, a leading global source for labour statistics that compiles indicators on employment, unemployment, wages, working time, child labour, informal economy, social protection, occupational injuries, and SDG decent work targets. The core indicator is EES_TEES_SEX_OCU_NB, representing employees by sex and occupation in thousands. Data is pulled directly from the ILOSTAT REST API and filtered to Asia ISO3 country codes. The dataset is in tabular format with columns such as country code, country name, source, indicator, sex disaggregation, year, observed value, etc., and is suitable for tabular classification, regression, and time-series forecasting tasks. It is harmonised using ICLS (International Conference of Labour Statisticians) definitions, with source and quality flags (e.g., provisional, unreliable). The dataset is repackaged by Electric Sheep Asia as part of a unified, ML-ready data layer for Asia, facilitating easy use by researchers and developers. The license is CC-BY-4.0, and users are required to cite both the original source and the repackaging. The dataset includes detailed schema, geographic coverage, methodology, usage examples, and citation information.




