遇见数据集

VSET: A harmonized European database of education, telework, labour market and earnings profiles

收藏
Zenodo2026-05-07 更新2026-05-26 收录
官方服务:

资源简介:

The VSET (Variables Sociales, Económicas y Teletrabajo) dataset is a harmonized European database derived from three major Eurostat microdata sources: the Adult Education Survey (AES) 2022, the Labour Force Survey (LFS) 2023, and the Structure of Earnings Survey (SES) 2018. Unlike person-level microdata, VSET uses a profile-based design where each row represents a unique demographic and occupational profile cell. Scope: Covers 22 European countries. Content: Contains 1,440,134 unique profile rows and 39 variables covering sociodemographic background, educational attainment (ISCED-based), occupational categories (ISCO-08 at 2-digit level), and 15 earnings summary measures. Usage: The variable N_instances represents the number of records each profile row represents and must be used as a frequency weight in all analytical summaries. Format: Provided as a partitioned Parquet dataset and a SQL-ready CSV with accompanying schema for relational database ingestion.

VSET(Variables Sociales, Económicas y Teletrabajo,即社会、经济与远程工作变量)数据集是一套经协调统一的欧洲数据库,其数据源自欧盟统计局(Eurostat)的三大主流微观数据源:2022年成人教育调查(Adult Education Survey, AES)、2023年劳动力调查(Labour Force Survey, LFS)以及2018年薪酬结构调查(Structure of Earnings Survey, SES)。与个体层面的微观数据不同,VSET采用基于画像的设计逻辑,每一行均代表一个独特的人口与职业画像单元格。 覆盖范围:涵盖22个欧洲国家。 数据内容:包含1,440,134条唯一画像行,以及39个变量,覆盖社会人口背景、基于国际教育标准分类(ISCED)的受教育程度、两位数层级的国际职业分类(ISCO-08)职业类别,以及15项薪酬汇总指标。 使用说明:变量N_instances代表每条画像行所对应的记录总数,在所有分析汇总工作中需将其作为频率权重使用。 数据格式:数据集以分区Parquet格式与可适配SQL导入的CSV格式提供,并附带适用于关系型数据库导入的配套架构文件。

提供机构:
Zenodo
创建时间:
2026-05-07
二维码
社区交流群
二维码
科研交流群
商业服务