遇见数据集

Greek Parliament Proceedings, 1989-2020

收藏
Zenodo2020-12-08 更新2026-05-25 收录
数据链接:
官方服务:

资源简介:

This dataset is produced on behalf of iMEdD by PhD Student in Machine Learning Konstantina Dritsa, with the contribution of data journalist and iMEdD Lab Project Manager Kelly Kiki. iMEdD (incubator for Media Education and Development) is a non-profit journalism organisation that supports and promotes transparency, credibility and independence in journalism. Lab is iMEdD’s content production division which publishes original interactive investigative and data-driven stories by experimenting with new forms and tools in journalism. This dataset is the next version of a previous upload, which originated from the work implemented during the course of the Master thesis entitled "Speech quality and sentiment analysis on the Hellenic Parliament proceedings" at the Athens University of Economics &amp; Business in 2018 under the supervision of the Associate Professor Panagiotis Louridas. This dataset includes 1,280,918 speeches (rows) of Greek parliament members with a total volume of 2.4 GB, that were exported from 5,355 parliamentary sitting record files. They extend chronologically from 1989 up to late July 2020. The dataset consists of a .csv file in UTF-8 encoding and includes the following columns of data: <strong>member_name</strong>: the official name of the parliament member who talked during a sitting. <strong>sitting_date</strong>: the date that the sitting took place. <strong>parliamentary_period</strong>: the name and/or number of the parliamentary period that the speech took place in. A parliamentary period includes multiple parliamentary sessions. <strong>parliamentary_session</strong>: the name and/or number of the parliamentary session that the speech took place in. A parliamentary session includes multiple parliamentary sittings. <strong>parliamentary_sitting</strong>: the name and/or number of the parliamentary sitting that the speech took place in. <strong>political_party</strong>: the political party that the speaker belonged to the moment of their speech. <strong>government</strong>: the government in force when the speech took place. <strong>member_region</strong>: the electoral district the speaker belonged to. <strong>roles</strong>: information about the parliamentary roles and/or government position of the speaker the moment of their speech. <strong>member_gender</strong>: the sex of the speaker <strong>speech</strong>: the speech that the member made during the parliamentary sitting The methodology followed for the production of this dataset is described in the iMEdD Lab's article entitled "The creation of a dataset with the parliament proceedings within 31 years". Scripts and relevant documentation are available on GitHub.

本数据集由机器学习领域博士生康斯坦蒂娜·德里察(Konstantina Dritsa)受iMEdD(媒体教育与发展孵化器,incubator for Media Education and Development)委托制作,数据记者兼iMEdD实验室项目主管凯莉·基基(Kelly Kiki)参与了本数据集的制作工作。iMEdD是一家非营利性新闻机构,致力于扶持并推动新闻行业的透明度、公信力与独立性。iMEdD实验室作为该机构的内容创作部门,通过探索新闻领域的新型形式与技术工具,发布原创交互式调查报道及数据驱动型新闻作品。本数据集为此前已上传版本的更新迭代版,其雏形源自2018年雅典经济与商业大学的一篇硕士学位论文研究工作,该论文题为《希腊议会议事过程中的语音质量与情感分析》,由帕纳约蒂斯·卢里达斯(Panagiotis Louridas)副教授指导。本数据集包含1,280,918条希腊议会议员的发言记录,对应数据集128万余行数据,总数据量达2.4 GB,源自5,355份议会议事记录文件,时间跨度覆盖1989年至2020年7月下旬。本数据集采用UTF-8编码的CSV格式存储,包含以下数据字段: - **member_name**:发言议员的官方姓名 - **sitting_date**:议事会议召开日期 - **parliamentary_period**:发言所属的议会会期名称及/或编号,一个议会会期包含多场议会会议 - **parliamentary_session**:发言所属的议会会议名称及/或编号,一场议会会议包含多场议事活动 - **parliamentary_sitting**:发言所属的议会议事场次名称及/或编号 - **political_party**:发言时议员所属的政党 - **government**:发言时任在任政府 - **member_region**:发言议员所属的选举选区 - **roles**:发言时议员的议会职务及/或政府职位信息 - **member_gender**:发言者的性别 - **speech**:议员在议事活动中发表的发言内容。本数据集的制作方法已发表于iMEdD Lab题为《31年议会议事记录数据集的构建》的文章中,相关脚本与文档可在GitHub平台获取。

提供机构:
Zenodo
创建时间:
2020-12-08
二维码
社区交流群
二维码
科研交流群
商业服务