Group Affect and Performance (GAP) Corpus
收藏资源简介:
GAP语料库是一个包含会议记录、注释、合并文件和会议数据的数据集,用于研究群体情感和表现。数据集中的会议和成员都有特定的标识码,会议记录包含在Transcripts/文件夹中,注释在Annotations/文件夹中,合并文件在Merged/文件夹中,会议数据在Group-Individual-Data/文件夹中。此外,还提供了音频记录和其他详细数据,如参与者的特征和会议特征。
The GAP corpus is a dataset comprising meeting transcripts, annotations, merged files, and meeting data, designed for the study of group emotions and performance. Each meeting and member within the dataset is assigned a specific identifier. The meeting transcripts are located in the Transcripts/ folder, annotations in the Annotations/ folder, merged files in the Merged/ folder, and meeting data in the Group-Individual-Data/ folder. Additionally, the dataset includes audio recordings and other detailed data such as participant characteristics and meeting features.
数据集概述
数据集名称
- GAP Corpus:Group Affect and Performance (GAP) Corpus
数据集内容
-
Meeting and Group Member Identification (ID) Codes:
- 会议使用基于数据收集时间顺序的数字编码。
- 小组成员使用五种颜色(蓝、绿、粉、橙、黄)进行编码。
-
Meeting Transcripts:
- 位于 "Transcripts/" 文件夹中的.txt文件。
- 文件名包含会议ID和数据收集的日期时间。
- 每条发言标注了小组成员ID、发言编号及起止时间。
-
Meeting Annotations:
- 位于 "Annotations/" 文件夹。
- 包含决策发言、情感发言、个人排名提及及生存物品提及的标注。
- 每条标注包含ID、编号及对应发言的起止时间。
-
Merged Transcripts and Annotations:
- 位于 "Merged/" 文件夹,包含合并的转录和标注文件。
-
Meeting Data:
- 位于 "Group-Individual-Data/" 文件夹。
- 包含冬季生存任务数据、任务后问卷响应、参与者特征及会议特征。
- 包含两个Excel文件:
- Group-Level Meeting Data:提供会议规模、时长、绝对小组得分等信息。
- Individual-Level Meeting Data:提供个人在大学的年级、性别、语言背景、绝对个人得分等信息。
-
Meeting Audio:
- 音频文件(.wav格式)可在数据集网页单独下载。
- 文件名包含会议ID和录制日期时间。
数据集许可
- Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0)




