遇见数据集

Dataset of Audiovisual Speech for AR Telepresence Studies (Speech Recordings)

收藏
Zenodo2025-03-16 更新2026-05-26 收录
官方服务:

资源简介:

Dataset of speech recordings made in the anechoic chamber "Lampio" at Aalto University.21 Participant ("P1" - "P21") Four parts are included:1) Conversations: Ten different scripted three-part conversations ("C1" - "C10"). Each participant is in two of them. All of the three parts is played by all three participants ("S1" - "S3"). See assignment_conversations.xlsx 2) Harvard_Sets: Sets 25 and 36 of the Harvard sentence lists3) Sentence 1 from List 25 in five different voice levels (from "barely not whispering" to "screaming as loud as you can")4) Native_Language: List 25 translated to native languages of 12 of the participants (French, Finnish, Hebrew, Hindi, Spanish (Mexico), Spanish (Chile), Catalan, Latvian, Italian, Polish, Romanian, German) Each file contain data from three receivers:Ch 1: GRAS 40 HF 1" low-noise meausurement microphone. 1.5 m away from the subjectCh 2: RØDE NT1 large diaphragm condenser microphone. 2 m away from the subjectCh 3: DPA 4060. Attached to the subject's clothesAccompanying video data can be obtained by personal request from nils.meyer-kahlen@aalto.fi

本数据集为阿尔托大学(Aalto University)"Lampio"消声室(anechoic chamber)录制的语音录音数据集,共涵盖21名参与者(编号"P1"至"P21")。 数据集包含四个组成部分: 1. 对话脚本集(Conversations):共10组预设三段式对话(编号"C1"至"C10"),每名参与者参与其中2组。所有三段对话的角色均由编号"S1"至"S3"的三名参与者分别扮演。详细分配规则请参见assignment_conversations.xlsx文件。 2. 哈佛语句集(Harvard_Sets):哈佛语句列表的第25组与第36组。 3. 可变语音强度语句:选取第25组语句列表中的语句1,分别以5种不同语音强度录制(覆盖从"几乎悄无声息"到"全力嘶吼"的范围)。 4. 母语录制内容(Native_Language):将第25组语句翻译为12名参与者的母语,涉及语言包括法语、芬兰语、希伯来语、印地语、墨西哥西班牙语、智利西班牙语、加泰罗尼亚语、拉脱维亚语、意大利语、波兰语、罗马尼亚语、德语。 每份音频文件均包含三路拾音数据: - 通道1:GRAS 40 HF 1英寸低噪声测量麦克风,与受试者间距1.5米; - 通道2:RØDE NT1大振膜电容麦克风,与受试者间距2米; - 通道3:DPA 4060麦克风,粘贴于受试者衣物表面。 如需获取配套视频数据,可通过个人申请的方式联系nils.meyer-kahlen@aalto.fi获取。

提供机构:
Zenodo
创建时间:
2025-03-16
二维码
社区交流群
二维码
科研交流群
商业服务