SD-Eval
收藏资源简介:
SD-Eval是一个专注于口语对话理解和生成的基准数据集,特别关注非语言和环境信息。它包含7,303个话语,总计8.76小时的语音数据,数据来源于八个公开数据集,涵盖情感、口音、年龄和背景声音四个视角。
SD-Eval is a benchmark dataset dedicated to spoken dialogue understanding and generation, with a special focus on non-verbal and environmental information. It contains 7,303 utterances, totaling 8.76 hours of speech data sourced from eight public datasets, and covers four perspectives: emotion, accent, age, and background sounds.
SD-Eval 数据集概述
数据集信息
- 许可证: cc-by-nc-4.0
- 任务类别:
- 文本生成
- 问答
- 语言: 英语
- 标签:
- 口语对话生成
- 牛角面包
- 数据规模: 1K<n<10K
数据集配置
- 配置名称: SD-Eval
- 特征:
- utt_id: 字符串
- transcript: 字符串
- audio:
- 采样率: 16000
- info: 字符串
- dataset_name: 字符串
- reply1: 字符串
- reply2: 字符串
- reply3: 字符串
- reply4: 字符串
- reply5: 字符串
数据集描述
- 目标: 多维度评估口语对话理解和生成
- 焦点: 副语言和环境信息
- 数据量: 7,303 个话语,总计 8.76 小时语音数据
- 来源: 来自八个公共数据集,代表四个视角:情感、口音、年龄和背景声音
引用
@article{ao2024sdeval, title = {SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words}, author = {Junyi Ao and Yuancheng Wang and Xiaohai Tian and Dekun Chen and Jun Zhang and Lu Lu and Yuxuan Wang and Haizhou Li and Zhizheng Wu}, eprint={2406.13340}, archivePrefix={arXiv}, primaryClass={cs.CL}, year={2024} }




