synthetic-emotional-en-v2
收藏资源简介:
Emotional English TTS 数据集是一个用于文本到语音转换(TTS)任务的情感语音数据集,特别针对跨语言情感迁移到非洲语言的应用场景(MALIBA-TTS)。数据集包含两个配置:默认配置(default)和已审核配置(reviewed)。默认配置包含619个样本,总大小约222MB;已审核配置包含182个样本,总大小约65MB。每个样本包含音频数据、文本内容、情感标签(包括chuckle、crying、fear、laugh、mixed、sigh等)、情感强度、语言信息和模型信息。数据集由Fish Audio S2 Pro生成,主要用于为非洲语言的跨语言情感迁移提供合成训练数据。情感分布方面,包含19个chuckle样本、114个crying样本、79个fear样本、98个laugh样本、103个mixed样本和56个sigh样本。数据集采用MIT许可协议。
The Emotional English TTS dataset is an emotional speech dataset for text-to-speech (TTS) tasks, specifically targeting the application scenario of cross-lingual emotional transfer to African languages (MALIBA-TTS). The dataset includes two configurations: the default configuration and the reviewed configuration. The default configuration contains 619 samples with a total size of approximately 222 MB, while the reviewed configuration has 182 samples with a total size of around 65 MB. Each sample includes audio data, text content, emotion labels (including chuckle, crying, fear, laugh, mixed, sigh, and more), emotional intensity, language information, and model information. The dataset was generated by Fish Audio S2 Pro, and is primarily used to provide synthetic training data for cross-lingual emotional transfer to African languages. In terms of emotional distribution, it contains 19 chuckle samples, 114 crying samples, 79 fear samples, 98 laugh samples, 103 mixed samples, and 56 sigh samples. The dataset is licensed under the MIT License.
数据集概述
基本信息
- 数据集名称: Emotional English TTS Dataset
- 托管地址: https://huggingface.co/datasets/sudoping01/synthetic-emotional-en-v2
- 语言: 英语 (en)
- 许可证: MIT
- 标签: 文本到语音 (text-to-speech)、情感语音合成 (emotional-tts)、副语言学 (paralinguistics)、MALIBA-TTS (maliba-tts)
数据集配置与结构
数据集包含两种配置:
1. 默认配置 (default)
- 数据文件路径: data/train-*
- 训练集样本数: 619
- 训练集大小: 221,959,802 字节 (约 212 MB)
- 下载大小: 216,275,218 字节 (约 206 MB)
2. 已审核配置 (reviewed)
- 数据文件路径: reviewed/train-*
- 训练集样本数: 182
- 训练集大小: 65,261,229.77221325 字节 (约 62 MB)
- 下载大小: 67,567,824 字节 (约 64 MB)
数据特征
所有配置均包含以下特征字段:
- audio: 音频数据 (audio 类型)
- text: 文本内容 (string 类型)
- emotion: 情感标签 (string 类型)
- intensity: 情感强度 (string 类型)
- language: 语言 (string 类型)
- model: 模型信息 (string 类型)
数据生成与内容
- 生成工具: 使用 Fish Audio S2 Pro 生成。
- 总样本量: 469 个样本(基于默认配置的分布统计)。
情感分布
- chuckle (轻笑): 19 个样本
- crying (哭泣): 114 个样本
- fear (恐惧): 79 个样本
- laugh (大笑): 98 个样本
- mixed (混合情感): 103 个样本
- sigh (叹息): 56 个样本
数据集目的
为 MALIBA-TTS 的跨语言情感迁移至非洲语言提供合成的情感训练数据。




