遇见数据集

Reza2kn/Common-Voice-25-Persian-Cleaned

收藏
Hugging Face2026-05-11 更新2026-05-31 收录
官方服务:

资源简介:

这是一个多语种语音数据集,包含音频文件(采样率为16kHz)及其相关元数据。数据特征包括客户端ID、音频路径、音频数据、句子ID、句子文本、句子领域、赞成票数、反对票数、说话人年龄、性别、口音、变体、地区、片段、持续时间(毫秒)、审核状态和审核错误信息。数据集分为开发集(dev)和验证集(validated),分别包含10个和340,608个样本,总大小约为28.99GB。数据集用于语音处理任务,支持多语言和多样化的说话人特征。

This is a multilingual speech dataset containing audio files with a sampling rate of 16 kHz and their associated metadata. The data features include client ID, audio path, audio data, sentence ID, sentence text, sentence domain, approval votes, disapproval votes, speaker age, gender, accent, variant, region, segment, duration (in milliseconds), review status, and review error information. The dataset is split into the development set (dev) and validation set (validated), which contain 10 and 340,608 samples respectively, with a total size of approximately 28.99 GB. This dataset is designed for speech processing tasks, supporting multilingual scenarios and diverse speaker characteristics.

提供机构:
Reza2kn
二维码
社区交流群
二维码
科研交流群
商业服务