遇见数据集

negfir/speech_commands_pitch_200hz

收藏
Hugging Face2025-11-11 更新2025-11-15 收录
官方服务:

资源简介:

该数据集是一个包含音频文件和对应标签的多功能数据集,适用于语音识别和命令控制等任务。数据集中的音频采样率为16000Hz,标签包括肯定、否定回答、方向指令、数字、动物名称等分类。数据集分为训练集、测试集和验证集,为模型的训练和评估提供了充足的资源。

This dataset is a multifunctional dataset containing audio files and corresponding labels, suitable for speech recognition and command control tasks. The audio sampling rate in the dataset is 16000Hz, and the labels include categories such as affirmations, negations, directional commands, numbers, animal names, etc. The dataset is divided into training, testing, and validation sets, providing abundant resources for model training and evaluation.

提供机构:
negfir
二维码
社区交流群
二维码
科研交流群
商业服务