thanhhieu2004/vivos-vie-speech2text
收藏资源简介:
这是一个名为vivos-vie-speech2text的越南语语音转文本数据集,主要用于语音识别任务。数据集包含音频文件和对应的转录文本,音频采样率为16000Hz。特征包括:audio(音频数据)、transcription(处理后的转录文本)和raw_transcription(原始转录文本)。数据集分为训练集(train)和测试集(test),其中训练集有11420个样本,测试集有1000个样本。总下载大小约为1.78GB,数据集大小约为1.81GB。该数据集可能源自VIVOS项目,适用于构建和评估越南语语音识别模型。
This is a Vietnamese speech-to-text dataset named vivos-vie-speech2text, primarily designed for speech recognition tasks. The dataset includes audio files and corresponding transcriptions, with audio sampled at 16000Hz. Features consist of: audio (audio data), transcription (processed transcription text), and raw_transcription (original transcription text). It is split into train and test sets, with 11420 examples in the train set and 1000 examples in the test set. The total download size is approximately 1.78GB, and the dataset size is approximately 1.81GB. This dataset likely originates from the VIVOS project and is suitable for building and evaluating Vietnamese speech recognition models.




