anatol_vyartsinski_pesnya_pra_hleb_output
收藏资源简介:
“Песня пра хлеб”是一个白俄罗斯语自动语音识别(ASR)数据集,属于Ministerskija系列的一部分,该系列是白俄罗斯语有声读物音频与转录文本的对齐集合。数据集基于作家Анатоль Вярцінскі的作品构建。原始有声读物被分割成约15秒的短音频片段,并利用Gemini和两个独立的ASR系统与转录文本进行了高置信度(≥ 0.95)的对齐处理。每个数据样本包含三个字段:`audio`(音频片段)、`text`(对应的转录文本)和`chunk_uid`(片段的唯一标识符)。该数据集在Hugging Face上公开发布了48个样本,总数据库包含69个样本,音频总时长约为14分钟。数据集规模属于1K到10K样本之间,适用于白俄罗斯语的自动语音识别模型训练与评估任务。
“Песня пра хлеб” is a Belarusian automatic speech recognition (ASR) dataset, part of the Ministerskija series, which is a collection of aligned audio and transcriptions from Belarusian audiobooks. The dataset is based on the works of the writer Anatol Vyarciński. In this dataset, the original audiobook is segmented into short audio clips of approximately 15 seconds, and high-confidence (≥ 0.95) alignment with transcriptions is performed using Gemini and two independent ASR systems. Each data sample includes three fields: `audio` (audio clip), `text` (corresponding transcription), and `chunk_uid` (unique identifier for the clip). The dataset is publicly released on Hugging Face with 48 samples, while the total database contains 69 samples, with an audio duration of approximately 14 minutes. The dataset scale falls between 1K to 10K samples and is suitable for training and evaluating automatic speech recognition models for the Belarusian language.
数据集概述:Песня пра хлеб (Anatol Vyartsinski)
- 语言: 白俄罗斯语 (be)
- 许可证: CC0-1.0
- 任务类别: 自动语音识别 (ASR)
- 数据集大小: 1,000 < n < 10,000 行
- 音频时长: 约 14 分钟
数据集来源与内容
该数据集是 Ministerskija 收藏集的一部分,包含白俄罗斯语有声读物的对齐音频片段及其转录文本。音频和文本来源于阿纳托利·维亚尔京斯基的作品《Песня пра хлеб》。
数据结构
每行数据包含以下字段:
- audio:音频片段(约15秒)
- text:转录文本(通过Gemini和ASR对齐生成)
- chunk_uid:片段的唯一标识符
数据处理方法
音频书被分割成短片段,并使用 Gemini 和两个独立的 ASR 系统进行转录对齐。对齐置信度阈值为 ≥ 0.95。
说话人信息
该数据集中的说话人被归类为 speaker_02 说话人集群。其平均相似度评分为 0.901,最接近的数据集是 eryh_maryya_remark_output(相似度 0.95)。此识别结果基于 WavLM-Base+ 模型(余弦相似度,阈值 0.82)。




