遇见数据集

youssefkhalil320/MedSynth-Combined

收藏
Hugging Face2026-04-27 更新2026-05-03 收录
官方服务:

资源简介:

该数据集是一个包含对话、笔记和音频的多模态数据集,专门用于训练目的。数据集包括2551个训练样本,总大小约为47.17GB。每个样本包含以下特征:Dialogue(对话文本)、Note(笔记文本)、audio(音频数据,采样率为24000Hz,未解码)和row_idx(整数行索引)。数据集仅提供训练分割,没有验证或测试集。音频特征以原始格式存储,适用于需要处理语音和文本结合的任务,如语音识别、对话生成或多模态学习。

This dataset is a multimodal dataset containing dialogues, notes and audio, specifically designed for training purposes. It includes 2551 training samples with a total size of approximately 47.17 GB. Each sample contains the following features: Dialogue (dialogue text), Note (note text), audio (raw audio data with a sampling rate of 24000 Hz, not decoded), and row_idx (integer row index). Only the training split is provided in this dataset, with no validation or test splits available. The audio features are stored in their raw format, which is suitable for tasks that require processing combined speech and text, such as speech recognition, dialogue generation, or multimodal learning.

提供机构:
youssefkhalil320
二维码
社区交流群
二维码
科研交流群
商业服务