遇见数据集

eureka1500/IFAO-lalm

收藏
Hugging Face2026-05-13 更新2026-05-31 收录
官方服务:

资源简介:

该数据集是一个包含353,554个训练样本的英语音频相关文本数据集,用于语言模型训练,涉及音频描述或响应生成任务。数据来源于文件captionstew400k_response_original.jsonl,并标注有audio和LALM标签,表明其专注于音频内容和语言模型应用。

This dataset is an English audio-related text dataset with 353,554 training examples, designed for language model training and tasks such as audio description or response generation. It is sourced from the file captionstew400k_response_original.jsonl and tagged with audio and LALM, indicating a focus on audio content and language model applications.

提供机构:
eureka1500
二维码
社区交流群
二维码
科研交流群
商业服务