eureka1500/IFAO-lalm
收藏官方服务:
资源简介:
该数据集是一个包含353,554个训练样本的英语音频相关文本数据集,用于语言模型训练,涉及音频描述或响应生成任务。数据来源于文件captionstew400k_response_original.jsonl,并标注有audio和LALM标签,表明其专注于音频内容和语言模型应用。
This dataset is an English audio-related text dataset with 353,554 training examples, designed for language model training and tasks such as audio description or response generation. It is sourced from the file captionstew400k_response_original.jsonl and tagged with audio and LALM, indicating a focus on audio content and language model applications.
提供机构:
eureka1500


