Aratako/LiquidAI-Hackathon-Tokyo-SFT-Data
收藏官方服务:
资源简介:
这是一个在Liquid AI Hackathon Tokyo中创建的数据集,用于支持语音识别和语音合成任务的模型训练。数据集包含了英文和日文两种语言的数据,分别用于自动语音识别、文本到语音、语音到语音的任务。每个配置都有对应的训练集,包含了输入ID、注意力掩码和标签信息。
This dataset was created for the Liquid AI Hackathon Tokyo, used for model training in tasks such as speech recognition and text-to-speech. The dataset includes data in both English and Japanese, designed for tasks like automatic speech recognition, text-to-speech, and audio-to-audio. Each configuration has its corresponding training set, which includes input IDs, attention masks, and label information.
提供机构:
Aratako


