遇见数据集

Aratako/LiquidAI-Hackathon-Tokyo-SFT-Data

收藏
Hugging Face2025-10-12 更新2025-10-25 收录
官方服务:

资源简介:

这是一个在Liquid AI Hackathon Tokyo中创建的数据集,用于支持语音识别和语音合成任务的模型训练。数据集包含了英文和日文两种语言的数据,分别用于自动语音识别、文本到语音、语音到语音的任务。每个配置都有对应的训练集,包含了输入ID、注意力掩码和标签信息。

This dataset was created for the Liquid AI Hackathon Tokyo, used for model training in tasks such as speech recognition and text-to-speech. The dataset includes data in both English and Japanese, designed for tasks like automatic speech recognition, text-to-speech, and audio-to-audio. Each configuration has its corresponding training set, which includes input IDs, attention masks, and label information.

提供机构:
Aratako
二维码
社区交流群
二维码
科研交流群
商业服务