遇见数据集

TucanoBR/ViTucano-Pretrain

收藏
Hugging Face2025-08-15 更新2025-04-12 收录
官方服务:

资源简介:

ViTucano-Pretrain数据集是一个经过葡萄牙语翻译的图像到文本和文本生成任务的数据集,用于训练ViTucano视觉助手模型。该数据集基于LLaVA-Pretrain LCS-558K子集,通过Google翻译API进行了翻译,并具有更加平衡的概念覆盖分布。

The ViTucano-Pretrain dataset is a Portuguese-translated image-to-text and text-generation task dataset used for training the ViTucano visual assistant model. This dataset is based on the LLaVA-Pretrain LCS-558K subset, translated via Googles translation API, and has a more balanced concept coverage distribution.

提供机构:
TucanoBR
二维码
社区交流群
二维码
科研交流群
商业服务