llm-lab/pixmotrain_caption_7_en
收藏官方服务:
资源简介:
这是一个包含消息内容和角色、图片路径以及数据有效性标记的数据集,适用于训练机器学习模型。数据集分为训练集,共有998235个样本,文件大小为1.22GB。
This dataset includes message content and role, image paths, and data validity markers, suitable for training machine learning models. The dataset is split into a training set with a total of 998,235 samples, with a file size of 1.22GB.
提供机构:
llm-lab


