zhixiangwei/VLM-1B
收藏官方服务:
资源简介:
VLM-1B是一个大规模的图像-文本数据集,使用增强的Qwen2VL模型(通过SFT增强)对图像进行了重新标注,以提高文本描述的对齐性和细节。
VLM-1B is a large-scale image-text dataset that has been recaptioned using an SFT-enhanced Qwen2VL model to enhance the alignment and detail of textual descriptions.
提供机构:
zhixiangwei


