$swint vision model training data
收藏官方服务:
资源简介:
Training language models to see using strings that represent pixelized images.
训练语言模型(language models)借助表征像素化图像(pixelized images)的字符串来实现视觉感知能力
提供机构:
Zenodo创建时间:
2024-03-30

Training language models to see using strings that represent pixelized images.
训练语言模型(language models)借助表征像素化图像(pixelized images)的字符串来实现视觉感知能力