VaseVQA-3D
收藏资源简介:
VaseVQA-3D数据集是首个针对古希腊陶罐分析的3D视觉问答数据集,收集了664个古希腊陶罐的3D模型及其对应的问题-答案数据,并建立了完整的数据构建流程。该数据集旨在解决当前视觉语言模型在3D陶罐文物数据稀缺和领域知识不足的问题,通过领域自适应训练,显著提高了模型在陶罐文物分析方面的性能,为数字遗产保护研究提供了新的技术路径。
The VaseVQA-3D dataset is the first 3D visual question answering (VQA) dataset dedicated to the analysis of ancient Greek vases. It includes 3D models of 664 ancient Greek vases along with their corresponding question-answer pairs, and establishes a complete data construction pipeline. This dataset aims to solve the problems of scarce 3D cultural relic data of vases and insufficient domain knowledge for current vision-language models. Through domain-adaptive training, it significantly improves the performance of models in the analysis of vase cultural relics, providing a new technical approach for digital heritage preservation research.
VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery
数据集概述
- 数据集名称:VaseVQA-3D
- 主要用途:评估3D视觉语言模型在古希腊陶器上的性能
- 核心特点:专注于3D视觉语言模型基准测试
研究领域
- 计算机视觉
- 三维视觉语言模型
- 文化遗产数字化
研究对象
- 古希腊陶器
数据集目标
- 建立评估3D视觉语言模型的基准
- 推动3D视觉语言模型在文化遗产领域的应用

- 1VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery北京大学, 北京交通大学, 拉筹伯大学 · 2025年



