SenseNova-SI-800K
收藏资源简介:
SenseNova-SI-800K 是一个多模态基础模型数据集,旨在提升空间智能能力。作为 SenseNova-SI 系列的一部分,该数据集包含 80 万个多样化的数据样本,用于训练和验证空间智能模型。数据集采用 JSONL 格式存储,每个条目包含唯一标识符、对话记录和图像路径。通过训练,模型在多个空间智能基准测试中表现出显著提升。该数据集是 SenseNova-SI-8M 的一个子集,用于研究数据规模对模型性能的影响。
SenseNova-SI-800K is a multimodal foundation model dataset designed to enhance spatial intelligence capabilities. As part of the SenseNova-SI series, this dataset contains 800,000 diverse data samples for training and validating spatial intelligence models. The dataset is stored in JSONL format, with each entry containing a unique identifier, conversation records, and image paths. Models trained on this dataset have exhibited significant performance improvements across multiple spatial intelligence benchmark tests. This dataset is a subset of SenseNova-SI-8M, which is utilized to investigate the impact of data scale on model performance.
SenseNova-SI-800K 数据集概述
基本信息
- 数据集名称: SenseNova-SI-800K
- 发布方: SenseNova
- 许可证: Apache-2.0
- 主要语言: 英语 (en)
- 数据规模: 100K < n < 1M
- 任务类别: 视觉问答、问答
- 数据格式: Parquet (通过
SenseNova-SI-800K.parquet文件提供) - 数据拆分: 训练集
数据集简介
SenseNova-SI-800K 是一个旨在提升多模态基础模型空间智能能力的数据集。它是更大规模数据集 SenseNova-SI-8M 的一个高效子集,用于研究数据规模扩展规律。该数据集通过系统化的空间能力分类法构建,包含多样化的数据样本。
核心内容与结构
数据集采用 JSONL 格式组织,每个数据条目包含三个主要字段:
id: 数据样本的唯一标识符。conversations: 对话轮次列表。每轮对话是一个包含from(说话者身份,如 human 或 gpt)和value(文本内容)的字典。在value中,<image>占位符用于标记图像插入位置。image: 图像路径字符串列表。路径是相对于根数据目录的相对路径。
图像占位符 <image> 的数量与 image 字段中列出的图像数量相匹配。
关联模型与评估
- 关联模型: 使用该数据集训练得到的模型示例为 SenseNova-SI-1.1-InternVL3-8B-800K。
- 性能表现: 该模型在 VSI、MMSI、MindCube-Tiny、ViewSpatial、SITE 等多个空间智能基准测试上相较于基础模型有显著提升,并与强基线模型性能相当。
- 评估工具: 建议使用 EASI 工具对训练后的模型在主流的空间智能基准上进行评估。
相关资源
- 论文: Scaling Spatial Intelligence with Multimodal Foundation Models
- 代码仓库: SenseNova_SI
- 评估排行榜: EASI-Leaderboard
引用信息
如需引用,请使用提供的 BibTeX 条目。




