OmniScene
收藏资源简介:
OmniScene是一个新发布的大规模合成数据集,由D-Robotics等机构构建,旨在支持异构相机的度量深度感知研究。该数据集包含约26.6万帧同步六视角图像,总计170万张独立图像,覆盖103个室内外场景,提供高质量渲染和精确标定的相机配置。数据通过合成渲染技术生成,包含多样化的场景布局和逼真的视觉内容,专门用于训练和评估跨相机模型的深度估计能力。该数据集主要应用于计算机视觉和机器人领域,旨在解决异构相机系统中度量深度估计的数据稀缺问题,推动实时空间感知算法的发展。
OmniScene is a newly released large-scale synthetic dataset constructed by institutions including D-Robotics, aiming to support metric depth perception research for heterogeneous camera systems. This dataset contains approximately 266,000 frames of synchronized six-view images, totaling 1.7 million independent images, covering 103 indoor and outdoor scenes, and provides high-quality rendering and precisely calibrated camera configurations. Generated via synthetic rendering technology, the dataset features diverse scene layouts and realistic visual content, and is specifically designed for training and evaluating the depth estimation capabilities of cross-camera models. Primarily applied in the fields of computer vision and robotics, this dataset aims to address the data scarcity issue of metric depth estimation in heterogeneous camera systems, and promote the development of real-time spatial perception algorithms.
数据集名称
OmniScene —— 大规模多视角异构相机合成数据集
基本统计
- 场景数量: 103 个室内外复杂场景
- 运动序列: 564 个
- 总帧数: 约 266K 帧(每帧六视图同步)
- 图像数量: 超过 170 万张已标定图像
相机配置
- 使用固定的六相机模组,包含:
- 4 个 180° 鱼眼相机
- 2 个针孔相机
- 所有相机已标定,提供每帧的 OpenCV 约定内参和外参
场景内容
- 场景来自 Kujiale(酷家乐)和 Unreal Engine 的专业资产
- 覆盖住宅、商业、工业、城市及科幻等类别
- 包含训练集(源自训练场景)、验证集(源自训练场景但单独划分)和测试集(完全未见过的场景)
标注信息
每一帧提供精确的噪声无关度量地面真值:
- 像素级 z 深度
- 有效掩码
- 天空指示器
- 每帧 OpenCV 约定的内参和外参
发布链接
- 数据集已公开发布于 Hugging Face,可通过 🤗 下载链接获取

- 1X-Lens: Real-Time Metric Depth Estimation with Heterogeneous CamerasD-Robotics; 东京大学; 苏州大学 · 2026年




