DRSet
收藏资源简介:
DRSet是由华中科技大学和中南民族大学联合构建的RGB-D指代多目标跟踪数据集,包含187个场景的同步RGB图像与深度图,以及240条语言描述(含56条深度相关描述)。该数据集通过系统化标注流程整合静态属性与动态行为特征,旨在解决传统RGB指代跟踪中因缺乏显式3D空间信息导致的语义歧义问题。其多模态特性支持机器人交互、自动驾驶等需融合视觉-语言-几何信息的应用场景,为3D感知的指代跟踪研究提供基准评估平台。
DRSet is an RGB-D referring multi-object tracking dataset jointly constructed by Huazhong University of Science and Technology and South-Central Minzu University. It contains synchronized RGB images and depth maps across 187 scenes, as well as 240 linguistic descriptions, including 56 depth-related ones. This dataset integrates static attributes and dynamic behavioral features through a systematic annotation pipeline, aiming to address the semantic ambiguity issue in traditional RGB referring tracking caused by the lack of explicit 3D spatial information. Its multimodal characteristics support application scenarios that require fusion of visual, linguistic and geometric information, such as robot interaction and autonomous driving, providing a benchmark evaluation platform for referring tracking research in 3D perception.
DRMOT数据集概述
数据集基本信息
- 数据集名称:DRMOT (RGBD Referring Multi-Object Tracking)
- 核心内容:一个用于RGBD参考多目标跟踪的数据集和框架
- 关联项目:与CRMOT项目类似(https://github.com/chen-si-jia/CRMOT)
数据集特点与目的
- 主要特点:整合RGB图像、语言描述和深度信息
- 解决的核心问题:解决仅依赖RGB图像和语言描述时,在深度相关空间描述下产生的歧义问题
- 技术优势:利用深度线索消除空间模糊性,实现准确的目标定位并保持时间身份一致性
数据与代码发布状态
- 发布计划:若论文被接受,将在一个月内完全开源DRSet数据集和DRTrack框架
- 开源内容:包括代码和模型权重
相关资源
- 论文地址:https://arxiv.org/pdf/2602.04692
- 示意图:展示了RMOT与DRMOT的对比(图示内容见数据集页面)

- 1DRMOT: A Dataset and Framework for RGBD Referring Multi-Object Tracking华中科技大学; 中南民族大学 · 2026年



