Virtual KITTI
收藏资源简介:
Virtual KITTI 是一个逼真的合成视频数据集,旨在学习和评估用于多个视频理解任务的计算机视觉模型:对象检测和多对象跟踪、场景级和实例级语义分割、光流和深度估计。 Virtual KITTI 包含 50 个高分辨率单目视频(21,260 帧),这些视频是从城市环境中的五个不同虚拟世界在不同成像和天气条件下生成的。这些世界是使用 Unity 游戏引擎和一种新颖的从真实到虚拟的克隆方法创建的。这些照片般逼真的合成视频会自动、准确且完全地注释 2D 和 3D 多对象跟踪,并在像素级别使用类别、实例、流和深度标签(参见下面的下载链接)。
Virtual KITTI is a photorealistic synthetic video dataset developed to train and evaluate computer vision models for a range of video understanding tasks: object detection and multi-object tracking, scene-level and instance-level semantic segmentation, optical flow, and depth estimation. It contains 50 high-resolution monocular videos (totaling 21,260 frames) generated from five distinct virtual urban environments under varying imaging and weather conditions. These virtual worlds were created using the Unity game engine and a novel real-to-virtual cloning approach. These photorealistic synthetic videos are automatically, accurately and fully annotated with 2D and 3D multi-object tracking labels, alongside pixel-level category, instance, flow, and depth labels (see the download links below).

- Virtual KITTI数据集首次发布,旨在提供一个高度逼真的虚拟环境,用于计算机视觉和自动驾驶研究。
- Virtual KITTI数据集首次应用于深度学习模型的训练,特别是在语义分割和目标检测任务中。
- Virtual KITTI数据集被广泛用于研究论文中,成为评估和比较不同算法性能的标准数据集之一。
- Virtual KITTI数据集的2.0版本发布,增加了更多的场景和光照条件,提升了数据集的多样性和复杂性。
- Virtual KITTI数据集被用于多个国际会议和竞赛中,推动了自动驾驶和计算机视觉领域的技术进步。



