vLAR/PhysInOne
收藏资源简介:
PhysInOne是一个大规模合成数据集,用于视觉物理学习和推理,包含153,810个动态3D场景和200万个带注释的视频,涵盖力学、光学、流体动力学和磁学中的71种物理现象。每个场景都具有复杂的多对象和多物理交互,并带有丰富的注释,包括RGB视频、深度图、对象掩码、3D轨迹、相机姿态、对象网格、材料特性和文本描述。该数据集支持物理感知视频生成、未来帧预测、物理属性估计、运动传递、物理推理和世界模型的研究。
PhysInOne is a large-scale synthetic dataset for visual physics learning and reasoning, containing 153,810 dynamic 3D scenes and 2 million annotated videos across 71 physical phenomena in mechanics, optics, fluid dynamics, and magnetism. Each scene features complex multi-object and multi-physics interactions with rich annotations, including RGB videos, depth maps, object masks, 3D trajectories, camera poses, object meshes, material properties, and textual descriptions. The dataset supports research on physics-aware video generation, future frame prediction, physical property estimation, motion transfer, physical reasoning, and world models.




