跳到正文
返回论文列表
cs.CV提交于 已译

Tetris3D:由可拼合物体构成的 3D 场景生成

Tetris3D: 3D Scene Generation With Objects That Fit Together

Jaeyeong Kim · Jinhyuk Jang · Jongmin Lee · Kyehong Park · Seungryong Kim

中文摘要

Tetris3D 是一个面向单图像 3D 场景重建的生成式框架,目标是恢复出在物理与几何层面彼此一致的物体,从而形成一个完整的场景。现有方法通常独立生成各个物体,或仅以隐式方式将它们耦合在一起,难以对相邻物体之间细粒度的空间兼容性提供有效约束。针对这一问题,Tetris3D 显式地将每个物体的生成条件建立在周围物体的几何与物理关系之上,从而使该物体在形状与位姿上均能在场景内保持几何与物理上的合理性。此外,论文还提出了 ComOb 数据集,基于物理仿真构建,包含 120 万个场景,涵盖多种类别物体之间的物理交互,并提供逐物体网格与成对物理关系标注。在合成与真实场景上的系统实验表明,Tetris3D 即使在交互区域被遮挡时,也能恢复出连贯一致的物体形状与位姿,并在生成质量与物理稳定性两个维度上均达到当前最优水平。

关键要点

  1. 01问题:现有单图像 3D 场景重建方法对物体独立生成或隐式耦合,缺乏对相邻物体细粒度空间兼容性的约束
  2. 02方法:Tetris3D 显式以周围物体的几何与物理关系为条件,引导其形状与位姿在场景中保持几何与物理合理
  3. 03数据:提出 ComOb,基于物理仿真的 120 万场景数据集,带逐物体网格与成对物理关系标注
  4. 04结果:在交互区域被遮挡时仍能恢复一致的物体形状与位姿,生成质量与物理稳定性达 SOTA
  5. 05局限:摘要未具体说明方法局限或失败情形

解读

尚无解读。

原始英文摘要

arXiv:2610.10539v1 Announce Type: new Abstract: We propose Tetris3D, a generative framework for single-image 3D scene reconstruction that recovers objects which are physically and geometrically coherent as a scene. Existing methods often generate objects independently or couple them implicitly, providing limited guidance for ensuring fine-grained spatial compatibility between neighboring objects that interact with one another. To address this, we explicitly condition the generation of each object on the geometry of surrounding objects and their physical relationships, guiding its shape and pose to remain geometrically and physically plausible within the scene. Moreover, we introduce ComOb, a physics simulation-based dataset of 1.2M scenes featuring physical interactions across diverse object categories, with per-object meshes and pairwise physical relation annotations. Comprehensive experiments on synthetic and realworld scenes show that Tetris3D recovers coherent object shapes and poses even when interacting regions are occluded, and achieves state-of-the-art performance in both generation quality and physical stability.

同方向论文 · cs.CV

查看全部 →