arXiv Computer Vision

SNAP3D: Physically Grounded 3D Parts for Assembly from a Single Image

arXiv Computer Vision
Aug 25

OmniCAD: A Large-Scale Benchmark for 3D Spatial Reasoning in Robotics Assemblies

arXiv:2608.22637v1 Announce Type: new Abstract: Recent vision-language models (VLMs) show strong capabilities in robotic perception and spatial reasoning, yet their ability to reason about complex me...

By Mingjia Wang, Taiting Lu, Ziwei Dong, Sisong Bei, Jingying Zeng, Runze Liu, Kaiyuan Lin, Hongxing Pan, Kai Zhang, Yizheng Hou, Yangshoudu Zheng, Chenchen Guo, Weiyuan Meng, Shubin Lyu, Zhijun Zheng, Dexu Wang, Xinyu Bai, Shurui Qian, Zhangzixin, Mengyu Pan, Guoliang Shi, Ling Ma, Yifan Yang, Qi He, Yi-Chao Chen, Yincheng Jin, Sung-Liang Chen, Mahanth Gowda
arXiv Machine Learning
Sep 2

CADKnitter: Compositional CAD Generation from Text and Geometry Guidance

CADKnitter is a compositional CAD generation framework that uses geometric-guiding cues to steer diffusion sampling, enabling the creation of complementary CAD parts that satisfy both geometric constraints of an existing model and semantic constraints from a text prompt. The authors introduce KnitCAD, a dataset of over 310,000 CAD models paired with textual prompts and assembly metadata to support training and evaluation. Experiments show that CADKnitter outperforms state‑of‑the‑art baselines by a clear margin.

By Tri Le, Khang Nguyen, Baoru Huang, Tung D. Ta, Anh Nguyen
Hugging Face Trending Papers
Aug 13

SCULPT: Subtractive Composition for 3D Part Generation

Part-aware 3D generation aims to create digital assets that are coherent as complete objects while exposing structural parts for editing, material assignment, animation, and reuse. Existing methods impose this structure outside the native generation loop: segmentation-based methods partition an already generated shape, while additive methods synthesize parts from predefined layouts, boxes, or tokens and then reconcile them into a whole.