← Back to all news
arXiv Computer Vision September 23, 2026 By Hongxuan Chen, Wenda Wang, Jiachen Lu, Qirui Shen, Zilong Huang, Lei He, Xinyue Dong, Weixin Huang

Evidence-gated multimodal parsing and vectorization of architectural floor plans

Read the original on arXiv Computer Vision →

The Flow has not summarised this story yet — read it at arXiv Computer Vision.

  • multimodal
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Computer Vision
Sep 1

MIRAGE-CAD: Construction-Mediated Multimodal Generation of Executable CAD Programs

arXiv:2608.28669v1 Announce Type: new Abstract: Recovering an executable parametric CAD program from an observed object is fundamentally ambiguous, because the same final geometry can result from dif...

By Jizong Zhan
multimodal
More like this →
arXiv AI
Jun 6

MUSE: Benchmarking Manufacturable, Functional, and Assemblable Text-to-CAD Generation

arXiv:2605. 28579v2 Announce Type: replace Abstract: Large language models (LLMs) have recently advanced text-driven 3D generation, yet Text-to-CAD remains far from supporting industrial product design.

By Xiaoyu Dong, Zhi Li, Xiao-Ming Wu
llmsmultimodalbenchmarkssafety
More like this →
arXiv AI
Jun 19

BIM-Edit: Benchmarking Large Language Models for IFC-Based Building Information Modeling

arXiv:2606. 20146v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to computer-aided design (CAD) to generate design artifacts from textual instructions.

By Bharathi Kannan Nithyanantham, Clemens Kujat, Tobias Sesterhenn, Stefan Telgmann, J\"orn Pl\"onnigs, Stefan L\"udtke, Christian Bartelt
llmsbenchmarks
More like this →
arXiv AI
Jun 10

Architect-Ant: Editable Automatic Furnishing of Architectural Floor Plans

arXiv:2606. 10953v1 Announce Type: new Abstract: Furnished floor plans are fundamental to real estate visualization, interior design, and architectural workflows.

By Fedor Rodionov, Aleksandar Cvejic, Michael Birsak, John Femiani, Peter Wonka
llmsfine-tuningmultimodalsafety
More like this →
arXiv AI
Jun 19

CADBench: A Multimodal Benchmark for AI-Assisted CAD Program Generation

arXiv:2605. 10873v2 Announce Type: replace-cross Abstract: Recovering editable CAD programs from images or 3D observations is central to AI-assisted design, but progress is difficult to measure because existing evaluations are fragmented across datasets, modalities, and metrics.

By Anna C. Doris, Jacob Thomas Sony, Ghadi Nehme, Era Syla, Amin Heyrani Nobari, Faez Ahmed
multimodalbenchmarks
More like this →
arXiv AI
Jul 20

DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings

arXiv:2607. 15418v1 Announce Type: new Abstract: We introduce DrawingVQA, the first benchmark designed to evaluate multimodal large language models (MLLMs) on real-world construction drawings -- a core media in architecture, civil, and many other engineering practices.

By Yoonhwa Jung, Junryu Fu, Mani Golparvar-Fard
llmsmultimodalbenchmarks
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea