arXiv AI By Zhihao Zhan, Le Tao, Yifei Tian, Xin Liu, Jie Yuan

ForestQuery: Boundary-Aware and Spatially Anchored Query Learning for Unified Forest Point Cloud Segmentation

Read the original on arXiv AI →

ForestQuery is a new framework for unified forest point cloud segmentation that incorporates boundary-aware and spatially anchored query learning. It explicitly models boundary uncertainty to improve instance query construction and uses learnable 3D anchors to encode forest vertical stratification for semantic queries. Experiments on public benchmarks and a real‑world dataset show consistent gains in both individual‑tree and semantic segmentation.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
Sep 22

Toward a foundation model for forest point clouds

arXiv:2609.24787v1 Announce Type: new Abstract: Forest inventories increasingly rely on artificial intelligence (AI) models to derive forest attributes from large-scale 3D point clouds. Current model...

By Yuanwen Yue, Stefano Puliti, Damien Robert, Atakan Topalo\u{g}lu, Binbin Xiang, Maciej Wielgosz, Jan Dirk Wegner, Rasmus Astrup, Christian Rupprecht, Konrad Schindler
arXiv Computer Vision
Aug 28

SelectAnyTree: A Promptable Instance Segmentation Model for 3D Forest LiDAR Point Clouds

SelectAnyTree is a promptable instance segmentation model designed for 3D forest LiDAR point clouds, enabling users to delineate individual trees with a few clicks. The architecture comprises a sparse voxel scene encoder, a click‑to‑query prompt encoder, and a state‑space query decoder that produces tree masks in linear time, requiring only 19.4 M parameters. Across seven forest regions and an independent dataset, the model achieves a 79.9 % IoU for a single‑click target tree, outperforming existing promptable baselines and requiring the fewest clicks to reach accuracy targets.

By Trung Thanh Nguyen, Daniel Lusk, Kilian Gerberding, Janusch Vajna-Jehle, Tuan-Anh Vu, Duc Viet Le, Tu Vo, Phi Le Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide, Julian Frey, Teja Kattenborn
Hugging Face Trending Papers
Aug 3

SpatialQuery: Benchmarking Geometry-Grounded Multi-Instance Spatial Reasoning in Vision-Language Models

Vision-language models (VLMs) achieve strong semantic understanding but remain unreliable in metric spatial reasoning, particularly when queries require comparing multiple instances of the same object category. We study this problem through the Closest-Instance Distance Query (CIDQ), where a model must identify the nearest visible candidate to a unique reference object and estimate their gravity-aligned floor-plane distance.