Hugging Face Trending Papers

DTIF: Robust Loop Closure Detection via Delaunay Triangle Topology in Complex Forests

Read the original on Hugging Face Trending Papers →

Accurate forest inventory and large-scale mapping are essential for ecosystem monitoring and sustainable forest management. Multiple low-cost edge platforms enable efficient large-area data acquisition, but merging independently constructed local maps in GNSS-denied understory environments still requires initialization-free loop closure detection and global registration.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.

arXiv Computer Vision
Sep 22

M3GA-Wild: A Large-Scale Dataset and Benchmark for Multi-Modal Multi-session Ground-to-Aerial Place Recognition in Forests

M3GA-Wild is a new benchmark for multi-modal, multi-session ground-to-aerial place recognition in forests, combining synchronized RGB imagery and LiDAR from ground traversals with high‑resolution aerial imagery and multi‑altitude LiDAR over 370 hectares. The dataset includes accurate geo‑referenced 6‑DoF poses and spans 36 km of forest traversals, enabling systematic evaluation of visual, LiDAR, cross‑modal, and multi‑modal methods. Baseline experiments show LiDAR outperforms vision‑only approaches under severe viewpoint changes, while current multi‑modal fusion offers limited gains due to poor cross‑modal alignment, highlighting challenges in cross‑platform localisation and domain gaps.

By Ethan Griffiths, Maryam Haghighat, Simon Denman, Clinton Fookes, Milad Ramezani
arXiv Computer Vision
Aug 28

SelectAnyTree: A Promptable Instance Segmentation Model for 3D Forest LiDAR Point Clouds

SelectAnyTree is a promptable instance segmentation model designed for 3D forest LiDAR point clouds, enabling users to delineate individual trees with a few clicks. The architecture comprises a sparse voxel scene encoder, a click‑to‑query prompt encoder, and a state‑space query decoder that produces tree masks in linear time, requiring only 19.4 M parameters. Across seven forest regions and an independent dataset, the model achieves a 79.9 % IoU for a single‑click target tree, outperforming existing promptable baselines and requiring the fewest clicks to reach accuracy targets.

By Trung Thanh Nguyen, Daniel Lusk, Kilian Gerberding, Janusch Vajna-Jehle, Tuan-Anh Vu, Duc Viet Le, Tu Vo, Phi Le Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide, Julian Frey, Teja Kattenborn
arXiv Computer Vision
Sep 21

PointLAM: Local Attentive Mamba for Efficient Point-based 3D Object Detection

PointLAM introduces a new point-based 3D object detection architecture that addresses efficiency and fidelity trade-offs inherent in LiDAR point cloud processing. It employs a Laplacian Point Sampler (LPS) to accelerate downsampling while preserving foreground structure, and a Local Hadamard Aggregator (LHA) that replaces costly continuous interactions with a topology‑aware gating mechanism. Combined with Bi‑Directional Mamba layers, the resulting Local Attentive Mamba (LAM) block delivers competitive performance on nuScenes and Waymo datasets, outperforming voxel‑based competitors in detecting small objects and handling extreme sparsity with a smaller computational footprint.

By Xuanming Shang, Weijia Zhang, Chao Ma