arXiv Computer Vision

Toward a foundation model for forest point clouds

arXiv AI
Aug 3

Leveraging Image Generators to Address Data Scarcity: The Gen4Regen Dataset for Forest Regeneration Mapping

arXiv:2605. 05627v2 Announce Type: replace-cross Abstract: Sustainable forest management relies on precise species composition mapping, yet traditional ground surveys are labour-intensive and geographically constrained.

By Gabriel Jeanson, David-Alexandre Duclos, William Larriv\'ee-Hardy, No\'e Cochet, Mat\v{e}j Boxan, Anthony Desch\^enes, Fran\c{c}ois Pomerleau, Philippe Gigu\`ere
arXiv AI
Jul 7

SilvaScenes: Tree Detection and Species Classification from Under-Canopy Images in Natural Forests

arXiv:2510. 09458v2 Announce Type: replace-cross Abstract: Interest in forestry automation is growing alongside rapid advances in deep learning.

By David-Alexandre Duclos, William Guimont-Martin, Gabriel Jeanson, Arthur Larochelle-Tremblay, Martine Lapointe, Th\'eo Defosse, Fr\'ed\'eric Moore, Philippe Nolet, Fran\c{c}ois Pomerleau, Philippe Gigu\`ere
arXiv Computer Vision
Sep 1

CedarCypress3D: an annotated UAV-LiDAR dataset of individual trees in planted cedar and cypress forests

arXiv:2608.30149v1 Announce Type: new Abstract: Individual tree measurements derived from Light Detection and Ranging (LiDAR) mounted on Unmanned Aerial Vehicles (UAV) provide valuable information fo...

By Katsuto Shimizu (Shikoku Research Center, Forestry and Forest Products Research Institute), Fumiaki Kitahara (Department of Forest Management, Forestry and Forest Products Research Institute), Tomohiro Nishizono (Department of Forest Management, Forestry and Forest Products Research Institute), Hideki Saito (Forestry and Forest Products Research Institute), Masayoshi Takahashi (Department of Forest Management, Forestry and Forest Products Research Institute), Shingo Obata (Hokkaido Research Center, Forestry and Forest Products Research Institute), Shunsuke Tei (Hokkaido Research Center, Forestry and Forest Products Research Institute), Naoyuki Furuya (Department of Forest Management, Forestry and Forest Products Research Institute), Tomoya Goto (Green Kogyo), Eiji Kodani (Department of Forest Management, Forestry and Forest Products Research Institute), Yusuke Yamada (Graduate School of Bioagricultural Sciences, Nagoya University)
arXiv Computer Vision
Aug 31

Bringing SAM to new heights: Leveraging elevation data for tree crown segmentation from drone imagery

The paper introduces BalSAM, a model that combines the Segment Anything Model (SAM) with Digital Surface Model (DSM) elevation data to improve tree crown instance segmentation from high‑resolution drone imagery. Experiments across boreal plantations, temperate forests, and tropical forests show that while off‑the‑shelf SAM does not beat a custom Mask R-CNN, fine‑tuning SAM end‑to‑end and incorporating DSM information yield promising results, especially for plantation sites.

By M\'elisande Teng, Arthur Ouaknine, Etienne Lalibert\'e, Yoshua Bengio, David Rolnick, Hugo Larochelle
arXiv Computer Vision
Aug 31

HyperVision: A Channel-Adaptive Ground-Based Hyperspectral Vision Pre-trained Backbone

HyperVision introduces the first ground‑based hyperspectral pre‑trained backbone, addressing challenges of varying spectral configurations, limited annotations, and dataset diversity. It employs a channel‑adaptive dynamic embedding to unify heterogeneous inputs, a multi‑source pseudo‑labeling strategy combining SAM2 spatial cues with HyperFree spectral details, and cross‑modal knowledge distillation from a pre‑trained RGB vision model. Trained on 15k images from 26 datasets, HyperVision achieves significant improvements—up to 16.3% relative gain in hyperspectral semantic segmentation, 2.1% in object tracking AUC, and 35.5% reduction in salient object detection MAE—while requiring only head‑only adaptation.

By Guanyiman Fu, Jingtao Li, Zihang Cheng, Zhuanfeng Li, Diqi Chen, Yan Xu, Xiangyu Liu, Fengchao Xiong, Jianfeng Lu, Chengrong Chen, Jun Zhou
arXiv AI
Jun 24

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

arXiv:2606. 20189v3 Announce Type: replace-cross Abstract: Leveraging Vision Foundation Models (VFMs) for camera-to-LiDAR knowledge distillation offers a promising solution to the scarcity of annotated data needed to represent the immense geometric and kinematic diversity of real-world autonomous driving (AD).

By Maciej Wozniak, Jesper Ericsson, Hariprasath Govindarajan, Truls Nyberg, Thomas Gustafsson, Patric Jensfelt, Olov Andersson