arXiv Computer Vision

Latent Commonality Expectation-Maximisation for Box-supervised Tree Crown Instance Segmentation

arXiv AI
Jul 7

SilvaScenes: Tree Detection and Species Classification from Under-Canopy Images in Natural Forests

arXiv:2510. 09458v2 Announce Type: replace-cross Abstract: Interest in forestry automation is growing alongside rapid advances in deep learning.

By David-Alexandre Duclos, William Guimont-Martin, Gabriel Jeanson, Arthur Larochelle-Tremblay, Martine Lapointe, Th\'eo Defosse, Fr\'ed\'eric Moore, Philippe Nolet, Fran\c{c}ois Pomerleau, Philippe Gigu\`ere
arXiv Computer Vision
Aug 31

Bringing SAM to new heights: Leveraging elevation data for tree crown segmentation from drone imagery

The paper introduces BalSAM, a model that combines the Segment Anything Model (SAM) with Digital Surface Model (DSM) elevation data to improve tree crown instance segmentation from high‑resolution drone imagery. Experiments across boreal plantations, temperate forests, and tropical forests show that while off‑the‑shelf SAM does not beat a custom Mask R-CNN, fine‑tuning SAM end‑to‑end and incorporating DSM information yield promising results, especially for plantation sites.

By M\'elisande Teng, Arthur Ouaknine, Etienne Lalibert\'e, Yoshua Bengio, David Rolnick, Hugo Larochelle
arXiv Computer Vision
Aug 28

Detection of Christmas tree plantations from high-resolution aerial imagery. A case study in the French Morvan

The study presents a new approach to detect Christmas tree plantations in high‑resolution aerial imagery, treating the task as a rare‑target semantic segmentation problem. It introduces a Hard Negative Mining strategy that significantly improves precision‑recall performance, achieving an IoU of 0.733 and an F1‑score of 0.846 on a 2020 test set. Temporal transfer experiments demonstrate the model’s ability to generalize across years, while large‑scale validation highlights the challenge posed by the plantations’ small spatial footprint.

By Francesca Razzano, Emanuele Dalsasso, Adrien Baysse-Lain\'e, Silvia Liberata Ullo, Gilda Schirinzi, Jocelyn Chanussot
arXiv Computer Vision
2d ago

DeepForestVisionV2: Ecology-Driven Taxonomy Expansion for Camera-Trap Monitoring in African Tropical Forests

arXiv:2606.20223v2 Announce Type: replace Abstract: Camera-trap monitoring in African tropical forests increasingly extends beyond closed-canopy interiors to riverbanks, clearings, and park edges. Am...

By Hugo Magaldi, Theau d'Audiffret, Etienne Francois Akomo-Okoue, Bala Amarasekaran, Naomi Anderson, Claire Auger, Noemie Cappelle, Daniel Cornelis, Raphael Cornette, Tobias Deschner, Gabriel Dubus, Davy Fonteyn, Rosa M. Garriga, Jennifer Hatlauf, Innocent Kasekendi, Raymond Katumba, Aram Kazandjian, Alfred Ngomanda, Stephan Ntie, Simone Pika, Xavier Rufray, Harold Rugonge, John Justice Tibesigwa, Peter van Lunteren, Hadrien Vanthomme, Joeri A. Zwerts, Sabrina Krief
arXiv Computer Vision
Aug 28

Learning Woody Clearing With Loss Alignment for Zero-Shot Regrowth and Woody Segmentation

The paper presents a deep learning approach for detecting woody clearing using bitemporal Sentinel‑2 imagery from New South Wales, Australia. By introducing a loss‑scaling coefficient, the authors align the model’s objective with end‑user metrics, boosting precision and recall. They further demonstrate zero‑shot transfer to woody regrowth and segmentation tasks, achieving significant error reductions and high F1 scores through image augmentation and generation techniques.

By Kal Backman, Jared Wood, Adam Roff
arXiv Machine Learning
Jun 26

Self-Supervised Tree-level Biomass Estimation in Urban Environments From Airborne LiDAR and Optical Observations

arXiv:2606. 26194v1 Announce Type: cross Abstract: Urban tree biomass remains less spatially explicitly quantified than biomass in managed forests because many estimates rely on inventories or coarse products that cannot resolve individual crowns or fine-scale heterogeneity.

By Jose Bermudez (McMaster University, Hamilton, Ontario, Canada), Zilong Zhong (McMaster University, Hamilton, Ontario, Canada), Dominic Cyr (, Environment and Climate Change Canada, Montreal, Quebec, Canada), Camile Sothe (Planet Labs PBC, San Francisco, California, USA), Alemu Gonsamo (McMaster University, Hamilton, Ontario, Canada)
arXiv Computer Vision
Aug 28

SelectAnyTree: A Promptable Instance Segmentation Model for 3D Forest LiDAR Point Clouds

SelectAnyTree is a promptable instance segmentation model designed for 3D forest LiDAR point clouds, enabling users to delineate individual trees with a few clicks. The architecture comprises a sparse voxel scene encoder, a click‑to‑query prompt encoder, and a state‑space query decoder that produces tree masks in linear time, requiring only 19.4 M parameters. Across seven forest regions and an independent dataset, the model achieves a 79.9 % IoU for a single‑click target tree, outperforming existing promptable baselines and requiring the fewest clicks to reach accuracy targets.

By Trung Thanh Nguyen, Daniel Lusk, Kilian Gerberding, Janusch Vajna-Jehle, Tuan-Anh Vu, Duc Viet Le, Tu Vo, Phi Le Nguyen, Yasutomo Kawanishi, Takahiro Komamizu, Ichiro Ide, Julian Frey, Teja Kattenborn