arXiv Computer Vision

Bayesian Fusion of Active Contour Models and ConvNet Priors for Standing Dead Tree Segmentation

arXiv AI
Jul 7

SilvaScenes: Tree Detection and Species Classification from Under-Canopy Images in Natural Forests

arXiv:2510. 09458v2 Announce Type: replace-cross Abstract: Interest in forestry automation is growing alongside rapid advances in deep learning.

By David-Alexandre Duclos, William Guimont-Martin, Gabriel Jeanson, Arthur Larochelle-Tremblay, Martine Lapointe, Th\'eo Defosse, Fr\'ed\'eric Moore, Philippe Nolet, Fran\c{c}ois Pomerleau, Philippe Gigu\`ere
arXiv Computer Vision
Aug 31

Bringing SAM to new heights: Leveraging elevation data for tree crown segmentation from drone imagery

The paper introduces BalSAM, a model that combines the Segment Anything Model (SAM) with Digital Surface Model (DSM) elevation data to improve tree crown instance segmentation from high‑resolution drone imagery. Experiments across boreal plantations, temperate forests, and tropical forests show that while off‑the‑shelf SAM does not beat a custom Mask R-CNN, fine‑tuning SAM end‑to‑end and incorporating DSM information yield promising results, especially for plantation sites.

By M\'elisande Teng, Arthur Ouaknine, Etienne Lalibert\'e, Yoshua Bengio, David Rolnick, Hugo Larochelle
arXiv Computer Vision
Aug 28

Learning Woody Clearing With Loss Alignment for Zero-Shot Regrowth and Woody Segmentation

The paper presents a deep learning approach for detecting woody clearing using bitemporal Sentinel‑2 imagery from New South Wales, Australia. By introducing a loss‑scaling coefficient, the authors align the model’s objective with end‑user metrics, boosting precision and recall. They further demonstrate zero‑shot transfer to woody regrowth and segmentation tasks, achieving significant error reductions and high F1 scores through image augmentation and generation techniques.

By Kal Backman, Jared Wood, Adam Roff
arXiv Computer Vision
Sep 17

Category Level 6D Object Pose Estimation from a Single RGB Image using Diffusion

The paper presents a generative framework that estimates category-level 6D pose and 3D size of objects from a single RGB image, using score-based diffusion models to produce a multi-hypothesis pose distribution. It replaces costly likelihood pruning with a Mean Shift approach to isolate the mode as the final pose estimate, achieving state-of-the-art results on the REAL275 benchmark. The method also decouples detection from pose estimation, enabling robust zero-shot generalisation on the Wild6D dataset and extending naturally to video sequences by propagating the pose distribution over time.

By Adam Bethell, Ravi Garg, Ian Reid
arXiv Machine Learning
Jul 8

Conformal Prediction Sets for Instance Segmentation

arXiv:2602. 10045v2 Announce Type: replace-cross Abstract: Current instance segmentation models achieve high performance on average predictions, but lack principled uncertainty quantification: their outputs are not calibrated, and there is no guarantee that a predicted mask is close to the ground truth.

By Kerri Lu, Dan M. Kluger, Stephen Bates, Sherrie Wang
Hugging Face Trending Papers
Jul 25

Investigating the Visual Cues of CNNs for Vascular Segmentation: A Case Study in Microscopy and Fundus Imaging

Vascular segmentation is a standard procedure for clinical diagnosis, yet the specific visual features determining model decisions remain poorly understood. This paper investigates the visual cues Convolutional Neural Networks (CNNs) use to segment blood vessels across two distinct imaging domains: fluorescence microscopy and retinal fundus photography.