Hugging Face Trending Papers

Scale Matters: Adaptive Granularity Selection for Cross-Species 3D Plant Organ Segmentation

The paper introduces AGS-PlantSeg, a few‑shot 3D plant organ segmentation method that uses the frozen Utonia foundation model and Adaptive Granularity Selection (AGS) to dynamically choose optimal spatial granularity for each plant. By extracting tailored geometric features for a lightweight MLP head, AGS-PlantSeg achieves superior cross‑species generalization, reaching an average mIoU of 88.9% and outperforming fixed‑granularity baselines by 2.5 points across PLANesT‑3D, Pheno4D, and Crops3D datasets. The approach requires minimal annotated data yet competes with fully supervised, plant‑specific architectures.

arXiv Computer Vision
Sep 3

PlantC2USeg: Cross-Scale Consistent Pre-Training for Few-Shot Unified Plant Point Cloud Segmentation

PlantC2USeg is a deep transfer‑learning framework that uses cross‑scale consistency learning and an information‑restricted decoder to improve plant point cloud segmentation. It achieves state‑of‑the‑art performance on Soybean3D and ShapeNet Part, and demonstrates strong few‑shot generalization across species and sensing conditions. The method reduces the need for large annotated datasets and lowers adaptation overhead for new plant species.

By Yu Tian, Xintong Jiang, Jan Franklin Adamowski, Shiv O. Prasher, Shangpeng Sun
Hugging Face Trending Papers
Jul 2

The Turning Point of 3D Plant Phenotyping: 3D Foundation Models Enable Minute-to-Second Cross-Crop Reconstruction and Beyond

3D plant phenotyping is notoriously known to be procedure-complicated and of low throughput due to the extensive multi-view imaging, the fragile 3D reconstruction pipeline, and the additional cost from reconstructed geometry to phenotypic extraction. These limitations are further amplified in low-cost data acquisition, where smartphone videos or sparsely sampled multi-view images provide limited view overlap and self-occlusion.

arXiv Machine Learning
Jun 15

NEST3D: A High-Resolution Multimodal Dataset of Sociable Weaver Tree Nests

arXiv:2606. 14562v1 Announce Type: cross Abstract: Sociable weaver nests function as complex ecological structures offering thermoregulatory microhabitats and sustaining diverse species; however, datasets used in prior studies lack fine-grained 3D structural detail.

By Constanza A. Molina Catricheo, Simon Boeder, Ting-Jia Guo, Giacomo May, Cl\'ement Berthelot, Devis Tuia, Friedrich Fedor Reinhard, Fabio Remondino, Benjamin Risse
arXiv Computer Vision
Sep 18

GAPrompt++: Multi-Granular Geometry-Aware Point Cloud Prompt for 3D Vision Model

GAPrompt++ is a multi-granular geometry-aware prompting method designed to adapt pre-trained 3D vision models to downstream tasks efficiently. It introduces a Point Shift Prompter for multi-scale geometric feature extraction, a Keypoint Prompter for local geometric saliency, and a Prompt Propagation mechanism to embed these cues throughout the model hierarchy. Experiments demonstrate that GAPrompt++ outperforms other prompting-based PEFT methods and even surpasses full fine-tuning while using less than 2% trainable parameters, and the authors provide two new challenging benchmarks for future research.

By Zixiang Ai, Zhenyu Cui, Yufei Guo, Wenwen Qiang, Lei Chen, Jiwen Lu, Jiahuan Zhou
arXiv AI
Aug 24

AT-ViT: Area-Targeted Multi-View Vision Transformer with Cross-Attention and Multi-Scale Patching for Plant Trait Recognition in Herbarium Images

AT‑ViT is a dual‑branch Vision Transformer that processes both raw herbarium scans and their segmentation masks through a multi‑scale, multi‑view cross‑attention fusion. It uses a mask‑guided patch weighting scheme to emphasize plant regions and suppress background artifacts, thereby encouraging plant‑centric representations. In trait classification tasks such as leaf base shape and thorns, AT‑ViT consistently outperforms baselines, improves spatial attention grounding (IoU_p +15.66 to +18.03 pp, IoU_b –27.92 to –31.02 pp), and shows greater robustness to synthetic background perturbations, surpassing ResNet101 by up to +32.32 accuracy points and CrossViT by up to +5.07 points. whyItMatters":"The model addresses shortcut learning caused by background cues in herbarium images, leading to more accurate and interpretable plant trait recognition."

By Amani Sedrat, Takieddine Chehhat, Youcef Sklab, Hanane Ariouat, Abderrazak Sebaa, Eric Chenin, Jean-Daniel Zucker, Edi Profiti
Hugging Face Trending Papers
Aug 4

Multimodal Plant Root Phenotyping with Integration of 3D Skeleton Extraction and Language Analysis

Plant root phenotyping is fundamental to understanding below-ground structures, optimizing crop management, and improving agricultural sustainability. This paper presents a multimodal robotic AI framework that integrates 3D skeleton extraction with language-guided reasoning for interpretable and data-efficient root analysis.

arXiv Computer Vision
Aug 31

Denoising-Aware Temporal Point Cloud Completion for 3D Crop Architecture Recovery and Phenotypic Trait Extraction

The paper introduces SynthCrop4D, a synthetic dataset of temporally evolving plant point clouds that includes controllable noise, occlusion, and complete geometry for benchmarking reconstruction methods. It proposes a two‑stage pipeline combining spatial denoising with an Adaptive Temporal PoinTr model to recover missing regions from self‑occlusion, achieving significant improvements in reconstruction quality on both SynthCrop4D and the real Pheno4D dataset. The completed point clouds are further used to extract phenotypic traits such as plant height, canopy width, and convex hull volume, demonstrating the pipeline’s utility for high‑throughput crop phenotyping.

By Mrudul Mittal, Soumyashree Kar