arXiv Computer Vision

LettuceVisSim: A Simulator That Generates Lettuce Image Time-series for Vision-Based Reinforcement Learning

arXiv Machine Learning
Aug 10

UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys

arXiv:2608. 06404v1 Announce Type: cross Abstract: Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and management response.

By Junxiong Zhou, Xuechen Li, Chonghao Qiu, Lang Qiao, Xiaowei Jia, Qi Yang, Chishan Zhang, Leikun Yin, Nanshan You, Vipin Kumar, David Mulla, Ce Yang, Zhenong Jin, Licheng Liu
arXiv Machine Learning
Sep 17

MCLC-NET: Multimodal Continual Learning for Leaf Counting

The paper introduces MCLC‑NET, a multimodal continual learning framework for leaf counting that sequentially learns tasks using a memory buffer to retain key samples. It also presents MMLC, a new real‑world dataset containing RGB, depth, and thermal images across different crops and environmental conditions, organized in crop‑wise, time‑wise, and mixed orderings. Experiments show that MCLC‑NET outperforms existing methods on all three task orderings, achieving the lowest average mean squared errors.

By Ruchi Bhatt, Pratibha Kumari, Shreya Bansal, Vedant Agnihotri, Dwarikanath Mahapatra, Mukesh Saini
arXiv Computer Vision
Sep 24

From greenhouse climate to individual leaves: an organ-resolved model of lettuce growth

A unified framework was created to simulate lettuce growth by modeling each leaf’s physiology and structure within a greenhouse environment. The model integrates leaf-level photosynthesis, carbon allocation, and 3‑D plant growth in NVIDIA Isaac Sim, allowing radiation interception to influence growth and vice versa. Validation against greenhouse data shows low prediction errors and demonstrates how variations in light, CO₂, and plant position affect dry weight, leaf number, and tipburn incidence.

By Md Hasibur Rahman, Faraz Ahmed, Hafiz Muhammad Bilal, Daniel Wells, Dylan Tobin, Tanzeel U. Rehman
arXiv AI
Sep 24

Using Vision Language Foundation Models to Generate Plant Simulation Configurations via In-Context Learning

The paper presents a benchmark to test whether vision‑language models can produce plant simulation configurations from images using in‑context learning. It focuses on cowpea plot reconstruction, requiring the models to output structured JSON that includes field and plant details. Open‑source multimodal models from the Gemma 4 and Qwen3.5 families are evaluated on synthetic and real drone datasets, using five in‑context methods, and the results show that while VLMs can generate valid JSON and estimate key agronomic metrics, their performance varies and often lags behind dataset baselines.

By Heesup Yun, Isaac Kazuo Uyehara, Earl Ranario, Lars Lundqvist, Christine H. Diepenbrock, Brian N. Bailey, J. Mason Earles
arXiv Computer Vision
Sep 21

Optimizing YOLO27, YOLO26, YOLO11, and YOLOv8 for Fine-Grained Small-Object Detection and Segmentation in Complex Orchard Environments

The paper compares Ultralytics YOLO27, YOLO26, YOLO11, and YOLOv8 for detecting and segmenting small fruit parts in orchard settings. It evaluates five model scales across 30 experiments, finding that YOLO11s-960 and YOLO26s-960 achieve the best mask and box mAP scores while maintaining efficient parameter counts. The study also highlights the difficulty of peduncle detection and provides publicly available code and models for reproducibility.

By Ranjan Sapkota, Manoj Karkee
arXiv AI
Jul 1

LeCropFollow: Latent Space Planning for Navigation in Unstructured Crop Fields

arXiv:2606. 31941v1 Announce Type: cross Abstract: Unstructured navigational features, such as irregular planting or discontinuities, remain the primary failure mode for under-canopy agricultural robots.

By Felipe Tommaselli, Francisco Affonso, Arthur Pompeu, Gianluca Capezzuto, Arun Narenthiran Sivakumar, Girish Chowdhary, Marcelo Becker