arXiv Machine Learning By Fernando Dupin da Cunha Mello (Stricto Sensu Department, SENAI CIMATEC University, Salvador, Bahia, Brazil), Prashant Kumar (Global Centre for Clean Air Research), Erick G. Sperandio Nascimento (Stricto Sensu Department, SENAI CIMATEC University, Salvador, Bahia, Brazil)

An Input-Frugal Deep Learning Framework for Weather-Driven National Crop-Yield Forecasting: A Case Study of Brazilian Soybean

Read the original on arXiv Machine Learning →

The paper introduces a lightweight deep learning framework that forecasts Brazilian soybean yields using only routine weather data and two simple static inputs (crop year and agro-environmental label). Across 20 seasons, transformer-based models achieved the highest accuracy, outperforming traditional ridge regression and a moving‑average baseline by nearly 48%. Ablation studies show that the static inputs and spatial expansion improve performance without adding complexity, and SHAP analysis highlights the importance of crop year and weather variables in driving yield variations.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Aug 19

Evaluating and improving crop-yield forecasting methods during extreme drought

The study evaluates crop‑yield forecasting methods for the 2012 Midwestern US drought, comparing non‑deep learning machine learning models with a deep learning model (VITA) using 16 meteorological predictors. It highlights challenges such as distributional dissimilarity between training and test data, spatial and temporal sparsity, and demonstrates that sample weighting and feature selection improve non‑deep learning models but not VITA. The work contrasts deep versus non‑deep learning approaches and shows how modifications can mitigate issues arising from extreme drought conditions.

By Shrey Gupta, Yi Ming, George Mohler
Hugging Face Trending Papers
Jul 22

Forecasting the Number of Harvest-ready Fruits of Sweet Peppers Using Multimodal Time-Series Data

Accurate yield forecasting at the individual-plant level is critical for precision agriculture and supply-chain planning, yet public datasets capturing both visual growth dynamics and per-plant measurement labels are scarce. In this paper, we introduce a novel, annotated image time-series dataset of 691 sweet pepper plants monitored over two growing seasons, comprising 4837 images with per-plant fruit counts categorized by maturity.

arXiv Computer Vision
Sep 24

AgroBench: A Reproducible Multimodal Benchmark for Weakly Supervised Crop Yield Learning from County Statistics and Pixel Observations

AgroBench is a reproducible benchmark that converts U.S. county-level crop yield statistics into weakly supervised pixel‑level crop time series. The data generation pipeline fuses USDA yield data with land cover masks, Sentinel‑2 and Sentinel‑1 imagery, climatic variables, and terrain information to produce multimodal sequences for individual crop pixels across the growing season. The benchmark includes over 13 million observations from 788,654 crop pixels, covering 5,107 county‑year combinations for five major U.S. crops from 2017 to 2024, and establishes a Leave‑One‑Year‑Out evaluation protocol with baseline machine learning results.

By Udaiveer Singh, Rajiv Ranjan, Shashank Tamaskar, Dharmendra Saraswat