arXiv:2608. 13141v1 Announce Type: cross Abstract: Vision Transformers (ViTs) demonstrate exceptional performance in computer vision but suffer from large parameter counts and quadratic computational complexity, severely limiting their deployment on resource-constrained edge hardware.
By Junseo Kim, Uraz Odyurt, Amirreza Yousefzadeh
arXiv:2608. 11053v1 Announce Type: cross Abstract: The application of computer vision in agriculture has shown significant potential for improving crop monitoring and precision farming.
By Ismail Ismail Tijjani, Sunusi Muhammad Ibrahim, Amina Ibrahim Khaleel, Lanre Olusegun Akinola, Fatima Isa Jibrin, Muhammad Bashir Aliyu, Abdullahi Abdussalam Dalhat, Abdullahi Suiudeen
STA‑Net is a lightweight neural network designed for plant disease classification on edge devices. It combines a training‑free neural architecture search (DeepMAD) to build an efficient backbone with a novel Shape‑Texture Attention Module (STAM) that separates shape and texture processing using deformable convolutions and a Gabor filter bank. On the CCMT plant disease dataset, STA‑Net achieved 89.00% accuracy and 88.96% F1 score with only 401K parameters and 51.1M FLOPs.
By Zongsen Qiu, Jianjun Wang, Yue Zhou, Zibo Zhou, Rui Chen
arXiv:2608.21454v1 Announce Type: new
Abstract: The same fruit appears in a bunch, unpicked, peeled, bagged in plastic, or sliced on a dish, so automated fruit classification in the wild (AFCW) must...
By Subhankar Chattoraj, Sawon Pratiher, Samiran Das, Hubert Konik
arXiv:2606. 03748v1 Announce Type: cross Abstract: Real-time vision demands models that are accurate, efficient, and simple to deploy across diverse hardware.
By Glenn Jocher, Jing Qiu, Mengyu Liu, Shuai Lyu, Fatih Cagatay Akyon, Muhammet Esat Kalfaoglu
arXiv:2606. 01503v1 Announce Type: cross Abstract: Unified vision-language models (VLMs) integrate visual understanding and visual generation within a single autoregressive backbone, but their joint training is computationally expensive and largely overlooked from an efficiency perspective.
By Siyi Chen, Weiming Zhuang, Jingtao Li, Lingjuan Lv