HyperVision introduces the first ground‑based hyperspectral pre‑trained backbone, addressing challenges of varying spectral configurations, limited annotations, and dataset diversity. It employs a channel‑adaptive dynamic embedding to unify heterogeneous inputs, a multi‑source pseudo‑labeling strategy combining SAM2 spatial cues with HyperFree spectral details, and cross‑modal knowledge distillation from a pre‑trained RGB vision model. Trained on 15k images from 26 datasets, HyperVision achieves significant improvements—up to 16.3% relative gain in hyperspectral semantic segmentation, 2.1% in object tracking AUC, and 35.5% reduction in salient object detection MAE—while requiring only head‑only adaptation.
By Guanyiman Fu, Jingtao Li, Zihang Cheng, Zhuanfeng Li, Diqi Chen, Yan Xu, Xiangyu Liu, Fengchao Xiong, Jianfeng Lu, Chengrong Chen, Jun Zhou
arXiv:2604.08884v2 Announce Type: replace-cross
Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on RGB image understanding, yet their ability to use spectral evide...
By Xinyu Zhang, Zurong Mai, Qingmei Li, Xiaoya Fan, Zjin Liao, Haoyuan Liang, Yibin Wen, Yuhang Chen, Chan Tsz Ho, Bi Tianyuan, Ruifeng Su, Zihao Qiang, Juepeng Zheng, Jianxi Huang, Yutong Lu, Haohuan Fu
The paper presents the first publicly available hyperspectral imaging (HSI) dataset of shredded black plastics from end‑of‑life vehicle waste, covering four industrial polymers across RGB, VNIR, SWIR, and MWIR modalities. It introduces a multi‑modal spectral‑spatial framework that combines foreground isolation, pixel‑wise classification, and object‑level majority voting, leveraging hyperspectral transformers and chemometric band selection to accurately classify complex black plastics. The study benchmarks nine processing methods—including chemometric, machine learning, and deep learning architectures—providing a reproducible, comprehensive benchmark for industrial hyperspectral object analysis.
By Elias Arbash, Andr\'ea de Lima Ribeiro, Filipa Sim\~oes, Ahmed Jamal Afifi, Aldino Rizaldy, Yuleika Madriz, Samuel Thiele, Sandra Lorenz, Margret Fuchs, Pedram Ghamisi, Paul Scheunders, Richard Gloaguen
arXiv:2608.30537v1 Announce Type: cross
Abstract: Rapid mineral characterization is essential for applications ranging from mineral exploration to industrial ore processing. To this end, Hyperspectra...
By Eleftheria Tetoula-Tsonga (Institute of Communication and Computer Systems, Athens, Greece), George Arvanitakis (Geonova, Athens, Greece), Theodoros Giannakas (Institute of Communication and Computer Systems, Athens, Greece)
arXiv:2603. 06673v2 Announce Type: replace-cross Abstract: Spectroscopic imaging (SI) has become central to heritage science because it enables non-invasive, spatially resolved characterisation of materials in artefacts.
By Shivam Pande, Nicolas Nadisic, Francisco Mederos-Henry, Aleksandra Pizurica
arXiv:2608. 00608v1 Announce Type: new Abstract: Visible and near-infrared (vis-NIR) and mid-infrared (MIR) spectroscopy enable rapid, cost-effective prediction of soil properties.
By Viacheslav Barkov, Jonas Schmidinger, Robin Gebbers, Martin Atzmueller
arXiv:2606. 18661v1 Announce Type: cross Abstract: Intelligent landslide hazard interpretation is critical for disaster prevention, yet current paradigms struggle to simultaneously extract visual features and high-level geoscientific semantics, while general-purpose vision-language models (VLMs) suffer from perceptual limitations and domain hallucinations in complex geological scenarios.
By Chengfu Liu, Dongyang Hou, Junwu Xiang, Cheng Yang, Xuezhi Cui, Zeyuan Wang, Liangtian Liu, Zelang Miao
Rapid mineral characterization is essential for applications ranging from mineral exploration to industrial ore processing. To this end, Hyperspectral Imaging (HSI) has emerged as a promising sensing...
The paper investigates a Multi-Scale Spectral Attention Module (MSAM) for hyperspectral image segmentation in autonomous driving. MSAM uses three parallel 1D convolutions with different kernel sizes (1–11) and adaptive feature aggregation, integrated into UNet’s skip connections. Experiments on urban driving datasets show that MSAM improves mIoU by 2.32% and mF1 by 2.88% over baseline UNet-SC while keeping GPU performance competitive, with optimal kernel combinations varying by dataset.
By Imad Ali Shah, Jiarong Li, Tim Brophy, Martin Glavin, Edward Jones, Enda Ward, Brian Deegan
arXiv:2607. 05207v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is designed to learn generic, transferable representations rather than representations optimized for a single task.
By Rohita Mocharla, Vishal M. Patel
arXiv:2606. 10819v1 Announce Type: cross Abstract: RS-MLLMs enable natural-language understanding and spatial reasoning over earth observation imagery.
By Miaoxin Cai, Guanqun Wang, Wei Zhang, Guangyao Zhou, Yin Zhuang, Tong Zhang, Hao Wang, He Chen, Jun Li
arXiv:2607. 25338v1 Announce Type: new Abstract: Achieving a coherent integration of spectral richness and spatial fidelity remains a central objective in hyperspectral image fusion.
By Chengxin Xie, Qiya Song, Yangbangyan Jiang, Renwei Dian, Xudong Kang