arXiv Computer Vision By Jiarong Li, Imad Ali Shah, Enda Ward, Martin Glavin, Edward Jones, Brian Deegan

Band-Selection Stability and Semantic Segmentation Performance: A Study on Hyperspectral City

Read the original on arXiv Computer Vision →

The paper investigates how stable band‑selection methods are and how that stability relates to semantic segmentation performance on the Hyperspectral City V2 dataset. Six band‑selection techniques were tested on ten different class‑balanced ROI sets, producing 60 top‑25 band subsets. Results show that Sim‑LP has the highest intra‑method stability, and together with JMIM+CSNR it also delivers the best segmentation results, achieving up to 2.01 mIoU improvement and 18–22× faster CPU inference for a 9‑band subset, though stability does not consistently predict segmentation quality.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv Computer Vision
1d ago

Hyperspectral Image Models: Technical Report

The technical report introduces Hyperspectral Image Models, a modular framework that unifies 55 deep‑learning models across six paradigms for hyperspectral remote sensing. It standardizes tensor conventions, evaluation protocols, and dataset handling, integrating 24 benchmark scenes from various sensors and providing tools to avoid train‑test overlap. Experiments across 1,320 model‑scene combinations show that scene difficulty outweighs architecture, with no single paradigm dominating and small models achieving performance comparable to much larger ones.

By Tanishq Rachamalla, Aryan Das, Srishti Kaushik, Swalpa Kumar Roy
arXiv Computer Vision
2d ago

HyperSAM: A Promptable Foundation Model for Hyperspectral Remote Sensing

HyperSAM is a promptable foundation model for hyperspectral remote sensing that integrates a data‑centric synthesis pipeline with a spectral adaptation architecture based on Segment Anything Model 3 (SAM3). The model generates full‑spectrum hyperspectral cubes from high‑resolution multispectral imagery using a physics‑informed abundance‑transfer generator, and employs SAM3‑derived pseudo‑masks for object‑centric supervision. With a frozen SAM3 RGB branch, a trainable hyperspectral encoder, ControlNet‑style feature injection, and a mixture‑of‑experts mask refiner, HyperSAM demonstrates strong generalization across diverse hyperspectral tasks such as classification, anomaly detection, change detection, target detection, and airborne oil‑spill mapping.

By Li Pang, Xinqiao Wu, Jing Yao, Pedram Ghamisi, Jun Zhou, Zhengchao Chen, Deyu Meng, Xiangyong Cao
arXiv AI
Aug 19

Multi-Scale Spectral Attention Module-based Hyperspectral Segmentation in Autonomous Driving Scenarios

The paper investigates a Multi-Scale Spectral Attention Module (MSAM) for hyperspectral image segmentation in autonomous driving. MSAM uses three parallel 1D convolutions with different kernel sizes (1–11) and adaptive feature aggregation, integrated into UNet’s skip connections. Experiments on urban driving datasets show that MSAM improves mIoU by 2.32% and mF1 by 2.88% over baseline UNet-SC while keeping GPU performance competitive, with optimal kernel combinations varying by dataset.

By Imad Ali Shah, Jiarong Li, Tim Brophy, Martin Glavin, Edward Jones, Enda Ward, Brian Deegan
arXiv Machine Learning
Sep 3

Ten Architectures, One Error: Shared Failure Modes in Hyperspectral Classification under Spatially Disjoint Evaluation

The paper critiques the common practice of random pixel splits in hyperspectral image classification, noting that such splits allow test pixels to be adjacent to training pixels, inflating accuracy. It proposes a leakage‑free evaluation protocol that enforces spatial separation based on the model’s receptive field and applies it to ten diverse architectures, finding a significant drop in Macro‑F1 (average 0.147) and substantial changes in model rankings. The study also shows that all ten models misclassify the same pixels, indicating a spectral ambiguity in the data that current methods cannot resolve.

By Ehsan Faghih, Fatemeh Ashrafi, Marguerite Moore, Zahra Saki