arXiv Machine Learning

Few-Shot Specific Emitter Identification via Integrated Complex Variational Mode Decomposition and Spatial Attention Transfer

The paper presents a new approach for specific emitter identification (SEI) that combines an integrated complex variational mode decomposition algorithm with a temporal convolutional network and a spatial attention mechanism. This method improves feature extraction from complex-valued signals and adaptively emphasizes informative segments, leading to higher identification accuracy. Experiments show the model reaches 96% accuracy using only 10 symbols and no prior knowledge, demonstrating its effectiveness in low‑data scenarios.

arXiv Computer Vision
Sep 22

CIG-MAE: Cross-Modal Information-Guided Masked Autoencoder for Self-Supervised WiFi Sensing

CIG-MAE is a self‑supervised framework for WiFi‑based human action recognition that uses a cross‑modal masked autoencoder to reconstruct both amplitude and phase of Channel State Information. It introduces an adaptive, information‑guided masking strategy that focuses on high‑density time‑frequency regions and employs a Barlow Twins regularizer to align cross‑modal representations without negative samples. Experiments on three public datasets show that CIG‑MAE outperforms state‑of‑the‑art SSL methods and even surpasses a fully supervised baseline, highlighting its data efficiency, robustness, and generalization.

By Gang Liu, Yanling Hao, Yixuan Zou
arXiv AI
Sep 1

Frequency Selective Neural Networks as a Foundation Architecture for Time Series Learning

The paper introduces the Frequency Selective Neural Network (FSNN), a new foundation architecture for time‑series learning that embeds advanced signal‑processing mathematics into its neural topology. By using a fully differentiable Wiener‑like filter bank optimized with complex‑domain backpropagation, FSNN autonomously discovers and isolates the precise physical modes of a given task, thereby avoiding the spectral entanglement that plagues CNNs, RNNs, and Transformers. Extensive evaluations show that FSNN achieves state‑of‑the‑art predictive performance, attaining 77.0 % average accuracy on the 10 multivariate UEA datasets and leading all major metrics on the imbalanced PTB‑XL ECG benchmark, while converging directly on physically meaningful frequency bands such as the cardiac QRS complex.

By Hui Huang, Ye Sun, Shiyan Hu
arXiv Computer Vision
Sep 25

A Hybrid CNN--State-Space--Attention Backbone with Joint-Embedding Predictive Pretraining for 12-Lead ECG Classification

The paper presents a hybrid CNN–state‑space–attention backbone designed for 12‑lead ECG classification, combining early waveform tokenization, mixed temporal dynamics modeling, and late global attention. It introduces an ECG‑oriented Joint‑Embedding Predictive Pretraining (JEPA) that samples span masks at latent resolution and predicts clean latent targets via a momentum encoder, avoiding waveform reconstruction. Experiments on CPSC2018, Chapman‑Shaoxing, and PTB‑XL, with pretraining on ~350K unlabeled CODE‑15 recordings, demonstrate strong supervised baselines and improved transfer, especially in low‑label scenarios and with LoRA adaptation.

By Yakoub Bazi, Sarah Aljuhani, Mohamad M. Al Rahhal, Mansour Zuair, Naif Alajlan
arXiv AI
Jun 24

From Spatial to Spectral: An Efficient, Frequency-Guided Feature Representation Learner for Small Object Detection

arXiv:2606. 23825v1 Announce Type: cross Abstract: Efficient small object detection is bottlenecked by the inherent feature scarcity of tiny targets, which is further aggravated by operations of spatial-domain detectors that indiscriminately discard critical high-frequency details.

By Yuhan Rui, Shihan Qiao, Yibin Lou, Mingxi Yu, Yutong Wan, Yanqiao Chen, Dongsheng Hou, Zhen Cao, Athena Zhuoming Zhong, Qi Hao