arXiv AI

Proto-LeakNet: Towards Signal-Leak Aware Attribution in Synthetic Human Face Imagery

arXiv:2511. 04260v3 Announce Type: replace-cross Abstract: The growing sophistication of synthetic image and deepfake generation models has turned source attribution and authenticity verification into a critical challenge for modern computer vision systems.

arXiv AI
Sep 3

Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics

The paper presents lightweight architectures for detecting GAN-generated synthetic faces, comparing a compact Swin Transformer, pre‑trained Swin‑Tiny and Swin‑Small models, and a hybrid EfficientNet‑B0 + Swin Transformer. Using the 140K Real and Fake Faces dataset, the hybrid model achieved 99% accuracy and 99.44% recall on 5,000 test images, outperforming both pure Swin variants and a CNN‑only baseline. The study demonstrates that combining hierarchical CNN features with shifted‑window self‑attention yields an efficient, computationally lightweight detection method.

By Sejuti Basu, Ashima Sood, Vijay Kumar, Sahil Sharma
arXiv Computer Vision
Sep 16

Unifying Semantic Priors and High-Frequency Traces: Enhancing V-JEPA with Mixture-of-Experts for Robust Synthetic Image Forensics

The paper introduces MoE-JEPA, a dual‑stream deepfake detection model that combines a V‑JEPA backbone with a Residual Mixture‑of‑Experts mechanism and a noise stream branch. It further incorporates a Gated Attention Multiple Instance Learning module to refine spatial semantic understanding. On the SID‑Set benchmark, MoE‑JEPA achieves a new state‑of‑the‑art accuracy of 95.54%, outperforming much larger models.

By Simone Teglia, Irene Amerini
arXiv Machine Learning
Jul 16

When T2I Synthetic Data Backfires: Amplified Privacy Risks in Real-Synthetic Mix Training

arXiv:2607. 13541v1 Announce Type: cross Abstract: To overcome data scarcity and privacy constraints in data collection, it has become standard practice across academia and industry to augment real training data with text-to-image (T2I)-generated synthetic data, a paradigm we term Real-Synthetic Mix-Training (RSMT).

By Na Li, Boyu Kuang, Hongsheng Hu, Liquan Chen, Hyoungshick Kim, Yansong Gao, Anmin Fu
arXiv Computer Vision
Sep 22

Style as Cover: Deep Image Steganography via Stylized Transmission

arXiv:2609.22392v1 Announce Type: new Abstract: Image steganography hides secret message within normal images, with most existing works relying on cover-preserving transmission. However, such a parad...

By Qi Li, Jidong Yang, Huaike Yu, Chunpeng Wang, Suo Gao, Herbert Ho-Ching Iu, Yuantian Miao, Bin Ma, Xiao Chen