arXiv Machine Learning By Orazio Pontorno, Mattia Litrico, Luca Guarnera, Mario Valerio Giuffrida, Sebastiano Battiato

$\mu$Flow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors

Read the original on arXiv Machine Learning →

arXiv:2606. 30528v1 Announce Type: cross Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy and security.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 3

Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics

The paper presents lightweight architectures for detecting GAN-generated synthetic faces, comparing a compact Swin Transformer, pre‑trained Swin‑Tiny and Swin‑Small models, and a hybrid EfficientNet‑B0 + Swin Transformer. Using the 140K Real and Fake Faces dataset, the hybrid model achieved 99% accuracy and 99.44% recall on 5,000 test images, outperforming both pure Swin variants and a CNN‑only baseline. The study demonstrates that combining hierarchical CNN features with shifted‑window self‑attention yields an efficient, computationally lightweight detection method.

By Sejuti Basu, Ashima Sood, Vijay Kumar, Sahil Sharma
arXiv Computer Vision
Sep 16

Unifying Semantic Priors and High-Frequency Traces: Enhancing V-JEPA with Mixture-of-Experts for Robust Synthetic Image Forensics

The paper introduces MoE-JEPA, a dual‑stream deepfake detection model that combines a V‑JEPA backbone with a Residual Mixture‑of‑Experts mechanism and a noise stream branch. It further incorporates a Gated Attention Multiple Instance Learning module to refine spatial semantic understanding. On the SID‑Set benchmark, MoE‑JEPA achieves a new state‑of‑the‑art accuracy of 95.54%, outperforming much larger models.

By Simone Teglia, Irene Amerini