arXiv Computer Vision

AniPrO: Interpretable Anime Image Provenance Detection via Multi-Dimensional Semantic Reasoning

arXiv AI
Sep 3

Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensics

The paper introduces a self‑referential retrosynthesis framework for explainable AI provenance forensics that works with fixed‑generator generative models. It uses a jointly optimized encoder‑decoder pair to embed client inputs, generate high‑fidelity outputs, and then verify provenance by comparing the resynthesized image to the original query. The method eliminates the need for watermarking or generator modifications while providing interpretable evidence of a model’s output origin.

By Yijie Lin, Ching-Chun Chang, Isao Echizen, Hui Li, Chin-Chen Chang
arXiv Computer Vision
Aug 31

Abstract4D: A Large-Scale Dataset and Framework for Understanding the Visual Language of Abstract Art

Abstract4D is the largest dataset of abstract paintings, containing over 120,000 images with rich metadata and multi‑dimensional prompts that capture perceptual attributes such as form, color, texture, and composition. The dataset is annotated via a hybrid human–VLM pipeline to ensure quality and consistency. Using Abstract4D, the authors analyze the semantic structure of abstract art through large‑scale embedding visualization and establish benchmark tasks for classification, cross‑modal retrieval, and text‑to‑image generation to evaluate AI models’ perception and reproduction of abstract visual language.

By Haowei Zhang, Yuanpei Zhao, Ji-Zhe Zhou, Mao Li
arXiv AI
Jun 2

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

arXiv:2606. 00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing new challenges to public discourse and societal security.

By Huidong Feng, Wentao Chen, Jie Chen, Xinqi Cai, Ruolong Ma, Yinglin Zheng, Yuxin Lin, Ming Zeng
Hugging Face Trending Papers
Sep 2

Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensics

The paper introduces a self-referential retrosynthesis framework for explainable AI provenance forensics that works with fixed generative models. It uses a jointly optimized encoder-decoder pair to embed client inputs, allowing the generator to produce high-fidelity outputs while enabling round-trip consistency checks. The method eliminates the need for watermarking or generator modifications and provides interpretable evidence linking generated images back to their source inputs.

arXiv AI
Aug 17

A Pathway to General-Purpose Scientific AI: Multimodal Comprehension of Scientific Images

arXiv:2608. 14075v1 Announce Type: new Abstract: Scientific figures and tables encode essential experimental evidence, yet remain difficult for digital libraries and multimodal AI systems to retrieve and interpret.

By Jennifer D'Souza, Fahad Ahmed, Cecilia Andrea Bustamante Andrade, Lina Frolova, Poorani Gnanasambandan, Dilshad Hussain, Muhammad Uzair Khan, Nkembeng Kevin Nkengfoa, Paul Praveen J., Fabio Priante, Sjoerd Franciscus van der Werf, Thomas Frederik Jan van Roeden