arXiv AI By Stefano Samele, Eugenio Lomurno, Teodora Jovanovic, Sanjay Shivakumar Manohar, Alberto Crivellaro, Matteo Matteucci

A Structured Benchmark for Text-Guided Anomaly Detection: When Language Stops Conditioning the Decision

Read the original on arXiv AI →

arXiv:2606. 01992v1 Announce Type: cross Abstract: Industrial anomaly detection has historically been a unimodal task.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
Sep 25

Industrial Anomaly Detection via Defect-Grounded Reasoning in Visual Latent Space

The paper introduces Anomaly‑LR, a defect‑grounded latent reasoning framework for industrial anomaly detection that builds a global understanding of an image and then refines anomaly‑relevant representations directly in visual latent space. It also presents IAD‑LR‑22K, a new instruction dataset with 22,228 image‑question pairs and detailed annotations. Experiments demonstrate that Anomaly‑LR outperforms comparable‑scale methods on multiple IAD benchmarks without needing external references or tools.

By Jaron Yeh, Yen-Wei Chang, Jiang Liu, Shao-Yuan Lo
arXiv AI
3d ago

TED:Text-Axis Evidence Decomposition for Prompted Anomaly Localization

The paper introduces TED (Text-Axis Evidence Decomposition), a post‑hoc scoring method that improves anomaly localization in CLIP‑based detectors without altering the backbone or prompts. TED evaluates whether ambiguous responses are better supported by defect patches or normal patches, thereby distinguishing true defects from visually complex normal regions. Experiments show that TED significantly enhances pixel‑level localization across frozen VLM backbones and adapted hosts, especially under hard‑false‑positive competition.

By JinYoung Kim, Geonho Kim, GiJeong Park, Geonu Lee, YoungJoon Yoo