arXiv:2508.08966v2 Announce Type: replace
Abstract: The attention mechanism lies at the core of the transformer architecture, providing an interpretable model-internal signal that has motivated a gro...
By Marte Eggen, Jacob Lysn{\ae}s-Larsen, Inga Str\"umke
arXiv:2609.37656v1 Announce Type: new
Abstract: Large vision-language models (LVLMs) exhibit strong reasoning capabilities, yet the visual and textual evidence supporting the generated responses rema...
By Bowen Yuan, Danny Wang, Ruihong Qiu, Zijian Wang, Zi Huang
arXiv:2606. 03885v1 Announce Type: new Abstract: Feature attribution methods explain predictions by assigning importance scores to input features.
By Kieran A. Murphy, Shameen Shrestha
arXiv:2605. 16651v2 Announce Type: replace-cross Abstract: Explanation mechanisms are increasingly used to support transparency and trust in vision-language models (VLMs), particularly in settings where model decisions require human oversight.
By Narges Babadi, Hadis Karimipour
arXiv:2606. 07180v1 Announce Type: cross Abstract: The growing demand for transparency in automated decision-making has propelled eXplainable Artificial Intelligence (XAI) to the forefront of machine learning research.
By Arthur Hoarau, Chenrui Zhu, Vu Linh Nguyen
arXiv:2604.05819v2 Announce Type: replace-cross
Abstract: Interpreting the decisions of complex computer vision models is crucial to establish trust and accountability, especially in safety-critical...
By David Schinagl, Christian Fruhwirth-Reisinger, Alexander Prutsch, Samuel Schulter, Horst Possegger
arXiv:2608. 16259v1 Announce Type: cross Abstract: The rapid progress of image generation models calls for AI-generated image (AIGI) detectors that are not only accurate but also explainable and reliable.
By Bowen Deng, Jiahui Zhan, Yikun Ji, Haozhen Yan, Jianfu Zhang
arXiv:2605. 28215v2 Announce Type: replace Abstract: In-context learning (ICL) enables multimodal large language models (MLLMs) to classify images from a few labelled examples.
By Carmen Quiles-Ram\'irez, Leticia L. Rodr\'iguez, Nicol\'as Martorell, Natalia D\'iaz-Rodr\'iguez
arXiv:2505. 03201v4 Announce Type: replace-cross Abstract: Integrated Gradients (IG) is a widely used attribution method in explainable AI, particularly in computer vision applications where reliable feature attribution is essential.
By Kien Tran Duc Tuan, Tam Nguyen Trong, Son Nguyen Hoang, Khoat Than, Anh Nguyen Duc
ProToMEx is a new explainability framework that uses Probabilistic Topic Models to learn latent topics representing high‑level reasons behind a classifier’s decisions, moving beyond simple feature attribution. It provides both global and local explanations, revealing multiple co‑existing reasons for individual predictions. Empirical results show that ProToMEx achieves comparable fidelity to SHAP and LIME while being 30–40× faster on standard tabular and synthetic datasets.
By Athina Georgara, Adarsh Valoor, Sarvapali D. Ramchurn
Det‑LIME is a detector‑aware, multi‑instance adaptation of LIME designed to explain black‑box object detectors used in marine mammal research. It generates instance‑specific, box‑aligned explanations by weighting detections, applying a proximity kernel, and using IoU‑based matching to track instances across perturbations. Evaluated on aerial drone imagery of harbor seals and a seabird case study, Det‑LIME outperformed vanilla LIME, Stabilized LIME, Deterministic LIME, and gradient‑based methods in Attribution Ratio and Max Saliency Hit Rate, offering higher‑resolution, instance‑aware explanations that aid debugging, data augmentation, and modeling improvements.
By Jiayi Zhou, David W. Johnston, Brinnae Bent
arXiv:2411. 05698v3 Announce Type: replace-cross Abstract: Convolutional Neural Networks (CNNs) have shown remarkable performance in image classification.
By Antonio De Santis, Riccardo Campi, Matteo Bianchi, Marco Brambilla