arXiv AI

HiRA-CAM: Preserving Fine-Grained Spatial Relevance in Gradient-Based Visual Explanations

arXiv:2608. 19407v1 Announce Type: cross Abstract: Deep Learning models can include billions of parameters or more, making it difficult to explain their internal transformations and outputs.

arXiv Machine Learning
Jun 9

Analysis of Information Theory for Explainable AI

arXiv:2507. 09092v2 Announce Type: replace-cross Abstract: With the intervention of machine vision in our crucial day to day necessities including healthcare and automated power plants, attention has been drawn to the internal mechanisms of convolutional neural networks, and the reason why the network provides specific inferences.

By Ram S Iyer
arXiv Computer Vision
3d ago

MIMONet: Multi-scale Input and Multi-scale Output Network for Salient Object Detection

MIMONet is a saliency detection model that uses multi‑scale inputs and outputs to better handle objects of varying sizes. It processes three differently sized images through separate encoder branches that exchange information, allowing each branch to learn size‑variation knowledge from the others. A Multi‑scale Perception module further refines features, and a Joint Saliency Loss ensures consistent, well‑preserved boundaries across the multiple saliency maps produced.

By Zhaojian Yao, Wei Gao, Tiesong Zhao, Hui Yuan, Sam Kwong
Hugging Face Trending Papers
Jun 24

Expresso-AI: Explainable Video-Based Deep Learning Models for Depression Diagnosis

Given the widespread prevalence of depression and its consequential impact on individuals and society, it is crucial to obtain objective measures for early diagnosis and intervention. As a multidisciplinary topic, these objective measures should be interpretable and accessible to health care professionals, ensuring effective collaboration and treatment planning in the realm of mental health care.

arXiv Machine Learning
Jul 28

Same Predictions, Different Reasons: The Effect of Quantization on Model Explanations

arXiv:2607. 22872v1 Announce Type: new Abstract: Post-training quantization (PTQ) has become a practical solution for deploying deep learning models on resource-constrained edge devices by compressing high-precision floating-point weights into low-precision representations without requiring retraining.

By Kazi Kamruzzaman Rabbi, Md. Zami Al Zunaed Farabe, M. Sohel Rahman
arXiv Computer Vision
5d ago

GradAttn: Replacing Fixed Residual Connections with Task-Modulated Attention Pathways

GradAttn replaces fixed residual connections in deep ConvNets with attention‑controlled gradient pathways, allowing the network to dynamically weight shallow texture features and deep semantic representations. The method extracts multi‑scale CNN features at different depths and regulates them through self‑attention, leading to improved performance over ResNet‑18 on five of eight evaluated datasets, including a +11.07% accuracy gain on FashionMNIST. Analysis of gradient flow shows that controlled instabilities introduced by attention can coincide with better generalization, while positional encoding proves to be dataset‑dependent.

By Soudeep Ghoshal, Himanshu Buckchash