GradAttn: Replacing Fixed Residual Connections with Task-Modulated Attention Pathways
Read the original on arXiv Computer Vision →GradAttn replaces fixed residual connections in deep ConvNets with attention‑controlled gradient pathways, allowing the network to dynamically weight shallow texture features and deep semantic representations. The method extracts multi‑scale CNN features at different depths and regulates them through self‑attention, leading to improved performance over ResNet‑18 on five of eight evaluated datasets, including a +11.07% accuracy gain on FashionMNIST. Analysis of gradient flow shows that controlled instabilities introduced by attention can coincide with better generalization, while positional encoding proves to be dataset‑dependent.
Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.