arXiv AI By Duong M. Nguyen, Tuan Nghia Nguyen, Xuan Truong Nguyen

ENAF: A Multi-Exit Network with an Adaptive Patch Fusion for Large Image Super Resolution

Read the original on arXiv AI →

arXiv:2608. 15349v1 Announce Type: cross Abstract: To accelerate single image super-resolution (SISR) networks on large images (2K-8K), many recent approaches decompose an image into small patches and dynamically determine an execution path according to its difficulty (referred to as a dynamic network).

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
Sep 11

FreeTransformSR: Efficient Lightweight Image Super-Resolution via Free Low-Rank Learnable Transform

FreeTransformSR is a lightweight image super‑resolution network that uses a channel‑wise free low‑rank learnable transform to adaptively modulate features with minimal parameters. It adds a local feature modulation branch with depthwise convolution and a soft complexity adaptive module that fuses local convolution and window self‑attention based on texture characteristics. The model also employs an adaptive intensity modulation strategy and achieves competitive PSNR/SSIM on five benchmark datasets while using only 595K parameters and running faster than competing methods.

By Hongji Li, Yunhui Li
arXiv Computer Vision
Sep 4

ProgResViT: Progressive Resolution and Width for Adaptive Vision Transformers

ProgResViT is an input‑adaptive Vision Transformer that processes images progressively across multiple rounds, starting with a low‑resolution image and a narrow subnetwork and refining the prediction with higher resolution and a wider subnetwork if needed. The method introduces Progress‑Conditioned Soft Gating (PSG) to share a single backbone across rounds while conditioning token fusion and layer outputs on the current round, block, and input resolution. Experiments on DeiT show improved accuracy‑compute trade‑offs compared to adaptive‑width, adaptive‑depth, and dynamic‑token baselines, and the design also benefits self‑supervised DINO representations and downstream semantic segmentation.

By Ali Hojjat, Janek Haberer, Olaf Landsiedel