arXiv:2607. 27372v1 Announce Type: new Abstract: The deep learning revolution, kicked off by AlexNet, taught us that end-to-end training beats decomposing a problem into hand-designed stages.
By Alexi Gladstone, Heng Ji, Yilun Du
arXiv:2607. 19332v1 Announce Type: new Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching.
By Chirag Vashist, Ke Li
Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques have become more complicated and various beliefs about what drives strong empirical performance have taken hold.
arXiv:2404. 06294v2 Announce Type: replace-cross Abstract: Super-Resolution (SR) is a time-hallowed image processing problem that aims to improve the quality of a Low-Resolution (LR) sample up to the standard of its High-Resolution (HR) counterpart.
By Arkaprabha Basu, Kushal Bose, Sankha Subhra Mullick, Anish Chakrabarty, Swagatam Das
EmbeddGAN introduces a new GAN framework that replaces the traditional discriminator with an embedding network trained to maximize statistical dependence between embeddings and real/fake labels using Gini distance correlation (gCor). The generator simultaneously minimizes this dependence, encouraging real and generated samples to become indistinguishable in the learned low‑dimensional embedding space. Experiments on MNIST, CIFAR‑10, and CelebA show competitive performance and notably more stable training dynamics compared to established baselines.
By MaTais Caldwell, Yixin Chen, Xin Dang, Charles Walter
The paper introduces StyleGANCA, a lightweight neural cellular automata (NCA) based generative adversarial network designed for medical image synthesis. By combining a StyleGAN-inspired mapping network with adaptive style modulation in a multi-scale NCA framework, the model achieves high-quality image generation with far fewer parameters than existing adversarial, variational, diffusion, and NCA baselines. Experiments on BloodMNIST and PathMNIST show competitive FID and KID scores, and the synthetic images preserve class-specific information, effectively supporting downstream multi-class classifier training.
By Anh Thi Luu, Nick Lemke, Anirban Mukhopadhyay