arXiv:2604. 05324v2 Announce Type: replace Abstract: Statistical evaluation aims to estimate the generalization performance of a model using held-out i.
By Shashaank Aiyer, Yishay Mansour, Shay Moran, Han Shao
arXiv:2606. 23872v1 Announce Type: cross Abstract: As generative models increasingly produce samples that are indistinguishable from human-created content, it becomes difficult to determine whether a given data point was part of a model's natural training set or was generated by the model itself, especially when models memorize and reproduce training data.
By Bihe Zhao, Michel Meintz, Juangui Xu, Franziska Boenisch, Adam Dziedzic
arXiv:2510. 22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models.
By Andreas Floros, Seyed-Mohsen Moosavi-Dezfooli, Pier Luigi Dragotti
arXiv:2402.04355v4 Announce Type: replace-cross
Abstract: We propose a likelihood-free method for comparing two distributions given samples from each, with the goal of assessing the quality of genera...
By Pablo Lemos, Sammy Sharief, Esmeralda S. Whitammer, Salma Salhi, Connor Stone, Laurence Perreault-Levasseur, Yashar Hezaveh
arXiv:2607. 19332v1 Announce Type: new Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching.
By Chirag Vashist, Ke Li
arXiv:2607. 08347v1 Announce Type: cross Abstract: Active testing provides a label--efficient approach to risk estimation by adaptively selecting which test points should be labelled.
By Kianoosh Ashouritaklimi, Valentin Kilian, Daolang Huang, Tom Rainforth, Fran\c{c}ois Caron