arXiv:2606. 30822v1 Announce Type: cross Abstract: In this paper, we attempt to enhance the theoretical understanding of convolutional neural networks (CNNs) as feature extractors in classification tasks by analyzing them through the lens of Cover's function-counting theory.
By Konstantin H\"aberle, Helmut B\"olcskei
arXiv:2012. 01982v3 Announce Type: replace Abstract: This paper proposes a standard way to represent sparse tensors.
By Wuming Pan
arXiv:2406. 08966v3 Announce Type: replace Abstract: The separation power of a machine learning model refers to its ability to distinguish between different inputs and is often used as a proxy for its expressivity.
By Marco Pacini, Xiaowen Dong, Bruno Lepri, Gabriele Santin
arXiv:2510. 15814v2 Announce Type: replace-cross Abstract: Universality results for equivariant neural networks remain rare.
By Marco Pacini, Mircea Petrache, Bruno Lepri, Shubhendu Trivedi, Robin Walters
The paper develops a theory for relocating a finite number of compact sets in ℝ^n to arbitrary target domains using diffeomorphisms of ℝ^n. It proves that any such collection can be embedded differentiably into ℝ^{n+1} so that the images become linearly separable. The authors apply this result to show that compact datasets in ℝ^n can be made linearly separable by width‑n deep neural networks with Leaky‑ReLU, ELU, or SELU activations, and that mutually disjoint compact datasets can be separated in ℝ^{n+1} by a width‑(n+1) DNN.
By Xiao-Song Yang, Xuan Zhou, Qi Zhou
arXiv:2607. 11938v1 Announce Type: cross Abstract: This book is about the mathematical foundations of data science.
By Afonso S. Bandeira, Amit Singer, Thomas Strohmer