arXiv Machine Learning

Any-Dimensional Learning by Sampling

arXiv:2607. 07680v1 Announce Type: cross Abstract: Many machine learning models are defined for inputs of different sizes, such as point clouds containing different numbers of points, sequences of tokens of different lengths, and graphs on different numbers of nodes.

arXiv Statistics ML
2d ago

Transferable Graph Metanetworks

arXiv:2610.00420v1 Announce Type: new Abstract: A weight space network (or metanetwork) takes the weights of another neural network as input and predicts properties of it. Most prior work trains such...

By Yuxin Ma, Adir Dayan, Yam Eitan, Haggai Maron, Soledad Villar
arXiv Machine Learning
2d ago

How Many Samples Are Enough for Learning Across Domains?

The paper investigates how many data samples per domain are needed for effective learning across multiple domains. It derives criteria from learning bounds that reveal an inverse linear relationship between the number of training domains and the required samples per domain, offering theoretical guidance for dataset adequacy and construction. The study also establishes a close link between in-domain learning and out-of-domain generalization through new generalization bounds.

By Hong Zheng
arXiv AI
Aug 25

Which Algorithms Can Graph Neural Networks Learn?

arXiv:2602.13106v2 Announce Type: replace-cross Abstract: In recent years, there has been growing interest in understanding neural architectures' ability to learn to execute discrete algorithms, a li...

By Solveig Wittig, Antonis Vasileiou, Robert R. Nerem, Timo Stoll, Floris Geerts, Yusu Wang, Christopher Morris
arXiv AI
Sep 24

Scalable Subgraph Sampling via Resistance Curvature

The paper introduces a scalable subgraph sampling method that uses resistance curvature to guide the selection of nodes and edges for graph neural network training. It builds on ERC‑LG, a curvature approximation technique that employs Johnson‑Lindenstrauss projections and regularized multi‑GPU batched conjugate gradient solvers, thereby avoiding costly Laplacian pseudoinverse calculations and large embedding storage. Experiments demonstrate that ERC‑LG‑based sampling matches pseudoinverse‑based curvature numerically, runs faster than conjugate‑gradient‑only approaches, and achieves the best mean accuracy on six of seven real‑world node‑classification datasets.

By Chaoqun Fei, Tinglve Zhou, Tianyong Hao, Yangyang Li