stat.MLSep 30, 2026

Transferable Graph Metanetworks

Authors: Yuxin Ma, Adir Dayan, Yam Eitan, Haggai Maron, Soledad Villar

Organizations: Department of Applied Mathematics and Statistics, Johns Hopkins University · Technion · NVIDIA

Abstract

A weight space network (or metanetwork) takes the weights of another neural network as input and predicts properties of it. Most prior work trains such models on input networks of one or a few fixed sizes and evaluates them in-distribution. The few attempts at out-of-distribution size generalization remain limited in scope and have achieved only modest success. Consequently, the potential efficiency gains of training on small networks and evaluating on much larger ones remain largely unrealized. We propose Transferable Graph Metanetworks, which extend the graph metanetwork paradigm with a set of modifications that make performance transferable across input networks of different widths. The modifications follow two principles: invariance to the ways in which networks of different widths represent the same function, and continuity, such that weights representing similar functions receive similar predictions. We further study whether size generalization is possible for input networks trained independently from random initialization. Empirically, our modifications significantly improve size generalization on every task we consider. Performance is strongest on input networks trained under the maximal-update parameterization (μμP), where it remains robust up to 42×42\times the training width. Theoretically, we explain these observations with infinite-width limit theory: we prove size-generalization guarantees for our model on μμP-trained inputs, and explain why it can fail under other parameterizations.

Figures & tables

Appendix figures & tables12 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Train Small, Deploy Large: Zero-Shot GNN Transfer Through Geometric Renormalization

    Jul 30, 2026Robert Jankowski, Pedro Almagro-Blanco, Marián Boguñá +2Graph Neural NetworksRenormalization Group

  2. Quasi-Equivariant Metanetworks

    Apr 26, 2026Viet-Hoang Tran, An Nguyen, Benoît Guérand +2Equivariant Neural NetworksSymmetry

  3. Architecture Generalization with MetaNCA

    Jul 8, 2026Meet Barot, Daniel Berenberg, Sina KhajehabdollahiNeural Cellular AutomataBrain-Inspired Synergistic Framework