cs.LGOct 5, 2026

Global Communication or Graph-Specific Memory?

Authors: Hamed Shirzad, Danica J. Sutherland

Organizations: University of British Columbia

Abstract

Scalable Graph Transformers are commonly trained and evaluated on static large graphs in a transductive setup. Many scalable Graph Transformer components can be formulated as a constant-size shared memory, similar to virtual nodes, providing compressed information about the whole graph. The counterpart of these models in language models and other domains is justified as the input changes, and this mechanism learns to compress some useful information about the input. In transductive learning on a single fixed graph, however, any shared memory can be seen as a constant at test time. This raises the question of what exactly this shared memory does in this static setup. We give preliminary evidence that optimizing a shared memory directly performs similarly to global communication methods, and so normal local message-passing models can embed similar information in their weights. Thus, these settings may be a poor fit for evaluating global communication in graph neural networks.

Figures & tables

Appendix figures & tables6 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Graph Memory Transformer (GMT)

    Apr 26, 2026Nicola Zanarini, Niccolò Ferrari, Evelina LammaMultiplex Graph TransformersTransformer Encoder

  2. MegaGraph: Towards Efficient Training of Large-Scale Graph Transformers with Automated Hybrid Parallelism

    Sep 28, 2026Tong Qiao, Ao Zhou, Yingjie Qi +2Multiplex Graph TransformersCpu-Gpu Hybrid Designs

  3. Graph Neural Networks Applications Across Domains: All Insights You Need

    Jun 25, 2026Abderaouf BahiGraph Neural NetworksNeural Network