arXiv · 2610.05874
Global Communication or Graph-Specific Memory?
Abstract
Scalable Graph Transformers are commonly trained and evaluated on static large graphs in a transductive setup. Many scalable Graph Transformer components can be formulated as a constant-size shared memory, similar to virtual nodes, providing compressed information about the whole graph. The counterpart of these models in language models and other domains is justified as the input changes, and this mechanism learns to compress some useful information about the input. In transductive learning on a single fixed graph, however, any shared memory can be seen as a constant at test time. This raises the question of what exactly this shared memory does in this static setup. We give preliminary evidence that optimizing a shared memory directly performs similarly to global communication methods, and so normal local message-passing models can embed similar information in their weights. Thus, these settings may be a poor fit for evaluating global communication in graph neural networks.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hamed Shirzad, Danica J. Sutherland. 2026-10-05. Global Communication or Graph-Specific Memory?. https://arxiv.org/abs/2610.05874
Cite the original work for its findings. Save a collection to share your selection of sources.