PaperScope
LIVE · 2026-10-06 05:40 UTC

Global Communication or Graph-Specific Memory?

Hamed Shirzad, Danica J. Sutherland

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2610.05874 v1
Category
Submitted
2026-10-05

Abstract

Scalable Graph Transformers are commonly trained and evaluated on static large graphs in a transductive setup. Many scalable Graph Transformer components can be formulated as a constant-size shared memory, similar to virtual nodes, providing compressed information about the whole graph. The counterpart of these models in language models and other domains is justified as the input changes, and this mechanism learns to compress some useful information about the input. In transductive learning on a single fixed graph, however, any shared memory can be seen as a constant at test time. This raises the question of what exactly this shared memory does in this static setup. We give preliminary evidence that optimizing a shared memory directly performs similarly to global communication methods, and so normal local message-passing models can embed similar information in their weights. Thus, these settings may be a poor fit for evaluating global communication in graph neural networks.

arXiv abs page · PDF · same-day batch