REVIEW 3 cited by
Exphormer: Sparse Transformers for Graphs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Graph transformers have emerged as a promising architecture for a variety of graph learning and representation tasks. Despite their successes, though, it remains challenging to scale graph transformers to large graphs while maintaining accuracy competitive with message-passing networks. In this paper, we introduce Exphormer, a framework for building powerful and scalable graph transformers. Exphormer consists of a sparse attention mechanism based on two mechanisms: virtual global nodes and expander graphs, whose mathematical characteristics, such as spectral expansion, pseduorandomness, and sparsity, yield graph transformers with complexity only linear in the size of the graph, while allowing us to prove desirable theoretical properties of the resulting transformer models. We show that incorporating Exphormer into the recently-proposed GraphGPS framework produces models with competitive empirical results on a wide variety of graph datasets, including state-of-the-art results on three datasets. We also show that Exphormer can scale to datasets on larger graphs than shown in previous graph transformer architectures. Code can be found at \url{https://github.com/hamed1375/Exphormer}.
Forward citations
Cited by 3 Pith papers
-
Fused3S: Fast Sparse Attention on Tensor Cores
A fused tensor-core sparse attention kernel (SDDMM, softmax, SpMM) that achieves large speedups over prior baselines on H100 and A30.
-
SFi-Former: Sparse Flow Induced Attention for Graph Transformer
SFi-Former replaces dense graph transformer attention with sparse flows from an l1-regularized energy minimization, improving long-range graph benchmark accuracy and generalization.
-
Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures
A review of Relational Deep Learning that formalizes relational entity graphs, catalogs datasets and methods, and advocates for unified GNN and foundation model research directions.
Discussion (0). Continue with ORCID to comment.