Pith. sign in

Paper Citation Record · LEDGER

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators

As of 14 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 0 inbound Pith citation observations for arXiv:2501.01951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.01951 v3

Coverage vector

measured 93 of 93 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:20:59.524307Z

measured 93 of 93 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

93 of 93 outbound references displayed

  • verified exact7
  • verified fuzzy48
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e46f153a-cbc2-41ab-a5d4-9574eadb917d · outbound

This paper cites TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.638267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.638267Z digest=sha256:d8eaf4092b5ff2eaf60c5518f13c5dfd65d09dc6f584dfc5a9578b4098067640

Observation d1e82c4d-b714-4cef-9c66-73a1f30aeca0 · outbound

This paper cites Hardware accel- eration of graph neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Hardware accel- eration of graph neural networks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.644621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.644621Z digest=sha256:26550af047e477b8d608693052f556c3441f2615a6cbb7d55ebb65ed5508a7b0

Observation 3274501f-e5b7-4c70-bcbc-1033bcbd988a · outbound

This paper cites Staleness-Alleviated Distributed GNN Training via Online Dynamic-Embedding Prediction.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Staleness-Alleviated Distributed GNN Training via Online Dynamic-Embedding Prediction

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:21:01.160978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.649361Z digest=sha256:15699c88b01d94e031586c10ebdaa1df459c8ccf23d1bbfa1c2a404b7628bc42

Observation e5c96f4a-508a-488b-b7d7-4d6674f254ff · outbound

This paper cites Pathways: Asynchronous distributed dataflow for ml.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Pathways: Asynchronous distributed dataflow for ml

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.654099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.654099Z digest=sha256:08af96b572024160bead0d39e700ba7813176ef33cf049ec29b141a0cb1a63ef

Observation 7245acf2-1add-4ba5-ab32-3ce031209894 · outbound

This paper cites Distributed Graph Neural Network Training with Periodic Stale Representation Synchronization.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Distributed Graph Neural Network Training with Periodic Stale Representation Synchronization

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:21:01.078748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.660107Z digest=sha256:6585717a187211f9b667adef0517a5729087c4ebd4b4dd66dfe02c9241c43846

Observation 81aa9a1f-8148-42a1-84b6-a145c7c84759 · outbound

This paper cites Dygnn: Algorithm and architecture support of dynamic pruning for graph neural net- works.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Dygnn: Algorithm and architecture support of dynamic pruning for graph neural net- works

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.668865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.668865Z digest=sha256:8c98c4a138c045a19a3601700c65231ecb587049f6cc3fabcfe2b42db1893e45

Observation e951a0f3-c9e8-4880-a91d-24df1d98633d · outbound

This paper cites Graph representation learning: a survey.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Graph representation learning: a survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.675487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.675487Z digest=sha256:250554ef3403600d894093eaff93ba8a53960e3b53d3f42875e7fdcc2211f86b

Observation a900a66c-853e-4fec-9910-0341ed2d1577 · outbound

This paper cites MXNet: A Flexible and Efficient Machine Learning Library for Heterogeneous Distributed Systems.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators MXNet: A Flexible and Efficient Machine Learning Library for Heterogeneous Distributed Systems

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.681665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.681665Z digest=sha256:973cb55a63353e609c82f802b34ffcbed09fbf3547a6eab961232d318f87902b

Observation 265278f4-a1ab-461e-9c46-63c206dce10c · outbound

This paper cites Rubik: A hierarchical architecture for efficient graph neural network training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Rubik: A hierarchical architecture for efficient graph neural network training

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.687941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.687941Z digest=sha256:aaba129031e7472f91451919b30ad6d63147810dc43882baf4f4c3982fb8fc12

Observation f1dfa1f2-f91a-4a3a-a18f-fe22a470139d · outbound

This paper cites The bandwidth problem for graphs and matrices—a survey.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators The bandwidth problem for graphs and matrices—a survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.695359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.695359Z digest=sha256:a29bec3d9ed09a67fdfb9c2a4cb79eb5f80a04e088b51e477ba78b3aa4aba49d

Observation 900c8cf9-a8eb-46bf-8836-e18c85657a60 · outbound

This paper cites Reducing the bandwidth of sparse symmetric matrices.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Reducing the bandwidth of sparse symmetric matrices

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.708177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.708177Z digest=sha256:17951bea9cad9fd4d995efe94b7fb748aab9f1932bc6a7a6d27ba59b3c46e228

Observation 3c5b3bc4-135b-4bf8-b227-e9ef3509220c · outbound

This paper cites Hardware acceleration of sparse and irregular tensor computations of ml models: A survey and insights.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Hardware acceleration of sparse and irregular tensor computations of ml models: A survey and insights

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.715413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.715413Z digest=sha256:95a27cc09e81ff33883643d44025026e81eec7a27f1ebe8bf79fda1a27935914

Observation 83f84d04-ce23-495c-a5e2-f37b943c8769 · outbound

This paper cites Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.720956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.720956Z digest=sha256:4682240dbedaaa4a5d4fc17a861dd8515d72196256c50abf58605794c91b2a2d

Observation f0113c2e-fa3e-447b-9752-ce615c03ca3a · outbound

This paper cites GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.727043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.727043Z digest=sha256:7229e147061a2f19118cf6cb23dd7b316e9e754ecd56fe383b62a15c4496a170

Observation 0651de6c-f8bc-472b-a9f7-ca3cd0786885 · outbound

This paper cites Tlpgnn: A lightweight two- level parallelism paradigm for graph neural network computation on gpu.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Tlpgnn: A lightweight two- level parallelism paradigm for graph neural network computation on gpu

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.733191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.733191Z digest=sha256:f8b3af36b26badfe17e3e6fc8369173124fcfe51d4daf8dc99136f5b1e2d75be

Observation bef0258f-a19c-4a8e-a99f-3d345a12397c · outbound

This paper cites P3: Distributed deep graph learning at scale.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators P3: Distributed deep graph learning at scale

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.743084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.743084Z digest=sha256:537cfec1930d6363e0da3b922c528a1f9eeb41813de12e37faec5d87ee65a65b

Observation ddb8f1d2-30fb-4930-9934-726a654b23c9 · outbound

This paper cites Understanding the Design-Space of Sparse/Dense Multiphase GNN dataflows on Spatial Accelerators.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Understanding the Design-Space of Sparse/Dense Multiphase GNN dataflows on Spatial Accelerators

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:21:00.914804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.752395Z digest=sha256:c67885c471e5103cb8102e0766e31878eab356d589023b4ba02d96df933fb301

Observation d30f8ec6-fb6f-4104-a5b9-a4e777ee08fb · outbound

This paper cites Awb-gcn: A graph convolutional network accelerator with runtime workload rebalancing.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Awb-gcn: A graph convolutional network accelerator with runtime workload rebalancing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.760701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.760701Z digest=sha256:091899c57b2018abd365978acc3b28a85f3cfbb53e84336427ffac9cd0c7060a

Observation 1dd4b653-4062-4595-b968-f8daa27bc569 · outbound

This paper cites I-gcn: A graph convolutional network accelerator with runtime locality enhancement through islandization.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators I-gcn: A graph convolutional network accelerator with runtime locality enhancement through islandization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.766827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.766827Z digest=sha256:f78f4ca47f19b64e546b909e5a7d9449b97a0cf50e7ab81d25cb17fcdb8bdaa8

Observation 1c2058d1-43f6-4f1b-8054-b49b731acfe8 · outbound

This paper cites Data-efficient graph grammar learning for molecu- lar generation.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Data-efficient graph grammar learning for molecu- lar generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.774832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.774832Z digest=sha256:48708d0c004bbc97f665acf2de562026667c3afdf2049aa354140277f8a34ad8

Observation d8d3438f-15be-42e0-a6a2-0bfc03f92b6c · outbound

This paper cites Inductive represen- tation learning on large graphs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Inductive represen- tation learning on large graphs

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:03.123452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.781848Z digest=sha256:cf52044b65baf77ca3b118f82748a018c9b9a42e925ab559a0327b73ebbf2b99

Observation 0d08922e-937e-4b4f-8655-414507764227 · outbound

This paper cites PipeDream: Fast and Efficient Pipeline Parallel DNN Training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators PipeDream: Fast and Efficient Pipeline Parallel DNN Training

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.789532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.789532Z digest=sha256:c1c5d5e5929cd3c7e990d3f3949a2a4f9772e6b0da21c1db771e16fe618cb765

Observation 08c0e4d8-8ded-42a4-9b44-487f86315581 · outbound

This paper cites Open Graph Benchmark: Datasets for Machine Learning on Graphs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Open Graph Benchmark: Datasets for Machine Learning on Graphs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.810421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.810421Z digest=sha256:fa131ff6cb37b1a4627b60cba8bda497df7fb56919cbfec8cd7bd04d09cea1ef

Observation 2934b5ee-8cfa-4ddd-b4e2-150e63895935 · outbound

This paper cites Recurrent graph convolutional network-based multi- task transient stability assessment framework in power system.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Recurrent graph convolutional network-based multi- task transient stability assessment framework in power system

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:03.089163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.818299Z digest=sha256:50e7f071d6a27f686beb4fa0f3a00cb8c5f09c133daa7e4c9484067d689f48f2

Observation a1a61f0a-9771-4f8d-a903-250a3949abdb · outbound

This paper cites Wisegraph: Optimizing gnn with joint workload partition of graph and operations.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Wisegraph: Optimizing gnn with joint workload partition of graph and operations

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:03.063763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.824637Z digest=sha256:ab7d6bbf43bf45bd3a831687cceb8be0422654673b3b6cae64f073eddac44bee

Observation 6638afb0-1936-4479-aa12-9858d99f90fc · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Gpipe: Efficient training of giant neural networks using pipeline parallelism

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:03.032277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.838644Z digest=sha256:c14c728069fa039676cbf263d9b1f867aacc8fc5decbdeaf5264e6513bc8e2be

Observation f14914ea-757b-45b0-943f-5a814e9559a1 · outbound

This paper cites GraphPipe: Improving Performance and Scalability of DNN Training with Graph Pipeline Parallelism.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GraphPipe: Improving Performance and Scalability of DNN Training with Graph Pipeline Parallelism

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.846946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.846946Z digest=sha256:4462ce7beaca45dc253945a8ec3752e4ccc42d0f650f4f71608cbbeea4c42670

Observation f248b074-31e6-4104-8cdd-03d2de625884 · outbound

This paper cites A survey on knowledge graphs: Representation, acquisition, and applications.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators A survey on knowledge graphs: Representation, acquisition, and applications

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:03.001880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.854077Z digest=sha256:8027934c8cf420705b739b8c303e0774bd913bf819551c4099f545c00c0e003b

Observation 142d96fe-707b-4af4-9af6-a87d94a4b1ab · outbound

This paper cites Improving the accuracy, scalability, and performance of graph neural networks with roc.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Improving the accuracy, scalability, and performance of graph neural networks with roc

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.968211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.863186Z digest=sha256:3970d8f2d8ccd98b3f498d07a3c0bc6bb11a585bc11f4e5bd80bda5fb76c56c4

Observation 5cb57111-06b2-458e-b51d-5e2e411d6aed · outbound

This paper cites A survey of frequent subgraph mining algorithms.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators A survey of frequent subgraph mining algorithms

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.919586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.886748Z digest=sha256:1cfd7eb59f04c5a3058c7e8838f0760b8b96965b6e1bf49bf7a7e3bc158bb971

Observation d437af7a-a70e-4d9e-b372-789939763c32 · outbound

This paper cites A unified architecture for accelerating distributed{DNN} 12 training in heterogeneous{GPU/CPU} clusters.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators A unified architecture for accelerating distributed{DNN} 12 training in heterogeneous{GPU/CPU} clusters

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.882539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.896202Z digest=sha256:bac7bc2ec65db34a57e1bbf841d67c2fd3a94d1b43fdad5578d0147637cd6af4

Observation b146ab98-1339-41ed-90f4-9b76ef32289f · outbound

This paper cites In-datacenter performance analysis of a tensor pro- cessing unit.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators In-datacenter performance analysis of a tensor pro- cessing unit

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.832052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.902286Z digest=sha256:e677539c226bcfa99f59ee0597467d6d2dd31c39a1a05e95fb43facf1a38ed4c

Observation 28f1548c-1997-4beb-a394-7b0b5127c010 · outbound

This paper cites A fast and high quality multilevel scheme for partitioning irregular graphs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators A fast and high quality multilevel scheme for partitioning irregular graphs

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.780051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.914019Z digest=sha256:9f13a0d9565971e32f5dd5f4577e80ba8562856b958c69173fba7472ab303190

Observation 6c470d37-b9c5-47f3-a7b0-fe75d054b951 · outbound

This paper cites GRIP: A Graph Neural Network Accelerator Architecture.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GRIP: A Graph Neural Network Accelerator Architecture

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.923650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.923650Z digest=sha256:5f45c3294ac9b1ed067e8182681ab7da65a5441b6f939bb8a5e208173d311a7f

Observation 91348ac9-53df-4283-873e-70b652883e0b · outbound

This paper cites Semi-Supervised Classification with Graph Convolutional Networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Semi-Supervised Classification with Graph Convolutional Networks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.935775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.935775Z digest=sha256:97791b28662f3c8f55c6a1e3c47fc18cbb9922a70e24d3102eded093c37e2018

Observation a5d7c67d-96a7-4e0e-b2ef-24c445cd0cf4 · outbound

This paper cites What is twitter, a social network or a news media? InProceedings of the 19th international conference on World wide web , pages 591–600, 2010.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators What is twitter, a social network or a news media? InProceedings of the 19th international conference on World wide web , pages 591–600, 2010

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.754966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.941452Z digest=sha256:a783eddfecbf9bbf74d5dcafff735afb776a72a953b60367176092e67670d48f

Observation 3916394e-82ee-480e-bedd-607716e4c432 · outbound

This paper cites Maeri: En- abling flexible dataflow mapping over dnn accelerators via reconfig- urable interconnects.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Maeri: En- abling flexible dataflow mapping over dnn accelerators via reconfig- urable interconnects

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.734867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.950658Z digest=sha256:63ab92c0d2002d3d337d863ea4b740a138c08db65e62b4dfc5dbda0d432d7cfd

Observation 85fb448c-8df8-422d-988e-ffa53899f03c · outbound

This paper cites GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.960665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.960665Z digest=sha256:2928f2b5832c67b4201f81a9a437b90036ce2863cd99033448b91c584953a8f4

Observation 98849045-ddec-4b1f-894f-826f52e88d75 · outbound

This paper cites Gcnax: A flexible and energy-efficient accelerator for graph convolutional neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Gcnax: A flexible and energy-efficient accelerator for graph convolutional neural networks

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.708857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:58.972154Z digest=sha256:2961313053790c4c49e10c8fd6fcb2cae889694dd351d0b9b5067bd50f3b3a72

Observation 20c0a3b0-4df1-44ea-969e-fd7f04788898 · outbound

This paper cites PyTorch Distributed: Experiences on Accelerating Data Parallel Training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators PyTorch Distributed: Experiences on Accelerating Data Parallel Training

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.982658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.982658Z digest=sha256:6317bd62850c258a9eb188a8946bc821d8c47f529f22b693a8aa4bd7fe8cd223

Observation 1f7421e2-3a0b-4046-a37d-35e0bb6c16f3 · outbound

This paper cites Terapipe: Token-level pipeline parallelism for training large-scale language models.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Terapipe: Token-level pipeline parallelism for training large-scale language models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:58.994822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:58.994822Z digest=sha256:87037ad2bc38cd978597aca2c856dcd484963925c5f1a6fadccf016bec7bd738

Observation 93ebb486-cd12-4c14-9b03-f6eab1dcceb8 · outbound

This paper cites Engn: A high-throughput and energy-efficient accelerator for large graph neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Engn: A high-throughput and energy-efficient accelerator for large graph neural networks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.632501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.016949Z digest=sha256:f573e2a68758696d2094fcefcf7b97e9fd9a6505941321678533257e6a862764

Observation 09e3ddb5-ab50-4d88-8004-35428db90a82 · outbound

This paper cites Nvidia tesla: A unified graphics and computing architecture.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Nvidia tesla: A unified graphics and computing architecture

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.600956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.030513Z digest=sha256:8a86e59f4d86e92492f681640a4e6faf96320a27d1cdc1b13da5e77b8b486d6c

Observation f98bde13-0dfc-4c84-a9cf-127e05790520 · outbound

This paper cites Flexflow: A flexible dataflow accelerator architecture for convolutional neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Flexflow: A flexible dataflow accelerator architecture for convolutional neural networks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.566211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.036783Z digest=sha256:06152d8d23b2127aa526600a1ada02fef0e0fd5a64e0fbd3852e3658901eea19

Observation 490a054e-7ca0-47b8-a6ec-a7c963318011 · outbound

This paper cites NeuGraph: Parallel deep neural network compu- tation on large graphs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators NeuGraph: Parallel deep neural network compu- tation on large graphs

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.539604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.042719Z digest=sha256:75b90f7de91cf29892d780b7ee410ad7214d9a3b91f707a0002616f6f7048fee

Observation 3978c9b6-e306-4a77-a3a5-a133550d9d88 · outbound

This paper cites All-to-all personalized communication on multi- stage interconnection networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators All-to-all personalized communication on multi- stage interconnection networks

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.494811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.052182Z digest=sha256:974f123fc0d0ca72d04bcd5b9ef8f3f033d60ccfd1fe5a6e4b9980691b86480d

Observation 6d4041c5-da6a-4b3c-8e98-05c1ebcd25d3 · outbound

This paper cites Distgnn: Scalable distributed training for large-scale graph neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Distgnn: Scalable distributed training for large-scale graph neural networks

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.447000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.060848Z digest=sha256:11990af3ea4477963dc82429ada1a3ecc1ce3bb43c247e6e0677f3604d1b6870

Observation 319f8905-194a-4c96-ad85-c462b4296de9 · outbound

This paper cites Device placement optimization with reinforce- ment learning.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Device placement optimization with reinforce- ment learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.418462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.071093Z digest=sha256:869cf73b45980fb463aff7a73398ad13e02f21e76fdc66480a30afbfd7883bc9

Observation 7e67f924-4cc7-4353-9169-75fe0bf7825b · outbound

This paper cites Pipedream: generalized pipeline parallelism for dnn train- ing.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Pipedream: generalized pipeline parallelism for dnn train- ing

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.364659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.079779Z digest=sha256:8033e76cec5666701d6cc2ee60e127834a6e8e362c42e57f7f625a7f861239a3

Observation 380d05f5-8580-4e87-a785-73e8589f9d2d · outbound

This paper cites Sancus: staleness-aware communication-avoiding full- graph decentralized training in large-scale graph neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Sancus: staleness-aware communication-avoiding full- graph decentralized training in large-scale graph neural networks

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.329105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.089244Z digest=sha256:95e2ac2523ac59fafd2ad77a81d74e011db5271b626427481e41316dc054c47f

Observation 65ed8d03-a91c-4252-b1bc-862013eeeb46 · outbound

This paper cites Fusedmm: A unified sddmm-spmm kernel for graph embedding and graph neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Fusedmm: A unified sddmm-spmm kernel for graph embedding and graph neural networks

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.274773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.095388Z digest=sha256:cb1687a25d2860792206b8d10a972189b04c6c3eae03eb3b12abcffbe04ad46b

Observation 65049aef-5c7d-47a0-8808-3a526f13763e · outbound

This paper cites DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.102053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.102053Z digest=sha256:2a90d9dc9085dc39afb955a237eaef4143c587b3b1f33429d120c421697b3b9b

Observation d8d35684-346d-4b49-8179-8094f6c961fa · outbound

This paper cites Zero: Memory optimizations toward training trillion parameter mod- els.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Zero: Memory optimizations toward training trillion parameter mod- els

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.108981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.108981Z digest=sha256:24fa6caf46d65d54e307f6e221ad859a48980eaa71e8d69e9b02ac7ec1490571

Observation 0a54229d-d23b-45e2-aa60-ea1ce748168c · outbound

This paper cites Learn Locally, Correct Globally: A Distributed Algorithm for Training Graph Neural Networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Learn Locally, Correct Globally: A Distributed Algorithm for Training Graph Neural Networks

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:21:00.479807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.117725Z digest=sha256:8eb3fb38c16cedf8eea06b1a4d9c081135d4c8f06b898f9afdbcc97c6b3c7a6a

Observation 8ee42cb4-310e-4f75-a8ce-ece2d785048e · outbound

This paper cites Deepspeed: System optimizations enable training deep learning mod- els with over 100 billion parameters.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Deepspeed: System optimizations enable training deep learning mod- els with over 100 billion parameters

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.207314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.126880Z digest=sha256:4804d062025b8deae9aed2f95955240f540f2e42ef48b67d5ca700cff029d12b

Observation c1a24337-6191-4dc6-9f27-cccc75667223 · outbound

This paper cites Algorithms for scheduling independent tasks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Algorithms for scheduling independent tasks

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.150058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.137669Z digest=sha256:4724dea2b232bb924932e665dfe3f8a40afe50ca0425e38909ae437e98b5392e

Observation 9189f8bc-bdda-46c3-82ba-50abd2636540 · outbound

This paper cites Horovod: fast and easy distributed deep learning in TensorFlow.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Horovod: fast and easy distributed deep learning in TensorFlow

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.150042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.150042Z digest=sha256:f2506ae053c64661b5075b7840896af4a4d40bef0a16fc59e6245d58fd0024b4

Observation 8a880ebf-9a11-49ce-8d5a-8122d0ec683e · outbound

This paper cites Mesh-tensorflow: Deep learning for supercomputers.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Mesh-tensorflow: Deep learning for supercomputers

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.121015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.161184Z digest=sha256:909e2278788051dd089fbf15c8dabcd3a8e05a25aa099f728228d13b9a6d80c2

Observation 40d352c3-29c8-49c0-aaf0-24533ccba5ce · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.171908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.171908Z digest=sha256:b9b55c76c2e77c771e7df8bfaa29629b22cadcf81f260f77072e319706e6cda1

Observation f4e393ec-e838-4319-9bfb-d610633b99d3 · outbound

This paper cites Synopsys design compiler.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Synopsys design compiler

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.091513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.181983Z digest=sha256:a1c52f98f5c42c16b7c5c665840d6fdb32c0b2fcaf2b4052aa4b1bc59979451a

Observation 9cd0c7a5-d439-4097-8b41-4bdc58782e50 · outbound

This paper cites Dorylus: affordable, scalable, and accurate gnn training with distributed cpu servers and serverless threads.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Dorylus: affordable, scalable, and accurate gnn training with distributed cpu servers and serverless threads

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.056078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.191790Z digest=sha256:4057150a7fa5265eaa322497ce9658b965c03c75bcdcdaf8d8542f542b4e1396

Observation a4d33c43-7177-494e-9fe3-f78a192c8417 · outbound

This paper cites Reducing Communication in Graph Neural Network Training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Reducing Communication in Graph Neural Network Training

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:21:00.319132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.199116Z digest=sha256:30f92569fb14ba837a7e26f36907f9e459ca67db5d940dbd3765a0c89376aa52

Observation 7e7b17ad-ad80-418d-8ba4-b88d7e979228 · outbound

This paper cites Graph Attention Networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Graph Attention Networks

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.211695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.211695Z digest=sha256:fad2e1da9fca158bb93dc273ad85a8b9924ccf268be6f4d4645c6f4db32961dc

Observation a770c2b6-0991-470a-9da3-cf781b35cc45 · outbound

This paper cites Adaptive message quan- tization and parallelization for distributed full-graph gnn training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Adaptive message quan- tization and parallelization for distributed full-graph gnn training

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:02.016523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.225730Z digest=sha256:8209283672b2fad7bb5ec12e9e65c3ce6b92914c4109e37042cca0b67c4a1bdb

Observation f98be4b5-ff80-4c04-8aea-3ea9058a0a02 · outbound

This paper cites BNS- GCN: Efficient full-graph training of graph convolutional networks with partition-parallelism and random boundary node sampling.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators BNS- GCN: Efficient full-graph training of graph convolutional networks with partition-parallelism and random boundary node sampling

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.978340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.236611Z digest=sha256:412d010b5c11d2950a38e91dd19361b0bc3927439f083c1f4b16e379a4800e30

Observation 4587520d-d135-4149-b62d-8d95c9b56882 · outbound

This paper cites Wolfe, Anastasios Kyrillidis, Nam Sung Kim, and Yingyan Lin.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Wolfe, Anastasios Kyrillidis, Nam Sung Kim, and Yingyan Lin

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.933270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.251968Z digest=sha256:ecf4cb0e8f80c7953f47efee06d6a6acdc0ef54054752c6c4c67b29c4cef82b2

Observation 1267dc0d-981c-4a75-82f6-b9fe313081cb · outbound

This paper cites Towards Cognitive AI Systems: a Survey and Prospective on Neuro-Symbolic AI.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Towards Cognitive AI Systems: a Survey and Prospective on Neuro-Symbolic AI

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.261699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.261699Z digest=sha256:9caafa094a3d5a956fe094279ac86e31e5907f9a2a819286db5b8af26461b4d7

Observation 835b8b77-41cd-4add-90f8-b678680feab1 · outbound

This paper cites Flexgraph: a flexible and efficient distributed framework for gnn training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Flexgraph: a flexible and efficient distributed framework for gnn training

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.890633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.271511Z digest=sha256:94152f90d250e4ed1ac1963fb362293633f62fbcfca5a1183b8ff14915ff31e1

Observation ca78c1c0-447d-4f2a-820c-5460288bee04 · outbound

This paper cites Supporting very large models using automatic dataflow graph partitioning.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Supporting very large models using automatic dataflow graph partitioning

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.856788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.278477Z digest=sha256:f9c76f432ac062552651bdb085e0af10e8126ed079e669a23ed945c9faebb6fe

Observation 5d290b2d-c270-4682-a6fe-5bba44dd1631 · outbound

This paper cites Deep Graph Library: A Graph-Centric, Highly-Performant Package for Graph Neural Networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Deep Graph Library: A Graph-Centric, Highly-Performant Package for Graph Neural Networks

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.290357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.290357Z digest=sha256:657ebbbe2ed0f7b117da12d7691b8e3885a84fb0ed483a91a656ff10d1326416

Observation 9362a877-359f-4af1-817d-259a0283cdaa · outbound

This paper cites Neutronstar: distributed gnn training with hybrid dependency management.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Neutronstar: distributed gnn training with hybrid dependency management

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.824164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.298190Z digest=sha256:cbaa3696d340fe8d40f462820717b6a68c38f0462a2f9300375c374e4b7b200d

Observation 6da99ef0-b5b7-497b-981c-529d4695541d · outbound

This paper cites GNNAdvisor: An Adaptive and Efficient Runtime System for GNN Acceleration on GPUs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GNNAdvisor: An Adaptive and Efficient Runtime System for GNN Acceleration on GPUs

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:21:00.130736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.304951Z digest=sha256:ed995ee9453e3807d14090e621c8bd1cdfb4ecfb5d07531eb7d50524a6f24f11

Observation 04ef5575-beb6-448a-97d3-5e090e4976d7 · outbound

This paper cites how graph neural networks go beyond weisfeiler-lehman?.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators how graph neural networks go beyond weisfeiler-lehman?

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.748058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.315146Z digest=sha256:cd4aba359bba27f58b9e046ef8667fe667063519bfbe88e7ff11dcdfc5687e11

Observation 58b0993f-a610-4bad-bfd2-354f248a1864 · outbound

This paper cites A comprehensive survey on graph neural networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators A comprehensive survey on graph neural networks

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.719851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.322158Z digest=sha256:9be6df99b22ed818102484c78e397d24700848408c1c220a091759f64d273f89

Observation 211a737e-8cfb-4d27-babd-abbddde3e4ca · outbound

This paper cites Graph learning: A survey.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Graph learning: A survey

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.683972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.328289Z digest=sha256:d35dca337b2c0e404a595c29c4c52fc8ec9da3866c40dc31cf4017dffd30dcd0

Observation 5ae5805c-cbec-452a-9a1d-dfdc7f84c6b4 · outbound

This paper cites How Powerful are Graph Neural Networks?.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators How Powerful are Graph Neural Networks?

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.341765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.341765Z digest=sha256:474b60646cc8d6ec7ee4a1acdae103c61b72b45011bd345dffa35d73d8c1473e

Observation 6f0b339d-f9ca-4afb-8fa3-9d4f07837ae4 · outbound

This paper cites GSPMD: General and Scalable Parallelization for ML Computation Graphs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GSPMD: General and Scalable Parallelization for ML Computation Graphs

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.348880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.348880Z digest=sha256:97c6b4289ac7241dbd1ab625a802611c6223b294bfa14d62e49a4c1976e7905c

Observation 1ae6a74e-3cab-4faa-bee7-24148c9f0f42 · outbound

This paper cites Hygcn: A gcn accelerator with hybrid architecture.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Hygcn: A gcn accelerator with hybrid architecture

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.640331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.354654Z digest=sha256:081cb726f28654232c28c1508e2d1ade5dbb56a943d0e141ec378514c5d34be2

Observation 1676b62e-3f12-4397-8fe0-94122f328542 · outbound

This paper cites Defining and evaluating network com- munities based on ground-truth.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Defining and evaluating network com- munities based on ground-truth

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.609457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.365045Z digest=sha256:16975b6b62e7d286a27fdad511dea29f527ce3d8c0f0da7616a26b1e6db554d3

Observation b7be38e9-e0d2-4528-a1cd-512d6665a03c · outbound

This paper cites Optimal all-to-all personalized exchange in self-routable multistage networks.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Optimal all-to-all personalized exchange in self-routable multistage networks

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.585646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.373416Z digest=sha256:34d466c02334de9e5bee8b8d56c10b94194397c8ac9982e33ffb047d73199ac7

Observation 36f83b9b-087d-4ade-a8c3-8a79ad3ff20a · outbound

This paper cites Graph convolutional neural networks for web-scale recommender systems.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Graph convolutional neural networks for web-scale recommender systems

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.563709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.380014Z digest=sha256:08ae3e3c81513ceac3e2a10645f2bee571ab87754967a95147b65ca186f028f6

Observation 5761de6c-1f3a-4b74-9879-388143157dc2 · outbound

This paper cites GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-Design.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-Design

Reference 82

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:20:59.949902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.388936Z digest=sha256:d91f313b5c79b59946892d6cbabca7af625092a2264fa26a2051bb6c8ff5876c

Observation 043d9979-41bb-4e62-8ef3-754470da660e · outbound

This paper cites Graphact: Accelerating gcn training on cpu-fpga heterogeneous platforms.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Graphact: Accelerating gcn training on cpu-fpga heterogeneous platforms

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.517837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.402570Z digest=sha256:6a7535a5324fd9f44be61ab128bbcedc0bc1790dff9e9fc005fa4ba3f73351d2

Observation f71bf75c-46be-4796-9987-846e59d84a8e · outbound

This paper cites Hardware accel- eration of large scale gcn inference.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Hardware accel- eration of large scale gcn inference

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.480358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.412065Z digest=sha256:25b2052b2671a856334e202a1b4ea0a7094bde3053b87c49b9535d1322700055

Observation d8bdd40a-449d-4617-b1fa-407f42a8deb7 · outbound

This paper cites Autosync: Learning to synchronize for data-parallel distributed deep learning.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Autosync: Learning to synchronize for data-parallel distributed deep learning

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.414892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.421597Z digest=sha256:0f160a32ae924e85ceac25d15ca2a56785f2d41593c5234a919c507069bdcab4

Observation 226251fb-c1b5-4d19-af6c-c9f5c1cb842f · outbound

This paper cites Understanding gnn computational graph: A coordinated computation, io, and memory perspective.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Understanding gnn computational graph: A coordinated computation, io, and memory perspective

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.380663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.428702Z digest=sha256:70d4836c78071aa47b1564b8afdd8269b99ccb09a485798ba7f24c216ab80579

Observation eef5f213-4459-43c6-8fcf-05e22b254284 · outbound

This paper cites Sylvie: 3d-adaptive and universal system for large-scale graph neural network training.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Sylvie: 3d-adaptive and universal system for large-scale graph neural network training

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.337455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.444800Z digest=sha256:ef7000845551ddb92734cbf204bd97a7ec155f584f0a28cdaa8aff7b7096ce40

Observation 074e3a38-5b1f-415a-87f7-6211cd42e502 · outbound

This paper cites A survey on graph neural network acceleration: Algorithms, systems, and customized hardware.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators A survey on graph neural network acceleration: Algorithms, systems, and customized hardware

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.466573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.466573Z digest=sha256:0be8289782000317f16813a653c46856fb653f8d44a00219984ce72a0de799a1

Observation 0017fa5e-48dd-41b9-893c-814409a72f4b · outbound

This paper cites G-cos: Gnn-accelerator co-search towards both better accuracy and efficiency.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators G-cos: Gnn-accelerator co-search towards both better accuracy and efficiency

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.295822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.474260Z digest=sha256:215ba4884ae2b4e39f7f77a1ee92276c398c38296b8bf8b8459f31d1ebbcdf98

Observation 354e0a3f-1189-409a-8e77-fc09bf6937e8 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.490966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.490966Z digest=sha256:76b356f5f875ca7a20b7e774135fd73696437435e63e53af1bedf622b9223c20

Observation cfbc0ac6-f3d0-4db7-bab7-145e04795291 · outbound

This paper cites Distdgl: dis- tributed graph neural network training for billion-scale graphs.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Distdgl: dis- tributed graph neural network training for billion-scale graphs

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:21:01.278205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T22:20:59.498436Z digest=sha256:2252d09b0e5a43e84f42b0d818329b8ac175bf0a7bee97a18c1417698594b11c

Observation 29e2cab1-2fc9-4a87-ad2f-824e7f23e79e · outbound

This paper cites Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.515950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.515950Z digest=sha256:326d7362dddd4da6402e6342fe2d2cc551a0b8c3ceb2a8567c5a7cb1de3f730d

Observation 4e0e9db2-931a-4e59-b5a0-117a1f9dcac7 · outbound

This paper cites AliGraph: A Comprehensive Graph Neural Network Platform.

MixGCN: Scalable GCN Training by Mixture of Parallelism and Mixture of Accelerators AliGraph: A Comprehensive Graph Neural Network Platform

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:59.524307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:59.524307Z digest=sha256:6d55746fa2dded86a11e4e8fef99354e10c33dc6b067eaf7ecf5dc8c1fba052f

Pith citing papers

No inbound Pith citation observations are available.