Pith. sign in

Paper Citation Record · LEDGER

Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2404.05019.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.05019 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:52:07.458601Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T19:23:54.128784Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c3b62ff0-2423-4cb8-b710-8a6971dab586 · inbound

A Survey on Large Language Models for Code Generation cites this paper.

A Survey on Large Language Models for Code Generation Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:06.709437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T20:18:06.304134Z digest=sha256:edd089a7de6d63d674d791eca4d395b5ca07c6158843f6e3195ac5365a3b436f

Observation 59daacf8-d7e6-4005-b239-25fded76cfef · inbound

MoETuner: Optimized Mixture of Expert Serving with Balanced Expert Placement and Token Routing cites this paper.

MoETuner: Optimized Mixture of Expert Serving with Balanced Expert Placement and Token Routing Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T14:52:07.458601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:52:07.458601Z digest=sha256:2db62f17665a713aff76a89727cebcc65dc3931fb75098d111ff7a2a63a91088

Observation 685b225d-df9a-4871-b6b0-8023773023be · inbound

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging cites this paper.

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:01.836956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:01.836956Z digest=sha256:288cf3a4a543a80797a69e1b8de4756f0d9a8f002413f99991ba5d10f8efb95d

Observation 42bc76e4-b3cf-475d-bf2b-430ed40f5930 · inbound

Chameleon: Adaptive Fault Tolerance for Distributed Training via Real-time Policy Selection cites this paper.

Chameleon: Adaptive Fault Tolerance for Distributed Training via Real-time Policy Selection Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:41:50.619868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T20:39:32.123038Z digest=sha256:75d888ba7b0e2ed5d2f3288c4c20a183bda0ecfbad2a0aaa74362853f7a77c98

Observation b511cd5d-fc47-4c18-b843-9628e6c05bcf · inbound

HD-MoE: Hybrid and Dynamic Parallelism for Mixture-of-Expert LLMs with 3D Near-Memory Processing cites this paper.

HD-MoE: Hybrid and Dynamic Parallelism for Mixture-of-Expert LLMs with 3D Near-Memory Processing Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T19:16:55.265118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:16:55.265118Z digest=sha256:c41fed24aa426d94a476e02b26f3e21f065d325c5d7325cf2cb73f90a8ea2c4b

Observation 1994da47-e34b-458d-a432-33f09c4263a5 · inbound

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference cites this paper.

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T01:45:52.013040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T01:29:26.298131Z digest=sha256:e63dcfd536769e33a9acbe7eacba9c0786ea6464389cc5f1b48a3b4e53ab1e12

Observation 32feb5ce-7ebc-45a8-9254-9b3bb704d926 · inbound

MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training cites this paper.

MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training Shortcut-connected Expert Parallelism for Accelerating Mixture-of-Experts

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:23:54.130290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-29T19:15:49.229099Z digest=sha256:b293f3f00ad5046e206baf694134fdc18243688489340269779efa5d977bdbcb