Pith. sign in

Paper Citation Record · LEDGER

Efficient Content-Based Sparse Attention with Routing Transformers

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2003.05997.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2003.05997 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T14:50:03.831572Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T19:45:36.414760Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e6bf9c7a-ee46-400a-93e2-c9a1bce4889b · inbound

Longformer: The Long-Document Transformer cites this paper.

Longformer: The Long-Document Transformer Efficient Content-Based Sparse Attention with Routing Transformers

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:29:58.812383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T13:29:58.719341Z digest=sha256:e5002b0c34a411186c8f3ec0967c31dd4cbc532a2ebbaf7264eb9de8bce66397

Observation 59d670d3-6313-40bb-a7d6-85ce5eef24f2 · inbound

Rethinking Attention with Performers cites this paper.

Rethinking Attention with Performers Efficient Content-Based Sparse Attention with Routing Transformers

Reference 147

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:16:14.467988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T09:16:14.336570Z digest=sha256:4e74a4ef7ee9ff8fd98fdc260ca0337599456a441f7f8ebc66b4e75a9ee65b99

Observation b1bf7893-dbbe-421f-8b55-f8978425fd85 · inbound

Deformable DETR: Deformable Transformers for End-to-End Object Detection cites this paper.

Deformable DETR: Deformable Transformers for End-to-End Object Detection Efficient Content-Based Sparse Attention with Routing Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:47:17.010120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T09:47:16.915936Z digest=sha256:1c6448d69a9d45a1e1e981fd777745fd974eef990d583e51f68c4c46655ee898

Observation 8747ea07-4842-4805-967b-f8a34b4a94f4 · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways Efficient Content-Based Sparse Attention with Routing Transformers

Reference 129

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:45:07.356120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:37136f87ac7ee6dee3ed446f3f5ce984cda5b4a824a48068149a882cae22f9ce

Observation 16a64e48-e4bb-40c1-aaf5-bb23935fdda5 · inbound

Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads cites this paper.

Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads Efficient Content-Based Sparse Attention with Routing Transformers

Reference 241

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T10:36:18.322637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T10:36:17.764761Z digest=sha256:2e2950c7553da8e458ab6b7eadade86605eebed69aca0e536657313915185ce3

Observation 27b91132-d50d-4583-87fa-e5e625d4ff0f · inbound

FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision cites this paper.

FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision Efficient Content-Based Sparse Attention with Routing Transformers

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:45:36.416769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T19:45:36.337956Z digest=sha256:c68a6ee57ebbf9d5cfbc7af5b87b1939451033b756d9d78715d61fd5a46021a7

Observation 6b6f5401-5fb3-4b33-9ff4-d87bb92c3783 · inbound

Remembering Distinct Items, Not Tokens: A Learnable Dirichlet-Process Cache Between State-Space Models and Attention cites this paper.

Remembering Distinct Items, Not Tokens: A Learnable Dirichlet-Process Cache Between State-Space Models and Attention Efficient Content-Based Sparse Attention with Routing Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T14:50:03.831572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T14:50:03.831572Z digest=sha256:61f3dc7662c67550207dfb9abbb44e94184a291db15d1e4b428fb3bf60bb28bb