Pith. sign in

Paper Citation Record · LEDGER

Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2406.16747.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.16747 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:13:38.199036Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:38:24.979853Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f266c08e-13e5-49ad-92b3-cc4fa45f6c38 · inbound

E2LLM: Encoder Elongated Large Language Models for Long-Context Understanding and Reasoning cites this paper.

E2LLM: Encoder Elongated Large Language Models for Long-Context Understanding and Reasoning Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:38:24.982446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T20:36:55.159302Z digest=sha256:110018c5fd28a7a78dbe5904af5d941452ee5d2e0e44fe2b66674213b61af802

Observation cec66b89-1446-4307-bca4-ce0a7f05dca7 · inbound

On Efficient Variants of Segment Anything Model: A Survey cites this paper.

On Efficient Variants of Segment Anything Model: A Survey Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 198

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T19:43:23.410696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T19:42:24.122342Z digest=sha256:35572bab0eaab2dbe7c3ce912939901016472eac9be28fb236085ff20985152a

Observation e4eb9add-3289-44cf-a011-86d0adc1ef83 · inbound

Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents cites this paper.

Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T14:13:38.199036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:13:38.199036Z digest=sha256:1771b41115c6633922bf15671ebe81b39ca0fa0880394366c2eac1f94b34f907

Observation f85ba02c-f795-45d9-ba9e-cbe62a4ad034 · inbound

ESPFormer: Doubly-Stochastic Attention with Expected Sliced Transport Plans cites this paper.

ESPFormer: Doubly-Stochastic Attention with Expected Sliced Transport Plans Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T11:23:58.555350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:23:58.555350Z digest=sha256:c438de8c0f33a9db95a8c9d388a146cd8b3973bcc480fc187ab0129bd2da0dc9

Observation d07a2776-cecf-4a79-b8e7-2e937c75b96f · inbound

A Survey of Scaling in Large Language Model Reasoning cites this paper.

A Survey of Scaling in Large Language Model Reasoning Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.285179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:901de2ba71e9151a010de39b2d78a059fef82390973a666d90508f8c5d3850ed

Observation b7b9b8d8-57db-4204-8c37-c22797ae4e76 · inbound

Lag-Relative Sparse Attention In Long Context Training cites this paper.

Lag-Relative Sparse Attention In Long Context Training Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:47.490600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:47.490600Z digest=sha256:a6647fa76b06193f189cfbf41af2231f6ca8c01a721af2513a7bf8bd2788f582

Observation 8255c281-94fe-4be6-ba08-cdcb5da006b4 · inbound

LIFELONG SOTOPIA: Evaluating Social Intelligence of Language Agents Over Lifelong Social Interactions cites this paper.

LIFELONG SOTOPIA: Evaluating Social Intelligence of Language Agents Over Lifelong Social Interactions Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:49:27.008971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:49:27.008971Z digest=sha256:6cd79109283065f24cda34cc245a4941f00ba61b8da9f9a3ef621e17bfc64944

Observation 00e02adc-9716-4445-82a2-57fffb6784c8 · inbound

OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference cites this paper.

OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:08:07.633297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:08:07.633297Z digest=sha256:58c6a36a9449765e700f91fc2f6480de07d41a7a0e71b6b51d79ef492c153773

Observation 146c5089-560e-422b-91e7-546d234cea56 · inbound

LOOM-Scope: a comprehensive and efficient LOng-cOntext Model evaluation framework cites this paper.

LOOM-Scope: a comprehensive and efficient LOng-cOntext Model evaluation framework Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T19:45:12.502627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:45:12.502627Z digest=sha256:e521ff7203c1a86f9cae5739b47c0e642341dd34cdaf961f9a7118b1be1b4470

Observation cd89a3ed-35f5-434a-b3af-665cc7116768 · inbound

Crisp Attention: Regularizing Transformers via Structured Sparsity cites this paper.

Crisp Attention: Regularizing Transformers via Structured Sparsity Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.855081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:03:07.855081Z digest=sha256:57ae8a63e5af683f73af2197e31f62898318fbd5dacd1c4b446fc0a56a879fb3

Observation 32dfe37a-6188-4150-9353-bf11be0219ca · inbound

SCOUT: Toward Sub-Quadratic Attention via Segment Compression for Optimized Utility in Transformers cites this paper.

SCOUT: Toward Sub-Quadratic Attention via Segment Compression for Optimized Utility in Transformers Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T13:07:40.842500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:07:40.842500Z digest=sha256:6b69b099c388aea5e219c02aa6305611b1bd766376e5df906984c95baedf3477

Observation 98028e8f-c1cc-4cfe-9d56-263da827de04 · inbound

MatchAttention: Embedding Explicit Matching Constraints into Attention for Efficient Stereo Matching cites this paper.

MatchAttention: Embedding Explicit Matching Constraints into Attention for Efficient Stereo Matching Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T09:40:29.873722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:40:29.873722Z digest=sha256:f9c36582af54d55c40f21d65f111e1128147112c113f41d74e409a931a73553a

Observation 9cf7cc62-2532-4f16-a0fd-cf5ae6433201 · inbound

Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence cites this paper.

Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.500950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T22:19:53.128437Z digest=sha256:c8719ddbcdf281e3a5ff2a763bc33a47a478f0e61be26b8d189698e3166e918b

Observation 3e0bbf0d-0e02-4d44-8c24-3c4d6e90474b · inbound

TempoNet: Slack-Quantized Transformer-Guided Reinforcement Scheduler for Adaptive Deadline-Centric Real-Time Dispatchs cites this paper.

TempoNet: Slack-Quantized Transformer-Guided Reinforcement Scheduler for Adaptive Deadline-Centric Real-Time Dispatchs Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:11:37.599868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T21:11:07.026338Z digest=sha256:1d16fcbae05df3e9c44cc0384d0314e86ff0f85598982d4867111781bda038d8

Observation 87bb23a7-ad60-48fc-ba7c-a1758dba391d · inbound

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices cites this paper.

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:43.881735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T20:00:27.987481Z digest=sha256:7a32b128f6d81780c364d395ad4253635d27073c22c8581ed03591082271ed95

Observation 8ad763a5-25d8-4db5-80d3-4fc1c6c07f85 · inbound

LISA: Linear-Indexed Sparse Attention for Efficient Long-Context Reasoning cites this paper.

LISA: Linear-Indexed Sparse Attention for Efficient Long-Context Reasoning Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T12:48:53.646647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:48:53.646647Z digest=sha256:8f3d89fa57d7419690cc6caedbb83cc55451d8a689fbd2f8814622b56d9a1c58

Observation dc488476-2c52-40c6-86d2-158451a75351 · inbound

Parameter-free Adaptive Sparse Attention via Compression-Based Content Selection cites this paper.

Parameter-free Adaptive Sparse Attention via Compression-Based Content Selection Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T06:53:20.427847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T06:53:20.427847Z digest=sha256:6780e0ab1246750e3adb9339a8ecaa9e682fc1e783ecbf7846dcec170a9bd5e1