Pith. sign in

Paper Citation Record · LEDGER

Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2409.17422.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.17422 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:51:12.623362Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3cdd9fd9-efc2-4b49-9c86-9d50509efbad · inbound

Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation cites this paper.

Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-09T18:51:12.623362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:51:12.623362Z digest=sha256:528e77f8ace0234fae2f9b7645022cfef094ddf97248ce3a77fdb3bd6178b7a8

Observation 7f7592e2-bbfe-468e-8737-9d5c95431b5b · inbound

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration cites this paper.

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.316490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T03:59:58.634512Z digest=sha256:349b1d6976313bbf578d8477a656712dd822a159a95fa30a2f6d220e5574ed0a

Observation 0e6dc874-1a7f-4625-9aea-9f54564b78d3 · inbound

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation cites this paper.

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T11:11:17.772628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:11:17.772628Z digest=sha256:16c9aabd63a4d9867f8d7942b8a2c72297035382853ee75d2d4a44b716e9818e

Observation 0e17897b-9f3b-4af4-bd78-0b15be0fe53d · inbound

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling cites this paper.

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:05.601985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:05.601985Z digest=sha256:c01baca270d735efc0e2d83676ed0441ed63aab54331ba3675fa096b8140bb60

Observation 88163e8c-8b7b-431e-85a9-a0678b69b679 · inbound

EARN: Efficient Inference Acceleration for LLM-based Generative Recommendation by Register Tokens cites this paper.

EARN: Efficient Inference Acceleration for LLM-based Generative Recommendation by Register Tokens Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:18:50.235479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:18:50.235479Z digest=sha256:ba2dfb7fd336f91c8ec6c3049c7f583952dcb64ec1781f4b20356c66d66ea521

Observation 0de5561e-521a-48b1-9f0f-5b1ec5ea8bb5 · inbound

Unifying Learning Dynamics and Generalization in Transformers Scaling Law cites this paper.

Unifying Learning Dynamics and Generalization in Transformers Scaling Law Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T14:02:56.724087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:02:56.724087Z digest=sha256:10c8d918bdb211f74133aed4f14d1dc37d63b8fbf4ef5c4c3e257d2083262d16

Observation d7b80cdd-66bb-46ee-abcd-7fb5b0da0ecd · inbound

Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection cites this paper.

Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T05:09:31.306514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:09:31.306514Z digest=sha256:5b5a6edb5a3c1d61db8b2724ec13675e3252d29bafd45618e51257e7339bdd7a

Observation 5fab0228-081c-4279-93da-8661da3c3267 · inbound

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference cites this paper.

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:40:44.112794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T18:37:10.526359Z digest=sha256:e5a1cfa4691fcf69278e9ca4d15aaa371009be4438bb33874afb3ddaa9789213

Observation a5846ad2-7d7f-4094-a949-38706c878d54 · inbound

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints cites this paper.

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:38:34.603206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:37:57.481847Z digest=sha256:7892fdb9fa780377221a9ff7eec8f083b8992229d0b2f7a9dda171b2abe9b859

Observation 012e4a30-cb49-4d45-8aa3-7b1cf9a6d113 · inbound

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management cites this paper.

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.133394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T01:35:01.565343Z digest=sha256:b9769eeaadc811f131ea6efda50e7a00dff154842d34685c1acad976131455d9

Observation 0828e6ad-9792-486b-b635-f8eed2672a5c · inbound

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM cites this paper.

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:21.357324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T07:19:48.530272Z digest=sha256:06f971ffda1185ac4842e250881712a909ffb88c858e0d1721ffefb9f2a7b1a8