Pith. sign in

Paper Citation Record · LEDGER

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference

As of 10 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2511.21702.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.21702 v2

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T22:05:05.627076Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4a3b168-a307-498c-ab29-7cdf42eea5df · outbound

This paper cites Attention is all you need,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Attention is all you need,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.591222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.591222Z digest=sha256:83fc3f47fd61639356fe8ebcf405186f3db739f53c390a4e56842ee3965fea07

Observation e40bd281-0255-4eca-9c3c-add394adf813 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.594492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.594492Z digest=sha256:7d2c4674397cabbf86b1780e8e6611e51ab149a2e748ae5677b06ebe302139c4

Observation 9da969fa-b914-4b1e-b3d8-fcd03967ae2e · outbound

This paper cites Mistral 7B.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Mistral 7B

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.597494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.597494Z digest=sha256:580db32966b02052b722632f65aa1f2d0bb7a87bbe50f0ed89d8b1021ce679ce

Observation 53c1ac62-143c-47c2-88b2-471c4cb9174c · outbound

This paper cites Efficient softmax ap- proximation for gpus,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Efficient softmax ap- proximation for gpus,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.600504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.600504Z digest=sha256:ff97024ab592e9320a388fcf27ad7aac4fe54c909ed4604e41df0b57294be52f

Observation 72ab3e34-582a-4c24-b05c-45f3ee453e17 · outbound

This paper cites Efficient estimation of word representations in vector space,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Efficient estimation of word representations in vector space,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.603392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.603392Z digest=sha256:6e7a81f5b78892cb9a42e524aa628b48897553ae93825e9b7c05f18ab9d6fbe0

Observation 357b694e-a262-45b9-8bd0-84ddad0d6793 · outbound

This paper cites Sparse convolutional neural networks,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Sparse convolutional neural networks,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.606140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.606140Z digest=sha256:64feb9443759038b7bd0cf0d003ba3600c706c3f77801cc0d860c9259447ca48

Observation f341e3fd-4c5f-44c2-8424-18c99136736d · outbound

This paper cites Reducing transformer depth on demand with structured dropout,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Reducing transformer depth on demand with structured dropout,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.609010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.609010Z digest=sha256:ac0adb5f2200e22f3696af32a1f3d697fe6d83ddb38612a0c3ef1334a4163a75

Observation 994039a7-02ed-4334-82c2-5372deded919 · outbound

This paper cites Fast inference from trans- formers via speculative decoding,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Fast inference from trans- formers via speculative decoding,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.611417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.611417Z digest=sha256:d3553cc27c249e51eb97c0f6fce4dfa12f1c611e78276df8fddeb6bb12f49e8e

Observation 18f5aabe-e6a3-4ec4-98fe-390be31f5dde · outbound

This paper cites Designing Large Foundation Models for Efficient Training and Inference: A Survey.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Designing Large Foundation Models for Efficient Training and Inference: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.613847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.613847Z digest=sha256:2fd7ef54118958e9dc7c74f1519551c8a4a20f4d9c0214c2e32baa59b4cc8cdd

Observation 81613c19-4385-4880-919b-d612802bda55 · outbound

This paper cites MKA: Memory-keyed attention for efficient long-context reasoning,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference MKA: Memory-keyed attention for efficient long-context reasoning,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.616895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.616895Z digest=sha256:0a89255f44f1c1a0a12c33d2f6b475ee330b4c4e7336197f47f7040c314b9592

Observation 9fb5c40e-6a18-41c8-b9e5-9d213d7c4ba3 · outbound

This paper cites Tinyserve: Query-aware cache selection for efficient llm serving,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Tinyserve: Query-aware cache selection for efficient llm serving,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.619389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.619389Z digest=sha256:f21c35a2e1c2cbdcb1c33b11e44ba74220e722030b1001c67bff11b54970cf45

Observation b9d4f2f3-e95f-4ce3-a933-9f1762b992c7 · outbound

This paper cites PiKV: KV Cache Management System for Mixture of Experts.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference PiKV: KV Cache Management System for Mixture of Experts

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.621861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.621861Z digest=sha256:07cb608d5bd74e0aec24f239878e88a00bde5b7b455b3a1f18eddfeb66c66f70

Observation 4f9a152f-3f27-4956-b755-346cfcb69f41 · outbound

This paper cites Llmeasyquant: Scalable quantization for parallel and distributed llm inference,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Llmeasyquant: Scalable quantization for parallel and distributed llm inference,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.624658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.624658Z digest=sha256:4c52fa72f43dabffc19cdee20bcce0bb3c9c7a5937c82ac370fe13c4d8a564f4

Observation 9036fda3-d7e0-4d49-8dd1-24fd1d3ae901 · outbound

This paper cites Flashattention- 2: Faster attention with better parallelism and work partitioning,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Flashattention- 2: Faster attention with better parallelism and work partitioning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.627076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.627076Z digest=sha256:cae80d0cc58678a71168a50c41f6add31b941b6d5343fb59124c999a897af56d

Pith citing papers

No inbound Pith citation observations are available.