Pith. sign in

Paper Citation Record · LEDGER

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts

As of 14 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 3 inbound Pith citation observations for arXiv:2412.01447.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.01447 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:26:59.641907Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T15:58:05.356765Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T21:57:38.315816Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7c041e5c-c5ee-438f-b824-906bedcf949b · outbound

This paper cites Hydra: Sequentially-Dependent Draft Heads for Medusa Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Hydra: Sequentially-Dependent Draft Heads for Medusa Decoding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.548823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.548823Z digest=sha256:7cb6401f7a38031534bd1db49194321e3e68654ff3a6a6217ac8adf5e3a11542

Observation 72239bfa-bf74-47b0-ad52-8c9f36a1ac3d · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.554432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.554432Z digest=sha256:18f6c8790f41511ee74ccaebcc1d8f72b5ff984227cc09b3f654f5efb32eadac

Observation a5889ef7-0e1b-4f18-a3a4-81cd68408397 · outbound

This paper cites Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.558035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.558035Z digest=sha256:dd7d47c55043b6781e64a7cf5331daf46ae9fc53761ffd61ba4473a30f9432a2

Observation f31cc967-2bf7-4145-95ad-afbe6b1d5328 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Accelerating Large Language Model Decoding with Speculative Sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.562551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.562551Z digest=sha256:a7858bc756f1b3e2a608ede8d5f4a96e3d6f16d2e6a1980245b97a49054c2e6e

Observation dc103ce0-8319-43b2-a2e8-2c5a98db2277 · outbound

This paper cites Break the Sequential Dependency of LLM Inference Using Lookahead Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Break the Sequential Dependency of LLM Inference Using Lookahead Decoding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.566128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.566128Z digest=sha256:c569d9ce9a7deb8df1265a022fb7a6cfeab80a7a91fc2f7740fa2320463eee44

Observation 66ae758d-d9fd-4db1-a6ff-87a38ad37b80 · outbound

This paper cites CodeEditorBench: Evaluating Code Editing Capability of Large Language Models.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts CodeEditorBench: Evaluating Code Editing Capability of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.569575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.569575Z digest=sha256:5f534d0f96dc8960bc95b5932a297d6aa93433e54aa0308e45a18295e89ecf7a

Observation 71fd460d-df3f-4a4d-91e7-f57d9d0207dd · outbound

This paper cites REST: Retrieval-Based Speculative Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts REST: Retrieval-Based Speculative Decoding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.573324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.573324Z digest=sha256:05910b12555e016020946fce3755a1f9c995512b37654c1ca923a56f3ee523f1

Observation 7a299b6a-fd7d-41bb-99d3-af084587349c · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:26:59.990407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T04:26:59.576927Z digest=sha256:8fab9026814a29843e50b8aa89f86b2ab364ed72bcd9cb74ac38f09a2826b745

Observation 57773ced-c4d9-48ca-9b17-9251afbdaffe · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.580077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.580077Z digest=sha256:9841a3beaf844f649d3ef39d495797c96e1b4c76f46928c7c7f391ad7790c020

Observation 34d71c21-e9f1-4f5d-a322-e0799c068ed6 · outbound

This paper cites A Survey of Large Language Models Attribution.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts A Survey of Large Language Models Attribution

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.583061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.583061Z digest=sha256:040e7229d9d263ba9c9671afe7159d3f3af4f5b9eada7b3320b50e9a6fc4f0f6

Observation 3dfe4135-9132-40f1-8083-b58527d2c6f5 · outbound

This paper cites EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.586480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.586480Z digest=sha256:004df2a8d3fde8e41436f5232b6979a77998ede06676022d07c0f6355654aa13

Observation 0e078163-0c18-4ab5-9e05-0fc8c0a9cc90 · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:26:59.975888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T04:26:59.590124Z digest=sha256:b1fee15d31e6be093948d0e4a06b780459a06d0c029ece81e8fa3f845df1656a

Observation 5c89d983-f2c5-487c-92fc-a090ca5e31f1 · outbound

This paper cites PaSS: Parallel Speculative Sampling.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts PaSS: Parallel Speculative Sampling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.593246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.593246Z digest=sha256:8eafc93449aef46f82cc6e39362862b1e75500f886985695aba8c6362b862dcb

Observation 30c431f7-c965-4238-9af7-ae4f21806233 · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.596732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.596732Z digest=sha256:b7dcc464367f00692cdd9302b24b2f9c173885710d21a0ad99d8c8a4fe1f4056

Observation b8842d11-f8e6-4f63-9662-0a137645b350 · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.599985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.599985Z digest=sha256:047b5a4b371581cc63cfaab887d42297079eefa4df6798f0271cc6e73571644b

Observation 22967f17-670e-4106-88ad-1d36c0485f37 · outbound

This paper cites Peering into the Mind of Language Models: An Approach for Attribution in Contextual Question Answering.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Peering into the Mind of Language Models: An Approach for Attribution in Contextual Question Answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.602918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.602918Z digest=sha256:947238a7109609a61d2eb214072d3507aebfd5db37051b0b9b65f3cc0e77108c

Observation 1a1b94c2-772d-4721-ac3e-468c37787079 · outbound

This paper cites Accelerating Transformer Inference for Translation via Parallel Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Accelerating Transformer Inference for Translation via Parallel Decoding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.606043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.606043Z digest=sha256:e4eb12d9e396161553460a35b3f55b23d16d8348e65abbf27778ff35e5199028

Observation bacb4164-4dbb-48b1-9ca3-e86c4ba1c671 · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:26:59.961653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T04:26:59.609507Z digest=sha256:c435ac07cc8dee6fb33236d6496a9a15e6e0a904178e923f8094097ce5065d20

Observation 5fc2a3db-201e-4a2b-9b4f-2939625a95f8 · outbound

This paper cites SEMQA: Semi-Extractive Multi-Source Question Answering.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts SEMQA: Semi-Extractive Multi-Source Question Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.612503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.612503Z digest=sha256:afe2ea51d0627bd1930d1c506c5da2b5e743e26178a590f089e8e496f179b9c9

Observation 62461f7a-7de6-4a8f-8634-bd6c52c3c91d · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.616025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.616025Z digest=sha256:efedd1b9dd8ef9329de2f27e192d4c812ec95fce7df479816fc65869e23825fb

Observation 1bfa167f-cdf0-4eb9-9c36-a736132907b6 · outbound

This paper cites Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.618955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.618955Z digest=sha256:983dd5716dbdd39dff7ec1e3b4b2e5b8159e67fcd4f6653ca96adae1895c2966

Observation 4e0058d6-38b9-40df-9a6c-606ef1746799 · outbound

This paper cites Inference with Reference: Lossless Acceleration of Large Language Models.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Inference with Reference: Lossless Acceleration of Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.622263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.622263Z digest=sha256:52563018b9acd02980cb90dd412744a1f79cb3d63001f087bef1f3b12063ef15

Observation f629380b-9983-4065-bf24-3eae929c4455 · outbound

This paper cites Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.625468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.625468Z digest=sha256:8db5fc868dbf350d995bfb3825cd81a75acb1afc3bc45baa6f0f2b9fb72d37ba

Observation 70bc2ee8-1378-4273-a1e7-7e03167a82b6 · outbound

This paper cites XATU: A Fine-grained Instruction-based Benchmark for Explainable Text Updates.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts XATU: A Fine-grained Instruction-based Benchmark for Explainable Text Updates

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-12T04:26:59.696939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-12T04:26:59.628545Z digest=sha256:b6cb40dcb1ed0eb9b7f578c4e62bf77334afb7382a43b4b086f22a0d1ae3da8d

Observation 6f8e579c-8b50-4ff5-8380-768aa5ef3a4d · outbound

This paper cites Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.631815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.631815Z digest=sha256:d858691e5b2679785fe159bee84e8acb035d3d8c21888ee731345c42a8f2db33

Observation 6bab6bb1-a413-43f8-b1b5-5cc13cd09350 · outbound

This paper cites an unresolved cited work.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.635085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.635085Z digest=sha256:18b154d02a6d7444f48e1feb2f52292973542d3edd032804b8ecded059c5b9b8

Observation dd9271dd-55e9-4020-83eb-1bc10e40ed25 · outbound

This paper cites online" 'onlinestring :=.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts online" 'onlinestring :=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.637975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.637975Z digest=sha256:e49935a067af1fcc8f16bc57888a096af50e051941b22894ae090e087fa2bae2

Observation 75979415-106c-4a32-b2f8-f245363b930b · outbound

This paper cites write newline.

PLD+: Accelerating LLM inference by leveraging Language Model Artifacts write newline

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T04:26:59.641907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:26:59.641907Z digest=sha256:5c28e2f60907252ba414cde99bd604616e49fc6601b4d6b382fc97f3daf74fba

Pith citing papers

Observation a68cddf1-9f91-4ee5-8579-1e988d9fe21a · inbound

JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting cites this paper.

JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting PLD+: Accelerating LLM inference by leveraging Language Model Artifacts

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:18:59.277568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T00:42:52.485768Z digest=sha256:2bc61c1d171667f6547f27caf3e09510a306a564e9e3a05be70f86b5dd9057dd

Observation b5961e8d-b4d5-47f2-8eda-10a18dedeb5e · inbound

Trees from Marginals: Autoregressive drafting with factorized priors cites this paper.

Trees from Marginals: Autoregressive drafting with factorized priors PLD+: Accelerating LLM inference by leveraging Language Model Artifacts

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-10T21:57:38.332138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-10T21:49:34.512370Z digest=sha256:89f375b4970264bae455d8f67c5b0a0cb1613d1a4d47f95f8ef96f470debeee1

Observation 1cb9ecee-da1a-424b-a225-49888d282dd6 · inbound

Trees from Marginals: Autoregressive drafting with factorized priors cites this paper.

Trees from Marginals: Autoregressive drafting with factorized priors PLD+: Accelerating LLM inference by leveraging Language Model Artifacts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T15:58:05.356765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:58:05.356765Z digest=sha256:4cd2c5e52fffdf28a617d980a5b0cc71e3db020ade33c666bdc910e9c5db982d