Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Speculative Decoding for Fast Ranking

As of 17 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 4 inbound Pith citation observations for arXiv:2505.20316.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20316 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:49.233352Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:01:18.473636Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T21:31:51.579040Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cd7e7623-affe-400e-af19-a87abb5b206a · outbound

This paper cites Learning to rank using gradient descent.

Reinforcement Speculative Decoding for Fast Ranking Learning to rank using gradient descent

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:54.309735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:44.807473Z digest=sha256:89159b281ace1d7b921add157232b0de7fa83bbabb9a6042003213a8ab82042d

Observation ab55c311-ad0f-4d2c-b5fb-1c1d418b5ce9 · outbound

This paper cites Medusa: Simple llm inference acceleration framework with multiple decoding heads.

Reinforcement Speculative Decoding for Fast Ranking Medusa: Simple llm inference acceleration framework with multiple decoding heads

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.981128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:44.877478Z digest=sha256:6b8b10fae643de057ae2bbd7e9e44b2a93d58fc13cf62313c4b605ab8cef36c3

Observation 951aa2c4-3b1b-48ab-b8b1-b65e47ceea8b · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Reinforcement Speculative Decoding for Fast Ranking Accelerating Large Language Model Decoding with Speculative Sampling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.008851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.008851Z digest=sha256:ff9f66055456b843d7278ad69ef0dad6688e8ad9542a42520dc5a826d786f1c3

Observation 5a260b83-81ef-4d10-b341-a15956c9039a · outbound

This paper cites Sequoia: Scalable, Robust, and Hardware-aware Speculative Decoding.

Reinforcement Speculative Decoding for Fast Ranking Sequoia: Scalable, Robust, and Hardware-aware Speculative Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.134263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.134263Z digest=sha256:21f702a270c91fc7192964ea9870d68347988ba1ff05664565a115b93bf302f8

Observation 2cb6909c-ba71-4577-ab3a-5da3a457fc6e · outbound

This paper cites Cascade speculative drafting for even faster llm inference.

Reinforcement Speculative Decoding for Fast Ranking Cascade speculative drafting for even faster llm inference

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.698037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:45.203410Z digest=sha256:70865181c639ce4632f875f6bbd322fcfb9e2fd5ba3f762af2ba5472ee168078

Observation eda251c9-b5ff-435b-a963-e0037c362d03 · outbound

This paper cites Glide with a cape: a low-hassle method to accelerate speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Glide with a cape: a low-hassle method to accelerate speculative decoding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.440746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:45.309951Z digest=sha256:516a6a4d3ec8982be76ed078aa64b00ca92588a2bc7a4e7185d2c24a424bf7b7

Observation cbc96cdc-c33b-471a-a449-d53158fdaa20 · outbound

This paper cites Quasi- metric learning for bilateral person-job fit.

Reinforcement Speculative Decoding for Fast Ranking Quasi- metric learning for bilateral person-job fit

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.165517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:45.427835Z digest=sha256:bb2e4cb8e6ea23907723c0c3914318f675c90a35fa2dcd92eb4e5b07ec56e8ba

Observation b3089c53-0cf5-492e-abc7-86b4c6618674 · outbound

This paper cites Enhancing job recommendation through llm-based generative adversarial networks.

Reinforcement Speculative Decoding for Fast Ranking Enhancing job recommendation through llm-based generative adversarial networks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.509766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.509766Z digest=sha256:b74d00c00f08e5f68e3b749dfa6aec521fed653d10f7ec300e0d87a583f18410

Observation 962a438b-8c4e-4740-a585-2a9b5cae7b62 · outbound

This paper cites Active large language model-based knowledge distillation for session-based recommendation.

Reinforcement Speculative Decoding for Fast Ranking Active large language model-based knowledge distillation for session-based recommendation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.827209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:45.595587Z digest=sha256:f3760acea052327cf7dd347b627ae8a4e7cd5444cc30a680ae9133aacaa8ba26

Observation 89c8e272-d3ee-4f42-a0b5-8714d252db41 · outbound

This paper cites Break the sequential dependency of llm inference using lookahead decoding.

Reinforcement Speculative Decoding for Fast Ranking Break the sequential dependency of llm inference using lookahead decoding

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.580518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:45.690591Z digest=sha256:3865e0f14cc7d90556ca3a8149bb3733e890d75f41559daf4f71195d48bae23a

Observation 1d7f1dd0-0f61-4d9a-b4ea-c07bfa6f41e0 · outbound

This paper cites The Llama 3 Herd of Models.

Reinforcement Speculative Decoding for Fast Ranking The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.758065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.758065Z digest=sha256:f3adf7363a38b3ab2e0b17b07eb738d2c26ed62be82502e897e605482e36da81

Observation f9821ace-78b5-4ccb-ba83-2d9a3b682592 · outbound

This paper cites Rest: Retrieval-based speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Rest: Retrieval-based speculative decoding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.343550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:45.829002Z digest=sha256:8a0694afccf5a8dd75374d64b0b5ee133f145435f51072e29007cac8857fc235

Observation df0725ad-3f72-4ec8-9388-84057782f793 · outbound

This paper cites SPEED: Speculative Pipelined Execution for Efficient Decoding.

Reinforcement Speculative Decoding for Fast Ranking SPEED: Speculative Pipelined Execution for Efficient Decoding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.915619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.915619Z digest=sha256:b7c7247cfe5edbef8286a55850e5c7e7c836fe4d8bb84cd1fe9221e92403e389

Observation a8fc6330-5401-4bd8-bdb5-520adafa04fd · outbound

This paper cites Large language models are zero-shot rankers for recommender systems.

Reinforcement Speculative Decoding for Fast Ranking Large language models are zero-shot rankers for recommender systems

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.030308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.027434Z digest=sha256:dd7738ea447f9af7c5f021eb41138b3810befe8fa4e425e56c10417a6d53685f

Observation 3b362025-19e4-4368-8fd2-366addca62b1 · outbound

This paper cites Neural input search for large scale recommendation models.

Reinforcement Speculative Decoding for Fast Ranking Neural input search for large scale recommendation models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:46.110093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:46.110093Z digest=sha256:e043273af1d6fbc6e35b6849a2756d51654db579f94a86b68754d0cb67523444

Observation 029d947d-c457-4564-b415-5c9a74fb6ac2 · outbound

This paper cites Speculative decoding with big little decoder.

Reinforcement Speculative Decoding for Fast Ranking Speculative decoding with big little decoder

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.858643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.201097Z digest=sha256:961b27348ce3e9a477dadfda29a8a1407c8cca287b4f0c4434cbb8eec2054f92

Observation fb07fa24-a7b0-4a81-99f8-db0e151e757a · outbound

This paper cites Ancestral gumbel-top-k sampling for sampling without replacement.

Reinforcement Speculative Decoding for Fast Ranking Ancestral gumbel-top-k sampling for sampling without replacement

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.645613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.305009Z digest=sha256:73107d649b3dc74b029e590482aa22d3e60018602f134959b45ef9e0aca7b138

Observation ac585f47-9ac5-4718-9f9b-35884071ad15 · outbound

This paper cites Fast inference from transformers via speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Fast inference from transformers via speculative decoding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:46.374306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:46.374306Z digest=sha256:22560cb8aadcf8786a53c4c38a7e8355bcb0f68733e711bcb04721db067ed9f3

Observation d596bb36-bcfc-407d-b4c4-dcb0513c0f27 · outbound

This paper cites Eagle: speculative sampling requires rethinking feature uncertainty.

Reinforcement Speculative Decoding for Fast Ranking Eagle: speculative sampling requires rethinking feature uncertainty

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.414693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.467020Z digest=sha256:6fdc0c04475e7267a1e32d125896b9fcedcc8e7d03ec535d483c617b8636a788

Observation 1184db20-b773-42a8-b66a-04d74ab532d8 · outbound

This paper cites Generalized ambiguity decomposition for ranking ensemble learning.

Reinforcement Speculative Decoding for Fast Ranking Generalized ambiguity decomposition for ranking ensemble learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.197273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.557451Z digest=sha256:b2269fcb9113b6f6992d78c808b2bc9363f9d7bc6e85960dcd16aa5633ee25fe

Observation b8570d3b-64b3-491e-a9f2-9008463571c2 · outbound

This paper cites Online speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Online speculative decoding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.105067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.700635Z digest=sha256:66a37b98467b29a79e29b07a16a5c8131ee53461d0227402238aeee19a8ed59c

Observation def8ea7e-72a2-4bd2-90f8-5d6597de3d3b · outbound

This paper cites Ranked list truncation for large language model-based re-ranking.

Reinforcement Speculative Decoding for Fast Ranking Ranked list truncation for large language model-based re-ranking

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.010746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:46.802330Z digest=sha256:e1f22fd64a89aed852c1d018aabb5ef93fc7c31d860a290871e12090c8b9a594

Observation 6f39294c-0397-4870-bea6-5bd41c1dca19 · outbound

This paper cites Specinfer: Accelerating large language model serving with tree-based speculative inference and verification.

Reinforcement Speculative Decoding for Fast Ranking Specinfer: Accelerating large language model serving with tree-based speculative inference and verification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:46.907373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:46.907373Z digest=sha256:4a19bbcbac017e17149314c438ee991061f5e1581b58eeafabe1a7ab98f37895

Observation 1232ac43-92d0-4be4-bdf3-4324e91ad33f · outbound

This paper cites PaSS: Parallel Speculative Sampling.

Reinforcement Speculative Decoding for Fast Ranking PaSS: Parallel Speculative Sampling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.016746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.016746Z digest=sha256:40cbfcf4fc602de50c9c703a436ca1a7ed331d2f198852d978e1370c005ad77a

Observation 0e1972d2-bad1-4713-adaf-c3c589e737b4 · outbound

This paper cites Machine learning: a probabilistic perspective.

Reinforcement Speculative Decoding for Fast Ranking Machine learning: a probabilistic perspective

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.125132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.125132Z digest=sha256:573d8827c8cd4ab809043b186a4d432d7baaf0342e0131238ffa89d97ef77213

Observation 1c218bff-cd9f-48df-b791-b8f3db60fdd4 · outbound

This paper cites Ms marco: A human-generated machine reading comprehension dataset.

Reinforcement Speculative Decoding for Fast Ranking Ms marco: A human-generated machine reading comprehension dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.197655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.197655Z digest=sha256:6b0b69d2093b74c9e1e2b8ee6ef686c23ce4fa5529bc32097797b7df1daa333f

Observation 25477b58-b72e-4a56-ab7f-b3a4cbfee646 · outbound

This paper cites Top-Down Partitioning for Efficient List-Wise Ranking.

Reinforcement Speculative Decoding for Fast Ranking Top-Down Partitioning for Efficient List-Wise Ranking

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.281945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.281945Z digest=sha256:03b13df965c279ee889d28030f935b1e1dc3fba6c7c309e3415390207315dca4

Observation fab25013-0afa-47e8-b14f-6c6605413863 · outbound

This paper cites RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models.

Reinforcement Speculative Decoding for Fast Ranking RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.416667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.416667Z digest=sha256:d8fb3318319ea451f9b5ef20479d7afb31ce94a04980f1ee82cafe9def9844bb

Observation 0fa9f67a-44a5-40cf-9f42-f42d2f6d5a56 · outbound

This paper cites RankZephyr: Effective and Robust Zero-Shot Listwise Reranking is a Breeze!.

Reinforcement Speculative Decoding for Fast Ranking RankZephyr: Effective and Robust Zero-Shot Listwise Reranking is a Breeze!

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.543984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.543984Z digest=sha256:33e0878659e67a1ef7a6612def3da356f1a9368cab95d5446b4d9ae6225407c3

Observation 38eb7458-6e76-4f52-832a-668851e71c9a · outbound

This paper cites First: Faster improved listwise reranking with single token decoding.

Reinforcement Speculative Decoding for Fast Ranking First: Faster improved listwise reranking with single token decoding

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.867711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:47.634430Z digest=sha256:d19d4fda0009ad91d5c7f65c7ed26b268b6a60e68111e72c15c4011a7c85efa0

Observation e0ca8bed-2e0d-4556-94f1-52948ba962f0 · outbound

This paper cites Accelerating transformer inference for translation via parallel decoding.

Reinforcement Speculative Decoding for Fast Ranking Accelerating transformer inference for translation via parallel decoding

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.719645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:47.742704Z digest=sha256:ba6da73091cce6b7eb9d8ab858f7f802bc09ede695f0ad25401ec0fe97e4f662

Observation c0831dec-029e-4720-a3f0-cd1b6d8bad79 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reinforcement Speculative Decoding for Fast Ranking DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.849400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.849400Z digest=sha256:4e1974065dd35d7266a823e38ca5bcdc3f488a0fe786579522eca4cc8037fe19

Observation 42cb5e4f-6e07-47cf-aa8b-216da521f1e6 · outbound

This paper cites Accelerating llm inference with staged speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Accelerating llm inference with staged speculative decoding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.581792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:47.938052Z digest=sha256:a7884a25ec5e4c40c35ff000a5fab7fb8635be891a44558409fc8c8d482abc10

Observation 2ed1d937-2e62-4e1d-b113-c74cc172ebed · outbound

This paper cites Blockwise parallel decoding for deep autoregressive models.

Reinforcement Speculative Decoding for Fast Ranking Blockwise parallel decoding for deep autoregressive models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.038119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.038119Z digest=sha256:b5d0cdbd93c3fc6eeb4c093c16ebecd2f93a3c3107110d515637f3209888fb06

Observation e324952c-13fd-4ee5-b4b4-9cb1bb572cc4 · outbound

This paper cites Spectr: Fast speculative decoding via optimal transport.

Reinforcement Speculative Decoding for Fast Ranking Spectr: Fast speculative decoding via optimal transport

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.454616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:48.131007Z digest=sha256:c97c8d508720f53a94f0112d48544f2eb8dd02be7563012eb124b24fac20f4e9

Observation 5543b4e6-97b7-44e7-8170-3a037c85c517 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Reinforcement Speculative Decoding for Fast Ranking LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.202843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.202843Z digest=sha256:d28295acddbc14a4a4dbc29dfcd2f3fbaf697c7fd2112ad663148d2182850476

Observation 00b9c292-7405-4106-9bae-a2ee1930512e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Reinforcement Speculative Decoding for Fast Ranking Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.302963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.302963Z digest=sha256:d4217e506a57868d11e6469c91556c9dd9ec58f8f5f902bc311c279488828f99

Observation f4907a0b-19fb-42fd-ad14-e7c5143efc21 · outbound

This paper cites Re2llm: Reflective reinforcement large language model for session-based recommendation.

Reinforcement Speculative Decoding for Fast Ranking Re2llm: Reflective reinforcement large language model for session-based recommendation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.328598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:48.394670Z digest=sha256:284f78f85136f5b55efa9819b957df996fbdf0adba9c14dde853ac25f5abd252

Observation 982ed65e-76c0-47ae-823b-8ac58dcdb0ba · outbound

This paper cites A survey on large language models for recommendation.

Reinforcement Speculative Decoding for Fast Ranking A survey on large language models for recommendation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.475112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.475112Z digest=sha256:cb49c657e4038e194215540a572e70de70eb31d298e0afffc9f99b07a748225a

Observation 2d3dab59-36fc-4ffc-89d2-527ca833f50d · outbound

This paper cites Speculative decoding: Exploiting speculative execution for accelerating seq2seq generation.

Reinforcement Speculative Decoding for Fast Ranking Speculative decoding: Exploiting speculative execution for accelerating seq2seq generation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.195461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:48.534931Z digest=sha256:8b92e52bab6677d8405f3ebc5e65ffa9e0de6aa8fd96bf2faddce5bd737d2fd1

Observation d903e4ea-3fee-46b8-851d-eaddab119383 · outbound

This paper cites Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.041078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:48.619763Z digest=sha256:0a59c5f29d613ba5f4b6e0349530c4dd68869c064584078a2538b6e3fe9773dc

Observation b00634c5-46fa-45d1-a089-0f120fd5358a · outbound

This paper cites Qwen2.5 Technical Report.

Reinforcement Speculative Decoding for Fast Ranking Qwen2.5 Technical Report

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.710769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.710769Z digest=sha256:6d0b2c325dbd1fa803de4eee6b070f8cb60a791994010aa53afcbc7f889e5517

Observation 3fc18fa0-9614-42a1-bc6c-c6d1c1c68ba5 · outbound

This paper cites Multi-Candidate Speculative Decoding.

Reinforcement Speculative Decoding for Fast Ranking Multi-Candidate Speculative Decoding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.798319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.798319Z digest=sha256:0583b0f553d64b15034a2d897f43ada88eb29a925326960cccb736a4c77efbc0

Observation d4381522-e859-41c0-8657-75ba81c155eb · outbound

This paper cites Predictive pipelined decoding: A compute-latency trade-off for exact llm decoding.Transactions on Machine Learning Research.

Reinforcement Speculative Decoding for Fast Ranking Predictive pipelined decoding: A compute-latency trade-off for exact llm decoding.Transactions on Machine Learning Research

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:49.919724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:48.853625Z digest=sha256:d08c81f6b3ba4b1b801997e72fe1b5fe12af7e80a6c6b073212c9859899b7dff

Observation c469d7b3-9ec3-40ef-98dd-a34d89f5821d · outbound

This paper cites Draft& verify: Lossless large language model acceleration via self-speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Draft& verify: Lossless large language model acceleration via self-speculative decoding

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:49.807720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:48.949300Z digest=sha256:31274a02269d465fc785f23d314b33c956834bc8b909feaf0b77ccf547edcca6

Observation f2f37503-db68-4a1d-8215-f8a7edfcb869 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Reinforcement Speculative Decoding for Fast Ranking OPT: Open Pre-trained Transformer Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:49.016380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:49.016380Z digest=sha256:42eec740b2ef8a5ff741d2140190824067c8c041201ca26de45149d438b6a6ac

Observation fb204efc-0b25-4d70-ae16-6aa166dcc0d1 · outbound

This paper cites Distillspec: Improving speculative decoding via knowledge distillation.

Reinforcement Speculative Decoding for Fast Ranking Distillspec: Improving speculative decoding via knowledge distillation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:49.683573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:55:49.138610Z digest=sha256:f8e69dc7b9b3e20236b0a621efcb8234d139863b720ffce8a7ef290eba98ee26

Observation a9040345-ff0a-4f93-975c-c3c4c446655a · outbound

This paper cites Large language models for information retrieval: A survey.

Reinforcement Speculative Decoding for Fast Ranking Large language models for information retrieval: A survey

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:49.233352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:49.233352Z digest=sha256:4bb4ece5874a0637d25448367218920cf95ec03259303f17366fb325e2f2167b

Pith citing papers

Observation 062f74ff-c5db-4046-844d-0f849a74c9dc · inbound

Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in Recommendation cites this paper.

Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in Recommendation Reinforcement Speculative Decoding for Fast Ranking

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:31:51.582307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-18T21:28:29.868292Z digest=sha256:19938817fd2d4c02402b9363d3b27e6b7fe5291552a60fb7130d4673329dacec

Observation fc09ac30-e460-4744-8903-82567dc2a135 · inbound

History Rhymes: Accelerating LLM Reinforcement Learning with RhymeRL cites this paper.

History Rhymes: Accelerating LLM Reinforcement Learning with RhymeRL Reinforcement Speculative Decoding for Fast Ranking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T17:01:18.473636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:01:18.473636Z digest=sha256:550ae966d7df51a76ea0abb9cf7274d9cbf5884d2e555234eece0e4beb60448a

Observation d650411c-fd95-43c3-9f9a-5d520b0ad975 · inbound

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs cites this paper.

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs Reinforcement Speculative Decoding for Fast Ranking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T01:33:51.671261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:33:51.671261Z digest=sha256:5cdfe74b46be0ca4d230cbcbdf56c01ba496827badb3b66e2f0725af6b368090

Observation 157a78b3-e3a3-4441-80ae-23fa8c7516c5 · inbound

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs cites this paper.

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs Reinforcement Speculative Decoding for Fast Ranking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T04:27:07.923007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:27:07.923007Z digest=sha256:3ccf830ba149cae43dfa84b1bd1bf9285164f337505dc45e8da3dc4f390a54bf