Pith. sign in

Paper Citation Record · LEDGER

Preble: Efficient Distributed Prompt Scheduling for LLM Serving

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2407.00023.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.00023 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:57.925106Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:29:39.037995Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a3c5bdaa-3e0d-480e-b037-eedc1031caf7 · inbound

CoDec: Prefix-Shared Decoding Kernel for LLMs cites this paper.

CoDec: Prefix-Shared Decoding Kernel for LLMs Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:57.925106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:57.925106Z digest=sha256:298b4bf0584a86650200c9b6fb1a88a97d50316d3983a20729675ddc0a08d283

Observation 914c9eb8-8f75-475d-8c57-35b59e533f5e · inbound

Nexus:Proactive Intra-GPU Disaggregation of Prefill and Decode in LLM Serving cites this paper.

Nexus:Proactive Intra-GPU Disaggregation of Prefill and Decode in LLM Serving Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T19:03:57.101107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:03:57.101107Z digest=sha256:63a9182fa533d42aa93e643d71f0091e8738bc7676eded92c826fb54b237e51b

Observation 1993b332-271a-4fe0-bffd-aa27d0248f65 · inbound

A Distributed Learned Hash Table cites this paper.

A Distributed Learned Hash Table Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T18:46:55.092252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:46:55.092252Z digest=sha256:5690f243a5dc512367c845d8de4fef16ef830bc4050c1b93d6fd63fab036fd79

Observation 80d92b4e-2a3b-48e4-896d-9da1d4791dab · inbound

Efficient Remote KV Cache Reuse with GPU-native Video Codec cites this paper.

Efficient Remote KV Cache Reuse with GPU-native Video Codec Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:22:22.865185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T05:21:04.555356Z digest=sha256:7ac101bebd7715149efe63307985deec798050730d6cf89b5a2c5e918b34f6db

Observation 3865934e-50fd-4c41-b948-300605a9c65f · inbound

GORGO: Online Tuning for Cross-Region Network-Aware LLM Serving cites this paper.

GORGO: Online Tuning for Cross-Region Network-Aware LLM Serving Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T00:06:24.883100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:06:24.883100Z digest=sha256:42465d018f4b6fa647d97439b793ab66e34281186541140985e8b528ea843d60

Observation 45e2baa5-fec3-43a3-ab5f-2b16fc3de31c · inbound

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion cites this paper.

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:46:42.257505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T04:24:53.023047Z digest=sha256:5f5f4653c24712ab4419026cab769e9dfb174a45596750dbc11808b2fbacab19

Observation 25a56a5a-1cd6-403a-9cef-c493fcbebe45 · inbound

Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference cites this paper.

Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:25:46.437799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:27:13.781231Z digest=sha256:ac091af28411f13324bec568eb2452fa5d8b1bf508353d83c7e84fff12135a3e

Observation 002efe1e-7cd1-48e2-b671-536c3537e4a3 · inbound

Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference cites this paper.

Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:35:10.211865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T00:31:12.965178Z digest=sha256:c71a749e8e690813e838c5ec8938d5f8390a34413bfd8116030d1587a9e28db0

Observation c53960ad-a788-4de5-819e-5e4c1d220ace · inbound

Sparse Prefix Caching for Hybrid and Recurrent LLM Serving cites this paper.

Sparse Prefix Caching for Hybrid and Recurrent LLM Serving Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:53:04.396252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T08:49:24.880528Z digest=sha256:ff5409a9a63580024b536e48fc69c50c9f5ad0658ccb16a1f134c026356d05ad

Observation 020d313f-09c5-4e0a-bfaf-11265397f8f2 · inbound

GoodServe: Towards High-Goodput Serving of Agentic LLM Inferences over Heterogeneous Resources cites this paper.

GoodServe: Towards High-Goodput Serving of Agentic LLM Inferences over Heterogeneous Resources Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T19:32:43.652815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T19:31:43.015127Z digest=sha256:ab66e195a6ac19e633478db213ffa266ca19bb7dc07fe0ca3cc232c1f6adfec9

Observation 7fd7a353-79a0-4b27-92f2-55a094ba0971 · inbound

Beyond Prediction: Tail-Aware Scheduling for LLM Inference cites this paper.

Beyond Prediction: Tail-Aware Scheduling for LLM Inference Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:58:57.853073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T01:01:34.655458Z digest=sha256:8f3bfd85fa5b078d762d636a818a8e2189ac4d9bb0812231aa1ea28400ed8e8d

Observation 41c2e940-26d8-4f40-bb56-6ab5cbedcd0f · inbound

Recency/Frequency Adaptive KV Caching for Large Language Model Serving cites this paper.

Recency/Frequency Adaptive KV Caching for Large Language Model Serving Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:29:39.040238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T13:24:44.218871Z digest=sha256:eef2bb81da6a822f766d2bbd319e7b868452b8ad81bbad865e16765d8f5fb4c8

Observation 9ee1c01c-1454-4f03-9492-0fc51311bf0a · inbound

Omni-Flow: A Unified Workflow Orchestration and Distributed KV Cache Sharing Framework for Multimodal Inference cites this paper.

Omni-Flow: A Unified Workflow Orchestration and Distributed KV Cache Sharing Framework for Multimodal Inference Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T11:35:43.630892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-01T04:13:53.722933Z digest=sha256:4c6167f9f993916198b44385d58d55a703b1144c870fa4645ba86bb305d17ea3

Observation 6115a4e9-75a5-45a0-81e6-1da64a5b2434 · inbound

ELDR: Expert-Locality-Aware Decode Routing for PD-Disaggregated MoE Serving cites this paper.

ELDR: Expert-Locality-Aware Decode Routing for PD-Disaggregated MoE Serving Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T06:56:43.959288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T06:40:53.505889Z digest=sha256:e07818faee95f3a9766c9d15b5fd147bee24be58e2a8ebd0a61170637fece0dc

Observation 7031347c-1ee0-42f7-a41f-9cd5ecf43824 · inbound

ELDR: Expert-Locality-Aware Decode Routing for PD-Disaggregated MoE Serving cites this paper.

ELDR: Expert-Locality-Aware Decode Routing for PD-Disaggregated MoE Serving Preble: Efficient Distributed Prompt Scheduling for LLM Serving

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:58:50.434903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T18:58:44.830766Z digest=sha256:736c50b69ef1f7bccbf0bfb4c11f112a1a3133ade0497ce939cff0bd110f4f54