Pith. sign in

Paper Citation Record · LEDGER

Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2404.16283.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.16283 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:35:08.205979Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T11:59:50.606807Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 77d1244f-8a61-41dc-9884-cd0b3331cf0b · inbound

Compass: SLO-aware Query Planner for Compound AI Serving at Scale cites this paper.

Compass: SLO-aware Query Planner for Compound AI Serving at Scale Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:15:03.461802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T19:14:58.225504Z digest=sha256:fd88670103eaf3f20d217887c56b1a18b8bf67b76aee1e6b4aa6b75b25795eb5

Observation 33b7247d-734e-4d43-ab8e-3604e6ee34dd · inbound

ServeGen: Workload Characterization and Generation of Large Language Model Serving in Production cites this paper.

ServeGen: Workload Characterization and Generation of Large Language Model Serving in Production Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-22T15:44:58.125103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T15:42:05.266854Z digest=sha256:5490ce68db889e99860c245039117b65e4a1e958cc28edee443986f7346db28c

Observation 6344b794-dfc0-401e-8af2-abd8daee2be0 · inbound

ConsumerBench: Benchmarking Generative AI Applications on End-User Devices cites this paper.

ConsumerBench: Benchmarking Generative AI Applications on End-User Devices Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:08.205979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:08.205979Z digest=sha256:3bac3dae93fdb88bd3dda1e1d432253f636c020d15812b7c5921cc9f3c2bdb0b

Observation 66d38722-36b3-4010-bb3f-fde70f928c6d · inbound

KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows cites this paper.

KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:45:42.375175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:45:42.375175Z digest=sha256:dd626447485b41820cd3d08a958ad2beb250502e13f2cc3e3863665005680cd4

Observation 0552ed18-9f58-4b02-b3a7-d9d40a6cc4c6 · inbound

GPU-to-Grid: Voltage Regulation via GPU Utilization Control cites this paper.

GPU-to-Grid: Voltage Regulation via GPU Utilization Control Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T04:26:57.612039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:26:57.612039Z digest=sha256:f5dc2d3459f3f271bd8f41000789fda8780989c1e1c1c5a9518267c3d37d5362

Observation 09b9884f-5a8f-4533-beb7-feccd839bb07 · inbound

CacheFlow: Efficient LLM Serving with 3D-Parallel KV Cache Restoration cites this paper.

CacheFlow: Efficient LLM Serving with 3D-Parallel KV Cache Restoration Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:21:14.044163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T15:32:32.808799Z digest=sha256:6dfefd6071dfb549a96a1e44002aaa4ff840775640615b323490e9005055d5f4

Observation e083e0e6-54e9-48b2-a6a9-24aa54a1a994 · inbound

Regulating Branch Parallelism in LLM Serving cites this paper.

Regulating Branch Parallelism in LLM Serving Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:50:58.495124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T01:00:53.308946Z digest=sha256:ecf19619518a17e07a8a9f5e99fbed9b968702855a6ce3e4ade93287e34e6df1

Observation b365884e-4e18-45ea-acbb-ccd89dfaaa44 · inbound

Beyond Prediction: Tail-Aware Scheduling for LLM Inference cites this paper.

Beyond Prediction: Tail-Aware Scheduling for LLM Inference Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:58:57.869067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T01:01:34.655458Z digest=sha256:85e48c8da51fa86f676189e7eb7055d4405d4a7f617c86b4d5d6fb7c2f96a8bd

Observation b186d42e-1a54-408f-8355-786c48aa6ce9 · inbound

LiveServe: Interaction-Aware Serving for Real-Time Omni-Modal LLMs cites this paper.

LiveServe: Interaction-Aware Serving for Real-Time Omni-Modal LLMs Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-04T11:59:50.608461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T07:26:07.356352Z digest=sha256:745cc6e4cb01bad3312b40ef6b4ce55f041db24107bc0887e2de1db5b7d6bf9c

Observation 4b379061-8c9a-480b-a59b-31ea0422ed20 · inbound

EnerInfer: Energy-Aware On-Device LLM Inference cites this paper.

EnerInfer: Energy-Aware On-Device LLM Inference Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-04T11:29:50.704798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T07:58:29.150228Z digest=sha256:3d7479b721a2bf27f75d9ffacaa996d0efda024e3fe9079651f935f62e8681e6

Observation 122d0a5f-1620-4849-a34a-e6de1496ce01 · inbound

Leaky Language Models: Stealing Architecture and Inference Optimizations via Per-Token Timing cites this paper.

Leaky Language Models: Stealing Architecture and Inference Optimizations via Per-Token Timing Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T09:35:56.582200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:35:56.582200Z digest=sha256:82957cb174ff49f0a18356b9f8a2e528de58d2046f342f86271622cdfe5e5fdb