Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2411.01783.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:59.378063Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T07:27:44.457158Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation de723323-0694-46f5-8a4c-dcd90f7cf9bd · inbound
CoDec: Prefix-Shared Decoding Kernel for LLMs Context Parallelism for Scalable Million-Token Inference
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a531bbb9-9266-4ee2-8fbd-2c505209c3fb · inbound
PRISM: Probabilistic Runtime Insights and Scalable Performance Modeling for Large-Scale Distributed Training Context Parallelism for Scalable Million-Token Inference
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 855e07e4-b108-4c24-86cd-a60d72f0414a · inbound
Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap Context Parallelism for Scalable Million-Token Inference
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04656ae5-6a52-436c-bd45-ecf99efbe174 · inbound
Neural Attention Search Linear: Towards Adaptive Token-Level Hybrid Attention Models Context Parallelism for Scalable Million-Token Inference
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c00ec107-823a-42a1-989a-50fcfe9ee556 · inbound
Scal3R: Scalable Test-Time Training for Large-Scale 3D Reconstruction Context Parallelism for Scalable Million-Token Inference
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bbf67b1-a67e-47ae-ba30-5b67cafbbf4c · inbound
Accuracy Is Speed: Towards Long-Context-Aware Routing for Distributed LLM Serving Context Parallelism for Scalable Million-Token Inference
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5af68555-97f5-4986-ad77-b6ede239c494 · inbound
GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Context Parallelism for Scalable Million-Token Inference
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 892ee340-59b7-4735-b4d8-5966013d2a57 · inbound
Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics Context Parallelism for Scalable Million-Token Inference
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4b8db0c8-f2d6-4491-91e5-8e39f5b70d46 · inbound
Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design Context Parallelism for Scalable Million-Token Inference
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3fc6e178-fbbf-49ca-9ec6-df41b844ad57 · inbound
Towards Distributed Inference of LLMs on a P2P Network Context Parallelism for Scalable Million-Token Inference
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 744dfe47-89ce-4be2-8d8f-d136776e890b · inbound
KernelFlume: Elastic Core-Attention Scaling for Agentic Long-Context Decoding Context Parallelism for Scalable Million-Token Inference
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.