Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2207.00032.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T05:18:07.705592Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T22:59:03.322421Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 39900118-976a-4971-8a9c-d9c4cbe878d9 · inbound
H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 08dccdd5-98d0-49e9-8fc2-68424e18710b · inbound
Efficient Memory Management for Large Language Model Serving with PagedAttention DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dd5d2662-76b1-44e0-91a3-18f9d06fb1ef · inbound
Aero-LLM: A Distributed Framework for Secure UAV Communication and Intelligent Decision-Making DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58157a7a-196f-4952-a0fe-cc8c9c97bf12 · inbound
SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8abc7c3-bda3-4139-8f81-e0cb39023140 · inbound
Hardware-Efficient Attention for Fast Decoding DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f5b486c-3f15-42ee-8066-5e3bcf0cabab · inbound
EvolveSearch: An Iterative Self-Evolving Search Agent DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bc7a03b-05fc-4f7c-abe1-632d5628a92a · inbound
Combating the Memory Walls: Optimization Pathways for Long-Context Agentic LLM Inference DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de56e279-cca7-41aa-b6f5-8e0a892fad08 · inbound
SweetSpot: An Analytical Model for Predicting Energy Efficiency of LLM Inference DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 48682b07-9fa5-4f44-9259-f8e962a6d74b · inbound
Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 97c1c357-ece0-4476-ad31-83fd25f2782d · inbound
Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4b7cea5c-8342-4858-b4d1-0379d2ba33f4 · inbound
Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e122ca3f-c6ac-4e43-8450-0c2acab501e7 · inbound
Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bafbf0de-24ce-459a-8c15-d717f12901b7 · inbound
ShardTensor: Domain Parallelism for Scientific Machine Learning DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a83bf6aa-9455-413c-b910-edc84e1334a9 · inbound
A Readiness-Driven Runtime for Pipeline-Parallel Training under Runtime Variability DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0bf0fca9-42f7-4096-af22-a01062156484 · inbound
A Paired Testing Protocol for Batch-Conditioned Refusal Robustness in LLM Serving DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2ca665c7-8363-4c6b-8dd6-93f5aa9b6272 · inbound
AoiZora: Topology-Aware Auto-Parallel Optimization for Inference of Diffusion Transformers DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 009c2579-6677-4a94-b3b7-aa8f7fa29de0 · inbound
Optimizing Teacher-Student Partitioning for Scalable Knowledge Distillation on HPC Systems DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 46693fd0-7386-4723-be2c-47a2a130340c · inbound
GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.