Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2307.02628.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T04:30:14.503529Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 5df63f8b-9d33-44e0-94ca-b8818c6cd662 · inbound
A Survey on Efficient Inference for Large Language Models SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 111
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a11ad5dc-7073-471b-a5e2-d670ee99f0ba · inbound
AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e26b19c1-73ca-4a02-901d-1d9b158496c8 · inbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ab3af2b-2f8d-4b4c-bcae-ff369ea093b4 · inbound
UniMoD: Efficient Unified Multimodal Transformers with Mixture-of-Depths SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76a8c7e2-caa3-4204-8c28-77bb60f2d59f · inbound
DASH: Input-Aware Dynamic Layer Skipping for Efficient LLM Inference with Markov Decision Policies SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac8afd31-eb27-4983-b3d3-4b83630c0fbf · inbound
System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a28ec175-a345-4258-aa1f-c00f2cc48c57 · inbound
CLaSp: In-Context Layer Skip for Self-Speculative Decoding SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8387e030-7012-406e-b94c-0a712bcafefc · inbound
AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e81d5a2-1bf9-4671-b9ad-37fb719fdde5 · inbound
SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4905ae9-ab05-421f-bcd0-23469d277d8a · inbound
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fe2c9d0-f948-465b-942a-3381ee8bc030 · inbound
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67315da6-4d1b-44f4-ab48-8590e629b0a2 · inbound
CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa2da869-017f-4f5f-8bc1-2d4616a095ca · inbound
VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f9d686ef-61f8-41bd-9c6d-53bf6b1067af · inbound
QTALE: Quantization-Robust Token-Adaptive Layer Execution for LLMs SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dfb2969-6b3f-41d5-bc01-e6d110b2d784 · inbound
Networking-Aware Energy Efficiency in Agentic AI Inference: A Survey SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 346aa319-838f-44e8-b307-fa6bad1c7577 · inbound
When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c95ffadd-1837-459b-af2e-d5db23303a91 · inbound
Depth Adaptive Efficient Visual Autoregressive Modeling SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7267a922-ce82-484f-b1ea-449b60807c5a · inbound
River-LLM: Large Language Model Seamless Exit Based on KV Share SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation de247a6c-11ba-4492-ac32-750217c5860b · inbound
Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 765d1458-1b04-4ec3-bec6-1fb1f8c05cd9 · inbound
N-vium: Mixture-of-Exits Transformer for Accelerated Exact Generation SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 28dd328d-0752-4925-a4a8-742b323b8eca · inbound
Constraint-Driven Model Optimization: An Industry Framework for Selecting Compression and Acceleration Techniques in Modern Machine Learning Systems SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf4ba506-0967-4a2d-ac28-8682fd074c09 · inbound
Per-Token Fixed-Point Convergence in Depth-Recurrent Transformers SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c403b17c-df75-49e6-a8f4-2bb4ad3f6767 · inbound
Learning to Predict Middle-Layer Attention in MLLMs for Visual Token Prunin SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.