Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T10:14:15.769111Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2607.20327.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T10:14:15.769111Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7e6c972b-025f-4e46-90f7-24537d427c2d · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0545cfab-e194-4e13-8ad2-0ea522a64c51 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Accelerating Large Language Model Decoding with Speculative Sampling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac93d4e0-5e35-4490-885f-7b9d42ff64db · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7222c22-3a5e-4dda-9082-13cbf2705f88 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2142e741-ccc8-4c55-82c7-e54f924f1ce9 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference MiniLLM: On-Policy Distillation of Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad5fdaf3-25e9-4fd7-9e57-51fcfe9a1fd5 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference doi: 10.18653/v1/2024.acl-long.211
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a10967e-af6f-4723-9c09-c4f692f58998 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference AIME 2024 dataset.https://huggingface.co/datasets/HuggingFaceH4/aime_2024,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 715ee1e4-58b9-4173-a258-1f06008d172c · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference arXiv:2309.06180
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa795d4c-df4d-440f-bea1-abce99604d42 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Fast Inference from Transformers via Speculative Decoding
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9e56ab1-c631-452a-9a7e-f119d3609eb5 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference AIME 2025 dataset.https://huggingface.co/datasets/yentinglin/aime_2025,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b91f3a8-d83a-4312-9ee4-4266a104c688 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Isaac Ong, Amjad Almahairi, Vincent Wu, Wei-Lin Chiang, Tianhao Wu, Joseph E
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 530deb04-d18a-42bf-a086-9ff81f68fe40 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference RouteLLM: Learning to Route LLMs with Preference Data
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1410c7ab-d4e5-4fee-afb6-e11cc1348204 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a57438e5-8adf-41b2-a66c-fc8ac83177bc · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Learning to Decode Collaboratively with Multiple Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 723212a3-74bc-4b8a-a9ae-27150fab658f · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference MixLLM: Dynamic Routing in Mixed Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9308cf6-41ea-4b30-9dfb-0195d6820f7c · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Heming Xia, Zhe Yang, Qingxiu Dong, Peiyi Wang, Yongqi Li, Tao Ge, Tianyu Liu, Wenjie Li, and Zhifang Sui
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d44352d-c718-4c95-81c8-019008783cb1 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b676662-3a8c-4c5a-88bc-d3b92cbbcff0 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f84fa8b-bdc9-4553-88f3-5f5630551f5c · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference GlimpRouter: Efficientcollaborativeinferencebyglimpsingonetokenofthoughts
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d168707e-35fe-4e43-95d1-751aaff63efd · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9791e570-20ef-4b91-9855-395ec8fb6379 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8ef10d5-80c2-4f41-b327-f3fbafc0deb2 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Takeaway.The trace separates model roles cleanly: the SLM supplies the governing equation, while the LLM completes the BCC-specific substitution, unit conversion, and rounding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f250f1a5-a5dd-4351-b8e0-89b29314acf0 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Xinyuan Wang, Yanchi Liu, Wei Cheng, Xujiang Zhao, Zhengzhang Chen, Wenchao Yu, Yanjie Fu, and Haifeng Chen
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a404b17-ab33-40ef-b445-4eff6cb665a6 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 550fd14e-acad-4d3e-8dde-855233cb7ced · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 293ae174-3f07-4053-a5c8-0d0d8fdb048d · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc175f25-7347-46ce-9df0-4ad2217144d1 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a15f6de-1c3f-45dc-a32c-53193cb650e6 · outbound
PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.