Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T13:40:56.623496Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2502.02040.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T13:40:56.623496Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4125d7d7-4910-4038-a307-f73154be051b · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Phi-3 technical report: A highly capable language model locally on your phone, 2024
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d0f03468-926a-41b4-8a48-0838424117a1 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Llm inference performance engineering: Best practices., 2023
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 71a53602-126b-415c-b675-abe2ed4fdebc · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fd5c2b9-a9c3-4750-aed2-84c18df135ad · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Colt5: Faster long-range transformers with conditional computation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ba66cd9e-843e-44c9-bdbe-c1c10b38aae4 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Hydra: Sequentially-dependent draft heads for medusa de- coding, 2024
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6d74ad29-afe5-4099-95e7-84ca590ae6c9 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Massively multilingual sentence embeddings for zero- shot cross-lingual transfer and beyond
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 41fab238-730c-4342-91a8-007410e1ba14 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b7497dd7-1fc9-440e-a26d-dc0aaff5af52 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Mt-bench-101: A fine-grained benchmark for evaluating large language models in multi-turn dialogues
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0d7e73a3-b19b-4020-8b21-4c906ee9a86a · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fd1e4e1-0fe1-44e3-ab53-3f16e10cdc7f · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Language models are few-shot learners
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4303cf81-200f-4f73-b870-0c06615b8231 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Medusa: Simple framework for accelerating llm generation with multiple decoding heads
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aac75849-a2df-4b96-a798-56abaf1bfb20 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Accelerating Large Language Model Decoding with Speculative Sampling
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 130c8be1-1ea0-485f-89b6-1dc16780f1fd · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ac1876a-3b63-4d21-a25c-97cf08a7e166 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference DialogSum: A real-life scenario dialogue summarization dataset
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 21915d66-c976-49e7-91c7-a3ebf8c8865f · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Koala instruction set documentation, 2023
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 227accd1-0895-45b8-bca0-66d1ef3531a3 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Dai and C
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e26b19c1-73ca-4a02-901d-1d9b158496c8 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference SkipDecode: Autoregressive Skip Decoding with Batching and Caching for Efficient LLM Inference
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e41a4539-c671-4783-bc6a-9b7406b46111 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference SpQR: A Sparse-Quantized Representation for Near-Lossless LLM Weight Compression
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e487e9-a792-4216-8855-6e30d3dee94a · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference GLaM: Efficient Scaling of Language Models with Mixture-of-Experts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 160c048d-2e2c-45f4-b846-0c76eae1e813 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Evaluating the State-of-the-Art of End-to-End Natural Language Generation: The E2E NLG Challenge
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 609d15ac-e549-4076-9b5a-8e8fab22c3dc · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Depth-adaptive transformer
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f0c36f0c-7c4e-4754-bc62-a18a23bbc08f · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Predictive Exit: Prediction of Fine-Grained Early Exits for Computation- and Energy-Efficient Inference
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f9c0b924-8153-49ff-ba7b-e93db0e1bb26 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Decoupled early time series classification using varied-length feature augmentation and gradient projection technique
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7084d49a-8e22-478c-a066-6a00445acd05 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4431a023-fc9e-4a15-b5d5-2dcf933839a3 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1b7f3946-3ba4-45cb-9ffe-e1ee086430ec · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Sparsegpt: Massive language models can be accurately pruned in one-shot
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb98c18-d580-4034-8917-807fbbb8ad18 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d4d0187-f8ff-4d4b-b8aa-da7a922e888a · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Breaking the sequential dependency of llm inference using lookahead decoding, November 2023
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4aa6fb13-2c4e-41b8-aacc-e8cb3ca8c476 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Garncarek and J
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ecfd5849-273a-42c5-98b4-890529b78c3a · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Koala: Dialogue-based fine-tuning improves factuality and safety of llms
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 85ed8cbd-fb4c-4221-b2c5-350088de0f74 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference MiniLLM: On-Policy Distillation of Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bca9d024-aab7-43b6-9906-3b91c0ed9afd · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Identity mappings in deep residual networks
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 26d78d7a-9e40-4440-85c6-0dca8b7a41ed · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Dynabert: dynamic bert with adaptive width and depth
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation faac1eb5-92f3-46ec-9546-12b43bf53dec · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Adaptive mixtures of local experts
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9f701b72-db33-40fa-87e4-9b0568cc5a4c · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Mixtral of Experts
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1bca0b5-6a53-4789-89b7-97a58e7df821 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Hierarchical mixtures of experts and the em algorithm
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8fdaa0e2-c1c9-4eea-b6c6-0c14f2a38fa0 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Ten lessons from three generations shaped google’s tpuv4i
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5e1ca558-39b8-43f0-b235-0172eaaed0ea · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference In-datacenter performance analysis of a tensor processing unit
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1e7dd89a-5e9b-4a86-9cba-04cb45604405 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Gpus and the future of parallel computing
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1fe90847-784b-4f93-afd3-b61115c1f186 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Ai and machine learning acceleration in mobile devices: A survey of architectures, hardware, and algorithms
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 178775d1-af33-4df4-9fab-7f32a86e552a · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c83a3b46-ca25-44a7-89b4-fdadeadc8250 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference GShard: Scaling Giant Models with Condi- tional Computation and Automatic Sharding
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bf0974c2-d4f8-4464-b951-dc8ea5df6241 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Fast inference from transformers via speculative decoding
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a989bc8e-992a-4d64-9e55-4178ae7d21e2 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Gemma: Open Models Based on Gemini Research and Technology
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7e6870a-1e77-4fe2-9c32-194f4a17984c · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference SpecInfer: Accelerating Generative Large Language Model Serving with Tree-based Speculative Inference and Verification
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a878b6c1-c1d3-4107-8066-86a96b964406 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference OLMoE: Open Mixture-of-Experts Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c104600-4e33-4238-8012-774f324cb833 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Efficient large-scale language model training on gpu clusters using megatron-lm
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f9aa0612-9a4f-46d0-9529-450ac437ab24 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Cuda c++ programming guide, 2021
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6e7f2e82-0775-4321-b5de-c29ebf5ee8af · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference GPT-4 Technical Report, 2023
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaa65c5c-e415-45b0-a151-d99761e0f97d · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Language models are unsupervised multitask learners, 2019
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4b394c27-284a-410a-9876-ede0ab73677d · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Mixture-of-Depths: Dynamically allocating compute in transformer-based language models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba1dab6-d238-47d8-b6e7-397f3f7d009d · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99140b0c-0b9b-45c6-bccf-dd6f79f30650 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Tran, Yi Tay, and Donald Metzler
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fc8612f8-4408-4a35-ba5e-20b685fb18db · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82d54c11-66cc-49ee-b739-1d0bc93339bc · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 88d5c12a-7214-4143-a264-5035e6bc7063 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f77c6337-9423-4108-8396-e2ff63b2e4f7 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Accelerating LLM Inference with Staged Speculative Decoding
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8082d3ff-2a22-46ec-b727-441fb821c775 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference A Simple and Effective Pruning Approach for Large Language Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 204e7fcc-a2aa-4a6b-a5f9-999fb832453c · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Manmatha
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aab5ecd6-5ddd-4812-9a99-a9d0e01cbdea · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Tuan Pham, S
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3340cb68-c9a3-433b-ac70-c5a9fc519aa3 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Model cascading: Towards jointly improving efficiency and accuracy of nlp systems
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a84b56e0-c4d7-43ac-9b3e-6fa0effefe23 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Investigating acceleration of LLaMA inference by enabling intermediate layer decoding via instruction tuning with ‘lite’
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eaccb3ce-68cd-4c3f-86a9-c4aa429352f3 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Attention is all you need
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 941f8570-0355-4682-9b53-c23bb9ffce91 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Self-Instruct: Aligning Language Models with Self-Generated Instructions
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb6434e6-a44b-44f0-8041-5c12645a8b5f · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Transformers: State-of-the-art natural language processing, 2020
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f6407729-53b6-45dd-acbe-16cf625e36cf · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Speculative decoding: Lossless speedup of autoregressive translation, 2023
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 09a028c5-939f-4d78-ba5c-e2473c964c4d · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Deebert: Dynamic early exiting for accelerating bert inference
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 772efeba-ed59-4ef5-a62b-88fe4ebbe676 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference WizardLM: Empowering large pre-trained language models to follow complex instructions
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3788be9-4e4b-4053-ad3c-d74088f4b257 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Zeroquant: Efficient and affordable post-training quantization for large-scale transformers
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 350e3e94-af68-4af6-9d1a-37eba339cc6f · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Spider: A Large-Scale Human-Labeled Dataset for Complex and Cross-Domain Semantic Parsing and Text-to-SQL Task
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e22e5582-7e5d-45ca-aeec-b77e78ccb127 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference P Xing, Hao Zhang, Joseph E
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 334aef28-f2e3-4be0-b144-280af8aee97c · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da1245c6-485f-4294-88a2-171311cd026f · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Zhou and et al
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e687d854-34be-4b62-9d06-1176b7fdb6f1 · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Designing efficient sparse expert models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f72b9031-d6d4-4fbb-9136-bdaa7c543a1b · outbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.