Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:19:11.474503Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 5 inbound Pith citation observations for arXiv:2506.05767.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:19:11.474503Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:55:09.763377Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T20:40:36.526819Z
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ef6c8e10-a1ae-4d1c-96dd-93af39aa0d70 · outbound
dots.llm1 Technical Report GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 874c6d55-412e-44e8-837b-22792e1024e3 · outbound
dots.llm1 Technical Report Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a929a264-7db0-4d92-973b-318edfa6424a · outbound
dots.llm1 Technical Report Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de6955d-0759-439f-88d1-fa8474b325bf · outbound
dots.llm1 Technical Report DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dda731e-f33b-44d3-8d76-0bfb6f1de269 · outbound
dots.llm1 Technical Report DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 466d1906-528a-4036-a4db-93f0a6e31ad2 · outbound
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70df57b9-3c64-48b7-86be-03d5b1a13380 · outbound
dots.llm1 Technical Report Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b325499-623b-42c7-a976-6ef01c17d7ee · outbound
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fdb3c03-4061-4d7a-9a7e-936015379f53 · outbound
dots.llm1 Technical Report Xiezhi: An Ever-Updating Benchmark for Holistic Domain Knowledge Evaluation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a16ed031-f900-4199-8887-d64ba2551741 · outbound
dots.llm1 Technical Report RULER: What's the Real Context Size of Your Long-Context Language Models?
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fd57e70-35bd-4eca-a815-f60b6b26379c · outbound
dots.llm1 Technical Report MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d0cfd26-094f-4b46-89d4-b83703665267 · outbound
dots.llm1 Technical Report LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4698fee8-303f-49a7-8c53-e933f36d40fe · outbound
dots.llm1 Technical Report Style Mixture of Experts for Expressive Text-To-Speech Synthesis
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f86bd160-19ff-4530-b1d6-2534e72d62a6 · outbound
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a3de6e9-767f-451e-a3a5-718ee58beeb3 · outbound
dots.llm1 Technical Report URLhttps://doi.org/10.1162/tacl a 00276
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9352a6-200d-43b2-8bf1-6ccd7310f589 · outbound
dots.llm1 Technical Report CMMLU: Measuring massive multitask language understanding in Chinese
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea5670b8-aafe-413c-89de-390a86bd6947 · outbound
dots.llm1 Technical Report From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b125ddbb-6f4d-44d5-8f15-a1a3eb7d7ef8 · outbound
dots.llm1 Technical Report A coordinated tiling and batching framework for efficient GEMM on gpus
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 15cae609-7370-44d3-b397-353c30ac1352 · outbound
dots.llm1 Technical Report URLhttps://doi.org/10.1145/3293883.3295734
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8527008f-023f-43a9-bbdc-fe0330356534 · outbound
dots.llm1 Technical Report Evaluating Language Models for Efficient Code Generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a816d02-a227-46da-b469-98334aa45ca4 · outbound
dots.llm1 Technical Report American invitational mathematics examination - aime
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de4206f7-58c5-4b9b-881c-f4cbb172ddd2 · outbound
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c1639d7-5a84-4d64-984c-650507aab5d5 · outbound
dots.llm1 Technical Report Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 199f002f-1d8a-42ba-ad1c-186077619a80 · outbound
dots.llm1 Technical Report Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08225807-ecd9-41ec-ac8d-8bc2997c788b · outbound
dots.llm1 Technical Report GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22c20e23-6217-4f49-9916-af3bbdd70b5e · outbound
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e802fe5-4bee-4ec8-8294-0adb838c78f0 · outbound
dots.llm1 Technical Report Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf8e0a40-cfe6-4325-a704-9ce7aabcf641 · outbound
dots.llm1 Technical Report SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55ab1894-7d2e-4258-b98e-5c51dfda6149 · outbound
dots.llm1 Technical Report Auxiliary-Loss-Free Load Balancing Strategy for Mixture-of-Experts
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e91f080-f5b3-4684-b3c8-6d239717794d · outbound
dots.llm1 Technical Report Small-scale proxies for large-scale Transformer training instabilities
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6f260a1-5374-492f-8e19-a98e7802c536 · outbound
dots.llm1 Technical Report Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4b98e07d-91b7-4e71-9c15-9aaaa0bc664a · outbound
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12a93e3-f33a-4404-8e35-4a89b11d0ae5 · outbound
dots.llm1 Technical Report Gated Linear Attention Transformers with Hardware-Efficient Training
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7fda504-eadb-42b7-8e76-ddf51d6378c0 · outbound
dots.llm1 Technical Report URL https://doi.org/10.1 007/s11227-022-04336-3
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 331151e7-c28b-4bb6-8c64-409cd28eb5db · outbound
dots.llm1 Technical Report URLhttps://doi.org/10.18653/v1/p19-1472
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32dcfcc-79c2-42d9-be34-be6939a940d3 · outbound
dots.llm1 Technical Report URL https://doi.org/10.110 9/HPCC-DSS-SmartCity-DependSys57074.2022.00143
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 075b5f16-5f97-4571-9558-5dd63cd38c43 · outbound
dots.llm1 Technical Report AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36f3b8fb-9001-4088-a469-8b4a76f468b4 · outbound
dots.llm1 Technical Report BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9949bfc3-4dfc-4637-8b6a-808984fff756 · outbound
dots.llm1 Technical Report Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 875f7de0-8f14-4d20-bc44-a8f2e8d682d7 · outbound
dots.llm1 Technical Report Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4d45c2e5-1bde-4045-b290-efc0c8da941a · outbound
dots.llm1 Technical Report The process involves the following steps: First, we apply the same text standardization as in the Identity Removal step
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c14e087c-596b-4339-8e96-85cebb7c55f8 · outbound
dots.llm1 Technical Report Quality ModelThe quality model performs comprehensive multidimensional analysis to evaluate and score training samples (Qwen, 2024a; Penedo et al., 2024)
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d0d39a29-5d2a-4235-a516-9c5f3ff028cd · outbound
dots.llm1 Technical Report The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07f4b476-928e-4ab7-affc-42a1d274c0a0 · outbound
dots.llm1 Technical Report FastText.zip: Compressing text classification models
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 229d4982-e67f-457b-ad3a-b22830717b8a · outbound
dots.llm1 Technical Report Training Verifiers to Solve Math Word Problems
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d21e61d-1cd5-4a46-a56e-4df876624cdc · outbound
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e046010-1d3e-4fb3-9e18-9fc4de538bc2 · outbound
dots.llm1 Technical Report URL https://doi.org/10.1609/aaai.v34i05.6239
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9968483b-4e90-42a4-b346-dbdfc6958140 · outbound
dots.llm1 Technical Report Yonatan Bisk, Rowan Zellers, Ronan Le Bras, Jianfeng Gao, and Yejin Choi
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 694b8fdb-094f-4618-9cf5-ffb90a1578b4 · outbound
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 014d4ab6-34bd-4b6c-bb88-9cc54102f850 · outbound
dots.llm1 Technical Report DROP: A reading comprehension benchmark requiring discrete reasoning over paragraphs
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d6f1e018-cf6d-4687-a980-ffa9bfeed1e7 · outbound
dots.llm1 Technical Report McEval: Massively Multilingual Code Evaluation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3381bbcd-06f1-4942-a951-f60497195e1f · outbound
dots.llm1 Technical Report Program Synthesis with Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a2fae88-c8b7-441e-a300-6c77883d71e3 · inbound
Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts dots.llm1 Technical Report
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f339dd7c-1b5f-4c73-b747-c9875d60c6f5 · inbound
UltraMemV2: Memory Networks Scaling to 120B Parameters with Superior Long-Context Learning dots.llm1 Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e67ca73-fa02-476a-bdb4-53b3e93a009e · inbound
Beyond Sunk Costs: Boosting LLM Pre-training Efficiency via Orthogonal Growth of Mixture-of-Experts dots.llm1 Technical Report
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 475a8df5-5fba-425a-be89-411437f1e31e · inbound
PolicyLong: Towards On-Policy Context Extension dots.llm1 Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1dc3dcc8-afca-4cf8-a372-b9febde852be · inbound
CuraWeb: Joint Optimization of Quality, Redundancy, and Diversity for Web-Scale Pretraining Data dots.llm1 Technical Report
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.