Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2312.02406.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:33.532810Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 0bead57a-b683-4937-8140-0549c6ce35a4 · inbound
DataComp-LM: In search of the next generation of training sets for language models Efficient Online Data Mixing For Language Model Pre-Training
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4bd057bc-eb32-4353-914d-1e15184bd867 · inbound
DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks Efficient Online Data Mixing For Language Model Pre-Training
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation de4514b8-5107-4316-909a-45c8d58ea7e1 · inbound
Merge to Mix: Mixing Datasets via Model Merging Efficient Online Data Mixing For Language Model Pre-Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f784259-f93c-48de-97e0-b32d974601ba · inbound
GRAPE: Optimize Data Mixture for Group Robust Multi-target Adaptive Pretraining Efficient Online Data Mixing For Language Model Pre-Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62f76901-e1cd-4dd5-b557-a4ee9abbe3a4 · inbound
Rethinking Data Mixture for Large Language Models: A Comprehensive Survey and New Perspectives Efficient Online Data Mixing For Language Model Pre-Training
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e214dd-d14f-4ad1-af5b-4b898e03b140 · inbound
MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning Efficient Online Data Mixing For Language Model Pre-Training
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a757b4-b7af-4777-8df1-10dfab028887 · inbound
AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs Efficient Online Data Mixing For Language Model Pre-Training
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15210643-f738-44b6-9b55-0efb7dff02c0 · inbound
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text Efficient Online Data Mixing For Language Model Pre-Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 917c7a5c-e1c5-43cc-ba0e-1d04c9a4f981 · inbound
Language Models Improve When Pretraining Data Matches Target Tasks Efficient Online Data Mixing For Language Model Pre-Training
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bda05d1-b601-47aa-ac2f-10d9cc52fb0d · inbound
MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining Efficient Online Data Mixing For Language Model Pre-Training
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 29625642-4fda-4ab9-a4d1-8621cc8980ab · inbound
Data Mixing for Large Language Models Pretraining: A Survey and Outlook Efficient Online Data Mixing For Language Model Pre-Training
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2866833-2709-4d6f-bfaf-a598fcbefb75 · inbound
Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling Efficient Online Data Mixing For Language Model Pre-Training
Reference 268
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 978961dc-3dc4-47d2-9e4f-f5450bcf27c9 · inbound
Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them Efficient Online Data Mixing For Language Model Pre-Training
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80073e1f-33d6-402a-9663-849359df43e1 · inbound
Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a3e95546-417f-496d-b6ee-c29f887664a7 · inbound
Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcbbb887-5ef7-4809-91a6-a7d7a2c0f8ee · inbound
Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b1c2692-3879-48e7-8864-d1c738009978 · inbound
DRIFT: Refining Instruction Data via On-Policy Data Attribution Efficient Online Data Mixing For Language Model Pre-Training
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2e0072df-b689-447b-a45f-cd91b1e4564a · inbound
Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning Efficient Online Data Mixing For Language Model Pre-Training
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4c3cf89b-205f-46bc-b4ad-a8646687d9e1 · inbound
Smooth Scaling Laws Hide Stepwise Token Learning Efficient Online Data Mixing For Language Model Pre-Training
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c6a2181e-ab22-42fa-94b5-ecccc8901c7d · inbound
Smooth Scaling Laws Hide Stepwise Token Learning Efficient Online Data Mixing For Language Model Pre-Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64423c53-a41e-4643-9677-4e878e0b15e2 · inbound
WARP: Weight-Space Analysis for Recovering Training Data Portfolios Efficient Online Data Mixing For Language Model Pre-Training
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.