Pith. sign in

Paper Citation Record · LEDGER

Efficient Online Data Mixing For Language Model Pre-Training

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2312.02406.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.02406 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:33.532810Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0bead57a-b683-4937-8140-0549c6ce35a4 · inbound

DataComp-LM: In search of the next generation of training sets for language models cites this paper.

DataComp-LM: In search of the next generation of training sets for language models Efficient Online Data Mixing For Language Model Pre-Training

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:58:16.837009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T22:58:16.523267Z digest=sha256:d3804508acb8b2808ac050bdf52b843e971e453630ee3d035addc566893e26d3

Observation 4bd057bc-eb32-4353-914d-1e15184bd867 · inbound

DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks cites this paper.

DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.417636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T03:58:48.967122Z digest=sha256:3b5929ed0fbce160434c64143ef3e41e4f34bc55b1305c95e3f3c331d796ab88

Observation de4514b8-5107-4316-909a-45c8d58ea7e1 · inbound

Merge to Mix: Mixing Datasets via Model Merging cites this paper.

Merge to Mix: Mixing Datasets via Model Merging Efficient Online Data Mixing For Language Model Pre-Training

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:33.532810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:33.532810Z digest=sha256:0e2b48652e77de11e5b9f500ae753d7270637a017244e7b550d6fb26cbd525bd

Observation 4f784259-f93c-48de-97e0-b32d974601ba · inbound

GRAPE: Optimize Data Mixture for Group Robust Multi-target Adaptive Pretraining cites this paper.

GRAPE: Optimize Data Mixture for Group Robust Multi-target Adaptive Pretraining Efficient Online Data Mixing For Language Model Pre-Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:24.571162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:04:24.571162Z digest=sha256:271755370cb3d9c22b5555b821b59920077d31be9028a669bfd1c9682004e320

Observation 62f76901-e1cd-4dd5-b557-a4ee9abbe3a4 · inbound

Rethinking Data Mixture for Large Language Models: A Comprehensive Survey and New Perspectives cites this paper.

Rethinking Data Mixture for Large Language Models: A Comprehensive Survey and New Perspectives Efficient Online Data Mixing For Language Model Pre-Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:48.487199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:48.487199Z digest=sha256:85e4df5a72ff535c4eb7cba3e3cafacfac6a4c32952204464ef0beddcb3492e0

Observation f3e214dd-d14f-4ad1-af5b-4b898e03b140 · inbound

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning cites this paper.

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:19:26.845188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:19:26.845188Z digest=sha256:41b2fb71e64d7eef38eb34f7141cc7903d54fd3a8db18d938935e0fadb8ce4d4

Observation 44a757b4-b7af-4777-8df1-10dfab028887 · inbound

AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs cites this paper.

AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:26.935097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:08:26.935097Z digest=sha256:b172687c1cdfa2dc654e7f6b2736e54c08a03414115e809eff4c784a921b8727

Observation 15210643-f738-44b6-9b55-0efb7dff02c0 · inbound

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text cites this paper.

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text Efficient Online Data Mixing For Language Model Pre-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:42.562902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:29:42.562902Z digest=sha256:c152abb04928f90911d70a97a40ac8b051515718600c4d071dc510efaf5186f0

Observation 917c7a5c-e1c5-43cc-ba0e-1d04c9a4f981 · inbound

Language Models Improve When Pretraining Data Matches Target Tasks cites this paper.

Language Models Improve When Pretraining Data Matches Target Tasks Efficient Online Data Mixing For Language Model Pre-Training

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:53:07.823581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:53:07.823581Z digest=sha256:5e3f0fa937427a229aff100073fe6da26815529c6bfe5a9092255c3fedf44fb7

Observation 7bda05d1-b601-47aa-ac2f-10d9cc52fb0d · inbound

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining cites this paper.

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining Efficient Online Data Mixing For Language Model Pre-Training

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:08:12.635236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:07:29.544613Z digest=sha256:75e7d17e78435ff712a822033fa2847bc012e382fd6ff12fecb0c0b08a2f1129

Observation 29625642-4fda-4ab9-a4d1-8621cc8980ab · inbound

Data Mixing for Large Language Models Pretraining: A Survey and Outlook cites this paper.

Data Mixing for Large Language Models Pretraining: A Survey and Outlook Efficient Online Data Mixing For Language Model Pre-Training

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:25.751611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:56:04.958757Z digest=sha256:90f453bd45d17655454515d0892ba61b95bb6238c99814782fa66e61b99c0cf4

Observation a2866833-2709-4d6f-bfaf-a598fcbefb75 · inbound

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling cites this paper.

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling Efficient Online Data Mixing For Language Model Pre-Training

Reference 268

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T03:08:59.654681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T03:05:36.871497Z digest=sha256:5269a455f9142ce975a9edd18fc7ce36493a04c0576e4870495881454cc10d14

Observation 978961dc-3dc4-47d2-9e4f-f5450bcf27c9 · inbound

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them cites this paper.

Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them Efficient Online Data Mixing For Language Model Pre-Training

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:22:46.209942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T23:19:22.355753Z digest=sha256:9ab71dde122e9609deb180e9a681afabb9c3cff143fd6437f917bc2ce1a47acc

Observation 80073e1f-33d6-402a-9663-849359df43e1 · inbound

Explaining Data Mixing Scaling Laws cites this paper.

Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:37:22.526097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T20:20:53.599064Z digest=sha256:481c8a52f0b6b512620eb5189e2f8ecd83e3a9cb543273b1b7188901508b3ef2

Observation a3e95546-417f-496d-b6ee-c29f887664a7 · inbound

Explaining Data Mixing Scaling Laws cites this paper.

Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-15T10:54:16.434702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:54:16.434702Z digest=sha256:70afcc90961124dd94bb151e62af0c57e1c25c7d5defd67b351202b256086b4b

Observation fcbbb887-5ef7-4809-91a6-a7d7a2c0f8ee · inbound

Explaining Data Mixing Scaling Laws cites this paper.

Explaining Data Mixing Scaling Laws Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T12:13:45.639396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:13:45.639396Z digest=sha256:a39732d24517cc10d2b952b4eb026f7acc0760d0d265f0c03f58ad70b3d80340

Observation 7b1c2692-3879-48e7-8864-d1c738009978 · inbound

DRIFT: Refining Instruction Data via On-Policy Data Attribution cites this paper.

DRIFT: Refining Instruction Data via On-Policy Data Attribution Efficient Online Data Mixing For Language Model Pre-Training

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T19:18:54.754640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T01:57:05.784589Z digest=sha256:85e6931cc5668b15c32adbf8c4d5a788e67cfccfd7c1c3b1641141e0885cfdc7

Observation 2e0072df-b689-447b-a45f-cd91b1e4564a · inbound

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning cites this paper.

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:19:57.827829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T00:46:13.587037Z digest=sha256:b6c2ada8f79d395c8775241f31d7112519a40de8044707c184d372fe6424e131

Observation 4c3cf89b-205f-46bc-b4ad-a8646687d9e1 · inbound

Smooth Scaling Laws Hide Stepwise Token Learning cites this paper.

Smooth Scaling Laws Hide Stepwise Token Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:14:26.553661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T06:08:22.652557Z digest=sha256:523b261261caa43880405ccf111f9f086b68456f7fcb80a0ba41315c6d77b50c

Observation c6a2181e-ab22-42fa-94b5-ecccc8901c7d · inbound

Smooth Scaling Laws Hide Stepwise Token Learning cites this paper.

Smooth Scaling Laws Hide Stepwise Token Learning Efficient Online Data Mixing For Language Model Pre-Training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-13T07:15:03.029543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:15:03.029543Z digest=sha256:81aae619c85524feff2ff24276a4f8cb4674c2f61d9c89d1dc079115ffee20c9

Observation 64423c53-a41e-4643-9677-4e878e0b15e2 · inbound

WARP: Weight-Space Analysis for Recovering Training Data Portfolios cites this paper.

WARP: Weight-Space Analysis for Recovering Training Data Portfolios Efficient Online Data Mixing For Language Model Pre-Training

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:58:46.729349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T17:54:31.856386Z digest=sha256:b656d29fdba2fde4380800e03631e1e8830f19f5007ae8433a840417d5e2aa54