Pith. sign in

Paper Citation Record · LEDGER

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning

As of 11 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 2 inbound Pith citation observations for arXiv:2505.24844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24844 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:44.945688Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T16:44:41.720388Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:48:39.449857Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact1
  • verified fuzzy30
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b3f3af47-ab0b-4944-8a1e-0d11b5d39ba9 · outbound

This paper cites write newline.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:38.503981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:38.503981Z digest=sha256:7fd4f7ad6c6366577f6d44857ee3a59ab5a63b0b9e3b3e749b215c38092a7850

Observation f79284cf-cb42-449d-838c-3a8f2c821a35 · outbound

This paper cites Fast Randomized Kernel Methods With Statistical Guarantees.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Fast Randomized Kernel Methods With Statistical Guarantees

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:38.631002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:38.631002Z digest=sha256:4133f15f0c414f6316a094ddd6c90631e1498ceb1355fd9515f18b2d954270a2

Observation dd4f53e5-1b52-40b0-8d45-6e3e3908d583 · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:53.333515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.718840Z digest=sha256:62884661925fad9052741c04854062070a714361f20a03af390785fcd40b6c81

Observation 9b3b7a62-c4f4-4b78-987d-d7cc769af6f0 · outbound

This paper cites Sharp analysis of low-rank kernel matrix approximations.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Sharp analysis of low-rank kernel matrix approximations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:53.080460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.787906Z digest=sha256:937f0cd7704f547183e4a75ee8f373f3e6feb6a43b1e24e0952d341295086b51

Observation a56dba4a-6129-49c8-8ce3-f3eb14834600 · outbound

This paper cites B., and Stylianopoulos, N.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning B., and Stylianopoulos, N

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.906128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.902143Z digest=sha256:a87ff70aceac34f33b260077cb9ee1227d44cccd9c331fbc8eb70aca1aa5d669

Observation 2f4e55df-f392-4929-a82a-d736c511e1c1 · outbound

This paper cites Piqa: Reasoning about physical commonsense in natural language.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Piqa: Reasoning about physical commonsense in natural language

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.731218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.976247Z digest=sha256:5589b41e1a07b5ba5fc1491936fcfcd72900259d484ef12a0528e2aec4d9ec85

Observation b30ac921-47fb-4036-abd2-5361d1de8069 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.047968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.047968Z digest=sha256:5df15a76ab124699287f8362d68127fce7217f29bd1aae5591c02d5e73ffb8c2

Observation 7fbd5401-2195-4b69-839d-07d6eea4a633 · outbound

This paper cites Analysis of nystr \"o m method with sequential ridge leverage score sampling.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Analysis of nystr \"o m method with sequential ridge leverage score sampling

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.574209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.122716Z digest=sha256:51d34f887d3b02d099a2321abd2530deeb496d2fc05417004bda68aa2af05cf6

Observation 278990ec-cd76-42d6-b790-d54e321149ba · outbound

This paper cites and Yang, Y.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning and Yang, Y

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.440369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.201995Z digest=sha256:f8794e50bc40d2228d39cf360700ba839fa771e81d1038f58aa7593e5580a14c

Observation 1c6eb76e-8c0a-4689-b2b3-782fa9d66a5a · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.280395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.280395Z digest=sha256:ff2bbbe5010e8a7a957fbda3646c267e4e67bb777b60d00114457468e34c306f

Observation e086f306-5235-4db8-b0df-acbfc899db2f · outbound

This paper cites Input Sparsity Time Low-Rank Approximation via Ridge Leverage Score Sampling.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Input Sparsity Time Low-Rank Approximation via Ridge Leverage Score Sampling

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.359944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.359944Z digest=sha256:bea707671e33c1dfe03f7b0d23340d808da829e0ee7e9f2e340f64d2ae72c02a

Observation 01ba9733-8ce0-4bf4-b1e7-3877415db091 · outbound

This paper cites B., Musco, C., and Musco, C.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning B., Musco, C., and Musco, C

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.263988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.456117Z digest=sha256:7c12f482cd80fa2c02f636e8d4877a39c4996c8fee0a425240c2a456fef43b05

Observation ce4338ec-32a4-4cb6-a4d2-f1892d733728 · outbound

This paper cites Multivariate christoffel functions and hyperinterpolation.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Multivariate christoffel functions and hyperinterpolation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.122473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.536150Z digest=sha256:3d8d4cc0f5ddf818a4e7beae0b19cb35b59fcd70fee43cf9dfaadef924995123

Observation d9b443af-97ae-4350-a8ab-5fd403456981 · outbound

This paper cites M., Tong, S., Lepikhin, D., Xu, Y., Krikun, M., Zhou, Y., Yu, A.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning M., Tong, S., Lepikhin, D., Xu, Y., Krikun, M., Zhou, Y., Yu, A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:51.922321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.652667Z digest=sha256:de14c8542f913720283cebfafbe719cfbbb1557d0e33a56fe377a179ffaa24ed

Observation 380f1443-4da2-442c-86b3-82838e03bb22 · outbound

This paper cites The Llama 3 Herd of Models.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.720002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.720002Z digest=sha256:cd8a1692d6f7814ac0db265461922fa941792f2e7ceb82f9e08d66d395df6cee

Observation d13ab9fc-b31e-41a0-ac62-e48ad12035af · outbound

This paper cites Leveraging the christoffel function for outlier detection in data streams.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Leveraging the christoffel function for outlier detection in data streams

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:51.720521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.845167Z digest=sha256:c78b2c9026614d59fb28b6ff41cdb2a481fcb51d105bf566d43e16f53247b102

Observation 00054f01-6e16-473e-9eeb-0913ae1d4dcd · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:51.501689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.919943Z digest=sha256:c573cb7de9eb3127cd3c0037595a86dd2a6772abb4dfb92ea7de4cd1d177c6d2

Observation 51e9940f-84a8-4e96-885d-66991935c67f · outbound

This paper cites Dynamic Gradient Alignment for Online Data Mixing.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Dynamic Gradient Alignment for Online Data Mixing

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:35:45.364390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.028666Z digest=sha256:e89909d0443703ec4e3095aad4d478e39d55398ef290fc59009ac8d4ac58c11d

Observation b3b65340-679b-42bb-bbc7-408c87847d30 · outbound

This paper cites DOGE : Domain reweighting with generalization estimation.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning DOGE : Domain reweighting with generalization estimation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:51.161312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.166611Z digest=sha256:65570fb92dfc44b4b5565d5c00852d4150e6bc20b9982d170b549832c011c093

Observation d7eacd52-3ee0-42af-957d-3736b71854a6 · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:50.927863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.351695Z digest=sha256:c0127d505b5ebb452cd96330fa2f0afcdb83faa41fee05187a9e4fae3671f8b7

Observation ed9a85c7-fce3-48df-8edd-4c624c12522a · outbound

This paper cites Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:40.495825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:40.495825Z digest=sha256:8353d54348ee29f3ec24e3869df3113dfd506e4f62fdac55046182ee57e8379d

Observation 33e7cd2f-c33a-4fa8-9d0c-3df679a513d1 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:40.638808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:40.638808Z digest=sha256:5fd2172dcd2b9b48a3e29a1a53791f1eb7f65ed7faf4b65e70ef0157a9aaa43b

Observation b2abb185-ba50-4701-a7cc-0209d60c8d74 · outbound

This paper cites A framework for few-shot language model evaluation, 2024.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning A framework for few-shot language model evaluation, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.506652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.839268Z digest=sha256:afb4c209c7f1ae3926be1bbfe37cc552a109306ead62a1e7314d955900bb34f1

Observation 043fde3c-f671-4c8e-af53-7afc339a4493 · outbound

This paper cites An introduction to statistical learning: with applications in R.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning An introduction to statistical learning: with applications in R

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.395008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.017439Z digest=sha256:be52f3794ece3d5e636d02ddb7ba031399dc50c28222c0a976b2d1828c7672d6

Observation 59c670a5-a10c-430d-b88e-1fe7c6a0de53 · outbound

This paper cites Wiki-40b: Multilingual language model dataset.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Wiki-40b: Multilingual language model dataset

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.249819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.157937Z digest=sha256:5afea9ce987d083db49f7a9543a92faf2b7f0aba37f1e1cc99d8b3e80c963be0

Observation 0ca0147c-e911-4038-9912-ae1e484d97a8 · outbound

This paper cites The elements of statistical learning: data mining, inference and prediction.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The elements of statistical learning: data mining, inference and prediction

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.088307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.258586Z digest=sha256:8e7a26f53e83924d93434fdfa693b33333409114dfbab036011eda482d52ffbc

Observation 1fb5bd19-70eb-4a72-9e70-11a649b5f5f6 · outbound

This paper cites Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.369147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.369147Z digest=sha256:1e102f5c76b2ff9c9a35e52fcbe83a738cd32a769e4b2f00b31e4b3024cf3243

Observation 7a0a942d-38d8-4917-a791-0fcef23d878c · outbound

This paper cites Autoscale: Automatic prediction of compute-optimal data composition for training llms.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Autoscale: Automatic prediction of compute-optimal data composition for training llms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.513536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.513536Z digest=sha256:c3b4a20c0f5d67e496d8febfa3cbb23c85ae062467ac533876b8e63a2201e6dc

Observation c592ddbb-c97b-4bbe-a438-535afb3b0a9f · outbound

This paper cites Looking beyond the surface: A challenge set for reading comprehension over multiple sentences.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Looking beyond the surface: A challenge set for reading comprehension over multiple sentences

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:49.919292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.639380Z digest=sha256:978cfd12ecaceadb82a2b2d60efbe3f759513b3c602f74f5a16a6edd753ad9a2

Observation 44b966dc-a8f5-4ef8-83e2-45e76205e9dc · outbound

This paper cites The Stack: 3 TB of permissively licensed source code.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Stack: 3 TB of permissively licensed source code

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.731039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.731039Z digest=sha256:0e45473f555e90b081aff7be6f839abe108d1b082359393211763fd77971f353

Observation 0199387e-b1d3-480b-af7d-0ae72d21cafb · outbound

This paper cites RACE: Large-scale ReAding Comprehension Dataset From Examinations.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning RACE: Large-scale ReAding Comprehension Dataset From Examinations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.838461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.838461Z digest=sha256:7b3e3e02e935800bce8519bf5c6bc501841849e1c20faed44686e2566914b092

Observation a4351c0a-630a-4f6e-9047-6a3a5400b16f · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:49.703606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.957391Z digest=sha256:4e83d78530c6a4ef4a6d57f5b4466edea9b6f4cc690948086361bd46bb63f2bd

Observation 7d0b0457-dd98-405f-b895-4bffbcd3e019 · outbound

This paper cites L., and Peng, R.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning L., and Peng, R

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:49.227083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.054731Z digest=sha256:41ff6da47728e98c09413efda506042ef96050149493a0fb6e67701ff53155a3

Observation d17e5040-a200-4cb6-b809-43f28a19b253 · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.123008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.123008Z digest=sha256:f2c3ef0df4ded5879b3b9f27496ffdaa16bab9387a1f0ae0b7a65925af1a5131

Observation 7ef517de-48b5-4365-bb9b-1415edd4590a · outbound

This paper cites RegMix: Data Mixture as Regression for Language Model Pre-training.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning RegMix: Data Mixture as Regression for Language Model Pre-training

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.207728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.207728Z digest=sha256:49b2fee1dd177522c4ef2ede461126524c36912a4ef3e28886724b0c11d5fb87

Observation 8cb0cd30-5f23-4188-a319-8bc17f062cd2 · outbound

This paper cites A pretrainer`s guide to training data: Measuring the effects of data age, domain coverage, quality, & toxicity.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning A pretrainer`s guide to training data: Measuring the effects of data age, domain coverage, quality, & toxicity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:48.684891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.302526Z digest=sha256:3a0b824930d003f021b873e4652129890399484e6ac738c75a8c064eba007b82

Observation b604cae5-2271-4340-bd58-428587ad586b · outbound

This paper cites At Which Training Stage Does Code Data Help LLMs Reasoning?.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning At Which Training Stage Does Code Data Help LLMs Reasoning?

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.400029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.400029Z digest=sha256:fb3d55b593044b15b2ac0971f6f59bc38d0784a787cdf591c3265e7fad5a303e

Observation 5cc82a3e-b6f7-45b5-a4f1-65d731824aed · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:48.414496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.551788Z digest=sha256:fb424b5d7f37736769c246bcb75afb50772e88908b28a5d5ccf4daeb86999c38

Observation 82154177-26b6-4222-85f4-a8772a31ffa7 · outbound

This paper cites When not to trust language models: Investigating effectiveness of parametric and non-parametric memories.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning When not to trust language models: Investigating effectiveness of parametric and non-parametric memories

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:48.154308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.648814Z digest=sha256:24d9e541e31960a4965d0f4358a3e7ec4ac8d5b02eec63986a0defc81d67ac12

Observation f337cb77-b714-4833-a708-daef5c4f64cd · outbound

This paper cites UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.804042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.804042Z digest=sha256:23dd594efffcf4204c71a3f325aad1e886ae25b164a74637d7be1f63df430834

Observation bad7de74-0c25-49fd-b56b-08022230fa8b · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.897037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.897037Z digest=sha256:2b872efeda474cc492904645f3784d40825bbab3d30744cc159d8f411295ebb0

Observation e1959f7f-8fbd-4119-9c60-c8f1a4536c34 · outbound

This paper cites and Musco, C.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning and Musco, C

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.868679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.996312Z digest=sha256:36b798fa3ad5dfddba1dca54e0edf46e55de7f1341c54912f643ce853f5ac519

Observation 7c59eaff-d040-4bda-be13-f0055e2cfb84 · outbound

This paper cites The LAMBADA dataset: Word prediction requiring a broad discourse context.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The LAMBADA dataset: Word prediction requiring a broad discourse context

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.087650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.087650Z digest=sha256:4244070ecfd629a505f7b24d14707325942befea759ad59ab98b6e595270e241

Observation e6ac58fd-dcf0-4aa1-a52c-a87c69a632a3 · outbound

This paper cites Data, Data Everywhere: A Guide for Pretraining Dataset Construction.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Data, Data Everywhere: A Guide for Pretraining Dataset Construction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.210706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.210706Z digest=sha256:bb9d20789aa7d3e0d8c49ce8552f2816945e260bb8d2ff7fd110a2b39654dc13

Observation 244660e8-c889-40d0-b79e-f02cbc813d6a · outbound

This paper cites The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.277388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.277388Z digest=sha256:c0707bcee4346d4a8894ebc182e0461e6565cb16c95303e90c459720881585a7

Observation 8d059221-f8d1-4d27-9d31-2c07972dee6b · outbound

This paper cites Relating leverage scores and density using regularized christoffel functions.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Relating leverage scores and density using regularized christoffel functions

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.678146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.394425Z digest=sha256:b1e2b3a628ded82dcf6551afa20cac4550d6f43f21b80f4ad0388904cee9c665

Observation 91ca0a5d-bda1-4d00-b818-82f9c0eae9f6 · outbound

This paper cites Falkon: An optimal large scale kernel method.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Falkon: An optimal large scale kernel method

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.531788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.481067Z digest=sha256:5188ae54676ad8264c3eaaec5329b01f861fa4ad31012507a09b2e9b6d0b8bda

Observation 44dda75b-1c18-4e14-979b-3125ce1b0ef2 · outbound

This paper cites On fast leverage score sampling and optimal learning.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning On fast leverage score sampling and optimal learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.334377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.550515Z digest=sha256:9d6105ba29828bd8f2e491bacd721a952e87f2bb0986b42514c6598dcee9f74e

Observation 13a410b8-2423-41a9-9932-d907f680f02e · outbound

This paper cites W., Hashimoto, T.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning W., Hashimoto, T

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.166816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.661039Z digest=sha256:e1153cf10a1e3e229b29d5c03407f4773fa4144c45420792e9204fbdfb4a8ec6

Observation a6f33091-d00f-4906-a6c2-95b580a7b1cd · outbound

This paper cites L., Bhagavatula, C., and Choi, Y.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning L., Bhagavatula, C., and Choi, Y

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.982119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.759872Z digest=sha256:2d81b9a7c4a3d23207ccef56820036c2a4305257d7ab6d4ec8140e2c2ebdca63

Observation c9862cc0-9dbd-45ba-9ff8-878343a3146f · outbound

This paper cites SocialIQA: Commonsense Reasoning about Social Interactions.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning SocialIQA: Commonsense Reasoning about Social Interactions

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.845142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.845142Z digest=sha256:267d60989fd1d5debba6bc57252c6cb24fc17e6716818d0ba0c82da7dedc2cbf

Observation 252658e2-bfaf-448a-b50b-a0d16113da04 · outbound

This paper cites Superglue: Learning feature matching with graph neural networks.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Superglue: Learning feature matching with graph neural networks

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.809132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.941226Z digest=sha256:300f78c99b0ad2761b4decb757aa3a8bdeca71c6138d24784d9e472927c69e79

Observation dca80466-c8a3-42bc-9c43-91a16adddafc · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:46.672149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.077211Z digest=sha256:b2f359e32dee35e03d6f2027e46becfde37962ff8c9efc9969e6456baa2b5529

Observation 68de9851-dcfe-4231-a816-3a668ff1a055 · outbound

This paper cites SlimPajama-DC: Understanding Data Combinations for LLM Training.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning SlimPajama-DC: Understanding Data Combinations for LLM Training

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.169610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.169610Z digest=sha256:9ffc91494d3f10198f6824d83744f1881ac3e474f4a30b8bd4dffa3ba51b1207

Observation b5b8a180-72cf-44ca-996f-2adf63905bf7 · outbound

This paper cites R., Hestness, J., and Dey, N.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning R., Hestness, J., and Dey, N

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.483514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.265934Z digest=sha256:c4b073b2ec1517b9f148e2323f545718f9bffb3d85e14c34126bdabdf9dd3dc9

Observation 260c8516-a786-4aa6-9a78-0a5b57cbb685 · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:46.303556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.370685Z digest=sha256:3abf70b5c82ee276bfe1e555794d8c8e36309f4c429792d1c5ea7b2410d9cdec

Observation 2c8a2fc8-5afd-4fb1-af99-59103a2155a1 · outbound

This paper cites N., Kaiser, L.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning N., Kaiser, L

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.106489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.430256Z digest=sha256:77aa784b4b1568199fc154c960dc8e631fb066400a62623429624e035bb5beed

Observation 33580535-e89e-4bd8-80f1-339ed89cfe22 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.518896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.518896Z digest=sha256:39d8a875e16b84fc9039482cfd9ab4ebe5cce559930e417e4724ca2ec2a2401b

Observation 07c17dd7-463b-4d89-a894-286945055dbf · outbound

This paper cites Crowdsourcing Multiple Choice Science Questions.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Crowdsourcing Multiple Choice Science Questions

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.600639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.600639Z digest=sha256:835a6331f090b4d00a4ae71697849a92e91059e92375ca127d8dce55915ecd67

Observation 3feeeabe-7f0d-4ce5-ae89-4dbf3c0c4ee5 · outbound

This paper cites M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P., Le, Q.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P., Le, Q

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:45.908361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.680872Z digest=sha256:5a4b607a807c1c9354810899b704c28b48017afd0ab4208729e149c5d1864b15

Observation a3642873-dedc-4f25-85fd-309da848bf11 · outbound

This paper cites N., and Mirzasoleiman, B.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning N., and Mirzasoleiman, B

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:45.740683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.748439Z digest=sha256:803d5900db98dc59f14d466c3312755d921f7f017bf9cb9632178d395296cf49

Observation a6e03e86-85cd-4257-912e-4fb29755758a · outbound

This paper cites Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.851956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.851956Z digest=sha256:9158db64014b0be9ad75657930cc2a8659227d127e1b3867b551646a8fa46aca

Observation cffc16bb-1f22-4850-bad5-7542e0b0237c · outbound

This paper cites H ella S wag: Can a machine really finish your sentence? In Annual Meeting of the Association for Computational Linguistics (ACL), pp.\ 4791--4800, 2019.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning H ella S wag: Can a machine really finish your sentence? In Annual Meeting of the Association for Computational Linguistics (ACL), pp.\ 4791--4800, 2019

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:45.575582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.945688Z digest=sha256:af5e19c42efcfce3e2ac0bc5839464ec6554d3610068da2fc16f2bc26c22f846

Pith citing papers

Observation daa3f0a0-bd0b-41b4-9534-615953804898 · inbound

Data Mixing for Large Language Models Pretraining: A Survey and Outlook cites this paper.

Data Mixing for Large Language Models Pretraining: A Survey and Outlook Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:25.742553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T00:56:04.958757Z digest=sha256:2002f767e4ff40de7e6f9fa707ef0da5f9913163079a8b5155258c1f19030d66

Observation 0d40c740-2663-4d1c-bb09-1788df5e4c44 · inbound

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures cites this paper.

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:48:39.451436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-07-03T16:44:41.720388Z digest=sha256:d2300d1f9e13710df02153d5c78ae6aeed54792a164ecdd7f60537f8c8c2854d