Pith. sign in

Paper Citation Record · LEDGER

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning

As of 8 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 2 inbound Pith citation observations for arXiv:2505.24844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24844 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:44.945688Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T16:44:41.720388Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:48:39.449857Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact1
  • verified fuzzy30
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b3f3af47-ab0b-4944-8a1e-0d11b5d39ba9 · outbound

This paper cites write newline.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:38.503981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:38.503981Z digest=sha256:c336ebf876e1132dcacdab9cdee95913df90fe58a2ec4680bf17fe99d170ff07

Observation f79284cf-cb42-449d-838c-3a8f2c821a35 · outbound

This paper cites Fast Randomized Kernel Methods With Statistical Guarantees.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Fast Randomized Kernel Methods With Statistical Guarantees

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:38.631002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:38.631002Z digest=sha256:c2b4aab166be594413c9e509f42d28d4eba7ef0b1f053f57d4342d0a05e04cea

Observation dd4f53e5-1b52-40b0-8d45-6e3e3908d583 · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:53.333515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.718840Z digest=sha256:6dc8c4a7cb21101621403222b14a259c3564f0e7b9b6c00b33bca2d6118481b1

Observation 9b3b7a62-c4f4-4b78-987d-d7cc769af6f0 · outbound

This paper cites Sharp analysis of low-rank kernel matrix approximations.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Sharp analysis of low-rank kernel matrix approximations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:53.080460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.787906Z digest=sha256:0707db18374a7db6b8cb16b5675746b7600e31853dd9c8b41d5adaef1e8423bd

Observation a56dba4a-6129-49c8-8ce3-f3eb14834600 · outbound

This paper cites B., and Stylianopoulos, N.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning B., and Stylianopoulos, N

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.906128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.902143Z digest=sha256:f63008f7623c864cca2ce3c2f81672ea0cb8c5a522180b8c6fff1e32f03db9fb

Observation 2f4e55df-f392-4929-a82a-d736c511e1c1 · outbound

This paper cites Piqa: Reasoning about physical commonsense in natural language.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Piqa: Reasoning about physical commonsense in natural language

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.731218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:38.976247Z digest=sha256:e75f09c89a3610c8e876ef4d6671df782fefc453cdc8985909ba77dab51bae98

Observation b30ac921-47fb-4036-abd2-5361d1de8069 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.047968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.047968Z digest=sha256:3f207d7fe8049c698f8a9b43bc9b97511e1a02326dd9c7c545e2f82fc31de3f3

Observation 7fbd5401-2195-4b69-839d-07d6eea4a633 · outbound

This paper cites Analysis of nystr \"o m method with sequential ridge leverage score sampling.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Analysis of nystr \"o m method with sequential ridge leverage score sampling

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.574209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.122716Z digest=sha256:4368967581c00615ce3ae4cd3a8d2fa99a13e0006bbae76995333c6e208a81f8

Observation 278990ec-cd76-42d6-b790-d54e321149ba · outbound

This paper cites and Yang, Y.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning and Yang, Y

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.440369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.201995Z digest=sha256:3353d9e45b768cd6a604bb98a0a7ceac76b4b7c2f455f678749e1e0f4e60d030

Observation 1c6eb76e-8c0a-4689-b2b3-782fa9d66a5a · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.280395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.280395Z digest=sha256:394f399b34e797ccc84e00ced6c4f97e9ce1ecda2c8b88113b7bfee6b2d9ad60

Observation e086f306-5235-4db8-b0df-acbfc899db2f · outbound

This paper cites Input Sparsity Time Low-Rank Approximation via Ridge Leverage Score Sampling.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Input Sparsity Time Low-Rank Approximation via Ridge Leverage Score Sampling

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.359944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.359944Z digest=sha256:a1a79d72ec9b139ba05d54ae321522dadd0a4b523195c164bef70e02946a7169

Observation 01ba9733-8ce0-4bf4-b1e7-3877415db091 · outbound

This paper cites B., Musco, C., and Musco, C.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning B., Musco, C., and Musco, C

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.263988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.456117Z digest=sha256:47bd6d5211a8e6e4b1ce27692648d1495d41f5d31c2e7511da4549dd203631e6

Observation ce4338ec-32a4-4cb6-a4d2-f1892d733728 · outbound

This paper cites Multivariate christoffel functions and hyperinterpolation.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Multivariate christoffel functions and hyperinterpolation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:52.122473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.536150Z digest=sha256:4c74e9b1a67e452ffa0b656564269c2f458cf0651a5e6a6afa5b80ce9b50a1d2

Observation d9b443af-97ae-4350-a8ab-5fd403456981 · outbound

This paper cites M., Tong, S., Lepikhin, D., Xu, Y., Krikun, M., Zhou, Y., Yu, A.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning M., Tong, S., Lepikhin, D., Xu, Y., Krikun, M., Zhou, Y., Yu, A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:51.922321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.652667Z digest=sha256:b503e01add9fd6d24233ae95dce34842439db17835293952034a3cb7e12ffc68

Observation 380f1443-4da2-442c-86b3-82838e03bb22 · outbound

This paper cites The Llama 3 Herd of Models.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:39.720002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:39.720002Z digest=sha256:a17489cea19b750cb23441cc1d434e278b4b7e44b9aafff13a8e5631796deb15

Observation d13ab9fc-b31e-41a0-ac62-e48ad12035af · outbound

This paper cites Leveraging the christoffel function for outlier detection in data streams.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Leveraging the christoffel function for outlier detection in data streams

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:51.720521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.845167Z digest=sha256:45efd94645bb2b5fe4a151136d34e9da31a21b1d1c5575c71bf3eb10b1f9ca15

Observation 00054f01-6e16-473e-9eeb-0913ae1d4dcd · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:51.501689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:39.919943Z digest=sha256:857781b91a9da1d5cc53fc8fabe172ba4a37a47966fde3354ec63b001dc6d603

Observation 51e9940f-84a8-4e96-885d-66991935c67f · outbound

This paper cites Dynamic Gradient Alignment for Online Data Mixing.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Dynamic Gradient Alignment for Online Data Mixing

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:35:45.364390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.028666Z digest=sha256:958ad738fcc7fc25a0432bd64ebfb52c41ce68eef7e6af8f5fca3d3b5d41ac2f

Observation b3b65340-679b-42bb-bbc7-408c87847d30 · outbound

This paper cites DOGE : Domain reweighting with generalization estimation.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning DOGE : Domain reweighting with generalization estimation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:51.161312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.166611Z digest=sha256:8eda7d1d987ee01dcedc5e0e7a227eb5ec5f676ee9b552302c01f163be6b3a47

Observation d7eacd52-3ee0-42af-957d-3736b71854a6 · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:50.927863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.351695Z digest=sha256:a4936dc96c1148f0ef99df1bb76bba56b2f8142d96f4a41d9a4efe1633c6f118

Observation ed9a85c7-fce3-48df-8edd-4c624c12522a · outbound

This paper cites Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:40.495825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:40.495825Z digest=sha256:7c57d98a70c26623a44edd125904dde424f28e64cdcdb82aafd089b01ef46370

Observation 33e7cd2f-c33a-4fa8-9d0c-3df679a513d1 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:40.638808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:40.638808Z digest=sha256:871e16b53204071fa5dec6a32d38f300d2cc3bfaf3bca3d2252aa15974ea1de0

Observation b2abb185-ba50-4701-a7cc-0209d60c8d74 · outbound

This paper cites A framework for few-shot language model evaluation, 2024.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning A framework for few-shot language model evaluation, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.506652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:40.839268Z digest=sha256:6d85810d6cde5d13c4fe221540e0f01358b099c257fedf5b457ce98dbe5488a9

Observation 043fde3c-f671-4c8e-af53-7afc339a4493 · outbound

This paper cites An introduction to statistical learning: with applications in R.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning An introduction to statistical learning: with applications in R

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.395008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.017439Z digest=sha256:48c445d57da298760e83885dcc16d84debb728db68d559b7dcbc2f12e508afd3

Observation 59c670a5-a10c-430d-b88e-1fe7c6a0de53 · outbound

This paper cites Wiki-40b: Multilingual language model dataset.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Wiki-40b: Multilingual language model dataset

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.249819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.157937Z digest=sha256:ab2690fcc78d22abd75348957d3919f0bca4130d7ea09565b65a41affa9c5c41

Observation 0ca0147c-e911-4038-9912-ae1e484d97a8 · outbound

This paper cites The elements of statistical learning: data mining, inference and prediction.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The elements of statistical learning: data mining, inference and prediction

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:50.088307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.258586Z digest=sha256:7b9e14b426ab3364d7f72b61a4beb42cf24a3d5706f4d0960f158be22c6f5a7b

Observation 1fb5bd19-70eb-4a72-9e70-11a649b5f5f6 · outbound

This paper cites Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.369147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.369147Z digest=sha256:db03ea1c8c32a9712e522f8cb77219f2b29c97ac69a3f4d22f7d29cf34b03859

Observation 7a0a942d-38d8-4917-a791-0fcef23d878c · outbound

This paper cites Autoscale: Automatic prediction of compute-optimal data composition for training llms.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Autoscale: Automatic prediction of compute-optimal data composition for training llms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.513536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.513536Z digest=sha256:65ad37aaf8061b392c4f0d453badb1c8cd76f95db47e1c0bffff1b18d11ff1ab

Observation c592ddbb-c97b-4bbe-a438-535afb3b0a9f · outbound

This paper cites Looking beyond the surface: A challenge set for reading comprehension over multiple sentences.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Looking beyond the surface: A challenge set for reading comprehension over multiple sentences

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:49.919292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.639380Z digest=sha256:f51074bd7e1270ed5070ba5be7830b715d67f67a3060956c7eadb6fa32f9340f

Observation 44b966dc-a8f5-4ef8-83e2-45e76205e9dc · outbound

This paper cites The Stack: 3 TB of permissively licensed source code.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Stack: 3 TB of permissively licensed source code

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.731039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.731039Z digest=sha256:f2a95f85fb42930ee0421703fb2cfd74172041079a489b475a6396f93bd37524

Observation 0199387e-b1d3-480b-af7d-0ae72d21cafb · outbound

This paper cites RACE: Large-scale ReAding Comprehension Dataset From Examinations.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning RACE: Large-scale ReAding Comprehension Dataset From Examinations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.838461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.838461Z digest=sha256:c7d15ec02f5bfac741f9a109eda595dbdd1212aa3fee21026a986081c0e84e1d

Observation a4351c0a-630a-4f6e-9047-6a3a5400b16f · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:49.703606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:41.957391Z digest=sha256:57e02f3943d07edc716468fe05ba9f66830188f2f4f25d869418c5dea4db6c55

Observation 7d0b0457-dd98-405f-b895-4bffbcd3e019 · outbound

This paper cites L., and Peng, R.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning L., and Peng, R

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:49.227083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.054731Z digest=sha256:c7255b9d3689bf93c0aaae44fa70c7746d4593bcae44df71f2775866d8028db3

Observation d17e5040-a200-4cb6-b809-43f28a19b253 · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.123008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.123008Z digest=sha256:85cf47868bb0056215235afde8aa5c94eea1a8f39dfdb1c868d4367c239b7842

Observation 7ef517de-48b5-4365-bb9b-1415edd4590a · outbound

This paper cites RegMix: Data Mixture as Regression for Language Model Pre-training.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning RegMix: Data Mixture as Regression for Language Model Pre-training

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.207728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.207728Z digest=sha256:c016c4994aecd2c6765b841bb770cefed9df78489e2c06b9bda8ddf779fda0d2

Observation 8cb0cd30-5f23-4188-a319-8bc17f062cd2 · outbound

This paper cites A pretrainer`s guide to training data: Measuring the effects of data age, domain coverage, quality, & toxicity.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning A pretrainer`s guide to training data: Measuring the effects of data age, domain coverage, quality, & toxicity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:48.684891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.302526Z digest=sha256:760be1c7d485ab0e08fc7fad2488479bf51cda3c09a1b5d3ad3e763a96d59740

Observation b604cae5-2271-4340-bd58-428587ad586b · outbound

This paper cites At Which Training Stage Does Code Data Help LLMs Reasoning?.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning At Which Training Stage Does Code Data Help LLMs Reasoning?

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.400029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.400029Z digest=sha256:81dda848894010e625ba2e8c6ebb550dffa80e149bf5cd670801a93872ba26ef

Observation 5cc82a3e-b6f7-45b5-a4f1-65d731824aed · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:48.414496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.551788Z digest=sha256:1901321ce37af9e8bb17c88b6f9623c4b2c2f0c5ac2b3f7892feb1e07f1a284e

Observation 82154177-26b6-4222-85f4-a8772a31ffa7 · outbound

This paper cites When not to trust language models: Investigating effectiveness of parametric and non-parametric memories.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning When not to trust language models: Investigating effectiveness of parametric and non-parametric memories

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:48.154308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.648814Z digest=sha256:993145bb5af96b83929b6adf7bc8fc9abbd4a6c2eecd8d9374bc8e53535f9879

Observation f337cb77-b714-4833-a708-daef5c4f64cd · outbound

This paper cites UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.804042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.804042Z digest=sha256:497878f86eb6f85abe5862f72583d62f0c8bbd3eca1776f6fb0e3d1cb40cb250

Observation bad7de74-0c25-49fd-b56b-08022230fa8b · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:42.897037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:42.897037Z digest=sha256:c7b5c83cc2a49b63ca6e58adf3eacac52d11d33ae93dcd3d916d2f89bfb5dc07

Observation e1959f7f-8fbd-4119-9c60-c8f1a4536c34 · outbound

This paper cites and Musco, C.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning and Musco, C

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.868679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:42.996312Z digest=sha256:55c888cdaa424dbf888aea169458445e8d6985ef0d21392878bb840fc2b8a406

Observation 7c59eaff-d040-4bda-be13-f0055e2cfb84 · outbound

This paper cites The LAMBADA dataset: Word prediction requiring a broad discourse context.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The LAMBADA dataset: Word prediction requiring a broad discourse context

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.087650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.087650Z digest=sha256:0108e2c6e286a4a8694a4b5f2348b1cb5d49e8c8bd5e38d2083c7e57968532d7

Observation e6ac58fd-dcf0-4aa1-a52c-a87c69a632a3 · outbound

This paper cites Data, Data Everywhere: A Guide for Pretraining Dataset Construction.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Data, Data Everywhere: A Guide for Pretraining Dataset Construction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.210706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.210706Z digest=sha256:2363bb3040d5465f7fae7fbaf0b32ffe0c45ffa2fa29b309321fe4d885bed560

Observation 244660e8-c889-40d0-b79e-f02cbc813d6a · outbound

This paper cites The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.277388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.277388Z digest=sha256:8dba7919f15ff0642500bf629192d625db347763faa0dab13d74fb85d5d0ca5d

Observation 8d059221-f8d1-4d27-9d31-2c07972dee6b · outbound

This paper cites Relating leverage scores and density using regularized christoffel functions.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Relating leverage scores and density using regularized christoffel functions

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.678146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.394425Z digest=sha256:330f1f77530c9cfbae96a2e043f390ea57db4de6333b50fd7d5e0beb67d1d161

Observation 91ca0a5d-bda1-4d00-b818-82f9c0eae9f6 · outbound

This paper cites Falkon: An optimal large scale kernel method.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Falkon: An optimal large scale kernel method

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.531788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.481067Z digest=sha256:086cf02cd90e023754fe74c2546ccff5164b83eb2cad92fae803d3a1d88ba275

Observation 44dda75b-1c18-4e14-979b-3125ce1b0ef2 · outbound

This paper cites On fast leverage score sampling and optimal learning.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning On fast leverage score sampling and optimal learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.334377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.550515Z digest=sha256:7fec6cf8d471bc10762f72b5937a4e88ed2dbc2eccd06d3aa64a605976aa2310

Observation 13a410b8-2423-41a9-9932-d907f680f02e · outbound

This paper cites W., Hashimoto, T.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning W., Hashimoto, T

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:47.166816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.661039Z digest=sha256:72be74c0e177f345420229825f99c612eca93c8f062631791d44697195a822b7

Observation a6f33091-d00f-4906-a6c2-95b580a7b1cd · outbound

This paper cites L., Bhagavatula, C., and Choi, Y.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning L., Bhagavatula, C., and Choi, Y

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.982119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.759872Z digest=sha256:5a46c8196c960a0a2302cf8ca2f5216e41a46f4c574dfb7ca9943d285c44ab02

Observation c9862cc0-9dbd-45ba-9ff8-878343a3146f · outbound

This paper cites SocialIQA: Commonsense Reasoning about Social Interactions.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning SocialIQA: Commonsense Reasoning about Social Interactions

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:43.845142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:43.845142Z digest=sha256:49c96937f64f3348551cd0233976011e80ac4aa8019df2fbe8269f7efd71c99c

Observation 252658e2-bfaf-448a-b50b-a0d16113da04 · outbound

This paper cites Superglue: Learning feature matching with graph neural networks.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Superglue: Learning feature matching with graph neural networks

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.809132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:43.941226Z digest=sha256:b0b365a12f18f54cddf5b17508a289a285ee3865077b875d11693c1924d6343a

Observation dca80466-c8a3-42bc-9c43-91a16adddafc · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:46.672149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.077211Z digest=sha256:d31faa6681786edfedde4dff2900cf3303934ec8859d33a506c93affe9959ef0

Observation 68de9851-dcfe-4231-a816-3a668ff1a055 · outbound

This paper cites SlimPajama-DC: Understanding Data Combinations for LLM Training.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning SlimPajama-DC: Understanding Data Combinations for LLM Training

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.169610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.169610Z digest=sha256:cd008f16aeea6d7c1b88a343b3d9aaf6c43febc45852d5ee36e4be13864ef284

Observation b5b8a180-72cf-44ca-996f-2adf63905bf7 · outbound

This paper cites R., Hestness, J., and Dey, N.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning R., Hestness, J., and Dey, N

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.483514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.265934Z digest=sha256:77910dce466c32a02ad17e97c1e23d1539e09031b72239e61f8773f7accf5631

Observation 260c8516-a786-4aa6-9a78-0a5b57cbb685 · outbound

This paper cites an unresolved cited work.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:46.303556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.370685Z digest=sha256:0d41caa8b32cf81e7fce5e5d21813548394f96fa448b75046aa861d7467ae260

Observation 2c8a2fc8-5afd-4fb1-af99-59103a2155a1 · outbound

This paper cites N., Kaiser, L.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning N., Kaiser, L

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:46.106489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.430256Z digest=sha256:085f7291085f7de5f789b14cac824cfcede97b2ded6c56277f72110d6da77346

Observation 33580535-e89e-4bd8-80f1-339ed89cfe22 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.518896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.518896Z digest=sha256:07d6d8c9fde8975782565afb6cce9502286f59f136bbade216ada9655c77cf5f

Observation 07c17dd7-463b-4d89-a894-286945055dbf · outbound

This paper cites Crowdsourcing Multiple Choice Science Questions.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Crowdsourcing Multiple Choice Science Questions

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.600639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.600639Z digest=sha256:5627691b63493d52c6990b7a2fffa069273c57947e55f8086f65a73040fe9ec3

Observation 3feeeabe-7f0d-4ce5-ae89-4dbf3c0c4ee5 · outbound

This paper cites M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P., Le, Q.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P., Le, Q

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:45.908361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.680872Z digest=sha256:03431c0abd86f091fda7cc5188aae99b68fef3dbb1fbf3ac28778ae92b9a1b7f

Observation a3642873-dedc-4f25-85fd-309da848bf11 · outbound

This paper cites N., and Mirzasoleiman, B.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning N., and Mirzasoleiman, B

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:45.740683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.748439Z digest=sha256:18ad953e54672d876307bb0a61716370e25e77df13f138341f4d815d53944460

Observation a6e03e86-85cd-4257-912e-4fb29755758a · outbound

This paper cites Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:44.851956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:44.851956Z digest=sha256:f370cfd132ba37d9b1bb82a671729f7fce2b25392a124aa9d943035077cbeda7

Observation cffc16bb-1f22-4850-bad5-7542e0b0237c · outbound

This paper cites H ella S wag: Can a machine really finish your sentence? In Annual Meeting of the Association for Computational Linguistics (ACL), pp.\ 4791--4800, 2019.

Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning H ella S wag: Can a machine really finish your sentence? In Annual Meeting of the Association for Computational Linguistics (ACL), pp.\ 4791--4800, 2019

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:35:45.575582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:44.945688Z digest=sha256:492459f79efeac1291960e0027651140ad992a252d7e2978444fcaf1ef33fb1b

Pith citing papers

Observation daa3f0a0-bd0b-41b4-9534-615953804898 · inbound

Data Mixing for Large Language Models Pretraining: A Survey and Outlook cites this paper.

Data Mixing for Large Language Models Pretraining: A Survey and Outlook Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:58:25.742553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:56:04.958757Z digest=sha256:4b8271fdd6310a2a4b64f47c61533947cca51aefe72e99ded0f6b9b6834490b1

Observation 0d40c740-2663-4d1c-bb09-1788df5e4c44 · inbound

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures cites this paper.

HERMES: A Multi-Granularity Labeling Substrate for Pre-training Data Mixtures Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:48:39.451436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-03T16:44:41.720388Z digest=sha256:9838a7d4c57065a2e42f2030a192271af87cbb60254f5b6ef3da1d2a765bd27b