Pith. sign in

Paper Citation Record · LEDGER

Does Training on Synthetic Data Make Models Less Robust?

As of 9 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2502.07164.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.07164 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:37:59.838990Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 95ac0ba6-2893-48cf-9431-04787c94990f · outbound

This paper cites online" 'onlinestring :=.

Does Training on Synthetic Data Make Models Less Robust? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.680598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.680598Z digest=sha256:950d00ad8b0f958d248288b99118c94094ef6391d011ddb5a95e5b3e876f646e

Observation 254ebd04-37a4-4619-8eb4-5c98d9bd3d1f · outbound

This paper cites write newline.

Does Training on Synthetic Data Make Models Less Robust? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.687043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.687043Z digest=sha256:eb811bc4e2423187381756546f9a72acdd573a2c77e8d4f91c90a48393ece087

Observation 70c0c8e6-d7eb-4e47-a555-1ed7638db7c2 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.462254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.692760Z digest=sha256:a08ad7f0c51d0cd8261a21327426c1c69d3a4253c645af8370d027b41b6e2870

Observation a5559ce3-9b09-4c79-9ed5-04163b996033 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 4

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.987145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.701941Z digest=sha256:727de8ba322daef6be2e8ac0600fd36a736deb4037523db408d9d78763d6d4e4

Observation 95a05bfa-baa6-4d72-bf40-7c25df19b0cc · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.707723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.707723Z digest=sha256:b0a51e49c24a53eb9e8b9fddda3075a1092d94da5e8c198de64cc7ff7ed1ecbc

Observation e9b7b0e7-72e5-4ae4-bf6e-42a4dffaf889 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.446916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.713056Z digest=sha256:d8a6f5262b70b57d659d984f97036db663ddd802e5d835d5b7733e7148643b90

Observation 3f373cbd-1e71-4493-b6c8-9c001703877a · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.718155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.718155Z digest=sha256:1228665c092078785b48691cef5d144b57aeeaecc4df5d0a0df0d710972ea4d0

Observation 7eeebd70-4095-40dc-a0ce-1cb1f6d52a40 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 8

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.969563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.723354Z digest=sha256:961d8b5d3af55a4c91108aeec1259459610a0df2f130950e36062203becd2122

Observation 1a4fcad8-8596-4f60-aa08-40902b6d85d0 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.421213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.729084Z digest=sha256:a25626dd5624b911de94aad07316cdd1fe35bcd24c48502048208761895e40cc

Observation b7d931b4-4745-45a2-a79e-8def6c9c585c · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.405095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.733991Z digest=sha256:929402fdc67da82d04d40d892e171d931f089241635d76f6c931426a3cd48da0

Observation 5abec63e-f68d-488f-a053-714f24163978 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.388597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.739060Z digest=sha256:c5532388c158189045c263ac8d42a40005fde7d57eb467ec4a1eb3776c943c88

Observation 19647b76-4479-4caa-b1b1-8f3f765ad942 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.371309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.744311Z digest=sha256:fb320afecd3081f4f7d20ffa1df38fa979f5bfc8163ce40e87b43ee8e422c16a

Observation 369bd7dc-46b7-4685-9dd2-8af633e9a604 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.749258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.749258Z digest=sha256:988c776684ff52345752a666846e39686e9f53db2b964a839350f67ee9dcf7a9

Observation 44822ea2-7766-4b95-bb75-21ccb7b8c5eb · outbound

This paper cites CultureLLM: Incorporating Cultural Differences into Large Language Models.

Does Training on Synthetic Data Make Models Less Robust? CultureLLM: Incorporating Cultural Differences into Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.754298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.754298Z digest=sha256:b5567c164a00115142292ca79f5247e795d20bd8e02ab3b1f4c30edc06ccd424

Observation 54f5fef4-4b64-4d5f-a951-4d6055bb1f09 · outbound

This paper cites CulturePark: Boosting Cross-cultural Understanding in Large Language Models.

Does Training on Synthetic Data Make Models Less Robust? CulturePark: Boosting Cross-cultural Understanding in Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.759730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.759730Z digest=sha256:d3a3a32974fce1603d2a7754cbe82990779b75e4996fc2af2261ed67c9779b39

Observation 604d6272-db2d-4758-bb56-69927d268923 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.765768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.765768Z digest=sha256:77463daed1a2e97ca5504bb4811f461f69870e1d6320ab481f5da0e5336215cb

Observation ad7ff3f0-33a4-44ed-bb8d-780662f0a1a5 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.771382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.771382Z digest=sha256:b157208e3cdda1b5e0400a9a3fb2be3141722971315168a786515265110f5ce9

Observation df7427f6-664a-43d3-8b8e-cdfb169a94fa · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.931976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.777149Z digest=sha256:483800ae43cb7bec163f9c236e7a065259dd3e236b0c7922576172bd1353c189

Observation ec290b92-ba12-448c-96de-17a340c60c59 · outbound

This paper cites Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan.

Does Training on Synthetic Data Make Models Less Robust? Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.782521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.782521Z digest=sha256:ace964548a946a848065053b47fdf5f556be6cb1236380e89a59a24d5e2617ef

Observation 5e50c65e-d5af-48cc-b8f2-bfc00c25b8d2 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 20

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.914189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.787562Z digest=sha256:63d1db3f2d9c3b12c2c32959ea9c2f423b3fe555c6b62decb02ba4a08c22bdea

Observation 42b58834-531f-48c9-86f1-92fc76eca5f1 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.343781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.793020Z digest=sha256:9f0133ca558656e71309cba1d484ef74a14baef779fe1914277d52c42a518523

Observation ea8234df-7a61-4250-a00f-8e4fed6303bf · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.797726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.797726Z digest=sha256:0ae678eb34a571fa04e4e9e0415f5214761c5bb43f975d17d4cb583732619b3d

Observation e80b1c8d-2696-4c11-b9ae-904933bd3099 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.316523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.802743Z digest=sha256:7ee55f9df1ba8f5fb92b280723d9dd21a90006f83804af5cae069aa3151a5c60

Observation 2363a3bc-a95b-4a78-86bb-a53fc0b955ff · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.300557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.807948Z digest=sha256:9cfab36a6bf1e7fcfbabd8ee21dc14c81ab6f19fdea2a19e6c252872a783434e

Observation 6edfa210-09d9-4a00-9195-c65bfb7820cf · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Does Training on Synthetic Data Make Models Less Robust? Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.813460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.813460Z digest=sha256:b5d549baa85a9bb8d9489452261d168faa716b12d4ee9ba370f112e81c8d3663

Observation 0b80f60c-680a-4e6f-97d1-7fbb9c8a3ea0 · outbound

This paper cites Smith, Daniel Khashabi, and Hannaneh Hajishirzi.

Does Training on Synthetic Data Make Models Less Robust? Smith, Daniel Khashabi, and Hannaneh Hajishirzi

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.818783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.818783Z digest=sha256:85e4edf42ca11bfe1b29f276ce75e02d36b9deaf28b279867da3715cabe95249

Observation 4c95426c-c9f6-4172-8bab-64626e90bce7 · outbound

This paper cites Simple synthetic data reduces sycophancy in large language models.

Does Training on Synthetic Data Make Models Less Robust? Simple synthetic data reduces sycophancy in large language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.823734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.823734Z digest=sha256:07b7aea4e23341807df6acd338040600eacd2e254003affe8df11db50be095f3

Observation f2830c54-a203-4718-9eb4-7b4144f65ac7 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.284384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.829463Z digest=sha256:765c47dda10fa1a8f7052a1f34326904f5925842877688ef7b2da513a9f1f1e1

Observation ad44ecad-aedc-4d05-81d3-6accab76d6f8 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.834054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.834054Z digest=sha256:9ab7b38cbeb34009f3bf5ad067ac0819b739c439c6f6db954ffb3ec2aa09fc66

Observation fa5f3cc2-b331-4642-97ea-b5c230eb5bef · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.838990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.838990Z digest=sha256:aa95449b5cd1e74357955a5c92aa6792da33ea1ca0758bf95b1a9f699919fc0f

Pith citing papers

No inbound Pith citation observations are available.