Pith. sign in

Paper Citation Record · LEDGER

Does Training on Synthetic Data Make Models Less Robust?

As of 10 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2502.07164.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.07164 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:37:59.838990Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 95ac0ba6-2893-48cf-9431-04787c94990f · outbound

This paper cites online" 'onlinestring :=.

Does Training on Synthetic Data Make Models Less Robust? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.680598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.680598Z digest=sha256:4de8ca3510e094ba5aabddb8d49fa8f995788a9c8377c7402dfdc0c5809ff6aa

Observation 254ebd04-37a4-4619-8eb4-5c98d9bd3d1f · outbound

This paper cites write newline.

Does Training on Synthetic Data Make Models Less Robust? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.687043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.687043Z digest=sha256:b00a2d290abe829cf614b4d7045901d3e26fac51009c3cd1cbf27b4f1045387e

Observation 70c0c8e6-d7eb-4e47-a555-1ed7638db7c2 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.462254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.692760Z digest=sha256:faa91a81b5e389886d1f8d6c7b8de119b31c23e99133cdd3ec1399e5bdac0ab5

Observation a5559ce3-9b09-4c79-9ed5-04163b996033 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 4

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.987145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.701941Z digest=sha256:87187d87f56266291662d6ee1c86d71f3d1fd3387bc7583fa29d328220f71a70

Observation 95a05bfa-baa6-4d72-bf40-7c25df19b0cc · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.707723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.707723Z digest=sha256:ccb8e28bce84cedc8c9ea5d3f76f734d8383d600b478f1931b7f69901e35ff27

Observation e9b7b0e7-72e5-4ae4-bf6e-42a4dffaf889 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.446916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.713056Z digest=sha256:c9135c9e87df1c5fa7b1d273210408c20bfc502b99a5a520a0bd8aee26a1831e

Observation 3f373cbd-1e71-4493-b6c8-9c001703877a · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.718155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.718155Z digest=sha256:2f37b12bb8c734389ac29cbe4ccd15e3ff35e3ab3f7cb85c4d9ec466c4bb2500

Observation 7eeebd70-4095-40dc-a0ce-1cb1f6d52a40 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 8

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.969563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.723354Z digest=sha256:ad34bb6b7b27456e8c1d495a421ce1c965598d51c4d577d066a7ae04254fe92a

Observation 1a4fcad8-8596-4f60-aa08-40902b6d85d0 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.421213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.729084Z digest=sha256:0c13017409e85b2964f0867e810687ceaf905440c3d18e826c636ee67e7cba6e

Observation b7d931b4-4745-45a2-a79e-8def6c9c585c · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.405095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.733991Z digest=sha256:514fafc08356b49874366fcd770ca73c4d4e13a92a74c69cbc3b2121b4eaaaa1

Observation 5abec63e-f68d-488f-a053-714f24163978 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.388597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.739060Z digest=sha256:f9673c122c02deb37dac52c1ac095a31ab27562c4de30f32e0a2fd59a373e6fa

Observation 19647b76-4479-4caa-b1b1-8f3f765ad942 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.371309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.744311Z digest=sha256:8e7c3d19a1ace84adc0e620919f5ae2c97ff436d5bb67638f1f91f9c04709b39

Observation 369bd7dc-46b7-4685-9dd2-8af633e9a604 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.749258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.749258Z digest=sha256:146ed0bb8282ca3002903412e45180bef639b1fbec53900598c14d44897e1bfe

Observation 44822ea2-7766-4b95-bb75-21ccb7b8c5eb · outbound

This paper cites CultureLLM: Incorporating Cultural Differences into Large Language Models.

Does Training on Synthetic Data Make Models Less Robust? CultureLLM: Incorporating Cultural Differences into Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.754298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.754298Z digest=sha256:fc933c1ae30987c4f1fc4e6c1fa474fd90ca7aee650f98e2a76895376ed188a2

Observation 54f5fef4-4b64-4d5f-a951-4d6055bb1f09 · outbound

This paper cites CulturePark: Boosting Cross-cultural Understanding in Large Language Models.

Does Training on Synthetic Data Make Models Less Robust? CulturePark: Boosting Cross-cultural Understanding in Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.759730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.759730Z digest=sha256:a99707d7cd6a52799da4aa3dad90247ffb4146c883bb6dca402822ee14a2b4fd

Observation 604d6272-db2d-4758-bb56-69927d268923 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.765768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.765768Z digest=sha256:062e277c18f022dc10619bc97d4ed6b55d044e867afedff896cd4953d154fe05

Observation ad7ff3f0-33a4-44ed-bb8d-780662f0a1a5 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.771382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.771382Z digest=sha256:18933f6e3b35b2f67ae8897d7f24a36aa091ea12a642c3d92b705e4285da9299

Observation df7427f6-664a-43d3-8b8e-cdfb169a94fa · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.931976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.777149Z digest=sha256:bc00321c8c0acf629b984660fea1efacd29ece5700d3f2d7953c5baaed28861d

Observation ec290b92-ba12-448c-96de-17a340c60c59 · outbound

This paper cites Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan.

Does Training on Synthetic Data Make Models Less Robust? Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.782521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.782521Z digest=sha256:0aa51289c4d9aae00f821ddee99b3ca2863d465f7451ec3d084db3efe7f0bfcb

Observation 5e50c65e-d5af-48cc-b8f2-bfc00c25b8d2 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 20

Resolution
verified exact
doi, observed 2026-08-08T13:37:59.914189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.787562Z digest=sha256:570d9e81a56ccb20dd00cb742c8a7f1ccbad531b9714d4ee7e957e6d0127856e

Observation 42b58834-531f-48c9-86f1-92fc76eca5f1 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.343781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.793020Z digest=sha256:3d16b28b30fd47ce6ed1dd5c2d837a6b51e14e3debea0f7e5f788b5cb84a816f

Observation ea8234df-7a61-4250-a00f-8e4fed6303bf · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.797726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.797726Z digest=sha256:b4bf0d44fe97d69670ae95b2b055c6176660ff6305b063b25ba657db889a50d1

Observation e80b1c8d-2696-4c11-b9ae-904933bd3099 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.316523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.802743Z digest=sha256:42cad2920fbda6b44b9053c77da18a3ff000d1f8dd5bc52488ce70e8ef5e142e

Observation 2363a3bc-a95b-4a78-86bb-a53fc0b955ff · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.300557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.807948Z digest=sha256:750be0a59b511e0f140abcf606f67efb886b0d95c0ab3f8e6f96dd4fd5560399

Observation 6edfa210-09d9-4a00-9195-c65bfb7820cf · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Does Training on Synthetic Data Make Models Less Robust? Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.813460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.813460Z digest=sha256:1e7745bd1600b72a9c6f2d6c7b21b28cbbfcf9a8e88cf2bb0e5aa2988748089c

Observation 0b80f60c-680a-4e6f-97d1-7fbb9c8a3ea0 · outbound

This paper cites Smith, Daniel Khashabi, and Hannaneh Hajishirzi.

Does Training on Synthetic Data Make Models Less Robust? Smith, Daniel Khashabi, and Hannaneh Hajishirzi

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.818783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.818783Z digest=sha256:8236b5a2112d12d0ba7e6d5987f31e14663d27415f535b7acb00b7fcaf6f0fe2

Observation 4c95426c-c9f6-4172-8bab-64626e90bce7 · outbound

This paper cites Simple synthetic data reduces sycophancy in large language models.

Does Training on Synthetic Data Make Models Less Robust? Simple synthetic data reduces sycophancy in large language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.823734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.823734Z digest=sha256:7bf6cdb33a9e53cfb9dc04be77ddd7bed2b6cc8e241d64af372f85aebddd35df

Observation f2830c54-a203-4718-9eb4-7b4144f65ac7 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-08T13:38:00.284384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T13:37:59.829463Z digest=sha256:84c3b317d63808e80e7cf4d31b277fdd46ad11e1ba83263faedb267b98719cd9

Observation ad44ecad-aedc-4d05-81d3-6accab76d6f8 · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.834054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.834054Z digest=sha256:839761029af63fb1c4f5fd9b30ce48c810f2934769179a2abb8ee565d53a5695

Observation fa5f3cc2-b331-4642-97ea-b5c230eb5bef · outbound

This paper cites an unresolved cited work.

Does Training on Synthetic Data Make Models Less Robust? Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:59.838990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:37:59.838990Z digest=sha256:a03c268cacd33d5386852e011a2f793e1f96a92598ec966762da65ad5348cdd5

Pith citing papers

No inbound Pith citation observations are available.