Pith. sign in

Paper Citation Record · LEDGER

Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2306.15895.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.15895 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:57:02.133770Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T19:05:00.314725Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eb7f56c0-61a8-4802-90b7-ae3206949ad3 · inbound

Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models cites this paper.

Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 227

Resolution
unresolved
no resolver link, observed 2026-08-11T22:57:02.133770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:57:02.133770Z digest=sha256:08097edae666edbaeabb01e259b9d03fdea94c22488e268f24a406a004942e1a

Observation 3c0f3e25-9dc7-4d59-98d4-6166554dcfc8 · inbound

Few-shot LLM Synthetic Data with Distribution Matching cites this paper.

Few-shot LLM Synthetic Data with Distribution Matching Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T17:20:36.612382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:20:36.612382Z digest=sha256:c58db2df19d26220e86d336bc3b9386b625852c9bef10184fd13729329ef87d7

Observation 59e5017b-f635-496d-8b23-47ff0b60d0d2 · inbound

Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning cites this paper.

Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:04.652198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:04.652198Z digest=sha256:9f0328e9dcbd629b52cec0d7e46379fd75c9d35ed870eb03ffe7faa2defd35ca

Observation 06e23b44-3803-4a57-b8d2-4579854aba6c · inbound

Agents of Diffusion: Enhancing Diffusion Language Models with Multi-Agent Reinforcement Learning for Structured Data Generation (Extended Version) cites this paper.

Agents of Diffusion: Enhancing Diffusion Language Models with Multi-Agent Reinforcement Learning for Structured Data Generation (Extended Version) Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T11:14:51.791406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:14:51.791406Z digest=sha256:d3e54865db5472f47c647736eaf5b2388f9d2cfa412a4bed5cee2b1cd89c16f9

Observation cc7d854a-6ba8-42a4-a3e1-239a1a35f62a · inbound

Towards Human-Level Book-Writing Capability cites this paper.

Towards Human-Level Book-Writing Capability Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:43:27.037322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T15:40:53.245419Z digest=sha256:953516c8837bd54699a7197d97d89f54665dd4b5bf32c874d032797fcdd07f60

Observation da8338da-fec3-4ce2-818c-24527a6582ab · inbound

Towards Human-Level Book-Writing Capability cites this paper.

Towards Human-Level Book-Writing Capability Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:05:00.316225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T19:03:27.257098Z digest=sha256:d25638c5dcd09cf6837b4d1807fcee7b436c3ee8c63847511e10b8c41092185c

Observation 285c1a1a-cc31-4a3c-af45-ac500176e3e6 · inbound

When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges cites this paper.

When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:13:59.405266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T21:13:44.154803Z digest=sha256:585fa0f3942ae08828aa9abfc88d2158f52a351c3141b220d25a8101a37008e2

Observation 0cfb7a20-c00e-4289-862f-39d11127aaeb · inbound

Make LLM Learn to Synthesize from Streaming Experiences through Feedback cites this paper.

Make LLM Learn to Synthesize from Streaming Experiences through Feedback Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.610684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-29T07:40:09.458698Z digest=sha256:b9ee5a411da5ee9ed6b522109219b08859d29310433af4a3b7af175719962711