Pith. sign in

Paper Citation Record · LEDGER

Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2409.14781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.14781 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:18:49.161581Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:57:23.314837Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9e7b95e0-4ba4-4d80-8a4b-3bf523483d1a · inbound

Self-Reflective Planning with Knowledge Graphs: Enhancing LLM Reasoning Reliability for Question Answering cites this paper.

Self-Reflective Planning with Knowledge Graphs: Enhancing LLM Reasoning Reliability for Question Answering Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:49.161581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:49.161581Z digest=sha256:13a6e94526b6c0d461f200fe318345cab24da64922961ecb0a08ac7599e43e68

Observation 637cefb0-2249-407d-885f-c228f6cf9551 · inbound

Identifying Pre-training Data in LLMs: A Neuron Activation-Based Detection Framework cites this paper.

Identifying Pre-training Data in LLMs: A Neuron Activation-Based Detection Framework Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T15:15:21.454509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:15:21.454509Z digest=sha256:d8d8da11294c7eb2b20bf5bddf64874fc58869016d3f4368a14286547bc011d8

Observation 7fc4b8d2-cf88-4f0a-b520-a16a19c6c187 · inbound

Investigating Training Data Detection in AI Coders cites this paper.

Investigating Training Data Detection in AI Coders Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T14:51:53.845119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:51:53.845119Z digest=sha256:00a36e969c76beec37f0f911783c70bf1ae21034dfa92aab05615a590a6eab50

Observation fe8683d4-a5ee-4a1b-8324-728973be40e7 · inbound

Filling the Gaps: Selective Knowledge Augmentation for LLM Recommenders cites this paper.

Filling the Gaps: Selective Knowledge Augmentation for LLM Recommenders Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:31:01.654876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:36:36.680191Z digest=sha256:4ff7ad4f07eca81ec4332ac42b70bbe79bc879a9d02fdd220e756663a5af42fd

Observation c8181a8a-0d7b-4f30-96d3-61fca2a7470a · inbound

MC-PDD: Masked Corpus-Level Pretraining Data Detection for Black-Box Large Language Models cites this paper.

MC-PDD: Masked Corpus-Level Pretraining Data Detection for Black-Box Large Language Models Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:57:23.317104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T20:02:50.169589Z digest=sha256:8a2936120fd9cffa7f78a2be3f627b610d9757905605fafc4f50f4c4d3421f52