Pith. sign in

Paper Citation Record · LEDGER

Continual Pre-training of Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2302.03241.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.03241 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:31.713925Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:24:01.963863Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 921063f1-6e6e-4ce5-8f3d-19c5d5d8f569 · inbound

Optimization Hyper-parameter Laws for Large Language Models cites this paper.

Optimization Hyper-parameter Laws for Large Language Models Continual Pre-training of Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:45:48.948200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T20:45:31.427677Z digest=sha256:3e9176d2cb4576024688276108f7a788118a43b3f8dc174f4385219bc7a75eb9

Observation ad19ebae-2d5a-4a6e-bb41-d08380558b00 · inbound

Data Doping or True Intelligence? Evaluating the Transferability of Injected Knowledge in LLMs cites this paper.

Data Doping or True Intelligence? Evaluating the Transferability of Injected Knowledge in LLMs Continual Pre-training of Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:31.713925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:31.713925Z digest=sha256:33f88dbf2f493cd7c44a1bc718b2c24f1671d65f22aadcfa0b4c417c81c08504

Observation 73ddad87-78bb-4a09-a9e0-731211447055 · inbound

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL cites this paper.

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL Continual Pre-training of Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:37.342138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:37.342138Z digest=sha256:a0e06293dc523e7f9d22024268433747f90bcce2b03b0aa2fc912a62b4b95126

Observation 5474a7be-41a2-4b96-b427-2c51bf97229d · inbound

Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection cites this paper.

Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection Continual Pre-training of Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:54.880750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:54.880750Z digest=sha256:4f58e78c5ad6082e2df43c4326af904749c1d9a885a4ba9e95ddcaabc1ee1181

Observation 9249228c-1fdd-4188-b2ac-11b50a54c596 · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Continual Pre-training of Language Models

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:23.861194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:23.861194Z digest=sha256:faeeeacc4aab67a0606acb209764d1320ebd5699ea4ea73af33bd32a88d9c211

Observation c3a28d15-4050-4a44-a6d8-1de3b97291e5 · inbound

Bisecle: Binding and Separation in Continual Learning for Video Language Understanding cites this paper.

Bisecle: Binding and Separation in Continual Learning for Video Language Understanding Continual Pre-training of Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:09.647212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:09.647212Z digest=sha256:087214288e53ca8185ffb8cbe91758ca47b0c8f16cb3c0e9124c667e14584139

Observation e439bcc1-23b8-4871-a4ca-dbc8711b2e83 · inbound

AI-Assisted Fixes to Code Review Comments at Scale cites this paper.

AI-Assisted Fixes to Code Review Comments at Scale Continual Pre-training of Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:29:47.708700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:29:47.708700Z digest=sha256:f31b9a1e436e7fce6e60ab564142e99c635ae3940bc32f1818d04e416a2d95a0

Observation 72de1ada-a28b-4c56-aa12-d5c0e60c7a24 · inbound

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector cites this paper.

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector Continual Pre-training of Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:42:47.455682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T17:39:17.456350Z digest=sha256:53de85ecc12ff6e3052fa3df8fff1bf66d647b8d3bb6db626aad0d1227cad665

Observation 911fdbd9-5e8c-43f8-aa34-e0a881a701ff · inbound

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization cites this paper.

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization Continual Pre-training of Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:51:30.401187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T14:46:45.598080Z digest=sha256:a6a2b13141dbbbb8a85e61b04473655413a69a337e363ef099358104d3cb699f

Observation e65f9345-c86b-4de6-9f95-d43377b06f5c · inbound

Learning to Discover at Test Time cites this paper.

Learning to Discover at Test Time Continual Pre-training of Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:16:04.138617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T05:16:04.001700Z digest=sha256:bb83bd0e3897e4e9a6da3a484f08d6fb84900120c52d322cea440b9d4d891ac6

Observation ae411208-6855-4ce0-93bb-4518282acf59 · inbound

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis cites this paper.

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis Continual Pre-training of Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.664096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T09:34:23.475942Z digest=sha256:f60b14211bc58ccc22cdd90d6c4ab0ee7b89b771e2ac78d79956d4d536fae899

Observation 24e3ec75-1195-4eac-84a1-ad6c5708bdbb · inbound

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration cites this paper.

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration Continual Pre-training of Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:36:58.134724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:36:58.134724Z digest=sha256:15a8499723d8f3650b7c4d054bc0238414c4c3016d210f66f9ad93f122aeb66f

Observation 1a5d88d4-be8b-4722-9808-66c2265a11b0 · inbound

HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning cites this paper.

HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning Continual Pre-training of Language Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:15.049738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T11:25:01.780099Z digest=sha256:ca4b170ee9b1f682c14e8535d2ac48b56a854ad3d40f7a6d9e8e53f11ff3ac06

Observation 55306f03-a062-40a7-b64f-62c43e096524 · inbound

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm cites this paper.

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm Continual Pre-training of Language Models

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:56:25.597688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:47:54.466097Z digest=sha256:f9073e131069f8aa4a739be93049ca293e19c59d6568a985e24e950fc0adf7e1

Observation f42279f7-5a37-4751-8db8-b59e099fcdc3 · inbound

MeMo: Memory as a Model cites this paper.

MeMo: Memory as a Model Continual Pre-training of Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:19:43.393031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T03:17:23.202604Z digest=sha256:cc6a6d9216131ee410cbb38ede15dfdf78dcfa5934398fc7a255a67aaa30857e

Observation e5d8e25b-7f4b-4b5a-84c9-9a33b867ef17 · inbound

MeMo: Memory as a Model cites this paper.

MeMo: Memory as a Model Continual Pre-training of Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.612682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T08:36:31.022046Z digest=sha256:c63fba1d26907dbfd47f7389176c1350d02b0b9eb65c37a215c4abe95dfedb7f

Observation 74830d9e-524c-4ee6-909d-e17a937d0c3b · inbound

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining cites this paper.

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining Continual Pre-training of Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:49:49.898714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T07:49:13.043266Z digest=sha256:c9de084698a10c642dd2469cbb642a326e8bd4837fcac7f815d05fc490a14649

Observation a63f0fc3-a9dc-4d0f-8893-f6a0ea715761 · inbound

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay cites this paper.

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay Continual Pre-training of Language Models

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:01.965352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T23:15:45.174086Z digest=sha256:2c6bbef50c8bedc759d223e55520ca1316cc13d4e517f4ddb33b305b0dfb2d59

Observation a6e68a9f-f59c-4174-9689-4b9adf2fb501 · inbound

Scaling Point-in-Time Language Models cites this paper.

Scaling Point-in-Time Language Models Continual Pre-training of Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T15:39:37.977391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:39:37.977391Z digest=sha256:53dcd5c5fa00dda4bbd3b76539e01ca2eb2621d637014dac7119d5c31178c473

Observation 89c1bf4c-5c67-45d0-a258-d09e19262ba6 · inbound

Learning to Prepare Molecular Ground States with Transformer Models cites this paper.

Learning to Prepare Molecular Ground States with Transformer Models Continual Pre-training of Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T04:45:23.593279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:45:23.593279Z digest=sha256:aa92b19df4cb9227708b1da3121023e1d0da33d02423fd9dbd862712b578a0df