Pith. sign in

Paper Citation Record · LEDGER

Continual Pre-training of Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2302.03241.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.03241 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:31.713925Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:24:01.963863Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 921063f1-6e6e-4ce5-8f3d-19c5d5d8f569 · inbound

Optimization Hyper-parameter Laws for Large Language Models cites this paper.

Optimization Hyper-parameter Laws for Large Language Models Continual Pre-training of Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:45:48.948200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T20:45:31.427677Z digest=sha256:fb262e5656d008d90ae048ee0124745a835da078c3a8a3f522efe8d73c96af93

Observation ad19ebae-2d5a-4a6e-bb41-d08380558b00 · inbound

Data Doping or True Intelligence? Evaluating the Transferability of Injected Knowledge in LLMs cites this paper.

Data Doping or True Intelligence? Evaluating the Transferability of Injected Knowledge in LLMs Continual Pre-training of Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:31.713925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:31.713925Z digest=sha256:d4176422451b4469b5f1c112db49c96db83d1e339722db30aee6fc13cb440bf2

Observation 73ddad87-78bb-4a09-a9e0-731211447055 · inbound

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL cites this paper.

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL Continual Pre-training of Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:37.342138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:37.342138Z digest=sha256:a0e06293dc523e7f9d22024268433747f90bcce2b03b0aa2fc912a62b4b95126

Observation 5474a7be-41a2-4b96-b427-2c51bf97229d · inbound

Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection cites this paper.

Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection Continual Pre-training of Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:54.880750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:54.880750Z digest=sha256:4f58e78c5ad6082e2df43c4326af904749c1d9a885a4ba9e95ddcaabc1ee1181

Observation 9249228c-1fdd-4188-b2ac-11b50a54c596 · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Continual Pre-training of Language Models

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:23.861194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:23.861194Z digest=sha256:faeeeacc4aab67a0606acb209764d1320ebd5699ea4ea73af33bd32a88d9c211

Observation c3a28d15-4050-4a44-a6d8-1de3b97291e5 · inbound

Bisecle: Binding and Separation in Continual Learning for Video Language Understanding cites this paper.

Bisecle: Binding and Separation in Continual Learning for Video Language Understanding Continual Pre-training of Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:09.647212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:09.647212Z digest=sha256:087214288e53ca8185ffb8cbe91758ca47b0c8f16cb3c0e9124c667e14584139

Observation e439bcc1-23b8-4871-a4ca-dbc8711b2e83 · inbound

AI-Assisted Fixes to Code Review Comments at Scale cites this paper.

AI-Assisted Fixes to Code Review Comments at Scale Continual Pre-training of Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:29:47.708700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:29:47.708700Z digest=sha256:471f1611c8ffc9193c09baacc5a2e717bcaa03aea7be4232a4d32783d2c420ba

Observation 72de1ada-a28b-4c56-aa12-d5c0e60c7a24 · inbound

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector cites this paper.

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector Continual Pre-training of Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:42:47.455682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T17:39:17.456350Z digest=sha256:746e38ddce4323f692c0387b7713d81f2a209f91b25794ca69def557569bd1c9

Observation 911fdbd9-5e8c-43f8-aa34-e0a881a701ff · inbound

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization cites this paper.

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization Continual Pre-training of Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:51:30.401187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T14:46:45.598080Z digest=sha256:51b2cc1d512b9df689005324b51c565bcdac769e139295cf6182471a535f1bcb

Observation e65f9345-c86b-4de6-9f95-d43377b06f5c · inbound

Learning to Discover at Test Time cites this paper.

Learning to Discover at Test Time Continual Pre-training of Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:16:04.138617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T05:16:04.001700Z digest=sha256:58763a64af9e0bb1230d5c5fdf022e49cf75821c1b4f074801f1b5475bdbe905

Observation ae411208-6855-4ce0-93bb-4518282acf59 · inbound

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis cites this paper.

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis Continual Pre-training of Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.664096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T09:34:23.475942Z digest=sha256:7f07321d78ca4dd96de8f3b4cb85021c58c6598b3c13fc54c37cda79ef7a9567

Observation 24e3ec75-1195-4eac-84a1-ad6c5708bdbb · inbound

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration cites this paper.

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration Continual Pre-training of Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:36:58.134724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:36:58.134724Z digest=sha256:15a8499723d8f3650b7c4d054bc0238414c4c3016d210f66f9ad93f122aeb66f

Observation 1a5d88d4-be8b-4722-9808-66c2265a11b0 · inbound

HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning cites this paper.

HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning Continual Pre-training of Language Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:15.049738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T11:25:01.780099Z digest=sha256:42a41bc1049946c20a4d98683166e2fe5a48b639d57285b13d17899c688fa65b

Observation 55306f03-a062-40a7-b64f-62c43e096524 · inbound

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm cites this paper.

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm Continual Pre-training of Language Models

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:56:25.597688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:47:54.466097Z digest=sha256:43dc93479cddc905c4064dbcac146ea7689fde23a6a742e8e5a335899a3460c7

Observation f42279f7-5a37-4751-8db8-b59e099fcdc3 · inbound

MeMo: Memory as a Model cites this paper.

MeMo: Memory as a Model Continual Pre-training of Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:19:43.393031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T03:17:23.202604Z digest=sha256:de6b2901d360c876b5d5f4efbe32cc233b2a9507eee78eb956b710db109cb131

Observation e5d8e25b-7f4b-4b5a-84c9-9a33b867ef17 · inbound

MeMo: Memory as a Model cites this paper.

MeMo: Memory as a Model Continual Pre-training of Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.612682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T08:36:31.022046Z digest=sha256:9c787dc85c0fe4593f589ea71eefc2e00a5c083fe72f24d72e93521f97361954

Observation 74830d9e-524c-4ee6-909d-e17a937d0c3b · inbound

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining cites this paper.

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining Continual Pre-training of Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:49:49.898714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T07:49:13.043266Z digest=sha256:fb0e9ef0331bb9ffdb77bc539cd1fd05884f57fcd38d67505f0ddceb02ad32c7

Observation a63f0fc3-a9dc-4d0f-8893-f6a0ea715761 · inbound

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay cites this paper.

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay Continual Pre-training of Language Models

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:01.965352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T23:15:45.174086Z digest=sha256:d72a10e563915349a58f33f5415f1f7a0b690cdb6dda44df964810d2f2194401

Observation a6e68a9f-f59c-4174-9689-4b9adf2fb501 · inbound

Scaling Point-in-Time Language Models cites this paper.

Scaling Point-in-Time Language Models Continual Pre-training of Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T15:39:37.977391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:39:37.977391Z digest=sha256:53dcd5c5fa00dda4bbd3b76539e01ca2eb2621d637014dac7119d5c31178c473

Observation 89c1bf4c-5c67-45d0-a258-d09e19262ba6 · inbound

Learning to Prepare Molecular Ground States with Transformer Models cites this paper.

Learning to Prepare Molecular Ground States with Transformer Models Continual Pre-training of Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T04:45:23.593279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:45:23.593279Z digest=sha256:aa92b19df4cb9227708b1da3121023e1d0da33d02423fd9dbd862712b578a0df