Pith. sign in

Paper Citation Record · LEDGER

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting

As of 10 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 3 inbound Pith citation observations for arXiv:2506.09428.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09428 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:52:56.078724Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T01:59:49.794527Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T12:46:24.236751Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d9983265-7526-4dd2-a7ca-a2e51a661b51 · outbound

This paper cites Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.022809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.022809Z digest=sha256:b21f7f286cef9e8bb1cf9b32e4d709f1243c6ec6ff2d80019783c6ae2244afbd

Observation be9a094f-9d7d-4544-bad3-c095b5a69b51 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Measuring Mathematical Problem Solving With the MATH Dataset

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.012378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.012378Z digest=sha256:7b5d04853d003c0b8dcc39b5b3c296c3f4a622adaf7d1785e720a90d19bf1b36

Observation e336d970-80ff-45f5-bbbb-71b308a37085 · outbound

This paper cites Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.017505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.017505Z digest=sha256:d92c55cf18b1604de469d8b5c614f5c77a025e7347baae495060d6b876835d97

Observation f86c8349-dc2f-4ae2-8ffb-112ed76d5de9 · outbound

This paper cites Mistral 7B.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Mistral 7B

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.028628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.028628Z digest=sha256:9e9033384119267cefb7bd899e2c515c6699235a09b47227737a3e7feca3ce68

Observation 982c1fff-25c3-4e19-9947-68e556012a58 · outbound

This paper cites Sequential Reptile: Inter-Task Gradient Alignment for Multilingual Learning.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Sequential Reptile: Inter-Task Gradient Alignment for Multilingual Learning

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:52:56.269150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:52:56.034222Z digest=sha256:0cd08461f3751d703abfc4df90d73e20dd6f05f0edfb4a5c23be3e02075d1921

Observation 50ebd32b-b888-4368-9177-dd1c1602f3a5 · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.039373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.039373Z digest=sha256:a2966b21e55c3695d871b1e223aa6165ad2c7f71a35a717c4801a7d5e2d61645

Observation 9ee831b8-9e8a-432b-80bb-144a89467036 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.049249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.049249Z digest=sha256:d57dfd8e3e7a40f6e86133c58986548384c528837d38df120b2eb2183c747d95

Observation 7fe85f42-63fb-43ba-a000-27e83d22f8a1 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.054436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.054436Z digest=sha256:726874212b28d5c9d459e1844f104b504404b114136a305132ed6f9e3ee1d5d4

Observation c6c7f2a2-3a73-44ca-950d-0ce514475618 · outbound

This paper cites Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.059943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.059943Z digest=sha256:931adb438c2bb4efad4dbdd14cf01a42ada4657084e6e8e28ccb79f650c3fbd4

Observation 6e830037-89fb-456b-986d-2cec45d0eb6c · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.069336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.069336Z digest=sha256:5a7377607869be3eff48c6c9a4a3f8d1b554f836fdd4b978c52e3501cb190bd0

Observation dbc8cea8-3982-45ee-979b-2caf821b3639 · outbound

This paper cites WizardLM: Empowering large pre-trained language models to follow complex instructions.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting WizardLM: Empowering large pre-trained language models to follow complex instructions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.073990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.073990Z digest=sha256:025e2c05420595ebd38b507d18af0d0109f56c1607a70336bf58edc605ee61a3

Observation 5ac1d499-a016-4a3f-a3db-5a51a661d15e · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Instruction-Following Evaluation for Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.078724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.078724Z digest=sha256:ebb129a1e2e3290179d09c88d2f4f683982ffdf9b6832f1205716f8ef51da6a6

Observation c2a3de47-eacf-4f3a-ada4-20aefe873105 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting LLaMA: Open and Efficient Foundation Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.064629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.064629Z digest=sha256:ae7262a1760a629d399641ab752b0127423589d4877c72a33354e0b83aabb196

Observation 2022c9d3-f28a-43f0-b226-0b7aa8035969 · outbound

This paper cites Continual Learning for Natural Language Generation in Task-oriented Dialog Systems.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Continual Learning for Natural Language Generation in Task-oriented Dialog Systems

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.044301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.044301Z digest=sha256:4ebbfd96d512c3906c9ce16dd6e7885c91c0296af0e78ff190a1d8934e97ffbf

Observation fe01d1e1-f818-49c1-addc-d3bd040e3d45 · outbound

This paper cites Continual Learning for Task-oriented Dialogue System with Iterative Network Pruning, Expanding and Masking.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Continual Learning for Task-oriented Dialogue System with Iterative Network Pruning, Expanding and Masking

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.006926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.006926Z digest=sha256:51daf384c79948c48f17f9d732ca03bf1fe8a5a95c85c4d201545e308bd4a075

Observation 6adee60d-47ea-46da-9f8f-d9ad82bcceb6 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:56.001442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:56.001442Z digest=sha256:b046680368d49d4f85879a36d86c74ff538debd2c57d0dac2d7e52afb8b969af

Observation 2fda7eb7-1f61-4008-b791-ef709c12bf72 · outbound

This paper cites GenQA: Generating Millions of Instructions from a Handful of Prompts.

Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting GenQA: Generating Millions of Instructions from a Handful of Prompts

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:55.994940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:55.994940Z digest=sha256:1e8db2dea8084f7e43bcdf2c379d600ece3ce8839bc88a221c304e1704d88588

Pith citing papers

Observation b7b5b330-6b44-4911-81f9-2d2ca43417c1 · inbound

Emergent Slow Thinking in LLMs as Inverse Tree Freezing cites this paper.

Emergent Slow Thinking in LLMs as Inverse Tree Freezing Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:46:24.240283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T12:43:49.628082Z digest=sha256:72abcaac1997df3afd0ff7dd988cd859be6bf45f3299b4adb1f8a05235ecf7dc

Observation 1991c15c-c5da-44ed-bfc2-23ccbf076010 · inbound

Crafting Reversible SFT Behaviors in Large Language Models cites this paper.

Crafting Reversible SFT Behaviors in Large Language Models Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:07.320848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T12:16:16.104176Z digest=sha256:647c1b275ad3357a1b78942a346e726421fbdc64919744679371939a86c3c768

Observation 66302a0f-e1f9-48af-8a1f-3809a69c10bb · inbound

MemSFT: Mitigating Alignment Tax with an External Parametric Memory cites this paper.

MemSFT: Mitigating Alignment Tax with an External Parametric Memory Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T01:59:49.794527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:59:49.794527Z digest=sha256:40cd5984363b498abd4ee26d5f64a0b46e4d2eeaf5898c3920790400cb8d5349