Pith. sign in

Paper Citation Record · LEDGER

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation

As of 10 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 1 inbound Pith citation observation for arXiv:2506.15702.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15702 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:40:25.825552Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-16T11:55:50.897500Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T11:55:51.079688Z

Reference resolution

36 of 36 outbound references displayed

  • verified exact5
  • verified fuzzy8
  • unresolved22
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b9924e1f-e1e8-4d68-ae66-05b90dd5e44a · outbound

This paper cites Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:28.392773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:21.901340Z digest=sha256:f184e2956e8623fe1561d70e32174e38503eb4fa54d340becb168a7d55d0f094

Observation 2166bd8d-13a3-4ede-b97d-d8758d1b375e · outbound

This paper cites Fine-Tuning can Distort Pretrained Features and Underperform Out-of-Distribution.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Fine-Tuning can Distort Pretrained Features and Underperform Out-of-Distribution

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:21.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:21.984686Z digest=sha256:ef3b96caa2dd2128b60b55777c0fd5be41615ed35413d5a2697070c9a19d8190

Observation 36d6d8aa-9349-4e21-82e5-716189de53a6 · outbound

This paper cites Distill and replay for continual language learning.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Distill and replay for continual language learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:28.237234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:22.104922Z digest=sha256:1a45babdb7c6a6b7ed681bcc1bdba30cf8fccdeb2a1d6ddd0ba9ffc9b8ba293d

Observation 0136af4c-2c5a-43b4-bca9-15aa968675cf · outbound

This paper cites Scalable Language Model with Generalized Continual Learning.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Scalable Language Model with Generalized Continual Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.174315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.174315Z digest=sha256:cc77f3a34133e79698824a2eba7208c25d3772a94dcbbe8ca31184d726bd9ed7

Observation d27f66ef-674d-47ad-9826-b7b39c8ef32e · outbound

This paper cites Continual Learning of Large Language Models: A Comprehensive Survey.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Continual Learning of Large Language Models: A Comprehensive Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.228275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.228275Z digest=sha256:041c769d65250f878c7c25e4133a73db2cf4f8af71f15abfb8791b332c68ba7a

Observation ada8e617-2b84-4a42-81c6-6ece30a7ae44 · outbound

This paper cites The Llama 3 Herd of Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.273316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.273316Z digest=sha256:e392dcd614877cb0002c34fbdd3a19431fd079770eef61c465260e663c2188ac

Observation 532801a0-376d-414e-9637-7faeb31f872b · outbound

This paper cites Parameter-efficient transfer learning for nlp.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Parameter-efficient transfer learning for nlp

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.325270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.325270Z digest=sha256:702f86a13df6df5db5c3c45e9ff9bd81e2be4b0089ce4b958dbcecca6beaf58b

Observation 03c095da-ffbb-4ca9-9bf5-dac60207c9b0 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation LoRA: Low-Rank Adaptation of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.403153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.403153Z digest=sha256:0a3b9dbc0162214c56b76aad19f54d9c3c2cbec58913dfab473fce63e4c8d692

Observation 7e690718-46dd-42f6-b1a3-406a06b9cc7a · outbound

This paper cites DoRA: Weight-Decomposed Low-Rank Adaptation.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation DoRA: Weight-Decomposed Low-Rank Adaptation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.478505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.478505Z digest=sha256:fada48fe020730c0584347da1707337e0be41b7cc5fa4ddacc40129d1f0a2d8b

Observation 4dce038a-5dd7-40d0-a31c-5a324fb5d4c7 · outbound

This paper cites Mitigating the alignment tax of rlhf, 2024.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Mitigating the alignment tax of rlhf, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:28.001975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:22.563915Z digest=sha256:b49d87e152e9e3a0b145354c6320bee6224a8a8d9cedc2f873b4571b5ad39443

Observation 131f8fd6-9ce0-4c47-9148-a405da58db28 · outbound

This paper cites Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.685924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.685924Z digest=sha256:b9a665cd5f3eaaec10f062d4fa4152ecb2fa4eac65b87d110bdc56aad98b89a8

Observation e32bdeb5-a6b0-4a28-9b14-b1b1c7ddcb8a · outbound

This paper cites Evaluating Language Model Finetuning Techniques for Low-resource Languages.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Evaluating Language Model Finetuning Techniques for Low-resource Languages

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:27.015634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:22.799461Z digest=sha256:eaedeeb9b59be0ac1721ea819935265a45b15331256afdea25da67447fd3ac89

Observation 56542d17-5577-4ebc-bd31-b1ba87260097 · outbound

This paper cites Fine-tuning and Utilization Methods of Domain-specific LLMs.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Fine-tuning and Utilization Methods of Domain-specific LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:22.930321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:22.930321Z digest=sha256:c452bd27f7a12df8e63a7ac19c9a5cfdda8c238f2ab37b2556216d83824c1326

Observation 534cfb67-86f9-4389-8742-ee595b0b2694 · outbound

This paper cites Harnessing pre-trained neural networks with rules for formality style transfer.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Harnessing pre-trained neural networks with rules for formality style transfer

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:27.784303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:23.069557Z digest=sha256:5778cfcdad99502013111d6e9da3fdf3f04275bc03048c214809df7a72142297

Observation b23a30d4-143b-4992-bb28-64b1b23c7ccf · outbound

This paper cites Zero: Memory optimizations toward training trillion parameter models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Zero: Memory optimizations toward training trillion parameter models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.184609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.184609Z digest=sha256:cacad29e0c24d40dc408d62caa0981d8e90edddea5965f5d161f008d7b1beb75

Observation 8d64e5a8-a154-4bab-af04-d1c6479caf9d · outbound

This paper cites OpenELM: An Efficient Language Model Family with Open Training and Inference Framework.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation OpenELM: An Efficient Language Model Family with Open Training and Inference Framework

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.283185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.283185Z digest=sha256:99c27b69cfffaedad482e26bc2d890bd0b48dc1471c6d201d5c36313b6af61ea

Observation 9bb6b675-a755-4d53-8141-790b5eed44ef · outbound

This paper cites GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow, March 2021.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow, March 2021

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:27.545350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:23.356559Z digest=sha256:6317ff5bd94270133b7def452943f52c1e7836a9f6e48692c4b807600b62efb6

Observation 11a82b97-dbbc-4169-b6ec-fca206eccbb0 · outbound

This paper cites Textbooks Are All You Need.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Textbooks Are All You Need

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.446594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.446594Z digest=sha256:3830fb45346d251476165490c079e0cfd65df0f95a5ab64de57e9ba759bb984e

Observation 21135899-76bd-4f7a-8647-ee4c8f05deb1 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.571712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.571712Z digest=sha256:c971e3d43b9ea90f10f3cf2ed18de033618107a098955b055ed64184c6678c36

Observation fe04fba2-d9cc-488d-b5e8-d4df9cfc286d · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Gemma: Open Models Based on Gemini Research and Technology

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.706659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.706659Z digest=sha256:e649eb2af9d42a3e293307122a73accf6696a13b918f9c48ec154899833b611e

Observation 7e664326-fcfb-4341-93bd-a253106a72d5 · outbound

This paper cites Compact Language Models via Pruning and Knowledge Distillation.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Compact Language Models via Pruning and Knowledge Distillation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.840008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.840008Z digest=sha256:2a650112767a60a6806c0594572e4842287f8cc26daf82add3abec871207f1e6

Observation 6cdce6dd-8b70-4fad-a826-1650997c087c · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:23.979348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:23.979348Z digest=sha256:b7521590856afef5c01bcfdd923a199ff92fddd59edf23b5ef138539a025a82f

Observation cc2d33b2-30b3-4a3d-9087-042ebf2fee62 · outbound

This paper cites Pmc open access subset, 2024.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Pmc open access subset, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:27.471214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:24.123062Z digest=sha256:1a71427fc6d602364540020819d4b2d017cfd46329bd5d71d064ff9e637677a5

Observation 524998d2-3a98-4b65-99c1-2dd8583bbe0a · outbound

This paper cites Pile of law: Learning responsible data filtering from the law and a 256gb open-source legal dataset.Advances in Neural Information Processing Systems , 35:29217–29234, 2022.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Pile of law: Learning responsible data filtering from the law and a 256gb open-source legal dataset.Advances in Neural Information Processing Systems , 35:29217–29234, 2022

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:27.459928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:24.342357Z digest=sha256:1ba9ef745d09f93bdd7376843e745239cedef1725c633a9fb7f346772d8e0406

Observation d09cf018-66b7-4013-89e9-aaaf55cc60bd · outbound

This paper cites OpenWebMath: An Open Dataset of High-Quality Mathematical Web Text.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation OpenWebMath: An Open Dataset of High-Quality Mathematical Web Text

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:24.492611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:24.492611Z digest=sha256:11aaa5713a6cc3308df92eef4f9e8b0fadd0fc48820b1832d32054870d65a87e

Observation a18c343b-6f9f-479d-bd70-10bd5032dfee · outbound

This paper cites Openwebtext corpus.http://Skylion007.github.io/OpenWebTextCorpus, 2019.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Openwebtext corpus.http://Skylion007.github.io/OpenWebTextCorpus, 2019

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:24.639960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:24.639960Z digest=sha256:55f7dcce532854cba5dd7181003af25af6c7263ae24fbab993b7d7920425b3b4

Observation 4eede3ac-a845-4b0c-9f0b-a20b36f93a73 · outbound

This paper cites Efficient Hierarchical Domain Adaptation for Pretrained Language Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Efficient Hierarchical Domain Adaptation for Pretrained Language Models

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:26.744067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:24.757051Z digest=sha256:eb71d4f023c5a728a286dca367467743947a3361aad5386e37abcba5c019f3e0

Observation f1304c0c-a8c7-4e9b-84bc-03d556617291 · outbound

This paper cites Unsupervised Domain Adaptation of a Pretrained Cross-Lingual Language Model.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Unsupervised Domain Adaptation of a Pretrained Cross-Lingual Language Model

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:26.595401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:24.874023Z digest=sha256:5934db0c218349cd438a510e7126369ba8f8539c598a24944c7b08cadea4845a

Observation 377d1690-53c8-4f96-bc5a-5f4d6375b1cd · outbound

This paper cites Effective Unsupervised Domain Adaptation with Adversarially Trained Language Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Effective Unsupervised Domain Adaptation with Adversarially Trained Language Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:26.349803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:24.989842Z digest=sha256:9003810f3aa76f27910625dee244304da244240d291af081e04641735e73738a

Observation f77d5e94-3af4-4021-ab8b-ad69c2b76e30 · outbound

This paper cites Taming pre-trained language models with n-gram representations for low-resource domain adaptation.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Taming pre-trained language models with n-gram representations for low-resource domain adaptation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:27.380758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:25.065622Z digest=sha256:58f6e5e0bf02a84403a9f48252ce4b2c1ca3e2aaf1310cf5c9d52180e6a36ef9

Observation c681321b-25d0-4ae3-aa35-cdf10623bde8 · outbound

This paper cites $k$NN-Adapter: Efficient Domain Adaptation for Black-Box Language Models.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation $k$NN-Adapter: Efficient Domain Adaptation for Black-Box Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:25.154848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:25.154848Z digest=sha256:b87679162a15e76c24de0ce91da22fd2fcf18f72c3e92a1b2d64eca1d3068399

Observation a36c054a-8148-4258-b397-d8431c44bc66 · outbound

This paper cites Maybe Only 0.5% Data is Needed: A Preliminary Exploration of Low Training Data Instruction Tuning.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Maybe Only 0.5% Data is Needed: A Preliminary Exploration of Low Training Data Instruction Tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:25.256358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:25.256358Z digest=sha256:bfc146a3cc4d2036cafc2c767bd89e95370f991afa9de989fa2b2101064ff8d3

Observation 85ad0bf9-ec45-4b43-851f-dc2ee4ff90f2 · outbound

This paper cites Unlocking Parameter-Efficient Fine-Tuning for Low-Resource Language Translation.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Unlocking Parameter-Efficient Fine-Tuning for Low-Resource Language Translation

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:26.067733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:25.450479Z digest=sha256:3761c42afe2db97ea62702b031810b7019d5fbc87922f71c14d5467d0e3e968f

Observation 0caa0043-5f84-482e-b6ae-cc4d4164c229 · outbound

This paper cites When Scaling Meets LLM Finetuning: The Effect of Data, Model and Finetuning Method.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation When Scaling Meets LLM Finetuning: The Effect of Data, Model and Finetuning Method

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:25.565467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:25.565467Z digest=sha256:235682b4bba8660a18849450e96b2e7c882b6722d0d0aefd7eeeab0cfad48eac

Observation 1b8f02e8-aba9-43d0-b476-ed70cce6d159 · outbound

This paper cites Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:25.692005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:25.692005Z digest=sha256:09a9ddb4bc5f3e323888a177db01748b6dfd4db47a04579156e105804d5edc10

Observation 2280a8ac-60dc-470b-82a0-579ba4d0cba4 · outbound

This paper cites Overcoming catastrophic forgetting in neural networks.

Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation Overcoming catastrophic forgetting in neural networks

Reference 36

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:40:27.239667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:25.825552Z digest=sha256:822f502a13bcf0e16e89c6012aba71edf6d164be36987fc15ba17288a6eb581d

Pith citing papers

Observation 7bfa0aa3-81b3-4c95-96b7-4eddc78029c9 · inbound

Small Language Models are the Future of Agentic AI cites this paper.

Small Language Models are the Future of Agentic AI Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:55:51.081789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T11:55:50.897500Z digest=sha256:dd03e5d3e05ae30f2e388875dd3160a7a1e4dbccc5d2726036d6201ddab0a25a