Pith. sign in

Paper Citation Record · LEDGER

Adaptive Supervised Anchoring for On-Policy Self-Distillation

As of 20 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2608.07935.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07935 v2

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:30:59.113734Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eb3f397f-f4ce-4197-9582-bac4b4577755 · outbound

This paper cites Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:58.977048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:58.977048Z digest=sha256:d0bf4255709ed1a093930eeca211b8f332a8f661bcd7181d5b73db3a13208944

Observation a4e7a171-046c-4ffa-8b7e-1d1ffe11b2e4 · outbound

This paper cites GPT-4 Technical Report.

Adaptive Supervised Anchoring for On-Policy Self-Distillation GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:58.983083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:58.983083Z digest=sha256:ac8d73283a121d3e6821b662a68a0d57ad81339abd9058091e046c7fcba53761

Observation eedeb8b6-9be7-4c0a-8980-e152a72fdd7b · outbound

This paper cites The Llama 3 Herd of Models.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:58.987144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:58.987144Z digest=sha256:7ae05507c5f2d7b13c6113372da52225ae486617facfdde13b842f649c25efa7

Observation c8e09a10-3630-47eb-8bc2-9e3b8dfa6624 · outbound

This paper cites Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:58.990466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:58.990466Z digest=sha256:d9aa9f5d23ae352d0b161b6d20a319e0ba099d304364b59d79a7a9f4e8084fe8

Observation 9abe9b71-a7ee-4838-83fa-6d3f5db86210 · outbound

This paper cites Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:58.993841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:58.993841Z digest=sha256:7f17923a26bb24a211d1e19090bb5bd60d67cd0a428841c8699fe0effbd341f2

Observation 956f6187-967a-4791-a592-97e3edf3f998 · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2023 , pages =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Findings of the Association for Computational Linguistics: ACL 2023 , pages =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.529142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:58.997553Z digest=sha256:c6469f754fa715df9f448dfe45da9dcad99c8fd6207906e990594eeebefb6f05

Observation 6184d41b-a8a8-4b47-9dea-6f2c3a28a5bc · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Distilling the Knowledge in a Neural Network

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.000910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.000910Z digest=sha256:149f0e63d62b7cc7ad7a7f5f53699cc0adee71ce6e23fdaa614f6116a54526df

Observation a55037bb-4c38-459a-9231-00a625671a58 · outbound

This paper cites International Journal of Computer Vision , volume =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation International Journal of Computer Vision , volume =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.517092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.004494Z digest=sha256:1dea0f152c88fb7e41c85758b4200fb6199cc59b2976019e27427fc6a3539e6c

Observation d95755b0-12bb-48e1-b83b-8cc663c3d1d1 · outbound

This paper cites The Twelfth International Conference on Learning Representations (ICLR) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Twelfth International Conference on Learning Representations (ICLR) , year =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.505303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.007682Z digest=sha256:0da6d486cb4cc6490ccc41bbe1e9d63f982bb11d334bab70e22f4ef5a3513a2d

Observation 1fb4efeb-2659-416d-9ad6-632848aa3bcc · outbound

This paper cites The Twelfth International Conference on Learning Representations (ICLR) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Twelfth International Conference on Learning Representations (ICLR) , year =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.494010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.011561Z digest=sha256:c5117d1614a9029d48b5f700aac4659a0aac9632a73634eda9025edf56be0492

Observation f4027dd6-8e3b-44f6-8fad-000e178ce8b5 · outbound

This paper cites Self-Distillation Enables Continual Learning.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Self-Distillation Enables Continual Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.015436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.015436Z digest=sha256:cbb20afc0400f73393211e60ccb1966b7e157ca000ddefeaf206789c2abb1add

Observation 297bf072-cc3d-4b96-a893-4347cdd737a3 · outbound

This paper cites OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning.

Adaptive Supervised Anchoring for On-Policy Self-Distillation OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.018986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.018986Z digest=sha256:f1afb7ef75da4c196b6a4a2bc615881d1397dabd557b8b3494e33e1a536f9769

Observation 83a821ad-fa66-4253-8098-709a3f82df36 · outbound

This paper cites The Forty-second International Conference on Machine Learning (ICML) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Forty-second International Conference on Machine Learning (ICML) , year =

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.483258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.022798Z digest=sha256:c5cebb29aeab21da38314cf08aa0a69393ee64306423544353ecbf9551507f72

Observation 803542d2-d62d-4886-a17a-8802722180e5 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , volume =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Advances in Neural Information Processing Systems (NeurIPS) , volume =

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.472256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.026426Z digest=sha256:9ebe3aa882a9e679d52e0af6a72e4f02f6a3d8beb3a6988d5ac712c33cfae005

Observation 71671b2e-43c4-4929-966c-f8484c4662bf · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Adaptive Supervised Anchoring for On-Policy Self-Distillation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.029979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.029979Z digest=sha256:5f918ae246cb1e6a22186903e62d273b7aa44d00ddd96008fbcf0087892be078

Observation 2fad940e-17a3-45fb-bbe3-ea2bd141b3aa · outbound

This paper cites The Forty-first International Conference on Machine Learning (ICML) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Forty-first International Conference on Machine Learning (ICML) , year =

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.461038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.033615Z digest=sha256:c656902e3e6c62ccb813b13a6aec297c61c20fc5b5b5973fff7ae02d73b81ee1

Observation 5c017229-56de-4653-94b6-953174184778 · outbound

This paper cites Rethinking K ullback- L eibler Divergence in Knowledge Distillation for Large Language Models.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Rethinking K ullback- L eibler Divergence in Knowledge Distillation for Large Language Models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.449719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.037234Z digest=sha256:cecd19e805100ad24a2de0118abb5aca4534429869275834aa06f86a384a535d

Observation d4ab564d-b91b-404a-b529-91f9e9aa9181 · outbound

This paper cites T o D i: Token-wise Distillation via Fine-Grained Divergence Control.

Adaptive Supervised Anchoring for On-Policy Self-Distillation T o D i: Token-wise Distillation via Fine-Grained Divergence Control

Reference 18

Resolution
verified exact
doi, observed 2026-08-15T14:30:59.149627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.041171Z digest=sha256:2b21c797438cc6bb2b7bd8a1a3d57177ab568b2d7c404642e39b000962c54a72

Observation 99a2bd73-9554-4472-8c58-5d90eceae9d4 · outbound

This paper cites Proceedings of the National Academy of Sciences , volume =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Proceedings of the National Academy of Sciences , volume =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.045368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.045368Z digest=sha256:c08eafee0bf36a628f31d8a1e3bdc46041001ab3b739f51aafc5d8c6e6ad0a0e

Observation 98e74a92-96df-48c8-a7e9-d85301fb13db · outbound

This paper cites An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning.

Adaptive Supervised Anchoring for On-Policy Self-Distillation An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.049773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.049773Z digest=sha256:3e491c00c1d94f409c0a33b42473813ffb29bf552b9fa3eee33302c43450af99

Observation be367df7-1f78-4904-9ed1-248876fa1083 · outbound

This paper cites The Fourteenth International Conference on Learning Representations (ICLR) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Fourteenth International Conference on Learning Representations (ICLR) , year =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.430401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.054139Z digest=sha256:ede8ed8d4869b7f93a9d31b5617d5a4dfcdb0bc43e2de8aba2b1a7426d136356

Observation 14172cd3-6456-4450-bc4c-84de55663785 · outbound

This paper cites ACM Computing Surveys , volume =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation ACM Computing Surveys , volume =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.420155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.057469Z digest=sha256:0aca56e451a8c9c8bb76422beb3bba96308cc16d530e4430d5665e7ad94a398c

Observation 6554449b-7dd4-47c6-990d-6e071b94ffc1 · outbound

This paper cites RL's Razor: Why Online Reinforcement Learning Forgets Less.

Adaptive Supervised Anchoring for On-Policy Self-Distillation RL's Razor: Why Online Reinforcement Learning Forgets Less

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.060832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.060832Z digest=sha256:3c199cff0224551dfcca32a69f6b288002ee01ad4e1f178164a65455e090dbf5

Observation 96cc048c-6a51-4d9d-8d19-c3c81990569e · outbound

This paper cites Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.064521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.064521Z digest=sha256:df1461f45793e829531cfd8ea993fb1fe1d4627119d39d5311f2161afc0366ff

Observation e6dccabe-05af-4ba3-aa04-99d1bfdcf144 · outbound

This paper cites LoRA Learns Less and Forgets Less.

Adaptive Supervised Anchoring for On-Policy Self-Distillation LoRA Learns Less and Forgets Less

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.068406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.068406Z digest=sha256:f0c313939ae5ced043f149e2a041f989b1d763581713385c7797a91233e4be74

Observation bd021fb8-4ca7-416d-8dbc-e14946274aa6 · outbound

This paper cites 2025 , publisher =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation 2025 , publisher =

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.410681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.072416Z digest=sha256:1cd4a7bbb35fa03a6cc81da88530206496723cdc05b8dc9d2b23c390b2d5fcad

Observation 8a34036f-e070-4c2b-97e6-452a0c32e547 · outbound

This paper cites 2025 , publisher =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation 2025 , publisher =

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.400237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.076481Z digest=sha256:19244f5142b461471543493c17c8d376e4663209df97b7a4d996a68041cd3d52

Observation b4b42150-e307-4c20-a796-dc0c7b587da5 · outbound

This paper cites ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases.

Adaptive Supervised Anchoring for On-Policy Self-Distillation ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.080228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.080228Z digest=sha256:9e6838ff31b910f741d400dde0e522822cb4e86875f9d989c364fb3f2403a112

Observation c48d9b77-1d45-44db-b66e-cd2b7a044386 · outbound

This paper cites HuggingFace repository , howpublished =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation HuggingFace repository , howpublished =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.389436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.084064Z digest=sha256:0e7e5fe66746019952ed163c6c440c591dcf4775b8d85688d60cc57781a17428

Observation e0949940-0dab-46c0-9699-6ffd8bbff241 · outbound

This paper cites an unresolved cited work.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.087343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.087343Z digest=sha256:a97af26182fa6931b7493fc6fefd80b8d5578f029ed9649c270e161b1edf8d72

Observation 9d625984-c4bd-48aa-a4ca-abcb1d48bcc3 · outbound

This paper cites The Twelfth International Conference on Learning Representations (ICLR) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Twelfth International Conference on Learning Representations (ICLR) , year =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.372655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.090430Z digest=sha256:5c116eed159d229b9627418e163c3e4e2e79416f7042a39b8ed4b6e150268a96

Observation 9fe2c3ec-dceb-4408-8214-de40686100e5 · outbound

This paper cites NeurIPS Datasets and Benchmarks Track , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation NeurIPS Datasets and Benchmarks Track , year =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.362082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.093520Z digest=sha256:b852139e1514f3e7f9e5ca7df60328bee84593942ed0c1c1e591a71862f5a790

Observation e3bb0cdb-52cc-4ea7-8d3c-335e809e1257 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Adaptive Supervised Anchoring for On-Policy Self-Distillation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.096546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.096546Z digest=sha256:c649e1612b59fcd926d286af736e088ae46d434a0928c720d168b3d58f5048d8

Observation 56ec0ebb-7dc8-4bf4-9d26-6610bf176eaa · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.100287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.100287Z digest=sha256:e814353c47b13118445a050eb9314101c96307fee6e51e951303e41db5245251

Observation 01dcbd3a-9392-44ed-84f2-2192eedd7f14 · outbound

This paper cites The Tenth International Conference on Learning Representations (ICLR) , year =.

Adaptive Supervised Anchoring for On-Policy Self-Distillation The Tenth International Conference on Learning Representations (ICLR) , year =

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:30:59.351887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T14:30:59.103436Z digest=sha256:e609ea38e2737cab8b2d6d3ab04b20a4cb3b38ede6ab9e03dd22215cee296c0e

Observation 4af330d0-18fa-446d-8941-f9aa0d809036 · outbound

This paper cites Qwen3 Technical Report.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Qwen3 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.106642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.106642Z digest=sha256:ff3cbea62118d128c6db9dd05e8354a3d9dee1082432b05d6808164d32167ad2

Observation 8c843bf2-b5d4-4618-9bea-18a4fdfe7842 · outbound

This paper cites SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning.

Adaptive Supervised Anchoring for On-Policy Self-Distillation SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.109816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.109816Z digest=sha256:5af7ec5617675dc87a5c302ce604006f013d6017cac88a868e4cb419feace715

Observation 26c43321-8eea-41be-8b5a-a3315a8ea8d9 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Adaptive Supervised Anchoring for On-Policy Self-Distillation Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T14:30:59.113734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:30:59.113734Z digest=sha256:80dfbb197a6b3fa2adc779c1125c6c7cbc4481034df1966190a09f0c86a11e51

Pith citing papers

No inbound Pith citation observations are available.