Pith. sign in

Paper Citation Record · LEDGER

Distilling Reasoning Capabilities into Smaller Language Models

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2212.00193.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.00193 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:45:20.076077Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cf79f997-8984-4103-93fc-7b75e0bf4c72 · inbound

Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models cites this paper.

Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models Distilling Reasoning Capabilities into Smaller Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T12:45:20.076077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:45:20.076077Z digest=sha256:7e9314f77e43ed674a4edbf30ea4b211dd32ca27a42f8a4e9890b0c968b5c92c

Observation 18f47611-1846-4d42-8e29-a4822852b890 · inbound

Enhancing Generalization in Chain of Thought Reasoning for Smaller Models cites this paper.

Enhancing Generalization in Chain of Thought Reasoning for Smaller Models Distilling Reasoning Capabilities into Smaller Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T19:42:22.814812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:42:22.814812Z digest=sha256:eaa4efbe3354e090856b0430a229e2507de9fd489e487f6aae5ca416720983ec

Observation c7785893-e39f-48a0-bc0d-a4e16d8f8e84 · inbound

Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation cites this paper.

Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation Distilling Reasoning Capabilities into Smaller Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T17:37:40.763610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T17:37:40.763610Z digest=sha256:cfdc73ea771debbf68d30b5b0ce0962c403d8a154b56185874f140c0c242998d

Observation b45c7894-61c3-4b03-b761-5fce8ea192e9 · inbound

Don't Just Demo, Teach Me the Principles: A Principle-Based Multi-Agent Prompting Strategy for Text Classification cites this paper.

Don't Just Demo, Teach Me the Principles: A Principle-Based Multi-Agent Prompting Strategy for Text Classification Distilling Reasoning Capabilities into Smaller Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T13:39:04.476069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:39:04.476069Z digest=sha256:923c087a7aa04a4df38380731ae37951d42d11aa69d52504f4116e5383a4b9ae

Observation 257dc47b-fe8a-44c9-a2f9-2b9ff886e6f3 · inbound

Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability cites this paper.

Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability Distilling Reasoning Capabilities into Smaller Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:24.689688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:40:24.689688Z digest=sha256:ad54ae043050c68ee92095797e0946210ed7779218c0e40122344daee50f704a

Observation f26e25c4-142e-4fb3-bd9d-8a960bff3f6b · inbound

Learning to Insert [PAUSE] Tokens for Better Reasoning cites this paper.

Learning to Insert [PAUSE] Tokens for Better Reasoning Distilling Reasoning Capabilities into Smaller Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:19.933207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:06:19.933207Z digest=sha256:9e835d3790c4b0f027131e14a63f6b2a5359309c76ca99b16ee89aa2ac2aa219

Observation 5e535acb-da1f-4c43-9b0c-23df43c80594 · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Distilling Reasoning Capabilities into Smaller Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:41.023501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:41.023501Z digest=sha256:b2799559feacf6cc978dc08518568a66f0acc6a1f95a4fa611cde4f0332cc92b

Observation f5735be7-56cc-43b7-8cfa-c2bad23c8ae5 · inbound

SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery cites this paper.

SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery Distilling Reasoning Capabilities into Smaller Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T21:37:22.374350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:37:22.374350Z digest=sha256:ac8d875ea56778187919f52485f0fe858a426203ed14567c3a2bdbc214852080

Observation 6de40924-e738-4283-918e-1dad23ad4ab0 · inbound

Deep sequence models tend to memorize geometrically; it is unclear why cites this paper.

Deep sequence models tend to memorize geometrically; it is unclear why Distilling Reasoning Capabilities into Smaller Language Models

Reference 166

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:40:36.235761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T20:38:18.005002Z digest=sha256:32a5f1c39aa04b61d83667fbc47862c2bb94a6f5310e6a2e6aae79996de68a71

Observation 7d5052cc-e4bf-4c50-bf7b-4a2c0948d171 · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces Distilling Reasoning Capabilities into Smaller Language Models

Reference 157

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.583158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:09c14b66ed0455f595e264836d973e162854ed014f8a0ee378a7949b22976a60

Observation 5d0924f7-af5f-4a83-9f8b-29861e27fa5b · inbound

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers cites this paper.

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers Distilling Reasoning Capabilities into Smaller Language Models

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:46.691403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T23:29:02.457697Z digest=sha256:727788fe5de42beff563eb701f21c0c27a25dd9ca5fcc08bc166d8b202d1d71d

Observation 191a9b54-8171-4add-b8d7-7cf65166ebc5 · inbound

Toward Calibrated, Fair, and accurate Deepfake Detection cites this paper.

Toward Calibrated, Fair, and accurate Deepfake Detection Distilling Reasoning Capabilities into Smaller Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-28T07:11:45.303811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T07:05:18.026601Z digest=sha256:232f4c4d801c21035fe2eafd4ab914cbb253ec0fe71e22525f69fe8400f220b3