Pith. sign in

Paper Citation Record · LEDGER

Distilling Reasoning Capabilities into Smaller Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2212.00193.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.00193 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:37:40.763610Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c7785893-e39f-48a0-bc0d-a4e16d8f8e84 · inbound

Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation cites this paper.

Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation Distilling Reasoning Capabilities into Smaller Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T17:37:40.763610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T17:37:40.763610Z digest=sha256:497e51920653e2b6aef81f9f653b780e11b4919f9b700df1635cb709f7d3a732

Observation b45c7894-61c3-4b03-b761-5fce8ea192e9 · inbound

Don't Just Demo, Teach Me the Principles: A Principle-Based Multi-Agent Prompting Strategy for Text Classification cites this paper.

Don't Just Demo, Teach Me the Principles: A Principle-Based Multi-Agent Prompting Strategy for Text Classification Distilling Reasoning Capabilities into Smaller Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T13:39:04.476069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:39:04.476069Z digest=sha256:741580fdc1bce07ef9f40f1a594497738b36f247a22ed42b052fa4d2860aae1a

Observation 257dc47b-fe8a-44c9-a2f9-2b9ff886e6f3 · inbound

Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability cites this paper.

Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability Distilling Reasoning Capabilities into Smaller Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:24.689688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:40:24.689688Z digest=sha256:8c54b9f46dedb02d282525074f4442c921fe4c7ed6fe0a2e92872cbdbba60a1a

Observation f26e25c4-142e-4fb3-bd9d-8a960bff3f6b · inbound

Learning to Insert [PAUSE] Tokens for Better Reasoning cites this paper.

Learning to Insert [PAUSE] Tokens for Better Reasoning Distilling Reasoning Capabilities into Smaller Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:19.933207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:06:19.933207Z digest=sha256:c93f2c95af8237742786a45ebbec4bdccf74bee5ffdb292dd3e2ccc58f42e722

Observation 5e535acb-da1f-4c43-9b0c-23df43c80594 · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Distilling Reasoning Capabilities into Smaller Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:41.023501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:41.023501Z digest=sha256:1af3088f2fb04454c3d50d4bc38f0763e33335c5e8ee3ddd480bcedeb32f8a75

Observation f5735be7-56cc-43b7-8cfa-c2bad23c8ae5 · inbound

SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery cites this paper.

SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery Distilling Reasoning Capabilities into Smaller Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T21:37:22.374350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:37:22.374350Z digest=sha256:25f3a0297ef92b9b99d5f1c05e1cd2eb9c0a28297990c4800d7daad8404d5d72

Observation 6de40924-e738-4283-918e-1dad23ad4ab0 · inbound

Deep sequence models tend to memorize geometrically; it is unclear why cites this paper.

Deep sequence models tend to memorize geometrically; it is unclear why Distilling Reasoning Capabilities into Smaller Language Models

Reference 166

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:40:36.235761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T20:38:18.005002Z digest=sha256:5de566d3ff342ed5c898f9eb3ca1d9d4d25b66be24b1f9e9ed51dc8ec562fe53

Observation 7d5052cc-e4bf-4c50-bf7b-4a2c0948d171 · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces Distilling Reasoning Capabilities into Smaller Language Models

Reference 157

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.583158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:ad607c44d4601c73dba186db7a198720a5d2d1f70d95d2ce5680c258e8ed82d6

Observation 5d0924f7-af5f-4a83-9f8b-29861e27fa5b · inbound

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers cites this paper.

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers Distilling Reasoning Capabilities into Smaller Language Models

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:46.691403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T23:29:02.457697Z digest=sha256:cc60489776e658307e1a6fda58454f6957c435fb4d8501b5ea9349b61f4f156c

Observation 191a9b54-8171-4add-b8d7-7cf65166ebc5 · inbound

Toward Calibrated, Fair, and accurate Deepfake Detection cites this paper.

Toward Calibrated, Fair, and accurate Deepfake Detection Distilling Reasoning Capabilities into Smaller Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-28T07:11:45.303811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T07:05:18.026601Z digest=sha256:1072e45631a13753545c836c5d1b40e89963a17b8cac20efa8e099ef2b3e64e9