Pith. sign in

Paper Citation Record · LEDGER

PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2305.13888.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.13888 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:46.868581Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T02:39:33.457522Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c023210b-b29c-4663-9930-aaad668ead97 · inbound

A Survey on Efficient Inference for Large Language Models cites this paper.

A Survey on Efficient Inference for Large Language Models PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:39:33.459253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T02:39:33.007894Z digest=sha256:209ff84b3119b1f463e5c40d73d7c48edbf43150d08526bc6b3cad24254a7aa8

Observation f1c867bc-ce99-4f8b-9bc7-a59cb9c8e73a · inbound

Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation cites this paper.

Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:46.868581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:46.868581Z digest=sha256:3fe1579b5b2ce223336a906cddf106cdd6292cf389ffa5edc049cb67f49f7cfc

Observation 6597f0b2-09ce-435a-8e16-6b832946d65a · inbound

Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models cites this paper.

Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T12:45:20.180921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:45:20.180921Z digest=sha256:e3e48c7ed087f801e267865130f2090b5a362d17762ccb9509ec3cf7c8f87fef

Observation 1b8cc4fd-add8-4b09-834c-ce222dc029a6 · inbound

Deploying Foundation Model Powered Agent Services: A Survey cites this paper.

Deploying Foundation Model Powered Agent Services: A Survey PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning

Reference 220

Resolution
unresolved
no resolver link, observed 2026-08-11T13:09:46.825042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:09:46.825042Z digest=sha256:4c1bc28f5a4c9a60fcbd05b56349511da70d0bba3db69bc09af9626c0a3c303a

Observation d6d1e800-8918-4325-9b43-ad8eeb96c8b2 · inbound

Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster cites this paper.

Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:34.203188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:32:34.203188Z digest=sha256:95174afa16924e929b65348251c986a216efc456c7f2ac40211598212eaf70c3