Pith. sign in

Paper Citation Record · LEDGER

Tuning Language Models by Proxy

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2401.08565.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.08565 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:44:46.598114Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T12:33:45.021126Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4a8f6943-03bf-4117-baa5-36fa9e476602 · inbound

A Roadmap to Pluralistic Alignment cites this paper.

A Roadmap to Pluralistic Alignment Tuning Language Models by Proxy

Reference 279

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:37:53.483023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-16T14:37:53.279275Z digest=sha256:7ea6089574ebbd4098bcbd1d21d8005f7536d664f9ffda836460b33bde1d2d06

Observation f5014e00-5819-44b8-abeb-3b4fe7587bdc · inbound

Best Practices for Large Language Models in Radiology cites this paper.

Best Practices for Large Language Models in Radiology Tuning Language Models by Proxy

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-12T04:35:59.543008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:35:59.543008Z digest=sha256:756620a816b0576256e5d3de16781333ebe0d26f9997a3fc3e68cee1909d8512

Observation 41ebfd6e-55d1-4d26-aa73-a81c4ff590db · inbound

Ensembling Large Language Models with Process Reward-Guided Tree Search for Better Complex Reasoning cites this paper.

Ensembling Large Language Models with Process Reward-Guided Tree Search for Better Complex Reasoning Tuning Language Models by Proxy

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T11:08:56.453684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:08:56.453684Z digest=sha256:6238d206a469b091100262e94470044fc4a26abfc163deb69f232e549ec50d0e

Observation efdfe643-d800-4bc3-926e-b785708aa619 · inbound

Stream Aligner: Efficient Sentence-Level Alignment via Distribution Induction cites this paper.

Stream Aligner: Efficient Sentence-Level Alignment via Distribution Induction Tuning Language Models by Proxy

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:17:48.408445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:17:48.408445Z digest=sha256:1169961fd05c088e0560ec3f919897acbb223e4edb67e233eb02d47da6dc15fa

Observation b1a42d7a-320b-4931-8a4a-d2ca0e46e4fa · inbound

Logits are All We Need to Adapt Closed Models cites this paper.

Logits are All We Need to Adapt Closed Models Tuning Language Models by Proxy

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T14:18:05.620575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:18:05.620575Z digest=sha256:1fd93c732fee577432b0130ac8cec5ccdbca240ff00350e59157ea4c6b866b05

Observation 69377ed7-c5cd-44dc-8fc3-e00252a25a9d · inbound

Towards Harnessing the Collaborative Power of Large and Small Models for Domain Tasks cites this paper.

Towards Harnessing the Collaborative Power of Large and Small Models for Domain Tasks Tuning Language Models by Proxy

Reference 138

Resolution
unresolved
no resolver link, observed 2026-08-16T10:44:46.598114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:44:46.598114Z digest=sha256:684294a251b0225144d30d842cb47ce76874f63c147bc635641717c56b1e067b

Observation e32aa794-f5d7-4a2e-92c8-0823552377f8 · inbound

Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models cites this paper.

Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models Tuning Language Models by Proxy

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:58.926261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:37:58.926261Z digest=sha256:79dc98e04ea617ff5fb6349060b5426f21d4353d116e9e428ddd5ed5e8e63d22

Observation fcb289a6-a2f4-4b35-b02e-0cc2d0a45638 · inbound

Scalable, Symbiotic, AI and Non-AI Agent Based Parallel Discrete Event Simulations cites this paper.

Scalable, Symbiotic, AI and Non-AI Agent Based Parallel Discrete Event Simulations Tuning Language Models by Proxy

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:49.678073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:49.678073Z digest=sha256:775388956a9683c7bdcc4d53815be0bc7e5e6e5ce039d8afc2727b458898af58

Observation 48185574-370d-49dd-8c9f-6ea8dab5867c · inbound

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training cites this paper.

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training Tuning Language Models by Proxy

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:58.459933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:45:58.459933Z digest=sha256:029a1d022eda60e5835b03d1a3627acc19711836b883f12e31ef8963f5e10d05

Observation a3f7c747-c499-49ab-a392-045e562f2926 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models Tuning Language Models by Proxy

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:43.942431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:43.942431Z digest=sha256:62c2f077f5fd906ab4860920ba700d6f2b13a5a187d08640b33d76e59db0b6cd

Observation 0447a07e-40a2-4891-b1f9-d0430be7c590 · inbound

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment cites this paper.

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment Tuning Language Models by Proxy

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:56.453878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:36:44.401045Z digest=sha256:89011d82bef37d8e750423af722b9f5e7ce77d7d63161ba82e9b5e88627860ae

Observation 6b425db2-3532-46b7-93d3-477c1b3e4efb · inbound

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration cites this paper.

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration Tuning Language Models by Proxy

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:16:04.584157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T16:07:20.180349Z digest=sha256:c4b03524734d8b3f7576026bdbe86079f8657faac2526640c841be2dc7a921ce

Observation 9ad932fa-1dbe-4f2f-a761-b03deac132c5 · inbound

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models cites this paper.

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models Tuning Language Models by Proxy

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:19:20.680811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T10:16:26.526986Z digest=sha256:56e7a3dae4c338736d81f20c6fc418c8880a977479adb2c6a5fd83786071cf65

Observation 72559b78-756f-4a17-ab44-4b65a2e8e7d7 · inbound

Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs cites this paper.

Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs Tuning Language Models by Proxy

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:39:43.651320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-21T06:39:06.365240Z digest=sha256:689bc91067755e3f7fc708a23982e98597df38f097b58e817e3d4167e4b15b1a

Observation ab6a4246-7973-4358-8b72-a7543faf22f8 · inbound

Failed Reasoning Traces Tell You What Is Fixable (But Not by Reading Them) cites this paper.

Failed Reasoning Traces Tell You What Is Fixable (But Not by Reading Them) Tuning Language Models by Proxy

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:45.672823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T06:48:28.788806Z digest=sha256:206515156a3d84a587a26fad794a0749d5bda7feb790f58890499f421dd7d261

Observation cde6781f-db19-42ab-8564-a536ac633c9c · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Tuning Language Models by Proxy

Reference 110

Resolution
verified exact
local_arxiv, observed 2026-07-07T12:33:45.022809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-07T12:31:42.224094Z digest=sha256:68f7e2974f12a693a4cc7b29e1dc2f9d9403bace6117213492f75c67ab04923d

Observation 1957543d-eade-4ff2-8eb8-308fa5535ebc · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Tuning Language Models by Proxy

Reference 106

Resolution
unresolved
no resolver link, observed 2026-07-11T07:01:56.628017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:01:56.628017Z digest=sha256:963d9076827b3adfa2b6f9ddda1fe8748c173fee502e4afb4a9f326e81e0fbdb

Observation 9780541d-452f-4e6e-86cf-3b72c5293382 · inbound

Proxy Exploration and Reusable Guidance: A Modular LLM Post-Training Paradigm via Proxy-Guided Update Signals cites this paper.

Proxy Exploration and Reusable Guidance: A Modular LLM Post-Training Paradigm via Proxy-Guided Update Signals Tuning Language Models by Proxy

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T05:09:00.865375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:09:00.865375Z digest=sha256:2d0365f02473ab2dac7c311f78e48d12ba3bca0b1e4b1edaec4048c407934ffb

Observation 8d2eae3d-df45-4cb0-b8b3-41f5bfc8df15 · inbound

Weak-to-Strong On-Policy Distillation cites this paper.

Weak-to-Strong On-Policy Distillation Tuning Language Models by Proxy

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T00:26:23.349392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:26:23.349392Z digest=sha256:ee372a0f7fbafb90b787512177fcb004d404376eb3063375bc79290880e9dab2

Observation 0854683a-1627-4ed5-b620-9a39239e3261 · inbound

Decoupled Contrastive Decoding via Expert-Aligned Drafting cites this paper.

Decoupled Contrastive Decoding via Expert-Aligned Drafting Tuning Language Models by Proxy

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T20:49:47.311952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:49:47.311952Z digest=sha256:cf55854a187f4670922526ea29f6b2e6aab2ff84b0fe7bbc3926ad183da3885e