Pith. sign in

Paper Citation Record · LEDGER

Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2501.05707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05707 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:01:12.434196Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:26:21.685055Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0afc715f-8713-4f01-8a59-1802273dd039 · inbound

Language Games as the Pathway to Artificial Superhuman Intelligence cites this paper.

Language Games as the Pathway to Artificial Superhuman Intelligence Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T22:01:12.434196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:01:12.434196Z digest=sha256:ab8f9f13c0d486d7f37eae7b2c0289836183c77d4413869c0e9d1178c9803fe7

Observation 8b1100ae-3f72-48ac-97ee-2a263b8f21e8 · inbound

Vision-Language Model Dialog Games for Self-Improvement cites this paper.

Vision-Language Model Dialog Games for Self-Improvement Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T11:21:38.288962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:21:38.288962Z digest=sha256:e46d3ff822cc6d8799a5906ab94e6b112a840ea69aba4bfd17e5bfbda7c3c2df

Observation dd056262-8853-4afe-bda8-0cca0aba7945 · inbound

When One LLM Drools, Multi-LLM Collaboration Rules cites this paper.

When One LLM Drools, Multi-LLM Collaboration Rules Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-08T22:33:09.605937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:33:09.605937Z digest=sha256:17aa4ac9dfef4ed8701a1f60fdb9ea7c2ff9e1ebfee07907f4762b10e94082e8

Observation d0a64fdc-e1cd-47f3-86ee-df9b3d89cd9f · inbound

The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants cites this paper.

The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:36.902960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:36.902960Z digest=sha256:b71dddfcbea2c255f9d3738ac2c2ffff659467ec44af117344bf715218fe4cd0

Observation cca9ab72-ddc9-4363-abc4-d0d10b5bd4bb · inbound

MaskSearch: A Universal Pre-Training Framework to Enhance Agentic Search Capability cites this paper.

MaskSearch: A Universal Pre-Training Framework to Enhance Agentic Search Capability Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:18.700217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:18.700217Z digest=sha256:c9ab539273fa02459ca8f432e073474f99c92bfbda1c168e9a53cf0c03144808

Observation 4c042a0f-f30e-41e7-a719-ba6232585ac4 · inbound

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation cites this paper.

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:34.120413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:34.120413Z digest=sha256:2efaebba3c61b90671fbf632e68e722894e075d15611d685ea5c0aebefdf781b

Observation 351222e8-1589-4ce1-9833-6a0d19e4e170 · inbound

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models cites this paper.

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:58.159431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:41:58.159431Z digest=sha256:020c2a3e2990abcaadda872b0799ef87bd6560364de781b5bde9455e98c18cf7

Observation 351c962c-0c08-473d-a4ee-702b72a5f176 · inbound

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems cites this paper.

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:21:42.706995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T23:21:42.029285Z digest=sha256:7a33d41f790a7e73ab76d60e4a6125362867471fc7766e071c2ece8d512cea71

Observation 406e8988-d288-4d62-bb5f-091bed1ec2c5 · inbound

Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs cites this paper.

Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T22:08:57.071164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:08:57.071164Z digest=sha256:971d6d178f634678a6baa83ae8093c26b0b70b464dcdd30be8b9ea9efca8f8ce

Observation 42ad7824-693b-46b6-be49-870abb151ffb · inbound

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control cites this paper.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.902002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.902002Z digest=sha256:5abd0a37bd1942f21ff08304f9066dffeb072318f61d593f5140b26a179a3706

Observation 5b5432d8-2288-4a53-bab1-23b993ac3dca · inbound

OSC: Cognitive Orchestration through Dynamic Knowledge Alignment in Multi-Agent LLM Collaboration cites this paper.

OSC: Cognitive Orchestration through Dynamic Knowledge Alignment in Multi-Agent LLM Collaboration Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T05:51:29.383333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:51:29.383333Z digest=sha256:7564cb59cf1137233937ee4899774bd00ef93935d5713b489d6a1a030e647096

Observation dfa6138b-2b02-4c85-8f7d-5515fa1b9891 · inbound

Emergent Coordination in Multi-Agent Language Models cites this paper.

Emergent Coordination in Multi-Agent Language Models Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:41:16.056498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T10:39:29.934529Z digest=sha256:d7190cc58c24b129df14e8cb27d3d18223ee38ec1a82f51fcc5d7ccb121299df

Observation 5170a9ed-d850-4eb1-aae6-a83eaca2c59f · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:18:20.311469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:95f24a291a167124c70dfe4ba7d6c55d67c580678aa7aeaa8c8b45dcc95cccd1

Observation 4af974d2-00a6-44ff-bcf1-2b5bc6274932 · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:50:49.360286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T09:47:47.051969Z digest=sha256:bf3233a11b275a544c1200ee179de43f59ef5b726cae238dbf8bc136726af622

Observation a36b1f5a-1999-49ff-9d2f-6d417e2d4145 · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-03T06:46:40.425080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:46:40.425080Z digest=sha256:a899f176388761a309e1910bbd8a21112a0fd4086d826890347527a29786966f

Observation b13b9ec4-2338-4f5d-a3c8-abff5daafd3c · inbound

Representing expertise accelerates learning from pedagogical interaction data cites this paper.

Representing expertise accelerates learning from pedagogical interaction data Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:59.337897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:16:06.474991Z digest=sha256:b4f176de6aacb773e43b53560296371298e46d02a0025664f7aa05d6e9b7a83a

Observation f51503d6-c5d2-457d-9499-042a6848065f · inbound

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms cites this paper.

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:07:27.293407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T07:03:59.974892Z digest=sha256:76f60dc90d773954bf07911702286b6b4c941173db55d2a86c4d58812efb32dc

Observation 400f0d37-8ed4-49b3-bcbd-6fc05d955a32 · inbound

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination cites this paper.

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-19T18:02:42.294497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T18:01:06.649723Z digest=sha256:8f6a0e178ec2df02266b4a53d3dd25b77c5aca5fbaacca1de1b526558d9c6f06

Observation cd37659e-a761-4bb0-aecd-e1366ea8b770 · inbound

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions cites this paper.

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:26:21.687466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T14:26:50.444070Z digest=sha256:1bcf978fe281f2861adf1a6534c3f85cde8995c93de27786ca808b16bd8fc140

Observation 0142a891-e66e-49af-bc1c-18998bbdf9d7 · inbound

Mixture of Debaters: Learn to Debate at Architectural Level in Multi-Agent Reasoning cites this paper.

Mixture of Debaters: Learn to Debate at Architectural Level in Multi-Agent Reasoning Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:14:20.891555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:11:02.464556Z digest=sha256:e2b613e14e587a5ae98ac796d2bf73ab90628977d0f3755eeb534fb520b3e5b3