Pith. sign in

Paper Citation Record · LEDGER

Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2501.05707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05707 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:36.902960Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:26:21.685055Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d0a64fdc-e1cd-47f3-86ee-df9b3d89cd9f · inbound

The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants cites this paper.

The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:36.902960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:36.902960Z digest=sha256:f2a66a9e395c6507cfa6c91c25c0e084221bd6cff4b01191c80423d583e44fba

Observation cca9ab72-ddc9-4363-abc4-d0d10b5bd4bb · inbound

MaskSearch: A Universal Pre-Training Framework to Enhance Agentic Search Capability cites this paper.

MaskSearch: A Universal Pre-Training Framework to Enhance Agentic Search Capability Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:18.700217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:18.700217Z digest=sha256:cf6855617accf7afd17fdc387f5945cb7f766f38a274c9e7defd6acd0f64d956

Observation 4c042a0f-f30e-41e7-a719-ba6232585ac4 · inbound

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation cites this paper.

OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:34.120413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:34.120413Z digest=sha256:dbfcde17389517da66723d902c09202ee398112fb7bc72295c7bdceea4c9cc0a

Observation 351222e8-1589-4ce1-9833-6a0d19e4e170 · inbound

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models cites this paper.

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:58.159431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:41:58.159431Z digest=sha256:020c2a3e2990abcaadda872b0799ef87bd6560364de781b5bde9455e98c18cf7

Observation 351c962c-0c08-473d-a4ee-702b72a5f176 · inbound

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems cites this paper.

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:21:42.706995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T23:21:42.029285Z digest=sha256:2abb35fa9782924bf4270117615a2937f781947de330d1ae09379c0963c54d92

Observation 406e8988-d288-4d62-bb5f-091bed1ec2c5 · inbound

Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs cites this paper.

Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T22:08:57.071164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:08:57.071164Z digest=sha256:66c99371df14eb127a46879ba63f9b0139e3933e052bc7f6ab006821eb3d34da

Observation 42ad7824-693b-46b6-be49-870abb151ffb · inbound

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control cites this paper.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:55.902002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:55.902002Z digest=sha256:7a0368bf6aa181d988ca7797c0ea2c470f4f074b1e341591cde13762c0b7928b

Observation 5b5432d8-2288-4a53-bab1-23b993ac3dca · inbound

OSC: Cognitive Orchestration through Dynamic Knowledge Alignment in Multi-Agent LLM Collaboration cites this paper.

OSC: Cognitive Orchestration through Dynamic Knowledge Alignment in Multi-Agent LLM Collaboration Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T05:51:29.383333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:51:29.383333Z digest=sha256:957c6a11c736457aafcb2224f72faa081eaa734c8996c26a9eb2702a9b4a78af

Observation dfa6138b-2b02-4c85-8f7d-5515fa1b9891 · inbound

Emergent Coordination in Multi-Agent Language Models cites this paper.

Emergent Coordination in Multi-Agent Language Models Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:41:16.056498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T10:39:29.934529Z digest=sha256:2f4a14eeba11a0ce2fc493c57b73ff7f56e1cb9268b81b612a362ebe533619b3

Observation 5170a9ed-d850-4eb1-aae6-a83eaca2c59f · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:18:20.311469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:c90843c6b30bd709eef34f061257247eade7e102e57b5b80a34a10444595ad8b

Observation 4af974d2-00a6-44ff-bcf1-2b5bc6274932 · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:50:49.360286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T09:47:47.051969Z digest=sha256:baf2e9b32d77b6c6a012500fa2149c784fd0fe225dcc3adb19db9aace2abd799

Observation a36b1f5a-1999-49ff-9d2f-6d417e2d4145 · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-03T06:46:40.425080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:46:40.425080Z digest=sha256:a899f176388761a309e1910bbd8a21112a0fd4086d826890347527a29786966f

Observation b13b9ec4-2338-4f5d-a3c8-abff5daafd3c · inbound

Representing expertise accelerates learning from pedagogical interaction data cites this paper.

Representing expertise accelerates learning from pedagogical interaction data Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:59.337897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:16:06.474991Z digest=sha256:c2119fa41c513006037eeabe00be9016574004b11c17e99f6230fe0617f6103b

Observation f51503d6-c5d2-457d-9499-042a6848065f · inbound

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms cites this paper.

GRAFT-ATHENA: Self-Improving Agentic Teams for Autonomous Discovery and Evolutionary Numerical Algorithms Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:07:27.293407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:03:59.974892Z digest=sha256:e8019e8f4e0cf8d811b43337df111c6a1f682334c015ddb2ec2fa245e8fdd9ec

Observation 400f0d37-8ed4-49b3-bcbd-6fc05d955a32 · inbound

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination cites this paper.

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-19T18:02:42.294497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T18:01:06.649723Z digest=sha256:f01dac6bd78f06b94d11a74d737b3ac66649b9465c885ee6f8e311cf45052e79

Observation cd37659e-a761-4bb0-aecd-e1366ea8b770 · inbound

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions cites this paper.

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:26:21.687466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T14:26:50.444070Z digest=sha256:da275c2e1c7715e8eaea1c7693f1a7338c4be8e003ece4f93ec4aa99a8e7b454

Observation 0142a891-e66e-49af-bc1c-18998bbdf9d7 · inbound

Mixture of Debaters: Learn to Debate at Architectural Level in Multi-Agent Reasoning cites this paper.

Mixture of Debaters: Learn to Debate at Architectural Level in Multi-Agent Reasoning Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:14:20.891555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T07:11:02.464556Z digest=sha256:ac7beca7f168c52eaa0c9e2bd3e0853976b7488c3ca728ede51a186ec3a4d962