Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:18:49.217321Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 2 inbound Pith citation observations for arXiv:2504.13145.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:18:49.217321Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:35:18.372917Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:37:49.973211Z
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 249b3db7-8aff-4e77-b34a-f77a8c814a26 · outbound
Exploring Expert Failures Improves LLM Agent Tuning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0d3ecae-0ccb-4ac0-9a07-f0eaa1188143 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84cfe1d2-0b4d-4f22-8fa9-e6bc399ca847 · outbound
Exploring Expert Failures Improves LLM Agent Tuning ATLaS: Agent Tuning via Learning Critical Steps
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72cc724b-ec18-4e01-8074-87fae4431a62 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Contextual Markov Decision Processes
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc2a8a1b-8fc2-475e-bf22-77fb5cbf60a5 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Inner Monologue: Embodied Reasoning through Planning with Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a0e70f-1438-40bc-8e6c-d88f0b49fba0 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Can Agents Run Relay Race with Strangers? Generalization of RL to Out-of-Distribution Trajectories
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8a194436-aeff-4e10-bc78-a7631828a1d7 · outbound
Exploring Expert Failures Improves LLM Agent Tuning AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b03e05ab-cc8b-4679-b438-1eb760f011d3 · outbound
Exploring Expert Failures Improves LLM Agent Tuning AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c65e9a6-bc17-4964-b91e-9929624fb5c5 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Let's reward step by step: Step-Level reward model as the Navigators for Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0734153d-3925-46e4-9365-2bd38a29105a · outbound
Exploring Expert Failures Improves LLM Agent Tuning Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44660778-45d2-4f54-9d0c-b38c82ffa363 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Co-Reyes, Rishabh Agarwal, Ankesh Anand, Piyush Patil, Peter J
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 76c7a648-bc90-4e61-a3c0-e7323f7616ec · outbound
Exploring Expert Failures Improves LLM Agent Tuning AgentBank: Towards Generalized LLM Agents via Fine-Tuning on 50000+ Interaction Trajectories
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d39c7cb4-c684-413d-937d-243cf74acd92 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28eeec0e-a312-48d7-88fc-48956ab785d8 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 592ad90b-a4d2-4c0a-a807-e590420720b9 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c99db4df-5e29-4fb1-8941-6487d8232cc0 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0395c9fb-f92f-411d-9e33-5911a80ef306 · outbound
Exploring Expert Failures Improves LLM Agent Tuning AgentTuning: Enabling Generalized Agent Abilities for LLMs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 038236ee-fc72-4ba4-8dc0-6b827899e086 · outbound
Exploring Expert Failures Improves LLM Agent Tuning AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21a63a2a-87fe-4779-8e77-355d2a965055 · outbound
Exploring Expert Failures Improves LLM Agent Tuning WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebf6be87-6040-4376-b23e-04bfd0ca8bc8 · outbound
Exploring Expert Failures Improves LLM Agent Tuning From Novice to Expert: LLM Agent Policy Optimization via Step-wise Reinforcement Learning
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ea9de0-b53f-46ba-a609-7c688668cfed · outbound
Exploring Expert Failures Improves LLM Agent Tuning GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83a868b4-051e-4786-b2b0-982c031b2ea0 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Multi-step Problem Solving Through a Verifier: An Empirical Analysis on Model-induced Process Supervision
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5dd1a83-41b9-4a2f-97c9-555ac768a72b · outbound
Exploring Expert Failures Improves LLM Agent Tuning FireAct: Toward Language Agent Fine-tuning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98cd6fc3-0cdb-4b1e-9952-546dc8930ff9 · outbound
Exploring Expert Failures Improves LLM Agent Tuning ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64679209-0df5-4952-8128-8e06271f7c27 · outbound
Exploring Expert Failures Improves LLM Agent Tuning Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58e481f7-a907-4cbb-9051-e260b61b71d4 · inbound
ProgRM: Build Better GUI Agents with Progress Rewards Exploring Expert Failures Improves LLM Agent Tuning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ac2fa8c3-3207-4f37-aa2d-4f7a74718ad6 · inbound
FailForge: Distilling Procedural Competence from Persistent Failures into Code Agents Exploring Expert Failures Improves LLM Agent Tuning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.