Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 71 inbound Pith citation observations for arXiv:2310.05915.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T15:01:44.601401Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation b87ab206-5067-4d51-9903-e4393ce36fea · inbound
Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security FireAct: Toward Language Agent Fine-tuning
Reference 247
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 646dde47-932f-4ed0-aa43-f7185d9a3607 · inbound
Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 47799336-aff9-4f60-b8de-51c448c20082 · inbound
Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training FireAct: Toward Language Agent Fine-tuning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f718996a-ad8b-406c-8ffd-65ecf8b4eb47 · inbound
InSTA: Towards Internet-Scale Training For Agents FireAct: Toward Language Agent Fine-tuning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22b0482f-0f63-45dc-9219-3a0fb7b7248f · inbound
Digi-Q: Learning Q-Value Functions for Training Device-Control Agents FireAct: Toward Language Agent Fine-tuning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a52da57-1d21-4abd-889f-50538f316495 · inbound
Effective Reinforcement Learning for Reasoning in Language Models FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d59dde0-aa66-4377-b779-ff008ad7da81 · inbound
Large Language Models for Planning: A Comprehensive and Systematic Survey FireAct: Toward Language Agent Fine-tuning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e07b629-aa8e-4ecc-a5f4-db57fc4ab08d · inbound
Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24ba30b1-3a1a-4744-9b32-ffb030550847 · inbound
Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acce814c-6101-497e-8130-b2be289ae051 · inbound
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution FireAct: Toward Language Agent Fine-tuning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34e23431-e069-445d-90ab-5951361985e9 · inbound
RRO: LLM Agent Optimization Through Rising Reward Trajectories FireAct: Toward Language Agent Fine-tuning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e42a0d3b-5c02-40f9-8b20-fe1b05831de5 · inbound
Agent-Environment Alignment via Automated Interface Generation FireAct: Toward Language Agent Fine-tuning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05a885bf-085a-4741-89e4-742cb11aae9e · inbound
WebDancer: Towards Autonomous Information Seeking Agency FireAct: Toward Language Agent Fine-tuning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da351dd0-93da-4672-b144-ced5533be68a · inbound
OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 622a48e0-de72-472a-a0c6-f43cbd5a0d5a · inbound
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7722ca1e-4be8-49b1-b8d0-65aaa2cb1e06 · inbound
LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5753f9fc-d847-4e26-a0c5-593e07aa9abe · inbound
Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games FireAct: Toward Language Agent Fine-tuning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1ac41ca7-a31c-4ac2-885e-2f056241fcbe · inbound
Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfbb0d56-e472-4344-ab30-6a7a09779e7c · inbound
Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57d5bf82-b0ed-40e6-a1f0-041237f1d47e · inbound
Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems FireAct: Toward Language Agent Fine-tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92f03b00-374b-4a14-9c7c-068dcff9429b · inbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfcb3673-566d-4a51-bd5f-6caf64e3b635 · inbound
WebSailor: Navigating Super-human Reasoning for Web Agent FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4ece5b4b-0d92-46bc-a299-ae61de169c11 · inbound
SAND: Boosting LLM Agents with Self-Taught Action Deliberation FireAct: Toward Language Agent Fine-tuning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6f50084-356e-42ed-b398-20cf7978c4cf · inbound
Initial Steps in Integrating Large Reasoning and Action Models for Service Composition FireAct: Toward Language Agent Fine-tuning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3baeeb8c-e4ef-4928-8cff-4e32bae54d62 · inbound
MMAT-1M: A Large Reasoning Dataset for Multimodal Agent Tuning FireAct: Toward Language Agent Fine-tuning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d149853-178e-4757-a71c-fdf5f94f7ea8 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey FireAct: Toward Language Agent Fine-tuning
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 233947b3-af51-43fb-aa22-d60d4dd230e7 · inbound
Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems FireAct: Toward Language Agent Fine-tuning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c363136c-8b56-418e-a591-d3f89be6daf9 · inbound
On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a703daeb-00fa-40df-9da0-698878cdb6e4 · inbound
MemVerse: Multimodal Memory for Lifelong Learning Agents FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 701c787b-a889-4547-acf4-0b55d6b846b8 · inbound
Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 215a4d4b-b014-49a7-b769-75b812a26f03 · inbound
SoK: Agentic Skills -- Beyond Tool Use in LLM Agents FireAct: Toward Language Agent Fine-tuning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2334525-f988-410c-bc05-b0a5cb16df15 · inbound
ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache FireAct: Toward Language Agent Fine-tuning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ebdc028a-c5d4-46cc-9c4e-61c78c258e34 · inbound
PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory FireAct: Toward Language Agent Fine-tuning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b0613023-41f0-471c-990b-2b4ecff093f3 · inbound
From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d76db5ba-0679-41f7-aad1-bd04ba976c6a · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0cfc38ba-4da7-40ec-b7a5-cd2a40a62e60 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2816bf74-7e82-4ed4-bfc5-c5539c5ab24a · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 414876f4-f394-4d5f-ae80-fb49c6d560d0 · inbound
SOD: Step-wise On-policy Distillation for Small Language Model Agents FireAct: Toward Language Agent Fine-tuning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a76083b1-83e4-42ea-9057-87b5a5a12829 · inbound
SOD: Step-wise On-policy Distillation for Small Language Model Agents FireAct: Toward Language Agent Fine-tuning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c713480d-1da4-4114-941f-bc60dfa7ca11 · inbound
Learning CLI Agents with Structured Action Credit under Selective Observation FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e71c9f61-b55d-40a9-ad30-20bcc08b828b · inbound
Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation FireAct: Toward Language Agent Fine-tuning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3701cabe-b463-4441-948b-0692179f4b81 · inbound
Verifiable Process Rewards for Agentic Reasoning FireAct: Toward Language Agent Fine-tuning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 972a20b1-7617-496f-a8cf-7bf7c87c0396 · inbound
Verifiable Process Rewards for Agentic Reasoning FireAct: Toward Language Agent Fine-tuning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 50bf99b1-d234-4603-9715-a2fd78b34f04 · inbound
SkillGen: Verified Inference-Time Agent Skill Synthesis FireAct: Toward Language Agent Fine-tuning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d945513a-b9ff-4c3d-8ce0-3cf7ca4ee900 · inbound
ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents FireAct: Toward Language Agent Fine-tuning
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03da1012-3696-4a93-90a2-e6fa14d1f4d9 · inbound
ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents FireAct: Toward Language Agent Fine-tuning
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c719705-040b-4d02-826f-e8d9c0595458 · inbound
TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination FireAct: Toward Language Agent Fine-tuning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 37d965ed-ee2c-4c4d-87e8-d9aac2c260eb · inbound
AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering FireAct: Toward Language Agent Fine-tuning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf1dca57-c7ec-4bf1-a977-a4b09a47e521 · inbound
Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost FireAct: Toward Language Agent Fine-tuning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87fd8886-651e-4c1b-9b21-1a64611edfbf · inbound
Test-Time Deep Thinking to Explore Implicit Rules FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 38045e09-7dbd-448d-89d7-bdbc8e7d98c3 · inbound
Agent Explorative Policy Optimization for Multimodal Agentic Reasoning FireAct: Toward Language Agent Fine-tuning
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 64d64355-8c2a-4094-8ea1-bf6143d4dfe2 · inbound
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents FireAct: Toward Language Agent Fine-tuning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71ae13f4-84af-4eac-86d7-6e31862d9f7f · inbound
SaliMory: Orchestrating Cognitive Memory for Conversational Agents FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee620fc9-0804-4521-bc91-bd3ab31fbd65 · inbound
Self-evolving LLM agents with in-distribution Optimization FireAct: Toward Language Agent Fine-tuning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 93edcb76-e681-48db-94cb-a54fbdee09f9 · inbound
Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents FireAct: Toward Language Agent Fine-tuning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d2f72c89-7104-4a4b-a26a-624139ce3b11 · inbound
Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents FireAct: Toward Language Agent Fine-tuning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 698a2d61-ed51-475a-b800-89780383ac57 · inbound
AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5074fd1-6551-4c79-8f50-a25b58c0258c · inbound
Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks FireAct: Toward Language Agent Fine-tuning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8096747b-5fb9-4f05-8ddc-2319995977ba · inbound
Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents FireAct: Toward Language Agent Fine-tuning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2444a382-c0b1-40c2-900f-f5c24b58fa44 · inbound
Atomic Task Graph: A Unified Framework for Agentic Planning and Execution FireAct: Toward Language Agent Fine-tuning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56cecc91-1805-4e49-912f-9797b1ea6101 · inbound
STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training FireAct: Toward Language Agent Fine-tuning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d8d235e-bea8-4cda-88cc-f30d8bc5c0a4 · inbound
Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fadb3a24-6dea-449d-8890-a67f0292a037 · inbound
Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories FireAct: Toward Language Agent Fine-tuning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0812fa76-c88c-4290-9848-71d20c002f2b · inbound
From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training FireAct: Toward Language Agent Fine-tuning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8650e116-c765-4a9c-bdf2-a1b5bc534758 · inbound
Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures FireAct: Toward Language Agent Fine-tuning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bda8c0e-88b7-4aa2-b1f1-f9856096f671 · inbound
Reason Before You Retrieve: Agentic Planning for Multi-modal RAG FireAct: Toward Language Agent Fine-tuning
Reference 129
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 212faca4-28ac-4214-93c4-009490a9726b · inbound
ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning FireAct: Toward Language Agent Fine-tuning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d4c95c5-497b-4e2b-9371-b8e6e72dd8c4 · inbound
AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents FireAct: Toward Language Agent Fine-tuning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5191ba9d-6e22-40fb-8c00-8ee41facf7f3 · inbound
AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents FireAct: Toward Language Agent Fine-tuning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b35807e4-fae5-419d-a086-825e7bb85ae3 · inbound
Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems FireAct: Toward Language Agent Fine-tuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f20d570-efd7-4c93-ab90-7707693ad3a7 · inbound
MemHarness: Memory Is Reconstructed, Not Replayed FireAct: Toward Language Agent Fine-tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.