Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2409.03797.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:04:46.846452Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T03:36:29.696220Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 845a9e89-81e7-41c7-a4e2-7771fdf4b6f8 · inbound
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57fe1622-e7d9-4369-90da-9e236bc5ade3 · inbound
A Survey of Context Engineering for Large Language Models NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 89aaede8-7ad7-492b-ace9-1b6eeb8d658c · inbound
LLM-based Agentic Reasoning Frameworks: A Survey from Methods to Scenarios NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47a21541-db06-42f7-a64d-4d957cf2a0e7 · inbound
How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on $\tau$-bench NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91a285a0-ab89-412b-9e4c-467a7bfe332a · inbound
ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 71622b61-2a6f-4e43-8e98-6b82048bf7f5 · inbound
RAG Strategies for Natural Language-Based SQL Query and REST API Call Generation NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a652ed7-9a22-404c-aa7b-60c8d909eb08 · inbound
Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.