Pith. sign in

Paper Citation Record · LEDGER

On the Tool Manipulation Capability of Open-source Large Language Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2305.16504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.16504 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:12:48.085263Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:47:17.727414Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 141acb86-5080-4989-8981-cd27ad4641eb · inbound

ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases cites this paper.

ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases On the Tool Manipulation Capability of Open-source Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:03:48.486935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T23:03:48.426204Z digest=sha256:dc9b391b72d468621e89226669842602cbff002357d74a95aea53d50ac3d2bd9

Observation f54a43fb-c740-4ab5-a05f-6f83d4a42093 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models On the Tool Manipulation Capability of Open-source Large Language Models

Reference 222

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:38.997640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:b54b1252aea4a45795eca10075d5cadab40c70cc89bbbf9c1bd021ee2d229c35

Observation f69583b9-7bb5-40f0-b517-4e201991b2a9 · inbound

GraphTool-Instruction: Revolutionizing Graph Reasoning in LLMs through Decomposed Subtask Instruction cites this paper.

GraphTool-Instruction: Revolutionizing Graph Reasoning in LLMs through Decomposed Subtask Instruction On the Tool Manipulation Capability of Open-source Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T18:19:00.818260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:19:00.818260Z digest=sha256:4d39be6579df4d53087f10b34230dfb04b0b592c68d10bc016ac2ce11a141df3

Observation 6e79476b-4f36-4000-b84b-79cd7883725f · inbound

Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline cites this paper.

Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline On the Tool Manipulation Capability of Open-source Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:27.115273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:17:27.115273Z digest=sha256:480647c63d77fda7a773f8de95f5b5df1fd1c93facf39f2ae15d810e78e10f52

Observation 762e91d1-ed07-44cc-8f27-572fac1edd02 · inbound

A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks cites this paper.

A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks On the Tool Manipulation Capability of Open-source Large Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:23.495454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:27:23.495454Z digest=sha256:6149ead123aa93c3089aeb8da1d09ced8b73af6801bc5b55624bba6b60fd7a0b

Observation f341fe92-734a-41a3-9164-f82bb9e4075c · inbound

AgentRec: Agent Recommendation Using Sentence Embeddings Aligned to Human Feedback cites this paper.

AgentRec: Agent Recommendation Using Sentence Embeddings Aligned to Human Feedback On the Tool Manipulation Capability of Open-source Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T16:17:39.527561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:17:39.527561Z digest=sha256:93dd85e3bb50d7fe20bdb8c17d993259ab027cc2c069cfae115bdc3adb37e7a2

Observation 846f67e8-7398-436f-8935-d577077775e1 · inbound

`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs cites this paper.

`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs On the Tool Manipulation Capability of Open-source Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T17:58:57.467117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:58:57.467117Z digest=sha256:8e2e82c5fe49facfeb6cb5a1b343eacdefb5f7aa3101a4230c6e4e428dc2a724

Observation 18706905-1040-47d4-9779-8211ee1d99f5 · inbound

When2Call: When (not) to Call Tools cites this paper.

When2Call: When (not) to Call Tools On the Tool Manipulation Capability of Open-source Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T10:12:48.085263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:12:48.085263Z digest=sha256:fb18e1fde81c718d4ab8487ec4dcd204fbda254c89f962da62bb9fa5f41ed173

Observation 223db9e6-9553-44c1-83e0-08db59ff02f7 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:52:16.379714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:57fcb670790c7b03005f0a7a95307e9ed43e07597efbb8637a8ee96f26b7396a

Observation e27cad62-82a1-4cf2-8979-786243445a7f · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey On the Tool Manipulation Capability of Open-source Large Language Models

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.615486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.615486Z digest=sha256:62ab5585e719878ed4b1b5a15a99f176b1a46f136047f4140b8facc3b4ab4a91

Observation 9a216b23-d2b7-4c7b-a3e1-e26f67b2ee8d · inbound

CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios cites this paper.

CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios On the Tool Manipulation Capability of Open-source Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:44:04.866566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:44:04.866566Z digest=sha256:a457ad7c3c979952947dd89090e0c8c82201f3a086fe2a26f55f452e9a85d6f9

Observation 5321c14a-8d40-4db5-a641-3352980a350c · inbound

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues cites this paper.

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues On the Tool Manipulation Capability of Open-source Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:40.749933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:02:40.749933Z digest=sha256:a341b531cc4fde37bd83cbfcc43fe5ec1000884fa410e23d4791a4236016d29f

Observation 06320fdd-fbd6-44de-969d-065a372ead7a · inbound

MassTool: A Multi-Task Search-Based Tool Retrieval Framework for Large Language Models cites this paper.

MassTool: A Multi-Task Search-Based Tool Retrieval Framework for Large Language Models On the Tool Manipulation Capability of Open-source Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:43.855623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:22:43.855623Z digest=sha256:4317b49a0ddb85fbce2555dffe5547a307964961b659ab30a6389d9270a3a93c

Observation a3b943cd-2ceb-45d2-bf5a-b805297b5a7e · inbound

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis cites this paper.

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis On the Tool Manipulation Capability of Open-source Large Language Models

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:40:51.380018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T00:37:11.945418Z digest=sha256:75ea8de13ca3c33b01cec3f57fedb0afde492b17f4311bce672d46eeede34ce9

Observation 56fe74af-f73f-4212-a055-c89c49c0ca8a · inbound

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems cites this paper.

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:21:42.199107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T23:21:42.029285Z digest=sha256:1f98dca6537292ba1be2899c84876fa7b1bc5b533880db4a26b874ea7295282e

Observation a8db8f16-c010-416d-b034-7cb28b28d022 · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 260

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:18:20.242995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:6a1341062d754317558304639ff650d3d964d09b917bb3e66b7ab05da09e5037

Observation dd321c8a-80c4-48f9-8e85-7e8113c85983 · inbound

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents cites this paper.

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T23:41:45.208041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:41:45.208041Z digest=sha256:fa33f95bd035205956063b71f180226443a3b16ba8935b2c40e81341a6ecdd2e

Observation c8d96b6f-6b18-422a-afd4-35ab12789168 · inbound

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents cites this paper.

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:35:51.573419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T18:59:59.383781Z digest=sha256:860256459a02008a393ac1fa4932023ae479c9c818e74a2290ea011b18b6cb48

Observation 56c9e580-ad0c-468d-bb26-9bbba19359ab · inbound

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation cites this paper.

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation On the Tool Manipulation Capability of Open-source Large Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:46:24.179373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T14:01:23.894966Z digest=sha256:06b47619112687cb9bb79421efc74da64939ba529f25eda8ef5fa62232cb7377

Observation cd26da40-5c61-4cbd-8219-8a6f79338f7f · inbound

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions cites this paper.

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions On the Tool Manipulation Capability of Open-source Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:31:29.337321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T05:49:57.957009Z digest=sha256:47bc61ee08d30522198d45081a896fb3595033e35b65dc51dab32c1a2d5788c2

Observation c9711691-ec57-4762-b5f1-c871a6dac7de · inbound

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems cites this paper.

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:21:27.819884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T04:19:32.031454Z digest=sha256:c6e5e8fc00f46a36c164924789d3c2e0a678f9b1078a30d67212d7cd2f15fb8b

Observation 029be727-401f-4c00-aa03-e8c5953c8e2a · inbound

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents cites this paper.

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.255056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T22:27:16.974169Z digest=sha256:5e86feff074a079d560ee69f1d714e39a2d27adc030bf1170bd4422df0dc3f8a

Observation 840da58a-95ab-4cc5-8812-1948d7c2153e · inbound

Continual Model Routing in Evolving Model Hubs cites this paper.

Continual Model Routing in Evolving Model Hubs On the Tool Manipulation Capability of Open-source Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:03:24.148298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-29T11:57:15.894309Z digest=sha256:4db0d9976a8b26bb0a3f9de799f58f0c28353632c236027bdfc2aedcb0dc2c88

Observation 9fb5f490-80ad-41e8-90a9-6f2d5f558164 · inbound

Capability Self-Assessment: Teaching LLMs to Know Their Limits cites this paper.

Capability Self-Assessment: Teaching LLMs to Know Their Limits On the Tool Manipulation Capability of Open-source Large Language Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.682649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T22:25:50.579196Z digest=sha256:58bbf1fc55c35fd7ee198f607e2edc2d5f3251b3597deee776a4e751e5b701c7

Observation 32594cdc-c444-4477-96cc-3f6f2062ddec · inbound

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems cites this paper.

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:47:17.728995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T21:53:37.616447Z digest=sha256:16e4096a54c922ffa4a16765128a0f503bd98c3a3c575e0f910d30b74542cdad

Observation ba7935e2-746e-4d98-ab8a-713c6f842280 · inbound

A Formal Hierarchical Architecture for Agentic Orchestration with Stack-Based Execution and Lazy Discovery cites this paper.

A Formal Hierarchical Architecture for Agentic Orchestration with Stack-Based Execution and Lazy Discovery On the Tool Manipulation Capability of Open-source Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T06:48:32.645976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:48:32.645976Z digest=sha256:462d0e3a0ec3d2647f4c1e79b80b7bb1e467628282601e0330038b665eb5f563

Observation b3683caa-8e72-47b0-9426-3861c75c768e · inbound

PAUSE: A User-Centric Benchmark for Personal AI Assistants in Unified Service Environments cites this paper.

PAUSE: A User-Centric Benchmark for Personal AI Assistants in Unified Service Environments On the Tool Manipulation Capability of Open-source Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T15:28:54.387803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:28:54.387803Z digest=sha256:4ff9d18d9b92ee73872bb4f51b59336d8c672f93a5c33290b9b0fab417e2139c

Observation fd84d87b-b936-4306-b40e-00b2d0c2948e · inbound

DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents cites this paper.

DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:15.951161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:15.951161Z digest=sha256:af079bde9ab7a0794c07b1ccbfb26fa3d3072ed938001babbed432fd8d69cfcb

Observation 0d8f8a41-0d7f-467f-9700-deacd409d182 · inbound

SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information cites this paper.

SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information On the Tool Manipulation Capability of Open-source Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T19:21:58.919706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:21:58.919706Z digest=sha256:29fe381ab550195a5a79d43a76eba89733320a981b488061dbab9da4eea7bc4b