Pith. sign in

Paper Citation Record · LEDGER

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses?

As of 10 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2608.04828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04828 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:38:43.670203Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a4266c2f-6900-4907-be5e-e97ba459722a · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.370321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.370321Z digest=sha256:7671fe298da71c2b25a6cb0f69df92ec4da06cb4bd9564e79368d75afb934f34

Observation 7589ae9d-d7f5-4bd1-90c8-da0b6ae5b8af · outbound

This paper cites Theagentcompany: benchmarking llm agents on consequen- tial real world tasks.Advances in Neural Information Processing Systems, 38, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Theagentcompany: benchmarking llm agents on consequen- tial real world tasks.Advances in Neural Information Processing Systems, 38, 2026

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.467109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:41.416490Z digest=sha256:792948b09ae88931aa98406700c81fa5d19a6c9ef2a5c67d33a5a3d78ca33f30

Observation 80bb1fca-ef73-4483-a08a-a46343d5056e · outbound

This paper cites Skill-R1: Agent Skill Evolution via Reinforcement Learning.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Skill-R1: Agent Skill Evolution via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.569596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.569596Z digest=sha256:236dd3960a93dec1d0b338e97fdfdd2c6665031901b7e06f7739258cd2f3e398

Observation 5f1ac2f4-bfda-4d8d-9d64-fb38090d1ccb · outbound

This paper cites Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.769513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.769513Z digest=sha256:2a041b5b483cd637ca05f70fd29b83e020b5755ead8d42e39ace03bf1e57046f

Observation 8a87d7b9-7bd6-48a7-96b4-fb039b38cb34 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.966146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.966146Z digest=sha256:156bcea6abe89efa01fbd6f8fbd0711931eff7cdf3916fc439f34bf944b823f2

Observation 8cf06d0b-ed6a-4d69-8d5c-d4b138ca2a04 · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.040035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.040035Z digest=sha256:725e08da1cc52dd27f725553ff7896da9a2d74d6f7a346e297e13bb9ea8d03ef

Observation 90ac718d-d84d-4824-acef-9aeb41f4943a · outbound

This paper cites Orga- nizing, orchestrating, and benchmarking agent skills at ecosystem scale.arXiv preprint arXiv:2603.02176, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Orga- nizing, orchestrating, and benchmarking agent skills at ecosystem scale.arXiv preprint arXiv:2603.02176, 2026

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.079299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.079299Z digest=sha256:adea3dd7139f371c603339673eecae28d1c690f4a1be2ecccf11038f24c42b31

Observation f0ee1b70-516b-47a8-9423-c4facdd1a8ff · outbound

This paper cites Skillnet: Create, evaluate, and connect ai skills.arXiv preprint arXiv:2603.04448, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Skillnet: Create, evaluate, and connect ai skills.arXiv preprint arXiv:2603.04448, 2026

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.169353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.169353Z digest=sha256:58662376e4a6a04262f21f305e3e4f51ba555d0af07d80807dc3186869bfd437

Observation a58fb1c1-d4fd-4fbf-9923-db935930c6b9 · outbound

This paper cites Autoskill: Experience-driven lifelong learning via skill self-evolution.arXiv preprint arXiv:2603.01145, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Autoskill: Experience-driven lifelong learning via skill self-evolution.arXiv preprint arXiv:2603.01145, 2026

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.251238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.251238Z digest=sha256:f9000c4fe885a29bd59fe8949b4efa91870a702a9d1ba35712e0db5350a0e5ef

Observation 075b6390-617e-4a08-86a9-68a5ad758c82 · outbound

This paper cites Memento-skills: Let agents design agents.arXiv preprint arXiv:2603.18743, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Memento-skills: Let agents design agents.arXiv preprint arXiv:2603.18743, 2026

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.343681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.343681Z digest=sha256:f2e2849ccca8316e787599b86b20a954aa856eab0df361eff0c1e1494957b422

Observation 8bc841d7-9ea7-4f04-a959-4cd49f36f544 · outbound

This paper cites SLBench: Evaluating How LLM Agents Follow Logical Relations in Skills.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SLBench: Evaluating How LLM Agents Follow Logical Relations in Skills

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:38:43.949998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:42.420308Z digest=sha256:a775e32852e21d95bb6614932077adbc99bfee30bb4d3544852e215185fcbec6

Observation 36817ce0-c3c4-4a23-871f-a55c30c2e68d · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Instruction-Following Evaluation for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.552670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.552670Z digest=sha256:48a8c328a5b7ec1457b6261e2ec1bced8c4aa3828eb433b192716b8db8993700

Observation 606233ca-dde7-41df-8d6f-d25fc75a3a00 · outbound

This paper cites Followbench: A multi-level fine-grained constraints following benchmark for large language models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Followbench: A multi-level fine-grained constraints following benchmark for large language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.246312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:42.639788Z digest=sha256:c11ffe4d9d761649a97317a96e9dee2c17b6cfccd4cca8e54a6ab291e9c00474

Observation d2ce020d-3290-40f0-86f7-c18c856f7429 · outbound

This paper cites Infobench: Evaluating instruction following ability in large language models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Infobench: Evaluating instruction following ability in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.724332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.724332Z digest=sha256:094abe4ca66b5b46d87840492d9292cc0da9cf2c012be09eb130c432834dc381

Observation cf969f85-d06d-4fc5-b477-7d19481f3089 · outbound

This paper cites Benchmarking complex instruction-following with multiple constraints composition.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Benchmarking complex instruction-following with multiple constraints composition

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.007953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:42.800877Z digest=sha256:0aab2352a25150c0b233a228d5b7a99ab4284f24e2a20c4c14e4e0a6e63a92a2

Observation 07c56e30-fdd0-4fbd-bf28-edf6934c4a1c · outbound

This paper cites Agentif: Benchmarking large language models instruction following ability in agentic scenarios.Advances in Neural Information Processing Systems, 38, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Agentif: Benchmarking large language models instruction following ability in agentic scenarios.Advances in Neural Information Processing Systems, 38, 2026

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:45.742772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:42.897685Z digest=sha256:a1ccc68c60f79b2871b819dabf6e58aa25a1face0d3eaf59dd265f043bc47549

Observation c2026e00-7ded-4a48-b435-ab6447975160 · outbound

This paper cites Stop Comparing LLM Agents Without Disclosing the Harness.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Stop Comparing LLM Agents Without Disclosing the Harness

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.965973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.965973Z digest=sha256:ec2fb4b8ccbc9368ec8f2a803123b445ea7bc320824df0f93a1c843f8ef8f620

Observation 395af4c8-0ea7-40dd-9ffb-33525113c553 · outbound

This paper cites Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.056799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.056799Z digest=sha256:60e69a341733d495175cce8f149c4d973b385d25948a41a24217c733d707eb6e

Observation 3470191b-dca1-4e18-8fc5-7d464a915668 · outbound

This paper cites Sysbench: Can llms follow system message? InThe Thirteenth International Conference on Learning Representations, 2024.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Sysbench: Can llms follow system message? InThe Thirteenth International Conference on Learning Representations, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:45.568923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:43.120628Z digest=sha256:73f714be9d9fd5c017aab756aa2e673049b36d6b8855f7fe4c0380ecb70f6574

Observation a5877146-a186-4571-8180-498013d79e94 · outbound

This paper cites Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.205719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.205719Z digest=sha256:3237d41d582a05af8c7110c8fbf7ac527f7fca6e7e0e38805320bebce6d1840e

Observation aa05d250-1d29-4d0f-89e4-cd52dd829559 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.274595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.274595Z digest=sha256:7d50ba97616b38c33bfd4dc9c84914a266d15afc3c80db7b81c5c3a191b4d283

Observation 581b78c4-ada0-4025-b4bf-4daf605c7e84 · outbound

This paper cites SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.353203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.353203Z digest=sha256:ee51b4b154010bc74f0204b49ccec170b54a4c60d691c1f24f682fb6b9c146e6

Observation 4200e5b6-42e8-4fd7-90d2-5fa0a37f6bd6 · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:45.416713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:43.436668Z digest=sha256:68489c75e8d5bfd5e83b564b66078d0a48387c02486d28f7b8faa03baa16bf64

Observation 486e0e27-5156-4bdc-a184-fe574b52ac5a · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:45.104227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:43.511861Z digest=sha256:c45598c68bd65997bfcaed6f1ed7cb053d00143ffb0c9556c17e7074e03649f3

Observation b6bc6ce9-2d16-4f92-a12c-2763c6f3590f · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:44.786035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:43.586596Z digest=sha256:39efd7f6c04286bbfc220bd91e0b5c5e4373576d14b2ca8fd55e3a8e1eb817f1

Observation 115f6e3a-ae6c-4593-bd56-6b6c61ec52f1 · outbound

This paper cites pdftk-server.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? pdftk-server

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:44.514436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:38:43.670203Z digest=sha256:28583beb517f45dda9113d412984bd181754b57ae17a8aafa76e3fc73b2f31ca

Pith citing papers

No inbound Pith citation observations are available.