Pith. sign in

Paper Citation Record · LEDGER

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

As of 12 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2608.03206.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03206 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:37:29.464625Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ef166bef-08fa-45f1-a334-15f164d3253e · outbound

This paper cites arXiv:2602.10620.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners arXiv:2602.10620

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-05T23:37:29.634063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T23:37:29.433527Z digest=sha256:6d6b5badf4e41745789bdea133cd884c8b9f41193c9ff28fa9ccb5413e62f55b

Observation cd85d86b-ccbc-4766-9cec-72aead192b20 · outbound

This paper cites Kargupta,P.;Agarwal,I.;Hakkani-Tur,D.;andHan,J.2024.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Kargupta,P.;Agarwal,I.;Hakkani-Tur,D.;andHan,J.2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.688705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T23:37:29.436887Z digest=sha256:585ab152f6cc070975152ab806b40905d758ff9e62d2609a2e14860e7cc15db5

Observation 673b6fae-be06-4ade-9a5e-03492290a0e5 · outbound

This paper cites Evaluating Gemini in an arena for learning.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Evaluating Gemini in an arena for learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.439716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.439716Z digest=sha256:4f80b2a4c83ab2de43d833cfe87e6da9c0be26e019c561072a0ceb696f62c6a4

Observation 9c015877-7920-4b9c-ba7a-c5903cad42ad · outbound

This paper cites arXiv:2601.14560.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners arXiv:2601.14560

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.442724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.442724Z digest=sha256:0935af800603e347442625e45a31f2e508c74b380ae3e847cf7676e97b051f29

Observation 4fadc8cb-ef1e-41ad-9769-7b8dacbc0cd3 · outbound

This paper cites Automated Knowledge Concept Annotation and Question Representation Learning for Knowledge Tracing.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Automated Knowledge Concept Annotation and Question Representation Learning for Knowledge Tracing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.448548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.448548Z digest=sha256:218282e0df011b6d40e741443a77bb98854c537aca61b06b54b531555ccbc2c3

Observation 8dca50ee-8600-45e7-9ded-bd3b2e85f451 · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2024, 13641–13650.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners In Findings of the Association for Computational Linguistics: EMNLP 2024, 13641–13650

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.663773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T23:37:29.454458Z digest=sha256:784ab8338dfcfb0d6906da0dd6d3d6e9b76f24ab7f898ff0b01d91cb4923cb0d

Observation e4ea5afa-e66c-408a-8615-e12908a76c01 · outbound

This paper cites Tutor CoPilot: A Human-AI Approach for Scaling Real-Time Expertise.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Tutor CoPilot: A Human-AI Approach for Scaling Real-Time Expertise

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.457145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.457145Z digest=sha256:e155346372125aad359f9ea92c55fcc8b2475f5cb59ae397d470c8cbcfc3c2ee

Observation e4519630-249f-437e-b2a8-7477f12ecc31 · outbound

This paper cites DeepTutor: Towards Agentic Personalized Tutoring.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners DeepTutor: Towards Agentic Personalized Tutoring

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.460581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.460581Z digest=sha256:1af3e70de7b896321bd1f49848512f5c573a3e9a4df393ab33e08b7b00634a07

Observation 09951edd-987c-46e1-b22c-7b7ba34034d8 · outbound

This paper cites A Unified Framework for the Evaluation of LLM Agentic Capabilities.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners A Unified Framework for the Evaluation of LLM Agentic Capabilities

Reference 13

Resolution
malformed identifier
no resolver link, observed 2026-08-05T23:37:29.464625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.464625Z digest=sha256:ae0ade6add589c821f3ea49e2a2457d71a3e32b3459c28a769d0c734caaf97b8

Observation df34748d-eaf8-464e-b968-449083be60a8 · outbound

This paper cites Sim- ulatedStudentsinTutoringDialogues:SubstanceorIllusion? InProceedingsofthe64thAnnualMeetingoftheAssociation for Computational Linguistics (ACL).

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Sim- ulatedStudentsinTutoringDialogues:SubstanceorIllusion? InProceedingsofthe64thAnnualMeetingoftheAssociation for Computational Linguistics (ACL)

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.672931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T23:37:29.451667Z digest=sha256:f9fe84c975eea4a517b56f73d6b88cd6c08dd3a523d26713a6a3ac24d08133be

Observation 394ebb12-1691-414f-a5b4-5df8bb4b31c8 · outbound

This paper cites InInternationalConferenceon Learning Representations (ICLR).

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners InInternationalConferenceon Learning Representations (ICLR)

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.680750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T23:37:29.445962Z digest=sha256:027bdb47640cfa631a5c97a1cd3111f3fde797ab6807775e7cc8c78e02114e55

Observation 80ad745e-3042-4085-859e-0bbbbdb9a0ac · outbound

This paper cites Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.426608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.426608Z digest=sha256:634478d4eda8d4dec91920ba509e82d7e163aa2287aea0769974bdd0ff59b3d1

Observation 025d87d5-3c05-48ac-b1d3-deb9dc49f8e7 · outbound

This paper cites Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators

Reference 2026

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T23:37:29.648290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-05T23:37:29.430382Z digest=sha256:862cb921154f38b4ee9eaf9b4e98d29dde4279f52f93a5f25738c92e94e023b2

Pith citing papers

No inbound Pith citation observations are available.