Pith. sign in

Paper Citation Record · LEDGER

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

As of 9 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2608.03206.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03206 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:37:29.464625Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ef166bef-08fa-45f1-a334-15f164d3253e · outbound

This paper cites arXiv:2602.10620.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners arXiv:2602.10620

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-05T23:37:29.634063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T23:37:29.433527Z digest=sha256:346932cb3f125425fca0529faf17cc15452e333f8a74ef273c6d86d891c9343a

Observation cd85d86b-ccbc-4766-9cec-72aead192b20 · outbound

This paper cites Kargupta,P.;Agarwal,I.;Hakkani-Tur,D.;andHan,J.2024.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Kargupta,P.;Agarwal,I.;Hakkani-Tur,D.;andHan,J.2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.688705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T23:37:29.436887Z digest=sha256:3d40e5ce8a792d46dbec415c95d55d44d4bfb4659becb0fc7f3b0e75d0e85076

Observation 673b6fae-be06-4ade-9a5e-03492290a0e5 · outbound

This paper cites Evaluating Gemini in an arena for learning.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Evaluating Gemini in an arena for learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.439716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.439716Z digest=sha256:1407a5ce0fa7503e14476f2241121bdbf39396f1ced4065ddd597a972e63e35b

Observation 9c015877-7920-4b9c-ba7a-c5903cad42ad · outbound

This paper cites arXiv:2601.14560.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners arXiv:2601.14560

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.442724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.442724Z digest=sha256:c4f496f004668d57855d77747bd5d1ab4c71827d772a84375a48452354e87400

Observation 4fadc8cb-ef1e-41ad-9769-7b8dacbc0cd3 · outbound

This paper cites Automated Knowledge Concept Annotation and Question Representation Learning for Knowledge Tracing.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Automated Knowledge Concept Annotation and Question Representation Learning for Knowledge Tracing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.448548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.448548Z digest=sha256:260ca554ed0261cd8a0f8945e935255e79ef1a8817739126fc36a3af57a88906

Observation 8dca50ee-8600-45e7-9ded-bd3b2e85f451 · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2024, 13641–13650.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners In Findings of the Association for Computational Linguistics: EMNLP 2024, 13641–13650

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.663773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T23:37:29.454458Z digest=sha256:6c623a2c18280fb222dcfaab02f0e17db8306e618efd730c08cfd639e7527f16

Observation e4ea5afa-e66c-408a-8615-e12908a76c01 · outbound

This paper cites Tutor CoPilot: A Human-AI Approach for Scaling Real-Time Expertise.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Tutor CoPilot: A Human-AI Approach for Scaling Real-Time Expertise

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.457145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.457145Z digest=sha256:a2e24bf7723eb91a7b0a23e6b54dce64330228eb7e95c4fffb61b20d0b092746

Observation e4519630-249f-437e-b2a8-7477f12ecc31 · outbound

This paper cites DeepTutor: Towards Agentic Personalized Tutoring.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners DeepTutor: Towards Agentic Personalized Tutoring

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.460581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.460581Z digest=sha256:1ab636a22e6f04917a5ff3fe62b4fe27460c4a58167221c84ca32c6c1f3e4624

Observation 09951edd-987c-46e1-b22c-7b7ba34034d8 · outbound

This paper cites A Unified Framework for the Evaluation of LLM Agentic Capabilities.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners A Unified Framework for the Evaluation of LLM Agentic Capabilities

Reference 13

Resolution
malformed identifier
no resolver link, observed 2026-08-05T23:37:29.464625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.464625Z digest=sha256:03615331b9d43db7091ba69b730497edcadbce6dd80b263fbe795628e1ab58a4

Observation df34748d-eaf8-464e-b968-449083be60a8 · outbound

This paper cites Sim- ulatedStudentsinTutoringDialogues:SubstanceorIllusion? InProceedingsofthe64thAnnualMeetingoftheAssociation for Computational Linguistics (ACL).

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Sim- ulatedStudentsinTutoringDialogues:SubstanceorIllusion? InProceedingsofthe64thAnnualMeetingoftheAssociation for Computational Linguistics (ACL)

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.672931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T23:37:29.451667Z digest=sha256:7dd5d9798b677400e869bc511bbd65327b60feeb359825fbe24bd8eab2292d67

Observation 394ebb12-1691-414f-a5b4-5df8bb4b31c8 · outbound

This paper cites InInternationalConferenceon Learning Representations (ICLR).

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners InInternationalConferenceon Learning Representations (ICLR)

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:37:29.680750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T23:37:29.445962Z digest=sha256:dcd873e98c6727b392e4f8e4937632b60fa49fe1aa78bf4ba0d3026df22cb19b

Observation 80ad745e-3042-4085-859e-0bbbbdb9a0ac · outbound

This paper cites Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T23:37:29.426608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:37:29.426608Z digest=sha256:98a2abc8fb963d6b5380002b207e2f090a409095959482190b473c26cec5261f

Observation 025d87d5-3c05-48ac-b1d3-deb9dc49f8e7 · outbound

This paper cites Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators.

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators

Reference 2026

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T23:37:29.648290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T23:37:29.430382Z digest=sha256:a08a9fbc4628a3ff4d938b0a8bf4bf9d7bfcf64831c46e480d2945791abf6cda

Pith citing papers

No inbound Pith citation observations are available.