Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T11:44:59.786646Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2502.07747.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T11:44:59.786646Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T17:57:23.282052Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T05:46:09.807404Z
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 39b2c24b-2545-4049-b454-603a1a4f1cb4 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5e90bf2-a2d1-41c1-943e-4a42362d8876 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 50b29953-3905-433e-9a69-512a0e5c8c4e · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Language Models are Few-Shot Learners
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 010dfd47-a8f6-406d-bebb-f6b29bc53777 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Measuring Massive Multitask Language Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 091c956d-2a2b-46be-bd39-1828c1575ef4 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 600c0ef7-64ef-43bd-9648-089d3d84fe33 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ccb9e5d-77dc-4605-97b2-3ff4ae39264b · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Holistic Evaluation of Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f06e9fa-0bdd-4c18-bd1b-7e6a95367a23 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9df960a-91e3-40da-8691-2bce538e669e · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ea275fd5-fdff-4702-9748-9a4287871438 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0f75bec-e159-4903-a39a-d7db722d934f · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories PlotMachines: Outline-Conditioned Generation with Dynamic Plot State Tracking
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f087741-f23f-44b2-91d8-e6093f350b0f · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f26113ed-0664-4453-8dee-31c96fbbe350 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 494b9779-6d1e-48d0-8252-c4273c606e7a · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1137e2f-cc6c-44c3-9429-dd16f6a4aebf · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db0776ab-3549-4225-b694-1d124fb1f7d8 · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92941092-748c-4246-9086-a0422cb9996c · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories online" 'onlinestring :=
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c46f2b9-10f3-40b8-bfb5-19a543093cba · outbound
WHODUNIT: Evaluation benchmark for culprit detection in mystery stories write newline
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5195f490-acfd-420c-9c88-98dbe2a01f82 · inbound
Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs WHODUNIT: Evaluation benchmark for culprit detection in mystery stories
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.