Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:32:33.689457Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 7 inbound Pith citation observations for arXiv:2507.23701.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:32:33.689457Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T21:20:37.943280Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T19:55:01.576018Z
16 of 16 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9cc5aa1a-e6e6-45a3-b825-0342e541670f · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? This is a fake!
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2c8caff6-8dcb-45c4-9c85-80b1ad948731 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9a2ed50-0cf9-4000-8aeb-05c722437768 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Interactive Fiction Games: A Colossal Adventure
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6931e197-79b0-4ee8-955f-8f83aa2a1dbf · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b427c8d8-807f-4f2c-8c40-6eb97c5780bf · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Humanity's Last Exam
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d032d437-81f3-4623-9314-d492177e97ed · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a513bc0-9837-42b3-8f72-4faf142f9ccd · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf4fa4a3-6d43-4ab8-b363-b39db566f4f3 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? PaperBench: Evaluating AI's Ability to Replicate AI Research
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a53f149a-7e84-4e6d-a52a-cb9a4b22006b · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf6e84a8-dbe5-4d56-bff1-3a6c33cbfc4b · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Keep CALM and Explore: Language Models for Action Generation in Text-based Games
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7ad646c-81c8-41ae-b2fd-771f9ec33677 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 936c68e0-6504-4eca-9798-42321d74d541 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Unresolved cited work
Reference 1983
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9dc0c724-c700-4f6c-81ad-33c8d9d18991 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Graph Constrained Reinforcement Learning for Natural Language Action Spaces
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f1e4e1fa-c232-434e-93ac-b54be87693bb · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? GAIA: a benchmark for General AI Assistants
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38731489-54f2-4d90-a894-9f3335f618d0 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d61648b-c074-4e2d-b1ed-e04f00aabbe7 · outbound
TextQuests: How Good are LLMs at Text-Based Video Games? Prithviraj Ammanabrolu and Matthew Hausknecht
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45cd63f8-f908-43a4-b57f-8de2389e77a0 · inbound
Mini Amusement Parks (MAPs): A Testbed for Modelling Business Decisions TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e43c9bd-e4ab-4596-ace1-ce68d015e3ad · inbound
RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6f946362-47eb-4357-acd2-754e05e1d385 · inbound
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 133
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 18599c8e-3c0d-43e6-829e-f524f7c865e1 · inbound
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 133
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eb0af504-3f2e-4f4a-95c6-464ffaad9468 · inbound
Muse Spark Safety & Preparedness Report TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dfdc32e5-0e37-4fb2-8617-9d77630073e2 · inbound
SysAdmin: Measuring Instrumental Power-Seeking in Frontier AI TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4489fd4-0148-4f49-af23-cc7ebf6ddca2 · inbound
Rushes: A Human Preference Dataset for Pluralistic Alignment TextQuests: How Good are LLMs at Text-Based Video Games?
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.