Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:08:14.068241Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 4 inbound Pith citation observations for arXiv:2602.05843.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:08:14.068241Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T01:10:21.243699Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T13:29:51.596096Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation abcf9bb2-67e3-4b98-a511-9cdc52c13b92 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07be5ed1-70db-4ffe-b7a9-17b09f240e50 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions J., Bethge, M., and Schulz, E
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94800fc7-0706-4bf2-8fb4-90643d4fc806 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions The claude 3 model family: Opus, sonnet, haiku
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57ffcf1c-e6a2-4347-b8e3-5696591186d4 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions H., and Bengio, Y
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16f03f94-b2e8-4fb0-b521-7d6d0098e075 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions On the Measure of Intelligence
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fe0ebb9-7175-4866-a10c-225c9e067939 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Evaluating long-context reasoning in llm-based webagents
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be7fd5c4-2ec1-406e-a03b-57bd71bd983c · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions The dynamical challenge
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62382c70-fc00-4411-b000-1b900451b63a · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Mind2web: Towards a generalist agent for the web
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44797934-6322-4048-b2ff-eb7641e983e9 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af6b907a-38c2-4a0f-be18-bc1e33218f2f · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e10462f-7d50-4565-9820-223e71e219ff · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions The Llama 3 Herd of Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed61fcd1-4de9-4f12-b221-08dfed5f1429 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions and Schmidhuber, J
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cec0ab3-7747-4553-af99-c627e6a1a99b · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions H., Gonzalez, J
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff668df-7ab4-408c-8124-43c317ff134d · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions M., Ullman, T
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ac8cd97-6d31-4f57-9d06-5cf1245b2007 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions State space models on temporal graphs: A first-principles study
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cd6a435-f4ba-42ab-b273-ca7db33800c4 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Y., Le Bras, R., Richardson, K., Sabharwal, A., Poovendran, R., Clark, P., and Choi, Y
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1852803b-9dba-413b-b13f-e8e029d63b47 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c604330b-a965-4290-9c52-a8cef996ec79 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Agentbench: Evaluating llms as agents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfa35dbd-6872-4218-ac85-b64b4971485e · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Gaia: a benchmark for general ai assistants
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85a10ad0-493f-4700-91ed-7e7dd255d1c3 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions gpt-oss-120b & gpt-oss-20b Model Card
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23bf426b-6ed8-4bef-ac19-2290650ab49b · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions G., Mao, H., Yan, F., Ji, C
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b09a2020-4b4a-47e9-954c-030b1dd6c2b7 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions E., Li, W., Campbell-Ajala, F., Toyama, D
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a45872a-b04e-43aa-ac18-15242df16764 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Reflexion: Language agents with verbal reinforcement learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6e9b9cd-6bbb-496e-b748-f03955026ff5 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Alfworld: Aligning text and embodied environments for interactive learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 631d89fc-a2a0-4684-9021-44f972d26f30 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Corex: Pushing the Boundaries of Complex Reasoning through Multi-Model Collaboration
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ba0fedd-aab4-41ee-9055-510c069421ca · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Os-genesis: Automating gui agent trajectory construction via reverse task synthesis
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b568cec5-36ea-4275-8471-e1aaa957c2bd · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5457a67c-5fef-457e-a935-c1c5ea5a7185 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Mars: Situated inductive reasoning in an open-world environment
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5db826d9-a19c-4df5-8b05-6c831e77eb1d · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Michelangelo: Long context evaluations beyond haystacks via latent structure queries
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b861cdd6-77b0-4f70-816b-14f279c5dd1e · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Large language models for robotics: Opportunities, challenges, and perspectives
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5072d6c0-9b4a-4c27-8feb-cd934da8508e · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions OdysseyBench: Evaluating LLM Agents on Long-Horizon Complex Office Application Workflows
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dc572a8-0df3-451e-b9ac-a00441bee7b8 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d64ee9c-600e-4045-91ef-3e82ee3fccb2 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions J., Cheng, Z., Shin, D., Lei, F., et al
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a7aa4d-0634-4392-a1d1-38bc028e9b73 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions -decoding: Adaptive foresight sampling for balanced inference-time exploration and exploitation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8845d5c2-7533-412e-9d91-966758b90c40 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Genius: A generalizable and purely unsupervised self-training framework for advanced reasoning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 75d7f364-33cf-4678-b519-71aa044782ed · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions F., Song, Y., Li, B., Tang, Y., Jain, K., Bao, M., Wang, Z
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd8d17e4-1205-4db8-a761-618e5d003638 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Tide: Trajectory-based diagnostic evaluation of test-time improvement in llm agents
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e1afbbe-60b3-4e4a-b98c-8b4a7d61c8ff · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Qwen3 Technical Report
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a70c3329-8764-4e79-8ab6-3bf806c448e3 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions R., and Cao, Y
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fcf34c6-6b1a-4498-9210-432e801688f0 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 610b432b-f0ae-41c2-9860-b129f41b0f91 · outbound
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions F., Zhu, H., Zhou, X., Lo, R., Sridhar, A., Cheng, X., Ou, T., Bisk, Y., Fried, D., et al
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1605ec29-59e7-4226-a88a-3b8ec39d37e8 · inbound
Data-Driven Boundary Control of Distributed Port-Hamiltonian Systems OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21c0901f-91b2-4d8e-9d21-c68fd5fcdc63 · inbound
OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation af3179df-4eb4-4276-ac8f-2bcd261284e6 · inbound
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e53c8bbd-d6b2-4c87-a4c7-57212f0f7113 · inbound
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.