Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T12:36:12.843668Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.19523.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T12:36:12.843668Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation afce5e30-3810-43fd-b616-0e2742d6513d · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e6af7b0-3c11-4fed-be2a-8c01d6caec69 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f89f6fef-cf17-449d-b882-b044d01dd1b0 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c829a6af-1648-4c32-b739-84b78005ebf6 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Understanding the Effects of RLHF on LLM Generalisation and Diversity
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16923aae-33af-4e0a-b388-7133ba93204c · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0ece43e-d7be-4773-9af5-a08a25aaca6a · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play One fish, two fish, but not the whole sea: Alignment reduces language models’ conceptual diversity
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd6c8850-31ea-49cf-8869-08b1b3f52da3 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Do llm agents have regret? a case study in online learning and games.arXiv preprint arXiv:2403.16843,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62ae9cd7-9a10-4c82-ba49-9080037c237c · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Offline Learning of Controllable Diverse Behaviors
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347ec6d5-84a2-47d0-b02c-4a838aa54c0a · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 362703cf-7d95-4f44-b663-1f445fe8c7fa · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Chessqa: Evaluating large language models for chess understanding.arXiv preprint arXiv:2510.23948,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1934517f-cac9-4a64-9ac3-74851e6dce68 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5e532c6-4ab3-43d1-9178-076df8ed43f8 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Qwen3 Technical Report
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 843f753f-d335-4f27-b7c2-eeeffabd27cc · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play How to Leverage Diverse Demonstrations in Offline Imitation Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 546ef691-ef62-4252-9512-6728fa6add49 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play The Price of Format: Diversity Collapse in LLMs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd8517d9-118c-4693-b629-f7ae8f4bfd2a · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e02c846-5d54-4d8e-a633-59e6e566deef · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play move": <action_label>}</action> The JSON key must be
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8cb85e9-1b8f-4505-bafe-81e03d552cd5 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play RvS: What is Essential for Offline RL via Supervised Learning?
Reference 1978
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85fdf8e-bbee-47d3-9028-4bcd64ee1471 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Preserving Diversity in Supervised Fine-Tuning of Large Language Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3c6f4a9-5b75-4101-be29-bcf99b26ebca · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Sed-sft: Selectively encouraging diversity in supervised fine-tuning.arXiv preprint arXiv:2602.07464,
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57a024b7-7231-4e20-8eba-43a32a9c4749 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Attributing mode collapse in the fine-tuning of large language models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc0bde1d-6aab-476d-853a-41364ac8003d · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Llm chess: Benchmarking reasoning and instruction- following in llms through chess.arXiv preprint arXiv:2512.01992,
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d203cf8-07a4-4a66-a48a-7c24f94d052e · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play TTT-Bench: A Benchmark for Evaluating Reasoning Ability with Simple and Novel Tic-Tac-Toe-style Games
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6b0c72c-0b3c-4188-9c25-0e11b169e95f · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c5fc3b4-ae13-44a5-94e6-d153a273df12 · outbound
When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.