Pith. sign in

Paper Citation Record · LEDGER

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play

As of 15 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.19523.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19523 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:36:12.843668Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation afce5e30-3810-43fd-b616-0e2742d6513d · outbound

This paper cites GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.267086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.267086Z digest=sha256:acf13a430e815ab46254259a8ab302fc69cdda6433c4b4e22963fc04a6c5f66c

Observation 0e6af7b0-3c11-4fed-be2a-8c01d6caec69 · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.939373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.939373Z digest=sha256:4891c867b87012a0a58beac862d5adb548d5c2a40cb45ff453f3332456e87237

Observation f89f6fef-cf17-449d-b882-b044d01dd1b0 · outbound

This paper cites MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.078426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.078426Z digest=sha256:b4d21dff1d0b3b7dc574ee986b79b65340a7ffa6f93fbac66415e1e39a54fa28

Observation c829a6af-1648-4c32-b739-84b78005ebf6 · outbound

This paper cites Understanding the Effects of RLHF on LLM Generalisation and Diversity.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Understanding the Effects of RLHF on LLM Generalisation and Diversity

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.245333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.245333Z digest=sha256:b7ff8244ca2aab6ec86bb5270d8268cf1bef4c7f9789921fa728a0e4b09f963d

Observation 16923aae-33af-4e0a-b388-7133ba93204c · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.470843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.470843Z digest=sha256:48e708dc846fc7c13990ecdf2a6489c6b72178484655b962d00cee3ae7ed72b9

Observation f0ece43e-d7be-4773-9af5-a08a25aaca6a · outbound

This paper cites One fish, two fish, but not the whole sea: Alignment reduces language models’ conceptual diversity.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play One fish, two fish, but not the whole sea: Alignment reduces language models’ conceptual diversity

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.887984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.887984Z digest=sha256:8acf3801703b498e982425e7467b4a92d8bbbdc681a559bacf2e74a3d76f4195

Observation fd6c8850-31ea-49cf-8869-08b1b3f52da3 · outbound

This paper cites Do llm agents have regret? a case study in online learning and games.arXiv preprint arXiv:2403.16843,.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Do llm agents have regret? a case study in online learning and games.arXiv preprint arXiv:2403.16843,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.113398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.113398Z digest=sha256:c1bb0bccd3c3287dbd694b80504585839684684365142726ef2136d052a36ad3

Observation 62ae9cd7-9a10-4c82-ba49-9080037c237c · outbound

This paper cites Offline Learning of Controllable Diverse Behaviors.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Offline Learning of Controllable Diverse Behaviors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.177892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.177892Z digest=sha256:bfa88441eec07562a848a9c5881981eea4b9d73c41a1e074f97c5211a67a13cc

Observation 347ec6d5-84a2-47d0-b02c-4a838aa54c0a · outbound

This paper cites Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.257522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.257522Z digest=sha256:900670c533f3795079e690cf8d6eb888dd725964e46d7aad959fd19418bd41dd

Observation 362703cf-7d95-4f44-b663-1f445fe8c7fa · outbound

This paper cites Chessqa: Evaluating large language models for chess understanding.arXiv preprint arXiv:2510.23948,.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Chessqa: Evaluating large language models for chess understanding.arXiv preprint arXiv:2510.23948,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.357853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.357853Z digest=sha256:67720fc576cc07eb7e94a42b2e7886491c040eeb31159212979ec8bb7e39acb0

Observation 1934517f-cac9-4a64-9ac3-74851e6dce68 · outbound

This paper cites AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.420962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.420962Z digest=sha256:4a0423e5ff2ac0af81e347278888f8d48df12acc683f1c24f262829c458f4e65

Observation f5e532c6-4ab3-43d1-9178-076df8ed43f8 · outbound

This paper cites Qwen3 Technical Report.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Qwen3 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.503627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.503627Z digest=sha256:105329dfbdf4d7f734e1f5c0da793e66f08a2f3939f2e02bb07f4bc4704cfe33

Observation 843f753f-d335-4f27-b7c2-eeeffabd27cc · outbound

This paper cites How to Leverage Diverse Demonstrations in Offline Imitation Learning.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play How to Leverage Diverse Demonstrations in Offline Imitation Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.597503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.597503Z digest=sha256:bce0e27fb2a0c5858e62d86c4fb7e91db0530f0ac06c78317c3a082d829754c5

Observation 546ef691-ef62-4252-9512-6728fa6add49 · outbound

This paper cites The Price of Format: Diversity Collapse in LLMs.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play The Price of Format: Diversity Collapse in LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.681644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.681644Z digest=sha256:7bdd5f2cbd8665b0d69260122492dda0a1c7e5d256178a7e87c26e7e4065adef

Observation dd8517d9-118c-4693-b629-f7ae8f4bfd2a · outbound

This paper cites Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.755535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.755535Z digest=sha256:22d047f1ce7e6d47ff310af6f4c1e29f6a100fa1ef57bfab2f7361a41ae391b6

Observation 2e02c846-5d54-4d8e-a633-59e6e566deef · outbound

This paper cites move": <action_label>}</action> The JSON key must be.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play move": <action_label>}</action> The JSON key must be

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.843668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.843668Z digest=sha256:1b0c22c5574a5105e4e8193fca5aa693debc4c88aed039a5aea0917ba9cddc14

Observation e8cb85e9-1b8f-4505-bafe-81e03d552cd5 · outbound

This paper cites RvS: What is Essential for Offline RL via Supervised Learning?.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play RvS: What is Essential for Offline RL via Supervised Learning?

Reference 1978

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.784085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.784085Z digest=sha256:f4a37c97e6138f54f399560e1fbda4a6f4460a8c4325e5f36a1135a835cf2ac6

Observation f85fdf8e-bbee-47d3-9028-4bcd64ee1471 · outbound

This paper cites Preserving Diversity in Supervised Fine-Tuning of Large Language Models.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Preserving Diversity in Supervised Fine-Tuning of Large Language Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.617634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.617634Z digest=sha256:edcd0a15758c9a029c5f8f7110e182d2f993eedbb0b675289dc68175c546cdb6

Observation d3c6f4a9-5b75-4101-be29-bcf99b26ebca · outbound

This paper cites Sed-sft: Selectively encouraging diversity in supervised fine-tuning.arXiv preprint arXiv:2602.07464,.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Sed-sft: Selectively encouraging diversity in supervised fine-tuning.arXiv preprint arXiv:2602.07464,

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.487813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.487813Z digest=sha256:4d09599791023e7f9838c1334dbea6bacadab8f95fd149f65e7c1929b56ab941

Observation 57a024b7-7231-4e20-8eba-43a32a9c4749 · outbound

This paper cites Attributing mode collapse in the fine-tuning of large language models.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Attributing mode collapse in the fine-tuning of large language models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.023276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.023276Z digest=sha256:91131a46208ddc43ad49bb0931991736e05c768accc303b94d811a13807734b0

Observation cc0bde1d-6aab-476d-853a-41364ac8003d · outbound

This paper cites Llm chess: Benchmarking reasoning and instruction- following in llms through chess.arXiv preprint arXiv:2512.01992,.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Llm chess: Benchmarking reasoning and instruction- following in llms through chess.arXiv preprint arXiv:2512.01992,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.356713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.356713Z digest=sha256:765ff3b544e52108329e1d9959317782c5f40a25ae2b274db73ef750cbd67f0f

Observation 8d203cf8-07a4-4a66-a48a-7c24f94d052e · outbound

This paper cites TTT-Bench: A Benchmark for Evaluating Reasoning Ability with Simple and Novel Tic-Tac-Toe-style Games.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play TTT-Bench: A Benchmark for Evaluating Reasoning Ability with Simple and Novel Tic-Tac-Toe-style Games

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:11.742193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:11.742193Z digest=sha256:ebd45cede920a58da4058e04da0a775e5f8c2cf4391f55fdaf43590f469ea90e

Observation e6b0c72c-0b3c-4188-9c25-0e11b169e95f · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.373328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.373328Z digest=sha256:df7886d67468fe39d8ca4a58f2503e4bca0ec2dd67a07808cde31a3869579749

Observation 3c5fc3b4-ae13-44a5-94e6-d153a273df12 · outbound

This paper cites Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.643992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.643992Z digest=sha256:3783fd2e500a58b8c37066bde8f00c198e5e7345530fcfd3c68f84e44cff6cdb

Pith citing papers

No inbound Pith citation observations are available.