Pith. sign in

Paper Citation Record · LEDGER

Measuring General Intelligence with Generated Games

As of 18 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 8 inbound Pith citation observations for arXiv:2505.07215.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.07215 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:26:35.889537Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:27.682101Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:57.651515Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy19
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 133e8599-d7df-4aa6-9141-146e2ffcc246 · outbound

This paper cites ZeroSumEval: An Extensible Framework For Scaling LLM Evaluation with Inter-Model Competition.

Measuring General Intelligence with Generated Games ZeroSumEval: An Extensible Framework For Scaling LLM Evaluation with Inter-Model Competition

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.489506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.489506Z digest=sha256:271cdefaa70697bdbd81edc339388cee91b6903b7be63dfe90f9d5a8bb8c2b31

Observation 82079a04-8e5d-43e5-9dfa-7af5d1442578 · outbound

This paper cites Claude 3.7 Sonnet, 2025.

Measuring General Intelligence with Generated Games Claude 3.7 Sonnet, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.506665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.506665Z digest=sha256:6f9e385ab96617a7f7af099b074b1428e106d308b9d2a563cb249dee7020f176

Observation 845ea991-ad7e-4f58-8af1-96187795ec15 · outbound

This paper cites Claude’s extended thinking.

Measuring General Intelligence with Generated Games Claude’s extended thinking

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.519112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.523508Z digest=sha256:eb12c15ef9596909130ed5ea59a0f443b3e68e8fba8963ec102805411dd52ac3

Observation 3b4bc024-e831-4352-a1de-19ccf2d0282c · outbound

This paper cites OpenAI Gym.

Measuring General Intelligence with Generated Games OpenAI Gym

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.537991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.537991Z digest=sha256:0a8b0048800ba2f5a54e9a8aac27f2a2de8386f3ce38a8d967b200f9b56cdd68

Observation f7b1ff31-f3b0-4d0b-8bbf-5d8e4540ec7f · outbound

This paper cites Superhuman AI for heads-up no-limit poker: Libratus beats top professionals.

Measuring General Intelligence with Generated Games Superhuman AI for heads-up no-limit poker: Libratus beats top professionals

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.548471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.548471Z digest=sha256:18d7f939184c2e9870fd157b96fe23f37e8d562a1a77cd21b897616b511bd262

Observation edcf14da-ba3d-49af-b08f-f8e24c9045ab · outbound

This paper cites Sparks of artificial general intelligence: Early experiments with GPT-4, 2023.

Measuring General Intelligence with Generated Games Sparks of artificial general intelligence: Early experiments with GPT-4, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.504453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.566647Z digest=sha256:7397e84a14023cff5ccaf943dbe9f07e96dd93b03a735c2024c0e054b7b5b340

Observation e8282534-51cd-4299-92c3-af854f325a23 · outbound

This paper cites Heuristic DENDRAL: A program for generating explanatory hypotheses.

Measuring General Intelligence with Generated Games Heuristic DENDRAL: A program for generating explanatory hypotheses

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.489938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.581270Z digest=sha256:3f9ce3e3b58999ec47f90b291465059a3dcd30d8dd394f5723a36ab0f7821dc6

Observation 3420c9a6-765f-485c-ad5e-1d5ab49d9a5e · outbound

This paper cites Deep Blue.

Measuring General Intelligence with Generated Games Deep Blue

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.598304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.598304Z digest=sha256:b9530565e39c25eae6c2831c29467d5c14d8b9dbef0351c5ddb4606f3a9a41e7

Observation 28334878-f7da-44d7-89f7-4c4bfda6c1dc · outbound

This paper cites Gonzalez, and Ion Stoica.

Measuring General Intelligence with Generated Games Gonzalez, and Ion Stoica

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.476525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.612932Z digest=sha256:75879c32474369cb4188b3c0b46a1d07555161c0daf2275cbc79ce0dc616a4ac

Observation 79e745c4-d5c0-440f-b691-7ed24f7d51df · outbound

This paper cites On the Measure of Intelligence.

Measuring General Intelligence with Generated Games On the Measure of Intelligence

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.628071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.628071Z digest=sha256:377e56939e3e32eeb9cb90c83cca73200fd35f6d33cfc3ea9b4703806093e454

Observation 2f4c514a-cebf-47b7-beab-ad45ea9b2ad2 · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

Measuring General Intelligence with Generated Games GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.639697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.639697Z digest=sha256:1e07b65e4b4e7b17b7fae019aa20b1e14e1b5babf1cb942e7256abaceb94f1c5

Observation 34ba052c-0c8a-46fc-8fc9-db3779152ca3 · outbound

This paper cites Gemini 2.5 Pro.

Measuring General Intelligence with Generated Games Gemini 2.5 Pro

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.465379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.649380Z digest=sha256:0783d757245eb366014edef8cc3a403df4bb2149696552130a0f72b47fd61184

Observation 69e182af-982f-45bd-9b81-7e57f494bf53 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Measuring General Intelligence with Generated Games DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.660993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.660993Z digest=sha256:81e2ff45915b67b77169f32a2dc0333bb53b6609ea0b5c9f34749edfd10d17a8

Observation 3db7325f-08a5-4276-9e4a-ec84cdb8e82a · outbound

This paper cites PAL: Program-aided language models.

Measuring General Intelligence with Generated Games PAL: Program-aided language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.452602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.684060Z digest=sha256:d5c3cb199c29c82da57ca3db58e51880fd3b0367bdc0b8183ccfce36aac15350

Observation 47c64b66-d5ec-4df8-b129-a7afe683f40c · outbound

This paper cites Frames of mind: The theory of multiple intelligences.

Measuring General Intelligence with Generated Games Frames of mind: The theory of multiple intelligences

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.697079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.697079Z digest=sha256:e71897d7e7216fa698788c06c8c36c5d66c844dc8ee8607a501b217a4def0d76

Observation 8ebb7676-7722-4d7c-afc8-701bc237ef4c · outbound

This paper cites Artificial general intelligence, volume 2.

Measuring General Intelligence with Generated Games Artificial general intelligence, volume 2

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.432174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.708897Z digest=sha256:36373f571badfad214aaaee25e53a6f99eee8799fe49f42d890b7f2944a90771

Observation ddcf07fd-d745-4b67-b751-6659d7c321c4 · outbound

This paper cites Interactive Fiction Games: A Colossal Adventure.

Measuring General Intelligence with Generated Games Interactive Fiction Games: A Colossal Adventure

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.717693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.717693Z digest=sha256:9bf6460037f3b38b0183cb12c462de0b1f5614e90f743342150cb0739e5b6c5c

Observation 947e82d6-ebcf-45b3-ba18-9d5d936480f1 · outbound

This paper cites Deep Reinforcement Learning from Self-Play in Imperfect-Information Games.

Measuring General Intelligence with Generated Games Deep Reinforcement Learning from Self-Play in Imperfect-Information Games

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.728726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.728726Z digest=sha256:7896ce48f994fa9a7a300c67aecfdb6c4a04ee0ccc89caa2445347ad5b496a76

Observation 61aa489d-b5f7-476b-9e6b-cb5c8af5250c · outbound

This paper cites Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan.

Measuring General Intelligence with Generated Games Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.740114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.740114Z digest=sha256:5754b6c21354100bc038b3dd689dfcdbbf9b23fdaee159ad9b8e4a87dcbaaf81

Observation c84e510d-e256-4345-944e-79a2541e376c · outbound

This paper cites Dynabench: Rethinking benchmarking in NLP.

Measuring General Intelligence with Generated Games Dynabench: Rethinking benchmarking in NLP

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.757320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.757320Z digest=sha256:a6c82e8c90179669675bf9ce120db96b661b4dc6a417dc6ccd6dac92c64c7cd2

Observation 3885d6f0-5714-4022-81bd-711b16c4d1b1 · outbound

This paper cites A collection of definitions of intelligence.

Measuring General Intelligence with Generated Games A collection of definitions of intelligence

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.406377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.768293Z digest=sha256:b8ef9e16d1e3a36ab85916abe1d1f46f10bc0eb8534eaeb54c1482edc691295c

Observation 51f296a6-74eb-4af6-bb61-4541346caba1 · outbound

This paper cites Guha, Karen Pittman, Dexter Pratt, and Mary Shepherd.

Measuring General Intelligence with Generated Games Guha, Karen Pittman, Dexter Pratt, and Mary Shepherd

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.394631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.776253Z digest=sha256:1124fe2e0ebf0aa982b9b01c141ff63e609918aab7741f3d1941c9cf431e074a

Observation 96887fc7-ed0b-4a72-a771-7ced05b51b5f · outbound

This paper cites Discovering and exploring cases of educational source code plagiarism with Dolos.

Measuring General Intelligence with Generated Games Discovering and exploring cases of educational source code plagiarism with Dolos

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.786483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.786483Z digest=sha256:c5f32db05739f7da53b3467912d2a87da4d5193052e5bc9143e14dd5721bf131

Observation 98cfb84d-6afe-41fd-8412-34a3a79e79ac · outbound

This paper cites A proposal for the Dartmouth summer research project on artificial intelligence.

Measuring General Intelligence with Generated Games A proposal for the Dartmouth summer research project on artificial intelligence

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.382760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.794744Z digest=sha256:baab257fda45ce6684f80c38c4f03b6bd492336b8eedf8719e18bee0b7760d36

Observation b88f4be3-1701-4b9d-b8d5-b9bc9f6e060c · outbound

This paper cites Min, Yangruibo Ding, Luca Buratti, Saurabh Pujar, Gail Kaiser, Suman Jana, and Baishakhi Ray.

Measuring General Intelligence with Generated Games Min, Yangruibo Ding, Luca Buratti, Saurabh Pujar, Gail Kaiser, Suman Jana, and Baishakhi Ray

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.370710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.800822Z digest=sha256:37ae1da3e0ec45f680e4a1cd25848f3be208199b4a5867b277798304bf93b0e3

Observation ed9248ad-cf24-4055-a9a7-488dbfdc86cf · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Measuring General Intelligence with Generated Games Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.812305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.812305Z digest=sha256:ca2af80e640ed628317dcb094466ae37e715aa1407390478822575ff9d691b9a

Observation 631c8233-6385-4b61-bfe8-f8f97f6fa1ed · outbound

This paper cites an unresolved cited work.

Measuring General Intelligence with Generated Games Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T22:26:36.358833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.805962Z digest=sha256:be321cf1b958e6285c9464a483e894393403f5fde20d5a592413be0c38bf87e6

Observation 98f28dba-20d3-4bf2-a6a0-a3344ee36131 · outbound

This paper cites OpenAI o3-mini, 2025.

Measuring General Intelligence with Generated Games OpenAI o3-mini, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.341564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.819273Z digest=sha256:9d504b5433abc3d6c914f0bd6e31d293df0dd2a9a65a0807df82590e710bae75

Observation da876027-ae3e-4578-ac19-4c36f74bad14 · outbound

This paper cites Introducing OpenAI o1, 2024.

Measuring General Intelligence with Generated Games Introducing OpenAI o1, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.815909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.815909Z digest=sha256:4bd1dd12c43862a3523c700849abec1d03a976f6f64906e9be0f11a20a7719ae

Observation 5cb7fdb3-c202-4ae9-8543-357bf9347835 · outbound

This paper cites Stable-Baselines3: Reliable reinforcement learning implementations.

Measuring General Intelligence with Generated Games Stable-Baselines3: Reliable reinforcement learning implementations

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.320828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.824886Z digest=sha256:06eae9a777fc49d60c26fee3e794c786423280765c18f07c3789f8191b1936e5

Observation d20ee3b0-c84d-471e-bd14-975a4d86827c · outbound

This paper cites Introducing o3 and o4-mini.

Measuring General Intelligence with Generated Games Introducing o3 and o4-mini

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.331968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.821973Z digest=sha256:d0d5e6cda865521d363a0f95bd71170b79c2523e18199f2c7861fdc5a15917e3

Observation 2652d443-5376-4930-ad95-24aa8d52da50 · outbound

This paper cites Neural theory-of-mind? on the limits of social intelligence in large LMs.

Measuring General Intelligence with Generated Games Neural theory-of-mind? on the limits of social intelligence in large LMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.830303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.830303Z digest=sha256:eb306f076367d5bb94697126b38a3361b0b546a9f69cd36c9567f9a059e570c6

Observation e348cb2e-7f2a-4584-8af7-b488e2075061 · outbound

This paper cites Artificial intelligence: a modern approach.

Measuring General Intelligence with Generated Games Artificial intelligence: a modern approach

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.827657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.827657Z digest=sha256:75dc3c8f1808f129f117f1e3dab32089c7ea37c30afdf9161963d635de6356fa

Observation 8cc3e7c0-e5bd-4c65-9dd9-b3c2bbddea8d · outbound

This paper cites Winnowing: local algorithms for document fingerprinting.

Measuring General Intelligence with Generated Games Winnowing: local algorithms for document fingerprinting

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.292482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.841442Z digest=sha256:495f50408608194c2d648a323d344d1be0eb314c220e33be236250ecf7a4826f

Observation a25b3ad0-77af-4059-80e4-9f345d87d6b2 · outbound

This paper cites Measuring intelligence through games,.

Measuring General Intelligence with Generated Games Measuring intelligence through games,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.302077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.834094Z digest=sha256:cb0f6c75f05c8aaae58618ddffbc50fe8ceea995f9fdb66116942aa5b566d29b

Observation a849fd1f-40e6-4731-9d9d-db623daf4265 · outbound

This paper cites Reflexion: Language Agents with Verbal Reinforcement Learning.

Measuring General Intelligence with Generated Games Reflexion: Language Agents with Verbal Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.847681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.847681Z digest=sha256:af67d8a9e0b63a59bef2b485361ee2292945343558bff17ddc6a84db4759037d

Observation bea15f2c-9585-4482-9127-3e0e4bf17d35 · outbound

This paper cites Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm.

Measuring General Intelligence with Generated Games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.851037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.851037Z digest=sha256:75c3cb04f7212f00ea124b04124b3f1a4d982b27539f321af17cd0f3901926d8

Observation 25176ad3-5ebc-4a9d-a1a7-1fb309d6bd72 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Measuring General Intelligence with Generated Games Proximal Policy Optimization Algorithms

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.844596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.844596Z digest=sha256:be6d86384cef6102fcc121df4c834c2967950d8128363ab84ed6f5254ef59323

Observation 48ce2e6f-28ec-4a44-8a84-35d8b492ca1c · outbound

This paper cites What is intelligence?: Contemporary viewpoints on its nature and definition.

Measuring General Intelligence with Generated Games What is intelligence?: Contemporary viewpoints on its nature and definition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.282733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.858755Z digest=sha256:361514eb7239ddc7f3df439b6f8ceb1d85f990f8574bea076aa5e7cb7b5dfcb2

Observation 79ecaa26-cecd-4a4e-9f23-e79ff1b2db3b · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Measuring General Intelligence with Generated Games Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.861284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.861284Z digest=sha256:92aae115ddbc36cd2edb8d53fc015abf14f47fe1c1ac938200629e3ee46fc5a6

Observation 72bc769b-fbcc-4611-a983-30a97404d409 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

Measuring General Intelligence with Generated Games Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.854776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.854776Z digest=sha256:81ccb38ff0eeb27714b920135deab20d9b5277f9b33ed4262d165521c5da29c4

Observation b4467a53-3f86-40ca-8de0-e64dcd7e4956 · outbound

This paper cites Goal-driven explainable clustering via language descriptions.

Measuring General Intelligence with Generated Games Goal-driven explainable clustering via language descriptions

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.871979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.871979Z digest=sha256:a0fac5c03ecb861e7781f10ed52d78fd0cf9fd79c11408aec65827dd551ca627

Observation 4f5e4639-fb1e-4ee1-9f4c-e3ff492a5716 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Measuring General Intelligence with Generated Games Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.875812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.875812Z digest=sha256:ce8c91a94807f5bfd7fe174848ca23c85ed080dab25029acef72ce16577ecf6a

Observation c9ebdd1b-ee3b-4df1-b5ce-c38f2778643c · outbound

This paper cites Evaluating large language models with grid-based game competitions: An extensible LLM benchmark and leaderboard,.

Measuring General Intelligence with Generated Games Evaluating large language models with grid-based game competitions: An extensible LLM benchmark and leaderboard,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.272810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.865418Z digest=sha256:7570afdb0ae069289e996bd17ea82c68180d2185e91b42187124053ee0e0f938

Observation c198cd6a-ed09-4756-805f-507792166ad6 · outbound

This paper cites Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard.

Measuring General Intelligence with Generated Games Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.868797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.868797Z digest=sha256:bb64dea1e3bd9504092e73eb7e5f5d98cdc93a16105326c3a3eb0d952e94615d

Observation 27550da4-078f-4621-a11c-cd4768757aee · outbound

This paper cites VideoGameBench: Research preview.

Measuring General Intelligence with Generated Games VideoGameBench: Research preview

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.262560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:26:35.886314Z digest=sha256:55b7b0771d7d43d12673984f2e4db22c02a0f4fa15697cac85e48e1cc5de552a

Observation 45a4d9c4-12de-4af2-962a-1fbacc3f3087 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Measuring General Intelligence with Generated Games Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.889537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.889537Z digest=sha256:ce57298bf87929688e92f533b796e53ae91e850685e20fd1d0da6cf434ae6312

Observation fc0fea58-d58e-4ca3-8e0c-4935f5fc9427 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Measuring General Intelligence with Generated Games ReAct: Synergizing Reasoning and Acting in Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.879068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.879068Z digest=sha256:3b32348c9f2a885336e65492fb2c64094517abcfd0fbef5997b372c43ea6dff5

Observation ac653952-9d44-4123-b12c-066b25484ba9 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

Measuring General Intelligence with Generated Games $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.882198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.882198Z digest=sha256:c5ab63dfa54721fef7fd8dbc0b39fa4a38cca898fe3e62ee4f691c63be4720ce

Observation 34f20f2d-1ecf-4ba9-8fa5-5f66ca68afd1 · outbound

This paper cites Measuring Intelligence through Games.

Measuring General Intelligence with Generated Games Measuring Intelligence through Games

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.837771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.837771Z digest=sha256:0448c51f92b74be7757aa75fc45b3ceb0020435fd50fb2db0a290618c0f3765f

Observation bfa34b4b-eb16-4c96-a133-3bd34e367920 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Measuring General Intelligence with Generated Games SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.747667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.747667Z digest=sha256:f11d6a8f15395f57c72e0bd574cde4416e299e6e0e90ce235939425ade226462

Pith citing papers

Observation b380370e-9acc-4f3e-a302-b99f85dadf03 · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation Measuring General Intelligence with Generated Games

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.682101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.682101Z digest=sha256:3c87ba5ee1d21a59ce8a65e4d6619ae95fa56de0a16b6b23fa3df54347cdf25d

Observation ebacd626-1aae-45ad-a2f4-29465a40b284 · inbound

Assessing Adaptive World Models in Machines with Novel Games cites this paper.

Assessing Adaptive World Models in Machines with Novel Games Measuring General Intelligence with Generated Games

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T16:41:19.917526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:41:19.917526Z digest=sha256:b361a3922febf571b2bfff147987dd34f95d9b012957eaf769137b7c6c51816e

Observation bae068f2-831c-4d1c-82fb-0e578cd9484d · inbound

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment cites this paper.

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment Measuring General Intelligence with Generated Games

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:25:55.254384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T05:23:08.478393Z digest=sha256:bf5365c97cd356aff2c2f4a562997e3abab8b5d52ba4e09115a20f03c2b9bde9

Observation 00cc2af3-2f7d-4d80-b951-208dc667a35b · inbound

Scalable Environments Drive Generalizable Agents cites this paper.

Scalable Environments Drive Generalizable Agents Measuring General Intelligence with Generated Games

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:23:12.204325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T10:19:14.829125Z digest=sha256:f73e37a5ac0ee2d4a5c4e578cc076516a0de4308fa2907304a49b7c0cdeea551

Observation 40d95776-71ee-4c53-9e12-e1115671faa6 · inbound

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models cites this paper.

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models Measuring General Intelligence with Generated Games

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:45:20.673786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T04:41:35.532363Z digest=sha256:b6c12245e378040551bd47737d0552a860914c68deed7f8148f6e48b5647ef23

Observation 0b83c40d-3bbc-4f5b-ade5-7efab9cdaee3 · inbound

Distilling Game Code World Model Generation into Lightweight Large Language Models cites this paper.

Distilling Game Code World Model Generation into Lightweight Large Language Models Measuring General Intelligence with Generated Games

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.687987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T13:58:37.956333Z digest=sha256:e931762520f6c260930d46e730fbdd96830c6b424a9ee296da94c01bf282262f

Observation 544c19b7-a014-44f3-b289-06bd07b418d8 · inbound

Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games cites this paper.

Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games Measuring General Intelligence with Generated Games

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:57.653000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T01:03:49.101568Z digest=sha256:f9ea9ca5a61ffb07cc865498ed50f1fb6e7560d064f2c294aba2284b89318666

Observation e4b33542-ef48-4d14-9e5a-ef52e875905c · inbound

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning cites this paper.

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning Measuring General Intelligence with Generated Games

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T10:58:16.945355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:58:16.945355Z digest=sha256:dd967d427c3d039cf62a48c5144045c63cdaacc8996145c19ef99f39ad641ea4