Pith. sign in

Paper Citation Record · LEDGER

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

As of 17 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2508.11944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11944 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:48:57.786431Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:38:14.922388Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e303be8b-d118-49da-83fd-debb922c0a42 · outbound

This paper cites GPT-4 Technical Report.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:53.816976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:53.816976Z digest=sha256:abd80a8b09a3ff599d2ded2d239ba71e62015f7d9e79fff2c1de328c0c9aeef5

Observation 59b3bbf1-b490-4d5c-bbf6-379836ff166d · outbound

This paper cites Akata, L.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Akata, L

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:02.677810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:53.901582Z digest=sha256:c11d8ec93aca75cf1a2648a3896e8c0b77cdfd1ed983bbc4d4b2c365c7abd7a2

Observation 8b04781d-4edc-44cf-ba33-23cc0ee101bf · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:53.984633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:53.984633Z digest=sha256:4df585271cc2f7a03fae3e032157d3641a5d66b845fbd745ae683822de369ac0

Observation 8d10f084-86a3-4d5e-8038-75417201380e · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:02.434647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:54.142406Z digest=sha256:470f15f6a05dc59b254e9ee0bfbebf881337100f037741a9885ac4c646832412

Observation 257e9564-3cfa-4a29-ab2f-289dcd0a884d · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:02.177388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:54.291423Z digest=sha256:92d38a152e764917aebdeeddf7aba8987ae8dc0ea7f52f9cb7cfec970c62d44c

Observation a21a191a-6f5b-40c0-9fdb-3b3ae9a3084e · outbound

This paper cites Chong, T.-H.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Chong, T.-H

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.912672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:54.446528Z digest=sha256:afc2820776ff1352f94eee8945b8e44f54bd42592057c595311bf2e4bb2d5d61

Observation 3843406b-2c53-4970-b085-fa8a20ecafa8 · outbound

This paper cites Costa-Gomes, V.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Costa-Gomes, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.645005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:54.541089Z digest=sha256:86ca66fe5a422feab0453e3d273b58e157812fca7275e0a50c604b4d39c0fd74

Observation df416bba-30db-46f5-8714-5221c907d163 · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:54.647303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:54.647303Z digest=sha256:4cb25e4afa1404991f82465858b9ebde2df7107b1a4163ebd5bc9a264eebf069

Observation 429390c3-b5d6-4da9-8863-ed5495f19bf6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:01.465721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:54.758623Z digest=sha256:99404ffda6a03132d88ba37771945812c48d8a819e246e436033c5caabfc7a7c

Observation 909c6723-8d06-42bc-977a-fb92d03ebb61 · outbound

This paper cites A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:54.905621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:54.905621Z digest=sha256:36fa9ddb26a9b3d1ada3add919ba792b292864b910ec42ac46f2e5eb19013a3f

Observation ec70a49a-e9c7-4d09-b7fe-9998a73e605b · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:01.256598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:55.063515Z digest=sha256:6de9e14372e59fa39e032d25aae92d77b7d53327eaa05f878269059d47971f3a

Observation 1a714281-9d29-48eb-b7b4-27f5823c0775 · outbound

This paper cites Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:55.222817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:55.222817Z digest=sha256:c451148bcebe93baf9c617af179ff705f6c67c066cd71689c25a5c5b0255d8a2

Observation c630cc8a-08cb-46c0-8133-c0ca8d957c88 · outbound

This paper cites Fudenberg and J.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Fudenberg and J

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.066870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:55.350398Z digest=sha256:5c0c0ae59223f73767d1e53f44e7fd3e8f7370c5d1ec8579f7806e478b1ced0b

Observation d3e8c5dc-09fd-49b6-a091-577ead17bf61 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.912683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:55.499041Z digest=sha256:9398d0315ae7a229f35bbcb7b5132aad0dc9cb2d3aac1805dff7f5b46ca30942

Observation d96cafc8-7b40-4594-89d4-bff81bf4cb91 · outbound

This paper cites The Llama 3 Herd of Models.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:55.653643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:55.653643Z digest=sha256:bd8183de88cb191a399c5354367a1f01ad77e6abc42a7391a63aa9e67d699d60

Observation ad6c2434-3691-40f2-aa21-6e23926de8b9 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.680640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:55.758196Z digest=sha256:50252df40a10e0e66bd36ea60aab0b7a2e85cb6c5e06a914b3f2b932e9e0d029

Observation 7162f826-bfe5-47c2-bd8d-c66074baf6d7 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.510380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:55.881168Z digest=sha256:c6233680482df4e3557ca31e85a9123e7ac4bd9a167c5e718f532c2f32a5dee3

Observation b88e6566-f4a1-4f4f-8ccc-a75650918512 · outbound

This paper cites A Survey on Large Language Model-Based Game Agents.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs A Survey on Large Language Model-Based Game Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.005817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.005817Z digest=sha256:0fdda5a3c7134c74522ec9074203703e29e72fc2311c662a5fbf1958236ba86a

Observation ce17b52e-199c-40a2-949f-6ca5081c00bc · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.126881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.126881Z digest=sha256:3a890b9e17ef0eaa34ab44ade1111b8eccb698504380bbf92e68e6d499010427

Observation 22f9f327-c4fb-4439-8656-ffee46d16aae · outbound

This paper cites Lorè and B.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Lorè and B

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:00.280558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.219061Z digest=sha256:a982e6894fb96b72ddd3f674d50825a68c37f2e9af0194c2a94a47c050ab7b98

Observation 0bf3f676-7b40-4777-b545-a9222a18c758 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.112476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.340141Z digest=sha256:c1035c217c6dbfe66ca0907ca6d0348b80199a4f7b783a5a052f848fe1fd6a79

Observation 58cc6ebe-2cba-432f-9598-7dea43a8b1c6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.906769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.417387Z digest=sha256:5b37235f8f2b9e819472336ba2041733782a9329fc954871c1bc3ef61c3acd8c

Observation 99134207-b27a-4412-abc1-1b21383bf8f4 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.595239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.536855Z digest=sha256:1c893e1594312979ac7899aa9b2e1c4aed323fd87bb2e5c6fd8949e4ee49b6f5

Observation 35fef22c-dd53-4fe4-8c62-2601bf32a03e · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.420270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.612362Z digest=sha256:fc6faf33d8f4968ed0c474fe8db2e4fcce7fe2503a668ab002fc2d7a1ef075ad

Observation c187cfa9-8834-46ac-a58a-cb99298d8be8 · outbound

This paper cites LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.677470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.677470Z digest=sha256:e3084bd1eef02be89af3cfff8ffa301411c5e54eb052c783a4b6cb945ce8142f

Observation a87b0631-deb3-4b2c-900b-a4e21463f7e5 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.188161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.793368Z digest=sha256:bafebec76b22ee6e846623fbd1fbde6b4f3a9284743ca52d823d0c87b6d260cc

Observation 934bacbe-ba0a-4994-ba64-9a68b3f426b6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.014871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:56.897295Z digest=sha256:6f812caa1d9d83c124f82c4cca4d253d41f221f0283f9cd4c9f82a2f893d3339

Observation 40691c42-96d4-4064-b748-14398244dee2 · outbound

This paper cites Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.972734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.972734Z digest=sha256:ad869938f810260445e633216f0d900d93b539b2f9d51d074f73a65b53b339f0

Observation d8542316-4165-4814-b143-c6a247efa951 · outbound

This paper cites Do Large Language Models Exhibit Spontaneous Rational Deception?.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Do Large Language Models Exhibit Spontaneous Rational Deception?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.044552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.044552Z digest=sha256:63557fe7344e8bf7749b08c40a110cafa221c16d3b8a8c671aefe7851593208a

Observation c4fe8893-7ef0-4a65-ab68-f1fff782b201 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.113858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.113858Z digest=sha256:8ff1b75180548a288b2f22de58fdb0635207f1edcbcfaa9f77e45d4d5e33c34e

Observation 8e62b07a-a745-4297-877c-b2f2fdde2ae1 · outbound

This paper cites V on Neumann and O.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs V on Neumann and O

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:48:58.793228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:57.188015Z digest=sha256:bf2ffef1ee2afd5f38153d8992100ee1851e1dbc05dd543d9734407ea8dc5e60

Observation 466212e9-ea9d-434d-95d4-9dec953ffccc · outbound

This paper cites TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.298899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.298899Z digest=sha256:99ea25dad0324dc95def23a485a4654f0d4be6929914a58e1fdb528fe50102f5

Observation e9dbb274-4c68-433e-9893-3fbd28f16a44 · outbound

This paper cites Wright and K.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Wright and K

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:48:58.622211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:57.369503Z digest=sha256:fa008fbe46692549ea4793351f2413f57037b1f87d5a17b9acca06fdca02dd0e

Observation cb3c3c2f-d1cf-40ec-a07f-edb370ef1b75 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.467609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:57.452686Z digest=sha256:b0cb46ff12da77c548eb50859d435e6951eb1552721f4f8b9822f5f314dc0072

Observation 8ac9ef5e-177e-4c42-98a2-6a93a1c701e2 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.283296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:57.519879Z digest=sha256:4685d46a60093df7eac93d04b22cbab02541676820b97c3af89b349ae4a34c7a

Observation 4d1f26ad-5048-4b03-afe6-e57a5495cd56 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.074219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T19:48:57.618873Z digest=sha256:d5b09b67d8f2987b0234010cee8a6b0e1b0b4f03c06096c782f99689a7a9f44f

Observation 70ba9204-49ab-4d69-95b3-db41edff3ada · outbound

This paper cites Qwen2.5 Technical Report.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Qwen2.5 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.703996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.703996Z digest=sha256:d5b914f9ef031ff29094d4ff86cbd8a9586d3109ed09a3ca422acd77db95fb5b

Observation ad174fc3-992c-4687-b16a-087d837f79bb · outbound

This paper cites LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.786431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.786431Z digest=sha256:886f146262e2a9d576c17a23f5f074a7b4e5b23b15a16610263ac4596e9b5023

Pith citing papers

Observation f03bcd7d-52dc-47aa-be67-13b89306fced · inbound

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks cites this paper.

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-26T00:38:42.728182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-26T00:38:14.922388Z digest=sha256:60e2f3df49d2162472b7506f0d26017ed60af79fdbc1cd3ed9156cd1af696977