Pith. sign in

Paper Citation Record · LEDGER

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2508.11944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11944 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:48:57.786431Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:38:14.922388Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e303be8b-d118-49da-83fd-debb922c0a42 · outbound

This paper cites GPT-4 Technical Report.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:53.816976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:53.816976Z digest=sha256:347ff7e81cc452f2751797763365962d7912a213d18fc7e85e6da7986575f5bb

Observation 59b3bbf1-b490-4d5c-bbf6-379836ff166d · outbound

This paper cites Akata, L.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Akata, L

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:02.677810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:53.901582Z digest=sha256:44a3f42962a6cf703d783080cc800ff60fd3605c881352dc9ebd5f1819c33432

Observation 8b04781d-4edc-44cf-ba33-23cc0ee101bf · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:53.984633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:53.984633Z digest=sha256:12a9e7b161e0f87ff679feb9d061baa2a19c0589110fe3f1e24ff09a594d3506

Observation 8d10f084-86a3-4d5e-8038-75417201380e · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:02.434647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.142406Z digest=sha256:43ebbfa681d53a965cd56cabf546b9cf9bc1c2ec86668ec4506a003b290c8318

Observation 257e9564-3cfa-4a29-ab2f-289dcd0a884d · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:02.177388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.291423Z digest=sha256:a84f8a0abd9b02aceed63f15b5d3332d801393aa81dceca2975431ec6348040d

Observation a21a191a-6f5b-40c0-9fdb-3b3ae9a3084e · outbound

This paper cites Chong, T.-H.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Chong, T.-H

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.912672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.446528Z digest=sha256:3de4d5f49b03ff0e06d828fe27c834c0ee89774954fe1878250750423205645e

Observation 3843406b-2c53-4970-b085-fa8a20ecafa8 · outbound

This paper cites Costa-Gomes, V.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Costa-Gomes, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.645005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.541089Z digest=sha256:35f84a5db20463457ccf84187d61f3535426102f8142b9da8649a639b6834a89

Observation df416bba-30db-46f5-8714-5221c907d163 · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:54.647303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:54.647303Z digest=sha256:cf00229ff4b6873813aa97e53aaa00c723a11756fbab08a2981cbbcedf2a2bf0

Observation 429390c3-b5d6-4da9-8863-ed5495f19bf6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:01.465721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.758623Z digest=sha256:76a2229bde5a9f5d58b0e891320ef2019e7c45d8804ef5f020baabdae826c222

Observation 909c6723-8d06-42bc-977a-fb92d03ebb61 · outbound

This paper cites A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:54.905621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:54.905621Z digest=sha256:b88ecbb08f4dbb72c2e7bc8a10fe48f76acc6f3ba4b61a93c3ee965e27a6eae4

Observation ec70a49a-e9c7-4d09-b7fe-9998a73e605b · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:01.256598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.063515Z digest=sha256:4fd667b1702fd1bf8ab6d7a8ba4961826032b7bc870b20e3d2e48401870062a9

Observation 1a714281-9d29-48eb-b7b4-27f5823c0775 · outbound

This paper cites Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:55.222817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:55.222817Z digest=sha256:392d3e72b0d6589196af2842f0bd0dfa6946fea381904f0672549601293fc7a9

Observation c630cc8a-08cb-46c0-8133-c0ca8d957c88 · outbound

This paper cites Fudenberg and J.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Fudenberg and J

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.066870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.350398Z digest=sha256:8baf21db0eacf33f42bc014aa1e1898f3d75a4dda79a77f0e50ede6d07ff762e

Observation d3e8c5dc-09fd-49b6-a091-577ead17bf61 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.912683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.499041Z digest=sha256:7c0e1825107973b8c37343c841de9c230d4370349e6464d3b7652b3f4b4a1344

Observation d96cafc8-7b40-4594-89d4-bff81bf4cb91 · outbound

This paper cites The Llama 3 Herd of Models.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:55.653643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:55.653643Z digest=sha256:98f081d9b681bacdbf1b3903f477b2b48cd6c5093fce2e764717788ebf6affca

Observation ad6c2434-3691-40f2-aa21-6e23926de8b9 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.680640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.758196Z digest=sha256:7003bd2772395b973bba1dae97901f6b1ee3f0e8a1bf545387f9c5cc23c5bc87

Observation 7162f826-bfe5-47c2-bd8d-c66074baf6d7 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.510380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.881168Z digest=sha256:e90adf14700c63405e8424154928e06eaf1992e2e237be794ff43f0ce0b9206f

Observation b88e6566-f4a1-4f4f-8ccc-a75650918512 · outbound

This paper cites A Survey on Large Language Model-Based Game Agents.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs A Survey on Large Language Model-Based Game Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.005817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.005817Z digest=sha256:72c384669173c63ed4c2ad98161ab51b2a97f00f48bce237de6d6048f79c415e

Observation ce17b52e-199c-40a2-949f-6ca5081c00bc · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.126881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.126881Z digest=sha256:8bf1373b4cfda600dfd8cdea913cff09cccdab1bf9f7e7ddd8478286b8a69c4f

Observation 22f9f327-c4fb-4439-8656-ffee46d16aae · outbound

This paper cites Lorè and B.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Lorè and B

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:00.280558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.219061Z digest=sha256:c8bca305a6dd5d8999d0cb3ba7cd974922a3f2c487ca4f740222ef111f36881b

Observation 0bf3f676-7b40-4777-b545-a9222a18c758 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.112476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.340141Z digest=sha256:600cfb166a9f513b00779bfb8e8c4ec6cb9eaf6b791f71f746f7d5011f092501

Observation 58cc6ebe-2cba-432f-9598-7dea43a8b1c6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.906769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.417387Z digest=sha256:725472c6e25260986cd771e31ceb488cb283519075e0a2171893f35bc7b48201

Observation 99134207-b27a-4412-abc1-1b21383bf8f4 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.595239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.536855Z digest=sha256:774f89d000b7e52acaa34ec597d7a549f7a0c91a293de19cf19be2f156f7b6e0

Observation 35fef22c-dd53-4fe4-8c62-2601bf32a03e · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.420270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.612362Z digest=sha256:3c87e8aca87829122e19935bcf9b514162e9464446311f64e2edb47fec73ad40

Observation c187cfa9-8834-46ac-a58a-cb99298d8be8 · outbound

This paper cites LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.677470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.677470Z digest=sha256:97d44136ebbe55575f965a31f0ca63342b190a58a0f807c24c72049621340ec5

Observation a87b0631-deb3-4b2c-900b-a4e21463f7e5 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.188161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.793368Z digest=sha256:3518ac58608a30d4b245edba800cb4b87ce77cb2c24ec449cfa30346c01ed2ef

Observation 934bacbe-ba0a-4994-ba64-9a68b3f426b6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.014871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.897295Z digest=sha256:7a0cd54e5dcf347fc81d65ddaab01439874e62d1e2f50150ea3d750b9fe9b926

Observation 40691c42-96d4-4064-b748-14398244dee2 · outbound

This paper cites Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.972734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.972734Z digest=sha256:0c84453af94b9870bda2296cc6f6e73a36f77d128baf7870faa3add1305491c4

Observation d8542316-4165-4814-b143-c6a247efa951 · outbound

This paper cites Do Large Language Models Exhibit Spontaneous Rational Deception?.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Do Large Language Models Exhibit Spontaneous Rational Deception?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.044552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.044552Z digest=sha256:e9d25987749fbf908a8a78fccfd82b909dcb050843ff5fb233465c0e0667db82

Observation c4fe8893-7ef0-4a65-ab68-f1fff782b201 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.113858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.113858Z digest=sha256:48c9c88e02025db5825233d470ecfe8a54a67cafc8605307a7c7e51b254af995

Observation 8e62b07a-a745-4297-877c-b2f2fdde2ae1 · outbound

This paper cites V on Neumann and O.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs V on Neumann and O

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:48:58.793228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.188015Z digest=sha256:3a15eab43f9234ac7e5f4373c3efa0b81fcfa0b5d3579db847fa63a1f33a83c6

Observation 466212e9-ea9d-434d-95d4-9dec953ffccc · outbound

This paper cites TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.298899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.298899Z digest=sha256:f7cc325ecb998677385407c3125332d16c5fa0bd1809ffdb7220a1333c33b2c4

Observation e9dbb274-4c68-433e-9893-3fbd28f16a44 · outbound

This paper cites Wright and K.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Wright and K

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:48:58.622211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.369503Z digest=sha256:7630df9790f67a86268f6cc25ac063ceb33754055ea32210317191bffbda128a

Observation cb3c3c2f-d1cf-40ec-a07f-edb370ef1b75 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.467609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.452686Z digest=sha256:fe5214b2e156e81fb92c716a58025540c68bc1abf7ef961ab82c79ca8c86deab

Observation 8ac9ef5e-177e-4c42-98a2-6a93a1c701e2 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.283296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.519879Z digest=sha256:2542d34ed6354b1bf012cfacc1f8d2e984b8de3ccb4dc77b3a3777b659b00fec

Observation 4d1f26ad-5048-4b03-afe6-e57a5495cd56 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.074219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.618873Z digest=sha256:d24238c7a14c7eba1aafa94affef07e5c23c55402cf107787b96c34276f82e06

Observation 70ba9204-49ab-4d69-95b3-db41edff3ada · outbound

This paper cites Qwen2.5 Technical Report.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Qwen2.5 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.703996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.703996Z digest=sha256:1c267b48560197aca364984b80ff454fa0d9058bb7b2bf7cd1e680ab3dfcc8a5

Observation ad174fc3-992c-4687-b16a-087d837f79bb · outbound

This paper cites LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.786431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.786431Z digest=sha256:74cf7479d34ddaf2ff6b54d6479212f6e8173354bbc77a301058295767618d5f

Pith citing papers

Observation f03bcd7d-52dc-47aa-be67-13b89306fced · inbound

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks cites this paper.

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-26T00:38:42.728182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T00:38:14.922388Z digest=sha256:3891fe2e04d9580b7e8572894b2c5d46177cac45208a12060add18b1e1518bd9