Pith. sign in

Paper Citation Record · LEDGER

LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2408.15778.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.15778 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:16:27.384439Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T12:44:40.220962Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4325c0cb-830e-4371-ba4b-2f1d1f7e18eb · inbound

CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance cites this paper.

CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T12:16:27.384439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:16:27.384439Z digest=sha256:7ef82b2d9c14899f605eb2017b6f77add07663ea784891609e0613b4ba4ccf78

Observation 05336bf8-4c47-4db1-8a32-1b9d12459245 · inbound

CryptoX : Compositional Reasoning Evaluation of Large Language Models cites this paper.

CryptoX : Compositional Reasoning Evaluation of Large Language Models LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T18:34:26.713246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:34:26.713246Z digest=sha256:d312cd52b88d1b41d34014f5e8a7d1264b7d448b6c28252a1c4a09c014ae3634

Observation 785a96bc-8fa0-4272-9202-7e439a8276d6 · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.488705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.488705Z digest=sha256:1fa3016825636651a777f492fc4f08c6864e25092a346cbecc02ad6cd9da58d8

Observation 76549ae5-f538-4580-a655-95ee05d74ed8 · inbound

OIBench: Benchmarking Strong Reasoning Models with Olympiad in Informatics cites this paper.

OIBench: Benchmarking Strong Reasoning Models with Olympiad in Informatics LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:33:10.456529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:33:10.456529Z digest=sha256:a259bdfaa664141bd04cc1046e54caecb14365da0dd1137698c385005b9da4ed

Observation 6c6955fe-5e2f-4772-93fb-cddf08be402c · inbound

Are Large Language Models Capable of Deep Relational Reasoning? Insights from DeepSeek-R1 and Benchmark Comparisons cites this paper.

Are Large Language Models Capable of Deep Relational Reasoning? Insights from DeepSeek-R1 and Benchmark Comparisons LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:50.183141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:50.183141Z digest=sha256:886a3c450c2da1ff8497690ff3eabbb6680ae362104971fc87681b3cfe88930f

Observation 0fbb90fd-3409-4f1f-a294-036a53e48e05 · inbound

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data cites this paper.

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T18:08:40.254125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:08:40.254125Z digest=sha256:edac4bd8d5e3994aa8ebece1dac7c3afcfd682872d527371b86603a0b642fdb3

Observation e546555a-3bc0-4ed5-b677-ec87a70695d5 · inbound

Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL cites this paper.

Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T04:42:05.864107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:42:05.864107Z digest=sha256:3fdef760e96f316502b0eaed474e122610e5a47462e38d5fb1f9f2a0fc1689c1

Observation ae0a7c6b-6068-4b31-a992-fb994c79c520 · inbound

BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards cites this paper.

BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:44:40.222479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T10:07:39.554999Z digest=sha256:2f577a555f5a478f60be8ee58fa3cc74e481710c97431bd4fa5b7be125c86f7b