Pith. sign in

Paper Citation Record · LEDGER

Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2304.03439.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.03439 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:29:05.173640Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T21:13:28.129474Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1aba025d-eef0-4033-8b81-d5a9713eedfd · inbound

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code cites this paper.

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 276

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:34:42.943084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T17:34:42.565806Z digest=sha256:a357e837a03298b594122f979b6c4dff6a3196785120c000c5862f357d5f4a37

Observation b4da8855-64c8-4b14-afed-79baffe23736 · inbound

Learning to Ask: When LLM Agents Meet Unclear Instruction cites this paper.

Learning to Ask: When LLM Agents Meet Unclear Instruction Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:13:28.132607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T21:08:42.276002Z digest=sha256:83524648463c6a4c38de8ddfd536022824bb44c1f81e20a66f47ff98671bb688

Observation e4519ae6-1558-49ff-a6d2-0ae583563b02 · inbound

S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency cites this paper.

S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T21:33:37.100770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:33:37.100770Z digest=sha256:e220fef90a314e8764ba9cdf96929e700333e01d811c33201f1819621a6d1834

Observation 27c8b0e7-b202-469a-9a3e-4de486c9a7b4 · inbound

Self-Rationalization in the Wild: A Large Scale Out-of-Distribution Evaluation on NLI-related tasks cites this paper.

Self-Rationalization in the Wild: A Large Scale Out-of-Distribution Evaluation on NLI-related tasks Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T21:29:44.366511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:29:44.366511Z digest=sha256:602e8da2a7a63ea35e98a1716ab8a2afd7e0810ca2daf9eb6c041e83a7f16886

Observation 3ac8284d-ce21-4762-8ab6-83c2d3f19bd0 · inbound

Reasoning-as-Logic-Units: Scaling Test-Time Reasoning in Large Language Models Through Logic Unit Alignment cites this paper.

Reasoning-as-Logic-Units: Scaling Test-Time Reasoning in Large Language Models Through Logic Unit Alignment Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T10:29:05.173640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:29:05.173640Z digest=sha256:1a670e989bf6ed0f659d11f200bbde7dbcc9d3df8063ce0679f0697e6f457c58

Observation 58c4c852-cbdc-4b7d-b95c-4db99928b980 · inbound

NLI under the Microscope: What Atomic Hypothesis Decomposition Reveals cites this paper.

NLI under the Microscope: What Atomic Hypothesis Decomposition Reveals Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T10:58:42.484105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T10:58:42.484105Z digest=sha256:2e499c6e0416f05db569104ebee04a179b8ad74be0d709ef4e269a94644b1c49

Observation 111969fe-0c94-429d-8c13-d85f6ef0b928 · inbound

Multilingual Question Answering in Low-Resource Settings: A Dzongkha-English Benchmark for Foundation Models cites this paper.

Multilingual Question Answering in Low-Resource Settings: A Dzongkha-English Benchmark for Foundation Models Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:25.222986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:32:25.222986Z digest=sha256:b6a44b1f9cca492f6394fa926c421f45faba55ab5a079d5ea5e9f6a751b2219e

Observation 32a1319c-7f7c-4da9-aad8-710c92c4af6e · inbound

A Goal-Oriented Chatbot for Engaging the Elderly Through Family Photo Conversations cites this paper.

A Goal-Oriented Chatbot for Engaging the Elderly Through Family Photo Conversations Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:20:39.871120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T05:17:40.538005Z digest=sha256:08e7b2761ee1f4c541005e6d39e3cd72f5b236ab1ad0d238e4157c487010b524

Observation b04e2b49-897c-4791-a12b-82411698dd5d · inbound

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning cites this paper.

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:26:02.778203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T21:53:51.983234Z digest=sha256:9dd26e44facf848828e81d663b1df5fe0e8e52ef6053ac2387d161c05b66029f

Observation 9e1292ec-c989-4531-9842-6b696a03a4cd · inbound

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization cites this paper.

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:58.747180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T02:17:42.123741Z digest=sha256:ab8d54b45331bc7f8e1efda6acb1d44cbde48b0d477dc776af807ea85c655a5f

Observation 28495df1-08e9-4d49-ae1b-7bfb49570115 · inbound

Novelty-based Tree-of-Thought Search for LLM Reasoning and Planning cites this paper.

Novelty-based Tree-of-Thought Search for LLM Reasoning and Planning Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:56:08.686089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T10:41:59.816073Z digest=sha256:11a26448c7ddf182daffd77d10ba6c5b951b42f7144fb3c2185734697f3cfb1b

Observation 27d53602-4a95-472c-8349-23118d12e48b · inbound

Computational models of pragmatic reasoning with flexible generation of meaning and expression alternatives cites this paper.

Computational models of pragmatic reasoning with flexible generation of meaning and expression alternatives Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4

Reference 193

Resolution
unresolved
no resolver link, observed 2026-08-01T15:27:11.534304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:27:11.534304Z digest=sha256:d95d99766bf9f4fa0c34b97d9815a9c360a0cd36bca7d229763ab2da82b9e8e6