Pith. sign in

Paper Citation Record · LEDGER

GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2310.12397.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.12397 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:33:37.395529Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

11
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2af25100-fe67-44ff-bece-e9f9a4e841fe · inbound

S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency cites this paper.

S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T21:33:37.395529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:33:37.395529Z digest=sha256:10eadd6036106cec2cd52d432fa43f8042c117bd8bd1d9d00472f877ebffad0b

Observation 5d064cd7-dda9-4876-b173-ef880458fcdb · inbound

Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning cites this paper.

Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T16:57:01.690429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:57:01.690429Z digest=sha256:a6a3c8b04bbfc785bf796b6739fea4911c97ff105a98b42de2ccaa7443364020

Observation c30d184e-8fd5-44b6-89be-c49c8358f46d · inbound

Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning cites this paper.

Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:12:13.428065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:12:13.428065Z digest=sha256:21e03f8ebb9e21dbbd77f61438f41785f3692cec8ea18e0e63db191752c86970

Observation 0d1cd1f9-3767-42b5-80d3-572571960587 · inbound

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models cites this paper.

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:00.039133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:00.039133Z digest=sha256:add432e82f76b360c435e26cdc0e52ad406ff4789a6ec6b80a37f4181011ee6f

Observation e490b109-c17f-4cd5-a8c2-c1fc2c6ca889 · inbound

It's Not That Simple. An Analysis of Simple Test-Time Scaling cites this paper.

It's Not That Simple. An Analysis of Simple Test-Time Scaling GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.621221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.621221Z digest=sha256:945508dd869c0c5fc052b71d62ce2ed938dff23b13d04cbe36197b26cc7ce7d7

Observation c2544d17-6930-4b8d-88d5-f83cfb1cdac9 · inbound

Lightweight Language Models are Prone to Reasoning Errors for Complex Computational Phenotyping Tasks cites this paper.

Lightweight Language Models are Prone to Reasoning Errors for Complex Computational Phenotyping Tasks GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:06:07.559969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:06:07.559969Z digest=sha256:bfc98c3693bd95c975661cc14edc3441d6350f0827b0f124cf3ff418fefd8f5c

Observation 56886bbe-0b59-4a90-8aae-f3d38fa406a8 · inbound

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces cites this paper.

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.605389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T02:57:15.521594Z digest=sha256:ec173135ad8ebc44a830cced332cb2f98cc54ebd620a7116dc709a62767ccba0

Observation cf2eba2c-f2d4-4fdf-897d-fe6536776d99 · inbound

Roll Out and Roll Back: Diffusion LLMs are Their Own Efficiency Teachers cites this paper.

Roll Out and Roll Back: Diffusion LLMs are Their Own Efficiency Teachers GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:42:46.111972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T20:41:48.063871Z digest=sha256:c4489c4f23891dc38e06924112324ab9c1f6e25511e533d2a618a0e0ff37b518

Observation 40139c4b-32fe-4dde-99d6-4ca72fa8a5ae · inbound

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning cites this paper.

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.375155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:23:24.536258Z digest=sha256:4985269cb51e441f188ea3345734be3d816c8e5c572c6a84e0b6f1d183434c99

Observation 06a3d703-496c-4a31-aaef-27d37cbbcfb3 · inbound

The Self-Correction Illusion: Role Relabeling Gates Explicit Error Flagging in Large Language Models cites this paper.

The Self-Correction Illusion: Role Relabeling Gates Explicit Error Flagging in Large Language Models GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T01:31:29.251945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T01:25:07.890796Z digest=sha256:c3dee7392dde0afe35554d21802ac28af3200b2490b80e42b6f1d0ad4f3e0026

Observation 3384582e-a352-49d0-a3ef-65f30119565d · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:07:23.497033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:103e8a20094f5f40f91b3a9a0f0257c74ee0d56d407b6b2ecb034627915a6451