Pith. sign in

Paper Citation Record · LEDGER

Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2406.02061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.02061 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T22:14:11.890053Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

21
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation aa4724bd-4dae-4859-90ad-8f213177259c · inbound

GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models cites this paper.

GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:42:12.053805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T00:42:11.891829Z digest=sha256:4363a1ec06790a4b5f05038703b7ad0a460ad2d892efb59ce491d2437804b39c

Observation f6171388-6057-4256-999d-17e645235902 · inbound

Reliable Conversational Agents under ASP Control that Understand Natural Language cites this paper.

Reliable Conversational Agents under ASP Control that Understand Natural Language Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T22:14:11.890053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:14:11.890053Z digest=sha256:22097fd955cbb78c8b850695cb8733de4e697eeb08f4054b2fecc1c1c27298ef

Observation e0b86c66-c7b7-48fd-a51d-a2a0725b8b4a · inbound

Reasoning Can Hurt the Inductive Abilities of Large Language Models cites this paper.

Reasoning Can Hurt the Inductive Abilities of Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:11.319203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:11.319203Z digest=sha256:208dbca93a3728a322fb5f66f952242c251b1f0befcfa5aabdcd3a55b9d70788

Observation c1b0bbe4-8cd5-40da-b9fa-64d14b087304 · inbound

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity cites this paper.

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:10:31.618613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T16:10:31.440921Z digest=sha256:806c63e0bfc254283c06f09d835560176253555e16eeef03df81748ef3de09f2

Observation 60db128b-f2c1-4cdf-8794-83ba9ad3f691 · inbound

Mechanistic Interpretability Needs Philosophy cites this paper.

Mechanistic Interpretability Needs Philosophy Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.398691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T23:49:19.683025Z digest=sha256:9559e2484f35a90a3ffde7971db789be8049ff6206b5eda49c0b8012d3e329d6

Observation 0b1a7d6d-2835-46cf-9ba1-f1bad85add1d · inbound

Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence cites this paper.

Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:15.310030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:15.310030Z digest=sha256:84fc6da46c6e4ef70818bfd7b4504c11ada23a0346ba595759593fa6c1d353bf

Observation d87bcd25-bf69-48cd-b7fa-c14ce7548102 · inbound

Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models cites this paper.

Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:28:57.217863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:28:57.217863Z digest=sha256:9280b5c51ff1c49eb780e792d6c2b0ee3e80ca096fb65dbe147d76dbff4155c6

Observation 1368fe56-c4ec-4ed7-b97a-ebb7bad285dd · inbound

Meanings are like Onions: a Layered Approach to Metaphor Processing cites this paper.

Meanings are like Onions: a Layered Approach to Metaphor Processing Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:37:15.112895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:37:15.112895Z digest=sha256:293953aff3b0c4ae76968c1da9c7df950d686aa9d090cd7588652b0483fd219b

Observation fb289db9-4163-4ce5-9019-af38378a7f74 · inbound

Beyond Behavior: Why AI Evaluation Needs a Cognitive Revolution cites this paper.

Beyond Behavior: Why AI Evaluation Needs a Cognitive Revolution Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:50:48.656399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:33:14.016012Z digest=sha256:1b52c8ae84846fb2ed29c9d59f4ad0613244cd13f2ea825497e37ec1e15aa6b3

Observation e087fb60-30a9-4289-80d9-855cfd1bb5b8 · inbound

Robust Reasoning Benchmark cites this paper.

Robust Reasoning Benchmark Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:08:21.755832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:06:08.545334Z digest=sha256:0c3a62b46bbe757d10cdf2f47ae3a55018db80332c0fb5f7e7f88fa10211b780

Observation dc3317ce-a4d3-426e-b2bf-58b1950dd02b · inbound

Robust Reasoning Benchmark cites this paper.

Robust Reasoning Benchmark Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:21:28.769149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T11:18:51.642405Z digest=sha256:6ae6f683fe273d1a1bd1435638fd3d56c09c53347fb376cba5fb422f8d310a63

Observation dc0b0e18-a13b-4ff4-b0c0-e1a3030eac64 · inbound

Robust Reasoning Benchmark cites this paper.

Robust Reasoning Benchmark Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T17:25:30.843496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:25:30.843496Z digest=sha256:7ebfcb15699d13dc834741faffe4204383f500cd245d420f87b7fd7c9b1bf1af

Observation 74629ac8-ff4d-42b9-ad10-2f397aaf8737 · inbound

Majority Voting for Code Generation cites this paper.

Majority Voting for Code Generation Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:36.920185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T09:22:33.721920Z digest=sha256:474d4dede6ea8a82c652641e188780c4ab84151d3d651382ba94cf8e0ca19bb5

Observation 491afd6f-31cf-4a43-93ca-7eb2796888c3 · inbound

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem cites this paper.

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.989517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:29:33.453354Z digest=sha256:e6efd36d13be4c709ccf584c0e10a969645e736bbe071b15dfbda21ffe5f95d9

Observation d205ede0-8888-4ccd-9241-607c323f9657 · inbound

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment cites this paper.

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:55:50.873760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:54:49.131461Z digest=sha256:5aed6e026b4eff570f5fe2f61b2b8d13afa166d021517c486d1e3a57bce3b2cb

Observation 14ee36dc-483b-45df-badf-8f1ff6a88212 · inbound

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities cites this paper.

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:25.946952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:02:06.788056Z digest=sha256:116ce96a4398acecd90c234c532f96885010c1786cbb744b39d9f2b6495cd01a

Observation 2fd4a101-057f-4b97-ac30-c73191cafc2b · inbound

Normative Robustness as a Frontier for Non-Verifiable Reasoning in LLMs cites this paper.

Normative Robustness as a Frontier for Non-Verifiable Reasoning in LLMs Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 113

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.395515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:50:24.891865Z digest=sha256:68d45b97ed3d536c8f3ac9718e7fa98b32072591db21597e13d12057febb1d8d