Pith. sign in

Paper Citation Record · LEDGER

Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2406.02061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.02061 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T22:14:11.890053Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

21
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation aa4724bd-4dae-4859-90ad-8f213177259c · inbound

GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models cites this paper.

GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:42:12.053805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T00:42:11.891829Z digest=sha256:4a103c4092b54fe2595b77f7eedd2fa485730ee74680faadab96a9194687a294

Observation f6171388-6057-4256-999d-17e645235902 · inbound

Reliable Conversational Agents under ASP Control that Understand Natural Language cites this paper.

Reliable Conversational Agents under ASP Control that Understand Natural Language Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T22:14:11.890053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:14:11.890053Z digest=sha256:64de481a3a9b5e7a0399fd7807c025ccf4246cb620ab03a3bfb9b45ce64c08ee

Observation e0b86c66-c7b7-48fd-a51d-a2a0725b8b4a · inbound

Reasoning Can Hurt the Inductive Abilities of Large Language Models cites this paper.

Reasoning Can Hurt the Inductive Abilities of Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:11.319203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:11.319203Z digest=sha256:299924071a8a094d0370a725a9b2c8b952f4638c9e1e3c446d3806265638f733

Observation c1b0bbe4-8cd5-40da-b9fa-64d14b087304 · inbound

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity cites this paper.

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:10:31.618613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:10:31.440921Z digest=sha256:bd061a2ad5203ec1f44e3e88d55a2af2220dce66be7406c5c8159ff1bed49817

Observation 60db128b-f2c1-4cdf-8794-83ba9ad3f691 · inbound

Mechanistic Interpretability Needs Philosophy cites this paper.

Mechanistic Interpretability Needs Philosophy Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.398691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T23:49:19.683025Z digest=sha256:101691b153dd7d6275737b2da683e5b983b038206511f24453367aa1e8bf5895

Observation 0b1a7d6d-2835-46cf-9ba1-f1bad85add1d · inbound

Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence cites this paper.

Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:15.310030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:15.310030Z digest=sha256:84fc6da46c6e4ef70818bfd7b4504c11ada23a0346ba595759593fa6c1d353bf

Observation d87bcd25-bf69-48cd-b7fa-c14ce7548102 · inbound

Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models cites this paper.

Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:28:57.217863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:28:57.217863Z digest=sha256:9280b5c51ff1c49eb780e792d6c2b0ee3e80ca096fb65dbe147d76dbff4155c6

Observation 1368fe56-c4ec-4ed7-b97a-ebb7bad285dd · inbound

Meanings are like Onions: a Layered Approach to Metaphor Processing cites this paper.

Meanings are like Onions: a Layered Approach to Metaphor Processing Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:37:15.112895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:37:15.112895Z digest=sha256:293953aff3b0c4ae76968c1da9c7df950d686aa9d090cd7588652b0483fd219b

Observation fb289db9-4163-4ce5-9019-af38378a7f74 · inbound

Beyond Behavior: Why AI Evaluation Needs a Cognitive Revolution cites this paper.

Beyond Behavior: Why AI Evaluation Needs a Cognitive Revolution Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:50:48.656399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:33:14.016012Z digest=sha256:3bd1cd556c0d775ef27b0b20a30b3463f0a85422c6024c96e258a790b51b9bd0

Observation e087fb60-30a9-4289-80d9-855cfd1bb5b8 · inbound

Robust Reasoning Benchmark cites this paper.

Robust Reasoning Benchmark Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:08:21.755832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:06:08.545334Z digest=sha256:38ff3ec92eb9cc3d55a2986af125131cc24fd9af5b26bce35eeab7f143b91b88

Observation dc3317ce-a4d3-426e-b2bf-58b1950dd02b · inbound

Robust Reasoning Benchmark cites this paper.

Robust Reasoning Benchmark Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:21:28.769149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T11:18:51.642405Z digest=sha256:df1bde2bea09d7d0b20ae56445ddbc285cb5a116dfe772e666e41477644d2f2b

Observation dc0b0e18-a13b-4ff4-b0c0-e1a3030eac64 · inbound

Robust Reasoning Benchmark cites this paper.

Robust Reasoning Benchmark Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T17:25:30.843496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:25:30.843496Z digest=sha256:129b3ed1281bdeeec2bfe0ff6cc4c48f40d6b51d7ba523ca104ff9ae21f1bd6c

Observation 74629ac8-ff4d-42b9-ad10-2f397aaf8737 · inbound

Majority Voting for Code Generation cites this paper.

Majority Voting for Code Generation Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:36.920185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T09:22:33.721920Z digest=sha256:5fc0187615112a4ec0bf5d8a30cae854702dd134bccd17c77ee5549008a28f0e

Observation 491afd6f-31cf-4a43-93ca-7eb2796888c3 · inbound

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem cites this paper.

How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.989517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:29:33.453354Z digest=sha256:f293dfd153139969694b2f792d2d1861e19e36cb13227d9f93828c09f2f63984

Observation d205ede0-8888-4ccd-9241-607c323f9657 · inbound

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment cites this paper.

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:55:50.873760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:54:49.131461Z digest=sha256:7410dda819e06f1a70605d513812925c6d65f41a81cb018fcb297e4aeeea67a9

Observation 14ee36dc-483b-45df-badf-8f1ff6a88212 · inbound

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities cites this paper.

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:25.946952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:02:06.788056Z digest=sha256:e4682e992e74c84ebbffa65445e5b2f02e0067b242eaface1e8361da8af49ffe

Observation 2fd4a101-057f-4b97-ac30-c73191cafc2b · inbound

Normative Robustness as a Frontier for Non-Verifiable Reasoning in LLMs cites this paper.

Normative Robustness as a Frontier for Non-Verifiable Reasoning in LLMs Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models

Reference 113

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.395515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T09:50:24.891865Z digest=sha256:35432b3e064293b4cbce3c62672612d3899319ff418cb79d52c589d26ec088c9