Pith. sign in

Paper Citation Record · LEDGER

Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2410.21333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.21333 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:40:35.314587Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 32593cf1-1e89-4c07-af2d-540706092296 · inbound

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models cites this paper.

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 114

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T01:29:57.415551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T01:29:56.480020Z digest=sha256:52487f923ad0e557766c6eca5d6a0e981c409ff39531f53ad2b4cfbaa09a6eab

Observation d8884055-f106-44e7-a599-64123dd7f2b3 · inbound

Are Large Language Models Reliable AI Scientists? Assessing Reverse-Engineering of Black-Box Systems cites this paper.

Are Large Language Models Reliable AI Scientists? Assessing Reverse-Engineering of Black-Box Systems Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:40:35.314587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:40:35.314587Z digest=sha256:fbcfe6aab996fab075796c1618263b84149b67487e16633ecd2b4d8ebf416cff

Observation b66d6b6f-068c-43f8-b0fa-1d08bfc82a9e · inbound

From Token to Action: State Machine Reasoning to Mitigate Overthinking in Information Retrieval cites this paper.

From Token to Action: State Machine Reasoning to Mitigate Overthinking in Information Retrieval Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:57:23.967398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:57:23.967398Z digest=sha256:430ab862ee18b4b8a626f276a6aa456586835e1e700dce1040e20e7e0acc3cc9

Observation 8447cec5-859c-4d40-84e3-723efe34fd4b · inbound

Knowing Before Saying: LLM Representations Encode Information About Chain-of-Thought Success Before Completion cites this paper.

Knowing Before Saying: LLM Representations Encode Information About Chain-of-Thought Success Before Completion Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:20.364514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:20.364514Z digest=sha256:4342e05a3f118f9b3e0a6272e78487c321588c0dcce460c8b98e22c0db387328

Observation 0539741c-71f3-4baf-8a1a-1e7beceaf4d8 · inbound

Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language Models cites this paper.

Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T06:09:18.801709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:09:18.801709Z digest=sha256:df07f9b3373c2c2d7620b4c92988730f0c9f9d338e4d2ad8fcbaecbcfb7a4396

Observation 5de6ad31-e6c1-478f-9106-a9a10891292e · inbound

Data Shifts Hurt CoT: A Theoretical Study cites this paper.

Data Shifts Hurt CoT: A Theoretical Study Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:30:39.274909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:30:39.274909Z digest=sha256:85ecebe9846d3d2032a803bf1745d77f64574c61c1c6238981b8d2afc731ef4f

Observation 871e67df-6eb3-4b60-b31e-774f944b4dec · inbound

Unveiling Confirmation Bias in Chain-of-Thought Reasoning cites this paper.

Unveiling Confirmation Bias in Chain-of-Thought Reasoning Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:59:56.705039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:59:56.705039Z digest=sha256:f8a6fb533086583068e0965e0314ed7200fa1cd38f8ea17c23c48c9c79538d87

Observation b1072eab-474a-4f5f-b15b-a627c17fdc1a · inbound

The Other Mind: How Language Models Exhibit Human Temporal Cognition cites this paper.

The Other Mind: How Language Models Exhibit Human Temporal Cognition Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:30:35.880034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:30:35.880034Z digest=sha256:7ba68d6cce4300bc0e585a16c495e41d0dde57c57264281125a398d0d91af167

Observation 98eea2dc-d6ec-496a-a79f-ccd80b5dd6cc · inbound

Before Humans Join the Team: Diagnosing Coordination Failures in Healthcare Robot Team Simulation cites this paper.

Before Humans Join the Team: Diagnosing Coordination Failures in Healthcare Robot Team Simulation Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T00:01:56.459681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T23:57:13.740711Z digest=sha256:b966123dbab5447d536817c6dffc987c674d9b4b094dcde98674e0234ec21b89

Observation 0880fb46-a954-4f60-8ee0-b3edf2265005 · inbound

Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation? cites this paper.

Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation? Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:30:30.201806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:30:30.201806Z digest=sha256:9b734a9ca77fda73101457d20cfd9a4488825320aa4b9892aec51223308b2ed9

Observation 3c168890-6b0a-49d1-b959-29764319d0ca · inbound

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences cites this paper.

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:41:37.992419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T16:39:03.794436Z digest=sha256:3e082f3b9a6df23b271123bded5d259b3f9241fe9c750367fef03bf584050bf9

Observation 0790eb7d-be21-4de4-8e4d-ec72e3366a15 · inbound

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences cites this paper.

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T16:41:37.497592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T16:39:03.794436Z digest=sha256:0bdef211961ed8b9ecab24403a1cf62e73426c5b7df86a1b4e038ea98c19e122

Observation 62900d50-7469-4982-a3bd-558b2c71a6ff · inbound

Artificial Phantasia: Emergent Mental Imagery in Large Language Models cites this paper.

Artificial Phantasia: Emergent Mental Imagery in Large Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T21:44:22.831192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T21:41:39.111769Z digest=sha256:fcf26da9efdf4c6d791f97ceee0589be009b39a208a1cf7e63d661441fe8dad3

Observation 8a46dd3c-f887-4b83-9cdf-c6f9509c4a9c · inbound

Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification? cites this paper.

Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification? Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:53:00.653876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T14:51:26.368439Z digest=sha256:5b78831235e533d5aa8a70e8606c3081df81a0aec4f3a996e24b9f00260e6256

Observation d149aad9-84d2-43ef-acbc-05810b350d41 · inbound

SinkTrack: Attention Sink based Context Anchoring for Large Language Models cites this paper.

SinkTrack: Attention Sink based Context Anchoring for Large Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:51:01.075927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:46:44.996260Z digest=sha256:b7f4f3bbe2a2c4b6068440e705d5fcb02b2bdf560e8b506371d13a9380002fd3

Observation 0de1dc43-2832-4448-ae20-42dfdc9fc68b · inbound

SinkTrack: Attention Sink based Context Anchoring for Large Language Models cites this paper.

SinkTrack: Attention Sink based Context Anchoring for Large Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T01:39:22.676699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T01:37:33.760985Z digest=sha256:3e6f5c526306c1fa8f4424c1b76016df0b24d8bb9a217ae2bb7fbe8d5e852849

Observation 3bfdf96b-9814-4a16-8590-8579a31e044e · inbound

Prism-Reranker: Beyond Relevance Scoring -- Jointly Producing Contributions and Evidence for Agentic Retrieval cites this paper.

Prism-Reranker: Beyond Relevance Scoring -- Jointly Producing Contributions and Evidence for Agentic Retrieval Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:31:15.662663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T05:16:19.037595Z digest=sha256:f714c7ca74955e485a73130922ca64595c2ffcebe7c1ddf070ae469818664521

Observation e44ff76b-12a4-490f-ac1a-44a2392f5741 · inbound

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost cites this paper.

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:06:09.151521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T10:19:08.451445Z digest=sha256:978a50a740487ca164183cd2ebe8bf6ed29aec3c71acea534b1e109b7f1cb5a1

Observation 616c15e9-9b2b-4ded-b561-e9662df67cc4 · inbound

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark cites this paper.

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:01:26.143631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T04:38:31.924896Z digest=sha256:6eeb3135ea3c5afc235b1927c88f5c9bacbbab5b85d1af49f6de5089e35a1948

Observation f91888a5-3589-4a0f-b4a3-6fa5b7d6b129 · inbound

To Whom Do Language Models Align? Measuring Principal Hierarchies Under High-Stakes Competing Demands cites this paper.

To Whom Do Language Models Align? Measuring Principal Hierarchies Under High-Stakes Competing Demands Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:07:22.268583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T06:06:42.656096Z digest=sha256:d5b66e3ba35d4a57b472137d4c0a0b5492c8817275c468f60bb16160af4b347e

Observation 136caf19-c65b-4c7c-9608-3cb4b412049d · inbound

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing cites this paper.

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:49:44.585762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T04:45:39.641535Z digest=sha256:d70abe31369a6e82632d220e0baec48107f90539a11ed4398513bd9ab144a8ac

Observation da34e246-c6a2-431c-9ea7-a32ef11adcee · inbound

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing cites this paper.

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T21:15:04.193446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T21:12:37.353935Z digest=sha256:2966c56e418dcd689a0acac12c2370a942807f50aaafd123c12d2d672581fe61

Observation e57530f8-53fb-4367-a9aa-998ce65f7a42 · inbound

TCARD: Nearly Balanced Two-Level Designs with Treatment Cardinality Constraints with an Application to LLM Prompt Engineering cites this paper.

TCARD: Nearly Balanced Two-Level Designs with Treatment Cardinality Constraints with an Application to LLM Prompt Engineering Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-21T03:19:28.987733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T03:15:18.040635Z digest=sha256:a7f35fc9932e1d8e6b55105141436610d0987543d1efb21124b844e313d49f48

Observation 42081810-4b80-4862-9263-db86feae0452 · inbound

When Do LLMs Reason? A Dynamical Systems View via Entropy Phase Transitions cites this paper.

When Do LLMs Reason? A Dynamical Systems View via Entropy Phase Transitions Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:15:23.546469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-25T06:14:49.430548Z digest=sha256:28dddfaf73e4afab4ee1a8046d1d402f44adf2806e7dd1f02e95551cf261b492

Observation 9cb02478-11db-4c69-ac8e-5f90f37b08a8 · inbound

Risk-aware Selective Prompting for Hallucination Mitigation in Large Vision-Language Models cites this paper.

Risk-aware Selective Prompting for Hallucination Mitigation in Large Vision-Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:13:27.176298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:12:31.665882Z digest=sha256:3df456faaeae3959ab19a1ff36a241b7c20eb81e410cf96421b41cce310c4709

Observation fb3f2b2f-0776-4ba5-a96b-6a186cf9980d · inbound

Forecasting Future Behavior as a Learning Task cites this paper.

Forecasting Future Behavior as a Learning Task Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T06:07:41.373518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T12:55:39.494339Z digest=sha256:69711391c7c5479ac78cc9650f5499286e9257ed61af9f42ab7c31bf500b22d6

Observation 00cefad6-f34a-41de-a391-21d9ac9d056a · inbound

Enabling Cloud-Level Accuracy in Edge AI through IoT Data Preprocessing cites this paper.

Enabling Cloud-Level Accuracy in Edge AI through IoT Data Preprocessing Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:49:44.088403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T09:34:00.058213Z digest=sha256:819a735e279116d79e66603db833d53335aa640ce2a964f7f43e6f7e5c0fddaa

Observation 6e5b49a3-774c-4feb-a8fe-78d6bfe35ca0 · inbound

Auto-Fill: Learning to Predict Missing Values Accurately with Specialist Language Models cites this paper.

Auto-Fill: Learning to Predict Missing Values Accurately with Specialist Language Models Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T11:40:00.889651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:40:00.889651Z digest=sha256:ad51b6fc1075061840a4461f80b2fea43e4e51f16c179d7169ab98ef8634ea34