Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:57:42.432385Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 23 inbound Pith citation observations for arXiv:2505.11423.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:57:42.432385Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:48:29.166327Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
35 of 35 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation f90a6148-be6a-4cfd-923e-6a5223f375da · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02d6ee03-dbdc-416b-87c8-57bf23427f18 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce271b57-27ef-4bfd-ae73-ae5af10aeaae · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Claude 3.7 sonnet and claude code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d8068035-37b6-4253-ae8c-e7cf702d529a · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdd3f9ce-e150-428c-8487-bcc27b70562a · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Language models are few-shot learners
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7b29e637-250c-4737-9862-6c12d6337d00 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Language Models as Compilers: Simulating Pseudocode Execution Improves Algorithmic Reasoning in Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5918775-7177-4e68-961b-b16de8d2bd2c · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Scaling Instruction-Finetuned Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c087069-ea70-4ed9-8122-d0d0ea0c2d1a · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Complexity-Based Prompting for Multi-Step Reasoning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37342762-b266-4b18-9002-dc49f4c9ca27 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs The Llama 3 Herd of Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 491d17a5-6f4e-4d4c-9253-d231937784cc · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Algorithmic Unfairness through the Lens of EU Non-Discrimination Law: Or Why the Law is not a Decision Tree
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ae3a7afc-466c-4897-81d3-025c667ec69d · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a32da1d3-e84d-4e10-8293-141c0a7e9515 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c20ee43-b435-4a45-9f41-a3fdf5340cc1 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Do LLMs "know" internally when they follow instructions?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc15cd24-b903-4832-bd9e-a88aa71d8bc6 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-Text Rationales
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f635579-8978-4f89-a70a-dc28eec3cfaf · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Position: Llms can’t plan, but can help planning in llm-modulo frameworks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d88d199-8b7b-42ea-bf36-ca470545355d · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Metrology of weak quantum perturbations
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d1b7c971-5168-4298-a2fc-bc96f6783a47 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f8d5873-075f-42a3-9818-892e8b75fcc2 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs The flan collection: Designing data and methods for effective instruction tuning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 967313ac-4c71-4cbb-90e5-c86893314412 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Crosslingual generalization through multitask finetuning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2bda7492-7187-4d93-beea-2b31732383b9 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Show your work: Scratchpads for intermediate computation with language models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d478df3-9534-41ea-9f13-20b35d7d6949 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Introducing openai o1
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a55ec9ce-6c62-4cad-b39b-42753f68db54 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Training language models to follow instructions with human feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d328994-16ad-4ea0-8fa7-a2eb1f762262 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74a1172e-28de-4dae-9393-70882d1ae31b · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Language models are unsupervised multitask learners, 2019
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65e23729-a29a-45e5-9fb6-cde33018e904 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74d28e30-eb18-4dab-a79b-2f59e3e88f20 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Self-instruct: Aligning language models with self-generated instructions
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 96fc1174-f857-46a0-a0c5-23fda575c566 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Mmlu-pro: A more robust and challenging multi-task language understanding benchmark
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1633b4b-9eed-4411-a180-59026c8659b4 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Emergent abilities of large language models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85c62210-a548-4206-9f3e-2fdbf057811b · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Chain-of-thought prompting elicits reasoning in large language models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c3b6a87-d713-4208-a50e-b99c48ac3477 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Benchmarking complex instruction-following with multiple constraints composition
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 91cbc755-63fc-49b7-8403-5e90d9fb4cfb · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs A Comparative Study on Reasoning Patterns of OpenAI's o1 Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94f97e73-01b9-4af6-a7c5-9230ce42c9ac · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Re-reading improves reasoning in language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 215052a1-2c78-4211-a939-400aa24a684d · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Instruction tuning for large language models: A survey
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0375dda8-faaf-4a1c-a1d0-04fff4f066f5 · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Least-to-Most Prompting Enables Complex Reasoning in Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b00f4192-28e5-4c92-8a3c-1774c8988cff · outbound
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs Instruction-Following Evaluation for Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab12ca32-08b3-4294-b6b8-3b87497cfeaf · inbound
Can LLMs Hack Enterprise Networks? Autonomous Assumed Breach Penetration-Testing Active Directory Networks When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb443729-0a07-4c96-bcc1-fb66c56a6af4 · inbound
From Token to Action: State Machine Reasoning to Mitigate Overthinking in Information Retrieval When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d5a8e18-0965-4709-b7b5-188421683c1b · inbound
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f6d90882-42a1-4a02-9b4d-b17aaff211e0 · inbound
On the Surprising Efficacy of LLMs for Penetration-Testing When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74af3947-339f-4281-8dfb-911079c03537 · inbound
AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28d7d64a-746a-41c4-a2ae-0c8af1df3622 · inbound
ProactiveEval: A Unified Evaluation Framework for Proactive Dialogue Agents When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96659764-f58e-4e3c-8f14-3da85b10e709 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 285
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a18bf195-318a-4464-8997-06b27c88909a · inbound
"GenAI Defaults to Bias!" Gamify AI Literacy Through Reflections on Prompts When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f11d1a3-2ba5-4015-b624-475e57d7c34b · inbound
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5912c812-75de-4777-bb03-f95516dfc9aa · inbound
Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c88f4463-35d5-4ac1-821e-ac6fbd29814c · inbound
Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6d6b204-c636-4ec0-ac89-45171800693c · inbound
M3CoTBench: Benchmark Chain-of-Thought of MLLMs in Medical Image Understanding When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed3d7412-c7c4-418a-bb5b-199f77ac88cd · inbound
LightThinker++: From Reasoning Compression to Memory Management When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3c62f23e-7361-4d4b-a9b9-1d85788b7df5 · inbound
Chain of Risk: Safety Failures in Large Reasoning Models and Mitigation via Adaptive Multi-Principle Steering When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 15fdfa06-9bde-47b7-94b5-529da32323fe · inbound
DataDignity: Training Data Attribution for Large Language Models When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e081aa12-7fc9-4bc5-b2ed-5278339ea8f1 · inbound
More Is Not Always Better: Cross-Component Interference in LLM Agent Scaffolding When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4fc68512-6158-45dc-bcad-f86df5b9b6ac · inbound
CLORE: Content-Level Optimization for Reasoning Efficiency When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e557b1c8-cf1e-4057-9d9a-c1152bca9336 · inbound
Prompt Governance? On Governing Technologies Governed by Natural Language When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 192
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 44e21807-6963-43e2-b65c-c109e9fd7e92 · inbound
When Built-in Thinking Helps and Hurts: Constraint-Level Error Shifts in Instruction Following When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 207644e1-4782-47a9-a0b4-48865ba5fce0 · inbound
Observable Patterns Are Not Explanations: A Causal-Geometric Analysis of Latent Reasoning Models When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 603ce71c-60d2-4c17-a412-fcf6d5368fa3 · inbound
Structured Thoughts For Improved Reasoning And Context Pruning When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e9cc34d-0a3f-4738-a879-ee75b7144c61 · inbound
When Reasoning Hurts Legal Drafting: The Verbalization Bottleneck in Patent Claim Generation When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e190c4f2-4725-41c1-a6ff-9c092fa92250 · inbound
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.