Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:52:57.142548Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2601.21909.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:52:57.142548Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 03949154-0246-4626-8627-8107394126c9 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d299b128-0f29-4165-b051-41baea0c6acf · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 597b6f69-0fa5-4b2c-a2ba-c6d6f82b2314 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning T., Feltovich, P
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0adb5e2d-88aa-4787-bc8d-36d6c8ffda88 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 824797b2-2eab-49ae-887d-e0890fd45205 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Training Verifiers to Solve Math Word Problems
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd41cb3a-75b2-4c37-9399-d6b65ca77288 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Supervised reinforcement learning: From expert trajectories to step-wise reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c088e9-c972-4317-a839-67884c173e3d · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Theory-based causal transfer: Integrating instance-level induction and abstract-level structure learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df9b6a1c-a1b5-441b-9afe-7731a3d85953 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning PAL: Program-aided Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 881a0dfa-63aa-45f2-b32b-e6a75935e05e · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Structure-mapping: A theoretical framework for analogy
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1055b0a2-c601-4c8f-b63e-53dd459cef29 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38085235-490f-481d-967a-19f282a016ca · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning The Llama 3 Herd of Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1b5cdaa-abd4-4682-ac5e-a3ad8e8a9f89 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df4db203-8bbe-46bf-8811-8df07b58f0c7 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3fd5b45-d18c-46ea-aba0-9bf7d3a6e83f · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Compositional generalization through abstract representations in human and artificial neural networks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84ea9a23-a26b-43c0-8f05-4286be73c63a · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Adam: A Method for Stochastic Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b567197-d207-410b-876c-2ca94fb4ada5 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Mawps: A math word problem repository
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad00f497-f171-44a3-ba37-c4661987a3df · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning LLM Post-Training: A Deep Dive into Reasoning Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 719c9849-4aec-4584-ade7-6c8d7cfd1fb0 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning D., Cohen, J
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b284252d-bd80-4853-b5c9-9039367201bd · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Let's verify step by step
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f778b08-1513-4b56-9814-1927e1be69d3 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Dynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical Reasoning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5234a885-8a0e-45ed-9088-f77498cf2e92 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Towards a unified view of large language model post-training
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 991f4cc9-d33c-47b8-ba09-2ce868d5f25c · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning W., and Behrens, T
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed24ea8-d1cd-4731-b99f-36337b5a6877 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning A diverse corpus for evaluating and developing english math word problem solvers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9b6fbf3-2b7f-4a9c-9775-b33ea1aee5ba · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d77c5eee-0496-4cb1-bdee-1941e1b767d8 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Abstraction and analogy-making in artificial intelligence
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cf0a90d-4661-4c33-a30c-9e30a453e41f · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Training language models to follow instructions with human feedback
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07e30e33-408e-4ecf-a777-7ff068f5df06 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Plan-tuning: Post-training language models to learn step-by-step planning for complex problem solving
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b807f3-2509-471b-bee4-cc3985adb1e6 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Are NLP Models really able to Solve Simple Math Word Problems?
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ada046a-e456-41ad-8116-866789fcc489 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6be17ceb-6d13-4a5d-98d0-563f1a98f517 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Proximal Policy Optimization Algorithms
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f0ce054-c423-4a49-890b-4e73028b022d · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dfbcecf-f88a-41a9-917c-a0789484b3b8 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d225e297-c130-4775-ae51-6de35305665d · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Cognitive load during problem solving: Effects on learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02ef53d2-8ade-4ccb-a886-6e490e394334 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Qwen2 Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25f6e867-6b54-4b22-9a09-7d0307f3f696 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Solving math word problems with process- and outcome-based feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b42ea3d-cb52-4bb2-a765-04e608af4a2d · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Post-Training Large Language Models via Reinforcement Learning from Self-Feedback
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57a344f1-34a7-4e3f-8772-3b5bc2852c5b · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning X., Kurth-Nelson, Z., Kumaran, D., Tirumala, D., Soyer, H., Leibo, J
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42edb151-c342-42d8-acae-edab792b795b · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning A Survey on Large Language Models for Mathematical Reasoning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9fa72b-c343-4503-a22c-74467f29fb60 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning UFT: Unifying Fine-Tuning of SFT and RLHF/DPO/UNA through a Generalized Implicit Reward Function
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c00f3af-7000-416b-bbea-f5d2c84d4de0 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning V., Zhou, D., et al
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98225bc7-fb84-4608-8e11-aab62dc43bc2 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning M., Meder, B., and Schulz, E
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f05880c-5c37-43ad-9651-56f7c529caa2 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 705591ad-5f76-428e-8bec-44a4f72c5fef · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Qwen3 Technical Report
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa4efabc-a5cb-4037-a9a2-d13caf639587 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Star: Bootstrapping reasoning with reasoning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f62c9ae-3988-48bd-8147-6254c80880f5 · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Groundedprm: Tree-guided and fidelity-aware process reward modeling for step-level reasoning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bbcb95d-0914-45b9-bdb9-a641c5ca952c · outbound
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Y., Garvert, M
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.