Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:42:06.520147Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 0 inbound Pith citation observations for arXiv:2509.06024.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:42:06.520147Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
71 of 71 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 22cd9d77-5bc3-43bf-9b6e-012dd366a55d · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL L1: Controlling how long a reasoning model thinks with reinforcement learning, 2025
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9f5930-0553-42f5-9e99-0b969903eda9 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Back to basics: Revisiting reinforce style optimization for learning from human feedback in llms, 2024
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c27a294e-71ce-4d0c-af54-188db47e8299 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69fdc157-bbdc-465c-b495-18e1571d5675 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e793c40-6723-4edc-a07b-443d5e24281c · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Empowering LLMs with Logical Reasoning: A Comprehensive Survey
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 212bd76f-6364-41ed-ac58-8ef82c2f2f98 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Self- playing adversarial language game enhances llm reasoning.Advances in Neural Information Processing Systems, 37:126515–126543, 2024
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0e1d43cc-a800-41fa-9bd3-ba158b18055a · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Transformers as Soft Reasoners over Language
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a96dd2b-7f59-414c-b992-b6f92d214188 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Training Verifiers to Solve Math Word Problems
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09d5c820-15d6-4f1a-9671-70f7e1198aab · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Gemini 2.0 flash thinking, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5269d331-add4-4aea-a46d-b61819dccfba · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a8ec324-7c9d-4337-b5a0-2ec0abbe9e96 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies.Transactions of the Association for Computational Linguistics, 9:346–361, 2021
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e546555a-3bc0-4ed5-b677-ec87a70695d5 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcd6ed91-71c9-4de7-b825-c40330044a9d · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1d3bab2-ec8e-4b06-b5ea-c4ff44d58900 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL FOLIO: Natural Language Reasoning with First-Order Logic
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a19a564f-598b-4599-a236-031e91b38bdf · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Measuring mathematical problem solving with the math dataset
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 355d0c2c-b29a-472a-a57c-407d9fa9959e · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18da2e70-102c-4c7f-8c81-4fd3f7fa4496 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Reasoning-Enhanced Healthcare Predictions with Knowledge Graph Community Retrieval
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6442c8e8-9681-4416-9224-33767f2bbff7 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Boardgameqa: A dataset for natural language reasoning with contradictory information
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2922c9ec-3339-4875-bddc-725e62d6c14e · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Buy 4 REINFORCE samples, get a baseline for free! In Deep Reinforcement Learning Meets Structured Prediction, ICLR 2019 Workshop, New Orleans, Louisiana, United States, May 6, 2019
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ee6cc1d8-bbdd-4b9b-9e66-7843aa934929 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Chain of code: Reasoning with a language model-augmented code emulator
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 03c8f9e5-1a42-40d5-b2da-16e5218e12b3 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a7f6c22-c1a8-4074-a35e-e00a8c6b8587 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Limr: Less is more for rl scaling, 2025
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c61012-3a83-4b03-9eea-bf583e5e64ee · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL From system 1 to system 2: A survey of reasoning large language models, 2025
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29969c6a-139c-4e09-a049-55ed9bb828c8 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8423668f-70ee-447b-833a-9dee553df8a4 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Natural language inference in context-investigating contextual reasoning over long texts
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3e1a4b64-531c-4654-89b3-9a8337c56d83 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Logiqa 2.0—an improved dataset for logical reasoning in natural language understanding.IEEE/ACM Transactions on Audio, Speech, and Language Processing, 31:2947–2962, 2023
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 95a44259-4307-4cc3-8ac1-b9ac02df23cf · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL 大模型逻辑推理研究综述 (survey on logical reasoning of large pre- trained language models)
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 273438f4-5775-49f5-885a-3fef33514fc5 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b5f3609-3501-456a-943d-3574cbc9e3f6 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Understanding r1-zero-like training: A critical perspective, 2025
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e4789af-807a-449b-9d8d-a0c69a8ccd72 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Reasoning with large language models for medical question answering.Journal of the American Medical Informatics Association, 31(9):1964–1975, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1e5db605-f610-426a-a1aa-ccd084413474 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Improve mathematical reasoning in language models by automated process supervision, 2024
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ca36034-17c8-4859-8c4b-7a18a21a1959 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Exploring the limit of outcome reward for learning mathematical reasoning, 2025
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a64b2246-ff11-42ec-b052-3c2b371c4d6d · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Real: Efficient rlhf training of large language models with parameter reallocation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89658dd5-f79b-4562-8d19-88c8603d5be1 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Enhancing reasoning capabilities of llms via principled synthetic logic corpus.Advances in Neural Information Processing Systems, 37:73572–73604, 2024
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9965ee69-3d72-4e05-9625-12449474ce55 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Advanced semantics for commonsense knowledge extraction
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a2abe552-b684-461c-8ee6-e02b32a85bb3 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Learning to reason with llms, 2024
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a124621e-0d8f-45cd-878f-6db8f00f9a15 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36668151-2c8f-4385-8f61-4c85f13e2149 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Making reasoning matter: Measuring and improving faithfulness of chain-of-thought reasoning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c84e8efc-0eb1-408a-b7e6-609ab6f1ef72 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Qwen2.5 technical report, 2025
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc4097e-25ec-4b96-a47e-faf186ca7e4d · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Qwq-32b: Embracing the power of reinforcement learning, 2024
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8703507-c9af-4c01-8e63-1957d0e91145 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0f0895b-9acb-461c-9d78-6c8844fc20b8 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Testing the general deductive reasoning capacity of large language models using ood examples.Advances in Neural Information Processing Systems, 36:3083–3105, 2023
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 53e22eda-0e18-4252-9fd6-6c8b7321b991 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL High-dimensional continuous control using generalized advantage estimation, 2018
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4665d72d-0ab1-45b7-9720-a7522d293317 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Proximal policy optimization algorithms, 2017
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76bf7573-9660-4de6-b51b-193c6c08f8cc · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Automatic solutions of logic puzzles
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e2a07ee8-ad48-43d5-bdb8-7c983b3c3e7d · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3877bd5-b1fc-4b05-8349-1711b53a8446 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8531279f-7ae6-4beb-933a-c0eb75e3c785 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8b7f272-204d-465c-a1aa-7444d57cf1f9 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL MIT press Cambridge, 1998
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f4937f-ea0b-4334-a053-bd4aaf3d7510 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dad3b24a-8875-480f-9a86-f853f97c303b · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Step-by-Step Reasoning to Solve Grid Puzzles: Where do LLMs Falter?
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c0cf8e43-d49e-4adf-a9be-78609e07433b · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d9c6894b-eabb-43b7-9ca1-1543ab1a513c · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Measuring multimodal mathematical reasoning with math-vision dataset.Advances in Neural Information Processing Systems, 37:95095–95169, 2024
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 543ede4a-f228-40e4-8e54-4343af2b16e0 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 71bd38f1-573f-4554-bff1-6b107809ad40 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Legalreasoner: A multi-stage framework for legal judgment prediction via large language models and knowledge integration
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ce955bc-2e1d-48be-a995-f1f6c09e22ea · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Mmlu-pro: A more robust and challenging multi-task language understanding benchmark
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c2e16b8-680b-44d5-8e29-414d7bb94455 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL On memorization of large language models in logical reasoning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36449521-a3f2-453d-a8b9-af3ee859ceb8 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Logic-rl: Unleashing llm reasoning with rule-based reinforcement learning, 2025
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94806828-92e3-4d58-bd00-648bf406b4d7 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95b6e9fb-a1bd-4f56-8651-e9a10cae643d · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Dapo: An open-source llm reinforcement learning system at scale, 2025
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 652cb997-1e1d-4751-bc58-5e4b95ddbf04 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Lawllm: Intelligent legal system with legal reasoning and verifiable retrieval
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a2e9bfbd-2315-4212-ae65-7a9fcd6943c1 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Rest-mcts*: Llm self-training via process reward guided tree search, 2024
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4404c5fd-2fa1-423a-a3b1-419b3d1c9fc1 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb0d3350-691f-4d83-82e6-406ba2ede228 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL R1-reward: Training multimodal reward model through stable reinforcement learning, 2025
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 921924ed-5215-4ebe-ba12-99a806a60fa3 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL AutoLogi: Automated Generation of Logic Puzzles for Evaluating Reasoning Abilities of Large Language Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8179a4a-a3d1-4047-9b5b-e87933e7816b · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 031ed804-4040-4d9e-b669-bb1e4e812f50 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c848a07-871c-470a-92bc-4a2048ff9b50 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 63fdc848-2390-47ce-b156-8f6f6ed1cfdf · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e0dcc09d-4cc0-470b-9de8-7480d9989c85 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5ceff70f-fa52-4110-b4aa-9ba3b33e5eb8 · outbound
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL Alice studies
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.