Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:43:32.296269Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 27 inbound Pith citation observations for arXiv:2411.19943.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:43:32.296269Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T19:37:15.914764Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T01:14:27.537549Z
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cff932a3-4fae-44d8-a0a8-e6e0e818b577 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3f33ca0-e460-4c96-a9c3-07cab5877c59 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Adversarial Contrastive Estimation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2e3a336b-d35e-412a-a212-196ae26a005e · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7af01e0a-c023-4038-90e9-ac3e09839277 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b99a00e-8979-476f-8f7d-adafd4504621 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Beyond Imitation: Leveraging Fine-grained Quality Signals for Alignment
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 232c31f4-a495-4526-85c3-39f190e2e66e · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dc7fcab-5400-449c-a40a-90a3ecd5c894 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability RewardBench: Evaluating Reward Models for Language Modeling
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c6fef85-2a72-43d5-aa2d-e7a57d611b01 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Let's Verify Step by Step
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f421e724-8314-48ab-9ee7-d48c232a2812 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Provably Mitigating Overoptimization in RLHF: Your SFT Loss is Implicitly an Adversarial Regularizer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acf1ed0b-174e-4538-a19d-d232d546d671 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Contrastive Decoding Improves Reasoning in Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cbd74b3-9cbb-4d9f-9654-84612b71472b · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a78b8149-e8ae-4083-94a3-7d9068f26c9c · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Iterative Reasoning Preference Optimization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17f4f9bd-254e-4139-8cfd-04f87749d09b · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Proximal Policy Optimization Algorithms
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bc083fb-f832-4448-b9f6-6aa3fbb794b1 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41207014-57b7-437e-a048-0c6030066147 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Unchosen Experts Can Contribute Too: Unleashing MoE Models' Power by Self-Contrast
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ee4dcdc-1f38-4535-82d2-4195ccc27e79 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Generalized Preference Optimization: A Unified Approach to Offline Alignment
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ae242d4-d980-40be-bae5-620781c02fea · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Improving Factuality in Large Language Models via Decoding-Time Hallucinatory and Truthful Comparators
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c79e017-8a33-482f-afa1-e6457421ce9a · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Tlcr: Token-level continuous reward for fine-grained reinforcement learning from human feedback
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 680747cf-b625-4563-b078-da4f0e69cde0 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 412575e6-f16c-4f16-b4a6-be6655f7d600 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Token-level Direct Preference Optimization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faf8ac9b-9b10-4ae9-b449-f34530718173 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Alleviating Hallucinations of Large Language Models through Induced Hallucinations
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0099b755-b59f-4b45-ae1c-3c8d4dfa0c37 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Adversarial contrastive decoding: Boosting safety alignment of large language models via opposite prompt optimization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afcbec82-3ff5-42d2-bbca-24375729ce36 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Fine-Tuning Language Models from Human Preferences
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bfd76e0-27fd-4260-830f-4685f2f4eb0b · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Training Verifiers to Solve Math Word Problems
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c7f1a85-a33b-4fb7-b807-b808c40c1571 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Decoding by Contrasting Knowledge: Enhancing LLMs' Confidence on Edited Facts
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b06d89d-3f53-4271-96e6-85480b26305a · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability The Llama 3 Herd of Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7e60f8e-66ab-458b-be00-c1b5ad2aabe8 · outbound
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability Direct Preference Optimization with an Offset
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 294c4aeb-1d41-4e1a-8897-da44984656a7 · inbound
Estimating LLM Uncertainty with Evidence Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dc331b4-4341-4494-8e12-08eb8254157a · inbound
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 570faee5-66db-4c4e-812c-0e3a9d8f6b1c · inbound
Reinforcing Video Reasoning with Focused Thinking Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8892f5b0-4655-4877-9ffb-5ba0cfa528ae · inbound
Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation be8cfd95-abeb-4958-a511-bf33a06779c9 · inbound
Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 926f0a81-4344-4a05-a07c-9d898abd7893 · inbound
RAST: Reasoning Activation in LLMs via Small-model Transfer Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e30123-4810-4801-b43d-cf1aaf244086 · inbound
Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5c76e98-2b08-4706-9725-3efa6118ceb1 · inbound
Embarrassingly Simple Self-Distillation Improves Code Generation Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ebc9c1-3b5b-4a3f-9af1-27f1c10791db · inbound
AtManRL: Towards Faithful Reasoning via Differentiable Attention Saliency Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c373ab1f-1fda-41f9-993f-7c6d710cb6e6 · inbound
Select to Think: Unlocking SLM Potential with Local Sufficiency Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4275dba6-da6f-4463-aef4-5c86ef9df97a · inbound
Select to Think: Unlocking SLM Potential with Local Sufficiency Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 129a9b1a-b312-4ae1-bb0f-4b6908089bc8 · inbound
When Are Experts Misrouted? Counterfactual Routing Analysis in Mixture-of-Experts Language Models Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 789b8bda-373b-47a3-8eb4-797ec833a81b · inbound
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bea7e6b6-1107-406c-85fd-1e971aa23b40 · inbound
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6acf0d4d-e4e7-4a93-bbe8-04ae7f564e4d · inbound
HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 29d311ec-69c6-4df4-8f62-a0ad9db35357 · inbound
Stateful Reasoning via Insight Replay Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1611c923-e0d4-48d7-a207-bd1c408213cc · inbound
Stateful Reasoning via Insight Replay Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 45fff654-df73-47ca-822f-caae21c4d517 · inbound
Token-weighted Direct Preference Optimization with Attention Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8f90ebb5-37ac-4ce0-8576-473783e46ae8 · inbound
Token-weighted Direct Preference Optimization with Attention Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a4016e76-460e-4422-a3b5-7515261b4ea6 · inbound
Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4eb04468-593f-41fe-84a8-8aa6c8dd8d53 · inbound
Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 78698ae4-996a-4c41-a85e-dd2adc39d311 · inbound
The Tell-Tale Norm: $\ell_2$ Magnitude as a Signal for Reasoning Dynamics in Large Language Models Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b56ec185-33d7-4752-8fb3-4665efe65114 · inbound
Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 581eb95a-5873-42cc-b4ab-c9132a3eee2e · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 149
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 48859137-e761-451f-8a44-ae33c494db81 · inbound
Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7831af7e-d003-4267-890d-3cfe74406dc9 · inbound
Beyond Entropy: Learning from Token-Level Distributional Deviations for LLM Reasoning Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bdadb457-eb31-4c85-8b4e-19aedac412ad · inbound
Rethinking On-Policy Self-Distillation for Thinking Models Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.