Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 58 inbound Pith citation observations for arXiv:2312.06585.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T14:27:51.152051Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T21:00:08.414065Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ab4126f8-e276-4acd-b497-4213e1f02631 · inbound
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 84baa3d2-b09e-4f5a-8843-044468b627a2 · inbound
Training Language Models to Self-Correct via Reinforcement Learning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eb350d8c-a9ab-4afd-9829-fd1f3060fba9 · inbound
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36080728-174d-4ffb-b838-7c44684f950c · inbound
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 206
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b96d62a6-a0d1-4148-bbf0-1a7fce3401f7 · inbound
Demystifying Long Chain-of-Thought Reasoning in LLMs Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2c04f306-b46e-4265-ba4b-088a53fde1bc · inbound
Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f91fb950-3512-4d62-9f44-2f9d72aaa330 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 192
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e247596f-c4d8-45a6-9b26-2b9f341b0783 · inbound
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2bcb275e-3172-40ef-9dad-2bf894550a60 · inbound
Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 159
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d1f3863e-7606-483e-a643-b3fc8dfbc7fe · inbound
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b550a2a-2d34-4a83-a7ac-7634953fc29d · inbound
Large Language Models for Planning: A Comprehensive and Systematic Survey Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 215
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6128fac-ad80-4c3e-8161-404773e5b32b · inbound
Maximizing Confidence Alone Improves Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d04703e-ce4e-42be-a82d-05ff6f566335 · inbound
HardTests: Synthesizing High-Quality Test Cases for LLM Coding Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5e4dfcb-47a4-4f45-a5dd-578774dff898 · inbound
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 607c30cf-eaaf-4c8f-8c72-1a6015fcdd00 · inbound
CIIR@LiveRAG 2025: Optimizing Multi-Agent Retrieval Augmented Generation through Self-Training Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 500a8237-3561-48aa-9da6-c5369e86188f · inbound
Spectra 1.1: Scaling Laws and Efficient Inference for Ternary Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60399d37-3763-4d34-9780-b5231395c56f · inbound
SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71059b91-6525-46b3-98d2-7379c09ed54a · inbound
Utilizing Training Data to Improve LLM Reasoning for Tabular Understanding Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87b836da-7df4-457d-a007-3dbb733fc341 · inbound
ReST-RL: Achieving Accurate Code Reasoning of LLMs with Optimized Self-Training and Decoding Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 804b351d-d527-4650-a4b3-d6ec3000e6d0 · inbound
Learn from What We HAVE: History-Aware VErifier that Reasons about Past Interactions Online Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ffe2ae4-5226-4f86-b0fb-2daf5a0086f6 · inbound
Bridging the Capability Gap: Joint Alignment Tuning for Harmonizing LLM-based Multi-Agent Systems Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5abc5f4-acb5-49b1-a537-180ae38e0bbb · inbound
rePIRL: Learn PRM with Inverse RL for LLM Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 790c996a-e57a-453a-a2cc-5e7228b8c615 · inbound
rePIRL: Learn PRM with Inverse RL for LLM Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f340510-3d13-47fd-b388-cdd51ecdeb84 · inbound
A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c37dfd8-83f4-4fe2-b809-033c678ab898 · inbound
Turbo Connection: Reasoning as Information Flow from Higher to Lower Layers Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 924abd59-3f5f-43c1-a9ff-3d048b97ebd4 · inbound
Space Syntax-guided Post-training for Residential Floor Plan Generation Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c391cbcd-6896-4f3d-856f-c02b0e90f75a · inbound
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 734bac73-6aee-4772-a492-e877d7328eae · inbound
$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation af2a7b0e-a535-4cb4-814f-64525ee14ee4 · inbound
$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 83deb363-64c7-4970-958d-fc6bd40d1a98 · inbound
Segment-Aligned Policy Optimization for Multi-Modal Reasoning Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b8c432ad-a1c4-4228-8f84-5e9a210351c7 · inbound
Beyond Negative Rollouts: Positive-Only Policy Optimization with Implicit Negative Gradients Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4df2d3ef-ea52-4954-ac9b-091668ef48f3 · inbound
SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2d8e60a0-23ff-491b-888b-6e971a037740 · inbound
RISE: Reliable Improvement in Self-Evolving Vision-Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7fbeceff-a5d9-481c-975e-d297f43c2bed · inbound
RISE: Reliable Improvement in Self-Evolving Vision-Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 66f7a843-37d7-49f4-bd12-4b296193d4ae · inbound
Self-Policy Distillation via Capability-Selective Subspace Projection Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5eb8abed-a77d-444c-800a-cad6063f909a · inbound
Peak-Then-Collapse and the Four Interface Channels of Knowledge-Graph Tool Use Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 162d4ffc-93b7-43b0-9ec3-d2daf9973525 · inbound
DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b1ad4b1d-ebb2-4778-83a0-a03e9fe21758 · inbound
Robust Reasoning via Dynamic Token Selection for Distribution-Aligned Self-Distillation Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 71a6711c-6d6e-4d83-8eba-186cf78d6b44 · inbound
Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 667f03c4-2391-4a3c-bcd0-3a76d4a86077 · inbound
Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e3ccd73e-cf5c-42c6-8855-488529460e81 · inbound
DRIFT: Refining Instruction Data via On-Policy Data Attribution Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f8137536-2793-4e73-828e-8f691d7d3b6d · inbound
Data Selection Through Iterative Self-Filtering for Vision-Language Settings Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dcf81913-45a8-4a5f-bf89-88c9345346a9 · inbound
On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c88cb04e-be10-40ae-bda7-3e86ffdc71c1 · inbound
Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6cb802be-2f28-43ef-a056-6c9296fdc956 · inbound
PHF: Privileged Hidden Flow for On-Policy Self-Distillation Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bd21824f-69a1-4582-8962-2589159d7811 · inbound
Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 90ab09a6-18ef-4af7-bb5b-b663d1979ec4 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 205
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5c001a8-f646-4831-b290-bdd4ceadd008 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 206
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d3c46f9-bc77-4557-bd9f-699103e384a7 · inbound
The Verifier is the Curriculum: Execution-Gated Self-Distillation for Cross-Family Game Generation Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13dd479b-e987-4d72-adad-948b0c3a7dfb · inbound
UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0ff153d-3cab-4e56-aead-9b7e7bd48a79 · inbound
Post-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape CoT Calibration Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 106
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1620c7f-8ad5-43ff-bc27-7b2489a9d2b6 · inbound
Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f75ed8b3-e8bc-4b2e-9605-2afc67667824 · inbound
CoTu at EXACT 2026: Neuro-Symbolic Reasoning for Transparent Educational QA Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8601a47d-6b23-4ca9-8ba9-1fc876d16795 · inbound
Off-Context GRPO: Learning to Reason on Hard Problems using Privileged Information Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0be574e1-6fbf-4c6b-96a8-61d630984dcf · inbound
FormulaSPIN: Self-Play Fine-Tuning for Natural Language to Spreadsheet Formula Generation Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05e6d4e8-23ee-4cdc-b797-a7c0281ffebe · inbound
Bridging Compute- and Data-Optimal Pretraining Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 171
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f08c263-545e-4637-b701-934097b26025 · inbound
From Scoring to Acting: Outcome-Verified Comparative Self-Distillation for LLM Agents Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2689870-d8ac-4522-a585-d2f1e16e5a9b · inbound
Recursive Synthesis for Long-Horizon Terminal Tasks Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.