Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-31T01:31:13.934392Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2607.27610.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-31T01:31:13.934392Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1c25651b-a4b8-4a13-aeb0-bdf29e2d307b · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Geometry3K : A large-scale multi-modal geometry reasoning dataset, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a2e794c-f80d-46eb-a986-786232dd9f50 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Lewkowycz, A
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e42e9094-2009-45f7-8f7d-221c4ca2ea4a · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dda7602f-c96d-4072-b9ad-4d6df2434769 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f2ced7f-357b-4639-a531-89f053522748 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning American mathematics competitions, 2023
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f05134d8-55f1-48e8-a3e1-31d07740462e · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning American invitational mathematics examination, 2024
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f33ec3-3cbe-474d-b9fd-6bda0f7826f5 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning American invitational mathematics examination, 2025
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10637e71-34e5-48c8-a440-83d3d5910da7 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8270d4cd-e7c3-406d-8fa7-f5f095e98d88 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b12c017-dd4e-4f1f-8514-87f5bc8b950c · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 315cf590-106e-40bb-96e7-5b1ee7ab8bac · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning RLHF Workflow: From Reward Modeling to Online RLHF
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb355ebc-f8b7-45be-baec-e2f1b4b86ec4 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Safe RLHF: Safe Reinforcement Learning from Human Feedback
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b45bfb3-0e3c-4249-b41b-dc6be31f009c · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Secrets of RLHF in Large Language Models Part I: PPO
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99c4092e-3461-4594-a35f-0f79c6d9c8b9 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning OpenAI o1 System Card
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d071ac7-eea3-4461-aa21-e75dd95511bb · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa2862f8-7d5d-4704-b836-75d1b55ce5a3 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Proximal Policy Optimization Algorithms
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6bfdf97-439a-4c01-b640-5c1342779ad2 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87f0f94c-469b-47e6-bc78-222da6881ad2 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd59f526-a25a-444f-b760-3d3d64719d89 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eec7c2b-382b-4f92-bdf7-c3502409e3d2 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3494809f-a207-4435-950f-a4ad60214071 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning VinePPO: Refining Credit Assignment in RL Training of LLMs
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cee83150-c3c7-40dd-9acf-6b0203a2add4 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning 2024 , journal =
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3f8a8f8-5aba-43a7-9ab0-4c93541a3f3c · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning LIMO: Less is More for Reasoning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd9d861c-ef6a-4fd8-b5e1-be805572c6a1 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning LIMR: Less is More for RL Scaling
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 702dcb8c-7dca-4c49-a568-d82c8a3d717a · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 792a1834-98fb-490f-9da1-117a075cb528 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ba65db4-5faa-4a7f-86b7-dcf9cc0c9b58 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning arXiv preprint arXiv:2504.05185 , year=
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7df0f2db-156f-4f71-9816-5ef0710f6e4a · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6dbc0bb-644e-4bdb-bb01-bed382450bf1 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Process Reinforcement through Implicit Rewards
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 930b2c14-c88c-4ef8-b913-27f7d821d268 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning CoRR , year=
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0747bebd-a698-4172-b9c2-dd88e52bec97 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning arXiv preprint arXiv:2504.03380 , year=
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b5556f5-4edb-4997-aeee-d5c16e842f80 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning arXiv preprint arXiv:2505.14970 , year=
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36364a43-6149-4376-b5fa-f611755097e8 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adb31b98-a7ef-4df7-8dc0-2e0aa0eb353a · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38ecddab-cd45-44c6-8b44-cf8bcc30b43e · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07682476-ca62-424f-b24f-6345fb6cb488 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning 2025 , note=
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfa8554c-6293-44e7-a104-36f1bc2a8a40 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning arXiv preprint arXiv:2506.06632 , year=
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 377c639e-e952-4a10-8880-d3d53aa9a529 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning arXiv preprint arXiv:2507.04632 , year=
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aebdfc7-abc6-4beb-b6c9-61d9447eab84 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning arXiv preprint arXiv:2510.26374 , year=
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1923581c-b5b5-43b4-9a48-b9c46c031552 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Skywork-R1V3 Technical Report
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 766c45ec-8a99-43e3-a963-464210332c12 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning SARI: Structured Audio Reasoning via Curriculum-Guided Reinforcement Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89327f97-d0b1-43a8-8056-6442326d0c24 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Measuring Mathematical Problem Solving With the MATH Dataset
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a77d461a-1e81-4245-af87-5d07bf159cc7 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Let's Verify Step by Step
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3a1e6b3-a337-4407-918e-62e938e91314 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33b5bc53-574a-4a61-ade5-2ae3eb70cddb · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60a3ae54-73f1-45a5-8f40-5a3bf36882b5 · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Advances in Neural Information Processing Systems , volume=
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1605788e-062c-4a43-bf1f-d7193c42c3ac · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f43e396d-7111-4f40-90fd-1522543e5d7d · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Unresolved cited work
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90c868fa-b460-46cc-b1cf-3190f1bd4b4e · outbound
Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.