Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:23:36.558304Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2601.22900.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:23:36.558304Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:27:32.645757Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-11T23:27:33.239830Z
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9cefae43-278d-4f2d-bf9b-26435b37dd92 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39808f75-d686-4975-b68e-79e1dfb4720f · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7018b70-7633-4544-b1c0-69a81084532d · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23a76f03-07a1-4066-a006-8e58f3aa03c0 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a52a348-46ea-4808-919b-8e4973d06bea · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop = <final answer>’ inside <feedback>.)\n - You MAY include tiny snippets (a short identity, a one-line correction),\n but avoid long derivations or long equations in <feedback>.\n
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fd936f3-d15e-4e22-9d72-b744531ef22a · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 641b4ea2-6f17-4862-817b-1b5daa71fd5e · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2b62e70-85b8-4c26-adcc-d68cd01d0def · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a69218a3-0eed-4330-87e6-b1003b24cecb · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ac0aa15-0fac-49c2-89bf-e2bb33accfd3 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afdb0b6d-24d4-475c-ac52-e69265380952 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a23b851c-55e9-452d-8ab8-4b36d0463768 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be49c4c1-f94c-40f4-9c2e-0d3b87be2e1a · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd0a76b0-391f-4067-bb9f-881d475c0ebf · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5714db7b-b9eb-41de-a31b-afebac613415 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b1d3473-6fe2-4b39-8013-20ff19281333 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c0eb349-cd56-40be-88ee-8b8e95beb06e · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 514016a6-2573-49dc-88bc-6b4cb1684cd8 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop SFT/RAFT system prompt and CITL-FT initial prompt You are a reasoning assistant.\n Solve the problem step by step.\n\n Output format (must follow exactly):\n
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 577a79cf-900d-4da7-8c73-40f546814446 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aac5df6-7992-41f9-b3e8-e06468f823f4 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b98c1b9e-9888-40c3-9cc6-306851cf896e · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 438ab60f-2539-4e78-93bd-ceb78d9abdf8 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8de2b3af-f033-4009-8cfb-b94c3b9c45ba · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c622e5a2-ba65-4ea6-8014-6331d6ba94e9 · outbound
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06f62b64-f665-46e8-a1f2-1cfca9c3763d · inbound
Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.