Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:24:27.883078Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2507.16864.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:24:27.883078Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-09T23:51:47.724033Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
30 of 30 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 9f291b92-9ff4-42fc-892d-ecaf5592c824 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Advancing Reasoning in Large Language Models: Promising Methods and Approaches
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e791a1a2-9549-4505-a103-4353d0e036f4 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 413e08c6-4fc0-45ed-8c58-1983936e23ed · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 03a384d6-6603-4d12-a5ca-d098dad036e3 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Multi-step reinforcement learning: A unifying algorithm
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6da6aa9e-b692-4407-a50d-263863717101 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaffbf05-6329-479d-bf10-f10098639611 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Process-Supervised Reinforcement Learning for Code Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c9a4a8c-d07e-4988-a050-86cb06ce2e57 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning System 2 Reasoning for Human-AI Alignment: Generality and Adaptivity via ARC-AGI
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47f297f6-ad03-4245-9655-2cd4e3090e17 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Mechanistic Interpretability for AI Safety -- A Review
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0356ed62-c5dd-4f6b-91f6-952dee763cc6 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Transformers in Reinforcement Learning: A Survey
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5700f723-f5e9-4840-b5d7-8340f3c30fef · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning A Survey on Transformers in Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fbb0660-e539-4f30-93da-fcf734f3e99a · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Deep Transformer Q-Networks for Partially Observable Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae8e3e84-d7d0-42fe-886a-3e73e6e824f5 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3febb1cf-f523-4cb9-8de8-7d08acdfd19c · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning TransDreamer: Reinforcement Learning with Transformer World Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f2d7c5d-0f5e-4bcf-94f6-8b4c9700e7ec · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3a5f4cac-de24-471d-8db3-3581bef3df7d · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a906c0a1-69e0-4378-930d-5c71a3242d7a · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Understanding the Language Model to Solve the Symbolic Multi-Step Reasoning Problem from the Perspective of Buffer Mechanism
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b41e75c3-aeea-4c77-91fd-e1d67ff4f3eb · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Enhancing Multi-Step Reasoning Abilities of Language Models through Direct Q-Function Optimization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de208ac0-a99d-4c4e-9d92-bd2b09d0416c · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f0e5f976-146d-40fe-9e6a-ced587c64e51 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation efb3d764-7090-4e84-8a54-65a4ec6c0b1c · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8cc50101-0e0b-45a8-81f1-4a2740de302c · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ee9628fd-abe5-4c13-92e9-51b449639072 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Poincar\'e GloVe: Hyperbolic Word Embeddings
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86128f63-7e99-4d89-b410-f9c31ad43f95 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 19b80823-8346-4964-95f5-fd171b1c823b · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning TransMLA: Multi-Head Latent Attention Is All You Need
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c910248f-3e95-48f7-b327-122eeb0e90bb · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0fd0795c-86f6-41e1-b087-1f98c1a8ac92 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cf8a0260-50bd-43bf-a72c-03e63185ed2c · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7e0152b2-7e3e-4348-be31-8bd04f407e02 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Offline Reinforcement Learning for LLM Multi-Step Reasoning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3495c8fb-34ed-46ca-86e3-adfeb0792238 · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Transformers in Reinforcement Learning: A Survey
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e63b8d5-7ed3-4a55-99b9-9c3cde6264ca · outbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d5b84770-5e8f-45f9-bf6d-9cdc8a10a36d · inbound
HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering Reinforcement Learning in hyperbolic space for multi-step reasoning
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.