Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion

As of 9 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2506.20036.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20036 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:04:22.954743Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T08:48:14.710291Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 333d578b-a6ab-4fbd-ba2e-17357ea098cb · outbound

This paper cites Learning to Walk via Deep Reinforcement Learning.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning to Walk via Deep Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.748967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.748967Z digest=sha256:eb0e4d7188ee28ecdc857bee4669eda4f0bd5fbed776b20f5917ea77cd421957

Observation daa5474c-f09e-4b37-b8b3-29fe1abf5c44 · outbound

This paper cites Learning to Walk in the Real World with Minimal Human Effort.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning to Walk in the Real World with Minimal Human Effort

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.760351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.760351Z digest=sha256:b0af11616b1ad4791b6e1e633ef386c9e1816bfdf686c7373691dad2c75df85f

Observation e9d06fc2-df31-45fa-be57-29e1d76e8880 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning to walk in minutes using massively parallel deep reinforcement learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.766963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.766963Z digest=sha256:162d625b751581c3df44862157278a18db599963ed3b7b80aa64c1429cc6ccac

Observation f0bfe37c-3866-452e-8475-47e3d2a54b7e · outbound

This paper cites Learning fast adapta- tion with meta strategy optimization,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning fast adapta- tion with meta strategy optimization,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:25.169003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.785144Z digest=sha256:e36c0019e621a21c929d0ef8de43dd024a1f32e8fa0852978955537a415cab50

Observation 0a0c4285-9394-447e-8d87-277170ab32fd · outbound

This paper cites Legged locomotion in challenging terrains using egocentric vision,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Legged locomotion in challenging terrains using egocentric vision,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.799138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.799138Z digest=sha256:e40037f4e2f738506c042b112ce795cbbe8c3e4be4aa99fe2bc321d8aa57cd60

Observation fd0a9e0c-9437-4f0f-8f90-e1d92ea97ea1 · outbound

This paper cites Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal Transformers.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.808203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.808203Z digest=sha256:ed54b127051189aa68b5dc8018a34850a9033f6c88525ea2ee7d2ca64c76acdb

Observation b0f45bd1-3c9a-43e8-872a-7ac3102dc1b4 · outbound

This paper cites Policies modulating trajectory generators,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Policies modulating trajectory generators,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.857451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.818994Z digest=sha256:d1643ea5f68f1e44de5f91b54bf5f7525b593467cd9aa0595cf999404bee9920

Observation 8c895a94-052b-4928-9aad-6c3a901fe411 · outbound

This paper cites Learning quadrupedal locomotion over challenging terrain,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning quadrupedal locomotion over challenging terrain,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.834747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.834747Z digest=sha256:d92c5a21a7f18848a004137cef3ae19377c90e7fcb6313e07dafc5b7a6857882

Observation 0397cb74-f44f-46c3-8611-9b99acfbc0fa · outbound

This paper cites Visual-locomotion: Learning to walk on complex terrains with vision,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Visual-locomotion: Learning to walk on complex terrains with vision,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.665324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.844742Z digest=sha256:95617a20c48019a5bdda3bbb5fe1553e7e82ac3ade5fe1fa369bc06a7157431a

Observation 7b72015c-7d05-470d-a77f-4deb24715e94 · outbound

This paper cites Zero-Shot Terrain Generalization for Visual Locomotion Policies.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Zero-Shot Terrain Generalization for Visual Locomotion Policies

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:04:23.198739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.852843Z digest=sha256:5052cc3e204597a92c34d3ae9bec6db3e57ff7e5f09f01f788a13fb01303d499

Observation c4170442-e519-4122-bfae-8d29ba4a321f · outbound

This paper cites Learning agile locomotion skills with a mentor,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning agile locomotion skills with a mentor,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.411864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.860434Z digest=sha256:cc3879b234fe6ec6287978da6eb7f7a0c2b0053abb78bc52a184cb3271ef3115

Observation 41f936ba-2ed3-4416-b529-93719e526df9 · outbound

This paper cites Learning Agile Robotic Locomotion Skills by Imitating Animals.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning Agile Robotic Locomotion Skills by Imitating Animals

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.865462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.865462Z digest=sha256:e1effefad601a10a9415b7d052618bc080476d1d005f7ecb9056a12670eb00d3

Observation 490190c5-0737-4616-9197-d5aa1f5a7118 · outbound

This paper cites Real-time trajectory adaptation for quadrupedal locomotion using deep reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Real-time trajectory adaptation for quadrupedal locomotion using deep reinforcement learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.154395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.871044Z digest=sha256:8f4bcc9819265b980ff6d9476b0658b5b5f2181c8cec5aa5d59799fa9ce8cbb2

Observation 73665f83-3395-4377-966d-2764f7468065 · outbound

This paper cites Guided constrained policy optimization for dynamic quadrupedal robot locomotion,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Guided constrained policy optimization for dynamic quadrupedal robot locomotion,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.876052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.876052Z digest=sha256:7315fc25f8edb4453139ebbb49d57a32e9a74664b001298b91b2747ea6eac2fe

Observation fac83b42-093f-4e70-9f79-15178228f663 · outbound

This paper cites Deepgait: Planning and control of quadrupedal gaits using deep reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Deepgait: Planning and control of quadrupedal gaits using deep reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.880917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.880917Z digest=sha256:a115adbc3161e6cf2126547c9624f4eb797d339f9a493120223648a96efe220b

Observation 659f3e14-aedc-4fec-a466-17373a2407d6 · outbound

This paper cites Allsteps: Curriculum-driven learning of stepping stone skills,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Allsteps: Curriculum-driven learning of stepping stone skills,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:23.744751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.887703Z digest=sha256:7d0a19db4c29fb9852e94989ccff9179dc1737a3cf8694586784bb2a5ceffb65

Observation 2dcdbbc2-1653-43c6-90c6-bc843fb53338 · outbound

This paper cites Learning gen- eralizable locomotion skills with hierarchical reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning gen- eralizable locomotion skills with hierarchical reinforcement learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:23.593917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.892820Z digest=sha256:6e7dfaf2964112e667e8e122467ab91a344a5259a77fea6c59b6894681bdf1bc

Observation a7aaf8dc-b9aa-4de5-86ba-380b572d6619 · outbound

This paper cites Scalable deep reinforcement learning for vision-based robotic manipulation,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Scalable deep reinforcement learning for vision-based robotic manipulation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:23.519591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:04:22.898822Z digest=sha256:77d29e8accdbc47ef5359be02ad27f20ceb9f4af546b5c409cc3a7655e891ad6

Observation c88ac69d-0779-4e24-b38a-7f1f285a8fd8 · outbound

This paper cites Hierarchical plan- ning through goal-conditioned offline reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Hierarchical plan- ning through goal-conditioned offline reinforcement learning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.904750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.904750Z digest=sha256:4d8330e19237201df0f5238235bdef41c95eb48fb5ced1adaa02dc525e50f995

Observation f58c29f1-7e1f-4e03-8604-376fb57eed20 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Proximal Policy Optimization Algorithms

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.917740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.917740Z digest=sha256:b2faf4d714a4d2293036def32b9fa635df7899fc521380af1b01aa9d49bf0658

Observation 11e69b46-db17-4390-b2c5-2987336495f1 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.923351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.923351Z digest=sha256:c02d24d39c7089286a55afa631792b4c722df002789ab8b239c2fe3923fefea9

Observation ab0e3be2-9288-4487-aa34-9c158f627305 · outbound

This paper cites Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.954743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.954743Z digest=sha256:ca2592ae35c990d726fc2d0d269d592bc4f8ecfe71ae9438056add743a422833

Pith citing papers

Observation 1f1692dc-f9fe-4483-9ca6-3963e6a63028 · inbound

PUMA: Perception-driven Unified Foothold Prior for Mobility Augmented Quadruped Parkour cites this paper.

PUMA: Perception-driven Unified Foothold Prior for Mobility Augmented Quadruped Parkour Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T08:48:14.710291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:48:14.710291Z digest=sha256:883cdde3cbad891c5c32561afa450dd24dc3c58dbc9ca50968be9a0caf55807e