Pith. sign in

Paper Citation Record · LEDGER

Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.08299.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.08299 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:01:24.053493Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T12:52:17.884464Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 856d1602-f593-4176-85e7-ac6be5519e6b · inbound

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion cites this paper.

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:52:17.886099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T12:50:17.902979Z digest=sha256:783bae8e39ef73e49c955eb7669e44ce6d2d1f34d8e697c06b235fdc43b8872b

Observation 2bb77020-12c5-4c6b-8420-129b419420fe · inbound

LOVON: Legged Open-Vocabulary Object Navigator cites this paper.

LOVON: Legged Open-Vocabulary Object Navigator Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:01:24.053493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:01:24.053493Z digest=sha256:0f981cf365d9497e82bfbfff649a8d91b9f5838e9afd9a2c5f3507d44d69e99e

Observation 2bdec20f-3c79-4083-aa19-c20f56dfd9e5 · inbound

Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots cites this paper.

Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T13:49:08.538358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:49:08.538358Z digest=sha256:a0cc02f4fbd3925aae68f7295cb843a2d46b7368f5293366a2be4194fa1f8d07

Observation 522cdbdd-760a-46ad-b5e2-7003e268af78 · inbound

DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction cites this paper.

DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:03:10.563965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:03:10.563965Z digest=sha256:cbdb89a10597fd47d7cfa9f95279a6a8f2d067d6daaa6cbbe25e4ba51b7f9a68

Observation 38f54183-f9cb-48f4-a584-dbf9ffd60c9a · inbound

Pretraining in Actor-Critic Reinforcement Learning for Locomotion cites this paper.

Pretraining in Actor-Critic Reinforcement Learning for Locomotion Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T10:03:33.681668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:03:33.681668Z digest=sha256:b93dd261f9d08e9175e0c2acc7038a381daa645d94b63de10b569aefb866a525

Observation 43010c18-8556-48e8-b9d2-b435ff33790b · inbound

Learning Agile Striker Skills for Humanoid Soccer Robots from Noisy Sensory Input cites this paper.

Learning Agile Striker Skills for Humanoid Soccer Robots from Noisy Sensory Input Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:21:23.433854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T00:20:43.058346Z digest=sha256:0b3b35b3d2a8c97b2ad8cd63ce555411804db2cafae0c9558a01df6128c0f82e