Pith. sign in

Paper Citation Record · LEDGER

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2603.13707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.13707 v3

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T18:15:50.566587Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T09:27:43.453005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:40:06.781891Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e6b7662f-cde7-401e-9e3a-835639d176c8 · outbound

This paper cites Tai- loring solution accuracy for fast whole-body model predictive control of legged robots,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Tai- loring solution accuracy for fast whole-body model predictive control of legged robots,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.683331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.683331Z digest=sha256:432cedbb12d2a72265b1af5ad8373fd0f97a9a91fcd93cecfd14fb6ede0c6c9f

Observation cb16dc8d-5e96-4cbc-976b-0b522170c050 · outbound

This paper cites Seec: Stable end- effector control with model-enhanced residual learning for humanoid loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Seec: Stable end- effector control with model-enhanced residual learning for humanoid loco-manipulation,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.732651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.732651Z digest=sha256:def95b1116ebacb3ddf09b17a3750350753cdb0ad56ee905aaab663f6d247bdd

Observation 2295d57f-360b-4992-a9d3-82cd7fa09cb2 · outbound

This paper cites Omnih2o: Universal and dexterous human-to-humanoid whole-body teleoperation and learning.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Omnih2o: Universal and dexterous human-to-humanoid whole-body teleoperation and learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.790300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.790300Z digest=sha256:7b578b277aa7622f5dcbb9ae608955f702c79284b59fd3d9df7d65167287d89e

Observation 84b3c53a-dbff-4d45-9410-8d34a65a5fb3 · outbound

This paper cites Humanplus: Hu- manoid shadowing and imitation from humans,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Humanplus: Hu- manoid shadowing and imitation from humans,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.842412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.842412Z digest=sha256:0956240b118ff23559e1735cacb93cbf9e50bca4f472a628849df3693b962dd0

Observation edd77b51-2099-48ca-a89d-46c35361b14d · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Diffusion policy: Visuomotor policy learning via action diffusion,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.914916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.914916Z digest=sha256:e7f8f9b8ead4e3e975fed329dcd54269662e058faa3920b55de8adfacb08d79b

Observation 7d07d19c-11c7-4a8e-a39e-c1fab143f97c · outbound

This paper cites Diffusion policy policy optimization,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Diffusion policy policy optimization,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.991325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.991325Z digest=sha256:6ae8e5c03ae390ee1b3febe6728f7506db48ee6f41c3aed98c12952ad650ad6a

Observation 07be161b-a45e-434d-99a7-ae6c5e9fdf43 · outbound

This paper cites Ppf: Pre-training and preservative fine-tuning of humanoid locomotion via model-assumption- based regularization,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Ppf: Pre-training and preservative fine-tuning of humanoid locomotion via model-assumption- based regularization,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.147433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.147433Z digest=sha256:4e385c3126ba4118809bca611c37aaffc5c956efb3214b02834fef8e07000c00

Observation 47997f94-1242-468f-a641-6d34497aa582 · outbound

This paper cites Humanoid locomotion and manipulation: Current progress and challenges in control, planning, and learning,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Humanoid locomotion and manipulation: Current progress and challenges in control, planning, and learning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.274577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.274577Z digest=sha256:9c03905b64d84f2d231f10a58441af04c3e241fe7602027ba8174ed63add1bdb

Observation a9a935df-5963-4bfc-951e-ffd88bd011e3 · outbound

This paper cites Twist2: Scalable, portable, and holistic humanoid data collection system,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Twist2: Scalable, portable, and holistic humanoid data collection system,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.413631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.413631Z digest=sha256:8e9fdee31db0589b600d5dc91a9838179099150de96709098209555f76a00015

Observation 7ca0f4b8-9443-4b8f-8845-4eded4a28ad2 · outbound

This paper cites Falcon: Learning force-adaptive humanoid loco- manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Falcon: Learning force-adaptive humanoid loco- manipulation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.556653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.556653Z digest=sha256:676e7b9192def54131e8e8324c48714b6b329106c18d86ded5a713eb4c292894

Observation 32751e62-7876-4883-b9af-9eee0fad104d · outbound

This paper cites Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.698722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.698722Z digest=sha256:afc06a50d3c9664374092ee9e231a33c5a10312bb78ec05c9a28371716274069

Observation dca8669d-b1e6-4fca-9040-c7b2c8ed063b · outbound

This paper cites HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.782052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.782052Z digest=sha256:fcbf58491d1bc53803ff453c73051f91618c892b8230b1b26deb9c6f8f79607b

Observation 5f0f052b-fb0e-4015-b6f2-5ba8b3f8715f · outbound

This paper cites Wococo: Learning whole-body humanoid control with sequential contacts,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Wococo: Learning whole-body humanoid control with sequential contacts,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.894940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.894940Z digest=sha256:41d06948dac1c1de3efbe1a2b9adf1f34754863e291181fb710b85f270cfa842

Observation 65a300ab-954f-470b-a9bf-cfa9e50c9245 · outbound

This paper cites Curiosity-driven learning of joint locomotion and manipulation tasks,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Curiosity-driven learning of joint locomotion and manipulation tasks,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.969528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.969528Z digest=sha256:ef22297cc9b07eac3ef862ed96cc289f8159e8a7ca4711593bb2be3f998c2710

Observation 9670c6cc-c8b7-4d52-840a-2308ebb17a2d · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Learning agile soccer skills for a bipedal robot with deep reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.052472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.052472Z digest=sha256:b03cfa51f866cc853ac5066a1a85670235bf36d722519878d5abfe1cf1754e9e

Observation 4ede54d1-788d-4f23-ae2f-730e90754b08 · outbound

This paper cites Opening the sim-to-real door for humanoid pixel-to- action policy transfer,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Opening the sim-to-real door for humanoid pixel-to- action policy transfer,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.158430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.158430Z digest=sha256:30de9eef34ee2c749fecb148fb4d2f220c8184fb9101889209df188e27ce4000

Observation ad6f6515-edfb-4643-b4e8-bb7021c761e3 · outbound

This paper cites Sim-to-real learning for humanoid box loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Sim-to-real learning for humanoid box loco-manipulation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.240523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.240523Z digest=sha256:e2981c591f597a8ca41070998698e331a941178e23a174ce1ff903e1a0d52629

Observation cda389a6-1076-4912-a836-ce51652aa1ef · outbound

This paper cites Opt2skill: Imitating dynamically-feasible whole-body trajectories for versatile humanoid loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Opt2skill: Imitating dynamically-feasible whole-body trajectories for versatile humanoid loco-manipulation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.351210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.351210Z digest=sha256:eb07832bdb94bc109dd6c667571bc487dc2071aa3ea4b7a074204ff59b1efb0e

Observation 5d206791-09fb-47d7-b75b-c5e671578c54 · outbound

This paper cites Hier- archical planning and control for box loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Hier- archical planning and control for box loco-manipulation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.460920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.460920Z digest=sha256:4d53398e5d8465299f2ca0e1522aadd081cc310198e099584a55574e7b03c191

Observation 5673059d-8830-4ad6-9cf8-3cadadb9e32b · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.602126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.602126Z digest=sha256:08e40187391e36483eae2e01109f136433f1d7319c3c21501d6821ebd0ba1f07

Observation f72cddf3-1bdd-4d01-a9ab-da6283afaff1 · outbound

This paper cites Visualmimic: Visual humanoid loco-manipulation via motion tracking and generation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Visualmimic: Visual humanoid loco-manipulation via motion tracking and generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.703596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.703596Z digest=sha256:093ae9d25688425ac6125df89b284a7a5b025ad08d85882caa2beb6eefddf2db

Observation b0343bb9-65fe-4e55-b041-d126507fe94e · outbound

This paper cites Hdmi: Learning interactive humanoid whole-body control from human videos,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Hdmi: Learning interactive humanoid whole-body control from human videos,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.879849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.879849Z digest=sha256:845e8eb00fb0a1455fa86d1c30a01aeaf153635583432a0866eec9778f396777

Observation 305bb488-6dba-4fc5-b99a-b45adbe7e23e · outbound

This paper cites From imitation to refinement – residual rl for precise assembly,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning From imitation to refinement – residual rl for precise assembly,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.992696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.992696Z digest=sha256:9e5591032517b0b2ee3d116bcfcd5d4bad668f045d2d9318e5757e6f8fef72dc

Observation db184c3c-63c1-4d1c-a119-5f2ecc1f6247 · outbound

This paper cites Rfs: Reinforce- ment learning with residual flow steering for dexterous manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Rfs: Reinforce- ment learning with residual flow steering for dexterous manipulation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.131081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.131081Z digest=sha256:96d6950a834245ff846097744b407f4a71c89dd5fe6aff0e5711c18f5241b4db

Observation 2b63d6d6-8220-4e3e-b889-07204454c797 · outbound

This paper cites Residual off-policy rl for finetuning behavior cloning policies,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Residual off-policy rl for finetuning behavior cloning policies,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.220783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.220783Z digest=sha256:a2cbda4516a3b3d26f960cdb950413f76faade34eed15a32b8bdef1d25a938f3

Observation ba798607-ea78-49cd-8cab-0fc7deef9f1d · outbound

This paper cites Proximal Policy Optimization Algorithms.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.360402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.360402Z digest=sha256:89e328e654ef4705a728539797dbab20d8e2bf964f04fa262759ec7f72cd9d5f

Observation a74d813b-ea64-4e86-8bcb-913bf209e99a · outbound

This paper cites Efficient online reinforcement learning for diffusion policy,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Efficient online reinforcement learning for diffusion policy,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.499122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.499122Z digest=sha256:d36939f7ea4a053a8cddeff200f97eefdcacb9e6c07cbe436d7bb0e3f97ab8f5

Observation 2d72b892-502b-4cf9-859d-f9f25451a1e1 · outbound

This paper cites π RL: Online rl fine-tuning for flow-based vision- language-action models,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning π RL: Online rl fine-tuning for flow-based vision- language-action models,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.582649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.582649Z digest=sha256:87116356e3b42a3ca5d527d972ea587b92d8ba1076ad427614d5f1dc6657d5b5

Observation 71651857-4908-47c2-b5e4-35b53fd4ce40 · outbound

This paper cites Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.694818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.694818Z digest=sha256:27ba76be6827ff0f583db6ff62edfef04887673604433c51ffa65a7d42d9b262

Observation c50a1bac-1116-4886-b20a-53a708757a20 · outbound

This paper cites BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.743659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.743659Z digest=sha256:51689530d631f19b2d10423e811b38e5e582b1813a53af3ec71e1addf5d15f4a

Observation 6331e3f7-12eb-42d1-bd60-c692afb82ef9 · outbound

This paper cites Re- inforcement learning-based footstep control for humanoid robots on complex terrain,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Re- inforcement learning-based footstep control for humanoid robots on complex terrain,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.850902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.850902Z digest=sha256:c96815c40d41ca81b0f052d54138b6ebd17ebd6540b847e79d98a5a2f04af706

Observation 9a56f533-4561-4441-948a-38df0a363957 · outbound

This paper cites Deepmimic: example-guided deep reinforcement learning of physics-based character skills,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Deepmimic: example-guided deep reinforcement learning of physics-based character skills,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.972201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.972201Z digest=sha256:d1294efe6f1b132ef60d8db22d37b1785700f181fb66920f45958c7ea0ccce7d

Observation 9b0f6f8a-4800-4fe7-bcaf-7c045b10541a · outbound

This paper cites Denoising diffusion probabilistic models,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Denoising diffusion probabilistic models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.018839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.018839Z digest=sha256:9d1de2c32e1fbb6348557e7f7a8f227c3879c86b27168fa9c154dcdeed3f9906

Observation e83cbd15-2b66-4150-b438-78460c0ee39d · outbound

This paper cites High- dimensional continuous control using generalized advantage estimation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning High- dimensional continuous control using generalized advantage estimation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.103256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.103256Z digest=sha256:f92c32e722a1818b211eb471daaebc22ff59a8064c6ed32be82339976e2e076b

Observation 8284605d-37e2-41f7-ad39-7cd5e70c5cb0 · outbound

This paper cites Improved Denoising Diffusion Proba- bilistic Models,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Improved Denoising Diffusion Proba- bilistic Models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.196798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.196798Z digest=sha256:cc5ec3c0b3dee4022c0bf0f85833152c6ba7de2b0a2d5534bdd3421934d2ef27

Observation 00433e31-41fc-4383-aaae-c9853f160f01 · outbound

This paper cites Stageact: Stage-conditioned imitation for robust humanoid door opening,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Stageact: Stage-conditioned imitation for robust humanoid door opening,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.306809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.306809Z digest=sha256:422dd10091b5569b2e6cccc80266f491d9d773c404718e6c054a01fea4a3cb23

Observation 1d79d6f4-75fa-4a49-bd1e-2b3d38e2fc76 · outbound

This paper cites A behavior architecture for fast humanoid robot door traversals,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning A behavior architecture for fast humanoid robot door traversals,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.566587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.566587Z digest=sha256:5c4c6b33c550d513a27a0b9c261a820238f41a58ffd446e90a95edeb2d29389e

Pith citing papers

Observation f2339b25-7a1a-432e-bffe-48e6c843d5f8 · inbound

EgoEngine: From Egocentric Human Videos to High-Fidelity Dexterous Robot Demonstrations cites this paper.

EgoEngine: From Egocentric Human Videos to High-Fidelity Dexterous Robot Demonstrations REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-31T02:03:17.213017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T09:27:43.453005Z digest=sha256:07bd4a868d3026c63e6faadd885c4cf440d9f4f2d603b21b7a5b5d5ea026166b

Observation a50faa90-7ed9-4a1f-bec1-50924d3641f0 · inbound

Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots cites this paper.

Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-31T02:03:17.213017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T21:08:43.127148Z digest=sha256:29466dae1cbabf227aa8cc4570878f9c1b7e52f38cbfd378eae2cec23911948a