Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation

As of 10 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2506.02206.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02206 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:50.317068Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T22:35:37.040638Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5e55a5b1-3656-4b35-a091-6a474d898f24 · outbound

This paper cites Introduction of the Foot Placement Estimator: A Dynamic Measure of Balance for Bipedal Robotics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Introduction of the Foot Placement Estimator: A Dynamic Measure of Balance for Bipedal Robotics,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.551834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:45.982721Z digest=sha256:8488b5b2f430cc798ac0d95e7580b0ba536e59cf7273c807a21c0c3213439344

Observation e184e266-55c1-446f-a646-fa74e451b99a · outbound

This paper cites Navigation planning for legged robots in challenging terrain,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Navigation planning for legged robots in challenging terrain,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.306213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.067955Z digest=sha256:e3b7e33ff6257dfe7538ff8dad17cb6cccb11f228a67531769caa177bcc6ec93

Observation 787cc79f-2b84-4667-857e-08cc2e5eca42 · outbound

This paper cites Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:46.171453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:46.171453Z digest=sha256:740e0e0712e1522a72466214d673a194766451892fffa54801dd70b278e0b8f5

Observation 43d458b8-00df-4988-b1d6-6210be1fec8c · outbound

This paper cites Exact cell decomposition of arrangements used for path planning in robotics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Exact cell decomposition of arrangements used for path planning in robotics,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.032791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.263030Z digest=sha256:ad70f9578431c59bd83b4ffd647583d142e6d953c6af34e86535822df4a47edf

Observation 91759f8a-8b67-4a89-a046-423966bdbc46 · outbound

This paper cites An overview of autonomous mobile robot path planning algorithms,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation An overview of autonomous mobile robot path planning algorithms,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.783635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.395990Z digest=sha256:a1f684e6a3d0c5fd93c8de6a5cea60870ef13b1e44e46fc31e640939a52fd893

Observation d231c70e-f21f-46b6-ad9c-740947232dfc · outbound

This paper cites Path planning and trajectory planning algorithms: A general overview,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Path planning and trajectory planning algorithms: A general overview,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.536460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.547721Z digest=sha256:0c312e7444997ae6d9774a580d55dcca0b8cf8e07bde98a6ddb50a5eb393fd7b

Observation 506751ae-76fb-4abb-bdf9-83af3c55f446 · outbound

This paper cites Confidence random tree- based algorithm for mobile robot path planning considering the path length and safety,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Confidence random tree- based algorithm for mobile robot path planning considering the path length and safety,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.205520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.715004Z digest=sha256:2cc7d170bf3dde098c2cda7873e192d0f5c9dbd71f154170a6c21a45d10cfc87

Observation cccc64f7-cb5b-4dcd-a7eb-6f09f0d3a155 · outbound

This paper cites Optimization-based motion planning for legged robots,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Optimization-based motion planning for legged robots,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.930491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.882468Z digest=sha256:2f5172e430e1c6ac1a35cc531d3345b855293e3017b10cc03b93c86495ab7ce9

Observation 1fe19592-9ca4-4e45-b327-3d3130da7bb5 · outbound

This paper cites Integrated task and motion planning for safe legged navigation in partially observable environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Integrated task and motion planning for safe legged navigation in partially observable environments,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.644239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:46.984997Z digest=sha256:bf85d93876fed4e3821b8f6a2fd48f65976c6829b6d4bc63b422a92264462ca5

Observation 185c0dd2-a53c-4d03-9b02-a89681136225 · outbound

This paper cites Unified Path and Gait Planning for Safe Bipedal Robot Navigation.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Unified Path and Gait Planning for Safe Bipedal Robot Navigation

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:33:50.614740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:47.137767Z digest=sha256:8e6465b0a459afd46d50dc7195cf67049908c0f5219da2cf8a5337ef4629bda3

Observation ee89f1e1-1319-4112-901e-3f9ca352ce98 · outbound

This paper cites Fast direct multiple shooting algorithms for optimal robot control,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Fast direct multiple shooting algorithms for optimal robot control,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.421021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:47.253223Z digest=sha256:2cf26a144aae74c99d0decc624d0b530045eacd6db01d0956bdcb5bec96c9f5c

Observation 4a83a47b-1afa-4501-b94b-438e51b99606 · outbound

This paper cites Using optimization to create self-stable human-like running,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Using optimization to create self-stable human-like running,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.146851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:47.401679Z digest=sha256:bd17b3b51c961f7b669e3c44b309aa6c2ce5bcd8134bdfb88763e23896da7ad8

Observation 4073e0d1-07d9-4e0d-81eb-8374028d13bf · outbound

This paper cites Whole-body motion planning with centroidal dynamics and full kinematics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Whole-body motion planning with centroidal dynamics and full kinematics,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.878671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:47.499247Z digest=sha256:29b005d3b423cd5292252046469edfe0f62b50aba514a5bb3165ba6dbe668dd9

Observation ad922e4f-c540-40af-b183-2aefc9c91cae · outbound

This paper cites The 3d linear inverted pendulum mode: A simple modeling for a biped walking pattern generation,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation The 3d linear inverted pendulum mode: A simple modeling for a biped walking pattern generation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.561134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:47.670448Z digest=sha256:ad0a0a344ea8ea8eb9795bbbbcc0768821faa864b2128c8105c543fbf9e331ef

Observation 339963da-d9e3-4b9f-8617-f27af015526f · outbound

This paper cites Bipedal walking control based on capture point dynamics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Bipedal walking control based on capture point dynamics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.298431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:47.823455Z digest=sha256:6301ffd68eae024f404ee8dc79835e53b7e368bf2cd389c774ad4d57624500e5

Observation a56aac47-df3a-4a49-aeea-6703e77f7eff · outbound

This paper cites Nonlinear model predictive control for rough-terrain robot hopping,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Nonlinear model predictive control for rough-terrain robot hopping,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.040334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:48.004964Z digest=sha256:cccad5c2486eceda268651179dd9ab483476e76cfd38cca2c6b5c9ae15f34bb5

Observation dbcb9620-946d-467a-9667-d332fcf5a736 · outbound

This paper cites Perceptive locomotion through nonlinear model-predictive control,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Perceptive locomotion through nonlinear model-predictive control,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.116002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.116002Z digest=sha256:459aed51385fb98c17ab1c98112a3e2a68bee62ba1632595bb7bc520dd27d0be

Observation c1f7cd1c-0a89-480b-b5ec-f3d9243b81f6 · outbound

This paper cites Model predictive control for dynamic footstep adjustment using the divergent component of motion,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Model predictive control for dynamic footstep adjustment using the divergent component of motion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.786744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:48.254270Z digest=sha256:8460cdd4a0f04a1cad7252b1d7d2f54956843bfa5f82663d270369209b3e72f3

Observation 566d52c9-0baa-45f1-ae10-292bcd563abf · outbound

This paper cites A sequential mpc approach to reactive planning for bipedal robots using safe corridors in highly cluttered environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation A sequential mpc approach to reactive planning for bipedal robots using safe corridors in highly cluttered environments,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.362816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.362816Z digest=sha256:ad5fe104d5cebf29dbb0582343902bb448708bcde6535dd3c145100297611767

Observation 4a85c47e-894c-43b0-a189-b32e9e7d9491 · outbound

This paper cites Real-time safe bipedal robot navigation using linear discrete control barrier functions,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Real-time safe bipedal robot navigation using linear discrete control barrier functions,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.514967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:48.481563Z digest=sha256:08f9f9dc7b7ff1c75b4772af6bb49b36b14d0ff59b4cc81532c1f4817bc79efb

Observation 4076a78d-f047-4957-b732-9d468378f1f2 · outbound

This paper cites Apprenticeship learning via inverse rein- forcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Apprenticeship learning via inverse rein- forcement learning,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.595799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.595799Z digest=sha256:9374579ac3426291e131b53924bf2d3034e0f87bd8473563d811100745ad4c36

Observation 0a5c7f4f-a9f7-4c4a-bd3d-8e2a1b241ac1 · outbound

This paper cites Apprenticeship learning using linear programming,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Apprenticeship learning using linear programming,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.238377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:48.692116Z digest=sha256:dd9abf6ae8d09ba20ce6a9bc373bba2fa032260f97c6b3d173e73948bf0683eb

Observation eaa33a93-53ab-4b61-918f-3256a4af3882 · outbound

This paper cites Generative adversarial imitation learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Generative adversarial imitation learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.843022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.843022Z digest=sha256:68d98d0aee30c2ae7fe2de8cb77c81318aa3893f09fe150683ac70ec76b1bb34

Observation a9d0184d-8bb5-4e04-be9a-4bd07c948efb · outbound

This paper cites Goal-oriented obstacle avoid- ance with deep reinforcement learning in continuous action space,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Goal-oriented obstacle avoid- ance with deep reinforcement learning in continuous action space,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.999533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:48.937422Z digest=sha256:ca8edcb254cd6ca12dd45ffb225309d910feee4d6ce7c226023d23fd4827faa8

Observation 32780fab-6113-4503-bb13-19ae90ccba2e · outbound

This paper cites Where to go next: Learning a subgoal recommendation policy for navigation in dynamic environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Where to go next: Learning a subgoal recommendation policy for navigation in dynamic environments,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:49.072122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:49.072122Z digest=sha256:c5ce9060ecd4a66aee99d412a692cc4d3ed587169969bc275200405a5d924393

Observation 5cb83d8d-a174-43ae-88c1-8d1898170a7e · outbound

This paper cites Robot navigation in constrained pedestrian environments using reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Robot navigation in constrained pedestrian environments using reinforcement learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.724840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:49.185856Z digest=sha256:0daa74203b1acf3e4268b0db004e38839bab479dc0fcfc4787837ed14c77ec16

Observation 072de955-c836-4511-9ba6-135d96e9f8a8 · outbound

This paper cites A hierarchical deep reinforcement learning framework with high efficiency and generalization for fast and safe navigation,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation A hierarchical deep reinforcement learning framework with high efficiency and generalization for fast and safe navigation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.486999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:49.305115Z digest=sha256:139f5019f584b1c2040023cdddc0161a93942c080e708e7be857532375254ded

Observation b2b485ca-9249-46b1-8473-6464a8bded44 · outbound

This paper cites Drl-vo: Learning to navigate through crowded dynamic scenes using velocity obstacles,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Drl-vo: Learning to navigate through crowded dynamic scenes using velocity obstacles,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.209570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:49.427620Z digest=sha256:78ee07e28b1f48be23278a553e261b400f0a030916e15a75d1df40f12e7b4d84

Observation 40c9287a-bc88-4cbb-8fe5-282da52f7ce6 · outbound

This paper cites Efficient deep reinforcement learning with imitative expert priors for autonomous driving,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Efficient deep reinforcement learning with imitative expert priors for autonomous driving,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.950891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:49.525766Z digest=sha256:7f27a3e2767f0d3499455e62c7429f8befa7153d6966121f9bda796f83b3fa56

Observation 7a35d118-85a1-466f-952a-eb6c61418154 · outbound

This paper cites Pre-training goal-based models for sample-efficient reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Pre-training goal-based models for sample-efficient reinforcement learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.691359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:49.640866Z digest=sha256:ed2246f26d109d07c022431003d85f20627c6cc163be30a59ac451e51bb00b80

Observation 80730197-6df9-4edf-a685-df9ee676e76f · outbound

This paper cites Deep q- learning from demonstrations,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Deep q- learning from demonstrations,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.413355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:49.819870Z digest=sha256:0183628075ee2e8ccdbe8250acb19f6f3f5aba498b2ed84b12329d2f9b7f3f8f

Observation 9c54604b-f17c-43a5-8cb0-8778ed0b5ecc · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:49.972886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:49.972886Z digest=sha256:62b7870cd35bbd9ea673477c65b4398388e3a169707e5f1fc2d3e000e8ed95be

Observation 5b7a965f-8aab-4297-bc93-4da4edbc2279 · outbound

This paper cites Template model inspired task space learning for robust bipedal locomotion,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Template model inspired task space learning for robust bipedal locomotion,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.141854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:50.057744Z digest=sha256:04f123a67f15571e912a4974b745658316301a5fb62f3bfb24c878862b7985a2

Observation e4b73840-65c4-4e0a-94c3-20e761fdb8b3 · outbound

This paper cites Conservative q- learning for offline reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Conservative q- learning for offline reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:50.895121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:33:50.168382Z digest=sha256:4650a870fdd878b7356a560614a819d6f8303b93738de9f516e7974fba6fc8a4

Observation a809ba34-adc7-45de-8f63-eb05135ab05c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Proximal Policy Optimization Algorithms

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:50.317068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:50.317068Z digest=sha256:3e905d86b40cf2bb27f4579b6ec567b34e6c8e1b5d6d7a3387a12109eaf215f0

Pith citing papers

Observation b3964702-4ab5-47bb-ba95-e3886eecbf7c · inbound

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC cites this paper.

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T22:35:37.040638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:35:37.040638Z digest=sha256:3e986916aedbe21ee43dc8b067c879e19c4eb04bb0f29bbee4a98c276f5fdf20