Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation

As of 23 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2506.02206.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02206 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:50.317068Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T22:35:37.040638Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5e55a5b1-3656-4b35-a091-6a474d898f24 · outbound

This paper cites Introduction of the Foot Placement Estimator: A Dynamic Measure of Balance for Bipedal Robotics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Introduction of the Foot Placement Estimator: A Dynamic Measure of Balance for Bipedal Robotics,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.551834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:45.982721Z digest=sha256:e40fba774775c9d9911edd84eb28562b4ed4f15568ba0617d19aea14b197c9c9

Observation e184e266-55c1-446f-a646-fa74e451b99a · outbound

This paper cites Navigation planning for legged robots in challenging terrain,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Navigation planning for legged robots in challenging terrain,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.306213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.067955Z digest=sha256:96728abefd561804f51ecdb9335e0c7480b328979f3361c84fcd01d19df26b07

Observation 787cc79f-2b84-4667-857e-08cc2e5eca42 · outbound

This paper cites Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:46.171453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:46.171453Z digest=sha256:cec1db79734592bf96bf0023e385c692f8fb653b24f8c020e706f3fadce99d82

Observation 43d458b8-00df-4988-b1d6-6210be1fec8c · outbound

This paper cites Exact cell decomposition of arrangements used for path planning in robotics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Exact cell decomposition of arrangements used for path planning in robotics,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.032791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.263030Z digest=sha256:c4b2f40c72f4de737f93a0dcbeca36095be99919a7250d6c2916e3d21a4487c4

Observation 91759f8a-8b67-4a89-a046-423966bdbc46 · outbound

This paper cites An overview of autonomous mobile robot path planning algorithms,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation An overview of autonomous mobile robot path planning algorithms,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.783635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.395990Z digest=sha256:f3e4106c6c735c7b945f52149c14cf9fa6a4ebc85f8e14eda6e4cb61403793d1

Observation d231c70e-f21f-46b6-ad9c-740947232dfc · outbound

This paper cites Path planning and trajectory planning algorithms: A general overview,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Path planning and trajectory planning algorithms: A general overview,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.536460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.547721Z digest=sha256:ba894cc415a7172ef7dfc33039a7964e14168404bac5e3f8c6863ac8adec9c8b

Observation 506751ae-76fb-4abb-bdf9-83af3c55f446 · outbound

This paper cites Confidence random tree- based algorithm for mobile robot path planning considering the path length and safety,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Confidence random tree- based algorithm for mobile robot path planning considering the path length and safety,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.205520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.715004Z digest=sha256:18335e41804ff6577322dfb1c6f945b8d08827c3bdb399405e2408835c89c56d

Observation cccc64f7-cb5b-4dcd-a7eb-6f09f0d3a155 · outbound

This paper cites Optimization-based motion planning for legged robots,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Optimization-based motion planning for legged robots,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.930491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.882468Z digest=sha256:6ef9565a1f822f974ad76c3a354c26441a44a453028b153c43d72ee7c0a7ce6a

Observation 1fe19592-9ca4-4e45-b327-3d3130da7bb5 · outbound

This paper cites Integrated task and motion planning for safe legged navigation in partially observable environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Integrated task and motion planning for safe legged navigation in partially observable environments,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.644239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:46.984997Z digest=sha256:ef4cc7acff817e21af8354df4a32ef4002d74550d4e8bcb6c869b03947b7e7dc

Observation 185c0dd2-a53c-4d03-9b02-a89681136225 · outbound

This paper cites Unified Path and Gait Planning for Safe Bipedal Robot Navigation.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Unified Path and Gait Planning for Safe Bipedal Robot Navigation

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:33:50.614740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:47.137767Z digest=sha256:8d58ef763da8744ebb0a8a79e46d66497bf59c4c0ee50f82d5e903ac081180d0

Observation ee89f1e1-1319-4112-901e-3f9ca352ce98 · outbound

This paper cites Fast direct multiple shooting algorithms for optimal robot control,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Fast direct multiple shooting algorithms for optimal robot control,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.421021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:47.253223Z digest=sha256:3a9fcfcd3e6f47a3b3e0d47c8e3bec70f5d6acab05a371f276ab21689a784bae

Observation 4a83a47b-1afa-4501-b94b-438e51b99606 · outbound

This paper cites Using optimization to create self-stable human-like running,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Using optimization to create self-stable human-like running,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.146851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:47.401679Z digest=sha256:cfe318ab22d05d7e63e4db18dd8c16b8c6fc7dc6dedc9d13dbed5b9b105056f1

Observation 4073e0d1-07d9-4e0d-81eb-8374028d13bf · outbound

This paper cites Whole-body motion planning with centroidal dynamics and full kinematics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Whole-body motion planning with centroidal dynamics and full kinematics,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.878671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:47.499247Z digest=sha256:6b2705b828b40310a608549ba3fe57392e06353a5d2304dcc59e50f0738ed419

Observation ad922e4f-c540-40af-b183-2aefc9c91cae · outbound

This paper cites The 3d linear inverted pendulum mode: A simple modeling for a biped walking pattern generation,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation The 3d linear inverted pendulum mode: A simple modeling for a biped walking pattern generation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.561134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:47.670448Z digest=sha256:8c05aa03e695911e9305f64639ad949ede6475979c15702d6b9b1eca6b518bad

Observation 339963da-d9e3-4b9f-8617-f27af015526f · outbound

This paper cites Bipedal walking control based on capture point dynamics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Bipedal walking control based on capture point dynamics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.298431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:47.823455Z digest=sha256:9d146e91f37ef6e4f4add7e64a97bc9f050a9509f9f677ab504795ce86c1fb31

Observation a56aac47-df3a-4a49-aeea-6703e77f7eff · outbound

This paper cites Nonlinear model predictive control for rough-terrain robot hopping,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Nonlinear model predictive control for rough-terrain robot hopping,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.040334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:48.004964Z digest=sha256:11bbb0e4d859cae0af14d1a9fd776dc55d6b9a697392c7555a236aa5ff9f9a71

Observation dbcb9620-946d-467a-9667-d332fcf5a736 · outbound

This paper cites Perceptive locomotion through nonlinear model-predictive control,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Perceptive locomotion through nonlinear model-predictive control,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.116002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.116002Z digest=sha256:6d5ad1ab0d75899829bcbba68ce8b3428e51488dc87a9b2e8e88680f0623cd1b

Observation c1f7cd1c-0a89-480b-b5ec-f3d9243b81f6 · outbound

This paper cites Model predictive control for dynamic footstep adjustment using the divergent component of motion,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Model predictive control for dynamic footstep adjustment using the divergent component of motion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.786744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:48.254270Z digest=sha256:91052e32c391f1995324e72d99664dad2383892b210073be06bf8a8236830ef4

Observation 566d52c9-0baa-45f1-ae10-292bcd563abf · outbound

This paper cites A sequential mpc approach to reactive planning for bipedal robots using safe corridors in highly cluttered environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation A sequential mpc approach to reactive planning for bipedal robots using safe corridors in highly cluttered environments,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.362816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.362816Z digest=sha256:c15c2dec89bcb86d918aec396d430dcc99f9d974cdae2c619aef2f52d69aa210

Observation 4a85c47e-894c-43b0-a189-b32e9e7d9491 · outbound

This paper cites Real-time safe bipedal robot navigation using linear discrete control barrier functions,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Real-time safe bipedal robot navigation using linear discrete control barrier functions,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.514967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:48.481563Z digest=sha256:48607378e1a2a23d3ece7e9d2eb066c730cd212ebe4cf9c48962453cfb5d70fd

Observation 4076a78d-f047-4957-b732-9d468378f1f2 · outbound

This paper cites Apprenticeship learning via inverse rein- forcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Apprenticeship learning via inverse rein- forcement learning,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.595799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.595799Z digest=sha256:76405dcd6c1dfae6aab4794a701c8e4429b497bb20c5f6cfbc5979807771a530

Observation 0a5c7f4f-a9f7-4c4a-bd3d-8e2a1b241ac1 · outbound

This paper cites Apprenticeship learning using linear programming,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Apprenticeship learning using linear programming,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.238377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:48.692116Z digest=sha256:4ddacb0dbed3656b4364c6827bb365604d807cf9ce8a808c51ec8510f153ed52

Observation eaa33a93-53ab-4b61-918f-3256a4af3882 · outbound

This paper cites Generative adversarial imitation learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Generative adversarial imitation learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.843022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.843022Z digest=sha256:10b6e3229fbf198b4955df9502a030262ee1e7d936ddabb6da9e81d0e26e0453

Observation a9d0184d-8bb5-4e04-be9a-4bd07c948efb · outbound

This paper cites Goal-oriented obstacle avoid- ance with deep reinforcement learning in continuous action space,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Goal-oriented obstacle avoid- ance with deep reinforcement learning in continuous action space,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.999533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:48.937422Z digest=sha256:0fcc7b3bb335636c6d225d746056ac658353775d3a28a548a424462b6997581f

Observation 32780fab-6113-4503-bb13-19ae90ccba2e · outbound

This paper cites Where to go next: Learning a subgoal recommendation policy for navigation in dynamic environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Where to go next: Learning a subgoal recommendation policy for navigation in dynamic environments,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:49.072122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:49.072122Z digest=sha256:8baa94a4e0d25a2838997d8d15f8bdfdcdf235240a020286f841570ee1dd46f7

Observation 5cb83d8d-a174-43ae-88c1-8d1898170a7e · outbound

This paper cites Robot navigation in constrained pedestrian environments using reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Robot navigation in constrained pedestrian environments using reinforcement learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.724840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:49.185856Z digest=sha256:08498a4cebd9bf06928f39ed17fe853e536ed6c2011f3d42980b0cb2cab56080

Observation 072de955-c836-4511-9ba6-135d96e9f8a8 · outbound

This paper cites A hierarchical deep reinforcement learning framework with high efficiency and generalization for fast and safe navigation,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation A hierarchical deep reinforcement learning framework with high efficiency and generalization for fast and safe navigation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.486999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:49.305115Z digest=sha256:022c8e3638219f3c7eae762538080784de6b9f2503c768fcf2dc9965f535f2f7

Observation b2b485ca-9249-46b1-8473-6464a8bded44 · outbound

This paper cites Drl-vo: Learning to navigate through crowded dynamic scenes using velocity obstacles,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Drl-vo: Learning to navigate through crowded dynamic scenes using velocity obstacles,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.209570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:49.427620Z digest=sha256:5e6cfd56a34b7a4239583153c58554cdbae267d90ccd625868668d47672e6d3b

Observation 40c9287a-bc88-4cbb-8fe5-282da52f7ce6 · outbound

This paper cites Efficient deep reinforcement learning with imitative expert priors for autonomous driving,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Efficient deep reinforcement learning with imitative expert priors for autonomous driving,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.950891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:49.525766Z digest=sha256:950367148ae0614d1b91f307e3c001bb1225cb445ce1081c6d5c29c5fb7d75e1

Observation 7a35d118-85a1-466f-952a-eb6c61418154 · outbound

This paper cites Pre-training goal-based models for sample-efficient reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Pre-training goal-based models for sample-efficient reinforcement learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.691359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:49.640866Z digest=sha256:a67a0c872514403492cf10ff6b213a09a08c3d0d4ee6b5191461ecad53de6aeb

Observation 80730197-6df9-4edf-a685-df9ee676e76f · outbound

This paper cites Deep q- learning from demonstrations,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Deep q- learning from demonstrations,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.413355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:49.819870Z digest=sha256:9f1e3cd0c5cd4430ad3d4dfe634a69f19ca8ef781592c7f5e138ba9cc382bee6

Observation 9c54604b-f17c-43a5-8cb0-8778ed0b5ecc · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:49.972886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:49.972886Z digest=sha256:b5f7d266378a8f1da3031efb783f4f360455a48a7f53529552c4624d69583137

Observation 5b7a965f-8aab-4297-bc93-4da4edbc2279 · outbound

This paper cites Template model inspired task space learning for robust bipedal locomotion,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Template model inspired task space learning for robust bipedal locomotion,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.141854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:50.057744Z digest=sha256:1fee9551ae49f10b62efe7006285c350c67e521e0132d6b554240fc145da1b00

Observation e4b73840-65c4-4e0a-94c3-20e761fdb8b3 · outbound

This paper cites Conservative q- learning for offline reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Conservative q- learning for offline reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:50.895121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T11:33:50.168382Z digest=sha256:ba0407b34523a05f889089b68d8df35e5b22d79c6bd0e98d3688637463d16648

Observation a809ba34-adc7-45de-8f63-eb05135ab05c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Proximal Policy Optimization Algorithms

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:50.317068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:50.317068Z digest=sha256:fbf368561713a81868c015e7b14a80eabc35cfabaf77cc8d03ca60b2bb60a0a9

Pith citing papers

Observation b3964702-4ab5-47bb-ba95-e3886eecbf7c · inbound

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC cites this paper.

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T22:35:37.040638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:35:37.040638Z digest=sha256:9a7593fcae9d1326c6e35a8a732700be9f5f0ff782293acecf8343942a8ff8c3