Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation

As of 7 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2506.02206.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02206 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:50.317068Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T22:35:37.040638Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5e55a5b1-3656-4b35-a091-6a474d898f24 · outbound

This paper cites Introduction of the Foot Placement Estimator: A Dynamic Measure of Balance for Bipedal Robotics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Introduction of the Foot Placement Estimator: A Dynamic Measure of Balance for Bipedal Robotics,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.551834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:45.982721Z digest=sha256:7370fff9cbdac0131170425cf1708469b96e3531927a9691307143fcec444856

Observation e184e266-55c1-446f-a646-fa74e451b99a · outbound

This paper cites Navigation planning for legged robots in challenging terrain,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Navigation planning for legged robots in challenging terrain,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.306213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.067955Z digest=sha256:676f8ab14fcbe87529040a35ec8b8ac94659b503c79dc6dcfce8cb099ede65a6

Observation 787cc79f-2b84-4667-857e-08cc2e5eca42 · outbound

This paper cites Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:46.171453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:46.171453Z digest=sha256:8500e30c7c822b2d64179c48f67609a273766cd1a1beecb5776c2b4783e2c2aa

Observation 43d458b8-00df-4988-b1d6-6210be1fec8c · outbound

This paper cites Exact cell decomposition of arrangements used for path planning in robotics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Exact cell decomposition of arrangements used for path planning in robotics,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:57.032791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.263030Z digest=sha256:4b922b7027f92cae9a35fbfdb1d521231df957cc76b97d71645a45753acabbd2

Observation 91759f8a-8b67-4a89-a046-423966bdbc46 · outbound

This paper cites An overview of autonomous mobile robot path planning algorithms,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation An overview of autonomous mobile robot path planning algorithms,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.783635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.395990Z digest=sha256:f439c4e4d8314c87b5d19cdfcf56eb5e54c49e53f3e58d42d6cb7b84b2554cb4

Observation d231c70e-f21f-46b6-ad9c-740947232dfc · outbound

This paper cites Path planning and trajectory planning algorithms: A general overview,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Path planning and trajectory planning algorithms: A general overview,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.536460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.547721Z digest=sha256:6dd73e548d23e3d724c99529aee38b33e7d5d8af79e8d89f69919d42178dfb7b

Observation 506751ae-76fb-4abb-bdf9-83af3c55f446 · outbound

This paper cites Confidence random tree- based algorithm for mobile robot path planning considering the path length and safety,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Confidence random tree- based algorithm for mobile robot path planning considering the path length and safety,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:56.205520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.715004Z digest=sha256:1a35ca2cd3df5842f79291d1ca27dace4318a44e228f3544d071789a2bd90477

Observation cccc64f7-cb5b-4dcd-a7eb-6f09f0d3a155 · outbound

This paper cites Optimization-based motion planning for legged robots,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Optimization-based motion planning for legged robots,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.930491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.882468Z digest=sha256:c6bddc832e4df3107717a21f3642e164f5eb8583f4f83f05664b0771df7ffaff

Observation 1fe19592-9ca4-4e45-b327-3d3130da7bb5 · outbound

This paper cites Integrated task and motion planning for safe legged navigation in partially observable environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Integrated task and motion planning for safe legged navigation in partially observable environments,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.644239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:46.984997Z digest=sha256:07c03bc2c0b9a6102f3dfe7f0078820ea610d93dd9cd69cfa88ba1f743242a7f

Observation 185c0dd2-a53c-4d03-9b02-a89681136225 · outbound

This paper cites Unified Path and Gait Planning for Safe Bipedal Robot Navigation.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Unified Path and Gait Planning for Safe Bipedal Robot Navigation

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:33:50.614740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:47.137767Z digest=sha256:528dda46ae977474be638e2275de2a252237d7a49f2004f944ec8c592eafe8f1

Observation ee89f1e1-1319-4112-901e-3f9ca352ce98 · outbound

This paper cites Fast direct multiple shooting algorithms for optimal robot control,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Fast direct multiple shooting algorithms for optimal robot control,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.421021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:47.253223Z digest=sha256:44344691397bfe2e58057036f67c8e12404644da6113071c75d84ce65ae21576

Observation 4a83a47b-1afa-4501-b94b-438e51b99606 · outbound

This paper cites Using optimization to create self-stable human-like running,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Using optimization to create self-stable human-like running,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:55.146851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:47.401679Z digest=sha256:422ef112d4865bd3a78c2bdd2578b191ff9043b45da6c6363ef5dd327e976c1a

Observation 4073e0d1-07d9-4e0d-81eb-8374028d13bf · outbound

This paper cites Whole-body motion planning with centroidal dynamics and full kinematics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Whole-body motion planning with centroidal dynamics and full kinematics,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.878671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:47.499247Z digest=sha256:9222ed31a63ac0262d84143a77313015b2504e00bef7b0e670418eed1b12202c

Observation ad922e4f-c540-40af-b183-2aefc9c91cae · outbound

This paper cites The 3d linear inverted pendulum mode: A simple modeling for a biped walking pattern generation,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation The 3d linear inverted pendulum mode: A simple modeling for a biped walking pattern generation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.561134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:47.670448Z digest=sha256:8c84f7bae13500bed9668f6ce2846f38257019f4ba7fdbef1067f7e72624994a

Observation 339963da-d9e3-4b9f-8617-f27af015526f · outbound

This paper cites Bipedal walking control based on capture point dynamics,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Bipedal walking control based on capture point dynamics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.298431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:47.823455Z digest=sha256:891e6a60f83886bc0e8ef1d20572279a0d55a6cafa6cef2b8f9da5802484402b

Observation a56aac47-df3a-4a49-aeea-6703e77f7eff · outbound

This paper cites Nonlinear model predictive control for rough-terrain robot hopping,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Nonlinear model predictive control for rough-terrain robot hopping,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:54.040334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:48.004964Z digest=sha256:a435e827fac2b03064ccdac4b95f01d50def9316c7c29bfca81bcb49c5f34e5f

Observation dbcb9620-946d-467a-9667-d332fcf5a736 · outbound

This paper cites Perceptive locomotion through nonlinear model-predictive control,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Perceptive locomotion through nonlinear model-predictive control,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.116002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.116002Z digest=sha256:bc1f4b6794e8f13686376f88df85e74a7a32571b4504b3585bb67531eeda7d05

Observation c1f7cd1c-0a89-480b-b5ec-f3d9243b81f6 · outbound

This paper cites Model predictive control for dynamic footstep adjustment using the divergent component of motion,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Model predictive control for dynamic footstep adjustment using the divergent component of motion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.786744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:48.254270Z digest=sha256:b8b3d7911804b9fc6fb09579cbb228d00eae1440de8004d618ea9f3314cdb08d

Observation 566d52c9-0baa-45f1-ae10-292bcd563abf · outbound

This paper cites A sequential mpc approach to reactive planning for bipedal robots using safe corridors in highly cluttered environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation A sequential mpc approach to reactive planning for bipedal robots using safe corridors in highly cluttered environments,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.362816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.362816Z digest=sha256:a3c4de0c3e6274ac0ce9f355a374f850d6068c9c73bb6e6d5084adea7c3ca397

Observation 4a85c47e-894c-43b0-a189-b32e9e7d9491 · outbound

This paper cites Real-time safe bipedal robot navigation using linear discrete control barrier functions,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Real-time safe bipedal robot navigation using linear discrete control barrier functions,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.514967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:48.481563Z digest=sha256:b72000e48d0e8b88d79b597d88abba0b4b42f0f8e3a6f7a52d616f5e03d04a99

Observation 4076a78d-f047-4957-b732-9d468378f1f2 · outbound

This paper cites Apprenticeship learning via inverse rein- forcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Apprenticeship learning via inverse rein- forcement learning,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.595799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.595799Z digest=sha256:d7125bd7885040753350fa82ff77db2be410a1621ce1e0efa1ac0fb1a3fce2cb

Observation 0a5c7f4f-a9f7-4c4a-bd3d-8e2a1b241ac1 · outbound

This paper cites Apprenticeship learning using linear programming,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Apprenticeship learning using linear programming,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:53.238377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:48.692116Z digest=sha256:8fc4be212980df15121d083e8dd4e76906ea83ba98df20a616fa82d8f1c3dab9

Observation eaa33a93-53ab-4b61-918f-3256a4af3882 · outbound

This paper cites Generative adversarial imitation learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Generative adversarial imitation learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:48.843022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:48.843022Z digest=sha256:c547fffdc886ea5bc02a97d7a1d577c8fef4240cefdb38bbdfd00864fac68b1a

Observation a9d0184d-8bb5-4e04-be9a-4bd07c948efb · outbound

This paper cites Goal-oriented obstacle avoid- ance with deep reinforcement learning in continuous action space,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Goal-oriented obstacle avoid- ance with deep reinforcement learning in continuous action space,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.999533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:48.937422Z digest=sha256:41409c05748cec0a355d46064eade69a51af235c1735585695232f0eb3b96e5c

Observation 32780fab-6113-4503-bb13-19ae90ccba2e · outbound

This paper cites Where to go next: Learning a subgoal recommendation policy for navigation in dynamic environments,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Where to go next: Learning a subgoal recommendation policy for navigation in dynamic environments,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:49.072122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:49.072122Z digest=sha256:38da9e2b2de762abaceaa7d84892b7dc6c5e1f7600286a0ff08f5c63548a20b4

Observation 5cb83d8d-a174-43ae-88c1-8d1898170a7e · outbound

This paper cites Robot navigation in constrained pedestrian environments using reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Robot navigation in constrained pedestrian environments using reinforcement learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.724840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:49.185856Z digest=sha256:c85e2ff4695d97c219ae9d48b930e46f05e081c3fea448d01097a63910046f3b

Observation 072de955-c836-4511-9ba6-135d96e9f8a8 · outbound

This paper cites A hierarchical deep reinforcement learning framework with high efficiency and generalization for fast and safe navigation,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation A hierarchical deep reinforcement learning framework with high efficiency and generalization for fast and safe navigation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.486999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:49.305115Z digest=sha256:f89a877267def3d3fd31fe5353f8d8fe045279bc9f87e6660e8eeec6c3b9a50c

Observation b2b485ca-9249-46b1-8473-6464a8bded44 · outbound

This paper cites Drl-vo: Learning to navigate through crowded dynamic scenes using velocity obstacles,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Drl-vo: Learning to navigate through crowded dynamic scenes using velocity obstacles,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:52.209570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:49.427620Z digest=sha256:a1bcd424a3f0c34691d5d8ebe61e4f33cdcb808d89f9774780802cf8573f52cc

Observation 40c9287a-bc88-4cbb-8fe5-282da52f7ce6 · outbound

This paper cites Efficient deep reinforcement learning with imitative expert priors for autonomous driving,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Efficient deep reinforcement learning with imitative expert priors for autonomous driving,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.950891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:49.525766Z digest=sha256:f01d2d61a3c5fd7916d2c3715a069e365c4c67aceb5107d674c8c187d32b23f5

Observation 7a35d118-85a1-466f-952a-eb6c61418154 · outbound

This paper cites Pre-training goal-based models for sample-efficient reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Pre-training goal-based models for sample-efficient reinforcement learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.691359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:49.640866Z digest=sha256:d125d3458fe601b535a89a1c45e6060ad571675a58adcde8228403bc0216bdaa

Observation 80730197-6df9-4edf-a685-df9ee676e76f · outbound

This paper cites Deep q- learning from demonstrations,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Deep q- learning from demonstrations,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.413355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:49.819870Z digest=sha256:25536847a7f0c4cb1bd6e9bc20451656c9928269912a47ae5d23a5df8df691ff

Observation 9c54604b-f17c-43a5-8cb0-8778ed0b5ecc · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:49.972886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:49.972886Z digest=sha256:77d0a6255d9ea47ccc5ab80ab59afeea01e880c281b96f3470a33c2ea3f57f0b

Observation 5b7a965f-8aab-4297-bc93-4da4edbc2279 · outbound

This paper cites Template model inspired task space learning for robust bipedal locomotion,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Template model inspired task space learning for robust bipedal locomotion,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:51.141854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:50.057744Z digest=sha256:2982f41fa5f30101e047dc71d45b31cf959ea2f7463b26df88ff899254b6294e

Observation e4b73840-65c4-4e0a-94c3-20e761fdb8b3 · outbound

This paper cites Conservative q- learning for offline reinforcement learning,.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Conservative q- learning for offline reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:50.895121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:50.168382Z digest=sha256:cfd0e97e97f8b3b044d871a75b7ebf7f1aecae1c7aa19b850dbacdc8bad7c345

Observation a809ba34-adc7-45de-8f63-eb05135ab05c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation Proximal Policy Optimization Algorithms

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:50.317068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:50.317068Z digest=sha256:fbdf96ce8f048b3744f81e897ad7bd0f283b9b0bcc1489402632c22d6d6b5924

Pith citing papers

Observation b3964702-4ab5-47bb-ba95-e3886eecbf7c · inbound

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC cites this paper.

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T22:35:37.040638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:35:37.040638Z digest=sha256:7d90a020313b8684c13979b55368a55e1156a8bf5f79a3e54593fadc01524848