Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:10.930242Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.18883.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:10.930242Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0ffdf3b3-c372-4e72-9bca-4fe0151eaa71 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Cipriano, P
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 37690468-5bf9-4288-8a48-9be8925a60c7 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Estevez, J
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c1dfbadf-792d-4715-8fb1-fee5982e6045 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3c434fbe-43c8-49a7-9291-b06c5a8fbd78 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4067b1c3-8aea-4b2d-878f-a8fca9b17162 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Bellman, Dynamic Programming (Princeton University Press, Princeton, NJ) (1957), intro- duces the formalism of Markov decision processes (MDPs) and the principle of optimality
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2bb8cd5e-fbac-41a8-9d65-c1e312c30cd5 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4d81ecb9-85bd-4e89-8303-950e820349c8 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f7e5f1c5-5923-487e-8baa-e09ab719a048 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d694c060-bbb5-4f71-964c-92cf6bb4486f · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 33c6fef4-d456-4592-92b7-21f8aa5f728b · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0f36d1c8-13eb-48aa-8ea0-da1530f32a02 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Hochreiter, J
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f16242e5-0751-4c0b-b493-40561be136db · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d5a7c946-349e-41ad-b066-855291a92ca0 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f7f490aa-2a40-420d-b0af-34528ce028f8 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 90d0627e-a914-4568-a179-8993aa403def · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Arcieri, et al., POMDP Inference and Robust Solution via Deep Reinforcement Learning: An Application to Railway Optimal Maintenance
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3ff6713b-873f-4e3a-abb9-738e86910c1c · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Lemmel, R
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a1ab6c8b-4f4a-4989-8c4c-9930e88ca20d · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f30e640-9f71-4dab-bb45-3cab7c529e1e · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c8306416-5b0e-4e2e-83ff-b8828f7659cf · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 118bf2cf-c6c2-43cb-936c-a8d46f7556d2 · outbound
Success in Humanoid Reinforcement Learning under Partial Observation MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 424d8515-7bf4-4613-b66e-754ae0e53ebd · outbound
Success in Humanoid Reinforcement Learning under Partial Observation Towers, et al., Gymnasium (2023), doi:10.5281/zenodo.8127026,https://zenodo.org/ record/8127025
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.