Pith. sign in

Paper Citation Record · LEDGER

Success in Humanoid Reinforcement Learning under Partial Observation

As of 21 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.18883.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18883 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:10.930242Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ffdf3b3-c372-4e72-9bca-4fe0151eaa71 · outbound

This paper cites Cipriano, P.

Success in Humanoid Reinforcement Learning under Partial Observation Cipriano, P

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.421259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.819046Z digest=sha256:ffc23c1f4641acb027e177c2de03b19a3523e647d4af1f27edeafea3dc12277d

Observation 37690468-5bf9-4288-8a48-9be8925a60c7 · outbound

This paper cites Estevez, J.

Success in Humanoid Reinforcement Learning under Partial Observation Estevez, J

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.401806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.826042Z digest=sha256:5fa7d7dc73ef0455acf3ed4869b9768b363a74ec582e5867ada3c9e28ae84a0b

Observation c1dfbadf-792d-4715-8fb1-fee5982e6045 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.382550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.831258Z digest=sha256:5ced37a41b269df28e35501262d697674430ae5a006092bc957a0b4b06bf65c4

Observation 3c434fbe-43c8-49a7-9291-b06c5a8fbd78 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.363776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.836379Z digest=sha256:85eb6f4eb536e2e57933e14e7e7365eaea825b92bceb132942d22b89091e66dd

Observation 4067b1c3-8aea-4b2d-878f-a8fca9b17162 · outbound

This paper cites Bellman, Dynamic Programming (Princeton University Press, Princeton, NJ) (1957), intro- duces the formalism of Markov decision processes (MDPs) and the principle of optimality.

Success in Humanoid Reinforcement Learning under Partial Observation Bellman, Dynamic Programming (Princeton University Press, Princeton, NJ) (1957), intro- duces the formalism of Markov decision processes (MDPs) and the principle of optimality

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.346834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.842521Z digest=sha256:37321cb4bcd7508655d5d1635b32bf951c0e5b9f2f8083afe70feaa70e9f4160

Observation 2bb8cd5e-fbac-41a8-9d65-c1e312c30cd5 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.327533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.847978Z digest=sha256:8a71647dc804b7e219e2c347f5191721541cec50b828d27fa1bd5a29a2f7632b

Observation 4d81ecb9-85bd-4e89-8303-950e820349c8 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.309457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.853493Z digest=sha256:6e79c1dc04b8b5a39e51ae5294a87e2e475d0faae1cd553d809991b4fad50d87

Observation f7e5f1c5-5923-487e-8baa-e09ab719a048 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.292533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.859193Z digest=sha256:9cc69c5480054b8f9d281c67a93fa8014049261c49709e8a4771aa29e42940e6

Observation d694c060-bbb5-4f71-964c-92cf6bb4486f · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.275683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.864080Z digest=sha256:99d64f412b25db876418fe42802771ec248fa9fcf5265cf94bddeced26ba5317

Observation 33c6fef4-d456-4592-92b7-21f8aa5f728b · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.254858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.870049Z digest=sha256:7a9ee35200017166aedf68e12760db6d72ea98977319354fde38feb4bc5bfaf2

Observation 0f36d1c8-13eb-48aa-8ea0-da1530f32a02 · outbound

This paper cites Hochreiter, J.

Success in Humanoid Reinforcement Learning under Partial Observation Hochreiter, J

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.237605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.875545Z digest=sha256:426b07693e45ac2dd07b683bab152b3af225645bd22e793fac9c41ebc333befa

Observation f16242e5-0751-4c0b-b493-40561be136db · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.217435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.880471Z digest=sha256:064ee9c2569e562c10bda4b5d026e478376cabab7d942b11522cc9c1721108bc

Observation d5a7c946-349e-41ad-b066-855291a92ca0 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.194171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.886132Z digest=sha256:da5e7a104b31ce97cad6ab4423c0bf549a62d23313f0c2514e9f1b6f980e7d79

Observation f7f490aa-2a40-420d-b0af-34528ce028f8 · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.172333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.891644Z digest=sha256:f0a13269bac4d172142876bcfbb58ebd7e148d9aea3998bcffefc89153f20698

Observation 90d0627e-a914-4568-a179-8993aa403def · outbound

This paper cites Arcieri, et al., POMDP Inference and Robust Solution via Deep Reinforcement Learning: An Application to Railway Optimal Maintenance.

Success in Humanoid Reinforcement Learning under Partial Observation Arcieri, et al., POMDP Inference and Robust Solution via Deep Reinforcement Learning: An Application to Railway Optimal Maintenance

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.153183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.896574Z digest=sha256:913a3efe81d5004dd0c17397bab77579c8fe8357f1a9ba00c9c26f839302ad13

Observation 3ff6713b-873f-4e3a-abb9-738e86910c1c · outbound

This paper cites Lemmel, R.

Success in Humanoid Reinforcement Learning under Partial Observation Lemmel, R

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:11:11.133464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.901994Z digest=sha256:74ddf26cfa536ef0a07f216eb0c086471e091f8e0f80852bf4b9ad631ca2de52

Observation a1ab6c8b-4f4a-4989-8c4c-9930e88ca20d · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Success in Humanoid Reinforcement Learning under Partial Observation Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.906816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.906816Z digest=sha256:f8c49e3abaa33c92fc95bd95f10d09ca78bf608f600468781aec790468606bef

Observation 3f30e640-9f71-4dab-bb45-3cab7c529e1e · outbound

This paper cites an unresolved cited work.

Success in Humanoid Reinforcement Learning under Partial Observation Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:11:11.113614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:11:10.913179Z digest=sha256:84458481260dba845309e9714c27737d1e8b1b7392ab633a4edbfbd7c4060030

Observation c8306416-5b0e-4e2e-83ff-b8828f7659cf · outbound

This paper cites Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces.

Success in Humanoid Reinforcement Learning under Partial Observation Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.919344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.919344Z digest=sha256:51e759a58059d60d955a81bfbf8dfd5379733c03a246f4cf8405d25fdf9ff2d4

Observation 118bf2cf-c6c2-43cb-936c-a8d46f7556d2 · outbound

This paper cites MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning.

Success in Humanoid Reinforcement Learning under Partial Observation MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.924642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.924642Z digest=sha256:6c92e1cca1ea2b81bdcf69be4b89b4b28e2606cab6f5e570458cb74504744693

Observation 424d8515-7bf4-4613-b66e-754ae0e53ebd · outbound

This paper cites Towers, et al., Gymnasium (2023), doi:10.5281/zenodo.8127026,https://zenodo.org/ record/8127025.

Success in Humanoid Reinforcement Learning under Partial Observation Towers, et al., Gymnasium (2023), doi:10.5281/zenodo.8127026,https://zenodo.org/ record/8127025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:11:10.930242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:11:10.930242Z digest=sha256:9dfadb53cdff82b9a98603850eb09178365ebf51637b32e1edf88c0005750b78

Pith citing papers

No inbound Pith citation observations are available.