Pith. sign in

Paper Citation Record · LEDGER

Skill-based Safe Reinforcement Learning with Risk Planning

As of 23 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2505.01619.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01619 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:20:07.679615Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b940988a-71e5-4797-ac48-ad5511145f17 · outbound

This paper cites write newline.

Skill-based Safe Reinforcement Learning with Risk Planning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:20:07.521180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:20:07.521180Z digest=sha256:27e822e280a9df949ec5e7dbfadfda0b822a1c6a55655ee1964b43b39fdc9f2d

Observation 2e652224-cab4-4ffb-91fd-ef858eef02fd · outbound

This paper cites Constrained policy optimization.

Skill-based Safe Reinforcement Learning with Risk Planning Constrained policy optimization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.202587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.529200Z digest=sha256:b0cfae70ec695fa38b43f68f190c0da390187c6f39277bcb79cbc7c8fdb95fc7

Observation 4e7de232-1e9d-463e-ba9e-2e7a7fd0f35f · outbound

This paper cites Constrained Markov decision processes: stochastic modeling.

Skill-based Safe Reinforcement Learning with Risk Planning Constrained Markov decision processes: stochastic modeling

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.185937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.534910Z digest=sha256:4ff25c4bdf1746ddc4605b4570feb9249aa331282ee09a23dd05b0e64aaefe5a

Observation 6ff90d14-c569-4cac-84e4-14d8371ca33a · outbound

This paper cites and Yarats, D.

Skill-based Safe Reinforcement Learning with Risk Planning and Yarats, D

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.170203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.541455Z digest=sha256:23a29431d94fc5624fd5646140a87546b92e0cd2d80812279a108946adc4307f

Observation e94de879-c529-44ad-9ba6-4b44a193f9b4 · outbound

This paper cites D., Chernova, S., Veloso, M., and Browning, B.

Skill-based Safe Reinforcement Learning with Risk Planning D., Chernova, S., Veloso, M., and Browning, B

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T04:20:07.546851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:20:07.546851Z digest=sha256:b82c5f362903f003a180f3f19085b4da0597f6f1096ff1e3a93cf35775655572

Observation 5a8e5b6e-6c4f-486d-ac96-33b8ac014e08 · outbound

This paper cites I., Kroese, D.

Skill-based Safe Reinforcement Learning with Risk Planning I., Kroese, D

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.144917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.552344Z digest=sha256:b7a0236a72e078abeba4a92fb5a0e22d42558a3d23da81422f67afb0820395ca

Observation ed966bc6-15dc-405f-9af9-de4b18f0e7a2 · outbound

This paper cites W., Yuan, Z., Zhou, S., Panerati, J., and Schoellig, A.

Skill-based Safe Reinforcement Learning with Risk Planning W., Yuan, Z., Zhou, S., Panerati, J., and Schoellig, A

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.129889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.558028Z digest=sha256:c87101db78600b041ac2f77fad1c2a2c51f78cf85e14c45575e88b9718f32021

Observation 8569d491-7140-44d3-a7c2-0247f0858564 · outbound

This paper cites B., Chernova, S., Taylor, M.

Skill-based Safe Reinforcement Learning with Risk Planning B., Chernova, S., Taylor, M

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.114330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.563471Z digest=sha256:be686968716c0983fbe4d88557b89b6b9f283459d2402499d3decb2578d62897

Observation b2c95931-e823-491d-94c5-c44bbdf5ce6a · outbound

This paper cites Class-prior estimation for learning from positive and unlabeled data.

Skill-based Safe Reinforcement Learning with Risk Planning Class-prior estimation for learning from positive and unlabeled data

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.097363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.568321Z digest=sha256:719df89302f6d97e55d287d6caac0aa7e9f8ded75740237dbe71aecf7a0ede02

Observation 1d23f29f-e721-44cb-929f-6de3b87b2ca8 · outbound

This paper cites Convex formulation for learning from positive and unlabeled data.

Skill-based Safe Reinforcement Learning with Risk Planning Convex formulation for learning from positive and unlabeled data

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.081625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.573370Z digest=sha256:153e72d514de97bd8576f1f9362812f2aedd4b6d2a786637a5b26a33dcbaaa86

Observation d6dc3696-36f3-4d8e-ac3b-e00a21dd537d · outbound

This paper cites C., Niu, G., and Sugiyama, M.

Skill-based Safe Reinforcement Learning with Risk Planning C., Niu, G., and Sugiyama, M

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.064154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.578248Z digest=sha256:0eb99299506669dcce79ebfefb9a1b502ecafe167435ced6ce1c6b2ffb1969db

Observation 6e6523a1-d1cf-4683-befe-ad65d4281149 · outbound

This paper cites and Fern \'a ndez, F.

Skill-based Safe Reinforcement Learning with Risk Planning and Fern \'a ndez, F

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.046820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.583120Z digest=sha256:7ddbc2534bd6a8c8f4aa5da9e9c5302bb8e41d606511cfdcd125bef810dd3d86

Observation ea193771-c9ce-4993-b89b-3fee9cce4a1c · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Skill-based Safe Reinforcement Learning with Risk Planning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.030374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.588244Z digest=sha256:953014ad989f9e84a984b370c7c945a9a474fe9a0b1247b65d95017549ac4d89

Observation b8685725-eac6-40df-af0c-20e6c70eb458 · outbound

This paper cites M., and Udluft, S.

Skill-based Safe Reinforcement Learning with Risk Planning M., and Udluft, S

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:08.010991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.593656Z digest=sha256:57a9ac25f4859447720b07b45c3400bea0578601ce69019f6a335dbaac7f4b06

Observation b3dbcf0d-0198-4476-87bd-72c5b72b9ba8 · outbound

This paper cites and Ermon, S.

Skill-based Safe Reinforcement Learning with Risk Planning and Ermon, S

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.994125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.599289Z digest=sha256:bb27a0bcb92757720e373c1c62c577c59ca007655d5323d89151a51be6f2f178

Observation 633bb447-93b2-44fb-8bc7-e68c9aee534e · outbound

This paper cites Estimating the class prior and posterior from noisy positives and unlabeled data.

Skill-based Safe Reinforcement Learning with Risk Planning Estimating the class prior and posterior from noisy positives and unlabeled data

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.978682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.604902Z digest=sha256:fe75507e5c920816bb54328c23e0a12d40b5849e2f498f5333aeee5788e96961

Observation 2ac314da-3507-4e5a-954f-4c0355d3aea1 · outbound

This paper cites C., and Sugiyama, M.

Skill-based Safe Reinforcement Learning with Risk Planning C., and Sugiyama, M

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.960424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.611375Z digest=sha256:21c16dad2335a130e73c3d6dd85462ceed68460281fc8db69f665e66a158b490

Observation 6501394d-9c11-4fdd-adde-9a0e8c8ba590 · outbound

This paper cites and Whiteson, S.

Skill-based Safe Reinforcement Learning with Risk Planning and Whiteson, S

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.943181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.616021Z digest=sha256:8b9af1b009a6acd44cb0ac4efdc4257f7e94fe9e45ea10dc3464dd2c517001f6

Observation 1a8ba416-5a3b-41c9-8587-6202ec5da25c · outbound

This paper cites Datasets and Benchmarks for Offline Safe Reinforcement Learning.

Skill-based Safe Reinforcement Learning with Risk Planning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:20:07.620593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:20:07.620593Z digest=sha256:01c7fa305aae019069b855a884eb0cda9682cf16449a3ceea4354f58cc21d3e5

Observation 41c9a7bc-00ea-4a1d-9189-dd3567acd8bd · outbound

This paper cites Accelerating reinforcement learning with learned skill priors.

Skill-based Safe Reinforcement Learning with Risk Planning Accelerating reinforcement learning with learned skill priors

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.926362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.625607Z digest=sha256:a9183c8c79ed4f33700915efeee33077c060c3113ef910a7e3b8af6a09bcad27

Observation 1537fd58-5cb5-45d2-8bc2-87cdf2612475 · outbound

This paper cites an unresolved cited work.

Skill-based Safe Reinforcement Learning with Risk Planning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:20:07.909372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.630272Z digest=sha256:003f03e8740730b6dd3d36451a40db48e8c0c97b01e2695df4a7c8466d5f774e

Observation 5eeecff8-20f1-4ea4-b35d-3f569ba98ded · outbound

This paper cites Z., Chow, Y., Dai, B., and Wichers, N.

Skill-based Safe Reinforcement Learning with Risk Planning Z., Chow, Y., Dai, B., and Wichers, N

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.892208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.634715Z digest=sha256:95be39d1ac08ee489cde39b9d5cf7efc4d524d00d311a018efae51892c40c4e4

Observation 766aca5d-9e73-4b97-9d9a-f0df56895cbe · outbound

This paper cites an unresolved cited work.

Skill-based Safe Reinforcement Learning with Risk Planning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T04:20:07.638867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:20:07.638867Z digest=sha256:569b2d03065510de4bfbfe80af47091969bdaa4a99ec51b877dda8d80660cc8e

Observation cdf8284f-1f32-49d0-8c10-159336ca78fa · outbound

This paper cites J., and Mannor, S.

Skill-based Safe Reinforcement Learning with Risk Planning J., and Mannor, S

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.864938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.643096Z digest=sha256:e30228a6db1ff9995e91aceb6859b994e370b7804854a549ed6e78604b77b03e

Observation 857a098a-4497-4070-a807-7829f023c6a3 · outbound

This paper cites E., Levine, S., Borrelli, F., and Goldberg, K.

Skill-based Safe Reinforcement Learning with Risk Planning E., Levine, S., Borrelli, F., and Goldberg, K

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.849796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.647242Z digest=sha256:ab1d922a32b5159e131101ef4f25b210e760d37a31d5d664ffc86c540513c455

Observation 9a31f6ff-12c5-44e9-b278-3d3902d08782 · outbound

This paper cites E., Ibarz, J., Finn, C., and Goldberg, K.

Skill-based Safe Reinforcement Learning with Risk Planning E., Ibarz, J., Finn, C., and Goldberg, K

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.835399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.651486Z digest=sha256:c48c8e32a65679f548c0aec3b967d121e159c7eafa838399ab4ba4c77a858fe4

Observation 365bd791-b859-4526-b9d5-99dfe3bb4365 · outbound

This paper cites Safe reinforcement learning by imagining the near future.

Skill-based Safe Reinforcement Learning with Risk Planning Safe reinforcement learning by imagining the near future

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.819937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.656103Z digest=sha256:1d2a9c451d871f3e6ef3075849c1361435e0e9ce8a642e01e8b54d16996d6364

Observation 7e7f6d9a-385b-43f1-b9d3-a6c71e49122c · outbound

This paper cites and Schwartz, A.

Skill-based Safe Reinforcement Learning with Risk Planning and Schwartz, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.802886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.660572Z digest=sha256:14c6291537f6354053b24ed95153123d7844ee0eba083f1447eef7718c941c0d

Observation e20d8d03-20a2-4e4d-971b-0b0210c2c1bc · outbound

This paper cites Mujoco: A physics engine for model-based control.

Skill-based Safe Reinforcement Learning with Risk Planning Mujoco: A physics engine for model-based control

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.786159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.665973Z digest=sha256:066974a27c023a40272f97182fd53fec2c268a722f2a17727ffe29f116966c5a

Observation a214cb6b-7bd7-4c46-a5ec-a741b7f63fd2 · outbound

This paper cites E., Xu, S., and Peng, H.

Skill-based Safe Reinforcement Learning with Risk Planning E., Xu, S., and Peng, H

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.770215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.670620Z digest=sha256:3959893b1636857f4ed3c10a6b2e56940bd8fcbfcc64105d7364eb9ee55a6db6

Observation dc18d33f-69e3-4cf7-9d3e-e8e148c35098 · outbound

This paper cites and Denil, M.

Skill-based Safe Reinforcement Learning with Risk Planning and Denil, M

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.753073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.675090Z digest=sha256:ed5f14340c08de39a0a796ff1aa7ec4f312a1128f68ae967a913185073515b75

Observation 0c785921-f0ee-4e6f-8554-c851a40d1f1d · outbound

This paper cites Constraints penalized q-learning for safe offline reinforcement learning.

Skill-based Safe Reinforcement Learning with Risk Planning Constraints penalized q-learning for safe offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:20:07.735682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-16T04:20:07.679615Z digest=sha256:01af9bdce556cb6b6179b5df679ae81f252cf0d9163375ffd6f7c94bc807fe29

Pith citing papers

No inbound Pith citation observations are available.