Pith. sign in

Paper Citation Record · LEDGER

Practical Risk Measures in Reinforcement Learning

As of 16 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:1908.08379.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.08379 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T11:47:00.466159Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c0622a28-0221-4059-bcb0-06077b7047cb · outbound

This paper cites Constrained Policy Optimization.

Practical Risk Measures in Reinforcement Learning Constrained Policy Optimization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.199339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.199339Z digest=sha256:09f35b7a8b46d611d9877592acd80838925776195ef05d2351ee4bb2f41c0b15

Observation f7a629dd-683e-47b0-ab95-ab1bd2d34ead · outbound

This paper cites an unresolved cited work.

Practical Risk Measures in Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:47:01.228320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.247482Z digest=sha256:dbdcddf5337931a85d41726bcbc8051c01a212fc82850a716b6fe5cd450065e1

Observation 6885c0a8-87ad-4c34-99eb-9d5749067e16 · outbound

This paper cites A comprehensive survey on safe reinforce- ment learning.

Practical Risk Measures in Reinforcement Learning A comprehensive survey on safe reinforce- ment learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.180579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.262981Z digest=sha256:31a908848f96b2de8e3277831c2f5096ece954bc7fbf249b44fa9aabc0306a35

Observation 5b494833-b55a-46a9-b9e4-fe4c6a965886 · outbound

This paper cites Risk-sensitive reinforcement learning applied to control under constraints.

Practical Risk Measures in Reinforcement Learning Risk-sensitive reinforcement learning applied to control under constraints

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.164611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.269387Z digest=sha256:aaf170cf2415e3eb27e5723ec039e449de7809604f79b8b83b95baa7ca5867dd

Observation 8dd704f5-210b-4d8e-a26e-adc390be3ede · outbound

This paper cites Likelihood ratio gradient es- timation for stochastic systems.

Practical Risk Measures in Reinforcement Learning Likelihood ratio gradient es- timation for stochastic systems

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.148182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.275005Z digest=sha256:9ef38c3b291ad161d4ada97ce73cd15c39886941d0de808a3c0abc1d2fa137e3

Observation 87f247a4-1757-4668-a409-092835200307 · outbound

This paper cites Risk-sensitive markov decision processes.

Practical Risk Measures in Reinforcement Learning Risk-sensitive markov decision processes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.114364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.285914Z digest=sha256:989e578cff7d50fbf67a3448ad80d888d3179144f3b3a813c2ca3522e9090e7c

Observation f1d830a1-cbc2-4f20-bc2a-93589db2b562 · outbound

This paper cites Func- tional value iteration for decision-theoretic planning with general utility functions.

Practical Risk Measures in Reinforcement Learning Func- tional value iteration for decision-theoretic planning with general utility functions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.000834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.317425Z digest=sha256:fb1cf752aa49a54c92a3d14995d46018081865040b53600cd3d7ad15e0581305

Observation 92173d96-c6d8-4931-8b53-7c840677d645 · outbound

This paper cites Mean-Variance Optimization in Markov Decision Processes.

Practical Risk Measures in Reinforcement Learning Mean-Variance Optimization in Markov Decision Processes

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.333595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.333595Z digest=sha256:50ee19c4d144a38a1124fee64036445e32b2b156e5e1a2df8f5dfc235e63bd33

Observation 864485fc-1928-468d-a3e3-b0f4becf540b · outbound

This paper cites Safe Exploration in Markov Decision Processes.

Practical Risk Measures in Reinforcement Learning Safe Exploration in Markov Decision Processes

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.351643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.351643Z digest=sha256:5f79fc0c3c0b31cf8bc9c023e84a6435f06ac3895967bcfbd0c7b5808b7d1389

Observation 366feeb0-c411-4045-b945-5c9f6fee9dd1 · outbound

This paper cites Parametric Return Density Estimation for Reinforcement Learning.

Practical Risk Measures in Reinforcement Learning Parametric Return Density Estimation for Reinforcement Learning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-14T11:47:00.520175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.358479Z digest=sha256:c3f312f29db72c4e0829a26d84d079f49c464f4d71b190a5dd028c3d20a7ecac

Observation 49cffd99-260f-496f-a1bc-ca950d320464 · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping.

Practical Risk Measures in Reinforcement Learning Policy invariance under reward transformations: Theory and application to reward shaping

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.932894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.365822Z digest=sha256:80e671c8f8779c02844a83afbf33e3993799d3902d38af400188beaf46abdf89

Observation 422180c0-9576-4cf1-9040-5734e0319dac · outbound

This paper cites Markov chains.

Practical Risk Measures in Reinforcement Learning Markov chains

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.914052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.371883Z digest=sha256:8583ecd5dc67b0fefedc70b73fa4634e8f9166ef8c7c7307c44e7c442715733e

Observation 80b9dabf-59c1-41ee-ac76-8122aa2ad39d · outbound

This paper cites Evaluation: from preci- sion, recall and f-measure to roc, informedness, markedness and correlation.

Practical Risk Measures in Reinforcement Learning Evaluation: from preci- sion, recall and f-measure to roc, informedness, markedness and correlation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.883270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.397172Z digest=sha256:422bdbba09aacd5c777c194d1340007b481c234a618df8d3730087c0b9a053a7

Observation 6d394dd3-f44e-4bad-a12a-7656b895fecb · outbound

This paper cites Actor-critic algorithms for risk- sensitive mdps.

Practical Risk Measures in Reinforcement Learning Actor-critic algorithms for risk- sensitive mdps

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.865002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.403592Z digest=sha256:a3c61c284fbb45717fe2675d60f388b73728b84c0c92a158596d1a506e759055

Observation aea5fe12-e1ec-4d56-9f28-027a2e8a4bbd · outbound

This paper cites Markov decision pro- cesses.

Practical Risk Measures in Reinforcement Learning Markov decision pro- cesses

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.843434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.411002Z digest=sha256:a566a514eb0cb7e422214431c273b76f3da7088de2d9be4112beb564742f0077

Observation fbd94178-7446-4867-a208-bb48b5074749 · outbound

This paper cites Td algorithm for the variance of return and mean-variance reinforcement learning.

Practical Risk Measures in Reinforcement Learning Td algorithm for the variance of return and mean-variance reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.802809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.423786Z digest=sha256:020501c37c254c8b40841e01032cd06819225208a9b3d46a1761904e8b2c5f8a

Observation 115460b0-48e8-41f0-9303-9c9092644f2b · outbound

This paper cites Lectures on stochastic program- ming: modeling and theory.

Practical Risk Measures in Reinforcement Learning Lectures on stochastic program- ming: modeling and theory

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.786428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.430564Z digest=sha256:0e792a0b26c087ed98c9e683fce11df780494ea047fe20ba4a44c8f38ee188af

Observation 60c9b40d-6e7d-4333-af3c-1fa87fb43545 · outbound

This paper cites Temporal difference methods for the variance of the reward to go.

Practical Risk Measures in Reinforcement Learning Temporal difference methods for the variance of the reward to go

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.696648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.460172Z digest=sha256:cb43fb29e5fe0e864be68865284d5761f94d06b3a3d5b636d7cc229fe49df0fe

Observation 1d70e6be-1833-424a-8977-010079b7e4bf · outbound

This paper cites The gambler’s ruin ap- proach to business risk.

Practical Risk Measures in Reinforcement Learning The gambler’s ruin ap- proach to business risk

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.676757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.466159Z digest=sha256:c80d72544adcc333882d53deddd54218669bcfa274fd1f341a561e4515edd9e6

Observation fbdd48ab-eb9d-43d0-aae3-f6c5ab4cb10c · outbound

This paper cites an unresolved cited work.

Practical Risk Measures in Reinforcement Learning Unresolved cited work

Reference 1186

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:47:00.982440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.322177Z digest=sha256:c0edb03b706e9baf091b6244b79591574c62edb71d5693e114a6641cb0116dca

Observation 9e94c784-f3a5-42a1-9e30-acec831f94d4 · outbound

This paper cites Dynamic probabilistic systems: Markov models, volume.

Practical Risk Measures in Reinforcement Learning Dynamic probabilistic systems: Markov models, volume

Reference 1972

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.090805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.290889Z digest=sha256:ac3629fd792424647f78cb0c52ec890deebdd4b2108e3c69bd31a9706813e3c6

Observation d0cb7324-dce7-49bf-bf05-d3ffd12d5fa2 · outbound

This paper cites Percentile performance criteria for limit- ing average markov decision processes.

Practical Risk Measures in Reinforcement Learning Percentile performance criteria for limit- ing average markov decision processes

Reference 1974

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.212764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.252471Z digest=sha256:3b39822df6ba649c6452191d2881856d8f6ae51c36ed5aaaa7a2eb02973f28e0

Observation 47e56ef4-44d9-4e50-8ad3-296d7d0ff9cc · outbound

This paper cites Reinforcement learning: An introduction.

Practical Risk Measures in Reinforcement Learning Reinforcement learning: An introduction

Reference 1982

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.442291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.442291Z digest=sha256:2825ab767ecda5257e1875338594ca69d449384bafce0009501a82abe3a69b4f

Observation e5516cde-bf17-4998-86c7-643b5d37f6c1 · outbound

This paper cites Policy gradients with variance related risk criteria.

Practical Risk Measures in Reinforcement Learning Policy gradients with variance related risk criteria

Reference 1988

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.715771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.454787Z digest=sha256:a4c3f036c661be84834fc3bf60ea9c60a5f9a06a4b90e22e4f27b2f3b278b9b2

Observation 24f94fc6-a7ea-4083-a22b-da5919f98226 · outbound

This paper cites Deep learning, vol- ume.

Practical Risk Measures in Reinforcement Learning Deep learning, vol- ume

Reference 1990

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.132329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.281120Z digest=sha256:2151f9839c952c358299159fa74f9237c19bbc8da1aa4d4a49c337a9cd357231

Observation fc743c7c-1e30-4516-a20a-4fe2e007a427 · outbound

This paper cites Optimization of conditional value-at-risk.

Practical Risk Measures in Reinforcement Learning Optimization of conditional value-at-risk

Reference 1994

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.824826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.416480Z digest=sha256:ad614acae23530f60cb71631af75682fccc6498aeca5579ed34828144be691de

Observation 7533c659-cf4a-4c3d-949a-65991729e8dd · outbound

This paper cites Monte Carlo: concepts, algorithms, and applications.

Practical Risk Measures in Reinforcement Learning Monte Carlo: concepts, algorithms, and applications

Reference 1995

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.196498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.257427Z digest=sha256:f113460bc075ef95e64759d9d72832c7439b17d9a972aaba355a4f195579dbf5

Observation fe836923-52ea-4ca8-991f-ffbd96e51f10 · outbound

This paper cites Safe policy iteration.

Practical Risk Measures in Reinforcement Learning Safe policy iteration

Reference 1998

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.898308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.384923Z digest=sha256:7aab7268ce9b9d500c74cbdd11611c1d346fc440e09482e5c38b4f2327b613a7

Observation 3e7b3fcc-7d6c-457f-9554-80bac6e4a827 · outbound

This paper cites Safe policy search for lifelong reinforce- ment learning with sublinear regret.

Practical Risk Measures in Reinforcement Learning Safe policy search for lifelong reinforce- ment learning with sublinear regret

Reference 1999

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.311760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.211287Z digest=sha256:b7fca4a30ef6ee545b0a426c16f18d4932c66665fd3b8ad82deecb1be4c08e70

Observation 9659f2d8-4bff-4b40-9f93-49a21b1bd4c9 · outbound

This paper cites Stochastic approximation and recursive algorithms and applications, volume.

Practical Risk Measures in Reinforcement Learning Stochastic approximation and recursive algorithms and applications, volume

Reference 2000

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.048918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.307149Z digest=sha256:5a09447fcfb88dca26eea45998c79a464939d9325f62807904a666539c65c699

Observation 99ac038d-8fbf-4f59-85c3-b88876d70d80 · outbound

This paper cites Dy- namic programming and optimal control, volume.

Practical Risk Measures in Reinforcement Learning Dy- namic programming and optimal control, volume

Reference 2001

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.276668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.229344Z digest=sha256:ced341cec63bcd54d9bd92aa22116a1521a7584ccb1e12efdb19c507c32e253e

Observation 7ad29f28-d649-425d-ac96-107b5674b402 · outbound

This paper cites Human-level control through deep reinforcement learning.

Practical Risk Measures in Reinforcement Learning Human-level control through deep reinforcement learning

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.346266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.346266Z digest=sha256:77d65338ec471f21757c19d4b07c9c410346d742df1288f33e905eb98a8f730c

Observation f102fa1f-3323-4da0-a0df-28fd66af8027 · outbound

This paper cites Deep learning.

Practical Risk Measures in Reinforcement Learning Deep learning

Reference 2003

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.026874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.312797Z digest=sha256:c61d9b2a0145ea0b9553b8f3e027fcfd2c3a19fa2cda76bedbb755959c368c94

Observation 82567b9a-80e6-4a58-ae6f-01dd20053b60 · outbound

This paper cites Model predictive control.

Practical Risk Measures in Reinforcement Learning Model predictive control

Reference 2005

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.261361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.235823Z digest=sha256:1f829d93b94bb2b549cb776961a4e52b7dbbc7f4b6eb1a66678a3586fee611c1

Observation c4994c29-6bdb-441c-a921-5c0017937f81 · outbound

This paper cites Existence and Finiteness Conditions for Risk-Sensitive Planning: Results and Conjectures.

Practical Risk Measures in Reinforcement Learning Existence and Finiteness Conditions for Risk-Sensitive Planning: Results and Conjectures

Reference 2006

Resolution
verified exact
local_arxiv, observed 2026-08-14T11:47:00.594354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.328004Z digest=sha256:b6d82ab06d8dea8e1fa6ccf6cb16f8d114eb09d9be5faa324cbc3e1d42d6e72b

Observation 98a3397f-b14f-427d-b44e-c62e1e69eb7c · outbound

This paper cites The variance of discounted markov decision processes.

Practical Risk Measures in Reinforcement Learning The variance of discounted markov decision processes

Reference 2009

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.766084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.436188Z digest=sha256:e32ca6b2cbe97011e829dd11cce597577f2ba0b325e5c6be4cb40e4faefe4662

Observation d88b01fb-298f-40c1-a244-7d19c6fc2aef · outbound

This paper cites Risk-sensitive reinforcement learning.

Practical Risk Measures in Reinforcement Learning Risk-sensitive reinforcement learning

Reference 2011

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.965665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.340096Z digest=sha256:276cedb7f85a1e2a0b3f83d60c6ebcab9838183d9898060db02f7bc37a546c05

Observation a8c73675-244b-4b88-bef3-af6ef33686bd · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Practical Risk Measures in Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.296273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.296273Z digest=sha256:6f91fad1e3709e48aee85749dabb95c22b59c400a6517bc7738b7a8145be8356

Observation f8cda374-623a-4141-b2c3-32946a77b97b · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria.

Practical Risk Measures in Reinforcement Learning Risk-constrained reinforcement learning with percentile risk criteria

Reference 2013

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.243517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.241793Z digest=sha256:a1ec6164dcac788bc09c24c079222b0d9fbba89c0df24b9981f97374f6d2ca3a

Observation f758d6f8-8c2f-47db-b1aa-d75bdddf43f3 · outbound

This paper cites Actor-critic algorithms.

Practical Risk Measures in Reinforcement Learning Actor-critic algorithms

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.072180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.302243Z digest=sha256:e38f97176ddd8b77ba17fa0ecec1162e136a95153f7210daba916933a37fe54b

Observation 1b2a7ee7-e7c6-4da9-b510-610e36ec7c36 · outbound

This paper cites Concrete Problems in AI Safety.

Practical Risk Measures in Reinforcement Learning Concrete Problems in AI Safety

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-14T11:47:00.216670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:47:00.216670Z digest=sha256:741bda6009109776c7beafd6c86042e6c247026020da59c747363485fd67763f

Observation 667f2c22-fb2c-40f6-b2e6-5872d5282a4e · outbound

This paper cites Infinite-horizon policy-gradient estimation.

Practical Risk Measures in Reinforcement Learning Infinite-horizon policy-gradient estimation

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.294961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.223403Z digest=sha256:6e0cb852b1403753c9777b8e7fd5fbcb602264de01e3feed48c3bed09a3c2f49

Observation c953d1a8-b7ee-4404-bf53-d5e7271a4d22 · outbound

This paper cites Constrained Markov decision processes, volume.

Practical Risk Measures in Reinforcement Learning Constrained Markov decision processes, volume

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:01.328689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.205811Z digest=sha256:18e1d87f0a05839adba5fd9d82399e9bcf06c08fe5df5044adcd95a65302b4d7

Observation 111b9753-ed5f-4a49-8c6d-1349bb516e05 · outbound

This paper cites Learning to predict by the methods of temporal differences.

Practical Risk Measures in Reinforcement Learning Learning to predict by the methods of temporal differences

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:47:00.734699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T11:47:00.448620Z digest=sha256:7125c312e6af6a699c160b345e53b490851eeb7d77591ddb82066de469c36ae1

Pith citing papers

No inbound Pith citation observations are available.