Pith. sign in

Paper Citation Record · LEDGER

State-wise Safe Reinforcement Learning: A Survey

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2302.03122.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.03122 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:00:05.724349Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T09:24:32.068752Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 350272cd-c349-4cbb-a677-6fefca8699a8 · inbound

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints cites this paper.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints State-wise Safe Reinforcement Learning: A Survey

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.999964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.999964Z digest=sha256:8f227ea15c4efb15633b1566a663eed9246c73b2003f39fa613d0987f82c35d5

Observation 54538de1-aa23-4c20-b6c2-5bd8e2cfe1de · inbound

Leveraging Constraint Violation Signals For Action-Constrained Reinforcement Learning cites this paper.

Leveraging Constraint Violation Signals For Action-Constrained Reinforcement Learning State-wise Safe Reinforcement Learning: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T19:00:00.440808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T19:00:00.440808Z digest=sha256:bf5852041dde12da2bb04176743df751cc07a84d38ad6a6451744aa15c3e0f80

Observation 79e8a757-6df0-4b3d-b3c7-3dc1547e5468 · inbound

Continuous World Coverage Path Planning for Fixed-Wing UAVs using Deep Reinforcement Learning cites this paper.

Continuous World Coverage Path Planning for Fixed-Wing UAVs using Deep Reinforcement Learning State-wise Safe Reinforcement Learning: A Survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:05.724349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:05.724349Z digest=sha256:eb73879c02e6db0b187bb522a45bc7b56d1f0807386c72fc8d0cd6ea129d9e86

Observation eaefb24a-b4ec-42bd-9bc9-8dbcf3644343 · inbound

Combee: Scaling Prompt Learning for Self-Improving Language Model Agents cites this paper.

Combee: Scaling Prompt Learning for Self-Improving Language Model Agents State-wise Safe Reinforcement Learning: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T10:33:19.825936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T10:33:19.825936Z digest=sha256:2da0778463e3c6f9a5a4d98a30a89dc971a56f2101b9890d6ccbecc77cd2f1ef

Observation bd5d61c3-a6c1-4275-a68b-2cb198109785 · inbound

Safe Deep Reinforcement Learning for Building Heating Control and Demand-side Flexibility cites this paper.

Safe Deep Reinforcement Learning for Building Heating Control and Demand-side Flexibility State-wise Safe Reinforcement Learning: A Survey

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:22:37.422246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T08:19:35.360048Z digest=sha256:08ba8d0df73d6d0c3b001517a85094b744cc1e9b30dd11309867bbebeaab4ab4

Observation f4b85175-ceaa-4575-ae0f-36d31971f233 · inbound

Hierarchical Decision Making with Structured Policies: A Principled Design via Inverse Optimization cites this paper.

Hierarchical Decision Making with Structured Policies: A Principled Design via Inverse Optimization State-wise Safe Reinforcement Learning: A Survey

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:24:32.070573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-30T09:24:25.190688Z digest=sha256:f88b7c142b53e8cbf8ec728f1bb29c29a35619fa27532aefd61c8ca86d5f1f5a