Pith. sign in

Paper Citation Record · LEDGER

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints

As of 12 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2412.04327.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.04327 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:39:12.005657Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved12
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 97fd7f80-2c1c-4987-8a9b-6814ec00a16a · outbound

This paper cites Safe Exploration in Continuous Action Spaces.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Safe Exploration in Continuous Action Spaces

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.928666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.928666Z digest=sha256:7116364dbd76ed9507848200ff84056eb72aeb5fdad73ab6a65c3b1a370e18f7

Observation 05999720-a7d4-4a56-9c5c-e66a0a6ebcc1 · outbound

This paper cites Feasible Actor-Critic: Constrained Reinforcement Learning for Ensuring Statewise Safety.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Feasible Actor-Critic: Constrained Reinforcement Learning for Ensuring Statewise Safety

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.955651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.955651Z digest=sha256:5f8221d2002a072d6fce1d742d6a2f045a5a891eee13b5fb6de100d0fe247bb4

Observation 96f35f8a-f1fd-4bb7-b509-1ae6ea2e49c5 · outbound

This paper cites Benchmarking Batch Deep Reinforcement Learning Algorithms.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Benchmarking Batch Deep Reinforcement Learning Algorithms

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.960949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.960949Z digest=sha256:e7319d0932fd9254a372078f4b3d9d2d1c98a224f26c0740bfe31c525be67a92

Observation 6a1a3f2b-7fd3-4ba6-ac6a-352b744262ee · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.966125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.966125Z digest=sha256:b1f273e6cdd63ed2d2b7ee7a876eb038dd24f10a49aa34a1a2e3c394dd1003dd

Observation d8bcd68a-5b75-45fe-9f23-fe56df76a5a0 · outbound

This paper cites Excluding the Irrelevant: Focusing Reinforcement Learning through Continuous Action Masking.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Excluding the Irrelevant: Focusing Reinforcement Learning through Continuous Action Masking

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-11T21:39:12.111309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T21:39:11.982996Z digest=sha256:a3b007fc1000b11203136ab34c8121d55d0c7bcf316e504701b83b92255a2b1f

Observation 083d31d0-75dc-4126-bd6a-00f468607559 · outbound

This paper cites RaceMOP: Mapless Online Path Planning for Multi-Agent Autonomous Racing using Residual Policy Learning.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints RaceMOP: Mapless Online Path Planning for Multi-Agent Autonomous Racing using Residual Policy Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-11T21:39:12.086025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T21:39:11.988838Z digest=sha256:38657dd5454e1d7ea287f887eaee35a5db9090fc38b063c30adf92c69e8ca681

Observation 2187593e-e1e6-4cd9-b695-80d81a08ed3d · outbound

This paper cites Penalized Proximal Policy Optimization for Safe Reinforcement Learning.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Penalized Proximal Policy Optimization for Safe Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.994473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.994473Z digest=sha256:f9ee6735423bf14ae527faf8acab5ab4c1bcf4d5442d9e140c86545108091ca8

Observation d53b5711-1645-427e-84c0-ffa6c9bbe208 · outbound

This paper cites an unresolved cited work.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Unresolved cited work

Reference 17

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T21:39:12.294747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T21:39:12.005657Z digest=sha256:5a81c3d46cc2a7dc7c10755d2c6668dc31f100b66af8d90d78e78cda1e86da01

Observation 9ec14cd6-f8f0-4220-b0b0-ed3bc72dc424 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Proximal Policy Optimization Algorithms

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.972205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.972205Z digest=sha256:2d1d7f25885ff2183ea1890eb55d102b8b1c54dc10cdaa8603b2bf1f17c6b3d1

Observation 34ee1234-8034-48d5-894a-9d9845beab7a · outbound

This paper cites Learning to be Safe: Deep RL with a Safety Critic.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Learning to be Safe: Deep RL with a Safety Critic

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.977292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.977292Z digest=sha256:676e95f8d74d8d3811df421ee1488b303d6321e3c051c8083d5c145f1728c57f

Observation c50f1d80-fa3f-41ac-af74-e2ad62556fc9 · outbound

This paper cites A Closer Look at Invalid Action Masking in Policy Gradient Algorithms.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints A Closer Look at Invalid Action Masking in Policy Gradient Algorithms

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.945279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.945279Z digest=sha256:61385232c6d1b96795270aa989612b01cef0324a44a1cad67a1a053781526d39

Observation 8feeebb4-3998-42f7-9843-17fff94009c1 · outbound

This paper cites Lyapunov-based Safe Policy Optimization for Continuous Control.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Lyapunov-based Safe Policy Optimization for Continuous Control

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.923300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.923300Z digest=sha256:85208f1e477878879b6b4253def61ff6ac29da2f9a2f0de876c3f942a2c78046

Observation e0592a37-9612-4e8e-9015-700333029c65 · outbound

This paper cites Safe reinforcement learning for autonomous lane changing using set-based prediction.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Safe reinforcement learning for autonomous lane changing using set-based prediction

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:39:12.311768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T21:39:11.950595Z digest=sha256:54b5eb7802d48149fdd7fbe98542d0c4c204212e287b8cef93a40150fd7a5166

Observation bd1b9821-a7db-4cbd-9802-90cf01a2c2b9 · outbound

This paper cites Conservative Safety Critics for Exploration.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Conservative Safety Critics for Exploration

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.917036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.917036Z digest=sha256:e09dcb7b06c2a22ce815c64914e2b43286c00da055c790b559351c03f60fcbde

Observation 1a9e4239-f93f-40a9-9e5a-a294b0382eac · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.939149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.939149Z digest=sha256:414035839f955246fb645502e7655b3a131c1c4db23629631b9adbe017ae919f

Observation 350272cd-c349-4cbb-a677-6fefca8699a8 · outbound

This paper cites State-wise Safe Reinforcement Learning: A Survey.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints State-wise Safe Reinforcement Learning: A Survey

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.999964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.999964Z digest=sha256:6db2ff1d109ee6dda4113953276f9aacaced81015d6bdd5b2494b9b82f87dbc5

Observation ead9282f-7d07-40b5-a522-5db4c326d1ea · outbound

This paper cites Niklas Funk, Georgia Chalvatzaki, Boris Belousov, and Jan Peters.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Niklas Funk, Georgia Chalvatzaki, Boris Belousov, and Jan Peters

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T21:39:12.327333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T21:39:11.934161Z digest=sha256:0ba64b50d6ba0c12a7aecbb9e709d62cbc2e056c9a68bb359e524acb0561da3d

Pith citing papers

No inbound Pith citation observations are available.