Pith. sign in

Paper Citation Record · LEDGER

Conservative Safety Critics for Exploration

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2010.14497.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.14497 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:39:11.917036Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

32
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bd1b9821-a7db-4cbd-9802-90cf01a2c2b9 · inbound

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints cites this paper.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Conservative Safety Critics for Exploration

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.917036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.917036Z digest=sha256:e7e6a827a5a8988d00b128b20de737959813cd17a7154004ede1246d0ae70bb6

Observation 33693b59-16e0-45ff-8b2c-7ff49c725d71 · inbound

Confidence-Guided Human-AI Collaboration: Reinforcement Learning with Distributional Proxy Value Propagation for Autonomous Driving cites this paper.

Confidence-Guided Human-AI Collaboration: Reinforcement Learning with Distributional Proxy Value Propagation for Autonomous Driving Conservative Safety Critics for Exploration

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:47.251327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:06:47.251327Z digest=sha256:035200d4cbc1255a9d66b03115fcfb3c181b57453ece722c6269238eae78904e

Observation 648534cc-54f6-48bc-be58-07809e82f975 · inbound

Safe and Performant Deployment of Autonomous Systems via Model Predictive Control and Hamilton-Jacobi Reachability Analysis cites this paper.

Safe and Performant Deployment of Autonomous Systems via Model Predictive Control and Hamilton-Jacobi Reachability Analysis Conservative Safety Critics for Exploration

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:50:27.426347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:50:27.426347Z digest=sha256:d14d75011b350517ada56d6906060d21264d84c14243aa9477b1c7437a4c75b3

Observation 334bd6ef-b2c9-4c2d-ba11-59d06e5ab319 · inbound

Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning cites this paper.

Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning Conservative Safety Critics for Exploration

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:16:20.283541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:16:20.283541Z digest=sha256:9819de26c3e15dd69dabb5ff74ae056ee8a2100b54b669ba345142815774c6ec

Observation 71fc7cc8-6839-4a25-9b29-4ff614bf22f4 · inbound

Safe-Support Q-Learning: Learning without Unsafe Exploration cites this paper.

Safe-Support Q-Learning: Learning without Unsafe Exploration Conservative Safety Critics for Exploration

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:36:38.584131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T16:36:35.746034Z digest=sha256:02d449ba9b6e23aa669552567709c8e0a83d18336800ffa21d17f73c9e6ae394

Observation 6ad14b2d-ed9d-4171-877a-979f3d4c58cb · inbound

Pre-VLA: Preemptive Runtime Verification for Reliable Vision-Language-Action and World-Model Rollouts cites this paper.

Pre-VLA: Preemptive Runtime Verification for Reliable Vision-Language-Action and World-Model Rollouts Conservative Safety Critics for Exploration

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:04:41.680504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T07:02:27.381182Z digest=sha256:c82f3b442c03d5688090b6d4f5b2f0f8a7f3d6188ce6f7eac8fed701a93d9407

Observation 00c8031e-9ab2-4d22-b2b1-83ada45ad16a · inbound

Safe Embodied AI for Long-horizon Tasks: A Cross-layer Analysis of Robotic Manipulation cites this paper.

Safe Embodied AI for Long-horizon Tasks: A Cross-layer Analysis of Robotic Manipulation Conservative Safety Critics for Exploration

Reference 79

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:56:56.755323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T01:46:36.081851Z digest=sha256:1bd781e1e3ba320d6e57c2f46feb3106271ec7f381c0976cd46b2fbd958e80d0

Observation dc6f4182-b742-4d66-9fd8-d342581fa8e5 · inbound

An Agency-Transferring Model-Free Policy Enhancement Technique cites this paper.

An Agency-Transferring Model-Free Policy Enhancement Technique Conservative Safety Critics for Exploration

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-27T17:21:06.998970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T17:11:45.240357Z digest=sha256:020582d580545704cab3dcad576fcd1eeed4ff6249833fc5fb18d84c641f1f8f

Observation 1530175c-0c1e-4df3-8196-e8a36b27e1f0 · inbound

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration cites this paper.

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration Conservative Safety Critics for Exploration

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:29.846675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T16:58:19.848530Z digest=sha256:df111eca9757760a87187ddf3673711914682982e7ea9f00a8856334fc8010f7

Observation d3db1e45-529a-400a-b05a-6e952f5c1b3d · inbound

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning cites this paper.

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning Conservative Safety Critics for Exploration

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.695906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T09:46:59.746745Z digest=sha256:be05ffc5f240a132bba36d141f2da4e5c46ea0538be6bc0851443a856058572d