Pith. sign in

Paper Citation Record · LEDGER

Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2104.09469.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.09469 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:23:42.450942Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:29:48.319428Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 41845b1c-2c55-4e83-898f-f811de7a69a1 · inbound

The Odyssey of the Fittest: Can Agents Survive and Still Be Good? cites this paper.

The Odyssey of the Fittest: Can Agents Survive and Still Be Good? Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T19:23:42.450942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T19:23:42.450942Z digest=sha256:cc7a3879c42b31100e936b3a925f75f423d3d71cf9267c88e75f58b2853cc6bd

Observation a3dbad0a-4ed1-45e7-97ba-f38945127bb1 · inbound

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning cites this paper.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:29:48.390968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:29:45.440912Z digest=sha256:3d1cdcf714c9e9cca180a395c5f15d134eff903579aa5b0779614efbded28415