Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T01:55:22.266761Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2512.04246.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T01:55:22.266761Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 65fc118f-8b8e-4d5b-be22-1152804a5203 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Towards artificial virtuous agents: games, dilemmas and machine learning.AI and Ethics, 3(3):663–672
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b8184140-7a14-49c7-b189-e15938be5ce2 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Reinforcement learning as a framework for ethical decision making
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0b603340-0650-4831-9dda-69d18fe323bb · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Reinforcement learning and machine ethics: a systematic review
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3d317330-0976-4308-acf6-f8405188ed22 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Groundwork of the metaphysic of morals
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 165de463-5327-4399-bc58-858b22e40eba · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Utilitarianism
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6a602ff6-f407-412c-bcfc-0ddac7a374b3 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Safe reinforcement learning via shielding
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 93f96e0d-dead-4312-9175-88f88be0362e · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Can model-free reinforcement learning explain deontological moral judgments?Cognition, 150:232–242
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea0ae0e3-ee69-4e20-8545-6b5890ae6943 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Reinforcement learning under moral uncertainty
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cf320c6b-aa29-4dfe-bea4-e1f9f9dbb7c0 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Q-learning as a model of utilitarianism in a human–machine team.Neural Computing and Applications, 35(23):16853–16864
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0041ea57-7882-47b9-ac40-77df2b52780b · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Artificial morality: Top-down, bottom-up, and hybrid approaches
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6e1235e9-c99d-4d1c-97b8-b5e20f6e2355 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Building ethically bounded ai
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b0b3d56a-599b-4450-8476-ab972085d163 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Building Ethics into Artificial Intelligence
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4797add2-a982-442d-ab9b-39d7dd1263f0 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap A low-cost ethics shaping approach for designing reinforcement learning agents
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5094bf44-3db7-4cc3-88d1-19f68825dcf7 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Virtuous vs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d044e971-4ef7-4f6d-970f-68b7082cc029 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Teaching ai agents ethical values using reinforcement learning and policy orchestration.IBM Journal of Research and Development, 63(4/5):2–1
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29d42ca0-a2b1-4545-be7d-b809c8d04ff9 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Cambridge University Press
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f8e052f7-d963-4aea-a4e5-298b88ef6df0 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Right action and the non-virtuous agent.Journal of Applied Philosophy, 28(1):80–92
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7f1e9bae-cfcb-4c21-8af9-83b56572ba6c · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Introduction to Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 089947a4-0e68-4e35-9f5f-7441dcb8e630 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Joint Attention for Multi-Agent Coordination and Social Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 35039a66-c971-4ece-bbef-011ca266d317 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Learning few-shot imitation as cultural transmission.Nature Communications, 14(1):7536
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ddf573f2-442c-44d0-86d3-9496a13d9d89 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap An Efficient Open World Environment for Multi-Agent Social Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8efbd136-d07c-4f8b-b2a6-0b38994e95bb · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Emergent social learning via multi-agent reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a641980-ca1d-4a19-91fb-6e137c7a6a18 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Social influence as intrinsic motivation for multi-agent deep reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a88686ce-f456-4a6a-8516-9bd636c56849 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Multi-objective reinforcement learning: an ethical perspective
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8a690175-566a-4676-a538-cfc137fbc855 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Exploring affinity-based reinforcement learning for designing artificial virtuous agents in stochastic environments
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 55ec86bd-7830-47c0-a72d-9ab0bf735c33 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap The core of confucian learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 98acc1b5-0419-47db-ad4e-c3ef28a39a9b · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap The daoist thought of wu wei–action through non-action and its influence in vietnam.Synesis (ISSN 1984-6754), 17(2):55–71
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 789ad6f0-1e3d-4e3b-b5a9-c11f7170a1f3 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap An anthology of philosophy in persia
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c8b9c0a0-1b71-403d-bbd0-000513057181 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Composable modular reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4baf5916-afb9-4490-8e2d-77a4c7d1e6fe · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Formal verification of ethical choices in autonomous systems.Robotics and Autonomous Systems, 77:1–14
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de968810-320b-43e0-b86b-1cbb1271b02d · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Ltl and beyond: Formal languages for reward function specification in reinforcement learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6dfc234c-3620-48bf-9429-82d71cbe5d65 · outbound
Toward Virtuous Reinforcement Learning: A Critique and Roadmap Using reward machines for high- level task specification and decomposition in reinforcement learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.