Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:13:43.733204Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2608.07280.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:13:43.733204Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f7fe8463-6b1b-439d-9b0a-a7a5a72249c2 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Advances in neural information processing systems , volume=
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ec1edc5a-9081-485d-800e-82d8864b1820 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction 2018 , eprint=
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7cc837c2-d7b2-4c0f-a5fb-dd00e1c11993 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Advances in neural information processing systems , volume=
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71d450c1-3f68-4754-9fa6-92b424713fe1 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction International conference on machine learning , pages=
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a2c1d3d1-5256-4209-90a0-2711984ca94e · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Advances in neural information processing systems , volume=
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8554099-b597-4900-881d-38e801740a5f · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Advances in Neural Information Processing Systems , volume=
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a59754-7dee-4f13-8c8b-39e3547ac778 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Artificial Intelligence Review , volume=
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81c53664-7fd8-4ccb-ae36-26b766c0b2a5 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Journal of Economic Literature , volume=
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e615f39e-6ab3-424d-adcf-e1ee62097986 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction PLOS Computational Biology , volume=
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c8d3e01d-e829-4f5f-bd51-9923dbbdaa9e · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Simulating social phenomena , pages=
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4db33246-0bc4-4d02-a42e-ff088c392bf9 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction 2019 , publisher =
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 466debc6-3962-4a6d-ba99-aed21feb3141 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 24f6ee4c-6639-48c4-a8dd-a2036bee64f4 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Science , volume=
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8138da85-2c4e-47b9-b167-844e603a99b3 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Colorado Technology Law Journal , volume=
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e3beb7c4-40dc-406f-9a9a-71ef3a6967dd · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Social theory re-wired , pages=
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 772b6fa1-17bd-43fe-be72-fb9252d1c23e · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Building a foundation for data-driven, interpretable, and robust policy design using the
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c6913b40-331a-4f1c-9445-4209f962abe8 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction and Socher, Richard , journal =
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca16c811-ef3c-4255-b55f-be74f77224d6 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Advancing the art of simulation in the social sciences
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e266f9cb-1136-4b27-a486-57b0b11f2f6d · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Agent-based modeling in economics and finance: Past, present, and future
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1855dab7-9c96-4c50-86f5-f6ef6092c94a · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Exposure to ideologically diverse news and opinion on facebook
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d0deff68-77db-4057-8991-b8f4651853d0 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Deep reinforcement learning from human preferences
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c152cc0-8519-4940-a48a-098ad4e570ee · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Multi-agent deep reinforcement learning: a survey
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 25b2c824-f475-4c3a-92b2-7dbd775a5725 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Leibo, Matthew G
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation be5384b4-43b1-4dac-8c38-86ff9242eb46 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Social influence as intrinsic motivation for multi-agent deep reinforcement learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 76f78d79-e76e-4b47-ad35-cc562439239e · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Covasim: an agent-based model of covid-19 dynamics and interventions
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9b38be00-60a9-4cd4-aa02-b1c2c6280777 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Multi-agent Reinforcement Learning in Sequential Social Dilemmas
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cc8b0a5-815d-48cf-9357-d58ee0e680b0 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Scalable agent alignment via reward modeling: a research direction
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1814b975-231d-4f06-9a20-ce47150dea6f · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction The Alignment Problem from a Deep Learning Perspective
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd68d971-23c1-4a40-a31a-ebe672ce84c4 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Training language models to follow instructions with human feedback
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b3aeb6c-a77a-48bb-b7d8-6653c949b0cc · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction A multi-agent reinforcement learning model of common-pool resource appropriation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 77083b68-abee-4377-80ee-8c649204d8c4 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Learning to summarize with human feedback
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a975993b-a186-492f-9bf0-9e3c34dd226f · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Building a Foundation for Data-Driven, Interpretable, and Robust Policy Design using the AI Economist
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d55d557-fc4d-49db-beb6-b7bb9aaeff6d · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Algorithmic harms beyond facebook and google: Emergent challenges of computational agency
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fbc3a772-0c88-4470-8403-e76f567df19a · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction An open source implementation of sequential social dilemma games
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation af4e319c-b131-4fb0-9454-55da045c73ca · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Rewarded Region Replay (R3) for Policy Learning with Discrete Action Space
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 651724a3-7dd8-4a79-a777-9f708a4fed78 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Parkes, and Richard Socher
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 190f5422-449a-42dd-bb65-f16d434bcae3 · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction Decoding global preferences: Temporal and cooperative dependency modeling in multi-agent preference-based reinforcement learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ecca495a-72c2-4db3-bffb-e35bd422615b · outbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction The age of surveillance capitalism
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.