Pith. sign in

Paper Citation Record · LEDGER

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis

As of 13 August 2026, this Paper Citation Record lists 100 of 146 outbound references and 0 inbound Pith citation observations for arXiv:2412.02091.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02091 v2

Coverage vector

measured 100 of 146 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:56:00.724470Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 146 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved73
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6217f445-406e-44b4-a06d-54d2ceeb8f01 · outbound

This paper cites an unresolved cited work.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.337546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.337546Z digest=sha256:988c712f417645b31be3ade74671b4c28f47e12628cd87cbe276d9f9cac23864

Observation 6c85ce51-012a-467c-8064-bf73762d4e4b · outbound

This paper cites A model of online misinformation.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A model of online misinformation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.342427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.342427Z digest=sha256:d644302f398378ed66566bdf2719e8723691d3473382fdd408a450b0f458d08d

Observation a1fffa89-d828-454b-adac-9f87c050b9eb · outbound

This paper cites The multiplicative weights updatemethod: ameta-algorithmandapplications.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The multiplicative weights updatemethod: ameta-algorithmandapplications

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.346995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.346995Z digest=sha256:53e8cb969ad18f3e11bea72fe3383e78a0991d03f9f9b4b9f27e788c4b59eb85

Observation 6a198f9b-9f1a-43ef-92b3-df9b37d6d00d · outbound

This paper cites Deep reinforcement learning: A brief survey.IEEE Signal Processing Magazine, 34(6):26–38, 2017.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Deep reinforcement learning: A brief survey.IEEE Signal Processing Magazine, 34(6):26–38, 2017

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.351316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.351316Z digest=sha256:bd71e5b64a7ad3091920c05be0d872d202cb529d6666b22672f6b9a909f0fe9f

Observation adc18232-2d67-458d-8305-8aae9710d867 · outbound

This paper cites An efficient dynamic mechanism.Economet- rica, 81(6):2463–2485, 2013.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis An efficient dynamic mechanism.Economet- rica, 81(6):2463–2485, 2013

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.355692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.355692Z digest=sha256:2418211f474ac9d2e7c57f4d65cb441976ec32dc47e33f42cc691222af03d5dd

Observation affb08b9-945c-4c10-b7c0-72377086f379 · outbound

This paper cites Using confidence bounds for exploitation-exploration trade- offs.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Using confidence bounds for exploitation-exploration trade- offs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.359629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.359629Z digest=sha256:b8c1132ccf50926559cab9d63d2cb1c1dfdf43c8ebe7dd5c205544c49ec429be

Observation 23e1bc6d-beef-4611-ac9b-63c33b908f4f · outbound

This paper cites The nonstochastic multiarmed bandit problem.SIAM Journal on Com- puting, 32(1):48–77, 2002.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The nonstochastic multiarmed bandit problem.SIAM Journal on Com- puting, 32(1):48–77, 2002

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.364212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.364212Z digest=sha256:83107e10b80f5e62fc75e7aafa762389213ccacbb603d7a5e55e23ef4f6594cf

Observation d691e754-5f9b-40ad-8f44-031399ce9f25 · outbound

This paper cites Correlated equilibrium as an expression of Bayesian rationality.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Correlated equilibrium as an expression of Bayesian rationality

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.368304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.368304Z digest=sha256:647c545dd184e8c864581c037a4a3cfe199596caf11cfc9b722c72f1d4c8c7ff

Observation 8aa510b6-94e2-4128-885c-0bdebbb46c8f · outbound

This paper cites The emergence of cooperation among egoists.American Political Science Review, 75(2):306–318, 1981.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The emergence of cooperation among egoists.American Political Science Review, 75(2):306–318, 1981

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.372307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.372307Z digest=sha256:08e399a534d21d7eed50a59d7f0272e806a26a40e1683bc8b16b56ce1f79c380

Observation 1ce94c58-1974-430b-8995-47379ed9790f · outbound

This paper cites The Formula: The Universal Laws of Success.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The Formula: The Universal Laws of Success

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.376120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.376120Z digest=sha256:f744f302ee9d6548e53ab1d7a8f4c5aeaaa9dc20bf28f624ae0fa1e75f119f67

Observation 1cb87b26-b958-4b6a-b781-43887567c6ee · outbound

This paper cites Dynamic incentives for congestion control.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Dynamic incentives for congestion control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.380923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.380923Z digest=sha256:a6ba37d3fd1d71da98c6f89c7ec1674a5add623eb1bad3f776471e1b6edb64dd

Observation 2bd0a905-a0f3-4799-a535-39ab03e0dae8 · outbound

This paper cites A neural probabilistic language model.J.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A neural probabilistic language model.J

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.385172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.385172Z digest=sha256:8a5a95d58223ab7411b13bcb3d5761c04eb6ff8fd085be4f241136b0aadc6d71

Observation 280058df-b74d-447e-b414-a556cc22816c · outbound

This paper cites Taming the Matthew effect in online markets with social influence.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Taming the Matthew effect in online markets with social influence

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.389340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.389340Z digest=sha256:bd2e41326ea33055ad1bb93610412ea85a768632f25e503b074e6b5e656cbaa0

Observation b0e9f2db-6e59-4e64-9874-2813c2813246 · outbound

This paper cites The dynamic pivot mechanism.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The dynamic pivot mechanism

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.393254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.393254Z digest=sha256:68e28b6d982ceb0c52746971291c5ca15007ebaea7594b3b382707072c870234

Observation 2ee5a282-4ee7-44ed-82a3-e57a9438f01c · outbound

This paper cites Dynamic mechanism design: An introduction.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Dynamic mechanism design: An introduction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.397020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.397020Z digest=sha256:e11cf8c7210533df3b14f0233e7e9376d507bc3c0b6e27046ff005186950d018

Observation 560cbd1e-87ca-40b5-8f2c-ddbabff06711 · outbound

This paper cites From external to internal regret.Jour- nal of Machine Learning Research, 8(6), 2007.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis From external to internal regret.Jour- nal of Machine Learning Research, 8(6), 2007

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.400660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.400660Z digest=sha256:3d61e0d155fb4d068fad9ecbd4355fbd8f046d177cad9443a98172710b991f3d

Observation 46cf5c75-9b58-49ab-baf8-da926974aaee · outbound

This paper cites An Introduction to the Theory of Mechanism Design.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis An Introduction to the Theory of Mechanism Design

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.404299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.404299Z digest=sha256:3189696f958287e897a66a59cccfae98859cf14b23528c496800909ec5852614

Observation a69de65e-3f50-4729-8bd6-430d598feae9 · outbound

This paper cites Superintelligence: Paths, Dangers, Strategies.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Superintelligence: Paths, Dangers, Strategies

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.408171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.408171Z digest=sha256:d11813d7b367ab93eefa7a31619b4ecd7e1565051f05e56906ff4d7e8aeb948d

Observation e395d10c-5de4-4fd6-b6c5-c7bf3e40bc55 · outbound

This paper cites Ethical issues in advanced artificial intelligence.Machine Ethics and Robot Ethics, pages 69–75, 2020.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Ethical issues in advanced artificial intelligence.Machine Ethics and Robot Ethics, pages 69–75, 2020

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.411991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.411991Z digest=sha256:91544f3d3929698964e56907986af7e4c6338cc6c61c943279dddf00689e287d

Observation 7dc6ae90-b268-44db-802a-5d3415663f28 · outbound

This paper cites Learn- ing to mitigate AI collusion on economic platforms.Advances in Neural Information Processing Systems, 35:37892–37904, 2022.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Learn- ing to mitigate AI collusion on economic platforms.Advances in Neural Information Processing Systems, 35:37892–37904, 2022

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.415614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.415614Z digest=sha256:4e16cbcf0b9412a2cf9255a0e9d7719f8f26644ad6d3ff9c49c3d8579596fec3

Observation 6731de7f-a2ef-43a3-8e22-6874b91483ae · outbound

This paper cites A survey of monte carlo tree search methods.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A survey of monte carlo tree search methods

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.419499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.419499Z digest=sha256:4545c8c0e9fab287f32a59c775a0fbe763d0e7b7089487e0a8e571a6c91e9902

Observation 2285585f-da61-4f76-9bb1-1dd29e1d5abd · outbound

This paper cites A comprehensive survey of graph embedding: Problems, techniques, and applications.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A comprehensive survey of graph embedding: Problems, techniques, and applications

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.423612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.423612Z digest=sha256:71f73cc744e05d074b2bd72b78cbeebf61f96c2ab4d1a6eeba08a98ee9752d41

Observation 72a1d687-08a5-48c9-a87b-e8e7b53c8a08 · outbound

This paper cites Artificial intelligence, algorithmic pricing, and collusion.American Economic Review, 110(10):3267–3297, 2020.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Artificial intelligence, algorithmic pricing, and collusion.American Economic Review, 110(10):3267–3297, 2020

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.427413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.427413Z digest=sha256:8c2a54e86d1a95f5de52a43fb1cd185a6c5d472a1b2d0b3d23006fb43db51f14

Observation ef6e43d6-345d-4334-b033-1770f70f3aad · outbound

This paper cites Self-predictive universal AI.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Self-predictive universal AI

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.430984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.430984Z digest=sha256:3b8d2e2c674b48d33e71edb1b95da7c6b9ef6d37a4f009a0597a8708453b25ea

Observation 10f84c2a-5ed3-458a-a767-21158954d741 · outbound

This paper cites Optimal coordi- nated planning amongst self-interested agents with private state.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Optimal coordi- nated planning amongst self-interested agents with private state

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.434733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.434733Z digest=sha256:284cd36be9b9f296f16c0cdc5f6a89c2552918da0b5856bc5996054ab1b1319a

Observation c2f5f229-1533-49ce-9a68-c094b43ee44d · outbound

This paper cites Cambridge University Press, 2006.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Cambridge University Press, 2006

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.438409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.438409Z digest=sha256:6466519100c0e715960bf25c8db69fa98e90b3ff9964ffec340ba51b5b001f09

Observation 6ba7a7c9-79a5-4aaa-b312-3422c562c94f · outbound

This paper cites Evolu- tionary dynamics of biological auctions.Theoretical Population Biology, 81(1):69–80, 2012.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Evolu- tionary dynamics of biological auctions.Theoretical Population Biology, 81(1):69–80, 2012

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.442224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.442224Z digest=sha256:f0c7488d9321e74a3e0ab4b572095af950abc84966a4afee237ab10516388a56

Observation def6f621-f15b-4cde-a4ee-dfdf52802156 · outbound

This paper cites Dynamic pricing in a labor market: Surge pricing and flexible work on the Uber platform.Ec, 16:455, 2016.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Dynamic pricing in a labor market: Surge pricing and flexible work on the Uber platform.Ec, 16:455, 2016

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.445778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.445778Z digest=sha256:33765a85652dcfd685cf587d0bb74834127650faac26a4342028e3618c51fae6

Observation 95490cd0-bbe9-4d93-b545-557981afda60 · outbound

This paper cites Hedging in games: Faster convergence of external and swap regrets.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Hedging in games: Faster convergence of external and swap regrets

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.449612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.449612Z digest=sha256:cb659f15c6f0c37da2d926e3dee24d000dad7c5cbd107ec33b733fd74c78342b

Observation 203acd08-516f-4a6f-aabf-190a9e5352c5 · outbound

This paper cites Prediction with expert evaluators’ advice.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Prediction with expert evaluators’ advice

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.453322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.453322Z digest=sha256:13c4b5b53ebf896487d7f02c3a27f09b184e936880d5c7e564f9ac4cd4e46755

Observation 74c9e1e0-4da9-44ea-9868-575381d29cbc · outbound

This paper cites an unresolved cited work.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.457197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.457197Z digest=sha256:ae69abc4498051cd8601fde516bdd2b97757baa8008148b515c4f878ecf198f1

Observation af3fdfba-3f9a-435e-ad48-fffda0e7bbe1 · outbound

This paper cites A formulation of the simple theory of types.Journal of Symbolic Logic, 5:56–68, 1940.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A formulation of the simple theory of types.Journal of Symbolic Logic, 5:56–68, 1940

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.460993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.460993Z digest=sha256:334846dc5b8b4d1eff12854c1f3d084589a68283d9f263989dff1a41a4e3a94a

Observation 9e51077a-1b86-4b7d-88ba-9b1fe2b78c11 · outbound

This paper cites The problem of social cost.Journal of Law and Economics, 3(1):1–44, 1960.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The problem of social cost.Journal of Law and Economics, 3(1):1–44, 1960

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.465002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.465002Z digest=sha256:9e1132271db95ef4e01b10205f13759e3f05865f169738397786a2150bca0c46

Observation bad48d59-e536-4348-9187-905218dc4bb2 · outbound

This paper cites The Firm, The Market, and The Law.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The Firm, The Market, and The Law

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.469204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.469204Z digest=sha256:e6852ee895023ff49bef22f16052775c75aabb5d51bd65b9cf34e4fb8d4482c9

Observation 14d07d0f-4d7c-4348-b4eb-20bdda64d7c2 · outbound

This paper cites Law for the platform economy.UCDL Rev., 51:133, 2017.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Law for the platform economy.UCDL Rev., 51:133, 2017

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.473652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.473652Z digest=sha256:8a25fe7a1735dbcd1aa8c92d433292af3fb1d71a9abb0842314586a7bc2adf77

Observation 9b1b5f03-b322-4923-9db5-96ade8787ba7 · outbound

This paper cites Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.477796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.477796Z digest=sha256:03fb9debe1a618587fd231ef85500af491a6142f4e38f98267ab7f55c345e43e

Observation 1040dfd4-fb2e-47d2-b49c-f86e17fec938 · outbound

This paper cites Cambridge University Press, 1996.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Cambridge University Press, 1996

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.482424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.482424Z digest=sha256:9df6982fc470fee7a73f2d163dcd9db5312cf7218bb54855d71f19d4430889c7

Observation dfe1b177-94ed-4836-bf51-1bf57c9e3842 · outbound

This paper cites A collusion-proof dynamic mechanism.SSRN, 2024.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A collusion-proof dynamic mechanism.SSRN, 2024

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.487015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.487015Z digest=sha256:0e19527bd90c647d7f266946377ea55cd642bc0b8b1ff030cdf5f39703f58d3c

Observation aeb00fe9-64d9-4d83-ba29-e0b6742b049b · outbound

This paper cites From external to swap regret 2.0: An efficient reduction for large action spaces.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis From external to swap regret 2.0: An efficient reduction for large action spaces

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.490732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.490732Z digest=sha256:24a9c279edc8672d02597ea05fd357b49f103e0b2c0fbf3dbac474c75c165c68

Observation 21e31179-4665-401c-a72c-6e029b107196 · outbound

This paper cites Logical and relational learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Logical and relational learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.494693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.494693Z digest=sha256:9bbc15e8c214dc05a3bdbf8c0035b2b8a909ed458595c59bdb68ad527064ae63

Observation dd52d47f-66f4-4baa-b6c7-654cc86ef84e · outbound

This paper cites The algorithmic foundations of differ- ential privacy.Found.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The algorithmic foundations of differ- ential privacy.Found

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.498928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.498928Z digest=sha256:1b645615edfd50f2e8f188189099e261e2374a8d92ecf5736185eb54781deaa9

Observation 38e0b50d-1752-47d5-9fc5-08256f631efa · outbound

This paper cites Relational reinforce- ment learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Relational reinforce- ment learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.502610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.502610Z digest=sha256:0085a3fc360cc6104d678cadcf42c85130b8cb8b12dc70c39e388fdb96f02b84

Observation 13ff3370-8792-4dde-bf84-71fa79b160a6 · outbound

This paper cites Matchmakers: the new economics of multisided platforms.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Matchmakers: the new economics of multisided platforms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.506290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.506290Z digest=sha256:0e582ac0cf6c60328a905d45409bed23152870b324735811351893a710bee3f9

Observation eb6ba134-f45b-491a-bc13-78baf0cc78a4 · outbound

This paper cites Reward tampering problems and solutions in reinforcement learning: A causal influence diagram perspective.Synthese, 198(Suppl 27):6435–6467, 2021.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Reward tampering problems and solutions in reinforcement learning: A causal influence diagram perspective.Synthese, 198(Suppl 27):6435–6467, 2021

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.510011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.510011Z digest=sha256:744f6f31b632f743986032d41bad9fda527b233f189a873c18afd5df67674da8

Observation 4d6a5807-2c83-40d2-88cc-b473e0a49b8d · outbound

This paper cites Re- inforcement learning with a corrupted reward channel.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Re- inforcement learning with a corrupted reward channel

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.513785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.513785Z digest=sha256:8b4d3186dad03d9690d72877ef42d60431c56b515b6a7c2c288b11e9aeaebdda

Observation f2698ef8-3976-4ce2-919c-ab6b274e234f · outbound

This paper cites AGI safety literature review.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis AGI safety literature review

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.517607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.517607Z digest=sha256:dd503f2859f3a2432940df046e3c7be7c9f39233ad7e2278636e8d77b94990fa

Observation 6d12efec-3c2a-482d-a66d-65f02b08094a · outbound

This paper cites Reflective or- acles: A foundation for game theory in artificial intelligence.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Reflective or- acles: A foundation for game theory in artificial intelligence

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.521276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.521276Z digest=sha256:c57e658d037f29ea1a67943633426823b1dfd2d382e13115f5a305226a0a1ae1

Observation af5359bf-0782-4a28-9534-4ea34a1a7360 · outbound

This paper cites an unresolved cited work.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.524751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.524751Z digest=sha256:d1992504021c403811d6b9f4df6a702793fee127c735030ca729b195698d733b

Observation 89de71b5-c4ab-48e9-bd2d-400ce6c7618b · outbound

This paper cites Economicsofoilrefining.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Economicsofoilrefining

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.528481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.528481Z digest=sha256:2c662a9665beff8de67029befd27bb60074de40e9fecba068d3ba12c50b6cc20

Observation f806c2e4-c365-4690-8c6b-c00a10868aba · outbound

This paper cites Calibrated learning and correlated equilibrium.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Calibrated learning and correlated equilibrium

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.532269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.532269Z digest=sha256:2f546b7455fea94ad31fff18d36e6adc33327c06e762b26ea672cc17a9d530d9

Observation 4f2752aa-76e0-415e-b201-1bd62ed557e8 · outbound

This paper cites A decision-theoretic generalization of on-line learning and an application to boosting.Journal of Computer and System Sciences, 55(1):119–139, 1997.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A decision-theoretic generalization of on-line learning and an application to boosting.Journal of Computer and System Sciences, 55(1):119–139, 1997

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.536357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.536357Z digest=sha256:0f3c3d6365e73c51904f9758eaca69cba2ec616bec3cdf8db9f46dd0c50f7c3e

Observation b0771398-efea-48b1-b9d7-531e99b07939 · outbound

This paper cites Schapire, Yoram Singer, and Manfred K.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Schapire, Yoram Singer, and Manfred K

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.540211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.540211Z digest=sha256:7c799b604898ff6372427f6ec6b15a1507cc8ee03e9d522e90e217095a36b698

Observation 734e2856-1e38-410c-a609-1098d2f6a77a · outbound

This paper cites The rationality of quali- fied lotteries.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The rationality of quali- fied lotteries

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.544224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.544224Z digest=sha256:d34929045e026f6affa64057b10b423717bc6f4530e506d2e6e801776c0699eb

Observation b6f5f842-092b-46a6-9378-6393f88afd31 · outbound

This paper cites MIT press, 1998.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis MIT press, 1998

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.548657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.548657Z digest=sha256:9061202b375189eb1f7bbaddf8e986cd8d5a1ede8fce34baf1200d12fa9e8c67

Observation de4b3e6e-966d-4059-ad85-ee3b97f48593 · outbound

This paper cites Artificial intelligence, values, and alignment.Minds and Machines, 30(3):411–437, 2020.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Artificial intelligence, values, and alignment.Minds and Machines, 30(3):411–437, 2020

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.552194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.552194Z digest=sha256:6921faecc15b5c8683bf9db926d97bccbdd6d737770fa2eb0ceff8c4b16192d5

Observation 1f28048b-f2f6-4fb3-b9a2-ec7be43d61dd · outbound

This paper cites Bayesian reinforcement learning: A survey.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Bayesian reinforcement learning: A survey

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.556143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.556143Z digest=sha256:ae94fc88ce66de51437886523730946030882a58090c5969685ef1088aff04bd

Observation 7f2751a5-8bef-451a-84d0-5efda07364c3 · outbound

This paper cites Graph embedding techniques, applica- tions, and performance: A survey.Knowledge-Based Systems, 151:78–94, 2018.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Graph embedding techniques, applica- tions, and performance: A survey.Knowledge-Based Systems, 151:78–94, 2018

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.560057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.560057Z digest=sha256:95c7fbb5ef4f2b60d222f896360e81aa723f64e97b6c77bafbde3eaba81ffc57

Observation 3b7274ae-7134-4432-adfb-6fbd1d2749ff · outbound

This paper cites Inconsistency of Bayesian in- ference for misspecified linear models, and a proposal for repairing it.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Inconsistency of Bayesian in- ference for misspecified linear models, and a proposal for repairing it

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.563841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.563841Z digest=sha256:fa8d39ab0906814f1aed0208f30bd1d9c374d452606f5ef36ea43d022eaa3631

Observation 32d209c2-436b-48a8-bb73-4da9b1458af9 · outbound

This paper cites The off-switch game.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The off-switch game

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.568217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.568217Z digest=sha256:e153da7e35be436f1ccdeb72ce1903493621d50452a47bee6c09c01d326c131a

Observation 4ab5d993-fd02-4ff2-bef8-71e847ea15de · outbound

This paper cites Cooperative inverse reinforcement learning.Advances in Neural Informa- tion Processing Systems, 29, 2016.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Cooperative inverse reinforcement learning.Advances in Neural Informa- tion Processing Systems, 29, 2016

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.572056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.572056Z digest=sha256:d976fb263adebe349e5997574010ad3c1d8e9d0e4c79507a36b8eec616df8e4c

Observation 3de8e233-2243-4902-9b2b-afabaf3aae6b · outbound

This paper cites The tragedy of the commons.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The tragedy of the commons

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.576001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.576001Z digest=sha256:8aff589fb098503492d78006aeb6d994a0141cb17c61ce4a2be66c49f836e666

Observation d0089685-5f65-4ff5-ac89-3132df8e7545 · outbound

This paper cites Completeness in the theory of types.Journal of Symbolic Logic, 15(2):81–91, 1950.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Completeness in the theory of types.Journal of Symbolic Logic, 15(2):81–91, 1950

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.579850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.579850Z digest=sha256:78d3d3b9554dfca1f6b071c4d81298cefafc80f6b74c2e11b7eaffcb0ef93765

Observation 241c5f4b-ba30-4405-8efa-f41aa3b0a07f · outbound

This paper cites Tracking the best expert.Ma- chine learning, 32(2):151–178, 1998.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Tracking the best expert.Ma- chine learning, 32(2):151–178, 1998

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.583839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.583839Z digest=sha256:8ac3f57881b57beaf63c25cee043ec20b1fa74d1d1fcfba5f92d2992411b8df6

Observation 1c3e2999-b6e2-4642-be3a-2296dc69c0cc · outbound

This paper cites The many faces of exponentialweightsinonlinelearning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The many faces of exponentialweightsinonlinelearning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.587304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.587304Z digest=sha256:b924abe4f254ca2c04f01ae4553ce95570a4333168df839e9d152419fc2a6a8b

Observation ae6fca03-0fa4-45bf-8a8b-86ed41623d39 · outbound

This paper cites The exponential mechanism for social welfare: Private, truthful, and nearly optimal.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The exponential mechanism for social welfare: Private, truthful, and nearly optimal

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.591019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.591019Z digest=sha256:b277377f82e1739ad57d692636ea2e6dd29124789611e6081e40ec8caa2d1830

Observation e41b946f-d2a9-4f18-aec3-8b8ba6e923ca · outbound

This paper cites Nash Incentive-compatible Online Mechanism Learning via Weakly Differentially Private Online Learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Nash Incentive-compatible Online Mechanism Learning via Weakly Differentially Private Online Learning

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:56:01.211238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.594642Z digest=sha256:d2381ac89b53202cb3d9b581f6c2f40014ed264d4f0165f00843c3d27dfe9c8d

Observation d1af85ec-7077-43b8-a0eb-af19d41d7921 · outbound

This paper cites Universal Artificial Intelligence: Sequential Decisions based on Algorithmic Probability.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Universal Artificial Intelligence: Sequential Decisions based on Algorithmic Probability

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.598968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.598968Z digest=sha256:2e6ecc4fc47e6d3a0b0b6cc7dab4696423494a200dfbe0190747ef7d91c75b35

Observation b2f60829-63bc-40e4-b756-836d44a65cc0 · outbound

This paper cites Feature reinforcement learning: Part I.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Feature reinforcement learning: Part I

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.602602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.602602Z digest=sha256:c0acfde55efa6bf11713984efa6ab84e77d38cd8dc81e3101dcff51df71723d5

Observation f2a59516-a8d6-4b10-bfa8-a10c41b4b133 · outbound

This paper cites CRC Press, 2024.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis CRC Press, 2024

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.179729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.606144Z digest=sha256:63df0b40c30b1c16168905e3568288358c8b199d2dfbec2fa00ff34765c67aba

Observation 593c4a24-bc2d-4fd1-b98c-b2ebc05da740 · outbound

This paper cites Principal-Agent Reinforcement Learning: Orchestrating AI Agents with Contracts.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Principal-Agent Reinforcement Learning: Orchestrating AI Agents with Contracts

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.609953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.609953Z digest=sha256:d8ea5ab43613c12eccf3fd7710caf62a98443f6a76c1a2574bb2a194555add25

Observation dd15f004-c4c4-4f94-85d8-6503d780457d · outbound

This paper cites Reward-free exploration for reinforcement learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Reward-free exploration for reinforcement learning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.167275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.614314Z digest=sha256:9db6d3f5a2144dc323cd9701889bf5d2d63145f4b1e6a5017c5e1a195472a97e

Observation 542e3856-66cc-4f4d-82e6-50489218de5c · outbound

This paper cites Provably efficient reinforcement learning with linear function approximation.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Provably efficient reinforcement learning with linear function approximation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.618166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.618166Z digest=sha256:2a9a6c46a1da83b5fc515b251a670fb05982b8849470f3f5d36160d2ed7767ed

Observation 1e1afb36-a4b4-488e-a518-f999dd481eaa · outbound

This paper cites Rational learning leads to Nash equilib- rium.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Rational learning leads to Nash equilib- rium

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.147664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.621863Z digest=sha256:ce0c392f05fea8879274f9b020e0e296effe7fa4031359af72ad4400a08b6852

Observation 3ced9075-2192-4b19-8c01-e9b52668bcca · outbound

This paper cites Gonzalez, Michael I.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Gonzalez, Michael I

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.136678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.625666Z digest=sha256:ade3f02f4a81f5743c300cfb6690281a4ed2075521c6fa65c30335a5052a7eae

Observation a88ebf50-8c57-418b-8b0a-3a91b5bc9d8b · outbound

This paper cites A survey of reinforcement learning from human feedback.arXiv:2312.14925, 10, 2023.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A survey of reinforcement learning from human feedback.arXiv:2312.14925, 10, 2023

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.629715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.629715Z digest=sha256:88b136d74db0fdc2972b415b6688aae17301b6ca46804399b256743076acfceb

Observation f0429409-8cec-4ad5-9f34-5312f0368b9f · outbound

This paper cites Mech- anism design in large games: Incentives and privacy.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Mech- anism design in large games: Incentives and privacy

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.125025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.633577Z digest=sha256:a28c8ce6cb9142454e1d36ab7ad3ae83f619d70fa5735604c22a083954898f90

Observation 6057d7ed-cf49-4855-9040-50dfabe934bf · outbound

This paper cites The rise of the platform economy.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis The rise of the platform economy

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.112429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.637394Z digest=sha256:060847a3c03bce5ef670a8fcbc3a28a74b8fc84cbfa1a3e68252263ba7f4aff8

Observation 40ec5578-7771-41e2-bb94-0d228b314f3d · outbound

This paper cites Bellman goes relational.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Bellman goes relational

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.099661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.641048Z digest=sha256:757b7056571820da64d1c9669669fd98eb46bf1303958dde4ba0105f6863d41a

Observation f534922c-bb39-4664-9de3-d49fe285958d · outbound

This paper cites Bandit based monte-carlo planning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Bandit based monte-carlo planning

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.086990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.644619Z digest=sha256:2ed0782996a6188b95a338f827f6af8a6c8ecba84268605624074d61f694e0ad

Observation 26c6102a-c9ae-43d8-8244-5ddc4813b053 · outbound

This paper cites Paying to do better: Games with payments between learning agents.arXiv:2405.20880, 2024.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Paying to do better: Games with payments between learning agents.arXiv:2405.20880, 2024

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.648645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.648645Z digest=sha256:8dc6c545cd6e78139e1233c80c80f24b708f31baaa578f67392ff81181951154

Observation e276eed5-bc61-4161-b383-cbbf2093d810 · outbound

This paper cites Universal codes from switching strategies.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Universal codes from switching strategies

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.073826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.652458Z digest=sha256:765799c6eb3be0c4580b74e994d6ca0830ee5b9ea28f201d2d5faa3b51072e20

Observation 456e641f-cee3-4c73-99f1-159d7fbdbec3 · outbound

This paper cites Krichevsky and V.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Krichevsky and V

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.061464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.656358Z digest=sha256:61b0881d6f9cb8ca795e5af70fe700440831828a54b2d2a2f77a7580d259a48d

Observation cb36b20f-8d48-4205-bf24-ad770218ee0a · outbound

This paper cites A unified game-theoretic approach to multiagent reinforcement learning.Advances in Neural Information Processing Systems, 30, 2017.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A unified game-theoretic approach to multiagent reinforcement learning.Advances in Neural Information Processing Systems, 30, 2017

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.048431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.660262Z digest=sha256:3966e9dbf20eab8a40a8c59f47cba0d2fa1b84c086b46fa70ee817fd0b033dcf

Observation c8c05919-8918-4a19-a339-a7a62b1035c1 · outbound

This paper cites Bandit Algorithms.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Bandit Algorithms

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T23:56:00.663954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:56:00.663954Z digest=sha256:8054e9b6f0a7403e1701a12e8dc4a9c133a6f796103270bdb5ad8044a2eecf5a

Observation 25a13960-66f4-4ec6-bcb6-736074a9099c · outbound

This paper cites Thomp- son sampling is asymptotically optimal in general environments.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Thomp- son sampling is asymptotically optimal in general environments

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.024580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.667799Z digest=sha256:9b0ce273ffc54b0d0f1e306e214e746ef0022cd478486aef408482e8c273289e

Observation 8345c7a4-cec9-4623-9c03-e25ed384c392 · outbound

This paper cites A formal solution to the grain of truth problem.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis A formal solution to the grain of truth problem

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:02.011678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.671607Z digest=sha256:7fc23e266b7cd35326ccede260afa0523f3cdf699c33d328f9a916570e3ec9b1

Observation e428798b-4a9b-4b95-87b0-bfaf050d9f62 · outbound

This paper cites Walsh, and Michael L.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Walsh, and Michael L

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.999348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.675744Z digest=sha256:088026a59186b2f4d0a75422e6124f09d686bfe3b214dc33e495362900524c58

Observation 1df2304d-fb51-411e-86a1-775d5095c3ed · outbound

This paper cites An Introduction to Kolmogorov Complexity and Its Applications.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis An Introduction to Kolmogorov Complexity and Its Applications

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.986452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.679584Z digest=sha256:9deb63b0e4733dd7b4708cace7877e743ea36c9ae8d7c9568d5ab5fc71d18549

Observation e20590f9-ba97-4875-880b-7294ab2241d2 · outbound

This paper cites Markov games as a framework for multi-agent rein- forcement learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Markov games as a framework for multi-agent rein- forcement learning

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.975202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.683289Z digest=sha256:4464e8ffd22558c343d0b79a5e107231dbef11c019afe3f59a2d2368c911683a

Observation 181ee530-d9e1-4ce0-990a-21317afb815d · outbound

This paper cites an unresolved cited work.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:56:01.963992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.687209Z digest=sha256:2da39cc7ab2a49879f1c22af3c56e8902c3cab3a28acc3509a1648ae126202b8

Observation a8a56564-3a59-4409-9d80-1c74351f735b · outbound

This paper cites Lloyd and Kee Siong Ng.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Lloyd and Kee Siong Ng

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.952311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.691113Z digest=sha256:71b3d07dd5ac4230c5b26096988207e5bd3e9304ea070d42c52768bf87e1460a

Observation 13d86ac0-2ee9-4b2b-9153-1e27c224aa40 · outbound

This paper cites Pes- simism meets VCG: Learning dynamic mechanism design via offline rein- forcement learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Pes- simism meets VCG: Learning dynamic mechanism design via offline rein- forcement learning

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.940697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.694562Z digest=sha256:9fd595f9d5ab3ab631be3c74897a3f6c8003f40632c7acc7b17ce20da1f4a2cb

Observation 2eb78945-7ed2-4e78-b17c-88f2a9763c59 · outbound

This paper cites Mechanism design via differential privacy.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Mechanism design via differential privacy

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.928234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.698198Z digest=sha256:781f253b788cf6805902ef5f8f3fe0792bbae7ce3b709a10de51faf7c4800699

Observation b4fe5859-b4c2-4c27-a59f-862ef65d6c35 · outbound

This paper cites Kenneth Arrow’s last theorem.The Journal of Mechanism and Institution Design, 9(1):7–11, 2024.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Kenneth Arrow’s last theorem.The Journal of Mechanism and Institution Design, 9(1):7–11, 2024

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.916339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.702062Z digest=sha256:9e163ac2c636d35fd7809073567131b13720302c0707a2f77d6914ed91e32c6b

Observation 8b8cd9ba-76a8-411e-92a0-51f5eb0b23bd · outbound

This paper cites Human-level control through deep re- inforcement learning.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Human-level control through deep re- inforcement learning

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.904186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.705819Z digest=sha256:14dd0e00b296c57992a94c938f9e499660cb587fec6647f2dd294f3eb173ebec

Observation 60fd1852-a849-474e-9128-7c669a8c1986 · outbound

This paper cites Efficient tracking of a growing number of experts.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Efficient tracking of a growing number of experts

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.891161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.709555Z digest=sha256:0fadfd138ba71da670658cc7fdb6a6a6ca1566eb068e19135e2685967883123b

Observation 534956a3-9058-4f25-8a44-63ab8767111a · outbound

This paper cites Exploratory engineering in artificial intelligence.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Exploratory engineering in artificial intelligence

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.877411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.713334Z digest=sha256:4001926c723db1735abc563d779da7d77591a4ccd3113d293ab71db3e1088b9a

Observation 4c834a32-4dce-456e-8f78-af4041217695 · outbound

This paper cites Lloyd, and William Uther.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Lloyd, and William Uther

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.863052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.716871Z digest=sha256:85454debaeccfdfb4c2cc10bb6440ab92d40062e7e45f2b9ef64d0b2df21d866

Observation 1dfc2b30-6ea0-42a7-8fcc-37bfe1605a43 · outbound

This paper cites Feature re- inforcement learning in practice.

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Feature re- inforcement learning in practice

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.849385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.720776Z digest=sha256:344fe2d56e4281c43063a39ad62c9d7f9da8653a405cb4a162e3d29d7241872f

Observation 477810e9-3361-4bc2-a97e-0daadef4a59d · outbound

This paper cites Introduction to mechanism design (for computer scientist).

The Problem of Social Cost in Multi-Agent General Reinforcement Learning: Survey and Synthesis Introduction to mechanism design (for computer scientist)

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:56:01.836606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:56:00.724470Z digest=sha256:4d6ab676dad39c128b2287de08718093a324420342a786fa8be4e204a21ad1c4

Pith citing papers

No inbound Pith citation observations are available.