Pith. sign in

REVIEW 32 cited by

Open Problems in Cooperative AI

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2012.08630 v1 pith:KJKGTU3D submitted 2020-12-15 cs.AI cs.MA

classification cs.AIcs.MA
keywords cooperationproblemsagentscooperativeresearchsocialareasartificial
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Problems of cooperation--in which agents seek ways to jointly improve their welfare--are ubiquitous and important. They can be found at scales ranging from our daily routines--such as driving on highways, scheduling meetings, and working collaboratively--to our global challenges--such as peace, commerce, and pandemic preparedness. Arguably, the success of the human species is rooted in our ability to cooperate. Since machines powered by artificial intelligence are playing an ever greater role in our lives, it will be important to equip them with the capabilities necessary to cooperate and to foster cooperation. We see an opportunity for the field of artificial intelligence to explicitly focus effort on this class of problems, which we term Cooperative AI. The objective of this research would be to study the many aspects of the problems of cooperation and to innovate in AI to contribute to solving these problems. Central goals include building machine agents with the capabilities needed for cooperation, building tools to foster cooperation in populations of (machine and/or human) agents, and otherwise conducting AI research for insight relevant to problems of cooperation. This research integrates ongoing work on multi-agent systems, game theory and social choice, human-machine interaction and alignment, natural-language processing, and the construction of social tools and platforms. However, Cooperative AI is not the union of these existing areas, but rather an independent bet about the productivity of specific kinds of conversations that involve these and other areas. We see opportunity to more explicitly focus on the problem of cooperation, to construct unified theory and vocabulary, and to build bridges with adjacent communities working on cooperation, including in the natural, social, and behavioural sciences.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 32 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Multi-Player Discrete-Bidding Games; Determinacy, Equilibria, and Complexity

    cs.GT 2026-07 accept novelty 7.0 of 10

    Under linear tie-breaking, multi-player discrete-bidding games are determined, admit pure Nash equilibria and mean-payoff values, and deciding the winner is already PSPACE-hard for unary reachability.

  2. Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety

    cs.AI 2026-07 conditional novelty 7.0 of 10

    Changing only the consequence-allocation rule in multi-agent AI shifts collective fatality by 22–58 percentage points across seven model populations, with identity salience in rule text causally driving targeted exploitation.

  3. Who Is Really Playing? Strategic Interaction in AI-Guided Populations

    cs.GT 2026-05 unverdicted novelty 7.0 of 10

    A folk theorem for LLMs proves that all feasible and individually rational outcomes can be sustained as ε-equilibria in repeated games where LLMs advise client populations, despite indirect observation.

  4. Verbalized Bayesian Persuasion

    cs.GT 2025-02 conditional novelty 7.0 of 10

    VBP solves Bayesian persuasion in natural language by treating LLMs as sender and receiver in a mediator-augmented game and searching prompt strategies with Prompt-PSRO.

  5. Multi-Agent AI Safety as an Institutional Design Problem

    cs.LG 2026-08 conditional novelty 6.0 of 10

    In synthetic delegation workflows, identical final violation rates hide different mechanisms: prompts prevent prohibited attempts, provenance-aware guards block and recover, and a local policy guard fails when transfo...

  6. Calibrating Artificial Guilt: Neurally Grounded Reward Shaping for Prosocial Multi-Agent Reinforcement Learning

    cs.AI 2026-08 reject novelty 6.0 of 10

    A happiness-regression contrast from the SoDec dataset is used as a reward-shaping weight in a two-agent Social Lottery, yielding a safe rate of 0.459 versus a human 0.484, but the contrast is statistically indistingu...

  7. Strategy, Not Payoffs: A Behavioural Embedding of Normal-Form Games

    cs.GT 2026-07 conditional novelty 6.0 of 10

    A two-feature game embedding (Nash entropy and best-response switching) predicts cross-game transfer of fine-tuned LLMs on held-out games, outperforming game identity and published structural embeddings.

  8. The Agentic Web Requires New Normative Infrastructure

    cs.CY 2026-06 unverdicted novelty 6.0 of 10

    The web's anti-bot regime should be replaced by a framework that presumptively lets user-authorized AI agents act for their principals, requires platforms to disclose access policies, and permits agent blocking only w...

  9. Inferring Hidden Motives: A Utility Bayesian Model of learning the values of others

    q-bio.NC 2025-11 conditional novelty 6.0 of 10

    People update beliefs about others' social preferences as graded, continuous values rather than discrete types, and the best-fitting account uses a seven-parameter utility function estimated from repeated dictator games.

  10. Emergence of Fair Leaders via Mediators in Multi-Agent Reinforcement Learning

    cs.MA 2025-08 conditional novelty 6.0 of 10

    A mediator that dynamically selects leaders in Stackelberg MARL can induce self-interested agents to adopt fair policies, improving fairness of returns.

  11. Network reciprocity turns cheap talk into a force for cooperation

    physics.soc-ph 2025-07 conditional novelty 6.0 of 10

    In spatial populations, conditional cooperators that pay a cognitive cost can act as catalysts that make cheap talk evolutionarily effective.

  12. How large language models judge and influence human cooperation

    physics.soc-ph 2025-06 conditional novelty 6.0 of 10

    LLMs' implicit social norms for judging cooperation vary by model and version, and these differences change predicted long-term cooperation in indirect reciprocity models.

  13. Learning from Active Human Involvement through Proxy Value Propagation

    cs.AI 2025-02 conditional novelty 6.0 of 10

    A reward-free human-in-the-loop RL method that labels human demonstrations with high Q values and intervened agent actions with low Q values, then propagates these values through TD learning to train policies across d...

  14. Governing AI Agents

    cs.AI 2025-01 conditional novelty 6.0 of 10

    Agency law and principal-agent theory can frame the governance problems posed by AI agents and justify new principles of inclusivity, visibility, and liability.

  15. Deterministic Model of Incremental Multi-Agent Boltzmann Q-Learning: Transient Cooperation, Metastability, and Oscillations

    cs.MA 2024-12 conditional novelty 6.0 of 10

    A frequency-aware mean-field model of incremental Boltzmann Q-learning in the Prisoner's Dilemma predicts that apparent stable cooperation is a long metastable transient and that high discount factors induce oscillati...

  16. Emergence of Reputation-Based Cooperation in LLM Agents

    cs.MA 2026-08 conditional novelty 5.0 of 10

    AI agents evolve donation strategies resembling Image Scoring, and the steepness of their discrimination against uncooperative opponents predicts resistance to free-riders.

  17. Steganalysis of Adaptive Covert Collusion in Tool-Using Agent Populations: A Black-Box, Cross-Principal Approach

    cs.CR 2026-08 reject novelty 5.0 of 10

    Proposes an encoding-agnostic, black-box detector for covert agent collusion and a capacity-theoretic frontier showing low-rate channels are undetectable, but all empirical numbers are placeholders pending measurement.

  18. Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives

    cs.MA 2026-07 conditional novelty 5.0 of 10

    LLM prosumers deplete a shared renewable reserve exactly when demand exceeds peak replacement, acting like impatient open-access users even when sustaining the reserve is feasible.

  19. Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems

    cs.AI 2026-03 conditional novelty 5.0 of 10

    In an evolutionary game where trust is reduced monitoring, safe and widely adopted AI is the stable outcome only when punishment for unsafe development exceeds the cost of safety and monitoring is affordable.

  20. Non-coercive extortion in game theory

    cs.GT 2025-07 conditional novelty 5.0 of 10

    An agent can profit by committing to give a co-player an outcome-contingent reward that worsens a target player's equilibrium, and win-win 2x2 games are the most vulnerable.

  21. HKGAI-V1: Towards Regional Sovereign Large Language Model for Hong Kong

    cs.CL 2025-07 reject novelty 5.0 of 10

    A DeepSeek-based model fine-tuned for Hong Kong outperforms general models on Hong Kong benchmarks, but most of those benchmarks are self-authored and unreleased.

  22. A theory of appropriateness with applications to generative artificial intelligence

    cs.AI 2024-12 conditional novelty 5.0 of 10

    A theory that human and AI behavior is guided by context-dependent appropriateness implemented as predictive pattern completion, with norms as conventional sanctioning patterns.

  23. The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem

    cs.AI 2026-04 unverdicted novelty 4.0 of 10

    Dominant control-based AI alignment falls short for potential AGI subjects; a parenting model drawing on Turing's child machines should foster gradual autonomy and cooperative coexistence.

  24. Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions

    cs.CY 2026-02 unverdicted novelty 4.0 of 10

    Current XAI methods for DNNs and LLMs rest on paradoxes and false assumptions that demand a paradigm shift to verification protocols, scientific foundations, context-aware design, and faithful model analysis rather th...

  25. The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

    cs.GT 2025-12 reject novelty 4.0 of 10

    A theory of strategic evolution says multi-level systems of self-reproducing optimizers are stable only under a small-gain condition, and stable AI alignment requires bounding self-modification.

  26. Language Games as the Pathway to Artificial Superhuman Intelligence

    cs.AI 2025-01 conditional novelty 4.0 of 10

    A position paper arguing that open-ended language games with fluid roles, varied rewards, and evolving rules can drive expanded data reproduction and thus a path to artificial superhuman intelligence.

  27. Towards a Theory of AI Personhood

    cs.AI 2025-01 accept novelty 4.0 of 10

    The paper outlines agency, theory of mind, and self-awareness as necessary conditions for AI personhood, reviews inconclusive evidence, and argues that AI personhood would make control-focused alignment ethically problematic.

  28. Agentic LLMs in the Supply Chain: Towards Autonomous Multi-Agent Consensus-Seeking

    cs.AI 2024-11 conditional novelty 4.0 of 10

    LLM-powered agents that negotiate with neighboring echelons reduce bullwhip and costs in a simulated supply chain, but the results rest on single runs and manually tuned prompts.

  29. Towards Transparent Ethical AI: A Roadmap for Trustworthy Robotic Systems

    cs.CY 2025-08 unverdicted novelty 3.0 of 10

    The paper argues transparency is fundamental to trustworthy robotics and proposes a framework connecting technical transparency tools to ethical outcomes such as accountability and informed consent.

  30. An Outlook on the Opportunities and Challenges of Multi-Agent AI Systems

    cs.MA 2025-05 conditional novelty 3.0 of 10

    The paper formalizes multi-agent AI systems and argues, with toy experiments, that they beat single agents only under narrow conditions on task decomposition, data diversity, and feedback.

  31. Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models

    cs.CL 2024-11 reject novelty 3.0 of 10

    A study measures how often GPT-4, Claude, LLaMA, and Gemini agree on PhD-level statistics questions, finding that Claude and GPT-4 produce questions with higher inter-model agreement, but the reliability metric relies...

  32. Modeling human reputation-seeking behavior in a spatio-temporally complex public good provision game

    cs.MA 2025-06 conditional novelty 2.0 of 10

    A reputation-motivated multi-agent RL model reproduces human groups' cooperation under identifiability and its collapse under anonymity in the Clean Up public goods game.

Pith tools