Pith. sign in

Human-centered mechanism design with Democratic AI

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Building artificial intelligence (AI) that aligns with human values is an unsolved problem. Here, we developed a human-in-the-loop research pipeline called Democratic AI, in which reinforcement learning is used to design a social mechanism that humans prefer by majority. A large group of humans played an online investment game that involved deciding whether to keep a monetary endowment or to share it with others for collective benefit. Shared revenue was returned to players under two different redistribution mechanisms, one designed by the AI and the other by humans. The AI discovered a mechanism that redressed initial wealth imbalance, sanctioned free riders, and successfully won the majority vote. By optimizing for human preferences, Democratic AI may be a promising method for value-aligned policy innovation.

citation-role summary

background 1

citation-polarity summary

fields

q-bio.NC 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

AI Agent Behavioral Science

q-bio.NC · 2025-06-04 · conditional · novelty 4.0

AI agents should be studied as behavioral entities shaped by context and interaction, not only as trained models.

citing papers explorer

Showing 1 of 1 citing paper.

  • AI Agent Behavioral Science q-bio.NC · 2025-06-04 · conditional · none · ref 78 · internal anchor

    AI agents should be studied as behavioral entities shaped by context and interaction, not only as trained models.