REVIEW 1 cited by
The Importance of Credo in Multiagent Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We propose a model for multi-objective optimization, a credo, for agents in a system that are configured into multiple groups (i.e., teams). Our model of credo regulates how agents optimize their behavior for the groups they belong to. We evaluate credo in the context of challenging social dilemmas with reinforcement learning agents. Our results indicate that the interests of teammates, or the entire system, are not required to be fully aligned for achieving globally beneficial outcomes. We identify two scenarios without full common interest that achieve high equality and significantly higher mean population rewards compared to when the interests of all agents are aligned.
Forward citations
Cited by 1 Pith paper
-
Modeling human reputation-seeking behavior in a spatio-temporally complex public good provision game
A reputation-motivated multi-agent RL model reproduces human groups' cooperation under identifiability and its collapse under anonymity in the Clean Up public goods game.
Discussion (0). Sign in to comment.