Pith. sign in

Generative ai in mafia-like game simulation,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.AI 1

years

2024 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Codenames as a Benchmark for Large Language Models

cs.AI · 2024-12-16 · conditional · novelty 6.0

In a full-rule Codenames benchmark, nine LLMs show distinct play styles, generalize better across teammates than word-vector agents, and still lose to self-matched word-vector agents on raw score.

citing papers explorer

Showing 1 of 1 citing paper.

  • Codenames as a Benchmark for Large Language Models cs.AI · 2024-12-16 · conditional · none · ref 7

    In a full-rule Codenames benchmark, nine LLMs show distinct play styles, generalize better across teammates than word-vector agents, and still lose to self-matched word-vector agents on raw score.