ClassicLogic is an open-source benchmark using four logic puzzles with a hierarchical strategy knowledge base to evaluate three forms of compositional generalization in AI agents.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
ClassicLogic: A Knowledge-Driven Benchmark of Classic Puzzle Games for Evaluating Compositional Generalization
ClassicLogic is an open-source benchmark using four logic puzzles with a hierarchical strategy knowledge base to evaluate three forms of compositional generalization in AI agents.