DEONTICBENCH is a new benchmark of 6,232 deontic reasoning tasks from U.S. legal domains where frontier LLMs reach only ~45% accuracy and symbolic Prolog assistance plus RL training still fail to solve tasks reliably.
Connecting Symbolic Statutory Reasoning with Legal Information Extraction
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
fields
cs.CL 2years
2026 2verdicts
UNVERDICTED 2representative citing papers
DAR lets LLMs interact dynamically with statutes for deontic reasoning, improving results on hard DeonticBench subsets but with uneven gains and higher token use for weaker models.
citing papers explorer
-
DeonticBench: A Benchmark for Reasoning over Rules
DEONTICBENCH is a new benchmark of 6,232 deontic reasoning tasks from U.S. legal domains where frontier LLMs reach only ~45% accuracy and symbolic Prolog assistance plus RL training still fail to solve tasks reliably.
-
DAR: Deontic Reasoning with Agentic Harnesses
DAR lets LLMs interact dynamically with statutes for deontic reasoning, improving results on hard DeonticBench subsets but with uneven gains and higher token use for weaker models.