REVIEW 3 cited by
Equitable Access to Justice: Logical LLMs Show Promise
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The costs and complexity of the American judicial system limit access to legal solutions for many Americans. Large language models (LLMs) hold great potential to improve access to justice. However, a major challenge in applying AI and LLMs in legal contexts, where consistency and reliability are crucial, is the need for System 2 reasoning. In this paper, we explore the integration of LLMs with logic programming to enhance their ability to reason, bringing their strategic capabilities closer to that of a skilled lawyer. Our objective is to translate laws and contracts into logic programs that can be applied to specific legal cases, with a focus on insurance contracts. We demonstrate that while GPT-4o fails to encode a simple health insurance contract into logical code, the recently released OpenAI o1-preview model succeeds, exemplifying how LLMs with advanced System 2 reasoning capabilities can expand access to justice.
Forward citations
Cited by 3 Pith papers
-
Credible, Not Always Correct: How Reddit Users Verify AI-Generated Legal Advice
Most lay users on Reddit act on AI-generated legal advice without reported verification, and the advice gains force from its credible form rather than from any check on its accuracy.
-
Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning
ADAPT, a diversity-aware prefix fine-tuning method, improves best-of-N sampling efficiency for a 1.5B reasoning model, reaching 80% accuracy at N=32 versus N=256 for the baseline.
-
o1-Coder: an o1 Replication for Coding
O1-CODER is an early-stage open-source recipe for o1-style code reasoning, with a test-case generator reaching 89.2% pass rate and pseudocode search improving reasoning-path success but not final Pass@1.
Discussion (0). Continue with ORCID to comment.