Programmers with LLM access solved fewer formally verified debugging tasks than a no-AI control group, though complete novices and strong language experts gained some benefit.
Next Steps in LLM-Supported Java Verification
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Recent work has shown that Large Language Models (LLMs) are not only a suitable tool for code generation but also capable of generating annotation-based code specifications. Scaling these methodologies may allow us to deduce provable correctness guarantees for large-scale software systems. In comparison to other LLM tasks, the application field of deductive verification has the notable advantage of providing a rigorous toolset to check LLM-generated solutions. This short paper provides early results on how this rigorous toolset can be used to reliably elicit correct specification annotations from an unreliable LLM oracle.
citation-role summary
citation-polarity summary
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Do AI models help produce verified bug fixes?
Programmers with LLM access solved fewer formally verified debugging tasks than a no-AI control group, though complete novices and strong language experts gained some benefit.