Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.SE 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Evaluating Large Language Models for Code Review

cs.SE · 2025-05-26 · conditional · novelty 6.0

GPT-4o and Gemini 2.0 Flash correctly judged code correctness in about 64-68% of cases and corrected 54-68% of faulty code, with better results when given problem descriptions.

citing papers explorer

Showing 1 of 1 citing paper.

  • Evaluating Large Language Models for Code Review cs.SE · 2025-05-26 · conditional · none · ref 5

    GPT-4o and Gemini 2.0 Flash correctly judged code correctness in about 64-68% of cases and corrected 54-68% of faulty code, with better results when given problem descriptions.