Across 7,624 ChatGPT code files from the DevGPT dataset, Kruskal-Wallis tests found no statistically significant differences in maintainability, reliability, or security issues among Zero-Shot, Few-Shot, Chain-of-Thought, and Persona prompts.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Do Prompt Patterns Affect Code Quality? A First Empirical Assessment of ChatGPT-Generated Code
Across 7,624 ChatGPT code files from the DevGPT dataset, Kruskal-Wallis tests found no statistically significant differences in maintainability, reliability, or security issues among Zero-Shot, Few-Shot, Chain-of-Thought, and Persona prompts.