A multivocal literature review finds ChatGPT's reported error rates range from single digits to over 80 percent depending on domain and task, yet its synthesized ranges are not backed by a released dataset.
Software Engineering Research Using Multivocal Literature Reviews,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Why you shouldn't fully trust ChatGPT: A synthesis of this AI tool's error rates across disciplines and the software engineering lifecycle
A multivocal literature review finds ChatGPT's reported error rates range from single digits to over 80 percent depending on domain and task, yet its synthesized ranges are not backed by a released dataset.