ChatGPT and DeepSeek correctly decompose only 8-9% of bug reports zero-shot and 19-23% with few-shot prompting, with over-decomposition as the dominant error.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
An Empirical Study on the Capability of LLMs in Decomposing Bug Reports
ChatGPT and DeepSeek correctly decompose only 8-9% of bug reports zero-shot and 19-23% with few-shot prompting, with over-decomposition as the dominant error.