Large new Python/Java benchmarks, diff-based metrics (BLEU-diff, CrystalBLEU-diff, LEMOD), and an LLM evaluation showing ~10% exact-match SATD repayment and larger models winning on fine-grained metrics.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Understanding the Effectiveness of LLMs in Automated Self-Admitted Technical Debt Repayment
Large new Python/Java benchmarks, diff-based metrics (BLEU-diff, CrystalBLEU-diff, LEMOD), and an LLM evaluation showing ~10% exact-match SATD repayment and larger models winning on fine-grained metrics.