GPT-4-generated unit tests matched manual tests on coverage and mutation scores but used less precise boundary values and needed human supervision.
IEEE Transactions on Software EngineeringSE-13(12), 1278–1296 (1987)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Evaluating Large Language Models for the Generation of Unit Tests with Equivalence Partitions and Boundary Values
GPT-4-generated unit tests matched manual tests on coverage and mutation scores but used less precise boundary values and needed human supervision.