ELABORATION provides a four-stage human-feedback taxonomy and an 8,320-problem dataset, with experiments showing human-LLM collaboration improves pass@1 by about 7 percent.
Code Error Identifications:
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
ELABORATION: A Comprehensive Benchmark on Human-LLM Competitive Programming
ELABORATION provides a four-stage human-feedback taxonomy and an 8,320-problem dataset, with experiments showing human-LLM collaboration improves pass@1 by about 7 percent.