A benchmark of 29 LLMs and four prompting strategies for classifying faulty computer components from user reports, finding that small models like gemma-2-2b match larger ones.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Evaluating LLMs and Prompting Strategies for Automated Hardware Diagnosis from Textual User-Reports
A benchmark of 29 LLMs and four prompting strategies for classifying faulty computer components from user reports, finding that small models like gemma-2-2b match larger ones.