Introduces SAKE benchmark with 2154 questions to assess LLMs on software architectural knowledge, showing high overall accuracy but marked gaps across categories.
In: Software Architecture: 19th European Conference, ECSA 2025, Proceedings
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2026 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
SAKE: Software Architectural Knowledge Evaluation Benchmark for Large Language Models
Introduces SAKE benchmark with 2154 questions to assess LLMs on software architectural knowledge, showing high overall accuracy but marked gaps across categories.