MM-Eval is a new Modern Mongolian benchmark showing LLMs perform best at syntax, worse at semantics and knowledge, and worst at reasoning.
In Findings of the Association for Computational Linguistics, ACL 2024, Bangkok, Thailand and virtual meeting, August 11-16, 2024 , pages 15670–15693
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
MM-Eval: A Hierarchical Benchmark for Modern Mongolian Evaluation in LLMs
MM-Eval is a new Modern Mongolian benchmark showing LLMs perform best at syntax, worse at semantics and knowledge, and worst at reasoning.