Jul 2026
IMO-CoT: a benchmark from International Mathematics Olympiads for evaluating chain-of-thought reasoning in large language models
IMO-CoT is introduced, a novel, selective, information-rich benchmark derived from International Mathematics Olympiad (IMO) problems, designed to evaluate CoT reasoning capabilities in LLMs.
Anurag Dutta, A. Ramamoorthy, M. G. Lakshmi et al.
· Iran Journal of Computer Sci... · 0 citations