IMO-CoT: a benchmark from International Mathematics Olympiads for evaluating chain-of-thought reasoning in large language models Anurag Dutta A. Ramamoorthy M. G. Lakshmi Pijush Kanti Kumar Jul 2026 · Iran Journal of Computer Science · Vol 9 · 0 citations · 39 references Computer Science DOI Semantic Scholar Save View source Cite { copied='apa'; setTimeout(() => { copied=null; open=false }, 1000) })" class="flex w-full items-center justify-between rounded-lg px-3 py-2 text-left text-sm hover:bg-gray-100 dark:hover:bg-ink-800"> Copy APA Copied ✓ { copied='mla'; setTimeout(() => { copied=null; open=false }, 1000) })" class="flex w-full items-center justify-between rounded-lg px-3 py-2 text-left text-sm hover:bg-gray-100 dark:hover:bg-ink-800"> Copy MLA Copied ✓ { copied='bibtex'; setTimeout(() => { copied=null; open=false }, 1000) })" class="flex w-full items-center justify-between rounded-lg px-3 py-2 text-left text-sm hover:bg-gray-100 dark:hover:bg-ink-800"> Copy BibTeX Copied ✓ Share