HyperCLOVA X SEED Think (32B) — benchmark results
Naver's open 32B Korean-centric vision-language reasoning model. Provider: Naver. Released 2025-04-24. Access: API.
Unified ELO 1588 ± 37, rank #508 of 1776 rated models, from 46 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA TAU-2 Bench | 87.43 | Accuracy (%) | 81.4 |
| AA LiveCodeBench | 62.86 | Pass@1 (%) | 69 |
| AA Omniscience - Software Engineering (SWE) - HTML | 38 | Accuracy (%) | 65.3 |
| AA MMLU-Pro | 78.49 | Accuracy (%) | 60.8 |
| AA AIME 2025 | 59 | Accuracy (%) | 56.1 |
| Artificial Analysis Intelligence Index | 17.02 | Intelligence Index | 53.1 |
| AA Terminal-Bench Hard | 12.12 | Accuracy (%) | 48 |
| AA Omniscience - Software Engineering (SWE) - R | 10 | Accuracy (%) | 45.4 |
| AA GPQA Diamond | 61.52 | Accuracy (%) | 40.3 |
| AA Humanity's Last Exam | 5.47 | Accuracy (%) | 40.2 |
| AA Omniscience - Humanities & Social Sciences | 17.2 | Accuracy (%) | 39 |
| AA Omniscience - Software Engineering (SWE) - Kotlin | 14 | Accuracy (%) | 38.6 |
Interactive version: theaggregate.ai/model?slug=hyperclova-x-seed-think-32b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.