HyperCLOVA X SEED Think (32B) — benchmark results

Naver's open 32B Korean-centric vision-language reasoning model. Provider: Naver. Released 2025-04-24. Access: API.

Unified ELO 1588 ± 37, rank #508 of 1776 rated models, from 46 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
AA TAU-2 Bench87.43Accuracy (%)81.4
AA LiveCodeBench62.86Pass@1 (%)69
AA Omniscience - Software Engineering (SWE) - HTML38Accuracy (%)65.3
AA MMLU-Pro78.49Accuracy (%)60.8
AA AIME 202559Accuracy (%)56.1
Artificial Analysis Intelligence Index17.02Intelligence Index53.1
AA Terminal-Bench Hard12.12Accuracy (%)48
AA Omniscience - Software Engineering (SWE) - R10Accuracy (%)45.4
AA GPQA Diamond61.52Accuracy (%)40.3
AA Humanity's Last Exam5.47Accuracy (%)40.2
AA Omniscience - Humanities & Social Sciences17.2Accuracy (%)39
AA Omniscience - Software Engineering (SWE) - Kotlin14Accuracy (%)38.6

Interactive version: theaggregate.ai/model?slug=hyperclova-x-seed-think-32b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.