Tsunami-1.0-14B-Instruct: benchmark results
Provider: Other. Access: Open.
Unified ELO 1574 ± 1, rank #344 of 1392 rated models, from 33 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - IFEval | 78.29 | Score | 94.7 |
| Open LLM Leaderboard - MATH Level 5 | 45.85 | Score | 94 |
| Thai LLM MC - thaiexam_qa | 61.06 | Accuracy (%) | 92.8 |
| Open LLM Leaderboard - MMLU-Pro | 47.21 | Score | 91.1 |
| Thai LLM MC - m3exam_tha_seacrowd_qa | 63.05 | Accuracy (%) | 90.6 |
| Open LLM Leaderboard - BBH | 49.15 | Score | 89.9 |
| Open LLM Leaderboard - GPQA | 14.21 | Score | 89.8 |
| Thai LLM NLU - xnli.tha_seacrowd_pairs | 42.57 | Accuracy (%) | 88.4 |
| Thai LLM NLG - flores200_tha_Thai_eng_Latn_seacrowd_t2t (chrF++) | 59.96 | chrF++ | 88.1 |
| Open LLM Leaderboard - MuSR | 16.34 | Score | 85.5 |
| Thai LLM NLU - wisesight_thai_sentiment_seacrowd_text | 51.18 | Accuracy (%) | 85.5 |
| Thai LLM - Math | 6.45 | Rating (0-10) | 79.6 |
Interactive version: theaggregate.ai/model?slug=tsunami-1-0-14b-instruct · How It Works · Data refreshed daily, snapshot 2026-09-05.