Tsunami-1.0-7B-Instruct: benchmark results
Provider: Other. Access: Open.
Unified ELO 1533 ± 1, rank #524 of 1392 rated models, from 33 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MATH Level 5 | 43.35 | Score | 92.4 |
| Open LLM Leaderboard - IFEval | 73.09 | Score | 87.3 |
| Open LLM Leaderboard - MuSR | 15.76 | Score | 83.8 |
| Thai LLM NLU - xnli.tha_seacrowd_pairs | 41.04 | Accuracy (%) | 82.6 |
| Open LLM Leaderboard - MMLU-Pro | 38.04 | Score | 82 |
| Open LLM Leaderboard - BBH | 35.86 | Score | 72.9 |
| Thai LLM NLG - xl_sum_tha_seacrowd_t2t (ROUGE2) | 10.14 | ROUGE2 | 71.6 |
| Thai LLM NLG - xl_sum_tha_seacrowd_t2t (ROUGE1) | 29.1 | ROUGE1 | 70.1 |
| Open LLM Leaderboard - GPQA | 8.39 | Score | 70 |
| Thai LLM NLU - wisesight_thai_sentiment_seacrowd_text | 47.47 | Accuracy (%) | 69.6 |
| Thai LLM NLG - xl_sum_tha_seacrowd_t2t (ROUGEL) | 19.44 | ROUGEL | 68.7 |
| Thai LLM NLU - xcopa_tha_seacrowd_qa | 88.4 | Accuracy (%) | 67.4 |
Interactive version: theaggregate.ai/model?slug=tsunami-1-0-7b-instruct · How It Works · Data refreshed daily, snapshot 2026-09-05.