Sailor2-20B-Chat: benchmark results
Provider: Sea AI Lab. Released 2025-02-18. Access: Open.
Unified ELO 1561 ± 1, rank #443 of 1553 rated models, from 29 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Thai LLM - Coding | 7.82 | Rating (0-10) | 93.7 |
| Thai LLM NLU - xcopa_tha_seacrowd_qa | 92.8 | Accuracy (%) | 92.8 |
| Thai LLM - Social Science | 8.4 | Rating (0-10) | 85.2 |
| Thai LLM - Roleplay | 7.6 | Rating (0-10) | 83.8 |
| Thai LLM - Math | 6.7 | Rating (0-10) | 83.1 |
| Thai LLM - STEM | 7 | Rating (0-10) | 83.1 |
| Thai LLM MC - m3exam_tha_seacrowd_qa | 60.89 | Accuracy (%) | 82.6 |
| Thai LLM NLU - belebele_tha_thai_seacrowd_qa | 85.78 | Accuracy (%) | 82.6 |
| Thai LLM - Writing | 7.2 | Rating (0-10) | 82.4 |
| SEA LLM Leaderboard - SeaExam | 73.1 | Private Average Score (%) | 82.2 |
| Thai LLM MC - thaiexam_qa | 58.05 | Accuracy (%) | 78.3 |
| Thai LLM - Knowledge III | 4.45 | Rating (0-10) | 71.8 |
Interactive version: theaggregate.ai/model?slug=sailor2-20b-chat · How It Works · Data refreshed daily, snapshot 2026-09-08.