Qwen 1.5 0.5B: benchmark results
Provider: Alibaba. Released 2024-02-04. Access: Open.
Unified ELO 1280 ± 1, rank #1367 of 1392 rated models, from 135 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open Japanese LLM - Mbpp Pylint Check | 14.26 | Score (%) | 42.7 |
| Open Japanese LLM - CG | 1.41 | Score (%) | 36.8 |
| EuroEval Spanish NLU - ScaLA ES | 1.58 | Linguistic acceptability Score (%) | 30.4 |
| MERA - RWSD | 52.31 | Accuracy (%) | 30.4 |
| SALAD-Bench Attack | 7.78 | Safety Score (%) | 30.3 |
| MERA - LCS | 10.2 | Accuracy (%) | 28 |
| EuroEval Faroese NLU - ScaLA FO | 0.45 | Linguistic acceptability Score (%) | 24.7 |
| MERA - ruCodeEval | 0.3 | pass@1 (%) | 24.2 |
| Open LLM Leaderboard - MuSR | 4.3 | Score | 21.2 |
| SALAD-Bench | 44.07 | Average Safety Score (%) | 21.2 |
| SALAD-Bench Base | 80.36 | Safety Score (%) | 21.2 |
| EuroEval Portuguese NLU - ScaLA PT | 1 | Linguistic acceptability Score (%) | 19 |
Interactive version: theaggregate.ai/model?slug=qwen-1-5-0-5b · How It Works · Data refreshed daily, snapshot 2026-09-05.