CodeQwen1.5-7B — benchmark results
Provider: Alibaba. Released 2024-04-16. Access: Open.
Unified ELO 1254 ± 20, rank #1608 of 1776 rated models, from 11 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| ShaderMatch | 16.75 | Clone Match Rate (%) | 90.5 |
| BigCode Models Leaderboard | 50.8 | HumanEval Python Pass@1 (%) | 60.2 |
| Big Code Memorization - HumanEval pass@50 | 38.84 | HumanEval pass@50 (%) | 58.8 |
| Big Code Memorization - HumanEval-ET pass@50 | 32.79 | HumanEval-ET pass@50 (%) | 58.8 |
| EvalPlus (HumanEval+ & MBPP+) | 53.2 | Pass@1 avg (%) | 50 |
| GSM8K | 37.7 | Accuracy (%) | 42.6 |
| Big Code Memorization - HumanEval pass@1 | 43.9 | HumanEval pass@1 (%) | 41.2 |
| Big Code Memorization - HumanEval-ET pass@1 | 38.41 | HumanEval-ET pass@1 (%) | 41.2 |
| MMLU | 40.5 | Accuracy (%) | 12.6 |
| ARC Challenge (AI2) | 35.7 | Accuracy (%) | 10.3 |
| WinoGrande | 59.8 | Accuracy (%) | 10 |
Interactive version: theaggregate.ai/model?slug=codeqwen1-5-7b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.