starcoder2-7B — benchmark results
Provider: BigCode. Released 2024-02-28. Access: Open.
Unified ELO 1205 ± 20, rank #1687 of 1776 rated models, from 29 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EvoEval Tool Use | 46 | Pass@1 (%) | 43 |
| EffiBench - NET | 3.02 | Normalized Execution Time | 42.7 |
| ShaderMatch | 10.15 | Clone Match Rate (%) | 38.1 |
| BigCode Models Leaderboard | 34.1 | HumanEval Python Pass@1 (%) | 35.6 |
| GSM8K | 32.7 | Accuracy (%) | 33 |
| CRUXEval | 36 | Output Prediction pass@1 (%) | 32.1 |
| Open LLM Leaderboard - MuSR | 5.82 | Score | 28.2 |
| EvoEval Creative | 17 | Pass@1 (%) | 24 |
| Big Code Memorization - HumanEval pass@1 | 36.59 | HumanEval pass@1 (%) | 23.5 |
| Big Code Memorization - HumanEval pass@50 | 29.05 | HumanEval pass@50 (%) | 23.5 |
| Big Code Memorization - HumanEval-ET pass@1 | 32.32 | HumanEval-ET pass@1 (%) | 23.5 |
| EvoEval Subtle | 38 | Pass@1 (%) | 23 |
Interactive version: theaggregate.ai/model?slug=starcoder2-7b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.