ruGPT3-large: benchmark results
Provider: Other. Access: Open.
Unified ELO 1164 ± 39, rank #1484 of 1537 rated models, from 31 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| MERA Original Paper - ruModAr | 0.1 | Exact Match (%; greedy generation) | 58.3 |
| MERA Original Paper - SimpleAr | 0.4 | Exact Match (%; 5-shot, greedy generation) | 50 |
| MERA - LCS | 12 | Accuracy (%) | 44.7 |
| MERA Original Paper - ruMultiAr | 0.7 | Exact Match (%; 5-shot, greedy generation) | 44.4 |
| MERA Original Paper - ruMMLU | 24.5 | Accuracy (%; 5-shot, log-likelihood) | 33.3 |
| MERA Original Paper - MathLogicQA | 25.1 | Accuracy (%; 5-shot, log-likelihood) | 27.8 |
| MERA - USE | 8.73 | Grade, normalized (%) | 25.8 |
| MERA Original Paper | 19.3 | Total Score (%; mean of 17 problem-solving and exam tasks, M | 19.4 |
| MERA - ruHateSpeech | 54.34 | Accuracy (%) | 13.2 |
| MERA - RWSD | 46.92 | Accuracy (%) | 12.4 |
| MERA Original Paper - ruWorldTree | 23.2 | Accuracy (%; 5-shot, log-likelihood) | 11.1 |
| MERA - MultiQ | 13.55 | F1 (%) | 11 |
Interactive version: theaggregate.ai/model?slug=rugpt3-large · How It Works · Data refreshed daily, snapshot 2026-09-25.