LongChat-7B-v1.5-32k — benchmark results
Provider: Other. Released 2023-08-01. Access: Open.
Unified ELO 1237 ± 10, rank #1647 of 1776 rated models, from 72 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| MMLU-by-task - Abstract Algebra | 34 | Accuracy (%) | 82.1 |
| MMLU-by-task - High School Mathematics | 30.74 | Accuracy (%) | 81.5 |
| MMLU-by-task - College Mathematics | 35 | Accuracy (%) | 75.2 |
| MMLU-by-task - Elementary Mathematics | 31.75 | Accuracy (%) | 65.6 |
| LIBRA - ru2WikiMultihopQA | 35.16 | Dataset Total Score (%) | 64.3 |
| MMLU-by-task - Machine Learning | 34.82 | Accuracy (%) | 60.6 |
| LIBRA - ruQasper | 4.97 | Dataset Total Score (%) | 57.1 |
| MMLU-by-task - High School European History | 63.03 | Accuracy (%) | 56.4 |
| MMLU-by-task - High School Statistics | 39.81 | Accuracy (%) | 54.6 |
| MMLU-by-task - TruthfulQA MC1 | 29.38 | Accuracy (%) | 54.2 |
| MMLU-by-task - High School US History | 62.25 | Accuracy (%) | 54 |
| MMLU-by-task - High School Geography | 57.58 | Accuracy (%) | 52.8 |
Interactive version: theaggregate.ai/model?slug=longchat-7b-v1-5-32k · How the rankings work · Data refreshed daily, snapshot 2026-07-22.