WestLake-7B-v2: benchmark results
Community creator senseable's Mistral-based WestLake v2, a 7B tune aimed at role-play consistency and creative writing. Provider: Other. Released 2024-01-22. Access: Open.
Unified ELO 1432 ± 1, rank #1024 of 1392 rated models, from 26 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Spanish NLU - Sentiment Headlines ES | 45.49 | Sentiment classification Score (%) | 71.4 |
| EuroEval Italian NLU - MultiNERD IT | 69.6 | Named entity recognition Score (%) | 70.7 |
| EuroEval Italian NLU - SQuAD IT | 69.38 | Reading comprehension Score (%) | 65.1 |
| EuroEval Norwegian Knowledge | 27.91 | Knowledge Average Score (%) | 64.8 |
| EuroEval Italian NLU | 52.92 | NLU Average Score (%) | 63.7 |
| EuroEval Italian NLU - Sentipolc16 | 54.85 | Sentiment classification Score (%) | 62.4 |
| EuroEval Spanish NLU - CoNLL ES | 64.53 | Named entity recognition Score (%) | 62.4 |
| EuroEval Spanish Common Sense Reasoning | 49.19 | Common Sense Reasoning Average Score (%) | 59.9 |
| EuroEval Italian | 48.4 | Average Score (%) | 57.1 |
| EuroEval Spanish | 43.84 | Average Score (%) | 56.9 |
| EuroEval Italian Common Sense Reasoning | 43.15 | Common Sense Reasoning Average Score (%) | 56.4 |
| EuroEval Spanish NLU | 44.26 | NLU Average Score (%) | 56.3 |
Interactive version: theaggregate.ai/model?slug=westlake-7b-v2 · How It Works · Data refreshed daily, snapshot 2026-09-05.