WestLake-7B-v2 — benchmark results
Community creator senseable's Mistral-based WestLake v2, a 7B tune aimed at role-play consistency and creative writing. Provider: Other. Released 2024-01-22. Access: Open.
Unified ELO 1435 ± 22, rank #1070 of 1776 rated models, from 26 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Spanish NLU - Sentiment Headlines ES | 45.49 | Sentiment classification Score (%) | 71.4 |
| EuroEval Italian NLU - MultiNERD IT | 69.6 | Named entity recognition Score (%) | 70.7 |
| EuroEval Italian NLU - SQuAD IT | 69.38 | Reading comprehension Score (%) | 65.1 |
| EuroEval Norwegian Knowledge | 27.91 | Knowledge Average Score (%) | 64.8 |
| EuroEval Italian NLU | 52.92 | NLU Average Score (%) | 63.7 |
| EuroEval Italian NLU - Sentipolc16 | 54.85 | Sentiment classification Score (%) | 62.4 |
| EuroEval Spanish NLU - CoNLL ES | 64.53 | Named entity recognition Score (%) | 62.4 |
| EuroEval Spanish Common Sense Reasoning | 49.19 | Common Sense Reasoning Average Score (%) | 59.9 |
| EuroEval Italian | 48.4 | Average Score (%) | 57.1 |
| EuroEval Spanish | 43.84 | Average Score (%) | 56.9 |
| EuroEval Italian Common Sense Reasoning | 43.15 | Common Sense Reasoning Average Score (%) | 56.4 |
| EuroEval Spanish NLU | 44.26 | NLU Average Score (%) | 56.3 |
Interactive version: theaggregate.ai/model?slug=westlake-7b-v2 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.