PLLuM-8x7B-chat: benchmark results
Provider: PLLuM. Access: Open.
Unified ELO 1461 ± 21, rank #1436 of 2656 rated models, from 19 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LLMZSZL Leaderboard | 52.8 | Score | 68.4 |
| MT-Bench PL - Extraction | 8 | Judge Score (0-10) | 42.9 |
| MT-Bench PL - Overall | 6.3 | Judge Score (0-10) | 40.8 |
| MT-Bench PL - STEM | 8.2 | Judge Score (0-10) | 39.8 |
| PLCC - Culture & Tradition | 60 | Accuracy (%) | 38.3 |
| MT-Bench PL - Reasoning | 4.9 | Judge Score (0-10) | 36.7 |
| MT-Bench PL - Coding | 4.55 | Judge Score (0-10) | 35.7 |
| MT-Bench PL - Humanities | 8.6 | Judge Score (0-10) | 34.7 |
| PLCC - Art & Entertainment | 45 | Accuracy (%) | 33.6 |
| PLCC - History | 68 | Accuracy (%) | 33.6 |
| PLCC - Geography | 66 | Accuracy (%) | 33 |
| CPTU Bench | 3.01 | Average Score (1-5) | 31.7 |
Interactive version: theaggregate.ai/model?slug=pllum-8x7b-chat · How It Works · Data refreshed daily, snapshot 2026-09-19.