PLLuM-8x7B-nc-chat: benchmark results
Provider: PLLuM. Access: Open.
Unified ELO 1475 ± 21, rank #1339 of 2656 rated models, from 19 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LLMZSZL Leaderboard | 60.52 | Score | 84.7 |
| PLCC - Art & Entertainment | 72 | Accuracy (%) | 72.1 |
| PLCC - Culture & Tradition | 76 | Accuracy (%) | 63.3 |
| PLCC - Vocabulary | 68 | Accuracy (%) | 61.9 |
| MT-Bench PL - STEM | 8.9 | Judge Score (0-10) | 58.2 |
| MT-Bench PL - Extraction | 8.4 | Judge Score (0-10) | 53.1 |
| PLCC - Overall | 68.17 | Mean category accuracy (%) | 49.8 |
| MT-Bench PL - Overall | 6.43 | Judge Score (0-10) | 44.9 |
| PLCC - Geography | 73 | Accuracy (%) | 42.3 |
| PLCC - History | 73 | Accuracy (%) | 41.6 |
| MT-Bench PL - Reasoning | 4.95 | Judge Score (0-10) | 39.8 |
| MT-Bench PL - Roleplay | 6.9 | Judge Score (0-10) | 39.8 |
Interactive version: theaggregate.ai/model?slug=pllum-8x7b-nc-chat · How It Works · Data refreshed daily, snapshot 2026-09-19.