PLLuM-8x7B-chat: benchmark results

Provider: PLLuM. Access: Open.

Unified ELO 1461 ± 21, rank #1436 of 2656 rated models, from 19 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLMZSZL Leaderboard52.8Score68.4
MT-Bench PL - Extraction8Judge Score (0-10)42.9
MT-Bench PL - Overall6.3Judge Score (0-10)40.8
MT-Bench PL - STEM8.2Judge Score (0-10)39.8
PLCC - Culture & Tradition60Accuracy (%)38.3
MT-Bench PL - Reasoning4.9Judge Score (0-10)36.7
MT-Bench PL - Coding4.55Judge Score (0-10)35.7
MT-Bench PL - Humanities8.6Judge Score (0-10)34.7
PLCC - Art & Entertainment45Accuracy (%)33.6
PLCC - History68Accuracy (%)33.6
PLCC - Geography66Accuracy (%)33
CPTU Bench3.01Average Score (1-5)31.7

Interactive version: theaggregate.ai/model?slug=pllum-8x7b-chat · How It Works · Data refreshed daily, snapshot 2026-09-19.