Qwen 3 14B — benchmark results
Alibaba Qwen 3 14B model row. Provider: Alibaba. Released 2025-04-28. Access: Open.
Unified ELO 1454 ± 14, rank #977 of 1776 rated models, from 444 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AgingBench | 7.9 | S1 kw_m HL (self-reported) | 100 |
| AppWorld Challenge | 67.6 | Task Goal Completion (%) | 100 |
| AppWorld Normal | 86.9 | Task Goal Completion (%) | 100 |
| ChLogic | 99.13 | English (self-reported) | 100 |
| TSCG | 90.2 | 20 Tools (json-text) (self-reported) | 100 |
| When Simulation Lies | 52.9 | Pert Acc (self-reported) | 100 |
| AGC-Bench - unfun_corpus | 1.44 | Dataset z-score | 98.7 |
| INCLUDE-base-44 European Languages | 63.03 | Average Accuracy (%) | 97.1 |
| EuroEval German NLU - GermEval | 73.49 | Named entity recognition Score (%) | 95.7 |
| EuroEval Lithuanian | 58.25 | Average Score (%) | 95.2 |
| EuroEval Lithuanian NLU | 55.05 | NLU Average Score (%) | 95.2 |
| AI Chess Leaderboard (Continuation) | 1409 | Elo | 94.7 |
Interactive version: theaggregate.ai/model?slug=qwen-3-14b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.