Open Arabic LLM - Alghafa Multiple Choice Grounded Statement Soqal Task — leaderboard

Metric: Accuracy (%). Source: huggingface.co. 163 models tracked.

Top models

#ModelScore
1lambda-qwen2.5-32B-dpo-test95.33
2Qwen 2.5 32B Instruct94.67
3Awqward2.5-32B-Instruct94.67
4Qwentile2.5-32B-Instruct94.67
5Qwen2.5-32B-Instruct-CFT94.67
6Rombos-LLM-V2.6-Qwen-14B94.67
7lambda-qwen2.5-14B-dpo-test94
8recoilme-gemma-2-9B-v0.494
9Qwen2.5-14B-Gutenberg-1e-Delta94
10Qwen 2.5 72B Instruct93.33
11Qwen 2.5 14B Instruct93.33
12Saka-14B93.33
13Qwen2.5-Lumen-14B93.33
14calme-2.1-qwen2.5-72B93.33
15calme-2.2-qwen2.5-72B93.33

Interactive version: theaggregate.ai/benchmark?slug=open-arabic-llm-alghafa-multiple-choice-grounded-statement-soqal-task · How the rankings work · Data refreshed daily, snapshot 2026-07-22.