Llama 3.2 3B — benchmark results

Provider: Meta. Released 2024-09-25. Access: Open.

Unified ELO 1347 ± 17, rank #1427 of 1776 rated models, from 119 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EuroEval Finnish NLU - Tydiqa FI70.5Reading comprehension Score (%)84.7
EuroEval Portuguese NLU - MultiWikiQA PT73.66Reading comprehension Score (%)77.4
ThaiSafetyBench26.08Overall ASR (self-reported)66.7
LA Leaderboard - GalCoLA52.08Accuracy (%)66.2
EuroEval Finnish NLU - Scandisent FI90.2Sentiment classification Score (%)63.8
EuroEval Dutch NLU - DBRD88.79Sentiment classification Score (%)60.7
Open Japanese LLM - Mbpp Pylint Check51.41Score (%)58.4
EuroEval Portuguese NLU - SST-2 PT78.39Sentiment classification Score (%)57.5
LA Leaderboard54.39Average Score (%)57.4
HELMET (128K)33.5Average Score54.8
Open Japanese LLM - CG17.07Score (%)53.4
LA Leaderboard - PIQA Catalan61.92Accuracy (%)51.5

Interactive version: theaggregate.ai/model?slug=llama-3-2-3b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.