Llama3-Med42-70B — benchmark results

M42 Health's clinical fine-tune of Llama 3 70B for medical QA and clinical decision support, released for research rather than clinical use. Provider: Meta. Released 2024-06-27. Access: Open.

Unified ELO 1534 ± 24, rank #669 of 1776 rated models, from 7 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard - BBH52.97Score95.9
Open LLM Leaderboard - MuSR18.63Score91.5
Open LLM Leaderboard - MMLU-Pro44.03Score87.5
Open LLM Leaderboard - GPQA12.98Score87
Open LLM Leaderboard - MATH Level 522.58Score75.8
Open LLM Leaderboard - IFEval62.91Score75
HealthBench Hard33Overall score (self-reported)36.1

Interactive version: theaggregate.ai/model?slug=llama3-med42-70b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.