Llama3-Med42-70B — benchmark results
M42 Health's clinical fine-tune of Llama 3 70B for medical QA and clinical decision support, released for research rather than clinical use. Provider: Meta. Released 2024-06-27. Access: Open.
Unified ELO 1534 ± 24, rank #669 of 1776 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - BBH | 52.97 | Score | 95.9 |
| Open LLM Leaderboard - MuSR | 18.63 | Score | 91.5 |
| Open LLM Leaderboard - MMLU-Pro | 44.03 | Score | 87.5 |
| Open LLM Leaderboard - GPQA | 12.98 | Score | 87 |
| Open LLM Leaderboard - MATH Level 5 | 22.58 | Score | 75.8 |
| Open LLM Leaderboard - IFEval | 62.91 | Score | 75 |
| HealthBench Hard | 33 | Overall score (self-reported) | 36.1 |
Interactive version: theaggregate.ai/model?slug=llama3-med42-70b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.