LongCat Flash Lite — benchmark results
Meituan LongCat's lite flash tier: 68.5B MoE with ~3B active params and a 256K context (early 2026). Provider: Meituan. Released 2026-01-29. Access: Open.
Unified ELO 1560 ± 27, rank #592 of 1776 rated models, from 38 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA TAU-2 Bench | 79.53 | Accuracy (%) | 69.6 |
| ZeroEval MATH-500 | 96.8 | MATH-500 Score | 69.4 |
| Artificial Analysis Intelligence Index | 17.22 | Intelligence Index | 53.5 |
| AA IFBench | 43.06 | Accuracy (%) | 46.4 |
| AA-LCR | 25.7 | Score (self-reported) | 45.6 |
| AA Terminal-Bench Hard | 10.61 | Accuracy (%) | 45.1 |
| AA Humanity's Last Exam | 6.02 | Accuracy (%) | 44.1 |
| AA GPQA Diamond | 63.64 | Accuracy (%) | 42.6 |
| AA Omniscience - Software Engineering (SWE) - Dart | 18 | Accuracy (%) | 40.7 |
| LLM Stats (CMMLU) | 82.48 | Score (%) | 40 |
| ZeroEval GPQA Diamond | 66.78 | GPQA Diamond Score | 39.8 |
| AA Omniscience - Software Engineering (SWE) - Kotlin | 14 | Accuracy (%) | 38.6 |
Interactive version: theaggregate.ai/model?slug=longcat-flash-lite · How the rankings work · Data refreshed daily, snapshot 2026-07-22.