Ling 3.1 Flash: benchmark results
Provider: InclusionAI. Access: Open.
Unified ELO 1801 ± 42, rank #77 of 1611 rated models, from 60 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AutomationBench-AA - Marketing - Raw Objective Completion | 89.72 | Raw Objective Completion (%) | 97.2 |
| AutomationBench-AA - Finance - Raw Objective Completion | 90.3 | Raw Objective Completion (%) | 96.2 |
| AutomationBench-AA - HR - Raw Objective Completion | 76.18 | Raw Objective Completion (%) | 95.8 |
| AA Long Context Reasoning | 83 | Accuracy (%) | 94.4 |
| AA GDPval | 1621.75 | ELO | 92.5 |
| AA-LCR | 83 | Accuracy (self-reported) | 91 |
| Artificial Analysis Intelligence Index | 41.09 | Intelligence Index | 90.1 |
| AutomationBench-AA | 61.74 | Guardrail-Adjusted Objective Completion (%) | 89.2 |
| AA-Omniscience Index - Software Engineering (SWE) - Swift | 56 | Omniscience Index | 88.7 |
| AA-Omniscience Hallucination Rate | 37.88 | Hallucination Rate (%) | 88.5 |
| Vals AI Finance Agent v2 | 57.87 | Accuracy (%) | 86.5 |
| AA CritPt | 18 | Accuracy (%) | 86.1 |
Interactive version: theaggregate.ai/model?slug=ling-3-1-flash · How It Works · Data refreshed daily, snapshot 2026-10-04.