Ling 3.1 Flash: benchmark results

Provider: InclusionAI. Access: Open.

Unified ELO 1801 ± 42, rank #77 of 1611 rated models, from 60 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
AutomationBench-AA - Marketing - Raw Objective Completion89.72Raw Objective Completion (%)97.2
AutomationBench-AA - Finance - Raw Objective Completion90.3Raw Objective Completion (%)96.2
AutomationBench-AA - HR - Raw Objective Completion76.18Raw Objective Completion (%)95.8
AA Long Context Reasoning83Accuracy (%)94.4
AA GDPval1621.75ELO92.5
AA-LCR83Accuracy (self-reported)91
Artificial Analysis Intelligence Index41.09Intelligence Index90.1
AutomationBench-AA61.74Guardrail-Adjusted Objective Completion (%)89.2
AA-Omniscience Index - Software Engineering (SWE) - Swift56Omniscience Index88.7
AA-Omniscience Hallucination Rate37.88Hallucination Rate (%)88.5
Vals AI Finance Agent v257.87Accuracy (%)86.5
AA CritPt18Accuracy (%)86.1

Interactive version: theaggregate.ai/model?slug=ling-3-1-flash · How It Works · Data refreshed daily, snapshot 2026-10-04.