IndicContextEval (Native-Script Entities): leaderboard

Metric: Word error rate (%) when the prompt gives the target language plus a list of 20-30 domain entities in the native script (L5); 56 hours of natural read and extempore speech from 555 speakers in 8 Indian languages (Hindi, Bengali, Telugu, Marathi, Gujarati, Malayalam, Odia, Urdu) and 23 professional domains, native-script output; lower is better. Source: arxiv.org. Saturation forecast: Around September 2027. 4 models tracked.

Top models

#ModelScore
1Gemini 3 Flash17.46
2Gemma 3n E4B43.11

Interactive version: theaggregate.ai/benchmark?slug=indiccontexteval-native-script-entities · How It Works · Data refreshed daily, snapshot 2026-09-29.