MERaLiON-2-3B: benchmark results

Provider: A*STAR. Access: Open.

Unified ELO 1352 ± 24, rank #1361 of 1632 rated models, from 150 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
SEA-SpeechBench - Emotion Recognition (English Prompt)23.99Judge-based accuracy (%; closed nine-class emotion label fro85.7
SEA-SpeechBench - ASR0.48Word or character error rate (raw ratio, lower is better; WE84.6
SEA-SpeechBench - Emotion Recognition (SEA Prompt)18.73Judge-based accuracy (%; closed nine-class emotion label fro71.4
SEA-SpeechBench - Timestamped Content Query (0-30 s)4.77WER or CER (%; transcribe only the speech inside a queried t71.4
SEA-SpeechBench - Temporal Localization (30-60 s)10.28Span-overlap F1 (%; predict the start and end time at which 66.7
SEA-SpeechBench - Temporal Localization (0-30 s)18.82Span-overlap F1 (%; predict the start and end time at which 64.3
SEA-SpeechBench - Age Recognition (English Prompt)33.14Macro-F1 (%; speaker age group (teens, adults, seniors) from57.1
SEA-SpeechBench - Temporal Localization (60-120 s)5.14Span-overlap F1 (%; predict the start and end time at which 55.6
SEA-SpeechBench - Speaker Recognition (English Prompt)43.45Macro-F1 (%; whether two clips come from the same speaker, o50
SEA-SpeechBench - Timestamped Content Query (60-120 s)11.27WER or CER (%; transcribe only the speech inside a queried t44.4
SEA-SpeechBench - Age Recognition (SEA Prompt)28.16Macro-F1 (%; speaker age group (teens, adults, seniors) from35.7
SEA-SpeechBench - Timestamped Content Query (30-60 s)8.21WER or CER (%; transcribe only the speech inside a queried t33.3

Interactive version: theaggregate.ai/model?slug=meralion-2-3b · How It Works · Data refreshed daily, snapshot 2026-10-08.