sherlock-dash-alpha: benchmark results

Provider: Other. Released 2025-11-15. Access: API.

Unified ELO 1671 ± 50, rank #397 of 2656 rated models, from 10 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
SpeechMap Compliance100% Requests Completed100
SvelteBench92.2Average pass@1 (%)68.9
EQ-Bench Creative Writing v31210.8Elo47.1
NYT Connections Older Models12.2Score (%)45.9
Creative Writing v31328.9Elo score (self-reported)42.7
EQ-Bench Longform Writing47.2Writing Score (0-100)30.1
LisanBench0.01Mean Path Length / Current Maximum26.8
Tinybird AI SQL Benchmark - Exactness38.24Result exactness vs human reference queries (0-100)16.3
Tinybird AI SQL Benchmark - First-Attempt Success Rate62Questions answered with a valid query on the first attempt (10
Tinybird AI SQL Benchmark - Success Rate74Questions answered with a valid query within 3 retries (%)9.2

Interactive version: theaggregate.ai/model?slug=sherlock-dash-alpha · How It Works · Data refreshed daily, snapshot 2026-09-19.