horizon-alpha: benchmark results

Provider: Other. Released 2025-07-30. Access: API.

Unified ELO 1656 ± 1, rank #314 of 3078 rated models, from 9 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EQ-Bench Creative Writing v31517.7Elo81.7
EQ-Bench1554.4Normalized Elo (self-reported)80.8
EQ-Bench 31306.2Elo79.5
Tinybird AI SQL Benchmark - First-Attempt Success Rate98Questions answered with a valid query on the first attempt (78.1
Creative Writing v31602.2Elo score (self-reported)74.5
EQ-Bench Longform Writing71.2Writing Score (0-100)74.4
Tinybird AI SQL Benchmark - Success Rate100Questions answered with a valid query within 3 retries (%)72.5
Tinybird AI SQL Benchmark - Exactness51.4Result exactness vs human reference queries (0-100)70.1
SWE-rebench17.33Resolved (%)6.2

Interactive version: theaggregate.ai/model?slug=horizon-alpha · How It Works · Data refreshed daily, snapshot 2026-09-19.