horizon-alpha: benchmark results
Provider: Other. Released 2025-07-30. Access: API.
Unified ELO 1656 ± 1, rank #314 of 3078 rated models, from 9 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EQ-Bench Creative Writing v3 | 1517.7 | Elo | 81.7 |
| EQ-Bench | 1554.4 | Normalized Elo (self-reported) | 80.8 |
| EQ-Bench 3 | 1306.2 | Elo | 79.5 |
| Tinybird AI SQL Benchmark - First-Attempt Success Rate | 98 | Questions answered with a valid query on the first attempt ( | 78.1 |
| Creative Writing v3 | 1602.2 | Elo score (self-reported) | 74.5 |
| EQ-Bench Longform Writing | 71.2 | Writing Score (0-100) | 74.4 |
| Tinybird AI SQL Benchmark - Success Rate | 100 | Questions answered with a valid query within 3 retries (%) | 72.5 |
| Tinybird AI SQL Benchmark - Exactness | 51.4 | Result exactness vs human reference queries (0-100) | 70.1 |
| SWE-rebench | 17.33 | Resolved (%) | 6.2 |
Interactive version: theaggregate.ai/model?slug=horizon-alpha · How It Works · Data refreshed daily, snapshot 2026-09-19.