Hunyuan-T1-20250711: benchmark results
Provider: Tencent. Released 2025-07-11. Access: API.
Unified ELO 1784 ± 35, rank #179 of 2656 rated models, from 102 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| OpenCompass Language - Instruction Following - Chinese (CompassBench 2507) | 83.1 | Score (%) | 96.3 |
| ReLE - Language and Instruction Following | 72.9 | Accuracy (%) | 93 |
| OpenCompass Agent - Hard Multi-Turn (CompassBench 2507) | 50 | Score (%) | 91.7 |
| OpenCompass Reasoning - Common - English (CompassBench 2507) | 75.6 | Score (%) | 91.7 |
| SuperCLUE General (July 2025) - Hallucination Control | 83.39 | Score | 91.3 |
| OpenCompass Code - Competition - Chinese (CompassBench 2507) | 59.4 | Score (%) | 88.9 |
| OpenCompass Knowledge - Humanities - English (CompassBench 2507) | 87 | Score (%) | 88.9 |
| OpenCompass Code - Competition - English (CompassBench 2507) | 60.9 | Score (%) | 87 |
| SuperCLUE General (July 2025) - Precise Instruction Following | 42.73 | Score | 87 |
| ReLE - Education - Primary School Subjects | 68 | Accuracy (%) | 85.4 |
| OpenCompass LLM - Code (CompassBench 2507) | 43.8 | Score (%) | 85.2 |
| OpenCompass LLM - Code - Chinese (CompassBench 2507) | 43.2 | Score (%) | 85.2 |
Interactive version: theaggregate.ai/model?slug=hunyuan-t1-20250711 · How It Works · Data refreshed daily, snapshot 2026-09-19.