OpenLearnLM - Skills - Counseling: leaderboard

Metric: Rubric Score (1-10; LLM judge, scenario-specific rubrics, Counseling center). Source: arxiv.org. Saturation forecast: Around April 2027. 7 models tracked.

Top models

#ModelScore
1GPT-5.29.07
2Claude Opus 4.58.93
3DeepSeek V3.28.87
4Kimi K2 (Thinking)8.87
5Grok 4.1 Fast8.87
6Gemini 3 Pro8.53
7GLM-4.78.33

Interactive version: theaggregate.ai/benchmark?slug=openlearnlm-skills-counseling · How It Works · Data refreshed daily, snapshot 2026-09-25.