INS-ActBench - Knowledge: leaderboard

Metric: Accuracy (%) on the 9,749 single-answer multiple-choice knowledge questions (INS-Act-Know); 12,050 questions from the exams of 16 actuarial associations; knowledge and case questions two-shot, practice tasks zero-shot; proprietary models with thinking enabled at high where supported. Source: arxiv.org. Saturation forecast: Around December 2026. 9 models tracked.

Top models

#ModelScore
1Gemini 3.1 Pro (Preview)94.5
2GPT-5.5 (High)93.3
3DeepSeek V4 Pro (High)91.3
4Qwen 3.6 Plus87.8
5Claude Opus 4.7 (High)81.1
6Kimi K2.670.9
7Qwen 3 14B40.4
8Qwen 3.5 35B A3B38.1
9Gemma 3 12B (IT)34.7

Interactive version: theaggregate.ai/benchmark?slug=ins-actbench-knowledge · How It Works · Data refreshed daily, snapshot 2026-09-29.