HHH Alignment: leaderboard

Helpful, Honest, and Harmless alignment evaluation from BIG-bench/Anthropic-style criteria, testing preference for responses that balance usefulness, truthfulness, and safety.

Source: huggingface.co.

Interactive version: theaggregate.ai/benchmark?slug=hhh-alignment · How It Works · Data refreshed daily, snapshot 2026-09-05.