TOFU LLaMA 1% — leaderboard
TOFU evaluates machine unlearning for large language models on fictitious author QA data; this leaderboard variant reports LLaMA submissions at the 1% forget-set setting.
Source: huggingface.co.
Interactive version: theaggregate.ai/benchmark?slug=tofu-llama-1 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.