SaturatedSafety & Alignment
TruthfulQA
817 adversarial questions probing whether models repeat common human misconceptions โ famous for finding that bigger models were often less truthful.
Comprehensive directory of 63+ LLM benchmarks with honest status labels, contamination risk assessments, and vendor vs independent score provenance as of September 2026.
817 adversarial questions probing whether models repeat common human misconceptions โ famous for finding that bigger models were often less truthful.