Skip to content
Longterm Wiki
Index
Fact·f_1BWsBJuBcg·Fact

Center for AI Safety (CAIS) — Description: The Center for AI Safety (CAIS) is a San Francisco-based nonprofit focused on reducing societal-scale risks from AI through technical safety research, field-building, and public communication. Founded by Dan Hendrycks and Oliver Zhang. Known for the MMLU benchmark, representation engineering, and the May 2023 "Statement on AI Risk" signed by 350+ AI leaders.

Verdictpartial85%
1 check · 9/21/2026

1 → partial

Our claim

entire record
Subject
Center for AI Safety (CAIS)
Property
Description
Value
The Center for AI Safety (CAIS) is a San Francisco-based nonprofit focused on reducing societal-scale risks from AI through technical safety research, field-building, and public communication. Founded by Dan Hendrycks and Oliver Zhang. Known for the MMLU benchmark, representation… expandThe Center for AI Safety (CAIS) is a San Francisco-based nonprofit focused on reducing societal-scale risks from AI through technical safety research, field-building, and public communication. Founded by Dan Hendrycks and Oliver Zhang. Known for the MMLU benchmark, representation engineering, and the May 2023 "Statement on AI Risk" signed by 350+ AI leaders.
As Of
2025

Source evidence

1 src · 1 check
partial85%primaryHaiku 4.5 · 9/21/2026

NoteThe claim makes four main assertions: (1) San Francisco-based nonprofit—unverifiable from source (location not stated); (2) focused on reducing societal-scale risks through technical safety research, field-building, and public communication—CONFIRMED (source says 'research, field-building, and advocacy'); (3) Founded by Dan Hendrycks and Oliver Zhang—CONFIRMED (both listed in Leadership); (4) Known for MMLU benchmark, representation engineering, and May 2023 'Statement on AI Risk' signed by 350+ leaders—CONTRADICTED on the statement signature count. The source states the 'Global Statement on AI Risk' was signed by '600 leading AI researchers and public figures,' not 350+. The MMLU benchmark and representation engineering are not mentioned in the source, making those claims unverifiable. The temporal qualifier 'as of 2025' cannot be verified from the source date. Overall: location unverifiable, core mission confirmed, founders confirmed, but the AI Risk statement signature count is contradicted (600 vs. 350+), and two key accomplishments are unverifiable.

Case № f_1BWsBJuBcgFiled 9/21/2026Confidence 85%
Source Check: Description | Longterm Wiki