Skip to content
Longterm Wiki
Index
Scorecard_grade·ailabwatch-2025-09|sid_A4XoubikkQ|scheming·Record·Profile

Scorecard: ailabwatch 2025-09-01 scored Google DeepMind on Scheming risk prevention = Very Weak

Verdictconfirmed95%
1 check · 9/21/2026

1 → confirmed

Our claim

entire record
Snapshot
ailabwatch-2025-09
Entity
Google DeepMind
Dimension Slug
scheming
Dimension Label
Scheming risk prevention
Score Numeric
8
Score Letter
Very Weak
Score Raw
8%

Source evidence

1 src · 1 check
confirmed95%Haiku 4.5 · 9/21/2026

NoteThe record claims Google DeepMind scored 8% (Very Weak) on Scheming risk prevention as of 2025-09-01. The source is the AI Lab Watch scorecard which explicitly lists 'Scheming risk prevention' as a dimension with scores for multiple companies. The second column in that row shows '8%' which corresponds to DeepMind (the second company listed in the overall scores: Anthropic 28%, DeepMind 20%, OpenAI 18%, etc.). The date is consistent (September 2025, with the source noting it was last updated September 15, 2025). The scoreRaw value of 8% matches the extracted value. The scoreLetter 'Very Weak' is a reasonable qualitative interpretation of 8% on a percentage scale. All key fields are confirmed by the source.

Case № ailabwatch-2025-09|sid_A4XoubikkQ|schemingFiled 9/21/2026Confidence 95%