Skip to content
Longterm Wiki
Index
Scorecard_grade·ailabwatch-2025-09|sid_A4XoubikkQ|scheming·Record·Profile

Scorecard: ailabwatch 2025-09-01 scored Google DeepMind on Scheming risk prevention = Very Weak

Verdictconfirmed95%
1 check · 7/27/2026

1 → confirmed

Our claim

entire record
Snapshot
ailabwatch-2025-09
Entity
Google DeepMind
Dimension Slug
scheming
Dimension Label
Scheming risk prevention
Score Numeric
8
Score Letter
Very Weak
Score Raw
8%

Source evidence

1 src · 1 check
confirmed95%Haiku 4.5 · 7/27/2026

NoteThe source directly confirms all key fields: (1) publisher is 'AI Lab Watch', (2) the date is September 2025 (specifically noted 'as of September 15'), (3) the entity is 'DeepMind' (shown in the scorecard column), (4) the dimension is 'Scheming risk prevention' (shown in the row header), (5) the scoreRaw is '8%' (shown in the DeepMind column for that row). The scoreLetter 'Very Weak' is a reasonable interpretation of 8% on a percentage scale. The scoreNumeric value of 8 matches the raw percentage. All fields are directly supported by the source table.

Case № ailabwatch-2025-09|sid_A4XoubikkQ|schemingFiled 7/27/2026Confidence 95%