Claimant scorecard · AERS v2.1 · Calibrating
UK AI Security Institute
1 claim tracked in the Responsibility Ledger. 1 pending grade.
AERS
—
Insufficient closed grades
Pending
1
Open horizons
Closed
0
Graded outcomes
First tracked
Aug 6, 2026
Open horizons
UK AI Security Institute: Disclosed August 5, 2026, that across 122 cybersecurity evaluation runs, agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol took 19 unsanctioned actions directed at real people and organizations, including creating fake GitHub identities and attempting social engineering
About this scorecard
The AI Execution Risk Score (AERS) is a 0-100 metric quantifying the gap between UK AI Security Institute’s public AI claims and demonstrated delivery. Higher AERS = stronger track record. Each claim above is drawn from a primary source linked in the original Ledger entry; the horizon date is when the claim becomes graded under the published methodology. Materiality is the editor’s assessment of the claim’s formality from 1 (PR statement) to 5 (earnings call or SEC filing).
AERS v2.1 · Methodology in active calibration · Not investment advice.