Claimant scorecard · AERS v2.1 · Calibrating
Cognition
1 claim tracked in the Responsibility Ledger. 1 pending grade.
AERS
—
Insufficient closed grades
Pending
1
Open horizons
Closed
0
Graded outcomes
First tracked
Jul 10, 2026
Open horizons
Cognition: Released SWE-1.7 on July 8, 2026, claiming near-frontier coding performance at lower cost by running reinforcement learning on an already-RL-trained base, challenging the "post-training ceiling"
Invalidator — If by January 2027, no other frontier lab or independent research group publishes results demonstrating comparable additional RL gains atop an already-RL-trained base, or if an independent audit finds FrontierCode 1.1 benchmark contamination that inflates SWE-1.7's reported scores.
About this scorecard
The AI Execution Risk Score (AERS) is a 0-100 metric quantifying the gap between Cognition’s public AI claims and demonstrated delivery. Higher AERS = stronger track record. Each claim above is drawn from a primary source linked in the original Ledger entry; the horizon date is when the claim becomes graded under the published methodology. Materiality is the editor’s assessment of the claim’s formality from 1 (PR statement) to 5 (earnings call or SEC filing).
AERS v2.1 · Methodology in active calibration · Not investment advice.