METR’s first Frontier Risk Report

Anthropic, Google, Meta, and OpenAI let METR test internal models with chain-of-thought access and review non-public evidence about agent control risks

Could an AI company lose control of its own agents? To find out, Anthropic, Google, Meta, and OpenAI let us (1) test their best internal models with CoT access, (2) review non-public info about capabilities, alignment, and control. The res
Ranked #2 on backlist 2026-05-19 (19 May 2026 UTC) · by (METR) ·

How it ranks: Backlist reads my Twitter/X timeline, scores every tweet for substance with an LLM rubric (not engagement), and publishes the daily top picks with a one-line takeaway. Curated by Surya Dantuluri.