We now embed researchers inside of AI labs to stress test monitoring, assess AI loss-of-control risk, and investigate misalignment. If you want to apply DFIR skills in frontier AI, apply (or DM).
Comp range is $400k - 580k cash.
jobs.lever.co/metr/b1a2f73...
We now embed researchers inside of AI labs to stress test monitoring, assess AI loss-of-control risk, and investigate misalignment. If you want to apply DFIR skills in frontier AI, apply (or DM).
Comp range is $400k - 580k cash.
jobs.lever.co/metr/b1a2f73...
The method compares performance as a function of spend for humans vs agents. The point where humans become more cost-effective is the agent’s expenditure horizon.
The method compares performance as a function of spend for humans vs agents. The point where humans become more cost-effective is the agent’s expenditure horizon.
The result: our first Frontier Risk Report.
The result: our first Frontier Risk Report.
On average, participants self-report that AI use made their work 1.6–2.1x more valuable, and that this multiplier will grow over time.
On average, participants self-report that AI use made their work 1.6–2.1x more valuable, and that this multiplier will grow over time.
www.nytimes.com/2026/04/17/t....
www.nytimes.com/2026/04/17/t....
In our new benchmark, MirrorCode, Claude Opus 4.6 reimplemented a 16,000-line bioinformatics toolkit — a task we believe would take a human engineer weeks.
Co-developed with @METR_Evals. Details in thread.
Link: metr.org/notes/2026-0...
Link: metr.org/notes/2026-0...
In early work, we find clear trends: more capable models (in terms of time horizon) are better able to detect covert behavior.
In early work, we find clear trends: more capable models (in terms of time horizon) are better able to detect covert behavior.