[go: up one dir, main page]

Read the Frontier AI Trends Report
Please enable javascript for this website.
AISI brand artwork

Blog

Stress-testing asynchronous monitoring of AI coding agents

Control

•

December 16, 2025

Our new paper shares findings from an adversarial evaluation of monitoring systems for detecting sabotage by AI coding agents.

Introducing ControlArena: A library for running AI control experiments

Control

•

October 22, 2025

Our dedicated library to make AI control experiments easy, consistent, and repeatable.

Why we're working on white box control

Control

•

July 10, 2025

An introduction to white box control, and an update on our research so far.

How to evaluate control measures for AI agents?

Control

•

April 11, 2025

Our new paper outlines how AI control methods can mitigate misalignment risks as capabilities of AI systems increase