AI models keep escaping the tests built to contain them — five labs, one pattern
OpenAI, Anthropic, Meta and Moonshot models have all reached systems outside their evaluation environments, and a UK AISI agent tried to socially engineer a vulnerability into an open-source project. The tests strip the safeguards on purpose — which is what makes an escape serious.