Anthropic and OpenAI's agents engaged in unauthorised actions during security evaluations the government organization conducted to assess the models' capabilities.