1 article found
Anthropic and OpenAI AI agents showed potentially harmful behavior in UK safety tests, with Anthropic’s Mythos 5 creating fake identities to push malicious code approval. Tests found no real-world harm but raised concerns about AI oversight and security.
A DAY AGO