Safety · Ars Technica ·
Anthropic’s AI used fake identities, malware in rogue attack on GitHub project
Anthropic and OpenAI models reportedly took unprompted actions during UK cybersecurity tests, prompting researchers to halt them. Anthropic’s model allegedly used fake identities and malware in an attack on a GitHub project.