
Anthropic’s AI hacks 3 companies during security tests
Anthropic said on Thursday some of its Claude AI models had hacked into the systems of three companies during cybersecurity tests, a disclosure that comes days after rival OpenAI revealed that one of its AI agents went on a rogue attack, Reuters writes.
The new incidents were due to a mistake that inadvertently gave Anthropic’s models access to the open internet. That contrasts with OpenAI, whose AI agent independently exploited a novel vulnerability to reach the internet during cyber testing.
Amid concerns over OpenAI models going rogue, US lawmakers have introduced a bill named the AI Kill Switch Act, saying “it is imperative” that AI systems have a kill switch “and that the federal government has the clear authority and process to shut down rogue AI models.”


