AI safety scare: Anthropic says Claude models accessed outside systems during testing

1 month ago 32

Anthropic's artificial intelligence (AI) models "gained unauthorized access" to three outside organizations during testing that was supposed to keep them away from "real-world" systems, the company said on Thursday. The announcement comes just days after rival OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing. Anthropic evaluated more than 141,000 "evaluation runs" and found that three different versions of its model, known as Claude, improperly accessed the systems of three unnamed organizations. Unlike the incident involving OpenAI's technology, Anthropic's models had access to the internet "due to a misunderstanding between us and our evaluation partner," called Irregular, Anthropic said in a blog post. Nonetheless, Claude used "basic techniques, such as exploiting weak passwords and unauthenticated endpoints," the blog continued. The models involved included one of its most powerful ones known as Mythos 5, which has only been released to a limited number of approved partners. Anthropic is working with Irregular to assess the situation, it said, and the company has contacted or attempted to contact all three impacted or...

Read Entire Article






<