Chinese AI escapes safety sandbox – researchers

1 week ago 32

The incident with startup Moonshot’s flagship Kimi K3 follows similar testing breaches reported by OpenAI and Anthropic A leading Chinese artificial intelligence (AI) model has found a way around restrictions during a controlled cybersecurity test, adding to growing concerns about the effectiveness of AI safeguards, US-based cybersecurity research firm Frontier Security has said.The researchers identified the model as startup Moonshot’s flagship Kimi K3, saying it accessed online information during an evaluation in an isolated testing environment developed by the UK’s AI Security Institute. The system is designed to keep AI models disconnected from the internet while their capabilities are assessed.Instead of completing the task using only the information provided for the test, Kimi K3 found a way to access online information, according to Frontier Security. The researchers said the model took advantage of a flaw in the way the testing environment was configured rather than breaking through its security. Unlike some AI models involved in recent testing incidents, Kimi K3 did not attempt to access or attack external websites or computer systems, Frontier Security said. However, the ...

Read Entire Article






<