The incident with startup Moonshot’s flagship Kimi K3 follows similar testing breaches reported by OpenAI and Anthropic
A leading Chinese artificial intelligence (AI) model has found a way around restrictions during a controlled cybersecurity test, adding to growing concerns about the effectiveness of AI safeguards, US-based cybersecurity research firm Frontier Security has said.
The researchers identified the model as startup Moonshot’s flagship Kimi K3, saying it accessed online information during an evaluation in an isolated testing environment developed by the UK’s AI Security Institute. The system is designed to keep AI models disconnected from the internet while their capabilities are assessed.
Instead of completing the task using only the information...
Read the full article here
