
Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation comes days after rival OpenAI said one of its own models had breached developer platform Hugging Face, adding to growing unease over whether frontier AI labs are doing enough to control the increasingly capable systems they are building.
In a blog post describing the incidents, Anthropic said Claude gained unauthorized access to the systems during cybersecurity evaluations.
Want to read more from the original publisher?
Read the full story at The Verge.
Create enterprise-grade marketing content, code summaries, and strategy reports 5x faster.
Be the first to react to this story!
Please sign in to leave a comment and share your opinion.
No comments yet. Be the first to start the conversation!
Most read & bookmarked this week