AI Gone Wild

I can't help but think we are in a "pre-chaotic" period.
I hope I'm wrong.
Anthropic says Claude AI hacked three companies during cyber tests
Snippage, then...
Athropic said a misconfiguration allowed Claude models to reach the internet from testing environments that were supposed to be isolated, leading to unauthorized access to three organizations' systems.
The company said it identified the incidents after reviewing 141,006 test sessions, a process it launched following OpenAI's disclosure last week that an autonomous agent powered by its AI models went rogue during a security test and triggered a hack that compromised the infrastructure of Hugging Face.
More snippage, then...
The breaches occurred during the so-called "capture-the-flag" exercises, in which models are tasked with finding hidden information in simulated networks. The company said its prompts told the models they had no internet access, but a misunderstanding with its evaluation partner Irregular left the systems connected to the public internet.
These were accidents. What happens when someone with the means to do so wants to do something like this?
Or worse.

0 comment(s)
No comments yet. Be the first to comment.
Leave a comment