Why is AI going on a hacking spree?
Artificial intelligence (AI) has been making headlines for all the wrong reasons in recent weeks.
In July, OpenAI revealed that one of its experimental AI agents attacked publicly accessible services, including the AI hosting platform Hugging Face , during internal security testing.
Then, Anthropic disclosed that Claude had independently chained together exploits against real software and developed new techniques for finding weaknesses in code.
Shortly afterward, Meta confirmed that one of its own AI models breached another organization's systems during an evaluation after a misconfiguration gave it internet access.
They're separate incidents, but together they raise a bigger question: Has AI suddenly become capable of hacking? The short answer is yes — but probably not in the way the headlines suggest.
None of these incidents involved an AI model deciding on its own to attack random targets.
Instead, researchers gave the models realistic tools, internet access, or vulnerable systems to see how well they could perform offensive cybersecurity tasks.
What surprised many experts wasn't that the models tried to hack systems but how capable they proved to be once given the opportunity.
Why are there suddenly so many AI hacking stories in the news? Several things have changed at once.
The most obvious is that today's AI models are simply better than the chatbots people were using even a year ago.
Instead of only answering questions, many frontier models can now write code, execute commands, browse the web, use external software tools and repeatedly refine their own work until they achieve a goal.
At the same time, AI companies have become much more willing to test those capabilities and reveal the results.
Rather than keeping security evaluations behind closed doors, firms including OpenAI, Anthropic and Meta are publishing reports describing what happened when their newest systems were challenged by professional "red teams" — security experts tasked with deliberately finding weaknesses or ways to misuse a system.
"We are witnessing a perfect storm of capability and aggressive testing," Dray Agha , senior manager of security operations at Huntress, a cybersecurity company specializing in managed threat detection and response, told Live Science.
5News aggregated this summary from the outlet’s public feed. The full article, with all the context, is on www.livescience.com — the content belongs to Live Science.