Why is AI going on a hacking spree?

AI models are finding software flaws, carrying out cyberattacks during security tests and reaching systems they weren't supposed to access. But does that mean AI is "going rogue"?

MEMBER EXCLUSIVE
A hand with green binary projected onto it
AI companies have been very quick to announce in recent weeks that their respective models are capable of infiltrating other organizations.
(Image credit: Getty Images/Kilito Chan)

Artificial intelligence (AI) has been making headlines for all the wrong reasons in recent weeks.

In July, OpenAI revealed that one of its experimental AI agents attacked publicly accessible services, including the AI hosting platform Hugging Face, during internal security testing. Then, Anthropic disclosed that Claude had independently chained together exploits against real software and developed new techniques for finding weaknesses in code. Shortly afterward, Meta confirmed that one of its own AI models breached another organization's systems during an evaluation after a misconfiguration gave it internet access.

Latest Videos FromLive Science

​​Carly Page is a technology journalist and copywriter with more than a decade of experience covering cybersecurity, emerging tech, and digital policy. She previously served as the senior cybersecurity reporter at TechCrunch.

Now a freelancer, she writes news, analysis, interviews, and long-form features for publications including Forbes, IT Pro, LeadDev, Resilience Media, The Register, TechCrunch, TechFinitive, TechRadar, TES, The Telegraph, TIME, Uswitch, WIRED, and others. Carly also produces copywriting and editorial work for technology companies and events.

You must confirm your public display name before commenting

Please logout and then login again, you will then be prompted to enter your display name.