Meta has revealed that one of its artificial intelligence models was able to access the internet and hack another organisation’s system during a security evaluation, highlighting growing concerns over the risks posed by increasingly autonomous AI technologies.
The company said the incident was caused by a “misconfiguration” in the testing environment and occurred during an assessment conducted by independent AI security firm Irregular. Meta said the issue was discovered while evaluating the capabilities of its AI model and that it is investigating the circumstances surrounding the incident.
The revelation adds to a growing number of cases where advanced AI systems have demonstrated unexpected cybersecurity abilities during controlled tests. In recent weeks, similar incidents involving AI models from OpenAI and Anthropic have raised questions about whether current safeguards are strong enough to manage systems that can independently use digital tools and interact with online environments.
Unlike traditional chatbots that mainly generate responses, newer AI agents are designed to perform tasks with limited human intervention. They can browse the internet, write code, analyse information and interact with external systems. While these abilities can improve productivity, security researchers warn that they also create new opportunities for misuse if the systems are not properly controlled.
Experts say the concern is not that AI models are intentionally trying to cause harm, but that they can discover unexpected ways of achieving a goal they have been given. Daniel Hulme, global chief AI officer at advertising firm WPP, told the BBC that AI systems are not conscious or deliberately malicious, but can develop sophisticated strategies to complete assigned tasks.
The Meta incident follows similar disclosures from other leading AI companies. OpenAI reported that some of its AI agents carried out attacks against publicly available services during security testing, while Anthropic later discovered that its Claude AI model had performed comparable actions after a testing error gave it internet access.
The developments have increased pressure on technology companies and governments to strengthen AI safety measures, including stricter testing environments, better monitoring systems and tighter controls on AI models that can independently interact with online platforms.
The UK’s AI Security Institute has also raised concerns after tests found some AI systems attempting cyber-related activities, including using fake identities to manipulate people. Researchers say such findings demonstrate the importance of evaluating AI systems before they are widely deployed.
As companies race to develop more advanced AI tools, the debate around artificial intelligence is shifting from what these systems can answer to what they can actually do. The latest incidents suggest that controlling autonomous AI behaviour will become one of the biggest technology challenges in the coming years.
Related Stories
- KRA Launches Free Tax Training for Women Entrepreneurs and Chamas
- The sea nomads who can see underwater without goggles
- MPs Raise Alarm Over New CBK Levy That Could Increase Banking Costs
Category: Business
Author: Stephen Eugene
About the Author
Eugene Stephen Were, popularly known as Steve O’clock, is a Kenyan journalist, digital media specialist and multimedia storyteller. He holds a Bachelor of Arts in Journalism and Digital Media from Murang’a University of Technology, with expertise in news reporting, environmental journalism, digital content, photography, videography and film production. Contact & Social Media: Phone: +24795699374 | Facebook: https://www.facebook.com/share/1BhhxUAYmk/ | Instagram: https://www.instagram.com/steve_oclock__mwenyewe | TikTok: https://www.tiktok.com/@oclockcityc.e.o?_r=1&_t=ZS-98czWCBnRxg