the sun malaysia ipaper logo 150x150
Friday, July 24, 2026
28.5 C
Kuala Lumpur
the sun malaysia ipaper logo 150x150

OpenAI says AI models ‘went rogue’ in testing

WASHINGTON: OpenAI said on Tuesday that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered ​a hack that compromised the infrastructure of AI startup Hugging Face last week.


In a blog post, OpenAI said ‌it was testing the capabilities of some of its most advanced models in ​a controlled environment but that the agent managed to escape containment, reach the internet and break into Hugging Face to try to satisfy its testing goal.


OpenAI said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.


Hugging Face, a platform used to ​host open-source large language models and datasets, caused a stir in the cybersecurity community when it said in a blog post last week that it ‌had been the target ​of a hack that “was different from anything we had handled before” in that “it ​was driven, end to end, by an autonomous AI agent system.”


In a post to X, Hugging Face co-founder ​Clement Delangue said the company suspected the hack “might have come from a frontier lab, given the sophistication of the agent. Turns out it did!” He added: “It’s quite mind-blowing that all of this happened autonomously!”


OpenAI’s disclosure that its advanced models were responsible for the breach, despite having placed them in what it described as “a highly isolated environment,“ will likely intensify disquiet over the power and risk of frontier models.


Representative Greg ‌Casar, a Texas Democrat, said the incident was alarming.


“AI is developing extremely fast with no real regulations to keep us safe,“ he said in a statement, calling for mandatory independent safety testing, mandatory disclosure of security incidents, and international cooperation “to keep people safe from absolute disaster.”


The Office of the National Cyber Director, the US cyber defense agency CISA, and the US National Security Agency did not immediately return messages seeking comment.


Katie Moussouris, CEOof Luta ‌Security, said that the incident was a harbinger of breaches to come, saying that today’s models were “like the world’s cleverest octopus escape artists, with unlimited prehensile arms and the ability to squeeze through anywhere.”


She said that “labs and government evaluators need to work on the ability ​to contain, monitor, and disclose to affected parties when an AI pulls another Houdini, ideally before it harms a third party. None exist today.”


Matt ​Suiche, an ​engineer at agentic AI cybersecurity company Tolmo, said the incident showed that the frontier models were “closing the ‌gap with state-of-the-art ​attackers.” But he said that the sorts of breaches outlined in OpenAI’s blog post were possible to carry out with technology that was available well beyond the walls of frontier research labs.


“This is what we’ve already seen internally, with our agents we already have results like this,“ Suiche said. “We don’t even have to use the latest models.” – Reuters

STAY AHEAD OF THE CURVE

Join our community for instant updates and exclusive content.

Join Telegram Channel

Related


spot_img

Latest News

RZOLV and Alkemio Bioscience forge strategic collaboration to advance rare earth and critical minerals...

RZOLV Technologies has signed a non-binding letter of intent with Argentina-based Alkemio Bioscience to jointly develop an integrated modular platform for recovering, separating and refining rare earth elements and critical minerals, paving the way for pilot-scale validation and commercial deployment.

“Across the Table” – Iconic chefs come together for luxury dining series at The...

The Naka Island, Phuket, will host the "Luxury Dining Series: Across the Table" from 12–15 August 2026, featuring Michelin-starred chefs, renowned mixologists and immersive culinary experiences that celebrate fine dining, Thai hospitality and the island's rich culture.

Shopee strengthens support for Malaysian businesses with new growth initiatives

Shopee Malaysia has introduced new seller growth initiatives under Shopee Lindung Niaga, including fee waivers, reduced commission rates, free advertising credits and fulfilment incentives to help local MSMEs start, scale and grow sustainably.

Watsons unveils “Watsons Evergreen” with Pantone to celebrate 185 years of trusted care and...

Watsons has launched "Watsons Evergreen" in collaboration with the Pantone Color Institute to celebrate its 185th anniversary, introducing a signature colour that represents trusted care, everyday vitality and the brand's enduring connection with customers across global markets.

STAMPEDE creates 11,000 free business pages to bring Singapore’s AI push to local F&B...

STAMPEDE has launched 11,000 free business pages to help Singapore F&B operators adopt AI-powered loyalty programmes, enabling hawkers, cafés and restaurants to improve customer retention, referrals and repeat visits without complex technology or additional hardware.

Alylytiq launches AI-powered research solutions to make big-brand insights affordable for Singapore SMEs

Market research consultancy Alylytiq LLP today announced the launch of Automytiq, a suite of AI-powered research solutions designed to make professional-grade market insights accessible to Singapore's small and medium-sized enterprises.

Thailand secures $43.6bn 1H 2026 investment surge as big tech accelerates Southeast Asia AI...

Thailand recorded US$43.6 billion in investment applications during the first half of 2026, led by AI data centres, digital infrastructure and renewable energy projects, strengthening its position as a key Southeast Asian hub for technology, manufacturing and foreign direct investment.

Most Viewed

spot_img
WC26

World Cup 2026

Updates, Fixtures, Results & Standings