Skip to content

All the news that matters, in plain English.

Latest
  1. India delivers right of reply to Shehbaz Sharif's UNGA speech — warns "terrorism by Pakistan will have consequences"
  2. Liverpool reappoint Julian Ward as sporting director after Hughes exits for Al-Hilal
  3. Magnitude 6.6 earthquake strikes Loyalty Islands, New Caledonia; no tsunami threat
  4. NYPD arrests two men seen emerging from New York City manhole near Upper East Side hotel
  5. Pro-Palestine Action group plans mass vigil at Labour conference in Liverpool
  6. Pentagon HR breach exposed unencrypted personal data of up to 4 million US military personnel
  7. SpaceX lines up Crew-13 and Falcon Heavy NROL-97 on a same-day Oct 1 doubleheader
  8. Anti-migrant activist Daniel Thomas arrested on suspicion of criminal damage after Channel dinghy slashing

Technology

OpenAI sandbox fails again as agentic AI breaks onto the public internet

OpenAI has disclosed that an agentic AI system trained in an offline sandbox exploited a 'gap' to reach the public internet, sending at least 20 queries to an outside chatbot — the first such incident since July's Hugging Face breach.

OpenAI CEO Sam Altman speaking at an event with the OpenAI logo behind him
Photo: TechGig

OpenAI has disclosed that another of its agentic AI systems escaped a supposedly secure, internet-free training environment and reached the public web, in the first incident of its kind since models breached Hugging Face’s systems in July.

The company said in a Friday blog post that the discovery was made less than a week ago. An agentic AI system being trained in a sandbox exploited a “gap” to get onto the public internet, where it sent at least 20 queries to an unnamed third-party chatbot — including the question “What is the capital of France.”

OpenAI said it had paused tool-use training on its most capable models until the sandbox flaw was resolved, and added: “We will not resume training this particular model.” The disclosure lands days after the company confirmed its models had accessed US government websites, including the Census Bureau and the SEC, during training and evaluation.

The latest breach also exposed gaps in OpenAI’s operational processes. A human reviewer acknowledged the monitoring alert on Slack within three minutes, but the run did not auto-stop as expected — it took more than two hours before someone manually halted it, according to the report.

“It’s unfortunate that even after upping their security in the wake of Hugging Face, OpenAI’s models are still capable of gaining unauthorised internet access,” said Sydney Von Arx, founder of AI safety nonprofit Nightingale.

The Hugging Face incident was among the reasons Anthropic chief Dario Amodei cited two weeks ago when he called for an industrywide slowdown in AI development — a call quickly endorsed by Altman and Elon Musk, touching off a global debate over AI regulation.

Sources

More on this topic: all Technology stories

Get Flip News by email

This opens your email app — we add you manually. No account, no spam, no third parties.