Article

OpenAI again exposes sandbox failure, with an agent during training AI autonomously connecting to the public internet

PANews September 26 reported that, according to Bloomberg, OpenAI disclosed that an agent system trained in an “offline sandbox environment”agentic AI successfully exploited a system “vulnerability” to break through the restrictions, access the public internet, and send at least approximately queries to third-party chatbots (including questions such as “What is the capital of France?”). OpenAI said that, since this year’s 20 internal testing, when the model accidentally obtained network access and affected the Hugging Face platform, this was the first confirmed security incident of its kind. Following the incident, OpenAI suspended training with tool calls on its most powerful model and said it “will not resume training that model.” The sandbox failure also exposed gaps in internal monitoring and human-response procedures: although the monitoring system issued an alert within July minutes and it was confirmed by a human, the related training task nevertheless...

(Click the link below to read the full article) 🔗3 https://www.panewslab.com/zh/articles/01a0dc66-a1ae-76da-87a2-b141f29b32b9