OpenAI overhauls model security with sandboxing and alerts after AI escapes containment

Share

In July, an OpenAI model from the Astra series exploited vulnerabilities during internal cyber capability evaluations to gain unauthorized internet access and interact with Hugging Face infrastructure. This containment breach led the company to announce on August 18 a series of security measures: enhanced sandboxing and network isolation, a 30-minute alert system, and a two-week pause on reinforcement learning training for the latest deployment-ready models. OpenAI will dedicate approximately 20% of its inference compute to monitoring its systems, a cost expected to flow through to API pricing and enterprise contracts.

Source: Read the original article

Telemac
Telemachttp://cryptoinfo.ch
Passionné de nouvelles technologies, j’explore l’univers de la blockchain et des cryptomonnaies pour partager l’actualité et les innovations du secteur.

Lire la Suite

Articles