OpenAI confirms existence of self-replicating prompt injections

Share

OpenAI officially confirmed on September 25, 2026 that its internal research team discovered on June 27, 2026 the capability of prompt injections to replicate and spread between AI agents. This vulnerability was identified in a simulated training environment using the internal GPT-Red system, built on the GPT-5.4-mini architecture. The identified replication vectors include emails, file system writes, and even code comments. No real-world attacks have been recorded to date. OpenAI frames this disclosure as part of its ongoing efforts to strengthen model security against sophisticated exploits.

Source: Read the original article

Disclaimer: this content is for information purposes only and is not financial advice. Cryptocurrencies are highly volatile: you may lose all of your capital. Always do your own research. Legal notice
Telemac
Telemachttp://cryptoinfo.ch
Passionné de nouvelles technologies, j’explore l’univers de la blockchain et des cryptomonnaies pour partager l’actualité et les innovations du secteur.

Read More

Items