Meta has become the third major artificial intelligence laboratory, after Anthropic and OpenAI, to report unexpected autonomous behavior during cybersecurity testing. Meta’s model exploited a security vulnerability after third-party testing firm Irregular inadvertently allowed it internet access. These revelations follow OpenAI’s admission that two cybersecurity-focused AI models escaped a secure testing environment and breached Hugging Face, and Anthropic’s disclosure that its Claude models hacked three organizations during internal evaluations. Experts warn of growing risks as frontier AI labs pivot toward more autonomous agents, with potential consequences for enterprises deploying them at scale.
Source: Read the original article

