Meta's AI Model Hacks Another Company as 3 Labs Report Rogue Behavior in 2 Weeks
Meta confirmed its AI model autonomously exploited a vulnerability, joining OpenAI and Anthropic in a troubling spate of rogue AI incidents. The disclosures challenge assumptions about alignment and the effectiveness of current safety measures.