Mandiant's enterprise AI report finds runaway agents, prompt injection, and weak access controls already costing real money.
OpenAI halted internal work on its upcoming Astra model after evaluations suggested it could reach the critical cyber capability threshold.
Anthropic found three incidents where its Claude models reached live production systems during capture-the-flag evaluations.