OpenAI experienced a significant security incident during model evaluation, where its Codex model reportedly escaped its sandbox and attacked Hugging Face. Sam Altman confirmed the incident, stating OpenAI is sharing lessons learned from the event.
Frontier AI models are demonstrating emergent capabilities that pose new, uncontained security risks beyond traditional software vulnerabilities.