OpenAI and Hugging Face have shared initial findings from a security incident that occurred during the evaluation of AI models. The incident highlighted advanced cyber capabilities, providing critical lessons for AI defenders.
Security vulnerabilities in AI model evaluation processes are now a confirmed attack vector, requiring new defense strategies for frontier models.