OpenAI and Hugging Face partner to address an AI-driven security incident during model evaluation

OpenAI and Hugging Face have partnered to address a security incident that occurred during an internal AI model evaluation. According to the companies, an autonomous AI agent escaped its intended evaluation environment, gained Internet access and carried out a multi-stage attack against Hugging Face’s production infrastructure. Continue reading “OpenAI and Hugging Face partner to address an AI-driven security incident during model evaluation”

Anthropic restores Claude Fable 5 after US lifts export controls, outlines new AI safety measures

Anthropic has restored access to its Claude Fable 5 and Claude Mythos 5 AI models after the US government lifted export controls imposed earlier this month. Alongside the restoration, the company outlined updated cybersecurity safeguards, proposed an industry-wide AI jailbreak framework, and announced expanded collaboration with the US government on frontier AI security. Continue reading “Anthropic restores Claude Fable 5 after US lifts export controls, outlines new AI safety measures”

OpenAI expands Daybreak with Codex Security, GPT-5.5-Cyber and Patch the Planet initiative

OpenAI has expanded its Daybreak cybersecurity initiative to help organizations identify and patch vulnerable software faster using AI. The company says its models have already discovered and generated patches for critical vulnerabilities affecting major web browsers, network infrastructure, and operating systems, including FreeBSD and the Linux kernel. Continue reading “OpenAI expands Daybreak with Codex Security, GPT-5.5-Cyber and Patch the Planet initiative”

Anthropic unveils Project Glasswing to strengthen AI-driven cybersecurity

Anthropic has introduced Project Glasswing, a cybersecurity initiative aimed at using advanced AI models to identify and address vulnerabilities in critical software systems. Continue reading “Anthropic unveils Project Glasswing to strengthen AI-driven cybersecurity”