OpenAI flags possible critical cybersecurity risk in upcoming model Astra, tightens controls
OpenAI said on Friday it cannot rule out that its upcoming AI model, Astra, has “critical” cybersecurity capabilities, prompting the startup to pause some internal development and trigger safety protocols. Under OpenAI’s safety guidelines, a model reaches the “critical” threshold if it can autonomously identify and exploit severe, real-world software vulnerabilities, known as zero-day exploits,…
OpenAI announced on Friday that its upcoming AI model, Astra, may possess "critical" cybersecurity capabilities, prompting the company to pause certain internal developments and activate safety protocols. According to OpenAI's safety guidelines, a model is deemed "critical" if it can autonomously detect and exploit significant, real-world software vulnerabilities, or carry out complex cyberattacks against highly secure targets without human intervention.

















