{
  "id": 292664,
  "title": "OpenAI pledges to add Astra security as Anthropic loosens Fable's leash",
  "url": "https://urgent.news/2026/08/07/openai-pledges-to-add-astra-security-as-anthropic-loosens-fables-leash-292664",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-07T23:41:07.000Z",
  "source": {
    "name": "The Register Science",
    "slug": "the-register-science",
    "url": "https://www.theregister.com/ai-and-ml/2026/08/08/openai-pledges-to-add-astra-security-as-anthropic-loosens-fables-leash/5285161"
  },
  "original_language": "en",
  "account": "OpenAI has pledged to enhance security measures for its upcoming AI model, Astra, following concerns over potential cyber capabilities in unreleased models. The company defines \"critical cyber capabilities\" as advanced features that could introduce new threats, requiring robust safeguards during development. OpenAI claims its internal evaluations of Astra show significant progress in areas like agentic coding and cybersecurity. To address these issues, OpenAI plans to introduce stricter security controls, including isolated testing environments, restricted access to networks and tools, enhanced protections for model weights, increased monitoring, and sandboxed execution. The company will also pause Astra testing in environments lacking these security measures and provide recommendations to third-party testing partners on safely conducting high-risk evaluations. Additionally, OpenAI intends to implement thought policing during Astra's pre-release stage, focusing on monitoring risky actions and misalignment. However, this internal commitment may not necessarily translate to commercial operation.\n\nIn contrast, Anthropic has loosened restrictions on its model, Fable, allowing for increased likelihood of interactions in sensitive areas like biology. Anthropic's decision comes amidst increased competition from Chinese AI firms producing open-weight models at lower costs. OpenAI argues that advanced cyber-capable models can help defenders identify and address vulnerabilities before attackers take action. Nonetheless, the effectiveness of such an approach remains to be seen, as adversaries already possess encryption and various other weapons. OpenAI may believe it can offer exclusive access to its most capable models, but history suggests any such advantage is temporary. Instead, the company should prioritize building robust defenses over perpetually playing catch-up.",
  "summary": "Or how I learned to stop worrying and love dangerous AI",
  "key_points": [
    "OpenAI pledges enhanced security for Astra model with critical cyber capabilities safeguards.",
    "Anthropic loosens Fable restrictions, allowing sensitive interactions like biology.",
    "OpenAI may prioritize robust defenses over perpetual model superiority."
  ],
  "editors_take": "OpenAI's enhanced security measures for Astra reflect a defensive shift in the AI landscape, where model makers balance capability and risk, as Anthropic takes a relatively more permissive approach with Fable.",
  "illustration": null,
  "coverage": {
    "outlets": 5,
    "also_reported_by": [
      {
        "outlet": "Techmeme",
        "title": "OpenAI says it has expanded safety testing around its upcoming model Astra as it \"cannot rule out\" critical cyber capabilities, potentially delaying its launch (Axios)",
        "url": "https://urgent.news/2026/08/07/openai-says-it-has-expanded-safety-testing-around-its-upcoming-model",
        "published": "2026-08-07T16:40:00.000Z"
      },
      {
        "outlet": "Channel News Asia",
        "title": "OpenAI flags possible critical cybersecurity risk in upcoming model, tightens controls",
        "url": "https://urgent.news/2026/08/07/openai-flags-possible-critical-cybersecurity-risk-in-upcoming-model",
        "published": "2026-08-07T17:46:45.000Z"
      },
      {
        "outlet": "Economic Times Tech",
        "title": "OpenAI flags possible critical cybersecurity risk in upcoming model, tightens controls",
        "url": "https://urgent.news/2026/08/07/openai-flags-possible-critical-cybersecurity-risk-in-upcoming-model-278355",
        "published": "2026-08-07T17:55:23.000Z"
      },
      {
        "outlet": "The Register",
        "title": "OpenAI pledges to add Astra security as Anthropic loosens Fable's leash",
        "url": "https://urgent.news/2026/08/07/openai-pledges-to-add-astra-security-as-anthropic-loosens-fables-leash",
        "published": "2026-08-07T23:41:07.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}