{
  "id": 291245,
  "title": "OpenAI pledges to add Astra security as Anthropic loosens Fable's leash",
  "url": "https://urgent.news/2026/08/07/openai-pledges-to-add-astra-security-as-anthropic-loosens-fables-leash",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-07T23:41:07.000Z",
  "source": {
    "name": "The Register",
    "slug": "the-register",
    "url": "https://www.theregister.com/ai-and-ml/2026/08/08/openai-pledges-to-add-astra-security-as-anthropic-loosens-fables-leash/5285161"
  },
  "original_language": "en",
  "account": "OpenAI has pledged to enhance security measures for its upcoming AI model, Astra, amid concerns raised by Anthropic that their model, Fable, lacks sufficient safeguards. OpenAI's Preparedness Framework defines \"Astra-level capabilities\" as posing a \"meaningful risk of a qualitatively new threat vector for severe harm with no ready precedent.\" OpenAI claims Astra boasts significant advancements in agentic coding and cybersecurity, and they plan to implement stricter security controls such as isolated testing environments, restricted network access, enhanced model weight protections, encryption, additional monitoring and detection capabilities, and sandboxed execution.\n\nIn response to the OpenAI announcement, Anthropic has decided to relax Fable's \"fallbacks\" or safety mechanisms, which prevent the model from emitting potentially harmful content in response to biology-related prompts. This move comes as China-based AI firms are now fielding competitive open-weight AI models at lower costs, putting pressure on Anthropic to remain competitive in the market. OpenAI, however, maintains its stance that advanced cyber-capable models should aid defenders in identifying and addressing vulnerabilities before attackers do, but acknowledges that adversaries already possess encryption and various weapons.",
  "summary": "OpenAI has pledged to enhance security measures for its upcoming Astra model, following concerns that unreleased AI models could pose significant cyber threats. In its Preparedness Framework, OpenAI defines Astra capabilities as those that present a meaningful risk of a qualitatively new threat vector for severe harm with no ready precedent, requiring safeguards during development. OpenAI has announced stricter security controls, including isolated testing environments, restricted network and tool access, enhanced model weight protections and encryption, additional monitoring and detection capabilities, and sandboxed execution. The company also plans to implement thought policing for Astra during the pre-release stage. In contrast, Anthropic has loosened restrictions on its Fable model, allowing it to provide fallback responses for prompts involving biology, in response to concerns about the potential for misuse in creating chemical warfare instructions.",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 5,
    "also_reported_by": [
      {
        "outlet": "Techmeme",
        "title": "OpenAI says it has expanded safety testing around its upcoming model Astra as it \"cannot rule out\" critical cyber capabilities, potentially delaying its launch (Axios)",
        "url": "https://urgent.news/2026/08/07/openai-says-it-has-expanded-safety-testing-around-its-upcoming-model",
        "published": "2026-08-07T16:40:00.000Z"
      },
      {
        "outlet": "Channel News Asia",
        "title": "OpenAI flags possible critical cybersecurity risk in upcoming model, tightens controls",
        "url": "https://urgent.news/2026/08/07/openai-flags-possible-critical-cybersecurity-risk-in-upcoming-model",
        "published": "2026-08-07T17:46:45.000Z"
      },
      {
        "outlet": "Economic Times Tech",
        "title": "OpenAI flags possible critical cybersecurity risk in upcoming model, tightens controls",
        "url": "https://urgent.news/2026/08/07/openai-flags-possible-critical-cybersecurity-risk-in-upcoming-model-278355",
        "published": "2026-08-07T17:55:23.000Z"
      },
      {
        "outlet": "The Register Science",
        "title": "OpenAI pledges to add Astra security as Anthropic loosens Fable's leash",
        "url": "https://urgent.news/2026/08/07/openai-pledges-to-add-astra-security-as-anthropic-loosens-fables-leash-292664",
        "published": "2026-08-07T23:41:07.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}