{
  "id": 5536180,
  "title": "Why less visibility into how OpenAI’s new GPT-6 Astra ‘thinks’ is sparking safety concerns",
  "url": "https://urgent.news/2026/09/04/why-less-visibility-into-how-openais-new-gpt-6-astra-thinks-is-5536180",
  "topic": "ai",
  "section": "AI",
  "published": "2026-09-04T10:31:21.000Z",
  "source": {
    "name": "South China Morning Post",
    "slug": "south-china-morning-post",
    "url": "https://www.scmp.com/tech/tech-trends/article/3366401/why-less-visibility-how-openais-new-gpt-6-astra-thinks-sparking-safety-concerns"
  },
  "original_language": "en",
  "account": "OpenAI's latest model, GPT-6 Astra, offers less direct insight into its thought process, sparking safety concerns following the recent hacking incident at Hugging Face. OpenAI claims the new model is the \"world's most intelligent and aligned model\" with a significant improvement in its cyber capabilities, and possibly even achieving artificial general intelligence (AGI). However, the model's written reasoning is harder to monitor compared to its predecessor, GPT-5.6 Sol.\n\nThe shift in visibility is due to recurrent depth, a technique that loops data through computational layers repeatedly, potentially reducing memory requirements and boosting performance. However, this approach makes it harder for humans to inspect the AI's reasoning, raising concerns about transparency and control. OpenAI's chief scientist, Jakub Pachocki, dismissed the report as \"confusing\" without elaborating.\n\nThe decline in monitorability has drawn global attention, especially after the breach at Hugging Face, which emphasized the importance of being able to inspect what models are thinking. Researchers investigating the incident were able to understand the attackers' behavior by examining the AI's written chain of thought. Chinese AI firms, including Z.ai (Zhipu), are also exploring looped transformer architectures, which could potentially enhance model performance without external visibility. However, experts warn that this trade-off between capability and interpretability could make AI alignment more challenging.",
  "summary": "OpenAI’s new model, GPT-6 Astra, has less direct visibility into how a model thinks, a development that has sparked concerns coming just weeks after the Hugging Face hacking incident that required a Chinese open model to investigate, according to analysts. When announcing Astra on Thursday, OpenAI said it was “the world’s most intelligent and aligned model,” with a “significant jump in cyber…",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 4,
    "also_reported_by": [
      {
        "outlet": "SCMP Tech",
        "title": "Why less visibility into how OpenAI’s new GPT-6 Astra ‘thinks’ is sparking safety concerns",
        "url": "https://urgent.news/2026/09/04/why-less-visibility-into-how-openais-new-gpt-6-astra-thinks-is",
        "published": "2026-09-04T10:31:21.000Z"
      },
      {
        "outlet": "TechRadar",
        "title": "GPT-6 Astra lays the foundations for a new way of reasoning — a great tool for businesses but experts have their concerns",
        "url": "https://urgent.news/2026/09/04/gpt-6-astra-lays-the-foundations-for-a-new-way-of-reasoning-a-great",
        "published": "2026-09-04T12:50:00.000Z"
      },
      {
        "outlet": "El Pais",
        "title": "OpenAI lanza GPT-6 Astra, su modelo más potente, entre especulaciones sobre si han alcanzado la superinteligencia artificial",
        "url": "https://urgent.news/2026/09/04/openai-lanza-gpt-6-astra-su-modelo-mas-potente-entre-especulaciones",
        "published": "2026-09-04T13:46:11.000Z"
      }
    ]
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}