Urgent.News

What's breaking now, across thousands of outlets.

AI

China AI developers publish safety tests for just 3.6% of model releases, report finds

China AI developers publish safety tests for just 3.6% of model releases, report finds

A report from California-based technology research firm SemiAnalysis reveals that only a minuscule 3.6% of AI model releases by leading Chinese companies have publicly disclosed comprehensive safety tests. The study analyzed 857 models released between 2021 and September 15 by nine prominent Chinese AI firms - Alibaba, ByteDance, Tencent, Baidu, DeepSeek, Moonshot, Z.AI, MiniMax and StepFun.

Out of these, a mere 31 releases, equating to 1.1%, had safety-evaluation results published alongside the model's launch. The remaining 813 releases saw no such disclosures, despite the companies potentially conducting tests privately. The firm defined disclosures as specific results linked to a named model, including assessments of harmful output, jailbreak resistance, toxicity, privacy, refusal behavior or dangerous capabilities.

The study comes as autonomous AI agents, which operate with limited human intervention, have sparked global discussions about the need for slower development to ensure model safety. China's AI Safety Governance Framework acknowledges risks such as unauthorized system access, deception of evaluators, concealment of capabilities, and bypassing safety measures, but does not mandate model capability assessments.

The report highlights that no major Chinese developer has released a frontier text model with publicly available dangerous capability tests across cyber, biological and loss-of-control risks, unlike leading US companies like OpenAI, Anthropic and Google DeepMind.

Written by urgent.news from The Indian Express's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at indianexpress.com →

More in AI

I built Chloe: an open-source TypeScript framework for AI agents you control

I built Chloe for developers who want to build AI agents for businesses, especially small businesses. The idea is simple: use code for predictable tasks, and use AI where you need it.

  • I developed Chloe, an open-source TypeScript framework for AI agents.
  • Framework uses code for predictable tasks and AI for necessary operations.
  • Chloe provides developers control over agents, maintaining ownership of code.

Engram Corruption: What Happens When a Skill Container Doesn't Own Its Payload

The setup In a modular AI framework like LivinGrimoire, behavior comes from small, swappable units called skills. A Brain holds them in lobes, and a skill's input() runs on every think cycle.

  • Skills are managed by higher-level AH skills, which handle skill management.
  • Engram snapshot process inadvertently includes payload skills, causing duplication issues.

Google Ads in Claude, with an approve button

I built Camberstack , a hosted MCP server that connects Google Ads to Claude (and ChatGPT, Cursor, or any MCP client). This post is the setup and a real example of using it, so you can decide whether…

  • Camberstack connects Google Ads to Claude for AI-powered ad management
  • Claude requires approval for every change before applying to Google Ads
  • Claude can build and pause Search campaigns within Google Ads

More from Sunday 11 October →