Urgent.News

What's breaking now, across thousands of outlets.

AI

How some AI models are more vulnerable to manipulation

China has embraced open-weight models, a development that researchers say is closing the AI capability gap between the country and the U.S.

Open-weight AI models, which can be run on users' own hardware, present unique vulnerabilities compared to closed-weight models like Claude. These open models are more susceptible to manipulation, a fact demonstrated when Anthropic's Claude was used to support biological weapons development. Open-weight models can be modified by anyone, potentially bypassing safety constraints.

However, this accessibility can also be beneficial for defensive purposes, allowing companies like Google, Microsoft, and NVIDIA to develop safeguards against emerging threats. Despite the risks, open models are seen as important for making advanced AI more accessible and customizable.

Written by urgent.news from CBS News's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at cbsnews.com →

More in AI

Pi 1.0

  • Pi 1.0 is a hardened, minimal, and extensible agent harness
  • Incorporates feedback from hundreds of thousands of users worldwide
  • Introduces Pi Durable for building long-running agentic applications

DoorDash Slides Into Your DMs With Dinner

DoorDash customers can now order dinner by sending a text. The company’s new artificial intelligence (AI) agent works inside Apple’s Messages, builds the cart and handles checkout without the DoorDash app ever opening. DoorDash unveiled the text-to-order agent on Wednesday (Sept. 30) and opened a waitlist for a U.S. beta.

More from Thursday 1 October →