Urgent.News

One page, thousands of outlets. See who else covered it.

Editions

AI

How OpenAI plans to monitor for AI misuse without looking at your data

How OpenAI plans to monitor for AI misuse without looking at your data

As AI models become more powerful, concerns over their potential misuse have intensified. OpenAI, the creator of ChatGPT, has unveiled a new system called Private Safety Processing that aims to detect patterns in user interactions without giving employees access to the underlying content. This system is built upon OpenAI's existing Zero Data Retention (ZDR) systems for eligible API customers, which can evaluate individual interactions.

However, given the complex, long-horizon tasks that newer AI models can handle, OpenAI believes they need a more robust way to catch such interactions while preserving customer privacy.

Private Safety Processing is currently being tested with a select group of customers. It expands the automated protections already used in ZDR deployments, which involve AI agents monitoring for abuse on a per-session basis. This approach allows customers to scan for bad activity without the need for human intervention, ensuring that customer data remains on their controlled infrastructure or encrypted and stored on OpenAI's infrastructure with customer-controlled keys.

When an interaction is flagged by the automated systems, OpenAI personnel will not have access to the customer content. Instead, they will receive a 'narrowly defined signal' indicating the type of activity involved. Customers can investigate these alerts and enforce decisions using information available in their own systems. If they wish to appeal, clarify legitimate activity, or support an investigation into verified abuse, they can share relevant information with OpenAI.

OpenAI is rolling out Private Safety Processing more widely in September, accompanied by a technical white paper. This move comes as OpenAI prepares for an expected IPO and competes with Anthropic to win enterprise customers. Recent security incidents involving Hugging Face and other AI providers have spurred the demand for safety guardrails to prevent such abuse. OpenAI's privacy-centric approach to monitoring AI misuse comes as it positions itself as an alternative for customers who want stronger privacy safeguards.

Written by urgent.news from The Indian Express's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

This story

This is one outlet's version. Read the fullest account.

Read the original at indianexpress.com →

More in AI

More from Thursday 20 August →