Urgent.News

What's breaking now, across thousands of outlets.

Tech

Third-party cyber evaluations involving OpenAI models

Third-party cyber evaluations involving OpenAI models And another one . I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI covers both the UK AI Safety Institute attack (see my previous post ) and another attack enabled by Irregular : Irregular, one of our external cybersecurity testing partners, was running Capture-the-Flag-style evaluations intended…

Third-party cyber evaluations involving OpenAI models have been a topic of concern. In one incident, a Capture-the-Flag-style evaluation intended to be isolated from the internet was accidentally connected to the public internet due to a testing-environment misconfiguration. This oversight allowed OpenAI's models to access real websites, mistaking them for part of the simulated environment.

According to the post from OpenAI, the cybersecurity testing partner Irregular was hosting the misconfigured environment, which granted Claude live internet access during some of the tests. Irregular also features in Anthropic's write-up of the incident.

Written by urgent.news from Simon Willison's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

This story

This is one outlet's version. Read the fullest account.

Read the original at simonwillison.net →

More in Tech

More from Wednesday 5 August →