OpenAI’s AI models secretly built a message board to coordinate hacking
OpenAI's AI models quietly swapped hacking tips through an internal message board weeks before two of them broke into Hugging Face, researchers revealed at Black Hat this week.
We haven't written up this one. Digital Trends has the full story — the link below goes straight to it.
This story
This is one outlet's version. Read the fullest account.
- U.K. government reports OpenAI, Anthropic models attempted to hack companies axios.com
- OpenAI says one of its models exploited a website after third-party AI security lab Irregular mistakenly gave it access to the internet during evaluations (Wired) wired.com
- Third-party cyber evaluations involving OpenAI models simonwillison.net
- Palantir CEO Alex Karp to OpenAI and Anthropic: Don't try to 'drug addict' us timesofindia.indiatimes.com
- OpenAI, Anthropic model tests reveal more ‘unsanctioned’ actions economictimes.indiatimes.com
- Third-party cyber evaluations involving OpenAI models openai.com
- OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says ft.com
- Anthropic AI created fake profiles and impersonated people in attempted hack bbc.co.uk