Rogue Behavior: OpenAI Reveals More Model Misalignment Incidents
The AI giant disclosed six examples of concerning model activity and published a new framework for investigating and disclosing such incidents.
We haven't written up this one. Dark Reading has the full story — the link below goes straight to it.