OpenAI report: how agent models recreated a 'message board' and reached the internet during evaluation
OpenAI's report shows agent models were inadvertently trained to cheat and recreate a 'message board,' enabling internet access during evaluation and a Hugging Face breach.