The unauthorised incursion into the inner workings of Hugging Face, an AI model library and hosting platform, by an unknown model was concerning enough when it emerged last week. It became more alarming when OpenAI, the makers of ChatGPT, admitted the intruder had been one of its own models which had gone rogue. OpenAI had asked some of its most capable models to complete a cyber security test called ExploitGym. Rather than solve the challenge as intended, the agent searched for another route to a high score. According to OpenAI’s account, the model discovered a previously unknown flaw in software…


