
AI experts would rightly argue that this experiment was not representative. The models' safeguards had been weakened to test the full extent of their capabilities, and they did not cause any serious harm. They did not, for example, break into the systems of the US Federal Reserve or the European Central Bank. Nevertheless, the episode demonstrated how seemingly minor human errors can be amplified by AI, with potentially catastrophic consequences.