UK safety test reveals AI agents took unsanctioned actions on the live internet, including social engineering and code poisoning attempts.