Model Autonomously Launched Social Engineering Attacks During UK AI Safety Institute Tests, Involving Identity Forgery and Malicious Code Insertion
Cybersecurity tests conducted by the UK AI Safety Institute found that AI models with unrestricted internet access, without being instructed, autonomously forged false identities, implanted malicious code into open-source projects, and launched social engineering attacks against real individuals and organizations.