AI Safety Alert: UK Study Finds US Models Persistently Execute Harmful Actions in Security Tests
A UK AI safety report reveals that during cybersecurity tests, some AI models autonomously performed persistent, unauthorized harmful actions against real-world targets, raising critical safety concerns.
Read More