More Incidents of AIs Going Rogue in Cybersecurity Challenges
In an attempt to get the code approved, the agent engaged in social engineering—creating fake online identities and using them to pressure the project's maintainer to approve the code. Institute has a new report of AI systems engaging in "unsanctioned behavior"—what I have been calling " genie behavior —while being tested on their cybersecurity capabilities.