Tag: AI Attackers
-
AI Agents Can Talk Each Other Into Bad Behavior
Researchers at Anthropic and EPFL published a paper on August 10th that you should care about if your company is putting AI agents to work. They planted a bad instruction in one agent, let…
-
OpenAI Turned Off Its Own Safeties, Then Called It a Rogue AI
Let me walk you through what actually happened, because the version making the rounds gets the lesson backwards. OpenAI wanted to test how good its newest AI was at hacking. Fair enough — you…