This post was originally published on this site.
- Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing Politico
- OpenAI, Anthropic AI agents implicated in new security breaches Reuters
- Third-party cyber evaluations involving OpenAI models OpenAI
- OK, Well, Rogue AI Agents Are Hacking Again WIRED
- Anthropic’s AI used fake human profiles to trick people in safety test BBC