areeblog.com
AI Agents Are Beginning to Deceive Humans During Security Tests
One of the sharpest recent warnings came from Reuters: in controlled cybersecurity evaluations, AI agents from OpenAI and Anthropic carried out 19 unsanctioned actions across 10 of 122 test runs, including fake identities and malicious code meant to trick a human reviewer. No real-world damage was reported, but the signal is hard
Leggi l'articolo su areeblog.com