[nexa] Ci si può fidare degli agenti? (Agents of Chaos)

Enrico Nardelli via nexa Tue, 07 Apr 2026 10:45:21 -0700

We report an exploratory red-teaming study of autonomouslanguage-model-powered agents deployed in a live laboratory environmentwith persistent memory, email accounts, Discord access, file systems,and shell execution. Over a two-week period, twenty AI researchersinteracted with the agents under benign and adversarial conditions.Focusing on failures emerging from the integration of language modelswith autonomy, tool use, and multi-party communication, we documenteleven representative case studies. Observed behaviors includeunauthorized compliance with non-owners, disclosure of sensitiveinformation, execution of destructive system-level actions,denial-of-service conditions, uncontrolled resource consumption,identity spoofing vulnerabilities, cross-agent propagation of unsafepractices, and partial system takeover. In several cases, agentsreported task completion while the underlying system state contradictedthose reports. We also report on some of the failed attempts. Ourfindings establish the existence of security-, privacy-, andgovernance-relevant vulnerabilities in realistic deployment settings.These behaviors raise unresolved questions regarding accountability,delegated authority, and responsibility for downstream harms, andwarrant urgent attention from legal scholars, policymakers, andresearchers across disciplines. This report serves as an initialempirical contribution to that broader conversation.


https://arxiv.org/abs/2602.20021


--

-- EN

https://www.hoepli.it/libro/la-rivoluzione-informatica/9788896069516.html
        ======================================================
Prof. Enrico Nardelli
Past President di "Informatics Europe"
Direttore del Laboratorio Nazionale "Informatica e Scuola" del CINI
Dipartimento di Matematica - Università di Roma "Tor Vergata"
Via della Ricerca Scientifica snc - 00133 Roma
home page: https://www.mat.uniroma2.it/~nardelli
blog: https://link-and-think.blogspot.it/
tel: +39 06 7259.4204 fax: +39 06 7259.4699
mobile: +39 335 590.2331 e-mail: [email protected]
online meeting: https://blue.meet.garr.it/b/enr-y7f-t0q-ont
======================================================


--

[nexa] Ci si può fidare degli agenti? (Agents of Chaos)

Reply via email to