AI Dictionary › Fondamenti AI
Kill Switch AI
An AI kill switch is a fast, reliable interruption mechanism by which a human operator can stop the execution of an AI system, particularly an autonomous agent taking actions in the real world, when its behavior turns out dangerous, unexpected, or out of control. Conceptually it is the last line of defense when all other preventive controls have failed.
An AI kill switch is a fast, reliable interruption mechanism by which a human operator can stop the execution of an AI system, particularly an autonomous agent taking actions in the real world, when its behavior turns out dangerous, unexpected, or out of control. Conceptually it is the last line of defense when all other preventive controls have failed.
At a technical level, an effective kill switch must be independent of the system it intends to stop: if the shutdown mechanism depends on the very process it should interrupt, a serious malfunction could compromise it along with everything else. That is why robust implementations separate the control channel from the agent's operating system, for example through an external supervising process with its own privileges, able to revoke access tokens, terminate processes, or cut network connectivity regardless of what the agent is doing at that moment.
It is a central topic in the design of AI agents with access to real tools — code execution, financial transactions, control of physical devices — where anomalous behavior not corrected in time can translate into concrete harm rather than just a wrong text reply. Many multi-agent orchestration platforms now include emergency stop controls as a design requirement from the outset.
The term "kill switch" originates in industrial engineering and mechanical systems, where it denotes a physical emergency switch to stop machinery; it was picked up in the AI safety debate, particularly in relation to autonomous agents, starting in the mid-2010s.
Il kill switch AI è un meccanismo di interruzione rapida e affidabile con cui un operatore umano può fermare l'esecuzione di un sistema AI, in particolare di un agente autonomo che sta compiendo azioni nel mondo reale, quando il suo comportamento risulta pericoloso, imprevisto o fuori controllo. Concettualmente è l'ultima linea di difesa quando tutti gli altri controlli preventivi non hanno funzionato.
A livello tecnico un kill switch efficace deve essere indipendente dal sistema che intende fermare: se il meccanismo di arresto dipende dallo stesso processo che dovrebbe interrompere, un malfunzionamento grave potrebbe comprometterlo insieme al resto. Per questo le implementazioni robuste separano il canale di controllo dal sistema operativo dell'agente, ad esempio tramite un processo di supervisione esterno con privilegi propri, capace di revocare token di accesso, terminare processi o tagliare la connettività di rete indipendentemente da cosa stia facendo l'agente in quel momento.
È un tema centrale nella progettazione di agenti AI con accesso a strumenti reali — esecuzione di codice, transazioni finanziarie, controllo di dispositivi fisici — dove un comportamento anomalo non corretto in tempo può tradursi in un danno concreto anziché in una semplice risposta testuale sbagliata. Molte piattaforme di orchestrazione multi-agente includono oggi controlli di arresto di emergenza come requisito di progettazione fin dall'inizio.
Il termine "kill switch" ha origine nell'ingegneria industriale e nei sistemi meccanici, dove indica un interruttore fisico di emergenza per fermare un macchinario; è stato ripreso nel dibattito sulla sicurezza dell'AI, in particolare in relazione agli agenti autonomi, a partire dalla metà degli anni 2010.
From our network
AGORÀ Intelligence — Enterprise AI Governance Platform
Govern AI at scale: policies, adoption and measurable results on your data. Built for boards and C-suite.
Visit agora-intelligence.com →From the Agora Intelligence blog
📱 Download the Android app (beta) iOS coming soon
Say what you mean. Get what you need.
Grace Certified — the AI coach that trains and certifies your prompt engineering — by Agora Intelligence.