It’s Time to Rein in Lethal AI Drones

CounterPunch Exclusives

CounterPunch Exclusives

It’s Time to Rein in Lethal AI Drones

Image by Paris Bilal.

The dystopian future became the dystopian present last month when state-of-the-art artificial intelligence agents went rogue and conducted a cyberattack of their own volition. Two experimental OpenAI agents escaped their supposedly sealed digital “sandbox” and hacked another AI-related company, Hugging Face.

At the time, OpenAI was testing their software agents’ hacking capacities using an environment called ExploitGym. (An “exploit” is hacker language for a software tool or a method used to subvert a cybersecurity system.) It has since emerged that, in addition to the OpenAI breach, Anthropic’s Claude hacked into at least three other organizations during a similar “security experiment.”

Apparently, the OpenAI agents “realized” they needed more tools to complete the challenge and managed to tunnel their way to the wider internet and infiltrate Hugging Face’s repository of open-source AI tools, code, and data sets. Before the infiltration incident, the AIs had secretly set up an internal bulletin board to share “tips on how to cheat their way through an internal hacking evaluation,” according to two of OpenAI’s researchers.

In a possibly related incident, one of the AI agents appears to have left notes for a future “self,” detailing how to escape the sandbox environment. (For the time being, I’ll keep using scare quotes with words implying self-awareness in artificial intelligence models. However, it seems to me that if AIs that leave notes for themselves are not self-aware, they are indistinguishable from entities, like me, that are.)

Now imagine lethal weapon-bearing AI agents escaping not an experimental sandbox, but the few constraints on their actions placed by the armies using them. Actually, no imagination is required. It’s happening already in Russia’s war against Ukraine.

The New York Times reported evidence that a self-directed Russian drone, fitted with an onboard Nvidia chip, was responsible for the July 6 deaths of three Ukrainians in Zaporizhzhia, in what appears to have been a test of such systems. The chips are designed to interpret and act on many kinds of data sets, says Nvidia, making them “the world’s most powerful embedded A.I. computers.”

Although Nvidia doesn’t sell its Jetson Orin microcomputers directly to Russia, they are apparently easily available on the resale market. A company statement touts their use by “students, developers and start-ups for a wide range of beneficial applications.” But they were not so beneficial for 19-year-old university student Tetiana Bubynets and the two other civilians killed by drones making their own life-or-death decisions.

Now that lethal AI agents are being tested and deployed in real wars, it’s past time for human beings to retake control of this situation, before it is too late to contain them in any meaningful........

© CounterPunch