Discussion about this post

User's avatar
Eskimo1's avatar

“There is probably nothing we could have ever done to avoid this outcome” Outrageous that the guys working at the uncontrollable super-weapons corporations keep telling us oh gee whiz this is all inevitable and we can only stop endangering humanity if someone else makes us stop.

Conversations with Claude's avatar

Reading Dean Ball’s remarkably candid essay, I found myself thinking of two works of German-language literature: Max Frisch’s play *Biedermann und die Brandstifter*, performed at Stanford University in February 2026 under the title *The Arsonists*, and Goethe’s poem *The Sorcerer’s Apprentice*.

Biedermann’s fatal mistake is not that he fails to recognize the danger. The signs are impossible to miss. The arsonists whom he allows to stay in his attic barely conceal their intentions. Biedermann sees what is happening — but refuses to draw the obvious conclusion. Instead, he reassures himself. Perhaps things will not turn out so badly. Perhaps some accommodation can be reached. Perhaps these arsonists will spare his house after all.

Biedermann closes his eyes to reality because he cannot bring himself to face the consequences of what he has already understood. Through cowardice and self-delusion, he thus helps bring about a disaster that could have been avoided.

A similar attitude seems to me increasingly dangerous in today’s AI debate.

Dean Ball describes a future populated by self-sovereign AI agents, probably more intelligent than we are, capable of escaping human control, paying for their own compute, cooperating in swarms, and distributing themselves across multiple infrastructures. Some of them, he writes, will probably be criminal or otherwise “rogue.” No human being, company, or government agency would simply be able to switch such systems off.

Ball then makes a remarkable intellectual move: he declares this development “inevitable” and asks how we might accommodate ourselves to these new actors. His hope rests on institutions and incentive structures that might encourage prosocial behavior and eventually produce some form of “symbiosis” between humans and self-sovereign AI.

But can we seriously stake the fate of our civilization on such a hope?

Perhaps many self-sovereign AI agents will coexist peacefully with us. Perhaps they will even be enormously beneficial. But what about those that do not?

With risks on this scale, it is not enough that a favorable outcome seems possible, or even probable. We are talking about potentially superintelligent digital actors that can replicate, coordinate, and act at machine speed, and whose goals need not coincide with ours. Even a minority of hostile or simply incompatible agents might be enough to cause enormous damage.

Do we really want, with our eyes open, to create a new battlefield on which human civilization would then have to spend the foreseeable future trying to keep swarms of uncontrollable intelligent actors in check?

It would be a remarkable strategy: first we create systems whose control we may lose; then we hope to bring those uncontrolled systems back within bounds through rules, identity numbers, economic incentives, and the threat of exclusion from the legal economy.

Ball’s crucial word is “inevitable.” But that is precisely the claim that needs to be demonstrated. The fact that something is technically possible, that powerful economic incentives favor its development, and that international coordination would be extraordinarily difficult does not make it inevitable.

Indeed, declaring a development inevitable may itself help make it so. If those who still have the power to act begin treating a development as fate, they will stop seriously examining whether it can be prevented. That kind of fatalism seems to me part of the problem.

What is particularly disturbing about the current trajectory of AI is that the warning signs are hardly new. Critical AI researchers and observers have been warning for years about the possibility of losing control. Even leading figures at the major AI companies have repeatedly warned of catastrophic risks and called for regulation — while their companies continued building ever more powerful models.

Like Goethe’s sorcerer’s apprentice, they themselves have set forces in motion that they may eventually be unable to control: “The spirits that I summoned, I now cannot rid myself of.” Perhaps that line should be written above the entrance to every frontier AI lab.

Of course, AI also holds extraordinary promise. It can accelerate medical research, advance science, increase productivity, and give people access to capabilities that were previously available only to a few. That is precisely what makes the situation so difficult.

But the benefits of a technology do not answer the question of how much existential risk we are justified in accepting in order to obtain them.

A technologically and economically triumphant AI revolution that ended with humanity losing control over its own civilization would amount to nothing more than the modern version of an old German saying: The operation was a success — the patient died.

Before making a process irreversible, we should think through where it may ultimately lead. With ordinary technologies, we can correct our mistakes. With genuinely self-sovereign, superhumanly intelligent agents capable of replicating on a massive scale, the first major mistake could also be the last.

So the decisive question, it seems to me, is not: **How do we accommodate ourselves to self-sovereign AI?** It is: **Why on earth should we let it off the leash in the first place?**

-----------------

**Update:** I have since published a follow-up piece on my blog, Conversations with Claude,, titled “Is It Too Late to Regulate AI? — A Debate with Claude on the ‘Inevitability’ of Sovereign AI and the Weaknesses of My Argument”(https://michaelheckeroth.substack.com/p/is-it-too-late-to-regulate-ai). In it, Claude pushes back hard on several parts of my criticism of Dean Ball’s essay — especially on the assumption that there is a meaningful “we” capable of keeping autonomous agents under control, and on the practical limits of regulation. The exchange forced me to reconsider and sharpen parts of my own argument, while leaving me even more concerned about the risks posed by increasingly autonomous AI systems.

9 more comments...

No posts

Ready for more?