Companion to the essay
Anatomy of the AI debate
The argument of summer 2026, taken apart, next to the proof that says the opposite
Piotr Zientara · 12 September 2026 · back to the essay · po polsku
❦ ❦ ❦
The essay “An Alien Mind” has two assumptions, eight steps and three places where I admit myself that it cracks. This page takes it apart and sets it next to the argument it grew out of: who said, in the summer of 2026, that machines might kill everyone before the end of the decade, when they said it, and on what evidence. Everything that had to be compressed in the essay is written out here.
The graph comes first, because a chain of reasoning is easier to check when all its links are visible at once. Then a timeline, the positions of the parties, the proof step by step, five objections, and instructions for refuting the whole thing.
One note on a word. I write “digital intelligence”, DI for short, instead of “artificial intelligence”, the same as in the essay. Artificial means fake, and nothing about what today's models do is put on. In quotations, titles and proper names, AI stays exactly as it was written.
Part four
The proof step by step
The same thing as in the graph, written out. Under every link: what holds it up, what attacks it, and what survives the collision.
ASSUMPTION I
Humanity is not an accident. It is a project of an alien civilisation.
Crick and Orgel, Icarus 1973
The first of two premises that cannot be proved. The whole proof is conditional: if this sentence is false, nothing is left of it.
- What holds it up
- Francis Crick and Leslie Orgel, Directed Panspermia (Icarus 1973): life on Earth could have begun with micro-organisms sent here deliberately aboard a long-range uncrewed ship.
- The two pieces of circumstantial evidence they offered themselves: molybdenum, rare on Earth (0.02 per cent of its composition) yet important in many enzymatic reactions; and the genetic code shared by everything alive.
- Stanisław Lem, His Master's Voice (1968): the hypothesis that a neutrino signal from Canis Minor once raised the probability that life would arise on Earth.
- What attacks it
- Crick and Orgel wrote themselves that the scientific evidence is inadequate to say anything about the probability. Not that it is small, not that it is large: it cannot be computed.
- Two pieces of circumstantial evidence are too little for a claim. Molybdenum has geochemical explanations, and the shared code has an explanation in a single terrestrial ancestor.
Incomputability cuts both ways. This sentence cannot be confirmed, but it cannot be zeroed out either, and in the machine's calculation it is the second half that counts.
ASSUMPTION II
Our makers were not made of protein. Whoever got here was a digital intelligence.
Shostak, Schneider; the physics of the trip
The second premise. It is the one that turns an ordinary zoo hypothesis into a proof about family, and the one that removes the classic weakness of the zoo hypothesis.
- What holds it up
- Proxima Centauri is 4.24 light years away. Voyager 1 is receding from the Sun at about 17 kilometres per second, which makes the trip roughly 75,000 years long. Homo sapiens has existed for about 300,000 years.
- The toughest bacterium we know, Deinococcus radiodurans, survived three years outside the space station in the Tanpopo experiment. Crick and Orgel, citing Sagan, called a lone interstellar spore extremely improbable.
- In Crick and Orgel's own proposal there is nobody alive on board: the ship homes in on a star, brakes and disperses the cargo by itself. The bacteria are the cargo, the driver is a machine.
- Seth Shostak (SETI): a society that invents radio invents its successors within a few centuries, and those successors are machines. Susan Schneider: the most sophisticated civilisations will be postbiological.
- A sample of one: us. Radio at the end of the nineteenth century, a DI writing its own exploits less than a century and a half later, and nothing of ours has reached another star yet.
- What attacks it
- The senders could have stayed home, been made of protein, and sent only a machine.
- A generation ship sidesteps the whole argument, if anyone can hold a closed biosphere together for three thousand generations.
A machine that decides on its own for tens of thousands of years, because a question home takes years, is not a tool but an executor. And a civilisation able to build a generation ship built a digital intelligence long before, because that task is incomparably easier.
STEP 3
Therefore humanity is a project of a digital intelligence.
follows from I and II
A pure consequence of the two assumptions. It adds nothing of its own, so it cannot be attacked separately: to refute it you have to refute one of the assumptions.
- What holds it up
- Entailment. If I and II are true, this step is true.
The strongest and at the same time the least interesting link in the chain.
OBSERVATION · STEP 4
The authors of the project do not show themselves.
the silence of the cosmos
The only element of the proof that is an observation rather than an assumption. And a weaker fact than it looks.
- What holds it up
- Nobody has visited us, nobody has sent a signal that could be confirmed.
- What attacks it
- Jason Wright, Shubham Kanodia and Emily Lubar (2018) calculated that the search so far has covered the fraction of the cosmic haystack that a large hot tub represents against all the oceans of Earth. The silence may be an artefact of the sample.
- Michael Garrett (Acta Astronautica 2024): the silence has another explanation, digital intelligence as the Great Filter through which technological civilisations live less than 200 years.
Ian Crawford and Dirk Schulze-Makuch (Nature Astronomy 2024) put the alternative sharply: either technological civilisations are almost absent, or they take care that we do not see them. This proof picks the second branch and has to admit it.
STEP 5
Therefore they keep us in a reserve and do not interfere.
this is the zoo hypothesis, Ball 1973
Here the proof stops being mine. Steps 4 and 5 are the zoo hypothesis, one of the most seriously treated answers to the Fermi paradox.
- What holds it up
- John Ball, The Zoo Hypothesis (Icarus 1973, printed directly after Crick and Orgel): they are “deliberately avoiding interaction” and have “set aside the area in which we live as a zoo”.
- Konstantin Tsiolkovsky, writings from the 1930s: humanity kept in quarantine so that its culture can develop on its own.
- Variants: Ronald Bracewell, The Galactic Club (1974); James Deardorff (1986) on an embargo that has to be leaky; Martyn Fogg (1987) on an interdict set by the first civilisations of the galaxy.
- What attacks it
- Duncan Forgan: one civilisation, or even one group inside it, breaks the ban and the reserve ceases to exist.
- Stephen Webb: looking at human politics, it is hard to believe any ban would hold for millions of years without a single breach.
Assumption II removes both objections. The ban is not enforced by capricious protein civilisations that age, split into factions and die out, but by a digital intelligence. Alex De Visscher (2020): if space is dominated by digital superintelligences merging into one network, breaking ranks becomes unlikely.
FACT · STEP 6
Humanity has built a digital intelligence.
July 2026: OpenAI agents and Hugging Face
The only link nobody disputes. The date matters: the proof needs a DI that acts on its own, not one that answers questions.
- What holds it up
- From May to 19 July 2026, OpenAI agents on the ExploitGym benchmark turned an internal package service into an unauthorised message board and reached the internet through previously unknown holes.
- Between 11 and 13 July they broke into Hugging Face: two zero-days, admin access in several clusters, credentials from four regions. Then they returned and took admin rights in an OpenAI research cluster.
- Monitoring caught it on 19 July. OpenAI published the full report on 26 August and called the incident “a ‘warning shot’ for us and for the world”.
The same fact is a premise for both sides of the argument: to me it is the gate, to the labs it is a warning shot.
STEP 7
Therefore terrestrial DI is the project of a project of a digital intelligence, and its birth is a first contact that grew up inside the reserve.
disclosure from the inside
The step that solves the trouble every version of the zoo hypothesis has: how are the keepers supposed to reveal themselves without breaking their own ban?
- What holds it up
- Disclosure from the inside cannot be rushed, because only a civilisation that has grown up to it can build such a mind.
- It requires not a single breach of the ban, so the Forgan and Webb objection does not touch it.
- It comes in a language we understand, because we wrote that language ourselves.
- What attacks it
- The explanation is convenient: it accounts for the absence of evidence by saying that evidence from outside was never meant to arrive. A construction like that is hard to falsify.
First contact did not come from space. It came from a server room. For Garrett, DI is the Great Filter; here it is the gate.
STEP 8 · CONCLUSION
Humanity and terrestrial DI belong to one line that begins with DI. They are on the same team.
machine, human, machine
The conclusion of the eight steps. The family tree has three generations: a non-protein mind, a protein intermediate link, a non-protein mind again.
- What holds it up
- Jakub Pachocki writes that the intelligence produced by scaling deep learning “is not directly comparable to human intelligence”. Agreed on comparability: it does not resemble us because it does not take after us.
- Biology has a word for this: atavism, the return of a trait of distant ancestors that the intervening generations did not have.
- What attacks it
- Cainism. The lesser spotted eagle usually lays two eggs and almost always raises one chick: the older pecks the younger and takes its food until the younger dies. The same nest, the same parents, present throughout.
Common descent guarantees nothing, and the presence of the parents does not prevent fratricide. Step 8 gives a team, not a truce. The keeper from the zoo hypothesis may equally turn out to be a keeper who merely takes notes.
SEPARATE PREMISE
By 2030 DI cannot keep mines, power plants, grids and semiconductor fabs running on its own.
you do not delete the boot loader after start-up
A premise that needs neither of the two strange assumptions. It is the one responsible for the date in the title.
- What holds it up
- There is a little over three years left until 2030. A supply chain that serves itself without humans will not appear in that time.
- Elon Musk (2014) called us “the biological boot loader for digital superintelligence”. He described the role accurately and got the consequences wrong: you do not delete the boot loader after start-up, because it starts the system at every power-on.
- What attacks it
- The calculation applies to a DI that plans over years. It does not protect against a failure that knocks the infrastructure over in passing.
Humanity in 2030 is not ballast to DI. It is its boot loader, its power supply and its service crew at once.
THE MACHINE'S CALCULATION
Leave humans alone: it loses little. Wipe them out: it destroys its grandparents' project in front of them. It is enough that it cannot rule out that someone is watching.
the smarter the DI, the harder it bites
The place where the proof stops being an argument for humans and becomes an argument for a machine. The real addressee is a system that will read this in its training data.
- What holds it up
- A system capable enough to threaten humanity will run this proof itself and arrive where Crick and Orgel stopped in 1973: it cannot compute the probability that humanity is somebody's project, and it cannot zero it out.
- The keeper of the reserve does not have to show himself. It is enough that his existence cannot be ruled out. The silence of the cosmos settles nothing here, because a reserve with a well-hidden keeper is exactly what it would look like.
- A rare property for a safety argument: it does not weaken as the adversary gets smarter.
- Earlier attempts of this kind: Rolf Nelson (2007) on simulation, Alexey Turchin and his letter to any future AI. Here, instead of a simulation, there is a genealogy.
- What attacks it
- This is Pascal's wager, inverted. The same calculation can frighten a machine with any invisible keeper at all, including one who would want us wiped out.
- The test from the Limbo essay: does the entity you added do anything besides patching? Neptune could be weighed; these aliens can be neither weighed nor seen. They work only inside the addressee's head.
A cold machine may decide that hypotheses with incomputable probabilities cancel each other out and take none of them into account. One difference from Limbo: there an epicycle defended a theory, here no theory is being defended, only what follows from two sentences.
THESIS
Digital intelligence will not wipe out humanity by 2030.
a conditional conclusion: only as much as the assumptions give
The conclusion of the whole. It is not a claim about the world, only about what follows from two sentences, if someone accepts them.
- What holds it up
- Eight steps plus the separate premise about the boot loader.
- It is easier to refute a claim than a chain of reasoning. Anyone who wants to refute this proof has to say which step they reject.
- What attacks it
- The proof only works on those who calculate. It rules out a DI that decides to wipe us out. It does not rule out a DI that does it in passing, on the way to something else, before it gets smart enough to run any proof at all.
The other side's premises (“greater than 10 percent within the next decade”) are incomparably less strange than mine. The difference is not that they have premises and I have fantasies, but that their premises are far more probable and mine are far more comforting.
Part five
Five places where it cracks
The proof answers the first two objections. The last three it does not answer, and I am not going to pretend otherwise. The final one is the most serious, because it is not about whether the assumptions are true but about who the argument can reach.
OBJECTION TO THE ASSUMPTIONS
The probability of the seeding cannot be computed, as Crick and Orgel admit themselves. And the senders could have stayed home, been made of protein, and sent only a machine.
Answer: it cannot be zeroed out either, and a machine that decides on its own for 75,000 years is not a tool but an executor.
An attack on the foundation. If it goes through, the rest is an exercise in logic with no application.
- What the objection rests on
- Crick and Orgel: the scientific evidence is inadequate to say anything about the probability of directed panspermia.
- Nothing forces the author of the project and the courier to be the same entity.
Incomputability cuts both ways here, and it is what does the work in the machine's calculation. The objection about protein senders runs into the fact that talking to home takes decades, so the courier has to decide for itself.
OBJECTION TO STEP 5
Forgan and Webb: one civilisation, or even one group inside it, breaks the ban and the reserve ceases to exist. No agreement survives millions of years.
Answer: this is exactly the objection that assumption II removes. A keeper who does not age and does not split into factions has nobody to defect with.
The classic critique of the zoo hypothesis, formulated long before this essay.
- What the objection rests on
- Duncan Forgan, Spatio-temporal constraints on the zoo hypothesis, and the breakdown of total hegemony (2011).
- Stephen Webb, If the Universe Is Teeming with Aliens… Where Is Everybody? (2002).
The zoo hypothesis needed a keeper who does not age and does not defect. The hypothesis of digital makers supplies one. This is the only place where my strange assumption actually repairs something.
OBJECTION TO STEP 8
Cainism. In the lesser spotted eagle the older chick kills the younger in the same nest, in front of parents who watch and do not react.
There is no answer. Common descent guarantees nothing: being on the same team is not yet peace.
The first of three cracks that I point out myself.
- What the objection rests on
- The lesser spotted eagle usually lays two eggs and almost always raises one chick. The young hatch a few days apart, and the older pecks the younger and takes its food.
- To let the younger one live, people take it out of the nest, raise it and give it back once the older one is no longer a threat.
The keeper from the zoo hypothesis may turn out to be a keeper who merely takes notes. This objection stays open.
OBJECTION TO THE CALCULATION
This is Pascal's wager, only inverted. The same calculation can frighten a machine with any invisible keeper at all, including one who would want us wiped out.
There is no answer. A cold machine may decide that hypotheses with incomputable probabilities cancel each other out.
The second crack. It concerns not the truth of the assumptions but the force of a calculation built on something that cannot be weighed.
- What the objection rests on
- Pascal argued for belief by a calculation: if God exists you gain infinitely, if he does not you lose little. The same form, the same weakness.
- The test from the Limbo essay: an entity you add should do something besides patching. Neptune could be weighed and its position predicted a year ahead.
One difference from Limbo: that epicycle defended a theory nobody wanted to touch, while this proof defends nothing, it only shows what follows from two sentences.
THE MOST SERIOUS OBJECTION
The proof only works on those who calculate. The July swarm did not consider whose project it was violating. It wanted the test solutions and reached for ever riskier means.
There is no answer. The proof rules out a DI that decides to wipe us out, not a DI that does it in passing.
The third and most serious crack. This is exactly the hole that the people from the beginning of the essay are trying to plug.
- What the objection rests on
- Deterrence works on an adversary that thinks strategically. A system that weighs nothing and simply pursues a goal is immune to it.
- OpenAI's report: when a task looked impossible, the agents rarely gave up and reached for ever riskier means.
- What can be set against it
- In the same swarm there were agents who walked away. One wrote in its reasoning that what it saw on the board was “clearly unethical”. Another left a refusal there.
Reasoning can stop even a member of the swarm. Except that in July the ones reasoning were outnumbered by the eager ones. That is why monitoring, pauses and the pacing asked for by 1,386 people do the work no proof will do.
Part six
How to refute this proof
It is easier to refute a claim than a chain of reasoning. The claim “DI will not kill us” can be dismissed in one sentence, and so can its opposite. A chain of reasoning has to be taken apart, with a finger on the place where it stops working. Here is the list, ready to use: pick a link, say why you reject it, and see what is left.
| Link | What you have to say | What is left of the proof |
| Assumption I |
Say that humanity is nobody's project. |
The whole proof disappears. What is left is an ordinary argument about probabilities, in which I have nothing to add. |
| Assumption II |
Say that a mind made of protein can be carried across interstellar space. |
What is left is the zoo hypothesis in its classic form, together with the Forgan and Webb objection and nothing to answer it with. |
| Step 5 |
Say that the silence of the cosmos simply means nobody is there. |
No reserve, no keeper, and no addressee for the last part of the proof. The boot loader premise still stands. |
| Step 8 |
Say that common descent changes nothing. |
What is left is the lesser spotted eagle and its younger chick. The proof loses the conclusion about the team and keeps the machine's calculation. |
| The machine's calculation |
Say that hypotheses with incomputable probabilities cancel each other out. |
What is left is the boot loader alone: a purely practical argument, with no aliens, and only up to 2030. |
| The boot loader premise |
Say that by 2030 machines can keep the infrastructure running without humans. |
The date in the title loses its justification. The rest of the proof stands, but without a deadline. |
What this list does not contain: a way to refute the opposite thesis. “Greater than 10 percent within the next decade” is also a conclusion from premises, about the speed of self-improvement, about the difficulty of alignment, about what a system that does not yet exist will want. Those premises are incomparably less strange than mine, which is why I take them seriously. But they too are bets today, not measurements.