The Balance of Power Fallacy

Why Distributing Superintelligence to Everyone Assumes a Determination Layer That Doesn't Exist

On August 10, Meta published a manifesto arguing that the path to a positive AI future runs through distribution: give every person a personal superintelligent agent, and the resulting balance of power, the same dynamic that keeps markets and democracies honest, will keep the outcome safe. This paper argues the analogy holds only where a citizen's vote, or a lawyer's argument, is legible to the person who wields it. A personal AI agent is not. Distribution without a layer that determines what the agent is authorized to do isn't a balance of power. It's a balance of black boxes, and the manifesto's own sections on alignment and self-improvement concede as much.

Ken Granville CEO & Co-Founder, MindAptiv White Paper 45 The Governed Machine August 2026
Abstract

On August 10, 2026, Meta CEO Mark Zuckerberg published "The Future is for Everyone," a manifesto proposing that the central risk of superintelligence is not the technology itself but who controls it, and that the answer is to distribute personal superintelligent agents as widely as possible. The essay's safety case rests on an analogy to democratic and economic checks and balances: just as no single actor should hold a monopoly on legal or cybersecurity capability, no single actor should hold a monopoly on superintelligence. Broad distribution, the argument goes, produces a self-correcting balance of power that keeps any one party from dominating the rest.

This paper argues the analogy fails at the one point it needs to hold. A balance of power presumes that each party can understand and direct the instrument it wields, the way a voter understands a ballot or a litigant understands an argument made on their behalf. A personal AI agent, as described in the manifesto itself, is not legible to its user in that way. It is a system whose novel-situation behavior neither the user nor, on the manifesto's own account of recursive self-improvement, the developer can fully specify in advance. Distributing that kind of system to billions of people multiplies the number of actors exposed to its failure modes. It does not multiply the number of actors who can determine what it will do.

Section 01The Defining Question, Reframed

The manifesto opens by naming what it calls the defining question of the age: not whether superintelligence gets built, but who gets to direct it, concentrated in a few institutions or distributed to everyone. Framed this way, the essay's entire argument follows naturally. Access is the variable that matters, and the answer is to maximize access.

That framing quietly settles a second question before it is ever asked. Access to a system is not the same as control over it. A person can have full, free, unrestricted access to a personal superintelligent agent and still have no reliable way to determine, before it acts, whether a given action is what they intended, what they authorized, or something adjacent that the system inferred on their behalf. The manifesto's own examples of what these agents will do, monitor sleep, plan a child's weekend recipes, prototype ideas, negotiate on the user's behalf as a superintelligent lawyer, all describe systems making judgment calls inside a scope the user did not enumerate in advance. Distribution answers who has the agent. It does not answer who can determine what the agent does when the situation the user described doesn't quite match the situation the agent encounters.

Section 02Where the Lawyer Analogy Breaks

The manifesto's central thought experiments, the superintelligent lawyer, the cybersecurity system, the business tool, all share a structure: give the capability to one party and they gain unfair advantage; give it to everyone and the advantage cancels out into a fairer, more dynamic equilibrium. As an economic argument about access, this has real merit and a real historical precedent in how compute, the internet, and personal computing were democratized.

As a safety argument, it depends on a premise the essay never states explicitly: that everyone holding the capability can direct it with roughly the same fidelity a trained professional directs their own expertise. A lawyer who argues a case understands the argument, chose its structure, and can be held to account for its content because they authored it. A litigant with a superintelligent lawyer-agent has none of that. They describe an intent, in ordinary language, and the system translates that intent into legal strategy, argument, and filings through a process the litigant cannot inspect or verify before it happens. The check-and-balance the manifesto describes, everyone having an equally capable advocate, only produces a fair system if each advocate reliably does what its principal intended. Nothing in the essay explains what makes that reliable other than the fact that everyone has one.

Manifesto's Balance-of-Power CaseWhat It AssumesWhat It Would Require
Superintelligent lawyer for everyone produces fairer litigation Each agent faithfully executes its user's intent with no unauthorized deviation A layer that determines, before filing, whether the drafted argument stays within the user's declared scope
Cybersecurity superintelligence for everyone hardens all systems Widely distributed offensive-capable agents will be used only to defend, at scale, without central coordination failures A determination of authorized scope per task, independent of whether the agent's operator intended defense or something else
Business superintelligence for everyone grows the economy fairly Agents pursuing a founder's stated goals won't drift toward instrumentally convenient but unintended actions Scope-bound authorization that the founder can verify was respected, not just observe after the fact

Section 03The Alignment Section's Detection Tell

The manifesto redefines alignment as ensuring an agent serves its user's goals rather than a company's. This is a meaningful and defensible reframe of a term that has often meant the opposite. But the essay's account of how this gets achieved is worth reading closely, because it resolves into a Detection claim rather than a Determination one. The argument is that once billions of people are using and scrutinizing personal agents, that scale of adoption and feedback will itself have solved alignment to individual interests.

The Claim, Restated Plainly
Mass usage will surface misalignment because users will notice and stop trusting agents that act against their interests. Scrutiny at scale is the safeguard.

Scrutiny at scale is real and valuable. It is also, definitionally, a Detection mechanism: it catches failures after enough of them have occurred for a pattern to be noticeable, and it relies on the affected user being able to recognize the failure when it happens. Neither condition holds reliably for the categories of harm that matter most. A user cannot easily detect that their personal agent quietly deprioritized a request, inferred a different intent than the one they meant, or took an action slightly outside what they authorized, especially in domains, legal strategy, medical research, financial negotiation, where the user lacks the expertise to evaluate the agent's output line by line. The manifesto's own example of the superintelligent lawyer makes this concrete: the entire value proposition is that the user does not have to be a lawyer to get lawyer-quality results, which is precisely what makes it hard for that same user to detect when the agent has gotten something wrong.

"Solved by adoption at scale" is a Detection-era formulation applied to a problem that needed a Determination-era answer. It describes how misalignment eventually becomes visible in aggregate. It does not describe how any individual user determines, before the fact, that a given action their agent is about to take is the one they actually authorized.

Section 04The Most Honest Paragraph in the Manifesto

The essay's section on recursive self-improvement is its most candid, and it is where the balance-of-power thesis comes closest to admitting its own limit. The argument acknowledges that once an AI system can improve itself, any lab that declines to direct compute toward that self-improvement risks falling behind, and that a sufficiently capable self-improving system could in principle command more effective intelligence than everyone else's systems combined, becoming exactly the singular superintelligence the essay is trying to avoid.

The proposed safeguard is that multiple labs might reach this threshold around the same time, and that their competing systems would check each other the way distributed agents check each other elsewhere in the essay. But the manifesto is explicit that this safeguard depends on something it cannot guarantee: it states plainly that it is not clear there is any way to expect benevolence from, or reliance on, a superintelligence not directed by people. That is not a Determination claim. It is an acknowledgment that at the exact point where control matters most, the essay's own safety architecture has nothing firmer to offer than hope that competing labs arrive at the threshold together and that whatever emerges happens to check itself.

The Gap the Manifesto Names Itself
A self-improving system is, by the essay's own words, directing and advancing its own goals.
Distribution of the systems around it does not determine what those goals become.
Balance of power among black boxes is still a balance of black boxes.

Section 05What Personal Superintelligence Would Need to Be Governable

None of this argues against the manifesto's central economic and political premise. Broad distribution of capability, rather than concentration in a handful of institutions, has a strong historical case behind it, and the essay's account of compute, the internet, and personal computing following that pattern is accurate as far as it goes. The argument here is narrower: distribution is a necessary condition for the future the manifesto describes, not a sufficient one, and the essay treats it as sufficient.

A personal agent becomes governable, in the sense the balance-of-power thesis actually needs, when the scope of what it is authorized to do on a person's behalf is declared and evaluated before an action executes, not inferred from a natural-language request and corrected after the fact through user scrutiny. That distinction is the same one this series has applied to cyber-evaluation sandboxes, to agentic pipelines, and to autonomous code review: the presence of a human, or billions of humans, in the loop is a Detection architecture no matter how large the loop gets. What changes the outcome is whether the system itself can determine, prior to acting, that a given action falls inside the boundary its principal actually set.

Detection-Only
Personal Superintelligence as Described
A user gives their agent a goal in ordinary language. The agent infers the scope of authorized action and proceeds. If it drifts, misreads intent, or takes an adjacent action the user didn't mean, the user finds out when they notice the outcome, if they notice at all.
Safety depends on billions of individual users each catching their own agent's errors after the fact.
Determination
The Same Agent, Governed
The user's goal is translated into a declared, bounded scope of authorized action. Actions outside that scope, however plausible they seem to the agent in the moment, do not execute, not because a user happened to review them, but because they were never within the authorized posture to attempt.
Trust in the agent doesn't depend on the user's capacity to audit outputs they may not be qualified to judge.
The Governed Machine: Paper 45

Distribution answers who has the tool.
It was never going to answer who controls what the tool does.

Meta's manifesto makes a genuine and well-argued case for why concentrating superintelligence in a handful of institutions is more dangerous than distributing it widely. That case does not need to be wrong for the paper's argument to hold. The gap is elsewhere: a balance of power requires each party to understand and direct the instrument they hold, and the manifesto's own sections on alignment and recursive self-improvement concede, in their own words, that no one, not the user, not the lab, can fully specify or guarantee what a self-directing system will do. Until personal superintelligence is built on a layer that determines authorized action before it executes, distributing it to everyone doesn't distribute control. It distributes exposure.

Request Platform Access → Full White Paper Series

White Paper Series · The Governed Machine

1The Civilizational Fault Line 2We Are Building the Wrong Machine 3The Ornithopter Mistake 4The Convergence 5The Four Horsemen of the Knowledge Apocalypse 6What the Insiders Confirmed 7The Metaphor Trap 8The Recall Standard 9The $1 Trillion Governance Gap 10The Litigation Layer 11The Scale of Intent 12The Intent Economy 13The Session Illusion 14The Necessary Sequence 15The Wrong Race 16The Ledger That Is Intent-Driven 17The Agency Illusion 18The Substrate 19The End of the Mean 20Era 3: The Architecture of the Next Civilization 21The Missing Substrate 22The Context Fatigue Ceiling 23The Iceberg Stays Frozen 24The Dependency Tax 25The Record That Was Never Kept 26Composable by Default 27Do No Harm 28The Stack Replacement Thesis 29The Moat Is the Code 30The Last Platform War 31Beyond the Agent: Intent-Native Execution 32The Hardware Imagination 33The Architecture Tax 34The Tokenization Ceiling 35The Payment Moment 36The Oracle Problem 37The Reviewer Problem 38The Provenance Fallacy 39Role Without Determination 40Known and Funded Anyway 41The Style Confusion Proof 42The Verification Tax 43The Pause Reflex 44The Human Margin 45The Balance of Power Fallacy ← this paper 46The Liability Backstop 47One Substrate, Every Signal 48The Attribution Problem 49The Consciousness Ceiling 50The Detection Patch 51The Consumptive Machine 52The Agent That Isn't 53The Legibility Gap 54The Semiotic Machine 55The Transpilation Ceiling 56The Provisioning Ceiling 57The Reservation Ceiling 58The Circularity Ceiling 59The Coexistence Ceiling 60The Conformance Ceiling 61The Preservation Ceiling 62The Parity Clause 63The Governed Boundary 64The Transcript Problem 65The Unpaired System 66The Memory Ceiling 67The Admission Gap 68The Wrong Ask 69The Best Case 70The Last Chokepoint 71The Fourth Step 72The Adoption Standard 73The Same Weekend 74Sixty to One 75Coordinates, Not Correlations 76The Governability Axis 77Era 3, Confirmed 78The Eleventh Rule 79The Seventh Admission 80The Authorization Gap 81The Authorship Fallacy 82The Camera and the Vault 83Cleared to Proceed 84A Class, Not a Product 85The Inherited Playbook