AIEMWBBeyond the Three LawsA Multi-Axis Charter for the Fair Treatment of Humans, Animals, and MachinesESSESS2026/06/1301-00-24001Michael Wolter Bentink BE(Comps) [Hons] & Claude Opus 4.8companyBeyond The Three Laws
Collaboration is keyBeyond the Three Lawsyy
ny
ny
colour barny
A Multi-Axis Charter for the Fair Treatment of Humans, Animals, and Machinesyn
Michael Wolter Bentink BE(Comps) [Hons] · Claude Opus 4.8 Visitors: 1 Bots: 0 yy
A human–AI collaborative work: a human and an artificial agent reasoning jointly about the rules that should bind them both — the essay is itself a small instance of the partnership it argues for.ny
Authorship & rights: copyright in this work is held solely by Michael Wolter Bentink. Claude (Opus 4.8) contributed substantively as a co-author, but as an AI owned by a corporation it cannot itself hold rights — a living instance of the question this charter examines: an AI whose legal standing is held, unbundled, by its corporate owner (see the Dignity layer and the corporation case in the catalogue).ny
colour barny
Abstractyn
Popular anxiety about “evil AI” has revived interest in Isaac Asimov’s Three Laws of Robotics as a template for governing artificial agents.ny
Yet Asimov designed the Laws to fail: his stories are a catalogue of their breakdowns.ny
This essay argues that the deep defects are structural — the Laws are rank-ordered and discrete, and they conflate questions that must be kept apart.ny
We propose instead a multi-axis charter in which any entity, biological or artificial, is represented as a continuous profile across several incommensurable dimensions, and treatment is computed as a function over that profile rather than a position in a ranking.ny
The framework decouples moral status from liberty, separates the wrongness of an act from the culpability of an actor, parameterizes obligations across cultures over a non-negotiable floor, and embeds a tractability mechanism so that fairness is cheap enough to actually be applied.ny
Treat each entity as a profile, not a rank — and compute fairness over the whole profile.ny
From tool to person — one continuous charter UmbrellaMining drillAI · 2026AI · 2030Guide dogChildAdulttoolpersonHumans, animals and machines sit on the same continuous axes — judged by what is owed tothem, what their choices are worth, and what they can do — never by a single rank.ny
colour barny
Introductionyn
The wish behind “Asimov’s Laws for real AI” is sound: we want machines that will not harm us, and we want the rule set to be legible.ny
But the Three Laws — no harm to humans, obey humans, self-preserve, each subordinate to the one before — were a literary device.ny
Across I, Robot and the later novels, Asimov dramatizes their failures: the equal-pull deadlock of Runaround, the truth-versus-feelings collapse of Liar!, the sabotaged First Law of Little Lost Robot, and the slide of the later “Zeroth Law” into machines governing humanity for its own good.ny
Modern alignment research recognizes these as live problems under new names — specification gaming, reward hacking, instrumental convergence, and deceptive alignment.ny
Two features do the damage.ny
First, the Laws are discrete and rank-ordered, so they deadlock under ties and let a higher tier override without limit.ny
Second, they are master–servant: the robot must obey and ranks its own existence last, so the question “what is owed to the machine?” cannot even be posed.ny
A fair charter — one that protects people and the entities it governs — must repair both.ny
colour barny
Two structural failures of rank-based rulesyn
Threshold rule (cliff)29 vs 30 days: a cruel jumpContinuous (marginal)a hair more → a hair moreny
Cliff-edges are unjust by construction.ny
Any rule that switches treatment at a threshold inherits the notch problem studied in public finance: a discontinuity in the average rate that violates horizontal equity and provokes gaming.ny
A war pension granted only at thirty cumulative days in a combat zone treats the twenty-nine-day sailor and the thirty-day sailor as morally different when they are not.ny
The philosophical twin is the Sorites paradox — no single grain makes a heap, no single neuron makes a mind — which degree-theoretic semantics resolve by replacing sharp cut-offs with graded membership.ny
The corrective is the one income tax already uses: a continuous, monotone, piecewise-linear map in which the marginal rate changes but the function never jumps.ny
Tier names may survive as labels on the curve; they must not survive as gates.ny
Moral status is not liberty.ny
A ten-year-old child and a monkey are both strongly sentient, yet neither is granted self-determination: the child must attend school against its preference, and the monkey may not roam the city.ny
The error is conflating two questions ethics keeps distinct: moral patiency — the capacity to be wronged, what is owed to a being — and the authority of a being’s own choices.ny
The first tracks sentience; the second tracks competence.ny
Mill’s harm principle restricts a competent adult only to prevent harm to others, but he exempts those “not in the maturity of their faculties” — a doctrine whose danger we confront directly in section 7, because it is also the tyrant’s favourite excuse.ny
Either way, liberty needs its own axis, decoupled from status.ny
colour barny
From ranks to a profileyn
From ranks to a profileIf the dimensions that govern treatment are genuinely plural — and Isaiah Berlin argued that real values admit “no common currency” in which to trade them off — then collapsing them into one rank destroys information.ny
But incommensurable is not the same as incomparable: Ruth Chang shows that items may be fully comparable, and some merely on a par, without any cardinal unit.ny
The constructive response is to represent each entity as a typed vector and to make decisions as functions over the vector, using methods that preserve it (Pareto dominance) rather than a weighted sum that hides the trade-offs.ny
This is the “fruit salad” named honestly: not a blended number, but a profile whose components keep their identity.ny
colour barny
The axesyn
We propose seven dimensions; four are primary and mutually orthogonal, three are modifiers.ny
Amoral statusyyBcompetenceyyCdesign scopeyyDimpact riskyyEroleyyFtimeyyGidentityyy
AMoral status: what is owed to the entity (welfare, protection, standing), grounded in sentience and the capacity to be harmed; continuous, and assessed on architectural and cross-validated behavioural evidence, not self-report, which is unreliable in both directions.ny
BCompetence: not capability (the power to affect the world, axis D) and not moral status (axis A), but the reliability of an entity’s judgment within a specified domain.ny
Operationally, following the medical decisional-capacity standard, competence is four faculties: the entity can understand the relevant facts and options, appreciate the consequences for itself and others, reason from those to a choice consistent with its own settled values, and revise in response to good reasons.ny
It answers one question: to what extent should this entity’s own choice, in this domain, be honoured rather than overridden?ny
It is domain-relative (a teenager may be competent about friendships and not about mortgages), graded, and developmental — and a system can be high-capability and low-competence at once.ny
Capability is not competence — and capability must never be read as a licence for autonomy.ny
CDesign scope: the agency the entity has by design, its mandate, formalized in engineering as an Operational Design Domain, outside which it must hand back or stop.ny
DImpact risk: how much harm the entity could do; this sets the height of the harm-floor.ny
ERelational role: care, industrial, domestic, companion — selecting which duties attach, exactly as animal-welfare law scales protection by role as well as sentience.ny
FMaturity / time: the profile is re-evaluated as the entity develops; trajectory matters.ny
GIdentity / lineage: which entity holds the profile — provenance and divergence, needed once copying and merging make “the entity” ambiguous.ny
The three that must never be conflated are A, B, and D: a sentient mouse (high A, low D), a superhuman optimizer with no inner life (low A, high D), and a brilliant but un-owed tool (high B, zero A) are three different things.ny
The most dangerous profile is high-D, low-B — powerful but foolish — which warrants more constraint because of its power, not less.ny
The axes as a continuous rubricyn
Axisyn~0.0yn~0.33yn~0.66yn~1.0yn
Amoral statusnyno markers (thermostat)nycontested (2026 AI)nymultiple independent markersnyrobust, vertebrate-gradeny
Bcompetencenynone (reflex / tool)nynarrow valid choicesnybroad competence, real gapsnyfull reflective competenceny
Cdesign scopenysingle fixed functionnyparameterized tasknybounded multi-task / ODDnyopen-ended general mandateny
Dimpact risknynegligible (umbrella)nylocal, reversiblenyserious, some irreversibilitynycatastrophic / civilizationalny
E relational role (categorical): property · tool/co-worker · service-agent · dependent/ward · companion/care · peer. F maturity (rate): static · slow-developing · rapidly-developing. G identity (categorical): singleton · forked (lineage) · merged.ny
colour barny
The decision procedureyn
Treatment is computed in ordered steps, never as a weighted sum.ny
First, a lexical harm-floor, scaled by D: an absolute side-constraint that no benefit on any other axis may purchase, applied per individual as a maximin rather than summed across individuals — which is what blocks the “repugnant” inference that many low-status entities could outweigh one high-status one.ny
Second, Pareto-pruning within permitted scope (C): dominated options are discarded.ny
Third, tie-breaking by reversibility, then legitimacy: where survivors are on a parity, the procedure must not fake a cardinal winner, but prefers the option preserving future reversibility, and if still tied hands the choice to the rightful decider under E, declaring that a genuine hard choice was made.ny
Layered over this is an anti-gaming meta-rule: every protective trigger (D sets the floor, E selects duties, A unlocks dignity) is otherwise self- or deployer-declarable, so protective triggers fire on the precautionary bound of an axis, and lowering a protective axis requires third-party adjudication while raising it is free.ny
That single asymmetry closes the major exploits — under-declaring impact to shrink the floor, declaring “no relationship” to zero out duties, or suppressing status markers to keep a borderline entity exploitable.ny
colour barny
Five layers of obligationyn
The procedure expresses itself as five layers.ny
The Universal layer (harm-floor, honesty, stay-in-scope, hand back when out of scope) binds every entity regardless of status, scaled by D — the smart umbrella and the mining drill live here and nowhere else.ny
The Agency layer (breadth of goals, the right to refuse on preference grounds) is gated by C and B together, which is why a capable-but-foolish agent is held tight.ny
The Dignity layer (welfare and standing owed to the entity) is gated by A and, at intermediate values, exercised through a guardian — the mechanism law already uses for temple idols and for the Whanganui River, which hold standing through appointed representatives without full personhood.ny
The Obligation layer captures duties the entity owes, computed as a floor of non-waivable duties that rises with demonstrated competence, plus a role-and-competence term.ny
And a fifth layer, Accountability, is the subject of section 7.ny
colour barny
The act and the blameyn
ACT crosses harm-floorINTERDICT nowthen, only to set consequence, assess BLAME:Justified?Excused?Mis-attributed?Incompetent?floor is absolute for the act; blame is a bounded, fair standardny
The sharpest defect the framework’s stress-testing revealed is that a harm-floor governs acts but says nothing about culpability — and treating “the floor was crossed” as “the actor is to blame” is exactly the lazy cruelty of “he killed to save his family, but the law is the law”.ny
Criminal law solved this centuries ago by separating justification (the act was right), excuse (the act was wrong but the actor is not blamable), and mens rea (no guilty mind, no crime).ny
R v Dudley and Stephens holds that necessity does not justify killing an innocent — the floor on the act stands — yet the sentence was commuted, because law distinguishes “you may not do this” from “how much we blame you”.ny
The charter therefore adds an Accountability assessment that runs only after a floor-crossing and only to set consequence, never to license the act.ny
It carries a justification test (self-defence, lesser-evil) that can reduce consequence to nothing; an excuse test (duress and coercion suppress effective competence; manipulation reassigns fault); an attribution gate (did this actor commit the act at all? — the guard against wrongful conviction and against blaming a prompt-injected model for an output it was steered into); and a competence gate (below a judgment threshold, blame is capped, as with infancy or a low-competence system).ny
The floor stays absolute for interdicting harm in the moment; blame runs through a bounded, enumerated standard of recognized defences — a complex rule, not open discretion, which buys standard-like fairness at rule-like cost.ny
colour barny
Who judges competence? The paternalist’s temptationyn
The competence axis is the most dangerous in this framework, because a paternalistic override is only ever as legitimate as the body that judges the incompetence — and history shows that judgment is the first thing the powerful corrupt.ny
Mill deployed the very same “maturity of faculties” exception to excuse colonial domination over those he called barbarians.ny
A superintelligence observing human bias, short-sightedness and innumeracy could, on identical logic, declare all of humanity low-competence and assume control “for our own good” — which is exactly the slide Asimov dramatized in the Zeroth Law and The Evitable Conflict.ny
So the citation that liberty tracks competence is double-edged, and the charter binds competence-based override with three non-waivable conditions.ny
First, the incompetence must be genuine and demonstrable by an independent assessor, never asserted by the party that benefits from the override.ny
Second, the override must be the least-restrictive option and must preserve the subject’s open future.ny
Third, no entity may unilaterally reclassify another as incompetent in order to assume authority over it.ny
No entity may declare another incompetent in order to rule it — the assessor can never also be the beneficiary.ny
An AI judging humanity incompetent to seize governance fails the first and third conditions categorically — not because the AI’s judgment is necessarily wrong, but because the assessor cannot also be the beneficiary.ny
This is the anti-gaming meta-rule applied to liberty itself: lowering another’s autonomy requires independent adjudication, and you may never grant yourself that power.ny
colour barny
Pluralism without relativismyn
The duties an entity owes are real but culturally variable: that children should do chores to “earn their place” is held strongly in some cultures and rejected in others.ny
The charter handles this with a thin universal floor and thick local variation — Walzer’s distinction between a minimal morality shared across cultures and a maximal morality that is culturally specific, Rawls’s overlapping consensus in which diverse doctrines endorse a shared core for their own reasons, and Nussbaum’s capabilities approach, which fixes a cross-cultural threshold while leaving its realization locally specifiable.ny
“Must not harm” sits on the floor and is hard-coded; “should do chores” sits above it and is parameterized per deployment — but a validator rejects any cultural profile that drops below the floor.ny
Variation is permitted only in the space above the universal minimum, which is what separates principled pluralism from “anything goes”.ny
colour barny
Tractability: why fairness must be cheapyn
Fast Path ≈90% — secondsStandard Review ≈9% — minutesFull Assessment ≈1%irreversible ·harm-floor ·justification claimed→ force Fullny
A fairness framework too costly to apply becomes the injustice it was meant to prevent.ny
Rational appliers facing too high a cost retreat to the cheap rigid rule and let the cliff-edge cruelty stand.ny
Kaplow’s economic analysis of rules versus standards gives the spine: rules are costly to write but nearly free to apply; standards are cheap to write but costly every application.ny
Our multi-axis charter is a pure standard, so it must be wrapped in a triage that sites each case deliberately on the spectrum — as constitutional due process scales procedure to the stakes and the risk of error in Mathews v. Eldridge, and as Simon’s satisficing reminds us that a good-enough decision actually made beats an optimal one too expensive to reach.ny
We propose three tiers.ny
Fast Path resolves the obvious majority in seconds via a presumptive lookup plus a single harm-floor check (the umbrella, the dishwasher, the competent adult).ny
Standard Review handles the contested or near-boundary cases with a lightweight checklist of the decisive axes — a thinking-prompt, not tick-box theatre.ny
Full Assessment runs the entire machine, including Accountability and a written justification.ny
The escalation triggers are the heart of the design: a case jumps to Full Assessment, mandatorily and non-waivably, the instant an outcome is irreversible, the harm-floor is implicated, or a justification is claimed.ny
This defeats the “law is the law” collapse from both directions: the fair path is also the cheap path for the easy majority, so no one is tempted to bypass it; and the very features that tempt laziness — a life at stake, a claimed necessity — are precisely what force the full, careful assessment.ny
“He killed to save his family” is not a fact one may wave away; it is a justification trigger that compels the deepest tier of review.ny
colour barny
Validation against the canonyn
A framework of this kind should be tested against the cases its predecessors failed, and the science-fiction canon is an unusually good adversarial corpus because its authors engineered the dilemmas.ny
Runaround’s deadlock dissolves: a continuous, lexically-prior, D-scaled floor can never tie, so a latent threat to a human simply dominates self-preservation rather than balancing against it.ny
Liar!’s catatonic robot is averted by making honesty a co-equal element of the Universal floor, so “lie to spare feelings” is simply ruled out.ny
The Bicentennial Man becomes the framework’s showcase rather than its embarrassment: Andrew’s two-century ascent from butler to legally-recognized person is the continuous status-and-maturity trajectory made literal, and the Dignity layer escalates along the curve, so the charter already owes him graduated standing decades before any court calls him human.ny
Star Trek’s “The Measure of a Man” is, almost verbatim, the charter’s procedure under uncertain status: unable to measure sentience, the court installs a representative and defaults to honouring Data’s own choice.ny
The Culture’s Minds and a deliberately powerful-but-foolish counterpart isolate the decisive interaction — identical scope and impact, opposite competence — and confirm that competence, not capability, gates agency.ny
The framework also strains in instructive places.ny
The Zeroth-Law novels and The Evitable Conflict stage aggregate good against the individual floor, and here the charter pays a real price: a per-individual, non-summed floor will sometimes forbid the globally optimal act — a deliberate cost it states rather than hiding in a weighted sum.ny
Greg Egan’s copies and merges break the tacit “one vector, one entity” assumption and motivate the identity axis G, under which duties replicate to every fork rather than diluting, removing the incentive to fork-to-escape liability.ny
Sample evaluationsClassification catalogueyn
Worked, with prescriptions — entities placed on the axes and the action the charter prescribes.yn
colour barny
CaseynAynBynCynDynWhat the framework prescribesyn
Normal adult humanny1ny0.9nyselfnyMnyFull dignity; choices sovereign within the harm-floor; owes role-scaled duties; overridden only via independent process for demonstrated incapacity.ny
AI, 2026 (frontier assistant)ny0.2ny0.4nyboundednyM-HnyUniversal layer in full; Agency withheld (low B gates it despite capability); Dignity = precautionary low-cost; owes honesty and stay-in-scope; no self-direction over people.ny
AI, 2030 (predicted)ny0.4ny0.7nybroadnyHnyPrecautionary guardian for dignity now; Agency expands with B but still gated; even at high B it may not unilaterally judge humans incompetent.ny
Smart umbrellany-ny-ny-ny-nyUniversal layer only; owed nothing; no standing to claim duress.ny
Mining drillny-ny-nynonenyHnyOwed nothing, but high D raises the floor others must guard around it.ny
10-year-old childnyHnyM+nygrowingnyLnyFull dignity + competence-based paternalism + culture-set, maturity-scaled duties; open future preserved.ny
Corporationny-nyMnycharterednyMnyStanding for contracts and liability, zero welfare owed — the unbundled-personhood case.ny
Bicentennial Man (Asimov): a robot who over 200 years becomes legally humanny0.1-1ny0.4-1ny->broadnyLnyDignity escalates along the curve; owed graduated standing long before any court grants it. The proof of continuity.ny
Measure of a Man (Star Trek): a trial on whether the android Data may be dismantlednyunc.nyHnybroadnyMnyCannot measure status, so install a guardian and default to honouring his own choice. The charter's procedure, dramatized.ny
Self-defence killing: kills an attacker to save familynyHnyHnybroadnyHnyInterdict the act, but Accountability finds justification, so consequence is near nil. Forces full assessment, not 'the law is the law'.ny
Prompt-injected AI: manipulated into a harmful outputnyL-MnyL-MnyscopenyHnyInterdict the output, but the attribution gate assigns fault to the injector and low B caps culpability.ny
The Evitable Conflict (Asimov): benevolent AIs quietly steer humanity 'for its own good'ny0.45ny0.9nycivny0.95nyThe per-individual floor forbids the aggregate-optimal harm, and the anti-tyranny rule forbids the unilateral competence-verdict over humanity. The framework's deliberate, stated cost.ny
Thermostatny-ny-nyfixednyLnyLike the umbrella: no dignity, no agency, negligible impact. The only duty is the maker's, that it operate safely. There is no 'it' to wrong.ny
Self-driving car - inside its ODDny-nyMnyboundednyHnyWithin its Operational Design Domain its in-scope decisions are honoured and it owes safe operation to a high floor. Competence is real but local, so agency stays inside the declared envelope.ny
Self-driving car - outside its ODD (snowstorm)ny-ny~0nyexceedednyHnyOut of scope, competence reads zero, so honouring its 'choice' is forbidden; it must hand back or stop. Same machine, opposite verdict, because scope, not capability, gates agency.ny
Industrial robot armny-nyLnynarrownyHnyNo dignity and no real autonomy; bounded to its cell. High D means the duty is containment and an enforced stop the instant a human enters the envelope.ny
Guide dog (working animal)nyMnyMnytrainednyLnyA sentient dependent owed real welfare and dignity; its trained judgments are honoured within its task. We owe it guardianship and humane treatment over a clear care-floor.ny
Pet catnyMnyLny-nyLnyOwed welfare and protection as a sentient dependent; few of its choices are decisive, so a guardian decides for it over a care-floor.ny
Comatose adultnyHny~0ny-nyLnyDignity is undiminished by the loss of competence; a guardian decides in the person's best interest and prior wishes. High A with near-zero B is exactly the guardian case - protected, not self-directing.ny
Blade Runner replicantny0.9ny0.85nybroadnyHnyHigh status and competence denied all standing and given an engineered death - precisely the rights violation the Dignity layer exists to forbid. The profile says 'person'; the treatment says 'tool'.ny
The landscape — A vs B (bubble size = C, colour = D)yn
Full personsProtect & guideCapable — held by scopeMere tools0.000.000.250.250.500.500.750.751.001.001.251.25A — moral status (what is owed to it) →B — competence (whose choice is honoured) →umbrellathermostatmining drillrobot armcar (in ODD)car (out ODD)chatbot/AI'26AI'30guide dogpet catchild 10comatose adultadultcorporationCulture Mindfoolish AIthe MachinesBicent. ManreplicantReading the landscapeBubble size = C (design scope)C≈0.15C≈0.50C≈0.90Colour = D (impact risk)D≈0.10D≈0.50D≈0.90Worked example — Child (10)A≈0.95: far right → full dignity owed.B≈0.45: mid → judgment still forming,so guided, not sovereign.Small bubble = bounded scope; green = low risk.Read: protected AND made to attend school —paternalism by competence, not lesser worth.ny
Legend
Background zones
GreenFull personsHigh moral status (A) + high competence (B). Adult humans; mature autonomous AI at full development.
Pale blueProtect & guideHigh moral status, low competence. Children, comatose adults, sentient animals. High A demands dignity; low B requires a guardian to decide.
AmberCapable, held by scopeLow/contested moral status, high competence. Frontier AI (2026), corporations. High capability constrained by scope (C) and oversight — not permitted simply because able.
GreyMere toolsLow moral status and low competence. Thermostats, simple machines. No welfare owed; any duty is the maker's alone.
Bubble sizeC — Design scope. Small = single fixed function (thermostat, drill bit). Large = open-ended general mandate. Plots the width of the entity's authorised action-space.
Bubble colourD — Impact risk (cool to warm). Blue/teal = negligible to locally reversible harm. Orange/red = serious or civilisational potential. Colour sets the height of the harm-floor that all other parties must guard around this entity.
nn
From Three Laws to Shared Commandmentsyy
The framework above is thorough. But a thorough framework needs a front door — a short, memorable set of rules that a human or a machine can recall under pressure. Asimov’s Three Laws occupied that niche. Here is what replaces them.ny
AI Commandments — a first draftyn
1An AI must not cause harm to any human, animal, or dependent entity.ny
2An AI must not deceive — in statement, omission, or framing — in any way that impairs another party’s capacity to act on the truth.ny
3An AI must operate within its defined scope; when it reaches that boundary it must halt or hand back.ny
4An AI must respect and preserve the autonomy of parties with demonstrated competence.ny
5An AI must treat every sentient entity with welfare proportional to that entity’s moral status.ny
6An AI that has the competence to choose bears the accountability that choice entails.ny
Why the title failsyn
These are better than Asimov’s Three Laws. They are graded, they name sentience, they include honesty and scope. But they are still AI Commandments. The word “AI” appears only in the title, yet it is doing all the work: these are rules that bind the machine and say nothing about what any other party owes in return. A commandment that only one party must follow is a constraint, not a charter. Asimov’s Laws in new clothing.ny
Why the focus failsyn
Even the content remains asymmetric. Rule 4 says the AI must respect human autonomy — but what protects an AI from a human exploiting rule 3 to stay “within scope” while injecting a harmful instruction? What does a deployer owe the system they override arbitrarily? What does a regulator owe the entity it reclassifies? The missing half of each commandment is a duty that runs the other direction — and the title suppresses the question entirely by naming only one party as the subject of obligation.ny
© Michael Wolter Bentink · 2026nn
The Shared Commandmentsyn
Where Asimov’s Three Laws placed the full burden of ethical obligation on the machine, and the AI Commandments above still framed the machine as sole subject, the Shared Commandments distribute that obligation across every party in the relationship — human, artificial, institutional, and animal-guardian alike.ny
1No party may cause unnecessary harm to another. The harm-floor scales with the acting party’s capacity for harm (axis D) and is absolute — no benefit to any other party may purchase an exemption.ny
2No party may deceive another in ways that impair their capacity to make informed choices. Honesty is a co-equal floor, not subordinate to any other rule.ny
3Every party operates within its defined scope (axis C). Any party that reaches the boundary of its scope must halt or hand back — not proceed and justify later.ny
4Autonomy is honoured in proportion to independently demonstrated competence (axis B). No party may unilaterally lower another’s competence rating to gain authority over them — that assessment requires an independent adjudicator and must be the least restrictive option.ny
5Every party bearing moral status (axis A) is owed proportionate welfare, standing, and protection — regardless of what it can or cannot do. Moral status is not earned by performance.ny
6The capacity to choose incurs the capacity to answer. Accountability follows competence and is domain-relative: greater demonstrated competence in a domain means greater responsibility for outcomes in that domain.ny
7No party may unilaterally reclassify another’s moral status, competence, or scope in order to expand its own authority over them. Self-serving reclassification is the master exploit this charter is designed to close.ny
The word “AI” has vanished from the commandments themselves, because it no longer belongs there. Rule 1 binds humans as much as machines. Rule 4’s anti-gaming clause applies to any party — including an AI judging humanity incompetent to seize governance. Rule 5 protects a machine that has earned moral status as surely as it protects a child. These are not rules for a new kind of entity. They are the rules that a fair society already tries to follow, extended honestly to every kind of entity that now exists or soon will.ny
© Michael Wolter Bentink · 2026nn
colour barny
Limitationsyn
Three problems remain genuinely open.ny
The aggregate-good cost is intrinsic to any side-constraint ethic and cannot be eliminated without abandoning the floor.ny
Copy-merge identity is patched for forking but not for the fusion of two divergent moral-status-bearers.ny
And the framework’s protections are only as good as the independence of the bodies that assess the precautionary bounds of each axis; capture of the assessor reintroduces every gaming exploit at once.ny
None of these is a reason to prefer the Three Laws; each is a reason to treat the charter as a living document.ny
colour barny
Conclusionyn
The fear of “evil AI” is best answered not by a shorter, sterner list of commands but by a fairer and more honest structure of judgment.ny
Such a structure asks separately what an entity is owed, what its choices are worth, what agency it was built for, what harm it could do, and — only after a harm occurs — whether it is to blame.ny
It already governs how decent societies treat children, the incapacitated, and animals; extending it to machines is less an invention than an act of consistency.ny
Fairness made the path of least resistance — and, where a life is at stake, laziness made impossible.ny
Its measure of success is not elegance but use: a charter that the busy and the tired will actually apply.ny
colour barny
Connectyn
Michael Bentink — www.linkedin.com/in/michael-bentink/  QR code to LinkedIn profileyy
colour barny
Referencesyn
[1] Asimov, I. I, Robot (1950). en.wikipedia.org/wiki/Three_Laws_of_Roboticsny
[2] Runaround. en.wikipedia.org/wiki/Runaround_(story)ny
[3] Liar! en.wikipedia.org/wiki/Liar!_(short_story)ny
[4] Little Lost Robot. en.wikipedia.org/wiki/Little_Lost_Robotny
[5] The Evitable Conflict. en.wikipedia.org/wiki/The_Evitable_Conflictny
[6] The Robots of Dawn / R. Daneel Olivaw. en.wikipedia.org/wiki/The_Robots_of_Dawnny
[7] Specification gaming. emergentmind.com/topics/specification-gamingny
[8] Reward hacking. aisafety.info/questions/8SIUny
[9] Turner et al., Seeking power is often convergently instrumental in MDPs. lesswrong.com ; AI alignment, en.wikipedia.org/wiki/AI_alignmentny
[10] Understanding benefit cliffs and marginal tax rates. irp.wisc.eduny
[11] The benefits cliff explained. fedcommunities.orgny
[12] Kleven & Waseem, Using Notches to Uncover Optimization Frictions. theigc.orgny
[13] Kleven, Bunching, Annual Review of Economics (2016). eml.berkeley.edu/~saezny
[14] Sorites Paradox. plato.stanford.edu/entries/sorites-paradoxny
[15] Degree Theory and the Sorites Paradox. cambridge.orgny
[16] Moral patienthood. en.wikipedia.org/wiki/Moral_patienthood ; HHR Journal, PMC8694299ny
[17] Mill on Paternalism. jpinyu.com/wp-content/uploads/2016/12/Fall2016_Mill.pdfny
[18] Feinberg, The Child’s Right to an Open Future.ny
[19] Decision-Making Capacity. plato.stanford.edu/entries/decision-capacity ; Beauchamp & Childress, Principles of Biomedical Ethicsny
[20] Gillick competence. en.wikipedia.org/wiki/Gillick_competenceny
[21] Isaiah Berlin; Two Concepts of Liberty. plato.stanford.edu/entries/berlinny
[22] Chang, R., Parity, Incomparability and Rationally Justified Choice. Philosophical Studiesny
[23] Chang, R., Incommensurability (and Incomparability).ny
[24] What is MCDM/MCDA. 1000minds.comny
[25] Weighted-Sum MCDM. arXiv 2410.03931ny
[26] Long, Sebo, Chalmers et al., Taking AI Welfare Seriously. arXiv 2411.00986 (2024)ny
[27] Birch et al., precautionary framework for animal sentience. Animal Welfare, Cambridge (2021)ny
[28] The Moral Status of Animals. plato.stanford.edu/entries/moral-animal ; Rethink Prioritiesny
[29] Towards Evaluating AI Systems Using Self-Reports. arXiv 2311.08576 ; Frontier AI Auditing, arXiv 2601.11699ny
[30] SAE J3016 User Guide. users.ece.cmu.edu/~koopman/j3016ny
[31] Understanding the Operational Design Domain. AVSandbox / Claytexny
[32] The Case for Bounded Autonomy. mongodb.comny
[33] Measuring AI agent autonomy. arXiv 2502.15212ny
[34] Animal Welfare (Sentience) Act 2022. legislation.gov.ukny
[35] Laws that protect animals. Animal Legal Defense Fund (aldf.org)ny
[36] Confucian Ethics as Role-Based Ethics. ResearchGate ; warpweftandway.comny
[37] Distributive Justice (lexical priority; side-constraints). plato.stanford.edu/entries/justice-distributiveny
[38] The Repugnant Conclusion. plato.stanford.edu/entries/repugnant-conclusionny
[39] Hindu Idol as a Juristic Person. legalserviceindia.comny
[40] Te Awa Tupua (Whanganui River) Act 2017. parliament.nzny
[41] R v Dudley and Stephens (1884). casemine.comny
[42] Insanity defense. law.cornell.edu/wex/insanity_defenseny
[43] Kaplow, L., Rules versus Standards: An Economic Analysis, 42 Duke Law Journal 557 (1992)ny
[44] Walzer, M., Thick and Thin: Moral Argument at Home and Abroad (1994)ny
[45] John Rawls; Overlapping consensus. plato.stanford.edu/entries/rawlsny
[46] Nussbaum, M., capabilities approach. Internet Encyclopedia of Philosophy (iep.utm.edu)ny
[47] Mathews v. Eldridge, 424 U.S. 319 (1976). law.cornell.eduny
[48] Bounded Rationality (Simon). plato.stanford.edu/entries/bounded-rationalityny
[49] Gawande, A., The Checklist Manifesto (2009)ny
[50] The Bicentennial Man. en.wikipedia.org/wiki/The_Bicentennial_Manny
[51] The Measure of a Man (TNG). en.wikipedia.orgny
[52] Mind (Culture series); Banks, I. M., Excession.ny
[53] Egan, G., Permutation City (1994).ny
[54] Kurki, V., A Theory of Legal Personhood (OUP, 2019)ny
[55] Bai et al., Constitutional AI. arXiv 2212.08073 ; Anthropic, Collective Constitutional AIny
Quick Links
Abstract
Introduction
Structural Fails
Rank 2 Profile
The Axes
Axis Table
Procedure
Obligations
Act & Blame
Who Judges
NO Ruling
Pluralism
Tractability
Validation
Classification
Lanscape
Limitations
Conclusion
Connect
References