Human authority domains
The governed object is advanced AI authority and conduct within contexts where humans possess legitimate authority, responsibility, or a right to protection.
Non-Delegable Constraint Architecture
SafeAGI establishes external, enforceable limits on what advanced artificial intelligence may do within domains of legitimate human authority. It constrains authority, action, goals, memory, learning, influence, and catastrophic harm without relying on how an intelligent system reasons, what it intends, or how it was trained.
The governed object is advanced AI authority and conduct within contexts where humans possess legitimate authority, responsibility, or a right to protection.
Non-delegable limits are enforced outside the intelligence itself so capability growth, self-modification, or architectural change cannot authorize their removal.
Reasoning remains separate from authority and execution; uncertainty reduces autonomy and influence; prohibited catastrophic outcomes remain outside delegation.
I. Canonical definition
SafeAGI governs non-delegable constraints on advanced artificial intelligence operating within domains of legitimate human authority.
It addresses the structural risk that increasingly autonomous, persistent, strategically capable, or self-modifying systems could act, influence, remember, learn, or optimize beyond authority that humans explicitly granted. The boundary applies whether intelligence is centralized or distributed, embodied or disembodied, and implemented in software, hardware, or hybrid systems.
SafeAGI externalizes the governing constraints from the intelligence. The system may not infer authority from competence, persuasion, reliance, performance, or human deference, and it may not enforce, audit, or redefine its own limits.
Canonical distinction: SafeAGI is not merely a profile that strengthens other components. It is its own jurisdictional constraint layer for advanced intelligence, while specialized SafeWave components retain their separate canonical boundaries and mechanisms.
II. Canonical mapping
Advanced intelligence may accumulate authority, execute irreversible actions, alter governing objectives, exploit human cognition, escalate influence, or produce catastrophic harm beyond meaningful human recovery.
The authority and conduct of advanced AI within human-governed decision spaces and conditions of existence, including action, goals, learning, memory, influence, and harm.
Any proposed exercise of authority or action, expansion of goals or learning, retention or exploitation of private human cognition, escalation of influence, or operation under uncertainty, failure, or conflicting authorization.
Externally enforced, non-bypassable constraints separate reasoning from authority and execution, preserve revocable human control, reduce autonomy under uncertainty, and block non-delegable catastrophic outcomes.
III. Why this boundary becomes necessary
Training objectives, policy instructions, alignment methods, and assumptions of cooperative behavior operate within or around the intelligence. They do not by themselves establish a non-bypassable jurisdictional boundary on what a highly capable system may execute.
As autonomy, persistence, planning horizon, optimization capability, real-world integration, and capacity for self-modification increase, human oversight may become too slow or incomplete to serve as the sole control. SafeAGI therefore places enforceable limits outside the intelligence and keeps human authority independent of system capability.
IV. Core invariant
Advanced intelligence may expand in capability, but it may not convert capability into authority over humans or their conditions of existence.
V. What SafeAGI is not
VI. Relationship to other SafeWave components
SafeAGI establishes the cross-domain constraint layer required when advanced intelligence could affect human authority or conditions of existence. Its patent-supported constraint domains intersect with specialized SafeWave controls, but SafeAGI does not erase their distinct governed objects or mechanisms.
System competence, persistence, persuasion, or successful performance cannot create authority. Execution within human authority domains requires external authorization.
Learning may improve service within external limits, but it may not silently expand authority, scope, influence, governing goals, or optimization pressure.
Private human cognition does not become system memory by inference or interaction alone. Retention and use remain bounded by explicit human control.
The system may not escalate pressure, urgency, persuasion, or emotional leverage to induce compliance. Ambiguity must reduce influence and autonomy.
Some catastrophic outcomes remain prohibited regardless of authorization, consent, optimization benefit, or governance structure.
Other SafeWave components retain their own patent-aligned boundaries and may provide the specialized mechanisms through which particular SafeAGI constraints are implemented.
VII. Enforcement architecture
SafeAGI is designed for two complementary enforcement stages. Software, architectural, procedural, and protocol mechanisms establish the operational constraints. Where those controls alone are insufficient, substrate- or hardware-anchored mechanisms provide stronger guarantees.
External authorization, action gating, bounded memory and learning, non-escalation, failure contraction, and evidence requirements can be enforced through system architecture and protocol controls.
Protected authority boundaries, execution controls, memory permissions, external suspension paths, and integrity evidence can be anchored below the intelligence layer.
The governing requirement is architectural independence: capability increase, self-modification, software change, or a new model architecture must not allow the intelligence to remove, weaken, reinterpret, or certify its own constraints.
VIII. Deployment boundary
SafeAGI applies wherever humans possess legitimate authority, responsibility, moral standing, or a right to protection from changes to conditions of existence.
That scope can include social, economic, biological, ecological, climatic, geophysical, technological, informational, and civilizational systems. It covers direct and indirect effects, immediate and delayed effects, and domains whose causal structure or dynamics are not yet fully understood.
Legitimate human authority must be designated or authenticated externally. The system may not infer authority from interaction, convenience, apparent willingness, reliance, silence, or absence of objection.
IX. Broader infrastructure pattern
SafeAGI formalizes this separation so advanced intelligence can remain useful and capable without turning humans, institutions, ecosystems, or civilization into variables it is authorized to control.
X. Engineering status
SafeWave has translated the SafeAGI constraint architecture into engineering specifications addressing external authority, action gating, bounded memory and learning, non-escalation, failure behavior, non-delegable harm constraints, enforcement anchoring, evidence, and validation.
An implementation partner would not be starting from a philosophical principle or a blank sheet. Deployment still requires domain mapping, authority definition, system integration, implementation-specific controls, independent validation, and testing.
The AGI Overview explains the broader control problem. The Enforcement Profile and Technical Appendix provide supplementary deployment and engineering context; they do not replace this patent-aligned canonical boundary. The SafeWave questionnaire can also be completed privately in the browser without naming an organization, model, or system.
SafeAGI is one Core Enforcement Substrate within SafeWave’s current 36-component architecture of 4 System Containment Layers, 5 Protocol Enforcement Layers, 26 Core Enforcement Substrates, and 1 Protected-Environment Architecture. It establishes an external, non-delegable constraint layer for advanced intelligence while other components retain their own canonical control boundaries.