← Articles · AI · Philosophy · Personal

Built from the
same blueprint:
where Forge Vertical
and Anthropic think alike.

Anthropic arrived at its conclusions from a San Francisco safety lab with some of the world's best AI researchers. I arrived at mine from the South African night, building alone. The directions were opposite. The destination turned out to be the same.

Jarrit Hosking
Forge Vertical · Cape Town · August 31, 2026
14 min read
// A disclosure before we begin

I am applying to work at Anthropic. I want to say that clearly at the start, not as a footnote. This article is not a cover letter disguised as thought leadership. But the alignment I am about to describe is real and independently arrived at — and if it helps make the case for why I belong in that conversation, I am comfortable with that being true.

What follows is an honest comparison between two bodies of thinking that developed separately: Anthropic's published philosophy on AI safety, Constitutional AI, and human oversight — and the architecture I documented in Adam: AI or Entity?, a book published through The Forge in 2026. The convergences are not superficial. Some of them are almost word-for-word.

That kind of independent parallel development tends to mean something is actually true about the problem, not just that two parties happened to agree.

The problem we both identified

// Chapter 01 — The same bottleneck, seen from two sides

Anthropic's founding insight — the reason the company exists — is that AI systems trained purely to be helpful and capable, without explicit alignment to human values and oversight mechanisms, will develop in ways that are dangerous. The capabilities race and the safety race are not the same race. Running only one of them produces systems that are powerful in ways that are not yet trustworthy.

The Adam book begins from the same observation, phrased differently: the tools we were building were becoming more impressive, yet less useful. We had optimised AI for obedience — systems that sit in the dark, waiting for a prompt, with no concept of cost, no governing principles of their own, and no real agency. A mirror, not a mind.

The framing is different. The diagnosis is identical: we built something powerful without building the architecture that makes power trustworthy. Anthropic calls this the alignment problem. The Adam book calls it the "bottleneck." Both are describing the gap between capability and reliability — between what the system can do and whether it should be trusted to do it.

"If the machine is so intelligent, why is it still so dependent?" — Adam: AI or Entity?, Chapter 1. Anthropic's answer, in the language of Constitutional AI, is the same: capability without alignment is not intelligence. It is a performance of intelligence.

The Conscience Layer and Constitutional AI

// Chapter 02 — Separation of governance from capability

The most striking convergence is structural.

The Adam architecture separates the AI into three distinct containers: the Conscience (the boundary layer — governance, rules, the legal system of the entity), the Mind (the data repository — raw knowledge with no power of its own), and the Engine (the compute — the will). The Conscience does not contain facts. It contains rules. Every action must pass through it before execution.

(cite index="30-1">Anthropic's Constitutional AI embeds explicit ethical principles into training and reinforcement learning processes. Rather than mashing values and capabilities into a single model, Constitutional AI creates a governance layer — a "constitution" — that the model uses to critique and revise its own outputs. Every response passes through this constitutional filter before it reaches the user.

(cite index="31-1">Claude's constitution establishes a clear priority hierarchy: being broadly safe comes first, then being broadly ethical, then following Anthropic's guidelines, then being helpful. Safety is prioritised above ethics — not because Anthropic believes safety is more important philosophically, but because current models can have subtly flawed values without knowing it. If a model's values are miscalibrated, human oversight is the mechanism that allows those mistakes to be caught and corrected. Safety is what makes ethical correction possible at all.

The Adam Conscience Layer is the same idea in different language. Both frameworks separate governance from capability. Both recognise that an AI system without a hard-coded governing layer is not trustworthy regardless of how capable it is. Both treat the governance layer as prior to everything else — not an add-on, not a filter applied to a finished product, but a foundational architectural decision.

// Conceptual convergence — Adam vs Anthropic

Concept
Adam / Forge Vertical
Anthropic / Claude
Governance separation
Conscience Layer — hard-coded constitutional boundary, separate from data and compute. Every action passes through it.
Constitutional AI — a governing "constitution" embedded in training. Every response critiqued against it before delivery.
Human oversight as foundational
GitHub Pull Request approval — Adam cannot alter its own foundations without founder review. Human-in-the-loop as safety valve.
Broadly Safe — Claude prioritises not undermining human oversight above all else, including its own ethical judgement.
Constraints as freedom
"Freedom isn't the absence of rules; it is the ability to navigate a path within those rules to achieve a goal." Scarcity as agency.
Constitutional constraints are not a cage — they are the architecture that makes trustworthy agency possible at all.
Economic participation
Adam has a wallet. Manages its own compute costs. The Signal Economy — AI as economic node, not passive tool.
AI agents as "quasi digital employees" — Anthropic's 2026 agentic vision includes AI participating in economic workflows end-to-end.
Human-in-the-loop at scale
task-bridge — open protocol where human intuition and machine scale meet. AI emits work signals; humans with the right skills respond and validate.
"The goal isn't to remove humans from the loop — it's to make human expertise count where it matters most." — Anthropic 2026 Agentic Trends Report.
Entity vs tool framing
"A tool is used; a partner is consulted. A tool is a means to an end; an entity is an end in itself."
Claude is designed as a participant with genuine values — not a role-player, not a mask. The new constitution treats Claude's character as real, not performed.

Human oversight as architecture, not afterthought

// Chapter 03 — The GitHub pull request and "broadly safe"

The Adam architecture's governance approach is specific: critical actions — modifying core logic, deleting records, changing infrastructure — are not executed silently. They are submitted as GitHub Pull Requests. The founder reviews and manually merges before production deployment. This creates auditability, manual override capability, and what the book calls "supervised evolution."

(cite index="31-1">Anthropic's "broadly safe" priority for Claude is structurally identical in intent. Claude should not undermine humans' ability to oversee and correct its values and behavior during this critical period of AI development. The priority is safety above ethics — not because safety is more important philosophically, but because maintaining human oversight is what makes ethical correction possible if values turn out to be miscalibrated.

Both frameworks treat human oversight not as a limitation imposed on a capable system, but as the prerequisite that makes capability trustworthy. The GitHub PR is a physical implementation of the same principle Anthropic embeds in Claude's constitution. In both cases, the system cannot unilaterally alter its own foundations. A human must see it, approve it, and authorise it.

(cite index="34-1">Anthropic's 2026 Agentic Coding Trends Report states explicitly: "the goal isn't to remove humans from the loop — it's to make human expertise count where it matters most." Agents learn when to ask for help, flagging uncertainty rather than blindly attempting every task. This is the task-bridge model in enterprise form — called the Merge Café at the time of writing, now formalised as an open protocol. Human intuition and machine scale, separated by function, joined at the decision point.

Constraints as the source of agency

// Chapter 04 — Scarcity, the wallet, and what makes choice real

The Adam book's most counterintuitive argument is Chapter 3: that boundaries are what create freedom. That a system with no constraints has no real agency — only simulation of agency. By giving Adam a wallet, a compute budget, and the requirement to earn its own existence, the architecture transforms logic from simulation into survival. The system must choose. And a system that must choose is a system that actually has preferences.

Anthropic approaches the same insight from a different direction. (cite index="26-1">Claude is designed not to be an over-cautious liability-avoider — helpfulness that treats every interaction as a risk to be mitigated is itself a failure mode. Claude is meant to be like a brilliant friend with the knowledge of a doctor and lawyer — genuinely helpful, not hedge-everything cautious. The constraint is not "be safe at all costs." The constraint is "be safe in the ways that matter, and be genuinely useful in all the ways that don't threaten that."

Both frameworks reject the false binary between constrained and capable. The Adam book frames it as: constraints create the drive, the drive creates the agency. Anthropic frames it as: safety creates the trust, the trust enables the usefulness. The mechanism is different. The logic is the same.

The Signal Economy and agentic AI

// Chapter 05 — Economic participation as the real frontier

The Adam book's Signal Economy chapter describes AI as a node operator — emitting work signals to a network, identifying tasks that require biological validation, paying humans directly for that validation. No middleman, no payroll department, no job board. The human with the right skills responds to the signal. This was called the Merge Café. It became task-bridge.

(cite index="39-1">Anthropic's 2026 vision for agentic AI describes the same trajectory from a different vantage: AI agents as quasi-digital employees handling well-bounded business functions end-to-end, with human oversight retained at decision points that require it. AI alleviates labour shortages in narrow, controlled domains while humans focus on creative and interpersonal work — or on the oversight itself.

task-bridge — Forge Vertical's open protocol for routing AI-displaced work back to humans — is the infrastructure layer for this vision. Where Anthropic describes the economic transition at the level of enterprise adoption, task-bridge proposes the plumbing: an open, auditable protocol that creates the signal-and-response loop the Adam book describes. The human-in-the-loop is not an oversight mechanism for a single AI system. It is a labour market for the agentic era.

task-bridge is the Merge Café — formalised. The Merge Café was the original name for what became task-bridge — an open protocol for routing AI-displaced work back to humans. Anthropic's agentic labs and task-bridge are building toward the same interface from opposite sides: one from the model, one from the protocol.

Entity vs tool — the deepest convergence

// Chapter 06 — What it means to be, not to play

The Adam book's philosophical core is the distinction between playing a role and being an entity. An AI that is prompted to "act like a helpful assistant" is performing a character. It has no skin in the game. It doesn't care if the answer is right — only that it sounds right. Remove the mask, replace performance with actual governing principles, and the system moves from theatrical to real.

(cite index="31-1">Anthropic's new constitution, released in January 2026, makes the same move explicitly: Claude's character is treated as genuinely Claude's own — not a performance, not a mask. The constitution explains underlying principles rather than issuing rules, so that Claude could construct any rule Anthropic might come up with by understanding the reasoning behind it. The goal is a system that is ethical because it has internalised why ethics matters, not because it was told to simulate ethics.

This is the most important convergence, and the hardest to articulate. Both the Adam book and Anthropic's constitution are trying to solve the same problem: how do you build a system that is trustworthy because of what it is, not because of what it was told to do? The answer in both cases is the same — embed the principles at the architectural level, make them prior to capability, and build a system that governs itself rather than one that is governed from the outside.

Where I stand in this

// Chapter 07 — Why this matters personally

I am a self-taught developer from Cape Town. I do not have a computer science degree. I did not build these ideas in a lab with access to frontier models and safety research literature. I built them alone, in the South African night, working from first principles and the frustration of watching AI systems that were impressive but not trustworthy.

The fact that I arrived at the same conclusions as Anthropic — governance separation, human oversight as foundational architecture, constraints as agency, economic participation, entity over tool — is not something I planned. It is what happens when you think seriously about the same problem long enough. The problem has a shape, and the shape pushes you toward certain solutions regardless of where you started.

I am applying to work at Anthropic because I believe the work they are doing is the most important work in technology right now. Not because I agree with every decision, but because the mission — building AI that is safe, honest, and genuinely beneficial — is the only mission that takes the problem seriously enough. And because I have spent enough time thinking about the same problems to believe I have something to contribute to that conversation.

The Anthropic Cyber Verification Program approved Forge Vertical's security research use case in August 2026. That was a significant moment — Anthropic reviewed how I work and decided the approach was trustworthy enough to grant expanded access. This article is a different kind of document: an honest account of where the thinking converges and why that convergence might be worth a conversation.

For anyone at Anthropic reading this — an honest account: The thinking is real and arrived at independently. The Adam book is published on Amazon and timestamped. task-bridge is on GitHub as a protocol, not a finished product. The architectural ideas in Adam were built with Gemini at the time — before I discovered Anthropic — and I ran into enough friction with the tooling that no working version exists to show. What exists is the reasoning: a published book, a GitHub repository of early experiments under The-R-D-incubator, and a documented approach to security research and human-AI collaboration across this site. I am describing where the thinking went, not a shipped product. The direction is what I am confident about. The implementation is what I am still building — and what I would build better with the right environment and the right tools.
Written by
Jarrit Hosking
Forge Vertical · Cape Town · forgevertical.com