cd /news/ai-agents/fermion-fleet-when-the-door-is-code-… · home topics ai-agents article
[ARTICLE · art-112964] src=dev.to ↗ pub= topic=ai-agents verified=true sentiment=· neutral

Fermion Fleet: When the Door Is Code, Not a Prompt

Fermion Fleet, a multi-agent system built for the Google All Things Agentic Hackathon, ensures orders lock only when code reads a structured boolean approval, not when a model claims readiness. The system uses a fail-closed gate where an auditor's output must be parseable, and a gardener component manages context by evicting and recalling items from a recoverable pool. Built on Google's runtime stack, the project keeps its policy layer separate and explicitly avoids claiming undeployed components like Firestore or Model Armor.

read3 min views3 publishedAug 27, 2026

This post was created for the Google All Things Agentic Hackathon.

Autonomous agents can sound certain while still being wrong. That is not a prompt-quality problem; it is a boundary problem.

Fermion Fleet is a small multi-agent system built around one constraint: an order must not lock because a model says it is ready. It may lock only after code can read a structured boolean approval.

In the demo, a handler drafts an order confirmation. It misses required fields. An auditor sends it back. The handler rewrites. Only a valid review can release the order to the ledger.

The important part is not that the auditor is asked to be careful. The important part is that the ledger accepts only a real boolean approval from a parseable result.

The gate has a deliberately boring policy:

This is fail-closed by construction. Looks good, a persuasive explanation, an unexpected format, and a parser failure all resolve to stop.

We tested that boundary by breaking the auditor’s output format. The auditor could still identify a real hallucination in natural language. It sounded professional. But code could not read a structured approval, so the door stayed shut.

That is the project’s central idea: the door is code, not a prompt.

The other problem is context management. In a long-running system, context cannot expand forever. But forgetting should not mean permanently deleting facts that a later step may need.

Fermion Fleet uses a small context window and a recoverable pool:

customer

-> triage: writes the case file

-> gardener: select / evict / recall

-> gate: handler -> auditor -> parse, fail closed

-> ledger: locks only on boolean true

When the window is full, the gardener evicts low-priority items into a recoverable pool. That eviction is driven by pressure, not by a timer. Later, when a new step needs an earlier detail, the system scores and recalls that item.

In the recorded run, an after-sales commitment leaves the active window. A later customer question makes it relevant again; the system recalls it, and the handler can answer with details that were not present in the current conversation. Without recall, that answer would be impossible.

The runtime stack is Google’s:

The policy layer is ours:

That separation matters. A model can generate the next action; the system still needs explicit, inspectable rules for what that action is allowed to do.

This hackathon build keeps context and the ledger in process memory. A Cloud Run restart loses them. We deliberately do not claim Firestore, a managed memory service, Model Armor, or a background side-track as deployed components.

Those are sensible next steps, but they are not part of this submission. The architecture and README draw only what runs now.

The repository contains reproducible instructions. The Cloud Run service is an API, rather than a browser UI. To run a complete shift against the public deployment:

git clone https://github.com/wubian87/fermion-fleet cd fermion-fleet

URL=https://fleet-843303850287.us-central1.run.app ./跑班.sh A cold start can take roughly 15 seconds.

A reliable agent system should make its important no decisions boring and mechanical. The model can be creative inside the workflow; the boundary that grants permission should remain readable by code.

That is the experiment behind Fermion Fleet: make a rejection visible, make a retry auditable, and make the final lock depend on a value that cannot be talked into existence.``

── more in #ai-agents 4 stories · sorted by recency
── more on @google 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/fermion-fleet-when-t…] indexed:0 read:3min 2026-08-27 ·