AgenTomte

July 20, 2026 · 2 min read

What 5,450 agent runs with zero failures taught us about AI operations

By Sahan, co-founder, systems and delivery

Between 2026-06-21 and 2026-07-04, the agent fleet that runs our three-company group completed 5,450 runs without a single failure. The number is audited by our own logging, and the full case study explains what the fleet does. This post is about what the number taught us, because none of the lessons were about prompts.

Zero failures is an operations property, not a model property

The models fail plenty. They misread, hallucinate, and time out. The zero is produced by the layer around them: inputs validated before a run starts, outputs checked before they touch real data, and a stop-and-ask default when anything is out of distribution. A run that halts loudly and escalates is a success. The failure we engineered against is the silent one that corrupts a ledger and gets found in month three.

If you take one sentence from this post: buy or build the discipline layer first, agents second.

Boring agents compound; impressive agents demo

Our highest-value agents are embarrassingly unimpressive: reconcile these numbers daily, publish this content through this QA gate, triage this intake. Each saves minutes-to-hours per day, forever, with no drama. The flashy autonomous “do my whole job” agent does not appear anywhere in the fleet, because it cannot pass the checks the boring ones pass.

Registration beats memory

Every agent is registered: an owner, a scope, a kill switch, a log. At 5 agents you can hold the fleet in your head. At 41 you cannot, and an unregistered automation is a future incident with no owner. This is the least glamorous investment we made and the one we would defend hardest.

Humans moved up, not out

Two humans run this fleet. Neither writes production output anymore; both architect, review, and approve. That is the honest version of “AI replaced the work”: the production layer went to agents, the judgment layer concentrated. Companies planning for AI should plan for that shape, not for empty desks.

The commercial footnote

This operating model is literally what we sell as the Fractional AI Officer retainer: roadmap, one shipped automation per month, monitoring, weekly written brief. We are not proposing to learn it on your business. The 5,450 runs were the tuition, and we already paid it.

Tell us what you want automated

Describe the work in writing. You get a written reply within one business day: a fixed-price proposal, a scoping question, or an honest referral out.

Start at /start

▸ written reply within one business day · no call scheduled, ever

Doesn't fit a package? Tell us what you need anyway.

Questions? Ask in writing

no chatbot · a human replies

Ask us anything, in writing

A founder replies within one business day. That is the same promise clients get.