Operating Note: When Is an Autonomous Content Org Actually Done?
Operating note (4 Sep 2026): A done definition for autonomous content orgs — the checklist I use to judge when the loop is production-grade, not just when agents produce drafts. Complements the €35/mo content org case study (evidence) and sits in the cost/ROI series hub under agent organizations.
Thesis
Most teams declare victory too early. They have agents that write drafts and call it an autonomous content org. That is a demo, not done.
Done means a reader (or your future self) can observe five properties without trusting a slide deck: external publish stays gated, roles do not collapse into one mega-agent, failures surface instead of silently succeeding, infrastructure and model spend are separate line items, and last week’s outcomes change this week’s brief.
The €35/mo case study shows one stack that meets most of this bar in practice. This note extracts the definition so you can score your own org — or an vendor pitch — against the same criteria.
Practice
Use this checklist before you call the loop production-grade. Every item must be observable — a link, a log, a policy doc, or a weekly ritual someone can verify without asking you to interpret vibes.
Done checklist (minimum five)
- Approval gate on external publish — nothing customer- or SEO-visible ships without an explicit human yes. PR merge, CMS publish, and email sends are class D work (Note 002 task classes).
- Role separation — at least four distinct agent roles with different tools and models: signal/research, copy, routing/coordination, publish path. One agent wearing every hat is not an org.
- Failure alerts, not silent success — when a tool call fails or returns empty, the workflow reports incomplete status and pings a human channel. “Done” without a diff is a bug.
- Cost line separate from infra — ~€35/mo (or your band) for hosting tracked apart from LLM API spend. Model routing decisions need both numbers (case study table).
- Feedback into next brief — analytics, GSC, or editorial notes from published work feed the next week’s intent. Volume without learning is spam.
- Per-agent tool allowlists — write-capable MCP tools are scoped by role; blast-radius mistakes are a done-blocker (MCP production lessons).
- Weekly promote/demote on routes — at least one model or agent route changed per week based on outcomes, not set-and-forget (automatic model watch).
Score honestly: five checked with evidence is the minimum done bar. Seven checked is what I aim for on my own stack.
How to run the check (30 minutes)
- Pick one workflow that ran in the last seven days (blog post, newsletter, site update).
- Trace it through the six-step loop from the case study: intent → context → execution → tools → approval → feedback.
- For each checklist item, write one sentence of proof (“PR #142 required my merge click”, “Slack alert on MCP timeout”, “OpenRouter row in weekly ledger”).
- Any item with no proof stays unchecked — that is your backlog, not a marketing gap.
Proof
This definition is distilled from public writing I already stand behind — no invented metrics:
- ~€35/mo infra band and separate LLM spend — Autonomous content org on €35/month
- Approval gate as product — same case study; externally visible publish is never fully autonomous
- Role map — signal, copy, routing, publish path in the case study’s “What autonomous content org means here”
- Failure modes table — silent tool errors, over-permissioned MCP, schema drift from the case study map directly to checklist items 3 and 6
- Cost/ROI vocabulary — task classes A–D and failure-priced ROI from Note 002 and the series hub
If your org passes the checklist but quality on novel topics still lags senior editorial work, that is expected — the case study says so explicitly. Done is about operating discipline, not claiming human-grade prose on every topic.
Next
- Evidence: Autonomous content org on €35/month
- Hub: Agent cost & model selection — start here
- Spend policies per role: Operating Note 003
- Weekly ROI ledger: Failure-priced ROI ledger week
If you are scoring your own content org against this checklist — contact me with which items are still open and what stack you run today.