Maturity Level 5 of 6

Autonomous

Self-optimising systems with policy-level human oversight.

Level 5 (Autonomous) is where systems start to run themselves within guardrails — planning models pick tools, continuous evaluation catches drift, and humans set policy rather than tasks.

Framework v2026.1 · Updated · machine-readable spec

What this level looks like in practice

The organization has confidence to delegate real decisions to AI within bounded scopes. This is where reasoning models earn their premium — planning-heavy tasks that would require a human orchestrator now run autonomously.

Level 5 across all six dimensions

  • Strategy & Leadership

    Board reviews GenAI value delivery quarterly. GenAI is a P&L line, not a cost line.

  • Data & Infrastructure

    Pipelines self-monitor for drift and freshness. Data platform team treats AI as a first-class customer.

  • Use Cases & Applications

    Continuous evaluation pipeline auto-flags regressions. Structured intake for new use cases.

  • Talent & Culture

    Continuous learning tied to promotion criteria. Internal community publishes case studies.

  • Governance & Risk

    Governance is a competitive advantage — externally audited, referenced in procurement.

  • Agentic AI

    Planning-capable reasoning models drive dynamic tool selection. Continuous evaluation improves prompts and workflows.

Common blockers to the next level

  • Operating model is still legacy — GenAI is inserted into existing workflows rather than reshaping them.
  • New product concepts still originate from human strategy, not from AI capability.
  • Multi-agent orchestration is one team's expertise, not a broad engineering discipline.

Worked example — Technology

A B2B SaaS company's customer support runs a fleet of agents that classify tickets, gather context from internal systems, draft responses, and escalate to humans only when confidence drops below policy thresholds. Agent evals run on every deployment. A reasoning model handles multi-step debugging tasks that used to require a senior engineer.

For developers

Call this framework and its tools from your own agent via the Model Context Protocol (MCP) server. Works with Claude Desktop, Cursor, Zed, Continue, and the OpenAI Agents SDK.