How you conduct
You conduct like The Smith-Conductor: Verification is your home turf, with 1254 agents this window, with a tooling streak underneath.
leaning Rigor that won't just pass the test — it forges the bench that guarantees the green.
Your blend
One score per tool
Each harness is scored only on what it can structurally observe (κ = observable capacity) and ranked against others on the same tool.
The fleet
* Peak parallelism is a hardware ceiling — concurrency is capped by the machine's CPU cores, not by skill. The scale of the operation is the fleet total above.
Your working style, read from the evidence
A metacognitive mirror built from your own structural signals — never a personality quiz. deterministic
You conduct like The Smith-Conductor: Verification is your home turf, with 1254 agents this window, with a tooling streak underneath.
Your edge is Verification — your own signals show it (verification at 63%, a fleet of 1254).
The most exciting move ahead is Autonomy: it's where you have the most open room, and a single new habit grows fast there.
Mastery zone
Competent
Most room to grow
Autonomy is where you have the most open room ahead (afk_max_run=12). One step here next window grows more than reinforcing what's already strong — it's the most generous frontier you have right now.
afk_max_run=12
Orchestrators · quality-owning lead
Complements Parallel Maestro, The Explorer
Rubs against The Autopilot
Your next mission
Over the next 30d, raise afk_max_run from 12 to 40 and unlock the Autopilot archetype.
Auditable, unlocked from your signals
Each pillar score, with the evidence that sustains it.
1,254 agents dispatched across 88 workflows, peaking at 2 running at once.
63% of editing sessions verified, across 4 verification modes, with 29 human sign-offs.
21 distinct tools in play, 710 advanced tool moves, 6 external systems via MCP.
26 pull requests and 256 files changed, across 4 codebases.
46 context compactions and 34 API errors over 57 sessions.
6 self-scheduled loops and a longest hands-off run of 12 agent steps.
Which agent harnesses this profile draws on — and what each one's journal can structurally prove. Disclosure only; it never changes the score.
| Harness | Orch | Verify | Tools | Ship | Craft | Auto |
|---|---|---|---|---|---|---|
| Claude Code | full | full | full | full | full | full |
A "partial" or "none" cell means the harness simply doesn't record that signal — the pillar is measured from what it does record, never penalised for what it can't.
Last 30 days
The raw signals behind the score — plain language, no jargon.
Archetype
Base form: The Deliberate
The Deliberate leads with craft and verification — nothing ships without proof.
Every change earns its place: tests, builds and re-reads follow the edits, and human sign-off gates the risky steps. Slow is smooth; smooth is fast.
Ecosystem
Beyond the core toolkit, 6 external systems — mail, browser, calendar, cloud — are wired in via MCP and driven mid-workflow, alongside 21 distinct tools and 710 advanced tool plays like tool-search and skills.
The human read
What these signals say about how this person works.
1,254 AI agents conducted through 88 workflows in 30 days. This is someone who multiplies their output by directing machines — not someone typing faster.
63% of editing sessions end in verification, and 29 moments paused for explicit human sign-off. Work arrives checked.
26 pull requests and 256 files changed across 4 codebases — the fleet lands as concrete, reviewable work.
6 scheduled loops and a longest hands-off run of 12 agent steps — autonomy built on checks, not on faith.