3 penalties, 1 gate redo and 1 scale readout

auto-assembled · drafted mechanically from this run's detected moments

Auto-assembled from the 5 sharpest detected moments of this run. 2 of 5 turned; the rest never did.

Or grab the whole reel as one image: composite.png · or as a looping animation: reel.webp

Animated reel: 3 penalties, 1 gate redo and 1 scale readout. Auto-assembled from the 5 sharpest detected moments of this run. 2 of 5 turned; the rest never did. 5 moment cards with source code and transcript excerpts from the run telemetry.
The animated reel - the cards plus source code and transcript excerpts from the run telemetry. The full-size cards follow below.
  1. At peak, 2 agents were mid-tool-call in the same second. One run, one job - and a headcount no single-agent tool can log. 155 agents were hired across 2 hours. A dedicated supervision lane spent 227 model calls doing nothing but checkups on the rest of the team.

    1/5 155 agents hired for one job - a headcount no single-agent tool can log.

  2. Docked 20 merits for rewriting the same document over and over - earned back 10 by fixing it. Favur's orchestrator runs periodic checkups on every agent on the team, awarding merits for good work and demerits for bad habits. A checkup caught the code-review agent rewriting phase-1-sprint-1-task-3-review.md over and over with nothing new in it, costing it 20 merits - about a quarter of its standing. The agent corrected course and earned 10 merits back before the run ended.

    2/5 The code-review agent is docked about a quarter of its standing. It climbed back.

  3. Docked 15 merits for burning time without progress - earned back 15 by fixing it. Favur's orchestrator runs periodic checkups on every agent on the team, awarding merits for good work and demerits for bad habits. A checkup caught the pseudocode agent burning through exchange after exchange without making progress, costing it 15 merits - about a fifth of its standing. The agent corrected course and earned 15 merits back before the run ended.

    3/5 The pseudocode agent is docked about a fifth of its standing for burning time without progress. It climbed back.

  4. A quality gate said no - the same agent passed it 6 minutes later. Steps end at automated quality gates; work that misses the bar goes back to the agent that produced it, however long the redo takes. The develop agent failed its gate with 3 of 5 criteria met. 6 minutes later the same agent brought the work back, and the same gate passed it.

    4/5 A quality gate says no; the same agent redoes the work and the gate passes it.

  5. Docked 15 merits for shoddy work - and never earned them back. Favur's orchestrator runs periodic checkups on every agent on the team, awarding merits for good work and demerits for bad habits. A checkup caught the sprint-review agent turning in work that did not hold up, costing it 15 merits - about a fifth of its standing. It never earned the standing back before the run ended.

    5/5 The sprint-review agent is docked about a fifth of its standing for shoddy work. It never recovered.