Summary
When I put AI agents in charge of polishing quality, the design that worked best was splitting the work into two agents: one that makes the fixes, and one that judges the result. If one agent both fixes and judges, it grades its own fixes too kindly, and you lose confidence in its verdict that quality has converged. Converged here means the state where new findings stop coming up for a run of rounds. So I built it so that the judging agent never sees a word of the fixing agent’s explanations, and decides pass or fail only from numbers measured by machine from the built product and from facts about the structure of the screen.
Here is what I measured over 10 rounds of this loop in one night. The count of new findings the judging agent detected went 4, 2, 3, 2, 2, 1, 0, 2, 2, 0, and it did not fall monotonically. After reaching zero once, it went back up. And at the end I stopped the loop myself, before it met the mechanical condition for convergence.
This article explains how to split the fixing role from the judging role so that self-judgment can be trusted. It also covers how the loop failed to stop on the mechanical condition for convergence alone. This article is a discussion based on the operating records of a quality loop I actually ran for one night.
Who this is for and what you can take away
This article is for people designing a mechanism that has AI agents raise quality through repetition. You will see why the fixing role and the judging role are split, and what the judging agent is shown and what it is kept away from. You can also take away how to combine a mechanical condition for convergence with human judgment.
The rest of this article is paid
You can read the rest by buying this article on its own, or with a subscription that covers every paid article.
The paid part is about 5,700 characters, roughly a 11-minute read.
Read just this article
From 300 JPY
Buy this article on its own. The exact price is shown at checkout. Purchased articles stay readable whenever you sign in with the email used at purchase.
Read every paid article
Standard is 490 JPY / month
Four new paid articles ship every month. For yearly billing and the full comparison, see the plans.
Add dialogue and columns
Premium is 980 JPY / month
Premium adds the subscriber-only columns and dialogue with matsumotory-kun on top of every paid article. See the plans for details.
Prices include tax. Purchases and subscriptions start after you sign in. Subscribers and readers who already bought this article can sign in and read the full article; for the plans, see the plans page.
Sales are available in supported regions only; the Terms list where we sell. All charges are in Japanese yen.