Even worse, when I ask one of the models to review the issue of another, the reviewing model invariably finds all kinds of problems with the first model’s work, leading to yet more multi-step fixes. I’m a Claude Pro and ChatGPT Plus subscriber ($20/month each), and lately I’m blowing half my weekly allowance chasing bugs found by Claude Opus 5 and GPT-5.6. Naturally, Opus 5 spat out a fairly lengthy six-point plan for the prompt, and when I asked GPT-5.6 Sol to review it, it gave me a dozen detailed criticisms. Telling GPT-5.6 to go back and follow the instructions in the original prompt, it quickly backtracked (“You’re right—I gold-plated the review”) and narrowed the list to three. Here’s the final prompt that Opus 5 and GPT-5.6 Sol agreed upon, pared down to four bullet points from the original six (and yes, it borrowed the “body shop” metaphor I originally gave it):