Skip to content
meirlabs

supercharge

Don't stop until every critic is wowed by the real benchmark

v0.1.0

Most "make this better" loops stop when the model runs out of ideas, not when the work is actually good. This one replaces that with a quality gate. The task is split into dimensions, one builder subagent owns each, and a separate harsh critic judges each dimension blind against a named real benchmark — Linear's dashboard, an actual published cold email, Stripe's docs page — fetched fresh rather than recalled. Wowed means you would pick ours in the blind comparison, or genuinely cannot tell which is the benchmark; "almost as good" is a fail, and "good enough" is not a verdict that exists. Invoking it is explicit opt-in to large fan-outs and ultracode-scale spend, and it overrides the usual proportionality rule for that one task.

npx @meir-labs/skill-supercharge
GitHub

What it does

  • names a real best-in-class benchmark up front, then decomposes the task into quality dimensions with one builder subagent each
  • pairs every builder with a separate critic whose default verdict is NOT wowed, judging blind side-by-side against the actual benchmark, never a rubric in isolation
  • loops on the quality gate, not a round counter — while (!wowed), with stalling escalating strategy (fresh opus re-diagnosis, a structurally different approach) instead of shipping
  • runs a final integration critic on the assembled whole, and reports a partial pass as NOT DONE with the failing verdicts quoted, never rounded up to a win