I am asking that partially as a rhetorical question but also wondering your thoughts on when to deploy 'loop engineering' versus 'one-shot' versus writing the code yourself.
Additionally the post seems to talk in broad strokes without a specific proposal on a 'scorer' / determining adherence to the state goal. Do you have a thought on how that should be expressed? Do you feel that a human in the loop slows things down unnecessarily?
AI slop tends to do that.
But at this point the AI most definitely understands the code and spots bugs better than I can, and it does seem like I can just let an AI bureaucracy run and then have a high-level look over its ideas and things will turn out just fine. If it isn't reliable enough, the problem nowadays is solvable just by adding more review agents.
"What is therefore required is not merely a more capable agent, but a closed-loop capability orchestration substrate capable of continuously reconciling observed product reality against an evolving demand manifold."