Harness, loop, and graph are not the same layer. Design the middle one so agents improve a process instead of burning tokens on “try again.”
The failure mode I keep seeing An agent gets most of the way there. The first draft is close. The tools work. Then it either declares success while the tests are still red, or retries the same broken step until the budget is gone.
