Trust, but Grep: The Plan Was Perfect and the Premise Was False
Part 8 of Building My Agentic Workflow in Public. Part 7 audited a painful bill and found my own specs were the culprit. This is what round two of testing revealed: the plan was not wrong because it thought badly. It was wrong because it never looked. After the retro in Part 7, I rebuilt my pipeline around a set of new rules and ran the whole thing again on the same two projects, SleemAI and Morpheus Observability. Round two went better, and I will share the full numbers in a coming post. But one finding from the audit was so simple, and so embarrassing, that it deserves its own short post. ...
Read blog post ›