Agentic QA: letting an LLM drive the regression suite
Six months of running an agent against a Playwright suite — what broke, what held, and the guardrails I keep.
Long-form notes on building and testing software with agents in the loop.
Six months of running an agent against a Playwright suite — what broke, what held, and the guardrails I keep.
Flakiness is rarely random. It is usually a race you have not named yet.
Prompting is scaffolding. The interesting part is what you refuse to accept back.
Compile-time route checking removed a whole class of bugs from my own site.
If the plan does not fit on a page, nobody reads it — including me.