Earlier quoted context omitted.
You didn't think they were the wrong tools when you wrote it. You said "this example is simple enough that both Claude and ChatGPT can implement it using their default code environments". From what I gather, a lot of people are using these code assistance tools because they too are in a hurry, under pressure from management forcing them to go faster with AI, and with limited ability to push back. You have significant…
I stand by what I originally wrote: the example was simple enough for ChatGPT and Claude do implement reasonably well. They didn't implement it well enough for people not to pick them apart though, which is a distraction from the concept I'm trying to demonstrate. This is honestly the biggest challenge in writing about this stuff, especially if you're doing it in public. Any example is an opportunity for people to fi…
It is not, however, something equivalent to the product of a disciplined TDD practitioner. Not even close.
You write that test-first development helps protect against two risks of code agents, but what does that mean for your specific example?
How is the final product better than the test-after prompt "Build a Python function to extract headers from a markdown string, then write a complete and robust test suite."
Otherwise, how do you know it's a "fantastic fit for coding agents" or that it gets "better results out of a coding agent"?