I Ran the Numbers: TDD Inside an AI Agent Loop Costs 3-8x More Tokens, With No Quality Win
Thoughtworks ran a controlled experiment on TDD inside an AI agent loop: red-green cycles cost up to 8.5x more tokens than spec-first prompting, and blind quality judging preferred the non-TDD output. Here's what that means for how you should actually prompt coding agents.