I've been doing TDD and clean architecture for years. When I started using Claude Code to build a real application, I wanted to create it the right way from day one - test-first, walking skeleton, steps sized for review.
And it worked. Partially. The agent could explain TDD better than most developers, but then skipped the red phase in the same session. And when I noticed, it apologized. And did that again.
It also avoided my recommendation to change only a few files at one go. And when I noticed - this one was hard to hide - it apologized. Then did it again.
Code agents do what we ask, until they decide not to. And that put our product at risk.
The good news: My way - using the classic hits - works. I just needed to teach it to my agent. And that requires putting reinforcements in place. Measure things and report them on the way, and make tasks small so they can be reviewed without putting the human to sleep.
In this session I'll show the steps, prompts, processes I set in the project to implement my ideal work process. I'll show how I enforced the TDD loop. How I made sure no extra code was created. The structured planning. And the thought process behind making the automation controllable and reviewable.
I'll also show how the project evolved to support my learning and retaining knowledge, performing diagnostics and tracking decisions.
I'm now producing code I can trust. And you can take those techniques and use them in your process too.
This talk has been presented at AI Coding Summit Berlin, check out the latest edition of this Tech Conference.



















