Dan Luu: naming TDD or formal methods rarely helps coding agents test
Across 26 prompt conditions on a Rust Zstd eval (Codex + GPT-5.6 Sol), telling agents to use TDD, formal methods, or popular test skills mostly underperformed a bare default prompt — agents used the named techniques poorly and still leaned on weak unit tests.