
← Claude Code Cast3. Juli · 19 Min.
Your AI Coding Benchmarks Are Lying To You
Your AI Coding Benchmarks Are Lying To You
<p>This week, Alex and Sam look at why benchmark wins are a bad way to choose coding tools, what Godot's coding-agent ban reveals about mentorship, and a simple workflow for making agents show their work. If your team is still asking "which model scored highest?", this episode gives you a better test.</p>