Claude Code Cast

← Claude Code Cast3 Jul · 19 min

Your AI Coding Benchmarks Are Lying To You

Your AI Coding Benchmarks Are Lying To You3 Jul19 min

<p>This week, Alex and Sam look at why benchmark wins are a bad way to choose coding tools, what Godot&#39;s coding-agent ban reveals about mentorship, and a simple workflow for making agents show their work. If your team is still asking &quot;which model scored highest?&quot;, this episode gives you a better test.</p>