
← The Modern Therapist's Survival Guide with Curt Widhalm and Katie Vernoy13. Juli · 47 Min.
Why AI Mental Health Chatbots Fail When It Matters Most: The Hidden Vulnerabilities Stress-Testing Reveals – An Interview with Shirali and Arul Nigam of Circuit Breaker Labs
Why AI Mental Health Chatbots Fail When It Matters Most: The Hidden Vulnerabilities Stress-Testing Reveals - An Interview with Shirali and Arul Nigam of Circuit Breaker Labs
Shirali and Arul Nigam of Circuit Breaker Labs on why AI mental health chatbots fail, how stress-testing exposes their hidden vulnerabilities, and what therapists need to know.
Curt and Katie talk with Shirali and Arul Nigam, the sibling co-founders of Circuit Breaker Labs, about what therapists tend to get wrong about AI, why the safety infrastructure behind many mental health chatbots is weaker than it looks, and how their team stress-tests these tools to find dangerous failures before real users ever encounter them.
Generative AI is probabilistic, so the same prompt can return a safe answer one moment and a harmful one the next. Shirali and Arul explain how guardrails get bypassed by a misspelled word, a teenager's slang, or the hundredth message in a long conversation, why mental health chatbots tend to fail in the moments that matter most, and what stress-testing hundreds of thousands of simulated conversations actually reveals about model safety.
The conversation closes on what clinicians can do now, why clinical insight is the missing ingredient in AI safety, and why third-party validation is becoming the standard regulators and developers expect. Used well, AI can be a supplement to care or a gateway to a human therapist, but it is not a replacement, and getting there safely starts with building clinical insight in from the foundation.
In this episode, we discuss:
- Why people usually turn to AI in place of no care, not in place of a therapist
- Why generative AI's unpredictability, not a single bad answer, is the real safety problem
- How a misspelling, slang, or a long conversation can slip past chatbot guardrails
- Why AI mental health chatbots tend to fail in the highest-risk moments
- What stress-testing hundreds of thousands of conversations reveals about model safety
- Why clinical insight is the missing ingredient, and what clinicians can do now