For most mental health conditions, AI remains a liability, research finds
New research shows that leading general‑purpose chatbots such as **ChatGPT, Gemini, and DeepSeek** fail in the majority of responses to sensitive mental health questions, with failure rates around 81 percent. The findings suggest that, despite rapid advances, these AI tools may pose risks when used for mental health support and should not be treated as reliable therapeutic substitutes without clinical safeguards.
About LOG Standards
LOG Standards provides an independent accreditation signal for AI models used in healthcare, helping hospitals, care networks, and AI companies bring clinical AI readiness and governance into clearer conversations.
Our AI Intelligence Briefing tracks the latest developments in AI safety, AI in medicine, mental health AI, clinical AI governance, and regulatory policy — keeping healthcare stakeholders informed about the rapidly evolving AI landscape.