{"id":5107,"date":"2026-08-04T23:10:01","date_gmt":"2026-08-04T23:10:01","guid":{"rendered":"https:\/\/lumeamara.online\/?p=5107"},"modified":"2026-08-04T23:10:02","modified_gmt":"2026-08-04T23:10:02","slug":"the-complexities-of-governing-mental-health-ai-stanford-hai","status":"publish","type":"post","link":"https:\/\/lumeamara.online\/?p=5107","title":{"rendered":"The Complexities of Governing Mental Health AI | Stanford HAI"},"content":{"rendered":"<p>Policymakers, academics, <a href=\"https:\/\/lumeamara.online\/?p=5091\" title=\"How Community Health Centers Expand Access to Care\">health<\/a>care providers, AI developers, and patient advocates convened by Stanford HAI identify critical gaps in how we regulate AI tools used for therapy and emotional support<\/p>\n<p>Millions of Americans <a href=\"https:\/\/www.nami.org\/mental-health-by-the-numbers\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>are affected<\/u><\/a> by mental illness every year, yet the cost of therapy remains <a href=\"https:\/\/pmc.ncbi.nlm.nih.gov\/articles\/PMC11786981\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>out of reach<\/u><\/a> for many, and a <a href=\"https:\/\/www.trillianthealth.com\/hubfs\/2026%20Behavioral%20Health%20Report\/2026%20Behavioral%20Health%20Report%20%7C%20Trilliant%20Health.pdf#page=48\" rel=\"nofollow noopener\" target=\"_blank\"><u>shortage<\/u><\/a> of licensed clinicians means that even those with insurance often wait months for an appointment. Into this void has stepped a rapidly expanding market of AI-powered tools, including chatbots that provide therapeutic counseling and apps that offer cognitive behavioral therapy exercises on demand. Children and adults also seek out general-purpose chatbots like ChatGPT and \u201ccompanion\u201d bots such as those offered by Character.ai and Replika in times of loneliness or emotional distress.<\/p>\n<p>There are promising potential upsides to the use of AI in <a href=\"https:\/\/lumeamara.online\/?p=5084\" title=\"Bedfordshire and London mental health NHS trust to make \u00a324m cuts\">mental health<\/a> care: greater access, lower cost, tools that could extend the reach of an <a href=\"https:\/\/www.parityindex.org\/home\" rel=\"nofollow noopener\" target=\"_blank\"><u>overstretched system<\/u><\/a>, and a form of social and emotional support for people experiencing loneliness. But these promises also entail risk. Absent clear regulation and standardized third-party testing, these tools risk delivering substandard care and putting users <u>in danger<\/u>. News headlines abound about minors developing <a href=\"https:\/\/www.npr.org\/2025\/10\/08\/nx-s1-5561981\/ai-students-schools-teachers\" rel=\"nofollow noopener\" target=\"_blank\"><u>unhealthy emotional attachments<\/u><\/a> to chatbots, users in crisis receiving <a href=\"https:\/\/www.bbc.com\/news\/articles\/ce3xgwyywe4o\" rel=\"nofollow noopener\" target=\"_blank\"><u>harmful or inadequate responses<\/u><\/a>, and research showing that general-purpose AI chatbots commonly <a href=\"https:\/\/www.commonsensemedia.org\/press-releases\/common-sense-media-finds-major-ai-chatbots-unsafe-for-teen-mental-health-support\" rel=\"nofollow noopener\" target=\"_blank\"><u>miss warning signs<\/u><\/a>.<\/p>\n<p>AI\u2019s role in mental health care is growing fast, and legislators are struggling to keep pace. To date, most legislative activities have happened in states, which have introduced <a href=\"https:\/\/pubmed.ncbi.nlm.nih.gov\/41172342\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>more than 140 bills<\/u><\/a> related to AI use in mental health contexts. Federal <a href=\"https:\/\/www.congress.gov\/bill\/119th-congress\/senate-bill\/2714\/text\" rel=\"nofollow noopener\" target=\"_blank\"><u>legislation<\/u><\/a><a href=\"https:\/\/www.congress.gov\/bill\/119th-congress\/senate-bill\/3062\/text\" rel=\"nofollow noopener\" target=\"_blank\"><u>is<\/u><\/a><a href=\"https:\/\/www.congress.gov\/bill\/119th-congress\/senate-bill\/4407\/text\" rel=\"nofollow noopener\" target=\"_blank\"><u>pending<\/u><\/a>, but so far has been narrowly focused on protection of minors. This <a href=\"https:\/\/mhaipolicy.org\/tracker\" rel=\"nofollow noopener\" target=\"_blank\"><u>fragmented<\/u><\/a> policy landscape is further hindered by a perpetually <a href=\"https:\/\/pubmed.ncbi.nlm.nih.gov\/40373033\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>lagging evidence base<\/u><\/a>: Many purpose-built AI mental health tools lack validated outcomes and representative samples, and are rarely evaluated with rigorous study designs. Meanwhile, the models powering general-purpose chatbots update so rapidly that safety research findings quickly become outdated.<\/p>\n<p>Recognizing these governance challenges, the Stanford Institute for Human-Centered AI (HAI) convened a select group of leading researchers, clinicians, policymakers, behavioral health experts, ethicists, AI developers, and patient advocates for a policy workshop on mental health and AI in June 2026. The meeting, which followed the inaugural <a href=\"https:\/\/med.stanford.edu\/psychiatry\/news\/spotlight\/ai4mhsymposium.html\" rel=\"nofollow noopener\" target=\"_blank\"><u>AI for Mental Health Symposium<\/u><\/a> held earlier that day, was hosted by <u>HAI\u2019s Healthcare AI Policy Steering Committee<\/u> in collaboration with the university\u2019s <a href=\"https:\/\/med.stanford.edu\/psychiatry\/special-initiatives\/ai4mh.html\" rel=\"nofollow noopener\" target=\"_blank\"><u>AI for Mental Health (AI4MH) Initiative<\/u><\/a>.<\/p>\n<p>Under the Chatham House Rule, participants had candid discussions about emerging efforts to regulate the use of AI in mental health care; evidentiary gaps that must be addressed to enable sound policymaking; and the technical feasibility of potential policy levers and guardrails. Below, we summarize three key policy challenges this group identified for further research<\/p>\n<h3>1. The Field Needs Clearer Definitions<\/h3>\n<p>Effective regulation will require policymakers, mental health practitioners, and AI developers to agree on the boundaries of \u201cmental health AI\u201d \u2013 a broad umbrella term that can refer to many different tools and applications (see table). Right now, there is no broad consensus on what specifically counts as an AI mental health tool, where the line falls between a clinical function that needs to be regulated and a <a href=\"https:\/\/lumeamara.online\/?p=5097\" title=\"ESL Commits $2M to Newark Health and Wellness Campus\">wellness<\/a> feature, or which policy levers to apply to which products. Should a chatbot that was clinically developed specifically for mental health contexts be treated the same as a general-purpose chatbot not originally designed for those purposes but that a user turns to in a mental health crisis?\u00a0<\/p>\n<p>While all mental health AI tools should meet some baseline safety expectations, different types of AI-driven mental health tools may create different expectations and risks, and thus require different regulation. Yet many legislative approaches are not sensitive to the differentiated implications of various mental health AI products. For example, <a href=\"https:\/\/www.ilga.gov\/Legislation\/BillStatus?DocNum=1806&amp;GAID=18&amp;GA=104&amp;DocTypeID=HB&amp;LegID=159064&amp;SessionID=114\" rel=\"nofollow noopener\" target=\"_blank\"><u>outright bans<\/u><\/a> on all AI used by human providers to deliver psychotherapy services don\u2019t address the reality that people may still turn to general-purpose chatbots for therapy that is not yet vetted or supervised by clinicians. In other words, a law meant to keep AI out of therapeutic relationships may instead push users toward less tailored and unregulated general tools. Until there is definitional clarity, regulation will remain fragmented, evaluation standards will be inconsistent, and companies will continue to operate in ambiguity.<\/p>\n<h4>An Illustrative Overview of Mental Health AI Categories<\/h4>\n<table>\n<tbody>\n<tr>\n<td><\/td>\n<td>\n<p><strong>Description<\/strong><\/p>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<p>General-purpose LLMs<\/p>\n<\/td>\n<td>\n<p><em>Examples: ChatGPT, Claude, Gemini<\/em><\/p>\n<p>AI tools that are not specifically designed for mental health purposes, but are commonly used for mental health support\u00a0<\/p>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<p>Companion chatbots<\/p>\n<\/td>\n<td>\n<p><em>Examples: ChatGPT\u2019s CounselorGPT, Character.ai\u2019s Trauma Therapist, and Meta AI Studio\u2019s My Therapist<\/em><\/p>\n<p>General-purpose LLMs that assume a particular persona or role, such as that of a therapist, friend, or trusted partner\u00a0<\/p>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<p>Wellness apps that use AI<\/p>\n<\/td>\n<td>\n<p><em>Examples: <a href=\"https:\/\/www.wysa.com\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>Wysa<\/u><\/a>\u00a0<\/em><\/p>\n<p>Applications that use AI to provide mental health support and promote emotional well-being more broadly but stop short of making medical claims \u2013 and therefore fall outside the regulatory scope of medical devices<\/p>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<p>Purpose-built LLMs for mental health support<\/p>\n<\/td>\n<td>\n<p><em>Examples: <a href=\"https:\/\/www.trytherabot.com\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>TheraBot<\/u><\/a>, Slingshot.AI\u2019s <a href=\"https:\/\/www.talktoash.com\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>Ash<\/u><\/a>, TalkSpace\u2019s <a href=\"https:\/\/investors.talkspace.com\/news-releases\/news-release-details\/talkspace-announces-tee-first-safe-ai-agent-specifically\" rel=\"nofollow noopener\" target=\"_blank\"><u>Tee<\/u><\/a><\/em><\/p>\n<p>LLMs developed specifically for application in mental health care, usually developed through clinical testing but varying in quality<\/p>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<p>AI tools used in non-treatment mental health care settings<\/p>\n<\/td>\n<td>\n<p><em>Examples: <a href=\"https:\/\/www.mentalyc.com\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>Mentalyc<\/u><\/a>, <a href=\"https:\/\/getscribe.commure.com\/lp\/ai-scribe-for-psychiatry\" rel=\"nofollow noopener\" target=\"_blank\"><u>Commure<\/u><\/a>, <a href=\"https:\/\/www.therapynotes.com\/features\/therapyfuel\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>TherapyNotes<\/u><\/a><\/em><\/p>\n<p>AI tools used by human therapists to assist with administrative work (e.g., clinical notetaking), training (e.g., upskilling novice counselors), or case management (e.g., re. These uses can still affect mental health outcomes and risk, even if they do not alone constitute psychotherapy<\/p>\n<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3>2. Evaluation Methods Are Lagging Behind<\/h3>\n<p>Although methods for systematically evaluating the performance of other healthcare AI tools are rapidly developing, methods to assess how effectively and safely mental health AI tools function in the real world are lagging behind. The <a href=\"https:\/\/openai.com\/index\/strengthening-chatgpt-responses-in-sensitive-conversations\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>highest-stakes interactions are rare<\/u><\/a> and hard to simulate, model behavior is unpredictable in real-world situations, and single-session testing reveals little about long-term effects. Meanwhile, the real-world chat data needed to study safety and efficacy at scale sits largely within industry, with no meaningful structures for sharing with independent researchers or regulators.<\/p>\n<p>This creates a policy bottleneck. Many reasonable safety requirements can\u2019t be justified or enforced without large-scale interaction data. Consider sycophancy: the tendency of AI tools to validate and affirm users, which <a href=\"https:\/\/www.centerforhopeandhealth.com\/blog\/when-ai-becomes-part-of-ocd-understanding-a-new-form-of-compulsion\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>may be particularly harmful<\/u><\/a> for people with OCD, where validation-seeking and prolonged engagement can reinforce compulsive patterns. Developing and enforcing suitable policy mechanisms to address the issue requires knowing how often it happens, to whom, and with what effect. We currently can\u2019t answer those questions.<\/p>\n<p>Even where data exists, the evaluation landscape is fragmented. While researchers have proposed many different approaches to benchmarking, there is a lack of consensus on what is most important to measure and how, and whether benchmarking is even the right tool for systems whose behavior changes week to week. Of the many already existing metrics, most reflect technical priorities set by developers (e.g., <a href=\"https:\/\/openai.com\/index\/strengthening-chatgpt-responses-in-sensitive-conversations\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>percentage of messages<\/u><\/a> indicating possible signs of mental health emergencies) rather than the goals and techniques of mental health care (e.g., patient-tailored treatments), which are themselves contested. Most evaluation frameworks currently focus on single, point-in-time analysis and thus are poorly suited to capturing how chatbot use affects users long term. This is a significant gap: Early evidence <a href=\"https:\/\/pmc.ncbi.nlm.nih.gov\/articles\/PMC12967755\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>suggests<\/u><\/a> prolonged use of some types of chatbots may worsen well-being, making longitudinal assessment essential. Policymakers, researchers, and industry must work further together to standardize evaluations of mental health AI and move them toward multimodal assessments that span <a href=\"https:\/\/deepeval.com\/guides\/guides-multi-turn-evaluation\" rel=\"nofollow noopener\" target=\"_blank\"><u>multiple back-and-forth chatbot exchanges<\/u><\/a>.<\/p>\n<h3>3. Target Low-Hanging Fruit While Tackling Deeper Issues<\/h3>\n<p>Workshop participants agreed that more comprehensive mental health AI policy is urgently needed. The good news is that there are some low-hanging fruit that policymakers can and should act on quickly. Transparency requirements, crisis response mechanisms, data protection measures, and parental controls for minors have broad agreement and urgency, which is why they\u2019re already among the <a href=\"https:\/\/governing-ai-in-mental-health.digitalpsychpapers.org\/\" rel=\"nofollow noopener\" target=\"_blank\"><u>most commonly passed<\/u><\/a> provisions in state mental health AI law. When a state enacts thoughtful regulations on such issues, it can set a precedent for other states and lead to some meaningful safety improvements.<\/p>\n<p>However, long-term, meaningful change will require resolving deeper policy challenges and resolving tensions across today\u2019s patchwork of state laws. Disparate state laws concerning therapy tools can be onerous for psychotherapists who are licensed in multiple states with different regulations. Perhaps the most underexamined and unresolved issue in this conversation is one of the most fundamental: Business models built around maximizing user engagement are structurally at odds with the goal of fostering a healthy relationship with chatbots. We have watched this play out before \u2013 <a href=\"https:\/\/www.bbc.com\/news\/articles\/c747x7gz249o\" rel=\"nofollow noopener\" target=\"_blank\"><u>courts<\/u><\/a> have already <a href=\"https:\/\/www.bbc.com\/news\/articles\/cql75dn07n2o\" rel=\"nofollow noopener\" target=\"_blank\"><u>tied<\/u><\/a> social media\u2019s \u201ckeep them on the platform\u201d logic to addiction and harm to minors, and <a href=\"https:\/\/jamanetwork-com.stanford.idm.oclc.org\/journals\/jama\/fullarticle\/2835481\" rel=\"nofollow noopener\" target=\"_blank\"><u>research<\/u><\/a> links heavy use to worse mental health outcomes and suicidal behavior. A chatbot optimized to simulate an intimate human relationship, too, can foster overuse and overreliance. Without mechanisms that reward responsible behavior, there is little reason to expect the industry to self-regulate differently.<\/p>\n<p>Finally, the policy conversation is currently too narrow. It is dominated by higher-income, commercially insured, and professionally licensed perspectives, with less representation of people with severe mental illness, young people, and individuals in the social welfare and criminal justice systems. Leaving them out <a href=\"https:\/\/www.sciencedirect.com\/science\/article\/pii\/S2666560321000177\" rel=\"nofollow noopener\" target=\"_blank\"><u>risks<\/u><\/a> compounding existing inequities at scale, especially at a time when policymakers are urgently seeking fast remedies.<\/p>\n<p><em>Authors: Caroline Yee is a research fellow at the Stanford McCoy Family Center for Ethics in Society; Caroline Meinhardt is the policy research manager at Stanford HAI; Michelle Mello a professor of law at Stanford Law School, professor of health policy at Stanford School of Medicine, and a faculty affiliate at Stanford HAI; Jane Paik Kim a clinical associate professor of psychiatry and behavioral sciences at Stanford School of Medicine.<\/em><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Policymakers, academics, healthcare providers, AI developers, and patient advocates convened by Stanford HAI identify critical gaps in how we regulate AI tools used for therapy and emotional support<\/p>\n","protected":false},"author":1,"featured_media":5109,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[23],"tags":[690,691,40,58,692],"class_list":["post-5107","post","type-post","status-publish","format-standard","has-post-thumbnail","category-mental-health","tag-complexities","tag-governing","tag-health","tag-mental","tag-stanford"],"_links":{"self":[{"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/posts\/5107","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/lumeamara.online\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=5107"}],"version-history":[{"count":1,"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/posts\/5107\/revisions"}],"predecessor-version":[{"id":5108,"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/posts\/5107\/revisions\/5108"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/lumeamara.online\/index.php?rest_route=\/wp\/v2\/media\/5109"}],"wp:attachment":[{"href":"https:\/\/lumeamara.online\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=5107"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/lumeamara.online\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=5107"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/lumeamara.online\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=5107"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}