LLM Reasoning 相关度: 9/10

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

Abdullah Mazhar, Het Riteshkumar Shah, Aseem Srivastava, Smriti Joshi, Md Shad Akhtar
arXiv: 2604.05795v1 发布: 2026-04-07 更新: 2026-04-07

AI 摘要

该论文提出FAITH-M基准和CARE框架,评估AI心理健康系统在治疗原则上的符合度。

主要贡献

  • 提出了FAITH-M基准,用于评估治疗原则
  • 提出了CARE框架,用于评估AI生成的回应
  • CARE在多个评估指标上优于基线模型

方法论

CARE框架结合了对话上下文、对比样例检索和知识蒸馏的CoT推理,以评估AI的回应。

原文摘要

The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psychotherapeutic best practices beyond surface-level fluency. While recent systems exhibit conversational competence, they lack structured mechanisms to evaluate adherence to core therapeutic principles. In this paper, we study the problem of evaluating AI-generated therapist-like responses for clinically grounded appropriateness and effectiveness. We assess each therapists utterance along six therapeutic principles: non-judgmental acceptance, warmth, respect for autonomy, active listening, reflective understanding, and situational appropriateness using a fine-grained ordinal scale. We introduce FAITH-M, a benchmark annotated with expert-assigned ordinal ratings, and propose CARE, a multi-stage evaluation framework that integrates intra-dialogue context, contrastive exemplar retrieval, and knowledge-distilled chain-of-thought reasoning. Experiments show that CARE achieves an F-1 score of 63.34 versus the strong baseline Qwen3 F-1 score of 38.56 which is a 64.26 improvement, which also serves as its backbone, indicating that gains arise from structured reasoning and contextual modeling rather than backbone capacity alone. Expert assessment and external dataset evaluations further demonstrate robustness under domain shift, while highlighting challenges in modelling implicit clinical nuance. Overall, CARE provides a clinically grounded framework for evaluating therapeutic fidelity in AI mental health systems.

标签

心理健康 大语言模型 评估 治疗原则

arXiv 分类

cs.CL