发表机构
Ahsanullah University of Science and Technology; Bangladesh University of Business and Technology; The University of New South Wales; Jashore University of Science and Technology; American International University - Bangladesh(阿山努拉科技大学; 孟加拉国商业与技术大学; 新南威尔士大学; 贾绍尔科技大学; 孟加拉国美国国际大学)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
研究探索大语言模型在低资源语言中生成同理心心理健康建议的能力,提出角色扮演反思性思维链咨询框架和基于Grok 4的评估框架,通过人机协作及特定策略,提升心理健康建议质量,使其更符合专业咨询。
AI 中文摘要
尽管大语言模型(LLMs)最近取得了进展,但它们在低资源语言中生成同理心心理健康咨询回复的能力在很大程度上仍未得到探索。为填补这一空白,我们从三个补充来源精心挑选了625个真实心理健康案例。基于这些案例构建了评估语料库。我们还提出了角色扮演反思性思维链咨询框架(RP - RCAF)和基于Grok 4的回复评估与评分框架(G - REFS)。实验结果表明,RP - RCAF在所有评估模型中始终优于传统提示,并产生与专业心理咨询更一致的回复。
英文摘要
Despite recent advances in large language models (LLMs), their ability to generate empathetic mental health counseling responses in low-resource languages remains largely unexplored. To address this gap, we curate 625 authentic mental health cases from three complementary sources: (1) publicly available Facebook posts discussing mental health concerns, (2) transcripts from the Bangladeshi television program "Ami Akhon Ki Korbo", and (3) anonymized student questionnaire responses covering diverse emotional and psychological challenges. Based on these cases, we build an evaluation corpus comprising advice written by licensed clinical psychologists and responses generated by three modern proprietary LLMs: GPT-4o Mini, Claude 4.5 Haiku, and Gemini 2.5 Pro. We further propose the Role-Playing Reflective Chain-of-Thought Advisory Framework (RP-RCAF), a task-specific prompting strategy that combines expert-authored few-shot examples with structured self-reflection to produce supportive, culturally aware, and ethically aligned counseling through a compassionate advisor persona. We also introduce the Grok 4-Based Response Evaluation and Scoring Framework (G-REFS), which integrates automated assessment with expert psychologist validation across emotional sensitivity, cultural appropriateness, linguistic clarity, and ethical soundness. Experimental results show that RP-RCAF consistently outperforms conventional prompting across all evaluated models and produces responses that more closely align with professional psychological counseling.