כתבה
arXiv cs.AI ·
Mental-R1: Aligning LLM Reasoning for Mental Health Assessment
תקציר מקורי באנגליתarXiv:2606.13176v2 Announce Type: replace Abstract: Mental health problems such as anxiety, depression, and suicide remain urgent global challenges, where timely and accurate assessment is critical for effective intervention. Recently, large language models have been explored for mental health assessment. However, existing general-purpose post-training methods do not align with the cognitive processes of human assessment, which may lead to unreliable reasoning outcomes. To bridge this gap, we propose Cognitive Relative Policy Optimization (CRPO), a reinforcement learning framework tailored for the mental health domain. CRPO extends group relative policy optimization by integrating stage-dependent uncertainty modeling into the policy optimization process. Specifically, we introduce a stage-
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית