arXiv · 2609.25287
Can LLMs identify and repair ruptures? Comparison between clinician practices and LLM behaviors
Abstract
Ruptures represent common albeit critical moments in interaction where relational alignment breaks down, making them essential for evaluating AI where trust and engagement matter most. In a scenario-driven empirical study, we examined the performance of three LLMs at identifying and resolving ruptures across 21 mental health conversations and 22 experts' evaluation of the strategies. For identification, LLMs relied on explicit linguistic cues within single turns whereas experts integrated implicit, relational, and contextual information across the conversation. For resolution, LLMs tended to produce more directive and scripted responses whereas experts adopted process-oriented strategies such as validation, open-ended exploration, and psychoeducation. Overall, LLMs showed higher agreement with predefined labels in identification, but not in resolution where experts rated their responses only moderately effective, with consistent limitations in timing, depth, and contextual sensitivity. We discuss implications for the design of mental health conversational agents emphasizing relational awareness, pacing, and human-in-the-loop support.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Jeongah Lee, Joy Qiuyue Zhong, Drishti Goel, Violeta J. Rodriguez, Dong Whi Yoo, Koustuv Saha, Ravi Karkar. 2026-09-21. Can LLMs identify and repair ruptures? Comparison between clinician practices and LLM behaviors. https://arxiv.org/abs/2609.25287
Cite the original work for its findings. Save a collection to share your selection of sources.