Multi-View Attention Multiple-Instance Learning Enhanced by LLM Reasoning for Cognitive Distortion Detection
Kim, Jun Seo, Kim, Hyemi, Oh, Woo Joo, Cho, Hongjin, Lee, Hochul, Kim, Hye Hyeon
–arXiv.org Artificial Intelligence
Cognitive distortions have been closely linked to mental health disorders, yet their automatic detection remained challenging due to contextual ambiguity, co-occurrence, and semantic overlap. We proposed a novel framework that combines Large Language Models (LLMs) with Multiple-Instance Learning (MIL) architecture to enhance interpretability and expression-level reasoning. Each utterance was decomposed into Emotion, Logic, and Behavior (ELB) components, which were processed by LLMs to infer multiple distortion instances, each with a predicted type, expression, and model-assigned salience score. These instances were integrated via a Multi-View Gated Attention mechanism for final classification. Experiments on Korean (KoACD) and English (Therapist QA) datasets demonstrate that incorporating ELB and LLM-inferred salience scores improves classification performance, especially for distortions with high interpretive ambiguity. Our results suggested a psychologically grounded and generalizable approach for fine-grained reasoning in mental health NLP.
arXiv.org Artificial Intelligence
Sep-23-2025
- Country:
- Asia
- Middle East
- Singapore (0.04)
- North America > United States (0.04)
- Asia
- Genre:
- Research Report > New Finding (1.00)
- Industry:
- Technology: