Search papers, labs, and topics across Lattice.
This paper introduces DynaBridge, a novel framework that integrates acoustic, visual, and textual data to enhance mental health assessments based on the DASS-21 questionnaire. By leveraging frozen-LLM-generated summaries and employing a confidence-aware refinement strategy, DynaBridge effectively predicts ordinal item distributions and reconstructs risk evidence for depression, anxiety, and stress. The framework outperforms existing multimodal methods, achieving a mean F1 score of 0.5012 for risk prediction, highlighting the importance of structured psychometric integration in mental health evaluations.
Bridging multimodal cues with psychometric structure, DynaBridge achieves unprecedented accuracy in predicting mental health risks.
Multimodal behavioral analysis offers a scalable approach to assessing depression, anxiety, and stress, yet generic fusion models often ignore the psychometric structure of questionnaire labels. In DASS-21, risk labels are derived from ordered symptom items through fixed item-to-subscale mappings. We propose \textbf{DynaBridge}, a dynamic summary-guided cross-task multimodal framework for DASS-structured mental health assessment. DynaBridge encodes acoustic, visual, and textual cues across multiple sessions and augments them with frozen-LLM-generated DASS-aware summaries as participant-level semantic evidence. It predicts ordinal item distributions, reconstructs depression, anxiety, and stress risk evidence from item-level soft scores, and fuses this evidence with direct multimodal risk predictions. A confidence-aware refinement strategy further incorporates high-confidence semantic cues conservatively. On the official AdoDAS validation split, DynaBridge outperforms the official baseline and representative multimodal methods, achieving 0.5012 mean F1 for D/A/S risk prediction and 0.3216 mean QWK for DASS-21 item prediction. These results show the value of bridging multimodal cues, semantic summaries, and DASS-21 psychometric structure.