Search papers, labs, and topics across Lattice.
This study explores the integration of egocentric vision with wearable inertial measurement units (IMUs) to improve freezing of gait (FOG) detection in Parkinson's disease, emphasizing the importance of contextual information in daily living activities. By analyzing synchronized video and IMU data from 13 participants, the researchers found that an IMU-based temporal convolutional network (TCN) outperformed ego-video features in event detection, achieving an F1 score of 42.3 and AUROC of 83.0. However, the ego-video approach demonstrated potential for capturing relevant contextual information, suggesting a complementary role in enhancing clinical motion understanding.
IMU-based sensing outperforms egocentric vision in detecting freezing of gait, but the latter reveals crucial contextual insights that could transform clinical assessments.
Understanding motion in daily living requires context beyond kinematics, because similar inertial patterns during activities of daily living (ADLs) can reflect intentional stopping, object interaction, or pathological movement impairment. Egocentric vision provides task-related context that may help disambiguate these cases. We investigate this challenge through freezing of gait (FOG) detection in Parkinson's disease (PD), a symptom strongly influenced by contextual factors during ADLs. Using synchronized egocentric video, wearable IMUs, and expert-annotated FOG labels collected from 13 PD participants in their homes, we evaluate frozen representations from pretrained ego-video and time-series foundation models, alongside an IMU-based TCN trained from scratch, under leave-one-subject-out evaluation. The IMU-based TCN achieved the strongest event-detection performance, reaching 42.3 F1 and 83.0 AUROC, compared with 32.6 F1 and 77.2 AUROC for V-JEPA2 ego-video features. Although ego-video alone did not outperform IMU-based sensing, it showed above-chance discrimination, and qualitative analyses suggest that egocentric vision may capture FOG-relevant information independent of IMUs. Together, these results support the use of pretrained ego-video representations to add contextual information to wearable-sensor-based clinical motion understanding in daily living.