Search papers, labs, and topics across Lattice.
This paper introduces XVAE-WMT, an unsupervised explainable generative AI algorithm that integrates a variational autoencoder with wavelet-based inputs and temporal consistency loss for the blind source separation of heart and lung sounds. By leveraging Continuous Wavelet Transform for enhanced time-frequency localization and utilizing SHAP for latent space interpretability, XVAE-WMT eliminates the need for paired clean recordings, addressing limitations of existing methods. The model achieves impressive separation metrics, with 26.8 dB SDR, 32.8 dB SIR, and 28.6 dB SAR across two datasets, highlighting its effectiveness in biomedical signal processing.
XVAE-WMT achieves superior sound separation without requiring paired clean recordings, setting a new standard for interpretability in biomedical signal processing.
The separation of cardiovascular sounds is a critical task in biomedical signal processing. In this paper, we introduce XVAE-WMT1, an unsupervised explainable generative AI algorithm combining a variational autoencoder (VAE) with explainable AI (XAI), wavelet-based inputs, a post-hoc output mask, and temporal consistency (TC) loss. Unlike existing supervised and VAE-based methods that rely on Short-Time Fourier Transform (STFT) and ignore latent interpretability, XVAE-WMT requires no paired clean recordings and integrates a Continuous Wavelet Transform (CWT) front-end for superior time-frequency localization. We assessed the latent space interpretability via different metrics, with SHAP (SHapley Additive exPlanations) enabling dimensionality reduction to the top 75% of latent features while preserving separation quality. Evaluated across two datasets using Signal-to-Distortion Ratio (SDR), Signal-to-Interference Ratio (SIR), and Signal-to-Artifacts Ratio (SAR), XVAE-WMT attains 26.8 dB SDR, 32.8 dB SIR, and 28.6 dB SAR.