Search papers, labs, and topics across Lattice.
This paper introduces Agent-Guided Concept Discovery, a novel framework that enables deep learning models to learn meaningful concepts from Rapid Evaporative Ionization Mass Spectrometry (REIMS) data without requiring predefined labels. By employing a reasoning agent to refine semantic descriptions and leveraging a biochemical knowledge graph, the model enhances interpretability and diagnostic relevance in surgical margin assessment. Results demonstrate improved balanced accuracy and sensitivity across cancer datasets, with fewer false positives in intraoperative scenarios, indicating better generalization to real-world conditions.
Learning surgical concepts directly from noisy intraoperative data could revolutionize how we assess surgical margins in real-time.
Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clinical adoption remains challenging due to limited generalization to operating room conditions. This difficulty arises because models are typically trained on labeled spectra collected from resected tissue samples, while they must operate on noisy, unlabeled data acquired directly during surgery. In addition, the black-box nature of deep learning models makes it difficult to understand and systematically improve their behavior. Concept-based learning offers a promising way to address these challenges by mapping raw measurements to human-understandable concepts. However, supervised concept-based approaches rely on concept annotations, which are difficult to obtain in complex mass spectrometry workflows. We propose Agent-Guided Concept Discovery, a framework that learns meaningful concepts directly from data without requiring predefined concept labels. During training, a reasoning agent refines semantic descriptions of the learned concepts and adaptively adjusts their weight based on diagnostic relevance. These concepts are further grounded using a biochemical knowledge graph to ensure consistency with known metabolic relationships. Across Skin and Breast Cancer datasets, our model improves balanced accuracy and sensitivity over the baseline. In a representative intraoperative case, it shows fewer false positives, indicating better generalization to surgical conditions.