Search papers, labs, and topics across Lattice.
This paper introduces "future querying," a novel approach that assesses the capability of large language models (LLMs) to act as implicit medical world models by answering time-indexed clinical queries regarding patient futures. By leveraging unstructured clinical documentation and employing endpoint-agnostic training, the framework allows a single model to handle a variety of clinical inquiries without the need for extensive manual feature engineering or retraining. The findings reveal that small, locally fine-tuned models can perform comparably to larger proprietary systems, highlighting the potential for privacy-preserving applications in clinical settings.
Small, fine-tuned LLMs can rival larger proprietary models in predicting patient outcomes, reshaping the landscape of clinical decision support.
Traditional clinical prediction models rely on task-specific pipelines and curated, structured data, which scale poorly and underutilize unstructured text. To address this, we introduce future querying, a paradigm that probes whether large language models (LLMs) can function as implicit medical world models by evaluating their ability to answer time-indexed clinical queries about a patient's future. Our framework operates on unstructured clinical documentation using endpoint-agnostic training, enabling a single model to answer diverse clinical queries over patient trajectories without manual feature engineering or task-specific retraining. We show that small, locally fine-tuned open-weight models can match or approach larger proprietary systems, making the framework suitable for privacy-preserving, on-premise deployment. Evaluated on a new synthetic medical reports dataset and real ICU notes from the MIMIC-IV dataset, our results provide encouraging evidence that LLMs can capture aspects of clinical dynamics.