Search papers, labs, and topics across Lattice.
This study evaluates the potential of large language models (LLMs) for early-stage chronic kidney disease (CKD) screening using zero-shot and few-shot in-context learning, contrasting their performance with traditional machine learning (ML) and deep learning (DL) methods. By leveraging clinically selected tabular features and structured prompt templates, the research demonstrates that LLMs can achieve competitive results in low-data scenarios without the need for task-specific training. However, while LLMs show promise in data efficiency, their performance is model-dependent and less stable compared to traditional approaches as input complexity increases.
LLMs can match or outperform traditional CKD screening methods with minimal examples, but their stability diminishes with increased input complexity.
Early screening of chronic kidney disease (CKD) is critical for timely intervention, yet most machine learning (ML) and deep learning (DL) approaches require labeled data and model training, limiting their use in real-world screening settings. This study evaluates the effectiveness of large language models (LLMs) for CKD screening under zero-shot and few-shot in-context learning settings and compares them with traditional ML and DL methods. We propose a framework that uses clinically selected tabular features and structured prompt templates to enable LLM-based inference without task-specific training. LLM performance is evaluated across multiple prompt styles, feature configurations, and data settings, and compared with standard ML, DL, and tabular foundation model (TFM) baselines, and existing CKD screening tools. The results show that LLMs can achieve competitive performance using only a small number of examples, often matching or outperforming traditional approaches in low-data settings. However, their performance remains model-dependent and less stable as input complexity increases. In contrast, ML, DL, and TFM models show more consistent improvement with larger training data. Overall, the findings highlight a trade-off between data efficiency and stability, suggesting that LLMs may serve as a flexible complementary approach for CKD screening when labeled data are limited.