Search papers, labs, and topics across Lattice.
This review analyzes the evolution of Color Fundus Photography (CFP) in the context of artificial intelligence, highlighting the interdependence of dataset development, preprocessing techniques, and modeling frameworks. It reveals a significant shift from small, task-specific datasets to large, multimodal collections that incorporate longitudinal clinical records, alongside advancements in preprocessing methods such as neural data-engineering and self-supervised imputation. The findings underscore that the future of CFP AI hinges on the synergistic enhancement of these components to achieve effective clinical applications and improved generalization across domains.
The evolution of Color Fundus Photography shows that integrating multimodal datasets with advanced preprocessing can transform clinical reasoning in ophthalmology.
Color Fundus Photography (CFP) is a primary non-invasive imaging modality for large-scale screening of ophthalmic and systemic diseases. Existing surveys mainly summarize task-specific algorithms, datasets, or preprocessing techniques independently, lacking a unified perspective on their co-evolution with modern artificial intelligence. This review provides an integrated overview of CFP AI through the interplay of dataset evolution, preprocessing paradigms, and modeling frameworks. We show that CFP datasets have evolved from small single-center collections with task-specific labels to large multi-center resources featuring multimodal pairings and longitudinal clinical records. Preprocessing has progressed from conventional image enhancement to neural data-engineering pipelines, hardware-aware token optimization, and self-supervised imputation for incomplete electronic health records (EHRs). Meanwhile, modeling has advanced from convolutional neural networks (CNNs) to vision foundation models, state space models (SSMs), and multimodal expert architectures. At the multimodal frontier, CFP is increasingly integrated with EHRs and longitudinal patient information, enabling more comprehensive clinical reasoning beyond isolated image analysis. We conclude that future progress depends on the collaborative optimization of datasets, preprocessing, and multimodal modeling, providing a roadmap toward robust clinical deployment, improved cross-domain generalization, and resource-efficient edge intelligence.