Search papers, labs, and topics across Lattice.
This paper introduces BioKERN, a multimodal spatial representation-learning framework that enhances histology-to-transcriptomics neighborhood retrieval by integrating biological structure as a learnable inductive bias. By constructing a biological kernel that combines transcriptomic similarity with spatial proximity, BioKERN provides graded neighborhood supervision, leading to significant improvements in retrieval performance across datasets like Mouse Brain Visium and Human Liver GSE240429. The results indicate that the gains in performance are primarily due to biological-kernel regularization rather than simply increasing model capacity, highlighting the importance of incorporating biological geometry in multimodal learning.
BioKERN achieves superior biological-neighborhood retrieval by leveraging an explicit biological kernel, outperforming existing methods without relying on larger models.
Spatially resolved biology requires representations that preserve biological neighborhood structure rather than only exact cross-modal correspondences. Existing histology--transcriptomics objectives can emphasize instance-level matching even when non-paired spots share molecular or spatial context. We introduce BioKERN, a multimodal spatial representation-learning framework that incorporates biological structure as an explicit, learnable inductive bias. BioKERN constructs a training-time biological kernel by combining transcriptomic similarity and spatial proximity, then uses it to provide graded neighborhood supervision and regularize embedding geometry. Evaluation uses a fixed, model-independent biological neighborhood definition shared by all methods. Across Mouse Brain Visium and Human Liver GSE240429, BioKERN consistently improves biological-neighborhood retrieval over BLEEP in both single- and multi-scale settings. Controlled shared-architecture experiments show that most of the improvement arises from biological-kernel regularization rather than increased model capacity. These results support explicit biological geometry as an interpretable inductive bias for multimodal learning in spatial biology.