Search papers, labs, and topics across Lattice.
This paper introduces a structure-guided molecular design framework that combines contrastive learning of 3D protein-ligand interactions with autoregressive molecular generation conditioned on commercially available compounds. An SE(3)-equivariant transformer encodes protein pockets and ligands into a shared embedding space via contrastive learning, enabling zero-shot virtual screening. These embeddings are then integrated into a multimodal Chemical Language Model (MCLM) to generate target-specific molecules, improving the predicted binding properties of generated candidates.
Generate novel drug candidates with favorable binding properties by steering a chemical language model with contrastive embeddings of protein-ligand complexes.
Structure-based drug discovery faces the dual challenge of accurately capturing 3D protein-ligand interactions while navigating ultra-large chemical spaces to identify synthetically accessible candidates. In this work, we present a unified framework that addresses these challenges by combining contrastive 3D structure encoding with autoregressive molecular generation conditioned on commercial compound spaces. First, we introduce an SE(3)-equivariant transformer that encodes ligand and pocket structures into a shared embedding space via contrastive learning, achieving competitive results in zero-shot virtual screening. Second, we integrate these embeddings into a multimodal Chemical Language Model (MCLM). The model generates target-specific molecules conditioned on either pocket or ligand structures, with a learned dataset token that steers the output toward targeted chemical spaces, yielding candidates with favorable predicted binding properties across diverse targets.