Search papers, labs, and topics across Lattice.
This study evaluates the ability of general-purpose language models (LLMs) to generate 3D molecular structures under various spatial constraints, comparing their performance to specialized diffusion models. Using a novel benchmarking strategy called 3D-Fit, the authors assess LLMs' capabilities in generating ligands conditioned on protein pockets and other spatial factors. The results indicate that while LLMs currently do not match the performance of diffusion models, they demonstrate significant potential in managing multiple spatial constraints simultaneously, suggesting a pathway for future advancements in molecular design.
LLMs can handle complex 3D spatial constraints in molecular design, though they still trail behind diffusion models in performance.
Structure-based drug design (SBDD) leverages the 3D structure of protein targets, often complemented by other spatial constraints, to generate candidate binding molecules. While diffusion models have dominated as a leading paradigm for high-quality 3D molecule generation, LLM-based methods are rapidly emerging in molecular design and have shown competitive performance in pocket-conditioned molecular generation. However, their ability to reason about physics and 3D spatial environments is largely underexplored. In this work, we systematically analyze whether current general-purpose LLMs are capable of navigating complex 3D constraints compared to established baselines such as specialized diffusion models. We consider 3D ligand generation conditioned on protein pockets together with ligand- and interaction-derived spatial constraints, including anchor fragments, pharmacophore points, and mandatory pocket-ligand interactions. To enable this evaluation, we introduce 3D-Fit - a token-efficient benchmarking strategy for assessing LLM performance on multi-conditioned spatial molecule generation. Our findings reveal a clear pattern in LLM spatial capabilities: while they still lag behind state-of-the-art approaches, they are promising and can handle multiple spatial constraints simultaneously, enabling scaling to heterogeneous setups.