Search papers, labs, and topics across Lattice.
This paper develops a comprehensive framework for molecular LLM agents that integrates architectural design with a scientific autonomy ladder, enabling these agents to effectively perceive, reason about, and act upon complex chemical data. By categorizing agents into four levels of autonomy, from assistive workflows to fully autonomous scientific agenda agents, the authors provide a structured approach to evaluate and enhance the capabilities of molecular LLMs. The framework not only identifies gaps in current agent capabilities but also outlines risks associated with their deployment in molecular discovery workflows, paving the way for more effective AI applications in molecular science.
A new framework categorizes molecular LLM agents into four levels of autonomy, revealing critical gaps and risks in their deployment for scientific discovery.
Molecular science represents an important frontier for LLM-based agents. Unlike general agents that mainly operate over natural language, code, or web environments, molecular LLM agents must perceive, reason about, and act upon chemical objects across symbolic strings, molecular graphs, 3D conformations, spectra, simulations, and wet-lab measurements. Their capabilities depend on chemically faithful molecular perception, an LLM-centered agent framework, domain-specific tool grounding, and computational or experimental feedback, in addition to planning and tool use. This work develops a conceptual framework for molecular LLM agents from two complementary perspectives. First, we introduce an architectural view of molecular-agent design, covering molecular representation and perception, the agent framework, domain-specific toolboxes, and learning and optimization. Second, we propose a scientific autonomy ladder inspired by staged autonomy in engineering systems, categorizing agents into four levels: L1 assistive or fixed workflows, L2 adaptive computational agents, L3 feedback-aware physical experiment agents, and L4 scientific-agenda agents. Together, these two perspectives establish a comprehensive framework for comparing existing molecular LLM agents, identifying missing capabilities and deployment risks, and guiding the design, evaluation, and deployment of future agents in molecular discovery workflows.