Search papers, labs, and topics across Lattice.
This study advances the field of planetary geology by embedding a unique methodology for geologic knowledge discovery into a multimodal vision-language architecture, effectively creating a "machine intelligence geologist." The model, trained on lunar basaltic mare volcanism, successfully integrates geological priors with local visual evidence to produce accurate stratigraphic interpretations, although it initially struggles with numeric age dating. By incorporating an open-book retrieval mechanism, the system can accurately reference published chronologies, highlighting the importance of combining visual interpretation with historical context in automated geologic inference.
A novel multimodal architecture allows a machine to generate accurate lunar geological interpretations while effectively citing scientific chronologies.
Planetary geology relies on historical, interpretive reasoning to reconstruct past events from diverse observations. Here, we present a step toward an automated "machine intelligence geologist" by embedding this distinct methodology of geologic knowledge discovery and inference into a multimodal vision-language architecture. Focusing on the stratigraphy of lunar basaltic mare volcanism, we train a model to generate verifiably grounded geologic interpretations directly from co-registered topographic, spectral, and geologic maps. We demonstrate that while the system successfully balances established geological priors with local visual evidence to accurately describe stratigraphy and terrain, numeric age dating derived solely from vision defaults to memorized priors. Integrating an open-book retrieval mechanism resolves this, enabling the model to faithfully cite published chronologies. Our findings delineate the necessary architecture for automated geologic inference: site evidence must be visually interpreted from local data, while quantitative historical context must be retrieved from the scientific record.