Search papers, labs, and topics across Lattice.
Mimir v1 is a 1-billion-parameter language model that utilizes the Hierarchical Reasoning Model (HRM) architecture and is trained exclusively on permissible post-training data, addressing ethical concerns in dataset sourcing. It achieves state-of-the-art performance for Danish and competes effectively with larger models like Qwen 3.5 4B and Gemma 4 E2B across 20 benchmarks in English, Math, and Code. This work not only sets a new standard for ethical AI development but also demonstrates that high performance can be achieved without relying on massive, non-permissible datasets.
Achieving state-of-the-art results for Danish with a model that uses only permissible data challenges the notion that larger datasets are necessary for competitive performance.
Current large language model development relies on massive, often non-permissible datasets, creating a high barrier for researchers committed to open-source and ethically sourced data. We introduce Mimir v1, a 1-billion-parameter language model based on the Hierarchical Reasoning Model (HRM) architecture, that is trained from scratch and delivers highly competitive performance for English and sets a new state of the art for Danish using only permissible post-training data. Trained on a mixture of 161 datasets, Mimir v1 outperforms the original HRM-Text 1B and competes with larger frontier models like Qwen 3.5 4B and Gemma 4 E2B, tested across 20 benchmarks for English, Math & Code and Danish. The model is available on the Hugging Face Hub: https://huggingface.co/danish-foundation-models/DFM-Mimir