Search papers, labs, and topics across Lattice.
The Palmyra x6 model is a large language model fine-tuned for enterprise agentic tasks using Anchored Supervised Fine-Tuning on a compact dataset of 626 synthetic tool-use trajectories. This approach, characterized by a conservative training strategy with a low learning rate and KL anchoring, yielded significant performance improvements over its predecessor, particularly in Writer Agent tasks, and achieved the highest scores on the BFCL Core benchmark. Additionally, the model demonstrated competitive performance in bias and safety evaluations compared to recent alternatives, highlighting its robustness in practical applications.
Palmyra x6 outperforms previous models in enterprise agentic tasks while maintaining a strong safety profile and low bias.
Palmyra x6 is a large language model optimized for use with enterprise-oriented agentic tasks. The model was built by post-training a Mixture-of-Experts base model with Anchored Supervised Fine-Tuning on a compact corpus of verified, synthetic tool-use trajectories, optimized with a Muon + Adam hybrid. The recipe is deliberately conservative and deliberately controlled: 626 trajectories, a single epoch, a low learning rate, and a KL anchor to the frozen base. The model shows substantial gains over the previous default model for Writer Agent, and compares favorably with several recent models on public benchmarks, scoring the highest on BFCL Core at $0.785$ and posts the highest six-benchmark mean of the cohort. Furthermore, the model has shown itself to be competitive or leading relative to comparators in our bias and safety evaluations.