Search papers, labs, and topics across Lattice.
Affiliation:
2
0
5
0
Trait-induced safety variation can lead to inconsistent safety decisions in LLMs, but a new tuning method stabilizes their behavior across different traits.
Z-1 boosts VLA model performance by over 13% using only publicly available demonstrations, showcasing the power of reinforcement learning in robotic manipulation.