Search papers, labs, and topics across Lattice.
2
0
3
0
Self-supervised finetuning can preserve a VLA model's instruction-following abilities while significantly boosting its performance on new robotic tasks.
ReToken boosts vision-language model performance by over 20% on key benchmarks while fitting training and inference within a single GPU's memory constraints.