Search papers, labs, and topics across Lattice.
Affiliation:
2
0
4
10
EXIMO achieves unprecedented sample efficiency in robotic policy finetuning by leveraging a vision language model to decompose complex tasks.
Non-stationary environments demand RL agents forget the past, or else they'll suffer regret.