Search papers, labs, and topics across Lattice.
2
1
4
4
By constraining reward functions to Control Barrier Functions, this approach achieves safe adversarial imitation learning that adapts directly from expert observations without requiring labeled data.
Robots can now learn manipulation tasks from scratch in under an hour using only visual observations and interaction, outperforming traditional IRL and RL methods in sample efficiency.