Search papers, labs, and topics across Lattice.
2
0
3
2
P^3 transforms VAE-based policy learning by slashing data inefficiency and convergence time, achieving over 96% efficiency in challenging tasks.
Robot RL training can be dramatically sped up (3-10x) by decoupling CPU-based simulation from GPU-based learning, challenging the assumption that GPU-resident physics is essential for efficiency.