Search papers, labs, and topics across Lattice.
Google Deepmind, New York, New York, USA
1
0
2
15
High sample efficiency in learning-to-rank can be achieved without the headaches of custom gradients, making RL more accessible for practitioners.