Search papers, labs, and topics across Lattice.
University of Amsterdam, Amsterdam, Netherlands
2
0
2
21
High sample efficiency in learning-to-rank can be achieved without the headaches of custom gradients, making RL more accessible for practitioners.
Capturing epistemic uncertainty in click predictions could revolutionize how we model user interactions with rankings, leading to more robust recommendation systems.