Search papers, labs, and topics across Lattice.
University, Seoul National University, Columbia University
1
0
2
Instance-optimality in MNL-based RL is now achievable, thanks to a new algorithm that adapts to the variance of learner-environment interactions.