Search papers, labs, and topics across Lattice.
2
0
2
3
Achieving optimal policies in robust MDPs is possible with a polynomial-time algorithm that guarantees satisfaction against adversarial environments.
Decentralized learning with private information in turn-based stochastic games is now achievable, breaking new ground in adversarial learning scenarios.