Search papers, labs, and topics across Lattice.
Universitat Pompeu Fabra
2
0
4
CG-CMARL scales to multi-agent scenarios without the exponential blowup in complexity, outperforming fixed reward-shaping methods in achieving optimal coordination.
Reinforcement learning can now handle active feature selection in high-dimensional datasets by intelligently pruning the feature search space and regularizing decision sequences, outperforming existing methods in accuracy and policy complexity.