Search papers, labs, and topics across Lattice.
Tencent YoutuLab
2
0
3
Categorical value learning can significantly enhance the performance of PPO critics in reinforcement learning, leading to better calibration and lower variance in advantage estimation.
Geometry-aware dataset condensation can dramatically enhance the fidelity of diffusion models, preserving essential distributional characteristics that traditional methods overlook.