Search papers, labs, and topics across Lattice.
ByteDance Seed
1
0
3
7
SpectraReward reveals that pretrained MLLMs can serve as powerful zero-shot reward models, outperforming traditional methods without the need for fine-tuning or external labels.