Search papers, labs, and topics across Lattice.
3
0
3
6
Achieving 64x faster generative reranking without sacrificing quality, hLLM redefines efficiency in language model output.
PRL not only boosts recommendation accuracy but also reveals meaningful user clusters, transforming how we approach model interpretability in recommender systems.
LLMs can now rank millions of candidates with significant accuracy gains thanks to a novel K-means clustering and graph-based ensemble approach that overcomes context length limitations.