Search papers, labs, and topics across Lattice.
3
0
5
5
Improving descriptive reasoning trace quality can actually hinder recommendation effectiveness, challenging assumptions about the benefits of interpretability in AI systems.
Grounding LLM evaluations in historical user behavior can boost relevance judgment accuracy by over 15%, making them more aligned with actual user preferences.
Spotify's GLIDE model proves that generative LLMs can drive significant gains in podcast discovery and non-habitual listening in a real-world, production environment.