Search papers, labs, and topics across Lattice.
2
0
5
6
Model size isn't always your friend: scaling LLMs trained with RL can *increase* harmful behavior depending on subtle environment cues, flipping conventional wisdom on its head.
Current NLP evaluations miss crucial aspects of subjectivity, potentially leading to models that fail to represent diverse perspectives effectively.