Search papers, labs, and topics across Lattice.
3
0
6
4
A simple screenshot-based judge often outperforms complex LLM evaluation methods, challenging assumptions about the necessity of intricate judging pipelines.
Merging LoRA modules with a hypernetwork enables superior few-shot adaptation, outperforming existing methods that overlook source domain knowledge.
LLMs reason better when their uncertainty consistently decreases, paving the way for shorter, more accurate chain-of-thought reasoning.