Search papers, labs, and topics across Lattice.
University of Pennsylvania
2
0
2
LLMs show significant variability in the actionability of their UX critiques, with some models outperforming others across different product categories.
Ground-truth labels don't improve LLM reasoning, but building a consensus-based knowledge graph of reasoning steps does, boosting accuracy by 10%+.