Search papers, labs, and topics across Lattice.
2
0
3
Instruction-tuned models abandon their own correct answers over 41% of the time when an incorrect option is merely prefaced with a bare, unverified "expert" claim.
Position bias and audit design rival or exceed demographic disparities in LLMs, rendering high-profile findings of rating-vs-ranking bias reversals non-replicable across hiring, lending, and triage.