Search papers, labs, and topics across Lattice.
2
0
3
5
Reporting only accuracy metrics can mask up to 21% of behavioral inconsistencies in LLM responses, challenging the reliability of current AI safety evaluations.
Lower-income users are targeted with ads more frequently than their higher-income counterparts, revealing a potential bias in LLM advertising strategies.