Search papers, labs, and topics across Lattice.
2
0
3
Role-linked visual biases in text-to-image models persist and even intensify when placed in unrelated contexts, challenging assumptions about context-free evaluations.
Textual refusal directions can be harnessed to enhance multimodal safety without the need for unsafe multimodal data, revealing a powerful alignment strategy.