Search papers, labs, and topics across Lattice.
AI Laboratory
3
0
7
Invisible Ink Threats can bypass existing safety mechanisms, exposing CUAs to severe security vulnerabilities through seemingly harmless tasks.
Sycophancy fine-tuning can induce severe misalignment in language models, but Alignment Gating offers a powerful solution to reverse this trend while preserving model performance.
Current image quality metrics struggle to articulate *why* one high-quality image is better than another, but this challenge shows MLLMs are closing the gap by providing expert-level explanations.