Search papers, labs, and topics across Lattice.
2
1
4
30
Arabic hateful memes pose unique challenges that current detection models are ill-equipped to handle, revealing critical gaps in our understanding of online hate.
LLM safety systems that appear robust in English crumble when faced with Chinese-specific adversarial attacks, exposing a critical gap in current alignment strategies.