Search papers, labs, and topics across Lattice.
Affiliation:
2
0
5
10
Traditional probing methods fail to reveal the true memorization capabilities of large code LLMs, leading to inflated performance scores that obscure their genuine understanding.
Simple image transformations can undermine even the most advanced AI-based content moderation systems, exposing significant vulnerabilities.