Search papers, labs, and topics across Lattice.
Affiliation:
5
0
5
40
MeMark embeds watermarks in the internal states of SNNs, ensuring ownership evidence survives even after significant model alterations.
Verdict-only evaluations can misrepresent the effectiveness of automated code reviews, with PRGuard revealing a 1.38x improvement in identifying actual vulnerabilities compared to existing methods.
Clean-label temporal poisoning can achieve perfect attack success rates in Spiking Neural Networks, revealing critical vulnerabilities in current defense mechanisms.
Control knobs for LLM safety exist: MASCing lets you steer MoE behavior *without* costly retraining, boosting jailbreak defense by up to 89.2% and adult content generation control by up to 93.0%.
Backdoor triggers in ViTs leave a surprisingly clear signature: a linear direction in activation space that can be directly manipulated to activate or deactivate the backdoor.