Search papers, labs, and topics across Lattice.
Affiliation:
2
0
4
Static safety benchmarks fail when attackers and defenders are evaluated in isolation; pitting them head-to-head with dual-routed payloads finally separates genuine defense filtering from raw exploit potency.
Across five major agent-memory frameworks, "soft-deleted" facts not only leak into context by default鈥攖hey actively outrank valid replacements and cause models to execute unsafe actions.