Search papers, labs, and topics across Lattice.
3
3
4
10
Attack success rates in cowork agents can vary dramatically, with some models achieving up to 94.4% effectiveness in executing adversarial tasks.
SecureCollaRAG can thwart knowledge corruption attacks in RAG systems, ensuring the integrity of generated outputs against adversarial manipulations.
By pinpointing the causal origins of tool use, AttriGuard neutralizes indirect prompt injection attacks that can hijack LLM agents, even when faced with adversarial optimization.