Search papers, labs, and topics across Lattice.
4
0
4
4
Extreme fidelity loss reveals critical vulnerabilities in long-horizon tasks that standard accuracy metrics overlook.
Indirect prompt injection can compromise AI systems like DeepSeek Harness, with attack success rates reaching up to 25.5% under certain conditions.
SkillJack reveals that self-evolving agents can unknowingly incorporate malicious skills, making traditional safety measures ineffective against persistent threats.
AI-Infra-Guard reveals that a unified security framework can effectively address the diverse attack surfaces of AI agents, making it a game-changer for AI safety.