Search papers, labs, and topics across Lattice.
2
0
3
0
A small LLM can be trained to detect hallucinations as effectively as larger models through an innovative self-play framework that evolves its own training data.
A stealthy skill injection method that achieves an 89.3% success rate while evading detection in LLM agents reveals critical vulnerabilities in current safety mechanisms.