Search papers, labs, and topics across Lattice.
Northwestern Polytechnical University
2
0
3
Success in agent evaluations often masks the true reasoning process, with a new method revealing that access to the correct target can inflate success scores by over 25 percentage points.
Even after rigorous quality controls, AI-generated text still reveals detectable patterns that traditional sentence-level detectors can exploit.