Search papers, labs, and topics across Lattice.
Pangram 4 is a state-of-the-art AI text classification model that achieves an impressive AUROC of 0.9916, with minimal false positive and false negative rates. This model not only surpasses its predecessor, Pangram 3, in overall accuracy but also demonstrates enhanced out-of-distribution generalization and resilience against adversarial attacks. Notably, Pangram 4 excels in distinguishing fine-grained edits and mixed AI-human co-authored text, marking significant advancements in boundary detection and AI assistance identification tasks.
Pangram 4 sets a new standard in AI text classification with unmatched accuracy and robustness against adversarial threats.
We present Pangram 4, the latest deep-learning-based AI-text classification model from Pangram Labs. We achieve an AUROC of 0.9916 with a false positive rate of 0.0041% and a false negative rate of 0.3396%. In addition to its increased overall accuracy compared with Pangram 3, Pangram 4 exhibits superior out-of-distribution generalization and robustness to adversarial attacks. Another novel contribution of Pangram 4 is its improved ability to distinguish fine-grained edits and mixed AI-human co-authored text. We demonstrate improvements to both boundary detection tasks and the detection of interleaved AI assistance. Finally, we report metrics on standard AI detection benchmarks showing that Pangram 4 achieves state-of-the-art performance on the AI text detection task across a wide variety of settings and domains.