Search papers, labs, and topics across Lattice.
This paper introduces TUE-Detector, a novel framework for detecting AI-generated videos by leveraging a multi-layered language model (MLLM) that acts as a tool-using expert. By focusing on tool-mediated evidence discovery, TUE-Detector effectively identifies subtle unnatural artifacts that are often missed by traditional detection methods. Experimental results show that TUE-Detector significantly improves detection accuracy compared to existing approaches, highlighting its potential as a reliable solution in the growing field of AI-generated content identification.
TUE-Detector outperforms traditional methods by using a multi-layered language model to invoke specialized tools for uncovering subtle artifacts in AI-generated videos.
AI-generated video detection, which aims to distinguish AI-generated videos from real ones, has recently received increasing research attention. To perform this task reliably, a key challenge lies in accurately identifying subtle-yet-measurable unnatural artifacts. In this work, we address this challenge from a novel perspective of tool-mediated evidence discovery and propose Tool-Using Expert MLLM-based AI-generated Video Detector (TUE-Detector), a novel framework for AI-generated video detection. TUE-Detector trains a general MLLM into a task-tailored tool-using expert detector that learns to invoke suitable tools, collect concrete evidence of unnaturalness, and reason over the evidence for reliable detection. Meanwhile, TUE-Detector further introduces novel designs to equip the expert detector with high-quality and suitable tools. Extensive experiments demonstrate the effectiveness of our framework.