Search papers, labs, and topics across Lattice.
This paper addresses the challenge of Multi-Tool Image Editing Attribution (MIEA) in facial forgery, where multiple editing tools are used in tandem, complicating the attribution process. The authors introduce a new dataset, MultiEdit, comprising over 500,000 edited facial images, and develop DPEC, a novel method that leverages locality-aware traces in both spatial and frequency domains to accurately identify the tools used. Experimental results demonstrate that DPEC significantly outperforms nine existing methods in attributing facial images edited with up to five tools.
Multi-Tool Image Editing Attribution reveals that existing methods fail to keep pace with the complexity of modern image edits, but a new approach can accurately identify multiple tools at work.
As generative AI tools become increasingly powerful and easy to use, people can easily edit portrait images with a prompt, necessitating the task of image editing attribution, which predicts the involved editing tools from the given image. Existing attribution methods hold the single-tool assumption and can only attribute a specific editing tool, but struggle to handle the more complex and increasingly common multi-tool editing scenarios, where artifacts left by different editing tools are composite and overlapped. To address this gap, we explore Multi-Tool Image Editing Attribution (MIEA), which aims to identify multiple editing tools involved in a multi-tool edited facial image. To simulate the real-life editing operations on facial images, we then construct a new dataset, MultiEdit, which contains 500k+ edited facial images and covers six types of editing tools that support face swapping (Deepfake) and various facial enhancements. Inspired by the findings from data analysis, we design DPEC, a multi-tool attribution method that can capture distinguishable, locality-aware editing tool traces from both spatial and frequency domains with the support of an error-based curriculum learning strategy. Experiments show \Method\ outperforms nine methods for facial images edited in at most five steps.