AutoSplice: A Text-prompt Manipulated Image Dataset for Media Forensics

04/14/2023
by   Shan Jia, et al.
0

Recent advancements in language-image models have led to the development of highly realistic images that can be generated from textual descriptions. However, the increased visual quality of these generated images poses a potential threat to the field of media forensics. This paper aims to investigate the level of challenge that language-image generation models pose to media forensics. To achieve this, we propose a new approach that leverages the DALL-E2 language-image model to automatically generate and splice masked regions guided by a text prompt. To ensure the creation of realistic manipulations, we have designed an annotation platform with human checking to verify reasonable text prompts. This approach has resulted in the creation of a new image dataset called AutoSplice, containing 5,894 manipulated and authentic images. Specifically, we have generated a total of 3,621 images by locally or globally manipulating real-world image-caption pairs, which we believe will provide a valuable resource for developing generalized detection methods in this area. The dataset is evaluated under two media forensic tasks: forgery detection and localization. Our extensive experiments show that most media forensic models struggle to detect the AutoSplice dataset as an unseen manipulation. However, when fine-tuned models are used, they exhibit improved performance in both tasks.

READ FULL TEXT

page 4

page 5

page 7

research
03/10/2023

New Benchmarks for Accountable Text-based Visual Re-creation

Given a command, humans can directly execute the action after thinking o...
research
01/28/2023

Towards Equitable Representation in Text-to-Image Synthesis Models with the Cross-Cultural Understanding Benchmark (CCUB) Dataset

It has been shown that accurate representation in media improves the wel...
research
09/21/2023

TextCLIP: Text-Guided Face Image Generation And Manipulation Without Adversarial Training

Text-guided image generation aimed to generate desired images conditione...
research
04/13/2018

The PS-Battles Dataset - an Image Collection for Image Manipulation Detection

The boost of available digital media has led to a significant increase i...
research
04/04/2023

Leveraging Deep Learning Approaches for Deepfake Detection: A Review

Conspicuous progression in the field of machine learning and deep learni...
research
05/23/2023

Generalizable Synthetic Image Detection via Language-guided Contrastive Learning

The heightened realism of AI-generated images can be attributed to the r...
research
04/13/2021

NewsCLIPpings: Automatic Generation of Out-of-Context Multimodal Media

The threat of online misinformation is hard to overestimate, with advers...

Please sign up or login with your details

Forgot password? Click here to reset