ViT
1 article · search the full text for this term
-
An Analysis of Multimodal Fusion in Deepfake Detection for Video Samples
Abstract: In today’s rapidly evolving digital landscape, deepfake technology stands as both a marvel and a threat to privacy and security. Deepfakes, hyper-realistic synthetic media created using artificial intelligence (AI), can deceive and manipulate on an unprecedented scale, from political propaganda to compromising videos of public figures. This research navigates deepfake detection, focusing on two advanced methodologies: the vision transformers (ViT) image classifier and the Meso4 method. The ViT model utilizes …
Published in Journal of Image Processing & Pattern Recognition Progress · Vol. 11, Issue 3, 2024 · pp. 19–27 Read article