Papers › Towards Quantitative Evaluation of Explainable AI Methods for Deepfake Detection

Towards Quantitative Evaluation of Explainable AI Methods for Deepfake Detection

29 Apr 2024arXiv:2404.18649archive 2025-07-28

Konstantinos Tsigos, Evlampios Apostolidis, Spyridon Baxevanakis, Symeon Papadopoulos, Vasileios Mezaris

In this paper we propose a new framework for evaluating the performance of explanation methods on the decisions of a deepfake detector. This framework assesses the ability of an explanation method to spot the regions of a fake image with the biggest influence on the decision of the deepfake detector, by examining the extent to which these regions can be modified through a set of adversarial attacks, in order to flip the detector's prediction or reduce its initial prediction; we anticipate a larger drop in deepfake detection accuracy and prediction, for methods that spot these regions more accurately. Based on this framework, we conduct a comparative study using a state-of-the-art model for deepfake detection that has been trained on the FaceForensics++ dataset, and five explanation methods from the literature. The findings of our quantitative and qualitative evaluations document the advanced performance of the LIME explanation method against the other compared ones, and indicate this method as the most appropriate for explaining the decisions of the utilized deepfake detector.

PaperPDFCode

Code

idt-iti/xai-deepfakes officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

DeepFake DetectionFace SwappingPrediction

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

FLIPLIMESET

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections