Divide and Conquer: Reliable Multi-View Evidential Learning for Deepfake Detection

📅 2026-06-01
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the limitations of existing single-view deepfake detection methods, which often exhibit poor generalization and overconfident predictions due to semantic features obscuring subtle forgery traces. To mitigate this issue, the authors propose DiCoME, a novel framework that decouples semantic content from forgery-related artifacts through geometric projection, thereby alleviating semantic masking via geometry-aware view purification. Furthermore, DiCoME integrates multi-view information and leverages uncertainty-aware evidential learning combined with cognitive conflict modeling to produce well-calibrated uncertainty estimates. Extensive experiments demonstrate that the proposed method significantly outperforms state-of-the-art approaches across multiple benchmarks, achieving both enhanced generalization and reliable quantification of detection confidence.
📝 Abstract
With the evolution of generative models, deepfakes have achieved near-perfect semantic realism, leaving forensic traces only in subtle structural anomalies. However, existing single-view paradigms often fail to generalize, as dominant semantic features overwhelm subtle artifact cues within entangled representations. This imbalance leads to overconfident yet brittle predictions -- a phenomenon we term the Semantic Masking Effect. To address this challenge, we propose a reliable framework called Divide-and-Conquer Multi-View Evidential Learning (DiCoME) for Deepfake Detection. In the "Divide" phase, we employ Geometric View Purification to decompose the entangled representation space through principled geometric projection. This process suppresses semantic interference within artifact-sensitive representations, forming the foundation for decorrelated yet complementary semantic and artifact views. In the "Conquer" phase, we leverage Uncertainty-Aware Evidential Learning to synthesize these distinct views. By explicitly modeling the "epistemic conflict" between semantic and artifact cues, this mechanism provides calibrated uncertainty estimates instead of forcing rigid deterministic decisions. Extensive experiments across multiple benchmarks demonstrate that our method consistently outperforms existing approaches in generalization performance, while providing reliable uncertainty estimation for trustworthy deepfake detection. Code is available at https://github.com/kxl0825/DiCoME.git.
Problem

Research questions and friction points this paper is trying to address.

Deepfake Detection
Semantic Masking Effect
Multi-View Learning
Generalization
Uncertainty Estimation
Innovation

Methods, ideas, or system contributions that make the work stand out.

Geometric View Purification
Uncertainty-Aware Evidential Learning
Semantic Masking Effect
Multi-View Representation
Epistemic Conflict
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
X
Xiaolu Kang
School of Computer Science, Wuhan University, Wuhan, China
Zhongyuan Wang
Zhongyuan Wang
Wuhan University
J
Jikang Cheng
School of Integrated Circuits, Peking University, Beijing, China
B
Baojin Huang
School of Information, Huazhong Agricultural University, Wuhan, China
Z
Zhanhe Lei
School of Computer Science, Wuhan University, Wuhan, China
Gang Wu
Gang Wu
Professor, University of Electronic Science and Technology of China
Energy-efficient wireless communicationsISACsemantic communications
Qin Zou
Qin Zou
Professor of Computer Science, Wuhan University
Computer VisionPattern RecognitionMachine Learning
Qian Wang
Qian Wang
IEEE Fellow, School of Cyber Science and Engineering, Wuhan University
AI system securitywireless system securityapplied cryptography