Journals
  Publication Years
  Keywords
Search within results Open Search
Please wait a minute...
For Selected: Toggle Thumbnails
Deepfake detection method based on fusion of multi-modal physical prior features
Renkun LYU, Peng SUN, Yubo LANG, Hong GUO, Zhe SHEN, Di TIAN
Journal of Computer Applications    2026, 46 (8): 2515-2523.   DOI: 10.11772/j.issn.1001-9081.2025070826
Abstract155)   HTML1)    PDF (2963KB)(18)       Save

The existing deepfake detection methods mainly model on the basis of pixel-level clues of images, and seldom consider the impact of the synthesis process on the forged images. Although good detection results are achieved, it is difficult to explain the detection process. Therefore, a multi-modal physical prior feature fusion-based explainable detection method for deepfakes was proposed. First, optical flow features, illumination features, edge features and DCT(Discrete Cosine Transform) features were used to describe the inter-frame motion differences in temporal videos, the illumination inconsistency in single-frame videos, and edge artifact information, respectively, so as to obtain multi-modal physical prior features with explainability. Second, a multi-modal mixture of experts network was proposed to construct expert sub-networks for different modalities, and after cross-modal attention weighting, the sub-networks were fused through a gated unit and input into the discriminative network for classification. Third, the SIAM (Spatial Intersection Attention Module) was introduced into the discriminative network, and the fully connected structure was replaced by the KAN (Kolmogorov-Arnold Network) structure. Finally, the multi-modal physical prior features were used to train different expert sub-networks, respectively, and the Shapley value analysis of different input features was given, thereby constructing a pre-feature-post-explanation explainable analysis framework to provide pixel-level explanations for model inference and prediction. Experimental results show that compared with algorithms such as CORE(COnsistent REpresentation learning), SRM(Rich Models for Steganalysis), and UCF(Uncovering Common Features), the proposed method achieves the best performance on AUC and accuracy, with an accuracy range of 97.35% to 98.75% and an average accuracy of 98.22% on FaceForensics++ dataset, and the model’s interpretability also is improved significantly.

Table and Figures | Reference | Related Articles | Metrics
Moving object detection based on background image set and sparse analysis
BAO Jinyu WANG Huibin CHEN Zhe SHEN Jie
Journal of Computer Applications    2013, 33 (05): 1401-1410.   DOI: 10.3724/SP.J.1087.2013.01401
Abstract1169)      PDF (989KB)(807)       Save
This paper proposed a moving target detection method based on the background image set and sparse representation. The method combined the Robust Principal Component Analysis (RPCA) and the image block analysis method based on sparse representation. The authors got a series of background images from a video sequence by RPCA, and combined these background images as the background image set, treating image block as basic unit, and moving target was extracted from input frame by image block analysis method based on sparse representation. The simulation results indicate that when the background illumination mutates, the proposed method can effectively eliminate the impact of environment noise and reduce the false detection rate of target detection.
Reference | Related Articles | Metrics