• Home
  • Search
  • Bi-Source Reconstruction-Based Classification Network for Face Forgery Video Detection
  • Cite Icon14
  • https://doi.org/10.1109/tcsvt.2023.3330390Copy DOI Icon

Bi-Source Reconstruction-Based Classification Network for Face Forgery Video Detection

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Current methods for detecting deep fakes concentrate on specific patterns of forgery like noise characteristics, local textures, or frequency statistics. These approaches assume training and test sets exhibit similar data distributions, which bring severe performance drops and further limit broader applications when migrating unseen domains. Existing works show that reconstruction learning is effective in capturing unseen forgery clues. However, 2D reconstruction is insufficient and can not handle non-frontal face reconstruction, while 3D reconstruction provides more critical details of facial structure and finds accurate forgery regions. In this paper, we propose a bi-source reconstruction based classification network (BRCNet) to incorporate 2D and 3D reconstruction as the supervisions and learn the optimal feature representation. In detail, we employ an encoder-decoder architecture to facilitate reconstruction learning, enhancing the learned representations to detect forgery patterns that are unknown. To further capture forgery evidence across multiple scales, instead of using encoder features from the reconstruction network only, we build a feature improvement network to combine feature details from encoder and decoder features in a multi-scale fashion. In addition, we use the reconstruction difference to supervise the feature aggregation, which enables detecting the subtle and trivial discrepancies between fake and real video frames. Extensive experiments are conducted to validate the performance of our proposed method on several deep fake benchmarks. The results demonstrate the efficacy of our approach, offering promising results and showcasing its potential for practical applications. The source code is available at https://github.com/cccvl/BRCNet.

Similar Papers
  • Book Chapter
  • Citations5

False Positive Suppression in Cervical Cell Screening via Attention-Guided Semi-supervised Learning

  • Jan 01, 2021
  • Xiaping Du +4
  • Conference Article
  • Citations2

Optimized Computation Combining Classification and Detection Networks with Distillation

  • Jul 18, 2021
  • Meng Xu +3
  • Research Article
  • Citations7

A Learnable and Explainable Wavelet Neural Network for EEG Artifacts Detection and Classification.

  • Jan 01, 2024
  • IEEE transactions on neural systems and rehabilitation engineering : a publication of the IEEE Engineering in Medicine and Biology Society
  • Yifei Yu +4
  • PDF
  • Research Article
  • Citations29

Generative Adversarial Networks for Abnormal Event Detection in Videos Based on Self-Attention Mechanism

  • Jan 01, 2021
  • IEEE Access
  • Weichao Zhang +4
  • Research Article
  • Citations24

Unsupervised Two-Path Neural Network for Cell Event Detection and Classification Using Spatiotemporal Patterns.

  • Dec 07, 2018
  • IEEE Transactions on Medical Imaging
  • Ha Tran Hong Phan +4
  • Conference Article
  • Citations14

Polyp Detection in Colonoscopy Videos by Bootstrapping Via Temporal Consistency

  • Apr 01, 2020
  • Yiting Ma +2
  • Book Chapter
  • Citations1

Missing Recover with Recurrent Neural Networks for Video Object Detection

  • Jan 01, 2018
  • Ranran Shen +3
  • Research Article

STFANet: A spatial and temporal feature aggregation network for fake face detection in videos

  • Dec 10, 2025
  • PLOS One
  • Guoren Yao +5
  • Research Article
  • Citations35

Salient object detection in video using deep non-local neural networks

  • Feb 19, 2020
  • Journal of Visual Communication and Image Representation
  • Mohammad Shokri +2
  • Research Article
  • Citations21

FTCF: Full temporal cross fusion network for violence detection in videos

  • Jun 06, 2022
  • Applied Intelligence
  • Tan Zhenhua +4
  • PDF
  • Research Article
  • Citations16

FLAME: fire detection in videos combining a deep neural network with a model-based motion analysis

  • Jan 09, 2025
  • Neural Computing and Applications
  • Diego Gragnaniello +3
  • Conference Article
  • Citations5

Part-based deep network for pedestrian detection in surveillance videos

  • Dec 01, 2015
  • Qi Chen +3
  • Conference Article
  • Citations2

Locality and Relative Distance-Aware Non-local Networks for Hand-Raising Detection in Classroom Video

  • Oct 01, 2021
  • Thu-Hien Le +7
  • Research Article
  • Citations1

ResNet50-deep affinity network for object detection and tracking in videos

  • Feb 29, 2024
  • International Journal of Advanced Technology and Engineering Exploration
  • Nandeeswar Sampigehalli Basavaraju +1
  • Research Article
  • Citations2

Visual-language collaborative multimodal transformer network for group activity detection in surveillance videos

  • Jul 09, 2025
  • Multimedia Systems
  • Fudong Nian +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.