• 综合
  • 标题
  • 关键词
  • 摘要
  • 学者
  • 期刊-刊名
  • 期刊-ISSN
  • 会议名称
搜索

作者:

Zhang, Dongming (Zhang, Dongming.) | Fu, Chenqin (Fu, Chenqin.) | Lu, Dingyu (Lu, Dingyu.) | Li, Jun (Li, Jun.) | Zhang, Yongdong (Zhang, Yongdong.)

收录:

EI Scopus SCIE

摘要:

Current methods for detecting deep fakes concentrate on specific patterns of forgery like noise characteristics, local textures, or frequency statistics. These approaches assume training and test sets exhibit similar data distributions, which bring severe performance drops and further limit broader applications when migrating unseen domains. Existing works show that reconstruction learning is effective in capturing unseen forgery clues. However, 2D reconstruction is insufficient and can not handle non-frontal face reconstruction, while 3D reconstruction provides more critical details of facial structure and finds accurate forgery regions. In this paper, we propose a bi-source reconstruction based classification network (BRCNet) to incorporate 2D and 3D reconstruction as the supervisions and learn the optimal feature representation. In detail, we employ an encoder-decoder architecture to facilitate reconstruction learning, enhancing the learned representations to detect forgery patterns that are unknown. To further capture forgery evidence across multiple scales, instead of using encoder features from the reconstruction network only, we build a feature improvement network to combine feature details from encoder and decoder features in a multi-scale fashion. In addition, we use the reconstruction difference to supervise the feature aggregation, which enables detecting the subtle and trivial discrepancies between fake and real video frames. Extensive experiments are conducted to validate the performance of our proposed method on several deep fake benchmarks. The results demonstrate the efficacy of our approach, offering promising results and showcasing its potential for practical applications. The source code is available at https://github.com/cccvl/BRCNet.

关键词:

multi-scale feature aggregation 3D reconstruction Face forgery video detection Image reconstruction Faces Feature extraction Three-dimensional displays multi-scale attention Forgery Decoding Face recognition feature alignment

作者机构:

  • [ 1 ] [Zhang, Dongming]Peoples Daily Online, State Key Lab Commun Content Cognit, Beijing 100733, Peoples R China
  • [ 2 ] [Fu, Chenqin]Peoples Daily Online, State Key Lab Commun Content Cognit, Beijing 100733, Peoples R China
  • [ 3 ] [Li, Jun]Peoples Daily Online, State Key Lab Commun Content Cognit, Beijing 100733, Peoples R China
  • [ 4 ] [Lu, Dingyu]Beijing Univ Technol, Fac Informat Technol, Beijing 100021, Peoples R China
  • [ 5 ] [Zhang, Yongdong]Univ Sci & Technol China, Sch Informat Sci & Technol, Hefei 230026, Peoples R China

通讯作者信息:

  • [Zhang, Dongming]Peoples Daily Online, State Key Lab Commun Content Cognit, Beijing 100733, Peoples R China;;

电子邮件地址:

查看成果更多字段

相关关键词:

相关文章:

来源 :

IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY

ISSN: 1051-8215

年份: 2024

期: 6

卷: 34

页码: 4257-4269

8 . 4 0 0

JCR@2022

被引次数:

WoS核心集被引频次:

SCOPUS被引频次: 2

ESI高被引论文在榜: 0 展开所有

万方被引频次:

中文被引频次:

近30日浏览量: 1

归属院系:

在线人数/总访问数:459/4978993
地址:北京工业大学图书馆(北京市朝阳区平乐园100号 邮编:100124) 联系我们:010-67392185
版权所有:北京工业大学图书馆 站点建设与维护:北京爱琴海乐之技术有限公司