AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (6.4 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access | Just Accepted

Towards Robust and Highly Stealthy Proactive Deepfake Defense on Consumer Electronic Devices

Chengsheng Yuan1Youqiang Cao1YuChen Wu1Fengjun Xiao2( )Xinting Li3( )Zhihua Xia4

1 Engineering Research Center of Digital Forensics, Ministry of Education, School of Computer Science, Nanjing University of Information Science and Technology, Nanjing 210044, China

2 Zhejiang Informatization Development Insti-tute, Hangzhou Dianzi University, Hangzhou 310018, China

3 School of Foreign Languages, National University of Defense Technology, Nanjing 210039, China

4 College of Cyber Security, Jinan University, Guangzhou 510632, China

Show Author Information

Abstract

The integration of AIGC technologies into consumer electronics has democratized the creation of highly realistic facial deepfakes, posing unprecedented threats to personal identity security and undermining trust in digital ecosystem. Current proactive defense methods primarily operate by injecting perturbations directly into the pixel space of facial images to disrupt deepfake synthesis. However, such pixel-level modifications often introduce noticeable artifacts, making them easily detectable by human observers and vulnerable to common image transformations. To tackle this issue, this paper proposes DiffDefend, a novel three-stage proactive defense framework based on low-frequency perturbations in the diffusion latent space. The framework first employs a stable diffusion model to encode the input image into a latent representation that preserves high-fidelity reconstruction capability. Adversarial perturbations are then introduced specifically into the low-frequency subband of the latent code, which is extracted via discrete wavelet transform (DWT/Discrete Wavelet Transform), and optimized through an adversarial game combining visual and adversarial losses, yielding an adversarial latent representation. Ultimately, the adversarial latent code is reconstructed into an adversarial image leveraging the powerful generative capabilities of the diffusion model, effectively defending against deepfake manipulations. Moreover, we design a multi-level loss that combines pixel-level and semantic-level constraints to enhance the visual quality and defense efficacy of the adversarial image. To evaluate the efficacy of our method, we conduct experiments on two representative deepfake tasks: attribute editing and face reenactment. Experimental results demonstrate that our method achieves superior performance over existing benchmarks in both visual quality and defensive performance.

References

【1】
【1】
 
 
Tsinghua Science and Technology

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Yuan C, Cao Y, Wu Y, et al. Towards Robust and Highly Stealthy Proactive Deepfake Defense on Consumer Electronic Devices. Tsinghua Science and Technology, 2026, https://doi.org/10.26599/TST.2026.90100020
Part of a topical collection:

1026

Views

103

Downloads

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 04 December 2025
Revised: 03 January 2026
Accepted: 16 January 2026
Available online: 10 February 2026

© The author(s) 2026.

The articles published in this open access journal are distributed under the terms of the Creative Commons Attribution 4.0 International License (http://creativecommons.org/licenses/by/4.0/).