AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (20.7 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

Edit-discrepancy-guided feature transformation in encoder-based GAN inversion for real image attribute editing

Wenbo Yan1Xing Xu1,2( )Yinglong Zhang1Xuewen Xia1Yuanxiang Li3,4
School of Physics and Information Engineering, Minnan Normal University, Zhangzhou 363000, China
Center for China-ASEAN Regional Collaborative Development, Minnan Normal University, Zhangzhou 363000, China
School of Computer Science, Wuhan University, Wuhan 430072, China
Digital Strategy Development Research Institute of Hechi University, Hechi 546399, China
Show Author Information

Abstract

Image attribute editing based on generative adversarial networks (GANs) typically begins by mapping real images to the latent space of a pretrained StyleGAN, followed by manipulating the corresponding latent codes. However, low-rate latent codes suffer from an information bottleneck, making it challenging to faithfully reconstruct complex real images. Recent encoder-based methods enhance reconstruction by injecting high-rate features into intermediate generator layers to better preserve fine details, but they often yield misaligned details in the edited images. The primary reason is that these methods still rely on global linear transformations of high-rate features, which overlook the nonlinear and spatially localized nature of real edits. To this end, we build on a high-fidelity encoder-based GAN inversion backbone and introduce an additional adaptive feature editor that is specifically trained to convert high-rate features during editing so that fine details are correctly aligned with the edited image. The backbone refines the feature through a cross-attention mechanism and residual enhancement. Building on this, the feature editor employs a window-based cross-attention mechanism to extract a discrepancy signal between the original and edited generator features, which specifies both where to modify and what content to change. This signal is then fused into the feature through spatially adaptive modulation techniques, enabling region-selective attribute changes while preserving irrelevant details. Experiments on face and car benchmarks demonstrate that our method improves both reconstruction fidelity and editing quality compared to existing GAN inversion methods.

References

【1】
【1】
 
 
Electronic Research Archive
Pages 4889-4912

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Yan W, Xu X, Zhang Y, et al. Edit-discrepancy-guided feature transformation in encoder-based GAN inversion for real image attribute editing. Electronic Research Archive, 2026, 34(7): 4889-4912. https://doi.org/10.3934/era.2026216

8

Views

0

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 13 December 2025
Revised: 26 April 2026
Accepted: 20 May 2026
Published: 15 July 2026
©2026 the Author(s), licensee AIMS Press.

This is an open access article distributed under the terms of the Creative Commons Attribution License (http://creativecommons.org/licenses/by/4.0)