AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (2.4 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

TSTBFuse: a two-stage three-branch feature extraction method for infrared and visible image fusion

Wangwei Zhang1Xinyue Qin1Menghao Dai1Bin Zhou2( )Changhai Wang1ZhiHeng Wang3SongZe Li4
Software Engineering College, Zhengzhou University of Light Industry, No.136 Science Road, Zhengzhou 450000, China
Electronics and Electrical Engineering College, Zhengzhou University of Science and Technology, No.1 Xueyuan Road, Zhengzhou 450064, China
Zhengzhou Xinda Institute of Advanced Technology, Zhengzhou, China
Henan Xindawangyu Science & Technology Co., Ltd., Zhengzhou, China
Show Author Information

Abstract

The purpose of image fusion is to combine information from different source images to produce a comprehensively representative image. Traditional autoencoder architectures often struggle to effectively extract both unique and shared features from these image types. A novel two-stage three-branch feature extraction method (TSTBFuse) was proposed in the study, specialized for the fusion of infrared and visible images. The proposed architecture employed a three-branch encoder that separately captured infrared-specific thermal radiation features, visible-specific texture details, and shared structural information. A two-stage end-to-end training strategy was introduced: the first stage focused on reconstructing the original input images to preserve modality-specific information, while the second stage leveraged the learned representations to generate high-quality fused images. we designed a comprehensive loss function combining mean squared error (MSE), structural similarity index (SSIM), and gradient loss, ensuring both pixel-level accuracy and structural integrity. Extensive experiments on public datasets (TNO, MSRS and RoadScene) demonstrated that TSTBFuse consistently outperformed seven state-of-the-art methods in both subjective and objective evaluations. Furthermore, the method exhibited strong generalization capabilities, successfully extending to challenging tasks such as magnetic resonance imaging-computed tomography (MRI-CT) medical image fusion and red-green-blue (RGB)-infrared image fusion without retraining. The code is publicly available at: https://github.com/QXinYue/TSTBFuse.

References

【1】
【1】
 
 
Electronic Research Archive
Pages 4045-4073

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Zhang W, Qin X, Dai M, et al. TSTBFuse: a two-stage three-branch feature extraction method for infrared and visible image fusion. Electronic Research Archive, 2025, 33(6): 4045-4073. https://doi.org/10.3934/era.2025180

377

Views

28

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 17 April 2025
Revised: 09 June 2025
Accepted: 13 June 2025
Published: 27 June 2025
©2025 the Author(s), licensee AIMS Press.

This is an open access article distributed under the terms of the Creative Commons Attribution License (http://creativecommons.org/licenses/by/4.0)