AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.1 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

Cross-modal deep fusion based on small samples for early rumor detection

Junqing Yang1Hongzhe Chen2Yao Zhao2Wing-Kuen Ling3Yang Zhou4( )
School of Photoelectric Engineering, Jiangxi Modern Polytechnic College, Jiangxi 330046, China
School of Automation, Guangdong University of Technology, Guangzhou 510006, China
Center for Integrated Circuits and Artificial Intelligence, Tsientang Institute for Advanced Study, Zhejiang 310024, China
School of Statistics, Beijing Normal University, Beijing 100875, China
Show Author Information

Abstract

Rumors circulating on social media can have significant adverse impacts on society, highlighting the urgent need for effective rumor detection. However, most existing methods predominantly focus on rumor identification only after widespread dissemination has occurred, when substantial harm has been inflicted. Early-stage rumor detection is a major challenge due to limited reach and small sample sizes, which restrict the use of large datasets and traditional propagation models. To overcome these challenges, a cross-modal deep fusion method based on small samples for early rumor detection is proposed, which includes a multimodal feature extraction network and a cross-modal deep fusion network. The multimodal feature extraction network captures features from multiple modalities, while the multimodal deep information extraction network derives deep representations from these modalities. The cross-modal deep fusion network integrates textual and visual features for rumor classification. Furthermore, an enhanced meta-learning training approach based on model-agnostic meta-learning is proposed to improve the efficiency of rumor detection by employing distinct learning rates both within and between tasks. Experimental results on two publicly available datasets demonstrate that the proposed cross-modal deep fusion method outperforms baseline methods and exhibits promising performance.

References

【1】
【1】
 
 
Electronic Research Archive
Pages 232-250

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Yang J, Chen H, Zhao Y, et al. Cross-modal deep fusion based on small samples for early rumor detection. Electronic Research Archive, 2026, 34(1): 232-250. https://doi.org/10.3934/era.2026012

191

Views

0

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 02 November 2025
Revised: 23 December 2025
Accepted: 25 December 2025
Published: 08 January 2026
©2026 the Author(s), licensee AIMS Press.

This is an open access article distributed under the terms of the Creative Commons Attribution License (http://creativecommons.org/licenses/by/4.0)