AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.9 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese

A Cross-Modal Face Retrieval Algorithm Based on Metric Learning

Yan WO( )Jiyun LIANGGuoqiang HAN
School of Computer Science and Engineering, South China University of Technology, Guangzhou 510006, Guangdong, China
Show Author Information

Abstract

The existing cross-modal retrieval algorithms based on metric learning ignore the pose differences and domain differences in cross-modal face retrieval tasks. In addition, these algorithms lack learning of global information in the process of metric learning and construct a large number of redundant triplets. Therefore, a cross-modal common representation generation algorithm based on metric learning was proposed in this paper. The algorithm uses the yaw angle equivariant module compensating for yaw angle differences to obtain the image features with robustness, uses the multi-layer attention mechanism to obtain video features with differentiability, uses global triplets and local triplets to jointly train the cross-modal common representation generation network, so as to improve the consistency and accuracy of metric learning. Then it accelerates the convergence of loss functions through the screening of semi-hard triplets. This study proposed a domain adaption algorithm which combines domain calibration and transfer learning to improve the generalization of common representations. The results of comparative experiments on three face video datasets, namely, PB, YTC, and UMD, demonstrate that the algorithm can improve the accuracy of cross-modal face retrieval, and fine-tuning the cross-modal common representation generation network with few samples can improve the accuracy of cross-modal retrieval using target domain images.

CLC number: TP391 Article ID: 1000-565X(2022)06-0001-09

References

【1】
【1】
 
 
Journal of South China University of Technology (Natural Science Edition)
Pages 1-9

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
WO Y, LIANG J, HAN G. A Cross-Modal Face Retrieval Algorithm Based on Metric Learning. Journal of South China University of Technology (Natural Science Edition), 2022, 50(6): 1-9. https://doi.org/10.12141/j.issn.1000-565X.210709

424

Views

1

Downloads

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 09 November 2021
Published: 25 June 2022
© Journal of South China University of Technology (Natural Science Edition)