AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (2 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Open Access

Long-tailed representation learning algorithm based on adaptive prototypes and semantic awareness

Tiantian LIZhen XUE( )Liangliang ZHANGXu LIAN
School of Mathematics, North University of China, Taiyuan 030051, China
Show Author Information

Abstract

Label scarcity and long-tailed distribution imbalance are significant challenges in industrial equipment monitoring. Currently, self-supervised learning methods are affected by sample quantity bias and semantic confusion under complex operating conditions, which limits their ability to represent sparse critical states. To address these issues, we propose a co-evolutionary prototypical contrastive learning (EPCL) framework. Through progressive learning from coarse-grained semantic discovery to fine-grained discriminative enhancement, this framework enables an in-depth analysis of the intrinsic structure of long-tailed data. Specifically, an adaptive prototype-based clustering algorithm based on optimal transport theory is introduced, thereby achieving unbiased representation learning through data-driven dynamic priors. Furthermore, a semantic-aware and hierarchical negative sample weighting scheme is designed to optimize discriminative boundaries while mitigating class imbalance by enforcing prototype consistency constraints and employing an adaptive weighting strategy. Extensive experiments were conducted on several public long-tailed visual benchmarks, including CIFAR10-LT, CIFAR100-LT, and ImageNet-100-LT, as well as the industrial fault diagnosis dataset. The results demonstrated that the EPCL achieved better performance than fifteen mainstream self-supervised methods (e.g., SimCLR and SwAV) in both linear evaluation and few-shot classification tasks. On the CIFAR100-LT dataset, the EPCL improved the tail-class accuracy by 4.56% compared to SimCLR. Ablation studies and visualization results verified the effectiveness and generalization ability of the framework. This work offers a promising insight and practical solution for representation learning from unlabeled long-tailed measurement data.

References

【1】
【1】
 
 
Journal of Measurement Science and Instrumentation
Pages 331-343

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
LI T, XUE Z, ZHANG L, et al. Long-tailed representation learning algorithm based on adaptive prototypes and semantic awareness. Journal of Measurement Science and Instrumentation, 2026, 17(2): 331-343. https://doi.org/10.62756/jmsi.1674-8042.2026028

248

Views

4

Downloads

0

Crossref

0

CSCD

Received: 30 December 2025
Revised: 20 January 2026
Accepted: 03 March 2026
Published: 01 June 2026
© The Author(s) 2026.

The articles published in this open access journal are distributed under the terms of the Creative Commons Attribution 4.0 International License (http://creativecommons.org/licenses/by/4.0/), which permits use, distribution and reproduction in any medium, provided the original work is properly cited.