AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.1 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese | Open Access

Reinforcement Learning-driven Multi-round Data Augmentation for Web Service Classification

School of Computer Science and Technology, Guangdong University of Technology, Guangzhou 510006, China
School of Automation, Guangdong University of Technology, Guangzhou 510006, China
Show Author Information

Abstract

To address the difficulty of recognizing tail classes in Web service classification caused by long-tailed data distributions, a reinforcement learning-enhanced framework is proposed, which integrates multi-round data augmentation with adaptive loss optimization. A large language model (LLM) is employed as the core semantic generator, and a reinforcement learning agent adaptively adjusts class-specific augmentation ratios and filtering thresholds at each iteration based on the observed environment state, driving a closed-loop multi-round process of “generation-evaluation-filtering-feedback.” In parallel, a Chain-of-Thought-based reasoning mechanism is introduced to evaluate generated samples from multiple dimensions-including novelty, semantic consistency, and reasoning quality-thereby filtering out template-like and semantically drifting instances and progressively improving the training data distribution. During classifier training, class weights are dynamically computed from the frequency statistics of the augmented dataset at each iteration. A Top-k Near-Miss Focal Loss is further designed to jointly emphasize long-tailed classes and near-miss boundary samples, penalizing ambiguous semantic regions and enabling adaptive loss optimization tailored to long-tailed and hard examples. Experiments conducted on a real-world long-tailed Web service dataset and the PMTD (Productive Math Tutoring Dialogue) instructional dialogue dataset demonstrate that the proposed method outperforms mainstream baselines such as NCAL (Neural-Collapse-Advanced Personalized Learning), RGPT, SRaSLR (Social Relation Aware Service Label Recommendation Model) and LLMEmbed across multiple evaluation metrics. In particular, substantial improvements are observed for tail-class recognition: on several lightweight models, Macro- F1 improves by up to 5 percentage points on the Web service dataset, and Weighted- F1 increases by approximately 2-3 percentage points. These results verify the effectiveness of the proposed approach in mitigating the bias introduced by long-tailed distributions and provide a practical solution for intelligent recognition and classification of semantically sparse services in open environments.

CLC number: TP391 Document code: A Article ID: 1007–7162(2026)4–80–11

References

【1】
【1】
 
 
Journal of Guangdong University of Technology
Pages 80-90

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
He H, Chen C, Wang T, et al. Reinforcement Learning-driven Multi-round Data Augmentation for Web Service Classification. Journal of Guangdong University of Technology, 2026, 43(4): 80-90. https://doi.org/10.12052/gdutxb.250184

8

Views

0

Downloads

0

Crossref

Received: 16 October 2025
Accepted: 23 December 2025
Published: 14 April 2026
© 2026 Editorial Office of Journal of Guangdong University of Technology

This is an open access article under the CC BY-NC-ND license (https://creativecommons.org/licenses/by-nc-nd/4.0/).