AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (6.7 MB)
Collect
AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Special Topic | Open Access

From TCMBank to translational hypotheses: A machine-learning framework for prioritizing Chinese medicinal herbs for liver diseases

Junqi CuiaWeijia LibEnoch Chi Ngai LimcXiaoqin WudFang CheneChi Eung Danforn Limc,e,f( )
School of Health Sciences, University of New South Wales, Kensington NSW 2052, Australia
School of Cardiovascular Medicine and Sciences, King's College London, London SE5 9NU, United Kingdom
Translational Research Department, Specialist Medical Services Group, Earlwood NSW 2206, Australia
Sydney Institute of Traditional Chinese Medicine, Haymarket NSW 2000, Australia
Data Science Institute, University of Technology Sydney, Ultimo NSW 2007, Australia
NICM Health Research Institute, Western Sydney University, Westmead NSW 2145, Australia

Peer review under responsibility of Beijing University of Chinese Medicine.

Show Author Information

Abstract

Objective

To develop a machine-learning framework integrating TCMBank-derived liver disease seed curation with structured herb-level annotations to prioritize herbs with potential relevance to liver disease.

Methods

Liver-focused non-tumor disease seeds were curated from TCMBank to identify 356 annotation-supported positive herbs. Herb-level features were derived from structured TCMBank annotations. Six unweighted machine-learning models were trained using the original 356-positive/8835-background positive-unlabeled dataset. Performance was assessed using the area under the receiver operating characteristic curve and the area under the precision-recall curve (PR-AUC), with PR-AUC interpreted relative to the baseline prevalence of 0.039. Candidate herbs were further refined through consensus prioritization, cross-model concordance, and translational evidence evaluation.

Results

A total of 124 curated liver disease seeds and 9191 TCMBank herb records were retained. The final modeling dataset comprised 356 annotation-defined positive herbs and 8835 unlabeled background herbs, corresponding to a positive prevalence of 0.039. Model performance was evaluated using the 356-positive positive-unlabeled dataset. PR-AUC baselines, candidate rankings, and validation-priority scores were generated within and aligned with the final modeling framework.

Conclusions

This study establishes a reproducible machine-learning framework for prioritizing candidate herbs in liver disease research and provides a data-driven strategy for translating large-scale database resources for Chinese medicine into experimentally-testable hypotheses. The prioritized candidates should be regarded as computational hypotheses requiring staged pharmacological, hepatobiliary, and safety validation rather than as evidence of established clinical efficacy.

References

【1】
【1】
 
 
Journal of Traditional Chinese Medical Sciences
Pages 310-318

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Cui J, Li W, Lim ECN, et al. From TCMBank to translational hypotheses: A machine-learning framework for prioritizing Chinese medicinal herbs for liver diseases. Journal of Traditional Chinese Medical Sciences, 2026, 13(3): 310-318. https://doi.org/10.1016/j.jtcms.2026.06.003

5

Views

0

Downloads

0

Crossref

0

Scopus

Received: 11 May 2026
Revised: 12 June 2026
Accepted: 14 June 2026
Published: 23 June 2026
© 2026 Beijing University of Chinese Medicine.

This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).