AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (12.5 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Review | Open Access

Transforming Healthcare with State-of-the-Art Medical-LLMs: A Comprehensive Evaluation of Current Advances Using Benchmarking Framework

Himadri Nath Saha1Dipanwita Chakraborty Bhattacharya2( )Sancharita Dutta3Arnab Bera3Srutorshi Basuray4Satyasaran Changdar5Saptarshi Banerjee6Jon Turdiev7
Department of Computer Science, SNEC, University of Calcutta, Kolkata, 700073, India
Department of Computer Science, PRTGC, West Bengal State University, Barasat, 700126, India
Department of Computer Science & Engineering, The Neotia University, Kolkata, 743368, India
Department of Computer Science & Engineering, University College of Science and Technology, University of Calcutta, Kolkata, 700009, India
Department of Food Science, University of Copenhagen, Copenhagen, 1165, Denmark
Department of Computer Science, Illinois Institute of Technology, 10 West 35th Street, Chicago, IL 60616, USA
Department of Computer Science, San Francisco State University, 1600 Holloway Avenue, San Francisco, CA 94132, USA
Show Author Information

Abstract

The emergence of Medical Large Language Models has significantly transformed healthcare. Medical Large Language Models (Med-LLMs) serve as transformative tools that enhance clinical practice through applications in decision support, documentation, and diagnostics. This evaluation examines the performance of leading Med-LLMs, including GPT-4Med, Med-PaLM, MEDITRON, PubMedGPT, and MedAlpaca, across diverse medical datasets. It provides graphical comparisons of their effectiveness in distinct healthcare domains. The study introduces a domain-specific categorization system that aligns these models with optimal applications in clinical decision-making, documentation, drug discovery, research, patient interaction, and public health. The paper addresses deployment challenges of Medical-LLMs, emphasizing trustworthiness and explainability as essential requirements for healthcare AI. It presents current evaluation techniques that improve model transparency in high-stakes medical contexts and analyzes regulatory frameworks using benchmarking datasets such as MedQA, MedMCQA, PubMedQA, and MIMIC. By identifying ongoing challenges in bias mitigation, reliability, and ethical compliance, this work serves as a resource for selecting appropriate Med-LLMs and outlines future directions in the field. This analysis offers a roadmap for developing Med-LLMs that balance technological innovation with the trust and transparency required for clinical integration, a perspective often overlooked in existing literature.

References

【1】
【1】
 
 
Computers, Materials & Continua
Pages 1-56

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Saha HN, Bhattacharya DC, Dutta S, et al. Transforming Healthcare with State-of-the-Art Medical-LLMs: A Comprehensive Evaluation of Current Advances Using Benchmarking Framework. Computers, Materials & Continua, 2026, 86(2): 1-56. https://doi.org/10.32604/cmc.2025.070507

5

Views

0

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 17 July 2025
Accepted: 16 September 2025
Published: 09 December 2025
© The Author 2025.

This work is licensed under a Creative Commons Attribution 4.0 International License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.