Accessibility settings

Published on in Vol 11 (2025)

Preprints (earlier versions) of this paper are available at https://preprints.jmir.org/preprint/70190, first published .
Doctor in scrubs typing on a laptop with stethoscope around neck

Enhancing Large Language Models for Improved Accuracy and Safety in Medical Question Answering: Comparative Study

Enhancing Large Language Models for Improved Accuracy and Safety in Medical Question Answering: Comparative Study

Journals

  1. Callens S. Effective prompt design for large language models in clinical practice. Acta Clinica Belgica 2026;81(2):118 View
  2. Pawlik L, Deniziak S. Reducing Hallucinations in Medical AI Through Citation Enforced Prompting in RAG Systems. Applied Sciences 2026;16(6):3013 View
  3. Li S, Wang X, Chen Y, Tian M, Lin P, Lai M, Jiang L. Large language models for primary care ophthalmic education: a systematic review. Frontiers in Medicine 2026;13 View
  4. Choi S, Kim D, Jeon J, Kim M, Lee D, Ahn D, Lee E, Kim Y, Youk H. Korean Medical Consultation With Open-Weight Large Language Models: Pilot Comparative Evaluation of Retrieval-Augmented Generation With Metadata Filtering. JMIR Formative Research 2026;10:e72604 View
  5. Garcia-Lopez I, Molina-Espinosa J, Ramirez-Montoya M. Optimizing Open Large Language Models for Equity and Accessibility in Higher Education Through Pruning and Retrieval-Augmented Generation. IEEE Revista Iberoamericana de Tecnologias del Aprendizaje 2026;21:466 View
  6. Canpolat A, Aksu Ö, Emral R, Canpolat U. Diagnostic Performance and Workup Efficiency of Large Language Models in Secondary Hypertension: A Blinded Comparative Study. Diagnostics 2026;16(14):2165 View
  7. Zhao Y, Miao Y, Guo R, Luo Y, Wang H, Wu Y. Evaluation Methods for Inference-Time Retrieval-Augmented and Graph Retrieval-Augmented Large Language Models in Health Care: Scoping Review. Journal of Medical Internet Research 2026;28:e90046 View
  8. Rajitha A. Emerging Risks of Generative AI in Healthcare: A Scientometric Analysis of Hallucinations and Medical Misinformation. The International Information & Library Review 2026:1 View
  9. Sun S, Niu D, Yuan M, Liu L, Li Y, Min H. Evaluating large language models in Helicobacter pylori-related question answering: From knowledge tests to patient queries. World Journal of Gastroenterology 2026;32(37) View
  10. Zallipalli S, Dendukuri N, Asokan G, Peri H, Kakarla M. AI before the doctor: Patients' use of artificial intelligence to interpret symptoms, explore diagnoses, and prepare for medical encounters. Intelligent Hospital 2026:100115 View
  11. Christopher Leela B, Jain N, Annambhotla T, Siddharthan A, Kanala H, Budhiraja A, Goutam Sinha M, Charan D, Ahmad Dar M, Patel T. Large Language Model Framework for Structured Biomarker Extraction and Triple Negative Breast Cancer Identification from Pathology Reports: Development and Evaluation Study (Preprint). JMIR Formative Research 2026 View

Conference Proceedings

  1. Chen J, Wang D, Hu Y, Yuan Q, Lin L, Zhang Y. Proceedings of the 2026 International Conference on Big Data and Informatization Education. Construction of a Retrieval-Augmented LLM Agent for Health Education and a Multi-dimensional Evaluation Framework: A Comparative Study of Chinese Large Language Models View
  2. Wang X, Li K, Hou S, Zhang Z. Proceedings of the 2026 2nd International Symposium on Bioinformatics and Computational Biology. A Multi-Agent Collaborative Chest X-Ray Diagnostic System Integrating DenseNet, Knowledge Graph, and Large Language Model View