Accessibility settings

Published on in Vol 6 (2025)

Preprints (earlier versions) of this paper are available at https://preprints.jmir.org/preprint/67661, first published .
Mobile app displaying patient data analysis and differential diagnosis for Cushing's disease.

Rapidly Benchmarking Large Language Models for Diagnosing Comorbid Patients: Comparative Study Leveraging the LLM-as-a-Judge Method

Rapidly Benchmarking Large Language Models for Diagnosing Comorbid Patients: Comparative Study Leveraging the LLM-as-a-Judge Method

Authors of this article:

Peter Sarvari1 Author Orcid Image ;   Zaid Al-fagih1 Author Orcid Image

Journals

  1. Sarvari P, Al-fagih Z. Authors’ Response to Peer Reviews of “Rapidly Benchmarking Large Language Models for Diagnosing Comorbid Patients: Comparative Study Leveraging the LLM-as-a-Judge Method”. JMIRx Med 2025;6:e81235 View
  2. Sarvari P, Al-fagih Z, Abou-Chedid A, Jewell P, Taylor R, Imtiaz A. Challenges and Solutions in Applying Large Language Models to Guideline-Based Management Planning and Automated Medical Coding in Health Care: Algorithm Development and Validation. JMIR Biomedical Engineering 2025;10:e66691 View

Books/Policy Documents

  1. Sanna L, Solinas E, Dragoni M. Artificial Intelligence in Medicine. View

Conference Proceedings

  1. Nanda S, Tripathy D, Ray N. 2025 OITS International Conference on Information Technology (OCIT). Automating Environmental Impact Assessments: An Agentic AI Framework for Comprehensive Report Generation View