-
Loading metrics
Evaluation of the performance of large language models in responding to medical questions related to multiple sclerosis: A case study of large language models including ChatGPT, Gemini, Grok and Copilot
- Meisam Dastani,
- Mohammad Shayan Sajjadi,
- Bassem Yamout,
- Melika Arab Bafrani,
- Amirreza Nasirzadeh
x
- Published: May 11, 2026
- https://doi.org/10.1371/journal.pone.0346445