Evaluation of DeepSeek-R1 and ChatGPT-4o on the Chinese national medical licensing examination: a multi-year comparative study
Abstract Large language models (LLMs) have demonstrated remarkable capabilities in natural language understanding and reasoning. However, their real-world applicability in high-stakes medical assessments remains underexplored, particularly in non-English contexts. This study aims to evaluate the per...
Gorde:
| Egile Nagusiak: | , , , , , , |
|---|---|
| Formatua: | Artigo |
| Hizkuntza: | Inglês |
| Argitaratua: |
Nature Portfolio
2026-01-01
|
| Saila: | Scientific Reports |
| Gaiak: | |
| Sarrera elektronikoa: | https://doi.org/10.1038/s41598-025-31874-6 |
| Etiketak: |
Etiketarik gabe, Izan zaitez lehena erregistro honi etiketa jartzen!
|
