Benchmarking the Confidence of Large Language Models in Answering Clinical Questions: Cross-Sectional Evaluation Study
Abstract BackgroundThe capabilities of large language models (LLMs) to self-assess their own confidence in answering questions within the biomedical realm remain underexplored. ObjectiveThis study evaluates the confidence levels of 12 LLMs across 5 medical specialt...
Salvato in:
| Autori principali: | , , , , |
|---|---|
| Natura: | Artigo |
| Lingua: | Inglês |
| Pubblicazione: |
JMIR Publications
2025-05-01
|
| Serie: | JMIR Medical Informatics |
| Accesso online: | https://medinform.jmir.org/2025/1/e66917 |
| Tags: |
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
