Abstract
Tan et al. proposed a S.C.O.R.E. framework to evaluate clinical responses from large language models (LLMs) in terms of Safety, Consensus & Context, Reproducibility, and Explainability. S.C.O.R.E. supports clinical LLM validation by providing structured, actionable insights to guide model optimization and refinement.
| Original language | English |
|---|---|
| Article number | 102883 |
| Journal | Cell Reports Medicine |
| Volume | 7 |
| Issue number | 7 |
| DOIs | |
| State | Published - Jul 21 2026 |
Keywords
- ChatGPT
- Claude
- DeepSeek
- chatbots
- evaluation
- framework
- healthcare
- large language models
- medicine
Fingerprint
Dive into the research topics of 'A S.C.O.R.E. framework for evaluating open-ended responses from large language models in healthcare'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver