// NATURE NEWS — SPAZIO & SCIENZA
Safety and security of large language models in healthcare
Nature
volume 656, pages 577–589 (2026) Cite this article
Integration of artificial intelligence methods into clinical care is proceeding rapidly, driven by advances in generative artificial intelligence, most notably large language models. Large language models trained on large amounts of text have shown potential across nearly every domain of healthcare. However, their broad applicability also comes with new responsibilities, vulnerabilities and threats. These need to be assessed and mitigated before widespread clinical adoption. Here we review the available literature on security and safety of large language models themselves as well as their integration with hospital workflows and interactions with human healthcare providers. We systematically map security hazards to development stages of clinical artificial intelligence systems (design, data, model, inference and environment), identify safety layers, from core optimization objectives, knowledge integrity and alignment, to interaction with humans and systems, and classify threats by their current clinical relevance. Finally, we provide a perspective on current mitigation techniques, illustrating respective stakeholders’ responsibilities.
This is a preview of subscription content, access via your institution
Access Nature and 54 other Nature Portfolio journals
Get Nature+, our best-value online-access subscription
Prices may be subject to local taxes which are calculated during checkout
Gommers, J. et al. Interval cancer, sensitivity, and specificity comparing AI-supported mammography screening with standard double reading without AI in the MASAI study: a randomised, controlled, non-inferiority, single-blinded, population-based, screening-accuracy trial. Lancet 407, 505–514 (2026).
Article
PubMed
Google Scholar
Lu, M. Y. et al. A multimodal generative AI copilot for human pathology. Nature 643, 466–473 (2024).
Article
ADS
CAS
PubMed
PubMed Central
Google Scholar