❌

Normal view

Context matching is not reasoning when performing generalized clinical evaluation of generative language models

npj Digital Medicine, Published online: 27 December 2025; doi:10.1038/s41746-025-02253-2

Context matching is not reasoning when performing generalized clinical evaluation of generative language models

Prompt Engineering in Clinical Practice: Tutorial for Clinicians

Large language models (LLMs), such as OpenAI’s GPT series and Google’s PaLM, are transforming healthcare by improving clinical decision-making, enhancing patient communication, and simplifying administrative tasks. However, their performance relies heavily on prompt design, where small changes in wording or structure can greatly impact output quality. This poses a challenge for clinicians who are not experts in natural language processing (NLP). This tutorial combines prompt engineering techniques tailored for clinical use, covering methods like zero-shot, few-shot, chain-of-thought, and meta-prompting. We examine four critical dimensions (accuracy, bias mitigation, privacy protection, and workflow integration) through clinical case studies grounded in real-world practice. We provide actionable guidance on defining objectives, applying core principles, iteratively refining prompts, and integrating them into interoperable electronic health record (EHR) systems. This framework helps clinicians leverage LLMs to improve decision-making, streamline documentation, and enhance patient communication while maintaining ethical standards and ensuring patient safety.

A pan-cancer single-cell panorama of human natural killer cells

Integrative single-cell RNA sequencing analyses on natural killer (NK) cells from over 700 patients across 24 tumor types depict shared and tumor-type-specific NK cell features and highlight the potential of specific myeloid cell subpopulations in regulating NK cell anti-tumor function.
❌