科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ PLOS digital health2026-09-01

Large language models and artificial intelligence for generating, simplifying, and enhancing outpatient clinic letters: A systematic review.

Rocco Sheldon, Morya Wadodkar, Andrew Gan, Rachael Yip, Yimeng Zhang, Jyoti Baharani

原始摘要(英文原文)· Original abstract
Outpatient clinic letters are a cornerstone of clinical communication but frequently exceed recommended reading levels, limiting patient comprehension and engagement. Large Language Models (LLMs) and Artificial Intelligence (AI) have been proposed as scalable tools for generating and simplifying these documents, but the evidence regarding their clinical accuracy, safety, and effectiveness remains fragmented. We conducted a systematic review of studies evaluating AI or LLMs for generating, simplifying, or enhancing outpatient clinic letters. Five databases (PubMed, EMBASE, Web of Science, CENTRAL, and CINAHL) were searched from inception to 1 November 2025. Two independent reviewers screened studies, extracted data, assessed quality using the Mixed Methods Appraisal Tool (MMAT), and certainty using the GRADE appraisal tool. Seven studies were included, comprising two studies using real-world clinic data and five using synthetic or hypothetical scenarios. Findings regarding readability were mixed; while some AI models improved readability scores compared to human-authored letters, most AI-generated content still failed to meet the recommended US Grade 6 reading level. The two studies that measured patient-reported outcomes reported high satisfaction and comprehension with AI-simplified letters, but these studies relied on subjective measures of understanding or lacked a human-authored control group. Clinical accuracy varied substantially, with information fidelity ranging from 10% to 100%. Risks of 'hallucinations', inappropriate medical advice, and tactless patient descriptions were documented. One study reported a ten-fold reduction in drafting time compared to human dictation. Overall certainty of evidence was rated very low across all primary outcomes due to methodological limitations, small sample sizes, and reliance on hypothetical scenarios. The current evidence supporting AI and LLMs use for outpatient clinic letters is preliminary and limited, providing insufficient evidence to support routine clinical implementation beyond supervised experimental and quality-improvement settings. Concerns regarding accuracy, safety, and generalisability remain. Further real-world patient-centred research is required before widespread adoption. PROSPERO registration: CRD420251181303.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Large language models and artificial intelligence for generating, simplifying, and enhancing outpatient clinic letters: A systematic review. — 科研速览 Science Skim