科研速览 · Science Skim继续刷下去 · Keep skimming →
◇ medRxiv2026-07-31· Computer science

Large language models enable consensus-level interpretation in metagenomic diagnostics

Eike Steinig, Marcelina Krysiak, Kirti Deo, Andrew Duncan, Jacqueline Prestedge, Jeremy Barr, Jean Moselen, Sadid F. Khan, Janath A. Fernando, Ivana Savic, Bhargavi Yellapu, Ammar Aziz, Wytamma Wirth, Jessica Parry, Angela McDonald, Chhay Lim, Sharon Trevor, Benjamin Aw-Yeong, Georgia McCluskey, Michael Moso, Eddie Chan, Sonia L. La Vita, Penelope A. Bryant, Amy Crowe, Ramla Maalim, Diana Velasquez Reyes, Maryza Graham, Eloise Williams, Jason C. Kwong, Rachel Woolstencroft, Monica Slavin, Lyndell L. Lim, Lachlan Coin, Leon Caly, Katherine Bond, Chuan Kok Lim, Timothy P. Stinear, Deborah A Williamson, Prashanth Ramachandran

原始摘要(英文原文)· Original abstract
Abstract Metagenomic sequencing can detect a broad range of pathogens, but interpreting which detections are clinically relevant requires expert adjudication that is difficult to scale and standardize. Here we present diagnostic classifiers that formalize expert adjudication by combining structured decision trees with large language model reasoning to assign diagnoses and select pathogen candidates. We first developed a short-read metagenomic assay for sterile-site specimens (cerebrospinal and ocular fluid) in the META-GP study (Victoria, Australia, 2024-2025) and evaluated classifiers on a validation dataset (n = 96; clinical samples, spike-ins and controls). Locally deployed, open-weight reasoning models (Ǫwen3) achieved diagnostic performance comparable to expert consensus, improving with clinical context (n = 79, above experimental limit-of-detection; without clinical notes, 94.4% sensitivity, 95.4% specificity; with clinical notes, 97.2% sensitivity, 100% specificity). Automated adjudication enabled systematic benchmarking of computational parameters and regression testing for pathogen detection tasks. In a heterogeneous development cohort (n = 78), reviewers and classifiers identified clinically significant pathogens missed during routine testing. By reproducing consensus detections without requiring a full review panel, diagnostic classifiers enable scalable, standardized metagenomic interpretation that complements expert adjudications.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Large language models enable consensus-level interpretation in metagenomic diagnostics — 科研速览 Science Skim