科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ eLife2026-07-31· Context (archaeology)

Benchmarking biochemical networks generated by large language models

Jeevan Tewari, Benjamin W Dahl, B Adam Bates, Jason A. Papin, Jeffrey J. Saucerman

原始摘要(英文原文)· Original abstract
Computational models of biochemical networks provide frameworks for predicting how molecular cues guide cell decisions. These models are typically limited by the time-intensive manual curation required to extract network mechanisms from incomplete literature. Here, we test whether general-purpose large language models (LLMs) can generate accurate models of signaling and metabolic networks. We find that general-purpose LLMs generate 24–65% of the reactions of literature-curated signaling networks for cardiomyocyte hypertrophy, myofibroblast activation, and mechanosignaling. Further, logic-based models based on these networks predict responses to perturbations with accuracies of 6–33%. In the context of metabolic modeling, LLMs are able to generate 64–91% of the reactions within the core Escherichia coli metabolic network and demonstrate highly variable accuracies in predicting substrate utilization. Current general-purpose LLMs generate biochemical networks with moderate accuracy, and this study provides a pipeline and benchmarks to guide future improvements.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Benchmarking biochemical networks generated by large language models — 科研速览 Science Skim