科研速览 · Science Skim继续刷下去 · Keep skimming →
◇ bioRxiv2026-09-18· biochemistry

Machine Learning for Toxicity Prediction in Low-Sample Molecular Classes

C. Barajas, L. Dunphy, L. Mullany, O. Tiburzi, E. Lloyd

原始摘要(英文原文)· Original abstract
Deep learning models such as Chemprop have advanced quantitative molecular property prediction, but their reliance on large training sets limits use in data-scarce domains. We propose a framework that fine-tunes a general baseline model trained on publicly available data on small, class-specific datasets. The resulting models retain the baseline's generalization ability while gaining class-specific accuracy and produce probabilistic outputs that capture uncertainty in the training data. We demonstrate the approach on three toxicity classes defined by a common core structure, target, or mode of action: (i) organophosphates, (ii) androgen receptor antagonists, and (iii) estrogen receptor beta antagonists. Each fine-tuned model outperforms classical machine-learning methods and the EPA TEST tool. The probabilistic nature of the predictions enables prioritization of compounds for experimental validation and seamless integration with data streams of varying quality, supporting iterative decision-making in chemical safety and drug discovery.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Machine Learning for Toxicity Prediction in Low-Sample Molecular Classes — 科研速览 Science Skim