科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Frontiers in bioinformatics2026-01-01

Architectural good practices for reproducible benchmarking in protein machine learning.

Julián García-Vinuesa, Diego Fernández-Villegas, Michelle Soto-García, Xavier Cadet, Mehdi D Davari, Frederic Cadet, Juan A Asenjo, Roberto Uribe-Paredes, David Medina-Ortiz

原始摘要(英文原文)· Original abstract
Reliable benchmarking in protein machine learning requires biological tasks, datasets, representations, execution conditions, and evaluation regimes to be explicitly defined and traceable. Protein-specific dependencies, including label semantics, negative-class construction, evolutionary relationships, similarity control, inferred structures, protein language model configurations, and potential pretraining contamination, can alter benchmark interpretation. This Perspective provides an operational, protein-specific synthesis of established reproducibility practices through an artefact-linked architecture. We define five core requirements for verifiable benchmarking, including deterministic curation, persistent and versioned artefacts, explicit interface contracts, reproducible execution, and transparent evaluation. Prediction provenance, uncertainty, calibration, and explainability are treated as an additional decision-readiness layer. These requirements are translated into minimum records, provenance links, and verification checks. An antimicrobial peptide classification case study demonstrates their application to alternative task formulations, negative-class definitions, dataset compositions, and partitioning strategies within a controlled workflow.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Architectural good practices for reproducible benchmarking in protein machine learning. — 科研速览 Science Skim