科研速览 · Science Skim继续刷下去 · Keep skimming →
◇ bioRxiv2026-08-12· bioinformatics

FrustrAI-Seq: Scaling Local Energetic Frustration to the Protein Sequence Space

J.-P. Leusch, M. Poley-Gil, M. Fernandez-Martin, J. Schlensok, F. L. Simonetti, N. Bordin, B. Rost, R. G. Parra, M. Heinzinger

原始摘要(英文原文)· Original abstract
Proteins fold into their native three-dimensional (3D) structures by navigating complex energy landscapes shaped by the biophysical and biochemical properties of their sequence. Once folded, some sequence positions (dubbed residues) remain locally frustrated, reflecting functional constraints incompatible with optimal packing. This local energetic frustration provides important insights into protein function and dynamics, but its analysis typically relies on structure-based energy calculations and remains energetically costly at scale. Here, we introduce an ultra-fast sequence-based prediction of local energetic frustration directly from protein sequences using embeddings from protein language models (pLMs). Our method, coined FrustrAI-Seq, enables proteome-wide frustration profiling in minutes (17 minutes for the entire human proteome on a single Nvidia H100 GPU) while retaining biologically relevant performance as shown for the alpha-globin and beta-lactamase family. By eliminating the need for explicit structural or evolutionary information, this approach expands frustration analysis to protein regions and classes that were previously inaccessible, including intrinsically disordered regions and high-throughput de novo designed protein datasets. To support reproducibility and large-scale applications, we provide the largest freely available resource of precomputed local frustration scores to date (10^6 proteins), along with model weights and complete training and inference code at: github.com/leuschjanphilipp/FrustrAI-Seq.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

FrustrAI-Seq: Scaling Local Energetic Frustration to the Protein Sequence Space — 科研速览 Science Skim