科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Energies2025-11-12· SCADA

A Cluster-Based Filtering Approach to SCADA Data Preprocessing for Wind Turbine Condition Monitoring and Fault Detection

Krzysztof Kijanowski, Tomasz Barszcz, Phong B. Dao

原始摘要(英文原文)· Original abstract
The high cost of wind turbine maintenance has intensified the need for reliable fault detection and condition monitoring methods. While Supervisory Control and Data Acquisition (SCADA) systems provide valuable operational data, the raw signals often contain noise, outliers, and missing or redundant entries, which can compromise analysis accuracy. This study presents a novel cluster-based outlier removal approach for SCADA data preprocessing, featuring a unique flexibility to include or exclude negative power values—a factor rarely investigated but potentially critical for fault detection performance. The method applies the K-Means++ unsupervised clustering algorithm to group data points along the wind speed–power curve. The number of clusters is determined heuristically using the elbow method, while outliers are identified through Mahalanobis distance with thresholds derived from Chebyshev’s inequality theorem. The approach was validated using SCADA data from a wind farm in Portugal and further assessed with a CUSUM test-based structural change detection method to study how preprocessing choices—outlier thresholds (5% vs. 1%) and inclusion/exclusion of negative power values—affect early fault identification. Results demonstrate reliable fault detection up to 14 days before failure, retaining over 99% of the original dataset. This work provides key insights into preprocessing impacts on model reliability and offers an open-source Python implementation for reproducibility.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

A Cluster-Based Filtering Approach to SCADA Data Preprocessing for Wind Turbine Condition Monitoring and Fault Detection — 科研速览 Science Skim