科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ HighTech and Innovation Journal2025-11-04· Computer science

Mathematical Approaches and Algorithms in Big Data Architecture and Hybrid System Efficiency

Serik Aliaskarov, Raissa Uskenbayeva, V.V. Serbin, Orazmuhamed Bekmurat, Umit Bazarbayeva, Yelena Bakhtiyarova, Kanibek Sansyzbay

原始摘要(英文原文)· Original abstract
This article presents a formal demonstration of a hybrid big data processing architecture that combines the fault tolerance and storage robustness of Hadoop with the speed and in-memory processing capabilities of Apache Spark. The proposed architecture is evaluated through test execution and performance benchmarking in real-world data centers across three regions in Kazakhstan. The model integrates distributed resource management components, Directed Acyclic Graph (DAG)-based scheduling mechanism, and Resilient Distributed Datasets (RDDs) to enable dynamic workload distribution and rapid failure recovery. The results demonstrate that the hybrid system consistently outperforms standalone Spark and Hadoop architectures under variable workloads, illustrating enhancements in execution time, task recovery, and resource utilization. Quantitative performance metrics allow for a structured comparison of architectures and help optimize deployments for diverse scenarios. The proposed hybrid architecture shows significant improvements, reducing average execution time by up to 38% and increasing resource efficiency by 25% compared to standalone Spark and Hadoop systems.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Mathematical Approaches and Algorithms in Big Data Architecture and Hybrid System Efficiency — 科研速览 Science Skim