Technical predictive analysis of big data for Apache Spark-based resource planning
Article Sidebar
Issue Vol. 16 No. 3 (2026)
-
Deep learning approach for automated skin cancer detection with comparative analysis of different batch sizes in dermatological image classification
Abini M.A.5-13
-
Modelling and management principles of combined granulated feed production through information-analytical systems
Mahil Mammadov, Tural Mammadli14-22
-
An intelligent IoT-based automatic control system for energy-efficient grain drying
Ainur Rustemova, Maria Yukhymchuk, Vladyslav Lesko, Marat Orynbet, Yurii Ivanov23-30
-
Design and experimental validation of an IoT-RPA-based smart refrigerator inventory system
Lyudmila Samchuk, Yuliia Povstiana, Nataliia Lishchyna, Mykyta Ponomarenko31-38
-
Intelligent soil monitoring system for sustainable agriculture using TinyML and ultra-low power wireless sensor networks
Muhammad Ovais Akhter39-48
-
Modelling the operation of a dynamic queue management system in supermarkets
Larysa Gumeniuk, Volodymyr Lotysh, Pavlo Humeniuk49-58
-
Experimental validation of error compensation techniques for fibre-optic gyroscopes in semi-natural conditions
Nurzhigit Smailov, Ainur Kuttybayeva, Yerlan Tashtay, Zhandos Dosbayev, Aruzhan Nazarova, Beibarys Sekenov, Aida Serbayeva, Akezhan Sabibolda59-63
-
Simulation of the operation of a four-channel ultrasonic flowmeter with identical signal trajectories in acoustic channels
Yosyp Bilynsky, Andrii Stetsenko64-68
-
Etching line with integrated electrochemical regeneration for copper recovery and sludge reduction
Anatoliy Nester, Vadim Romanuke, Tetiana Yakovyshyna69-76
-
Method for HCI users identification based on their "polyfactor portraits" of perception subjectivization of HCI object
Andrii Pukach, Vasyl Teslyuk, Artem Kazarian77-87
-
Enhancing diagnostic systems for analysing and controlling operating modes in electrical distribution networks
Igor Khomenko, Ruslan Lozhkin, Andrii Shkrebela, Oleksandr Miroshnyk, Anatolii Sereda, Taras Shchur, Mariana Bohach88-94
-
A modular monolith architecture for high-availability electric vehicle charging station management software
Volodymyr Horshkov, Nataliia Lishchyna, Andrii Yashchuk, Olena Surynovych, Valerii Lishchyna95-102
-
Investigation of a multilevel inverter based on IGBT transistors for power supply of a tethered UAV with telemetry-based monitoring of power parameters
Kyrmyzy Taissariyeva, Kuanysh Muslimov, Gulim Jobalayeva, Ingkar Issakozhayeva, Zhansaya Ayapbergen103-107
-
Dynamic obstacle avoidance for UAVs using fuzzy logic control and neural network
Anton Makohonov, Ivan Marynych108-113
-
Hybrid simulation and experimental framework for real-time fault detection in PV boost converters using fuzzy logic and LoRa connectivity
Oussama Sait, Mabrouk Khemliche, Samia Latreche, Sait Belkacem, Hamza Khemliche114-122
-
Benchmarking machine learning algorithms for high-fidelity power forecasting in utility-scale PV plants
Ahmed Saidi, Abdelghani Draoui, Touhami Abdelouahed123-128
-
Study of the dynamic variation in the security level of an information protection system
Olha Saliieva, Yurii Yaremchuk129-133
-
A new hyperchaotic generator: circuit realization and analysis
Volodymyr Rusyn, Petro Kindrachuk, Bogdan Markovych134-138
-
Strategies for ensuring functional stability of a complex sensor network based on the research of the dynamics of the behavior of evolutionary equations
Valentyn Sobchuk, Yurii Kravchenko, Mykhaylo Sharapov, Oleksandr Laptiev, Andrii Sobchuk139-143
-
Development of robust DVB-T2 OFDM transmission: BER assessment and channel impairment mitigation for reliable high-definition broadcasting
Olarewaju Peter Ayeoribe144-149
-
Real-time network anomaly detection in O-RAN using deep learning on streaming big data
Satya Sumanth Vanapalli, Rajesh Polepogu, Parish Venkata Kumar K, Vijayasankar Anumala, Vinodh Babu Panguluri, Lakshmi Narayana Jammula, Brahmaiah Madamanchi, Sravani Duvvu, Bhanusree Nanduri, Syam Sundar Musinala150-158
-
Technical predictive analysis of big data for Apache Spark-based resource planning
Bakhshali Bakhtiyarov, Aynur Jabiyeva, Ulkar Musavi159-166
-
Conversion of voxel models into polygonal meshes with ensuring structural integrity and editability
Semen Duvanov, Iryna Baranova167-172
-
Fuzzy Delphi method for identifying key parameters in the optimization of intelligent control systems for oil refining processes
Kamala Aliyeva173-179
-
Method of determining the grounds for relating data to official information and the degree of restriction access "For official use"
Yurii Dreis, Oleh Harasymchuk180-183
-
Model and tools for content generation and publication based on large language models
Artem Kazarian, Vasyl Teslyuk, Andrii Pukach184-190
-
Evaluation of the usability of web services with interactive maps
Lukasz Sendecki, Sergiusz Skalski, Mariusz Dzienkowski191-197
-
Fuzzy model for assessing digital security literacy across population groups with different social profiles
Volodymyr Polishchuk, Vasyl Sehlianyk, Inna Polishchuk, Andrii Shafar198-203
-
An intelligent module for productivity assessment of remote employees in working time monitoring systems
Aigul Moldakalykova, Maria Yukhymchuk, Gulzhan Kashaganova, Vladyslav Lesko, Yurii Ivanov, Natalia Sachaniuk-Kavets’ka204-209
-
Hybrid machine learning framework for forecasting and evaluating university rankings
Nurzhigit Smailov, Nursultan Kuldeyev, Akezhan Sabibolda, Raigul Ustemirova210-216
Archives
-
Vol. 16 No. 3
2026-09-30 30
-
Vol. 16 No. 2
2026-06-30 27
-
Vol. 16 No. 1
2026-03-30 27
-
Vol. 15 No. 4
2025-12-20 27
-
Vol. 15 No. 3
2025-09-30 24
-
Vol. 15 No. 2
2025-06-27 24
-
Vol. 15 No. 1
2025-03-31 26
-
Vol. 14 No. 4
2024-12-21 25
-
Vol. 14 No. 3
2024-09-30 24
-
Vol. 14 No. 2
2024-06-30 24
-
Vol. 14 No. 1
2024-03-31 23
-
Vol. 13 No. 4
2023-12-20 24
-
Vol. 13 No. 3
2023-09-30 25
-
Vol. 13 No. 2
2023-06-30 14
-
Vol. 13 No. 1
2023-03-31 12
-
Vol. 12 No. 4
2022-12-30 16
-
Vol. 12 No. 3
2022-09-30 15
-
Vol. 12 No. 2
2022-06-30 16
-
Vol. 12 No. 1
2022-03-31 9
Main Article Content
Authors
Abstract
The large database analysis is a substantial form of research that is presently expanding immensely in current technologies, and it also affects numerous aspects of life. Powerful, scalable, and customizable platforms are required to win the challenges in the field and increase effectiveness of analysing the data. Apache Spark is one of the most trending high-performance computing engines to process big data and thus deliver a revolutionary approach to data science and engineering. Big data analytics involve Apache Spark and its application is increasing at lightning speed both in academic and business spheres. It has established itself in the field of data analytics as it can support various kinds of loads within a single architecture. The development of Apache Spark continues with new developments of data analysis, hence bringing it to the notice of researchers and practitioners as a high utility tool to solve the problem of big data. The Apache Spark is a big data processing system which has proven effective in numerous applications of analysing big data; this paper examines the working process of Apache Spark, its advantages and area of application and provides an examination of the future outlook of this platform.
Keywords:
Sustainable Development Goal (SDG)
- Industry, Innovation, Technology and Infrastructure
References
[1] Akram, A. W., & Alamgir, Z. (2022). Distributed fuzzy clustering algorithm for mixed-mode data in Apache SPARK. Journal of Big Data, 9(1), 121. https://doi.org/10.1186/s40537-022-00671-7 DOI: https://doi.org/10.1186/s40537-022-00671-7
[2] Alam, M. A., Nabil, A. R., Mintoo, A. A., & Islam, A. (2024). Real-Time Analytics In Streaming Big Data: Techniques And Applications. Journal of Science and Engineering Research, 1(01), 104–122. https://doi.org/10.70008/jeser.v1i01.56 DOI: https://doi.org/10.70008/jeser.v1i01.56
[3] Bakhtiyarov, B., Jabiyeva, A., Mutallimova, A., Novruzova, R., & Khudaverdiyeva, M. (2025). Adaptive Gaussian-Based Kernel K-Means: Scalable Adaptive Kernel-Based Clustering for Big Data. Journal of Computational and Cognitive Engineering, 5(2), 258–273. https://doi.org/10.47852/bonviewJCCE52026511 DOI: https://doi.org/10.47852/bonviewJCCE52026511
[4] Bao, Z. (2025). Research on an Efficient Data Statistics Upgrade Algorithm Based on the Spark Framework. 2025 IEEE 4th International Conference of Safe Production and Informatization (IICSPI), 619–624. https://doi.org/10.1109/IICSPI66775.2025.11437736 DOI: https://doi.org/10.1109/IICSPI66775.2025.11437736
[5] Belcastro, L., Cantini, R., Marozzo, F., Orsino, A., Talia, D., & Trunfio, P. (2022). Programming big data analysis: Principles and solutions. Journal of Big Data, 9(1), 4. https://doi.org/10.1186/s40537-021-00555-2 DOI: https://doi.org/10.1186/s40537-021-00555-2
[6] Coimbra, M. E., Francisco, A. P., & Veiga, L. (2021). An analysis of the graph processing landscape. Journal of Big Data, 8(1), 55. https://doi.org/10.1186/s40537-021-00443-9 DOI: https://doi.org/10.1186/s40537-021-00443-9
[7] Dubuc, T., Stahl, F., & Roesch, E. B. (2021). Mapping the Big Data Landscape: Technologies, Platforms and Paradigms for Real-Time Analytics of Data Streams. IEEE Access, 9, 15351–15374. https://doi.org/10.1109/ACCESS.2020.3046132 DOI: https://doi.org/10.1109/ACCESS.2020.3046132
[8] Hafsa, M., & Jemili, F. (2018). Comparative Study between Big Data Analysis Techniques in Intrusion Detection. Big Data and Cognitive Computing, 3(1), 1. https://doi.org/10.3390/bdcc3010001 DOI: https://doi.org/10.3390/bdcc3010001
[9] Hasan, Z., Jie Xing, H., & M. Idrees Magray. (2022). Big Data Machine Learning Using Apache Spark Mllib. Mesopotamian Journal of Big Data, 2022, 1–11. https://doi.org/10.58496/MJBD/2022/001 DOI: https://doi.org/10.58496/MJBD/2022/001
[10] Hedayati, S., Maleki, N., Olsson, T., Ahlgren, F., Seyednezhad, M., & Berahmand, K. (2023). MapReduce scheduling algorithms in Hadoop: A systematic study. Journal of Cloud Computing, 12(1), 143. https://doi.org/10.1186/s13677-023-00520-9 DOI: https://doi.org/10.1186/s13677-023-00520-9
[11] Jiang, J., Xiao, P., Yu, L., Li, X., Cheng, J., Miao, X., Zhang, Z., & Cui, B. (2020). PSGraph: How Tencent trains extremely large-scale graphs with Spark? 2020 IEEE 36th International Conference on Data Engineering (ICDE), 1549–1557. https://doi.org/10.1109/ICDE48307.2020.00137 DOI: https://doi.org/10.1109/ICDE48307.2020.00137
[12] Mandelli, A., & Mari, A. (2012). The relationship between social media conversations and reputation during a crisis: The Toyota case. International Journal of Management Cases, 14(1), 456–489. https://doi.org/10.5848/APBJ.2012.00041 DOI: https://doi.org/10.5848/APBJ.2012.00041
[13] Noor, S., Awan, H. H., Hashmi, A. S., Saeed, A., Khan, S., & AlQahtani, S. A. (2025). Optimizing performance of parallel computing platforms for large-scale genome data analysis. Computing, 107(3), 86. https://doi.org/10.1007/s00607-025-01441-y DOI: https://doi.org/10.1007/s00607-025-01441-y
[14] Rajpurohit, A. M., Kumar, P., Kumar, R. R., & Kumar, R. (2024). A Review on Apache Spark. SSRN. https://doi.org/10.2139/ssrn.4492445 DOI: https://doi.org/10.2139/ssrn.4492445
[15] Saha, B., Shah, H., Seth, S., Vijayaraghavan, G., Murthy, A., & Curino, C. (2015). Apache Tez: A Unifying Framework for Modeling and Building Data Processing Applications. Proceedings of the 2015 ACM SIGMOD International Conference on Management of Data, 1357–1369. https://doi.org/10.1145/2723372.2742790 DOI: https://doi.org/10.1145/2723372.2742790
[16] Seidel, M.-D. L. (2022). Open-Source Software. In L. A. Schintler & C. L. McNeely (Eds), Encyclopedia of Big Data (pp. 720–723). Springer International Publishing. https://doi.org/10.1007/978-3-319-32010-6_157 DOI: https://doi.org/10.1007/978-3-319-32010-6_157
[17] Singh, S., Alam, M. N., Kaur, B., Kaur, K., Kaur, S., & Hossain, S. (2025). Comparative analysis of Apache Hadoop and Apache Spark for business intelligence. 020040. https://doi.org/10.1063/5.0246086 DOI: https://doi.org/10.1063/5.0246086
Article Details
Abstract views: 7
Downloads: 10

