https://www.gphjournal.org/index.php/cse/issue/feed GPH-International Journal of Computer Science and Engineering 2026-07-23T11:21:34+00:00 Sahrukh Khan info@gphjournal.org Open Journal Systems <p><strong>GPH-International Journal of Computer Science and Engineering (<a title="Online ISSN Number" href="https://portal.issn.org/resource/ISSN/3050-9610" target="_blank" rel="noopener">e-ISSN 3050-9661</a>)</strong> is a peer-reviewed, open-access journal dedicated to advancing cutting-edge research in computer science, engineering, and emerging digital technologies. The journal publishes original research papers, technical studies, and comprehensive reviews in areas such as artificial intelligence, machine learning, data science, algorithms, software engineering, computer networks, cybersecurity, cloud computing, hardware design, and intelligent systems.&nbsp;With a focus on innovation and interdisciplinary collaboration, the journal provides a global platform for academics, researchers, and professionals to share new ideas, explore real-world applications, and contribute to the evolving landscape of computing and AI-driven technologies.</p> https://www.gphjournal.org/index.php/cse/article/view/2518 A Scalable Distributed and Fault-Tolerant Architecture for Cloud-Based Machine Learning and Data Analysis 2026-07-23T11:21:34+00:00 GBOR, Grace Dooshima noreplygphjournals@gmail.com Emmanuel Ogala noreplygphjournals@gmail.com Donald Douglas Atsa'am noreplygphjournals@gmail.com Iorshashe Agaji noreplygphjournals@gmail.com <p>The rapid growth of data-intensive applications has necessitated the development of scalable and efficient architectures for cloud-based machine learning and data analysis. This study proposes a <strong>scalable, distributed, and fault-tolerant architecture</strong> designed to address the challenges of processing large-scale and dynamic datasets in cloud environments. The architecture integrates key components, including data ingestion, distributed storage, parallel processing frameworks, machine learning pipelines, and application deployment layers, enabling seamless data flow and modular system design. It supports both batch and real-time data processing, making it adaptable to diverse analytical workloads. A <strong>design science and experimental research methodology</strong> was adopted to develop and evaluate the proposed system. Mathematical modeling and performance analysis were employed to assess system scalability, throughput, and latency under varying load conditions. Experimental results demonstrated that the architecture achieves significant improvements in processing efficiency and resource utilization through horizontal scaling. However, the findings also revealed sub-linear scalability behavior due to factors such as communication overhead, synchronization delays, and resource contention, which are inherent in distributed systems. The architecture exhibited strong <strong>fault tolerance and resilience</strong><strong>,</strong> ensuring continuous system operation through redundancy and dynamic resource management. Performance evaluations highlighted an optimal operating region where throughput is maximized and latency remains within acceptable limits, beyond which system performance begins to degrade. The proposed architecture provides a robust and flexible framework for large-scale machine learning and data analysis in cloud environments. It offers a balance between scalability, performance, and reliability, making it suitable for modern data-driven applications. Future research may focus on enhancing auto-scaling strategies, optimizing workload distribution, and incorporating intelligent resource management techniques to further improve system efficiency.</p> 2026-07-23T11:21:34+00:00 ##submission.copyrightStatement##