CORTEXA
← Browse
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-23Cited by 0

A Scalable Distributed and Fault-Tolerant Architecture for Cloud-Based Machine Learning and Data Analysis

Grace Dooshima GBOR, Emmanuel Ogala, Donald Douglas Atsa’am, Iorshashe Agaji

Abstract The rapid growth of data-intensive applications has necessitated the development of scalable and efficient architectures for cloud-based machine learning and data analysis. This study proposes a scalable, distributed, and fault-tolerant architecture designed to address the challenges of processing large-scale and dynamic datasets in cloud environments. The architecture integrates key components, including data ingestion, distributed storage, parallel processing frameworks, machine learning pipelines, and application deployment layers, enabling seamless data flow and modular system design. It supports both batch and real-time data processing, making it adaptable to diverse analytical workloads. A design science and experimental research methodology was adopted to develop and evaluate the proposed system. Mathematical modeling and performance analysis were employed to assess system scalability, throughput, and latency under varying load conditions. Experimental results demonstrated that the architecture achieves significant improvements in processing efficiency and resource utilization through horizontal scaling. However, the findings also revealed sub-linear scalability behavior due to factors such as communication overhead, synchronization delays, and resource contention, which are inherent in distributed systems. The architecture exhibited strong fault tolerance and resilience, ensuring continuous system operation through redundancy and dynamic resource management. Performance evaluations highlighted an optimal operating region where throughput is maximized and latency remains within acceptable limits, beyond which system performance begins to degrade. The proposed architecture provides a robust and flexible framework for large-scale machine learning and data analysis in cloud environments. It offers a balance between scalability, performance, and reliability, making it suitable for modern data-driven applications. Future research may focus on enhancing auto-scaling strategies, optimizing workload distribution, and incorporating intelligent resource management techniques to further improve system efficiency. Keywords: Cloud-Based Machine Learning, Scalable Architecture, Fault tolerance, Parallel processing, Resource management

View free PDFSource page

Related papers

openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

Serverless Hyperspectral Image Processing: A Cloud-Native Machine Learning Architecture for UAV-Assisted Agricultural and Disaster Remote Sensing

YINKA ADERIBIGBE

The integration of Unmanned Aerial Vehicles equipped with hyperspectral and multispectral sensors has revolutionized remote sensing in agriculture, forestry, and disaster management. However, hyperspectral imaging generates extraordinarily dense, high-dimensional datasets that ov…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

Cloud-Native Accounting Measurement: A Serverless Machine Learning Architecture for Integrating Climate Risk into Real-Time Equity Valuation

YINKA ADERIBIGBE

The measurement of climate risk and its influence on accounting-based equity valuation has become a critical mandate in empirical financial research. Traditional methodologies utilize log-linear valuation models and historical panel data to observe how investors adjust their rela…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Before the Model: Why Datasets and Data Representation Define What Machine Learning Can Learn

Jean Franck Loa Rojas

Machine learning systems do not learn reality directly; they learn from the representations preserved in their datasets. This structured narrative review examines how dataset purpose, coverage, integrity, labeling, independence, reproducibility, governance, and continuity determi…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Cloud-Native Adversarial Machine Learning: A Serverless Cybersecurity Architecture for Neutralizing Prompt Injections in Large Language Models

YINKA ADERIBIGBE

The integration of Large Language Models into enterprise network architectures has introduced severe cybersecurity vulnerabilities, most notably adversarial prompt injection and zero-day data extraction attacks. Traditional network security protocols are fundamentally ill-equippe…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-26

Cloud-Native Water Markets: A Serverless Machine Learning Architecture for Supply-Side Water Trading and Dynamic Catchment Pricing

YINKA ADERIBIGBE

The efficient allocation of freshwater resources in small catchments represents a critical challenge in environmental economics. While theoretical models propose supply-side water trading and group-level "water clubs" to mitigate resource depletion, the empirical testing of these…

View free PDFSource page