Accelerate Literature Icon
Want to do a literature review? Try our new Literature Review workflow

Applied Distributed Systems For Large-Scale Telecom Network Optimization

  • Abstract
  • Literature Map
  • Similar Papers
Abstract
Translate article icon Translate Article Star icon

Given the continued growth of global telecommunications, the engineering of telecom networks has far outgrown the customary centralized planning architecture. This article covers the application of distributed systems to architecture design, artificial intelligence, real-time processing, and fault tolerance, which are the key operational domains of large-scale telecom network optimization. With the unprecedented growth of mobile subscriptions and mobile capacity under intense pressure from data-based applications, scalable, automated and reliable planning platforms have become a critical challenge for operators. The convergence of big-data technologies, distributed machine learning, streaming observability frameworks, and self-healing infrastructure patterns has made possible a new generation of planning platforms that, at scale and speed, can replace human-centric engineering. The principles and applied patterns discussed in this article are representative of a maturing discipline and relevant to the global telecom community.

Similar Papers
  • Book Chapter
  • 10.62311/nesx/92472
Quantum AI, edge AI for IoT, blockchain for privacy-preserving analytics, AGI, and quantum cryptography for finance and communications
  • Oct 1, 2024
  • Murali Krishna Pasupuleti

Abstract: This chapter explores the convergence of cutting-edge technologies—Quantum AI, Edge AI for IoT, Blockchain for Privacy-Preserving Analytics, Artificial General Intelligence (AGI), and Quantum Cryptography—and their transformative potential in finance and communications. Quantum AI accelerates machine learning and decision-making through quantum computing, while Edge AI enables real-time data processing at the network edge for IoT systems. Blockchain ensures secure and decentralized privacy-preserving analytics, and AGI aims to develop AI systems capable of human-like general intelligence, potentially revolutionizing industries. Quantum Cryptography offers ultra-secure communication channels, safeguarding financial transactions and sensitive data. The chapter discusses these technologies' impact, applications, challenges, and ethical considerations, emphasizing the need for responsible development and interdisciplinary collaboration to harness their full potential. Keywords: Quantum AI, Edge AI, IoT, Blockchain, Privacy-Preserving Analytics, AGI, Quantum Cryptography, Finance, Communications, Quantum Computing, Machine Learning, Secure Data, Cryptography, Artificial Intelligence, Real-Time Processing, Data Security.

  • Research Article
  • Cite Count Icon 10
  • 10.1016/j.neucom.2023.127009
Topologies in distributed machine learning: Comprehensive survey, recommendations and future directions
  • Nov 10, 2023
  • Neurocomputing
  • Ling Liu + 6 more

Topologies in distributed machine learning: Comprehensive survey, recommendations and future directions

  • Research Article
  • Cite Count Icon 8
  • 10.1007/s44163-025-00266-0
Ethical challenges and solutions in AI-driven medical data management: a focus on distributed machine learning
  • May 10, 2025
  • Discover Artificial Intelligence
  • Martin Hähnel

This article addresses the complexities and ethical concerns surrounding data collection and artificial intelligence (AI) in general and particularly in medical contexts. It highlights the challenges posed by AI in adhering to the “5 Ps of ethical data handling” (provenance, protection, purpose, preparation, and privacy) and emphasizes the need of supplementation of existing regulations such as the GDPR to fully ensure ethical data practices. This article explores both technical and nontechnical methods for managing sensitive data, with a particular focus on distributed machine learning (DML). DML, while promising for secure and collaborative medical data management, raises unique ethical issues that require thorough examination. The paper underscores the need for a synergy between technological advancements and ethical considerations to uphold values such as patient autonomy, data privacy, and justice in AI applications. It also provides an in-depth analysis of different DML methods (split learning, federated learning, and swarm learning) and their potential applications and drawbacks in healthcare, stressing the importance of developing ethical, secure, and transparent AI systems to prevent misuse and ensure patient trust.

  • Conference Article
  • Cite Count Icon 11
  • 10.1109/weconf48837.2020.9131518
Ubiquitous Computing and Distributed Machine Learning in Smart Cities
  • Jun 1, 2020
  • 2020 Wave Electronics and its Application in Information and Telecommunication Systems (WECONF)
  • D R Mukhametov

The article is devoted to the analysis of the use of ubiquitous computing and distributed machine learning in smart cities. Smart city is characterized by the introduction of high-tech infrastructure, digital services, integrated information monitoring systems that allow to optimize the environment and processes of urban management. The most promising direction of smart cities development is the implementation of ubiquitous computing systems. Ubiquitous computing involves the introduction of a significant number of technologies, including sensors, artificial intelligence, Internet of Things, network robots. Since ubiquitous computing is based on the processing of data generated by different devices, the new solutions are needed to structure and ensure data compatibility. Such solutions are the distributed machine learning methods: stochastic gradient descent and K-means method. The work separately considers the use of federated training, which has advantages in data privacy and mobile computing. The article deals with the main provisions of the concept of smart city, technologies of ubiquitous computing, features of methods of distributed machine learning and their introduction into urban systems management.

  • Conference Article
  • Cite Count Icon 1
  • 10.1109/icdt57929.2023.10151116
Distributed Consensus and Fault Tolerance Mechanisms Using Distributed Machine Learning
  • May 11, 2023
  • P Ramesh Babu + 3 more

In this paper, we develop a distributed consensus model to improve the process of fault tolerance in cloud. The distributed consensus mechanism uses distributed machine learning as its base estimator to predict the fault instances when a task is allocated for possible offloading of user contents in the cloud. The distributed machine learning model senses the number of tasks and available nodes to complete the possible offloading of task in the cloud. The python simulator is conducted to test the efficacy of the distributed consensus mechanism in allocating the offloaded task without faults in the cloud servers. The results of simulation shows an increased prediction accuracy, reduced latency in offloading the task and classifying the fault tolerant VMs or task in process the task update.

  • Research Article
  • 10.54809/galla.2025.003
Distributed Machine Learning Algorithms: A Comprehensive Review
  • Jan 1, 2025
  • Galla Journal
  • Nashma Taha Muhammed + 1 more

Artificial intelligence has expanded significantly over the last decade due to growing user demand and has achieved major advancements in managing complex tasks. Processing and analyzing this large volume of data is time-consuming and requires substantial computational resources. To address these limitations, distributed machine learning (DML) has emerged as an effective solution, enabling parallelization of tasks by distributing data, models, or both across multiple servers. This review paper thoroughly examines various strategies and methodologies used in DML, with a particular emphasis on data parallelism and model parallelism. These methods significantly enhance scalability and computational efficiency, which in turn accelerate AI advancements in sectors such as autonomous driving, healthcare, and recommendation systems. Additionally, this paper provides an extensive overview of key DML algorithms and frameworks, exploring their advantages, practical applications, and limitations. Furthermore, it identifies and examines important challenges—such as security concerns and communication overhead—and offers recommendations for future research to develop DML systems that are more reliable, scalable, and efficient.

  • PDF Download Icon
  • Research Article
  • Cite Count Icon 27
  • 10.1109/access.2020.2971519
Distributed Machine Learning Oriented Data Integrity Verification Scheme in Cloud Computing Environment
  • Jan 1, 2020
  • IEEE Access
  • Xiao-Ping Zhao + 1 more

Distributed Machine Learning (DML) is one of the core technologies for Artificial Intelligence (AI). However, in the existing distributed machine learning framework, the data integrity is not taken into account. If network attackers forge the data, modify the data, or destroy the data, the training model in the distributed machine learning system will be greatly affected, and the training results are led to be wrong. Therefore, it is crucial to guarantee the data integrity in the DML. In this paper, we propose a distributed machine learning oriented data integrity verification scheme (DML-DIV) to ensure the integrity of training data. Firstly, we adopt the idea of Provable Data Possession (PDP) sampling auditing algorithm to achieve data integrity verification so that our DML-DIV scheme can resist forgery attacks and tampering attacks. Secondly, we generate a random number, namely blinding factor, and apply the discrete logarithm problem (DLP) to construct proof and ensure privacy protection in the TPA verification process. Thirdly, we employ identity-based cryptography and two-step key generation technology to generate data owner's public/private key pair so that our DML-DIV scheme can solve the key escrow problem and reduce the cost of managing the certificates. Finally, formal theoretical analysis and experimental results show the security and efficiency of our DML-DIV scheme.

  • Research Article
  • 10.5281/zenodo.5856301
The FORA Fog Computing Platform for Industrial IoT
  • Jul 2, 2020
  • Zenodo (CERN European Organization for Nuclear Research)
  • Paul Pop + 6 more

Industry 4.0 will only become a reality through the convergence of Operational and Information Technologies (OT & IT), which use different computation and communication technologies. Cloud Computing cannot be used for OT involving industrial applications, since it cannot guar-antee stringent non-functional requirements, e.g., dependability, trustworthiness and timeliness. Instead, a new computing paradigm, called Fog Computing, is envisioned as an architectural means to realize the IT/OT convergence. In this paper we propose a Fog Computing Platform (FCP) reference architecture targeting Industrial IoT applications. The FCP is based on: deter-ministic virtualization that reduces the effort required for safety and security assurance; middle-ware for supporting both critical control and dynamic Fog applications; deterministic networking and interoperability, using open standards such as IEEE 802.1 Time-Sensitive Networking (TSN) and OPC Unified Architecture (OPC UA); mechanisms for resource management and or-chestration; and services for security, fault tolerance and distributed machine learning. We pro-pose a methodology for the definition and the evaluation of the reference architecture. We use the Architecture Analysis Design Language (AADL) to model the FCP reference architecture, and a set of industrial use cases to evaluate its suitability for the Industrial IoT area.

  • Research Article
  • Cite Count Icon 81
  • 10.1016/j.is.2021.101727
The FORA Fog Computing Platform for Industrial IoT
  • Feb 2, 2021
  • Information Systems
  • Paul Pop + 6 more

Industry 4.0 will only become a reality through the convergence of Operational and Information Technologies (OT & IT), which use different computation and communication technologies. Cloud Computing cannot be used for OT involving industrial applications, since it cannot guarantee stringent non-functional requirements, e.g., dependability, trustworthiness and timeliness. Instead, a new computing paradigm, called Fog Computing, is envisioned as an architectural means to realize the IT/OT convergence. In this paper we propose a Fog Computing Platform (FCP) reference architecture targeting Industrial IoT applications. The FCP is based on: deterministic virtualization that reduces the effort required for safety and security assurance; middleware for supporting both critical control and dynamic Fog applications; deterministic networking and interoperability, using open standards such as IEEE 802.1 Time-Sensitive Networking (TSN) and OPC Unified Architecture (OPC UA); mechanisms for resource management and orchestration; and services for security, fault tolerance and distributed machine learning. We propose a methodology for the definition and the evaluation of the reference architecture. We use the Architecture Analysis Design Language (AADL) to model the FCP reference architecture, and a set of industrial use cases to evaluate its suitability for the Industrial IoT area.

  • PDF Download Icon
  • Research Article
  • Cite Count Icon 5
  • 10.3390/electronics12183761
Distributed Machine Learning and Native AI Enablers for End-to-End Resources Management in 6G
  • Sep 6, 2023
  • Electronics
  • Orfeas Agis Karachalios + 3 more

6G targets a broad and ambitious range of networking scenarios with stringent and diverse requirements. Such challenging demands require a multitude of computational and communication resources and means for their efficient and coordinated management in an end-to-end fashion across various domains. Conventional approaches cannot handle the complexity, dynamicity, and end-to-end scope of the problem, and solutions based on artificial intelligence (AI) become necessary. However, current applications of AI to resource management (RM) tasks provide partial ad hoc solutions that largely lack compatibility with notions of native AI enablers, as foreseen in 6G, and either have a narrow focus, without regard for an end-to-end scope, or employ non-scalable representations/learning. This survey article contributes a systematic demonstration that the 6G vision promotes the employment of appropriate distributed machine learning (ML) frameworks that interact through native AI enablers in a composable fashion towards a versatile and effective end-to-end RM framework. We start with an account of 6G challenges that yields three criteria for benchmarking the suitability of candidate ML-powered RM methodologies for 6G, also in connection with an end-to-end scope. We then proceed with a focused survey of appropriate methodologies in light of these criteria. All considered methodologies are classified in accordance with six distinct methodological frameworks, and this approach invites broader insight into the potential and limitations of the more general frameworks, beyond individual methodologies. The landscape is complemented by considering important AI enablers, discussing their functionality and interplay, and exploring their potential for supporting each of the six methodological frameworks. The article culminates with lessons learned, open issues, and directions for future research.

  • Research Article
  • Cite Count Icon 1
  • 10.13107/jcorth.2023.v08i02.578
ORTHO AI : The Dawn Of A New Era: Artificial Intelligence In Orthopaedics
  • Jan 1, 2023
  • Journal of Clinical Orthopaedics
  • Parag Sancheti + 4 more

ORTHO AI : The Dawn Of A New Era: Artificial Intelligence In Orthopaedics

  • Research Article
  • Cite Count Icon 7
  • 10.1088/1757-899x/912/2/022029
Comparison of deep learning models for predictive maintenance
  • Aug 1, 2020
  • IOP Conference Series: Materials Science and Engineering
  • R Naren + 1 more

There is a clear intersection between the Internet of Things (IoT) and Artificial Intelligence (AI). IoT is about connecting machines and making use of the data generated from those machines. AI is about simulating intelligent behaviour in machines of all kinds. As IoT devices will generate vast amounts of data, then AI will be functionally necessary to deal with these huge volumes if we’re to have any chance of making sense of the data. AI is beneficial for both real-time and post event processing: Post event processing – identifying patterns in data sets and running predictive analytics, e.g. the correlation between traffic congestion, air pollution and chronic respiratory illnesses within a city centre. Real-time processing – responding quickly to conditions and building up knowledge of decisions about those events, e.g. remote video camera reading license plates for parking payments.

  • Conference Article
  • Cite Count Icon 61
  • 10.1109/icdcs.2019.00159
Applying Differential Privacy Mechanism in Artificial Intelligence
  • Jul 1, 2019
  • Tianqing Zhu + 1 more

Artificial Intelligence (AI) has attracted a large amount of attention in recent years. However, several new problems, such as privacy violations, security issues, or effectiveness, have been emerging. Differential privacy has several attractive properties that make it quite valuable for AI, such as privacy preservation, security, randomization, composition, and stability. Therefore, this paper presents differential privacy mechanisms for multi-agent systems, reinforcement learning, and knowledge transfer based on those properties, which proves that current AI can benefit from differential privacy mechanisms. In addition, the previous usage of differential privacy mechanisms in private machine learning, distributed machine learning, and fairness in models is discussed, bringing several possible avenues to use differential privacy mechanisms in AI. The purpose of this paper is to deliver the initial idea of how to integrate AI with differential privacy mechanisms and to explore more possibilities to improve AIs performance.

  • Conference Article
  • Cite Count Icon 21
  • 10.1145/3041021.3051099
Distributed Machine Learning
  • Jan 1, 2017
  • Tie-Yan Liu + 2 more

In recent years, artificial intelligence has achieved great success in many important applications. Both novel machine learning algorithms (e.g., deep neural networks), and their distributed implementations play very critical roles in the success. In this tutorial, we will first review popular machine learning algorithms and the optimization techniques they use. Second, we will introduce widely used ways of parallelizing machine learning algorithms (including both data parallelism and model parallelism, both synchronous and asynchronous parallelization), and discuss their theoretical properties, strengths, and weakness. Third, we will present some recent works that try to improve standard parallelization mechanisms. Last, we will provide some practical examples of parallelizing given machine learning algorithms in online application (e.g. Recommendation and Ranking) by using popular distributed platforms, such as Spark MlLib, DMTK, and Tensorflow. By listening to this tutorial, the audience can form a clear knowledge framework about distributed machine learning, and gain some hands-on experiences on parallelizing a given machine learning algorithm using popular distributed systems.

  • Research Article
  • Cite Count Icon 1
  • 10.54254/2755-2721/2025.tj23594
Communication-Efficient Distributed Machine Learning: Techniques and Innovations
  • Jun 9, 2025
  • Applied and Computational Engineering
  • Juntao Wang

As Artificial Intelligence (AI) technologies continue to advance, the size and complexity of machine learning models are rapidly increasing. Distributed Machine Learning (DML) has been proposed to improve the limitations of centralized training in terms of computational power and memory. However, communication overhead remains a significant obstacle in DML, restricting training efficiency. This paper proposes a hybrid approach combining Adaptive Gradient Compression (AGC) and Locally Updated Stochastic Gradient Descent (LU-SGD) to maintain model performance while reducing communication overhead. Specifically, the communication load is first reduced by compressing the gradient during each transmission round via AGC. Second, this study uses LU-SGD to minimize the number of communication phases by executing several local updates before synchronizing the gradients. Extensive experiments are conducted on Modified National Institute of Standards and Technology (MNIST), Canadian Institute for Advanced Research (CIFAR)-10, and ImageNet datasets with LeNet, Residual Neural Network (ResNet), and Visual Geometry Group (VGG). Experimental results show the hybrid approach reduces communication data while maintaining efficiency and model accuracy. This approach effectively optimizes communication and demonstrates its potential to improve distributed machine learning frameworks' expansion capability and efficiency.

Save Icon
Up Arrow
Open/Close
Notes

Save Important notes in documents

Highlight text to save as a note, or write notes directly

You can also access these Documents in Paperpal, our AI writing tool

Powered by our AI Writing Assistant