Skip to main navigation Skip to search Skip to main content

Graceful performance degradation in Apache Storm

  • Mohammad Reza HoseinyFarahabady*
  • , Javid Taheri
  • , Albert Y. Zomaya
  • , Zahir Tari
  • *Corresponding author for this work

Research output: Chapter in Book/Report/Conference proceedingConference contribution

Abstract

The concept of stream data processing is becoming challenging in most business sectors where try to improve their operational efficiency by deriving valuable information from unstructured, yet, contentiously generated high volume raw data in an expected time spans. A modern streamlined data processing platform is required to execute analytical pipelines over a continues flow of data-items that might arrive in a high rate. In most cases, the platform is also expected to dynamically adapt to dynamic characteristics of the incoming traffic rates and the ever-changing condition of underlying computational resources while fulfill the tight latency constraints imposed by the end-users. Apache Storm has emerged as an important open source technology for performing stream processing with very tight latency constraints over a cluster of computing nodes. To increase the overall resource utilization, however, the service provider might be tempted to use a consolidation strategy to pack as many applications as possible in a (cloud-centric) cluster with limited number of working nodes. However, collocated applications can negatively compete with each other, for obtaining the resource capacity in a shared platform that, in turn, the result may lead to a severe performance degradation among all running applications. The main objective of this work is to develop an elastic solution in a modern stream processing ecosystem, for addressing the shared resource contention problem among collocated applications. We propose a mechanism, based on design principles of Model Predictive Control theory, for coping with the extreme conditions in which the collocated analytical applications have different quality of service (QoS) levels while the shared-resource interference is considered as a key performance limiting parameter. Experimental results confirm that the proposed controller can successfully enhance the p -99 latency of high priority applications by 67%, compared to the default round robin resource allocation strategy in Storm, during the high traffic load, while maintaining the requested quality of service levels.

Original languageEnglish
Title of host publicationParallel and distributed computing, applications and technologies: 21st International Conference, PDCAT 2020, proceedings
EditorsYong Zhang, Yicheng Xu, Hui Tian
PublisherSpringer Cham
Pages389-400
Number of pages12
ISBN (Electronic)9783030692445
ISBN (Print)9783030692438
DOIs
Publication statusPublished - 21 Feb 2021
Externally publishedYes
Event21st International Conference on Parallel and Distributed Computing, Applications, and Technologies 2020 - Shenzhen, China
Duration: 28 Dec 202030 Dec 2020

Publication series

NameLecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
Volume12606 LNCS
ISSN (Print)0302-9743
ISSN (Electronic)1611-3349

Conference

Conference21st International Conference on Parallel and Distributed Computing, Applications, and Technologies 2020
Abbreviated titlePDCAT 2020
Country/TerritoryChina
CityShenzhen
Period28/12/202030/12/2020

Keywords

  • Apache storm streaming processing platform
  • Elastic resource controller
  • Performance modeling of computer system
  • Quality of Services (QoS)

ASJC Scopus subject areas

  • Theoretical Computer Science
  • General Computer Science

Fingerprint

Dive into the research topics of 'Graceful performance degradation in Apache Storm'. Together they form a unique fingerprint.

Cite this