Umeå universitets logga

umu.sePublikationer
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Accurate and low-overhead workload prediction for cloud management
Umeå universitet, Teknisk-naturvetenskapliga fakulteten, Institutionen för datavetenskap. (Autonomous Distributed Systems Lab)ORCID-id: 0000-0002-8097-1143
2025 (Engelska)Doktorsavhandling, sammanläggning (Övrigt vetenskapligt)Alternativ titel
Noggrann och effektiv prediktering av last för resurshantering i datormoln (Svenska)
Abstract [en]

Cloud computing has transformed the IT landscape by offering users and orga-nizations on-demand access to computing power, storage, data processing, andmachine learning resources. Despite the benefits, cloud resource managementfaces challenges due to the heterogeneous and dynamic nature of workloads.Inefficient provisioning manifests in two critical forms: underprovisioning leadsto degraded Quality of Service (QoS) and unmet Service-Level Agreements(SLAs), while overprovisioning results in unnecessary energy consumption andhigh operational costs. With the current rise of AI and machine learning in-novations, machine learning-based workload prediction for resource provisionplays a vital role in predicting future scenarios and identifying new occurrences,enabling service providers to prepare ahead of time. However, various challengesare associated with machine learning-based workload prediction.This thesis addresses the challenges of machine learning-based workloadprediction in cloud environments, including data drift due to dynamic workloads,high computational overhead, and storage overhead. Firstly, cloud workloads aredynamic, and models trained with old historical data can become obsolete overtime. We addressed the challenge of accurate prediction and data drift by incor-porating machine learning and streaming data processing algorithms to assistadaptive prediction. Secondly, constantly training and updating deep learningmodels adds significant computational overhead to the cloud infrastructure. Weaddressed this problem by proposing a solution that incorporates a knowledgebase repository with transfer learning-based adaptation. Moreover, we exploredthe tradeoff between model accuracy and computational overhead. Finally, wepropose a data compression mechanism that leverages an autoencoder to reducestorage overhead resulting from the continuous generation of monitoring datain cloud management systems.Our findings reveal that the proposed methods have significantly improvedthe machine learning-based cloud management system. Extensive evaluationusing real-world datasets reveals that the proposed methods facilitate thecreation of accurate predictions, even in the face of ever-changing patterns incloud workloads. Moreover, the methods reduced computation overhead byleveraging existing knowledge and highlighting the tradeoff required to achievea balance between prediction accuracy and computation overhead.

Ort, förlag, år, upplaga, sidor
Umeå: Umeå University, 2025. , s. 38
Serie
Report / UMINF, ISSN 0348-0542 ; 25.09
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
URN: urn:nbn:se:umu:diva-238533ISBN: 978-91-8070-713-8 (digital)ISBN: 978-91-8070-712-1 (tryckt)OAI: oai:DiVA.org:umu-238533DiVA, id: diva2:1956902
Disputation
2025-05-30, MIT.A.121, MIT-huset, Umeå,, 13:15 (Engelska)
Opponent
Handledare
Tillgänglig från: 2025-05-09 Skapad: 2025-05-07 Senast uppdaterad: 2025-05-08Bibliografiskt granskad
Delarbeten
1. When and How to Retrain Machine Learning-based Cloud Management Systems
Öppna denna publikation i ny flik eller fönster >>When and How to Retrain Machine Learning-based Cloud Management Systems
2022 (Engelska)Ingår i: 2022 IEEE International Parallel and Distributed Processing Symposium Workshops (IPDPSW), IEEE, 2022, s. 688-698Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

Cloud management systems increasingly rely on machine learning (ML) models to predict incoming workload rates, load, and other system behaviors for efficient dynamic resource management. Current state-of-the-art prediction models demonstrate high accuracy, but assume that data patterns remain stable. However, in production use, systems may face hardware upgrades, changes in user behavior etc. that lead to concept drifts - significant changes in characteristics of data streams over time. To mitigate prediction deterioration, ML models need to be updated - but questions of when and how to best retrain these models are unsolved in the context of cloud management. We present a pilot study that address these questions for one of the most common models for adaptive prediction - Long Short Term Memory (LSTM) - using synthetic and real-world workload data. Our analysis of when to retrain explores approaches for detecting when retraining is required using both concept drift detection and prediction error thresholds, and at what point of retraining should actually take place. Our analysis of how to retrain focuses on the data required for retraining, and what proportion should be taken from before and after the need for retraining is detected. We present initial results that indicate that retraining of existing models can achieve prediction accuracy close to that of newly trained models but for much less cost, and present initial advice for how to provide cloud management systems with support for automatic retraining of ML-based methods.

Ort, förlag, år, upplaga, sidor
IEEE, 2022
Nyckelord
cloud computing, cloud workload prediction, concept drift, machine learning, time series prediction
Nationell ämneskategori
Datorsystem
Identifikatorer
urn:nbn:se:umu:diva-198541 (URN)10.1109/IPDPSW55747.2022.00120 (DOI)000855041000086 ()2-s2.0-85136190866 (Scopus ID)9781665497473 (ISBN)9781665497480 (ISBN)
Konferens
2022 IEEE International Parallel and Distributed Processing Symposium, 30 May 2022-03 June 2022, Lyon, France
Forskningsfinansiär
Knut och Alice Wallenbergs StiftelseeSSENCE - An eScience Collaboration
Tillgänglig från: 2022-08-09 Skapad: 2022-08-09 Senast uppdaterad: 2025-05-07Bibliografiskt granskad
2. Efficient retraining of machine learning algorithms in cloud management systems
Öppna denna publikation i ny flik eller fönster >>Efficient retraining of machine learning algorithms in cloud management systems
(Engelska)Manuskript (preprint) (Övrigt vetenskapligt)
Nationell ämneskategori
Datorsystem
Identifikatorer
urn:nbn:se:umu:diva-238555 (URN)
Tillgänglig från: 2025-05-08 Skapad: 2025-05-08 Senast uppdaterad: 2025-05-08Bibliografiskt granskad
3. Automated hyperparameter tuning for adaptive cloud workload prediction
Öppna denna publikation i ny flik eller fönster >>Automated hyperparameter tuning for adaptive cloud workload prediction
2023 (Engelska)Ingår i: UCC '23: Proceedings of the IEEE/ACM 16th International Conference on Utility and Cloud Computing, New York: Association for Computing Machinery (ACM), 2023Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

Efficient workload prediction is essential for enabling timely resource provisioning in cloud computing environments. However, achieving accurate predictions, ensuring adaptability to changing conditions, and minimizing computation overhead pose significant challenges for workload prediction models. Furthermore, the continuous streaming nature of workload metrics requires careful consideration when applying machine learning and data mining algorithms, as manual hyperparameter optimization can be time-consuming and suboptimal. We propose an automated parameter tuning and adaptation approach for workload prediction models and concept drift detection algorithms utilized in predicting future workload. Our method leverages a pre-built knowledge-base based on historical data statistical features, enabling automatic adjustment of model weights and concept drift detection parameters. Additionally, model adaptation is facilitated through a transfer learning approach. We evaluate the effectiveness of our automated approach by comparing it with static approaches using synthetic and real-world datasets. By automating the parameter tuning process and integrating concept drift detection, in our experiments the proposed method enhances the accuracy and efficiency of workload prediction models by 50%.

Ort, förlag, år, upplaga, sidor
New York: Association for Computing Machinery (ACM), 2023
Nyckelord
Cloud computing, Hyperparameter optimization, Workload prediction, Concept drift, Data mining
Nationell ämneskategori
Datorsystem
Identifikatorer
urn:nbn:se:umu:diva-223451 (URN)10.1145/3603166.3632244 (DOI)001211822800044 ()2-s2.0-85191659681 (Scopus ID)979-8-4007-0234-1 (ISBN)
Konferens
CC '23: IEEE/ACM 16th International Conference on Utility and Cloud Computing, Taormina (Messina), Italy, December 4-7, 2023
Forskningsfinansiär
Knut och Alice Wallenbergs Stiftelse, 2019.0352eSSENCE - An eScience Collaboration
Tillgänglig från: 2024-04-16 Skapad: 2024-04-16 Senast uppdaterad: 2025-05-07Bibliografiskt granskad
4. A hybrid autoencoder-LSTM framework for efficient workload prediction
Öppna denna publikation i ny flik eller fönster >>A hybrid autoencoder-LSTM framework for efficient workload prediction
(Engelska)Manuskript (preprint) (Övrigt vetenskapligt)
Nationell ämneskategori
Data- och informationsvetenskap
Identifikatorer
urn:nbn:se:umu:diva-238556 (URN)
Tillgänglig från: 2025-05-08 Skapad: 2025-05-08 Senast uppdaterad: 2025-05-08Bibliografiskt granskad
5. A data-driven framework for efficient and automated workload prediction in cloud computing
Öppna denna publikation i ny flik eller fönster >>A data-driven framework for efficient and automated workload prediction in cloud computing
(Engelska)Manuskript (preprint) (Övrigt vetenskapligt)
Nationell ämneskategori
Data- och informationsvetenskap
Identifikatorer
urn:nbn:se:umu:diva-238557 (URN)
Tillgänglig från: 2025-05-08 Skapad: 2025-05-08 Senast uppdaterad: 2025-05-08Bibliografiskt granskad

Open Access i DiVA

fulltext(2642 kB)358 nedladdningar
Filinformation
Filnamn FULLTEXT01.pdfFilstorlek 2642 kBChecksumma SHA-512
89981e8efb575d76cc57614a4e3e619c5f489263bc003e514297824bfd1d855ad8f9a96a402775fe80ea6b41b9ef85d22ec333a702213ad13cfb7e928c17f121
Typ fulltextMimetyp application/pdf
spikblad(216 kB)146 nedladdningar
Filinformation
Filnamn FULLTEXT02.pdfFilstorlek 216 kBChecksumma SHA-512
8a354a743aa52c9c0d5632ec5f1973a88d58764775d06a4b3af643b7ac943e9843eeb35a94ba8fc8462e0f085036ed0a76816bd238ac8ea6f146d412576a0cfc
Typ fulltextMimetyp application/pdf

Person

Kidane, Lidia

Sök vidare i DiVA

Av författaren/redaktören
Kidane, Lidia
Av organisationen
Institutionen för datavetenskap
Datavetenskap (datalogi)

Sök vidare utanför DiVA

GoogleGoogle Scholar
Totalt: 511 nedladdningar
Antalet nedladdningar är summan av nedladdningar för alla fulltexter. Det kan inkludera t.ex tidigare versioner som nu inte längre är tillgängliga.

isbn
urn-nbn

Altmetricpoäng

isbn
urn-nbn
Totalt: 6092 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf