paper-with-me

홈 › Papers

Predictive Auto-scaling with OpenStack Monasca

2021-11-03 · Giacomo Lanciano, Filippo Galli, Tommaso Cucinotta, Davide Bacciu, Andrea Passarella

Cloud auto-scaling mechanisms are typically based on reactive automation rules that scale a cluster whenever some metric, e.g., the average CPU usage among instances, exceeds a predefined threshold. Tuning these rules becomes particularly cumbersome when scaling-up a cluster involves non-negligible times to bootstrap new instances, as it happens frequently in production cloud services. To deal with this problem, we propose an architecture for auto-scaling cloud services based on the status in which the system is expected to evolve in the near future. Our approach leverages on time-series forecasting techniques, like those based on machine learning and artificial neural networks, to predict the future dynamics of key metrics, e.g., resource consumption metrics, and apply a threshold-based scaling policy on them. The result is a predictive automation policy that is able, for instance, to automatically anticipate peaks in the load of a cloud application and trigger ahead of time appropriate scaling actions to accommodate the expected increase in traffic. We prototyped our approach as an open-source OpenStack component, which relies on, and extends, the monitoring capabilities offered by Monasca, resulting in the addition of predictive metrics that can be leveraged by orchestration components like Heat or Senlin. We show experimental results using a recurrent neural network and a multi-layer perceptron as predictor, which are compared with a simple linear regression and a traditional non-predictive auto-scaling policy. However, the proposed framework allows for the easy customization of the prediction policy as needed.

📄 PDF Abstract BibTeX arXiv:2111.02133

Code (1)

giacomolanciano/UCC2021-predictive-auto-scaling-openstack 공식 구현 pytorch

Tasks

CPUTime Series AnalysisTime Series Forecasting

Methods 이 논문이 사용한 방법론

Linear Regression Linear Regression is a method for modelling a relationship between a dependent variable and independent variables. These models can be fit with numerous approaches. The most…

Similar Papers 제목 키워드 기반

Isoflat: Flat Provider Network Multiplexing and Firewalling in OpenStack Cloud

2019-07-15 · IEEE International Conference on Communications (ICC) 2019 7 · Ruipeng Zhang, Mengjun Xie, Li Yang

Networking is one of the key enablers of cloud computing and its security is essential for multi-tenant clouds. As a widely used open source solution to cloud computing, OpenStack allows computing resources to connect to…

Cloud Computing

A Comparison of Reinforcement Learning Techniques for Fuzzy Cloud Auto-Scaling

2017-05-19 · Hamid Arabnejad, Claus Pahl, Pooyan Jamshidi, Giovani Estrada

A goal of cloud service management is to design self-adaptable auto-scaler to react to workload fluctuations and changing the resources assigned. The key problem is how and when to add/remove resources in order to meet a…

ManagementQ-Learningreinforcement-learningReinforcement Learning+1

An ML-based Approach to Predicting Software Change Dependencies: Insights from an Empirical Study on OpenStack

2025-08-07 · Ali Arabat, Mohammed Sayagh, Jameleddine Hassine arxiv

As software systems grow in complexity, accurately identifying and managing dependencies among changes becomes increasingly critical. For instance, a change that leverages a function must depend on the change that introd…

Online Ensemble Transformer for Accurate Cloud Workload Forecasting in Predictive Auto-Scaling

2025-08-18 · Jiadong Chen, Xiao He, Hengyu Ye, Fuxin Jiang 외 arxiv

In the swiftly evolving domain of cloud computing, the advent of serverless systems underscores the crucial need for predictive auto-scaling systems. This necessity arises to ensure optimal resource allocation and mainta…

A Meta Reinforcement Learning Approach for Predictive Autoscaling in the Cloud

2022-05-31 · Siqiao Xue, Chao Qu, Xiaoming Shi, Cong Liao 외

Predictive autoscaling (autoscaling with workload forecasting) is an important mechanism that supports autonomous adjustment of computing resources in accordance with fluctuating workload demands in the Cloud. In recent …

CPUDecision MakingManagementMeta Reinforcement Learning+2