Network management and monitoring for cloud systems
George Dan Suciu, Simona Viorica Halunga, Adelina Ochian, Victor Suciu · 2014
Monitoring represents an important factor in improving the quality of the services provided in cloud computing, given the fact that it allows scaling resource utilization in an adaptive manner. It is widely used for detecting critical events and abnormalities of distributed systems and also it helps identifying the faults within the system, discovering application patterns for the users. As cloud systems increase their architecture, the degree of workload also grows in datacenters, causing node failures and performance issues. This paper aims to provide a solution for the monitoring of cloud computing systems and services, allowing users and also providers to optimize the usage of the computational resources according to the constantly changing business requirements inside an organization. The main contribution of the paper consists of the integration of the monitoring system, which is based on Nagios and NConf with a test cloud architecture. Finally, the paper discusses the main findings for a reference implementation using the OpenStack cloud platform.