A Theory of Auto-Scaling for Resource Reservation in Cloud Services-Reference-Cited by-同舟云学术

A Theory of Auto-Scaling for Resource Reservation in Cloud Services

Published:2021-03-05 Issue:3 Volume:48 Page:27-32
ISSN:0163-5999
Container-title:ACM SIGMETRICS Performance Evaluation Review
language:en
Short-container-title:SIGMETRICS Perform. Eval. Rev.

Author:

Psychas Konstantinos¹,Ghaderi Javad¹

Affiliation:

1. Columbia University

Abstract

We consider a distributed server system consisting of a large number of servers, each with limited capacity on multiple resources (CPU, memory, disk, etc.). Jobs with different rewards arrive over time and require certain amounts of resources for the duration of their service. When a job arrives, the system must decide whether to admit it or reject it, and if admitted, in which server to schedule the job. The objective is to maximize the expected total reward received by the system. This problem is motivated by control of cloud computing clusters, in which, jobs are requests for Virtual Machines or Containers that reserve resources for various services, and rewards represent service priority of requests or price paid per time unit of service by clients. We study this problem in an asymptotic regime where the number of servers and jobs' arrival rates scale by a factor L, as L becomes large. We propose a resource reservation policy that asymptotically achieves at least 1/2, and under certain monotone property on jobs' rewards and resources, at least 11/4 of the optimal expected reward. The policy automatically scales the number of VM slots for each job type as the demand changes, and decides in which servers the slots should be created in advance, without the knowledge of traffic rates. It effectively tracks a low-complexity greedy packing of existing jobs in the system while maintaining only a small number, g(L) = w(logL), of reserved VM slots for high priority jobs that pack well.

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Networks and Communications,Hardware and Architecture,Software

Link

https://dl.acm.org/doi/pdf/10.1145/3453953.3453958

Reference34 articles.

1. AWS container 2019. Amazon AWS Containers. https://aws.amazon.com/ containers/ AWS container 2019. Amazon AWS Containers. https://aws.amazon.com/ containers/

2. Asymptotic analysis of single resource loss systems in heavy traffic, with applications to integrated networks

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. ReSQoV: A Scalable Resource Allocation Model for QoS-Satisfied Cloud Services;Future Internet;2022-04-26