Dynamic resource allocation in the cloud with near-optimal efficiency
From MaRDI portal
Abstract: Cloud computing has motivated renewed interest in resource allocation problems with new consumption models. A common goal is to share a resource, such as CPU or I/O bandwidth, among distinct users with different demand patterns as well as different quality of service requirements. To ensure these service requirements, cloud offerings often come with a service level agreement (SLA) between the provider and the users. An SLA specifies the amount of a resource a user is entitled to utilize. In many cloud settings, providers would like to operate resources at high utilization while simultaneously respecting individual SLAs. There is typically a tradeoff between these two objectives; for example, utilization can be increased by shifting away resources from idle users to "scavenger" workload, but with the risk of the former then becoming active again. We study this fundamental tradeoff by formulating a resource allocation model that captures basic properties of cloud computing systems, including SLAs, highly limited feedback about the state of the system, and variable and unpredictable input sequences. Our main result is a simple and practical algorithm that achieves near-optimal performance on the above two objectives. First, we guarantee nearly optimal utilization of the resource even if compared to the omniscient offline dynamic optimum. Second, we simultaneously satisfy all individual SLAs up to a small error. The main algorithmic tool is a multiplicative weight update algorithm, and a primal-dual argument to obtain its guarantees. We also provide numerical validation on real data to demonstrate the performance of our algorithm in practical applications.
Recommendations
- Dual time-scale distributed capacity allocation and load redirect algorithms for cloud systems
- Resource Allocation in Cloud Computing Via Optimal Control to Queuing Systems
- Service provisioning problem in cloud and multi-cloud systems
- Extended efficiency and soft-fairness multiresource allocation in a cloud computing system
- SLA based resource allocation policies in autonomic environments
Cites work
- A decision-theoretic generalization of on-line learning and an application to boosting
- Dynamic server allocation to parallel queues with randomly varying connectivity
- Fast Approximation Algorithms for Fractional Packing and Covering Problems
- Foundations of machine learning
- scientific article; zbMATH DE number 3126031 (Why is no real title available?)
- scientific article; zbMATH DE number 2107836 (Why is no real title available?)
- Lectures on modern convex optimization. Analysis, algorithms, and engineering applications
- Online algorithms: a survey
- Online learning and online convex optimization
- Optimal Energy and Delay Tradeoffs for Multiuser Wireless Downlinks
- Regret analysis of stochastic and nonstochastic multi-armed bandit problems
- Service provisioning problem in cloud and multi-cloud systems
- The multiplicative weights update method: a meta-algorithm and applications
Cited in
(27)- Optimization-based resource allocation for software as a service application in cloud computing
- Resource allocation based on redundancy models for high availability cloud
- Utilizing unreliable public resources for higher profit and better sla compliance in computing utilities
- Minimum congestion mapping in a cloud
- Constructing Performance-Predictable Clusters with Performance-Varying Resources of Clouds
- Service provisioning problem in cloud and multi-cloud systems
- Dual time-scale distributed capacity allocation and load redirect algorithms for cloud systems
- scientific article; zbMATH DE number 2043505 (Why is no real title available?)
- \texttt{DEPAS}: a decentralized probabilistic algorithm for auto-scaling
- An index policy for dynamic pricing in cloud computing under price commitments
- Automated provisioning of fairly priced resources
- A backfilling algorithm based on EDF and LWF in cloud environment
- Collaborative Resource Allocation Over a Hybrid Cloud Center and Edge Server Network
- A strong law for the rate of growth of long latency periods in a cloud computing service
- A theory of auto-scaling for resource reservation in cloud services
- Resource Allocation in Cloud Computing Via Optimal Control to Queuing Systems
- scientific article; zbMATH DE number 7108665 (Why is no real title available?)
- Minimum congestion mapping in a cloud
- A scheduling algorithm towards bandwidth guarantee for virtual cluster in the cloud
- A filter based resource demand estimation for on-demand provision
- A combinatorial auction mechanism for time-varying multidimensional resource allocation and pricing in fog computing
- Extended efficiency and soft-fairness multiresource allocation in a cloud computing system
- Technical Note—Cloud Cost Optimization: Model, Bounds, and Asymptotics
- Resource allocation algorithms for virtualized service hosting platforms
- A distributed primal-dual hybrid gradient algorithm for fair resource allocation
- Distributionally robust resource allocation using Wasserstein distance
- SLA based resource allocation policies in autonomic environments
This page was built for publication: Dynamic resource allocation in the cloud with near-optimal efficiency
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5106381)