首页 | 本学科首页   官方微博 | 高级检索  
     


Multi-stage resource-aware scheduling for data centers with heterogeneous servers
Authors:Tony T. Tran  Meghana Padmanabhan  Peter Yun Zhang  Heyse Li  Douglas G. Down  J. Christopher Beck
Affiliation:1.Department of Mechanical and Industrial Engineering,University of Toronto,Toronto,Canada;2.Engineering Systems Division,Massachusetts Institute of Technology,Cambridge,USA;3.Department of Computing and Software,McMaster University,Hamilton,Canada
Abstract:This paper presents a three-stage algorithm for resource-aware scheduling of computational jobs in a large-scale heterogeneous data center. The algorithm aims to allocate job classes to machine configurations to attain an efficient mapping between job resource request profiles and machine resource capacity profiles. The first stage uses a queueing model that treats the system in an aggregated manner with pooled machines and jobs represented as a fluid flow. The latter two stages use combinatorial optimization techniques to solve a shorter-term, more accurate representation of the problem using the first-stage, long-term solution for heuristic guidance. In the second stage, jobs and machines are discretized. A linear programming model is used to obtain a solution to the discrete problem that maximizes the system capacity given a restriction on the job class and machine configuration pairings based on the solution of the first stage. The final stage is a scheduling policy that uses the solution from the second stage to guide the dispatching of arriving jobs to machines. We present experimental results of our algorithm on both Google workload trace data and generated data and show that it outperforms existing schedulers. These results illustrate the importance of considering heterogeneity of both job and machine configuration profiles in making effective scheduling decisions.
Keywords:
本文献已被 SpringerLink 等数据库收录!
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号