Job-Aware Scheduling for Big Data Processing
At a glance
- الاستشهادات
- 5
- المراجع
- 14
- Comments
- 0
Abstract
Most big data jobs are network-bound, which involve large amount of data transfers among the nodes in a cluster. Optimizing the scheduling of flows can improve big data job performance. Traditional techniques are mostly flow-based scheduling, without considering the flow correlations. In this paper, we take the dependency of the flows into account and propose traffic forecasting and job-aware priority scheduling for big data processing. First, we forecast the network traffic for flows of the same job through run-time monitoring, and assign a unique priority for each job and tag every packet in this job. Then we schedule flows of the same priority (often the same job) in a FIFO order. We implement our proposed scheme using NS-2 simulator and show that our system can increase the network utilization and reduce the job completion time.
Publication details
- DOI
- 10.1109/ccbd.2015.14
- OpenAlex
- W2327972540
- Document type
- conference-paper
- Language
- EN
- Last metadata update
Comments
تسجيل الدخول للانضمام إلى النقاش.