• Home
  • Search
  • Adaptive work-stealing with parallelism feedback
  • Cite Icon74
  • https://doi.org/10.1145/1394441.1394443Copy DOI Icon

Adaptive work-stealing with parallelism feedback

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Multiprocessor scheduling in a shared multiprogramming environment can be structured as two-level scheduling, where a kernel-level job scheduler allots processors to jobs and a user-level thread scheduler schedules the work of a job on its allotted processors. We present a randomized work-stealing thread scheduler for fork-join multithreaded jobs that provides continual parallelism feedback to the job scheduler in the form of requests for processors. Our A-STEAL algorithm is appropriate for large parallel servers where many jobs share a common multiprocessor resource and in which the number of processors available to a particular job may vary during the job's execution. Assuming that the job scheduler never allots a job more processors than requested by the job's thread scheduler, A-STEAL guarantees that the job completes in near-optimal time while utilizing at least a constant fraction of the allotted processors. We model the job scheduler as the thread scheduler's adversary, challenging the thread scheduler to be robust to the operating environment as well as to the job scheduler's administrative policies. For example, the job scheduler might make a large number of processors available exactly when the job has little use for them. To analyze the performance of our adaptive thread scheduler under this stringent adversarial assumption, we introduce a new technique called trim analysis, which allows us to prove that our thread scheduler performs poorly on no more than a small number of time steps, exhibiting near-optimal behavior on the vast majority. More precisely, suppose that a job has work T 1 and span T ∞ . On a machine with P processors, A-STEAL completes the job in an expected duration of O ( T 1 / P˜ + T ∞ + L lg P ) time steps, where L is the length of a scheduling quantum, and P˜ denotes the O ( T ∞ + L lg P )-trimmed availability. This quantity is the average of the processor availability over all time steps except the O ( T ∞ + L lg P ) time steps that have the highest processor availability. When the job's parallelism dominates the trimmed availability, that is, P˜ < T 1 / T ∞ , the job achieves nearly perfect linear speedup. Conversely, when the trimmed mean dominates the parallelism, the asymptotic running time of the job is nearly the length of its span, which is optimal. We measured the performance of A-STEAL on a simulated multiprocessor system using synthetic workloads. For jobs with sufficient parallelism, our experiments confirm that A-STEAL provides almost perfect linear speedup across a variety of processor availability profiles. We compared A-STEAL with the ABP algorithm, an adaptive work-stealing thread scheduler developed by Arora et al. [1998] which does not employ parallelism feedback. On moderately to heavily loaded machines with large numbers of processors, A-STEAL typically completed jobs more than twice as quickly as ABP, despite being allotted the same number or fewer processors on every step, while wasting only 10% of the processor cycles wasted by ABP.

Similar Papers
  • Conference Article
  • Citations3

A Truthful Mechanism for Scheduling and Pricing Pleasingly Parallel Jobs in a Service Cloud

  • Jul 01, 2018
  • Bingbing Zheng +3
  • Book Chapter

Parallel Job Scheduling Using Bacterial Foraging Optimization for Heterogeneous Multi-cluster Environment

  • Jan 01, 2019
  • Navjot Kaur +2
  • Conference Article
  • Citations31

Program-based static allocation policies for highly parallel computers

  • Mar 28, 1995
  • I.M Ismail +1
  • Conference Article
  • Citations18

A Comparative Study of Job Scheduling Strategies in Large-Scale Parallel Computational Systems

  • Jul 01, 2013
  • Aftab Ahmed Chandio +4
  • Conference Article
  • Citations1

An Adaptive Scheduler Framework for Complex Workflow Jobs on Grid Systems

  • Sep 01, 2011
  • G M Siddesh +1
  • Book Chapter
  • Citations58

Effective Selection of Partition Sizes for Moldable Scheduling of Parallel Jobs

  • Jan 01, 2002
  • Srividya Srinivasan +4
  • PDF
  • Research Article
  • Citations3

Optimizing Big Data Retrieval and Job Scheduling Using Deep Learning Approaches

  • Jan 01, 2023
  • Computer Modeling in Engineering & Sciences
  • Bao Rong Chang +2
  • Conference Article
  • Citations19

Goodbye to Fixed Bandwidth Reservation: Job Scheduling with Elastic Bandwidth Reservation in Clouds

  • Dec 01, 2016
  • Haiying Shen +3
  • Book Chapter
  • Citations5

Prediction Based Job Scheduling Strategy for a Volunteer Desktop Grid

  • Jan 01, 2013
  • Shaik Naseera +1
  • Book Chapter
  • Citations29

Hybrid Performance-Oriented Scheduling of Moldable Jobs with QoS Demands in Multiclusters and Grids

  • Jan 01, 2004
  • Ligang He +4
  • Research Article
  • Citations17

A flexible resource investment problem based on project splitting for aircraft moving assembly line

  • Jul 25, 2019
  • Assembly Automation
  • Yifei Ren +1
  • Conference Article
  • Citations19

Partitioned multiprocessor scheduling of mixed-criticality parallel jobs

  • Aug 01, 2014
  • Guangdong Liu +3
  • Research Article
  • Citations7

Cloud computing and big data: Technologies and applications

  • May 20, 2018
  • Concurrency and Computation: Practice and Experience
  • Mostapha Zbakh +3
  • Book Chapter
  • Citations7

Online Scheduling of Parallel Jobs with Dependencies on 2-Dimensional Meshes

  • Jan 01, 2003
  • Deshi Ye +1
  • Conference Article
  • Citations39

Architectural support for enhanced SMT job scheduling

  • Nov 08, 2004
  • A Settle +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.