• Home
  • Search
  • Compiler Optimization for Irregular Memory Access Patterns in PGAS Programs
  • Cite Icon1
  • https://doi.org/10.1007/978-3-031-31445-2_1Copy DOI Icon

Compiler Optimization for Irregular Memory Access Patterns in PGAS Programs

  • Jan 1, 2023
  • Thomas B Rolinger +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Irregular memory access patterns pose performance and user productivity challenges on distributed-memory systems. They can lead to fine-grained remote communication and the data access patterns are often not known until runtime. The Partitioned Global Address Space (PGAS) programming model addresses these challenges by providing users with a view of a distributed-memory system that resembles a single shared address space. However, this view often leads programmers to write code that causes fine-grained remote communication, which can result in poor performance. Prior work has shown that the performance of irregular applications written in Chapel, a high-level PGAS language, can be improved by manually applying optimizations. However, applying such optimizations by hand reduces the productivity advantages provided by Chapel and the PGAS model. We present an inspector-executor based compiler optimization for Chapel programs that automatically performs remote data replication. While there have been similar compiler optimizations implemented for other PGAS languages, high-level features in Chapel such as implicit processor affinity lead to new challenges for compiler optimization. We evaluate the performance of our optimization across two irregular applications. Our results show that the total runtime can be improved by as much as 52x on a Cray XC system with a low-latency interconnect and 364x on a standard Linux cluster with an Infiniband interconnect, demonstrating that significant performance gains can be achieved without sacrificing user productivity.

Similar Papers
  • Conference Article

Scalable PGAS Metadata Management on Extreme Scale Systems

  • May 01, 2013
  • D Chavarria-Miranda +2
  • Conference Article
  • Citations5

DART-CUDA: A PGAS Runtime System for Multi-GPU Systems

  • Jun 01, 2015
  • Lei Zhou +1
  • Conference Article
  • Citations6

Enabling Multi-physics Coupled Simulations within the PGAS Programming Framework

  • May 23, 2011
  • Fan Zhang +3
  • Conference Article
  • Citations22

Parallel performance wizard: A performance analysis tool for partitioned global-address-space programming

  • Apr 01, 2008
  • Proceedings - IEEE International Parallel and Distributed Processing Symposium
  • Hung-Hsun Su +2
  • Conference Article
  • Citations4

Locality-aware power optimization and measurement methodology for PGAS workloads on SMP clusters

  • Jun 01, 2013
  • David K Newsom +3
  • Conference Article
  • Citations2

Predictive energy management techniques for PGAS programming

  • May 01, 2013
  • David K Newsom +3
  • Research Article

PCI Express 기반 OpenSHMEM 초기 설계 및 구현

  • Mar 31, 2017
  • KIPS Transactions on Computer and Communication Systems
  • Young-Woong Joo +1
  • Research Article
  • Citations5

Extending a message passing runtime to support partitioned, global logical address spaces

  • Nov 13, 2016
  • D Brian Larkins +1
  • Research Article
  • Citations3

Hybrid-view programming of nuclear fusion simulation code in the PGAS parallel programming language XcalableMP

  • Jun 01, 2016
  • Parallel Computing
  • Keisuke Tsugane +5
  • Conference Article
  • Citations8

Distributed Shared Memory Programming in the Cloud

  • May 01, 2012
  • Ahmad Anbar +2
  • Conference Article
  • Citations6

High performance OpenSHMEM for Xeon Phi clusters: Extensions, runtime designs and application co-design

  • Sep 01, 2014
  • Jithin Jose +6
  • Conference Article
  • Citations6

On the performance and energy efficiency of the PGAS programming model on multicore architectures

  • Jul 01, 2016
  • Jeremie Lagraviere +4
  • Research Article

Using the PGAS Programming Paradigm for Biological Sequence Alignment on a Chip Multi-Threading Architecture

  • Feb 29, 2008
  • Zenodo (CERN European Organization for Nuclear Research)
  • Mohamed Bakhouya +2
  • Conference Article
  • Citations196

Productivity and performance using partitioned global address space languages

  • Jul 27, 2007
  • Katherine Yelick +15
  • Conference Article
  • Citations2

Paving the way for Distributed Non-Blocking Algorithms and Data Structures in the Partitioned Global Address Space model

  • May 01, 2020
  • Garvit Dewan +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.