• Home
  • Search
  • Affinity-Based Thread and Data Mapping in Shared Memory Systems
  • Cite Icon43
  • https://doi.org/10.1145/3006385Copy DOI Icon

Affinity-Based Thread and Data Mapping in Shared Memory Systems

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Shared memory architectures have recently experienced a large increase in thread-level parallelism, leading to complex memory hierarchies with multiple cache memory levels and memory controllers. These new designs created a Non-Uniform Memory Access (NUMA) behavior, where the performance and energy consumption of memory accesses depend on the place where the data is located in the memory hierarchy. Accesses to local caches or memory controllers are generally more efficient than accesses to remote ones. A common way to improve the locality and balance of memory accesses is to determine the mapping of threads to cores and data to memory controllers based on the affinity between threads and data. Such mapping techniques can operate at different hardware and software levels, which impacts their complexity, applicability, and the resulting performance and energy consumption gains. In this article, we introduce a taxonomy to classify different mapping mechanisms and provide a comprehensive overview of existing solutions.

Similar Papers
  • PDF
  • Research Article
  • Citations5

Analysis of Memory System of Tiled Many-Core Processors

  • Jan 01, 2019
  • IEEE Access
  • Ye Liu +2
  • Conference Article
  • Citations232

A case for NUMA-aware contention management on multicore systems

  • Sep 11, 2010
  • Sergey Blagodurov +3
  • Research Article
  • Citations4

Algorithmic optimizations of a conjugate gradient solver on shared memory architectures

  • Jan 01, 2006
  • International Journal of Parallel, Emergent and Distributed Systems
  • Henrik Löf +1
  • PDF
  • Research Article

NUMA-AWARE DATA MANAGEMENT FOR NEUTRON CROSS SECTION DATA IN CONTINUOUS ENERGY MONTE CARLO NEUTRON TRANSPORT SIMULATION

  • Jan 01, 2021
  • EPJ Web of Conferences
  • Nicolas Denoyelle +4
  • Research Article
  • Citations107

OpenMP task scheduling strategies for multicore NUMA systems

  • Feb 07, 2012
  • The International Journal of High Performance Computing Applications
  • Stephen L Olivier +4
  • Conference Article
  • Citations13

Model-based, memory-centric performance and power optimization on NUMA multiprocessors

  • Nov 01, 2012
  • Chunyi Su +5
  • Research Article
  • Citations1

Performance visualization for distributed shared memory systems

  • Jan 01, 2001
  • Scalable Computing Practice and Experience
  • James E Lumpp +3
  • Research Article
  • Citations42

Fairness via source throttling

  • Mar 05, 2010
  • ACM SIGPLAN Notices
  • Eiman Ebrahimi +3
  • Conference Article
  • Citations3

A prototype sampling interface for PAPI

  • Jan 01, 2015
  • Ivonne Lopez +2
  • Conference Article
  • Citations20

Low-overhead load-balanced scheduling for sparse tensor computations

  • Sep 01, 2014
  • Muthu Baskaran +2
  • Book Chapter
  • Citations10

Explicit Management of Memory Hierarchy

  • Jan 01, 1997
  • Jarek Nieplocha +2
  • Conference Article
  • Citations7

Revisiting parallel rendering for shared memory machines

  • Apr 10, 2011
  • Boonthanome Nouanesengsy +3
  • Conference Article
  • Citations5

PufferFish: NUMA-Aware Work-stealing Library using Elastic Tasks

  • Dec 01, 2020
  • Vivek Kumar
  • Research Article
  • Citations1

ON THE PERFORMANCE AND TECHNOLOGICAL IMPACT OF ADDING MEMORY CONTROLLERS IN MULTI-CORE PROCESSORS

  • Dec 01, 2010
  • Parallel Processing Letters
  • Jose Carlos Sancho +2
  • Conference Article
  • Citations3

Extending FPGA based teaching boards into the area of distributed memory multiprocessors

  • Jan 01, 2004
  • Michael Manzke +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.