• Home
  • Search
  • MetaMorph: A Library Framework for Interoperable Kernels on Multi- and Many-Core Clusters
  • Open Access IconOpen Access
  • Cite Icon5
  • https://doi.org/10.1109/sc.2016.10Copy DOI Icon

MetaMorph: A Library Framework for Interoperable Kernels on Multi- and Many-Core Clusters

  • Nov 1, 2016
  • Ahmed E Helal +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

To attain scalable performance efficiently, the HPC community expects future exascale systems to consist of multiple nodes, each with different types of hardware accelerators. In addition to GPUs and Intel MICs, additional candidate accelerators include embedded multiprocessors and FPGAs. End users need appropriate tools to efficiently use the available compute resources in such systems, both within a compute node and across compute nodes. As such, we present MetaMorph, a library framework designed to (automatically) extract as much computational capability as possible from HPC systems. Its design centers around three core principles: abstraction, interoperability, and adaptivity. To demonstrate its efficacy, we present a case study that uses the structured grids design pattern, which is heavily used in computational fluid dynamics. We show how MetaMorph significantly reduces the development time, while delivering performance and interoperability across an array of heterogeneous devices, including multicore CPUs, Intel MICs, AMD GPUs, and NVIDIA GPUs.

Similar Papers
  • Conference Article

Clone_n(): Parallel Thread Creation for Upcoming Many-Core Architectures

  • Sep 01, 2012
  • Balazs Gerofi +2
  • Research Article
  • Citations24

Manycore Algorithms for Batch Scalar and Block Tridiagonal Solvers

  • Jun 30, 2016
  • ACM Transactions on Mathematical Software
  • Endre László +2
  • Research Article
  • Citations20

New capabilities of the Monte Carlo dose engine ARCHER-RT: Clinical validation of the Varian TrueBeam machine for VMAT external beam radiotherapy.

  • Apr 13, 2020
  • Medical Physics
  • David P Adam +4
  • Conference Article
  • Citations3

Applications of many-core technologies to on-line event reconstruction in High Energy Physics experiments

  • Oct 01, 2013
  • A Gianelle +12
  • Conference Article
  • Citations8

Performance Portable Applications for Hardware Accelerators: Lessons Learned from SPEC ACCEL

  • May 01, 2015
  • Guido Juckeland +2
  • PDF
  • Research Article
  • Citations48

The ESCAPE project: Energy-efficient Scalable Algorithms for Weather Prediction at Exascale

  • Oct 22, 2019
  • Geoscientific Model Development
  • Andreas Müller +58
  • Conference Article
  • Citations3

Accelerating Lagrangian particle dispersion in the atmosphere with OpenCL across multiple platforms

  • Jan 01, 2014
  • Paul Harvey +2
  • PDF
  • Research Article
  • Citations4

Reconstruction of Charged Particle Tracks in Realistic Detector Geometry Using a Vectorized and Parallelized Kalman Filter Algorithm

  • Jan 01, 2020
  • EPJ Web of Conferences
  • Giuseppe Cerati +16
  • PDF
  • Research Article

Performance of GeantV EM Physics Models

  • Oct 01, 2017
  • Journal of Physics: Conference Series
  • G Amádio +29
  • Research Article
  • Citations37

Time-domain seismic modeling in viscoelastic media for full waveform inversion on heterogeneous computing platforms with OpenCL

  • Dec 08, 2016
  • Computers & Geosciences
  • Gabriel Fabien-Ouellet +2
  • Research Article
  • Citations4

Enabling Parallel Performance and Portability of Solid Mechanics Simulations Across CPU and GPU Architectures

  • Nov 07, 2024
  • Information
  • Nathaniel Morgan +12
  • Research Article
  • Citations11

Sparse matrix–vector multiplication on the Single-Chip Cloud Computer many-core processor

  • Aug 14, 2013
  • Journal of Parallel and Distributed Computing
  • Juan C Pichel +1
  • Book Chapter
  • Citations10

Evaluating Performance Portability of OpenMP for SNAP on NVIDIA, Intel, and AMD GPUs Using the Roofline Methodology

  • Jan 01, 2021
  • Neil A Mehta +4
  • Book Chapter

LU Factorisation on Xeon and Xeon Phi Processors

  • Jan 01, 2016
  • Jackson Adrian +1
  • Conference Article
  • Citations3

Methods to utilize SIMT and SIMD instruction level parallelism in tridiagonal solvers

  • Jul 01, 2014
  • Endre Laszlo +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.