• Open Access IconOpen Access
  • https://doi.org/10.5194/gmd-2021-382-rc2Copy DOI Icon

Comment on gmd-2021-382

  • Feb 8, 2022
  • Lars Hoffmann +13 more
Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Similar Papers
Abstract

Lagrangian models are fundamental tools to study atmospheric transport processes and for practical applications such as dispersion modeling for anthropogenic and natural emission sources. However, conducting large-scale Lagrangian transport simulations with millions of air parcels or more can become numerically rather costly. In this study, we assessed the potential of exploiting graphics processing units (GPUs) to accelerate Lagrangian transport simulations. We ported the Massive-Parallel Trajectory Calculations (MPTRAC) model to GPUs using the open accelerator (OpenACC) programming model. The trajectory calculations conducted within the MPTRAC model were fully ported to GPUs, i.e., except for feeding in the meteorological input data and for extracting the particle output data, the code operates entirely on the GPU devices without frequent data transfers between CPU and GPU memory. Model verification, performance analyses, and scaling tests of the MPI/OpenMP/OpenACC hybrid parallelization of MPTRAC were conducted on the JUWELS Booster supercomputer operated by the Jülich Supercomputing Centre, Germany. The JUWELS Booster comprises 3744 NVIDIA A100 Tensor Core GPUs, providing a peak performance of 71.0 PFlop/s. As of June 2021, it is the most powerful supercomputer in Europe and listed among the most energy-efficient systems internationally. For large-scale simulations comprising 108 particles driven by the European Centre for Medium-Range Weather Forecasts' ERA5 reanalysis, the performance evaluation showed a maximum speedup of a factor of 16 due to the utilization of GPUs compared to CPU-only runs on the JUWELS Booster. In the large-scale GPU run, about 67 % of the runtime is spent on the physics calculations, conducted on the GPUs. Another 15 % of the runtime is required for file-I/O, mostly to read the large ERA5 data set from disk. Meteorological data preprocessing on the CPUs also requires about 15 % of the runtime. Although this study identified potential for further improvements of the GPU code, we consider the MPTRAC model ready for production runs on the JUWELS Booster in its present form. The GPU code provides a much faster time to solution than the CPU code, which is particularly relevant for near-real-time applications of a Lagrangian transport model.

Loading PDF

Similar Papers
  • PDF
  • Research Article
  • Citations23

Massive-Parallel Trajectory Calculations version 2.2 (MPTRAC-2.2): Lagrangian transport simulations on graphics processing units (GPUs)

  • Apr 05, 2022
  • Geoscientific Model Development
  • Lars Hoffmann +13
  • Preprint Article
  • Citations2

From ERA-Interim to ERA5: considerable impact of ECMWF's next-generation reanalysis on Lagrangian transport simulations

  • Mar 23, 2020
  • Lars Hoffmann +10
  • Conference Article
  • Citations3

Efficient GPU-Based Query Processing with Pruned List Caching in Search Engines

  • Dec 01, 2017
  • Dongdong Wang +6
  • Conference Article
  • Citations3

FlexGPU: A Flexible and Efficient Scheduler for GPU Sharing Systems

  • May 01, 2020
  • Qichen Chen +3
  • Research Article
  • Citations109

Multi-phase SPH modelling of violent hydrodynamics on GPUs

  • Jul 07, 2015
  • Computer Physics Communications
  • Athanasios Mokos +3
  • Research Article
  • Citations6

MO‐F‐213CD‐01: GPU‐Based Monte Carlo Methods for Accelerating Radiographic and CT Imaging Dose Calculations: Feasibility and Scalability

  • Jun 01, 2012
  • Medical Physics
  • T Liu +2
  • Research Article
  • Citations20

Accelerating the discontinuous Galerkin method for seismic wave propagation simulations using the graphic processing unit (GPU)—single-GPU implementation

  • Jul 31, 2012
  • Computers & Geosciences
  • Dawei Mu +2
  • Preprint Article

Modeling of hydroxyl oxidation and wet deposition in volcanic sulfur dioxide plumes: the July 2018 Ambae eruption

  • Mar 27, 2022
  • Mingzhao Liu +2
  • Conference Article
  • Citations3

Victream

  • Dec 05, 2017
  • Jun Suzuki +6
  • Research Article
  • Citations2

Gpubased, microsecond latency, hectochannel mimo feedback control of magnetically confined plasmas

  • Jan 01, 2013
  • Columbia Academic Commons (Columbia University)
  • M E Mauel +1
  • Research Article
  • Citations34

OpenACC acceleration of an unstructured CFD solver based on a reconstructed discontinuous Galerkin method for compressible flows

  • Feb 09, 2015
  • International Journal for Numerical Methods in Fluids
  • Yidong Xia +4
  • Conference Article
  • Citations20

ConVGPU: GPU Management Middleware in Container Based Virtualized Environment

  • Sep 01, 2017
  • Daeyoun Kang +4
  • Research Article
  • Citations21

MoDNN: Memory Optimal Deep Neural Network Training on Graphics Processing Units

  • Mar 01, 2019
  • IEEE Transactions on Parallel and Distributed Systems
  • Xiaoming Chen +3
  • Conference Article
  • Citations1

Automatic Parallelization of GPU Applications Using OpenCL

  • Jul 01, 2015
  • Lizandro D Solano-Quinde +2
  • Conference Article
  • Citations114

Stealing Webpages Rendered on Your Browser by Exploiting GPU Vulnerabilities

  • May 01, 2014
  • Sangho Lee +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.