• Home
  • Search
  • Fast parallel cutoff pair interactions for molecular dynamics on heterogeneous systems
  • Cite Icon8
  • https://doi.org/10.1109/tst.2012.6216756Copy DOI Icon

Fast parallel cutoff pair interactions for molecular dynamics on heterogeneous systems

Show More
  • Abstract
  • Literature Map
  • Citations
  • Similar Papers
Abstract

Heterogeneous systems with both Central Processing Units (CPUs) and Graphics Processing Units (GPUs) are frequently used to accelerate short-ranged Molecular Dynamics (MD) simulations. The most time-consuming task in short-ranged MD simulations is the computation of particle-to-particle interactions. Beyond a certain distance, these interactions decrease to zero. To minimize the operations to investigate distance, previous works have tiled interactions by employing the spatial attribute, which increases the memory access and GPU computations, hence decreasing performance. Other studies ignore the spatial attribute and construct an all-versus-all interaction matrix, which has poor scalability. This paper presents an improved algorithm. The algorithm first bins particles into voxels according to the spatial attributes, and then tiles the all-versus-all matrix into voxel-versus-voxel sub-matrixes. Only the sub-matrixes between neighboring voxels are computed on the GPU. Therefore, the algorithm reduces the distance examine operations and limits additional memory access and GPU computations. This paper also adopts a multi-level programming model to implement the algorithm on multi-nodes of Tianhe-lA. By employing (1) a patch design to exploit parallelism across the simulation domain, (2) a communication overlapping method to overlap the communications between CPUs and GPUs, and (3) a dynamic workload balancing method to adjust the workloads among compute nodes, the implementation achieves a speedup of 4.16× on one NVIDIA Tesla M2050 GPU compared to a 2.93 GHz six-core Intel Xeon X5670 CPU. In addition, it runs 2.41× faster on 256 compute nodes of Tianhe-lA (with two CPUs and one GPU inside a node) than on 256 GPU-excluded nodes.

Similar Papers
  • Research Article
  • Citations31

GPU acceleration of MPAS microphysics WSM6 using OpenACC directives: Performance and verification

  • Oct 14, 2020
  • Computers & Geosciences
  • Jae Youp Kim +2
  • Research Article

SU‐E‐T‐423: Fast Photon Convolution Calculation with a 3D‐Ideal Kernel On the GPU

  • Jun 01, 2015
  • Medical Physics
  • S Moriya +2
  • Book Chapter
  • Citations3

Utilizing GPU Virtualization to Protect the Private Keys of GPU Cryptographic Computation

  • Jan 01, 2018
  • Ziyang Wang +3
  • Book Chapter
  • Citations15

GPU Computation in Bioinspired Algorithms: A Review

  • Jan 01, 2011
  • M G Arenas +3
  • Conference Article
  • Citations12

High-Order Spectral Difference: Verification and Acceleration using GPU Computing

  • Jun 22, 2013
  • Ben J Zimmerman +2
  • Book Chapter
  • Citations1

GPU Computing in Biomolecular Modeling and Nanodesign

  • Jan 01, 2012
  • Tibor Kožár
  • Research Article
  • Citations12

HPMaX: heterogeneous parallel matrix multiplication using CPUs and GPUs

  • Oct 11, 2020
  • Computing
  • Homin Kang +2
  • Research Article
  • Citations13

Practical parallel AES algorithms on cloud for massive users and their performance evaluation

  • Dec 17, 2015
  • Concurrency and Computation: Practice and Experience
  • Xiongwei Fei +3
  • Research Article
  • Citations26

Efficient methods for implementation of multi-level nonrigid mass-preserving image registration on GPUs and multi-threaded CPUs

  • Jan 06, 2016
  • Computer Methods and Programs in Biomedicine
  • Nathan D Ellingwood +3
  • PDF
  • Research Article
  • Citations25

DdcMD: A fully GPU-accelerated molecular dynamics program for the Martini force field.

  • Jul 23, 2020
  • The Journal of Chemical Physics
  • Xiaohua Zhang +8
  • Research Article
  • Citations2

Gpubased, microsecond latency, hectochannel mimo feedback control of magnetically confined plasmas

  • Jan 01, 2013
  • Columbia Academic Commons (Columbia University)
  • M E Mauel +1
  • PDF
  • Research Article
  • Citations5

Parallel Thomas approach development for solving tridiagonal systems in GPU programming − steady and unsteady flow simulation

  • Jan 01, 2020
  • Mechanics & Industry
  • Milad Souri +2
  • Research Article

Improving MSAProbs Algorithm performance and Parallel Computing using GPU

  • Apr 18, 2023
  • International Journal of Computer Applications
  • Sally Zaki El-Hadary +2
  • Research Article
  • Citations4

A systematic parallel strategy for generating contours from large-scale DEM data using collaborative CPUs and GPUs

  • Feb 18, 2021
  • Cartography and Geographic Information Science
  • Chen Zhou +1
  • PDF
  • Research Article
  • Citations7

LTTng CLUST: A System-Wide Unified CPU and GPU Tracing Tool for OpenCL Applications

  • Aug 19, 2015
  • Advances in Software Engineering
  • David Couturier +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.