• Home
  • Search
  • Optimal Loop Tiling for Minimizing Write Operations on NVMs with Complete Memory Latency Hiding
  • Cite Icon2
  • https://doi.org/10.1109/asp-dac52403.2022.9712532Copy DOI Icon

Optimal Loop Tiling for Minimizing Write Operations on NVMs with Complete Memory Latency Hiding

  • Jan 17, 2022
  • Rui Xu +4 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Non-volatile memory (NVM) is expected to be the second level memory (named remote memory) in two-level memory hierarchy in the future. However, NVM has the limited write endurance, thus it is vital to reduce the number of write operations on NVM. Meanwhile, in two-level memory hierarchy, prefetch is widely used for fetching certain data before it is actually required, to hide the remote memory access latency. In general, large-scale nested loop is the performance bottleneck in one program due to the write operations on NVM caused by the first level memory (named local memory) miss and data reuse. Loop tiling is the key technique for grouping iterations so as to reduce the communication with remote memory used in compiler. In this paper, we propose a new loop tiling approach for minimizing the write operations on NVMs and completely hiding the NVM access latency. Specifically, we introduce a series of theorems to help loop tiling. Then, a legal tile shape and an optimal tile size selection strategy is proposed according to data dependency and local memory capacity. Furthermore, we propose a pipeline scheduling policy to completely hide the remote memory latency. Extensive experiments show that the proposed techniques can reduce write operations on NVMs by 95.1% on average, and NVM latency can be completely hidden.

Similar Papers
  • Research Article
  • Citations5

A Combined Optimization Method for Tuning Two-Level Memory Hierarchy Considering Energy Consumption

  • Sep 28, 2010
  • EURASIP Journal on Embedded Systems
  • Abel Guilhermino Silva-Filho +1
  • Conference Article
  • Citations36

Exploiting inter- and intra-memory asymmetries for data mapping in hybrid tiered-memories

  • Jun 16, 2020
  • Shihao Song +2
  • Research Article
  • Citations16

A Latency-Optimized and Energy-Efficient Write Scheme in NVM-Based Main Memory

  • Jan 01, 2020
  • IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems
  • Yuncheng Guo +2
  • Research Article
  • Citations3

Improving Bank-Level Parallelism for In-Memory Checkpointing in Hybrid Memory Systems

  • Apr 01, 2022
  • IEEE Transactions on Big Data
  • Xiaofei Liao +3
  • Conference Article

An Approach of Spatial Usage Optimization for NVM-Based Storage System

  • Dec 01, 2017
  • Zheng Zhang +6
  • Conference Article
  • Citations40

Bridging the Latency Gap between NVM and DRAM for Latency-bound Operations

  • Jul 01, 2019
  • Georgios Psaropoulos +4
  • Conference Article
  • Citations1

A Memory Access Performance Detection and Optimization Driven by Address Mapping

  • May 20, 2022
  • Weijia Hua +1
  • Conference Article
  • Citations7

An FPGA-based platform for non volatile memory emulation

  • Aug 01, 2017
  • Taemin Lee +1
  • Conference Article
  • Citations29

Switch cache: a framework for improving the remote memory access latency of CC-NUMA multiprocessors

  • Jan 01, 1999
  • R Iyer +1
  • Conference Article
  • Citations958

A durable and energy efficient main memory using phase change memory technology

  • Jun 20, 2009
  • Ping Zhou +3
  • Dissertation

Towards new memory paradigms : integrating non-volatile main memory and remote direct memory access in modern systems

  • Jan 01, 2024
  • Rémi Dulong
  • Conference Article
  • Citations2

Analyzing and modeling the impact of memory latency and bandwidth on application performance

  • Apr 09, 2018
  • Myonghoon Oh +5
  • Conference Article
  • Citations2

Powering-off DRAM with aggressive page-out to storage-class memory in low power virtual memory system

  • Apr 01, 2016
  • Yusuke Shirota +3
  • Conference Article
  • Citations56

Multi-GPU System Design with Memory Networks

  • Dec 01, 2014
  • Gwangsun Kim +3
  • Conference Article

Dual-layered file cache on cc-NUMA system

  • Apr 25, 2006
  • Zhou Yingchao +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.