• Home
  • Search
  • CATCH: A Cloud-Based Adaptive Data Transfer Service for HPC
  • Open Access IconOpen Access
  • Cite Icon32
  • https://doi.org/10.1109/ipdps.2011.118Copy DOI Icon

CATCH: A Cloud-Based Adaptive Data Transfer Service for HPC

  • May 1, 2011
  • Henry M Monti +2 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Modern High Performance Computing (HPC) applications process very large amounts of data. A critical research challenge lies in transporting input data to the HPC center from a number of distributed sources, e.g., scientific experiments and web repositories, etc., and offloading the result data to geographically distributed, intermittently available end-users, often over under-provisioned connections. Such end-user data services are typically performed using point-to-point transfers that are designed for well-endowed sites and are unable to reconcile the center's resource usage and users' delivery deadlines, unable to adapt to changing dynamics in the end-to-end data path and are not fault-tolerant. To overcome these inefficiencies, decentralized HPC data services are emerging as viable alternatives. In this paper, we develop and enhance such distributed data services by designing CATCH, a Cloud-based Adaptive data Transfer service for HPC. CATCH leverages a bevy of cloud storage resources to orchestrate a decentralized data transport with fail-over capabilities. Our results demonstrate that CATCH is a feasible approach, and can help improve the data transfer times at the HPC center by as much as 81.1\% for typical HPC workloads.

Similar Papers
  • Conference Article
  • Citations10

SciDP: Support HPC and Big Data Applications via Integrated Scientific Data Processing

  • Sep 01, 2018
  • Kun Feng +3
  • Conference Article
  • Citations44

Predicting the Energy-Consumption of MPI Applications at Scale Using Only a Single Node

  • Sep 01, 2017
  • Franz Christian Heinrich +7
  • Conference Article
  • Citations3

Expressing and Managing Network Policies for Emerging HPC Systems

  • Jul 28, 2019
  • Proceedings of the Practice and Experience in Advanced Research Computing on Rise of the Machines (learning)
  • Sergio Rivera +3
  • Research Article
  • Citations3

Parallel Performance Modeling using a Genetic Programming-based Error Correction Procedure

  • Jul 01, 2007
  • SIMULATION
  • Kavitha Raghavachar +4
  • Research Article
  • Citations12

Managing high-performance computing applications as an on-demand service on federated clouds

  • Mar 16, 2018
  • Computers & Electrical Engineering
  • Zhengxiong Hou +5
  • Research Article

I/O patterns modeling of HPC applications with call stacks for predictive prefetch

  • Feb 01, 2026
  • Future Generation Computer Systems
  • Louis-Marie Nicolas +3
  • Conference Article
  • Citations5

Combining cooling technology and facility design to improve HPC data center energy efficiency

  • May 01, 2016
  • Lynn A Parnell +2
  • PDF
  • Research Article
  • Citations8

Overcoming Challenges to Continuous Integration in HPC

  • Nov 01, 2022
  • Computing in Science & Engineering
  • Todd Gamblin +1
  • Conference Article
  • Citations8

Correlation-wise Smoothing: Lightweight Knowledge Extraction for HPC Monitoring Data

  • May 01, 2021
  • Alessio Netti +3
  • Research Article

PerfSuite: An Automatic Performance Suite for HPC Applications

  • Jan 14, 2026
  • Journal of Circuits, Systems and Computers
  • Qiong Jiang +5
  • Conference Article
  • Citations2

Employing Augmented Reality for Cybersecurity Operations in High Performance Computing Environments

  • Jul 28, 2019
  • Proceedings of the Practice and Experience in Advanced Research Computing on Rise of the Machines (learning)
  • Nitin Sukhija +2
  • Conference Article
  • Citations7

Performance and Cost-aware HPC in Clouds: A Network Interconnection Assessment

  • Jul 01, 2020
  • Anderson M Maliszewski +5
  • Research Article
  • Citations6

HDF5eis: A storage and input/output solution for big multidimensional time series data from environmental sensors

  • Apr 12, 2023
  • Geophysics
  • Malcolm C A White +5
  • Conference Article
  • Citations5

Proposal of MPI Operation Level Checkpoint/Rollback and One Implementation

  • Jan 01, 2006
  • Yuan Tang +2
  • Conference Article
  • Citations11

A Case Study of Designing Efficient Algorithm-based Fault Tolerant Application for Exascale Parallelism

  • May 01, 2012
  • Erlin Yao +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.