• Home
  • Search
  • Modeling Instability for Large Scale Processing Tasks Within HEP Distributed Computing Environments
  • Open Access IconOpen Access
  • https://doi.org/10.1007/978-3-319-69835-9_32Copy DOI Icon

Modeling Instability for Large Scale Processing Tasks Within HEP Distributed Computing Environments

  • Nov 3, 2017
  • Olga Datskova +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

The scale of large processing tasks being run within the scientific Grid and Cloud environments have introduced a need for stability guarantees from geographically spanning resources, to ensure that failures are detected and handled preemptively. Performance inefficiencies within stacked service environments are a challenge to detect, where failures stem from a multitude of causes, often requiring expert intervention. Online reporting and classification of performance fluctuations can aid experts, central services, and users to target service areas where optimizations can be introduced. This paper describes an approach for modeling performance states for production tasks running within the ALICE Grid. We first provide an overview for the ALICE data and software workflow, focusing on the production job computational profile. Data center event state is then developed, based on data center job, computing, storage, and user behavior. With across site analysis, we then train groups to classify service domain states. Our approach is able to detect periods of service instability and the affected service domains. This can guide users, central, and data center experts to take action in advance of service failure effects.

Similar Papers
  • Research Article
  • Citations4

SEED: solar energy‐aware efficient scheduling for data centers

  • Dec 05, 2013
  • Concurrency and Computation: Practice and Experience
  • Chao Jing +2
  • Conference Article

From Network Perspective to Optimize Shuffling Phase in MapReduce

  • Aug 01, 2018
  • Haifeng Wang +1
  • Conference Article
  • Citations56

Improving spark application throughput via memory aware task co-location

  • Dec 11, 2017
  • Vicent Sanz Marco +3
  • Book Chapter

Environmental Impact Study on Carbon Footprint Emission and Development of Software Architectural Framework to Measure the Level of Emission in Cloud Services

  • Jan 01, 2020
  • P Vaishnavi +1
  • Research Article

Optimum Selection of Virtual Machine Using Improved Particle Swarm Optimization in Cloud Environment

  • Feb 28, 2022
  • International Journal of Computer Networks And Applications
  • R Jeena +1
  • PDF
  • Research Article
  • Citations43

Electricity Price Forecasting for Cloud Computing Using an Enhanced Machine Learning Model

  • Jan 01, 2020
  • IEEE Access
  • Saleh Albahli +2
  • Book Chapter
  • Citations2

Comparing Single Tier and Three Tier Infrastructure Designs against DDoS Attacks

  • Jan 01, 2021
  • Akashdeep Bhardwaj +1
  • Research Article
  • Citations9

An efficient priority-driven congestion control algorithm for data center networks

  • Jun 01, 2020
  • China Communications
  • Jiahua Zhu +6
  • Research Article
  • Citations564

GreenCloud: a packet-level simulator of energy-aware cloud computing data centers

  • Nov 09, 2010
  • The Journal of Supercomputing
  • Dzmitry Kliazovich +2
  • Conference Article
  • Citations312

GreenCloud: A Packet-Level Simulator of Energy-Aware Cloud Computing Data Centers

  • Dec 01, 2010
  • Dzmitry Kliazovich +3
  • Research Article

Multi-Objective Low-Carbon Scheduling Method for Data Centers Based on Ensemble Reinforcement Learning

  • Jan 01, 2026
  • IEEE Transactions on Smart Grid
  • Yifan Wang +3
  • Research Article
  • Citations2

Intelligent and compliant dynamic software license consolidation in cloud environment

  • Jan 01, 2022
  • Computing
  • Leila Helali +1
  • Research Article
  • Citations3

Optimizing Data Centre Energy Efficiency with Dynamic Resource Allocation and Intelligent Cooling Management through Machine Learning

  • Apr 04, 2024
  • Journal of Electrical Systems
  • Niranchana Radhakrishnan
  • Research Article
  • Citations1

Load balancing in cloud environment with switching mechanism and token-based algorithm

  • Jan 01, 2019
  • International Journal of Public Sector Performance Management
  • M Aruna +2
  • Research Article
  • Citations1

Artificial intelligence-based cloud-internet of things resource management for energy conservation

  • Sep 01, 2024
  • International Journal of Advances in Applied Sciences
  • Soukaina Ouhame +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.