Cargando…

Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment

The unprecedented computing resource needs of the ATLAS experiment have motivated the Collaboration to become a leader in exploiting High Performance Computers (HPCs). To meet the requirements of HPCs, the PanDA system has been equipped with two new components; Pilot 2 and Harvester, that were desig...

Descripción completa

Detalles Bibliográficos
Autores principales: Nilsson, Paul, Anisenkov, Alexey, Benjamin, Douglas, Guan, Wen, Javurek, Tomas, Oleynik, Danila
Lenguaje:eng
Publicado: 2020
Materias:
Acceso en línea:https://dx.doi.org/10.1051/epjconf/202024503025
http://cds.cern.ch/record/2713594
_version_ 1780965334007676928
author Nilsson, Paul
Anisenkov, Alexey
Benjamin, Douglas
Guan, Wen
Javurek, Tomas
Oleynik, Danila
author_facet Nilsson, Paul
Anisenkov, Alexey
Benjamin, Douglas
Guan, Wen
Javurek, Tomas
Oleynik, Danila
author_sort Nilsson, Paul
collection CERN
description The unprecedented computing resource needs of the ATLAS experiment have motivated the Collaboration to become a leader in exploiting High Performance Computers (HPCs). To meet the requirements of HPCs, the PanDA system has been equipped with two new components; Pilot 2 and Harvester, that were designed with HPCs in mind. While Harvester is a resource-facing service which provides resource provisioning and workload shaping, Pilot 2 is responsible for payload execution on the resource. The presentation focuses on Pilot 2, which is a complete rewrite of the original PanDA Pilot used by ATLAS and other experiments for well over a decade. Pilot 2 has a flexible and adaptive design that allows for plugins to be defined with streamlined workflows. In particular, it has plugins for specific hardware infrastructures (HPC/GPU clusters) as well as for dedicated workflows defined by the needs of an experiment. Examples of dedicated HPC workflows are discussed in which the Pilot either uses an MPI application for processing fine-grained event level service under the control of the Harvester service or acts like an MPI application itself and runs a set of job in an assemble. In addition to describing the technical details of these workflows, results are shown from its deployment on Titan (OLCF) and other HPCs in ATLAS.
id cern-2713594
institution Organización Europea para la Investigación Nuclear
language eng
publishDate 2020
record_format invenio
spelling cern-27135942021-03-22T22:08:54Zdoi:10.1051/epjconf/202024503025http://cds.cern.ch/record/2713594engNilsson, PaulAnisenkov, AlexeyBenjamin, DouglasGuan, WenJavurek, TomasOleynik, DanilaHarnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS ExperimentParticle Physics - ExperimentThe unprecedented computing resource needs of the ATLAS experiment have motivated the Collaboration to become a leader in exploiting High Performance Computers (HPCs). To meet the requirements of HPCs, the PanDA system has been equipped with two new components; Pilot 2 and Harvester, that were designed with HPCs in mind. While Harvester is a resource-facing service which provides resource provisioning and workload shaping, Pilot 2 is responsible for payload execution on the resource. The presentation focuses on Pilot 2, which is a complete rewrite of the original PanDA Pilot used by ATLAS and other experiments for well over a decade. Pilot 2 has a flexible and adaptive design that allows for plugins to be defined with streamlined workflows. In particular, it has plugins for specific hardware infrastructures (HPC/GPU clusters) as well as for dedicated workflows defined by the needs of an experiment. Examples of dedicated HPC workflows are discussed in which the Pilot either uses an MPI application for processing fine-grained event level service under the control of the Harvester service or acts like an MPI application itself and runs a set of job in an assemble. In addition to describing the technical details of these workflows, results are shown from its deployment on Titan (OLCF) and other HPCs in ATLAS.ATL-SOFT-PROC-2020-032oai:cds.cern.ch:27135942020-03-23
spellingShingle Particle Physics - Experiment
Nilsson, Paul
Anisenkov, Alexey
Benjamin, Douglas
Guan, Wen
Javurek, Tomas
Oleynik, Danila
Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment
title Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment
title_full Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment
title_fullStr Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment
title_full_unstemmed Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment
title_short Harnessing the power of supercomputers using the PanDA Pilot 2 in the ATLAS Experiment
title_sort harnessing the power of supercomputers using the panda pilot 2 in the atlas experiment
topic Particle Physics - Experiment
url https://dx.doi.org/10.1051/epjconf/202024503025
http://cds.cern.ch/record/2713594
work_keys_str_mv AT nilssonpaul harnessingthepowerofsupercomputersusingthepandapilot2intheatlasexperiment
AT anisenkovalexey harnessingthepowerofsupercomputersusingthepandapilot2intheatlasexperiment
AT benjamindouglas harnessingthepowerofsupercomputersusingthepandapilot2intheatlasexperiment
AT guanwen harnessingthepowerofsupercomputersusingthepandapilot2intheatlasexperiment
AT javurektomas harnessingthepowerofsupercomputersusingthepandapilot2intheatlasexperiment
AT oleynikdanila harnessingthepowerofsupercomputersusingthepandapilot2intheatlasexperiment