Cargando…

The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale

The EU-funded ESCAPE project aims at enabling a prototype federated storage infrastructure, a Data Lake, that would handle data on the exabyte-scale, address the FAIR data management principles and provide science projects a unified scalable data management solution for accessing and analyzing large...

Descripción completa

Detalles Bibliográficos
Autores principales: Dona, Rizart, Maria, Riccardo Di
Lenguaje:eng
Publicado: 2021
Materias:
Acceso en línea:https://dx.doi.org/10.1051/epjconf/202125102060
http://cds.cern.ch/record/2780368
_version_ 1780971866434830336
author Dona, Rizart
Maria, Riccardo Di
author_facet Dona, Rizart
Maria, Riccardo Di
author_sort Dona, Rizart
collection CERN
description The EU-funded ESCAPE project aims at enabling a prototype federated storage infrastructure, a Data Lake, that would handle data on the exabyte-scale, address the FAIR data management principles and provide science projects a unified scalable data management solution for accessing and analyzing large volumes of scientific data. In this respect, data transfer and management technologies such as Rucio, FTS and GFAL are employed along with monitoring enabling solutions such as Grafana, Elasticsearch and perf- SONAR. This paper presents and describes the technical details behind the machinery of testing and monitoring of the Data Lake – this includes continuous automated functional testing, network monitoring and development of insightful visualizations that reflect the current state of the system. Topics that are also addressed include the integration with the CRIC information system as well as the initial support for token based authentication / authorization by using OpenID Connect. The current architecture of these components is provided and future enhancements are discussed.
id cern-2780368
institution Organización Europea para la Investigación Nuclear
language eng
publishDate 2021
record_format invenio
spelling cern-27803682021-09-07T19:17:05Zdoi:10.1051/epjconf/202125102060http://cds.cern.ch/record/2780368engDona, RizartMaria, Riccardo DiThe ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scaleComputing and ComputersThe EU-funded ESCAPE project aims at enabling a prototype federated storage infrastructure, a Data Lake, that would handle data on the exabyte-scale, address the FAIR data management principles and provide science projects a unified scalable data management solution for accessing and analyzing large volumes of scientific data. In this respect, data transfer and management technologies such as Rucio, FTS and GFAL are employed along with monitoring enabling solutions such as Grafana, Elasticsearch and perf- SONAR. This paper presents and describes the technical details behind the machinery of testing and monitoring of the Data Lake – this includes continuous automated functional testing, network monitoring and development of insightful visualizations that reflect the current state of the system. Topics that are also addressed include the integration with the CRIC information system as well as the initial support for token based authentication / authorization by using OpenID Connect. The current architecture of these components is provided and future enhancements are discussed.oai:cds.cern.ch:27803682021
spellingShingle Computing and Computers
Dona, Rizart
Maria, Riccardo Di
The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
title The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
title_full The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
title_fullStr The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
title_full_unstemmed The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
title_short The ESCAPE Data Lake: The machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
title_sort escape data lake: the machinery behind testing, monitoring and supporting a unified federated storage infrastructure of the exabyte-scale
topic Computing and Computers
url https://dx.doi.org/10.1051/epjconf/202125102060
http://cds.cern.ch/record/2780368
work_keys_str_mv AT donarizart theescapedatalakethemachinerybehindtestingmonitoringandsupportingaunifiedfederatedstorageinfrastructureoftheexabytescale
AT mariariccardodi theescapedatalakethemachinerybehindtestingmonitoringandsupportingaunifiedfederatedstorageinfrastructureoftheexabytescale
AT donarizart escapedatalakethemachinerybehindtestingmonitoringandsupportingaunifiedfederatedstorageinfrastructureoftheexabytescale
AT mariariccardodi escapedatalakethemachinerybehindtestingmonitoringandsupportingaunifiedfederatedstorageinfrastructureoftheexabytescale