$content.nome.text $content.cognome.text

Framing Apache Spark in life sciences

Manconi, Andrea;Gnocchi, Matteo;Milanesi, Luciano;Marullo, Osvaldo;Armano, Giuliano

2023-01-01

Abstract

Advances in high-throughput and digital technologies have required the adoption of big data for handling complex tasks in life sciences. However, the drift to big data led researchers to face technical and infrastructural challenges for storing, sharing, and analysing them. In fact, this kind of tasks requires distributed computing systems and algorithms able to ensure efficient processing. Cutting edge distributed programming frameworks allow to implement flexible algorithms able to adapt the computation to the data over on-premise HPC clusters or cloud architectures. In this context, Apache Spark is a very powerful HPC engine for large-scale data processing on clusters. Also thanks to specialised libraries for working with structured and relational data, it allows to support machine learning, graph-based computation, and stream processing. This review article is aimed at helping life sciences researchers to ascertain the features of Apache Spark and to assess whether it can be successfully used in their research activities.

Short Card

Tab complete

Full Sheet(DC)

         Anno di pubblicazione 
       
        2023 
       
         Parole chiave 
       
        Apache Spark; Big data; Parallel computing; HPC 
       
         Type: 
       
        1.1 Articolo in rivista

Files in This Item:

File	Size	Format
2023 - Heliyon (Manconi).pdf open access Type: versione editoriale Size 817.86 kB Format Adobe PDF View/Open	817.86 kB	Adobe PDF	View/Open

University of Cagliari

University of Cagliari

Framing Apache Spark in life sciences

Manconi, Andrea;Gnocchi, Matteo;Milanesi, Luciano;Marullo, Osvaldo;Armano, Giuliano

2023-01-01

Abstract

Short Card

Tab complete

Full Sheet(DC)

Framing Apache Spark in life sciences

Manconi, Andrea;Gnocchi, Matteo;Milanesi, Luciano;Marullo, Osvaldo;Armano, Giuliano

2023-01-01

Abstract

Short Card Tab complete Full Sheet(DC)

Questionnaire and social

Short Card

Tab complete

Full Sheet(DC)