arXiv · 2312.07090
Scaling a Variant Calling Genomics Pipeline with FaaS
Abstract
With the escalating complexity and volume of genomic data, the capacity of biology institutions' HPC faces limitations. While the Cloud presents a viable solution for short-term elasticity, its intricacies pose challenges for bioinformatics users. Alternatively, serverless computing allows for workload scalability with minimal developer burden. However, porting a scientific application to serverless is not a straightforward process. In this article, we present a Variant Calling genomics pipeline migrated from single-node HPC to a serverless architecture. We describe the inherent challenges of this approach and the engineering efforts required to achieve scalability. We contribute by open-sourcing the pipeline for future systems research and as a scalable user-friendly tool for the bioinformatics community.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Aitor Arjona, Arnau Gabriel-Atienza, Sara Lanuza-Orna, Xavier Roca-Canals, Ayman Bourramouss, Tyler K. Chafin, Lucio Marcello, Paolo Ribeca, Pedro García-López. 2023-12-12. Scaling a Variant Calling Genomics Pipeline with FaaS. https://doi.org/10.1145/3631295.3631403
Cite the original work for its findings. Save a collection to share your selection of sources.