arXiv · 2409.05886
Optimizing STAR Aligner for High Throughput Computing in the Cloud
Abstract
We propose a scalable, cloud-native architecture designed for Transcriptomics Atlas Pipeline, using a resource-intensive STAR aligner and processing tens or hundreds of terabytes of RNA-seq data. We implement the pipeline using AWS cloud services, introduce performance optimizations and perform experimental evaluation in the cloud. Our optimization techniques result in computational savings thanks to the "early stopping" approach, selection of right-sized resources, and using newer version of Ensembl genome.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Piotr Kica, Sabina Lichołai, Michał Orzechowski, Maciej Malawski. 2024-08-26. Optimizing STAR Aligner for High Throughput Computing in the Cloud. https://arxiv.org/abs/2409.05886
Cite the original work for its findings. Save a collection to share your selection of sources.