Working paper Open Access

Scaling of Biological Data Work ows to Large HPC Systems - A Case Study in Marine Genomics -

Thomas Röblitz

      Thomas Röblitz
      Department for Research Computing, University Center for Information Technology (USIT), University of Oslo, P.O. Box 1059, Blindern, 0316 Oslo, Norway
    Scaling of Biological Data Work ows to Large HPC Systems - A Case Study in Marine Genomics -
    workflows, magnitude scaling
    2014-06-04
  Working paper
    Creative Commons Attribution 4.0 International
    Open Access
    Sequencing projects, like the Aqua Genome project, generate vast amounts of data which is processed through dif-
ferent work ows composed of several steps linked together. Currently, such workflows are often run manually on
large servers. With the increasing amount of raw data that approach is no longer feasible. The successful imple-
mentation of the project's goals requires 2-3 orders of magnitude scaling of computing, while achieving high reli-
ability on and supporting ease-of-use of super computing resources at the same time. We describe two example
use cases, the implementation challenges and constraints, the actual application enabling and report our ndings.
      European Commission
      10.13039/501100000780
      312763
      PRACE - Third Implementation Phase Project
