Data in the genomics field is booming. In just a few years, organizations such as the National Institutes of Health (NIH) will host 50+ petabytesâ??or over 50 million gigabytesâ??of genomic data, and theyâ??re turning to cloud infrastructure to make that data available to the research community. How do you adapt analysis tools and protocols to access and analyze that volume of data in the cloud?
With this practical book, researchers will learn how to work with genomics algorithms using open source tools including the Genome Analysis Toolkit (GATK), Docker, WDL, and Terra. Geraldine Van der Auwera, longtime custodian of the GATK user community, and Brian Oâ??Connor of the UC Santa Cruz Genomics Institute, guide you through the process. Youâ??ll learn by working with real data and genomics algorithms from the field.
This book covers:
ISBN: | 9781491975190 |
Publication date: | 24th April 2020 |
Author: | Geraldine van der Auwera, Brian D OConnor |
Publisher: | O'Reilly an imprint of O'Reilly Media |
Format: | Paperback |
Pagination: | 493 pages |
Genres: |
Epidemiology and Medical statistics Molecular biology Algorithms and data structures Computer science |