Sequence Read Archive (SRA) Data Technical User Manual
- Stine, Adam1, 2, 3
-
Caetano-Anollés, Derek1, 2, 3
- Durbrow, Kenneth1, 2, 3
- Raetz, Wolfgang1, 2, 3
- Boshkin, Anatoly1, 2, 3
- Klymenko, Andrew1, 2, 3
-
Pannu, Ravinder1, 2, 3
- Iwig, Laura1, 2, 3
-
Skripchenko, Yuriy1, 2, 3
- Hicks, Denise1, 2, 3
- Yaschenko, Eugene1, 2, 3
-
O'Sullivan, Christopher1, 2, 3
-
Ryan, Connor1, 2, 3
-
Zalunin, Vadim1, 2, 3
- Domrachev, Mikhail1, 2, 3
- Katz, Kenneth1, 2, 3
- Brister, J. Rodney1, 2, 3
Description
The Sequence Read Archive (SRA) is the largest publicly available repository of high throughput sequencing data.
This Technical User Manual provides a comprehensive guide to SRA data formats, compression strategies, and tools for managing Next-Generation Sequencing (NGS) data. It details the SRA Normalized and SRA Lite formats, which standardize and optimize storage for large-scale sequencing data, including basecalls, quality scores, and alignment information. The manual also covers the Virtual Database (VDB) system, a structured and efficient framework for storing and accessing biological data, along with practical examples and scripts for data conversion and manipulation.
This resource is essential for researchers and bioinformaticians seeking to understand or utilize SRA data for interoperability, analysis, and archival purposes.
Files
SRA-Technical-Manual-v1.1.pdf
Files
(966.6 kB)
| Name | Size | Download all |
|---|---|---|
|
md5:017616593a98e643b87bf48439a7d763
|
966.6 kB | Preview Download |
Additional details
Software
- Development Status
- Active