Bioinformatics with Python Cookbook
上QQ阅读APP看书,第一时间看更新

Working with modern sequence formats

Here, we will work with FASTQ files, the standard format output used by modern sequencers. You will learn how to work with quality scores per base and also consider the variations in output coming from different sequencing machines and databases. This is the first recipe that will use real data (big data) from the Human 1,000 Genomes Project. We will start with a brief description of the project.