Accurate identification of SNPs from next-generation sequencing data is crucial for high-quality downstream analysis. Whole genome sequence data of 65 key ancestors of genotyped Swiss dairy populations were available for investigation (24 billion reads, 96.8% mapped to UMD31, 12x coverage). Four publically available variant calling programmes were assessed and different levels of pre-calling handling for each method were tested and compared. SNP concordance was examined with Illuminas BovineHD Genotyping BeadChip. Depending on variant calling software used, between 16,894,054 and 22,048,382 SNP were identified (multi-sample calling). A total of 14,644,310 SNP were identified by all four variantcallers (multi-sample calling). InDel counts ranged from 1,997,791 to 2,857,754; 1,708,649 InDels were identified by all four variant callers. A minimum of pre-calling data handling resulted in the highest non-reference sensitivity and the lowest non-reference discrepancy rate.
Proceedings of the World Congress on Genetics Applied to Livestock Production, Volume Methods and Tools: Genome sequencing (Posters), , 668, 2014
Download Full PDF
Search the Proceedings
This work is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License.