Exploiting next-generation sequencing to solve the haplotyping puzzle in polyploids: a simulation study

E. Motazedi, H.J. Finkers, C.A. Maliepaard, D. de Ridder*

*Corresponding author for this work

Research output: Contribution to journalArticleAcademicpeer-review

40 Citations (Scopus)


Haplotypes are the units of inheritance in an organism, and many genetic analyses depend on their precise determination. Methods for haplotyping single individuals use the phasing information available in next-generation sequencing reads, by matching overlapping single-nucleotide polymorphisms while penalizing post hoc nucleotide corrections made. Haplotyping diploids is relatively easy, but the complexity of the problem increases drastically for polyploid genomes, which are found in both model organisms and in economically relevant plant and animal species. Although a number of tools are available for haplotyping polyploids, the effects of the genomic makeup and the sequencing strategy followed on the accuracy of these methods have hitherto not been thoroughly evaluated.

We developed the simulation pipeline haplosim to evaluate the performance of three haplotype estimation algorithms for polyploids: HapCompass, HapTree and SDhaP, in settings varying in sequencing approach, ploidy levels and genomic diversity, using tetraploid potato as the model. Our results show that sequencing depth is the major determinant of haplotype estimation quality, that 1 kb PacBio circular consensus sequencing reads and Illumina reads with large insert-sizes are competitive and that all methods fail to produce good haplotypes when ploidy levels increase. Comparing the three methods, HapTree produces the most accurate estimates, but also consumes the most resources. There is clearly room for improvement in polyploid haplotyping algorithms.
Original languageEnglish
Pages (from-to)387-403
JournalBriefings in Bioinformatics
Issue number3
Early online date8 Jan 2017
Publication statusPublished - May 2018


  • next-generation sequencing; haplotyping; polyploids; benchmarking


Dive into the research topics of 'Exploiting next-generation sequencing to solve the haplotyping puzzle in polyploids: a simulation study'. Together they form a unique fingerprint.

Cite this