Skip NavigationSkip to Content

ISHAPE: new rapid and accurate software for haplotyping

  1. Author:
    Delaneau, O.
    Coulonges, C.
    Boelle, P. Y.
    Nelson, G.
    Spadoni, J. L.
    Zagury, J. F.
  2. Author Address

    Conservatoire Natl Arts & Metiers, Chaire Bioinformat, F-75003 Paris, France. INSERM, Unite U736, Ctr Rech Cordeliers, F-75006 Paris, France. INSERM, Unite U707, F-75012 Paris, France. SAIC Frederick, Lab Genom Divers, Frederick, MD USA. INSERM, Unite U841, F-94000 Creteil, France.;Zagury, JF, Conservatoire Natl Arts & Metiers, Chaire Bioinformat, 292 Rue St Martin, F-75003 Paris, France.;olivier.delaneau@gmail.com cedcoul@gmail.com boelle@u707.jussieu.fr nelsong@ncifcrf.gov jean-louis.spadoni@cnam.fr zagury@cnam.fr
    1. Year: 2007
    2. Date: Jun
  1. Journal: Bmc Bioinformatics
    1. 8
  2. Type of Article: Article
  3. Article Number: 205
  4. ISSN: 1471-2105
  1. Abstract:

    Background: We have developed a new haplotyping program based on the combination of an iterative multiallelic EM algorithm (IEM), bootstrap resampling and a pseudo Gibbs sampler. The use of the IEM-bootstrap procedure considerably reduces the space of possible haplotype configurations to be explored, greatly reducing computation time, while the adaptation of the Gibbs sampler with a recombination model on this restricted space maintains high accuracy. On large SNP datasets (> 30 SNPs), we used a segmented approach based on a specific partition-ligation strategy. We compared this software, Ishape (Iterative Segmented HAPlotyping by Em), with reference programs such as Phase, Fastphase, and PL-EM. Analogously with Phase, there are 2 versions of Ishape: Ishape1 which uses a simple coalescence model for the pseudo Gibbs sampler step, and Ishape2 which uses a recombination model instead. Results: We tested the program on 2 types of real SNP datasets derived from Hapmap: adjacent SNPs (high LD) and SNPs spaced by 5 Kb (lower level of LD). In both cases, we tested 100 replicates for each size: 10, 20, 30, 40, 50, 60, and 80 SNPs. For adjacent SNPs Ishape2 is superior to the other software both in terms of speed and accuracy. For SNPs spaced by 5 Kb, Ishape2 yields similar results to Phase2.1 in terms of accuracy, and both outperform the other software. In terms of speed, Ishape2 runs about 4 times faster than Phase2.1 with 10 SNPs, and about 10 times faster with 80 SNPs. For the case of 5kb-spaced SNPs, Fastphase may run faster with more than 100 SNPs. Conclusion: These results show that the Ishape heuristic approach for haplotyping is very competitive in terms of accuracy and speed and deserves to be evaluated extensively for possible future widespread use.

    See More

External Sources

  1. DOI: 10.1186/1471-2105-8-205
  2. WOS: 000248130100001

Library Notes

  1. No notes added.
NCI at Frederick

You are leaving a government website.

This external link provides additional information that is consistent with the intended purpose of this site. The government cannot attest to the accuracy of a non-federal site.

Linking to a non-federal site does not constitute an endorsement by this institution or any of its employees of the sponsors or the information and products presented on the site. You will be subject to the destination site's privacy policy when you follow the link.

ContinueCancel