首页 | 本学科首页   官方微博 | 高级检索  
     


ANGSD: Analysis of Next Generation Sequencing Data
Authors:Thorfinn Sand Korneliussen  Anders Albrechtsen  Rasmus Nielsen
Affiliation:.Centre for GeoGenetics, Natural History Museum of Denmark, Copenhagen, Denmark ;.Bioinformatics Centre, Department of Biology, University of Copenhagen, Ole Maaloes Vej 5, Copenhagen, DK-2200 Denmark ;.Department of Integrative Biology and Statistics, UC-Berkeley, 4098 VLSB, Berkeley, California, 94720 USA
Abstract:

Background

High-throughput DNA sequencing technologies are generating vast amounts of data. Fast, flexible and memory efficient implementations are needed in order to facilitate analyses of thousands of samples simultaneously.

Results

We present a multithreaded program suite called ANGSD. This program can calculate various summary statistics, and perform association mapping and population genetic analyses utilizing the full information in next generation sequencing data by working directly on the raw sequencing data or by using genotype likelihoods.

Conclusions

The open source c/c++ program ANGSD is available at http://www.popgen.dk/angsd. The program is tested and validated on GNU/Linux systems. The program facilitates multiple input formats including BAM and imputed beagle genotype probability files. The program allow the user to choose between combinations of existing methods and can perform analysis that is not implemented elsewhere.

Electronic supplementary material

The online version of this article (doi:10.1186/s12859-014-0356-4) contains supplementary material, which is available to authorized users.
Keywords:Next-generation sequencing   Bioinformatics   Population genetics   Association studies
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号