ANGSD: analysis of next generation sequencing data

670 Citationer (Scopus)
1817 Downloads (Pure)

Abstract

Background: High-throughput DNA sequencing technologies are generating vast amounts of data. Fast, flexible and memory efficient implementations are needed in order to facilitate analyses of thousands of samples simultaneously. Results: We present a multithreaded program suite called ANGSD. This program can calculate various summary statistics, and perform association mapping and population genetic analyses utilizing the full information in next generation sequencing data by working directly on the raw sequencing data or by using genotype likelihoods. Conclusions: The open source c/c++ program ANGSD is available at . The program is tested and validated on GNU/Linux systems. The program facilitates multiple input formats including BAM and imputed beagle genotype probability files. The program allow the user to choose between combinations of existing methods and can perform analysis that is not implemented elsewhere.

OriginalsprogEngelsk
Artikelnummer356
TidsskriftB M C Bioinformatics
Vol/bind15
Antal sider13
ISSN1471-2105
DOI
StatusUdgivet - 25 nov. 2014

Fingeraftryk

Dyk ned i forskningsemnerne om 'ANGSD: analysis of next generation sequencing data'. Sammen danner de et unikt fingeraftryk.

Citationsformater