AdmixPipe v3: facilitating population structure delimitation from SNP data

Bioinform Adv. 2023 Nov 23;3(1):vbad168. doi: 10.1093/bioadv/vbad168. eCollection 2023.

Abstract

Summary: Quantifying genetic clusters (=populations) from genotypic data is a fundamental, but non-trivial task for population geneticists that is compounded by: hierarchical population structure, diverse analytical methods, and complex software dependencies. AdmixPipe v3 ameliorates many of these issues in a single bioinformatic pipeline that facilitates all facets of population structure analysis by integrating outputs generated by several popular packages (i.e. CLUMPAK, EvalAdmix). The pipeline interfaces disparate software packages to parse Admixture outputs and conduct EvalAdmix analyses in the context of multimodal population structure results identified by CLUMPAK. We further streamline these tasks by packaging AdmixPipe v3 within a Docker container to create a standardized analytical environment that allows for complex analyses to be replicated by different researchers. This also grants operating system flexibility and mitigates complex software dependencies.

Availability and implementation: Source code, documentation, example files, and usage examples are freely available at https://github.com/stevemussmann/admixturePipeline. Installation is facilitated via Docker container available from https://hub.docker.com/r/mussmann/admixpipe. Usage under Windows operating systems requires the Windows Subsystem for Linux.

Associated data

  • Dryad/10.5061/dryad.51c59zw62