Back to results
Bibliographic record · Consultation and access
Artículo

Second-generation PLINK: rising to the challenge of larger and richer datasets

Christopher Chang; Carson C. Chow; Laurent CAM Tellier; Shashaank Vattikuti; Shaun Purcell; James J. Lee · GigaScience · 2015

Resource page
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.

Resource access

Open the content from the main option or choose another available source.

OpenAlex OpenAlex Works
Entrar por OpenAlex
Main access

Resource page

Resource reference page. Full text availability has not been automatically confirmed.
Open resource

Summary

Descripción general del contenido del recurso.

BACKGROUND: PLINK 1 is a widely used open-source C/C++ toolset for genome-wide association studies (GWAS) and research in population genetics. However, the steady accumulation of data from imputation and whole-genome sequencing studies has exposed a strong need for faster and scalable implementations of key functions, such as logistic regression, linkage disequilibrium estimation, and genomic distance evaluation. In addition, GWAS and population-genetic data now frequently contain genotype likelihoods, phase information, and/or multiallelic variants, none of which can be represented by PLINK 1's primary data format. FINDINGS: To address these issues, we are developing a second-generation codebase for PLINK. The first major release from this codebase, PLINK 1.9, introduces extensive use of bit-level parallelism, [Formula: see text]-time/constant-space Hardy-Weinberg equilibrium and Fisher's exact tests, and many other algorithmic improvements. In combination, these changes accelerate most operations by 1-4 orders of magnitude, and allow the program to handle datasets too large to fit in RAM. We have also developed an extension to the data format which adds low-overhead support for genotype likelihoods, phase, multiallelic variants, and reference vs. alternate alleles, which is the basis of our planned second release (PLINK 2.0). CONCLUSIONS: The second-generation versions of PLINK will offer dramatic improvements in performance and compatibility. For the first time, users without access to high-end computing resources can perform several essential analyses of the feature-rich and very large genetic datasets coming into use.

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

Chang, C, Chow, C. C, Tellier, L. C, Vattikuti, S, Purcell, S, & Lee, J. J. (2015). Second-generation PLINK: rising to the challenge of larger and richer datasets. https://doi.org/10.1186/s13742-015-0047-8

MLA

Chang, Christopher, et al. "Second-generation PLINK: rising to the challenge of larger and richer datasets." 2015. https://doi.org/10.1186/s13742-015-0047-8.

Chicago

Chang, Christopher, Carson C. Chow, Laurent CAM Tellier, Shashaank Vattikuti, Shaun Purcell, and James J. Lee. 2015. "Second-generation PLINK: rising to the challenge of larger and richer datasets.". https://doi.org/10.1186/s13742-015-0047-8.

Harvard

Chang, C. et al. 2015, Second-generation PLINK: rising to the challenge of larger and richer datasets, GigaScience, available at: https://doi.org/10.1186/s13742-015-0047-8 [Accessed 8 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
Second-generation PLINK: rising to the challenge of larger and richer datasets
Author / contributors
Christopher Chang; Carson C. Chow; Laurent CAM Tellier; Shashaank Vattikuti; Shaun Purcell; James J. Lee
Publisher
GigaScience
Publication year
2015
Language
English

Subjects

Explore related resources through these subjects.

Copied