controlling for population structure and genotyping platform bias in the emerge multi-institutional biobank linked to electronic health records

Clicks: 280
ID: 156225
2014
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This article has not been analysed, so there is no overall score — reader engagement is measured and shown alongside.
AI Quality Assessment
Not analyzed
Readership in this journal
Steady

Ranked #99 of 264 articles by views in chemical record (new york, ny)

Most read Least read

Bar heights use a square-root scale. Only the 120 most-read articles are drawn; the journal has 264 in total.

Mint this article as an NFT
Not yet minted

Create a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.

5 SUSD one-off · no wallet required
Abstract
Combining samples across multiple cohorts in large-scale scientific research programs is often required to achieve the necessary power for genome-wide association studies. Controlling for genomic ancestry through principal component analysis (PCA) to address the effect of population stratification is a common practice. In addition to local genomic variation, such as copy number variation and inversions, other factors directly related to combining multiple studies, such as platform and site recruitment bias, can drive the correlation patterns in PCA. In this report, we describe combination and analysis of multi-ethnic cohort with biobanks linked to electronic health records for large-scale genomic association discovery analyses. First, we outline the observed site and platform bias, in addition to ancestry differences. Second, we outline a general protocol for selecting variants for input into the subject variance-covariance matrix, the conventional PCA approach. Finally, we introduce an alternative approach to PCA by deriving components from subject loadings calculated from a reference sample. This alternative approach of generating principal components controlled for site and platform bias, in addition to ancestry differences, with the advantage of fewer covariates and degrees of freedom.principal component analysis, ancestry, biobank, loadings, genetic association study
Reference Key
crosslin2014frontierscontrolling Use this key to autocite in the manuscript while using SciMatic Manuscript Manager or Thesis Manager
Authors ;David Russell Crosslin;David Russell Crosslin;Gerard eTromp;Amber eBurt;Daniel Seung Kim;Shefali S Verma;Anastasia M. Lucas;Yuki eBradford;Dana C. Crawford;Dana C. Crawford;Sebastian M. Armasu;John A. Heit;M. Geoffrey Hayes;Helena eKuivaniemi;Marylyn D Ritchie;Gail P. Jarvik;Gail P. Jarvik;Mariza eDe Andrade
Journal chemical record (new york, ny)
Year 2014
DOI
10.3389/fgene.2014.00352
URL
Keywords

Citations

No citations found. To add a citation, contact the admin at info@scimatic.org

No comments yet. Be the first to comment on this article.