Large Scale Maximum Average Power Multiple Inference on Time-Course Count Data with Application to RNA-Seq Analysis.
Clicks: 313
ID: 36158
2019
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This
article has not been analysed, so there is no overall score —
reader engagement is measured and shown alongside.
Reader Engagement
Emerging Content
72.5
/100
313 views
224 readers
Trending
AI Quality Assessment
Not analyzed
Readership in this journal
EmergingRanked #8 of 238 articles by views in biometrics
Most read
Least read
Bar heights use a square-root scale. Only the 120 most-read articles are drawn; the journal has 238 in total.
Mint this article as an NFT
Not yet mintedCreate a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.
5
SUSD
one-off · no wallet required
Abstract
Experiments that longitudinally collect RNA sequencing (RNA-seq) data can provide transformative insights in biology research by revealing dynamic patterns of genes. Such experiments create great demands for new analytic approaches to identify differentially expressed (DE) genes based on large-scale time-course count data. Existing methods, however, are sub-optimal with respect to power and may lack theoretical justification. Furthermore, most existing tests are designed to distinguish among conditions based on overall differential patterns across time, though in practice, a variety of composite hypotheses are of more scientific interest. Lastly, some current methods may fail to control the false discovery rate (FDR). In this paper, we propose a new model and testing procedure to address the above issues simultaneously. Specifically, conditional on a latent Gaussian mixture with evolving means, we model the data by negative binomial distributions. Motivated by Storey (2007) and Hwang and Liu (2010), we introduce a general testing framework based on the proposed model and show that the proposed test enjoys the optimality property of maximum average power. The test allows not only identification of traditional DE genes but also testing of a variety of composite hypotheses of biological interest. We establish the identifiability of the proposed model, implement the proposed method via efficient algorithms, and demonstrate its good performance via simulation studies. The procedure reveals interesting biological insights when applied to data from an experiment that examines the effect of varying light environments on the fundamental physiology of the marine diatom Phaeodactylum tricornutum. This article is protected by copyright. All rights reserved.
| Reference Key |
cao2019largebiometrics
Use this key to autocite in the manuscript while using
SciMatic Manuscript Manager or Thesis Manager
|
|---|---|
| Authors | Cao, Meng;Zhou, Wen;Breidt, F Jay;Peers, Graham; |
| Journal | biometrics |
| Year | 2019 |
| DOI |
10.1111/biom.13144
|
| URL | |
| Keywords | Keywords not found |
Citations
No citations found. To add a citation, contact the admin at info@scimatic.org
Comments
No comments yet. Be the first to comment on this article.