Source:http://linkedlifedata.com/resource/pubmed/id/15840707
Switch to
Predicate | Object |
---|---|
rdf:type | |
lifeskim:mentions | |
pubmed:issue |
13
|
pubmed:dateCreated |
2005-6-16
|
pubmed:abstractText |
MOTIVATION: In microarray data studies most researchers are keenly aware of the potentially high rate of false positives and the need to control it. One key statistical shift is the move away from the well-known P-value to false discovery rate (FDR). Less discussion perhaps has been spent on the sensitivity or the associated false negative rate (FNR). The purpose of this paper is to explain in simple ways why the shift from P-value to FDR for statistical assessment of microarray data is necessary, to elucidate the determining factors of FDR and, for a two-sample comparative study, to discuss its control via sample size at the design stage. RESULTS: We use a mixture model, involving differentially expressed (DE) and non-DE genes, that captures the most common problem of finding DE genes. Factors determining FDR are (1) the proportion of truly differentially expressed genes, (2) the distribution of the true differences, (3) measurement variability and (4) sample size. Many current small microarray studies are plagued with large FDR, but controlling FDR alone can lead to unacceptably large FNR. In evaluating a design of a microarray study, sensitivity or FNR curves should be computed routinely together with FDR curves. Under certain assumptions, the FDR and FNR curves coincide, thus simplifying the choice of sample size for controlling the FDR and FNR jointly.
|
pubmed:language |
eng
|
pubmed:journal | |
pubmed:citationSubset |
IM
|
pubmed:status |
MEDLINE
|
pubmed:month |
Jul
|
pubmed:issn |
1367-4803
|
pubmed:author | |
pubmed:issnType |
Print
|
pubmed:day |
1
|
pubmed:volume |
21
|
pubmed:owner |
NLM
|
pubmed:authorsComplete |
Y
|
pubmed:pagination |
3017-24
|
pubmed:dateRevised |
2006-11-15
|
pubmed:meshHeading |
pubmed-meshheading:15840707-Algorithms,
pubmed-meshheading:15840707-Computer Simulation,
pubmed-meshheading:15840707-False Positive Reactions,
pubmed-meshheading:15840707-Gene Expression Profiling,
pubmed-meshheading:15840707-Models, Biological,
pubmed-meshheading:15840707-Models, Statistical,
pubmed-meshheading:15840707-Oligonucleotide Array Sequence Analysis,
pubmed-meshheading:15840707-Reproducibility of Results,
pubmed-meshheading:15840707-Sample Size,
pubmed-meshheading:15840707-Sensitivity and Specificity,
pubmed-meshheading:15840707-Software
|
pubmed:year |
2005
|
pubmed:articleTitle |
False discovery rate, sensitivity and sample size for microarray studies.
|
pubmed:affiliation |
Department of Medical Epidemiology and Biostatistics, Karolinska Institutet 17177 Stockholm, Sweden. yudi.pawitan@meb.ki.se
|
pubmed:publicationType |
Journal Article,
Comparative Study,
Evaluation Studies
|