arrow
返回

Statistical similarities between transcriptomics and quantitative shotgun proteomics data

delete2008-04-01
delete163
delete
OA
AI
N
Norman Pavelka
M
Marjorie Fournier
S
Selene K. Swanson
M
Mattia Pelizzola
P
Paola Ricciardi‐Castagnoli
L
Laurence Florens
M
Michael P. Washburn *
DOI:10.1074/mcp.M700240-MCP200delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
If the large collection of microarray-specific statistical tools was applicable to the analysis of quantitative shotgun proteomics datasets, it would certainly foster an important advancement of proteomics research. Here we analyze two large multidimensional protein identification technology datasets, one containing eight replicates of the soluble fraction of a yeast whole-cell lysate and one containing nine replicates of a human immunoprecipitate, to test whether normalized spectral abundance factor (NSAF) values share substantially similar statistical properties with transcript abundance values from Affymetrix GeneChip data. First we show similar dynamic range and distribution properties of these two types of numeric values. Next we show that the standard deviation (S.D.) of a protein's NSAF values was dependent on the average NSAF value of the protein itself, following a power law. This relationship can be modeled by a power law global error model (PLGEM), initially developed to describe the variance-versus-mean dependence that exists in GeneChip data. PLGEM parameters obtained from NSAF datasets proved to be surprisingly similar to the typical parameters observed in GeneChip datasets. The most important common feature identified by this approach was that, although in absolute terms the S.D. of replicated abundance values increases as a function of increasing average abundance, the coefficient of variation, a relative measure of variability, becomes progressively smaller under the same conditions. We next show that PLGEM parameters were reasonably stable to decreasing numbers of replicates. We finally illustrate one possible application of PLGEM in the identification of differentially abundant proteins that might potentially outperform standard statistical tests. In summary, we believe that this body of work lays the foundation for the application of microarray-specific tools in the analysis of NSAF datasets.
Keyword:
DIFFERENTIALLY-EXPRESSED GENES
PROTEIN IDENTIFICATION TECHNOLOGY
MISSING VALUE ESTIMATION
CDNA MICROARRAY EXPERIMENTS
MASS-SPECTROMETRY
SACCHAROMYCES-CEREVISIAE
MEASUREMENT ERROR
STATIONARY-PHASE
SYSTEMS BIOLOGY
MODEL
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

M
Molecular and Cellular Proteomics
IF:
5.5
论文数:
4.8K
被引数:
1.7W

机构

U
university of milano-bicocca
学者数:
2.0W
论文数: 1.5W
被引数: 22
A
agency for science technology & research (a*star)
学者数:
2.2W
论文数: 1.9W
被引数: 57
Stowers Institute for Medical Research 封面图
Stowers Institute for Medical Research
学者数:
1.2K
论文数: 894
被引数: 1.3K
学者 查看更多机构