An investigation into the population abundance distribution of mRNAs, proteins, and metabolites in biological systems

Chuan Lu; Ross D King

doi:10.1093/bioinformatics/btp360

An investigation into the population abundance distribution of mRNAs, proteins, and metabolites in biological systems

Bioinformatics. 2009 Aug 15;25(16):2020-7. doi: 10.1093/bioinformatics/btp360. Epub 2009 Jun 17.

Authors

Chuan Lu¹, Ross D King

Affiliation

¹ Department of Computer Science, Aberystwyth University, Ceredigion SY23 3DB, UK.

PMID: 19535531
DOI: 10.1093/bioinformatics/btp360

Abstract

Motivation: Distribution analysis is one of the most basic forms of statistical analysis. Thanks to improved analytical methods, accurate and extensive quantitative measurements can now be made of the mRNA, protein and metabolite from biological systems. Here, we report a large-scale analysis of the population abundance distributions of the transcriptomes, proteomes and metabolomes from varied biological systems.

Results: We compared the observed empirical distributions with a number of distributions: power law, lognormal, loglogistic, loggamma, right Pareto-lognormal (PLN) and double PLN (dPLN). The best-fit for mRNA, protein and metabolite population abundance distributions was found to be the dPLN. This distribution behaves like a lognormal distribution around the centre, and like a power law distribution in the tails. To better understand the cause of this observed distribution, we explored a simple stochastic model based on geometric Brownian motion. The distribution indicates that multiplicative effects are causally dominant in biological systems. We speculate that these effects arise from chemical reactions: the central-limit theorem then explains the central lognormal, and a number of possible mechanisms could explain the long tails: positive feedback, network topology, etc. Many of the components in the central lognormal parts of the empirical distributions are unidentified and/or have unknown function. This indicates that much more biology awaits discovery.

Publication types

Research Support, Non-U.S. Gov't

MeSH terms

Computer Simulation
Metabolome*
Models, Biological
Proteins / chemistry
Proteins / metabolism*
RNA, Messenger / metabolism*

Substances

Proteins
RNA, Messenger

Grants and funding

BB/F008228/1/BB_/Biotechnology and Biological Sciences Research Council/United Kingdom