A critical assessment of storytelling: Gene ontology categories and the importance of validating genomic scans

Pavlos Pavlidis, Jeffrey D. Jensen, Wolfgang Stephan, Alexandros Stamatakis

Research output: Contribution to journalArticle

123 Scopus citations

Abstract

In the age of whole-genome population genetics, so-called genomic scan studies often conclude with a long list of putatively selected loci. These lists are then further scrutinized to annotate these regions by gene function, corresponding biological processes, expression levels, or gene networks. Such annotations are often used to assess and/or verify the validity of the genome scan and the statistical methods that have been used to perform the analyses. Furthermore, these results are frequently considered to validate true-positives if the identified regions make biological sense a posteriori. Here, we show that this approach can be potentially misleading. By simulating neutral evolutionary histories, we demonstrate that it is possible not only to obtain an extremely high false-positive rate but also to make biological sense out of the false-positives and construct a sensible biological narrative. Results are compared with a recent polymorphism data set from Drosophila melanogaster.

Original languageEnglish (US)
Pages (from-to)3237-3248
Number of pages12
JournalMolecular biology and evolution
Volume29
Issue number10
DOIs
StatePublished - Oct 2012

Keywords

  • gene ontology
  • genome scanning
  • literature mining
  • positive selection
  • validation

ASJC Scopus subject areas

  • Ecology, Evolution, Behavior and Systematics
  • Molecular Biology
  • Genetics

Fingerprint Dive into the research topics of 'A critical assessment of storytelling: Gene ontology categories and the importance of validating genomic scans'. Together they form a unique fingerprint.

  • Cite this