ORegAnno: an open-access community-driven resource for regulatory annotation

Nucleic Acids Res. 2008 Jan;36(Database issue):D107-13. doi: 10.1093/nar/gkm967. Epub 2007 Nov 15.

Abstract

ORegAnno is an open-source, open-access database and literature curation system for community-based annotation of experimentally identified DNA regulatory regions, transcription factor binding sites and regulatory variants. The current release comprises 30 145 records curated from 922 publications and describing regulatory sequences for over 3853 genes and 465 transcription factors from 19 species. A new feature called the 'publication queue' allows users to input relevant papers from scientific literature as targets for annotation. The queue contains 4438 gene regulation papers entered by experts and another 54 351 identified by text-mining methods. Users can enter or 'check out' papers from the queue for manual curation using a series of user-friendly annotation pages. A typical record entry consists of species, sequence type, sequence, target gene, binding factor, experimental outcome and one or more lines of experimental evidence. An evidence ontology was developed to describe and categorize these experiments. Records are cross-referenced to Ensembl or Entrez gene identifiers, PubMed and dbSNP and can be visualized in the Ensembl or UCSC genome browsers. All data are freely available through search pages, XML data dumps or web services at: http://www.oreganno.org.

Publication types

  • Research Support, Non-U.S. Gov't

MeSH terms

  • Access to Information
  • Animals
  • Binding Sites
  • Databases, Nucleic Acid*
  • Humans
  • Internet
  • Regulatory Elements, Transcriptional*
  • Transcription Factors / metabolism*
  • User-Computer Interface

Substances

  • Transcription Factors