Title
ORegAnno: an open-access community-driven resource for regulatory annotation
Abstract
ORegAnno is an open-source, open-access database and literature curation system for community-based annotation of experimentally identified DNA regulatory regions, transcription factor binding sites and regulatory variants. The current release comprises 30 145 records curated from 922 publications and describing regulatory sequences for over 3853 genes and 465 transcription factors from 19 species. A new feature called the publication queue allows users to input relevant papers from scientific literature as targets for annotation. The queue contains 4438 gene regulation papers entered by experts and another 54 351 identified by text-mining methods. Users can enter or check out papers from the queue for manual curation using a series of user-friendly annotation pages. A typical record entry consists of species, sequence type, sequence, target gene, binding factor, experimental outcome and one or more lines of experimental evidence. An evidence ontology was developed to describe and categorize these experiments. Records are cross-referenced to Ensembl or Entrez gene identifiers, PubMed and dbSNP and can be visualized in the Ensembl or UCSC genome browsers. All data are freely available through search pages, XML data dumps or web services at: http://www.oreganno.org.
Year
DOI
Venue
2008
10.1093/nar/gkm967
NUCLEIC ACIDS RESEARCH
Keywords
Field
DocType
gene regulation,transcription factor binding site,binding sites,text mining,transcription factors,transcription factor,medicine,web service,internet
Annotation,Information retrieval,XML,DNA binding site,Biology,Ensembl,Queue,dbSNP,Entrez Gene,Bioinformatics,Genetics,Web service
Journal
Volume
Issue
ISSN
36
Database-I
0305-1048
Citations 
PageRank 
References 
49
3.04
13
Authors
29