PA-GOSUB: A searchable database of model organism protein sequences with their predicted GO molecular function and subcellular localization

  • Author(s) / Creator(s)
  • PA-GOSUB (Proteome Analyst: Gene Ontology Molecular Function and Subcellular Localization) is a publicly available, web-based, searchable and downloadable database that contains the sequences, predicted GO molecular functions and predicted subcellular localizations of more than 107 000 proteins from 10 model organisms (and growing), covering the major kingdoms and phyla for which annotated proteomes exist ( PA/GOSUB). The PA-GOSUB database effectively expands the coverage of subcellular localization and GO function annotations by a significant factor (already over five for subcellular localization, compared with Swiss-Prot v42.7), and more model organisms are being added to PA-GOSUB as their sequenced proteomes become available. PA-GOSUB can be used in three main ways. First, a researcher can browse the pre-computed PA-GOSUB annotations on a per-organism and per-protein basis using annotation-based and text-based filters. Second, a user can perform BLAST searches against the PA-GOSUB database and use the annotations from the homologs as simple predictors for the new sequences. Third, the whole of PA-GOSUB can be downloaded in either FASTA or comma-separated values (CSV) formats.

  • Date created
  • Subjects / Keywords
  • Type of Item
    Article (Published)
  • DOI
  • License
    The online version of this article has been published under an open access model. Users are entitled to use, reproduce, disseminate, or display the open access version of this article for non-commercial purposes provided that: the original authorship is properly and fully attributed; the Journal and Oxford University Press are attributed as the original place of publication with the correct citation details given; if an article is subsequently reproduced or disseminated not in its entirety but only in part or as a derivative work this must be clearly indicated. For commercial re-use permissions, please contact Copyright 2005, the authors
  • Language
  • Citation for previous publication
    • P Lu, D Szafron, R Greiner, DS Wishart, A Fyshe, B Pearcy, B Poulin, R Eisner, D Ngo and N Lamb. "PA-GOSUB: A searchable database of model organism protein sequences with their predicted GO molecular function and subcellular localization." Nucleic Acids Research 33 Database Issue (2005): D147-153.