HAL will be down for maintenance from Friday, June 10 at 4pm through Monday, June 13 at 9am. More information
Skip to Main content Skip to Navigation
Conference papers

Associating Gene Ontology Terms with Pfam Protein Domains

Seyed Ziaeddin Alborzi 1 Marie-Dominique Devignes 1 David Ritchie 1
1 CAPSID - Computational Algorithms for Protein Structures and Interactions
Inria Nancy - Grand Est, LORIA - AIS - Department of Complex Systems, Artificial Intelligence & Robotics
Abstract : With the growing number of three-dimensional protein structures in the protein data bank (PDB), there is a need to annotate these structures at the domain level in order to relate protein structure to protein function. Thanks to the SIFTS database, many PDB chains are now cross-referenced with Pfam domains and Gene ontology (GO) terms. However, these annotations do not include any explicit relationship between individual Pfam domains and GO terms. Therefore, creating a direct mapping between GO terms and Pfam domains will provide a new and more detailed level of protein structure annotation. This article presents a novel content-based filtering method called GODM that can automatically infer associations between GO terms and Pfam domains directly from existing GO-chain/Pfam-chain associations from the SIFTS database and GO-sequence/Pfam-sequence associations from the UniProt databases. Overall, GODM finds a total of 20,318 non-redundant GO-Pfam associations with a F-measure of 0.98 with respect to the InterPro database, which is treated here as a “Gold Standard”. These associations could be used to annotate thousands of PDB chains or protein sequences for which their domain composition is known but which currently lack any GO annotation. The GODM database is publicly available at http://godm.loria.fr/
Document type :
Conference papers
Complete list of metadata

Cited literature [18 references]  Display  Hide  Download

Contributor : David Ritchie Connect in order to contact the contributor
Submitted on : Friday, June 2, 2017 - 1:49:58 PM
Last modification on : Thursday, April 28, 2022 - 3:11:31 AM
Long-term archiving on: : Wednesday, December 13, 2017 - 10:14:34 AM


Files produced by the author(s)




Seyed Ziaeddin Alborzi, Marie-Dominique Devignes, David Ritchie. Associating Gene Ontology Terms with Pfam Protein Domains. 5th International Work-Conference on Bioinformatics and Biomedical Engineering - IWBBIO 2017, Apr 2017, Granada, Spain. pp.127-138, ⟨10.1007/978-3-319-56154-7_13⟩. ⟨hal-01531204⟩



Record views


Files downloads