Skip to Main content Skip to Navigation
Conference papers

Associating Gene Ontology Terms with Pfam Protein Domains

Seyed Ziaeddin Alborzi 1 Marie-Dominique Devignes 1 David Ritchie 1
1 CAPSID - Computational Algorithms for Protein Structures and Interactions
Inria Nancy - Grand Est, LORIA - AIS - Department of Complex Systems, Artificial Intelligence & Robotics
Abstract : With the growing number of three-dimensional protein structures in the protein data bank (PDB), there is a need to annotate these structures at the domain level in order to relate protein structure to protein function. Thanks to the SIFTS database, many PDB chains are now cross-referenced with Pfam domains and Gene ontology (GO) terms. However, these annotations do not include any explicit relationship between individual Pfam domains and GO terms. Therefore, creating a direct mapping between GO terms and Pfam domains will provide a new and more detailed level of protein structure annotation. This article presents a novel content-based filtering method called GODM that can automatically infer associations between GO terms and Pfam domains directly from existing GO-chain/Pfam-chain associations from the SIFTS database and GO-sequence/Pfam-sequence associations from the UniProt databases. Overall, GODM finds a total of 20,318 non-redundant GO-Pfam associations with a F-measure of 0.98 with respect to the InterPro database, which is treated here as a “Gold Standard”. These associations could be used to annotate thousands of PDB chains or protein sequences for which their domain composition is known but which currently lack any GO annotation. The GODM database is publicly available at http://godm.loria.fr/
Document type :
Conference papers
Complete list of metadata

Cited literature [18 references]  Display  Hide  Download

https://hal.inria.fr/hal-01531204
Contributor : David Ritchie <>
Submitted on : Friday, June 2, 2017 - 1:49:58 PM
Last modification on : Sunday, January 3, 2021 - 5:00:03 PM
Long-term archiving on: : Wednesday, December 13, 2017 - 10:14:34 AM

File

godm_02_feb_2017.pdf
Files produced by the author(s)

Identifiers

Collections

Citation

Seyed Ziaeddin Alborzi, Marie-Dominique Devignes, David Ritchie. Associating Gene Ontology Terms with Pfam Protein Domains. 5th International Work-Conference on Bioinformatics and Biomedical Engineering - IWBBIO 2017, Apr 2017, Granada, Spain. pp.127-138, ⟨10.1007/978-3-319-56154-7_13⟩. ⟨hal-01531204⟩

Share

Metrics

Record views

446

Files downloads

1050