ADDA: a domain database with global coverage of the protein universe
Author(s) -
Andreas Heger,
Christopher Wilton,
Ashwin Sivakumar,
Liisa Holm
Publication year - 2004
Publication title -
nucleic acids research
Language(s) - English
Resource type - Journals
SCImago Journal Rank - 9.008
H-Index - 537
eISSN - 1362-4954
pISSN - 0305-1048
DOI - 10.1093/nar/gki096
Subject(s) - biology , domain (mathematical analysis) , database , protein domain , upload , computational biology , interface (matter) , protein–protein interaction , bioinformatics , computer science , genetics , world wide web , mathematical analysis , mathematics , gene , pulmonary surfactant , biochemistry , gibbs isotherm
We used the Automatic Domain Decomposition Algorithm (ADDA) to generate a database of protein domain families with complete coverage of all protein sequences. Sequences are split into domains and domains are grouped into protein domain families in a completely automated process. The current database contains domains for more than 1.5 million sequences in more than 40,000 domain families. In particular, there are 3828 novel domain families that do not overlap with the curated domain databases Pfam, SCOP and InterPro. The data are freely available for downloading and querying via a web interface (http://ekhidna.biocenter.helsinki.fi:9801/sqgraph/pairsdb).
Accelerating Research
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom
Address
John Eccles HouseRobert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom