Back to Search Start Over

Pfam: the protein families database.

Authors :
Finn RD
Bateman A
Clements J
Coggill P
Eberhardt RY
Eddy SR
Heger A
Hetherington K
Holm L
Mistry J
Sonnhammer EL
Tate J
Punta M
Source :
Nucleic acids research [Nucleic Acids Res] 2014 Jan; Vol. 42 (Database issue), pp. D222-30. Date of Electronic Publication: 2013 Nov 27.
Publication Year :
2014

Abstract

Pfam, available via servers in the UK (http://pfam.sanger.ac.uk/) and the USA (http://pfam.janelia.org/), is a widely used database of protein families, containing 14 831 manually curated entries in the current release, version 27.0. Since the last update article 2 years ago, we have generated 1182 new families and maintained sequence coverage of the UniProt Knowledgebase (UniProtKB) at nearly 80%, despite a 50% increase in the size of the underlying sequence database. Since our 2012 article describing Pfam, we have also undertaken a comprehensive review of the features that are provided by Pfam over and above the basic family data. For each feature, we determined the relevance, computational burden, usage statistics and the functionality of the feature in a website context. As a consequence of this review, we have removed some features, enhanced others and developed new ones to meet the changing demands of computational biology. Here, we describe the changes to Pfam content. Notably, we now provide family alignments based on four different representative proteome sequence data sets and a new interactive DNA search interface. We also discuss the mapping between Pfam and known 3D structures.

Details

Language :
English
ISSN :
1362-4962
Volume :
42
Issue :
Database issue
Database :
MEDLINE
Journal :
Nucleic acids research
Publication Type :
Academic Journal
Accession number :
24288371
Full Text :
https://doi.org/10.1093/nar/gkt1223