Toward completion of the Earth’s proteome: an update a decade later
Explore this paper's citation graph
Summary
It is proposed that new sequencing projects can be made more useful if they are driven to sequencing voids, parts of the tree of life far from already sequenced species or model organisms, and these voids are present in the Archaea and Eukarya domains of life.
- Type
- article
- Published
- 2019-03-25
- Cited by
- 3
- References
- 21
- OpenAlex
- https://openalex.org/W2762751282
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:24491953
Keywords
UniProt, Redundancy (engineering), Protein sequencing, Sequence database, Protein superfamily
References
- How Many Species Are There on Earth and in the Ocean?
- New approaches narrow global species estimates for beetles, insects, and terrestrial arthropods
- Emerging roles of pseudokinases.
- Towards completion of the Earth's proteome
- Biodiversity. 8.7 million: a new estimate for all the complex species on Earth.
- The others: our biased perspective of eukaryotic genomes
- Twilight zone of protein sequence alignments.
- UniRef clusters: a comprehensive and scalable alternative for improving sequence similarity searches
- Practical limits of function prediction
- Insights from 20 years of bacterial genome sequencing
- UniProt: the Universal Protein knowledgebase
- FastaHerder2: Four Ways to Research Protein Function and Evolution with Clustering and Clustered Databases
- Genomes OnLine Database (GOLD) v.6: data updates and feature enhancements
- Uniclust databases of clustered and deeply annotated protein sequences and alignments
- Minimizing proteome redundancy in the UniProt Knowledgebase
- Asgard archaea illuminate the origin of eukaryotic cellular complexity
- A Review of Bioinformatics Tools for Bio-Prospecting from Metagenomic Sequence Data
- The secret life of kinases: insights into non-catalytic signalling functions from pseudokinases.
- Duplicates, redundancies and inconsistencies in the primary nucleotide databases: a descriptive study
- Database resources of the National Center for Biotechnology Information
Cited by
Related papers
- Uniclust databases of clustered and deeply annotated protein sequences and alignments
- Protein sequence databases.
- A database of unique protein sequence identifiers for proteome studies
- The Universal Protein Resource (UniProt)
- The SWISS-PROT protein sequence database and its supplement TrEMBL in 2000
- The SWISS-PROT protein sequence data bank and its new supplement TREMBL
- The role SWISS-PROT and TrEMBL play in the genome research environment.
- ProGMap: an integrated annotation resource for protein orthology