CRISPR sequences are sometimes erroneously translated and can contaminate public databases with spurious proteins containing spaced repeats. (18th November 2020)
- Record Type:
- Journal Article
- Title:
- CRISPR sequences are sometimes erroneously translated and can contaminate public databases with spurious proteins containing spaced repeats. (18th November 2020)
- Main Title:
- CRISPR sequences are sometimes erroneously translated and can contaminate public databases with spurious proteins containing spaced repeats
- Authors:
- Rubio, Alejandro
Mier, Pablo
Andrade-Navarro, Miguel A
Garzón, Andrés
Jiménez, Juan
Pérez-Pulido, Antonio J - Abstract:
- Abstract: The genomics era is resulting in the generation of a plethora of biological sequences that are usually stored in public databases. There are many computational tools that facilitate the annotation of these sequences, but sometimes they produce mistakes that enter the databases and can be propagated when erroneous data are used for secondary analyses, such as gene prediction or homology searching. While developing a computational gene finder based on protein-coding sequences, we discovered that the reference UniProtKB protein database is contaminated with some spurious sequences translated from DNA containing clustered regularly interspaced short palindromic repeats. We therefore encourage developers of prokaryotic computational gene finders and protein database curators to consider this source of error.
- Is Part Of:
- Database. Volume 2020(2020)
- Journal:
- Database
- Issue:
- Volume 2020(2020)
- Issue Display:
- Volume 2020, Issue 2020 (2020)
- Year:
- 2020
- Volume:
- 2020
- Issue:
- 2020
- Issue Sort Value:
- 2020-2020-2020-0000
- Page Start:
- Page End:
- Publication Date:
- 2020-11-18
- Subjects:
- Biology -- Databases -- Periodicals
Bioinformatics -- Periodicals
570.285 - Journal URLs:
- http://database.oxfordjournals.org/ ↗
http://ukcatalogue.oup.com/ ↗ - DOI:
- 10.1093/database/baaa088 ↗
- Languages:
- English
- ISSNs:
- 1758-0463
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 26038.xml