Assessing DNA Barcoding as a Tool for Species Identification and Data Quality Control
Open Access
- 19 February 2013
- journal article
- research article
- Published by Public Library of Science (PLoS) in PLOS ONE
- Vol. 8 (2), e57125
- https://doi.org/10.1371/journal.pone.0057125
Abstract
In recent years, the number of sequences of diverse species submitted to GenBank has grown explosively and not infrequently the data contain errors. This problem is extensively recognized but not for invalid or incorrectly identified species, sample mixed-up, and contamination. DNA barcoding is a powerful tool for identifying and confirming species and one very important application involves forensics. In this study, we use DNA barcoding to detect erroneous sequences in GenBank by evaluating deep intraspecific and shallow interspecific divergences to discover possible taxonomic problems and other sources of error. We use the mitochondrial DNA gene encoding cytochrome b (Cytb) from turtles to test the utility of barcoding for pinpointing potential errors. This gene is widely used in phylogenetic studies of the speciose group. Intraspecific variation is usually less than 2.0% and in most cases it is less than 1.0%. In comparison, most species differ by more than 10.0% in our dataset. Overlapping intra- and interspecific percentages of variation mainly involve problematic identifications of species and outdated taxonomies. Further, we detect identical problems in Cytb from Insectivora and Chiroptera. Upon applying this strategy to 47,524 mammalian CoxI sequences, we resolve a suite of potentially problematic sequences. Our study reveals that erroneous sequences are not rare in GenBank and that the DNA barcoding can serve to confirm sequencing accuracy and discover problems such as misidentified species, inaccurate taxonomies, contamination, and potential errors in sequencing.Keywords
This publication has 47 references indexed in Scilit:
- Assessment of Three Mitochondrial Genes (16S, Cytb, CO1) for Identifying Species in the Praomyini Tribe (Rodentia: Muridae)PLOS ONE, 2012
- The dazed and confused identity of Agassiz’s land tortoise, Gopherus agassizii (Testudines: Testudinidae) with the description of a new species and its consequences for conservationZooKeys, 2011
- Parallelization of the MAFFT multiple sequence alignment programBioinformatics, 2010
- Application of DNA Bar Codes for Screening of Industrially Important Fungi: the Haplotype of Trichoderma harzianum Sensu Stricto Indicates Superior Chitinase FormationApplied and Environmental Microbiology, 2007
- Comprehensive DNA barcode coverage of North American birdsMolecular Ecology Notes, 2007
- Prospects for fungus identification usingCO1DNA barcodes, withPenicilliumas a test caseProceedings of the National Academy of Sciences of the United States of America, 2007
- Assessing the effect of varying sequence length on DNA barcoding of fungiMolecular Ecology Notes, 2007
- Identification of Birds through DNA BarcodesPLoS Biology, 2004
- Modernizing the Tree of LifeScience, 2003
- A simple method for estimating evolutionary rates of base substitutions through comparative studies of nucleotide sequencesJournal of Molecular Evolution, 1980