دورية أكاديمية

Taxonomic identification accuracy from BOLD and GenBank databases using over a thousand insect DNA barcodes from Colombia.

التفاصيل البيبلوغرافية
العنوان: Taxonomic identification accuracy from BOLD and GenBank databases using over a thousand insect DNA barcodes from Colombia.
المؤلفون: Baena-Bejarano, Nathalie, Reina, Catalina, Martínez-Revelo, Diego Esteban, Medina, Claudia A., Tovar, Eduardo, Uribe-Soto, Sandra, Neita-Moreno, Jhon Cesar, Gonzalez, Mailyn A.
المصدر: PLoS ONE; 4/24/2023, Vol. 17 Issue 4, p1-19, 19p
مصطلحات موضوعية: IDENTIFICATION, INSECTS, DNA data banks, INSECT populations, DATABASES, SYSTEM identification, BEETLES, DNA
مصطلحات جغرافية: COLOMBIA
مستخلص: Recent declines of insect populations at high rates have resulted in the need to develop a quick method to determine their diversity and to process massive data for the identification of species of highly diverse groups. A short sequence of DNA from COI is widely used for insect identification by comparing it against sequences of known species. Repositories of sequences are available online with tools that facilitate matching of the sequences of interest to a known individual. However, the performance of these tools can differ. Here we aim to assess the accuracy in identification of insect taxonomic categories from two repositories, BOLD Systems and GenBank. This was done by comparing the sequence matches between the taxonomist identification and the suggested identification from the platforms. We used 1,160 COI sequences representing eight orders of insects from Colombia. After the comparison, we reanalyzed the results from a representative subset of the data from the subfamily Scarabaeinae (Coleoptera). Overall, BOLD systems outperformed GenBank, and the performance of both engines differed by orders and other taxonomic categories (species, genus and family). Higher rates of accurate identification were obtained at family and genus levels. The accuracy was higher in BOLD for the order Coleoptera at family level, for Coleoptera and Lepidoptera at genus and species level. Other orders performed similarly in both repositories. Moreover, the Scarabaeinae subset showed that species were correctly identified only when BOLD match percentage was above 93.4% and a total of 85% of the samples were correctly assigned to a taxonomic category. These results accentuate the great potential of the identification engines to place insects accurately into their respective taxonomic categories based on DNA barcodes and highlight the reliability of BOLD Systems for insect identification in the absence of a large reference database for a highly diverse country. [ABSTRACT FROM AUTHOR]
Copyright of PLoS ONE is the property of Public Library of Science and its content may not be copied or emailed to multiple sites or posted to a listserv without the copyright holder's express written permission. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
قاعدة البيانات: Complementary Index
الوصف
تدمد:19326203
DOI:10.1371/journal.pone.0277379