دورية أكاديمية

Taxonomic identification accuracy from BOLD and GenBank databases using over a thousand insect DNA barcodes from Colombia

التفاصيل البيبلوغرافية
العنوان: Taxonomic identification accuracy from BOLD and GenBank databases using over a thousand insect DNA barcodes from Colombia
المؤلفون: Baena-Bejarano, Nathalie, Reina, Catalina, Martínez-Revelo, Diego Esteban, Medina, Claudia A., Tovar, Eduardo, Uribe-Soto, Sandra, Neita-Moreno, Jhon Cesar, Gonzalez, Mailyn A.
المساهمون: ZHANG, Feng, Minciencias, Colombia BIO Agreement, Santander BIO inter-agency Agreement
المصدر: PLOS ONE ; volume 18, issue 4, page e0277379 ; ISSN 1932-6203
بيانات النشر: Public Library of Science (PLoS)
سنة النشر: 2023
المجموعة: PLOS Publications (via CrossRef)
الوصف: Recent declines of insect populations at high rates have resulted in the need to develop a quick method to determine their diversity and to process massive data for the identification of species of highly diverse groups. A short sequence of DNA from COI is widely used for insect identification by comparing it against sequences of known species. Repositories of sequences are available online with tools that facilitate matching of the sequences of interest to a known individual. However, the performance of these tools can differ. Here we aim to assess the accuracy in identification of insect taxonomic categories from two repositories, BOLD Systems and GenBank. This was done by comparing the sequence matches between the taxonomist identification and the suggested identification from the platforms. We used 1,160 COI sequences representing eight orders of insects from Colombia. After the comparison, we reanalyzed the results from a representative subset of the data from the subfamily Scarabaeinae (Coleoptera). Overall, BOLD systems outperformed GenBank, and the performance of both engines differed by orders and other taxonomic categories (species, genus and family). Higher rates of accurate identification were obtained at family and genus levels. The accuracy was higher in BOLD for the order Coleoptera at family level, for Coleoptera and Lepidoptera at genus and species level. Other orders performed similarly in both repositories. Moreover, the Scarabaeinae subset showed that species were correctly identified only when BOLD match percentage was above 93.4% and a total of 85% of the samples were correctly assigned to a taxonomic category. These results accentuate the great potential of the identification engines to place insects accurately into their respective taxonomic categories based on DNA barcodes and highlight the reliability of BOLD Systems for insect identification in the absence of a large reference database for a highly diverse country.
نوع الوثيقة: article in journal/newspaper
اللغة: English
DOI: 10.1371/journal.pone.0277379
الإتاحة: https://doi.org/10.1371/journal.pone.0277379Test
حقوق: http://creativecommons.org/licenses/by/4.0Test/
رقم الانضمام: edsbas.14FCC2BC
قاعدة البيانات: BASE