Development of a Reference Standard Library of Chloroplast Genome Sequences, GenomeTrakrCP
received 07 March 2017
revised 21 May 2017
accepted 05 June 2017
26 June 2017 (eFirst)
Precise, species-level identification of plants in foods and dietary supplements is difficult. While the use of DNA barcoding regions (short regions of DNA with diagnostic utility) has been effective for many inquiries, it is not always a robust approach for closely related species, especially in highly processed products. The use of fully sequenced chloroplast genomes, as an alternative to short diagnostic barcoding regions, has demonstrated utility for closely related species. The U. S. Food and Drug Administration (FDA) has also developed species-specific DNA-based assays targeting plant species of interest by utilizing chloroplast genome sequences. Here, we introduce a repository of complete chloroplast genome sequences called GenomeTrakrCP, which will be publicly available at the National Center for Biotechnology Information (NCBI). Target species for inclusion are plants found in foods and dietary supplements, toxin producers, common contaminants and adulterants, and their close relatives. Publicly available data will include annotated assemblies, raw sequencing data, and voucher information with each NCBI accession associated with an authenticated reference herbarium specimen. To date, 40 complete chloroplast genomes have been deposited in GenomeTrakrCP (https://www.ncbi.nlm.nih.gov/bioproject/PRJNA325670/), and this will be expanded in the future.