Sequencing and Comparative Analysis of the Chloroplast Genome of Angelica polymorpha and the Development of a Novel Indel Marker for Species Identification.
Inkyu ParkSungyu YangWook Jin KimJun-Ho SongHyun-Sook LeeHyun Oh LeeJung-Hyun LeeSang-Nag AhnByeong-Cheol MoonPublished in: Molecules (Basel, Switzerland) (2019)
The genus Angelica (Apiaceae) comprises valuable herbal medicines. In this study, we determined the complete chloroplast (CP) genome sequence of A. polymorpha and compared it with that of Ligusticum officinale (GenBank accession no. NC039760). The CP genomes of A. polymorpha and L. officinale were 148,430 and 147,127 bp in length, respectively, with 37.6% GC content. Both CP genomes harbored 113 unique functional genes, including 79 protein-coding, four rRNA, and 30 tRNA genes. Comparative analysis of the two CP genomes revealed conserved genome structure, gene content, and gene order. However, highly variable regions, sufficient to distinguish between A. polymorpha and L. officinale, were identified in hypothetical chloroplast open reading frame1 (ycf1) and ycf2 genic regions. Nucleotide diversity (Pi) analysis indicated that ycf4⁻chloroplast envelope membrane protein (cemA) intergenic region was highly variable between the two species. Phylogenetic analysis revealed that A. polymorpha and L. officinale were well clustered at family Apiaceae. The ycf4-cemA intergenic region in A. polymorpha carried a 418 bp deletion compared with L. officinale. This region was used for the development of a novel indel marker, LYCE, which successfully discriminated between A. polymorpha and L. officinale accessions. Our results provide important taxonomic and phylogenetic information on herbal medicines and facilitate their authentication using the indel marker.