A high-quality genome assembly of Morinda officinalis, a famous native southern herb in the Lingnan region of southern China

Jihua Wang , Shiqiang Xu , Yu Mei , Shike Cai , Yan Gu , Minyang Sun , Zhan Liang , Yong Xiao , Muqing Zhang , Shaohai Yang

Horticulture Research ›› 2021, Vol. 8 ›› Issue (1) : 135

PDF (2556KB)
Horticulture Research ›› 2021, Vol. 8 ›› Issue (1) :135 DOI: 10.1038/s41438-021-00551-w
Article
research-article
A high-quality genome assembly of Morinda officinalis, a famous native southern herb in the Lingnan region of southern China
Author information +
History +
PDF (2556KB)

Abstract

Morinda officinalis is a well-known medicinal and edible plant that is widely cultivated in the Lingnan region of southern China. Its dried roots (called bajitian in traditional Chinese medicine) are broadly used to treat various diseases, such as impotence and rheumatism. Here, we report a high-quality chromosome-scale genome assembly of M. officinalis using Nanopore single-molecule sequencing and Hi-C technology. The assembled genome size was 484.85 Mb with a scaffold N50 of 40.97 Mb, and 90.77% of the assembled sequences were anchored on eleven pseudochromosomes. The genome includes 27,698 protein-coding genes, and most of the assemblies are repetitive sequences. Genome evolution analysis revealed that M. officinalis underwent core eudicot γ genome triplication events but no recent whole-genome duplication (WGD). Likewise, comparative genomic analysis showed no large-scale structural variation after species divergence between M. officinalis and Coffea canephora. Moreover, gene family analysis indicated that gene families associated with plant–pathogen interactions and sugar metabolism were significantly expanded in M. officinalis. Furthermore, we identified many candidate genes involved in the biosynthesis of major active components such as anthraquinones, iridoids and polysaccharides. In addition, we also found that the DHQS, GGPPS, TPS-Clin, TPS04, sacA, and UGDH gene families-which include the critical genes for active component biosynthesis-were expanded in M. officinalis. This study provides a valuable resource for understanding M. officinalis genome evolution and active component biosynthesis. This work will facilitate genetic improvement and molecular breeding of this commercially important plant.

Cite this article

Download citation ▾
Jihua Wang, Shiqiang Xu, Yu Mei, Shike Cai, Yan Gu, Minyang Sun, Zhan Liang, Yong Xiao, Muqing Zhang, Shaohai Yang. A high-quality genome assembly of Morinda officinalis, a famous native southern herb in the Lingnan region of southern China. Horticulture Research, 2021, 8 (1) : 135 DOI:10.1038/s41438-021-00551-w

登录浏览全文

4963

注册一个新账户 忘记密码

References

[1]

Lim, T. K. Edible Medicinal and Non-Medicinal Plants Vol. 5 (Springer, 2013).

[2]

Zhang, J. H. et al. Morinda officinalis how-a comprehensive review of traditional uses, phytochemistry and pharmacology. J. Ethnopharmacol. 213, 230-255 (2018).

[3]

Wu, Z. Q. et al. Effect of bajijiasu isolated from Morinda officinalis FC how on sexual function in male mice and its antioxidant protection of human sperm. J. Ethnopharmacol. 164, 283-292 (2015).

[4]

Bao, L. et al. Anthraquinone compounds from Morinda officinalis inhibit osteoclastic bone resorption in vitro. Chem. Biol. Interact. 194, 97-105 (2011).

[5]

Zhao, X., Wei, J. & Yang, M. Simultaneous analysis of iridoid glycosides and anthraquinones in Morinda officinalis using UPLC-QqQ-MS/MS and UPLC-Q/TOF-MSE. Molecules 23, 1070 (2018).

[6]

Yong, J., Lu, C., Huang, S. & Wu, X. Chemical components isolated from the roots of Morinda officinalis. Chem. Nat. Compd. 51, 548-549 (2015).

[7]

Malik, E. M. & Müller, C. E. Anthraquinones as pharmacological tools and drugs. Med. Res. Rev. 36, 705-748 (2016).

[8]

PHAKHODEE, W. Distribution of naturally occurring anthraquinones, iridoids and flavonoids from Morinda genus: chemistry and biological activity. Walailak J. Sci. Technol. 9, 173-188 (2012).

[9]

Choi, J. et al. Antinociceptive anti-inflammatory effect of monotropein isolated from the root of Morinda officinalis. Biol. Pharm. Bull. 28, 1915-1918 (2005).

[10]

Zhang, H. L. et al. Structural characterization and anti-fatigue activity of polysaccharides from the roots of Morinda officinalis. Int. J. Biol. Macromol. 44, 257-261 (2009).

[11]

Zhu, M. et al. Extraction of polysaccharides from Morinda officinalis by response surface methodology and effect of the polysaccharides on bone-related genes. Carbohydr. Polym. 85, 23-28 (2011).

[12]

Li, Y. F., Gong, Z. H., Yang, M., Zhao, Y. M. & Luo, Z. P. Inhibition of the oligosaccharides extracted from Morinda officinalis, a Chinese traditional herbal medicine, on the corticosterone induced apoptosis in PC12 cells. Life Sci. 72, 933-942 (2003).

[13]

Yan, C. et al. Identification and characterization of a polysaccharide from the roots of Morinda officinalis, as an inducer of bone formation by up-regulation of target gene expression. Int. J. Biol. Macromol. 133, 446-456 (2019).

[14]

Li, X. et al. Research of Morinda officinalis how’s oligosaccharide extraction and antidepressant effects. Bulgar. Chem. Commun. 49, 162-167 (2017).

[15]

Zhang, R. et al. Investigation on germplasm resources of Morinda officinalis how. Mod. Chin. Med 18, 482-487 (2016).

[16]

Liu, J., Ding, P., Zhan, R. T. & Chen, W. W. Resource survey of medicinal plant of Morinda officinalis how in Guangdong and Fujian Provinces. J. Guangzhou Univ. Trad. Chin. Med. 26, 485-487 (2009).

[17]

Luo, M., Shu, Y., Zhang, W. & Dong, Z. Research progress of Morinda officinalis disease. J. Agric. Sci. 8, 1312-1317 (2018).

[18]

Yang, J., Jia, M. & Guo, J. Functional genome of medicinal plants. Mol. Pharmacogn. 11, 191-234 (2019).

[19]

Zhao, Q. et al. The reference genome sequence of Scutellaria baicalensis provides insights into the evolution of wogonin biosynthesis. Mol. Plant 12, 935-950 (2019).

[20]

Kang, M. et al. A chromosome-scale genome assembly of Isatis indigotica, an important medicinal plant used in traditional Chinese medicine. Hortic. Res. 7, 1-10 (2020).

[21]

Li, L. F., Cushman, S. A., He, Y. X. & Li, Y. Genome sequencing and population genomics modeling provide insights into the local adaptation of weeping forsythia. Hortic. Res. 7, 1-12 (2020).

[22]

Liu, X. et al. The genome of medicinal plant Macleaya cordata provides new insights into benzylisoquinoline alkaloids metabolism. Mol. Plant 10, 975-989 (2017).

[23]

Gao, S. et al. A high-quality reference genome of wild Cannabis sativa. Hortic. Res. 7, 1-11 (2020).

[24]

Tu, L. et al. Genome of Tripterygium wilfordii and identification of cytochrome P450 involved in triptolide biosynthesis. Nat. Commun. 11, 1-12 (2020).

[25]

Denoeud, F. et al. The coffee genome provides insight into the convergent evolution of caffeine biosynthesis. Science 345, 1181-1184 (2014).

[26]

Christianson, D. W. Roots of biosynthetic diversity. Science 316, 60-61 (2007).

[27]

El Baidouri, M. & Panaud, O. Comparative genomic paleontology across plant kingdom reveals the dynamics of TE-driven genome evolution. Genome Biol. Evol. 5, 954-965 (2013).

[28]

Zou, Y. et al. Transcriptional regulation of the immune receptor FLS2 controls the ontogeny of plant innate immunity. Plant Cell 30, 2779-2794 (2018).

[29]

Meng, X. et al. Phosphorylation of an ERF transcription factor by Arabidopsis MPK3/MPK6 regulates plant defense gene induction and fungal resistance. Plant Cell 25, 1126-1142 (2013).

[30]

Saand, M. A., Xu, Y. P., Li, W., Wang, J. & Cai, X. Z. Cyclic nucleotide gated channel gene family in tomato: genome-wide identification and functional analyses in disease resistance. Front. Plant Sci. 6, 303 (2015).

[31]

Yamazaki, M. et al. Coupling deep transcriptome analysis with untargeted metabolic profiling in Ophiorrhiza pumila to further the understanding of the biosynthesis of the anti-cancer alkaloid camptothecin and anthraquinones. Plant Cell Physiol. 54, 686-696 (2013).

[32]

Kang, S. H. et al. Genome-enabled discovery of anthraquinone biosynthesis in Senna tora. Nat. Commun. 11, 1-11 (2020).

[33]

Ye, P. et al. Transcriptome analysis and targeted metabolic profiling for pathway elucidation and identification of a geraniol synthase involved in iridoid biosynthesis from Gardenia jasminoides. Ind. Crops Prod. 132, 48-58 (2019).

[34]

Han, X. et al. Transcriptome analysis reveals the molecular mechanisms of mucilage biosynthesis during Artemisia sphaerocephala seed development. Ind. Crops Prod. 145, 111991 (2020).

[35]

Han, Y. S., Van der Heijden, R. & Verpoorte, R. Biosynthesis of anthraquinones in cell cultures of the Rubiaceae. Plant Cell Tissue Organ Cult. 67, 201-220 (2001).

[36]

Liang, W. et al. Research progress on synthesis of anthraquinones based on shikimic acid/o-succinylbenzoic acid pathway. Chin. Trad. Herb. Drugs 51, 1939-1950 (2020).

[37]

Sun, W. et al. The genome of the medicinal plant Andrographis paniculata provides insight into the biosynthesis of the bioactive diterpenoid neoandrographolide. Plant J. 97, 841-857 (2019).

[38]

Xu, S., Yao, S., Huang, R., Tan, Y. & Huang, D. Transcriptome-wide analysis of the AP2/ERF transcription factor gene family involved in the regulation of gypenoside biosynthesis in Gynostemma pentaphyllum. Plant Physiol. Biochem. 154, 238-247 (2020).

[39]

Li, H. B. et al. Two new iridoid glycosides from the fruit of Gardenia jasminoides. Nat. Prod. Res. 2015, 1-7 (2020).

[40]

Xia, Z. et al. Chromosome-scale genome assembly provides insights into the evolution and flavor synthesis of passion fruit (Passiflora edulis Sims). Hortic. Res. 8, 1-14 (2021).

[41]

Xu, Z. et al. Comparative genome analysis of Scutellaria baicalensis and Scutellaria barbata reveals the evolution of active flavonoid biosynthesis. Genomics Proteomics Bioinformatics 18, 230-240 (2020).

[42]

Fernie, A. R., Carrari, F. & Sweetlove, L. J. Respiratory metabolism: glycolysis, the TCA cycle and mitochondrial electron transport. Curr. Opin. Plant Biol. 7, 254-261 (2004).

[43]

Patel, R. K. & Jain, M. NGS QC toolkit: a toolkit for quality control of next generation sequencing data. PLoS ONE 7, e30619 (2012).

[44]

Chen, S., Zhou, Y., Chen, Y. & Gu, J. fastp: an ultra-fast all-in-one FASTQ preprocessor. Bioinformatics 34, i884-i890 (2018).

[45]

Wick, R. R., Judd, L. M. & Holt, K. E. Performance of neural network basecalling tools for Oxford Nanopore sequencing. Genome Biol. 20, 129 (2019).

[46]

Senol Cali, D., Kim, J. S., Ghose, S., Alkan, C. & Mutlu, O. Nanopore sequencing technology and tools for genome assembly: computational analysis of the current state, bottlenecks and future directions. Brief Bioinformatics 20, 1542-1559 (2019).

[47]

Marçais, G. & Kingsford, C. A fast, lock-free approach for efficient parallel counting of occurrences of k-mers. Bioinformatics 27, 764-770 (2011).

[48]

Li, H. Minimap and miniasm: fast mapping and de novo assembly for noisy long sequences. Bioinformatics 32, 2103-2110 (2016).

[49]

Hu, J., Fan, J., Sun, Z. & Liu, S. NextPolish: a fast and efficient genome polishing tool for long read assembly. Bioinformatics 36, 2253-2255 (2020).

[50]

Li, H. Aligning sequence reads, clone sequences and assembly contigs with BWA-MEM. Preprint at https://arxiv.org/abs/1303.3997 (2013).

[51]

Walker, B. J. et al. Pilon: an integrated tool for comprehensive microbial variant detection and genome assembly improvement. PLoS ONE 9, e112963 (2014).

[52]

Pryszcz, L. P. & Gabaldón, T. Redundans: an assembly pipeline for highly heterozygous genomes. Nucleic Acids Res. 44, e113- e113 (2016).

[53]

Langmead, B. & Salzberg, S. Fast gapped-read alignment with Bowtie 2. Nat. Methods 9, 357-359 (2012).

[54]

Servant, N. et al. HiC-Pro: an optimized and flexible pipeline for Hi-C data processing. Genome Biol. 16, 259 (2015).

[55]

Burton, J. N. et al. Chromosome-scale scaffolding of de novo genome assemblies based on chromatin interactions. Nat. Biotechnol. 31, 1119-1125 (2013).

[56]

Kim, D., Paggi, J. M., Park, C., Bennett, C. & Salzberg, S. L. Graph-based genome alignment and genotyping with HISAT2 and HISAT-genotype. Nat. Biotechnol. 37, 907-915 (2019).

[57]

Simão, F. A., Waterhouse, R. M., Ioannidis, P., Kriventseva, E. V. & Zdobnov, E. M. BUSCO: assessing genome assembly and annotation completeness with single-copy orthologs. Bioinformatics 31, 3210-3212 (2015).

[58]

Wang, X. & Wang, L. GMATA: an integrated software package for genome-scale SSR mining, marker development and viewing. Front. Plant Sci. 7, 1350 (2016).

[59]

Benson, G. Tandem repeats finder: a program to analyze DNA sequences. Nucleic Acids Res. 27, 573-580 (1999).

[60]

Bedell, J. A., Korf, I. & Gish, W. MaskerAid: a performance enhancement to RepeatMasker. Bioinformatics 16, 1040-1041 (2000).

[61]

Abrusán, G., Grundmann, N., DeMester, L. & Makalowski, W. TEclass-a tool for automated classification of unknown eukaryotic transposable elements. Bioinformatics 25, 1329-1330 (2009).

[62]

Jurka, J. et al. Repbase update, a database of eukaryotic repetitive elements. Cytogenet. Genome Res. 110, 462-467 (2005).

[63]

Han, Y. & Wessler, S. R. MITE-Hunter: a program for discovering miniature inverted-repeat transposable elements from genomic sequences. Nucleic Acids Res. 38, e199 (2010).

[64]

Xu, Z. & Wang, H. LTR-FINDER: an efficient tool for the prediction of full-length LTR retrotransposons. Nucleic Acids Res. 35, W265-W268 (2007).

[65]

Ellinghaus, D., Kurtz, S. & Willhoeft, U. LTRharvest, an efficient and flexible software for de novo detection of LTR retrotransposons. BMC Bioinformatics 9, 18 (2008).

[66]

Ou, S. & Jiang, N. LTR_retriever: a highly accurate and sensitive program for identification of long terminal repeat retrotransposons. Plant Physiol. 176, 1410-1422 (2018).

[67]

Lowe, T. M. & Chan, P. P. tRNAscan-SE On-line: integrating search and context for analysis of transfer RNA genes. Nucleic Acids Res. 44, W54-W57 (2016).

[68]

Gardner, P. P. et al. Rfam: updates to the RNA families database. Nucleic Acids Res. 37, D136-D140 (2009).

[69]

Nawrocki, E. P., Kolbe, D. L. & Eddy, S. R. Infernal 1.0: inference of RNA alignments. Bioinformatics 25, 1335-1337 (2009).

[70]

Lagesen, K. et al. RNAmmer: consistent and rapid annotation of ribosomal RNA genes. Nucleic Acids Res. 35, 3100-3108 (2007).

[71]

De Nardi, B. et al. Differential responses of Coffea arabica L. leaves and roots to chemically induced systemic acquired resistance. Genome 49, 1594-1605 (2006).

[72]

Franke, J. et al. Gene discovery in Gelsemium highlights conserved gene clusters in monoterpene indole alkaloid biosynthesis. ChemBioChem 20, 83-87 (2019).

[73]

Kaul, S. et al. Analysis of the genome sequence of the flowering plant Arabidopsis thaliana. Nature 408, 796-815 (2000).

[74]

Keilwagen, J. et al. Using intron position conservation for homology-based gene prediction. Nucleic Acids Res. 44, e89- e89 (2016).

[75]

Haas, B. J. et al. Improving the Arabidopsis genome annotation using maximal transcript alignment assemblies. Nucleic Acids Res. 31, 5654-5666 (2003).

[76]

Haas, B. J. et al. De novo transcript sequence reconstruction from RNA-seq using the Trinity platform for reference generation and analysis. Nat. Protoc. 8, 1494-1512 (2013).

[77]

Stanke, M., Diekhans, M., Baertsch, R. & Haussler, D. Using native and syntentically mapped cDNA alignments to improve de novo gene finding. Bioinformatics 24, 637-644 (2008).

[78]

Haas, B. J. et al. Automated eukaryotic gene structure annotation using EVidenceModeler and the Program to Assemble Spliced Alignments. Genome Biol. 9, R7 (2008).

[79]

Urasaki, N. et al. Draft genome sequence of bitter gourd (Momordica charantia), a vegetable and medicinal plant in tropical and subtropical regions. DNA Res. 24, 51-58 (2017).

[80]

Hunter, S. et al. InterPro: the integrative protein signature database. Nucleic Acids Res. 37, D211-D215 (2009).

[81]

Li, L., Stoeckert, C. J. & Roos, D. S. OrthoMCL: identification of ortholog groups for eukaryotic genomes. Genome Res. 13, 2178-2189 (2003).

[82]

Katoh, K. & Standley, D. M. MAFFT multiple sequence alignment software version 7: improvements in performance and usability. Mol. Biol. Evol. 30, 772-780 (2013).

[83]

Castresana, J. Selection of conserved blocks from multiple alignments for their use in phylogenetic analysis. Mol. Biol. Evol. 17, 540-552 (2000).

[84]

Stamatakis, A. RAxML-VI-HPC: maximum likelihood-based phylogenetic analyses with thousands of taxa and mixed models. Bioinformatics 22, 2688-2690 (2006).

[85]

Yang, Z. PAML: a program package for phylogenetic analysis by maximum likelihood. Comput. Appl. Biosci. 13, 555-556 (1997).

[86]

De Bie, T., Cristianini, N., Demuth, J. P. & Hahn, M. W. CAFE: a computational tool for the study of gene family evolution. Bioinformatics 22, 1269-1271 (2006).

[87]

Gao, F. et al. EasyCodeML: a visual tool for analysis of selection using CodeML. Ecol. Evol. 9, 3891-3898 (2019).

[88]

Wang, Y. et al. MCScanX: a toolkit for detection and evolutionary analysis of gene synteny and collinearity. Nucleic Acids Res. 40, e49- e49 (2012).

[89]

Robinson, M. D., McCarthy, D. J. & Smyth, G K . edgeR: a Bioconductor package for differential expression analysis of digital gene expression data. Bioinformatics 26, 139-140 (2010).

PDF (2556KB)

0

Accesses

0

Citation

Detail

Sections
Recommended

/