The genome of Magnolia biondii Pamp. provides insights into the evolution of Magnoliales and biosynthesis of terpenoids

Shanshan Dong , Min Liu , Yang Liu , Fei Chen , Ting Yang , Lu Chen , Xingtan Zhang , Xing Guo , Dongming Fang , Linzhou Li , Tian Deng , Zhangxiu Yao , Xiaoan Lang , Yiqing Gong , Ernest Wu , Yaling Wang , Yamei Shen , Xun Gong , Huan Liu , Shouzhou Zhang

Horticulture Research ›› 2021, Vol. 8 ›› Issue (1) : 38

PDF (1449KB)
Horticulture Research ›› 2021, Vol. 8 ›› Issue (1) :38 DOI: 10.1038/s41438-021-00471-9
Article
research-article
The genome of Magnolia biondii Pamp. provides insights into the evolution of Magnoliales and biosynthesis of terpenoids
Author information +
History +
PDF (1449KB)

Abstract

Magnolia biondii Pamp. (Magnoliaceae, magnoliids) is a phylogenetically, economically, and medicinally important ornamental tree species widely grown and cultivated in the north-temperate regions of China. Determining the genome sequence of M. biondii would help resolve the phylogenetic uncertainty of magnoliids and improve the understanding of individual trait evolution within the Magnolia genus. We assembled a chromosome-level reference genome of M. biondii using ~67, ~175, and ~154 Gb of raw DNA sequences generated via Pacific Biosciences single-molecule real-time sequencing, 10X Genomics Chromium, and Hi-C scaffolding strategies, respectively. The final genome assembly was ~2.22 Gb, with a contig N50 value of 269.11 kb and a BUSCO complete gene percentage of 91.90%. Approximately 89.17% of the genome was organized into 19 chromosomes, resulting in a scaffold N50 of 92.86 Mb. The genome contained 47,547 protein-coding genes, accounting for 23.47% of the genome length, whereas 66.48% of the genome length consisted of repetitive elements. We confirmed a WGD event that occurred very close to the time of the split between the Magnoliales and Laurales. Functional enrichment of the Magnolia-specific and expanded gene families highlighted genes involved in the biosynthesis of secondary metabolites, plant–pathogen interactions, and responses to stimuli, which may improve the ecological fitness and biological adaptability of the lineage. Phylogenomic analyses revealed a sister relationship of magnoliids and Chloranthaceae, which are sister to a clade comprising monocots and eudicots. The genome sequence of M. biondii could lead to trait improvement, germplasm conservation, and evolutionary studies on the rapid radiation of early angiosperms.

Cite this article

Download citation ▾
Shanshan Dong, Min Liu, Yang Liu, Fei Chen, Ting Yang, Lu Chen, Xingtan Zhang, Xing Guo, Dongming Fang, Linzhou Li, Tian Deng, Zhangxiu Yao, Xiaoan Lang, Yiqing Gong, Ernest Wu, Yaling Wang, Yamei Shen, Xun Gong, Huan Liu, Shouzhou Zhang. The genome of Magnolia biondii Pamp. provides insights into the evolution of Magnoliales and biosynthesis of terpenoids. Horticulture Research, 2021, 8 (1) : 38 DOI:10.1038/s41438-021-00471-9

登录浏览全文

4963

注册一个新账户 忘记密码

References

[1]

Rivers, M., Beech, E., Murphy, L. & Oldfield, S. The red list of Magnoliaceae-revised and extended. https://globaltrees.org/resources/the-red-list-of-magnoliaceae-revised-and-extended/ (2016).

[2]

Figlar, R. B. & Nooteboom, H. P. Notes on Magnoliaceae IV. Blumea 49, 87-100 (2004).

[3]

Kim, S. & Suh, Y. Phylogeny of Magnoliaceae based on ten chloroplast DNA regions. J. Plant Biol. 56, 290-305 (2013).

[4]

Azuma, H., García‐Franco, J. G., Rico‐Gray, V. & Thien, L. B. Molecular phylogeny of the Magnoliaceae: the biogeography of tropical and temperate disjunctions. Am. J. Bot. 88, 2275-2285 (2001).

[5]

Soltis, D. E. & Soltis, P. S. Nuclear genomes of two magnoliids. Nat. Plants 5, 6-7 (2019).

[6]

Soltis, D., Bell, C., Kim, S. & Soltis, P. S. Origin and early evolution of angiosperms. Ann. N. Y. Acad. Sci. 1133, 3 (2008).

[7]

Kersey, P. J. Plant genome sequences: past, present, future. Curr. Opin. Plant Biol. 48, 1-8 (2019).

[8]

Hu, L. et al. The chromosome-scale reference genome of black pepper provides insight into piperine biosynthesis. Nat. Commun. 10, 4702 (2019).

[9]

Rendón-Anaya, M. et al. The avocado genome informs deep angiosperm phylogeny, highlights introgressive hybridization, and reveals pathogen influenced gene space adaptation. Proc. Natl Acad. Sci. USA 116, 17081-17089 (2019).

[10]

Strijk, J. S. et al. The soursop genome and comparative genomics of basal angiosperms provide new insights on evolutionary incongruence. Preprint at bioRxiv https://doi.org/10.1101/639153 (2019).

[11]

Chaw, S. M. et al. Stout camphor tree genome fills gaps in understanding of flowering plant genome evolution. Nat. Plants 5, 63-73 (2019).

[12]

Chen, J. et al. Liriodendron genome sheds light on angiosperm phylogeny and species-pair differentiation . Nat. Plants 5, 18-25 (2018).

[13]

Wickett, N. J. et al. Phylotranscriptomic analysis of the origin and early diversification of land plants. Proc. Natl Acad. Sci. USA 111, 4859-4868 (2014).

[14]

Zeng, L. et al. Resolution of deep angiosperm phylogeny using conserved nuclear genes and estimates of early divergence times. Nat. Commun. 5, 4956 (2014).

[15]

Gitzendanner, M. A., Soltis, P. S., Wong, G. K.-S., Ruhfel, B. R. & Soltis, D. E. Plastid phylogenomic analysis of green plants: a billion years of evolutionary history. Am. J. Bot. 105, 291-301 (2018).

[16]

Ruhfel, B. R., Gitzendanner, M. A., Soltis, P. S., Soltis, D. E. & Burleigh, J. G. From algae to angiosperms-inferring the phylogeny of green plants (Viridiplantae) from 360 plastid genomes. BMC Evol. Biol. 14, 23 (2014).

[17]

Li, H. T. et al. Origin of angiosperms and the puzzle of the Jurassic gap. Nature Plants 5, 461-470 (2019).

[18]

Wang, Y. L. & Zhang, S. Z. Studies on the microsporogenesis and development of the male gametophyte of Magnolia championii Benth. J. Wuhan. Bot. Res 26, 547-553 (2008).

[19]

Hirayama, K., Ishida, K. & Tomaru, N. Effects of pollen shortage and self-pollination on seed production of an endangered tree, Magnolia stellata . Ann. Bot. 95, 1009-1015 (2005).

[20]

Yang, X., Yang, Z. L., Wang, J., Tan, G. Y. & He, Z. S. Floral syndrome and breeding system of endangered species Magnolia officinalis subsp. biloba . Chinese J. Ecol. 3, 551-556 (2012).

[21]

Wang, X. et al. Development of EST-SSR markers and their application in an analysis of the genetic diversity of the endangered species Magnolia sinostellata . Mol. Genet. Genomics 294, 135-147 (2019).

[22]

Jiang, W., Cao, J., Li, G. & Weng, M. Development of new ornamental tree species of Magnolia family in China and its application in landscaping. Acta Agriculturae Shanghai 21, 68-73 (2005).

[23]

Zhao, L. The terpenoid biosynthesis pathway in Magnolia and their significance for taxonomy in the genus. Guihaia 4, 7 (2005).

[24]

Ho, K. Y., Tsai, C. C., Chen, C. P., Huang, J. S. & Lin, C. C. Antimicrobial activity of honokiol and magnolol isolated from Magnolia officinalis . Phytother. Res. 15, 139-141 (2001).

[25]

China Pharmacopoeia Committee . Pharmacopoeia of the People’s Republic of China The first Division of 2000 English edn (Chemical Industry Press, 2000).

[26]

Qu, L., Qi, Y., Fan, G. & Wu, Y. Determination of the volatile oil of Magnolia biondii pamp by GC-MS combined with chemometric techniques. Chromatographia 70, 905-914 (2009).

[27]

Zhao, W., Zhou, T., Fan, G., Chai, Y. & Wu, Y. Isolation and purification of lignans from Magnolia biondii pamp by isocratic reversed-phase two-dimensional liquid chromatography following microwave-assisted extraction. J. Sep. Sci. 30, 2370-2381 (2015).

[28]

Chen, Y., Gao, B. C., Qiao, L. & Han, G. Q. Study on the hydrophilic components of Magnolia biondii pamp. Acta Pharm. Sin. 29, 506-510 (1994).

[29]

Porebski, S., Bailey, L. G. & Bernard, R. B. Modification of a CTAB DNA extraction protocol for plants containing high polysaccharide and polyphenol components. Plant Mol. Biol. Report 15, 8-15 (1997).

[30]

Chen, Y. et al. SOAPnuke: a MapReduce acceleration-supported software for integrated quality control and preprocessing of high-throughput sequencing data. Gigascience 7, gix120 (2017).

[31]

Belaghzal, H., Dekker, J. & Gibcus, J. H. Hi-C 2.0: an optimized Hi-C procedure for high-resolution genome-wide mapping of chromosome conformation. Methods 123, 56-65 (2017).

[32]

Bolger, A. M., Lohse, M. & Usadel, B. Trimmomatic: a flexible trimmer for Illumina sequence data. Bioinformatics 30, 2114-2120 (2014).

[33]

Moscone, E. A. et al. Analysis of nuclear DNA content in Capsicum (Solanaceae) by flow cytometry and Feulgen densitometry. Ann. Bot. 92, 21-29 (2003).

[34]

Chang, Y. et al. The draft genomes of five agriculturally important African orphan crops. Gigascience 8, 1-16 (2019).

[35]

Liu, B. et al. Estimation of genomic characteristics by analyzing k-mer frequency in de novo genome projects. arXiv Prepr. 1308, 2012 (2013).

[36]

Koren, S. et al. Canu: scalable and accurate long-read assembly via adaptive k-mer weighting and repeat separation. Genome Res. 27, 722 (2017).

[37]

Li, H. Minimap and miniasm: fast mapping and de novo assembly for noisy long sequences. Bioinformatics 32, 2103-2110 (2016).

[38]

Kolmogorov, M., Yuan, J., Lin, Y. & Pevzner, P. A. Assembly of long, error-prone reads using repeat graphs. Nat. Biotechnol. 37, 540 (2019).

[39]

Simao, F. A., Waterhouse, R. M., Ioannidis, P., Kriventseva, E. V. & Zdobnov, E. M. BUSCO: assessing genome assembly and annotation completeness with single-copy orthologs. Bioinformatics 31, 3210-3212 (2015).

[40]

Vaser, R., Sović, I., Nagarajan, N. & Šikić, M. Fast and accurate de novo genome assembly from long uncorrected reads. Genome Res. 27, 737-746 (2016).

[41]

Walker, B. J. et al. Pilon: an integrated tool for comprehensive microbial variant detection and genome assembly improvement. PLoS ONE 9, e112963 (2014).

[42]

Durand, N. C. et al. Juicer provides a one-click system for analyzing loop-resolution Hi-C experiments. Cell Syst. 3, 95-98 (2016).

[43]

Dudchenko, O. et al. De novo assembly of the Aedes aegypti genome using Hi-C yields chromosome-length scaffolds. Science 356, 92-95 (2017).

[44]

Dudchenko, O. et al. The Juicebox Assembly Tools module facilitates de novo assembly of mammalian genomes with chromosome-length scaffolds for under $1000. Preprint at bioRxiv https://doi.org/10.1101/254797 (2018).

[45]

Li, H. Aligning sequence reads, clone sequences and assembly contigs with BWA-MEM. arXiv 1303, 3997v2 (2013).

[46]

Kim, D. et al. Tophat2: accurate alignment of transcriptomes in the presence of insertions, deletions and gene fusions. Genome Biol. Evol. 14, R36 (2013).

[47]

Chang, Z. et al. Bridger: a new framework for de novo transcriptome assembly using RNA-seq data. Genome Biol. 16, 30 (2015).

[48]

Kent, W. J. BLAT-the BLAST-like alignment tool. Genome Res. 12, 656-664 (2002).

[49]

Jerzy, J . Repbase update: a database and an electronic journal of repetitive elements. Trends Genet. 16, 418-420 (2000).

[50]

Tarailo‐Graovac, M. & Chen, N. Using RepeatMasker to identify repetitive elements in genomic sequences. Curr. Protoc. Bioinforma. 25, 4-10 (2009).

[51]

Hubley, R. & Smit, A. RepeatModeler. http://www.repeatmasker.org/RepeatModeler/ (2019).

[52]

Xu, Z. & Wang, H. LTR_FINDER: an efficient tool for the prediction of full-length LTR retrotransposons. Nucleic Acids Res. 35, W265-W268 (2007).

[53]

Benson, G. Tandem repeats finder: a program to analyze DNA sequences. Nucleic Acids Res. 1999, 2 (1999).

[54]

Campbell, M. S., Holt, C., Moore, B. & Yandell, M. Genome annotation and curation using MAKER and MAKER‐P. Curr. Protoc. Bioinformatics 48, 4-11 (2014).

[55]

Lomsadze, A., Ter-Hovhannisyan, V., Chernoff, Y. & Borodovsky, M. Gene identification in novel eukaryotic genomes by self-training algorithm. Nucleic Acids Res. 33, 6494-6506 (2005).

[56]

Stanke, M., Schöffmann, O., Morgenstern, B. & Waack, S. Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources. BMC Bioinforma. 7, 62 (2006).

[57]

Johnson, A. D. et al. SNAP: a web-based tool for identification and annotation of proxy SNPs using HapMap. Bioinformatics 24, 2938-2939 (2008).

[58]

Haas, B. J. et al. De novo transcript sequence reconstruction from RNA-seq using the Trinity platform for reference generation and analysis. Nat. Protoc. 8, 1494 (2013).

[59]

Lowe, T. M. & Chan, P. P. tRNAscan-SE On-line: integrating search and context for analysis of transfer RNA genes. Nucleic Acids Res. 44, W54-W57 (2016).

[60]

Kalvari, I. et al. Rfam 13.0: shifting to a genome-centric resource for non-coding RNA families. Nucleic Acids Res. 46, D335-D342 (2017).

[61]

Altschul, S. F., Gish, W., Miller, W., Myers, E. W. & Lipman, D. J. Basic local alignment search tool. J. Mol. Biol. 215, 403-410 (1990).

[62]

Vitales, D., D’Ambrosio, U., Gálvez, F., Kovařík, A. & Garcia, S. Third release of the plant rDNA database with updated content and information on telomere composition and sequenced plant genomes. Plant Syst. Evol. 303, 1115-1121 (2017).

[63]

Aoki, K. F. & Kanehisa, M. Using the KEGG database resource. Curr. Protoc. Bioinforma. 11, 1-12 (2005).

[64]

Tatusov, R. L., Koonin, E. V. & Lipman, D. J. A genomic perspective on protein families. Science 278, 631-637 (1997).

[65]

Boeckmann, B. et al. The SWISS-PROT protein knowledgebase and its supplement TrEMBL in 2003. Nucleic Acids Res. 31, 365-370 (2003).

[66]

Jones, P. et al. InterProScan 5: genome-scale protein function classification. Bioinformatics 30, 1236-1240 (2014).

[67]

Finn, R. D. et al. The Pfam protein families database. Nucleic Acids Res. 36, D281-D288 (2007).

[68]

Letunic, I., Doerks, T. & Bork, P. SMART 6: recent updates and new developments. Nucleic Acids Res. 37, D229-D232 (2009).

[69]

Mi, H., Muruganujan, A., Casagrande, J. T. & Thomas, P. D. Large-scale gene function analysis with the PANTHER classification system. Nat. Protoc. 8, 1551 (2013).

[70]

Attwood, T. K. et al. PRINTS and its automatic supplement, prePRINTS. Nucleic Acids Res. 31, 400-402 (2003).

[71]

Corpet, F., Servant, F., Gouzy, J. & Kahn, D. ProDom and ProDom-CG: tools for protein domain analysis and whole genome comparisons. Nucleic Acids Res. 28, 267-269 (2000).

[72]

Emms, D. M. & Kelly, S. OrthoFinder: phylogenetic orthology inference for comparative genomics. Genome Biol. 20, 238 (2019).

[73]

Katoh, K., Kuma, K., Toh, H. & Miyata, T. MAFFT version 5: improvement in accuracy of multiple sequence alignment. Nucleic Acids Res. 33, 511-518 (2005).

[74]

Talavera, G. & Castresana, J. Improvement of phylogenies after removing divergent and ambiguously aligned blocks from protein sequence alignments. Syst. Biol. 56, 564-577 (2007).

[75]

Lanfear, R., Calcott, B., Ho, S. Y. & Guindon, S. PartitionFinder: combined selection of partitioning schemes and substitution models for phylogenetic analyses. Mol. Biol. Evol. 29, 1695-1701 (2012).

[76]

Stamatakis, A. RAxML-VI-HPC: maximum likelihood-based phylogenetic analyses with thousands of taxa and mixed models. Bioinformatics 22, 2688-2690 (2006).

[77]

Yang, Z. PAML4: phylogenetic analysis by maximum likelihood. Mol. Biol. Evol. 24, 1586-1591 (2007).

[78]

De Bie, T., Cristianini, N., Demuth, J. P. & Hahn, M. W. Cafe: a computational tool for the study of gene family evolution. Bioinformatics 22, 1269-1271 (2006).

[79]

Wang, Y. et al. MCScanX: a toolkit for detection and evolutionary analysis of gene synteny and collinearity. Nucleic Acids Res. 40, e49- e49 (2012).

[80]

Li, L., Stoeckert, C. J. & Roos, D. S. OrthoMCL: identification of ortholog groups for eukaryotic genomes. Genome Res. 13, 2178-2189 (2003).

[81]

Tang, H. et al. Unraveling ancient hexaploidy through multiply-aligned angiosperm gene maps. Genome Res. 18, 1944-1954 (2008).

[82]

Finn, R. D., Clements, J. & Eddy, S. R. HMMER web server: interactive sequence similarity searching. Nucleic Acids Res. 39, 29-37 (2011).

[83]

Tamura, K. MEGA6: molecular evolutionary genetics analysis version 6.0. Mol. Biol. Evol. 30, 2725-2729 (2013).

[84]

Nguyen, L. T., Schmidt, H. A., von Haeseler, A. & Minh, B. Q. IQ-TREE: a fast and effective stochastic algorithm for estimating maximum-likelihood phylogenies. Mol. Biol. Evol. 32, 268-274 (2014).

[85]

Langmead, B. & Salzberg, S. L. Fast gapped-read alignment with Bowtie 2. Nat. Methods 9, 357-1359 (2012).

[86]

Roberts, A. & Pachter, L. Streaming fragment assignment for real-time analysis of sequencing experiments. Nat. Methods 10, 71-73 (2013).

[87]

Robinson, M. D., McCarthy, D. J. & Smyth, G. K. edgeR: a Bioconductor package for differential expression analysis of digital gene expression data. Bioinformatics 26, 139-140 (2010).

[88]

Vanburen, R. et al. Single-molecule sequencing of the desiccation-tolerant grass Oropetium thomaeum . Nature 527, 508 (2015).

[89]

Chen, Y. et al. The Litsea genome and the evolution of the laurel family. Nat. Commun. 11, 1675 (2020).

[90]

Lu, J. et al. Analysis of the chemical constituents of essential oil from Magnolia biondii by GC-MS. J. Chin. Med. Mater. 31, 1649-1651 (2008).

PDF (1449KB)

0

Accesses

0

Citation

Detail

Sections
Recommended

/