Chromosome-scale genome assembly of sweet cherry (Prunus avium L.) cv. Tieton obtained using long-read and Hi-C sequencing

Jiawei Wang , Weizhen Liu , Dongzi Zhu , Po Hong , Shizhong Zhang , Shijun Xiao , Yue Tan , Xin Chen , Li Xu , Xiaojuan Zong , Lisi Zhang , Hairong Wei , Xiaohui Yuan , Qingzhong Liu

Horticulture Research ›› 2020, Vol. 7 ›› Issue (1) : 122

PDF (1345KB)
Horticulture Research ›› 2020, Vol. 7 ›› Issue (1) :122 DOI: 10.1038/s41438-020-00343-8
Article
research-article
Chromosome-scale genome assembly of sweet cherry (Prunus avium L.) cv. Tieton obtained using long-read and Hi-C sequencing
Author information +
History +
PDF (1345KB)

Abstract

Sweet cherry (Prunus avium) is an economically significant fruit species in the genus Prunus. However, in contrast to other important fruit trees in this genus, only one draft genome assembly is available for sweet cherry, which was assembled using only Illumina short-read sequences. The incompleteness and low quality of the current sweet cherry draft genome limit its use in genetic and genomic studies. A high-quality chromosome-scale sweet cherry reference genome assembly is therefore needed. A total of 65.05 Gb of Oxford Nanopore long reads and 46.24 Gb of Illumina short reads were generated, representing ~190x and 136x coverage, respectively, of the sweet cherry genome. The final de novo assembly resulted in a phased haplotype assembly of 344.29 Mb with a contig N50 of 3.25 Mb. Hi-C scaffolding of the genome resulted in eight pseudochromosomes containing 99.59% of the bases in the assembled genome. Genome annotation revealed that more than half of the genome (59.40%) was composed of repetitive sequences, and 40,338 protein-coding genes were predicted, 75.40% of which were functionally annotated. With the chromosome-scale assembly, we revealed that gene duplication events contributed to the expansion of gene families for salicylic acid/jasmonic acid carboxyl methyltransferase and ankyrin repeat-containing proteins in the genome of sweet cherry. Four auxin-responsive genes (two GH3s and two SAURs) were induced in the late stage of fruit development, indicating that auxin is crucial for the sweet cherry ripening process. In addition, 772 resistance genes were identified and functionally predicted in the sweet cherry genome. The high-quality genome assembly of sweet cherry obtained in this study will provide valuable genomic resources for sweet cherry improvement and molecular breeding.

Cite this article

Download citation ▾
Jiawei Wang, Weizhen Liu, Dongzi Zhu, Po Hong, Shizhong Zhang, Shijun Xiao, Yue Tan, Xin Chen, Li Xu, Xiaojuan Zong, Lisi Zhang, Hairong Wei, Xiaohui Yuan, Qingzhong Liu. Chromosome-scale genome assembly of sweet cherry (Prunus avium L.) cv. Tieton obtained using long-read and Hi-C sequencing. Horticulture Research, 2020, 7 (1) : 122 DOI:10.1038/s41438-020-00343-8

登录浏览全文

4963

注册一个新账户 忘记密码

References

[1]

Quero-García, J., Iezzoni, A., Puławska, J. & Lang, G. Cherries: Botany, Production and Uses 1-533 (CABI PUBLISHING, Boston, MA, 2017).

[2]

Xuwei, D. et al. Fruit scientific research in New China in the past 70 years: cherry. J. Fruit. Sci. 36, 1339-1351 (2019).

[3]

Garcia, J. Q. Cherry breeding in the world: current analysis and future perspectives. Italus Hortus 26, 9-20 (2019).

[4]

Shirasawa, K. et al. The genome sequence of sweet cherry (Prunus avium) for use in genomics-assisted breeding. DNA Res. 24, 499-508 (2017).

[5]

Zhang, Q. et al. The genome of Prunus mume. Nat. Commun. 3, 1318 (2012).

[6]

Verde, I. et al. The Peach v2.0 release: high-resolution linkage mapping and deep resequencing improve chromosome-scale assembly and contiguity. BMC Genomics. 18, 225 (2017).

[7]

International Peach Genome, I. et al. The high-quality draft genome of peach (Prunus persica) identifies unique patterns of genetic diversity, domestication and genome evolution. Nat. Genet. 45, 487-494 (2013).

[8]

Sanchez-Perez, R. et al. Mutation of a bHLH transcription factor allowed almond domestication. Science 364, 1095-1098 (2019).

[9]

Baek, S. et al. Draft genome sequence of wild Prunus yedoensis reveals massive inter-specific hybridization between sympatric flowering cherries. Genome Biol. 19, 127 (2018).

[10]

Shirasawa, K. et al. Phased genome sequence of an interspecific hybrid flowering cherry, ‘Somei-Yoshino’ (Cerasus × yedoensis). DNA Res. 26, 379-389 (2019).

[11]

Jiang, F. et al. The apricot (Prunus armeniaca L.) genome elucidates Rosaceae evolution and beta-carotenoid synthesis. Horticulture Res. 6, 128 (2019).

[12]

Zhebentyayeva, T. et al. Genetic characterization of worldwide Prunus domestica (plum) germplasm using sequence-based genotyping. Hortic. Res. 6, 12 (2019).

[13]

Seppey, M., Manni, M. & Zdobnov, E. M. BUSCO: assessing genome assembly and annotation completeness. Methods Mol. Biol. 1962, 227-245 (2019).

[14]

Zhang, X., Zhang, S., Zhao, Q., Ming, R. & Tang, H. Assembly of allele-aware, chromosomal-scale autopolyploid genomes based on Hi-C data. Nat. Plants 5, 833-845 (2019).

[15]

Ou, S., Chen, J. & Jiang, N. Assessing genome assembly quality using the LTR Assembly Index (LAI). Nucleic Acids Res. 46, e126 (2018).

[16]

Arumuganathan, K. & Earle, E. D. Nuclear DNA content of some important plant species. Plant Mol. Biol. Report. 9, 208-218 (1991).

[17]

Huerta-Cepas, J. et al. Fast genome-wide functional annotation through orthology assignment by eggNOG-Mapper. Mol. Biol. Evol. 34, 2115-2122 (2017).

[18]

El-Gebali, S. et al. The Pfam protein families database in 2019. Nucleic Acids Res. 47, D427-D432 (2019).

[19]

UniProt, C . UniProt: a worldwide hub of protein knowledge. Nucleic Acids Res. 47, D506-D515 (2019).

[20]

Kanehisa, M. et al. Data, information, knowledge and principle: back to metabolism in KEGG. Nucleic Acids Res. 42, D199-D205 (2014).

[21]

Harris, M. A. et al. The Gene Ontology (GO) database and informatics resource. Nucleic Acids Res. 32, D258-D261 (2004).

[22]

Koonin, E. V. et al. A comprehensive evolutionary classification of proteins encoded in complete eukaryotic genomes. Genome Biol. 5, R7 (2004).

[23]

Rawlings, N. D. et al. The MEROPS database of proteolytic enzymes, their substrates and inhibitors in 2017 and a comparison with peptidases in the PANTHER database. Nucleic Acids Res. 46, D624-D632 (2018).

[24]

Kall, L., Krogh, A . & Sonnhammer, E. L. A combined transmembrane topology and signal peptide prediction method. J. Mol. Biol. 338, 1027-1036 (2004).

[25]

Almagro Armenteros, J. J. et al. SignalP 5.0 improves signal peptide predictions using deep neural networks. Nat. Biotechnol. 37, 420-423 (2019).

[26]

Lombard, V., Golaconda Ramulu, H., Drula, E., Coutinho, P. M. & Henrissat, B. The carbohydrate-active enzymes database (CAZy) in 2013. Nucleic Acids Res. 42, D490-D495 (2014).

[27]

Paajanen, P. et al. A critical comparison of technologies for a plant genome sequencing project. Gigascience 8, giy163 (2019).

[28]

Yu, Y. et al. Genome re-sequencing reveals the evolutionary history of peach fruit edibility. Nat. Commun. 9, 5404 (2018).

[29]

Aranzana, M. J. et al. Prunus genetics and applications after de novo genome sequencing: achievements and prospects. Hortic. Res. 6, 58 (2019).

[30]

Li, Y. X. et al. Salicylic acid in Populus tomentosa is a remote signalling molecule induced by Botryosphaeria dothidea infection. Sci. Rep. 8, 14059 (2018).

[31]

Sharma, M. & Pandey, G. K. Expansion and function of repeat domain proteins during stress and development in plants. Front Plant Sci. 6, 1218 (2015).

[32]

Wei, H. et al. Comparative transcriptome analysis of genes involved in anthocyanin biosynthesis in the red and yellow fruits of sweet cherry (Prunus avium L.). PLoS ONE 10, e0121164 (2015).

[33]

Shen, X. J. et al. A role for PacMYBA in ABA-regulated anthocyanin biosynthesis in red-colored sweet cherry cv. Hong Deng (Prunus avium L.). Plant Cell Physiol. 55, 862-880 (2014).

[34]

Teribia, N., Tijero, V. & Munne-Bosch, S. Linking hormonal profiles with variations in sugar and anthocyanin contents during the natural development and ripening of sweet cherries. N. Biotechnol. 33, 824-833 (2016).

[35]

Wang, Y. P. et al. Transcriptional regulation of PaPYLs, PaPP2Cs and PaSnRK2s during sweet cherry fruit development and in response to abscisic acid and auxin at onset of fruit ripening. Plant Growth Regul. 75, 455-464 (2015).

[36]

Stern, R. A., Flaishman, M., Applebaum, S. & Ben-Arie, R. Effect of synthetic auxins on fruit development of ‘Bing’ cherry (Prunus avium L.). Sci. Horticulturae. 114, 275-280 (2007).

[37]

Berni, R., Cai, G., Xu, X., Hausman, J. F. & Guerriero, G. Identification of jasmonic acid biosynthetic genes in sweet cherry and expression analysis in four ancient varieties from tuscany. Int. J. Mol. Sci. 20, 10 (2019).

[38]

Kondo, S., Motoyama, M., Michiyama, H. & Kim, M. Roles of jasmonic acid in the development of sweet cherries as measured from fruit or disc samples. Plant Growth Regul. 37, 37-44 (2002).

[39]

Kondo, S., Tomiyama, A. & Seto, H. Changes of endogenous jasmonic acid and methyl jasmonate in apples and sweet cherries during fruit development. J. Am. Soc. Horticultural Sci. 125, 282-287 (2000).

[40]

Zimin, A. V. et al. The MaSuRCA genome assembler. Bioinformatics 29, 2669-2677 (2013).

[41]

Koren, S. et al. Canu: scalable and accurate long-read assembly via adaptive k-mer weighting and repeat separation. Genome Res. 27, 722-736 (2017).

[42]

Chen, Y. et al. Fast and accurate assembly of Nanopore reads via progressive error correction and adaptive read selection. bioRxiv. https://doi.org/10.1101/2020.02.01.930107 (2020).

[43]

Waterhouse, R. M. et al. BUSCO applications from quality assessments to gene prediction and phylogenomics. Mol. Biol. Evol 35, 543-548 (2017).

[44]

Roach, M. J., Schmidt, S. A. & Borneman, A. R. Purge Haplotigs: allelic contig reassignment for third-gen diploid genome assemblies. BMC Bioinforma. 19, 460 (2018).

[45]

Durand, N. C. et al. Juicer provides a one-click system for analyzing loop-resolution Hi-C experiments. Cell Syst. 3, 95-98 (2016).

[46]

Dudchenko, O. et al. De novo assembly of the Aedes aegypti genome using Hi-C yields chromosome-length scaffolds. Science 356, 92-95 (2017).

[47]

Marcais, G. & Kingsford, C. A fast, lock-free approach for efficient parallel counting of occurrences of k-mers. Bioinformatics 27, 764-770 (2011).

[48]

Vurture, G. W. et al. GenomeScope: fast reference-free genome profiling from short reads. Bioinformatics 33, 2202-2204 (2017).

[49]

Flynn, J. M. et al. RepeatModeler2: automated genomic discovery of transposable element families. bioRxiv. https://doi.org/10.1101/856591 (2019).

[50]

Tarailo-Graovac, M. & Chen, N. Using RepeatMasker to identify repetitive elements in genomic sequences. Curr. Protoc. Bioinformatics 25, Unit 4-10 (2009).

[51]

Hubley, R. et al. The Dfam database of repetitive DNA families. Nucleic Acids Res. 44, D81-D89 (2016).

[52]

Nawrocki, E. P. & Eddy, S. R. Infernal 1.1: 100-fold faster RNA homology searches. Bioinformatics 29, 2933-2935 (2013).

[53]

Kalvari, I. et al. Non-coding RNA analysis using the Rfam database. Curr. Protoc. Bioinforma. 62, e51 (2018).

[54]

Lagesen, K. et al. RNAmmer: consistent and rapid annotation of ribosomal RNA genes. Nucleic Acids Res. 35, 3100-3108 (2007).

[55]

Love, J. et al. nextgenusfs/funannotate: funannotate v1.7.2. Zenodo. https://doi.org/10.5281/zenodo.3594559 (2019).

[56]

Lomsadze, A., Ter-Hovhannisyan, V., Chernoff, Y. O. & Borodovsky, M. Gene identification in novel eukaryotic genomes by self-training algorithm. Nucleic Acids Res. 33, 6494-506 (2005).

[57]

Hoff, K. J. & Stanke, M. Predicting genes in single genomes with AUGUSTUS. Curr. Protoc. Bioinforma. 65, e57 (2019).

[58]

Testa, A. C., Hane, J. K., Ellwood, S. R. & Oliver, R. P. CodingQuarry: highly accurate hidden Markov model gene prediction in fungal genomes using RNA-seq transcripts. BMC Genomics. 16, 170 (2015).

[59]

Korf, I. Gene finding in novel genomes. BMC Bioinforma. 5, 59 (2004).

[60]

Majoros, W. H., Pertea, M. & Salzberg, S. L. TigrScan and GlimmerHMM: two open source ab initio eukaryotic gene-finders. Bioinformatics 20, 2878-2879 (2004).

[61]

Haas, B. J. et al. Improving the Arabidopsis genome annotation using maximal transcript alignment assemblies. Nucleic Acids Res. 31, 5654-5666 (2003).

[62]

Jones, P. et al. InterProScan 5: genome-scale protein function classification. Bioinformatics 30, 1236-1240 (2014).

[63]

Mitchell, A. L. et al. InterPro in 2019: improving coverage, classification and access to protein sequence annotations. Nucleic Acids Res. 47, D351-D360 (2019).

[64]

Moriya, Y., Itoh, M., Okuda, S., Yoshizawa, A. C. & Kanehisa, M. KAAS: an automatic genome annotation and pathway reconstruction server. Nucleic Acids Res. 35, W182-W185 (2007).

[65]

Emms, D. M. & Kelly, S. OrthoFinder: phylogenetic orthology inference for comparative genomics. Genome Biol. 20, 238 (2019).

[66]

Ye, J. et al. WEGO 2.0: a web tool for analyzing and plotting GO annotations, 2018 update. Nucleic Acids Res. 46, W71-W75 (2018).

[67]

Han, M. V., Thomas, G. W., Lugo-Martinez, J. & Hahn, M. W. Estimating gene gain and loss rates in the presence of error in genome assembly and annotation using CAFE 3. Mol. Biol. Evol. 30, 1987-1997 (2013).

[68]

Su, W., Sun, J., Shimizu, K. & Kadota, K. TCC-GUI: a Shiny-based application for differential expression analysis of RNA-Seq count data. BMC Res Notes 12, 133 (2019).

[69]

Xie, C. et al. KOBAS 2.0: a web server for annotation and identification of enriched pathways and diseases. Nucleic Acids Res. 39, W316-W322 (2011).

[70]

Tang, H. et al. Synteny and collinearity in plant genomes. Science 320, 486-488 (2008).

[71]

Krzywinski, M. et al. Circos: an information aesthetic for comparative genomics. Genome Res. 19, 1639-1645 (2009).

PDF (1345KB)

0

Accesses

0

Citation

Detail

Sections
Recommended

/