Genome-wide characterization of simple sequence repeats in cucumber (Cucumis sativus L.)
- PMID: 20950470
- PMCID: PMC3091718
- DOI: 10.1186/1471-2164-11-569
Genome-wide characterization of simple sequence repeats in cucumber (Cucumis sativus L.)
Abstract
Background: Cucumber, Cucumis sativus L. is an important vegetable crop worldwide. Until very recently, cucumber genetic and genomic resources, especially molecular markers, have been very limited, impeding progress of cucumber breeding efforts. Microsatellites are short tandemly repeated DNA sequences, which are frequently favored as genetic markers due to their high level of polymorphism and codominant inheritance. Data from previously characterized genomes has shown that these repeats vary in frequency, motif sequence, and genomic location across taxa. During the last year, the genomes of two cucumber genotypes were sequenced including the Chinese fresh market type inbred line '9930' and the North American pickling type inbred line 'Gy14'. These sequences provide a powerful tool for developing markers in a large scale. In this study, we surveyed and characterized the distribution and frequency of perfect microsatellites in 203 Mbp assembled Gy14 DNA sequences, representing 55% of its nuclear genome, and in cucumber EST sequences. Similar analyses were performed in genomic and EST data from seven other plant species, and the results were compared with those of cucumber.
Results: A total of 112,073 perfect repeats were detected in the Gy14 cucumber genome sequence, accounting for 0.9% of the assembled Gy14 genome, with an overall density of 551.9 SSRs/Mbp. While tetranucleotides were the most frequent microsatellites in genomic DNA sequence, dinucleotide repeats, which had more repeat units than any other SSR type, had the highest cumulative sequence length. Coding regions (ESTs) of the cucumber genome had fewer microsatellites compared to its genomic sequence, with trinucleotides predominating in EST sequences. AAG was the most frequent repeat in cucumber ESTs. Overall, AT-rich motifs prevailed in both genomic and EST data. Compared to the other species examined, cucumber genomic sequence had the highest density of SSRs (although comparable to the density of poplar, grapevine and rice), and was richest in AT dinucleotides. Using an electronic PCR strategy, we investigated the polymorphism between 9930 and Gy14 at 1,006 SSR loci, and found unexpectedly high degree of polymorphism (48.3%) between the two genotypes. The level of polymorphism seems to be positively associated with the number of repeat units in the microsatellite. The in silico PCR results were validated empirically in 660 of the 1,006 SSR loci. In addition, primer sequences for more than 83,000 newly-discovered cucumber microsatellites, and their exact positions in the Gy14 genome assembly were made publicly available.
Conclusions: The cucumber genome is rich in microsatellites; AT and AAG are the most abundant repeat motifs in genomic and EST sequences of cucumber, respectively. Considering all the species investigated, some commonalities were noted, especially within the monocot and dicot groups, although the distribution of motifs and the frequency of certain repeats were characteristic of the species examined. The large number of SSR markers developed from this study should be a significant contribution to the cucurbit research community.
Figures





Similar articles
-
A second generation framework for the analysis of microsatellites in expressed sequence tags and the development of EST-SSR markers for a conifer, Cryptomeria japonica.BMC Genomics. 2012 Apr 16;13:136. doi: 10.1186/1471-2164-13-136. BMC Genomics. 2012. PMID: 22507374 Free PMC article.
-
Transcriptome sequencing and comparative analysis of cucumber flowers with different sex types.BMC Genomics. 2010 Jun 17;11:384. doi: 10.1186/1471-2164-11-384. BMC Genomics. 2010. PMID: 20565788 Free PMC article.
-
Eucalyptus microsatellites mined in silico: survey and evaluation.J Genet. 2008 Apr;87(1):21-5. doi: 10.1007/s12041-008-0003-9. J Genet. 2008. PMID: 18560170
-
Recent progress on the molecular breeding of Cucumis sativus L. in China.Theor Appl Genet. 2020 May;133(5):1777-1790. doi: 10.1007/s00122-019-03484-0. Epub 2019 Nov 21. Theor Appl Genet. 2020. PMID: 31754760 Review.
-
Streamlining of Simple Sequence Repeat Data Mining Methodologies and Pipelines for Crop Scanning.Plants (Basel). 2024 Sep 19;13(18):2619. doi: 10.3390/plants13182619. Plants (Basel). 2024. PMID: 39339594 Free PMC article. Review.
Cited by
-
A 1,681-locus consensus genetic map of cultivated cucumber including 67 NB-LRR resistance gene homolog and ten gene loci.BMC Plant Biol. 2013 Mar 25;13:53. doi: 10.1186/1471-2229-13-53. BMC Plant Biol. 2013. PMID: 23531125 Free PMC article.
-
Genome-wide characterization and linkage mapping of simple sequence repeats in mei (Prunus mume Sieb. et Zucc.).PLoS One. 2013;8(3):e59562. doi: 10.1371/journal.pone.0059562. Epub 2013 Mar 28. PLoS One. 2013. PMID: 23555708 Free PMC article.
-
Cytogenetic Diversity of Simple Sequences Repeats in Morphotypes of Brassica rapa ssp. chinensis.Front Plant Sci. 2016 Jul 26;7:1049. doi: 10.3389/fpls.2016.01049. eCollection 2016. Front Plant Sci. 2016. PMID: 27507974 Free PMC article.
-
Development and characterization of genomic and expressed SSRs in citrus by genome-wide analysis.PLoS One. 2013 Oct 28;8(10):e75149. doi: 10.1371/journal.pone.0075149. eCollection 2013. PLoS One. 2013. PMID: 24204572 Free PMC article.
-
Exploring the genes of yerba mate (Ilex paraguariensis A. St.-Hil.) by NGS and de novo transcriptome assembly.PLoS One. 2014 Oct 16;9(10):e109835. doi: 10.1371/journal.pone.0109835. eCollection 2014. PLoS One. 2014. PMID: 25330175 Free PMC article.
References
-
- Powell W, Morgante M, Andre C, Henfey M, Vogel J, Tingy S, Rafalsky A. The comparison of RFLP, RAPD, AFLP and SSR (microsatellite) markers for germplasm analysis. Mol Breed. 1996;2:225–238. doi: 10.1007/BF00564200. - DOI
MeSH terms
Substances
LinkOut - more resources
Full Text Sources
Research Materials
Miscellaneous