The genome sequence and transcriptome of Potentilla micrantha and their comparison to Fragaria vesca (the woodland strawberry).
ABSTRACT: Background:The genus Potentilla is closely related to that of Fragaria, the economically important strawberry genus. Potentilla micrantha is a species that does not develop berries but shares numerous morphological and ecological characteristics with Fragaria vesca. These similarities make P. micrantha an attractive choice for comparative genomics studies with F. vesca. Findings:In this study, the P. micrantha genome was sequenced and annotated, and RNA-Seq data from the different developmental stages of flowering and fruiting were used to develop a set of gene predictions. A 327 Mbp sequence and annotation of the genome of P. micrantha, spanning 2674 sequence contigs, with an N50 size of 335,712, estimated to cover 80% of the total genome size of the species was developed. The genus Potentilla has a characteristically larger genome size than Fragaria, but the recovered sequence scaffolds were remarkably collinear at the micro-syntenic level with the genome of F. vesca, its closest sequenced relative. A total of 33,602 genes were predicted, and 95.1% of bench-marking universal single-copy orthologous genes were complete within the presented sequence. Thus, we argue that the majority of the gene-rich regions of the genome have been sequenced. Conclusions:Comparisons of RNA-Seq data from the stages of floral and fruit development revealed genes differentially expressed between P. micrantha and F. vesca.The data presented are a valuable resource for future studies of berry development in Fragaria and the Rosaceae and they also shed light on the evolution of genome size and organization in this family.
Project description:The woodland strawberry, Fragaria vesca (2n = 2x = 14), is a versatile experimental plant system. This diminutive herbaceous perennial has a small genome (240 Mb), is amenable to genetic transformation and shares substantial sequence identity with the cultivated strawberry (Fragaria × ananassa) and other economically important rosaceous plants. Here we report the draft F. vesca genome, which was sequenced to ×39 coverage using second-generation technology, assembled de novo and then anchored to the genetic linkage map into seven pseudochromosomes. This diploid strawberry sequence lacks the large genome duplications seen in other rosids. Gene prediction modeling identified 34,809 genes, with most being supported by transcriptome mapping. Genes critical to valuable horticultural traits including flavor, nutritional value and flowering time were identified. Macrosyntenic relationships between Fragaria and Prunus predict a hypothetical ancestral Rosaceae genome that had nine chromosomes. New phylogenetic analysis of 154 protein-coding genes suggests that assignment of Populus to Malvidae, rather than Fabidae, is warranted.
Project description:There is an increasing interest in berries, especially blackberries in the diet, because of recent reports of their health benefits due to their high content of flavonoids. A broad range of genomic tools are available for other Rosaceae species but these tools are still lacking in the Rubus genus, thus limiting gene discovery and the breeding of improved varieties.De novo RNA-seq of ripe blackberries grown under field conditions was performed using Illumina Hiseq 2000. Almost 9 billion nucleotide bases were sequenced in total. Following assembly, 42,062 consensus sequences were detected. For functional annotation, 33,040 (NR), 32,762 (NT), 21,932 (Swiss-Prot), 20,134 (KEGG), 13,676 (COG), 24,168 (GO) consensus sequences were annotated using different databases; in total 34,552 annotated sequences were identified. For protein prediction analysis, the number of coding DNA sequences (CDS) that mapped to the protein database was 32,540. Non redundant (NR), annotation showed that 25,418 genes (73.5%) has the highest similarity with Fragaria vesca subspecies vesca. Reanalysis was undertaken by aligning the reads with this reference genome for a deeper analysis of the transcriptome. We demonstrated that de novo assembly, using Trinity and later annotation with Blast using different databases, were complementary to alignment to the reference sequence using SOAPaligner/SOAP2. The Fragaria reference genome belongs to a species in the same family as blackberry (Rosaceae) but to a different genus. Since blackberries are tetraploids, the possibility of artefactual gene chimeras resulting from mis-assembly was tested with one of the genes sequenced by RNAseq, Chalcone Synthase (CHS). cDNAs encoding this protein were cloned and sequenced. Primers designed to the assembled sequences accurately distinguished different contigs, at least for chalcone synthase genes.We prepared and analysed transcriptome data from ripe blackberries, for which prior genomic information was limited. This new sequence information will improve the knowledge of this important and healthy fruit, providing an invaluable new tool for biological research.
Project description:BACKGROUND: The cultivated strawberry Fragaria xananassa is one of the most economically-important soft-fruit species. Few structural genomic resources have been reported for Fragaria and there exists an urgent need for the development of physical mapping resources for the genus. The first stage in the development of a physical map for Fragaria is the construction and characterisation of a high molecular weight bacterial artificial chromosome (BAC) library. METHODS: A BAC library, consisting of 18,432 clones was constructed from Fragaria vesca f. semperflorens accession 'Ali Baba'. BAC DNA from individual library clones was pooled to create a PCR-based screening assay for the library, whereby individual clones could be identified with just 34 PCR reactions. These pools were used to screen the BAC library and anchor individual clones to the diploid Fragaria reference map (FVxFN). FINDINGS: Clones from the BAC library developed contained an average insert size of 85 kb, representing over seven genome equivalents. The pools and superpools developed were used to identify a set of BAC clones containing 70 molecular markers previously mapped to the diploid Fragaria FVxFN reference map. The number of positive colonies identified for each marker suggests the library represents between 4x and 10x coverage of the diploid Fragaria genome, which is in accordance with the estimate of library coverage based on average insert size. CONCLUSION: This BAC library will be used for the construction of a physical map for F. vesca and the superpools will permit physical anchoring of molecular markers using PCR.
Project description:Whole-genome duplications are radical evolutionary events that have driven speciation and adaptation in many taxa. Higher-order polyploids have complex histories often including interspecific hybridization and dynamic genomic changes. This chromosomal reshuffling is poorly understood for most polyploid species, despite their evolutionary and agricultural importance, due to the challenge of distinguishing homologous sequences from each other. Here, we use dense linkage maps generated with targeted sequence capture to improve the diploid strawberry (Fragaria vesca) reference genome and to disentangle the subgenomes of the wild octoploid progenitors of cultivated strawberry, Fragaria virginiana and Fragaria chiloensis. Our novel approach, POLiMAPS (Phylogenetics Of Linkage-Map-Anchored Polyploid Subgenomes), leverages sequence reads to associate informative interhomeolog phylogenetic markers with linkage groups and reference genome positions. In contrast to a widely accepted model, we find that one of the four subgenomes originates with the diploid cytoplasm donor F. vesca, one with the diploid Fragaria iinumae, and two with an unknown ancestor close to F. iinumae. Extensive unidirectional introgression has converted F. iinumae-like subgenomes to be more F. vesca-like, but never the reverse, due either to homoploid hybridization in the F. iinumae-like diploid ancestors or else strong selection spreading F. vesca-like sequence among subgenomes through homeologous exchange. In addition, divergence between homeologous chromosomes has been substantially augmented by interchromosomal rearrangements. Our phylogenetic approach reveals novel aspects of the complicated web of genetic exchanges that occur during polyploid evolution and suggests a path forward for unraveling other agriculturally and ecologically important polyploid genomes.
Project description:BACKGROUND:The diploid woodland strawberry (Fragaria vesca) is an attractive system for functional genomics studies. Its small stature, fast regeneration time, efficient transformability and small genome size, together with substantial EST and genomic sequence resources make it an ideal reference plant for Fragaria and other herbaceous perennials. Most importantly, this species shares gene sequence similarity and genomic microcolinearity with other members of the Rosaceae family, including large-statured tree crops (such as apple, peach and cherry), and brambles and roses as well as with the cultivated octoploid strawberry, F. xananassa. F. vesca may be used to quickly address questions of gene function relevant to these valuable crop species. Although some F. vesca lines have been shown to be substantially homozygous, in our hands plants in purportedly homozygous populations exhibited a range of morphological and physiological variation, confounding phenotypic analyses. We also found the genotype of a named variety, thought to be well-characterized and even sold commercially, to be in question. An easy to grow, standardized, inbred diploid Fragaria line with documented genotype that is available to all members of the research community will facilitate comparison of results among laboratories and provide the research community with a necessary tool for functionally testing the large amount of sequence data that will soon be available for peach, apple, and strawberry. RESULTS:A highly inbred line, YW5AF7, of a diploid strawberry Fragaria vesca f. semperflorens line called "Yellow Wonder" (Y2) was developed and examined. Botanical descriptors were assessed for morphological characterization of this genotype. The plant line was found to be rapidly transformable using established techniques and media formulations. CONCLUSION:The development of the documented YW5AF7 line provides an important tool for Rosaceae functional genomic analyses. These day-neutral plants have a small genome, a seed to seed cycle of 3.0 - 3.5 months, and produce fruit in 7.5 cm pots in a growth chamber. YW5AF7 is runnerless and therefore easy to maintain in the greenhouse, forms abundant branch crowns for vegetative propagation, and produces highly aromatic yellow fruit throughout the year in the greenhouse. F. vesca can be transformed with Agrobacterium tumefaciens, making these plants suitable for insertional mutagenesis, RNAi and overexpression studies that can be compared against a stable baseline of phenotypic descriptors and can be readily genetically substantiated.
Project description:The cultivated strawberry (Fragaria × ananassa) is an octoploid (2n = 8x = 56) of the Rosaceae family whose genomic architecture is still controversial. Several recent studies support the AAA'A'BBB'B' model, but its complexity has hindered genetic and genomic analysis of this important crop. To overcome this difficulty and to assist genome-wide analysis of F. × ananassa, we constructed an integrated linkage map by organizing a total of 4474 of simple sequence repeat (SSR) markers collected from published Fragaria sequences, including 3746 SSR markers [Fragaria vesca expressed sequence tag (EST)-derived SSR markers] derived from F. vesca ESTs, 603 markers (F. × ananassa EST-derived SSR markers) from F. × ananassa ESTs, and 125 markers (F. × ananassa transcriptome-derived SSR markers) from F. × ananassa transcripts. Along with the previously published SSR markers, these markers were mapped onto five parent-specific linkage maps derived from three mapping populations, which were then assembled into an integrated linkage map. The constructed map consists of 1856 loci in 28 linkage groups (LGs) that total 2364.1 cM in length. Macrosynteny at the chromosome level was observed between the LGs of F. × ananassa and the genome of F. vesca. Variety distinction on 129 F. × ananassa lines was demonstrated using 45 selected SSR markers.
Project description:Mikania micrantha is one of the top 100 worst invasive species that can cause serious damage to natural ecosystems and substantial economic losses. Here, we present its 1.79?Gb chromosome-scale reference genome. Half of the genome is composed of long terminal repeat retrotransposons, 80% of which have been derived from a significant expansion in the past one million years. We identify a whole genome duplication event and recent segmental duplications, which may be responsible for its rapid environmental adaptation. Additionally, we show that M. micrantha achieves higher photosynthetic capacity by CO2 absorption at night to supplement the carbon fixation during the day, as well as enhanced stem photosynthesis efficiency. Furthermore, the metabolites of M. micrantha can increase the availability of nitrogen by enriching the microbes that participate in nitrogen cycling pathways. These findings collectively provide insights into the rapid growth and invasive adaptation.
Project description:The genus Fragaria encompasses species at ploidy levels ranging from diploid to decaploid. The cultivated strawberry, Fragaria×ananassa, and its two immediate progenitors, F. chiloensis and F. virginiana, are octoploids. To elucidate the ancestries of these octoploid species, we performed a phylogenetic analysis using intron-containing sequences of the nuclear ADH-1 gene from 39 germplasm accessions representing nineteen Fragaria species and one outgroup species, Dasiphora fruticosa. All trees from Maximum Parsimony and Maximum Likelihood analyses showed two major clades, Clade A and Clade B. Each of the sampled octoploids contributed alleles to both major clades. All octoploid-derived alleles in Clade A clustered with alleles of diploid F. vesca, with the exception of one octoploid allele that clustered with the alleles of diploid F. mandshurica. All octoploid-derived alleles in clade B clustered with the alleles of only one diploid species, F. iinumae. When gaps encoded as binary characters were included in the Maximum Parsimony analysis, tree resolution was improved with the addition of six nodes, and the bootstrap support was generally higher, rising above the 50% threshold for an additional nine branches. These results, coupled with the congruence of the sequence data and the coded gap data, validate and encourage the employment of sequence sets containing gaps for phylogenetic analysis. Our phylogenetic conclusions, based upon sequence data from the ADH-1 gene located on F. vesca linkage group II, complement and generally agree with those obtained from analyses of protein-encoding genes GBSSI-2 and DHAR located on F. vesca linkage groups V and VII, respectively, but differ from a previous study that utilized rDNA sequences and did not detect the ancestral role of F. iinumae.
Project description:Two fundamental questions on how invasive species are able to rapidly colonize novel habitat have emerged. One asks whether a negative correlation exists between the genetic diversity of invasive populations and their geographic distance from the origin of introduction. The other is whether selection on the chloroplast genome is important driver of adaptation to novel soil environments. Here, we addressed these questions in a study of the noxious invasive weed, <i>Mikania micrantha</i>, which has rapidly expanded in to southern China after being introduced to Hong Kong in 1884. Seven chloroplast simple sequence repeats (cpSSRs) were used to investigate population genetics in 28 populations of <i>M. micrantha</i>, which produced 39 loci. The soil compositions for these populations, including Mg abundance, were measured. The results showed that <i>M. micrantha</i> possessed relatively high cpSSR variation and differentiation among populations. Multiple diversity indices were quantified, and none was significantly correlated with distance from the origin of introduction. No evidence for "isolation by distance," significant spatial structure, bottlenecks, nor linkage disequilibrium was detected. We also were unable to identify loci on the chloroplast genome that exhibited patterns of differentiation that would suggest adaptive evolution in response to soil attributes. Soil Mg had only a genome-wide effect instead of being a selective factor, which highlighted the association between Mg and the successful invasion. This study characterizes the role of the chloroplast genome of <i>M. micrantha</i> during its recent invasion of southern China.