首页 | 本学科首页   官方微博 | 高级检索  
相似文献
 共查询到20条相似文献,搜索用时 31 毫秒
1.
To improve the efficiency of breeding of Miscanthus for biomass yield, there is a need to develop genomics‐assisted selection for this long‐lived perennial crop by relating genotype to phenotype and breeding value across a broad range of environments. We present the first genome‐wide association (GWA) and genomic prediction study of Miscanthus that utilizes multilocation phenotypic data. A panel of 568 Miscanthus sinensis accessions was genotyped with 46,177 single nucleotide polymorphisms (SNPs) and evaluated at one subtropical and five temperate locations over 3 years for biomass yield and 14 yield‐component traits. GWA and genomic prediction were performed separately for different years of data in order to assess reproducibility. The analyses were also performed for individual field trial locations, as well as combined phenotypic data across groups of locations. GWA analyses identified 27 significant SNPs for yield, and a total of 504 associations across 298 unique SNPs across all traits, sites, and years. For yield, the greatest number of significant SNPs was identified by combining phenotypic data across all six locations. For some of the other yield‐component traits, greater numbers of significant SNPs were obtained from single site data, although the number of significant SNPs varied greatly from site to site. Candidate genes were identified. Accounting for population structure, genomic prediction accuracies for biomass yield ranged from 0.31 to 0.35 across five northern sites and from 0.13 to 0.18 for the subtropical location, depending on the estimation method. Genomic prediction accuracies of all traits were similar for single‐location and multilocation data, suggesting that genomic selection will be useful for breeding broadly adapted M. sinensis as well as M. sinensis optimized for specific climates. All of our data, including DNA sequences flanking each SNP, are publicly available. By facilitating genomic selection in M. sinensis and Miscanthus × giganteus, our results will accelerate the breeding of these species for biomass in diverse environments.  相似文献   

2.
Next‐generation sequencing and the collection of genome‐wide data allow identifying adaptive variation and footprints of directional selection. Using a large SNP data set from 259 RAD‐sequenced European eel individuals (glass eels) from eight locations between 34 and 64oN, we examined the patterns of genome‐wide genetic diversity across locations. We tested for local selection by searching for increased population differentiation using FST‐based outlier tests and by testing for significant associations between allele frequencies and environmental variables. The overall low genetic differentiation found (FST = 0.0007) indicates that most of the genome is homogenized by gene flow, providing further evidence for genomic panmixia in the European eel. The lack of genetic substructuring was consistent at both nuclear and mitochondrial SNPs. Using an extensive number of diagnostic SNPs, results showed a low occurrence of hybrids between European and American eel, mainly limited to Iceland (5.9%), although individuals with signatures of introgression several generations back in time were found in mainland Europe. Despite panmixia, a small set of SNPs showed high genetic differentiation consistent with single‐generation signatures of spatially varying selection acting on glass eels. After screening 50 354 SNPs, a total of 754 potentially locally selected SNPs were identified. Candidate genes for local selection constituted a wide array of functions, including calcium signalling, neuroactive ligand–receptor interaction and circadian rhythm. Remarkably, one of the candidate genes identified is PERIOD, possibly related to differences in local photoperiod associated with the >30° difference in latitude between locations. Genes under selection were spread across the genome, and there were no large regions of increased differentiation as expected when selection occurs within just a single generation due to panmixia. This supports the conclusion that most of the genome is homogenized by gene flow that removes any effects of diversifying selection from each new generation.  相似文献   

3.
Detecting genetic variants under selection using FST outlier analysis (OA) and environmental association analyses (EAAs) are popular approaches that provide insight into the genetic basis of local adaptation. Despite the frequent use of OA and EAA approaches and their increasing attractiveness for detecting signatures of selection, their application to field‐based empirical data have not been synthesized. Here, we review 66 empirical studies that use Single Nucleotide Polymorphisms (SNPs) in OA and EAA. We report trends and biases across biological systems, sequencing methods, approaches, parameters, environmental variables and their influence on detecting signatures of selection. We found striking variability in both the use and reporting of environmental data and statistical parameters. For example, linkage disequilibrium among SNPs and numbers of unique SNP associations identified with EAA were rarely reported. The proportion of putatively adaptive SNPs detected varied widely among studies, and decreased with the number of SNPs analysed. We found that genomic sampling effort had a greater impact than biological sampling effort on the proportion of identified SNPs under selection. OA identified a higher proportion of outliers when more individuals were sampled, but this was not the case for EAA. To facilitate repeatability, interpretation and synthesis of studies detecting selection, we recommend that future studies consistently report geographical coordinates, environmental data, model parameters, linkage disequilibrium, and measures of genetic structure. Identifying standards for how OA and EAA studies are designed and reported will aid future transparency and comparability of SNP‐based selection studies and help to progress landscape and evolutionary genomics.  相似文献   

4.
Anadromous Atlantic salmon (Salmo salar) is a species of major conservation and management concern in North America, where population abundance has been declining over the past 30 years. Effective conservation actions require the delineation of conservation units to appropriately reflect the spatial scale of intraspecific variation and local adaptation. Towards this goal, we used the most comprehensive genetic and genomic database for Atlantic salmon to date, covering the entire North American range of the species. The database included microsatellite data from 9142 individuals from 149 sampling locations and data from a medium‐density SNP array providing genotypes for >3000 SNPs for 50 sampling locations. We used neutral and putatively selected loci to integrate adaptive information in the definition of conservation units. Bayesian clustering with the microsatellite data set and with neutral SNPs identified regional groupings largely consistent with previously published regional assessments. The use of outlier SNPs did not result in major differences in the regional groupings, suggesting that neutral markers can reflect the geographic scale of local adaptation despite not being under selection. We also performed assignment tests to compare power obtained from microsatellites, neutral SNPs and outlier SNPs. Using SNP data substantially improved power compared to microsatellites, and an assignment success of 97% to the population of origin and of 100% to the region of origin was achieved when all SNP loci were used. Using outlier SNPs only resulted in minor improvements to assignment success to the population of origin but improved regional assignment. We discuss the implications of these new genetic resources for the conservation and management of Atlantic salmon in North America.  相似文献   

5.
High‐density SNP genotyping arrays can be designed for any species given sufficient sequence information of high quality. Two high‐density SNP arrays relying on the Infinium iSelect technology (Illumina) were designed for use in the conifer white spruce (Picea glauca). One array contained 7338 segregating SNPs representative of 2814 genes of various molecular functional classes for main uses in genetic association and population genetics studies. The other one contained 9559 segregating SNPs representative of 9543 genes for main uses in population genetics, linkage mapping of the genome and genomic prediction. The SNPs assayed were discovered from various sources of gene resequencing data. SNPs predicted from high‐quality sequences derived from genomic DNA reached a genotyping success rate of 64.7%. Nonsingleton in silico SNPs (i.e. a sequence polymorphism present in at least two reads) predicted from expressed sequenced tags obtained with the Roche 454 technology and Illumina GAII analyser resulted in a similar genotyping success rate of 71.6% when the deepest alignment was used and the most favourable SNP probe per gene was selected. A variable proportion of these SNPs was shared by other nordic and subtropical spruce species from North America and Europe. The number of shared SNPs was inversely proportional to phylogenetic divergence and standing genetic variation in the recipient species, but positively related to allele frequency in P. glauca natural populations. These validated SNP resources should open up new avenues for population genetics and comparative genetic mapping at a genomic scale in spruce species.  相似文献   

6.
Arabidopsis thaliana inhabits diverse climates and exhibits varied phenology across its range. Although A. thaliana is an extremely well‐studied model species, the relationship between geography, growing season climate and its genetic variation is poorly characterized. We used redundancy analysis (RDA) to quantify the association of genomic variation [214 051 single nucleotide polymorphisms (SNPs)] with geography and climate among 1003 accessions collected from 447 locations in Eurasia. We identified climate variables most correlated with genomic variation, which may be important selective gradients related to local adaptation across the species range. Climate variation among sites of origin explained slightly more genomic variation than geographical distance. Large‐scale spatial gradients and early spring temperatures explained the most genomic variation, while growing season and summer conditions explained the most after controlling for spatial structure. SNP variation in Scandinavia showed the greatest climate structure among regions, possibly because of relatively consistent phenology and life history of populations in this region. Climate variation explained more variation among nonsynonymous SNPs than expected by chance, suggesting that much of the climatic structure of SNP correlations is due to changes in coding sequence that may underlie local adaptation.  相似文献   

7.
Single-nucleotide polymorphisms (SNPs) and insertion–deletions (INDELs) are currently the important classes of genetic markers for major crop species. In this study, methods for developing SNP markers in rapeseed (Brassica napus L.) and their in silico mapping and use for genotyping are demonstrated. For the development of SNP and INDEL markers, 181 fragments from 121 different gene sequences spanning 86 kb were examined. A combination of different selection methods (genome-specific amplification, hetero-duplex analysis and sequence analysis) allowed the detection of 18 singular fragments that showed a total of 87 SNPs and 6 INDELs between 6 different rapeseed varieties. The average frequency of sequence polymorphism was estimated to be one SNP every 247 bp and one INDEL every 3,583 bp. Most SNPs and INDELs were found in non-coding regions. Polymorphism information content values for SNP markers ranged between 0.02 and 0.50 in a set of 86 varieties. Using comparative genetics data for B. napus and Arabidopsis thaliana, an allocation of SNP markers to linkage groups in rapeseed was achieved: a unique location was determined for seven gene sequences; two and three possible locations were found for six and four sequences, respectively. The results demonstrate the usefulness of existing genomic resources for SNP discovery in rapeseed.  相似文献   

8.
Knowledge about the sequence-based genetic diversity of a crop species is important in order to develop highly informative genotyping assays, which will eventually positively impact breeding practice. Diversity data were obtained from two pools of 185 and 75 accessions each, representing most of the species belonging to the genus Malus, by re-sequencing 27 gene-specific amplicons and by screening 237 Malus × domestica SNPs using the multiplex genotyping technology SNPlex™. Nucleotide diversity and insertion/deletion rates in M. × domestica were estimated as π = 0.0037 and 1/333 bp, respectively. The SNP frequency was estimated as 0.0194 (1 SNP/52 bp) while within a single apple cultivar an average of one SNP in every 455 bp was found. We also investigated transferability (T SNP) of the heterozygous state of SNPs across the species M. × domestica and the genus Malus. Raw re-sequencing showed that 12–15% of M. × domestica SNPs are transferable to a second M. × domestica cultivar, however T SNP rose to ∼41% with SNPs selected for high minor allele frequency. T SNP of chosen SNPs averaged ∼27% in the two M. × domestica-related species, Malus sieversii and Malus sylvestris, but was much lower in more distantly related species. On the basis of T SNP, simulations, and empirical results, we calculated that a close-design, multiplexed genotyping array with at least 2,000 SNPs is required for building a highly saturated linkage maps within any M. × domestica cross. The same array would gradually lose informativeness in increasingly phylogenetically distant Malus species.  相似文献   

9.
10.
Four custom Axiom genotyping arrays were designed for a genome-wide association (GWA) study of 100,000 participants from the Kaiser Permanente Research Program on Genes, Environment and Health. The array optimized for individuals of European race/ethnicity was previously described. Here we detail the development of three additional microarrays optimized for individuals of East Asian, African American, and Latino race/ethnicity. For these arrays, we decreased redundancy of high-performing SNPs to increase SNP capacity. The East Asian array was designed using greedy pairwise SNP selection. However, removing SNPs from the target set based on imputation coverage is more efficient than pairwise tagging. Therefore, we developed a novel hybrid SNP selection method for the African American and Latino arrays utilizing rounds of greedy pairwise SNP selection, followed by removal from the target set of SNPs covered by imputation. The arrays provide excellent genome-wide coverage and are valuable additions for large-scale GWA studies.  相似文献   

11.
Thyroid hormones play an important role in regulating metabolism and can affect homeostasis of fat deposition. The gene encoding thyroglobulin (TG), producing the precursor for thyroid hormones, has been proposed as a positional and functional candidate gene for a QTL with an effect on fat deposition. In the present study, we identified 6 novel SNPs at the 3′ flanking region of the TG gene. The SNP marker association analysis indicated that the T354C, G392A, A430G and T433G SNP markers were significantly associated with marbling score (P < 0.05). Animals with the new homozygote genotype had higher marbling score than those with the other genotypes. Otherwise, the linkage disequilibrium analysis indicated that these four SNPs were completely linked (r 2 = 1). Results from this study suggest that TG gene-specific SNP may be a useful marker for meat quality traits in future marker assisted selection programs in beef cattle.  相似文献   

12.
There are two categories of immune responses – innate and adaptive immunity – both having polygenic backgrounds and a significant environmental component. In our study, adaptive immunity was represented by the specific antibody response toward keyhole limpet hemocyanin (KLH); innate immunity was represented by natural antibodies toward lipopolysaccharide (LPS) and lipoteichoic acid (LTA). Defining genetic bases of immune responses leads from defining quantitative trait loci (QTL) toward a single mutation responsible for variation in the phenotypic trait. The goal of the reported study was to define candidate genes and mutations for the immune traits of interest in chicken by performing an association study of SNPs located in candidate genes defined in QTL regions. Candidate genes and SNPs in QTL regions were selected in silico. SNP association was based on a custom SNP panel, GoldenGate genotyping assay (Illumina) and two statistical models: random mixed model and CAR score. The most significant SNP for immune response toward KLH was located in the JMJD6 gene located on GGA18. Four SNPs in candidate genes FOXJ1 (GGA18), EPHB1 (GGA9), PTGER4 (GGAZ) and PRKCB (GGA14) showed association with natural antibodies for LPS. A single SNP in ITGB4 (GGA18) was associated with natural antibodies for LTA. All associated SNPs mentioned above showed additive effects.  相似文献   

13.
Domestication and selection for important performance traits can impact the genome, which is most often reflected by reduced heterozygosity in and surrounding genes related to traits affected by selection. In this study, analysis of the genomic impact caused by domestication and artificial selection was conducted by investigating the signatures of selection using single nucleotide polymorphisms (SNPs) in channel catfish (Ictalurus punctatus). A total of 8.4 million candidate SNPs were identified by using next generation sequencing. On average, the channel catfish genome harbors one SNP per 116 bp. Approximately 6.6 million, 5.3 million, 4.9 million, 7.1 million and 6.7 million SNPs were detected in the Marion, Thompson, USDA103, Hatchery strain, and wild population, respectively. The allele frequencies of 407,861 SNPs differed significantly between the domestic and wild populations. With these SNPs, 23 genomic regions with putative selective sweeps were identified that included 11 genes. Although the function for the majority of the genes remain unknown in catfish, several genes with known function related to aquaculture performance traits were included in the regions with selective sweeps. These included hypoxia-inducible factor 1β· HIFιβ ¨ and the transporter gene ATP-binding cassette sub-family B member 5 (ABCB5). HIF1β· is important for response to hypoxia and tolerance to low oxygen levels is a critical aquaculture trait. The large numbers of SNPs identified from this study are valuable for the development of high-density SNP arrays for genetic and genomic studies of performance traits in catfish.  相似文献   

14.
Advances in sequencing technology have led to a rapid rise in the genomic data available for plants, driving new insights into the evolution, domestication and improvement of crops. Single nucleotide polymorphisms (SNPs) are a major component of crop genomic diversity, and are invaluable as genetic markers in research and breeding programs. High‐throughput SNP arrays, or ‘SNP chips’, can generate reproducible sets of informative SNP markers and have been broadly adopted. Although there are many public repositories for sequencing data, which are routinely uploaded, there are no formal repositories for crop SNP array data. To make SNP array data more easily accessible, we have developed CropSNPdb ( http://snpdb.appliedbioinformatics.com.au ), a database for SNP array data produced by the Illumina Infinium? hexaploid bread wheat (Triticum aestivum) 90K and Brassica 60K arrays. We currently host SNPs from datasets covering 526 Brassica lines and 309 bread wheat lines, and provide search, download and upload utilities for users. CropSNPdb provides a useful repository for these data, which can be applied for a range of genomics and molecular crop‐breeding activities.  相似文献   

15.
Recent advances in high‐throughput sequencing technologies have offered the possibility to generate genomewide sequence data to delineate previously unidentified genetic structure, obtain more accurate estimates of demographic parameters and to evaluate potential adaptive divergence. Here, we identified 27 556 single nucleotide polymorphisms for the small yellow croaker (Larimichthys polyactis) using restriction‐site‐associated DNA (RAD) sequencing of 24 individuals from two populations. Significant sources of genetic variation were identified, with an average nucleotide diversity (π) of 0.00105 ± 0.000425 across individuals, and long‐term effective population size was thus estimated to range between 26 172 and 261 716. According to the results, no differentiation between the two populations was detected based on the SNP data set of top quality score per contig or neutral loci. However, the two analysed populations were highly differentiated based on SNP data set of both top FST value per contig and the outlier SNPs. Moreover, local adaptation was highlighted by an FST‐based outlier tests implemented in LOSITAN and a total of 538 potentially locally selected SNPs were identified. blast2go annotation of contigs containing the outlier SNPs yielded hits for 37 (66%) of 56 significant blastx matches. Candidate genes for local adaptation constituted a wide array of biological functions, including cellular response to oxidative stress, actin filament binding, ion transmembrane transport and synapse assembly. The generated SNP resources in this study provided a valuable tool for future population genetics and genomics studies of L. polyactis.  相似文献   

16.
17.
The genus Agapornis, or lovebirds, are popular pet parrots worldwide. Currently, breeders are dependent on pedigree records as a selection tool as no molecular parentage verification test is available for any of the nine species. The A. roseicollis reference genome was recently assembled. This was followed by the sequencing of the whole genomes of the parents of the reference genome individual at 30× coverage. The parents’ reads were mapped against the reference genome to identify SNPs. Over 1.6 million SNPs, shared between the parents, were discovered using the Genome Analysis Toolkit pipeline. SNPs were filtered to a panel of 480 SNPs based on Genome Analysis Toolkit parameters. The panel of 480 SNPs was genotyped in a population of 960 lovebirds across seven species. A panel of 262 SNPs was compiled that included SNPs successfully amplified across all species. The 262‐SNP panel was reduced based on the observed heterozygosity (HO) and minor allele frequency (MAF) values per SNP to include the lowest number of SNPs with the highest exclusion power for parentage verification. Two smaller panels consisting of 195 SNPs with MAF and HO values >0.1 and 40 SNPs with MAF and HO values >0.3, were constructed. The panels were verified using 43 families from different species with known relationships to evaluate the exclusion power of each panel. The 195 SNP panel with an average exclusion probability of 99.9% and MAF and HO values >0.1 was proposed as the routine Agapornis parentage verification panel.  相似文献   

18.
Next-generation sequencing has transformed the fields of ecological and evolutionary genetics by allowing for cost-effective identification of genome-wide variation. Single nucleotide polymorphism (SNP) arrays, or “SNP chips”, enable very large numbers of individuals to be consistently genotyped at a selected set of these identified markers, and also offer the advantage of being able to analyse samples of variable DNA quality. We used reduced representation restriction-aided digest sequencing (RAD-seq) of 31 birds of the threatened hihi (Notiomystis cincta; stitchbird) and low-coverage whole genome sequencing (WGS) of 10 of these birds to develop an Affymetrix 50 K SNP chip. We overcame the limitations of having no hihi reference genome and a low quantity of sequence data by separate and pooled de novo assembly of each of the 10 WGS birds. Reads from all individuals were mapped back to these de novo assemblies to identify SNPs. A subset of RAD-seq and WGS SNPs were selected for inclusion on the chip, prioritising SNPs with the highest quality scores whose flanking sequence uniquely aligned to the zebra finch (Taeniopygia guttata) genome. Of the 58,466 SNPs manufactured on the chip, 72% passed filtering metrics and were polymorphic. By genotyping 1,536 hihi on the array, we found that SNPs detected in multiple assemblies were more likely to successfully genotype, representing a cost-effective approach to identify SNPs for genotyping. Here, we demonstrate the utility of the SNP chip by describing the high rates of linkage disequilibrium in the hihi genome, reflecting the history of population bottlenecks in the species.  相似文献   

19.
Whole genome resequencing of 51 Populus nigra (L.) individuals from across Western Europe was performed using Illumina platforms. A total number of 1 878 727 SNPs distributed along the P. nigra reference sequence were identified. The SNP calling accuracy was validated with Sanger sequencing. SNPs were selected within 14 previously identified QTL regions, 2916 expressional candidate genes related to rust resistance, wood properties, water‐use efficiency and bud phenology and 1732 genes randomly spread across the genome. Over 10 000 SNPs were selected for the construction of a 12k Infinium Bead‐Chip array dedicated to association mapping. The SNP genotyping assay was performed with 888 P. nigra individuals. The genotyping success rate was 91%. Our high success rate was due to the discovery panel design and the stringent parameters applied for SNP calling and selection. In the same set of P. nigra genotypes, linkage disequilibrium throughout the genome decayed on average within 5–7 kb to half of its maximum value. As an application test, ADMIXTURE analysis was performed with a selection of 600 SNPs spread throughout the genome and 706 individuals collected along 12 river basins. The admixture pattern was consistent with genetic diversity revealed by neutral markers and the geographical distribution of the populations. These newly developed SNP resources and genotyping array provide a valuable tool for population genetic studies and identification of QTLs through natural‐population based genetic association studies in P. nigra.  相似文献   

20.
Use of SNPs has been favoured due to their abundance in plant and animal genomes, accompanied by the falling cost and rising throughput capacity for detection and genotyping. Here, we present in vitro (obtained from targeted sequencing) and in silico discovery of SNPs, and the design of medium‐throughput genotyping arrays for two oyster species, the Pacific oyster, Crassostrea gigas, and European flat oyster, Ostrea edulis. Two sets of 384 SNP markers were designed for two Illumina GoldenGate arrays and genotyped on more than 1000 samples for each species. In each case, oyster samples were obtained from wild and selected populations and from three‐generation families segregating for traits of interest in aquaculture. The rate of successfully genotyped polymorphic SNPs was about 60% for each species. Effects of SNP origin and quality on genotyping success (Illumina functionality Score) were analysed and compared with other model and nonmodel species. Furthermore, a simulation was made based on a subset of the C. gigas SNP array with a minor allele frequency of 0.3 and typical crosses used in shellfish hatcheries. This simulation indicated that at least 150 markers were needed to perform an accurate parental assignment. Such panels might provide valuable tools to improve our understanding of the connectivity between wild (and selected) populations and could contribute to future selective breeding programmes.  相似文献   

设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号