4.6 Article

Identification of single nucleotide polymorphisms from the transcriptome of an organism with a whole genome duplication

期刊

BMC BIOINFORMATICS
卷 14, 期 -, 页码 -

出版社

BMC
DOI: 10.1186/1471-2105-14-325

关键词

SNP; Polyploid; Rainbow trout; Genome duplication

资金

  1. USDA
  2. Agriculture and Food Research Initiative Competitive Grants from the USDA National Institute of Food and Agriculture [2009-35205-05067, 2011-67015-30091]
  3. National Institute of General Medical Sciences [T32GM083864]
  4. NIFA [2009-35205-05067, 582549] Funding Source: Federal RePORTER

向作者/读者索取更多资源

Background: The common ancestor of salmonid fishes, including rainbow trout (Oncorhynchus mykiss), experienced a whole genome duplication between 20 and 100 million years ago, and many of the duplicated genes have been retained in the trout genome. This retention complicates efforts to detect allelic variation in salmonid fishes. Specifically, single nucleotide polymorphism (SNP) detection is problematic because nucleotide variation can be found between the duplicate copies (paralogs) of a gene as well as between alleles. Results: We present a method of differentiating between allelic and paralogous (gene copy) sequence variants, allowing identification of SNPs in organisms with multiple copies of a gene or set of genes. The basic strategy is to: 1) identify windows of unique cDNA sequences with homology to each other, 2) compare these unique cDNAs if they are not shared between individuals (i.e. the cDNA is homozygous in one individual and homozygous for another cDNA in the other individual), and 3) give a SNP score value between zero and one to each candidate sequence variant based on six criteria. Using this strategy we were able to detect about seven thousand potential SNPs from the transcriptomes of several clonal lines of rainbow trout. When directly compared to a pre-validated set of SNPs in polyploid wheat, we were also able to estimate the false-positive rate of this strategy as 0 to 28% depending on parameters used. Conclusions: This strategy has an advantage over traditional techniques of SNP identification because another dimension of sequencing information is utilized. This method is especially well suited for identifying SNPs in polyploids, both outbred and inbred, but would tend to be conservative for diploid organisms.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.6
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据