Tag Archives: SNP

Aqua Gen and Center for Integrative Genomics (CIGENE) Collaborate With Affymetrix on a Salmon Genoty

By Business Wirevia The Motley Fool

Filed under:

Aqua Gen and Center for Integrative Genomics (CIGENE) Collaborate With Affymetrix on a Salmon Genotyping Array to Achieve the World’s First High-Density View of the Marker Patterns in the Atlantic Salmon

ÅS, Norway & SANTA CLARA, Calif.–(BUSINESS WIRE)– Aqua Gen, a member of the Erich Wesjohann (EW) Group GmBH and Center for Integrative Genomics (CIGENE) at the Norwegian University of Life Sciences (UMB) in collaboration with Affymetrix, Inc. (NAS: AFFX) announced today that they are the first to genotype more than 900,000 markers per sample from the Atlantic salmon (Salmo salar), thereby achieving the capability to implement genomic selection and improve their salmon breeding program at Aqua Gen.

Aqua Gen, a leader in selective breeding, manages a large scale selective breeding program for Atlantic salmon, a major contributor to the world’s aquaculture production. Aqua Gen has also been pioneering the use of marker-assisted selection in aquaculture breeding through their highly successful QTL-innOva products. CIGENE, located at the Norwegian University of Life Sciences, maintains research programs that contribute to the understanding of the biology of aquaculture and plant production, focusing in particular on the genetics of complex traits of economic and ecological significance. Aqua Gen and CIGENE partnered with Affymetrix to develop the salmon genotyping screening array which consists of 923,627 SNP (Single Nucleotide Polymorphism) markers and includes both diploid and tetraploid sequence variants. The goal of the ongoing study is to identify relevant and polymorphic high resolution markers that can be used downstream for marker trait association studies, genomic selection programs, as well as for a wide variety of applications in genetics and ecology.

“We are thrilled to have access to this first of its kind and groundbreaking marker map of the Atlantic salmon in such a short time,” said Dr. Nina Santi, Director, Research and Development at Aqua Gen. “Aqua Gen is always looking for new technologies that can assist in the development of genetic material to meet the increasing demand for cost-effective and sustainable seafood production. This high-density SNP array gives us entirely new possibilities for improving the disease resistance and robustness of farmed Atlantic salmon. In particular, we are enthusiastic about the prospects of implementing so-called genomic selection in our breeding program. This array will also facilitate the identification of causal genetic variants underlying complex traits, thus contributing greatly to our understanding of salmonid biology.”

“Development of SNP arrays and automated genotyping in Atlantic salmon is complicated by the autotetraploid whole genome duplication that occurred in the common ancestor of extant salmonids,” said Dr. Sigbjørn Lien, Professor and Assistant …read more
Source: FULL ARTICLE at DailyFinance

Genetic variation controls predation: Benefits of being a mosaic

A genetically mosaic Eucalyptus tree is able to control which leaves are saved from predation because of alterations in its genes, finds an study published in BioMed Central’s open access journal BMC Plant Biology. Between two leaves of the same tree there can be many genetic differences – this study found ten SNP, including ones in genes that regulate terpene production, which influence whether or not a leaf is edible. …read more
Source: FULL ARTICLE at Phys.org

Comparing multiple substrings for a match

I have a tab-delimited file containing a large genetic dataset with binary base calls, in this format:

Code:
Chr7 26021407 1/1:0,0,0:5 1/1:0,0,0:5 1/1:0,0,0:5
Chr7 26022023 1/1:0,0,0:3 1/1:0,0,0:3 1/1:28,3,0:5
Chr7 26022087 1/1:0,0,0:6 1/1:25,3,0:9 1/1:25,3,0:9
Chr7 26022656 1/1:0,0,0:3 1/1:27,3,0:5 1/1:0,0,0:3
Chr7 26022752 1/1:21,3,0:5 0/1:0,0,0:3 1/1:24,3,0:5
Chr7 26022759 0/1:15,3,0:4 0/1:0,0,0:3 0/1:18,3,0:4
Chr7 26022873 1/1:36,3,0:7 1/1:0,0,0:4 1/1:16,3,0:7
Chr7 26022940 1/1:0,0,0:5 1/1:28,3,0:8 1/1:14,3,0:8
Chr7 26023652 1/1:0,0,0:6 1/1:0,0,0:6 1/1:25,3,0:8
The 2 leading columns are coordinate information, then there are some columns with other information (omitted from this example), and the SNP data (the data of interest, columns beginning 0/0, 0/1, or 1/1) begin at column 10. The number of subsequent columns matches the number of samples.

I would like to identify and eliminate any line in which all of the SNP call data (the leading 0/0,0/1, or 1/1) match across the dataset, and are therefore uninformative in my analyses. What I have in mind would function like this:

Code:
awk ‘{if (substr($10,1,3)==substr($11,1,3) && substr($10,1,3)==substr($12,1,3)…. [include all columns from $10 to end]) {next}} {print}’ [infile]
But I want to be able to specify a range of substrings to compare across, rather than entering each one manually, which seems tedious and unnecessary. Any help with this would be hugely appreciated.

thanks a lot!
Source: The UNIX and Linux Forums