Improving Imputation Quality in BEAGLE for Crop and Livestock Data. Issue 1 (1st January 2020)
- Record Type:
- Journal Article
- Title:
- Improving Imputation Quality in BEAGLE for Crop and Livestock Data. Issue 1 (1st January 2020)
- Main Title:
- Improving Imputation Quality in BEAGLE for Crop and Livestock Data
- Authors:
- Pook, Torsten
Mayer, Manfred
Geibel, Johannes
Weigend, Steffen
Cavero, David
Schoen, Chris C
Simianer, Henner - Abstract:
- Abstract: Imputation is one of the key steps in the preprocessing and quality control protocol of any genetic study. Most imputation algorithms were originally developed for the use in human genetics and thus are optimized for a high level of genetic diversity. Different versions of BEAGLE were evaluated on genetic datasets of doubled haploids of two European maize landraces, a commercial breeding line and a diversity panel in chicken, respectively, with different levels of genetic diversity and structure which can be taken into account in BEAGLE by parameter tuning. Especially for phasing BEAGLE 5.0 outperformed the newest version (5.1) which in turn also lead to improved imputation. Earlier versions were far more dependent on the adaption of parameters in all our tests. For all versions, the parameter ne (effective population size) had a major effect on the error rate for imputation of ungenotyped markers, reducing error rates by up to 98.5%. Further improvement was obtained by tuning of the parameters affecting the structure of the haplotype cluster that is used to initialize the underlying Hidden Markov Model of BEAGLE. The number of markers with extremely high error rates for the maize datasets were more than halved by the use of a flint reference genome (F7, PE0075 etc.) instead of the commonly used B73. On average, error rates for imputation of ungenotyped markers were reduced by 8.5% by excluding genetically distant individuals from the reference panel for theAbstract: Imputation is one of the key steps in the preprocessing and quality control protocol of any genetic study. Most imputation algorithms were originally developed for the use in human genetics and thus are optimized for a high level of genetic diversity. Different versions of BEAGLE were evaluated on genetic datasets of doubled haploids of two European maize landraces, a commercial breeding line and a diversity panel in chicken, respectively, with different levels of genetic diversity and structure which can be taken into account in BEAGLE by parameter tuning. Especially for phasing BEAGLE 5.0 outperformed the newest version (5.1) which in turn also lead to improved imputation. Earlier versions were far more dependent on the adaption of parameters in all our tests. For all versions, the parameter ne (effective population size) had a major effect on the error rate for imputation of ungenotyped markers, reducing error rates by up to 98.5%. Further improvement was obtained by tuning of the parameters affecting the structure of the haplotype cluster that is used to initialize the underlying Hidden Markov Model of BEAGLE. The number of markers with extremely high error rates for the maize datasets were more than halved by the use of a flint reference genome (F7, PE0075 etc.) instead of the commonly used B73. On average, error rates for imputation of ungenotyped markers were reduced by 8.5% by excluding genetically distant individuals from the reference panel for the chicken diversity panel. To optimize imputation accuracy one has to find a balance between representing as much of the genetic diversity as possible while avoiding the introduction of noise by including genetically distant individuals. … (more)
- Is Part Of:
- G3. Volume 10:Issue 1(2020)
- Journal:
- G3
- Issue:
- Volume 10:Issue 1(2020)
- Issue Display:
- Volume 10, Issue 1 (2020)
- Year:
- 2020
- Volume:
- 10
- Issue:
- 1
- Issue Sort Value:
- 2020-0010-0001-0000
- Page Start:
- 177
- Page End:
- 188
- Publication Date:
- 2020-01-01
- Subjects:
- imputation -- BEAGLE -- reference panel -- reference genome
Genetics -- Research -- Periodicals
Genomics -- Periodicals
Genetics
Genomics
Genes
Genetics -- Research
Genomics
Electronic journals
Periodical
Periodicals
Fulltext
Internet Resources
Periodicals
572.8 - Journal URLs:
- https://academic.oup.com/g3journal ↗
http://bibpurl.oclc.org/web/43467 ↗
http://www.g3journal.org ↗
http://www.oxfordjournals.org/ ↗ - DOI:
- 10.1534/g3.119.400798 ↗
- Languages:
- English
- ISSNs:
- 2160-1836
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 22172.xml