Proteogenomic strategies for identification of aberrant cancer peptides using large‐scale next‐generation sequencing data. Issue 23 (17th November 2014)

Record Type:: Journal Article
Title:: Proteogenomic strategies for identification of aberrant cancer peptides using large‐scale next‐generation sequencing data. Issue 23 (17th November 2014)
Main Title:: Proteogenomic strategies for identification of aberrant cancer peptides using large‐scale next‐generation sequencing data
Authors:: Woo, Sunghee
Cha, Seong Won
Na, Seungjin
Guest, Clark
Liu, Tao
Smith, Richard D.
Rodland, Karin D.
Payne, Samuel
Bafna, Vineet
Pandey, Akhilesh
Pevzner, Pavel A.
Abstract:: <abstract abstract-type="main"> <title> <x xml:space="preserve">Abstract</x> </title> <p>Cancer is driven by the acquisition of somatic DNA lesions. Distinguishing the early driver mutations from subsequent passenger mutations is key to molecular subtyping of cancers, understanding cancer progression, and the discovery of novel biomarkers. The advances of genomics technologies (whole‐genome exome, and transcript sequencing, collectively referred to as NGS (next‐generation sequencing)) have fueled recent studies on somatic mutation discovery. However, the vision is challenged by the complexity, redundancy, and errors in genomic data, and the difficulty of investigating the proteome translated portion of aberrant genes using only genomic approaches. Combination of proteomic and genomic technologies are increasingly being employed. Various strategies have been employed to allow the usage of large‐scale NGS data for conventional MS/MS searches. This paper provides a discussion of applying different strategies relating to large database search, and FDR (false discovery rate) ‐based error control, and their implication to cancer proteogenomics. Moreover, it extends and develops the idea of a unified genomic variant database that can be searched by any MS sample. A total of 879 BAM files downloaded from TCGA repository were used to create a 4.34 GB unified FASTA database that contained <inline-formula><alternatives><inline-graphic mimetype="image" … (more)
Is Part Of:: Proteomics. Volume 14:Issue 23/24(2014)
Journal:: Proteomics
Issue:: Volume 14:Issue 23/24(2014)
Issue Display:: Volume 14, Issue 23/24 (2014)
Year:: 2014
Volume:: 14
Issue:: 23/24
Issue Sort Value:: 2014-0014-NaN-0000
Page Start:: 2719
Page End:: 2730
Publication Date:: 2014-11-17
Subjects:: Proteins -- Separation -- Periodicals
Bioinformatics -- Periodicals
Proteomics -- Periodicals
Genomes -- Periodicals
Molecular genetics -- Periodicals
572.605
Journal URLs:: http://onlinelibrary.wiley.com/journal/10.1002/(ISSN)1615-9861 ↗
http://onlinelibrary.wiley.com/ ↗
DOI:: 10.1002/pmic.201400206 ↗
Languages:: English
ISSNs:: 1615-9853
Deposit Type:: Legaldeposit
View Content:: Available online (eLD content is only available in our Reading Rooms) ↗
Physical Locations:: British Library DSC - 6936.178000
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store
Ingest File:: 3582.xml