Protein Sequencing and Identification by Mass Spectrometry

Size: px
Start display at page:

Download "Protein Sequencing and Identification by Mass Spectrometry"

Transcription

1 Protein Sequencing and Identification by Mass Spectrometry

2 Tandem Mass Spectrometry De Novo Peptide Sequencing Spectrum Graph Protein Identification via Database Search Identifying Post Translationally Modified Peptides Spectral Convolution Spectral Alignment

3

4 H...-HN-CH-CO-NH-CH-CO-NH-CH-CO- OH R i-1 R i R i+1 N-terminus C-terminus AA residue i-1 AA residue i AA residue i+1

5 Collision Induced Dissociation H + H...-HN-CH-CO... NH-CH-CO-NH-CH-CO- OH R i-1 R i R i+1 Prefix Fragment Suffix Fragment Peptides tend to fragment along the backbone. Fragments can also loose neutral chemical groups like NH 3 and H 2 O.

6 Proteases, e.g. trypsin, break protein into peptides. A Tandem Mass Spectrometer further breaks the peptides down into fragment ions and measures the mass of each piece. Mass Spectrometer accelerates the fragmented ions; heavier ions accelerate slower than lighter ones. Mass Spectrometer measure mass/charge ratio of an ion.

7

8 Peptide Mass (D) = 415 Peptide without Mass (D) = 397

9

10

11

12 Reconstruct peptide from the set of masses of fragment ions (mass-spectrum)

13 b 2 -H 2 O b 3 - NH 3 a b a b 3 HO NH + 3 R O R O R O R H -- N --- C --- C --- N --- C --- C --- N --- C --- C --- N --- C -- COOH H H H H H H H y 3 y 2 y 1 y 3 -H 2 O y 2 - NH 3

14 G V D L K H 2 O K L D V G 57 Da = G 99 Da = V mass 0 The peaks in the mass spectrum: Prefix and Suffix Fragments. Fragments with neutral losses (-H 2 O, -NH 3 ) Noise and missing peaks.

15 G V D L K Peptide MS/MS Identification: Intensity mass 0

16

17 GTDIMR MPSERGTDIMRPAKID... PAKID MPSER HPLC To MS/MS protein peptides

18 Matrix-Assisted Laser Desorption/Ionization (MALDI) From lectures by Vineet Bafna (UCSD)

19 Relative Abundance RT: LC Time(min) NL: 1.52E8 BasePeak F: + c Full ms [ ] S#: 1707 RT: AV: 1 NL: 2.41E7 F: + c Full ms [ ] Relative Abundance m/z MS Scan S#: 1708 RT: AV: 1 NL: 5.27E6 T: + c d Full ms [ ] collision Ion MS-1 cell MS-2 Relative Abundance MS/MS Source Scan m/z

20 S e MS/MS instrument S#: 1708 RT: AV: 1 NL: 5.27E6 T: + c d Full ms [ ] q u e n c e Database search Sequest de Novo interpretation Sherenga Relative Abundance m/z

21 Tandem Mass Spectrometry (MS/MS): mainly generates partial N- and C-terminal peptides Spectrum consists of different ion types because peptides can be broken in several places. Chemical noise often complicates the spectrum. Represented in 2-D: mass/charge axis vs. intensity axis

22 S # : R T : AV: 1 N L : E 6 T : + c d F u ll m s [ ] Database Relative Abundance De Novo 3 0 Search m /z Mass, Score Database of known peptides MDERHILNM, KLQWVCSDL, PTYWASDL, ENQIKRSACVM, TLACHGGEM, NGALPQWRT, HLLERTKMNVV, GGPASSDA, GGLITGMQSD, MQPLMNWE, ALKIIMNVRT, AVGELTK, HEWAILF, GHNLWAMNAC, GVFGSVLRA, EKLNKAATYIN.. A A L C V V W P G G W Database of all peptides = 20 n AAAAAAAA,AAAAAAAC,AAAAAAAD,AAAAAAAE, AAAAAAAG,AAAAAAAF,AAAAAAAH,AAAAAAI, AVGELTI, AVGELTK, AVGELTL, AVGELTM, YYYYYYYS,YYYYYYYT,YYYYYYYV,YYYYYYYY E E T L L D T R T K K AVGELTK

23 De Novo vs. Database Search: A Paradox The database of all peptides is huge O(20 n ). The database of all known peptides is much smaller O(10 8 ). However, de novo algorithms can be much faster, even though their search space is much larger! A database search scans all peptides in the database of all known peptides search space to find best one. De novo eliminates the need to scan database of all peptides by modeling the problem as a graph search.

24 S#: 1708 RT: AV: 1 NL: 5.27E6 T: + c d Full ms [ ] Relative Abundance m/z Sequence

25

26 (cont d)

27 (cont d)

28 How to create vertices (from masses) How to create edges (from mass differences) How to score paths How to find best path

29 b Mass/Charge (M/Z)

30 a Mass/Charge (M/Z)

31 a is an ion type shift in b Mass/Charge (M/Z)

32 y Mass/Charge (M/Z)

33 Mass/Charge (M/Z)

34 Mass/Charge (M/Z)

35 noise Mass/Charge (M/Z)

36 Mass/Charge (M/z)

37 Some Mass Differences between Peaks Correspond to Amino Acids u q e s e u q e n c n e e s e q u e n c e s c e

38 Some masses correspond to fragment ions, others are just random noise Knowing ion types ={δ 1, δ 2,, δ k } lets us distinguish fragment ions from noise We can learn ion types δ i and their probabilities q i by analyzing a large test sample of annotated spectra.

39 ={δ 1, δ 2,, δ k } Ion types {b, b-nh 3, b-h 2 O} correspond to ={0, 17, 18} *Note: In reality the δ value of ion type b is -1 but we will hide it for the sake of simplicity

40 The match between two spectra is the number of masses (peaks) they share (Shared Peak Count or SPC) In practice mass-spectrometrists use the weighted SPC that reflects intensities of the peaks Match between experimental and theoretical spectra is defined similarly

41 Goal: Find a peptide with maximal match between an experimental and theoretical spectrum. Input: S: experimental spectrum : set of possible ion types m: parent mass Output: P: peptide with mass m, whose theoretical spectrum matches the experimental S spectrum the best

42 Masses of potential N-terminal peptides Vertices are generated by reverse shifts corresponding to ion types ={δ 1, δ 2,, δ k } Every N-terminal peptide can generate up to k ions m-δ 1, m-δ 2,, m-δ k Every mass s in an MS/MS spectrum generates k vertices V(s) = {s+δ, s+δ,, s+δ } 1 2 k corresponding to potential N-terminal peptides Vertices of the spectrum graph: {initial vertex} V(s ) V(s )... V(s ) {terminal vertex} 1 2 m

43 Shift in H 2 O Shift in H 2 O+NH 3

44 Two vertices with mass difference corresponding to an amino acid A: Connect with an edge labeled by A Gap edges for di- and tri-peptides

45 Path in the labeled graph spell out amino acid sequences There are many paths, how to find the correct one? We need scoring to evaluate paths

46 p(p,s) = probability that peptide P produces spectrum S= {s,s, s } 1 2 q p(p, s) = the probability that peptide P generates a peak s Scoring = computing probabilities p(p,s) = π sєs p(p, s)

47 For a position t that represents ion type d j : p(p,s t ) = q j, if peak is generated at t 1-q j, otherwise

48 (cont d) For a position t that is not associated with an ion type: q, if peak is generated at t R p (P,s ) = R t 1-q, otherwise R q = the probability of a noisy peak that does R not correspond to any ion type

49 Finding Optimal Paths in the Spectrum Graph For a given MS/MS spectrum S, find a peptide P maximizing p(p,s) over all possible peptides P: p(p',s) = max P p(p,s) Peptides = paths in the spectrum graph P = the optimal path in the spectrum graph

50 Tandem mass spectrometry is characterized by a set of ion types {δ,δ,..,δ } and their 1 2 k probabilities {q,...,q } 1 k δ -ions of a partial peptide are produced i independently with probabilities q i

51 A peptide has all k peaks with probability k i= 1 q i k and no peaks with probability i= 1 (1 q i ) A peptide also produces a ``random noise'' with uniform probability q R in any position.

52 Ratio Test Scoring for Partial Peptides Incorporates premiums for observed ions and penalties for missing ions. Example: for k=4, assume that for a partial peptide P we only see ions δ 1,δ 2,δ 4. The score is calculated as: q q (1 q ) q q q (1 q ) q R R R R

53 T- set of all positions. T i ={t δ1,, t δ2,...,,t δk, }- set of positions that represent ions of partial peptides P i. A peak at position t δj is generated with probability q j. R=T- U T i - set of positions that are not associated with any partial peptides (noise).

54 For a position t δj T i the probability p(t, P,S) that peptide P produces a peak at position t. P( t, P, S ) = q j if a peak is generatedat position t δ j 1 q j otherwise Similarly, for t R, the probability that P produces a random noise peak at t is: P R ( t ) = 1 q R q R if a peak is generated at position t otherwise

55 For a peptide P with n amino acids, the score for the whole peptides is expressed by the following ratio test: = = = n i k j i R i R j j t p S P t p S p S P p 1 1 ) ( ),, ( ) ( ), ( δ δ

56 How can we find the amino acid sequence that best explains the spectrum? F S N A M S D I a) Explains the largest number of peaks V SGQ L I D b) Explains the most intensity Exhaustive enumeration? 1. Generate every possible sequence with the same peptide mass 2. Match each sequence to the spectrum 3. Choose the sequence that explains the most intensity in the spectrum Takes too long!

57 Computer programs for de-novo interpretation of MS/MS spectra date as far back as 1966 when Biemman et al. proposed a prefix extension algorithm: Tries every possible prefix extension Eliminates solutions with missing peaks Outputs every peptide with a matching parent mass ALL prefix peaks must be present in the spectrum In an attempt to include interpretations with missing peaks, Sakurai et al. (1984) proposed: Exhaustive search over the space of all possible permutations of amino acid multisets where the total mass equals the parent mass of the MS/MS spectrum Score peptide/spectrum matches by counting prefix and suffix mass matches Naturally more sensitive but also very slow

58 The second half of the eighties saw a few better designed approaches to this problem, based on the same type of algorithm: Prefix extension, one amino acid at a time. Tolerate missing peaks. Include both prefix and suffix peaks in the score. User-specified maximum number of candidates in memory at any point of the execution. (sub-optimal) Ishikawa and Niwa 86, Siegel and Baumann 88, Johnson and Biemann 89, Zidarov et al. 90

59 In 1990 Bartels introduced a graph representation of an MS/MS spectrum Every peak in the spectrum defines a vertex Vertices connected by an edge if peak mass difference is an amino acid mass v 0 v M(S) Best peptide is defined as the best path between the two endpoint vertices: v to v 0 M(S) No detailed algorithm was given for finding the best peptide; interactive exploration tool was made available.

60 What is the DP recursion? Score(i) = intensity(i) + max( Score(j) ), for all j with mass(i)-mass(j) Amino acid masses Recovered de-novo sequence? ESESE

Lecture 15: Realities of Genome Assembly Protein Sequencing

Lecture 15: Realities of Genome Assembly Protein Sequencing Lecture 15: Realities of Genome Assembly Protein Sequencing Study Chapter 8.10-8.15 1 Euler s Theorems A graph is balanced if for every vertex the number of incoming edges equals to the number of outgoing

More information

Mass spectrometry in proteomics

Mass spectrometry in proteomics I519 Introduction to Bioinformatics, Fall, 2013 Mass spectrometry in proteomics Haixu Tang School of Informatics and Computing Indiana University, Bloomington Modified from: www.bioalgorithms.info Outline

More information

Protein Sequencing and Identification by Mass Spectrometry

Protein Sequencing and Identification by Mass Spectrometry Protein Sequencing and Identification by Mass Spectrometry Tandem Mass Spectrometry De Novo Peptide Sequencing Spectrum Graph Protein Identification via Database Search Identifying Post Translationally

More information

Protein Sequencing and Identification by Mass Spectrometry

Protein Sequencing and Identification by Mass Spectrometry Protein Sequencing and Identification by Mass Spectrometry Outline Tandem Mass Spectrometry De Novo Peptide Sequencing Spectrum Graph Protein Identification via Database Search Identifying Post Translationally

More information

Was T. rex Just a Big Chicken? Computational Proteomics

Was T. rex Just a Big Chicken? Computational Proteomics Was T. rex Just a Big Chicken? Computational Proteomics Phillip Compeau and Pavel Pevzner adjusted by Jovana Kovačević Bioinformatics Algorithms: an Active Learning Approach 215 by Compeau and Pevzner.

More information

De Novo Peptide Sequencing

De Novo Peptide Sequencing De Novo Peptide Sequencing Outline A simple de novo sequencing algorithm PTM Other ion types Mass segment error De Novo Peptide Sequencing b 1 b 2 b 3 b 4 b 5 b 6 b 7 b 8 A NELLLNVK AN ELLLNVK ANE LLLNVK

More information

Computational Methods For Identification Of Cyclic Peptides Using Mass Spectrometry. Julio Ng Bioinformatics Program, UCSD March, 26 th 2010

Computational Methods For Identification Of Cyclic Peptides Using Mass Spectrometry. Julio Ng Bioinformatics Program, UCSD March, 26 th 2010 Computational Methods or Identification Of Cyclic Peptides Using Mass Spectrometry Julio Ng Bioinformatics Program, UCSD March, 26 th 2010 Outline Importance of natural products Mass spectrometry on cyclic

More information

De novo Protein Sequencing by Combining Top-Down and Bottom-Up Tandem Mass Spectra. Xiaowen Liu

De novo Protein Sequencing by Combining Top-Down and Bottom-Up Tandem Mass Spectra. Xiaowen Liu De novo Protein Sequencing by Combining Top-Down and Bottom-Up Tandem Mass Spectra Xiaowen Liu Department of BioHealth Informatics, Department of Computer and Information Sciences, Indiana University-Purdue

More information

Protein Identification Using Tandem Mass Spectrometry. Nathan Edwards Informatics Research Applied Biosystems

Protein Identification Using Tandem Mass Spectrometry. Nathan Edwards Informatics Research Applied Biosystems Protein Identification Using Tandem Mass Spectrometry Nathan Edwards Informatics Research Applied Biosystems Outline Proteomics context Tandem mass spectrometry Peptide fragmentation Peptide identification

More information

Proteomics. November 13, 2007

Proteomics. November 13, 2007 Proteomics November 13, 2007 Acknowledgement Slides presented here have been borrowed from presentations by : Dr. Mark A. Knepper (LKEM, NHLBI, NIH) Dr. Nathan Edwards (Center for Bioinformatics and Computational

More information

Tandem Mass Spectrometry: Generating function, alignment and assembly

Tandem Mass Spectrometry: Generating function, alignment and assembly Tandem Mass Spectrometry: Generating function, alignment and assembly With slides from Sangtae Kim and from Jones & Pevzner 2004 Determining reliability of identifications Can we use Target/Decoy to estimate

More information

SPECTRA LIBRARY ASSISTED DE NOVO PEPTIDE SEQUENCING FOR HCD AND ETD SPECTRA PAIRS

SPECTRA LIBRARY ASSISTED DE NOVO PEPTIDE SEQUENCING FOR HCD AND ETD SPECTRA PAIRS SPECTRA LIBRARY ASSISTED DE NOVO PEPTIDE SEQUENCING FOR HCD AND ETD SPECTRA PAIRS 1 Yan Yan Department of Computer Science University of Western Ontario, Canada OUTLINE Background Tandem mass spectrometry

More information

Computational Methods for Mass Spectrometry Proteomics

Computational Methods for Mass Spectrometry Proteomics Computational Methods for Mass Spectrometry Proteomics Eidhammer, Ingvar ISBN-13: 9780470512975 Table of Contents Preface. Acknowledgements. 1 Protein, Proteome, and Proteomics. 1.1 Primary goals for studying

More information

CSE182-L8. Mass Spectrometry

CSE182-L8. Mass Spectrometry CSE182-L8 Mass Spectrometry Project Notes Implement a few tools for proteomics C1:11/2/04 Answer MS questions to get started, select project partner, select a project. C2:11/15/04 (All but web-team) Plan

More information

via Tandem Mass Spectrometry and Propositional Satisfiability De Novo Peptide Sequencing Renato Bruni University of Perugia

via Tandem Mass Spectrometry and Propositional Satisfiability De Novo Peptide Sequencing Renato Bruni University of Perugia De Novo Peptide Sequencing via Tandem Mass Spectrometry and Propositional Satisfiability Renato Bruni bruni@diei.unipg.it or bruni@dis.uniroma1.it University of Perugia I FIMA International Conference

More information

Mass spectrometry has been used a lot in biology since the late 1950 s. However it really came into play in the late 1980 s once methods were

Mass spectrometry has been used a lot in biology since the late 1950 s. However it really came into play in the late 1980 s once methods were Mass spectrometry has been used a lot in biology since the late 1950 s. However it really came into play in the late 1980 s once methods were developed to allow the analysis of large intact (bigger than

More information

12/3/13. The Dynamic Nature of the Proteome

12/3/13. The Dynamic Nature of the Proteome I519 Introduction to Bioinformatics, Fall, 213 Mass spctromtry in protomics Haixu Tang School of Informatics and Computing Indiana Univrsity, Bloomington Modifid from: www.bioalgorithms.info Outlin Protomics

More information

Nature Methods: doi: /nmeth Supplementary Figure 1. Fragment indexing allows efficient spectra similarity comparisons.

Nature Methods: doi: /nmeth Supplementary Figure 1. Fragment indexing allows efficient spectra similarity comparisons. Supplementary Figure 1 Fragment indexing allows efficient spectra similarity comparisons. The cost and efficiency of spectra similarity calculations can be approximated by the number of fragment comparisons

More information

Other Methods for Generating Ions 1. MALDI matrix assisted laser desorption ionization MS 2. Spray ionization techniques 3. Fast atom bombardment 4.

Other Methods for Generating Ions 1. MALDI matrix assisted laser desorption ionization MS 2. Spray ionization techniques 3. Fast atom bombardment 4. Other Methods for Generating Ions 1. MALDI matrix assisted laser desorption ionization MS 2. Spray ionization techniques 3. Fast atom bombardment 4. Field Desorption 5. MS MS techniques Matrix assisted

More information

MassHunter TOF/QTOF Users Meeting

MassHunter TOF/QTOF Users Meeting MassHunter TOF/QTOF Users Meeting 1 Qualitative Analysis Workflows Workflows in Qualitative Analysis allow the user to only see and work with the areas and dialog boxes they need for their specific tasks

More information

Efficiency of Database Search for Identification of Mutated and Modified Proteins via Mass Spectrometry

Efficiency of Database Search for Identification of Mutated and Modified Proteins via Mass Spectrometry Methods Efficiency of Database Search for Identification of Mutated and Modified Proteins via Mass Spectrometry Pavel A. Pevzner, 1,3 Zufar Mulyukov, 1 Vlado Dancik, 2 and Chris L Tang 2 Department of

More information

De Novo Peptide Sequencing: Informatics and Pattern Recognition applied to Proteomics

De Novo Peptide Sequencing: Informatics and Pattern Recognition applied to Proteomics De Novo Peptide Sequencing: Informatics and Pattern Recognition applied to Proteomics John R. Rose Computer Science and Engineering University of South Carolina 1 Overview Background Information Theoretic

More information

MS-MS Analysis Programs

MS-MS Analysis Programs MS-MS Analysis Programs Basic Process Genome - Gives AA sequences of proteins Use this to predict spectra Compare data to prediction Determine degree of correctness Make assignment Did we see the protein?

More information

A Dynamic Programming Approach to De Novo Peptide Sequencing via Tandem Mass Spectrometry

A Dynamic Programming Approach to De Novo Peptide Sequencing via Tandem Mass Spectrometry A Dynamic Programming Approach to De Novo Peptide Sequencing via Tandem Mass Spectrometry Ting Chen Department of Genetics arvard Medical School Boston, MA 02115, USA Ming-Yang Kao Department of Computer

More information

DE NOVO PEPTIDE SEQUENCING FOR MASS SPECTRA BASED ON MULTI-CHARGE STRONG TAGS

DE NOVO PEPTIDE SEQUENCING FOR MASS SPECTRA BASED ON MULTI-CHARGE STRONG TAGS DE NOVO PEPTIDE SEQUENCING FO MASS SPECTA BASED ON MULTI-CHAGE STONG TAGS KANG NING, KET FAH CHONG, HON WAI LEONG Department of Computer Science, National University of Singapore, 3 Science Drive 2, Singapore

More information

Tutorial 1: Setting up your Skyline document

Tutorial 1: Setting up your Skyline document Tutorial 1: Setting up your Skyline document Caution! For using Skyline the number formats of your computer have to be set to English (United States). Open the Control Panel Clock, Language, and Region

More information

De novo peptide sequencing methods for tandem mass. spectra

De novo peptide sequencing methods for tandem mass. spectra De novo peptide sequencing methods for tandem mass spectra A Thesis Submitted to the College of Graduate Studies and Research in Partial Fulfillment of the Requirements for the degree of Doctor of Philosophy

More information

QuasiNovo: Algorithms for De Novo Peptide Sequencing

QuasiNovo: Algorithms for De Novo Peptide Sequencing University of South Carolina Scholar Commons Theses and Dissertations 2013 QuasiNovo: Algorithms for De Novo Peptide Sequencing James Paul Cleveland University of South Carolina Follow this and additional

More information

TUTORIAL EXERCISES WITH ANSWERS

TUTORIAL EXERCISES WITH ANSWERS TUTORIAL EXERCISES WITH ANSWERS Tutorial 1 Settings 1. What is the exact monoisotopic mass difference for peptides carrying a 13 C (and NO additional 15 N) labelled C-terminal lysine residue? a. 6.020129

More information

MS-based proteomics to investigate proteins and their modifications

MS-based proteomics to investigate proteins and their modifications MS-based proteomics to investigate proteins and their modifications Francis Impens VIB Proteomics Core October th 217 Overview Mass spectrometry-based proteomics: general workflow Identification of protein

More information

Effective Strategies for Improving Peptide Identification with Tandem Mass Spectrometry

Effective Strategies for Improving Peptide Identification with Tandem Mass Spectrometry Effective Strategies for Improving Peptide Identification with Tandem Mass Spectrometry by Xi Han A thesis presented to the University of Waterloo in fulfillment of the thesis requirement for the degree

More information

Introduction to spectral alignment

Introduction to spectral alignment SI Appendix C. Introduction to spectral alignment Due to the complexity of the anti-symmetric spectral alignment algorithm described in Appendix A, this appendix provides an extended introduction to the

More information

TANDEM MASS SPECTROSCOPY

TANDEM MASS SPECTROSCOPY TANDEM MASS SPECTROSCOPY 1 MASS SPECTROMETER TYPES OF MASS SPECTROMETER PRINCIPLE TANDEM MASS SPECTROMETER INSTRUMENTATION QUADRAPOLE MASS ANALYZER TRIPLE QUADRAPOLE MASS ANALYZER TIME OF FLIGHT MASS ANALYSER

More information

De Novo Peptide Identification Via Mixed-Integer Linear Optimization And Tandem Mass Spectrometry

De Novo Peptide Identification Via Mixed-Integer Linear Optimization And Tandem Mass Spectrometry 17 th European Symposium on Computer Aided Process Engineering ESCAPE17 V. Plesu and P.S. Agachi (Editors) 2007 Elsevier B.V. All rights reserved. 1 De Novo Peptide Identification Via Mixed-Integer Linear

More information

Fundamentals of Mass Spectrometry. Fundamentals of Mass Spectrometry. Learning Objective. Proteomics

Fundamentals of Mass Spectrometry. Fundamentals of Mass Spectrometry. Learning Objective. Proteomics Mass spectrometry (MS) is the technique for protein identification and analysis by production of charged molecular species in vacuum, and their separation by magnetic and electric fields based on mass

More information

WADA Technical Document TD2015IDCR

WADA Technical Document TD2015IDCR MINIMUM CRITERIA FOR CHROMATOGRAPHIC-MASS SPECTROMETRIC CONFIRMATION OF THE IDENTITY OF ANALYTES FOR DOPING CONTROL PURPOSES. The ability of a method to identify an analyte is a function of the entire

More information

NPTEL VIDEO COURSE PROTEOMICS PROF. SANJEEVA SRIVASTAVA

NPTEL VIDEO COURSE PROTEOMICS PROF. SANJEEVA SRIVASTAVA LECTURE-25 Quantitative proteomics: itraq and TMT TRANSCRIPT Welcome to the proteomics course. Today we will talk about quantitative proteomics and discuss about itraq and TMT techniques. The quantitative

More information

Parallel Algorithms For Real-Time Peptide-Spectrum Matching

Parallel Algorithms For Real-Time Peptide-Spectrum Matching Parallel Algorithms For Real-Time Peptide-Spectrum Matching A Thesis Submitted to the College of Graduate Studies and Research in Partial Fulfillment of the Requirements for the degree of Master of Science

More information

Mass Spectrometry. Hyphenated Techniques GC-MS LC-MS and MS-MS

Mass Spectrometry. Hyphenated Techniques GC-MS LC-MS and MS-MS Mass Spectrometry Hyphenated Techniques GC-MS LC-MS and MS-MS Reasons for Using Chromatography with MS Mixture analysis by MS alone is difficult Fragmentation from ionization (EI or CI) Fragments from

More information

Mass Spectrometry Based De Novo Peptide Sequencing Error Correction

Mass Spectrometry Based De Novo Peptide Sequencing Error Correction Mass Spectrometry Based De Novo Peptide Sequencing Error Correction by Chenyu Yao A thesis presented to the University of Waterloo in fulfillment of the thesis requirement for the degree of Master of Mathematics

More information

MASS SPECTROMETRY. Topics

MASS SPECTROMETRY. Topics MASS SPECTROMETRY MALDI-TOF AND ESI-MS Topics Principle of Mass Spectrometry MALDI-TOF Determination of Mw of Proteins Structural Information by MS: Primary Sequence of a Protein 1 A. Principles Ionization:

More information

LECTURE-13. Peptide Mass Fingerprinting HANDOUT. Mass spectrometry is an indispensable tool for qualitative and quantitative analysis of

LECTURE-13. Peptide Mass Fingerprinting HANDOUT. Mass spectrometry is an indispensable tool for qualitative and quantitative analysis of LECTURE-13 Peptide Mass Fingerprinting HANDOUT PREAMBLE Mass spectrometry is an indispensable tool for qualitative and quantitative analysis of proteins, drugs and many biological moieties to elucidate

More information

Supporting Information for: A. Portz 1, C. R. Gebhardt 2, and M. Dürr 1, Giessen, D Giessen, Germany

Supporting Information for: A. Portz 1, C. R. Gebhardt 2, and M. Dürr 1, Giessen, D Giessen, Germany Supporting Information for: Real-time Investigation of the H/D Exchange Kinetics of Porphyrins and Oligopeptides by means of Neutral Cluster Induced Desorption/Ionization Mass Spectrometry A. Portz 1,

More information

BST 226 Statistical Methods for Bioinformatics David M. Rocke. January 22, 2014 BST 226 Statistical Methods for Bioinformatics 1

BST 226 Statistical Methods for Bioinformatics David M. Rocke. January 22, 2014 BST 226 Statistical Methods for Bioinformatics 1 BST 226 Statistical Methods for Bioinformatics David M. Rocke January 22, 2014 BST 226 Statistical Methods for Bioinformatics 1 Mass Spectrometry Mass spectrometry (mass spec, MS) comprises a set of instrumental

More information

Learning Score Function Parameters for Improved Spectrum Identification in Tandem Mass Spectrometry Experiments

Learning Score Function Parameters for Improved Spectrum Identification in Tandem Mass Spectrometry Experiments pubs.acs.org/jpr Learning Score Function Parameters for Improved Spectrum Identification in Tandem Mass Spectrometry Experiments Marina Spivak, Michael S. Bereman, Michael J. MacCoss, and William Stafford

More information

An SVM Scorer for More Sensitive and Reliable Peptide Identification via Tandem Mass Spectrometry

An SVM Scorer for More Sensitive and Reliable Peptide Identification via Tandem Mass Spectrometry An SVM Scorer for More Sensitive and Reliable Peptide Identification via Tandem Mass Spectrometry Haipeng Wang, Yan Fu, Ruixiang Sun, Simin He, Rong Zeng, and Wen Gao Pacific Symposium on Biocomputing

More information

Modeling Mass Spectrometry-Based Protein Analysis

Modeling Mass Spectrometry-Based Protein Analysis Chapter 8 Jan Eriksson and David Fenyö Abstract The success of mass spectrometry based proteomics depends on efficient methods for data analysis. These methods require a detailed understanding of the information

More information

Figure S1. Interaction of PcTS with αsyn. (a) 1 H- 15 N HSQC NMR spectra of 100 µm αsyn in the absence (0:1, black) and increasing equivalent

Figure S1. Interaction of PcTS with αsyn. (a) 1 H- 15 N HSQC NMR spectra of 100 µm αsyn in the absence (0:1, black) and increasing equivalent Figure S1. Interaction of PcTS with αsyn. (a) 1 H- 15 N HSQC NMR spectra of 100 µm αsyn in the absence (0:1, black) and increasing equivalent concentrations of PcTS (100 µm, blue; 500 µm, green; 1.5 mm,

More information

Tandem mass spectra were extracted from the Xcalibur data system format. (.RAW) and charge state assignment was performed using in house software

Tandem mass spectra were extracted from the Xcalibur data system format. (.RAW) and charge state assignment was performed using in house software Supplementary Methods Software Interpretation of Tandem mass spectra Tandem mass spectra were extracted from the Xcalibur data system format (.RAW) and charge state assignment was performed using in house

More information

(Refer Slide Time 00:09) (Refer Slide Time 00:13)

(Refer Slide Time 00:09) (Refer Slide Time 00:13) (Refer Slide Time 00:09) Mass Spectrometry Based Proteomics Professor Sanjeeva Srivastava Department of Biosciences and Bioengineering Indian Institute of Technology, Bombay Mod 02 Lecture Number 09 (Refer

More information

Last updated: Copyright

Last updated: Copyright Last updated: 2012-08-20 Copyright 2004-2012 plabel (v2.4) User s Manual by Bioinformatics Group, Institute of Computing Technology, Chinese Academy of Sciences Tel: 86-10-62601016 Email: zhangkun01@ict.ac.cn,

More information

Making Sense of Differences in LCMS Data: Integrated Tools

Making Sense of Differences in LCMS Data: Integrated Tools Making Sense of Differences in LCMS Data: Integrated Tools David A. Weil Agilent Technologies MassHunter Overview Page 1 March 2008 How Clean is our Water?... Page 2 Chemical Residue Analysis.... From

More information

WADA Technical Document TD2003IDCR

WADA Technical Document TD2003IDCR IDENTIFICATION CRITERIA FOR QUALITATIVE ASSAYS INCORPORATING CHROMATOGRAPHY AND MASS SPECTROMETRY The appropriate analytical characteristics must be documented for a particular assay. The Laboratory must

More information

Structure of the α-helix

Structure of the α-helix Structure of the α-helix Structure of the β Sheet Protein Dynamics Basics of Quenching HDX Hydrogen exchange of amide protons is catalyzed by H 2 O, OH -, and H 3 O +, but it s most dominated by base

More information

Chemistry Instrumental Analysis Lecture 37. Chem 4631

Chemistry Instrumental Analysis Lecture 37. Chem 4631 Chemistry 4631 Instrumental Analysis Lecture 37 Most analytes separated by HPLC are thermally stable and non-volatile (liquids) (unlike in GC) so not ionized easily by EI or CI techniques. MS must be at

More information

Proteome Informatics. Brian C. Searle Creative Commons Attribution

Proteome Informatics. Brian C. Searle Creative Commons Attribution Proteome Informatics Brian C. Searle searleb@uw.edu Creative Commons Attribution Section structure Class 1 Class 2 Homework 1 Mass spectrometry and de novo sequencing Database searching and E-value estimation

More information

sample was a solution that was evaporated in the spectrometer (such as with ESI-MS) ions such as H +, Na +, K +, or NH 4

sample was a solution that was evaporated in the spectrometer (such as with ESI-MS) ions such as H +, Na +, K +, or NH 4 Introduction to Spectroscopy V: Mass Spectrometry Basic Theory: Unlike other forms of spectroscopy used in structure elucidation of organic molecules mass spectrometry does not involve absorption/emission

More information

Protein Post-translational Modifications Mapping with MS/MS based Frequent Interval Pattern Mining

Protein Post-translational Modifications Mapping with MS/MS based Frequent Interval Pattern Mining Protein Post-translational Modifications Mapping with MS/MS based Frequent Interval Pattern Mining Han Liu Department of Computer Science University of Illinois at Urbana-Champaign Email: hanliu@ncsa.uiuc.edu

More information

1. Prepare the MALDI sample plate by spotting an angiotensin standard and the test sample(s).

1. Prepare the MALDI sample plate by spotting an angiotensin standard and the test sample(s). Analysis of a Peptide Sequence from a Proteolytic Digest by MALDI-TOF Post-Source Decay (PSD) and Collision-Induced Dissociation (CID) Standard Operating Procedure Purpose: The following procedure may

More information

TOMAHAQ Method Construction

TOMAHAQ Method Construction TOMAHAQ Method Construction Triggered by offset mass accurate-mass high-resolution accurate quantitation (TOMAHAQ) can be performed in the standard method editor of the instrument, without modifications

More information

Protein Quantitation II: Multiple Reaction Monitoring. Kelly Ruggles New York University

Protein Quantitation II: Multiple Reaction Monitoring. Kelly Ruggles New York University Protein Quantitation II: Multiple Reaction Monitoring Kelly Ruggles kelly@fenyolab.org New York University Traditional Affinity-based proteomics Use antibodies to quantify proteins Western Blot RPPA Immunohistochemistry

More information

Lecture 15: Introduction to mass spectrometry-i

Lecture 15: Introduction to mass spectrometry-i Lecture 15: Introduction to mass spectrometry-i Mass spectrometry (MS) is an analytical technique that measures the mass/charge ratio of charged particles in vacuum. Mass spectrometry can determine masse/charge

More information

Tandem MS = MS / MS. ESI-MS give information on the mass of a molecule but none on the structure

Tandem MS = MS / MS. ESI-MS give information on the mass of a molecule but none on the structure Tandem MS = MS / MS ESI-MS give information on the mass of a molecule but none on the structure In tandem MS (MSMS) (pseudo-)molecular ions are selected in MS1 and fragmented by collision with gas. collision

More information

DIA-Umpire: comprehensive computational framework for data independent acquisition proteomics

DIA-Umpire: comprehensive computational framework for data independent acquisition proteomics DIA-Umpire: comprehensive computational framework for data independent acquisition proteomics Chih-Chiang Tsou 1,2, Dmitry Avtonomov 2, Brett Larsen 3, Monika Tucholska 3, Hyungwon Choi 4 Anne-Claude Gingras

More information

Mass Spectrometry. Electron Ionization and Chemical Ionization

Mass Spectrometry. Electron Ionization and Chemical Ionization Mass Spectrometry Electron Ionization and Chemical Ionization Mass Spectrometer All Instruments Have: 1. Sample Inlet 2. Ion Source 3. Mass Analyzer 4. Detector 5. Data System http://www.asms.org Ionization

More information

Workflow concept. Data goes through the workflow. A Node contains an operation An edge represents data flow The results are brought together in tables

Workflow concept. Data goes through the workflow. A Node contains an operation An edge represents data flow The results are brought together in tables PROTEOME DISCOVERER Workflow concept Data goes through the workflow Spectra Peptides Quantitation A Node contains an operation An edge represents data flow The results are brought together in tables Protein

More information

PeptideProphet: Validation of Peptide Assignments to MS/MS Spectra. Andrew Keller

PeptideProphet: Validation of Peptide Assignments to MS/MS Spectra. Andrew Keller PeptideProphet: Validation of Peptide Assignments to MS/MS Spectra Andrew Keller Outline Need to validate peptide assignments to MS/MS spectra Statistical approach to validation Running PeptideProphet

More information

Protein analysis using mass spectrometry

Protein analysis using mass spectrometry Protein analysis using mass spectrometry Michael Stadlmeier 2017/12/18 Literature http://www.carellgroup.de/teaching/master 3 What is Proteomics? The proteome is: the entire set of proteins in a given

More information

Biological Mass Spectrometry

Biological Mass Spectrometry Biochemistry 412 Biological Mass Spectrometry February 13 th, 2007 Proteomics The study of the complete complement of proteins found in an organism Degrees of Freedom for Protein Variability Covalent Modifications

More information

Protein Quantitation II: Multiple Reaction Monitoring. Kelly Ruggles New York University

Protein Quantitation II: Multiple Reaction Monitoring. Kelly Ruggles New York University Protein Quantitation II: Multiple Reaction Monitoring Kelly Ruggles kelly@fenyolab.org New York University Traditional Affinity-based proteomics Use antibodies to quantify proteins Western Blot Immunohistochemistry

More information

Chemistry 311: Topic 3 - Mass Spectrometry

Chemistry 311: Topic 3 - Mass Spectrometry Mass Spectroscopy: A technique used to measure the mass-to-charge ratio of molecules and atoms. Often characteristic ions produced by an induced unimolecular dissociation of a molecule are measured. These

More information

Key questions of proteomics. Bioinformatics 2. Proteomics. Foundation of proteomics. What proteins are there? Protein digestion

Key questions of proteomics. Bioinformatics 2. Proteomics. Foundation of proteomics. What proteins are there? Protein digestion s s Key questions of proteomics What proteins are there? Bioinformatics 2 Lecture 2 roteomics How much is there of each of the proteins? - Absolute quantitation - Stoichiometry What (modification/splice)

More information

Comprehensive support for quantitation

Comprehensive support for quantitation Comprehensive support for quantitation One of the major new features in the current release of Mascot is support for quantitation. This is still work in progress. Our goal is to support all of the popular

More information

A Better Scoring Model for De Novo Peptide Sequencing: The Symmetric Difference between Explained and Measured Masses Supplementary Figures

A Better Scoring Model for De Novo Peptide Sequencing: The Symmetric Difference between Explained and Measured Masses Supplementary Figures A Better Scoring Model for De Novo Peptide Sequencing: The Symmetric Difference between Explained and Measured Masses Supplementary Figures Thomas Tschager *, Simon Rösch *, Ludovic Gillet 2 and Peter

More information

What is Tandem Mass Spectrometry? (MS/MS)

What is Tandem Mass Spectrometry? (MS/MS) What is Tandem Mass Spectrometry? (MS/MS) Activate MS1 MS2 Gas Surface Photons Electrons http://www.iupac.org/goldbook/t06250.pdf Major difference between MS and MS/MS: two analysis steps Ionization Analysis

More information

On Optimizing the Non-metric Similarity Search in Tandem Mass Spectra by Clustering

On Optimizing the Non-metric Similarity Search in Tandem Mass Spectra by Clustering On Optimizing the Non-metric Similarity Search in Tandem Mass Spectra by Clustering Jiří Novák, David Hoksza, Jakub Lokoč, and Tomáš Skopal Siret Research Group, Faculty of Mathematics and Physics, Charles

More information

Identification of proteins by enzyme digestion, mass

Identification of proteins by enzyme digestion, mass Method for Screening Peptide Fragment Ion Mass Spectra Prior to Database Searching Roger E. Moore, Mary K. Young, and Terry D. Lee Beckman Research Institute of the City of Hope, Duarte, California, USA

More information

LC-MS Based Metabolomics

LC-MS Based Metabolomics LC-MS Based Metabolomics Analysing the METABOLOME 1. Metabolite Extraction 2. Metabolite detection (with or without separation) 3. Data analysis Metabolite Detection GC-MS: Naturally volatile or made volatile

More information

for the Novice Mass Spectrometry (^>, John Greaves and John Roboz yc**' CRC Press J Taylor & Francis Group Boca Raton London New York

for the Novice Mass Spectrometry (^>, John Greaves and John Roboz yc**' CRC Press J Taylor & Francis Group Boca Raton London New York Mass Spectrometry for the Novice John Greaves and John Roboz (^>, yc**' CRC Press J Taylor & Francis Group Boca Raton London New York CRC Press is an imprint of the Taylor & Francis Croup, an informa business

More information

ASCQ_ME: a new engine for peptide mass fingerprint directly from mass spectrum without mass list extraction

ASCQ_ME: a new engine for peptide mass fingerprint directly from mass spectrum without mass list extraction ASCQ_ME: a new engine for peptide mass fingerprint directly from mass spectrum without mass list extraction Jean-Charles BOISSON1, Laetitia JOURDAN1, El-Ghazali TALBI1, Cécile CREN-OLIVE2 et Christian

More information

A graph-based filtering method for top-down mass spectral identification

A graph-based filtering method for top-down mass spectral identification Yang and Zhu BMC Genomics 2018, 19(Suppl 7):666 https://doi.org/10.1186/s12864-018-5026-x METHODOLOGY Open Access A graph-based filtering method for top-down mass spectral identification Runmin Yang and

More information

Key Words Q Exactive, Accela, MetQuest, Mass Frontier, Drug Discovery

Key Words Q Exactive, Accela, MetQuest, Mass Frontier, Drug Discovery Metabolite Stability Screening and Hotspot Metabolite Identification by Combining High-Resolution, Accurate-Mass Nonselective and Selective Fragmentation Tim Stratton, Caroline Ding, Yingying Huang, Dan

More information

Mass Spectrometry and Proteomics - Lecture 5 - Matthias Trost Newcastle University

Mass Spectrometry and Proteomics - Lecture 5 - Matthias Trost Newcastle University Mass Spectrometry and Proteomics - Lecture 5 - Matthias Trost Newcastle University matthias.trost@ncl.ac.uk Previously Proteomics Sample prep 144 Lecture 5 Quantitation techniques Search Algorithms Proteomics

More information

Lecture 14 Organic Chemistry 1

Lecture 14 Organic Chemistry 1 CHEM 232 Organic Chemistry I at Chicago Lecture 14 Organic Chemistry 1 Professor Duncan Wardrop February 25, 2010 1 CHEM 232 Organic Chemistry I at Chicago Mass Spectrometry Sections: 13.24-13.25 2 Spectroscopy

More information

Peter A. DiMaggio, Jr., Nicolas L. Young, Richard C. Baliban, Benjamin A. Garcia, and Christodoulos A. Floudas. Research

Peter A. DiMaggio, Jr., Nicolas L. Young, Richard C. Baliban, Benjamin A. Garcia, and Christodoulos A. Floudas. Research Research A Mixed Integer Linear Optimization Framework for the Identification and Quantification of Targeted Post-translational Modifications of Highly Modified Proteins Using Multiplexed Electron Transfer

More information

Supplementary Material for: Clustering Millions of Tandem Mass Spectra

Supplementary Material for: Clustering Millions of Tandem Mass Spectra Supplementary Material for: Clustering Millions of Tandem Mass Spectra Ari M. Frank 1 Nuno Bandeira 1 Zhouxin Shen 2 Stephen Tanner 3 Steven P. Briggs 2 Richard D. Smith 4 Pavel A. Pevzner 1 October 4,

More information

A Suboptimal Algorithm for De Novo Peptide Sequencing via Tandem Mass Spectrometry. BINGWEN LU and TING CHEN ABSTRACT

A Suboptimal Algorithm for De Novo Peptide Sequencing via Tandem Mass Spectrometry. BINGWEN LU and TING CHEN ABSTRACT JOURNAL OF COMPUTATIONAL BIOLOGY Volume 10, Number 1, 2003 Mary Ann Liebert, Inc. Pp. 1 12 A Suboptimal Algorithm for De Novo Peptide Sequencing via Tandem Mass Spectrometry BINGWEN LU and TING CHEN ABSTRACT

More information

SRM assay generation and data analysis in Skyline

SRM assay generation and data analysis in Skyline in Skyline Preparation 1. Download the example data from www.srmcourse.ch/eupa.html (3 raw files, 1 csv file, 1 sptxt file). 2. The number formats of your computer have to be set to English (United States).

More information

LECTURE-11. Hybrid MS Configurations HANDOUT. As discussed in our previous lecture, mass spectrometry is by far the most versatile

LECTURE-11. Hybrid MS Configurations HANDOUT. As discussed in our previous lecture, mass spectrometry is by far the most versatile LECTURE-11 Hybrid MS Configurations HANDOUT PREAMBLE As discussed in our previous lecture, mass spectrometry is by far the most versatile technique used in proteomics. We had also discussed some of the

More information

High-Field Orbitrap Creating new possibilities

High-Field Orbitrap Creating new possibilities Thermo Scientific Orbitrap Elite Hybrid Mass Spectrometer High-Field Orbitrap Creating new possibilities Ultrahigh resolution Faster scanning Higher sensitivity Complementary fragmentation The highest

More information

ABSTRACT 1. INTRODUCTION

ABSTRACT 1. INTRODUCTION JOURNAL OF COMPUTATIONAL BIOLOGY Volume 6, Numbers 3/4, 1999 Mary Ann Liebert, Inc. Pp. 327 342 De Novo Peptide Sequencing via Tandem Mass Spectrometry VLADO DAN Ï CÍK, 1, 2 THERESA A. ADDONA, 1 KARL R.

More information

UC San Diego UC San Diego Electronic Theses and Dissertations

UC San Diego UC San Diego Electronic Theses and Dissertations UC San Diego UC San Diego Electronic Theses and Dissertations Title Algorithms for tandem mass spectrometry-based proteomics Permalink https://escholarship.org/uc/item/89f7x81r Author Frank, Ari Michael

More information

Computationally Analyzing Mass Spectra of Hydrogen Deuterium Exchange Experiments Kevin S. Drew University of Chicago May 21, 2005

Computationally Analyzing Mass Spectra of Hydrogen Deuterium Exchange Experiments Kevin S. Drew University of Chicago May 21, 2005 Computationally Analyzing Mass Spectra of Hydrogen Deuterium Exchange Experiments Kevin S. Drew University of Chicago May 21, 2005 1 Abstract Hydrogen deuterium exchange (HDX) using Mass Spectrometers

More information

6 x 5 Ways to Ensure Your LC-MS/MS is Healthy

6 x 5 Ways to Ensure Your LC-MS/MS is Healthy 6 x 5 Ways to Ensure Your LC-MS/MS is Healthy (Also known as - Tracking Performance with the 6 x 5 LC-MS/MS Peptide Reference Mixture) Mike Rosenblatt, Ph.D. Group Leader Mass Spec Reagents 215. We monitor

More information

BENG 183 Trey Ideker. Protein Sequencing

BENG 183 Trey Ideker. Protein Sequencing BENG 183 Trey Ideker Protein Sequencing The following slides borrowed from Hong Li s Biochemistry Course: www.sb.fsu.edu/~hongli/4053notes Introduction to Proteins Proteins are of vital importance to biological

More information

BIOINF 4120 Bioinformatics 2 - Structures and Systems - Oliver Kohlbacher Summer Systems Biology Exp. Methods

BIOINF 4120 Bioinformatics 2 - Structures and Systems - Oliver Kohlbacher Summer Systems Biology Exp. Methods BIOINF 4120 Bioinformatics 2 - Structures and Systems - Oliver Kohlbacher Summer 2013 14. Systems Biology Exp. Methods Overview Transcriptomics Basics of microarrays Comparative analysis Interactomics:

More information

MASS ANALYSER. Mass analysers - separate the ions according to their mass-to-charge ratio. sample. Vacuum pumps

MASS ANALYSER. Mass analysers - separate the ions according to their mass-to-charge ratio. sample. Vacuum pumps ION ANALYZERS MASS ANALYSER sample Vacuum pumps Mass analysers - separate the ions according to their mass-to-charge ratio MASS ANALYSER Separate the ions according to their mass-to-charge ratio in space

More information

Proteome-wide label-free quantification with MaxQuant. Jürgen Cox Max Planck Institute of Biochemistry July 2011

Proteome-wide label-free quantification with MaxQuant. Jürgen Cox Max Planck Institute of Biochemistry July 2011 Proteome-wide label-free quantification with MaxQuant Jürgen Cox Max Planck Institute of Biochemistry July 2011 MaxQuant MaxQuant Feature detection Data acquisition Initial Andromeda search Statistics

More information

Introduction to Comparative Protein Modeling. Chapter 4 Part I

Introduction to Comparative Protein Modeling. Chapter 4 Part I Introduction to Comparative Protein Modeling Chapter 4 Part I 1 Information on Proteins Each modeling study depends on the quality of the known experimental data. Basis of the model Search in the literature

More information

Quantitation of a target protein in crude samples using targeted peptide quantification by Mass Spectrometry

Quantitation of a target protein in crude samples using targeted peptide quantification by Mass Spectrometry Quantitation of a target protein in crude samples using targeted peptide quantification by Mass Spectrometry Jon Hao, Rong Ye, and Mason Tao Poochon Scientific, Frederick, Maryland 21701 Abstract Background:

More information