Sequence comparison: Score matrices. Genome 559: Introduction to Statistical and Computational Genomics Prof. James H. Thomas
|
|
- Meredith McCoy
- 6 years ago
- Views:
Transcription
1 Sequence comparison: Score matrices Genome 559: Introduction to Statistical and omputational Genomics Prof James H Thomas
2 FYI - informal inductive proof of best alignment path onsider the last step in the best alignment path to node a below This path must come from one of the three nodes shown, where X, Y, and Z are the cumulative scores of the best alignments up to those nodes We can reach node a by three possible paths: an A-B match, a gap in sequence A or a gap in sequence B: seq B X Z seq A match gap Y gap a The best-scoring path to a is the maximum of: X + match Y + gap Z + gap BUT the best paths to X, Y, and Z are analogously the max of their three upstream possibilities, etc Inductively QED
3 Local alignment - review A G T A G T d = -5 A A G A F i 1, j 1 0 s x i, y j F i, j 1 d G F i 1, j d F i, j (no arrow means no preceding alignment)
4 Local alignment Two differences from global alignment: If a score is negative, replace with 0 Traceback from the highest score in the matrix and continue until you reach 0 Global alignment algorithm: eedleman- Wunsch Local alignment algorithm: Smith- Waterman
5 dot plot of two DA sequences overlay of the global DP alignment path
6 Protein score matrices Quantitatively represent the degree of conservation of typical amino acid residues over evolutionary time All possible amino acid changes are represented (matrix of size at least 20 x 20) Most commonly used are several different BLOSUM matrices derived for different degrees of evolutionary divergence DA score matrices are much simpler (and conceptually similar)
7 # BLOSUM lustered Scoring Matrix in 1/2 Bit Units # luster Percentage: >= 62 BLOSUM62 Score Matrix regular 20 amino acids ambiguity codes and stop
8 Hydrophobic Amino acid structures Polar harged phenylalanine F
9 BLOSUM62 Score Matrix Good scores chemically similar Bad scores chemically dissimilar
10 Amino acid structures Hydrophobic Polar harged glycine G H H SH H cysteine histidine H H alanine A H H 3 OH serine S H H 3 valine V H OH lysine K H H 3 threonine T H H 3 arginine R H H leucine L tyrosine Y H OH H 3 H 3 O isoleucine I H H 2 aspartate D H H 3 H asparagine H 2 glutamate E methionine M H S H 3 glutamine Q H O H O + H H 2 O - O - O H + 3 H 2 + proline P H H tryptophan W H
11 Deriving BLOSUM scores Find sets of sequences whose alignment is thought to be correct (this is partly bootstrapped by alignment) Measure how often various amino acid pairs occur in the alignments ormalize this to the expected frequency of such pairs randomly in the same set of alignments Derive a log-odds score for aligned vs random
12 Example of alignment block (the BLO part of BLOSUM) 31 positions (columns) 61 sequences (rows) Thousands of such blocks go into computing a single BLOSUM matrix Represent full diversity of sequences Results are summed over all columns of all blocks
13 Pair frequency vs expectation Actual aligned pair frequency: q ij 1 T c where cij is the count of ij pairs and T is the total pair count ij this is called the sum of pairs (the SUM part of BLOSUM) Sample column from an alignment block: D E D D D etc Randomly expected pair frequency: e p p aa a a e p p p p 2p p ab a b b a a b where pa and pb are the overall probabilities (frequencies) of specific residues a and b 6 D-D pairs 4 D-E pairs 4 D- pairs 1 E- pair (a multiple alignment of sequences is the equivalent of all the pairwise alignments, which number ()(-1)/2)
14 Log-odds score calculation (so adding scores == multiplying probabilities) s ij log 2 q e ij ij For computational speed often rounded to nearest integer and (to reduce round-off error) they are often multiplied by 2 (or more) first, giving a half-bit score: matrixscore (rounded) 2log 2 q e ij ij (computers can add integers faster than floats)
15 BLOSUM62 matrix (half-bit scores) ( 9 half-bits = 45 bits ) Frequency of residue over all proteins: (you have to look this up) Reverse calculation of aligned - pair frequency in BLOSUM data set: - q e cc cc thus e cc q cc
16 onstructing Blocks Blocks are ungapped alignments of multiple sequences, usually 20 to 100 amino acids long luster the members of each block according to their percent identity Make pair counts and score matrix from a large collection of similarly clustered blocks Each BLOSUM matrix is named for the percent identity cutoff in step 2 (eg BLOSUM70 for 70% identity)
17 Probabilistic Interpretation of Scores (ungapped) matrixscore (rounded) 2log 2 By converting scores back to probabilities, we can give a probabilistic interpretation to an alignment score q e ij ij (BLOSUM62) this alignment has a score of 16 ( ) by BLOSUM 62, meaning an alignment with this score or more is 2 8 (256) times more likely than expected from a random alignment FIAP FLSP this 15 amino acid alignment has a score of 75, meaning that it is ~10 11 times more likely to be seen in a real alignment than in a random alignment(!!) VHRDLKPELLLASK VHRDLKPELLLASK ( )
18 Randomly Distributed Gaps if then p g P( g k 1 ) (probability of a gap at each position in the sequence) k, P( g 2 ) k 2,, P( g n ) k n [note - the slope of the line on a log-linear plot will vary according to the frequency of gaps, but it will always be linear]
19 Distribution of real alignment gap lengths in large set of structurally-aligned proteins log-linear plot owhere near linear - hence the use of affine gap penalties (there ideally would be several levels of decreasing affine penalties)
20 Summary How a score matrix is derived What the scores mean probablistically Why gap penalties should be affine
Sequence comparison: Score matrices
Sequence comparison: Score matrices http://facultywashingtonedu/jht/gs559_2013/ Genome 559: Introduction to Statistical and omputational Genomics Prof James H Thomas FYI - informal inductive proof of best
More informationSequence comparison: Score matrices. Genome 559: Introduction to Statistical and Computational Genomics Prof. James H. Thomas
Sequence comparison: Score matrices Genome 559: Introduction to Statistical and omputational Genomics Prof James H Thomas Informal inductive proof of best alignment path onsider the last step in the best
More informationProteins: Characteristics and Properties of Amino Acids
SBI4U:Biochemistry Macromolecules Eachaminoacidhasatleastoneamineandoneacidfunctionalgroupasthe nameimplies.thedifferentpropertiesresultfromvariationsinthestructuresof differentrgroups.thergroupisoftenreferredtoastheaminoacidsidechain.
More informationScoring Matrices. Shifra Ben-Dor Irit Orr
Scoring Matrices Shifra Ben-Dor Irit Orr Scoring matrices Sequence alignment and database searching programs compare sequences to each other as a series of characters. All algorithms (programs) for comparison
More informationFinding the Best Biological Pairwise Alignment Through Genetic Algorithm Determinando o Melhor Alinhamento Biológico Através do Algoritmo Genético
Finding the Best Biological Pairwise Alignment Through Genetic Algorithm Determinando o Melhor Alinhamento Biológico Através do Algoritmo Genético Paulo Mologni 1, Ailton Akira Shinoda 2, Carlos Dias Maciel
More informationC h a p t e r 2 A n a l y s i s o f s o m e S e q u e n c e... methods use different attributes related to mis sense mutations such as
C h a p t e r 2 A n a l y s i s o f s o m e S e q u e n c e... 2.1Introduction smentionedinchapter1,severalmethodsareavailabletoclassifyhuman missensemutationsintoeitherbenignorpathogeniccategoriesandthese
More informationProtein Secondary Structure Prediction
part of Bioinformatik von RNA- und Proteinstrukturen Computational EvoDevo University Leipzig Leipzig, SS 2011 the goal is the prediction of the secondary structure conformation which is local each amino
More information8 Grundlagen der Bioinformatik, SS 09, D. Huson, April 28, 2009
8 Grundlagen der Bioinformatik, SS 09, D. Huson, April 28, 2009 2 Pairwise alignment We will discuss: 1. Strings 2. Dot matrix method for comparing sequences 3. Edit distance and alignment 4. The number
More information8 Grundlagen der Bioinformatik, SoSe 11, D. Huson, April 18, 2011
8 Grundlagen der Bioinformatik, SoSe 11, D. Huson, April 18, 2011 2 Pairwise alignment We will discuss: 1. Strings 2. Dot matrix method for comparing sequences 3. Edit distance and alignment 4. The number
More informationPROTEIN STRUCTURE AMINO ACIDS H R. Zwitterion (dipolar ion) CO 2 H. PEPTIDES Formal reactions showing formation of peptide bond by dehydration:
PTEI STUTUE ydrolysis of proteins with aqueous acid or base yields a mixture of free amino acids. Each type of protein yields a characteristic mixture of the ~ 20 amino acids. AMI AIDS Zwitterion (dipolar
More informationRange of Certified Values in Reference Materials. Range of Expanded Uncertainties as Disseminated. NMI Service
Calibration and Capabilities Amount of substance,, Russian Federation (Ural Scientific and Research Institiute Metrology, Rosstandart) (D.I. Mendeleyev Institute Metrology, Rosstandart) The uncertainty
More informationSEQUENCE ALIGNMENT BACKGROUND: BIOINFORMATICS. Prokaryotes and Eukaryotes. DNA and RNA
SEQUENCE ALIGNMENT BACKGROUND: BIOINFORMATICS 1 Prokaryotes and Eukaryotes 2 DNA and RNA 3 4 Double helix structure Codons Codons are triplets of bases from the RNA sequence. Each triplet defines an amino-acid.
More informationPeriodic Table. 8/3/2006 MEDC 501 Fall
Periodic Table 8/3/2006 MEDC 501 Fall 2006 1 rbitals Shapes of rbitals s - orbital p -orbital 8/3/2006 MEDC 501 Fall 2006 2 Ionic Bond - acl Electronic Structure 11 a :: 1s 2 2s 2 2p x2 2p y2 2p z2 3s
More informationChapter 5. Proteomics and the analysis of protein sequence Ⅱ
Proteomics Chapter 5. Proteomics and the analysis of protein sequence Ⅱ 1 Pairwise similarity searching (1) Figure 5.5: manual alignment One of the amino acids in the top sequence has no equivalent and
More informationProtein structure. Protein structure. Amino acid residue. Cell communication channel. Bioinformatics Methods
Cell communication channel Bioinformatics Methods Iosif Vaisman Email: ivaisman@gmu.edu SEQUENCE STRUCTURE DNA Sequence Protein Sequence Protein Structure Protein structure ATGAAATTTGGAAACTTCCTTCTCACTTATCAGCCACCT...
More informationProteome Informatics. Brian C. Searle Creative Commons Attribution
Proteome Informatics Brian C. Searle searleb@uw.edu Creative Commons Attribution Section structure Class 1 Class 2 Homework 1 Mass spectrometry and de novo sequencing Database searching and E-value estimation
More informationQuantifying sequence similarity
Quantifying sequence similarity Bas E. Dutilh Systems Biology: Bioinformatic Data Analysis Utrecht University, February 16 th 2016 After this lecture, you can define homology, similarity, and identity
More informationLecture 14 - Cells. Astronomy Winter Lecture 14 Cells: The Building Blocks of Life
Lecture 14 Cells: The Building Blocks of Life Astronomy 141 Winter 2012 This lecture describes Cells, the basic structural units of all life on Earth. Basic components of cells: carbohydrates, lipids,
More informationProperties of amino acids in proteins
Properties of amino acids in proteins one of the primary roles of DNA (but not the only one!) is to code for proteins A typical bacterium builds thousands types of proteins, all from ~20 amino acids repeated
More informationHypergraphs, Metabolic Networks, Bioreaction Systems. G. Bastin
Hypergraphs, Metabolic Networks, Bioreaction Systems. G. Bastin PART 1 : Metabolic flux analysis and minimal bioreaction modelling PART 2 : Dynamic metabolic flux analysis of underdetermined networks 2
More informationLecture 4: Evolutionary Models and Substitution Matrices (PAM and BLOSUM)
Bioinformatics II Probability and Statistics Universität Zürich and ETH Zürich Spring Semester 2009 Lecture 4: Evolutionary Models and Substitution Matrices (PAM and BLOSUM) Dr Fraser Daly adapted from
More informationPROTEIN SECONDARY STRUCTURE PREDICTION: AN APPLICATION OF CHOU-FASMAN ALGORITHM IN A HYPOTHETICAL PROTEIN OF SARS VIRUS
Int. J. LifeSc. Bt & Pharm. Res. 2012 Kaladhar, 2012 Research Paper ISSN 2250-3137 www.ijlbpr.com Vol.1, Issue. 1, January 2012 2012 IJLBPR. All Rights Reserved PROTEIN SECONDARY STRUCTURE PREDICTION:
More informationAmino Acids and Peptides
Amino Acids Amino Acids and Peptides Amino acid a compound that contains both an amino group and a carboxyl group α-amino acid an amino acid in which the amino group is on the carbon adjacent to the carboxyl
More informationLecture 15: Realities of Genome Assembly Protein Sequencing
Lecture 15: Realities of Genome Assembly Protein Sequencing Study Chapter 8.10-8.15 1 Euler s Theorems A graph is balanced if for every vertex the number of incoming edges equals to the number of outgoing
More informationViewing and Analyzing Proteins, Ligands and their Complexes 2
2 Viewing and Analyzing Proteins, Ligands and their Complexes 2 Overview Viewing the accessible surface Analyzing the properties of proteins containing thousands of atoms is best accomplished by representing
More informationUsing Higher Calculus to Study Biologically Important Molecules Julie C. Mitchell
Using Higher Calculus to Study Biologically Important Molecules Julie C. Mitchell Mathematics and Biochemistry University of Wisconsin - Madison 0 There Are Many Kinds Of Proteins The word protein comes
More informationScoring Matrices. Shifra Ben Dor Irit Orr
Scoring Matrices Shifra Ben Dor Irit Orr Scoring matrices Sequence alignment and database searching programs compare sequences to each other as a series of characters. All algorithms (programs) for comparison
More informationPairwise sequence alignment
Department of Evolutionary Biology Example Alignment between very similar human alpha- and beta globins: GSAQVKGHGKKVADALTNAVAHVDDMPNALSALSDLHAHKL G+ +VK+HGKKV A+++++AH+D++ +++++LS+LH KL GNPKVKAHGKKVLGAFSDGLAHLDNLKGTFATLSELHCDKL
More informationAlgorithms in Bioinformatics FOUR Pairwise Sequence Alignment. Pairwise Sequence Alignment. Convention: DNA Sequences 5. Sequence Alignment
Algorithms in Bioinformatics FOUR Sami Khuri Department of Computer Science San José State University Pairwise Sequence Alignment Homology Similarity Global string alignment Local string alignment Dot
More informationBioinformatics (GLOBEX, Summer 2015) Pairwise sequence alignment
Bioinformatics (GLOBEX, Summer 2015) Pairwise sequence alignment Substitution score matrices, PAM, BLOSUM Needleman-Wunsch algorithm (Global) Smith-Waterman algorithm (Local) BLAST (local, heuristic) E-value
More informationProtein Identification Using Tandem Mass Spectrometry. Nathan Edwards Informatics Research Applied Biosystems
Protein Identification Using Tandem Mass Spectrometry Nathan Edwards Informatics Research Applied Biosystems Outline Proteomics context Tandem mass spectrometry Peptide fragmentation Peptide identification
More informationPatrick: An Introduction to Medicinal Chemistry 5e Chapter 03
01) Which of the following statements is not true regarding the active site of an enzyme? a. An active site is normally on the surface of an enzyme. b. An active site is normally hydrophobic in nature.
More informationProteomics. November 13, 2007
Proteomics November 13, 2007 Acknowledgement Slides presented here have been borrowed from presentations by : Dr. Mark A. Knepper (LKEM, NHLBI, NIH) Dr. Nathan Edwards (Center for Bioinformatics and Computational
More informationLocal Alignment: Smith-Waterman algorithm
Local Alignment: Smith-Waterman algorithm Example: a shared common domain of two protein sequences; extended sections of genomic DNA sequence. Sensitive to detect similarity in highly diverged sequences.
More informationThe Select Command and Boolean Operators
The Select Command and Boolean Operators Part of the Jmol Training Guide from the MSOE Center for BioMolecular Modeling Interactive version available at http://cbm.msoe.edu/teachingresources/jmol/jmoltraining/boolean.html
More information3. SEQUENCE ANALYSIS BIOINFORMATICS COURSE MTAT
3. SEQUENCE ANALYSIS BIOINFORMATICS COURSE MTAT.03.239 25.09.2012 SEQUENCE ANALYSIS IS IMPORTANT FOR... Prediction of function Gene finding the process of identifying the regions of genomic DNA that encode
More informationPart 4 The Select Command and Boolean Operators
Part 4 The Select Command and Boolean Operators http://cbm.msoe.edu/newwebsite/learntomodel Introduction By default, every command you enter into the Console affects the entire molecular structure. However,
More informationLecture 2, 5/12/2001: Local alignment the Smith-Waterman algorithm. Alignment scoring schemes and theory: substitution matrices and gap models
Lecture 2, 5/12/2001: Local alignment the Smith-Waterman algorithm Alignment scoring schemes and theory: substitution matrices and gap models 1 Local sequence alignments Local sequence alignments are necessary
More informationExam III. Please read through each question carefully, and make sure you provide all of the requested information.
09-107 onors Chemistry ame Exam III Please read through each question carefully, and make sure you provide all of the requested information. 1. A series of octahedral metal compounds are made from 1 mol
More informationA Theoretical Inference of Protein Schemes from Amino Acid Sequences
A Theoretical Inference of Protein Schemes from Amino Acid Sequences Angel Villahoz-Baleta angel_villahozbaleta@student.uml.edu ABSTRACT Proteins are based on tri-dimensional dispositions generated from
More informationEXAM 1 Fall 2009 BCHS3304, SECTION # 21734, GENERAL BIOCHEMISTRY I Dr. Glen B Legge
EXAM 1 Fall 2009 BCHS3304, SECTION # 21734, GENERAL BIOCHEMISTRY I 2009 Dr. Glen B Legge This is a Scantron exam. All answers should be transferred to the Scantron sheet using a #2 pencil. Write and bubble
More informationEvidence from Evolution Activity 75 Points. Fossils Use your textbook and the diagrams on the next page to answer the following questions.
Name(s): Biology Evidence from Evolution Activity 75 Points Fossils Use your textbook and the diagrams on the next page to answer the following questions. 1. What are fossils? How are most fossils formed?
More informationAlgorithms in Bioinformatics
Algorithms in Bioinformatics Sami Khuri Department of omputer Science San José State University San José, alifornia, USA khuri@cs.sjsu.edu www.cs.sjsu.edu/faculty/khuri Pairwise Sequence Alignment Homology
More informationLecture 5,6 Local sequence alignment
Lecture 5,6 Local sequence alignment Chapter 6 in Jones and Pevzner Fall 2018 September 4,6, 2018 Evolution as a tool for biological insight Nothing in biology makes sense except in the light of evolution
More informationBioinformatics. Scoring Matrices. David Gilbert Bioinformatics Research Centre
Bioinformatics Scoring Matrices David Gilbert Bioinformatics Research Centre www.brc.dcs.gla.ac.uk Department of Computing Science, University of Glasgow Learning Objectives To explain the requirement
More informationSara C. Madeira. Universidade da Beira Interior. (Thanks to Ana Teresa Freitas, IST for useful resources on this subject)
Bioinformática Sequence Alignment Pairwise Sequence Alignment Universidade da Beira Interior (Thanks to Ana Teresa Freitas, IST for useful resources on this subject) 1 16/3/29 & 23/3/29 27/4/29 Outline
More informationPairwise & Multiple sequence alignments
Pairwise & Multiple sequence alignments Urmila Kulkarni-Kale Bioinformatics Centre 411 007 urmila@bioinfo.ernet.in Basis for Sequence comparison Theory of evolution: gene sequences have evolved/derived
More informationDiscussion Section (Day, Time):
Chemistry 27 Spring 2005 Exam 3 Chemistry 27 Professor Gavin MacBeath arvard University Spring 2005 our Exam 3 Friday April 29 th, 2005 11:07 AM 12:00 PM Discussion Section (Day, Time): TF: Directions:
More informationChemistry Chapter 22
hemistry 2100 hapter 22 Proteins Proteins serve many functions, including the following. 1. Structure: ollagen and keratin are the chief constituents of skin, bone, hair, and nails. 2. atalysts: Virtually
More informationChemical Properties of Amino Acids
hemical Properties of Amino Acids Protein Function Make up about 15% of the cell and have many functions in the cell 1. atalysis: enzymes 2. Structure: muscle proteins 3. Movement: myosin, actin 4. Defense:
More informationBioinformatics and BLAST
Bioinformatics and BLAST Overview Recap of last time Similarity discussion Algorithms: Needleman-Wunsch Smith-Waterman BLAST Implementation issues and current research Recap from Last Time Genome consists
More informationSimilarity or Identity? When are molecules similar?
Similarity or Identity? When are molecules similar? Mapping Identity A -> A T -> T G -> G C -> C or Leu -> Leu Pro -> Pro Arg -> Arg Phe -> Phe etc If we map similarity using identity, how similar are
More informationPairwise sequence alignments
Pairwise sequence alignments Volker Flegel VI, October 2003 Page 1 Outline Introduction Definitions Biological context of pairwise alignments Computing of pairwise alignments Some programs VI, October
More informationTranslation. A ribosome, mrna, and trna.
Translation The basic processes of translation are conserved among prokaryotes and eukaryotes. Prokaryotic Translation A ribosome, mrna, and trna. In the initiation of translation in prokaryotes, the Shine-Dalgarno
More informationLecture 4: Evolutionary models and substitution matrices (PAM and BLOSUM).
1 Bioinformatics: In-depth PROBABILITY & STATISTICS Spring Semester 2011 University of Zürich and ETH Zürich Lecture 4: Evolutionary models and substitution matrices (PAM and BLOSUM). Dr. Stefanie Muff
More informationStructures in equilibrium at point A: Structures in equilibrium at point B: (ii) Structure at the isoelectric point:
ame 21 F10-Final Exam Page 2 I. (42 points) (1) (16 points) The titration curve for L-lysine is shown below. Provide (i) the main structures in equilibrium at each of points A and B indicated below and
More informationPairwise sequence alignments. Vassilios Ioannidis (From Volker Flegel )
Pairwise sequence alignments Vassilios Ioannidis (From Volker Flegel ) Outline Introduction Definitions Biological context of pairwise alignments Computing of pairwise alignments Some programs Importance
More informationA rapid and highly selective colorimetric method for direct detection of tryptophan in proteins via DMSO acceleration
A rapid and highly selective colorimetric method for direct detection of tryptophan in proteins via DMSO acceleration Yanyan Huang, Shaoxiang Xiong, Guoquan Liu, Rui Zhao Beijing National Laboratory for
More informationINTRODUCTION. Amino acids occurring in nature have the general structure shown below:
Biochemistry I Laboratory Amino Acid Thin Layer Chromatography INTRODUCTION The primary importance of amino acids in cell structure and metabolism lies in the fact that they serve as building blocks for
More informationSequence Alignment: Scoring Schemes. COMP 571 Luay Nakhleh, Rice University
Sequence Alignment: Scoring Schemes COMP 571 Luay Nakhleh, Rice University Scoring Schemes Recall that an alignment score is aimed at providing a scale to measure the degree of similarity (or difference)
More informationTHEORY. Based on sequence Length According to the length of sequence being compared it is of following two types
Exp 11- THEORY Sequence Alignment is a process of aligning two sequences to achieve maximum levels of identity between them. This help to derive functional, structural and evolutionary relationships between
More informationCHEMISTRY ATAR COURSE DATA BOOKLET
CHEMISTRY ATAR COURSE DATA BOOKLET 2018 2018/2457 Chemistry ATAR Course Data Booklet 2018 Table of contents Periodic table of the elements...3 Formulae...4 Units...4 Constants...4 Solubility rules for
More informationInvestigating Evolutionary Relationships between Species through the Light of Graph Theory based on the Multiplet Structure of the Genetic Code
07 IEEE 7th International Advance Computing Conference Investigating Evolutionary Relationships between Species through the Light of Graph Theory based on the Multiplet Structure of the Genetic Code Antara
More informationDiscussion Section (Day, Time):
Chemistry 27 Spring 2005 Exam 1 Chemistry 27 Professor Gavin MacBeath arvard University Spring 2005 our Exam 1 Friday, February 25, 2005 11:07 AM 12:00 PM Discussion Section (Day, Time): TF: Directions:
More informationDiscussion Section (Day, Time): TF:
ame: Chemistry 27 Professor Gavin MacBeath arvard University Spring 2004 Final Exam Thursday, May 28, 2004 2:15 PM - 5:15 PM Discussion Section (Day, Time): Directions: TF: 1. Do not write in red ink.
More informationSequence analysis and Genomics
Sequence analysis and Genomics October 12 th November 23 rd 2 PM 5 PM Prof. Peter Stadler Dr. Katja Nowick Katja: group leader TFome and Transcriptome Evolution Bioinformatics group Paul-Flechsig-Institute
More informationCHAPTER 29 HW: AMINO ACIDS + PROTEINS
CAPTER 29 W: AMI ACIDS + PRTEIS For all problems, consult the table of 20 Amino Acids provided in lecture if an amino acid structure is needed; these will be given on exams. Use natural amino acids (L)
More informationGlobal alignments - review
Global alignments - review Take two sequences: X[j] and Y[j] M[i-1, j-1] ± 1 M[i, j] = max M[i, j-1] 2 M[i-1, j] 2 The best alignment for X[1 i] and Y[1 j] is called M[i, j] X[j] Initiation: M[,]= pply
More informationNational Nutrient Database for Standard Reference Release 28 slightly revised May, 2016
National base for Standard Reference Release 28 slightly revised May, 206 Full Report (All s) 005, Cheese, cottage, lowfat, 2% milkfat Report Date: February 23, 208 02:7 EST values and weights are for
More informationSequence Alignments. Dynamic programming approaches, scoring, and significance. Lucy Skrabanek ICB, WMC January 31, 2013
Sequence Alignments Dynamic programming approaches, scoring, and significance Lucy Skrabanek ICB, WMC January 31, 213 Sequence alignment Compare two (or more) sequences to: Find regions of conservation
More information1. Wings 5.. Jumping legs 2. 6 Legs 6. Crushing mouthparts 3. Segmented Body 7. Legs 4. Double set of wings 8. Curly antennae
Biology Cladogram practice Name Per Date What is a cladogram? It is a diagram that depicts evolutionary relationships among groups. It is based on PHYLOGENY, which is the study of evolutionary relationships.
More informationDiscussion Section (Day, Time):
Chemistry 27 pring 2005 Exam 3 Chemistry 27 Professor Gavin MacBeath arvard University pring 2005 our Exam 3 Friday April 29 th, 2005 11:07 AM 12:00 PM Discussion ection (Day, Time): TF: Directions: 1.
More informationHow did they form? Exploring Meteorite Mysteries
Exploring Meteorite Mysteries Objectives Students will: recognize that carbonaceous chondrite meteorites contain amino acids, the first step towards living plants and animals. conduct experiments that
More informationEdward Susko Department of Mathematics and Statistics, Dalhousie University. Introduction. Installation
1 dist est: Estimation of Rates-Across-Sites Distributions in Phylogenetic Subsititution Models Version 1.0 Edward Susko Department of Mathematics and Statistics, Dalhousie University Introduction The
More informationPractical considerations of working with sequencing data
Practical considerations of working with sequencing data File Types Fastq ->aligner -> reference(genome) coordinates Coordinate files SAM/BAM most complete, contains all of the info in fastq and more!
More informationAll Proteins Have a Basic Molecular Formula
All Proteins Have a Basic Molecular Formula Homa Torabizadeh Abstract This study proposes a basic molecular formula for all proteins. A total of 10,739 proteins belonging to 9 different protein groups
More informationPAM-1 Matrix 10,000. From: Ala Arg Asn Asp Cys Gln Glu To:
119-1 atrix 10,000 rom: la rg sn sp ys ln lu o: la 9867 2 9 10 3 8 17 rg 1 9913 1 0 1 10 0 sn 4 1 9822 36 0 4 6 sp 6 0 42 9859 0 6 53 ys 1 1 0 0 9973 0 0 ln 3 9 4 5 0 9876 27 lu 10 0 7 56 0 35 9865 120
More informationC E N T R. Introduction to bioinformatics 2007 E B I O I N F O R M A T I C S V U F O R I N T. Lecture 5 G R A T I V. Pair-wise Sequence Alignment
C E N T R E F O R I N T E G R A T I V E B I O I N F O R M A T I C S V U Introduction to bioinformatics 2007 Lecture 5 Pair-wise Sequence Alignment Bioinformatics Nothing in Biology makes sense except in
More informationProtein Structure Bioinformatics Introduction
1 Swiss Institute of Bioinformatics Protein Structure Bioinformatics Introduction Basel, 27. September 2004 Torsten Schwede Biozentrum - Universität Basel Swiss Institute of Bioinformatics Klingelbergstr
More informationCISC 889 Bioinformatics (Spring 2004) Sequence pairwise alignment (I)
CISC 889 Bioinformatics (Spring 2004) Sequence pairwise alignment (I) Contents Alignment algorithms Needleman-Wunsch (global alignment) Smith-Waterman (local alignment) Heuristic algorithms FASTA BLAST
More informationSystematic approaches to study cancer cell metabolism
Systematic approaches to study cancer cell metabolism Kivanc Birsoy Laboratory of Metabolic Regulation and Genetics The Rockefeller University, NY Cellular metabolism is complex ~3,000 metabolic genes
More informationComputational Biology
Computational Biology Lecture 6 31 October 2004 1 Overview Scoring matrices (Thanks to Shannon McWeeney) BLAST algorithm Start sequence alignment 2 1 What is a homologous sequence? A homologous sequence,
More informationStudent Handout 2. Human Sepiapterin Reductase mrna Gene Map A 3DMD BioInformatics Activity. Genome Sequencing. Sepiapterin Reductase
Project-Based Learning ctivity Human Sepiapterin Reductase mrn ene Map 3DMD BioInformatics ctivity 498 ---+---------+--------- ---------+---------+---------+---------+---------+---------+---------+---------+---------+---------
More informationEnzyme Catalysis & Biotechnology
L28-1 Enzyme Catalysis & Biotechnology Bovine Pancreatic RNase A Biochemistry, Life, and all that L28-2 A brief word about biochemistry traditionally, chemical engineers used organic and inorganic chemistry
More informationStudies Leading to the Development of a Highly Selective. Colorimetric and Fluorescent Chemosensor for Lysine
Supporting Information for Studies Leading to the Development of a Highly Selective Colorimetric and Fluorescent Chemosensor for Lysine Ying Zhou, a Jiyeon Won, c Jin Yong Lee, c * and Juyoung Yoon a,
More informationSolutions In each case, the chirality center has the R configuration
CAPTER 25 669 Solutions 25.1. In each case, the chirality center has the R configuration. C C 2 2 C 3 C(C 3 ) 2 D-Alanine D-Valine 25.2. 2 2 S 2 d) 2 25.3. Pro,, Trp, Tyr, and is, Trp, Tyr, and is Arg,
More informationMoreover, the circular logic
Moreover, the circular logic How do we know what is the right distance without a good alignment? And how do we construct a good alignment without knowing what substitutions were made previously? ATGCGT--GCAAGT
More informationBiochemistry 324 Bioinformatics. Pairwise sequence alignment
Biochemistry 324 Bioinformatics Pairwise sequence alignment How do we compare genes/proteins? When we have sequenced a genome, we try and identify the function of unknown genes by finding a similar gene
More informationTools and Algorithms in Bioinformatics
Tools and Algorithms in Bioinformatics GCBA815, Fall 2013 Week3: Blast Algorithm, theory and practice Babu Guda, Ph.D. Department of Genetics, Cell Biology & Anatomy Bioinformatics and Systems Biology
More information12/6/12. Dr. Sanjeeva Srivastava IIT Bombay. Primary Structure. Secondary Structure. Tertiary Structure. Quaternary Structure.
Dr. anjeeva rivastava Primary tructure econdary tructure Tertiary tructure Quaternary tructure Amino acid residues α Helix Polypeptide chain Assembled subunits 2 1 Amino acid sequence determines 3-D structure
More informationUsing an Artificial Regulatory Network to Investigate Neural Computation
Using an Artificial Regulatory Network to Investigate Neural Computation W. Garrett Mitchener College of Charleston January 6, 25 W. Garrett Mitchener (C of C) UM January 6, 25 / 4 Evolution and Computing
More informationProtein Struktur (optional, flexible)
Protein Struktur (optional, flexible) 22/10/2009 [ 1 ] Andrew Torda, Wintersemester 2009 / 2010, AST nur für Informatiker, Mathematiker,.. 26 kt, 3 ov 2009 Proteins - who cares? 22/10/2009 [ 2 ] Most important
More informationCollected Works of Charles Dickens
Collected Works of Charles Dickens A Random Dickens Quote If there were no bad people, there would be no good lawyers. Original Sentence It was a dark and stormy night; the night was dark except at sunny
More information1 large 50g. 1 extra large 56g
National base for Standard Reference Release 28 slightly revised May, 206 Full Report (All s) 023, Egg, whole, raw, fresh Report Date: February 08, 208 06:8 EST values and weights are for edible portion.
More informationSequence analysis and comparison
The aim with sequence identification: Sequence analysis and comparison Marjolein Thunnissen Lund September 2012 Is there any known protein sequence that is homologous to mine? Are there any other species
More information1. Amino Acids and Peptides Structures and Properties
1. Amino Acids and Peptides Structures and Properties Chemical nature of amino acids The!-amino acids in peptides and proteins (excluding proline) consist of a carboxylic acid ( COOH) and an amino ( NH
More informationCollision Cross Section: Ideal elastic hard sphere collision:
Collision Cross Section: Ideal elastic hard sphere collision: ( r r 1 ) Where is the collision cross-section r 1 r ) ( 1 Where is the collision distance r 1 r These equations negate potential interactions
More informationAdvanced topics in bioinformatics
Feinberg Graduate School of the Weizmann Institute of Science Advanced topics in bioinformatics Shmuel Pietrokovski & Eitan Rubin Spring 2003 Course WWW site: http://bioinformatics.weizmann.ac.il/courses/atib
More informationRead more about Pauling and more scientists at: Profiles in Science, The National Library of Medicine, profiles.nlm.nih.gov
2018 Biochemistry 110 California Institute of Technology Lecture 2: Principles of Protein Structure Linus Pauling (1901-1994) began his studies at Caltech in 1922 and was directed by Arthur Amos oyes to
More informationMultiple Sequence Alignment, Gunnar Klau, December 9, 2005, 17:
Multiple Sequence Alignment, Gunnar Klau, December 9, 2005, 17:50 5001 5 Multiple Sequence Alignment The first part of this exposition is based on the following sources, which are recommended reading:
More information