Araport, a community portal for Arabidopsis. Data integration, sharing and reuse. sergio contrino University of Cambridge

Size: px
Start display at page:

Download "Araport, a community portal for Arabidopsis. Data integration, sharing and reuse. sergio contrino University of Cambridge"

Transcription

1 Araport, a community portal for Arabidopsis. Data integration, sharing and reuse sergio contrino University of Cambridge

2 Acknowledgements J Craig Venter Institute Chris Town Agnes Chan Vivek Krishnakumar University of Cambridge Gos Micklem Sergio Contrino Former members Jason Miller Ben Rosen Svetlana Karamycheva Eleanor Pence Maria Kim Seth Schobel Chia-Yi Cheng Irina Belyaeva Matt Hanlon Walter Moreira Texas Advanced Computing Center Matt Vaughn Steve Mock Rion Dooley Joe Stubbs Josue Coronel Erik Ferlanti TAIR Eva Huala Bob Muller

3

4 Araport11 Col-0 reannotation Cheng, Krishnakumar et al 2016, Plant Journal Protein-coding Genes Non-coding Genes Genomic Features Gene Loci mirna 27,416 -> 27,655 (NCBI, Maker, UniProt; Retired loci: 452 short coding sequences) 177 -> 325 (mirbase) Novel Transcribed Regions NAT 223 -> 1,115 (Publication) Transcript Variants AC 35,386 -> 48,359 (113 RNA-seq Datasets) lincrna Latest gene models from Maker, NCBI, UniProt RNA-seq supported isoforms Updated functional names uorf 58 -> 84 (Publication) 36 -> 2,444 (Publication) snorna 71 -> 287 (Publication) TAIR10 counts -> Araport11 counts (Evidence type) na -> 508 (RNA-seq) snrna Small RNAs na -> 35,846 (Prediction) 13 -> 82 (Publication) trna, rrna No change Data available via NCBI, ThaleMine, JBrowse, FTP, and Araport web services

5

6 eplant: Home

7 eplant: example Note: Uses ThaleMine web services

8

9 Araport App Store APIs developed by the community and the Araport team

10 Araport App Store: Dev Zone APIs developed by the community and the Araport team

11

12 AC JBrowse Over 100 Data Tracks Araport11 RNA-seq T-DNA Phytozome Vista plot

13 JBrowse - Self-Service Data Sharing AC CyVerse Workflow Variants (VCF) Genes (GFF) Features (BED) Your genomic feature files Data Store User upload Publication App Araport JBrowse Publish Output

14 JBrowse Self-Service Prototype Your genomic feature data will be displayed in the JBrowse menu and data tracks Araport11 gene models Community Data Tracks Seeding mirna-seq Seedling RNA-Seq reads Pollen RNA-Seq junctions AC

15

16

17 from: ThaleMine: A Warehouse for Arabidopsis Data Integration and Discovery Plant Cell Physiol. 58(1): e4(1-11) 2017 doi: /pcp/pcw200 araport@jcvi.org

18 AC ThaleMine - Gene Report (FT gene, AT1G65480)

19 ThaleMine - Gene Report (Expression, Interaction, Sequence) FT gene, AT1G65480 AC

20 ThaleMine - Gene Report (Seed Stock, Genotype, Phenotype) FT gene, AT1G65480 AC

21 What is InterMine? Open Source Data Warehouse github.com/intermine InterMine has three parts: 1. Database a. Load data from different data sets into single database 2. Webapp a. Mine the data b. Visualise! 3. Web services InterMine: a flexible data warehouse system for the integration and analysis of heterogeneous biological data Bioinformatics December 1; 28(23): YY

22 What sort of data is in InterMine? and in ThaleMine also - transcriptome profiles - GeneRif - Seed stocks and germplasms - Gene Locus History Genomic Pathways Interactions (complex and binary) Protein structures GO and other ontologies (Mammalian phenotype, etc) Protein domains Variation Expression and regulation Lots more

23 How does one get the data from an InterMine? 1. Search tools in UI a. Keyword search b. Sophisticated query tools 2. Web services a. Public API b. - interactive documentation YY

24 ThaleMine Pre-defined Queries Browse and export data in ThaleMine For example: List seed stocks related to flowering List all interactors (genetic and physical) for the FT gene Get gene expression values from different treatment conditions for a gene set AC

25 ThaleMine Build Your Own Query Transparent data model for flexible query construction AC

26 ThaleMine Region search

27 ThaleMine Region search

28 InterMine automatic (query) code generation xml example:

29 InterMine Lists analysis and operations

30 ThaleMine simple use case from: ThaleMine: A Warehouse for Arabidopsis Data Integration and Discovery Plant Cell Physiol. 58(1): e4(1-11) 2017 doi: /pcp/pcw200 araport@jcvi.org

31 Common Data Model Enables Crosstalk Between Resources ThaleMine Orthology Synteny Araport Model Organisms Human Legume Federation Cowpea Mine JGI SoyMine Yeast PhytoMine Bean Mine and more AC Medic Mine Peanut Mine and more

32 ThaleMine Recently Added Heatmap for 113 RNA-Seq Studies Ortholog links to other InterMines Gene Report Gene Report (e.g. dicer-like 1/AT1G01040) Gene Human RNA-seq study Medicago Gene List Analysis Genes JGI PhytoMine RNA-seq study AC

33 GO tool: an example using queries across different mines Summary view

34 GO tool: an example using queries across different mines Ontology Ontologyview view

35 Plant InterMines (afawk) ThaleMine - Arabidopsis thaliana PhytoMine - plants MedicMine - Medicago truncatula [HymenopteraMine - Bees, Ants & Wasps] SoyMine - Soybase soy bean data BeanMine - LegFed chado bean data LegumeMine - String bean, Soy, and Peanut PeanutMine - Peanut chado/gff data Wheat3BMine - Wheat chromosome 3B GrapeMine - Grapevine araport@jcvi.org

36 Next: CerealsMine? - choose an icon - choose a colour scheme - follow instructions at happy to help, and there is an active community at dev@lists.intermine.org demo? more info? questions? get in touch during the meeting sergio@intermine.org araport@jcvi.org

Browsing Genomic Information with Ensembl Plants

Browsing Genomic Information with Ensembl Plants Browsing Genomic Information with Ensembl Plants Etienne de Villiers, PhD (Adapted from slides by Bert Overduin EMBL-EBI) Outline of workshop Brief introduction to Ensembl Plants History Content Tutorial

More information

GEP Annotation Report

GEP Annotation Report GEP Annotation Report Note: For each gene described in this annotation report, you should also prepare the corresponding GFF, transcript and peptide sequence files as part of your submission. Student name:

More information

Introduction to the EMBL-EBI Ontology Lookup Service

Introduction to the EMBL-EBI Ontology Lookup Service Introduction to the EMBL-EBI Ontology Lookup Service Simon Jupp jupp@ebi.ac.uk, @simonjupp Samples, Phenotypes and Ontologies Team European Bioinformatics Institute Cambridge, UK. Ontologies in the life

More information

Supplemental Data. Perea-Resa et al. Plant Cell. (2012) /tpc

Supplemental Data. Perea-Resa et al. Plant Cell. (2012) /tpc Supplemental Data. Perea-Resa et al. Plant Cell. (22)..5/tpc.2.3697 Sm Sm2 Supplemental Figure. Sequence alignment of Arabidopsis LSM proteins. Alignment of the eleven Arabidopsis LSM proteins. Sm and

More information

Comparative genomics of gene families in relation with metabolic pathways for gene candidates highlighting

Comparative genomics of gene families in relation with metabolic pathways for gene candidates highlighting Comparative genomics of gene families in relation with metabolic pathways for gene candidates highlighting Delphine Larivière & David Couvin Under the supervision of Dominique This, Jean-François Dufayard

More information

EBI web resources II: Ensembl and InterPro. Yanbin Yin Spring 2013

EBI web resources II: Ensembl and InterPro. Yanbin Yin Spring 2013 EBI web resources II: Ensembl and InterPro Yanbin Yin Spring 2013 1 Outline Intro to genome annotation Protein family/domain databases InterPro, Pfam, Superfamily etc. Genome browser Ensembl Hands on Practice

More information

Synteny Portal Documentation

Synteny Portal Documentation Synteny Portal Documentation Synteny Portal is a web application portal for visualizing, browsing, searching and building synteny blocks. Synteny Portal provides four main web applications: SynCircos,

More information

Introduction to Bioinformatics. Shifra Ben-Dor Irit Orr

Introduction to Bioinformatics. Shifra Ben-Dor Irit Orr Introduction to Bioinformatics Shifra Ben-Dor Irit Orr Lecture Outline: Technical Course Items Introduction to Bioinformatics Introduction to Databases This week and next week What is bioinformatics? A

More information

86 Part 4 SUMMARY INTRODUCTION

86 Part 4 SUMMARY INTRODUCTION 86 Part 4 Chapter # AN INTEGRATION OF THE DESCRIPTIONS OF GENE NETWORKS AND THEIR MODELS PRESENTED IN SIGMOID (CELLERATOR) AND GENENET Podkolodny N.L. *1, 2, Podkolodnaya N.N. 1, Miginsky D.S. 1, Poplavsky

More information

Ensembl Genomes (non-chordates): Quick tour. This quick tour provides a brief introduction to Ensembl Genomes [2], the non-chordate genome browser.

Ensembl Genomes (non-chordates): Quick tour. This quick tour provides a brief introduction to Ensembl Genomes [2], the non-chordate genome browser. Paul Kersey [1] DNA & RNA Beginner 0.5 hour This quick tour provides a brief introduction to Ensembl Genomes [2], the non-chordate genome browser. Learning objectives: Basic understanding of Ensembl Genomes

More information

Proteomics. 2 nd semester, Department of Biotechnology and Bioinformatics Laboratory of Nano-Biotechnology and Artificial Bioengineering

Proteomics. 2 nd semester, Department of Biotechnology and Bioinformatics Laboratory of Nano-Biotechnology and Artificial Bioengineering Proteomics 2 nd semester, 2013 1 Text book Principles of Proteomics by R. M. Twyman, BIOS Scientific Publications Other Reference books 1) Proteomics by C. David O Connor and B. David Hames, Scion Publishing

More information

Small RNA in rice genome

Small RNA in rice genome Vol. 45 No. 5 SCIENCE IN CHINA (Series C) October 2002 Small RNA in rice genome WANG Kai ( 1, ZHU Xiaopeng ( 2, ZHONG Lan ( 1,3 & CHEN Runsheng ( 1,2 1. Beijing Genomics Institute/Center of Genomics and

More information

SoyBase, the USDA-ARS Soybean Genetics and Genomics Database

SoyBase, the USDA-ARS Soybean Genetics and Genomics Database SoyBase, the USDA-ARS Soybean Genetics and Genomics Database David Grant Victoria Carollo Blake Steven B. Cannon Kevin Feeley Rex T. Nelson Nathan Weeks SoyBase Site Map and Navigation Video Tutorials:

More information

Annotation of Plant Genomes using RNA-seq. Matteo Pellegrini (UCLA) In collaboration with Sabeeha Merchant (UCLA)

Annotation of Plant Genomes using RNA-seq. Matteo Pellegrini (UCLA) In collaboration with Sabeeha Merchant (UCLA) Annotation of Plant Genomes using RNA-seq Matteo Pellegrini (UCLA) In collaboration with Sabeeha Merchant (UCLA) inuscu1-35bp 5 _ 0 _ 5 _ What is Annotation inuscu2-75bp luscu1-75bp 0 _ 5 _ Reconstruction

More information

Procedure to Create NCBI KOGS

Procedure to Create NCBI KOGS Procedure to Create NCBI KOGS full details in: Tatusov et al (2003) BMC Bioinformatics 4:41. 1. Detect and mask typical repetitive domains Reason: masking prevents spurious lumping of non-orthologs based

More information

Mapping connections between the genome, ionome and the physical landscape. Photo by Bruce Bohm. David E Salt Purdue University, USA

Mapping connections between the genome, ionome and the physical landscape. Photo by Bruce Bohm. David E Salt Purdue University, USA Mapping connections between the genome, ionome and the physical landscape Photo by Bruce Bohm David E Salt Purdue University, USA What is the Ionome Environment Transcriptome Proteome Ionome The elemental

More information

EBI web resources II: Ensembl and InterPro

EBI web resources II: Ensembl and InterPro EBI web resources II: Ensembl and InterPro Yanbin Yin http://www.ebi.ac.uk/training/online/course/ 1 Homework 3 Go to http://www.ebi.ac.uk/interpro/training.htmland finish the second online training course

More information

Computational Structural Bioinformatics

Computational Structural Bioinformatics Computational Structural Bioinformatics ECS129 Instructor: Patrice Koehl http://koehllab.genomecenter.ucdavis.edu/teaching/ecs129 koehl@cs.ucdavis.edu Learning curve Math / CS Biology/ Chemistry Pre-requisite

More information

Networks & pathways. Hedi Peterson MTAT Bioinformatics

Networks & pathways. Hedi Peterson MTAT Bioinformatics Networks & pathways Hedi Peterson (peterson@quretec.com) MTAT.03.239 Bioinformatics 03.11.2010 Networks are graphs Nodes Edges Edges Directed, undirected, weighted Nodes Genes Proteins Metabolites Enzymes

More information

Gene Ontology and overrepresentation analysis

Gene Ontology and overrepresentation analysis Gene Ontology and overrepresentation analysis Kjell Petersen J Express Microarray analysis course Oslo December 2009 Presentation adapted from Endre Anderssen and Vidar Beisvåg NMC Trondheim Overview How

More information

Host_microbe_PPI - R package to analyse intra-species and interspecies protein-protein interactions in the model plant Arabidopsis thaliana

Host_microbe_PPI - R package to analyse intra-species and interspecies protein-protein interactions in the model plant Arabidopsis thaliana Host_microbe_PPI - R package to analyse intra-species and interspecies protein-protein interactions in the model plant Arabidopsis thaliana Thomas Nussbaumer 1,2 1 Institute of Network Biology (INET),

More information

Computational Biology: Basics & Interesting Problems

Computational Biology: Basics & Interesting Problems Computational Biology: Basics & Interesting Problems Summary Sources of information Biological concepts: structure & terminology Sequencing Gene finding Protein structure prediction Sources of information

More information

IUCLID Substance Data

IUCLID Substance Data 1 Workshop on CEFIC LRI Project EEM9.4 LRI AMBIT with IUCLID6 support and extended search capabilities IUCLID Substance Data Nikolay Kochev Ideaconsult Ltd. Sofia,Bulgaria 2 Chemical structure vs. Substance

More information

Genome Browsers And Genome Databases. Andy Conley Computational Genomics 2009

Genome Browsers And Genome Databases. Andy Conley Computational Genomics 2009 Genome Browsers And Genome Databases Andy Conley Computational What is a Genome Browser Genome browsers facilitate genomic analysis by presenting alignment, experimental and annotation data in the context

More information

Introduction Biology before Systems Biology: Reductionism Reduce the study from the whole organism to inner most details like protein or the DNA.

Introduction Biology before Systems Biology: Reductionism Reduce the study from the whole organism to inner most details like protein or the DNA. Systems Biology-Models and Approaches Introduction Biology before Systems Biology: Reductionism Reduce the study from the whole organism to inner most details like protein or the DNA. Taxonomy Study external

More information

BMD645. Integration of Omics

BMD645. Integration of Omics BMD645 Integration of Omics Shu-Jen Chen, Chang Gung University Dec. 11, 2009 1 Traditional Biology vs. Systems Biology Traditional biology : Single genes or proteins Systems biology: Simultaneously study

More information

Bioinformatics tools for phylogeny and visualization. Yanbin Yin

Bioinformatics tools for phylogeny and visualization. Yanbin Yin Bioinformatics tools for phylogeny and visualization Yanbin Yin 1 Homework assignment 5 1. Take the MAFFT alignment http://cys.bios.niu.edu/yyin/teach/pbb/purdue.cellwall.list.lignin.f a.aln as input and

More information

Androgen-independent prostate cancer

Androgen-independent prostate cancer The following tutorial walks through the identification of biological themes in a microarray dataset examining androgen-independent. Visit the GeneSifter Data Center (www.genesifter.net/web/datacenter.html)

More information

Predicting Protein Functions and Domain Interactions from Protein Interactions

Predicting Protein Functions and Domain Interactions from Protein Interactions Predicting Protein Functions and Domain Interactions from Protein Interactions Fengzhu Sun, PhD Center for Computational and Experimental Genomics University of Southern California Outline High-throughput

More information

SUSTAINABLE AND INTEGRAL EXPLOITATION OF AGAVE

SUSTAINABLE AND INTEGRAL EXPLOITATION OF AGAVE SUSTAINABLE AND INTEGRAL EXPLOITATION OF AGAVE Editor Antonia Gutiérrez-Mora Compilers Benjamín Rodríguez-Garay Silvia Maribel Contreras-Ramos Manuel Reinhart Kirchmayr Marisela González-Ávila Index 1.

More information

Arcgis Tutorial Manual READ ONLINE

Arcgis Tutorial Manual READ ONLINE Arcgis Tutorial Manual READ ONLINE ArcGIS Desktop 10 Trial Help - Note: The Quick Start Guide contains instructions that do not pertain to the trial edition. Many tutorials are found in the ArcGIS Desktop

More information

Bio-Medical Text Mining with Machine Learning

Bio-Medical Text Mining with Machine Learning Sumit Madan Department of Bioinformatics - Fraunhofer SCAI Textual Knowledge PubMed Journals, Books Patents EHRs What is Bio-Medical Text Mining? Phosphorylation of glycogen synthase kinase 3 beta at Threonine,

More information

Lesson Overview. Gene Regulation and Expression. Lesson Overview Gene Regulation and Expression

Lesson Overview. Gene Regulation and Expression. Lesson Overview Gene Regulation and Expression 13.4 Gene Regulation and Expression THINK ABOUT IT Think of a library filled with how-to books. Would you ever need to use all of those books at the same time? Of course not. Now picture a tiny bacterium

More information

Algorithms in Computational Biology (236522) spring 2008 Lecture #1

Algorithms in Computational Biology (236522) spring 2008 Lecture #1 Algorithms in Computational Biology (236522) spring 2008 Lecture #1 Lecturer: Shlomo Moran, Taub 639, tel 4363 Office hours: 15:30-16:30/by appointment TA: Ilan Gronau, Taub 700, tel 4894 Office hours:??

More information

USDA-DOE Plant Feedstock Genomics for Bioenergy

USDA-DOE Plant Feedstock Genomics for Bioenergy USDA-DOE Plant Feedstock Genomics for Bioenergy BERAC Thursday, June 7, 2012 Cathy Ronning, DOE-BER Ed Kaleikau, USDA-NIFA Plant Feedstock Genomics for Bioenergy Joint competitive grants program initiated

More information

Mathangi Thiagarajan Rice Genome Annotation Workshop May 23rd, 2007

Mathangi Thiagarajan Rice Genome Annotation Workshop May 23rd, 2007 -2 Transcript Alignment Assembly and Automated Gene Structure Improvements Using PASA-2 Mathangi Thiagarajan mathangi@jcvi.org Rice Genome Annotation Workshop May 23rd, 2007 About PASA PASA is an open

More information

Grundlagen der Bioinformatik Summer semester Lecturer: Prof. Daniel Huson

Grundlagen der Bioinformatik Summer semester Lecturer: Prof. Daniel Huson Grundlagen der Bioinformatik, SS 10, D. Huson, April 12, 2010 1 1 Introduction Grundlagen der Bioinformatik Summer semester 2010 Lecturer: Prof. Daniel Huson Office hours: Thursdays 17-18h (Sand 14, C310a)

More information

Statistical Inferences for Isoform Expression in RNA-Seq

Statistical Inferences for Isoform Expression in RNA-Seq Statistical Inferences for Isoform Expression in RNA-Seq Hui Jiang and Wing Hung Wong February 25, 2009 Abstract The development of RNA sequencing (RNA-Seq) makes it possible for us to measure transcription

More information

Principles of QTL Mapping. M.Imtiaz

Principles of QTL Mapping. M.Imtiaz Principles of QTL Mapping M.Imtiaz Introduction Definitions of terminology Reasons for QTL mapping Principles of QTL mapping Requirements For QTL Mapping Demonstration with experimental data Merit of QTL

More information

Systematic prediction of gene function in Arabidopsis thaliana using a probabilistic functional gene network

Systematic prediction of gene function in Arabidopsis thaliana using a probabilistic functional gene network Systematic prediction of gene function in Arabidopsis thaliana using a probabilistic functional gene network Sohyun Hwang 1, Seung Y Rhee 2, Edward M Marcotte 3,4 & Insuk Lee 1 protocol 1 Department of

More information

Genome Annotation. Qi Sun Bioinformatics Facility Cornell University

Genome Annotation. Qi Sun Bioinformatics Facility Cornell University Genome Annotation Qi Sun Bioinformatics Facility Cornell University Some basic bioinformatics tools BLAST PSI-BLAST - Position-Specific Scoring Matrix HMM - Hidden Markov Model NCBI BLAST How does BLAST

More information

Chapter 2: Extensions to Mendel: Complexities in Relating Genotype to Phenotype.

Chapter 2: Extensions to Mendel: Complexities in Relating Genotype to Phenotype. Chapter 2: Extensions to Mendel: Complexities in Relating Genotype to Phenotype. please read pages 38-47; 49-55;57-63. Slide 1 of Chapter 2 1 Extension sot Mendelian Behavior of Genes Single gene inheritance

More information

Wendy T Vu 1*, Peter L Chang 1, Ken S Moriuchi 2 and Maren L Friesen 1,3*

Wendy T Vu 1*, Peter L Chang 1, Ken S Moriuchi 2 and Maren L Friesen 1,3* Vu et al. BMC Evolutionary Biology (2015) 15:59 DOI 10.1186/s12862-015-0322-4 RESEARCH ARTICLE Open Access Genetic variation of transgenerational plasticity of offspring germination in response to salinity

More information

Model plants and their Role in genetic manipulation. Mitesh Shrestha

Model plants and their Role in genetic manipulation. Mitesh Shrestha Model plants and their Role in genetic manipulation Mitesh Shrestha Definition of Model Organism Specific species or organism Extensively studied in research laboratories Advance our understanding of Cellular

More information

MIP543 RNA Biology Fall 2015

MIP543 RNA Biology Fall 2015 MIP543 RNA Biology Fall 2015 Credits: 3 Term Offered: Day and Time: Fall (odd years) Mondays and Wednesdays, 4:00-5:15 pm Classroom: MRB 123 Course Instructor: Dr. Jeffrey Wilusz, Professor, MIP Office:

More information

SUPPLEMENTARY INFORMATION

SUPPLEMENTARY INFORMATION Supplementary Discussion Rationale for using maternal ythdf2 -/- mutants as study subject To study the genetic basis of the embryonic developmental delay that we observed, we crossed fish with different

More information

Bioinformatics. Dept. of Computational Biology & Bioinformatics

Bioinformatics. Dept. of Computational Biology & Bioinformatics Bioinformatics Dept. of Computational Biology & Bioinformatics 3 Bioinformatics - play with sequences & structures Dept. of Computational Biology & Bioinformatics 4 ORGANIZATION OF LIFE ROLE OF BIOINFORMATICS

More information

Nature Genetics: doi: /ng Supplementary Figure 1. The phenotypes of PI , BR121, and Harosoy under short-day conditions.

Nature Genetics: doi: /ng Supplementary Figure 1. The phenotypes of PI , BR121, and Harosoy under short-day conditions. Supplementary Figure 1 The phenotypes of PI 159925, BR121, and Harosoy under short-day conditions. (a) Plant height. (b) Number of branches. (c) Average internode length. (d) Number of nodes. (e) Pods

More information

Introduction to Bioinformatics

Introduction to Bioinformatics CSCI8980: Applied Machine Learning in Computational Biology Introduction to Bioinformatics Rui Kuang Department of Computer Science and Engineering University of Minnesota kuang@cs.umn.edu History of Bioinformatics

More information

Correlation between flowering time, circadian rhythm and gene expression in Capsella bursa-pastoris

Correlation between flowering time, circadian rhythm and gene expression in Capsella bursa-pastoris Correlation between flowering time, circadian rhythm and gene expression in Capsella bursa-pastoris Johanna Nyström Degree project in biology, Bachelor of science, 2013 Examensarbete i biologi 15 hp till

More information

The Developmental Transcriptome of the Mosquito Aedes aegypti, an invasive species and major arbovirus vector.

The Developmental Transcriptome of the Mosquito Aedes aegypti, an invasive species and major arbovirus vector. The Developmental Transcriptome of the Mosquito Aedes aegypti, an invasive species and major arbovirus vector. Omar S. Akbari*, Igor Antoshechkin*, Henry Amrhein, Brian Williams, Race Diloreto, Jeremy

More information

Application integration: Providing coherent drug discovery solutions

Application integration: Providing coherent drug discovery solutions Application integration: Providing coherent drug discovery solutions Mitch Miller, Manish Sud, LION bioscience, American Chemical Society 22 August 2002 Overview 2 Introduction: exploring application integration

More information

Chapter 15 Active Reading Guide Regulation of Gene Expression

Chapter 15 Active Reading Guide Regulation of Gene Expression Name: AP Biology Mr. Croft Chapter 15 Active Reading Guide Regulation of Gene Expression The overview for Chapter 15 introduces the idea that while all cells of an organism have all genes in the genome,

More information

Prospecting for Green Revolution Genes

Prospecting for Green Revolution Genes Learning Objectives: Prospecting for Green Revolution Genes 1) Discover how changes in individual genes produce phenotypic change 2) Learn to apply bioinformatics tools to identify groups of related genes

More information

Hands-On Nine The PAX6 Gene and Protein

Hands-On Nine The PAX6 Gene and Protein Hands-On Nine The PAX6 Gene and Protein Main Purpose of Hands-On Activity: Using bioinformatics tools to examine the sequences, homology, and disease relevance of the Pax6: a master gene of eye formation.

More information

SnoPatrol: How many snorna genes are there? Supplementary

SnoPatrol: How many snorna genes are there? Supplementary SnoPatrol: How many snorna genes are there? Supplementary materials. Paul P. Gardner 1, Alex G. Bateman 1 and Anthony M. Poole 2,3 1 Wellcome Trust Sanger Institute, Wellcome Trust Genome Campus, Hinxton,

More information

Genomes and Their Evolution

Genomes and Their Evolution Chapter 21 Genomes and Their Evolution PowerPoint Lecture Presentations for Biology Eighth Edition Neil Campbell and Jane Reece Lectures by Chris Romero, updated by Erin Barley with contributions from

More information

Leveraging GIS data and tools for maintaining hydraulic sewer models

Leveraging GIS data and tools for maintaining hydraulic sewer models Leveraging GIS data and tools for maintaining hydraulic sewer models Ben Gamble & Joseph Koran Metropolitan Sewer District of Greater Cincinnati Carl C. Chan & Michael York CDM Smith Ben Gamble Senior

More information

GCD3033:Cell Biology. Transcription

GCD3033:Cell Biology. Transcription Transcription Transcription: DNA to RNA A) production of complementary strand of DNA B) RNA types C) transcription start/stop signals D) Initiation of eukaryotic gene expression E) transcription factors

More information

Genetic dissection of the Arabidopsis thaliana ionome

Genetic dissection of the Arabidopsis thaliana ionome Genetic dissection of the Arabidopsis thaliana ionome Genome Ionome Landscape distribution David E Salt Purdue University, USA What is the Ionome Environment Transcriptome Proteome Ionome The elemental

More information

TRANSCRIPTOMICS. (or the analysis of the transcriptome) Mario Cáceres. Main objectives of genomics. Determine the entire DNA sequence of an organism

TRANSCRIPTOMICS. (or the analysis of the transcriptome) Mario Cáceres. Main objectives of genomics. Determine the entire DNA sequence of an organism TRANSCRIPTOMICS (or the analysis of the transcriptome) Mario Cáceres Main objectives of genomics Determine the entire DNA sequence of an organism Identify and annotate the complete set of genes encoded

More information

The Saguaro Genome. Toward the Ecological Genomics of a Sonoran Desert Icon. Dr. Dario Copetti June 30, 2015 STEMAZing workshop TCSS

The Saguaro Genome. Toward the Ecological Genomics of a Sonoran Desert Icon. Dr. Dario Copetti June 30, 2015 STEMAZing workshop TCSS The Saguaro Genome Toward the Ecological Genomics of a Sonoran Desert Icon Dr. Dario Copetti June 30, 2015 STEMAZing workshop TCSS Why study a genome? - the genome contains the genetic information of an

More information

Heterosis and inbreeding depression of epigenetic Arabidopsis hybrids

Heterosis and inbreeding depression of epigenetic Arabidopsis hybrids Heterosis and inbreeding depression of epigenetic Arabidopsis hybrids Plant growth conditions The soil was a 1:1 v/v mixture of loamy soil and organic compost. Initial soil water content was determined

More information

Administering your Enterprise Geodatabase using Python. Jill Penney

Administering your Enterprise Geodatabase using Python. Jill Penney Administering your Enterprise Geodatabase using Python Jill Penney Assumptions Basic knowledge of python Basic knowledge enterprise geodatabases and workflows You want code Please turn off or silence cell

More information

Предсказание и анализ промотерных последовательностей. Татьяна Татаринова

Предсказание и анализ промотерных последовательностей. Татьяна Татаринова Предсказание и анализ промотерных последовательностей Татьяна Татаринова Eukaryotic Transcription 2 Initiation Promoter: the DNA sequence that initially binds the RNA polymerase The structure of promoter-polymerase

More information

Bayesian Clustering of Multi-Omics

Bayesian Clustering of Multi-Omics Bayesian Clustering of Multi-Omics for Cardiovascular Diseases Nils Strelow 22./23.01.2019 Final Presentation Trends in Bioinformatics WS18/19 Recap Intermediate presentation Precision Medicine Multi-Omics

More information

Environmental Chemistry through Intelligent Atmospheric Data Analysis (EnChIlADA): A Platform for Mining ATOFMS and Other Atmospheric Data

Environmental Chemistry through Intelligent Atmospheric Data Analysis (EnChIlADA): A Platform for Mining ATOFMS and Other Atmospheric Data Environmental Chemistry through Intelligent Atmospheric Data Analysis (EnChIlADA): A Platform for Mining ATOFMS and Other Atmospheric Data Katie Barton, John Choiniere, Melanie Yuen, and Deborah Gross

More information

Algorithmics and Bioinformatics

Algorithmics and Bioinformatics Algorithmics and Bioinformatics Gregory Kucherov and Philippe Gambette LIGM/CNRS Université Paris-Est Marne-la-Vallée, France Schedule Course webpage: https://wikimpri.dptinfo.ens-cachan.fr/doku.php?id=cours:c-1-32

More information

Utilizing Illumina high-throughput sequencing technology to gain insights into small RNA biogenesis and function

Utilizing Illumina high-throughput sequencing technology to gain insights into small RNA biogenesis and function Utilizing Illumina high-throughput sequencing technology to gain insights into small RNA biogenesis and function Brian D. Gregory Department of Biology Penn Genome Frontiers Institute University of Pennsylvania

More information

Chemical Data Retrieval and Management

Chemical Data Retrieval and Management Chemical Data Retrieval and Management ChEMBL, ChEBI, and the Chemistry Development Kit Stephan A. Beisken What is EMBL-EBI? Part of the European Molecular Biology Laboratory International, non-profit

More information

Investigation 3: Comparing DNA Sequences to Understand Evolutionary Relationships with BLAST

Investigation 3: Comparing DNA Sequences to Understand Evolutionary Relationships with BLAST Investigation 3: Comparing DNA Sequences to Understand Evolutionary Relationships with BLAST Introduction Bioinformatics is a powerful tool which can be used to determine evolutionary relationships and

More information

Tree Building Activity

Tree Building Activity Tree Building Activity Introduction In this activity, you will construct phylogenetic trees using a phenotypic similarity (cartoon microbe pictures) and genotypic similarity (real microbe sequences). For

More information

In Genomes, Two Types of Genes

In Genomes, Two Types of Genes In Genomes, Two Types of Genes Protein-coding: [Start codon] [codon 1] [codon 2] [ ] [Stop codon] + DNA codons translated to amino acids to form a protein Non-coding RNAs (NcRNAs) No consistent patterns

More information

BIOINFORMATICS LAB AP BIOLOGY

BIOINFORMATICS LAB AP BIOLOGY BIOINFORMATICS LAB AP BIOLOGY Bioinformatics is the science of collecting and analyzing complex biological data. Bioinformatics combines computer science, statistics and biology to allow scientists to

More information

HEREDITY: Objective: I can describe what heredity is because I can identify traits and characteristics

HEREDITY: Objective: I can describe what heredity is because I can identify traits and characteristics Mendel and Heredity HEREDITY: SC.7.L.16.1 Understand and explain that every organism requires a set of instructions that specifies its traits, that this hereditary information. Objective: I can describe

More information

Computational Genomics. Systems biology. Putting it together: Data integration using graphical models

Computational Genomics. Systems biology. Putting it together: Data integration using graphical models 02-710 Computational Genomics Systems biology Putting it together: Data integration using graphical models High throughput data So far in this class we discussed several different types of high throughput

More information

Supplemental Information

Supplemental Information Molecular Cell, Volume 52 Supplemental Information The Translational Landscape of the Mammalian Cell Cycle Craig R. Stumpf, Melissa V. Moreno, Adam B. Olshen, Barry S. Taylor, and Davide Ruggero Supplemental

More information

13.4 Gene Regulation and Expression

13.4 Gene Regulation and Expression 13.4 Gene Regulation and Expression Lesson Objectives Describe gene regulation in prokaryotes. Explain how most eukaryotic genes are regulated. Relate gene regulation to development in multicellular organisms.

More information

VCE BIOLOGY Relationship between the key knowledge and key skills of the Study Design and the Study Design

VCE BIOLOGY Relationship between the key knowledge and key skills of the Study Design and the Study Design VCE BIOLOGY 2006 2014 Relationship between the key knowledge and key skills of the 2000 2005 Study Design and the 2006 2014 Study Design The following table provides a comparison of the key knowledge (and

More information

Principles of Genetics

Principles of Genetics Principles of Genetics Snustad, D ISBN-13: 9780470903599 Table of Contents C H A P T E R 1 The Science of Genetics 1 An Invitation 2 Three Great Milestones in Genetics 2 DNA as the Genetic Material 6 Genetics

More information

Related Courses He who asks is a fool for five minutes, but he who does not ask remains a fool forever.

Related Courses He who asks is a fool for five minutes, but he who does not ask remains a fool forever. CSE 527 Computational Biology http://www.cs.washington.edu/527 Lecture 1: Overview & Bio Review Autumn 2004 Larry Ruzzo Related Courses He who asks is a fool for five minutes, but he who does not ask remains

More information

Big Idea 3: Living systems store, retrieve, transmit, and respond to information essential to life processes.

Big Idea 3: Living systems store, retrieve, transmit, and respond to information essential to life processes. Big Idea 3: Living systems store, retrieve, transmit, and respond to information essential to life processes. Enduring understanding 3.A: Heritable information provides for continuity of life. Essential

More information

Chapter 9 DNA recognition by eukaryotic transcription factors

Chapter 9 DNA recognition by eukaryotic transcription factors Chapter 9 DNA recognition by eukaryotic transcription factors TRANSCRIPTION 101 Eukaryotic RNA polymerases RNA polymerase RNA polymerase I RNA polymerase II RNA polymerase III RNA polymerase IV Function

More information

Introduction to molecular biology. Mitesh Shrestha

Introduction to molecular biology. Mitesh Shrestha Introduction to molecular biology Mitesh Shrestha Molecular biology: definition Molecular biology is the study of molecular underpinnings of the process of replication, transcription and translation of

More information

11/24/13. Science, then, and now. Computational Structural Bioinformatics. Learning curve. ECS129 Instructor: Patrice Koehl

11/24/13. Science, then, and now. Computational Structural Bioinformatics. Learning curve. ECS129 Instructor: Patrice Koehl Computational Structural Bioinformatics ECS129 Instructor: Patrice Koehl http://www.cs.ucdavis.edu/~koehl/teaching/ecs129/index.html koehl@cs.ucdavis.edu Learning curve Math / CS Biology/ Chemistry Pre-requisite

More information

GENETICS - CLUTCH CH.1 INTRODUCTION TO GENETICS.

GENETICS - CLUTCH CH.1 INTRODUCTION TO GENETICS. !! www.clutchprep.com CONCEPT: HISTORY OF GENETICS The earliest use of genetics was through of plants and animals (8000-1000 B.C.) Selective breeding (artificial selection) is the process of breeding organisms

More information

training workshop 2015

training workshop 2015 TransPLANT user training workshop 2015 Slides: http://tinyurl.com/transplant2015 Workshop on variation data EMBL-EBI Hinxton-UK 2nd July 2015 Ensembl Genomes Team Notes: This workshop is based on Ensembl

More information

Introduction to Bioinformatics

Introduction to Bioinformatics Introduction to Bioinformatics Jianlin Cheng, PhD Department of Computer Science Informatics Institute 2011 Topics Introduction Biological Sequence Alignment and Database Search Analysis of gene expression

More information

Bioinformatics Chapter 1. Introduction

Bioinformatics Chapter 1. Introduction Bioinformatics Chapter 1. Introduction Outline! Biological Data in Digital Symbol Sequences! Genomes Diversity, Size, and Structure! Proteins and Proteomes! On the Information Content of Biological Sequences!

More information

ToxiCat: Hybrid Named Entity Recognition services to support curation of the Comparative Toxicogenomic Database

ToxiCat: Hybrid Named Entity Recognition services to support curation of the Comparative Toxicogenomic Database ToxiCat: Hybrid Named Entity Recognition services to support curation of the Comparative Toxicogenomic Database Dina Vishnyakova 1,2, 4, *, Julien Gobeill 1,3,4, Emilie Pasche 1,2,3,4 and Patrick Ruch

More information

Genetics 275 Notes Week 7

Genetics 275 Notes Week 7 Cytoplasmic Inheritance Genetics 275 Notes Week 7 Criteriafor recognition of cytoplasmic inheritance: 1. Reciprocal crosses give different results -mainly due to the fact that the female parent contributes

More information

How much non-coding DNA do eukaryotes require?

How much non-coding DNA do eukaryotes require? How much non-coding DNA do eukaryotes require? Andrei Zinovyev UMR U900 Computational Systems Biology of Cancer Institute Curie/INSERM/Ecole de Mine Paritech Dr. Sebastian Ahnert Dr. Thomas Fink Bioinformatics

More information

10-810: Advanced Algorithms and Models for Computational Biology. microrna and Whole Genome Comparison

10-810: Advanced Algorithms and Models for Computational Biology. microrna and Whole Genome Comparison 10-810: Advanced Algorithms and Models for Computational Biology microrna and Whole Genome Comparison Central Dogma: 90s Transcription factors DNA transcription mrna translation Proteins Central Dogma:

More information

Evaluating Physical, Chemical, and Biological Impacts from the Savannah Harbor Expansion Project Cooperative Agreement Number W912HZ

Evaluating Physical, Chemical, and Biological Impacts from the Savannah Harbor Expansion Project Cooperative Agreement Number W912HZ Evaluating Physical, Chemical, and Biological Impacts from the Savannah Harbor Expansion Project Cooperative Agreement Number W912HZ-13-2-0013 Annual Report FY 2018 Submitted by Sergio Bernardes and Marguerite

More information

Lecture Materials are available on the 321 web site

Lecture Materials are available on the 321 web site MENDEL AND MODELS Explore the Course Web Site http://fire.biol.wwu.edu/trent/trent/biol321index.html Lecture Materials are available on the 321 web site http://fire.biol.wwu.edu/trent/trent/321lectureindex.html

More information

RNA & PROTEIN SYNTHESIS. Making Proteins Using Directions From DNA

RNA & PROTEIN SYNTHESIS. Making Proteins Using Directions From DNA RNA & PROTEIN SYNTHESIS Making Proteins Using Directions From DNA RNA & Protein Synthesis v Nitrogenous bases in DNA contain information that directs protein synthesis v DNA remains in nucleus v in order

More information

Genomes Comparision via de Bruijn graphs

Genomes Comparision via de Bruijn graphs Genomes Comparision via de Bruijn graphs Student: Ilya Minkin Advisor: Son Pham St. Petersburg Academic University June 4, 2012 1 / 19 Synteny Blocks: Algorithmic challenge Suppose that we are given two

More information

GENOME DUPLICATION AND GENE ANNOTATION: AN EXAMPLE FOR A REFERENCE PLANT SPECIES.

GENOME DUPLICATION AND GENE ANNOTATION: AN EXAMPLE FOR A REFERENCE PLANT SPECIES. GENOME DUPLICATION AND GENE ANNOTATION: AN EXAMPLE FOR A REFERENCE PLANT SPECIES. Alessandra Vigilante, Mara Sangiovanni, Chiara Colantuono, Luigi Frusciante and Maria Luisa Chiusano Dept. of Soil, Plant,

More information

Leveraging Web GIS: An Introduction to the ArcGIS portal

Leveraging Web GIS: An Introduction to the ArcGIS portal Leveraging Web GIS: An Introduction to the ArcGIS portal Derek Law Product Management DLaw@esri.com Agenda Web GIS pattern Product overview Installation and deployment Configuration options Security options

More information

Geodatabase Best Practices. Dave Crawford Erik Hoel

Geodatabase Best Practices. Dave Crawford Erik Hoel Geodatabase Best Practices Dave Crawford Erik Hoel Geodatabase best practices - outline Geodatabase creation Data ownership Data model Data configuration Geodatabase behaviors Data integrity and validation

More information