Genotyping by sequencing and data analysis. Ross Whetten North Carolina State University

Size: px
Start display at page:

Download "Genotyping by sequencing and data analysis. Ross Whetten North Carolina State University"

Transcription

1 Genotyping by sequencing and data analysis Ross Whetten North Carolina State University

2 Stein (2010) Genome Biology 11:207

3 More New Technology on the Horizon

4 Genotyping By Sequencing Timeline 2007 Complexity Reduction of Polymorphic Sequences van Orsouw et al., PLoS ONE 2(11): e1172. SNP discovery using 454 sequencing, genotyping via Keygene SNPWave > patent application 2008 Rapid SNP Discovery and Genetic Mapping Using Sequenced RAD Markers. Baird et al. PLoS ONE 3(10): e3376 Direct SNP genotyping by Illumina sequencing High-throughput genotyping by whole-genome resequencing. Huang et al., Genome Res 19: Low-coverage whole genome resequencing of rice RILs

5 Genotyping By Sequencing Timeline 2011 Multiplex shotgun genotyping for rapid and efficient genetic mapping. Andolfatto et al., Genome Res. 21(4): Single restriction enzyme digest, HMM model for data analysis 2011 A Robust, Simple Genotyping-by-Sequencing (GBS) Approach for High Diversity Species. Elshire et al., PLoS ONE 6(5): e Simplified protocol for high throughput 2012 Development of High-Density Genetic Maps for Barley and Wheat Using a Novel Two-Enzyme Genotyping-by-Sequencing Approach. Poland et al., PLoS ONE 7(2): e32253 Two-enzyme version of Elshire et al protocol

6 Genotyping By Sequencing Timeline 2012 Double-digest RAD-seq. Peterson et al., PLoS ONE 7(5): e37135 Two-enzyme method similar to Poland et al, with size-selection to increase reproducibility of genotyping 2013 RESTseq Efficient Benchtop Population Genomics with RESTriction Fragment SEQuencing. Stolle & Moritz, PLoS ONE 8(5): e63960 Complexity reduction for fewer markers, higher multiplexing, reduced costs

7 Multiplexing Strategies Multiplexing = pooling multiple different samples, each identified by a unique tag, and sequencing the mixture First approach: Barcodes in the sequence reads. User-designed, no modification to sequencing protocol Illumina index adapters. Variable sequence in the middle of one sequencing adapter, requires an independent sequence read

8 Sequencing technology overview Illumina Glass flowcell, with 8 separate lanes. GAIIx ~ million clusters per lane, 150-nt reads HiSeq ~ 200 million clusters per lane, 100-nt reads MiSeq 25 million clusters, 300-nt reads

9 Sequencing technology overview Illumina Fragment DNA, ligate adaptor oligo Single-stranded DNA binds to surface

10 Sequencing technology overview Illumina Extend surfacebound primer, denature, wash away template New strand anneals to complementary primer, which is extended to copy it Many cycles create ~ 1000 molecules in a cluster. After PCR, free ends are blocked

11 Sequencing technology overview Illumina Another perspective of the amplification process, showing the clusters of products

12 Sequencing technology overview Illumina

13 Sequencing technology overview Illumina

14 Sequencing technology overview Illumina GCTGA CTTAG Cycle 1 Cycle 2 Cycle 3 Cycle 4 Cycle 5 AGCCG TAAGT Although four different colors are used for the fluorescent nucleotides, only two lasers are used to excite the fluorescence. The fluorescent labels are grouped in pairs - labels on A and C are excited by a red laser, and labels on G and T are excited by a green laser. The software assumes that signal from both lasers will be balanced at each cycle. This means that distinguishing between the A signal and the C signal is more difficult for the instrument than A versus G or A versus T. Base substitution errors are the most common type of sequencing error for Illumina instruments.

15 Sequencing technology overview Illumina The software assumes that signal from both lasers will be balanced at each cycle. It also uses patterns of base addition in adjacent clusters to determine which clusters are a single template, and which are a mixture of templates, and rejects data from clusters that are mixed templates. Barcode RE site A T C G A A T T C C G T A A A T T C G C A T A A T T C T A G C A A T T C Fixed-length barcodes Poor yield of sequence A T C G A A T T C u v w C G T A G A A T T C x y G C A T C T A A T T C z T A G C T G T A A T T C Variable-length barcodes Good yield of sequence

16 Experimental design Why is this important? It determines the value of the data for analysis Doesn t the statistical software take care of that? No amount of statistical sophistication can separate confounded factors after data have been collected Auer and Doerge, 2010 Sources of variation Nuisance factors technical problems, random noise Experimental factors variation among individuals

17 Experimental design What are the sources of variation in the experiment? Among individuals Among treatments Among sequencing runs Among lanes/sectors within run Among library preparations Which sources of variation are of interest? Avoid confounding effects of interest with nuisance effects Allocate effort to estimate effects of greatest interest Block to replicate measurements across nuisance effects Exploit barcoding/multiplexing tools for blocking Balanced or partially-balanced designs possible, using complete or incomplete blocking

18 Experimental design Genetics 185(2):405-16, 2010 Suppose 21 libraries (each containing different pools of multiplexed samples) are each sequenced in a separate lane Nuisance effects are confounded with sample differences

19 Auer & Doerge, Genetics 185:405, 2010 Experimental design

20 DNA sequencing provides options

21 Just the facts

22 Sense from sequence reads: methods for alignment and assembly. Flicek & Birney, Nat Methods 6(11 Suppl):S6-S12, 2009 the individual outputs of the sequence machines are essentially worthless by themselves once analyzed collectively DNA sequencing reads have tremendous versatility... Biologists interested in sequencing to answer their experimental questions should prepare themselves to join a fast-moving field and embrace the tools being developed specifically for it. As more sequence is generated, effective use of computational resources will be more and more important.

23 STACKS software data analysis workflow Catchen et al., Mol Ecol 22, , 2013

24 Analysis of Deep Sequencing Data Summary Data are not information Information is not knowledge Knowledge is not wisdom Anon.

25 Conclusions Sequencing technology continues to develop rapidly Data analysis methods are also developing, but almost all use Linux a potential barrier for biologists The costs of DNA sequencing are dropping; experiments are likely to be less expensive next year than this year Skills for computational data analysis are a key component of successful GBS experiments

26 Thank you!

27 Computational thinking four general principles Decompose a complex problem to simple steps. Linux is based on simple tools that do one thing well; these tools require problems to be framed in simple terms. Look for patterns. Recognizing similarities among different types of problems allows re-use of the same tools in new contexts. Generalize patterns to create abstract versions. A tool is most powerful when it can be applied to a variety of problems that all share common features Combine simple tools into more complex pipelines. Repetitive tasks are what computers are good at our job is to build the algorithms, or sequences of simple steps, that allow the computer to do those repetitive tasks so we don t have to.

28 Key principles from the Eric Raymond book chapter Clarity is better than cleverness. Document everything you do, because you won t remember what you did, or why Programmer time is more expensive than machine time. Don t worry about optimizing things unless it is necessary Prototype before polishing get it working before you optimize it. It is often easiest to start with something very simple, then add complexity and capability in steps Design for simplicity; add complexity only where you must. Make things as simple as possible, but no simpler paraphrased from Albert Einstein

29 Illumina flowcell geometry (Hiseq) Tiles within a Hiseq lane are numbered using a system with three separate numbers. The first digit denotes which surface (1 = lower, 2 = upper), the second denotes a vertical swath (1 = left, 2 = middle, 3 = right), and the last two digits denote a tile within that swath (01 means closest to the outflow end of the lane; 08 or 16 means closest to the inflow end of the lane

30 Understanding FASTQ format or what do all these symbols mean? Instrument ID lane tile X Y barcode read# flowcell Header lines sequence quality scores Quality scores are numbers that represent the probability that the given base call is an error. These probabilities are always less than 1, so the value is given as 10 times minus log(10) of the probability For example, an error probability of (1x10-3 ) is represented as a quality score of 30. The numbers are converted into text characters so they occupy less space a single character is as meaningful as 2 numbers plus a space between adjacent values

31 Understanding FASTQ format Illumina v1.8 header 1:N:ATCACG Instrument /flowcell ID lane tile X Y index read# Header lines sequence quality scores Unfortunately, at least four different ways of converting numbers to characters have been used, and header line formats have also changed, so one aspect of data analysis is knowing what you have.

Introduction to next-generation sequencing data

Introduction to next-generation sequencing data Introduction to next-generation sequencing data David Simpson Centre for Experimental Medicine Queens University Belfast http://www.qub.ac.uk/research-centres/cem/ Outline History of DNA sequencing NGS

More information

Illumina Sequencing Technology

Illumina Sequencing Technology Illumina Sequencing Technology Highest data accuracy, simple workflow, and a broad range of applications. Introduction Figure 1: Illumina Flow Cell Illumina sequencing technology leverages clonal array

More information

SEQUENCING. From Sample to Sequence-Ready

SEQUENCING. From Sample to Sequence-Ready SEQUENCING From Sample to Sequence-Ready ACCESS ARRAY SYSTEM HIGH-QUALITY LIBRARIES, NOT ONCE, BUT EVERY TIME The highest-quality amplicons more sensitive, accurate, and specific Full support for all major

More information

TruSeq Custom Amplicon v1.5

TruSeq Custom Amplicon v1.5 Data Sheet: Targeted Resequencing TruSeq Custom Amplicon v1.5 A new and improved amplicon sequencing solution for interrogating custom regions of interest. Highlights Figure 1: TruSeq Custom Amplicon Workflow

More information

Illumina TruSeq DNA Adapters De-Mystified James Schiemer

Illumina TruSeq DNA Adapters De-Mystified James Schiemer 1 of 5 Illumina TruSeq DNA Adapters De-Mystified James Schiemer The key to sequencing random fragments of DNA is by the addition of short nucleotide sequences which allow any DNA fragment to: 1) Bind to

More information

MiSeq: Imaging and Base Calling

MiSeq: Imaging and Base Calling MiSeq: Imaging and Page Welcome Navigation Presenter Introduction MiSeq Sequencing Workflow Narration Welcome to MiSeq: Imaging and. This course takes 35 minutes to complete. Click Next to continue. Please

More information

Technical Note. Roche Applied Science. No. LC 18/2004. Assay Formats for Use in Real-Time PCR

Technical Note. Roche Applied Science. No. LC 18/2004. Assay Formats for Use in Real-Time PCR Roche Applied Science Technical Note No. LC 18/2004 Purpose of this Note Assay Formats for Use in Real-Time PCR The LightCycler Instrument uses several detection channels to monitor the amplification of

More information

Introduction to transcriptome analysis using High Throughput Sequencing technologies (HTS)

Introduction to transcriptome analysis using High Throughput Sequencing technologies (HTS) Introduction to transcriptome analysis using High Throughput Sequencing technologies (HTS) A typical RNA Seq experiment Library construction Protocol variations Fragmentation methods RNA: nebulization,

More information

Core Facility Genomics

Core Facility Genomics Core Facility Genomics versatile genome or transcriptome analyses based on quantifiable highthroughput data ascertainment 1 Topics Collaboration with Harald Binder and Clemens Kreutz Project: Microarray

More information

PreciseTM Whitepaper

PreciseTM Whitepaper Precise TM Whitepaper Introduction LIMITATIONS OF EXISTING RNA-SEQ METHODS Correctly designed gene expression studies require large numbers of samples, accurate results and low analysis costs. Analysis

More information

Deep Sequencing Data Analysis

Deep Sequencing Data Analysis Deep Sequencing Data Analysis Ross Whetten Professor Forestry & Environmental Resources Background Who am I, and why am I teaching this topic? I am not an expert in bioinformatics I started as a biologist

More information

Data Processing of Nextera Mate Pair Reads on Illumina Sequencing Platforms

Data Processing of Nextera Mate Pair Reads on Illumina Sequencing Platforms Data Processing of Nextera Mate Pair Reads on Illumina Sequencing Platforms Introduction Mate pair sequencing enables the generation of libraries with insert sizes in the range of several kilobases (Kb).

More information

Analysis of gene expression data. Ulf Leser and Philippe Thomas

Analysis of gene expression data. Ulf Leser and Philippe Thomas Analysis of gene expression data Ulf Leser and Philippe Thomas This Lecture Protein synthesis Microarray Idea Technologies Applications Problems Quality control Normalization Analysis next week! Ulf Leser:

More information

Introduction to NGS data analysis

Introduction to NGS data analysis Introduction to NGS data analysis Jeroen F. J. Laros Leiden Genome Technology Center Department of Human Genetics Center for Human and Clinical Genetics Sequencing Illumina platforms Characteristics: High

More information

Go where the biology takes you. Genome Analyzer IIx Genome Analyzer IIe

Go where the biology takes you. Genome Analyzer IIx Genome Analyzer IIe Go where the biology takes you. Genome Analyzer IIx Genome Analyzer IIe Go where the biology takes you. To published results faster With proven scalability To the forefront of discovery To limitless applications

More information

The Power of Next-Generation Sequencing in Your Hands On the Path towards Diagnostics

The Power of Next-Generation Sequencing in Your Hands On the Path towards Diagnostics The Power of Next-Generation Sequencing in Your Hands On the Path towards Diagnostics The GS Junior System The Power of Next-Generation Sequencing on Your Benchtop Proven technology: Uses the same long

More information

DNA Sequence Analysis

DNA Sequence Analysis DNA Sequence Analysis Two general kinds of analysis Screen for one of a set of known sequences Determine the sequence even if it is novel Screening for a known sequence usually involves an oligonucleotide

More information

Next Generation Sequencing

Next Generation Sequencing Next Generation Sequencing DNA sequence represents a single format onto which a broad range of biological phenomena can be projected for high-throughput data collection Over the past three years, massively

More information

1. Molecular computation uses molecules to represent information and molecular processes to implement information processing.

1. Molecular computation uses molecules to represent information and molecular processes to implement information processing. Chapter IV Molecular Computation These lecture notes are exclusively for the use of students in Prof. MacLennan s Unconventional Computation course. c 2013, B. J. MacLennan, EECS, University of Tennessee,

More information

How Sequencing Experiments Fail

How Sequencing Experiments Fail How Sequencing Experiments Fail v1.0 Simon Andrews simon.andrews@babraham.ac.uk Classes of Failure Technical Tracking Library Contamination Biological Interpretation Something went wrong with a machine

More information

Cluster Generation. Module 2: Overview

Cluster Generation. Module 2: Overview Cluster Generation Module 2: Overview Sequencing Workflow Sample Preparation Cluster Generation Sequencing Data Analysis 2 Cluster Generation 3 5 DNA (0.1-5.0 μg) Library preparation Single Cluster molecule

More information

How many of you have checked out the web site on protein-dna interactions?

How many of you have checked out the web site on protein-dna interactions? How many of you have checked out the web site on protein-dna interactions? Example of an approximately 40,000 probe spotted oligo microarray with enlarged inset to show detail. Find and be ready to discuss

More information

FOR REFERENCE PURPOSES

FOR REFERENCE PURPOSES BIOO LIFE SCIENCE PRODUCTS FOR REFERENCE PURPOSES This manual is for Reference Purposes Only. DO NOT use this protocol to run your assays. Periodically, optimizations and revisions are made to the kit

More information

Tutorial for Windows and Macintosh. Preparing Your Data for NGS Alignment

Tutorial for Windows and Macintosh. Preparing Your Data for NGS Alignment Tutorial for Windows and Macintosh Preparing Your Data for NGS Alignment 2015 Gene Codes Corporation Gene Codes Corporation 775 Technology Drive, Ann Arbor, MI 48108 USA 1.800.497.4939 (USA) 1.734.769.7249

More information

Next Generation Sequencing

Next Generation Sequencing Next Generation Sequencing Technology and applications 10/1/2015 Jeroen Van Houdt - Genomics Core - KU Leuven - UZ Leuven 1 Landmarks in DNA sequencing 1953 Discovery of DNA double helix structure 1977

More information

Analysis of DNA methylation: bisulfite libraries and SOLiD sequencing

Analysis of DNA methylation: bisulfite libraries and SOLiD sequencing Analysis of DNA methylation: bisulfite libraries and SOLiD sequencing An easy view of the bisulfite approach CH3 genome TAGTACGTTGAT TAGTACGTTGAT read TAGTACGTTGAT TAGTATGTTGAT Three main problems 1.

More information

July 7th 2009 DNA sequencing

July 7th 2009 DNA sequencing July 7th 2009 DNA sequencing Overview Sequencing technologies Sequencing strategies Sample preparation Sequencing instruments at MPI EVA 2 x 5 x ABI 3730/3730xl 454 FLX Titanium Illumina Genome Analyzer

More information

Data Analysis for Ion Torrent Sequencing

Data Analysis for Ion Torrent Sequencing IFU022 v140202 Research Use Only Instructions For Use Part III Data Analysis for Ion Torrent Sequencing MANUFACTURER: Multiplicom N.V. Galileilaan 18 2845 Niel Belgium Revision date: August 21, 2014 Page

More information

Single-Cell DNA Sequencing with the C 1. Single-Cell Auto Prep System. Reveal hidden populations and genetic diversity within complex samples

Single-Cell DNA Sequencing with the C 1. Single-Cell Auto Prep System. Reveal hidden populations and genetic diversity within complex samples DATA Sheet Single-Cell DNA Sequencing with the C 1 Single-Cell Auto Prep System Reveal hidden populations and genetic diversity within complex samples Single-cell sensitivity Discover and detect SNPs,

More information

Speed Matters - Fast ways from template to result

Speed Matters - Fast ways from template to result qpcr Symposium 2007 - Weihenstephan Speed Matters - Fast ways from template to result March 28, 2007 Dr. Thorsten Traeger Senior Scientist, Research and Development - 1 - Overview Ạgenda Fast PCR The Challenges

More information

An Introduction to Next-Generation Sequencing for in vitro Fertilization

An Introduction to Next-Generation Sequencing for in vitro Fertilization An Introduction to Next-Generation Sequencing for in vitro Fertilization www.illumina.com/ivfprimer Table of Contents Part I. Welcome to Next-Generation Sequencing 3 NGS for in vitro Fertilization 3 Part

More information

Introduction To Real Time Quantitative PCR (qpcr)

Introduction To Real Time Quantitative PCR (qpcr) Introduction To Real Time Quantitative PCR (qpcr) SABiosciences, A QIAGEN Company www.sabiosciences.com The Seminar Topics The advantages of qpcr versus conventional PCR Work flow & applications Factors

More information

Illumina GAIIx Sequencing Service

Illumina GAIIx Sequencing Service Illumina GAIIx Sequencing Service As researchers continue to develop novel applications for next generation sequencers, the technology landscape of the industry continues to advance at an unprecedented

More information

Next generation DNA sequencing technologies. theory & prac-ce

Next generation DNA sequencing technologies. theory & prac-ce Next generation DNA sequencing technologies theory & prac-ce Outline Next- Genera-on sequencing (NGS) technologies overview NGS applica-ons NGS workflow: data collec-on and processing the exome sequencing

More information

Bioruptor NGS: Unbiased DNA shearing for Next-Generation Sequencing

Bioruptor NGS: Unbiased DNA shearing for Next-Generation Sequencing STGAAC STGAACT GTGCACT GTGAACT STGAAC STGAACT GTGCACT GTGAACT STGAAC STGAAC GTGCAC GTGAAC Wouter Coppieters Head of the genomics core facility GIGA center, University of Liège Bioruptor NGS: Unbiased DNA

More information

Genetic Analysis. Phenotype analysis: biological-biochemical analysis. Genotype analysis: molecular and physical analysis

Genetic Analysis. Phenotype analysis: biological-biochemical analysis. Genotype analysis: molecular and physical analysis Genetic Analysis Phenotype analysis: biological-biochemical analysis Behaviour under specific environmental conditions Behaviour of specific genetic configurations Behaviour of progeny in crosses - Genotype

More information

Welcome to Pacific Biosciences' Introduction to SMRTbell Template Preparation.

Welcome to Pacific Biosciences' Introduction to SMRTbell Template Preparation. Introduction to SMRTbell Template Preparation 100 338 500 01 1. SMRTbell Template Preparation 1.1 Introduction to SMRTbell Template Preparation Welcome to Pacific Biosciences' Introduction to SMRTbell

More information

Advances in RainDance Sequence Enrichment Technology and Applications in Cancer Research. March 17, 2011 Rendez-Vous Séquençage

Advances in RainDance Sequence Enrichment Technology and Applications in Cancer Research. March 17, 2011 Rendez-Vous Séquençage Advances in RainDance Sequence Enrichment Technology and Applications in Cancer Research March 17, 2011 Rendez-Vous Séquençage Presentation Overview Core Technology Review Sequence Enrichment Application

More information

SeqScape Software Version 2.5 Comprehensive Analysis Solution for Resequencing Applications

SeqScape Software Version 2.5 Comprehensive Analysis Solution for Resequencing Applications Product Bulletin Sequencing Software SeqScape Software Version 2.5 Comprehensive Analysis Solution for Resequencing Applications Comprehensive reference sequence handling Helps interpret the role of each

More information

New generation sequencing: current limits and future perspectives. Giorgio Valle CRIBI - Università di Padova

New generation sequencing: current limits and future perspectives. Giorgio Valle CRIBI - Università di Padova New generation sequencing: current limits and future perspectives Giorgio Valle CRIBI Università di Padova Around 2004 the Race for the 1000$ Genome started A few questions... When? How? Why? Standard

More information

A guide to the analysis of KASP genotyping data using cluster plots

A guide to the analysis of KASP genotyping data using cluster plots extraction sequencing genotyping extraction sequencing genotyping extraction sequencing genotyping extraction sequencing A guide to the analysis of KASP genotyping data using cluster plots Contents of

More information

Complete Genomics Sequencing

Complete Genomics Sequencing TECHNOLOGY OVERVIEW Complete Genomics Sequencing Introduction With advances in nanotechnology, high-throughput instruments, and large-scale computing, it has become possible to sequence a complete human

More information

The Techniques of Molecular Biology: Forensic DNA Fingerprinting

The Techniques of Molecular Biology: Forensic DNA Fingerprinting Revised Fall 2011 The Techniques of Molecular Biology: Forensic DNA Fingerprinting The techniques of molecular biology are used to manipulate the structure and function of molecules such as DNA and proteins

More information

A greedy algorithm for the DNA sequencing by hybridization with positive and negative errors and information about repetitions

A greedy algorithm for the DNA sequencing by hybridization with positive and negative errors and information about repetitions BULLETIN OF THE POLISH ACADEMY OF SCIENCES TECHNICAL SCIENCES, Vol. 59, No. 1, 2011 DOI: 10.2478/v10175-011-0015-0 Varia A greedy algorithm for the DNA sequencing by hybridization with positive and negative

More information

Sanger Sequencing and Quality Assurance. Zbigniew Rudzki Department of Pathology University of Melbourne

Sanger Sequencing and Quality Assurance. Zbigniew Rudzki Department of Pathology University of Melbourne Sanger Sequencing and Quality Assurance Zbigniew Rudzki Department of Pathology University of Melbourne Sanger DNA sequencing The era of DNA sequencing essentially started with the publication of the enzymatic

More information

How is genome sequencing done?

How is genome sequencing done? How is genome sequencing done? Using 454 Sequencing on the Genome Sequencer FLX System, DNA from a genome is converted into sequence data through four primary steps: Step One DNA sample preparation; Step

More information

An example of bioinformatics application on plant breeding projects in Rijk Zwaan

An example of bioinformatics application on plant breeding projects in Rijk Zwaan An example of bioinformatics application on plant breeding projects in Rijk Zwaan Xiangyu Rao 17-08-2012 Introduction of RZ Rijk Zwaan is active worldwide as a vegetable breeding company that focuses on

More information

DNA Sequencing. Ben Langmead. Department of Computer Science

DNA Sequencing. Ben Langmead. Department of Computer Science DN Sequencing Ben Langmead Department of omputer Science You are free to use these slides. If you do, please sign the guestbook (www.langmead-lab.org/teaching-materials), or email me (ben.langmead@gmail.com)

More information

HiPer RT-PCR Teaching Kit

HiPer RT-PCR Teaching Kit HiPer RT-PCR Teaching Kit Product Code: HTBM024 Number of experiments that can be performed: 5 Duration of Experiment: Protocol: 4 hours Agarose Gel Electrophoresis: 45 minutes Storage Instructions: The

More information

An Overview of DNA Sequencing

An Overview of DNA Sequencing An Overview of DNA Sequencing Prokaryotic DNA Plasmid http://en.wikipedia.org/wiki/image:prokaryote_cell_diagram.svg Eukaryotic DNA http://en.wikipedia.org/wiki/image:plant_cell_structure_svg.svg DNA Structure

More information

Application Guide... 2

Application Guide... 2 Protocol for GenomePlex Whole Genome Amplification from Formalin-Fixed Parrafin-Embedded (FFPE) tissue Application Guide... 2 I. Description... 2 II. Product Components... 2 III. Materials to be Supplied

More information

Innovations in Molecular Epidemiology

Innovations in Molecular Epidemiology Innovations in Molecular Epidemiology Molecular Epidemiology Measure current rates of active transmission Determine whether recurrent tuberculosis is attributable to exogenous reinfection Determine whether

More information

Nazneen Aziz, PhD. Director, Molecular Medicine Transformation Program Office

Nazneen Aziz, PhD. Director, Molecular Medicine Transformation Program Office 2013 Laboratory Accreditation Program Audioconferences and Webinars Implementing Next Generation Sequencing (NGS) as a Clinical Tool in the Laboratory Nazneen Aziz, PhD Director, Molecular Medicine Transformation

More information

Chapter 8: Recombinant DNA 2002 by W. H. Freeman and Company Chapter 8: Recombinant DNA 2002 by W. H. Freeman and Company

Chapter 8: Recombinant DNA 2002 by W. H. Freeman and Company Chapter 8: Recombinant DNA 2002 by W. H. Freeman and Company Genetic engineering: humans Gene replacement therapy or gene therapy Many technical and ethical issues implications for gene pool for germ-line gene therapy what traits constitute disease rather than just

More information

Next Generation Sequencing for DUMMIES

Next Generation Sequencing for DUMMIES Next Generation Sequencing for DUMMIES Looking at a presentation without the explanation from the author is sometimes difficult to understand. This document contains extra information for some slides that

More information

PrimeSTAR HS DNA Polymerase

PrimeSTAR HS DNA Polymerase Cat. # R010A For Research Use PrimeSTAR HS DNA Polymerase Product Manual Table of Contents I. Description...3 II. III. IV. Components...3 Storage...3 Features...3 V. General Composition of PCR Reaction

More information

NGS data analysis. Bernardo J. Clavijo

NGS data analysis. Bernardo J. Clavijo NGS data analysis Bernardo J. Clavijo 1 A brief history of DNA sequencing 1953 double helix structure, Watson & Crick! 1977 rapid DNA sequencing, Sanger! 1977 first full (5k) genome bacteriophage Phi X!

More information

Forensic DNA Testing Terminology

Forensic DNA Testing Terminology Forensic DNA Testing Terminology ABI 310 Genetic Analyzer a capillary electrophoresis instrument used by forensic DNA laboratories to separate short tandem repeat (STR) loci on the basis of their size.

More information

Gene Mapping Techniques

Gene Mapping Techniques Gene Mapping Techniques OBJECTIVES By the end of this session the student should be able to: Define genetic linkage and recombinant frequency State how genetic distance may be estimated State how restriction

More information

Standards, Guidelines and Best Practices for RNA-Seq V1.0 (June 2011) The ENCODE Consortium

Standards, Guidelines and Best Practices for RNA-Seq V1.0 (June 2011) The ENCODE Consortium Standards, Guidelines and Best Practices for RNA-Seq V1.0 (June 2011) The ENCODE Consortium I. Introduction: Sequence based assays of transcriptomes (RNA-seq) are in wide use because of their favorable

More information

SNP genotyping. Gene expression. And now Solexa sequencing.

SNP genotyping. Gene expression. And now Solexa sequencing. SNP genotyping. Gene expression. And now Solexa sequencing. Let s find the answers together. It s your research. You question. You test. You want answers quickly, accurately, and at a good value. Illumina

More information

Lecture 13: DNA Technology. DNA Sequencing. DNA Sequencing Genetic Markers - RFLPs polymerase chain reaction (PCR) products of biotechnology

Lecture 13: DNA Technology. DNA Sequencing. DNA Sequencing Genetic Markers - RFLPs polymerase chain reaction (PCR) products of biotechnology Lecture 13: DNA Technology DNA Sequencing Genetic Markers - RFLPs polymerase chain reaction (PCR) products of biotechnology DNA Sequencing determine order of nucleotides in a strand of DNA > bases = A,

More information

Introduction Bioo Scientific

Introduction Bioo Scientific Next Generation Sequencing Catalog 2014-2015 Introduction Bioo Scientific Bioo Scientific is a global life science company headquartered in Austin, TX, committed to providing innovative products and superior

More information

Computational Genomics. Next generation sequencing (NGS)

Computational Genomics. Next generation sequencing (NGS) Computational Genomics Next generation sequencing (NGS) Sequencing technology defies Moore s law Nature Methods 2011 Log 10 (price) Sequencing the Human Genome 2001: Human Genome Project 2.7G$, 11 years

More information

4.2.1. What is a contig? 4.2.2. What are the contig assembly programs?

4.2.1. What is a contig? 4.2.2. What are the contig assembly programs? Table of Contents 4.1. DNA Sequencing 4.1.1. Trace Viewer in GCG SeqLab Table. Box. Select the editor mode in the SeqLab main window. Import sequencer trace files from the File menu. Select the trace files

More information

Concepts and methods in sequencing and genome assembly

Concepts and methods in sequencing and genome assembly BCM-2004 Concepts and methods in sequencing and genome assembly B. Franz LANG, Département de Biochimie Bureau: H307-15 Courrier électronique: Franz.Lang@Umontreal.ca Outline 1. Concepts in DNA and RNA

More information

UGENE Quick Start Guide

UGENE Quick Start Guide Quick Start Guide This document contains a quick introduction to UGENE. For more detailed information, you can find the UGENE User Manual and other special manuals in project website: http://ugene.unipro.ru.

More information

Troubleshooting for PCR and multiplex PCR

Troubleshooting for PCR and multiplex PCR Page 1 of 5 Page designed and maintained by Octavian Henegariu (Email: Tavi's Yale email or Tavi's Yahoo email). As I am currently pursuing a new junior faculty position, the Yale URL and email may change

More information

Recombinant DNA & Genetic Engineering. Tools for Genetic Manipulation

Recombinant DNA & Genetic Engineering. Tools for Genetic Manipulation Recombinant DNA & Genetic Engineering g Genetic Manipulation: Tools Kathleen Hill Associate Professor Department of Biology The University of Western Ontario Tools for Genetic Manipulation DNA, RNA, cdna

More information

biology Genotyping-by-Sequencing in Plants Biology 2012, 1, 460-483; doi:10.3390/biology1030460 ISSN 2079-7737 www.mdpi.com/journal/biology Review

biology Genotyping-by-Sequencing in Plants Biology 2012, 1, 460-483; doi:10.3390/biology1030460 ISSN 2079-7737 www.mdpi.com/journal/biology Review Biology 2012, 1, 460-483; doi:10.3390/biology1030460 Review OPEN ACCESS biology ISSN 2079-7737 www.mdpi.com/journal/biology Genotyping-by-Sequencing in Plants Stéphane Deschamps 1, *, Victor Llaca 1 and

More information

Systematic discovery of regulatory motifs in human promoters and 30 UTRs by comparison of several mammals

Systematic discovery of regulatory motifs in human promoters and 30 UTRs by comparison of several mammals Systematic discovery of regulatory motifs in human promoters and 30 UTRs by comparison of several mammals Xiaohui Xie 1, Jun Lu 1, E. J. Kulbokas 1, Todd R. Golub 1, Vamsi Mootha 1, Kerstin Lindblad-Toh

More information

Real-time PCR: Understanding C t

Real-time PCR: Understanding C t APPLICATION NOTE Real-Time PCR Real-time PCR: Understanding C t Real-time PCR, also called quantitative PCR or qpcr, can provide a simple and elegant method for determining the amount of a target sequence

More information

Single Nucleotide Polymorphisms (SNPs)

Single Nucleotide Polymorphisms (SNPs) Single Nucleotide Polymorphisms (SNPs) Additional Markers 13 core STR loci Obtain further information from additional markers: Y STRs Separating male samples Mitochondrial DNA Working with extremely degraded

More information

Single-Cell Whole Genome Sequencing on the C1 System: a Performance Evaluation

Single-Cell Whole Genome Sequencing on the C1 System: a Performance Evaluation PN 100-9879 A1 TECHNICAL NOTE Single-Cell Whole Genome Sequencing on the C1 System: a Performance Evaluation Introduction Cancer is a dynamic evolutionary process of which intratumor genetic and phenotypic

More information

LightCycler 480 Real-Time PCR System

LightCycler 480 Real-Time PCR System Roche Applied Science LightCycler 480 Real-Time PCR System Planned introduction of the LightCycler 480 System: September 2005 Rapid by nature, accurate by design The LightCycler 480 Real-Time PCR System

More information

GenScript BloodReady TM Multiplex PCR System

GenScript BloodReady TM Multiplex PCR System GenScript BloodReady TM Multiplex PCR System Technical Manual No. 0174 Version 20040915 I Description.. 1 II Applications 2 III Key Features.. 2 IV Shipping and Storage. 2 V Simplified Procedures. 2 VI

More information

TCRG TCRA/D IGH IGK/L

TCRG TCRA/D IGH IGK/L Assays immunoseq Assay The inquiry to insight solution for profiling T- and B-cell s Immunosequencing solutions for multiple species and loci Illuminate the adaptive immune system with bias-controlled

More information

All your base(s) are belong to us

All your base(s) are belong to us All your base(s) are belong to us The dawn of the high-throughput DNA sequencing era 25C3 Magnus Manske The place Sanger Center, Cambridge, UK Basic biology Level of complexity Genome Single (all chromosomes

More information

HENIPAVIRUS ANTIBODY ESCAPE SEQUENCING REPORT

HENIPAVIRUS ANTIBODY ESCAPE SEQUENCING REPORT HENIPAVIRUS ANTIBODY ESCAPE SEQUENCING REPORT Kimberly Bishop Lilly 1,2, Truong Luu 1,2, Regina Cer 1,2, and LT Vishwesh Mokashi 1 1 Naval Medical Research Center, NMRC Frederick, 8400 Research Plaza,

More information

2. True or False? The sequence of nucleotides in the human genome is 90.9% identical from one person to the next. False (it s 99.

2. True or False? The sequence of nucleotides in the human genome is 90.9% identical from one person to the next. False (it s 99. 1. True or False? A typical chromosome can contain several hundred to several thousand genes, arranged in linear order along the DNA molecule present in the chromosome. True 2. True or False? The sequence

More information

Focusing on results not data comprehensive data analysis for targeted next generation sequencing

Focusing on results not data comprehensive data analysis for targeted next generation sequencing Focusing on results not data comprehensive data analysis for targeted next generation sequencing Daniel Swan, Jolyon Holdstock, Angela Matchan, Richard Stark, John Shovelton, Duarte Mohla and Simon Hughes

More information

High Performance Compu2ng Facility

High Performance Compu2ng Facility High Performance Compu2ng Facility Center for Health Informa2cs and Bioinforma2cs Accelera2ng Scien2fic Discovery and Innova2on in Biomedical Research at NYULMC through Advanced Compu2ng Efstra'os Efstathiadis,

More information

Lectures 1 and 8 15. February 7, 2013. Genomics 2012: Repetitorium. Peter N Robinson. VL1: Next- Generation Sequencing. VL8 9: Variant Calling

Lectures 1 and 8 15. February 7, 2013. Genomics 2012: Repetitorium. Peter N Robinson. VL1: Next- Generation Sequencing. VL8 9: Variant Calling Lectures 1 and 8 15 February 7, 2013 This is a review of the material from lectures 1 and 8 14. Note that the material from lecture 15 is not relevant for the final exam. Today we will go over the material

More information

Notice. DNA Sequencing Module User Guide

Notice. DNA Sequencing Module User Guide GenomeStudio TM DNA Sequencing Module v1.0 User Guide An Integrated Platform for Data Visualization and Analysis FOR RESEARCH ONLY DS ILLUMINA PROPRIETARY Part # 11319092, Rev. A Notice This publication

More information

Biotechnology: DNA Technology & Genomics

Biotechnology: DNA Technology & Genomics Chapter 20. Biotechnology: DNA Technology & Genomics 2003-2004 The BIG Questions How can we use our knowledge of DNA to: diagnose disease or defect? cure disease or defect? change/improve organisms? What

More information

G E N OM I C S S E RV I C ES

G E N OM I C S S E RV I C ES GENOMICS SERVICES THE NEW YORK GENOME CENTER NYGC is an independent non-profit implementing advanced genomic research to improve diagnosis and treatment of serious diseases. capabilities. N E X T- G E

More information

Chapter 6 DNA Replication

Chapter 6 DNA Replication Chapter 6 DNA Replication Each strand of the DNA double helix contains a sequence of nucleotides that is exactly complementary to the nucleotide sequence of its partner strand. Each strand can therefore

More information

Reading DNA Sequences:

Reading DNA Sequences: Reading DNA Sequences: 18-th Century Mathematics for 21-st Century Technology Michael Waterman University of Southern California Tsinghua University DNA Genetic information of an organism Double helix,

More information

The RNAi Consortium (TRC) Broad Institute

The RNAi Consortium (TRC) Broad Institute TRC Laboratory Protocols Protocol Title: One Step PCR Preparation of Samples for Illumina Sequencing Current Revision Date: 11/10/2012 RNAi Platform,, trc_info@broadinstitute.org Brief Description: This

More information

Removing Sequential Bottlenecks in Analysis of Next-Generation Sequencing Data

Removing Sequential Bottlenecks in Analysis of Next-Generation Sequencing Data Removing Sequential Bottlenecks in Analysis of Next-Generation Sequencing Data Yi Wang, Gagan Agrawal, Gulcin Ozer and Kun Huang The Ohio State University HiCOMB 2014 May 19 th, Phoenix, Arizona 1 Outline

More information

PolyLens: Software for Map-based Visualization and Analysis of Genome-scale Polymorphism Data

PolyLens: Software for Map-based Visualization and Analysis of Genome-scale Polymorphism Data PolyLens: Software for Map-based Visualization and Analysis of Genome-scale Polymorphism Data Ryhan Pathan Department of Electrical Engineering and Computer Science University of Tennessee Knoxville Knoxville,

More information

A Primer of Genome Science THIRD

A Primer of Genome Science THIRD A Primer of Genome Science THIRD EDITION GREG GIBSON-SPENCER V. MUSE North Carolina State University Sinauer Associates, Inc. Publishers Sunderland, Massachusetts USA Contents Preface xi 1 Genome Projects:

More information

Beginner s Guide to Real-Time PCR

Beginner s Guide to Real-Time PCR Beginner s Guide to Real-Time PCR 02 Real-time PCR basic principles PCR or the Polymerase Chain Reaction has become the cornerstone of modern molecular biology the world over. Real-time PCR is an advanced

More information

Whole genome Bisulfite Sequencing for Methylation Analysis Preparing Samples for the Illumina Sequencing Platform

Whole genome Bisulfite Sequencing for Methylation Analysis Preparing Samples for the Illumina Sequencing Platform Whole genome Bisulfite Sequencing for Methylation Analysis Preparing Samples for the Illumina Sequencing Platform Introduction, 2 Sample Prep Workflow, 3 Best Practices, 4 DNA Input Recommendations, 6

More information

Real-Time PCR Vs. Traditional PCR

Real-Time PCR Vs. Traditional PCR Real-Time PCR Vs. Traditional PCR Description This tutorial will discuss the evolution of traditional PCR methods towards the use of Real-Time chemistry and instrumentation for accurate quantitation. Objectives

More information

Molecular and Cell Biology Laboratory (BIOL-UA 223) Instructor: Ignatius Tan Phone: 212-998-8295 Office: 764 Brown Email: ignatius.tan@nyu.

Molecular and Cell Biology Laboratory (BIOL-UA 223) Instructor: Ignatius Tan Phone: 212-998-8295 Office: 764 Brown Email: ignatius.tan@nyu. Molecular and Cell Biology Laboratory (BIOL-UA 223) Instructor: Ignatius Tan Phone: 212-998-8295 Office: 764 Brown Email: ignatius.tan@nyu.edu Course Hours: Section 1: Mon: 12:30-3:15 Section 2: Wed: 12:30-3:15

More information

Description: Molecular Biology Services and DNA Sequencing

Description: Molecular Biology Services and DNA Sequencing Description: Molecular Biology s and DNA Sequencing DNA Sequencing s Single Pass Sequencing Sequence data only, for plasmids or PCR products Plasmid DNA or PCR products Plasmid DNA: 20 100 ng/μl PCR Product:

More information

BacReady TM Multiplex PCR System

BacReady TM Multiplex PCR System BacReady TM Multiplex PCR System Technical Manual No. 0191 Version 10112010 I Description.. 1 II Applications 2 III Key Features.. 2 IV Shipping and Storage. 2 V Simplified Procedures. 2 VI Detailed Experimental

More information

Just the Facts: A Basic Introduction to the Science Underlying NCBI Resources

Just the Facts: A Basic Introduction to the Science Underlying NCBI Resources 1 of 8 11/7/2004 11:00 AM National Center for Biotechnology Information About NCBI NCBI at a Glance A Science Primer Human Genome Resources Model Organisms Guide Outreach and Education Databases and Tools

More information