SlideShare une entreprise Scribd logo
1  sur  59
Télécharger pour lire hors ligne
Surya Saha, Ph.D.
Cornell University & Boyce Thompson Institute
suryasaha@cornell.edu @SahaSurya
Centre for Agricultural Bioinformatics
Pusa, New Delhi
June 13,2014
Slides: http://bit.ly/CABin_Pusa_2014
http://www.acgt.me/blog/2014/3/7/next-generation-sequencing-must-die
Genome Assembly
Jason Chin http://www.bit.ly/SZPKIG
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 2
You are free to:
Copy, share, adapt, or re-mix;
Photograph, film, or broadcast;
Blog, live-blog, or post video of;
This presentation. Provided that:
You attribute the work to its author and respect the rights
and licenses associated with its components.
Slide Concept by Cameron Neylon, who has waived all copyright and related or neighbouring rights. This slide only ccZero. Social Media Icons adapted with
permission from originals by Christopher Ross. Original images are available under GPL at
http://www.thisismyurl.com/free-downloads/15-free-speech-bubble-icons-for-popular-websites
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 3
Sequencing
1953
DNA Structure
discovery
1977
2012
Sanger DNA sequencing by
chain-terminating inhibitors
1984
Epstein-Barr
virus
(170 Kb)
1987Abi370
Sequencer
1995
2001
Homo
sapiens
(3.0 Gb)
2005
454
Solexa
Solid
2007
2011
Ion
Torrent
PacBio
Haemophilus
influenzae
(1.83 Mb)
2013
Slide credit: Aureliano Bombarely
Sequencing over the Ages
Illumina
Illumina
Hiseq X
454
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 4
Pinus
taeda
(24 Gb)
2014
MinION
The Next Generation
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 5
Its all about the $£€¥
http://www.genome.gov/sequencingcosts/
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 6
First generation sequencing
Sanger method
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 7
Frederick Sanger
13 Aug 1918 – 19 Nov 2013
Won the Nobel Prize for Chemistry in 1958 and
1980. Published the dideoxy chain termination
method or “Sanger method” in 1977
http://dailym.ai/1f1XeTB
Sanger method
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 8
http://bit.ly/1g6Cudq
http://bit.ly/1lcQO4J
First generation sequencing
• Very high quality sequences (99.999%)
• Very low throughput
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 9
Run Time Read Length Reads / Run
Total
nucleotides
sequenced
Cost / MB
Capillary
Sequencing
(ABI3730xl)
20m-3h 400-900 bp 96 or 386 1.9-84 Kb $2400
http://bit.ly/1clLps3
http://1.usa.gov/1cLqIRd
Use the specific technology used
to generate the data
– Illumina Hiseq/Miseq/NextSeq
– Pacific Biosciences RS I/RS II
– Ion Torrent Proton/PGM
– SOLiD
– 454
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 10
http://www.acgt.me/blog/2014/3/10/next-generation-
sequencing-must-diepart-2
454 Pyrosequencing
One purified DNA
fragment, to one bead, to
one read.
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 11
http://bit.ly/1ehwxWN
GS FLX
Titanium
http://bit.ly/1ehAcEh
Illumina
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 12
Output 15 Gb 120 GB 1000 GB 1800 GB
Number
of Reads
25 Million 400 Million 4 Billion 6 Billion
Read
Length
2x300 bp 2x150 bp 2x125 bp
(2x250 update mid-2014)
2x150 bp
Cost $99K $250K $740K $10M
Source: Illumina
$1000 human
genome??
Illumina
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 13
http://1.usa.gov/1fP9ybl
Illumina:Moleculo
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 14
http://bit.ly/1aEPOBn
Pacific Biosciences SMRT sequencing
Single Molecule Real
Time sequencing
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 15
http://bit.ly/1naxgTe
Pacific Biosciences SMRT sequencing
Error correction methods
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 16
Hierarchical genome-assembly
process (HGAP)
PBJelly
Enlish et al., PLOS One. 2012
PBJelly
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 17
Pacific Biosciences SMRT sequencing
Read Lengths
Oxford Nanopore
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 18
https://www.nanoporetech.com/
• No data yet??
• Error model
http://erlichya.tumblr.com/post/66376172948/hands-on-
experience-with-oxford-nanopore-minion
Others
• Ion Torrent Proton/PGM
• Nabsys
• SOLiD
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 19
Comparison
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 20
Next generation sequencing
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 21
Run Time Read Length Quality
Total
nucleotides
sequenced
Cost /MB
454
Pyrosequencing
24h 700 bp Q20-Q30 0.7 GB $10
Illumina Miseq 27h 2x250bp > Q30 15 GB $0.15
Illumina Hiseq
2500
11days 2x125bp >Q30 1000 GB $0.05
Ion torrent 2h 400bp >Q20 50MB-1GB $1
Pacific
Biosciences
2h 10-20kb
>Q30 consensus
>Q10 single
400-800MB
/SMRT cell
$0.33-$1
http://bit.ly/1clLps3
http://1.usa.gov/1cLqIRd
http://omicsmaps.com/
Next Generation Genomics:
World Map of High-throughput Sequencers
Centre for Agricultural Bioinformatics, Pusa6/15/2014 22
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 23
http://bit.ly/18pfUId
Real cost of Sequencing!!
Sboner, Genome Biology, 2011
6/15/2014 24Centre for Agricultural Bioinformatics, Pusa
Library Types
Single end
Pair end (PE, 150-800 bp, Fwd:/1, Rev:/2)
Mate pair (MP, 2Kb to 20 Kb)
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 25
F
F R
F R 454/Roche
FR Illumina
Illumina
Slide credit: Aureliano Bombarely
Implications of Choice of Library
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 26
Slide credit: Aureliano Bombarely
Consensus sequence
(Contig)
Reads
Scaffold
(or Supercontig)
Pair Read information
NNNNN
Pseudomolecule
(or ultracontig)
F
Genetic information (markers)
NNNNN NN
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 27
Quality control: Encoding
http://bit.ly/N28yUd
Phred score of a base is:
Qphred = -10 log10 (e)
where e is the estimated probability of a base
being incorrect
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 28
Genome Assembly
Whole Genome Shotgun Sequencing
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 29
Slide credit: cbcb.umd.edu
Genome Sequencing Strategies
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 30
Slide credit: Aureliano Bombarely
Genome Sequencing Strategies
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 31
International Human Genome Sequencing Consortium 2001
Overlap Layout Consensus
http://contig.wordpress.com/
cbcb.umd.edu
DeBruijnGraph
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 32
Ingredient for a Good Assembly
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 33
Slide credit: Mike Schatz
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 34
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 35
Bird Snake
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 36
• You have the expertise to install and run
• You have the suitable infrastructure (CPU & RAM) to run the assembler
• You have sufficient time to run the assembler
• Is designed to work with the specific mix of NGS data that you have
generated
• Best addresses what you want to get out of a genome assembly (bigger
overall assembly, more genes, most accuracy, longer scaffolds, most
resolution of haplotypes, most tolerant of repeats, etc.)
The BEST?? Genome Assembler for YOU
http://haldanessieve.org/2013/01/28/our-paper-making-pizzas-and-genome-assemblies/
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 37
Which technology to use??
• Microbial genomes
• Eukaryotic genomes
• Resequencing genomes
• RNAseq and other XXXseq methods
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 38
http://bit.ly/1ko9Kgh
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 39
SOL Genomics Network
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 40
The SGN Team!!
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 41
Surya Saha, Tom Fisher-York, Hartmut Foerster, Suzy Strickler, Jeremy Edwards,
Noe Fernandez, Naama Menda, Aure Bombarely, Aimin Yan, Isaak Tecle
SGN Website
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 42
http://solgenomics.net
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 43
Main web page (front page):
WEB ICONS
TOOL BAR
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 44
Main web page (front page):
TOOL BAR
(MENUS)
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 45
But the DATA also can be
edited
LocusLocus Editor Data
Community Data Curation
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 46
You need
• SGN account.
• Activate submitter / Locus Editor privileges by SGN curator
LocusLocus Editor Data
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 47
Tools
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 48
Genome Browser: GBrowse
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 49
Genome Browser: JBrowse
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 50
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 51
CassavaBase
http://cassavabase.org/
Slide credit: Jeremy Edwards
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 52
NextGen Cassava Project
● Project: Adapt SGN database for Cassava Breeding
● Goal: Apply Genomic Selection to cassava breeding
● Predict breeding values from genotype information
● Shorten the breeding cycle
● Massive amounts of genotypic data (GBS)
● Phenotypic data
● Data management challenge
● Improve flowering
● http://nextgencassava.org
Slide credit: Jeremy Edwards
SGN/Cassavabase behind the scenes
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 53
● Perl/Catalyst MVC Framework
● PostgreSQL Database
● Generic Model Organism Database (GMOD)
– Chado relational database schema
– GBrowse
– JBrowse
● R
– Experimental design
– QTL mapping
– Genomic selection
Slide credit: Jeremy Edwards
Objectives
Provide cassava breeders and researchers access
to data and tools in a centralized, user-friendly
and reliable database.
– Improve partner breeding program information
tracking
– Streamline management of genotypic and
phenotypic data
– Pipeline genotypic and phenotypic data through
Genomic Selection prediction analyses
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 54
Slide credit: Jeremy Edwards
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 55
Genomic Selection
The 'training population' is genotyped and phenotyped to 'train'
the genomic selection (GS) prediction model. Genotypic
information from the breeding material is then fed into the
model to calculate genomic estimated breeding values (GEBV)
for these lines. From Heffner et al. 2009 Crop Sci. 49:1–12
Information from a majority of lines in the breeding population (the training set) is used to create the
prediction model. The model is then used to predict the phenotypes of the remaining lines (the validation
set), using genotypic information only. The results from the model are compared to the actual data to give
the prediction accuracy. Image courtesy of Martha Hamblin, Cornell University
Flow diagram of a genomic selection breeding program.
Breeding cycle time is shortened by removing phenotypic
evaluation of lines before selection as parents for the next
cycle. From Heffner et al. 2009 Crop Sci. 49:1–12
Slide credit: Jeremy Edwards
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 56
Data collection in the field
● Android tablets
● Field book app
– Jesse Poland's group at
USDA-ARS / Kansas
State University
Slide credit: Jeremy Edwards
Cassava Trait Ontology
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 57
Kulakow et al. 2011
Kulakow et al. 2011
● Standard terminology
● Facilitate the sharing of information
● Allow users to query keywords related to traits
Slide credit: Jeremy Edwards
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 58
Position available at Solgenomics
Cassavabase project
Plant Breeding + Bioinformatician
● Familiar with breeding
● Programming in Perl, R, SQL, Hadoop
● Linux
● Africa
● Genius
http://www.cassavabase.org/forum/posts
.pl?topic_id=9
Thank you!!
Questions??
6/15/2014 Centre for Agricultural Bioinformatics, Pusa 59

Contenu connexe

Tendances

BioSMACK - Linux Live CD for GWAS
BioSMACK - Linux Live CD for GWASBioSMACK - Linux Live CD for GWAS
BioSMACK - Linux Live CD for GWASHong ChangBum
 
Next-generation sequencing from 2005 to 2020
Next-generation sequencing from 2005 to 2020Next-generation sequencing from 2005 to 2020
Next-generation sequencing from 2005 to 2020Christian Frech
 
Partial thesis defence presentation
Partial thesis defence presentationPartial thesis defence presentation
Partial thesis defence presentationSanjeewaRupasinghe
 
H Mishima - Biogem, Ruby UCSC API, and BioRuby
H Mishima - Biogem, Ruby UCSC API, and BioRubyH Mishima - Biogem, Ruby UCSC API, and BioRuby
H Mishima - Biogem, Ruby UCSC API, and BioRubyJan Aerts
 
BioRuby -- Bioinformatics Library
BioRuby -- Bioinformatics LibraryBioRuby -- Bioinformatics Library
BioRuby -- Bioinformatics Libraryngotogenome
 
CSU Next Generation Sequencing Core 06/09/2015
CSU Next Generation Sequencing Core 06/09/2015CSU Next Generation Sequencing Core 06/09/2015
CSU Next Generation Sequencing Core 06/09/2015Richard Casey
 
Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...
Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...
Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...Joe Parker
 
How scientists and artists are discovering musical patterns in protein struct...
How scientists and artists are discovering musical patterns in protein struct...How scientists and artists are discovering musical patterns in protein struct...
How scientists and artists are discovering musical patterns in protein struct...Willard Van De Bogart
 
Browsing Genes, Variation and Regulation data with Ensembl
Browsing Genes, Variation and Regulation data with EnsemblBrowsing Genes, Variation and Regulation data with Ensembl
Browsing Genes, Variation and Regulation data with EnsemblDenise Carvalho-Silva, PhD
 
Variation and the VEP: Ensembl Online Webinar series
Variation and the VEP: Ensembl Online Webinar seriesVariation and the VEP: Ensembl Online Webinar series
Variation and the VEP: Ensembl Online Webinar seriesDenise Carvalho-Silva, PhD
 

Tendances (11)

BioSMACK - Linux Live CD for GWAS
BioSMACK - Linux Live CD for GWASBioSMACK - Linux Live CD for GWAS
BioSMACK - Linux Live CD for GWAS
 
Next-generation sequencing from 2005 to 2020
Next-generation sequencing from 2005 to 2020Next-generation sequencing from 2005 to 2020
Next-generation sequencing from 2005 to 2020
 
Partial thesis defence presentation
Partial thesis defence presentationPartial thesis defence presentation
Partial thesis defence presentation
 
H Mishima - Biogem, Ruby UCSC API, and BioRuby
H Mishima - Biogem, Ruby UCSC API, and BioRubyH Mishima - Biogem, Ruby UCSC API, and BioRuby
H Mishima - Biogem, Ruby UCSC API, and BioRuby
 
BioRuby -- Bioinformatics Library
BioRuby -- Bioinformatics LibraryBioRuby -- Bioinformatics Library
BioRuby -- Bioinformatics Library
 
CSU Next Generation Sequencing Core 06/09/2015
CSU Next Generation Sequencing Core 06/09/2015CSU Next Generation Sequencing Core 06/09/2015
CSU Next Generation Sequencing Core 06/09/2015
 
Bioinformatica 06-10-2011-t2-databases
Bioinformatica 06-10-2011-t2-databasesBioinformatica 06-10-2011-t2-databases
Bioinformatica 06-10-2011-t2-databases
 
Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...
Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...
Single-molecule real-time (SMRT) Nanopore sequencing for Plant Pathology appl...
 
How scientists and artists are discovering musical patterns in protein struct...
How scientists and artists are discovering musical patterns in protein struct...How scientists and artists are discovering musical patterns in protein struct...
How scientists and artists are discovering musical patterns in protein struct...
 
Browsing Genes, Variation and Regulation data with Ensembl
Browsing Genes, Variation and Regulation data with EnsemblBrowsing Genes, Variation and Regulation data with Ensembl
Browsing Genes, Variation and Regulation data with Ensembl
 
Variation and the VEP: Ensembl Online Webinar series
Variation and the VEP: Ensembl Online Webinar seriesVariation and the VEP: Ensembl Online Webinar series
Variation and the VEP: Ensembl Online Webinar series
 

En vedette

A Comparison of NGS Platforms.
A Comparison of NGS Platforms.A Comparison of NGS Platforms.
A Comparison of NGS Platforms.mkim8
 
Ngs microbiome
Ngs microbiomeNgs microbiome
Ngs microbiomejukais
 
NGS - Basic principles and sequencing platforms
NGS - Basic principles and sequencing platformsNGS - Basic principles and sequencing platforms
NGS - Basic principles and sequencing platformsAnnelies Haegeman
 
Molecular characterization of Pst isolates from Western Canada
Molecular characterization of Pst isolates from Western CanadaMolecular characterization of Pst isolates from Western Canada
Molecular characterization of Pst isolates from Western CanadaBorlaug Global Rust Initiative
 
New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...
New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...
New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...Lex Nederbragt
 
Studying the microbiome
Studying the microbiomeStudying the microbiome
Studying the microbiomeMick Watson
 
140127 abrf interlaboratory study proposal
140127 abrf interlaboratory study proposal140127 abrf interlaboratory study proposal
140127 abrf interlaboratory study proposalGenomeInABottle
 
Next-Generation Sequencing Commercial Milestones Infographic
Next-Generation Sequencing Commercial Milestones InfographicNext-Generation Sequencing Commercial Milestones Infographic
Next-Generation Sequencing Commercial Milestones InfographicQIAGEN
 
Phylogenomic methods for comparative evolutionary biology - University Colleg...
Phylogenomic methods for comparative evolutionary biology - University Colleg...Phylogenomic methods for comparative evolutionary biology - University Colleg...
Phylogenomic methods for comparative evolutionary biology - University Colleg...Joe Parker
 
Aug2014 abrf interlaboratory study plans
Aug2014 abrf interlaboratory study plansAug2014 abrf interlaboratory study plans
Aug2014 abrf interlaboratory study plansGenomeInABottle
 
Molecular QC: Interpreting your Bioinformatics Pipeline
Molecular QC: Interpreting your Bioinformatics PipelineMolecular QC: Interpreting your Bioinformatics Pipeline
Molecular QC: Interpreting your Bioinformatics PipelineCandy Smellie
 
Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...
Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...
Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...John Blue
 
Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016
Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016
Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016Prof. Wim Van Criekinge
 
Galaxy RNA-Seq Analysis: Tuxedo Protocol
Galaxy RNA-Seq Analysis: Tuxedo ProtocolGalaxy RNA-Seq Analysis: Tuxedo Protocol
Galaxy RNA-Seq Analysis: Tuxedo ProtocolHong ChangBum
 

En vedette (20)

Introduction to next generation sequencing
Introduction to next generation sequencingIntroduction to next generation sequencing
Introduction to next generation sequencing
 
A Comparison of NGS Platforms.
A Comparison of NGS Platforms.A Comparison of NGS Platforms.
A Comparison of NGS Platforms.
 
Ngs microbiome
Ngs microbiomeNgs microbiome
Ngs microbiome
 
Clinical Applications of Next Generation Sequencing
Clinical Applications of Next Generation SequencingClinical Applications of Next Generation Sequencing
Clinical Applications of Next Generation Sequencing
 
NGS - Basic principles and sequencing platforms
NGS - Basic principles and sequencing platformsNGS - Basic principles and sequencing platforms
NGS - Basic principles and sequencing platforms
 
Illumina Sequencing
Illumina SequencingIllumina Sequencing
Illumina Sequencing
 
Molecular characterization of Pst isolates from Western Canada
Molecular characterization of Pst isolates from Western CanadaMolecular characterization of Pst isolates from Western Canada
Molecular characterization of Pst isolates from Western Canada
 
New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...
New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...
New High Throughput Sequencing technologies at the Norwegian Sequencing Centr...
 
Studying the microbiome
Studying the microbiomeStudying the microbiome
Studying the microbiome
 
140127 abrf interlaboratory study proposal
140127 abrf interlaboratory study proposal140127 abrf interlaboratory study proposal
140127 abrf interlaboratory study proposal
 
Next-Generation Sequencing Commercial Milestones Infographic
Next-Generation Sequencing Commercial Milestones InfographicNext-Generation Sequencing Commercial Milestones Infographic
Next-Generation Sequencing Commercial Milestones Infographic
 
Phylogenomic methods for comparative evolutionary biology - University Colleg...
Phylogenomic methods for comparative evolutionary biology - University Colleg...Phylogenomic methods for comparative evolutionary biology - University Colleg...
Phylogenomic methods for comparative evolutionary biology - University Colleg...
 
Aug2014 abrf interlaboratory study plans
Aug2014 abrf interlaboratory study plansAug2014 abrf interlaboratory study plans
Aug2014 abrf interlaboratory study plans
 
Molecular QC: Interpreting your Bioinformatics Pipeline
Molecular QC: Interpreting your Bioinformatics PipelineMolecular QC: Interpreting your Bioinformatics Pipeline
Molecular QC: Interpreting your Bioinformatics Pipeline
 
Ngs part i 2013
Ngs part i 2013Ngs part i 2013
Ngs part i 2013
 
Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...
Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...
Dr. Douglas Marthaler - Use of Next Generation Sequencing for Whole Genome An...
 
Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016
Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016
Galaxy dna-seq-variant calling-presentationandpractical_gent_april-2016
 
Galaxy RNA-Seq Analysis: Tuxedo Protocol
Galaxy RNA-Seq Analysis: Tuxedo ProtocolGalaxy RNA-Seq Analysis: Tuxedo Protocol
Galaxy RNA-Seq Analysis: Tuxedo Protocol
 
2016 iHT2 San Diego Health IT Summit
2016 iHT2 San Diego Health IT Summit2016 iHT2 San Diego Health IT Summit
2016 iHT2 San Diego Health IT Summit
 
Biz model for ion proton dna sequencer
Biz model for ion proton dna sequencerBiz model for ion proton dna sequencer
Biz model for ion proton dna sequencer
 

Similaire à Sequencing, Genome Assembly and the SGN Platform

ICAR Soybean Indore 2014
ICAR Soybean Indore 2014ICAR Soybean Indore 2014
ICAR Soybean Indore 2014Surya Saha
 
Improving Torsion Library Patterns with SMARTScompare
Improving Torsion Library Patterns with SMARTScompareImproving Torsion Library Patterns with SMARTScompare
Improving Torsion Library Patterns with SMARTScomparePatrick Penner
 
Plant Pathology Seminar
Plant Pathology SeminarPlant Pathology Seminar
Plant Pathology SeminarBongsoo Park
 
BOKU BILS 2016
BOKU BILS 2016BOKU BILS 2016
BOKU BILS 2016GBX Events
 
Bioinformatic tools in Pheromone technology
Bioinformatic tools in Pheromone technologyBioinformatic tools in Pheromone technology
Bioinformatic tools in Pheromone technologyTHILAKAR MANI
 
Powering Scientific Discovery with the Semantic Web (VanBUG 2014)
Powering Scientific Discovery with the Semantic Web (VanBUG 2014)Powering Scientific Discovery with the Semantic Web (VanBUG 2014)
Powering Scientific Discovery with the Semantic Web (VanBUG 2014)Michel Dumontier
 

Similaire à Sequencing, Genome Assembly and the SGN Platform (6)

ICAR Soybean Indore 2014
ICAR Soybean Indore 2014ICAR Soybean Indore 2014
ICAR Soybean Indore 2014
 
Improving Torsion Library Patterns with SMARTScompare
Improving Torsion Library Patterns with SMARTScompareImproving Torsion Library Patterns with SMARTScompare
Improving Torsion Library Patterns with SMARTScompare
 
Plant Pathology Seminar
Plant Pathology SeminarPlant Pathology Seminar
Plant Pathology Seminar
 
BOKU BILS 2016
BOKU BILS 2016BOKU BILS 2016
BOKU BILS 2016
 
Bioinformatic tools in Pheromone technology
Bioinformatic tools in Pheromone technologyBioinformatic tools in Pheromone technology
Bioinformatic tools in Pheromone technology
 
Powering Scientific Discovery with the Semantic Web (VanBUG 2014)
Powering Scientific Discovery with the Semantic Web (VanBUG 2014)Powering Scientific Discovery with the Semantic Web (VanBUG 2014)
Powering Scientific Discovery with the Semantic Web (VanBUG 2014)
 

Plus de Surya Saha

An open access resource portal for arthropod vectors and agricultural pathosy...
An open access resource portal for arthropod vectors and agricultural pathosy...An open access resource portal for arthropod vectors and agricultural pathosy...
An open access resource portal for arthropod vectors and agricultural pathosy...Surya Saha
 
Functional annotation of invertebrate genomes
Functional annotation of invertebrate genomesFunctional annotation of invertebrate genomes
Functional annotation of invertebrate genomesSurya Saha
 
Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...
Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...
Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...Surya Saha
 
Updates on Citrusgreening.org database from USDA NIFA project meeting
Updates on Citrusgreening.org database from USDA NIFA project meetingUpdates on Citrusgreening.org database from USDA NIFA project meeting
Updates on Citrusgreening.org database from USDA NIFA project meetingSurya Saha
 
Updates on the ACP v3 genome and annotation from USDA NIFA project meeting
Updates on the ACP v3 genome and annotation from USDA NIFA project meetingUpdates on the ACP v3 genome and annotation from USDA NIFA project meeting
Updates on the ACP v3 genome and annotation from USDA NIFA project meetingSurya Saha
 
AgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant Diseases
AgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant DiseasesAgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant Diseases
AgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant DiseasesSurya Saha
 
Visualization of insect vector-plant pathogen interactions in the citrus gree...
Visualization of insect vector-plant pathogen interactions in the citrus gree...Visualization of insect vector-plant pathogen interactions in the citrus gree...
Visualization of insect vector-plant pathogen interactions in the citrus gree...Surya Saha
 
Deciphering the genome of Diaphorina citri to develop solutions for the citru...
Deciphering the genome of Diaphorina citri to develop solutions for the citru...Deciphering the genome of Diaphorina citri to develop solutions for the citru...
Deciphering the genome of Diaphorina citri to develop solutions for the citru...Surya Saha
 
Quality Control of Sequencing Data
Quality Control of Sequencing Data Quality Control of Sequencing Data
Quality Control of Sequencing Data Surya Saha
 
Community resources for all y’all Omics
Community resources for all y’all OmicsCommunity resources for all y’all Omics
Community resources for all y’all OmicsSurya Saha
 
CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis...
 CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis... CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis...
CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis...Surya Saha
 
Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...
Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...
Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...Surya Saha
 
Tomato Genome Build SL3.0
Tomato Genome Build SL3.0Tomato Genome Build SL3.0
Tomato Genome Build SL3.0Surya Saha
 
Quality Control of Sequencing Data
Quality Control of Sequencing DataQuality Control of Sequencing Data
Quality Control of Sequencing DataSurya Saha
 
Tomato Genome SL2.50 and Beyond…
Tomato Genome SL2.50 and Beyond…Tomato Genome SL2.50 and Beyond…
Tomato Genome SL2.50 and Beyond…Surya Saha
 
Quality Control of NGS Data
Quality Control of NGS Data Quality Control of NGS Data
Quality Control of NGS Data Surya Saha
 
Quality Control of NGS Data Solutions
Quality Control of NGS Data  SolutionsQuality Control of NGS Data  Solutions
Quality Control of NGS Data SolutionsSurya Saha
 
Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...
Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...
Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...Surya Saha
 
Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...
Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...
Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...Surya Saha
 
Tools for Metagenomics with 16S/ITS and Whole Genome Shotgun Sequences
Tools for Metagenomics with 16S/ITS and Whole Genome Shotgun SequencesTools for Metagenomics with 16S/ITS and Whole Genome Shotgun Sequences
Tools for Metagenomics with 16S/ITS and Whole Genome Shotgun SequencesSurya Saha
 

Plus de Surya Saha (20)

An open access resource portal for arthropod vectors and agricultural pathosy...
An open access resource portal for arthropod vectors and agricultural pathosy...An open access resource portal for arthropod vectors and agricultural pathosy...
An open access resource portal for arthropod vectors and agricultural pathosy...
 
Functional annotation of invertebrate genomes
Functional annotation of invertebrate genomesFunctional annotation of invertebrate genomes
Functional annotation of invertebrate genomes
 
Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...
Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...
Saha UC Davis Plant Pathology seminar Infrastructure for battling the Citrus ...
 
Updates on Citrusgreening.org database from USDA NIFA project meeting
Updates on Citrusgreening.org database from USDA NIFA project meetingUpdates on Citrusgreening.org database from USDA NIFA project meeting
Updates on Citrusgreening.org database from USDA NIFA project meeting
 
Updates on the ACP v3 genome and annotation from USDA NIFA project meeting
Updates on the ACP v3 genome and annotation from USDA NIFA project meetingUpdates on the ACP v3 genome and annotation from USDA NIFA project meeting
Updates on the ACP v3 genome and annotation from USDA NIFA project meeting
 
AgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant Diseases
AgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant DiseasesAgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant Diseases
AgriVectors: A Data and Systems Resource for Arthropod Vectors of Plant Diseases
 
Visualization of insect vector-plant pathogen interactions in the citrus gree...
Visualization of insect vector-plant pathogen interactions in the citrus gree...Visualization of insect vector-plant pathogen interactions in the citrus gree...
Visualization of insect vector-plant pathogen interactions in the citrus gree...
 
Deciphering the genome of Diaphorina citri to develop solutions for the citru...
Deciphering the genome of Diaphorina citri to develop solutions for the citru...Deciphering the genome of Diaphorina citri to develop solutions for the citru...
Deciphering the genome of Diaphorina citri to develop solutions for the citru...
 
Quality Control of Sequencing Data
Quality Control of Sequencing Data Quality Control of Sequencing Data
Quality Control of Sequencing Data
 
Community resources for all y’all Omics
Community resources for all y’all OmicsCommunity resources for all y’all Omics
Community resources for all y’all Omics
 
CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis...
 CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis... CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis...
CitrusCyc: Metabolic Pathway Databases for the C. clementina and C. sinensis...
 
Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...
Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...
Using Long Reads, Optical Maps and Long-Range Scaffolding to improve the Diap...
 
Tomato Genome Build SL3.0
Tomato Genome Build SL3.0Tomato Genome Build SL3.0
Tomato Genome Build SL3.0
 
Quality Control of Sequencing Data
Quality Control of Sequencing DataQuality Control of Sequencing Data
Quality Control of Sequencing Data
 
Tomato Genome SL2.50 and Beyond…
Tomato Genome SL2.50 and Beyond…Tomato Genome SL2.50 and Beyond…
Tomato Genome SL2.50 and Beyond…
 
Quality Control of NGS Data
Quality Control of NGS Data Quality Control of NGS Data
Quality Control of NGS Data
 
Quality Control of NGS Data Solutions
Quality Control of NGS Data  SolutionsQuality Control of NGS Data  Solutions
Quality Control of NGS Data Solutions
 
Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...
Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...
Mining Eukaryotic Meta-Genomes for Endosymbionts using Next-Generation Sequen...
 
Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...
Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...
Endosymbiont hunting in the metagenome of Asian citrus psyllid (Diaphorina ci...
 
Tools for Metagenomics with 16S/ITS and Whole Genome Shotgun Sequences
Tools for Metagenomics with 16S/ITS and Whole Genome Shotgun SequencesTools for Metagenomics with 16S/ITS and Whole Genome Shotgun Sequences
Tools for Metagenomics with 16S/ITS and Whole Genome Shotgun Sequences
 

Dernier

Pests of Bengal gram_Identification_Dr.UPR.pdf
Pests of Bengal gram_Identification_Dr.UPR.pdfPests of Bengal gram_Identification_Dr.UPR.pdf
Pests of Bengal gram_Identification_Dr.UPR.pdfPirithiRaju
 
Radiation physics in Dental Radiology...
Radiation physics in Dental Radiology...Radiation physics in Dental Radiology...
Radiation physics in Dental Radiology...navyadasi1992
 
LIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptx
LIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptxLIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptx
LIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptxmalonesandreagweneth
 
Vision and reflection on Mining Software Repositories research in 2024
Vision and reflection on Mining Software Repositories research in 2024Vision and reflection on Mining Software Repositories research in 2024
Vision and reflection on Mining Software Repositories research in 2024AyushiRastogi48
 
Davis plaque method.pptx recombinant DNA technology
Davis plaque method.pptx recombinant DNA technologyDavis plaque method.pptx recombinant DNA technology
Davis plaque method.pptx recombinant DNA technologycaarthichand2003
 
《Queensland毕业文凭-昆士兰大学毕业证成绩单》
《Queensland毕业文凭-昆士兰大学毕业证成绩单》《Queensland毕业文凭-昆士兰大学毕业证成绩单》
《Queensland毕业文凭-昆士兰大学毕业证成绩单》rnrncn29
 
Base editing, prime editing, Cas13 & RNA editing and organelle base editing
Base editing, prime editing, Cas13 & RNA editing and organelle base editingBase editing, prime editing, Cas13 & RNA editing and organelle base editing
Base editing, prime editing, Cas13 & RNA editing and organelle base editingNetHelix
 
Harmful and Useful Microorganisms Presentation
Harmful and Useful Microorganisms PresentationHarmful and Useful Microorganisms Presentation
Harmful and Useful Microorganisms Presentationtahreemzahra82
 
THE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptx
THE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptxTHE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptx
THE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptxNandakishor Bhaurao Deshmukh
 
Pests of jatropha_Bionomics_identification_Dr.UPR.pdf
Pests of jatropha_Bionomics_identification_Dr.UPR.pdfPests of jatropha_Bionomics_identification_Dr.UPR.pdf
Pests of jatropha_Bionomics_identification_Dr.UPR.pdfPirithiRaju
 
Topic 9- General Principles of International Law.pptx
Topic 9- General Principles of International Law.pptxTopic 9- General Principles of International Law.pptx
Topic 9- General Principles of International Law.pptxJorenAcuavera1
 
ALL ABOUT MIXTURES IN GRADE 7 CLASS PPTX
ALL ABOUT MIXTURES IN GRADE 7 CLASS PPTXALL ABOUT MIXTURES IN GRADE 7 CLASS PPTX
ALL ABOUT MIXTURES IN GRADE 7 CLASS PPTXDole Philippines School
 
Microphone- characteristics,carbon microphone, dynamic microphone.pptx
Microphone- characteristics,carbon microphone, dynamic microphone.pptxMicrophone- characteristics,carbon microphone, dynamic microphone.pptx
Microphone- characteristics,carbon microphone, dynamic microphone.pptxpriyankatabhane
 
FREE NURSING BUNDLE FOR NURSES.PDF by na
FREE NURSING BUNDLE FOR NURSES.PDF by naFREE NURSING BUNDLE FOR NURSES.PDF by na
FREE NURSING BUNDLE FOR NURSES.PDF by naJASISJULIANOELYNV
 
User Guide: Capricorn FLX™ Weather Station
User Guide: Capricorn FLX™ Weather StationUser Guide: Capricorn FLX™ Weather Station
User Guide: Capricorn FLX™ Weather StationColumbia Weather Systems
 
Thermodynamics ,types of system,formulae ,gibbs free energy .pptx
Thermodynamics ,types of system,formulae ,gibbs free energy .pptxThermodynamics ,types of system,formulae ,gibbs free energy .pptx
Thermodynamics ,types of system,formulae ,gibbs free energy .pptxuniversity
 
Bioteknologi kelas 10 kumer smapsa .pptx
Bioteknologi kelas 10 kumer smapsa .pptxBioteknologi kelas 10 kumer smapsa .pptx
Bioteknologi kelas 10 kumer smapsa .pptx023NiWayanAnggiSriWa
 
Observational constraints on mergers creating magnetism in massive stars
Observational constraints on mergers creating magnetism in massive starsObservational constraints on mergers creating magnetism in massive stars
Observational constraints on mergers creating magnetism in massive starsSérgio Sacani
 
Dubai Calls Girl Lisa O525547819 Lexi Call Girls In Dubai
Dubai Calls Girl Lisa O525547819 Lexi Call Girls In DubaiDubai Calls Girl Lisa O525547819 Lexi Call Girls In Dubai
Dubai Calls Girl Lisa O525547819 Lexi Call Girls In Dubaikojalkojal131
 

Dernier (20)

Pests of Bengal gram_Identification_Dr.UPR.pdf
Pests of Bengal gram_Identification_Dr.UPR.pdfPests of Bengal gram_Identification_Dr.UPR.pdf
Pests of Bengal gram_Identification_Dr.UPR.pdf
 
Radiation physics in Dental Radiology...
Radiation physics in Dental Radiology...Radiation physics in Dental Radiology...
Radiation physics in Dental Radiology...
 
LIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptx
LIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptxLIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptx
LIGHT-PHENOMENA-BY-CABUALDIONALDOPANOGANCADIENTE-CONDEZA (1).pptx
 
Vision and reflection on Mining Software Repositories research in 2024
Vision and reflection on Mining Software Repositories research in 2024Vision and reflection on Mining Software Repositories research in 2024
Vision and reflection on Mining Software Repositories research in 2024
 
Davis plaque method.pptx recombinant DNA technology
Davis plaque method.pptx recombinant DNA technologyDavis plaque method.pptx recombinant DNA technology
Davis plaque method.pptx recombinant DNA technology
 
《Queensland毕业文凭-昆士兰大学毕业证成绩单》
《Queensland毕业文凭-昆士兰大学毕业证成绩单》《Queensland毕业文凭-昆士兰大学毕业证成绩单》
《Queensland毕业文凭-昆士兰大学毕业证成绩单》
 
Base editing, prime editing, Cas13 & RNA editing and organelle base editing
Base editing, prime editing, Cas13 & RNA editing and organelle base editingBase editing, prime editing, Cas13 & RNA editing and organelle base editing
Base editing, prime editing, Cas13 & RNA editing and organelle base editing
 
Harmful and Useful Microorganisms Presentation
Harmful and Useful Microorganisms PresentationHarmful and Useful Microorganisms Presentation
Harmful and Useful Microorganisms Presentation
 
THE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptx
THE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptxTHE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptx
THE ROLE OF PHARMACOGNOSY IN TRADITIONAL AND MODERN SYSTEM OF MEDICINE.pptx
 
Pests of jatropha_Bionomics_identification_Dr.UPR.pdf
Pests of jatropha_Bionomics_identification_Dr.UPR.pdfPests of jatropha_Bionomics_identification_Dr.UPR.pdf
Pests of jatropha_Bionomics_identification_Dr.UPR.pdf
 
Let’s Say Someone Did Drop the Bomb. Then What?
Let’s Say Someone Did Drop the Bomb. Then What?Let’s Say Someone Did Drop the Bomb. Then What?
Let’s Say Someone Did Drop the Bomb. Then What?
 
Topic 9- General Principles of International Law.pptx
Topic 9- General Principles of International Law.pptxTopic 9- General Principles of International Law.pptx
Topic 9- General Principles of International Law.pptx
 
ALL ABOUT MIXTURES IN GRADE 7 CLASS PPTX
ALL ABOUT MIXTURES IN GRADE 7 CLASS PPTXALL ABOUT MIXTURES IN GRADE 7 CLASS PPTX
ALL ABOUT MIXTURES IN GRADE 7 CLASS PPTX
 
Microphone- characteristics,carbon microphone, dynamic microphone.pptx
Microphone- characteristics,carbon microphone, dynamic microphone.pptxMicrophone- characteristics,carbon microphone, dynamic microphone.pptx
Microphone- characteristics,carbon microphone, dynamic microphone.pptx
 
FREE NURSING BUNDLE FOR NURSES.PDF by na
FREE NURSING BUNDLE FOR NURSES.PDF by naFREE NURSING BUNDLE FOR NURSES.PDF by na
FREE NURSING BUNDLE FOR NURSES.PDF by na
 
User Guide: Capricorn FLX™ Weather Station
User Guide: Capricorn FLX™ Weather StationUser Guide: Capricorn FLX™ Weather Station
User Guide: Capricorn FLX™ Weather Station
 
Thermodynamics ,types of system,formulae ,gibbs free energy .pptx
Thermodynamics ,types of system,formulae ,gibbs free energy .pptxThermodynamics ,types of system,formulae ,gibbs free energy .pptx
Thermodynamics ,types of system,formulae ,gibbs free energy .pptx
 
Bioteknologi kelas 10 kumer smapsa .pptx
Bioteknologi kelas 10 kumer smapsa .pptxBioteknologi kelas 10 kumer smapsa .pptx
Bioteknologi kelas 10 kumer smapsa .pptx
 
Observational constraints on mergers creating magnetism in massive stars
Observational constraints on mergers creating magnetism in massive starsObservational constraints on mergers creating magnetism in massive stars
Observational constraints on mergers creating magnetism in massive stars
 
Dubai Calls Girl Lisa O525547819 Lexi Call Girls In Dubai
Dubai Calls Girl Lisa O525547819 Lexi Call Girls In DubaiDubai Calls Girl Lisa O525547819 Lexi Call Girls In Dubai
Dubai Calls Girl Lisa O525547819 Lexi Call Girls In Dubai
 

Sequencing, Genome Assembly and the SGN Platform

  • 1. Surya Saha, Ph.D. Cornell University & Boyce Thompson Institute suryasaha@cornell.edu @SahaSurya Centre for Agricultural Bioinformatics Pusa, New Delhi June 13,2014 Slides: http://bit.ly/CABin_Pusa_2014 http://www.acgt.me/blog/2014/3/7/next-generation-sequencing-must-die Genome Assembly Jason Chin http://www.bit.ly/SZPKIG
  • 2. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 2 You are free to: Copy, share, adapt, or re-mix; Photograph, film, or broadcast; Blog, live-blog, or post video of; This presentation. Provided that: You attribute the work to its author and respect the rights and licenses associated with its components. Slide Concept by Cameron Neylon, who has waived all copyright and related or neighbouring rights. This slide only ccZero. Social Media Icons adapted with permission from originals by Christopher Ross. Original images are available under GPL at http://www.thisismyurl.com/free-downloads/15-free-speech-bubble-icons-for-popular-websites
  • 3. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 3 Sequencing
  • 4. 1953 DNA Structure discovery 1977 2012 Sanger DNA sequencing by chain-terminating inhibitors 1984 Epstein-Barr virus (170 Kb) 1987Abi370 Sequencer 1995 2001 Homo sapiens (3.0 Gb) 2005 454 Solexa Solid 2007 2011 Ion Torrent PacBio Haemophilus influenzae (1.83 Mb) 2013 Slide credit: Aureliano Bombarely Sequencing over the Ages Illumina Illumina Hiseq X 454 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 4 Pinus taeda (24 Gb) 2014 MinION The Next Generation
  • 5. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 5 Its all about the $£€¥ http://www.genome.gov/sequencingcosts/
  • 6. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 6 First generation sequencing
  • 7. Sanger method 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 7 Frederick Sanger 13 Aug 1918 – 19 Nov 2013 Won the Nobel Prize for Chemistry in 1958 and 1980. Published the dideoxy chain termination method or “Sanger method” in 1977 http://dailym.ai/1f1XeTB
  • 8. Sanger method 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 8 http://bit.ly/1g6Cudq http://bit.ly/1lcQO4J
  • 9. First generation sequencing • Very high quality sequences (99.999%) • Very low throughput 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 9 Run Time Read Length Reads / Run Total nucleotides sequenced Cost / MB Capillary Sequencing (ABI3730xl) 20m-3h 400-900 bp 96 or 386 1.9-84 Kb $2400 http://bit.ly/1clLps3 http://1.usa.gov/1cLqIRd
  • 10. Use the specific technology used to generate the data – Illumina Hiseq/Miseq/NextSeq – Pacific Biosciences RS I/RS II – Ion Torrent Proton/PGM – SOLiD – 454 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 10 http://www.acgt.me/blog/2014/3/10/next-generation- sequencing-must-diepart-2
  • 11. 454 Pyrosequencing One purified DNA fragment, to one bead, to one read. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 11 http://bit.ly/1ehwxWN GS FLX Titanium http://bit.ly/1ehAcEh
  • 12. Illumina 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 12 Output 15 Gb 120 GB 1000 GB 1800 GB Number of Reads 25 Million 400 Million 4 Billion 6 Billion Read Length 2x300 bp 2x150 bp 2x125 bp (2x250 update mid-2014) 2x150 bp Cost $99K $250K $740K $10M Source: Illumina $1000 human genome??
  • 13. Illumina 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 13 http://1.usa.gov/1fP9ybl
  • 14. Illumina:Moleculo 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 14 http://bit.ly/1aEPOBn
  • 15. Pacific Biosciences SMRT sequencing Single Molecule Real Time sequencing 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 15 http://bit.ly/1naxgTe
  • 16. Pacific Biosciences SMRT sequencing Error correction methods 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 16 Hierarchical genome-assembly process (HGAP) PBJelly Enlish et al., PLOS One. 2012 PBJelly
  • 17. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 17 Pacific Biosciences SMRT sequencing Read Lengths
  • 18. Oxford Nanopore 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 18 https://www.nanoporetech.com/ • No data yet?? • Error model http://erlichya.tumblr.com/post/66376172948/hands-on- experience-with-oxford-nanopore-minion
  • 19. Others • Ion Torrent Proton/PGM • Nabsys • SOLiD 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 19
  • 20. Comparison 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 20
  • 21. Next generation sequencing 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 21 Run Time Read Length Quality Total nucleotides sequenced Cost /MB 454 Pyrosequencing 24h 700 bp Q20-Q30 0.7 GB $10 Illumina Miseq 27h 2x250bp > Q30 15 GB $0.15 Illumina Hiseq 2500 11days 2x125bp >Q30 1000 GB $0.05 Ion torrent 2h 400bp >Q20 50MB-1GB $1 Pacific Biosciences 2h 10-20kb >Q30 consensus >Q10 single 400-800MB /SMRT cell $0.33-$1 http://bit.ly/1clLps3 http://1.usa.gov/1cLqIRd
  • 22. http://omicsmaps.com/ Next Generation Genomics: World Map of High-throughput Sequencers Centre for Agricultural Bioinformatics, Pusa6/15/2014 22
  • 23. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 23 http://bit.ly/18pfUId
  • 24. Real cost of Sequencing!! Sboner, Genome Biology, 2011 6/15/2014 24Centre for Agricultural Bioinformatics, Pusa
  • 25. Library Types Single end Pair end (PE, 150-800 bp, Fwd:/1, Rev:/2) Mate pair (MP, 2Kb to 20 Kb) 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 25 F F R F R 454/Roche FR Illumina Illumina Slide credit: Aureliano Bombarely
  • 26. Implications of Choice of Library 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 26 Slide credit: Aureliano Bombarely Consensus sequence (Contig) Reads Scaffold (or Supercontig) Pair Read information NNNNN Pseudomolecule (or ultracontig) F Genetic information (markers) NNNNN NN
  • 27. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 27 Quality control: Encoding http://bit.ly/N28yUd Phred score of a base is: Qphred = -10 log10 (e) where e is the estimated probability of a base being incorrect
  • 28. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 28 Genome Assembly
  • 29. Whole Genome Shotgun Sequencing 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 29 Slide credit: cbcb.umd.edu
  • 30. Genome Sequencing Strategies 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 30 Slide credit: Aureliano Bombarely
  • 31. Genome Sequencing Strategies 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 31 International Human Genome Sequencing Consortium 2001 Overlap Layout Consensus http://contig.wordpress.com/ cbcb.umd.edu
  • 32. DeBruijnGraph 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 32
  • 33. Ingredient for a Good Assembly 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 33 Slide credit: Mike Schatz
  • 34. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 34
  • 35. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 35 Bird Snake
  • 36. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 36 • You have the expertise to install and run • You have the suitable infrastructure (CPU & RAM) to run the assembler • You have sufficient time to run the assembler • Is designed to work with the specific mix of NGS data that you have generated • Best addresses what you want to get out of a genome assembly (bigger overall assembly, more genes, most accuracy, longer scaffolds, most resolution of haplotypes, most tolerant of repeats, etc.) The BEST?? Genome Assembler for YOU http://haldanessieve.org/2013/01/28/our-paper-making-pizzas-and-genome-assemblies/
  • 37. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 37
  • 38. Which technology to use?? • Microbial genomes • Eukaryotic genomes • Resequencing genomes • RNAseq and other XXXseq methods 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 38 http://bit.ly/1ko9Kgh
  • 39. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 39 SOL Genomics Network
  • 40. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 40
  • 41. The SGN Team!! 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 41 Surya Saha, Tom Fisher-York, Hartmut Foerster, Suzy Strickler, Jeremy Edwards, Noe Fernandez, Naama Menda, Aure Bombarely, Aimin Yan, Isaak Tecle
  • 42. SGN Website 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 42 http://solgenomics.net
  • 43. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 43 Main web page (front page): WEB ICONS TOOL BAR
  • 44. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 44 Main web page (front page): TOOL BAR (MENUS)
  • 45. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 45 But the DATA also can be edited LocusLocus Editor Data Community Data Curation
  • 46. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 46 You need • SGN account. • Activate submitter / Locus Editor privileges by SGN curator LocusLocus Editor Data
  • 47. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 47 Tools
  • 48. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 48 Genome Browser: GBrowse
  • 49. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 49 Genome Browser: JBrowse
  • 50. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 50
  • 51. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 51 CassavaBase http://cassavabase.org/ Slide credit: Jeremy Edwards
  • 52. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 52 NextGen Cassava Project ● Project: Adapt SGN database for Cassava Breeding ● Goal: Apply Genomic Selection to cassava breeding ● Predict breeding values from genotype information ● Shorten the breeding cycle ● Massive amounts of genotypic data (GBS) ● Phenotypic data ● Data management challenge ● Improve flowering ● http://nextgencassava.org Slide credit: Jeremy Edwards
  • 53. SGN/Cassavabase behind the scenes 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 53 ● Perl/Catalyst MVC Framework ● PostgreSQL Database ● Generic Model Organism Database (GMOD) – Chado relational database schema – GBrowse – JBrowse ● R – Experimental design – QTL mapping – Genomic selection Slide credit: Jeremy Edwards
  • 54. Objectives Provide cassava breeders and researchers access to data and tools in a centralized, user-friendly and reliable database. – Improve partner breeding program information tracking – Streamline management of genotypic and phenotypic data – Pipeline genotypic and phenotypic data through Genomic Selection prediction analyses 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 54 Slide credit: Jeremy Edwards
  • 55. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 55 Genomic Selection The 'training population' is genotyped and phenotyped to 'train' the genomic selection (GS) prediction model. Genotypic information from the breeding material is then fed into the model to calculate genomic estimated breeding values (GEBV) for these lines. From Heffner et al. 2009 Crop Sci. 49:1–12 Information from a majority of lines in the breeding population (the training set) is used to create the prediction model. The model is then used to predict the phenotypes of the remaining lines (the validation set), using genotypic information only. The results from the model are compared to the actual data to give the prediction accuracy. Image courtesy of Martha Hamblin, Cornell University Flow diagram of a genomic selection breeding program. Breeding cycle time is shortened by removing phenotypic evaluation of lines before selection as parents for the next cycle. From Heffner et al. 2009 Crop Sci. 49:1–12 Slide credit: Jeremy Edwards
  • 56. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 56 Data collection in the field ● Android tablets ● Field book app – Jesse Poland's group at USDA-ARS / Kansas State University Slide credit: Jeremy Edwards
  • 57. Cassava Trait Ontology 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 57 Kulakow et al. 2011 Kulakow et al. 2011 ● Standard terminology ● Facilitate the sharing of information ● Allow users to query keywords related to traits Slide credit: Jeremy Edwards
  • 58. 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 58 Position available at Solgenomics Cassavabase project Plant Breeding + Bioinformatician ● Familiar with breeding ● Programming in Perl, R, SQL, Hadoop ● Linux ● Africa ● Genius http://www.cassavabase.org/forum/posts .pl?topic_id=9
  • 59. Thank you!! Questions?? 6/15/2014 Centre for Agricultural Bioinformatics, Pusa 59