scispace - formally typeset
Open AccessJournal ArticleDOI

FLASH: Fast Length Adjustment of Short Reads to Improve Genome Assemblies

Tanja Magoc, +1 more
- 01 Nov 2011 - 
- Vol. 27, Iss: 21, pp 2957-2963
TLDR
FLASH is a fast computational tool to extend the length of short reads by overlapping paired-end reads from fragment libraries that are sufficiently short and when FLASH was used to extend reads prior to assembly, the resulting assemblies had substantially greater N50 lengths for both contigs and scaffolds.
Abstract
Motivation: Next-generation sequencing technologies generate very large numbers of short reads. Even with very deep genome coverage, short read lengths cause problems in de novo assemblies. The use of paired-end libraries with a fragment size shorter than twice the read length provides an opportunity to generate much longer reads by overlapping and merging read pairs before assembling a genome. Results: We present FLASH, a fast computational tool to extend the length of short reads by overlapping paired-end reads from fragment libraries that are sufficiently short. We tested the correctness of the tool on one million simulated read pairs, and we then applied it as a pre-processor for genome assemblies of Illumina reads from the bacterium Staphylococcus aureus and human chromosome 14. FLASH correctly extended and merged reads >99% of the time on simulated reads with an error rate of <1%. With adequately set parameters, FLASH correctly merged reads over 90% of the time even when the reads contained up to 5% errors. When FLASH was used to extend reads prior to assembly, the resulting assemblies had substantially greater N50 lengths for both contigs and scaffolds. Availability and Implementation: The FLASH system is implemented in C and is freely available as open-source code at http://www.cbcb.umd.edu/software/flash. Contact: moc.liamg@cogam.t

read more

Content maybe subject to copyright    Report

Citations
More filters
Journal ArticleDOI

Consortia of low-abundance bacteria drive sulfate reduction-dependent degradation of fermentation products in peat soil microcosms.

TL;DR: It is shown that diverse consortia of low-abundance microorganisms can perform peat soil sulfate reduction, a process that exerts control on methane production in these climate-relevant ecosystems.
Journal ArticleDOI

Mechanism, kinetics and microbiology of inhibition caused by long-chain fatty acids in anaerobic digestion of algal biomass.

TL;DR: It is demonstrated that inoculum concentration has a more significant effect on alleviating LCFA inhibition than calcium concentration, while calcium only played a role when inoculum concentrations met a threshold level.
Journal ArticleDOI

Quantitative microbiome profiling disentangles inflammation- and bile duct obstruction-associated microbiota alterations across PSC/IBD diagnoses.

TL;DR: The authors apply quantitative microbiome profiling to a metagenomics data set comprising patients with primary sclerosing cholangitis and/or inflammatory bowel disease and identify microbial taxa associated with inflammation or specific disease indicators, which were validated in an independentinflammatory bowel disease cohort.
Journal ArticleDOI

Cultivation and functional characterization of 79 planctomycetes uncovers their unique biology

TL;DR: Diversity-driven cultivation, characterization and genome sequencing of 79 bacterial strains from all major taxonomic clades of the conspicuous bacterial phylum Planctomycetes are reported, identified previously unknown modes of bacterial cell division and illustrated how ‘microbial dark matter’ can be accessed by cultivation techniques, expanding the organismic background for small-molecule research and drug-target detection.
Journal ArticleDOI

Rescue of Fructose-Induced Metabolic Syndrome by Antibiotics or Faecal Transplantation in a Rat Model of Obesity.

TL;DR: In rats fed a fructose-rich diet the development of metabolic syndrome is directly correlated with variations of the gut content of specific bacterial taxa, pointing to a correlation between their abundance and the developing of the metabolic syndrome.
References
More filters
Journal ArticleDOI

The Sequence Alignment/Map format and SAMtools

TL;DR: SAMtools as discussed by the authors implements various utilities for post-processing alignments in the SAM format, such as indexing, variant caller and alignment viewer, and thus provides universal tools for processing read alignments.
Journal ArticleDOI

Ultrafast and memory-efficient alignment of short DNA sequences to the human genome

TL;DR: Bowtie extends previous Burrows-Wheeler techniques with a novel quality-aware backtracking algorithm that permits mismatches and can be used simultaneously to achieve even greater alignment speeds.
Journal ArticleDOI

Versatile and open software for comparing large genomes

TL;DR: The newest version of MUMmer easily handles comparisons of large eukaryotic genomes at varying evolutionary distances, as demonstrated by applications to multiple genomes.
Journal ArticleDOI

De novo assembly of human genomes with massively parallel short read sequencing

TL;DR: The development of this de novo short read assembly method creates new opportunities for building reference sequences and carrying out accurate analyses of unexplored genomes in a cost-effective way.
Journal ArticleDOI

High-quality draft assemblies of mammalian genomes from massively parallel sequence data

TL;DR: The development of an algorithm for genome assembly, ALLPATHS-LG, and its application to massively parallel DNA sequence data from the human and mouse genomes, generated on the Illumina platform, have good accuracy, short-range contiguity, long-range connectivity, and coverage of the genome.
Related Papers (5)