Home
/
Authors
/
Michelle L. Gwinn

Author

Michelle L. Gwinn

Bio: Michelle L. Gwinn is an academic researcher from J. Craig Venter Institute. The author has contributed to research in topics: Genome & Gene. The author has an hindex of 25, co-authored 28 publications receiving 19142 citations.

Topics: Genome, Gene, Plasmid, Whole genome sequencing, Virulence ...read more

Papers

PDF

Open Access

More filters

Journal Article•DOI•

Genome analysis of multiple pathogenic isolates of Streptococcus agalactiae: Implications for the microbial “pan-genome”

[...]

Hervé Tettelin, Vega Masignani, Michael J. Cieslewicz¹, Claudio Donati, Duccio Medini, Naomi L. Ward², Samuel V. Angiuoli³, Jonathan Crabtree³, Amanda L. Jones⁴, A. Scott Durkin³, Robert T. DeBoy³, Tanja M. Davidsen³, Marirosa Mora, Maria Scarselli, Immaculada Margarit Y Ros, Jeremy Peterson³, Christopher R. Hauser³, Jaideep P. Sundaram³, William C. Nelson³, Ramana Madupu³, Lauren M. Brinkac³, Robert J. Dodson³, M. J. Rosovitz³, Steven A. Sullivan³, Sean C. Daugherty³, Daniel H. Haft³, Jeremy D. Selengut³, Michelle L. Gwinn³, Liwei Zhou³, Nikhat Zafar³, Hoda Khouri³, Diana Radune³, George Dimitrov³, Kisha Watkins³, Kevin J. B. O'Connor⁵, Shannon Smith³, Teresa Utterback³, Owen White³, Craig E. Rubens⁴, Guido Grandi, Lawrence C. Madoff¹, Dennis L. Kasper¹, John L. Telford, Michael R. Wessels¹, Rino Rappuoli, Claire M. Fraser⁶ - Show less +42 more•Institutions (6)

Harvard University¹, University of Maryland, Baltimore County², J. Craig Venter Institute³, Boston Children's Hospital⁴, Johns Hopkins University⁵, George Washington University⁶

27 Sep 2005-Proceedings of the National Academy of Sciences of the United States of America

TL;DR: The genomic sequence of six strains representing the five major disease-causing serotypes of Streptococcus agalactiae, the main cause of neonatal infection in humans, was generated and Mathematical extrapolation of the data suggests that the gene reservoir available for inclusion in the S. agalactic pan-genome is vast and that unique genes will continue to be identified even after sequencing hundreds of genomes.

...read moreread less

Abstract: The development of efficient and inexpensive genome sequencing methods has revolutionized the study of human bacterial pathogens and improved vaccine design. Unfortunately, the sequence of a single genome does not reflect how genetic variability drives pathogenesis within a bacterial species and also limits genome-wide screens for vaccine candidates or for antimicrobial targets. We have generated the genomic sequence of six strains representing the five major disease-causing serotypes of Streptococcus agalactiae, the main cause of neonatal infection in humans. Analysis of these genomes and those available in databases showed that the S. agalactiae species can be described by a pan-genome consisting of a core genome shared by all isolates, accounting for ≈80% of any single genome, plus a dispensable genome consisting of partially shared and strain-specific genes. Mathematical extrapolation of the data suggests that the gene reservoir available for inclusion in the S. agalactiae pan-genome is vast and that unique genes will continue to be identified even after sequencing hundreds of genomes.

...read moreread less

2,092 citations

Journal Article•DOI•

Genomic sequence of a Lyme disease spirochaete, Borrelia burgdorferi

[...]

Claire M. Fraser¹, Sherwood R. Casjens², Wai Mun Huang², Granger G. Sutton¹, Rebecca A. Clayton¹, Raju Lathigra³, Owen White¹, Karen A. Ketchum¹, Robert J. Dodson¹, Erin Hickey¹, Michelle L. Gwinn¹, Brian Dougherty¹, J F Tomb¹, Robert D. Fleischmann¹, Delwood Richardson¹, Jeremy Peterson¹, Anthony R. Kerlavage¹, John Quackenbush¹, Steven L. Salzberg¹, Mark S. Hanson³, René Van Vugt², Nanette Palmer², Mark Raymond Adams¹, Jeannine D. Gocayne¹, Janice Weidman¹, Teresa Utterback¹, Larry Watthey¹, Lisa McDonald¹, Patricia Artiach¹, Cheryl Bowman¹, Stacey Garland¹, Claire Fujii¹, Matthew D. Cotton¹, Kurt Horst¹, Kevin Roberts¹, Bonnie Hatch¹, Hamilton O. Smith¹, J. Craig Venter¹ - Show less +34 more•Institutions (3)

J. Craig Venter Institute¹, University of Utah², MedImmune³

11 Dec 1997-Nature

TL;DR: The genome of the bacterium Borrelia burgdorferi B31, the aetiologic agent of Lyme disease, contains a linear chromosome of 910,725 base pairs and at least 17 linear and circular plasmids with a combined size of more than 533,000 base pairs, which suggest their limited metabolic capacities reflect convergent evolution by gene loss from more metabolically competent progenitors.

...read moreread less

Abstract: The genome of the bacterium Borrelia burgdorferi B31, the aetiologic agent of Lyme disease, contains a linear chromosome of 910,725 base pairs and at least 17 linear and circular plasmids with a combined size of more than 533,000 base pairs. The chromosome contains 853 genes encoding a basic set of proteins for DNA replication, transcription, translation, solute transport and energy metabolism, but, like Mycoplasma genitalium, it contains no genes for cellular biosynthetic reactions. Because B. burgdorferi and M. genitalium are distantly related eubacteria, we suggest that their limited metabolic capacities reflect convergent evolution by gene loss from more metabolically competent progenitors. Of 430 genes on 11 plasmids, most have no known biological function; 39% of plasmid genes are paralogues that form 47 gene families. The biological significance of the multiple plasmid-encoded genes is not clear, although they may be involved in antigenic variation or immune evasion.

...read moreread less

2,025 citations

Journal Article•DOI•

DNA sequence of both chromosomes of the cholera pathogen Vibrio cholerae

[...]

John F. Heidelberg, Jonathan A. Eisen, William C. Nelson, Rebecca A. Clayton, Michelle L. Gwinn, Robert J. Dodson, Daniel H. Haft, Erin Hickey, Jeremy Peterson, Lowell Umayam, Steven R. Gill, Karen E. Nelson, Timothy D. Read, Hervé Tettelin, Delwood Richardson, Maria D. Ermolaeva, Jessica Vamathevan, Steven Bass, Haiying Qin, Ioana Dragoi, Patrick Sellers, Lisa McDonald, Teresa Utterback, Robert D. Fleishmann, William C. Nierman, Owen White, Steven L. Salzberg, Hamilton O. Smith¹, Rita R. Colwell², Rita R. Colwell³, John J. Mekalanos⁴, J. Craig Venter¹, Claire M. Fraser - Show less +29 more•Institutions (4)

Celera Corporation¹, University of Maryland, College Park², University of Maryland Biotechnology Institute³, Harvard University⁴

03 Aug 2000-Nature

TL;DR: The V. cholerae genomic sequence provides a starting point for understanding how a free-living, environmental organism emerged to become a significant human bacterial pathogen.

...read moreread less

Abstract: Here we determine the complete genomic sequence of the Gram negative, g-Proteobacterium Vibrio cholerae El Tor N16961 to be 4,033,460 base pairs (bp). The genome consists of two circular chromosomes of 2,961,146 bp and 1,072,314 bp that together encode 3,885 open reading frames. The vast majority of recognizable genes for essential cell functions (such as DNA replication, transcription, translation and cell-wall biosynthesis) and pathogenicity (for example, toxins, surface antigens and adhesins) are located on the large chromosome. In contrast, the small chromosome contains a larger fraction (59%) of hypothetical genes compared with the large chromosome (42%), and also contains many more genes that appear to have origins other than the g-Proteobacteria. The small chromosome also carries a gene capture system (the integron island) and host ‘addiction’ genes that are typically found on plasmids; thus, the small chromosome may have originally been a megaplasmid that was captured by an ancestral Vibrio species. The V. cholerae genomic sequence provides a starting point for understanding how a free-living, environmental organism emerged to become a significant human bacterial pathogen.

...read moreread less

1,785 citations

Journal Article•DOI•

Evidence for lateral gene transfer between Archaea and Bacteria from genome sequence of Thermotoga maritima

[...]

Karen E. Nelson¹, Rebecca A. Clayton¹, Steven R. Gill¹, Michelle L. Gwinn¹, Robert J. Dodson¹, Daniel H. Haft¹, Erin Hickey¹, Jeremy Peterson¹, William C. Nelson¹, Karen A. Ketchum¹, Lisa McDonald¹, Teresa Utterback¹, Joel A. Malek¹, Katja D. Linher¹, Mina M. Garrett¹, Ashley M. Stewart¹, Matthew D. Cotton¹, Matthew S. Pratt¹, Cheryl Phillips¹, Delwood Richardson¹, John F. Heidelberg¹, Granger G. Sutton¹, Robert D. Fleischmann¹, Jonathan A. Eisen¹, Owen White¹, Steven L. Salzberg¹, Hamilton O. Smith¹, J. Craig Venter¹, Claire M. Fraser¹ - Show less +25 more•Institutions (1)

J. Craig Venter Institute¹

27 May 1999-Nature

TL;DR: Genome analysis reveals numerous pathways involved in degradation of sugars and plant polysaccharides, and 108 genes that have orthologues only in the genomes of other thermophilic Eubacteria and Archaea.

...read moreread less

Abstract: The 1,860,725-base-pair genome of Thermotoga maritima MSB8 contains 1,877 predicted coding regions, 1,014 (54%) of which have functional assignments and 863 (46%) of which are of unknown function. Genome analysis reveals numerous pathways involved in degradation of sugars and plant polysaccharides, and 108 genes that have orthologues only in the genomes of other thermophilic Eubacteria and Archaea. Of the Eubacteria sequenced to date, T. maritima has the highest percentage (24%) of genes that are most similar to archaeal genes. Eighty-one archaeal-like genes are clustered in 15 regions of the T. maritima genome that range in size from 4 to 20 kilobases. Conservation of gene order between T. maritima and Archaea in many of the clustered regions suggests that lateral gene transfer may have occurred between thermophilic Eubacteria and Archaea.

...read moreread less

1,486 citations

Journal Article•DOI•

Complete Genome Sequence of a Virulent Isolate of Streptococcus pneumoniae

[...]

Hervé Tettelin¹, Karen E. Nelson¹, Ian T. Paulsen¹, Jonathan A. Eisen¹, Timothy D. Read¹, Scott N. Peterson¹, John F. Heidelberg¹, Robert T. DeBoy¹, Daniel H. Haft¹, Robert J. Dodson¹, Anthony S. Durkin¹, Michelle L. Gwinn¹, James F. Kolonay¹, William C. Nelson¹, Jeremy Peterson¹, Lowell Umayam¹, Owen White¹, Steven L. Salzberg¹, Matthew R. Lewis¹, Diana Radune¹, Erik Holtzapple¹, Hoda Khouri¹, Alex M. Wolf¹, T. Utterback¹, Cheryl L. Hansen¹, Lisa McDonald¹, Tamara Feldblyum¹, Samuel V. Angiuoli¹, T. Dickinson¹, Erin Hickey¹, Ingeborg Holt¹, Brendan J. Loftus¹, Fan Yang¹, Hamilton O. Smith¹, J. C. Venter¹, Brian Dougherty¹, Donald A. Morrison¹, S. K. Hollingshead¹, Claire M. Fraser¹ - Show less +35 more•Institutions (1)

J. Craig Venter Institute¹

20 Jul 2001-Science

TL;DR: A motif identified within the signal peptide of proteins is potentially involved in targeting these proteins to the cell surface of low–guanine/cytosine Gram-positive species.

...read moreread less

Abstract: The 2,160,837-base pair genome sequence of an isolate of Streptococcus pneumoniae, a Gram-positive pathogen that causes pneumonia, bacteremia, meningitis, and otitis media, contains 2236 predicted coding regions; of these, 1440 (64%) were assigned a biological role. Approximately 5% of the genome is composed of insertion sequences that may contribute to genome rearrangements through uptake of foreign DNA. Extracellular enzyme systems for the metabolism of polysaccharides and hexosamines provide a substantial source of carbon and nitrogen for S. pneumoniae and also damage host tissues and facilitate colonization. A motif identified within the signal peptide of proteins is potentially involved in targeting these proteins to the cell surface of low-guanine/cytosine (GC) Gram-positive species. Several surface-exposed proteins that may serve as potential vaccine candidates were identified. Comparative genome hybridization with DNA arrays revealed strain differences in S. pneumoniae that could contribute to differences in virulence and antigenicity.

...read moreread less

1,409 citations

1
2
3
4
…
5
6

Collapse

Cited by

PDF

Open Access

More filters

疟原虫var基因转换速率变化导致抗原变异[英]／Paul H, Robert P, Christodoulou Z, et al//Proc Natl Acad Sci U S A

[...]

宁北芳, 朱淮民

28 Jul 2005

TL;DR: PfPMP1）与感染红细胞、树突状组胞以及胎盘的单个或多个受体作用，在黏附及免疫逃避中起关键的作�ly.

...read moreread less

Abstract: 抗原变异可使得多种致病微生物易于逃避宿主免疫应答。表达在感染红细胞表面的恶性疟原虫红细胞表面蛋白1（PfPMP1）与感染红细胞、内皮细胞、树突状细胞以及胎盘的单个或多个受体作用，在黏附及免疫逃避中起关键的作用。每个单倍体基因组var基因家族编码约60种成员，通过启动转录不同的var基因变异体为抗原变异提供了分子基础。

...read moreread less

18,940 citations

Journal Article•DOI•

The sequence of the human genome.

[...]

J. Craig Venter¹, Mark Raymond Adams¹, Eugene W. Myers¹, Peter W. Li¹ +269 more•Institutions (12)

16 Feb 2001-Science

TL;DR: Comparative genomic analysis indicates vertebrate expansions of genes associated with neuronal function, with tissue-specific developmental regulation, and with the hemostasis and immune systems are indicated.

...read moreread less

Abstract: A 2.91-billion base pair (bp) consensus sequence of the euchromatic portion of the human genome was generated by the whole-genome shotgun sequencing method. The 14.8-billion bp DNA sequence was generated over 9 months from 27,271,853 high-quality sequence reads (5.11-fold coverage of the genome) from both ends of plasmid clones made from the DNA of five individuals. Two assembly strategies-a whole-genome assembly and a regional chromosome assembly-were used, each combining sequence data from Celera and the publicly funded genome effort. The public data were shredded into 550-bp segments to create a 2.9-fold coverage of those genome regions that had been sequenced, without including biases inherent in the cloning and assembly procedure used by the publicly funded group. This brought the effective coverage in the assemblies to eightfold, reducing the number and size of gaps in the final assembly over what would be obtained with 5.11-fold coverage. The two assembly strategies yielded very similar results that largely agree with independent mapping data. The assemblies effectively cover the euchromatic regions of the human chromosomes. More than 90% of the genome is in scaffold assemblies of 100,000 bp or more, and 25% of the genome is in scaffolds of 10 million bp or larger. Analysis of the genome sequence revealed 26,588 protein-encoding transcripts for which there was strong corroborating evidence and an additional approximately 12,000 computationally derived genes with mouse matches or other weak supporting evidence. Although gene-dense clusters are obvious, almost half the genes are dispersed in low G+C sequence separated by large tracts of apparently noncoding sequence. Only 1.1% of the genome is spanned by exons, whereas 24% is in introns, with 75% of the genome being intergenic DNA. Duplications of segmental blocks, ranging in size up to chromosomal lengths, are abundant throughout the genome and reveal a complex evolutionary history. Comparative genomic analysis indicates vertebrate expansions of genes associated with neuronal function, with tissue-specific developmental regulation, and with the hemostasis and immune systems. DNA sequence comparisons between the consensus sequence and publicly funded genome data provided locations of 2.1 million single-nucleotide polymorphisms (SNPs). A random pair of human haploid genomes differed at a rate of 1 bp per 1250 on average, but there was marked heterogeneity in the level of polymorphism across the genome. Less than 1% of all SNPs resulted in variation in proteins, but the task of determining which SNPs have functional consequences remains an open challenge.

...read moreread less

12,098 citations

Journal Article•DOI•

Genome sequencing in microfabricated high-density picolitre reactors

[...]

Marcel Margulies, Michael Egholm, William E. Altman, Said Attiya, Joel S. Bader, Lisa A. Bemben, Jan Berka, Michael S. Braverman, Yi-Ju Chen, Zhoutao Chen, Scott Dewell, Lei Du, J. M. Fierro, Xavier V. Gomes, Brian C. Godwin, Wen He, Scott Edward Helgesen, Chun Heen Ho, Gerard P. Irzyk, Szilveszter C. Jando, Maria L. I. Alenquer, Thomas P. Jarvie, Kshama B. Jirage, Jong-Bum Kim, James R. Knight, Janna R. Lanza, John H. Leamon, Steven Lefkowitz, Ming Lei, Jing Li, Kenton Lohman, Hong Lu, Vinod Makhijani, Keith Mcdade, Michael P. McKenna, Eugene W. Myers¹, Elizabeth Nickerson, John Nobile, Ramona Plant, Bernard P. Puc, Michael T. Ronan, George T. Roth, Gary J. Sarkis, Jan Fredrik Simons, John Simpson, Maithreyan Srinivasan, Karrie R. Tartaro, Alexander Tomasz², Kari A. Vogt, Greg A. Volkmer, Shally H. Wang, Yong Wang, Michael P. Weiner³, Pengguang Yu, Richard F. Begley, Jonathan M. Rothberg - Show less +52 more•Institutions (3)

University of California, Berkeley¹, Rockefeller University², Rothberg Institute For Childhood Diseases³

15 Sep 2005-Nature

TL;DR: A scalable, highly parallel sequencing system with raw throughput significantly greater than that of state-of-the-art capillary electrophoresis instruments with 96% coverage at 99.96% accuracy in one run of the machine is described.

...read moreread less

Abstract: The proliferation of large-scale DNA-sequencing projects in recent years has driven a search for alternative methods to reduce time and cost. Here we describe a scalable, highly parallel sequencing system with raw throughput significantly greater than that of state-of-the-art capillary electrophoresis instruments. The apparatus uses a novel fibre-optic slide of individual wells and is able to sequence 25 million bases, at 99% or better accuracy, in one four-hour run. To achieve an approximately 100-fold increase in throughput over current Sanger sequencing technology, we have developed an emulsion method for DNA amplification and an instrument for sequencing by synthesis using a pyrosequencing protocol optimized for solid support and picolitre-scale volumes. Here we show the utility, throughput, accuracy and robustness of this system by shotgun sequencing and de novo assembly of the Mycoplasma genitalium genome with 96% coverage at 99.96% accuracy in one run of the machine.

...read moreread less

8,434 citations

Journal Article•DOI•

Pilon: An Integrated Tool for Comprehensive Microbial Variant Detection and Genome Assembly Improvement

[...]

Bruce J. Walker¹, Thomas Abeel², Terrance Shea¹, Margaret Priest¹, Amr Abouelliel¹, Sharadha Sakthikumar¹, Christina A. Cuomo¹, Qiandong Zeng¹, Jennifer R. Wortman¹, Sarah Young¹, Ashlee M. Earl¹ - Show less +7 more•Institutions (2)

Broad Institute¹, Ghent University²

19 Nov 2014-PLOS ONE

TL;DR: Pilon is a fully automated, all-in-one tool for correcting draft assemblies and calling sequence variants of multiple sizes, including very large insertions and deletions, which is being used to improve the assemblies of thousands of new genomes and to identify variants from thousands of clinically relevant bacterial strains.

...read moreread less

Abstract: Advances in modern sequencing technologies allow us to generate sufficient data to analyze hundreds of bacterial genomes from a single machine in a single day. This potential for sequencing massive numbers of genomes calls for fully automated methods to produce high-quality assemblies and variant calls. We introduce Pilon, a fully automated, all-in-one tool for correcting draft assemblies and calling sequence variants of multiple sizes, including very large insertions and deletions. Pilon works with many types of sequence data, but is particularly strong when supplied with paired end data from two Illumina libraries with small e.g., 180 bp and large e.g., 3-5 Kb inserts. Pilon significantly improves draft genome assemblies by correcting bases, fixing mis-assemblies and filling gaps. For both haploid and diploid genomes, Pilon produces more contiguous genomes with fewer errors, enabling identification of more biologically relevant genes. Furthermore, Pilon identifies small variants with high accuracy as compared to state-of-the-art tools and is unique in its ability to accurately identify large sequence variants including duplications and resolve large insertions. Pilon is being used to improve the assemblies of thousands of new genomes and to identify variants from thousands of clinically relevant bacterial strains. Pilon is freely available as open source software.

...read moreread less

5,659 citations

Journal Article•DOI•

Versatile and open software for comparing large genomes

[...]

Stefan Kurtz¹, Adam M. Phillippy, Arthur L. Delcher, Michael E. Smoot², Martin Shumway, Corina Antonescu, Steven L. Salzberg - Show less +3 more•Institutions (2)

University of Hamburg¹, University of Virginia²

30 Jan 2004-Genome Biology

TL;DR: The newest version of MUMmer easily handles comparisons of large eukaryotic genomes at varying evolutionary distances, as demonstrated by applications to multiple genomes.

...read moreread less

Abstract: The newest version of MUMmer easily handles comparisons of large eukaryotic genomes at varying evolutionary distances, as demonstrated by applications to multiple genomes. Two new graphical viewing tools provide alternative ways to analyze genome alignments. The new system is the first version of MUMmer to be released as open-source software. This allows other developers to contribute to the code base and freely redistribute the code. The MUMmer sources are available at http://www.tigr.org/software/mummer.

...read moreread less

4,886 citations

1
2
3
4
…
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
66
67
68
69
70
71
72
73
74
75
76
77
78
79
80
81
82
83
84
85
86
87
88
89
90
91
92
93
94
95
96
97
98
99
100
101
102
103
104
105
106
107
108
109
110
111
112
113
114
115
116
117
118
119
120
121
122
123
124
125
126
127
128
129
130
131
132
133
134
135
136
137
138
139
140
141
142
143
144
145
146
147
148
149
150
151
152
153
154
155
156
157
158
159
160
161
162
163
164
165
166
167
168
169
170
171
172
173
174
175
176
177
178
179
180
181
182
183
184
185
186
187
188
189
190
191
192
193
194
195
196
197
198
199
200

Collapse