BACKGROUND: The fungal genus Aspergillus is of critical importance to humankind. Species include those with industrial applications, important pathogens of humans, animals and crops, a source of potent carcinogenic contaminants of food, and an important genetic model. The genome sequences of eight aspergilli have already been explored to investigate aspects of fungal biology, raising questions about evolution and specialization within this genus. RESULTS: We have generated genome sequences for ten novel, highly diverse Aspergillus species and compared these in detail to sister and more distant genera. Comparative studies of key aspects of fungal biology, including primary and secondary metabolism, stress response, biomass degradation, and signal transduction, revealed both conservation and diversity among the species. Observed genomic differences were validated with experimental studies. This revealed several highlights, such as the potential for sex in asexual species, organic acid production genes being a key feature of black aspergilli, alternative approaches for degrading plant biomass, and indications for the genetic basis of stress response. A genome-wide phylogenetic analysis demonstrated in detail the relationship of the newly genome sequenced species with other aspergilli. CONCLUSIONS: Many aspects of biological differences between fungal species cannot be explained by current knowledge obtained from genome sequences. The comparative genomics and experimental study, presented here, allows for the first time a genus-wide view of the biological diversity of the aspergilli and in many, but not all, cases linked genome differences to phenotype. Insights gained could be exploited for biotechnological and medical applications of fungi.
Halorubrum lacusprofundi is an extreme halophile within the archaeal phylum Euryarchaeota. The type strain ACAM 34 was isolated from Deep Lake, Antarctica. H. lacusprofundi is of phylogenetic interest because it is distantly related to the haloarchaea that have previously been sequenced. It is also of interest because of its psychrotolerance. We report here the complete genome sequence of H. lacusprofundi type strain ACAM 34 and its annotation. This genome is part of a 2006 Joint Genome Institute Community Sequencing Program project to sequence genomes of diverse Archaea.
The ecosystem roles of fungi have been extensively studied by targeting one organism and/or biological process at a time, but the full metabolic potential of fungi has rarely been captured in an environmental context. We hypothesized that fungal genome sequences could be assembled directly from the environment using metagenomics and that transcriptomics and proteomics could simultaneously reveal metabolic differentiation across habitats. We reconstructed the near-complete 27 Mbp genome of a filamentous fungus, Acidomyces richmondensis, and evaluated transcript and protein expression in floating and streamer biofilms from an acid mine drainage (AMD) system. A. richmondensis transcripts involved in denitrification and in the degradation of complex carbon sources (including cellulose) were up-regulated in floating biofilms, whereas central carbon metabolism and stress-related transcripts were significantly up-regulated in streamer biofilms. These findings suggest that the biofilm niches are distinguished by distinct carbon and nitrogen resource utilization, oxygen availability, and environmental challenges. An isolated A. richmondensis strain from this environment was used to validate the metagenomics-derived genome and confirm nitrous oxide production at pH 1. Overall, our analyses defined mechanisms of fungal adaptation and identified a functional shift related to different roles in carbon and nitrogen turnover for the same species of fungi growing in closely located but distinct biofilm niches.
Ascomycete yeasts are metabolically diverse, with great potential for biotechnology. Here, we report the comparative genome analysis of 29 taxonomically and biotechnologically important yeasts, including 16 newly sequenced. We identify a genetic code change, CUG-Ala, in Pachysolen tannophilus in the clade sister to the known CUG-Ser clade. Our well-resolved yeast phylogeny shows that some traits, such as methylotrophy, are restricted to single clades, whereas others, such as l-rhamnose utilization, have patchy phylogenetic distributions. Gene clusters, with variable organization and distribution, encode many pathways of interest. Genomics can predict some biochemical traits precisely, but the genomic basis of others, such as xylose utilization, remains unresolved. Our data also provide insight into early evolution of ascomycetes. We document the loss of H3K9me2/3 heterochromatin, the origin of ascomycete mating-type switching, and panascomycete synteny at the MAT locus. These data and analyses will facilitate the engineering of efficient biosynthetic and degradative pathways and gateways for genomic manipulation.
Spirochaeta caldaria Pohlschroeder et al. 1995 is an obligately anaerobic, spiral-shaped bacterium that is motile via periplasmic flagella. The type strain, H1(T), was isolated in 1990 from cyanobacterial mat samples collected at a freshwater hot spring in Oregon, USA, and is of interest because it enhances the degradation of cellulose when grown in co-culture with Clostridium thermocellum. Here we provide a taxonomic re-evaluation for S. caldaria based on phylogenetic analyses of 16S rRNA sequences and whole genomes, and propose the reclassification of S. caldaria and two other Spirochaeta species as members of the emended genus Treponema. Whereas genera such as Borrelia and Sphaerochaeta possess well-distinguished genomic features related to their divergent lifestyles, the physiological and functional genomic characteristics of Spirochaeta and Treponema appear to be intermixed and are of little taxonomic value. The 3,239,340 bp long genome of strain H1(T) with its 2,869 protein-coding and 59 RNA genes is a part of the G enomic E ncyclopedia of Bacteria and Archaea project.
The complete genome sequence of Methylomicrobium album strain BG8, a methane-oxidizing gammaproteobacterium isolated from freshwater, is reported. Aside from a conserved inventory of genes for growth on single-carbon compounds, M. album BG8 carries a range of gene inventories for additional carbon and nitrogen transformations but no genes for growth on multicarbon substrates or for N fixation.
Spirochaeta africana Zhilina et al. 1996 is an anaerobic, aerotolerant, spiral-shaped bacterium that is motile via periplasmic flagella. The type strain of the species, Z-7692(T), was isolated in 1993 or earlier from a bacterial bloom in the brine under the trona layer in a shallow lagoon of the alkaline equatorial Lake Magadi in Kenya. Here we describe the features of this organism, together with the complete genome sequence, and annotation. Considering the pending reclassification of S. caldaria to the genus Treponema, S. africana is only the second 'true' member of the genus Spirochaeta with a genome-sequenced type strain to be published. The 3,285,855 bp long genome of strain Z-7692(T) with its 2,817 protein-coding and 57 RNA genes is a part of the G enomic E ncyclopedia of B acteria and A rchaea project.
Coriobacterium glomerans Haas and Konig 1988, is the only species of the genus Coriobacterium, family Coriobacteriaceae, order Coriobacteriales, phylum Actinobacteria. The bacterium thrives as an endosymbiont of pyrrhocorid bugs, i.e. the red fire bug Pyrrhocoris apterus L. The rationale for sequencing the genome of strain PW2(T) is its endosymbiotic life style which is rare among members of Actinobacteria. Here we describe the features of this symbiont, together with the complete genome sequence and its annotation. This is the first complete genome sequence of a member of the genus Coriobacterium and the sixth member of the order Coriobacteriales for which complete genome sequences are now available. The 2,115,681 bp long single replicon genome with its 1,804 protein-coding and 54 RNA genes is part of the G enomic E ncyclopedia of Bacteria and Archaea project.
Holophaga foetida Liesack et al. 1995 is a member of the phylum Acidobacteria and is of interest for its ability to anaerobically degrade aromatic compounds and for its production of volatile sulfur compounds through a unique pathway. The genome of H. foetida strain TMBS4(T) is the first to be sequenced for a representative of the class Holophagae. Here we describe the features of this organism, together with the complete genome sequence (improved high quality draft), and annotation. The 4,127,237 bp long chromosome with its 3,615 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Niabella soli Weon et al. 2008 is a member of the Chitinophagaceae, a family within the class Sphingobacteriia that is poorly characterized at the genome level, thus far. N. soli strain JS13-8(T) is of interest for its ability to produce a variety of glycosyl hydrolases. The genome of N. soli strain JS13-8(T) is only the second genome sequence of a type strain from the family Chitinophagaceae to be published, and the first one from the genus Niabella. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 4,697,343 bp long chromosome with its 3,931 protein-coding and 49 RNA genes is a part of the Genomic Encyclopedia ofBacteria andArchaea project.
Halopiger xanaduensis is the type species of the genus Halopiger and belongs to the euryarchaeal family Halobacteriaceae. H. xanaduensis strain SH-6, which is designated as the type strain, was isolated from the sediment of a salt lake in Inner Mongolia, Lake Shangmatala. Like other members of the family Halobacteriaceae, it is an extreme halophile requiring at least 2.5 M salt for growth. We report here the sequencing and annotation of the 4,355,268 bp genome, which includes one chromosome and three plasmids. This genome is part of a Joint Genome Institute (JGI) Community Sequencing Program (CSP) project to sequence diverse haloarchaeal genomes.
Marinithermus hydrothermalis Sako et al. 2003 is the type species of the monotypic genus Marinithermus. M. hydrothermalis T1(T) was the first isolate within the phylum "Thermus-Deinococcus" to exhibit optimal growth under a salinity equivalent to that of sea water and to have an absolute requirement for NaCl for growth. M. hydrothermalis T1(T) is of interest because it may provide a new insight into the ecological significance of the aerobic, thermophilic decomposers in the circulation of organic compounds in deep-sea hydrothermal vent ecosystems. This is the first completed genome sequence of a member of the genus Marinithermus and the seventh sequence from the family Thermaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 2,269,167 bp long genome with its 2,251 protein-coding and 59 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Efficient lignin depolymerization is unique to the wood decay basidiomycetes, collectively referred to as white rot fungi. Phanerochaete chrysosporium simultaneously degrades lignin and cellulose, whereas the closely related species, Ceriporiopsis subvermispora, also depolymerizes lignin but may do so with relatively little cellulose degradation. To investigate the basis for selective ligninolysis, we conducted comparative genome analysis of C. subvermispora and P. chrysosporium. Genes encoding manganese peroxidase numbered 13 and five in C. subvermispora and P. chrysosporium, respectively. In addition, the C. subvermispora genome contains at least seven genes predicted to encode laccases, whereas the P. chrysosporium genome contains none. We also observed expansion of the number of C. subvermispora desaturase-encoding genes putatively involved in lipid metabolism. Microarray-based transcriptome analysis showed substantial up-regulation of several desaturase and MnP genes in wood-containing medium. MS identified MnP proteins in C. subvermispora culture filtrates, but none in P. chrysosporium cultures. These results support the importance of MnP and a lignin degradation mechanism whereby cleavage of the dominant nonphenolic structures is mediated by lipid peroxidation products. Two C. subvermispora genes were predicted to encode peroxidases structurally similar to P. chrysosporium lignin peroxidase and, following heterologous expression in Escherichia coli, the enzymes were shown to oxidize high redox potential substrates, but not Mn(2+). Apart from oxidative lignin degradation, we also examined cellulolytic and hemicellulolytic systems in both fungi. In summary, the C. subvermispora genetic inventory and expression patterns exhibit increased oxidoreductase potential and diminished cellulolytic capability relative to P. chrysosporium.
Sulfuricurvum kujiense Kodama and Watanabe 2004 is the type species of the monotypic genus Sulfuricurvum, which belongs to the family Helicobacteraceae in the class Epsilonproteobacteria. The species is of interest because it is frequently found in crude oil and oil sands where it utilizes various reduced sulfur compounds such as elemental sulfur, sulfide and thiosulfate as electron donors. Members of the species do not utilize sugars, organic acids or hydrocarbons as carbon and energy sources. This genome sequence represents the type strain of the only species in the genus Sulfuricurvum. The genome, which consists of a circular chromosome of 2,574,824 bp length and four plasmids of 118,585 bp, 71,513 bp, 51,014 bp, and 3,421 bp length, respectively, harboring a total of 2,879 protein-coding and 61 RNA genes and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Starkeya novella (Starkey 1934) Kelly et al. 2000 is a member of the family Xanthobacteraceae in the order 'Rhizobiales', which is thus far poorly characterized at the genome level. Cultures from this species are most interesting due to their facultatively chemolithoautotrophic lifestyle, which allows them to both consume carbon dioxide and to produce it. This feature makes S. novella an interesting model organism for studying the genomic basis of regulatory networks required for the switch between consumption and production of carbon dioxide, a key component of the global carbon cycle. In addition, S. novella is of interest for its ability to grow on various inorganic sulfur compounds and several C1-compounds such as methanol. Besides Azorhizobium caulinodans, S. novella is only the second species in the family Xanthobacteraceae with a completely sequenced genome of a type strain. The current taxonomic classification of this group is in significant conflict with the 16S rRNA data. The genomic data indicate that the physiological capabilities of the organism might have been underestimated. The 4,765,023 bp long chromosome with its 4,511 protein-coding and 52 RNA genes was sequenced as part of the DOE Joint Genome Institute Community Sequencing Program (CSP) 2008.
Saccharomonospora azurea Runmao et al. 1987 is a member of the genus Saccharomonospora, which is in the family Pseudonocardiaceae and thus far poorly characterized genomically. Members of the genus Saccharomonospora are of interest because they originate from diverse habitats, such as leaf litter, manure, compost, the surface of peat, and moist and over-heated grain, and may play a role in the primary degradation of plant material by attacking hemicellulose. Next to S. viridis, S. azurea is only the second member in the genus Saccharomonospora for which a completely sequenced type strain genome will be published. Here we describe the features of this organism, together with the complete genome sequence with project status 'Improved high quality draft', and the annotation. The 4,763,832 bp long chromosome with its 4,472 protein-coding and 58 RNA genes was sequenced as part of the DOE funded Community Sequencing Program (CSP) 2010 at the Joint Genome Institute (JGI).
Saccharomonospora marina Liu et al. 2010 is a member of the genus Saccharomonospora, in the family Pseudonocardiaceae that is poorly characterized at the genome level thus far. Members of the genus Saccharomonospora are of interest because they originate from diverse habitats, such as leaf litter, manure, compost, surface of peat, moist, over-heated grain, and ocean sediment, where they might play a role in the primary degradation of plant material by attacking hemicellulose. Organisms belonging to the genus are usually Gram-positive staining, non-acid fast, and classify among the actinomycetes. Here we describe the features of this organism, together with the complete genome sequence (permanent draft status), and annotation. The 5,965,593 bp long chromosome with its 5,727 protein-coding and 57 RNA genes was sequenced as part of the DOE funded Community Sequencing Program (CSP) 2010 at the Joint Genome Institute (JGI).
Saprospira grandis Gross 1911 is a member of the Saprospiraceae, a family in the class 'Sphingobacteria' that remains poorly characterized at the genomic level. The species is known for preying on other marine bacteria via 'ixotrophy'. S. grandis strain Sa g1 was isolated from decaying crab carapace in France and was selected for genome sequencing because of its isolated location in the tree of life. Only one type strain genome has been published so far from the Saprospiraceae, while the sequence of strain Sa g1 represents the second genome to be published from a non-type strain of S. grandis. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 4,495,250 bp long Improved-High-Quality draft of the genome with its 3,536 protein-coding and 62 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Paenibacillus sp.Y412MC10 was one of a number of organisms isolated from Obsidian Hot Spring, Yellowstone National Park, Montana, USA under permit from the National Park Service. The isolate was initially classified as a Geobacillus sp. Y412MC10 based on its isolation conditions and similarity to other organisms isolated from hot springs at Yellowstone National Park. Comparison of 16 S rRNA sequences within the Bacillales indicated that Geobacillus sp.Y412MC10 clustered with Paenibacillus species, and the organism was most closely related to Paenibacillus lautus. Lucigen Corp. prepared genomic DNA and the genome was sequenced, assembled, and annotated by the DOE Joint Genome Institute. The genome sequence was deposited at the NCBI in October 2009 (NC_013406). The genome of Paenibacillus sp. Y412MC10 consists of one circular chromosome of 7,121,665 bp with an average G+C content of 51.2%. Comparison to other Paenibacillus species shows the organism lacks nitrogen fixation, antibiotic production and social interaction genes reported in other paenibacilli. The Y412MC10 genome shows a high level of synteny and homology to the draft sequence of Paenibacillus sp. HGF5, an organism from the Human Microbiome Project (HMP) Reference Genomes. This, combined with genomic CAZyme analysis, suggests an intestinal, rather than environmental origin for Y412MC10.
Polynucleobacter necessarius subsp. asymbioticus strain QLW-P1DMWA-1(T) is a planktonic freshwater bacterium affiliated with the family Burkholderiaceae (class Betaproteobacteria). This strain is of interest because it represents a subspecies with cosmopolitan and ubiquitous distribution in standing freshwater systems. The 16S-23S ITS genotype represented by the sequenced strain comprised on average more than 10% of bacterioplankton in its home habitat. While all strains of the subspecies P. necessarius asymbioticus are free-living freshwater bacteria, strains belonging to the only other subspecies, P. necessarius subsp. necessarius are obligate endosymbionts of the ciliate Euplotes aediculatus. The two subspecies of P. necessarius are the instances of two closely related subspecies that differ in their lifestyle (free-living vs. obligate endosymbiont), and they are the only members of the genus Polynucleobacter with completely sequenced genomes. Here we describe the features of P. necessarius subsp. asymbioticus, together with the complete genome sequence and annotation. The 2,159,490 bp long chromosome with a total of 2,088 protein-coding and 48 RNA genes is the first completed genome sequence of the genus Polynucleobacter to be published and was sequenced as part of the DOE Joint Genome Institute Community Sequencing Program 2006.
Agaricus bisporus is the model fungus for the adaptation, persistence, and growth in the humic-rich leaf-litter environment. Aside from its ecological role, A. bisporus has been an important component of the human diet for over 200 y and worldwide cultivation of the "button mushroom" forms a multibillion dollar industry. We present two A. bisporus genomes, their gene repertoires and transcript profiles on compost and during mushroom formation. The genomes encode a full repertoire of polysaccharide-degrading enzymes similar to that of wood-decayers. Comparative transcriptomics of mycelium grown on defined medium, casing-soil, and compost revealed genes encoding enzymes involved in xylan, cellulose, pectin, and protein degradation are more highly expressed in compost. The striking expansion of heme-thiolate peroxidases and beta-etherases is distinctive from Agaricomycotina wood-decayers and suggests a broad attack on decaying lignin and related metabolites found in humic acid-rich environment. Similarly, up-regulation of these genes together with a lignolytic manganese peroxidase, multiple copper radical oxidases, and cytochrome P450s is consistent with challenges posed by complex humic-rich substrates. The gene repertoire and expression of hydrolytic enzymes in A. bisporus is substantially different from the taxonomically related ectomycorrhizal symbiont Laccaria bicolor. A common promoter motif was also identified in genes very highly expressed in humic-rich substrates. These observations reveal genetic and enzymatic mechanisms governing adaptation to the humic-rich ecological niche formed during plant degradation, further defining the critical role such fungi contribute to soil structure and carbon sequestration in terrestrial ecosystems. Genome sequence will expedite mushroom breeding for improved agronomic characteristics.
Serratia plymuthica AS13 is a plant-associated Gammaproteobacteria, isolated from rapeseed roots. It is of special interest because of its ability to inhibit fungal pathogens of rapeseed and to promote plant growth. The complete genome of S. plymuthica AS13 consists of a 5,442,549 bp circular chromosome. The chromosome contains 4,951 protein-coding genes, 87 tRNA genes and 7 rRNA operons. This genome was sequenced as part of the project entitled "Genomics of four rapeseed plant growth promoting bacteria with antagonistic effect on plant pathogens" within the 2010 DOE-JGI Community Sequencing Program (CSP2010).
A plant-associated member of the family Enterobacteriaceae, Serratia plymuthica strain AS12 was isolated from rapeseed roots. It is of scientific interest because it promotes plant growth and inhibits plant pathogens. The genome of S. plymuthica AS12 comprises a 5,443,009 bp long circular chromosome, which consists of 4,952 protein-coding genes, 87 tRNA genes and 7 rRNA operons. This genome was sequenced within the 2010 DOE-JGI Community Sequencing Program (CSP2010) as part of the project entitled "Genomics of four rapeseed plant growth promoting bacteria with antagonistic effect on plant pathogens".
Serratia plymuthica are plant-associated, plant beneficial species belonging to the family Enterobacteriaceae. The members of the genus Serratia are ubiquitous in nature and their life style varies from endophytic to free-living. S. plymuthica AS9 is of special interest for its ability to inhibit fungal pathogens of rapeseed and to promote plant growth. The genome of S. plymuthica AS9 comprises a 5,442,880 bp long circular chromosome that consists of 4,952 protein-coding genes, 87 tRNA genes and 7 rRNA operons. This genome is part of the project entitled "Genomics of four rapeseed plant growth promoting bacteria with antagonistic effect on plant pathogens" awarded through the 2010 DOE-JGI Community Sequencing Program (CSP2010).
Wallemia (Wallemiales, Wallemiomycetes) is a genus of xerophilic Fungi of uncertain phylogenetic position within Basidiomycota. Most commonly found as food contaminants, species of Wallemia have also been isolated from hypersaline environments. The ability to tolerate environments with reduced water activity is rare in Basidiomycota. We sequenced the genome of W. sebi in order to understand its adaptations for surviving in osmotically challenging environments, and we performed phylogenomic and ultrastructural analyses to address its systematic placement and reproductive biology. W. sebi has a compact genome (9.8 Mb), with few repeats and the largest fraction of genes with functional domains compared with other Basidiomycota. We applied several approaches to searching for osmotic stress-related proteins. In silico analyses identified 93 putative osmotic stress proteins; homology searches showed the HOG (High Osmolarity Glycerol) pathway to be mostly conserved. Despite the seemingly reduced genome, several gene family expansions and a high number of transporters (549) were found that also provide clues to the ability of W. sebi to colonize harsh environments. Phylogenetic analyses of a 71-protein dataset support the position of Wallemia as the earliest diverging lineage of Agaricomycotina, which is confirmed by septal pore ultrastructure that shows the septal pore apparatus as a variant of the Tremella-type. Mating type gene homologs were identified although we found no evidence of meiosis during conidiogenesis, suggesting there may be aspects of the life cycle of W. sebi that remain cryptic.
Desulfosporosinus species are sulfate-reducing bacteria belonging to the Firmicutes. Their genomes will give insights into the genetic repertoire and evolution of sulfate reducers typically thriving in terrestrial environments and able to degrade toluene (Desulfosporosinus youngiae), to reduce Fe(III) (Desulfosporosinus meridiei, Desulfosporosinus orientis), and to grow under acidic conditions (Desulfosporosinus acidiphilus).
Owenweeksia hongkongensis Lau et al. 2005 is the sole member of the monospecific genus Owenweeksia in the family Cryomorphaceae, a poorly characterized family at the genome level thus far. This family comprises seven genera within the class Flavobacteria. Family members are known to be psychrotolerant, rod-shaped and orange pigmented (beta-carotene), typical for Flavobacteria. For growth, seawater and complex organic nutrients are necessary. The genome of O. hongkongensis UST20020801(T) is only the second genome of a member of the family Cryomorphaceae whose sequence has been deciphered. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 4,000,057 bp long chromosome with its 3,518 protein-coding and 45 RNA genes is a part of the GenomicEncyclopedia ofBacteriaandArchaea project.
Gillisia limnaea Van Trappen et al. 2004 is the type species of the genus Gillisia, which is a member of the well characterized family Flavobacteriaceae. The genome of G. limnea R-8282(T) is the first sequenced genome (permanent draft) from a type strain of the genus Gillisia. Here we describe the features of this organism, together with the permanent-draft genome sequence and annotation. The 3,966,857 bp long chromosome (two scaffolds) with its 3,569 protein-coding and 51 RNA genes is a part of the GenomicEncyclopedia of Bacteria and Archaea project.
BACKGROUND: Softwood is the predominant form of land plant biomass in the Northern hemisphere, and is among the most recalcitrant biomass resources to bioprocess technologies. The white rot fungus, Phanerochaete carnosa, has been isolated almost exclusively from softwoods, while most other known white-rot species, including Phanerochaete chrysosporium, were mainly isolated from hardwoods. Accordingly, it is anticipated that P. carnosa encodes a distinct set of enzymes and proteins that promote softwood decomposition. To elucidate the genetic basis of softwood bioconversion by a white-rot fungus, the present study reports the P. carnosa genome sequence and its comparative analysis with the previously reported P. chrysosporium genome. RESULTS: P. carnosa encodes a complete set of lignocellulose-active enzymes. Comparative genomic analysis revealed that P. carnosa is enriched with genes encoding manganese peroxidase, and that the most divergent glycoside hydrolase families were predicted to encode hemicellulases and glycoprotein degrading enzymes. Most remarkably, P. carnosa possesses one of the largest P450 contingents (266 P450s) among the sequenced and annotated wood-rotting basidiomycetes, nearly double that of P. chrysosporium. Along with metabolic pathway modeling, comparative growth studies on model compounds and chemical analyses of decomposed wood components showed greater tolerance of P. carnosa to various substrates including coniferous heartwood. CONCLUSIONS: The P. carnosa genome is enriched with genes that encode P450 monooxygenases that can participate in extractives degradation, and manganese peroxidases involved in lignin degradation. The significant expansion of P450s in P. carnosa, along with differences in carbohydrate- and lignin-degrading enzymes, could be correlated to the utilization of heartwood and sapwood preparations from both coniferous and hardwood species.
We sequenced and compared the genomes of the Dothideomycete fungal plant pathogens Cladosporium fulvum (Cfu) (syn. Passalora fulva) and Dothistroma septosporum (Dse) that are closely related phylogenetically, but have different lifestyles and hosts. Although both fungi grow extracellularly in close contact with host mesophyll cells, Cfu is a biotroph infecting tomato, while Dse is a hemibiotroph infecting pine. The genomes of these fungi have a similar set of genes (70% of gene content in both genomes are homologs), but differ significantly in size (Cfu >61.1-Mb; Dse 31.2-Mb), which is mainly due to the difference in repeat content (47.2% in Cfu versus 3.2% in Dse). Recent adaptation to different lifestyles and hosts is suggested by diverged sets of genes. Cfu contains an alpha-tomatinase gene that we predict might be required for detoxification of tomatine, while this gene is absent in Dse. Many genes encoding secreted proteins are unique to each species and the repeat-rich areas in Cfu are enriched for these species-specific genes. In contrast, conserved genes suggest common host ancestry. Homologs of Cfu effector genes, including Ecp2 and Avr4, are present in Dse and induce a Cf-Ecp2- and Cf-4-mediated hypersensitive response, respectively. Strikingly, genes involved in production of the toxin dothistromin, a likely virulence factor for Dse, are conserved in Cfu, but their expression differs markedly with essentially no expression by Cfu in planta. Likewise, Cfu has a carbohydrate-degrading enzyme catalog that is more similar to that of necrotrophs or hemibiotrophs and a larger pectinolytic gene arsenal than Dse, but many of these genes are not expressed in planta or are pseudogenized. Overall, comparison of their genomes suggests that these closely related plant pathogens had a common ancestral host but since adapted to different hosts and lifestyles by a combination of differentiated gene content, pseudogenization, and gene regulation.
Leadbetterella byssophila Weon et al. 2005 is the type species of the genus Leadbetterella of the family Cytophagaceae in the phylum Bacteroidetes. Members of the phylum Bacteroidetes are widely distributed in nature, especially in aquatic environments. They are of special interest for their ability to degrade complex biopolymers. L. byssophila occupies a rather isolated position in the tree of life and is characterized by its ability to hydrolyze starch and gelatine, but not agar, cellulose or chitin. Here we describe the features of this organism, together with the complete genome sequence, and annotation. L. byssophila is already the 16(th) member of the family Cytophagaceae whose genome has been sequenced. The 4,059,653 bp long single replicon genome with its 3,613 protein-coding and 53 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Cellulophaga algicola Bowman 2000 belongs to the family Flavobacteriaceae within the phylum 'Bacteroidetes' and was isolated from Melosira collected from the Eastern Antarctic coastal zone. The species is of interest because its members produce a wide range of extracellular enzymes capable of degrading proteins and polysaccharides with temperature optima of 20-30 degrees C. This is the first completed genome sequence of a member of the genus Cellulophaga. The 4,888,353 bp long genome with its 4,285 protein-coding and 62 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
The genus Caldicellulosiruptor contains the most thermophilic, plant biomass-degrading bacteria isolated to date. Previously, genome sequences from three cellulolytic members of this genus were reported (C. saccharolyticus, C. bescii, and C. obsidiansis). To further explore the physiological and biochemical basis for polysaccharide degradation within this genus, five additional genomes were sequenced: C. hydrothermalis, C. kristjanssonii, C. kronotskyensis, C. lactoaceticus, and C. owensensis. Taken together, the seven completed and one draft-phase Caldicellulosiruptor genomes suggest that, while central metabolism is highly conserved, significant differences in glycoside hydrolase inventories and numbers of carbohydrate transporters exist, a finding which likely relates to variability observed in plant biomass degradation capacity.
Ktedonobacter racemifer corrig. Cavaletti et al. 2007 is the type species of the genus Ktedonobacter, which in turn is the type genus of the family Ktedonobacteraceae, the type family of the order Ktedonobacterales within the class Ktedonobacteria in the phylum 'Chloroflexi'. Although K. racemifer shares some morphological features with the actinobacteria, it is of special interest because it was the first cultivated representative of a deep branching unclassified lineage of otherwise uncultivated environmental phylotypes tentatively located within the phylum 'Chloroflexi'. The aerobic, filamentous, non-motile, spore-forming Gram-positive heterotroph was isolated from soil in Italy. The 13,661,586 bp long non-contiguous finished genome consists of ten contigs and is the first reported genome sequence from a member of the class Ktedonobacteria. With its 11,453 protein-coding and 87 RNA genes, it is the largest prokaryotic genome reported so far. It comprises a large number of over-represented COGs, particularly genes associated with transposons, causing the genetic redundancy within the genome being considerably larger than expected by chance. This work is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Thermomonospora curvata Henssen 1957 is the type species of the genus Thermomonospora. This genus is of interest because members of this clade are sources of new antibiotics, enzymes, and products with pharmacological activity. In addition, members of this genus participate in the active degradation of cellulose. This is the first complete genome sequence of a member of the family Thermomonosporaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 5,639,016 bp long genome with its 4,985 protein-coding and 76 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Nocardioides sp. strain JS614 grows on ethene and vinyl chloride (VC) as sole carbon and energy sources and is of interest for bioremediation and biocatalysis. Sequencing of the complete genome of JS614 provides insight into the genetic basis of alkene oxidation, supports ongoing research into the physiology and biochemistry of growth on ethene and VC, and provides biomarkers to facilitate detection of VC/ethene oxidizers in the environment. This is the first genome sequence from the genus Nocardioides and the first genome of a VC/ethene-oxidizing bacterium.
Haliscomenobacter hydrossis van Veen et al. 1973 is the type species of the genus Haliscomenobacter, which belongs to order "Sphingobacteriales". The species is of interest because of its isolated phylogenetic location in the tree of life, especially the so far genomically uncharted part of it, and because the organism grows in a thin, hardly visible hyaline sheath. Members of the species were isolated from fresh water of lakes and from ditch water. The genome of H. hydrossis is the first completed genome sequence reported from a member of the family "Saprospiraceae". The 8,771,651 bp long genome with its three plasmids of 92 kbp, 144 kbp and 164 kbp length contains 6,848 protein-coding and 60 RNA genes, and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Brown rot decay removes cellulose and hemicellulose from wood--residual lignin contributing up to 30% of forest soil carbon--and is derived from an ancestral white rot saprotrophy in which both lignin and cellulose are decomposed. Comparative and functional genomics of the "dry rot" fungus Serpula lacrymans, derived from forest ancestors, demonstrated that the evolution of both ectomycorrhizal biotrophy and brown rot saprotrophy were accompanied by reductions and losses in specific protein families, suggesting adaptation to an intercellular interaction with plant tissue. Transcriptome and proteome analysis also identified differences in wood decomposition in S. lacrymans relative to the brown rot Postia placenta. Furthermore, fungal nutritional mode diversification suggests that the boreal forest biome originated via genetic coevolution of above- and below-ground biota.
A large region of suppressed recombination surrounds the sex-determining locus of the self-fertile fungus Neurospora tetrasperma. This region encompasses nearly one-fifth of the N. tetrasperma genome and suppression of recombination is necessary for self-fertility. The similarity of the N. tetrasperma mating chromosome to plant and animal sex chromosomes and its recent origin (<5 MYA), combined with a long history of genetic and cytological research, make this fungus an ideal model for studying the evolutionary consequences of suppressed recombination. Here we compare genome sequences from two N. tetrasperma strains of opposite mating type to determine whether structural rearrangements are associated with the nonrecombining region and to examine the effect of suppressed recombination for the evolution of the genes within it. We find a series of three inversions encompassing the majority of the region of suppressed recombination and provide evidence for two different types of rearrangement mechanisms: the recently proposed mechanism of inversion via staggered single-strand breaks as well as ectopic recombination between transposable elements. In addition, we show that the N. tetrasperma mat a mating-type region appears to be accumulating deleterious substitutions at a faster rate than the other mating type (mat A) and thus may be in the early stages of degeneration.
Clostridium thermocellum DSM1313 is a thermophilic, anaerobic bacterium with some of the highest rates of cellulose hydrolysis reported. The complete genome sequence reveals a suite of carbohydrate-active enzymes and demonstrates a level of diversity at the species level distinguishing it from the type strain ATCC 27405.
Recent research has provided mechanistic insight into the important contributions of the gut microbiota to vertebrate biology, but questions remain about the evolutionary processes that have shaped this symbiosis. In the present study, we showed in experiments with gnotobiotic mice that the evolution of Lactobacillus reuteri with rodents resulted in the emergence of host specialization. To identify genomic events marking adaptations to the murine host, we compared the genome of the rodent isolate L. reuteri 100-23 with that of the human isolate L. reuteri F275, and we identified hundreds of genes that were specific to each strain. In order to differentiate true host-specific genome content from strain-level differences, comparative genome hybridizations were performed to query 57 L. reuteri strains originating from six different vertebrate hosts in combination with genome sequence comparisons of nine strains encompassing five phylogenetic lineages of the species. This approach revealed that rodent strains, although showing a high degree of genomic plasticity, possessed a specific genome inventory that was rare or absent in strains from other vertebrate hosts. The distinct genome content of L. reuteri lineages reflected the niche characteristics in the gastrointestinal tracts of their respective hosts, and inactivation of seven out of eight representative rodent-specific genes in L. reuteri 100-23 resulted in impaired ecological performance in the gut of mice. The comparative genomic analyses suggested fundamentally different trends of genome evolution in rodent and human L. reuteri populations, with the former possessing a large and adaptable pan-genome while the latter being subjected to a process of reductive evolution. In conclusion, this study provided experimental evidence and a molecular basis for the evolution of host specificity in a vertebrate gut symbiont, and it identified genomic events that have shaped this process.
BACKGROUND: Sinorhizobium meliloti is a model system for the studies of symbiotic nitrogen fixation. An extensive polymorphism at the genetic and phenotypic level is present in natural populations of this species, especially in relation with symbiotic promotion of plant growth. AK83 and BL225C are two nodule-isolated strains with diverse symbiotic phenotypes; BL225C is more efficient in promoting growth of the Medicago sativa plants than strain AK83. In order to investigate the genetic determinants of the phenotypic diversification of S. meliloti strains AK83 and BL225C, we sequenced the complete genomes for these two strains. RESULTS: With sizes of 7.14 Mbp and 6.97 Mbp, respectively, the genomes of AK83 and BL225C are larger than the laboratory strain Rm1021. The core genome of Rm1021, AK83, BL225C strains included 5124 orthologous groups, while the accessory genome was composed by 2700 orthologous groups. While Rm1021 and BL225C have only three replicons (Chromosome, pSymA and pSymB), AK83 has also two plasmids, 260 and 70 Kbp long. We found 65 interesting orthologous groups of genes that were present only in the accessory genome, consequently responsible for phenotypic diversity and putatively involved in plant-bacterium interaction. Notably, the symbiosis inefficient AK83 lacked several genes required for microaerophilic growth inside nodules, while several genes for accessory functions related to competition, plant invasion and bacteroid tropism were identified only in AK83 and BL225C strains. Presence and extent of polymorphism in regulons of transcription factors involved in symbiotic interaction were also analyzed. Our results indicate that regulons are flexible, with a large number of accessory genes, suggesting that regulons polymorphism could also be a key determinant in the variability of symbiotic performances among the analyzed strains. CONCLUSIONS: In conclusions, the extended comparative genomics approach revealed a variable subset of genes and regulons that may contribute to the symbiotic diversity.
Isosphaera pallida (ex Woronichin 1927) Giovannoni et al. 1995 is the type species of the genus Isosphaera. The species is of interest because it was the first heterotrophic bacterium known to be phototactic, and it occupies an isolated phylogenetic position within the Planctomycetaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first complete genome sequence of a member of the genus Isosphaera and the third of a member of the family Planctomycetaceae. The 5,472,964 bp long chromosome and the 56,340 bp long plasmid with a total of 3,763 protein-coding and 60 RNA genes are part of the Genomic Encyclopedia of Bacteria and Archaea project.
Bacteroides salanitronis Lan et al. 2006 is a species of the genus Bacteroides, which belongs to the family Bacteroidaceae. The species is of interest because it was isolated from the gut of a chicken and the growing awareness that the anaerobic microflora of the cecum is of benefit for the host and may impact poultry farming. The 4,308,663 bp long genome consists of a 4.24 Mbp chromosome and three plasmids (6 kbp, 19 kbp, 40 kbp) containing 3,737 protein-coding and 101 RNA genes and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Paludibacter propionicigenes Ueki et al. 2006 is the type species of the genus Paludibacter, which belongs to the family Porphyromonadaceae. The species is of interest because of the position it occupies in the tree of life where it can be found in close proximity to members of the genus Dysgonomonas. This is the first completed genome sequence of a member of the genus Paludibacter and the third sequence from the family Porphyromonadaceae. The 3,685,504 bp long genome with its 3,054 protein-coding and 64 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Truepera radiovictrix Albuquerque et al. 2005 is the type species of the genus Truepera within the phylum "Deinococcus/Thermus". T. radiovictrix is of special interest not only because of its isolated phylogenetic location in the order Deinococcales, but also because of its ability to grow under multiple extreme conditions in alkaline, moderately saline, and high temperature habitats. Of particular interest is the fact that, T. radiovictrix is also remarkably resistant to ionizing radiation, a feature it shares with members of the genus Deinococcus. This is the first completed genome sequence of a member of the family Trueperaceae and the fourth type strain genome sequence from a member of the order Deinococcales. The 3,260,398 bp long genome with its 2,994 protein-coding and 52 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Arthrobacter phenanthrenivorans is the type species of the genus, and is able to metabolize phenanthrene as a sole source of carbon and energy. A. phenanthrenivorans is an aerobic, non-motile, and Gram-positive bacterium, exhibiting a rod-coccus growth cycle which was originally isolated from a creosote polluted site in Epirus, Greece. Here we describe the features of this organism, together with the complete genome sequence, and annotation.
Chthoniobacter flavus Ellin428 is the first isolate from the class Spartobacteria of the bacterial phylum Verrucomicrobia. C. flavus Ellin428 can metabolize many of the saccharide components of plant biomass but is incapable of growth on amino acids or organic acids other than pyruvate.
"Pedosphaera parvula" Ellin514 is an aerobically grown verrucomicrobial isolate from pasture soil. It is one of the few cultured representatives of subdivision 3 of the phylum Verrucomicrobia. Members of this group are widespread in terrestrial environments.
Herpetosiphon aurantiacus Holt and Lewin 1968 is the type species of the genus Herpetosiphon, which in turn is the type genus of the family Herpetosiphonaceae, type family of the order Herpetosiphonales in the phylum Chloroflexi. H. aurantiacus cells are organized in filaments which can rapidly glide. The species is of interest not only because of its rather isolated position in the tree of life, but also because Herpetosiphon ssp. were identified as predators capable of facultative predation by a wolf pack strategy and of degrading the prey organisms by excreted hydrolytic enzymes. The genome of H. aurantiacus strain 114-95(T) is the first completely sequenced genome of a member of the family Herpetosiphonaceae. The 6,346,587 bp long chromosome and the two 339,639 bp and 99,204 bp long plasmids with a total of 5,577 protein-coding and 77 RNA genes was sequenced as part of the DOE Joint Genome Institute Program DOEM 2005.
Bacillus tusciae Bonjour & Aragno 1994 is a hydrogen-oxidizing, thermoacidophilic spore former that lives as a facultative chemolithoautotroph in solfataras. Although 16S rRNA gene sequencing was well established at the time of the initial description of the organism, 16S sequence data were not available and the strain was placed into the genus Bacillus based on limited chemotaxonomic information. Despite the now obvious misplacement of strain T2 as a member of the genus Bacillus in 16S rRNA-based phylogenetic trees, the misclassification remained uncorrected for many years, which was likely due to the extremely difficult, analysis-hampering cultivation conditions and poor growth rate of the strain. Here we provide a taxonomic re-evaluation of strain T2T (= DSM 2912 = NBRC 15312) and propose its reclassification as the type strain of a new species, Kyrpidia tusciae, and the type species of the new genus Kyrpidia, which is a sister-group of Alicyclobacillus. The family Alicyclobacillaceae da Costa and Rainey, 2010 is emended. The 3,384,766 bp genome with its 3,323 protein-coding and 78 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Bacteroides coprosuis Whitehead et al. 2005 belongs to the genus Bacteroides, which is a member of the family Bacteroidaceae. Members of the genus Bacteroides in general are known as beneficial protectors of animal guts against pathogenic microorganisms, and as contributors to the degradation of complex molecules such as polysaccharides. B. coprosuis itself was isolated from a manure storage pit of a swine facility, but has not yet been found in an animal host. The species is of interest solely because of its isolated phylogenetic location. The genome of B. coprosuis is already the 5(th) sequenced type strain genome from the genus Bacteroides. The 2,991,798 bp long genome with its 2,461 protein-coding and 78 RNA genes and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Weeksella virosa Holmes et al. 1987 is the sole member and type species of the genus Weeksella which belongs to the family Flavobacteriaceae of the phylum Bacteroidetes. Twenty-nine isolates, collected from clinical specimens provided the basis for the taxon description. While the species seems to be a saprophyte of the mucous membranes of healthy man and warm-blooded animals a causal relationship with disease has been reported in a few instances. Except for the ability to produce indole and to hydrolyze Tween and proteins such as casein and gelatin, this aerobic, non-motile, non-pigmented bacterial species is metabolically inert in most traditional biochemical tests. The 2,272,954 bp long genome with its 2,105 protein-coding and 76 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Cellulosilyticum lentocellum DSM 5427 is an anaerobic, endospore-forming member of the Firmicutes. We describe the complete genome sequence of this cellulose-degrading bacterium, which was originally isolated from estuarine sediment of a river that received both domestic and paper mill waste. Comparative genomics of cellulolytic clostridia will provide insight into factors that influence degradation rates.
Tsukamurella paurometabola corrig. (Steinhaus 1941) Collins et al. 1988 is the type species of the genus Tsukamurella, which is the type genus to the family Tsukamurellaceae. The species is not only of interest because of its isolated phylogenetic location, but also because it is a human opportunistic pathogen with some strains of the species reported to cause lung infection, lethal meningitis, and necrotizing tenosynovitis. This is the first completed genome sequence of a member of the genus Tsukamurella and the first genome sequence of a member of the family Tsukamurellaceae. The 4,479,724 bp long genome contains a 99,806 bp long plasmid and a total of 4,335 protein-coding and 56 RNA genes, and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
"Thioalkalivibrio sulfidophilus" HL-EbGr7 is an obligately chemolithoautotrophic, haloalkaliphilic sulfur-oxidizing bacterium (SOB) belonging to the Gammaproteobacteria. The strain was found to predominate a full-scale bioreactor, removing sulfide from biogas. Here we report the complete genome sequence of strain HL-EbGr7 and its annotation. The genome was sequenced within the Joint Genome Institute Community Sequencing Program, because of its relevance to the sustainable removal of sulfide from bio- and industrial waste gases.
Desulfobulbus propionicus Widdel 1981 is the type species of the genus Desulfobulbus, which belongs to the family Desulfobulbaceae. The species is of interest because of its great implication in the sulfur cycle in aquatic sediments, its large substrate spectrum and a broad versatility in using various fermentation pathways. The species was the first example of a pure culture known to disproportionate elemental sulfur to sulfate and sulfide. This is the first completed genome sequence of a member of the genus Desulfobulbus and the third published genome sequence from a member of the family Desulfobulbaceae. The 3,851,869 bp long genome with its 3,351 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Marivirga tractuosa (Lewin 1969) Nedashkovskaya et al. 2010 is the type species of the genus Marivirga, which belongs to the family Flammeovirgaceae. Members of this genus are of interest because of their gliding motility. The species is of interest because representative strains show resistance to several antibiotics, including gentamicin, kanamycin, neomycin, polymixin and streptomycin. This is the first complete genome sequence of a member of the family Flammeovirgaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 4,511,574 bp long chromosome and the 4,916 bp plasmid with their 3,808 protein-coding and 49 RNA genes are a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Cellulophaga lytica (Lewin 1969) Johansen et al. 1999 is the type species of the genus Cellulophaga, which belongs to the family Flavobacteriaceae within the phylum 'Bacteroidetes' and was isolated from marine beach mud in Limon, Costa Rica. The species is of biotechnological interest because its members produce a wide range of extracellular enzymes capable of degrading proteins and polysaccharides. After the genome sequence of Cellulophaga algicola this is the second completed genome sequence of a member of the genus Cellulophaga. The 3,765,936 bp long genome with its 3,303 protein-coding and 55 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Bacteroides helcogenes Benno et al. 1983 is of interest because of its isolated phylogenetic location and, although it has been found in pig feces and is known to be pathogenic for pigs, occurrence of this bacterium is rare and it does not cause significant damage in intensive animal husbandry. The genome of B. helcogenes P 36-108(T) is already the fifth completed and published type strain genome from the genus Bacteroides in the family Bacteroidaceae. The 3,998,906 bp long genome with its 3,353 protein-coding and 83 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Oceanithermus profundus Miroshnichenko et al. 2003 is the type species of the genus Oceanithermus, which belongs to the family Thermaceae. The genus currently comprises two species whose members are thermophilic and are able to reduce sulfur compounds and nitrite. The organism is adapted to the salinity of sea water, is able to utilize a broad range of carbohydrates, some proteinaceous substrates, organic acids and alcohols. This is the first completed genome sequence of a member of the genus Oceanithermus and the fourth sequence from the family Thermaceae. The 2,439,291 bp long genome with its 2,391 protein-coding and 54 RNA genes consists of one chromosome and a 135,351 bp long plasmid, and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Calditerrivibrio nitroreducens Iino et al. 2008 is the type species of the genus Calditerrivibrio. The species is of interest because of its important role in the nitrate cycle as nitrate reducer and for its isolated phylogenetic position in the Tree of Life. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the third complete genome sequence of a member of the family Deferribacteraceae. The 2,216,552 bp long genome with its 2,128 protein-coding and 50 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Pseudonocardia dioxanivorans CB1190 is the first bacterium reported to be capable of growth on the environmental contaminant 1,4-dioxane and the first member of the genus Pseudonocardia for which there is an annotated genome sequence. Preliminary analysis of the genome (chromosome and three plasmids) indicates that strain CB1190 possesses several multicomponent monooxygenases that could be involved in the aerobic degradation of 1,4-dioxane and other environmental contaminants.
Methylocystis sp. strain Rockwell (ATCC 49242) is an aerobic methane-oxidizing alphaproteobacterium isolated from an aquifer in southern California. Unlike most methanotrophs in the Methylocystaceae family, this strain has a single pmo operon encoding particulate methane monooxygenase but no evidence of the genes encoding soluble methane monooxygenase. This is the first reported genome sequence of a member of the Methylocystis species of the Methylocystaceae family in the order Rhizobiales.
Ruminococcus albus 7 is a highly cellulolytic ruminal bacterium that is a member of the phylum Firmicutes. Here, we describe the complete genome of this microbe. This genome will be useful for rumen microbiology and cellulosome biology and in biofuel production, as one of its major fermentation products is ethanol.
BACKGROUND: Chloroflexus aurantiacus is a thermophilic filamentous anoxygenic phototrophic (FAP) bacterium, and can grow phototrophically under anaerobic conditions or chemotrophically under aerobic and dark conditions. According to 16S rRNA analysis, Chloroflexi species are the earliest branching bacteria capable of photosynthesis, and Cfl. aurantiacus has been long regarded as a key organism to resolve the obscurity of the origin and early evolution of photosynthesis. Cfl. aurantiacus contains a chimeric photosystem that comprises some characters of green sulfur bacteria and purple photosynthetic bacteria, and also has some unique electron transport proteins compared to other photosynthetic bacteria. METHODS: The complete genomic sequence of Cfl. aurantiacus has been determined, analyzed and compared to the genomes of other photosynthetic bacteria. RESULTS: Abundant genomic evidence suggests that there have been numerous gene adaptations/replacements in Cfl. aurantiacus to facilitate life under both anaerobic and aerobic conditions, including duplicate genes and gene clusters for the alternative complex III (ACIII), auracyanin and NADH:quinone oxidoreductase; and several aerobic/anaerobic enzyme pairs in central carbon metabolism and tetrapyrroles and nucleic acids biosynthesis. Overall, genomic information is consistent with a high tolerance for oxygen that has been reported in the growth of Cfl. aurantiacus. Genes for the chimeric photosystem, photosynthetic electron transport chain, the 3-hydroxypropionate autotrophic carbon fixation cycle, CO2-anaplerotic pathways, glyoxylate cycle, and sulfur reduction pathway are present. The central carbon metabolism and sulfur assimilation pathways in Cfl. aurantiacus are discussed. Some features of the Cfl. aurantiacus genome are compared with those of the Roseiflexus castenholzii genome. Roseiflexus castenholzii is a recently characterized FAP bacterium and phylogenetically closely related to Cfl. aurantiacus. According to previous reports and the genomic information, perspectives of Cfl. aurantiacus in the evolution of photosynthesis are also discussed. CONCLUSIONS: The genomic analyses presented in this report, along with previous physiological, ecological and biochemical studies, indicate that the anoxygenic phototroph Cfl. aurantiacus has many interesting and certain unique features in its metabolic pathways. The complete genome may also shed light on possible evolutionary connections of photosynthesis.
Cellulosic biomass is an abundant and underused substrate for biofuel production. The inability of many microbes to metabolize the pentose sugars abundant within hemicellulose creates specific challenges for microbial biofuel production from cellulosic material. Although engineered strains of Saccharomyces cerevisiae can use the pentose xylose, the fermentative capacity pales in comparison with glucose, limiting the economic feasibility of industrial fermentations. To better understand xylose utilization for subsequent microbial engineering, we sequenced the genomes of two xylose-fermenting, beetle-associated fungi, Spathaspora passalidarum and Candida tenuis. To identify genes involved in xylose metabolism, we applied a comparative genomic approach across 14 Ascomycete genomes, mapping phenotypes and genotypes onto the fungal phylogeny, and measured genomic expression across five Hemiascomycete species with different xylose-consumption phenotypes. This approach implicated many genes and processes involved in xylose assimilation. Several of these genes significantly improved xylose utilization when engineered into S. cerevisiae, demonstrating the power of comparative methods in rapidly identifying genes for biomass conversion while reflecting on fungal ecology.
Victivallis vadensis ATCC BAA-548 represents the first cultured representative from the novel phylum Lentisphaerae, a deep-branching bacterial lineage. Few cultured bacteria from this phylum are known, and V. vadensis therefore represents an important organism for evolutionary studies. V. vadensis is a strictly anaerobic sugar-fermenting isolate from the human gastrointestinal tract.
Bacteria of the deeply branching phylum Verrucomicrobia are rarely cultured yet commonly detected in metagenomic libraries from aquatic, terrestrial, and intestinal environments. We have sequenced the genome of Opitutus terrae PB90-1, a fermentative anaerobe within this phylum, isolated from rice paddy soil and capable of propionate production from plant-derived polysaccharides.
Cellulomonas flavigena (Kellerman and McBeth 1912) Bergey et al. 1923 is the type species of the genus Cellulomonas of the actinobacterial family Cellulomonadaceae. Members of the genus Cellulomonas are of special interest for their ability to degrade cellulose and hemicellulose, particularly with regard to the use of biomass as an alternative energy source. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the genus Cellulomonas, and next to the human pathogen Tropheryma whipplei the second complete genome sequence within the actinobacterial family Cellulomonadaceae. The 4,123,179 bp long single replicon genome with its 3,735 protein-coding and 53 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Methanothermus fervidus Stetter 1982 is the type strain of the genus Methanothermus. This hyperthermophilic genus is of a thought to be endemic in Icelandic hot springs. M. fervidus was not only the first characterized organism with a maximal growth temperature (97 degrees C) close to the boiling point of water, but also the first archaeon in which a detailed functional analysis of its histone protein was reported and the first one in which the function of 2,3-cyclodiphosphoglycerate in thermoadaptation was characterized. Strain V24S(T) is of interest because of its very low substrate ranges, it grows only on H(2) + CO(2). This is the first completed genome sequence of the family Methanothermaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 1,243,342 bp long genome with its 1,311 protein-coding and 50 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Methanoplanus petrolearius Ollivier et al. 1998 is the type strain of the genus Methanoplanus. The strain was originally isolated from an offshore oil field from the Gulf of Guinea. Members of the genus Methanoplanus are of interest because they play an important role in the carbon cycle and also because of their significant contribution to the global warming by methane emission in the atmosphere. Like other archaea of the family Methanomicrobiales, the members of the genus Methanoplanus are able to use CO(2) and H(2) as a source of carbon and energy; acetate is required for growth and probably also serves as carbon source. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first complete genome sequence of a member of the family Methanomicrobiaceae and the sixth complete genome sequence from the order Methanomicrobiales. The 2,843,290 bp long genome with its 2,824 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Acidaminococcus fermentans (Rogosa 1969) is the type species of the genus Acidaminococcus, and is of phylogenetic interest because of its isolated placement in a genomically little characterized region of the Firmicutes. A. fermentans is known for its habitation of the gastrointestinal tract and its ability to oxidize trans-aconitate. Its anaerobic fermentation of glutamate has been intensively studied and will now be complemented by the genomic basis. The strain described in this report is a nonsporulating, nonmotile, Gram-negative coccus, originally isolated from a pig alimentary tract. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the family Acidaminococcaceae, and the 2,329,769 bp long genome with its 2,101 protein-coding and 81 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Aminobacterium colombiense Baena et al. 1999 is the type species of the genus Aminobacterium. This genus is of large interest because of its isolated phylogenetic location in the family Synergistaceae, its strictly anaerobic lifestyle, and its ability to grow by fermentation of a limited range of amino acids but not carbohydrates. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the second completed genome sequence of a member of the family Synergistaceae and the first genome sequence of a member of the genus Aminobacterium. The 1,980,592 bp long genome with its 1,914 protein-coding and 56 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Intrasporangium calvum Kalakoutskii et al. 1967 is the type species of the genus Intrasporangium, which belongs to the actinobacterial family Intrasporangiaceae. The species is a Gram-positive bacterium that forms a branching mycelium, which tends to break into irregular fragments. The mycelium of this strain may bear intercalary vesicles but does not contain spores. The strain described in this study is an airborne organism that was isolated from a school dining room in 1967. One particularly interesting feature of I. calvum is that the type of its menaquinone is different from all other representatives of the family Intrasporangiaceae. This is the first completed genome sequence from a member of the genus Intrasporangium and also the first sequence from the family Intrasporangiaceae. The 4,024,382 bp long genome with its 3,653 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Syntrophothermus lipocalidus Sekiguchi et al. 2000 is the type species of the genus Syntrophothermus. The species is of interest because of its strictly anaerobic lifestyle, its participation in the primary step of the degradation of organic maters, and for releasing products which serve as substrates for other microorganisms. It also contributes significantly to maintain a regular pH in its environment by removing the fatty acids through beta-oxidation. The strain is able to metabolize isobutyrate and butyrate, which are the substrate and the product of degradation of the substrate, respectively. This is the first complete genome sequence of a member of the genus Syntrophothermus and the second in the family Syntrophomonadaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 2,405,559 bp long genome with its 2,385 protein-coding and 55 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Xylanimonas cellulosilytica Rivas et al. 2003 is the type species of the genus Xylanimonas of the actinobacterial family Promicromonosporaceae. The species X. cellulosilytica is of interest because of its ability to hydrolyze cellulose and xylan. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the large family Promicromonosporaceae, and the 3,831,380 bp long genome (one chromosome plus an 88,604 bp long plasmid) with its 3485 protein-coding and 61 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Chitinophaga pinensis Sangkhobol and Skerman 1981 is the type strain of the species which is the type species of the rapidly growing genus Chitinophaga in the sphingobacterial family 'Chitinophagaceae'. Members of the genus Chitinophaga vary in shape between filaments and spherical bodies without the production of a fruiting body, produce myxospores, and are of special interest for their ability to degrade chitin. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the family 'Chitinophagaceae', and the 9,127,347 bp long single replicon genome with its 7,397 protein-coding and 95 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Ignisphaera aggregans Niederberger et al. 2006 is the type and sole species of genus Ignisphaera. This archaeal species is characterized by a coccoid-shape and is strictly anaerobic, moderately acidophilic, heterotrophic hyperthermophilic and fermentative. The type strain AQ1.S1(T) was isolated from a near neutral, boiling spring in Kuirau Park, Rotorua, New Zealand. This is the first completed genome sequence of the genus Ignisphaera and the fifth genome (fourth type strain) sequence in the family Desulfurococcaceae. The 1,875,953 bp long genome with its 2,009 protein-coding and 52 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Olsenella uli (Olsen et al. 1991) Dewhirst et al. 2001 is the type species of the genus Olsenella, which belongs to the actinobacterial family Coriobacteriaceae. The species is of interest because it is frequently isolated from dental plaque in periodontitis patients and can cause primary endodontic infection. The species is a Gram-positive, non-motile and non-sporulating bacterium. The strain described in this study was isolated from human gingival crevices. This is the first completed sequence of the genus Olsenella and the fifth sequence from a member of the family Coriobacteriaceae. The 2,051,896 bp long genome with its 1,795 protein-coding and 55 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Thermaerobacter marianensis Takai et al. 1999 is the type species of the genus Thermaerobacter, which belongs to the Clostridiales family Incertae Sedis XVII. The species is of special interest because T. marianensis is an aerobic, thermophilic marine bacterium, originally isolated from the deepest part in the western Pacific Ocean (Mariana Trench) at the depth of 10.897m. Interestingly, the taxonomic status of the genus has not been clarified until now. The genus Thermaerobacter may represent a very deep group within the Firmicutes or potentially a novel phylum. The 2,844,696 bp long genome with its 2,375 protein-coding and 60 RNA genes consists of one circular chromosome and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Sebaldella termitidis (Sebald 1962) Collins and Shah 1986, is the only species in the genus Sebaldella within the fusobacterial family 'Leptotrichiaceae'. The sole and type strain of the species was first isolated about 50 years ago from intestinal content of Mediterranean termites. The species is of interest for its very isolated phylogenetic position within the phylum Fusobacteria in the tree of life, with no other species sharing more than 90% 16S rRNA sequence similarity. The 4,486,650 bp long genome with its 4,210 protein-coding and 54 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Haliangium ochraceum Fudou et al. 2002 is the type species of the genus Haliangium in the myxococcal family 'Haliangiaceae'. Members of the genus Haliangium are the first halophilic myxobacterial taxa described. The cells of the species follow a multicellular lifestyle in highly organized biofilms, called swarms, they decompose bacterial and yeast cells as most myxobacteria do. The fruiting bodies contain particularly small coccoid myxospores. H. ochraceum encodes the first actin homologue identified in a bacterial genome. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the myxococcal suborder Nannocystineae, and the 9,446,314 bp long single replicon genome with its 6,898 protein-coding and 53 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Gordonia bronchialis Tsukamura 1971 is the type species of the genus. G. bronchialis is a human-pathogenic organism that has been isolated from a large variety of human tissues. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of the family Gordoniaceae. The 5,290,012 bp long genome with its 4,944 protein-coding and 55 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Denitrovibrio acetiphilus Myhr and Torsvik 2000 is the type species of the genus Denitrovibrio in the bacterial family Deferribacteraceae. It is of phylogenetic interest because there are only six genera described in the family Deferribacteraceae. D. acetiphilus was isolated as a representative of a population reducing nitrate to ammonia in a laboratory column simulating the conditions in off-shore oil recovery fields. When nitrate was added to this column undesirable hydrogen sulfide production was stopped because the sulfate reducing populations were superseded by these nitrate reducing bacteria. Here we describe the features of this marine, mesophilic, obligately anaerobic organism respiring by nitrate reduction, together with the complete genome sequence, and annotation. This is the second complete genome sequence of the order Deferribacterales and the class Deferribacteres, which is the sole class in the phylum Deferribacteres. The 3,222,077 bp genome with its 3,034 protein-coding and 51 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
'Thermobaculum terrenum' Botero et al. 2004 is the sole species within the proposed genus 'Thermobaculum'. Strain YNP1(T) is the only cultivated member of an acid tolerant, extremely thermophilic species belonging to a phylogenetically isolated environmental clone group within the phylum Chloroflexi. At present, the name 'Thermobaculum terrenum' is not yet validly published as it contravenes Rule 30 (3a) of the Bacteriological Code. The bacterium was isolated from a slightly acidic extreme thermal soil in Yellowstone National Park, Wyoming (USA). Depending on its final taxonomic allocation, this is likely to be the third completed genome sequence of a member of the class Thermomicrobia and the seventh type strain genome from the phylum Chloroflexi. The 3,101,581 bp long genome with its 2,872 protein-coding and 58 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Dethiosulfovibrio peptidovorans Magot et al. 1997 is the type species of the genus Dethiosulfovibrio of the family Synergistaceae in the recently created phylum Synergistetes. The strictly anaerobic, vibriod, thiosulfate-reducing bacterium utilizes peptides and amino acids, but neither sugars nor fatty acids. It was isolated from an offshore oil well where it was been reported to be involved in pitting corrosion of mild steel. Initially, this bacterium was described as a distant relative of the genus Thermoanaerobacter, but was not assigned to a genus, it was subsequently placed into the novel phylum Synergistetes. A large number of repeats in the genome sequence prevented an economically justifiable closure of the last gaps. This is only the third published genome from a member of the phylum Synergistetes. The 2,576,359 bp long genome consists of three contigs with 2,458 protein-coding and 59 RNA genes and is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Planctomyces limnophilus Hirsch and Muller 1986 belongs to the order Planctomycetales, which differs from other bacterial taxa by several distinctive features such as internal cell compartmentalization, multiplication by forming buds directly from the spherical, ovoid or pear-shaped mother cell and a cell wall which is stabilized by a proteinaceous layer rather than a peptidoglycan layer. Besides Pirellula staleyi, this is the second completed genome sequence of the family Planctomycetaceae. P. limnophilus is of interest because it differs from Pirellula by the presence of a stalk and its structure of fibril bundles, its cell shape and size, the formation of multicellular rosettes, low salt tolerance and red pigmented colonies. The 5,460,085 bp long genome with its 4,304 protein-coding and 66 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
BACKGROUND: Cupriavidus necator JMP134 is a Gram-negative beta-proteobacterium able to grow on a variety of aromatic and chloroaromatic compounds as its sole carbon and energy source. METHODOLOGY/PRINCIPAL FINDINGS: Its genome consists of four replicons (two chromosomes and two plasmids) containing a total of 6631 protein coding genes. Comparative analysis identified 1910 core genes common to the four genomes compared (C. necator JMP134, C. necator H16, C. metallidurans CH34, R. solanacearum GMI1000). Although secondary chromosomes found in the Cupriavidus, Ralstonia, and Burkholderia lineages are all derived from plasmids, analyses of the plasmid partition proteins located on those chromosomes indicate that different plasmids gave rise to the secondary chromosomes in each lineage. The C. necator JMP134 genome contains 300 genes putatively involved in the catabolism of aromatic compounds and encodes most of the central ring-cleavage pathways. This strain also shows additional metabolic capabilities towards alicyclic compounds and the potential for catabolism of almost all proteinogenic amino acids. This remarkable catabolic potential seems to be sustained by a high degree of genetic redundancy, most probably enabling this catabolically versatile bacterium with different levels of metabolic responses and alternative regulation necessary to cope with a challenging environment. From the comparison of Cupriavidus genomes, it is possible to state that a broad metabolic capability is a general trait for Cupriavidus genus, however certain specialization towards a nutritional niche (xenobiotics degradation, chemolithoautotrophy or symbiotic nitrogen fixation) seems to be shaped mostly by the acquisition of "specialized" plasmids. CONCLUSIONS/SIGNIFICANCE: The availability of the complete genome sequence for C. necator JMP134 provides the groundwork for further elucidation of the mechanisms and regulation of chloroaromatic compound biodegradation.
Coraliomargarita akajimensis Yoon et al. 2007 is the type species of the genus Coraliomargarita. C. akajimensis is an obligately aerobic, Gram-negative, non-spore-forming, non-motile, spherical bacterium that was isolated from seawater surrounding the hard coral Galaxea fascicularis. C. akajimensis is of special interest because of its phylogenetic position in a genomically under-studied area of the bacterial diversity. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the family Puniceicoccaceae. The 3,750,771 bp long genome with its 3,137 protein-coding and 55 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Alicyclobacillus acidocaldarius (Darland and Brock 1971) is the type species of the larger of the two genera in the bacillal family 'Alicyclobacillaceae'. A. acidocaldarius is a free-living and non-pathogenic organism, but may also be associated with food and fruit spoilage. Due to its acidophilic nature, several enzymes from this species have since long been subjected to detailed molecular and biochemical studies. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of the family 'Alicyclobacillaceae'. The 3,205,686 bp long genome (chromosome and three plasmids) with its 3,153 protein-coding and 82 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Vulcanisaeta distributa Itoh et al. 2002 belongs to the family Thermoproteaceae in the phylum Crenarchaeota. The genus Vulcanisaeta is characterized by a global distribution in hot and acidic springs. This is the first genome sequence from a member of the genus Vulcanisaeta and seventh genome sequence in the family Thermoproteaceae. The 2,374,137 bp long genome with its 2,544 protein-coding and 49 RNA genes is a part of the Genomic Encyclopedia of Bacteriaand Archaea project.
Spirochaeta smaragdinae Magot et al. 1998 belongs to the family Spirochaetaceae. The species is Gram-negative, motile, obligately halophilic and strictly anaerobic and is of interest because it is able to ferment numerous polysaccharides. S. smaragdinae is the only species of the family Spirochaetaceae known to reduce thiosulfate or element sulfur to sulfide. This is the first complete genome sequence in the family Spirochaetaceae. The 4,653,970 bp long genome with its 4,363 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Ferrimonas balearica Rossello-Mora et al. 1996 is the type species of the genus Ferrimonas, which belongs to the family Ferrimonadaceae within the Gammaproteobacteria. The species is a Gram-negative, motile, facultatively anaerobic, non spore-forming bacterium, which is of special interest because it is a chemoorganotroph and has a strictly respiratory metabolism with oxygen, nitrate, Fe(III)-oxyhydroxide, Fe(III)-citrate, MnO(2), selenate, selenite and thiosulfate as electron acceptors. This is the first completed genome sequence of a member of the genus Ferrimonas and also the first sequence from a member of the family Ferrimonadaceae. The 4,279,159 bp long genome with its 3,803 protein-coding and 144 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Streptosporangium roseum Crauch 1955 is the type strain of the species which is the type species of the genus Streptosporangium. The 'pinkish coiled Streptomyces-like organism with a spore case' was isolated from vegetable garden soil in 1955. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a member of the family Streptosporangiaceae, and the second largest microbial genome sequence ever deciphered. The 10,369,518 bp long genome with its 9421 protein-coding and 80 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Arcobacter nitrofigilis (McClung et al. 1983) Vandamme et al. 1991 is the type species of the genus Arcobacter in the family Campylobacteraceae within the Epsilonproteobacteria. The species was first described in 1983 as Campylobacter nitrofigilis [1] after its detection as a free-living, nitrogen-fixing Campylobacter species associated with Spartina alterniflora Loisel roots [2]. It is of phylogenetic interest because of its lifestyle as a symbiotic organism in a marine environment in contrast to many other Arcobacter species which are associated with warm-blooded animals and tend to be pathogenic. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a type stain of the genus Arcobacter. The 3,192,235 bp genome with its 3,154 protein-coding and 70 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Brachyspira murdochii Stanton et al. 1992 is a non-pathogenic, host-associated spirochete of the family Brachyspiraceae. Initially isolated from the intestinal content of a healthy swine, the 'group B spirochaetes' were first described as Serpulina murdochii. Members of the family Brachyspiraceae are of great phylogenetic interest because of the extremely isolated location of this family within the phylum 'Spirochaetes'. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a type strain of a member of the family Brachyspiraceae and only the second genome sequence from a member of the genus Brachyspira. The 3,241,804 bp long genome with its 2,893 protein-coding and 40 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Sphaerobacter thermophilus Demharter et al. 1989 is the sole and type species of the genus Sphaerobacter, which is the type genus of the family Sphaerobacteraceae, the order Sphaerobacterales and the subclass Sphaerobacteridae. Phylogenetically, it belongs to the genomically little studied class of the Thermomicrobia in the bacterial phylum Chloroflexi. Here, the genome of strain S 6022(T) is described which is an obligate aerobe that was originally isolated from an aerated laboratory-scale fermentor that was pulse fed with municipal sewage sludge. We describe the features of this organism, together with the complete genome and annotation. This is the first complete genome sequence of the thermomicrobial subclass Sphaerobacteridae, and the second sequence from the chloroflexal class Thermomicrobia. The 3,993,764 bp genome with its 3,525 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Aminomonas paucivorans Baena et al. 1999 is the type species of the genus Aminomonas, which belongs to the family Synergistaceae. The species is of interest because it is an asaccharolytic chemoorganotrophic bacterium which ferments quite a number of amino acids. This is the first finished genome sequence (with one gap in a rDNA region) of a member of the genus Aminomonas and the third sequence from the family Synergistaceae. The 2,630,120 bp long genome with its 2,433 protein-coding and 61 RNA genes is a part of the GenomicEncyclopedia ofBacteria andArchaea project.
The genus Conexibacter (Monciardini et al. 2003) represents the type genus of the family Conexibacteraceae (Stackebrandt 2005, emend. Zhi et al. 2009) with Conexibacter woesei as the type species of the genus. C. woesei is a representative of a deep evolutionary line of descent within the class Actinobacteria. Strain ID131577(T) was originally isolated from temperate forest soil in Gerenzano (Italy). Cells are small, short rods that are motile by peritrichous flagella. They may form aggregates after a longer period of growth and, then as a typical characteristic, an undulate structure is formed by self-aggregation of flagella with entangled bacterial cells. Here we describe the features of the organism, together with the complete sequence and annotation. The 6,359,369 bp long genome of C. woesei contains 5,950 protein-coding and 48 RNA genes and is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Ensifer (Sinorhizobium) medicae is an effective nitrogen fixing microsymbiont of a diverse range of annual Medicago (medic) species. Strain WSM419 is an aerobic, motile, non-spore forming, Gram-negative rod isolated from a M. murex root nodule collected in Sardinia, Italy in 1981. WSM419 was manufactured commercially in Australia as an inoculant for annual medics during 1985 to 1993 due to its nitrogen fixation, saprophytic competence and acid tolerance properties. Here we describe the basic features of this organism, together with the complete genome sequence, and annotation. This is the first report of a complete genome sequence for a microsymbiont of the group of annual medic species adapted to acid soils. We reveal that its genome size is 6,817,576 bp encoding 6,518 protein-coding genes and 81 RNA only encoding genes. The genome contains a chromosome of size 3,781,904 bp and 3 plasmids of size 1,570,951 bp, 1,245,408 bp and 219,313 bp. The smallest plasmid is a feature unique to this medic microsymbiont.
Haloterrigena turkmenica (Zvyagintseva and Tarasov 1987) Ventosa et al. 1999, comb. nov. is the type species of the genus Haloterrigena in the euryarchaeal family Halobacteriaceae. It is of phylogenetic interest because of the yet unclear position of the genera Haloterrigena and Natrinema within the Halobacteriaceae, which created some taxonomic problems historically. H. turkmenica, was isolated from sulfate saline soil in Turkmenistan, is a relatively fast growing, chemoorganotrophic, carotenoid-containing, extreme halophile, requiring at least 2 M NaCl for growth. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the genus Haloterrigena, but the eighth genome sequence from a member of the family Halobacteriaceae. The 5,440,782 bp genome (including six plasmids) with its 5,287 protein-coding and 63 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Segniliparus rotundus Butler 2005 is the type species of the genus Segniliparus, which is currently the only genus in the corynebacterial family Segniliparaceae. This family is of large interest because of a novel late-emerging genus-specific mycolate pattern. The type strain has been isolated from human sputum and is probably an opportunistic pathogen. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of the family Segniliparaceae. The 3,157,527 bp long genome with its 3,081 protein-coding and 52 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Meiothermus silvanus (Tenreiro et al. 1995) Nobre et al. 1996 belongs to a thermophilic genus whose members share relatively low degrees of 16S rRNA gene sequence similarity. Meiothermus constitutes an evolutionary lineage separate from members of the genus Thermus, from which they can generally be distinguished by their slightly lower temperature optima. M. silvanus is of special interest as it causes colored biofilms in the paper making industry and may thus be of economic importance as a biofouler. This is the second completed genome sequence of a member of the genus Meiothermus and only the third genome sequence to be published from a member of the family Thermaceae. The 3,721,669 bp long genome with its 3,667 protein-coding and 55 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Ilyobacter polytropus Stieb and Schink 1984 is the type species of the genus Ilyobacter, which belongs to the fusobacterial family Fusobacteriaceae. The species is of interest because its members are able to ferment quite a number of sugars and organic acids. I. polytropus has a broad versatility in using various fermentation pathways. Also, its members do not degrade poly-beta-hydroxybutyrate but only the monomeric 3-hydroxybutyrate. This is the first completed genome sequence of a member of the genus Ilyobacter and the second sequence from the family Fusobacteriaceae. The 3,132,314 bp long genome with its 2,934 protein-coding and 108 RNA genes consists of two chromosomes (2 and 1 Mbp long) and one plasmid, and is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Acetohalobium arabaticum Zhilina and Zavarzin 1990 is of special interest because of its physiology and its participation in the anaerobic C(1)-trophic chain in hypersaline environments. This is the first completed genome sequence of the family Halobacteroidaceae and only the second genome sequence in the order Halanaerobiales. The 2,469,596 bp long genome with its 2,353 protein-coding and 90 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Sulfurimonas autotrophica Inagaki et al. 2003 is the type species of the genus Sulfurimonas. This genus is of interest because of its significant contribution to the global sulfur cycle as it oxidizes sulfur compounds to sulfate and by its apparent habitation of deep-sea hydrothermal and marine sulfidic environments as potential ecological niche. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the second complete genome sequence of the genus Sulfurimonas and the 15(th) genome in the family Helicobacteraceae. The 2,153,198 bp long genome with its 2,165 protein-coding and 55 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Thermosphaera aggregans Huber et al. 1998 is the type species of the genus Thermosphaera, which comprises at the time of writing only one species. This species represents archaea with a hyperthermophilic, heterotrophic, strictly anaerobic and fermentative phenotype. The type strain M11TL(T) was isolated from a water-sediment sample of a hot terrestrial spring (Obsidian Pool, Yellowstone National Park, Wyoming). Here we describe the features of this organism, together with the complete genome sequence and annotation. The 1,316,595 bp long single replicon genome with its 1,410 protein-coding and 47 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Desulfohalobium retbaense (Ollivier et al. 1991) is the type species of the polyphyletic genus Desulfohalobium, which comprises, at the time of writing, two species and represents the family Desulfohalobiaceae within the Deltaproteobacteria. D. retbaense is a moderately halophilic sulfate-reducing bacterium, which can utilize H(2) and a limited range of organic substrates, which are incompletely oxidized to acetate and CO(2), for growth. The type strain HR(100) (T) was isolated from sediments of the hypersaline Retba Lake in Senegal. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a member of the family Desulfohalobiaceae. The 2,909,567 bp genome (one chromosome and a 45,263 bp plasmid) with its 2,552 protein-coding and 57 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Methylosinus trichosporium OB3b (for "oddball" strain 3b) is an obligate aerobic methane-oxidizing alphaproteobacterium that was originally isolated in 1970 by Roger Whittenbury and colleagues. This strain has since been used extensively to elucidate the structure and function of several key enzymes of methane oxidation, including both particulate and soluble methane monooxygenase (sMMO) and the extracellular copper chelator methanobactin. In particular, the catalytic properties of soluble methane monooxygenase from M. trichosporium OB3b have been well characterized in context with biodegradation of recalcitrant hydrocarbons, such as trichloroethylene. The sequence of the M. trichosporium OB3b genome is the first reported from a member of the Methylocystaceae family in the order Rhizobiales.
Rhodobacter capsulatus SB 1003 belongs to the group of purple nonsulfur bacteria. Its genome consists of a 3.7-Mb chromosome and a 133-kb plasmid. The genome encodes genes for photosynthesis, nitrogen fixation, utilization of xenobiotic organic substrates, and synthesis of polyhydroxyalkanoates. These features made it a favorite research tool for studying these processes. Here we report its complete genome sequence.
Nocardiopsis dassonvillei (Brocq-Rousseau 1904) Meyer 1976 is the type species of the genus Nocardiopsis, which in turn is the type genus of the family Nocardiopsaceae. This species is of interest because of its ecological versatility. Members of N. dassonvillei have been isolated from a large variety of natural habitats such as soil and marine sediments, from different plant and animal materials as well as from human patients. Moreover, representatives of the genus Nocardiopsis participate actively in biopolymer degradation. This is the first complete genome sequence in the family Nocardiopsaceae. Here we describe the features of this organism, together with the complete genome sequence and annotation. The 6,543,312 bp long genome consist of a 5.77 Mbp chromosome and a 0.78 Mbp plasmid and with its 5,570 protein-coding and 77 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Desulfarculus baarsii (Widdel 1981) Kuever et al. 2006 is the type and only species of the genus Desulfarculus, which represents the family Desulfarculaceae and the order Desulfarculales. This species is a mesophilic sulfate-reducing bacterium with the capability to oxidize acetate and fatty acids of up to 18 carbon atoms completely to CO(2). The acetyl-CoA/CODH (Wood-Ljungdahl) pathway is used by this species for the complete oxidation of carbon sources and autotrophic growth on formate. The type strain 2st14(T) was isolated from a ditch sediment collected near the University of Konstanz, Germany. This is the first completed genome sequence of a member of the order Desulfarculales. The 3,655,731 bp long single replicon genome with its 3,303 protein-coding and 52 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Meiothermus ruber (Loginova et al. 1984) Nobre et al. 1996 is the type species of the genus Meiothermus. This thermophilic genus is of special interest, as its members share relatively low degrees of 16S rRNA gene sequence similarity and constitute a separate evolutionary lineage from members of the genus Thermus, from which they can generally be distinguished by their slightly lower temperature optima. The temperature related split is in accordance with the chemotaxonomic feature of the polar lipids. M. ruber is a representative of the low-temperature group. This is the first completed genome sequence of the genus Meiothermus and only the third genome sequence to be published from a member of the family Thermaceae. The 3,097,457 bp long genome with its 3,052 protein-coding and 53 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Thermocrinis albus Eder and Huber 2002 is one of three species in the genus Thermocrinis in the family Aquificaceae. Members of this family have become of significant interest because of their involvement in global biogeochemical cycles in high-temperature ecosystems. This interest had already spurred several genome sequencing projects for members of the family. We here report the first completed genome sequence a member of the genus Thermocrinis and the first type strain genome from a member of the family Aquificaceae. The 1,500,577 bp long genome with its 1,603 protein-coding and 47 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Arcanobacterium haemolyticum (ex MacLean et al. 1946) Collins et al. 1983 is the type species of the genus Arcanobacterium, which belongs to the family Actinomycetaceae. The strain is of interest because it is an obligate parasite of the pharynx of humans and farm animal; occasionally, it causes pharyngeal or skin lesions. It is a Gram-positive, nonmotile and non-sporulating bacterium. The strain described in this study was isolated from infections amongst American soldiers of certain islands of the North and West Pacific. This is the first completed sequence of a member of the genus Arcanobacterium and the ninth type strain genome from the family Actinomycetaceae. The 1,986,154 bp long genome with its 1,821 protein-coding and 64 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Micrococcus luteus (NCTC2665, "Fleming strain") has one of the smallest genomes of free-living actinobacteria sequenced to date, comprising a single circular chromosome of 2,501,097 bp (G+C content, 73%) predicted to encode 2,403 proteins. The genome shows extensive synteny with that of the closely related organism, Kocuria rhizophila, from which it was taxonomically separated relatively recently. Despite its small size, the genome harbors 73 insertion sequence (IS) elements, almost all of which are closely related to elements found in other actinobacteria. An IS element is inserted into the rrs gene of one of only two rrn operons found in M. luteus. The genome encodes only four sigma factors and 14 response regulators, a finding indicative of adaptation to a rather strict ecological niche (mammalian skin). The high sensitivity of M. luteus to beta-lactam antibiotics may result from the presence of a reduced set of penicillin-binding proteins and the absence of a wblC gene, which plays an important role in the antibiotic resistance in other actinobacteria. Consistent with the restricted range of compounds it can use as a sole source of carbon for energy and growth, M. luteus has a minimal complement of genes concerned with carbohydrate transport and metabolism and its inability to utilize glucose as a sole carbon source may be due to the apparent absence of a gene encoding glucokinase. Uniquely among characterized bacteria, M. luteus appears to be able to metabolize glycogen only via trehalose and to make trehalose only via glycogen. It has very few genes associated with secondary metabolism. In contrast to most other actinobacteria, M. luteus encodes only one resuscitation-promoting factor (Rpf) required for emergence from dormancy, and its complement of other dormancy-related proteins is also much reduced. M. luteus is capable of long-chain alkene biosynthesis, which is of interest for advanced biofuel production; a three-gene cluster essential for this metabolism has been identified in the genome.
Archaeoglobus profundus (Burggraf et al. 1990) is a hyperthermophilic archaeon in the euryarchaeal class Archaeoglobi, which is currently represented by the single family Archaeoglobaceae, containing six validly named species and two strains ascribed to the genus 'Geoglobus' which is taxonomically challenged as the corresponding type species has no validly published name. All members were isolated from marine hydrothermal habitats and are obligate anaerobes. Here we describe the features of the organism, together with the complete genome sequence and annotation. This is the second completed genome sequence of a member of the class Archaeoglobi. The 1,563,423 bp genome with its 1,858 protein-coding and 52 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
BACKGROUND: Staphylothermus marinus is an anaerobic, sulfur-reducing peptide fermenter of the archaeal phylum Crenarchaeota. It is the third heterotrophic, obligate sulfur reducing crenarchaeote to be sequenced and provides an opportunity for comparative analysis of the three genomes. RESULTS: The 1.57 Mbp genome of the hyperthermophilic crenarchaeote Staphylothermus marinus has been completely sequenced. The main energy generating pathways likely involve 2-oxoacid:ferredoxin oxidoreductases and ADP-forming acetyl-CoA synthases. S. marinus possesses several enzymes not present in other crenarchaeotes including a sodium ion-translocating decarboxylase likely to be involved in amino acid degradation. S. marinus lacks sulfur-reducing enzymes present in the other two sulfur-reducing crenarchaeotes that have been sequenced -- Thermofilum pendens and Hyperthermus butylicus. Instead it has three operons similar to the mbh and mbx operons of Pyrococcus furiosus, which may play a role in sulfur reduction and/or hydrogen production. The two marine organisms, S. marinus and H. butylicus, possess more sodium-dependent transporters than T. pendens and use symporters for potassium uptake while T. pendens uses an ATP-dependent potassium transporter. T. pendens has adapted to a nutrient-rich environment while H. butylicus is adapted to a nutrient-poor environment, and S. marinus lies between these two extremes. CONCLUSION: The three heterotrophic sulfur-reducing crenarchaeotes have adapted to their habitats, terrestrial vs. marine, via their transporter content, and they have also adapted to environments with differing levels of nutrients. Despite the fact that they all use sulfur as an electron acceptor, they are likely to have different pathways for sulfur reduction.
Methanocorpusculum labreanum is a methanogen belonging to the order Methanomicrobiales within the archaeal kingdom Euryarchaeota. The type strain Z was isolated from surface sediments of Tar Pit Lake in the La Brea Tar Pits in Los Angeles, California. M. labreanum is of phylogenetic interest because at the time the sequencing project began only one genome had previously been sequenced from the order Methanomicrobiales. We report here the complete genome sequence of M. labreanum type strain Z and its annotation. This is part of a 2006 Joint Genome Institute Community Sequencing Program project to sequence genomes of diverse Archaea.
Methanoculleus marisnigri Romesser et al. 1981 is a methanogen belonging to the order Methanomicrobiales within the archaeal phylum Euryarchaeota. The type strain, JR1, was isolated from anoxic sediments of the Black Sea. M. marisnigri is of phylogenetic interest because at the time the sequencing project began only one genome had previously been sequenced from the order Methanomicrobiales. We report here the complete genome sequence of M. marisnigri type strain JR1 and its annotation. This is part of a Joint Genome Institute 2006 Community Sequencing Program to sequence genomes of diverse Archaea.
Staphylothermus marinus Fiala and Stetter 1986 belongs to the order Desulfurococcales within the archaeal phylum Crenarchaeota. S. marinus is a hyperthermophilic, sulfur-dependent, anaerobic heterotroph. Strain F1 was isolated from geothermally heated sediments at Vulcano, Italy, but S. marinus has also been isolated from a hydrothermal vent on the East Pacific Rise. We report the complete genome of S. marinus strain F1, the type strain of the species. This is the fifth reported complete genome sequence from the order Desulfurococcales.
Halorhabdus utahensis Waino et al. 2000 is the type species of the genus, which is of phylogenetic interest because of its location on one of the deepest branches within the very extensive euryarchaeal family Halobacteriaceae. H. utahensis is a free-living, motile, rod shaped to pleomorphic, Gram-negative archaeon, which was originally isolated from a sediment sample collected from the southern arm of Great Salt Lake, Utah, USA. When grown on appropriate media, H. utahensis can form polyhydroxybutyrate (PHB). Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the a member of halobacterial genus Halorhabdus, and the 3,116,795 bp long single replicon genome with its 3027 protein-coding and 48 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Acidimicrobium ferrooxidans (Clark and Norris 1996) is the sole and type species of the genus, which until recently was the only genus within the actinobacterial family Acidimicrobiaceae and in the order Acidomicrobiales. Rapid oxidation of iron pyrite during autotrophic growth in the absence of an enhanced CO(2) concentration is characteristic for A. ferrooxidans. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the order Acidomicrobiales, and the 2,158,157 bp long single replicon genome with its 2038 protein coding and 54 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Pirellula staleyi Schlesner and Hirsch 1987 is the type species of the genus Pirellula of the family Planctomycetaceae. Members of this pear- or teardrop-shaped bacterium show a clearly visible pointed attachment pole and can be distinguished from other Planctomycetes by a lack of true stalks. Strains closely related to the species have been isolated from fresh and brackish water, as well as from hypersaline lakes. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of the order Planctomyces and only the second sequence from the phylum Planctobacteria/Planctomycetes. The 6,196,199 bp long genome with its 4773 protein-coding and 49 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Catenulispora acidiphila Busti et al. 2006 is the type species of the genus Catenulispora, and is of interest because of the rather isolated phylogenetic location it occupies within the scarcely explored suborder Catenulisporineae of the order Actinomycetales. C. acidiphilia is known for its acidophilic, aerobic lifestyle, but can also grow scantly under anaerobic conditions. Under regular conditions, C. acidiphilia grows in long filaments of relatively short aerial hyphae with marked septation. It is a free living, non motile, Gram-positive bacterium isolated from a forest soil sample taken from a wooded area in Gerenzano, Italy. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first complete genome sequence of the actinobacterial family Catenulisporaceae, and the 10,467,782 bp long single replicon genome with its 9056 protein-coding and 69 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Atopobium parvulum (Weinberg et al. 1937) Collins and Wallbanks 1993 comb. nov. is the type strain of the species and belongs to the genomically yet unstudied Atopobium/Olsenella branch of the family Coriobacteriaceae. The species A. parvulum is of interest because its members are frequently isolated from the human oral cavity and are found to be associated with halitosis (oral malodor) but not with periodontitis. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the genus Atopobium, and the 1,543,805 bp long single replicon genome with its 1369 protein-coding and 49 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Desulfomicrobium baculatum is the type species of the genus Desulfomicrobium, which is the type genus of the family Desulfomicrobiaceae. It is of phylogenetic interest because of the isolated location of the family Desulfomicrobiaceae within the order Desulfovibrionales. D. baculatum strain X(T) is a Gram-negative, motile, sulfate-reducing bacterium isolated from water-saturated manganese carbonate ore. It is strictly anaerobic and does not require NaCl for growth, although NaCl concentrations up to 6% (w/v) are tolerated. The metabolism is respiratory or fermentative. In the presence of sulfate, pyruvate and lactate are incompletely oxidized to acetate and CO(2). Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a member of the deltaproteobacterial family Desulfomicrobiaceae, and this 3,942,657 bp long single replicon genome with its 3494 protein-coding and 72 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Kangiella koreensis (Yoon et al. 2004) is the type species of the genus and is of phylogenetic interest because of the very isolated location of the genus Kangiella in the gammaproteobacterial order Oceanospirillales. K. koreensis SW-125(T) is a Gram-negative, non-motile, non-spore-forming bacterium isolated from tidal flat sediments at Daepo Beach, Yellow Sea, Korea. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first completed genome sequence from the genus Kangiella and only the fourth genome from the order Oceanospirillales. This 2,852,073 bp long single replicon genome with its 2647 protein-coding and 48 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Pedobacter heparinus (Payza and Korn 1956) Steyn et al. 1998 comb. nov. is the type species of the rapidly growing genus Pedobacter within the family Sphingobacteriaceae of the phylum 'Bacteroidetes'. P. heparinus is of interest, because it was the first isolated strain shown to grow with heparin as sole carbon and nitrogen source and because it produces several enzymes involved in the degradation of mucopolysaccharides. All available data about this species are based on a sole strain that was isolated from dry soil. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first report on a complete genome sequence of a member of the genus Pedobacter, and the 5,167,383 bp long single replicon genome with its 4287 protein-coding and 54 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Organisms of the candidate phylum termite group 1 (TG1) are regularly encountered in termite hindguts but are present also in many other habitats. Here, we report the complete genome sequence (1.64 Mbp) of "Elusimicrobium minutum" strain Pei191(T), the first cultured representative of the TG1 phylum. We reconstructed the metabolism of this strictly anaerobic bacterium isolated from a beetle larva gut, and we discuss the findings in light of physiological data. E. minutum has all genes required for uptake and fermentation of sugars via the Embden-Meyerhof pathway, including several hydrogenases, and an unusual peptide degradation pathway comprising transamination reactions and leading to the formation of alanine, which is excreted in substantial amounts. The presence of genes encoding lipopolysaccharide biosynthesis and the presence of a pathway for peptidoglycan formation are consistent with ultrastructural evidence of a gram-negative cell envelope. Even though electron micrographs showed no cell appendages, the genome encodes many genes putatively involved in pilus assembly. We assigned some to a type II secretion system, but the function of 60 pilE-like genes remains unknown. Numerous genes with hypothetical functions, e.g., polyketide synthesis, nonribosomal peptide synthesis, antibiotic transport, and oxygen stress protection, indicate the presence of hitherto undiscovered physiological traits. Comparative analysis of 22 concatenated single-copy marker genes corroborated the status of "Elusimicrobia" (formerly TG1) as a separate phylum in the bacterial domain, which was so far based only on 16S rRNA sequence analysis.
Leptotrichia buccalis (Robin 1853) Trevisan 1879 is the type species of the genus, and is of phylogenetic interest because of its isolated location in the sparsely populated and neither taxonomically nor genomically adequately accessed family 'Leptotrichiaceae' within the phylum 'Fusobacteria'. Species of Leptotrichia are large, fusiform, non-motile, non-sporulating rods, which often populate the human oral flora. L. buccalis is anaerobic to aerotolerant, and saccharolytic. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first complete genome sequence of the order 'Fusobacteriales' and no more than the second sequence from the phylum 'Fusobacteria'. The 2,465,610 bp long single replicon genome with its 2306 protein-coding and 61 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Sanguibacter keddieii is the type species of the genus Sanguibacter, the only genus within the family of Sanguibacteraceae. Phylogenetically, this family is located in the neighborhood of the genus Oerskovia and the family Cellulomonadaceae within the actinobacterial suborder Micrococcineae. The strain described in this report was isolated from blood of apparently healthy cows. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the family Sanguibacteraceae, and the 4,253,413 bp long single replicon genome with its 3735 protein-coding and 70 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Anaerococcus prevotii (Foubert and Douglas 1948) Ezaki et al. 2001 is the type species of the genus, and is of phylogenetic interest because of its arguable assignment to the provisionally arranged family 'Peptostreptococcaceae'. A. prevotii is an obligate anaerobic coccus, usually arranged in clumps or tetrads. The strain, whose genome is described here, was originally isolated from human plasma; other strains of the species were also isolated from clinical specimen. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a member of the genus. Next to Finegoldia magna, A. prevotii is only the second species from the family 'Peptostreptococcaceae' for which a complete genome sequence is described. The 1,998,633 bp long genome (chromosome and one plasmid) with its 1852 protein-coding and 61 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Actinosynnema mirum Hasegawa et al. 1978 is the type species of the genus, and is of phylogenetic interest because of its central phylogenetic location in the Actino-synnemataceae, a rapidly growing family within the actinobacterial suborder Pseudo-nocardineae. A. mirum is characterized by its motile spores borne on synnemata and as a producer of nocardicin antibiotics. It is capable of growing aerobically and under a moderate CO(2) atmosphere. The strain is a Gram-positive, aerial and substrate mycelium producing bacterium, originally isolated from a grass blade collected from the Raritan River, New Jersey. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first complete genome sequence of a member of the family Actinosynnemataceae, and only the second sequence from the actinobacterial suborder Pseudonocardineae. The 8,248,144 bp long single replicon genome with its 7100 protein-coding and 77 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Beutenbergia cavernae (Groth et al. 1999) is the type species of the genus and is of phylogenetic interest because of its isolated location in the actinobacterial suborder Micrococcineae. B. cavernae HKI 0122(T) is a Gram-positive, non-motile, non-spore-forming bacterium isolated from a cave in Guangxi (China). B. cavernae grows best under aerobic conditions and shows a rod-coccus growth cycle. Its cell wall peptidoglycan contains the diagnostic L-lysine <-- L-glutamate interpeptide bridge. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first completed genome sequence from the poorly populated micrococcineal family Beutenbergiaceae, and this 4,669,183 bp long single replicon genome with its 4225 protein-coding and 53 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Dyadobacter fermentans (Chelius and Triplett, 2000) is the type species of the genus Dyadobacter. It is of phylogenetic interest because of its location in the Cytophagaceae, a very diverse family within the order 'Sphingobacteriales'. D. fermentans has a mainly respiratory metabolism, stains Gram-negative, is non-motile and oxidase and catalase positive. It is characterized by the production of cell filaments in aging cultures, a flexirubin-like pigment and its ability to ferment glucose, which is almost unique in the aerobically living members of this taxonomically difficult family. Here we describe the features of this organism, together with the complete genome sequence, and its annotation. This is the first complete genome sequence of the sphingobacterial genus Dyadobacter, and this 6,967,790 bp long single replicon genome with its 5804 protein-coding and 50 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Brachybacterium faecium Collins et al. 1988 is the type species of the genus, and is of phylogenetic interest because of its location in the Dermabacteraceae, a rather isolated family within the actinobacterial suborder Micrococcineae. B. faecium is known for its rod-coccus growth cycle and the ability to degrade uric acid. It grows aerobically or weakly anaerobically. The strain described in this report is a free-living, nonmotile, Gram-positive bacterium, originally isolated from poultry deep litter. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the actinobacterial family Dermabacteraceae, and the 3,614,992 bp long single replicon genome with its 3129 protein-coding and 69 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Many marine bacteria have evolved to grow optimally at either high (copiotrophic) or low (oligotrophic) nutrient concentrations, enabling different species to colonize distinct trophic habitats in the oceans. Here, we compare the genome sequences of two bacteria, Photobacterium angustum S14 and Sphingopyxis alaskensis RB2256, that serve as useful model organisms for copiotrophic and oligotrophic modes of life and specifically relate the genomic features to trophic strategy for these organisms and define their molecular mechanisms of adaptation. We developed a model for predicting trophic lifestyle from genome sequence data and tested >400,000 proteins representing >500 million nucleotides of sequence data from 126 genome sequences with metagenome data of whole environmental samples. When applied to available oceanic metagenome data (e.g., the Global Ocean Survey data) the model demonstrated that oligotrophs, and not the more readily isolatable copiotrophs, dominate the ocean's free-living microbial populations. Using our model, it is now possible to define the types of bacteria that specific ocean niches are capable of sustaining.
Halogeometricum borinquense Montalvo-Rodriguez et al. 1998 is the type species of the genus, and is of phylogenetic interest because of its distinct location between the halobacterial genera Haloquadratum and Halosarcina. H. borinquense requires extremely high salt (NaCl) concentrations for growth. It can not only grow aerobically but also anaerobically using nitrate as electron acceptor. The strain described in this report is a free-living, motile, pleomorphic, euryarchaeon, which was originally isolated from the solar salterns of Cabo Rojo, Puerto Rico. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the halobacterial genus Halogeometricum, and this 3,944,467 bp long six replicon genome with its 3937 protein-coding and 57 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Halothermothirx orenii is a strictly anaerobic thermohalophilic bacterium isolated from sediment of a Tunisian salt lake. It belongs to the order Halanaerobiales in the phylum Firmicutes. The complete sequence revealed that the genome consists of one circular chromosome of 2578146 bps encoding 2451 predicted genes. This is the first genome sequence of an organism belonging to the Haloanaerobiales. Features of both Gram positive and Gram negative bacteria were identified with the presence of both a sporulating mechanism typical of Firmicutes and a characteristic Gram negative lipopolysaccharide being the most prominent. Protein sequence analyses and metabolic reconstruction reveal a unique combination of strategies for thermophilic and halophilic adaptation. H. orenii can serve as a model organism for the study of the evolution of the Gram negative phenotype as well as the adaptation under thermohalophilic conditions and the development of biotechnological applications under conditions that require high temperatures and high salt concentrations.
Capnocytophaga ochracea (Prevot et al. 1956) Leadbetter et al. 1982 is the type species of the genus Capnocytophaga. It is of interest because of its location in the Flavobacteriaceae, a genomically not yet charted family within the order Flavobacteriales. The species grows as fusiform to rod shaped cells which tend to form clumps and are able to move by gliding. C. ochracea is known as a capnophilic (CO(2)-requiring) organism with the ability to grow under anaerobic as well as aerobic conditions (oxygen concentration larger than 15%), here only in the presence of 5% CO(2). Strain VPI 2845(T), the type strain of the species, is portrayed in this report as a gliding, Gram-negative bacterium, originally isolated from a human oral cavity. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first completed genome sequence from the flavobacterial genus Capnocytophaga, and the 2,612,925 bp long single replicon genome with its 2193 protein-coding and 59 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Cryptobacterium curtum Nakazawa etal. 1999 is the type species of the genus, and is of phylogenetic interest because of its very distant and isolated position within the family Coriobacteriaceae. C. curtum is an asaccharolytic, opportunistic pathogen with a typical occurrence in the oral cavity, involved in dental and oral infections like periodontitis, inflammations and abscesses. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the actinobacterial family Coriobacteriaceae, and this 1,617,804 bp long single replicon genome with its 1364 protein-coding and 58 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
The 6.10-Mb genome sequence of the aerobic chitin-digesting gliding bacterium Flavobacterium johnsoniae (phylum Bacteroidetes) is presented. F. johnsoniae is a model organism for studies of bacteroidete gliding motility, gene regulation, and biochemistry. The mechanism of F. johnsoniae gliding is novel, and genome analysis confirms that it does not involve well-studied motility organelles, such as flagella or type IV pili. The motility machinery is composed of Gld proteins in the cell envelope that are thought to comprise the "motor" and SprB, which is thought to function as a cell surface adhesin that is propelled by the motor. Analysis of the genome identified genes related to sprB that may encode alternative adhesins used for movement over different surfaces. Comparative genome analysis revealed that some of the gld and spr genes are found in nongliding bacteroidetes and may encode components of a novel protein secretion system. F. johnsoniae digests proteins, and 125 predicted peptidases were identified. F. johnsoniae also digests numerous polysaccharides, and 138 glycoside hydrolases, 9 polysaccharide lyases, and 17 carbohydrate esterases were predicted. The unexpected ability of F. johnsoniae to digest hemicelluloses, such as xylans, mannans, and xyloglucans, was predicted based on the genome analysis and confirmed experimentally. Numerous predicted cell surface proteins related to Bacteroides thetaiotaomicron SusC and SusD, which are likely involved in binding of oligosaccharides and transport across the outer membrane, were also identified. Genes required for synthesis of the novel outer membrane flexirubin pigments were identified by a combination of genome analysis and genetic experiments. Genes predicted to encode components of a multienzyme nonribosomal peptide synthetase were identified, as were novel aspects of gene regulation. The availability of techniques for genetic manipulation allows rapid exploration of the features identified for the polysaccharide-digesting gliding bacteroidete F. johnsoniae.
Vinyl chloride (VC) is a human carcinogen and widespread priority pollutant. Here we report the first, to our knowledge, complete genome sequences of microorganisms able to respire VC, Dehalococcoides sp. strains VS and BAV1. Notably, the respective VC reductase encoding genes, vcrAB and bvcAB, were found embedded in distinct genomic islands (GEIs) with different predicted integration sites, suggesting that these genes were acquired horizontally and independently by distinct mechanisms. A comparative analysis that included two previously sequenced Dehalococcoides genomes revealed a contextually conserved core that is interrupted by two high plasticity regions (HPRs) near the Ori. These HPRs contain the majority of GEIs and strain-specific genes identified in the four Dehalococcoides genomes, an elevated number of repeated elements including insertion sequences (IS), as well as 91 of 96 rdhAB, genes that putatively encode terminal reductases in organohalide respiration. Only three core rdhA orthologous groups were identified, and only one of these groups is supported by synteny. The low number of core rdhAB, contrasted with the high rdhAB numbers per genome (up to 36 in strain VS), as well as their colocalization with GEIs and other signatures for horizontal transfer, suggests that niche adaptation via organohalide respiration is a fundamental ecological strategy in Dehalococccoides. This adaptation has been exacted through multiple mechanisms of recombination that are mainly confined within HPRs of an otherwise remarkably stable, syntenic, streamlined genome among the smallest of any free-living microorganism.
Stackebrandtia nassauensis Labeda and Kroppenstedt (2005) is the type species of the genus Stackebrandtia, and a member of the actinobacterial family Glycomycetaceae. Stackebrandtia currently contains two species, which are differentiated from Glycomyces spp. by cellular fatty acid and menaquinone composition. Strain LLR-40K-21(T) is Gram-positive, aerobic, and nonmotile, with a branched substrate mycelium and on some media an aerial mycelium. The strain was originally isolated from a soil sample collected from a road side in Nassau, Bahamas. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first complete genome sequence of the actinobacterial suborder Glycomycineae. The 6,841,557 bp long single replicon genome with its 6487 protein-coding and 53 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Rhodothermus marinus Alfredsson et al. 1995 is the type species of the genus and is of phylogenetic interest because the Rhodothermaceae represent the deepest lineage in the phylum Bacteroidetes. R. marinus R-10(T) is a Gram-negative, non-motile, non-spore-forming bacterium isolated from marine hot springs off the coast of Iceland. Strain R-10(T) is strictly aerobic and requires slightly halophilic conditions for growth. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the genus Rhodothermus, and only the second sequence from members of the family Rhodothermaceae. The 3,386,737 bp genome (including a 125 kb plasmid) with its 2914 protein-coding and 48 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Streptobacillus moniliformis Levaditi et al. 1925 is the type and sole species of the genus Streptobacillus, and is of phylogenetic interest because of its isolated location in the sparsely populated and neither taxonomically nor genomically much accessed family 'Leptotrichiaceae' within the phylum Fusobacteria. The 'Leptotrichiaceae' have not been well characterized, genomically or taxonomically. S. moniliformis,is a Gram-negative, non-motile, pleomorphic bacterium and is the etiologic agent of rat bite fever and Haverhill fever. Strain 9901(T), the type strain of the species, was isolated from a patient with rat bite fever. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is only the second completed genome sequence of the order Fusobacteriales and no more than the third sequence from the phylum Fusobacteria. The 1,662,578 bp long chromosome and the 10,702 bp plasmid with a total of 1511 protein-coding and 55 RNA genes are part of the Genomic Encyclopedia of Bacteria and Archaea project.
Saccharomonospora viridis (Schuurmans et al. 1956) Nonomurea and Ohara 1971 is the type species of the genus Saccharomonospora which belongs to the family Pseudonocardiaceae. S. viridis is of interest because it is a Gram-negative organism classified among the usually Gram-positive actinomycetes. Members of the species are frequently found in hot compost and hay, and its spores can cause farmer's lung disease, bagassosis, and humidifier fever. Strains of the species S. viridis have been found to metabolize the xenobiotic pentachlorophenol (PCP). The strain described in this study has been isolated from peat-bog in Ireland. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the family Pseudonocardiaceae, and the 4,308,349 bp long single replicon genome with its 3906 protein-coding and 64 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Slackia heliotrinireducens (Lanigan 1983) Wade et al. 1999 is of phylogenetic interest because of its location in a genomically yet uncharted section of the family Coriobacteriaceae, within the deep branching Actinobacteria. Strain RHS 1(T) was originally isolated from the ruminal flora of a sheep. It is a proteolytic anaerobic coccus, able to reductively cleave pyrrolizidine alkaloids. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the genus Slackia, and the 3,165,038 bp long single replicon genome with its 2798 protein-coding and 60 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Jonesia denitrificans (Prevot 1961) Rocourt et al. 1987 is the type species of the genus Jonesia, and is of phylogenetic interest because of its isolated location in the actinobacterial suborder Micrococcineae. J. denitrificans is characterized by a typical coryneform morphology and is able to form irregular nonsporulating rods showing branched and club-like forms. Coccoid cells occur in older cultures. J. denitrificans is classified as a pathogenic organism for animals (vertebrates). The type strain whose genome is described here was originally isolated from cooked ox blood. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a member of the genus for which a complete genome sequence is described. The 2,749,646 bp long genome with its 2558 protein-coding and 71 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Eggerthella lenta (Eggerth 1935) Wade et al. 1999, emended Wurdemann et al. 2009 is the type species of the genus Eggerthella, which belongs to the actinobacterial family Coriobacteriaceae. E. lenta is a Gram-positive, non-motile, non-sporulating pathogenic bacterium that can cause severe bacteremia. The strain described in this study has been isolated from a rectal tumor in 1935. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of the genus Eggerthella, and the 3,632,260 bp long single replicon genome with its 3123 protein-coding and 58 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Kytococcus sedentarius (ZoBell and Upham 1944) Stackebrandt et al. 1995 is the type strain of the species, and is of phylogenetic interest because of its location in the Dermacoccaceae, a poorly studied family within the actinobacterial suborder Micrococcineae. Kytococcus sedentarius is known for the production of oligoketide antibiotics as well as for its role as an opportunistic pathogen causing valve endocarditis, hemorrhagic pneumonia, and pitted keratolysis. It is strictly aerobic and can only grow when several amino acids are provided in the medium. The strain described in this report is a free-living, nonmotile, Gram-positive bacterium, originally isolated from a marine environment. Here we describe the features of this organism, together with the complete genome sequence, and annotation. This is the first complete genome sequence of a member of the family Dermacoccaceae and the 2,785,024 bp long single replicon genome with its 2639 protein-coding and 64 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
Desulfotomaculum acetoxidans Widdel and Pfennig 1977 was one of the first sulfate-reducing bacteria known to grow with acetate as sole energy and carbon source. It is able to oxidize substrates completely to carbon dioxide with sulfate as the electron acceptor, which is reduced to hydrogen sulfide. All available data about this species are based on strain 5575(T), isolated from piggery waste in Germany. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence of a Desulfotomaculum species with validly published name. The 4,545,624 bp long single replicon genome with its 4370 protein-coding and 100 RNA genes is a part of the Genomic Encyclopedia of Bacteria and Archaea project.
Halomicrobium mukohataei (Ihara et al. 1997) Oren et al. 2002 is the type species of the genus Halomicrobium. It is of phylogenetic interest because of its isolated location within the large euryarchaeal family Halobacteriaceae. H. mukohataei is an extreme halophile that grows essentially aerobically, but can also grow anaerobically under a change of morphology and with nitrate as electron acceptor. The strain, whose genome is described in this report, is a free-living, motile, Gram-negative euryarchaeon, originally isolated from Salinas Grandes in Jujuy, Andes highlands, Argentina. Its genome contains three genes for the 16S rRNA that differ from each other by up to 9%. Here we describe the features of this organism, together with the complete genome sequence and annotation. This is the first completed genome sequence from the poorly populated genus Halomicrobium, and the 3,332,349 bp long genome (chromosome and one plasmid) with its 3416 protein-coding and 56 RNA genes is part of the Genomic Encyclopedia of Bacteria and Archaea project.
We report the complete genome of Thermofilum pendens, a deeply branching, hyperthermophilic member of the order Thermoproteales in the archaeal kingdom Crenarchaeota. T. pendens is a sulfur-dependent, anaerobic heterotroph isolated from a solfatara in Iceland. It is an extracellular commensal, requiring an extract of Thermoproteus tenax for growth, and the genome sequence reveals that biosynthetic pathways for purines, most amino acids, and most cofactors are absent. In fact, T. pendens has fewer biosynthetic enzymes than obligate intracellular parasites, although it does not display other features that are common among obligate parasites and thus does not appear to be in the process of becoming a parasite. It appears that T. pendens has adapted to life in an environment rich in nutrients. T. pendens was known previously to utilize peptides as an energy source, but the genome revealed a substantial ability to grow on carbohydrates. T. pendens is the first crenarchaeote and only the second archaeon found to have a transporter of the phosphotransferase system. In addition to fermentation, T. pendens may obtain energy from sulfur reduction with hydrogen and formate as electron donors. It may also be capable of sulfur-independent growth on formate with formate hydrogen lyase. Additional novel features are the presence of a monomethylamine:corrinoid methyltransferase, the first time that this enzyme has been found outside the Methanosarcinales, and the presence of a presenilin-related protein. The predicted highly expressed proteins do not include proteins encoded by housekeeping genes and instead include ABC transporters for carbohydrates and peptides and clustered regularly interspaced short palindromic repeat-associated proteins.
The candidate division Korarchaeota comprises a group of uncultivated microorganisms that, by their small subunit rRNA phylogeny, may have diverged early from the major archaeal phyla Crenarchaeota and Euryarchaeota. Here, we report the initial characterization of a member of the Korarchaeota with the proposed name, "Candidatus Korarchaeum cryptofilum," which exhibits an ultrathin filamentous morphology. To investigate possible ancestral relationships between deep-branching Korarchaeota and other phyla, we used whole-genome shotgun sequencing to construct a complete composite korarchaeal genome from enriched cells. The genome was assembled into a single contig 1.59 Mb in length with a G + C content of 49%. Of the 1,617 predicted protein-coding genes, 1,382 (85%) could be assigned to a revised set of archaeal Clusters of Orthologous Groups (COGs). The predicted gene functions suggest that the organism relies on a simple mode of peptide fermentation for carbon and energy and lacks the ability to synthesize de novo purines, CoA, and several other cofactors. Phylogenetic analyses based on conserved single genes and concatenated protein sequences positioned the korarchaeote as a deep archaeal lineage with an apparent affinity to the Crenarchaeota. However, the predicted gene content revealed that several conserved cellular systems, such as cell division, DNA replication, and tRNA maturation, resemble the counterparts in the Euryarchaeota. In light of the known composition of archaeal genomes, the Korarchaeota might have retained a set of cellular features that represents the ancestral archaeal form.
The Bacillus cereus group represents sporulating soil bacteria containing pathogenic strains which may cause diarrheic or emetic food poisoning outbreaks. Multiple locus sequence typing revealed a presence in natural samples of these bacteria of about 30 clonal complexes. Application of genomic methods to this group was however biased due to the major interest for representatives closely related to Bacillus anthracis. Albeit the most important food-borne pathogens were not yet defined, existing data indicate that they are scattered all over the phylogenetic tree. The preliminary analysis of the sequences of three genomes discussed in this paper narrows down the gaps in our knowledge of the B. cereus group. The strain NVH391-98 is a rare but particularly severe food-borne pathogen. Sequencing revealed that the strain should be a representative of a novel bacterial species, for which the name Bacillus cytotoxis or Bacillus cytotoxicus is proposed. This strain has a reduced genome size compared to other B. cereus group strains. Genome analysis revealed absence of sigma B factor and the presence of genes encoding diarrheic Nhe toxin, not detected earlier. The strain B. cereus F837/76 represents a clonal complex close to that of B. anthracis. Including F837/76, three such B. cereus strains had been sequenced. Alignment of genomes suggests that B. anthracis is their common ancestor. Since such strains often emerge from clinical cases, they merit a special attention. The third strain, KBAB4, is a typical facultative psychrophile generally found in soil. Phylogenic studies show that in nature it is the most active group in terms of gene exchange. Genomic sequence revealed high presence of extra-chromosomal genetic material (about 530kb) that may account for this phenomenon. Genes coding Nhe-like toxin were found on a big plasmid in this strain. This may indicate a potential mechanism of toxicity spread from the psychrophile strain community. The results of this genomic work and ecological compartments of different strains incite to consider a necessity of creating prophylactic vaccines against bacteria closely related to NVH391-98 and F837/76. Presumably developing of such vaccines can be based on the properties of non-pathogenic strains such as KBAB4 or ATCC14579 reported here or earlier. By comparing the protein coding genes of strains being sequenced in this project to others we estimate the shared proteome, or core genome, in the B. cereus group to be 3000+/-200 genes and the total proteome, or pan-genome, to be 20-25,000 genes.
Following birth, the breast-fed infant gastrointestinal tract is rapidly colonized by a microbial consortium often dominated by bifidobacteria. Accordingly, the complete genome sequence of Bifidobacterium longum subsp. infantis ATCC15697 reflects a competitive nutrient-utilization strategy targeting milk-borne molecules which lack a nutritive value to the neonate. Several chromosomal loci reflect potential adaptation to the infant host including a 43 kbp cluster encoding catabolic genes, extracellular solute binding proteins and permeases predicted to be active on milk oligosaccharides. An examination of in vivo metabolism has detected the hallmarks of milk oligosaccharide utilization via the central fermentative pathway using metabolomic and proteomic approaches. Finally, conservation of gene clusters in multiple isolates corroborates the genomic mechanism underlying milk utilization for this infant-associated phylotype.
Sulfur-oxidizing epsilonproteobacteria are common in a variety of sulfidogenic environments. These autotrophic and mixotrophic sulfur-oxidizing bacteria are believed to contribute substantially to the oxidative portion of the global sulfur cycle. In order to better understand the ecology and roles of sulfur-oxidizing epsilonproteobacteria, in particular those of the widespread genus Sulfurimonas, in biogeochemical cycles, the genome of Sulfurimonas denitrificans DSM1251 was sequenced. This genome has many features, including a larger size (2.2 Mbp), that suggest a greater degree of metabolic versatility or responsiveness to the environment than seen for most of the other sequenced epsilonproteobacteria. A branched electron transport chain is apparent, with genes encoding complexes for the oxidation of hydrogen, reduced sulfur compounds, and formate and the reduction of nitrate and oxygen. Genes are present for a complete, autotrophic reductive citric acid cycle. Many genes are present that could facilitate growth in the spatially and temporally heterogeneous sediment habitat from where Sulfurimonas denitrificans was originally isolated. Many resistance-nodulation-development family transporter genes (10 total) are present; of these, several are predicted to encode heavy metal efflux transporters. An elaborate arsenal of sensory and regulatory protein-encoding genes is in place, as are genes necessary to prevent and respond to oxidative stress.
Prochlorococcus is a marine cyanobacterium that numerically dominates the mid-latitude oceans and is the smallest known oxygenic phototroph. Numerous isolates from diverse areas of the world's oceans have been studied and shown to be physiologically and genetically distinct. All isolates described thus far can be assigned to either a tightly clustered high-light (HL)-adapted clade, or a more divergent low-light (LL)-adapted group. The 16S rRNA sequences of the entire Prochlorococcus group differ by at most 3%, and the four initially published genomes revealed patterns of genetic differentiation that help explain physiological differences among the isolates. Here we describe the genomes of eight newly sequenced isolates and combine them with the first four genomes for a comprehensive analysis of the core (shared by all isolates) and flexible genes of the Prochlorococcus group, and the patterns of loss and gain of the flexible genes over the course of evolution. There are 1,273 genes that represent the core shared by all 12 genomes. They are apparently sufficient, according to metabolic reconstruction, to encode a functional cell. We describe a phylogeny for all 12 isolates by subjecting their complete proteomes to three different phylogenetic analyses. For each non-core gene, we used a maximum parsimony method to estimate which ancestor likely first acquired or lost each gene. Many of the genetic differences among isolates, especially for genes involved in outer membrane synthesis and nutrient transport, are found within the same clade. Nevertheless, we identified some genes defining HL and LL ecotypes, and clades within these broad ecotypes, helping to demonstrate the basis of HL and LL adaptations in Prochlorococcus. Furthermore, our estimates of gene gain events allow us to identify highly variable genomic islands that are not apparent through simple pairwise comparisons. These results emphasize the functional roles, especially those connected to outer membrane synthesis and transport that dominate the flexible genome and set it apart from the core. Besides identifying islands and demonstrating their role throughout the history of Prochlorococcus, reconstruction of past gene gains and losses shows that much of the variability exists at the "leaves of the tree," between the most closely related strains. Finally, the identification of core and flexible genes from this 12-genome comparison is largely consistent with the relative frequency of Prochlorococcus genes found in global ocean metagenomic databases, further closing the gap between our understanding of these organisms in the lab and the wild.
Thermobifida fusca is a moderately thermophilic soil bacterium that belongs to Actinobacteria. It is a major degrader of plant cell walls and has been used as a model organism for the study of secreted, thermostable cellulases. The complete genome sequence showed that T. fusca has a single circular chromosome of 3,642,249 bp predicted to encode 3,117 proteins and 65 RNA species with a coding density of 85%. Genome analysis revealed the existence of 29 putative glycoside hydrolases in addition to the previously identified cellulases and xylanases. The glycosyl hydrolases include enzymes predicted to exhibit mainly dextran/starch- and xylan-degrading functions. T. fusca possesses two protein secretion systems: the sec general secretion system and the twin-arginine translocation system. Several of the secreted cellulases have sequence signatures indicating their secretion may be mediated by the twin-arginine translocation system. T. fusca has extensive transport systems for import of carbohydrates coupled to transcriptional regulators controlling the expression of the transporters and glycosylhydrolases. In addition to providing an overview of the physiology of a soil actinomycete, this study presents insights on the transcriptional regulation and secretion of cellulases which may facilitate the industrial exploitation of these systems.
Soil bacteria that also form mutualistic symbioses in plants encounter two major levels of selection. One occurs during adaptation to and survival in soil, and the other occurs in concert with host plant speciation and adaptation. Actinobacteria from the genus Frankia are facultative symbionts that form N(2)-fixing root nodules on diverse and globally distributed angiosperms in the "actinorhizal" symbioses. Three closely related clades of Frankia sp. strains are recognized; members of each clade infect a subset of plants from among eight angiosperm families. We sequenced the genomes from three strains; their sizes varied from 5.43 Mbp for a narrow host range strain (Frankia sp. strain HFPCcI3) to 7.50 Mbp for a medium host range strain (Frankia alni strain ACN14a) to 9.04 Mbp for a broad host range strain (Frankia sp. strain EAN1pec.) This size divergence is the largest yet reported for such closely related soil bacteria (97.8%-98.9% identity of 16S rRNA genes). The extent of gene deletion, duplication, and acquisition is in concert with the biogeographic history of the symbioses and host plant speciation. Host plant isolation favored genome contraction, whereas host plant diversification favored genome expansion. The results support the idea that major genome expansions as well as reductions can occur in facultative symbiotic soil bacteria as they respond to new environments in the context of their symbioses.
We report here a comparative analysis of the genome sequence of Methanosarcina barkeri with those of Methanosarcina acetivorans and Methanosarcina mazei. The genome of M. barkeri is distinguished by having an organization that is well conserved with respect to the other Methanosarcina spp. in the region proximal to the origin of replication, with interspecies gene similarities as high as 95%. However, it is disordered and marked by increased transposase frequency and decreased gene synteny and gene density in the distal semigenome. Of the 3,680 open reading frames (ORFs) in M. barkeri, 746 had homologs with better than 80% identity to both M. acetivorans and M. mazei, while 128 nonhypothetical ORFs were unique (nonorthologous) among these species, including a complete formate dehydrogenase operon, genes required for N-acetylmuramic acid synthesis, a 14-gene gas vesicle cluster, and a bacterial-like P450-specific ferredoxin reductase cluster not previously observed or characterized for this genus. A cryptic 36-kbp plasmid sequence that contains an orc1 gene flanked by a presumptive origin of replication consisting of 38 tandem repeats of a 143-nucleotide motif was detected in M. barkeri. Three-way comparison of these genomes reveals differing mechanisms for the accrual of changes. Elongation of the relatively large M. acetivorans genome is the result of uniformly distributed multiple gene scale insertions and duplications, while the M. barkeri genome is characterized by localized inversions associated with the loss of gene content. In contrast, the short M. mazei genome most closely approximates the putative ancestral organizational state of these species.