WorldWideScience

Sample records for saccharomyces genome database

  1. Saccharomyces genome database informs human biology

    OpenAIRE

    Skrzypek, Marek S; Nash, Robert S; Wong, Edith D; MacPherson, Kevin A; Hellerstedt, Sage T; Engel, Stacia R; Karra, Kalpana; Weng, Shuai; Sheppard, Travis K; Binkley, Gail; Simison, Matt; Miyasato, Stuart R; Cherry, J Michael

    2017-01-01

    Abstract The Saccharomyces Genome Database (SGD; http://www.yeastgenome.org) is an expertly curated database of literature-derived functional information for the model organism budding yeast, Saccharomyces cerevisiae. SGD constantly strives to synergize new types of experimental data and bioinformatics predictions with existing data, and to organize them into a comprehensive and up-to-date information resource. The primary mission of SGD is to facilitate research into the biology of yeast and...

  2. The Saccharomyces Genome Database Variant Viewer.

    Science.gov (United States)

    Sheppard, Travis K; Hitz, Benjamin C; Engel, Stacia R; Song, Giltae; Balakrishnan, Rama; Binkley, Gail; Costanzo, Maria C; Dalusag, Kyla S; Demeter, Janos; Hellerstedt, Sage T; Karra, Kalpana; Nash, Robert S; Paskov, Kelley M; Skrzypek, Marek S; Weng, Shuai; Wong, Edith D; Cherry, J Michael

    2016-01-04

    The Saccharomyces Genome Database (SGD; http://www.yeastgenome.org) is the authoritative community resource for the Saccharomyces cerevisiae reference genome sequence and its annotation. In recent years, we have moved toward increased representation of sequence variation and allelic differences within S. cerevisiae. The publication of numerous additional genomes has motivated the creation of new tools for their annotation and analysis. Here we present the Variant Viewer: a dynamic open-source web application for the visualization of genomic and proteomic differences. Multiple sequence alignments have been constructed across high quality genome sequences from 11 different S. cerevisiae strains and stored in the SGD. The alignments and summaries are encoded in JSON and used to create a two-tiered dynamic view of the budding yeast pan-genome, available at http://www.yeastgenome.org/variant-viewer. © The Author(s) 2015. Published by Oxford University Press on behalf of Nucleic Acids Research.

  3. Exploring Protein Function Using the Saccharomyces Genome Database.

    Science.gov (United States)

    Wong, Edith D

    2017-01-01

    Elucidating the function of individual proteins will help to create a comprehensive picture of cell biology, as well as shed light on human disease mechanisms, possible treatments, and cures. Due to its compact genome, and extensive history of experimentation and annotation, the budding yeast Saccharomyces cerevisiae is an ideal model organism in which to determine protein function. This information can then be leveraged to infer functions of human homologs. Despite the large amount of research and biological data about S. cerevisiae, many proteins' functions remain unknown. Here, we explore ways to use the Saccharomyces Genome Database (SGD; http://www.yeastgenome.org ) to predict the function of proteins and gain insight into their roles in various cellular processes.

  4. Biocuration at the Saccharomyces genome database.

    Science.gov (United States)

    Skrzypek, Marek S; Nash, Robert S

    2015-08-01

    Saccharomyces Genome Database is an online resource dedicated to managing information about the biology and genetics of the model organism, yeast (Saccharomyces cerevisiae). This information is derived primarily from scientific publications through a process of human curation that involves manual extraction of data and their organization into a comprehensive system of knowledge. This system provides a foundation for further analysis of experimental data coming from research on yeast as well as other organisms. In this review we will demonstrate how biocuration and biocurators add a key component, the biological context, to our understanding of how genes, proteins, genomes and cells function and interact. We will explain the role biocurators play in sifting through the wealth of biological data to incorporate and connect key information. We will also discuss the many ways we assist researchers with their various research needs. We hope to convince the reader that manual curation is vital in converting the flood of data into organized and interconnected knowledge, and that biocurators play an essential role in the integration of scientific information into a coherent model of the cell. © 2015 Wiley Periodicals, Inc.

  5. Large-scale functional genomic analysis of sporulation and meiosis in Saccharomyces cerevisiae.

    OpenAIRE

    Enyenihi, Akon H; Saunders, William S

    2003-01-01

    We have used a single-gene deletion mutant bank to identify the genes required for meiosis and sporulation among 4323 nonessential Saccharomyces cerevisiae annotated open reading frames (ORFs). Three hundred thirty-four sporulation-essential genes were identified, including 78 novel ORFs and 115 known genes without previously described sporulation defects in the comprehensive Saccharomyces Genome (SGD) or Yeast Proteome (YPD) phenotype databases. We have further divided the uncharacterized sp...

  6. Genomic insights into the Saccharomyces sensu stricto complex.

    Science.gov (United States)

    Borneman, Anthony R; Pretorius, Isak S

    2015-02-01

    The Saccharomyces sensu stricto group encompasses species ranging from the industrially ubiquitous yeast Saccharomyces cerevisiae to those that are confined to geographically limited environmental niches. The wealth of genomic data that are now available for the Saccharomyces genus is providing unprecedented insights into the genomic processes that can drive speciation and evolution, both in the natural environment and in response to human-driven selective forces during the historical "domestication" of these yeasts for baking, brewing, and winemaking. Copyright © 2015 by the Genetics Society of America.

  7. Mitochondrial genome evolution in the Saccharomyces sensu stricto complex.

    Science.gov (United States)

    Ruan, Jiangxing; Cheng, Jian; Zhang, Tongcun; Jiang, Huifeng

    2017-01-01

    Exploring the evolutionary patterns of mitochondrial genomes is important for our understanding of the Saccharomyces sensu stricto (SSS) group, which is a model system for genomic evolution and ecological analysis. In this study, we first obtained the complete mitochondrial sequences of two important species, Saccharomyces mikatae and Saccharomyces kudriavzevii. We then compared the mitochondrial genomes in the SSS group with those of close relatives, and found that the non-coding regions evolved rapidly, including dramatic expansion of intergenic regions, fast evolution of introns and almost 20-fold higher rearrangement rates than those of the nuclear genomes. However, the coding regions, and especially the protein-coding genes, are more conserved than those in the nuclear genomes of the SSS group. The different evolutionary patterns of coding and non-coding regions in the mitochondrial and nuclear genomes may be related to the origin of the aerobic fermentation lifestyle in this group. Our analysis thus provides novel insights into the evolution of mitochondrial genomes.

  8. Fungal genomics beyond Saccharomyces cerevisiae?

    DEFF Research Database (Denmark)

    Hofmann, Gerald; Mcintyre, Mhairi; Nielsen, Jens

    2003-01-01

    Fungi are used extensively in both fundamental research and industrial applications. Saccharomyces cerevisiae has been the model organism for fungal research for many years, particularly in functional genomics. However, considering the diversity within the fungal kingdom, it is obvious...

  9. Genomics and Biochemistry of Saccharomyces cerevisiae Wine Yeast Strains.

    Science.gov (United States)

    Eldarov, M A; Kishkovskaia, S A; Tanaschuk, T N; Mardanov, A V

    2016-12-01

    Saccharomyces yeasts have been used for millennia for the production of beer, wine, bread, and other fermented products. Long-term "unconscious" selection and domestication led to the selection of hundreds of strains with desired production traits having significant phenotypic and genetic differences from their wild ancestors. This review summarizes the results of recent research in deciphering the genomes of wine Saccharomyces strains, the use of comparative genomics methods to study the mechanisms of yeast genome evolution under conditions of artificial selection, and the use of genomic and postgenomic approaches to identify the molecular nature of the important characteristics of commercial wine strains of Saccharomyces. Succinctly, data concerning metagenomics of microbial communities of grapes and wine and the dynamics of yeast and bacterial flora in the course of winemaking is provided. A separate section is devoted to an overview of the physiological, genetic, and biochemical features of sherry yeast strains used to produce biologically aged wines. The goal of the review is to convince the reader of the efficacy of new genomic and postgenomic technologies as tools for developing strategies for targeted selection and creation of new strains using "classical" and modern techniques for improving winemaking technology.

  10. Incorporating Protein Biosynthesis into the Saccharomyces cerevisiae Genome-scale Metabolic Model

    DEFF Research Database (Denmark)

    Olivares Hernandez, Roberto

    Based on stoichiometric biochemical equations that occur into the cell, the genome-scale metabolic models can quantify the metabolic fluxes, which are regarded as the final representation of the physiological state of the cell. For Saccharomyces Cerevisiae the genome scale model has been construc......Based on stoichiometric biochemical equations that occur into the cell, the genome-scale metabolic models can quantify the metabolic fluxes, which are regarded as the final representation of the physiological state of the cell. For Saccharomyces Cerevisiae the genome scale model has been...

  11. The Sequenced Angiosperm Genomes and Genome Databases.

    Science.gov (United States)

    Chen, Fei; Dong, Wei; Zhang, Jiawei; Guo, Xinyue; Chen, Junhao; Wang, Zhengjia; Lin, Zhenguo; Tang, Haibao; Zhang, Liangsheng

    2018-01-01

    Angiosperms, the flowering plants, provide the essential resources for human life, such as food, energy, oxygen, and materials. They also promoted the evolution of human, animals, and the planet earth. Despite the numerous advances in genome reports or sequencing technologies, no review covers all the released angiosperm genomes and the genome databases for data sharing. Based on the rapid advances and innovations in the database reconstruction in the last few years, here we provide a comprehensive review for three major types of angiosperm genome databases, including databases for a single species, for a specific angiosperm clade, and for multiple angiosperm species. The scope, tools, and data of each type of databases and their features are concisely discussed. The genome databases for a single species or a clade of species are especially popular for specific group of researchers, while a timely-updated comprehensive database is more powerful for address of major scientific mysteries at the genome scale. Considering the low coverage of flowering plants in any available database, we propose construction of a comprehensive database to facilitate large-scale comparative studies of angiosperm genomes and to promote the collaborative studies of important questions in plant biology.

  12. Inheritance and organisation of the mitochondrial genome differ between two Saccharomyces yeasts

    DEFF Research Database (Denmark)

    Petersen, Randi Føns; Langkjær, Rikke Breinhold; Hvidtfeldt, J.

    2002-01-01

    Petite-positive Saccharomyces yeasts can be roughly divided into the sensu stricto, including Saccharomyces cerevisiae, and sensu lato group, including Saccharomyces castellii; the latter was recently studied for transmission and the organisation of its mitochondrial genome. S. castellii mitochon......Petite-positive Saccharomyces yeasts can be roughly divided into the sensu stricto, including Saccharomyces cerevisiae, and sensu lato group, including Saccharomyces castellii; the latter was recently studied for transmission and the organisation of its mitochondrial genome. S. castellii...... mitochondrial molecules (mtDNA) carrying point mutations, which confer antibiotic resistance, behaved in genetic crosses as the corresponding point mutants of S. cerevisiae. While S. castellii generated spontaneous petite mutants in a similar way as S. cerevisiae, the petites exhibited a different inheritance...... pattern. In crosses with the wild type strains a majority of S. castellii petites was neutral, and the suppressivity in suppressive petites was never over 50%. The two yeasts also differ in organisation of their mtDNA molecules. The 25,753 bp sequence of S. castellii mtDNA was determined and the coding...

  13. Non-introgressive genome chimerisation by malsegregation in autodiploidised allotetraploids during meiosis of Saccharomyces kudriavzevii x Saccharomyces uvarum hybrids.

    Science.gov (United States)

    Karanyicz, Edina; Antunovics, Zsuzsa; Kallai, Z; Sipiczki, M

    2017-06-01

    Saccharomyces strains with chimerical genomes consisting of mosaics of the genomes of different species ("natural hybrids") occur quite frequently among industrial and wine strains. The most widely endorsed hypothesis is that the mosaics are introgressions acquired via hybridisation and repeated backcrosses of the hybrids with one of the parental species. However, the interspecies hybrids are sterile, unable to mate with their parents. Here, we show by analysing synthetic Saccharomyces kudriavzevii x Saccharomyces uvarum hybrids that mosaic (chimeric) genomes can arise without introgressive backcrosses. These species are biologically separated by a double sterility barrier (sterility of allodiploids and F1 sterility of allotetraploids). F1 sterility is due to the diploidisation of the tetraploid meiosis resulting in MAT a /MAT α heterozygosity which suppresses mating in the spores. This barrier can occasionally be broken down by malsegregation of autosyndetically paired chromosomes carrying the MAT loci (loss of MAT heterozygosity). Subsequent malsegregation of additional autosyndetically paired chromosomes and occasional allosyndetic interactions chimerise the hybrid genome. Chromosomes are preferentially lost from the S. kudriavzevii subgenome. The uniparental transmission of the mitochondrial DNA to the hybrids indicates that nucleo-mitochondrial interactions might affect the direction of the genomic changes. We propose the name GARMe (Genome AutoReduction in Meiosis) for this process of genome reduction and chimerisation which involves no introgressive backcrossings. It opens a way to transfer genetic information between species and thus to get one step ahead after hybridisation in the production of yeast strains with beneficial combinations of properties of different species.

  14. Genome-scale reconstruction of the Saccharomyces cerevisiae metabolic network

    DEFF Research Database (Denmark)

    Förster, Jochen; Famili, I.; Fu, P.

    2003-01-01

    The metabolic network in the yeast Saccharomyces cerevisiae was reconstructed using currently available genomic, biochemical, and physiological information. The metabolic reactions were compartmentalized between the cytosol and the mitochondria, and transport steps between the compartments...

  15. Gleaning evolutionary insights from the genome sequence of a probiotic yeast Saccharomyces boulardii.

    Science.gov (United States)

    Khatri, Indu; Akhtar, Akil; Kaur, Kamaldeep; Tomar, Rajul; Prasad, Gandham Satyanarayana; Ramya, Thirumalai Nallan Chakravarthy; Subramanian, Srikrishna

    2013-10-22

    The yeast Saccharomyces boulardii is used worldwide as a probiotic to alleviate the effects of several gastrointestinal diseases and control antibiotics-associated diarrhea. While many studies report the probiotic effects of S. boulardii, no genome information for this yeast is currently available in the public domain. We report the 11.4 Mbp draft genome of this probiotic yeast. The draft genome was obtained by assembling Roche 454 FLX + shotgun data into 194 contigs with an N50 of 251 Kbp. We compare our draft genome with all other Saccharomyces cerevisiae genomes. Our analysis confirms the close similarity of S. boulardii to S. cerevisiae strains and provides a framework to understand the probiotic effects of this yeast, which exhibits unique physiological and metabolic properties.

  16. The YH database: the first Asian diploid genome database

    DEFF Research Database (Denmark)

    Li, Guoqing; Ma, Lijia; Song, Chao

    2009-01-01

    genome consensus. The YH database is currently one of the three personal genome database, organizing the original data and analysis results in a user-friendly interface, which is an endeavor to achieve fundamental goals for establishing personal medicine. The database is available at http://yh.genomics.org.cn....

  17. Mycobacteriophage genome database.

    Science.gov (United States)

    Joseph, Jerrine; Rajendran, Vasanthi; Hassan, Sameer; Kumar, Vanaja

    2011-01-01

    Mycobacteriophage genome database (MGDB) is an exclusive repository of the 64 completely sequenced mycobacteriophages with annotated information. It is a comprehensive compilation of the various gene parameters captured from several databases pooled together to empower mycobacteriophage researchers. The MGDB (Version No.1.0) comprises of 6086 genes from 64 mycobacteriophages classified into 72 families based on ACLAME database. Manual curation was aided by information available from public databases which was enriched further by analysis. Its web interface allows browsing as well as querying the classification. The main objective is to collect and organize the complexity inherent to mycobacteriophage protein classification in a rational way. The other objective is to browse the existing and new genomes and describe their functional annotation. The database is available for free at http://mpgdb.ibioinformatics.org/mpgdb.php.

  18. Karyotypes of Saccharomyces sensu lato species

    DEFF Research Database (Denmark)

    Petersen, Randi Føns; Nilsson-Tilgren, Torsten; Piskur, Jure

    1999-01-01

    An improved pulsed-field electrophoresis program was developed to study differently sized chromosomes within the genus Saccharomyces. The number of chromosomes in the type strains was shown to be nine in Saccharomyces castellii and Saccharomyces dairenensis, 12 in Saccharomyces servazzii...... and Saccharomyces unisporus, 16 in Saccharomyces exiguus and seven in Saccharomyces kluyveri. The sizes of individual chromosomes were resolved and the approximate genome sizes were determined by the addition of individual chromosomes of the karyotypes. Apparently. the genome of S. exiguus, which is the only...... Saccharomyces sensu late yeast to contain small chromosomes, is larger than that of Saccharomyces cerevisiae. On the other hand, other species exhibited genome sizes that were 10-25% smaller than that of S. cerevisiae. Well-defined karyotypes represent the basis for future genome mapping and sequencing projects...

  19. Full Data of Yeast Interacting Proteins Database (Original Version) - Yeast Interacting Proteins Database | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available List Contact us Yeast Interacting Proteins Database Full Data of Yeast Interacting Proteins Database (Origin...al Version) Data detail Data name Full Data of Yeast Interacting Proteins Database (Original Version) DOI 10....18908/lsdba.nbdc00742-004 Description of data contents The entire data in the Yeast Interacting Proteins Database...eir interactions are required. Several sources including YPD (Yeast Proteome Database, Costanzo, M. C., Hoga...ematic name in the SGD (Saccharomyces Genome Database; http://www.yeastgenome.org /). Bait gene name The gen

  20. i-Genome: A database to summarize oligonucleotide data in genomes

    Directory of Open Access Journals (Sweden)

    Chang Yu-Chung

    2004-10-01

    Full Text Available Abstract Background Information on the occurrence of sequence features in genomes is crucial to comparative genomics, evolutionary analysis, the analyses of regulatory sequences and the quantitative evaluation of sequences. Computing the frequencies and the occurrences of a pattern in complete genomes is time-consuming. Results The proposed database provides information about sequence features generated by exhaustively computing the sequences of the complete genome. The repetitive elements in the eukaryotic genomes, such as LINEs, SINEs, Alu and LTR, are obtained from Repbase. The database supports various complete genomes including human, yeast, worm, and 128 microbial genomes. Conclusions This investigation presents and implements an efficiently computational approach to accumulate the occurrences of the oligonucleotides or patterns in complete genomes. A database is established to maintain the information of the sequence features, including the distributions of oligonucleotide, the gene distribution, the distribution of repetitive elements in genomes and the occurrences of the oligonucleotides. The database can provide more effective and efficient way to access the repetitive features in genomes.

  1. Rat Genome Database (RGD)

    Data.gov (United States)

    U.S. Department of Health & Human Services — The Rat Genome Database (RGD) is a collaborative effort between leading research institutions involved in rat genetic and genomic research to collect, consolidate,...

  2. MIPS: a database for genomes and protein sequences.

    Science.gov (United States)

    Mewes, H W; Frishman, D; Güldener, U; Mannhaupt, G; Mayer, K; Mokrejs, M; Morgenstern, B; Münsterkötter, M; Rudd, S; Weil, B

    2002-01-01

    The Munich Information Center for Protein Sequences (MIPS-GSF, Neuherberg, Germany) continues to provide genome-related information in a systematic way. MIPS supports both national and European sequencing and functional analysis projects, develops and maintains automatically generated and manually annotated genome-specific databases, develops systematic classification schemes for the functional annotation of protein sequences, and provides tools for the comprehensive analysis of protein sequences. This report updates the information on the yeast genome (CYGD), the Neurospora crassa genome (MNCDB), the databases for the comprehensive set of genomes (PEDANT genomes), the database of annotated human EST clusters (HIB), the database of complete cDNAs from the DHGP (German Human Genome Project), as well as the project specific databases for the GABI (Genome Analysis in Plants) and HNB (Helmholtz-Netzwerk Bioinformatik) networks. The Arabidospsis thaliana database (MATDB), the database of mitochondrial proteins (MITOP) and our contribution to the PIR International Protein Sequence Database have been described elsewhere [Schoof et al. (2002) Nucleic Acids Res., 30, 91-93; Scharfe et al. (2000) Nucleic Acids Res., 28, 155-158; Barker et al. (2001) Nucleic Acids Res., 29, 29-32]. All databases described, the protein analysis tools provided and the detailed descriptions of our projects can be accessed through the MIPS World Wide Web server (http://mips.gsf.de).

  3. GenColors-based comparative genome databases for small eukaryotic genomes.

    Science.gov (United States)

    Felder, Marius; Romualdi, Alessandro; Petzold, Andreas; Platzer, Matthias; Sühnel, Jürgen; Glöckner, Gernot

    2013-01-01

    Many sequence data repositories can give a quick and easily accessible overview on genomes and their annotations. Less widespread is the possibility to compare related genomes with each other in a common database environment. We have previously described the GenColors database system (http://gencolors.fli-leibniz.de) and its applications to a number of bacterial genomes such as Borrelia, Legionella, Leptospira and Treponema. This system has an emphasis on genome comparison. It combines data from related genomes and provides the user with an extensive set of visualization and analysis tools. Eukaryote genomes are normally larger than prokaryote genomes and thus pose additional challenges for such a system. We have, therefore, adapted GenColors to also handle larger datasets of small eukaryotic genomes and to display eukaryotic gene structures. Further recent developments include whole genome views, genome list options and, for bacterial genome browsers, the display of horizontal gene transfer predictions. Two new GenColors-based databases for two fungal species (http://fgb.fli-leibniz.de) and for four social amoebas (http://sacgb.fli-leibniz.de) were set up. Both new resources open up a single entry point for related genomes for the amoebozoa and fungal research communities and other interested users. Comparative genomics approaches are greatly facilitated by these resources.

  4. Mms1 binds to G-rich regions in Saccharomyces cerevisiae and influences replication and genome stability

    NARCIS (Netherlands)

    Wanzek, Katharina; Schwindt, Eike; Capra, John A.; Paeschke, Katrin

    2017-01-01

    The regulation of replication is essential to preserve genome integrity. Mms1 is part of the E3 ubiquitin ligase complex that is linked to replication fork progression. By identifying Mms1 binding sites genome-wide in Saccharomyces cerevisiae we connected Mms1 function to genome integrity and

  5. Systematic discovery of unannotated genes in 11 yeast species using a database of orthologous genomic segments

    LENUS (Irish Health Repository)

    OhEigeartaigh, Sean S

    2011-07-26

    Abstract Background In standard BLAST searches, no information other than the sequences of the query and the database entries is considered. However, in situations where two genes from different species have only borderline similarity in a BLAST search, the discovery that the genes are located within a region of conserved gene order (synteny) can provide additional evidence that they are orthologs. Thus, for interpreting borderline search results, it would be useful to know whether the syntenic context of a database hit is similar to that of the query. This principle has often been used in investigations of particular genes or genomic regions, but to our knowledge it has never been implemented systematically. Results We made use of the synteny information contained in the Yeast Gene Order Browser database for 11 yeast species to carry out a systematic search for protein-coding genes that were overlooked in the original annotations of one or more yeast genomes but which are syntenic with their orthologs. Such genes tend to have been overlooked because they are short, highly divergent, or contain introns. The key features of our software - called SearchDOGS - are that the database entries are classified into sets of genomic segments that are already known to be orthologous, and that very weak BLAST hits are retained for further analysis if their genomic location is similar to that of the query. Using SearchDOGS we identified 595 additional protein-coding genes among the 11 yeast species, including two new genes in Saccharomyces cerevisiae. We found additional genes for the mating pheromone a-factor in six species including Kluyveromyces lactis. Conclusions SearchDOGS has proven highly successful for identifying overlooked genes in the yeast genomes. We anticipate that our approach can be adapted for study of further groups of species, such as bacterial genomes. More generally, the concept of doing sequence similarity searches against databases to which external

  6. BGD: a database of bat genomes.

    Science.gov (United States)

    Fang, Jianfei; Wang, Xuan; Mu, Shuo; Zhang, Shuyi; Dong, Dong

    2015-01-01

    Bats account for ~20% of mammalian species, and are the only mammals with true powered flight. For the sake of their specialized phenotypic traits, many researches have been devoted to examine the evolution of bats. Until now, some whole genome sequences of bats have been assembled and annotated, however, a uniform resource for the annotated bat genomes is still unavailable. To make the extensive data associated with the bat genomes accessible to the general biological communities, we established a Bat Genome Database (BGD). BGD is an open-access, web-available portal that integrates available data of bat genomes and genes. It hosts data from six bat species, including two megabats and four microbats. Users can query the gene annotations using efficient searching engine, and it offers browsable tracks of bat genomes. Furthermore, an easy-to-use phylogenetic analysis tool was also provided to facilitate online phylogeny study of genes. To the best of our knowledge, BGD is the first database of bat genomes. It will extend our understanding of the bat evolution and be advantageous to the bat sequences analysis. BGD is freely available at: http://donglab.ecnu.edu.cn/databases/BatGenome/.

  7. BGD: a database of bat genomes.

    Directory of Open Access Journals (Sweden)

    Jianfei Fang

    Full Text Available Bats account for ~20% of mammalian species, and are the only mammals with true powered flight. For the sake of their specialized phenotypic traits, many researches have been devoted to examine the evolution of bats. Until now, some whole genome sequences of bats have been assembled and annotated, however, a uniform resource for the annotated bat genomes is still unavailable. To make the extensive data associated with the bat genomes accessible to the general biological communities, we established a Bat Genome Database (BGD. BGD is an open-access, web-available portal that integrates available data of bat genomes and genes. It hosts data from six bat species, including two megabats and four microbats. Users can query the gene annotations using efficient searching engine, and it offers browsable tracks of bat genomes. Furthermore, an easy-to-use phylogenetic analysis tool was also provided to facilitate online phylogeny study of genes. To the best of our knowledge, BGD is the first database of bat genomes. It will extend our understanding of the bat evolution and be advantageous to the bat sequences analysis. BGD is freely available at: http://donglab.ecnu.edu.cn/databases/BatGenome/.

  8. GOBASE: an organelle genome database

    OpenAIRE

    O?Brien, Emmet A.; Zhang, Yue; Wang, Eric; Marie, Veronique; Badejoko, Wole; Lang, B. Franz; Burger, Gertraud

    2008-01-01

    The organelle genome database GOBASE, now in its 21st release (June 2008), contains all published mitochondrion-encoded sequences (?913 000) and chloroplast-encoded sequences (?250 000) from a wide range of eukaryotic taxa. For all sequences, information on related genes, exons, introns, gene products and taxonomy is available, as well as selected genome maps and RNA secondary structures. Recent major enhancements to database functionality include: (i) addition of an interface for RNA editing...

  9. Saccharomyces pastorianus: genomic insights inspiring innovation for industry.

    Science.gov (United States)

    Gibson, Brian; Liti, Gianni

    2015-01-01

    A combination of biological and non-biological factors has led to the interspecific hybrid yeast species Saccharomyces pastorianus becoming one of the world's most important industrial organisms. This yeast is used in the production of lager-style beers, the fermentation of which requires very low temperatures compared to other industrial fermentation processes. This group of organisms has benefited from both the whole-genome duplication in its ancestral lineage and the subsequent hybridization event between S. cerevisiae and S. eubayanus, resulting in strong fermentative ability. The hybrid has key traits, such as cold tolerance and good maltose- and maltotriose-utilizing ability, inherited either from the parental species or originating from genetic interactions between the parent genomes. Instability in the nascent allopolyploid hybrid genome may have contributed to rapid evolution of the yeast to tolerate conditions prevalent in the brewing environment. The recent discovery of S. eubayanus has provided new insights into the evolutionary history of S. pastorianus and may offer new opportunities for generating novel industrially-beneficial lager yeast strains. Copyright © 2014 John Wiley & Sons, Ltd.

  10. Complete genome sequence and comparative genomics of the probiotic yeast Saccharomyces boulardii.

    Science.gov (United States)

    Khatri, Indu; Tomar, Rajul; Ganesan, K; Prasad, G S; Subramanian, Srikrishna

    2017-03-23

    The probiotic yeast, Saccharomyces boulardii (Sb) is known to be effective against many gastrointestinal disorders and antibiotic-associated diarrhea. To understand molecular basis of probiotic-properties ascribed to Sb we determined the complete genomes of two strains of Sb i.e. Biocodex and unique28 and the draft genomes for three other Sb strains that are marketed as probiotics in India. We compared these genomes with 145 strains of S. cerevisiae (Sc) to understand genome-level similarities and differences between these yeasts. A distinctive feature of Sb from other Sc is absence of Ty elements Ty1, Ty3, Ty4 and associated LTR. However, we could identify complete Ty2 and Ty5 elements in Sb. The genes for hexose transporters HXT11 and HXT9, and asparagine-utilization are absent in all Sb strains. We find differences in repeat periods and copy numbers of repeats in flocculin genes that are likely related to the differential adhesion of Sb as compared to Sc. Core-proteome based taxonomy places Sb strains along with wine strains of Sc. We find the introgression of five genes from Z. bailii into the chromosome IV of Sb and wine strains of Sc. Intriguingly, genes involved in conferring known probiotic properties to Sb are conserved in most Sc strains.

  11. Genomic Sequence of Saccharomyces cerevisiae BAW-6, a Yeast Strain Optimal for Brewing Barley Shochu.

    Science.gov (United States)

    Kajiwara, Yasuhiro; Mori, Kazuki; Tashiro, Kosuke; Higuchi, Yujiro; Takegawa, Kaoru; Takashita, Hideharu

    2018-04-05

    Here, we report the draft genome sequence of Saccharomyces cerevisiae strain BAW-6, which is used for the production of barley shochu, a traditional Japanese spirit. This genomic information can be used to elucidate the genetic basis underlying the high alcohol production capacity and citric acid tolerance of shochu yeast. Copyright © 2018 Kajiwara et al.

  12. Private and Efficient Query Processing on Outsourced Genomic Databases.

    Science.gov (United States)

    Ghasemi, Reza; Al Aziz, Md Momin; Mohammed, Noman; Dehkordi, Massoud Hadian; Jiang, Xiaoqian

    2017-09-01

    Applications of genomic studies are spreading rapidly in many domains of science and technology such as healthcare, biomedical research, direct-to-consumer services, and legal and forensic. However, there are a number of obstacles that make it hard to access and process a big genomic database for these applications. First, sequencing genomic sequence is a time consuming and expensive process. Second, it requires large-scale computation and storage systems to process genomic sequences. Third, genomic databases are often owned by different organizations, and thus, not available for public usage. Cloud computing paradigm can be leveraged to facilitate the creation and sharing of big genomic databases for these applications. Genomic data owners can outsource their databases in a centralized cloud server to ease the access of their databases. However, data owners are reluctant to adopt this model, as it requires outsourcing the data to an untrusted cloud service provider that may cause data breaches. In this paper, we propose a privacy-preserving model for outsourcing genomic data to a cloud. The proposed model enables query processing while providing privacy protection of genomic databases. Privacy of the individuals is guaranteed by permuting and adding fake genomic records in the database. These techniques allow cloud to evaluate count and top-k queries securely and efficiently. Experimental results demonstrate that a count and a top-k query over 40 Single Nucleotide Polymorphisms (SNPs) in a database of 20 000 records takes around 100 and 150 s, respectively.

  13. Ginseng Genome Database: an open-access platform for genomics of Panax ginseng.

    Science.gov (United States)

    Jayakodi, Murukarthick; Choi, Beom-Soon; Lee, Sang-Choon; Kim, Nam-Hoon; Park, Jee Young; Jang, Woojong; Lakshmanan, Meiyappan; Mohan, Shobhana V G; Lee, Dong-Yup; Yang, Tae-Jin

    2018-04-12

    The ginseng (Panax ginseng C.A. Meyer) is a perennial herbaceous plant that has been used in traditional oriental medicine for thousands of years. Ginsenosides, which have significant pharmacological effects on human health, are the foremost bioactive constituents in this plant. Having realized the importance of this plant to humans, an integrated omics resource becomes indispensable to facilitate genomic research, molecular breeding and pharmacological study of this herb. The first draft genome sequences of P. ginseng cultivar "Chunpoong" were reported recently. Here, using the draft genome, transcriptome, and functional annotation datasets of P. ginseng, we have constructed the Ginseng Genome Database http://ginsengdb.snu.ac.kr /, the first open-access platform to provide comprehensive genomic resources of P. ginseng. The current version of this database provides the most up-to-date draft genome sequence (of approximately 3000 Mbp of scaffold sequences) along with the structural and functional annotations for 59,352 genes and digital expression of genes based on transcriptome data from different tissues, growth stages and treatments. In addition, tools for visualization and the genomic data from various analyses are provided. All data in the database were manually curated and integrated within a user-friendly query page. This database provides valuable resources for a range of research fields related to P. ginseng and other species belonging to the Apiales order as well as for plant research communities in general. Ginseng genome database can be accessed at http://ginsengdb.snu.ac.kr /.

  14. The UCSC Genome Browser Database: 2008 update

    DEFF Research Database (Denmark)

    Karolchik, D; Kuhn, R M; Baertsch, R

    2007-01-01

    The University of California, Santa Cruz, Genome Browser Database (GBD) provides integrated sequence and annotation data for a large collection of vertebrate and model organism genomes. Seventeen new assemblies have been added to the database in the past year, for a total coverage of 19 vertebrat...

  15. Recent updates and developments to plant genome size databases

    Science.gov (United States)

    Garcia, Sònia; Leitch, Ilia J.; Anadon-Rosell, Alba; Canela, Miguel Á.; Gálvez, Francisco; Garnatje, Teresa; Gras, Airy; Hidalgo, Oriane; Johnston, Emmeline; Mas de Xaxars, Gemma; Pellicer, Jaume; Siljak-Yakovlev, Sonja; Vallès, Joan; Vitales, Daniel; Bennett, Michael D.

    2014-01-01

    Two plant genome size databases have been recently updated and/or extended: the Plant DNA C-values database (http://data.kew.org/cvalues), and GSAD, the Genome Size in Asteraceae database (http://www.asteraceaegenomesize.com). While the first provides information on nuclear DNA contents across land plants and some algal groups, the second is focused on one of the largest and most economically important angiosperm families, Asteraceae. Genome size data have numerous applications: they can be used in comparative studies on genome evolution, or as a tool to appraise the cost of whole-genome sequencing programs. The growing interest in genome size and increasing rate of data accumulation has necessitated the continued update of these databases. Currently, the Plant DNA C-values database (Release 6.0, Dec. 2012) contains data for 8510 species, while GSAD has 1219 species (Release 2.0, June 2013), representing increases of 17 and 51%, respectively, in the number of species with genome size data, compared with previous releases. Here we provide overviews of the most recent releases of each database, and outline new features of GSAD. The latter include (i) a tool to visually compare genome size data between species, (ii) the option to export data and (iii) a webpage containing information about flow cytometry protocols. PMID:24288377

  16. De-anonymizing Genomic Databases Using Phenotypic Traits

    Directory of Open Access Journals (Sweden)

    Humbert Mathias

    2015-06-01

    Full Text Available People increasingly have their genomes sequenced and some of them share their genomic data online. They do so for various purposes, including to find relatives and to help advance genomic research. An individual’s genome carries very sensitive, private information such as its owner’s susceptibility to diseases, which could be used for discrimination. Therefore, genomic databases are often anonymized. However, an individual’s genotype is also linked to visible phenotypic traits, such as eye or hair color, which can be used to re-identify users in anonymized public genomic databases, thus raising severe privacy issues. For instance, an adversary can identify a target’s genome using known her phenotypic traits and subsequently infer her susceptibility to Alzheimer’s disease. In this paper, we quantify, based on various phenotypic traits, the extent of this threat in several scenarios by implementing de-anonymization attacks on a genomic database of OpenSNP users sequenced by 23andMe. Our experimental results show that the proportion of correct matches reaches 23% with a supervised approach in a database of 50 participants. Our approach outperforms the baseline by a factor of four, in terms of the proportion of correct matches, in most scenarios. We also evaluate the adversary’s ability to predict individuals’ predisposition to Alzheimer’s disease, and we observe that the inference error can be halved compared to the baseline. We also analyze the effect of the number of known phenotypic traits on the success rate of the attack. As progress is made in genomic research, especially for genotype-phenotype associations, the threat presented in this paper will become more serious.

  17. Benchmarking database performance for genomic data.

    Science.gov (United States)

    Khushi, Matloob

    2015-06-01

    Genomic regions represent features such as gene annotations, transcription factor binding sites and epigenetic modifications. Performing various genomic operations such as identifying overlapping/non-overlapping regions or nearest gene annotations are common research needs. The data can be saved in a database system for easy management, however, there is no comprehensive database built-in algorithm at present to identify overlapping regions. Therefore I have developed a novel region-mapping (RegMap) SQL-based algorithm to perform genomic operations and have benchmarked the performance of different databases. Benchmarking identified that PostgreSQL extracts overlapping regions much faster than MySQL. Insertion and data uploads in PostgreSQL were also better, although general searching capability of both databases was almost equivalent. In addition, using the algorithm pair-wise, overlaps of >1000 datasets of transcription factor binding sites and histone marks, collected from previous publications, were reported and it was found that HNF4G significantly co-locates with cohesin subunit STAG1 (SA1).Inc. © 2015 Wiley Periodicals, Inc.

  18. CyanoBase: the cyanobacteria genome database update 2010

    OpenAIRE

    Nakao, Mitsuteru; Okamoto, Shinobu; Kohara, Mitsuyo; Fujishiro, Tsunakazu; Fujisawa, Takatomo; Sato, Shusei; Tabata, Satoshi; Kaneko, Takakazu; Nakamura, Yasukazu

    2009-01-01

    CyanoBase (http://genome.kazusa.or.jp/cyanobase) is the genome database for cyanobacteria, which are model organisms for photosynthesis. The database houses cyanobacteria species information, complete genome sequences, genome-scale experiment data, gene information, gene annotations and mutant information. In this version, we updated these datasets and improved the navigation and the visual display of the data views. In addition, a web service API now enables users to retrieve the data in var...

  19. Benchmark data for identifying N6-methyladenosine sites in the Saccharomyces cerevisiae genome

    Directory of Open Access Journals (Sweden)

    Wei Chen

    2015-12-01

    Full Text Available This data article contains the benchmark dataset for training and testing iRNA-Methyl, a web-server predictor for identifying N6-methyladenosine sites in RNA (Chen et al., 2015 [15]. It can also be used to develop other predictors for identifying N6-methyladenosine sites in the Saccharomyces cerevisiae genome.

  20. High-efficiency genome editing and allele replacement in prototrophic and wild strains of Saccharomyces.

    Science.gov (United States)

    Alexander, William G; Doering, Drew T; Hittinger, Chris Todd

    2014-11-01

    Current genome editing techniques available for Saccharomyces yeast species rely on auxotrophic markers, limiting their use in wild and industrial strains and species. Taking advantage of the ancient loss of thymidine kinase in the fungal kingdom, we have developed the herpes simplex virus thymidine kinase gene as a selectable and counterselectable marker that forms the core of novel genome engineering tools called the H: aploid E: ngineering and R: eplacement P: rotocol (HERP) cassettes. Here we show that these cassettes allow a researcher to rapidly generate heterogeneous populations of cells with thousands of independent chromosomal allele replacements using mixed PCR products. We further show that the high efficiency of this approach enables the simultaneous replacement of both alleles in diploid cells. Using these new techniques, many of the most powerful yeast genetic manipulation strategies are now available in wild, industrial, and other prototrophic strains from across the diverse Saccharomyces genus. Copyright © 2014 by the Genetics Society of America.

  1. Database Description - TMBETA-GENOME | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available ENOME is a database for transmembrane β-barrel proteins in complete genomes. For each genome, calculations with machine learning algo...rithms and statistical methods have been perfumed and th

  2. CyanoBase: the cyanobacteria genome database update 2010.

    Science.gov (United States)

    Nakao, Mitsuteru; Okamoto, Shinobu; Kohara, Mitsuyo; Fujishiro, Tsunakazu; Fujisawa, Takatomo; Sato, Shusei; Tabata, Satoshi; Kaneko, Takakazu; Nakamura, Yasukazu

    2010-01-01

    CyanoBase (http://genome.kazusa.or.jp/cyanobase) is the genome database for cyanobacteria, which are model organisms for photosynthesis. The database houses cyanobacteria species information, complete genome sequences, genome-scale experiment data, gene information, gene annotations and mutant information. In this version, we updated these datasets and improved the navigation and the visual display of the data views. In addition, a web service API now enables users to retrieve the data in various formats with other tools, seamlessly.

  3. Genome scale models of yeast: towards standardized evaluation and consistent omic integration

    DEFF Research Database (Denmark)

    Sanchez, Benjamin J.; Nielsen, Jens

    2015-01-01

    Genome scale models (GEMs) have enabled remarkable advances in systems biology, acting as functional databases of metabolism, and as scaffolds for the contextualization of high-throughput data. In the case of Saccharomyces cerevisiae (budding yeast), several GEMs have been published and are curre......Genome scale models (GEMs) have enabled remarkable advances in systems biology, acting as functional databases of metabolism, and as scaffolds for the contextualization of high-throughput data. In the case of Saccharomyces cerevisiae (budding yeast), several GEMs have been published...... in which all levels of omics data (from gene expression to flux) have been integrated in yeast GEMs. Relevant conclusions and current challenges for both GEM evaluation and omic integration are highlighted....

  4. MIPS: a database for protein sequences and complete genomes.

    Science.gov (United States)

    Mewes, H W; Hani, J; Pfeiffer, F; Frishman, D

    1998-01-01

    The MIPS group [Munich Information Center for Protein Sequences of the German National Center for Environment and Health (GSF)] at the Max-Planck-Institute for Biochemistry, Martinsried near Munich, Germany, is involved in a number of data collection activities, including a comprehensive database of the yeast genome, a database reflecting the progress in sequencing the Arabidopsis thaliana genome, the systematic analysis of other small genomes and the collection of protein sequence data within the framework of the PIR-International Protein Sequence Database (described elsewhere in this volume). Through its WWW server (http://www.mips.biochem.mpg.de ) MIPS provides access to a variety of generic databases, including a database of protein families as well as automatically generated data by the systematic application of sequence analysis algorithms. The yeast genome sequence and its related information was also compiled on CD-ROM to provide dynamic interactive access to the 16 chromosomes of the first eukaryotic genome unraveled. PMID:9399795

  5. The UCSC Genome Browser Database: update 2006

    DEFF Research Database (Denmark)

    Hinrichs, A S; Karolchik, D; Baertsch, R

    2006-01-01

    The University of California Santa Cruz Genome Browser Database (GBD) contains sequence and annotation data for the genomes of about a dozen vertebrate species and several major model organisms. Genome annotations typically include assembly data, sequence composition, genes and gene predictions, ...

  6. Brassica ASTRA: an integrated database for Brassica genomic research.

    Science.gov (United States)

    Love, Christopher G; Robinson, Andrew J; Lim, Geraldine A C; Hopkins, Clare J; Batley, Jacqueline; Barker, Gary; Spangenberg, German C; Edwards, David

    2005-01-01

    Brassica ASTRA is a public database for genomic information on Brassica species. The database incorporates expressed sequences with Swiss-Prot and GenBank comparative sequence annotation as well as secondary Gene Ontology (GO) annotation derived from the comparison with Arabidopsis TAIR GO annotations. Simple sequence repeat molecular markers are identified within resident sequences and mapped onto the closely related Arabidopsis genome sequence. Bacterial artificial chromosome (BAC) end sequences derived from the Multinational Brassica Genome Project are also mapped onto the Arabidopsis genome sequence enabling users to identify candidate Brassica BACs corresponding to syntenic regions of Arabidopsis. This information is maintained in a MySQL database with a web interface providing the primary means of interrogation. The database is accessible at http://hornbill.cspp.latrobe.edu.au.

  7. The UCSC genome browser database: update 2007

    DEFF Research Database (Denmark)

    Kuhn, R M; Karolchik, D; Zweig, A S

    2006-01-01

    The University of California, Santa Cruz Genome Browser Database contains, as of September 2006, sequence and annotation data for the genomes of 13 vertebrate and 19 invertebrate species. The Genome Browser displays a wide variety of annotations at all scales from the single nucleotide level up t...

  8. Chromosomal Copy Number Variation in Saccharomyces pastorianus Is Evidence for Extensive Genome Dynamics in Industrial Lager Brewing Strains.

    Science.gov (United States)

    van den Broek, M; Bolat, I; Nijkamp, J F; Ramos, E; Luttik, M A H; Koopman, F; Geertman, J M; de Ridder, D; Pronk, J T; Daran, J-M

    2015-09-01

    Lager brewing strains of Saccharomyces pastorianus are natural interspecific hybrids originating from the spontaneous hybridization of Saccharomyces cerevisiae and Saccharomyces eubayanus. Over the past 500 years, S. pastorianus has been domesticated to become one of the most important industrial microorganisms. Production of lager-type beers requires a set of essential phenotypes, including the ability to ferment maltose and maltotriose at low temperature, the production of flavors and aromas, and the ability to flocculate. Understanding of the molecular basis of complex brewing-related phenotypic traits is a prerequisite for rational strain improvement. While genome sequences have been reported, the variability and dynamics of S. pastorianus genomes have not been investigated in detail. Here, using deep sequencing and chromosome copy number analysis, we showed that S. pastorianus strain CBS1483 exhibited extensive aneuploidy. This was confirmed by quantitative PCR and by flow cytometry. As a direct consequence of this aneuploidy, a massive number of sequence variants was identified, leading to at least 1,800 additional protein variants in S. pastorianus CBS1483. Analysis of eight additional S. pastorianus strains revealed that the previously defined group I strains showed comparable karyotypes, while group II strains showed large interstrain karyotypic variability. Comparison of three strains with nearly identical genome sequences revealed substantial chromosome copy number variation, which may contribute to strain-specific phenotypic traits. The observed variability of lager yeast genomes demonstrates that systematic linking of genotype to phenotype requires a three-dimensional genome analysis encompassing physical chromosomal structures, the copy number of individual chromosomes or chromosomal regions, and the allelic variation of copies of individual genes. Copyright © 2015, van den Broek et al.

  9. INE: a rice genome database with an integrated map view.

    Science.gov (United States)

    Sakata, K; Antonio, B A; Mukai, Y; Nagasaki, H; Sakai, Y; Makino, K; Sasaki, T

    2000-01-01

    The Rice Genome Research Program (RGP) launched a large-scale rice genome sequencing in 1998 aimed at decoding all genetic information in rice. A new genome database called INE (INtegrated rice genome Explorer) has been developed in order to integrate all the genomic information that has been accumulated so far and to correlate these data with the genome sequence. A web interface based on Java applet provides a rapid viewing capability in the database. The first operational version of the database has been completed which includes a genetic map, a physical map using YAC (Yeast Artificial Chromosome) clones and PAC (P1-derived Artificial Chromosome) contigs. These maps are displayed graphically so that the positional relationships among the mapped markers on each chromosome can be easily resolved. INE incorporates the sequences and annotations of the PAC contig. A site on low quality information ensures that all submitted sequence data comply with the standard for accuracy. As a repository of rice genome sequence, INE will also serve as a common database of all sequence data obtained by collaborating members of the International Rice Genome Sequencing Project (IRGSP). The database can be accessed at http://www. dna.affrc.go.jp:82/giot/INE. html or its mirror site at http://www.staff.or.jp/giot/INE.html

  10. The Ensembl genome database project.

    Science.gov (United States)

    Hubbard, T; Barker, D; Birney, E; Cameron, G; Chen, Y; Clark, L; Cox, T; Cuff, J; Curwen, V; Down, T; Durbin, R; Eyras, E; Gilbert, J; Hammond, M; Huminiecki, L; Kasprzyk, A; Lehvaslaiho, H; Lijnzaad, P; Melsopp, C; Mongin, E; Pettett, R; Pocock, M; Potter, S; Rust, A; Schmidt, E; Searle, S; Slater, G; Smith, J; Spooner, W; Stabenau, A; Stalker, J; Stupka, E; Ureta-Vidal, A; Vastrik, I; Clamp, M

    2002-01-01

    The Ensembl (http://www.ensembl.org/) database project provides a bioinformatics framework to organise biology around the sequences of large genomes. It is a comprehensive source of stable automatic annotation of the human genome sequence, with confirmed gene predictions that have been integrated with external data sources, and is available as either an interactive web site or as flat files. It is also an open source software engineering project to develop a portable system able to handle very large genomes and associated requirements from sequence analysis to data storage and visualisation. The Ensembl site is one of the leading sources of human genome sequence annotation and provided much of the analysis for publication by the international human genome project of the draft genome. The Ensembl system is being installed around the world in both companies and academic sites on machines ranging from supercomputers to laptops.

  11. gEVE: a genome-based endogenous viral element database provides comprehensive viral protein-coding sequences in mammalian genomes.

    Science.gov (United States)

    Nakagawa, So; Takahashi, Mahoko Ueda

    2016-01-01

    In mammals, approximately 10% of genome sequences correspond to endogenous viral elements (EVEs), which are derived from ancient viral infections of germ cells. Although most EVEs have been inactivated, some open reading frames (ORFs) of EVEs obtained functions in the hosts. However, EVE ORFs usually remain unannotated in the genomes, and no databases are available for EVE ORFs. To investigate the function and evolution of EVEs in mammalian genomes, we developed EVE ORF databases for 20 genomes of 19 mammalian species. A total of 736,771 non-overlapping EVE ORFs were identified and archived in a database named gEVE (http://geve.med.u-tokai.ac.jp). The gEVE database provides nucleotide and amino acid sequences, genomic loci and functional annotations of EVE ORFs for all 20 genomes. In analyzing RNA-seq data with the gEVE database, we successfully identified the expressed EVE genes, suggesting that the gEVE database facilitates studies of the genomic analyses of various mammalian species.Database URL: http://geve.med.u-tokai.ac.jp. © The Author(s) 2016. Published by Oxford University Press.

  12. Genomic Evolution of Saccharomyces cerevisiae under Chinese Rice Wine Fermentation

    Science.gov (United States)

    Li, Yudong; Zhang, Weiping; Zheng, Daoqiong; Zhou, Zhan; Yu, Wenwen; Zhang, Lei; Feng, Lifang; Liang, Xinle; Guan, Wenjun; Zhou, Jingwen; Chen, Jian; Lin, Zhenguo

    2014-01-01

    Rice wine fermentation represents a unique environment for the evolution of the budding yeast, Saccharomyces cerevisiae. To understand how the selection pressure shaped the yeast genome and gene regulation, we determined the genome sequence and transcriptome of a S. cerevisiae strain YHJ7 isolated from Chinese rice wine (Huangjiu), a popular traditional alcoholic beverage in China. By comparing the genome of YHJ7 to the lab strain S288c, a Japanese sake strain K7, and a Chinese industrial bioethanol strain YJSH1, we identified many genomic sequence and structural variations in YHJ7, which are mainly located in subtelomeric regions, suggesting that these regions play an important role in genomic evolution between strains. In addition, our comparative transcriptome analysis between YHJ7 and S288c revealed a set of differentially expressed genes, including those involved in glucose transport (e.g., HXT2, HXT7) and oxidoredutase activity (e.g., AAD10, ADH7). Interestingly, many of these genomic and transcriptional variations are directly or indirectly associated with the adaptation of YHJ7 strain to its specific niches. Our molecular evolution analysis suggested that Japanese sake strains (K7/UC5) were derived from Chinese rice wine strains (YHJ7) at least approximately 2,300 years ago, providing the first molecular evidence elucidating the origin of Japanese sake strains. Our results depict interesting insights regarding the evolution of yeast during rice wine fermentation, and provided a valuable resource for genetic engineering to improve industrial wine-making strains. PMID:25212861

  13. GEMMER: GEnome-wide tool for Multi-scale Modeling data Extraction and Representation for Saccharomyces cerevisiae.

    Science.gov (United States)

    Mondeel, Thierry D G A; Crémazy, Frédéric; Barberis, Matteo

    2018-02-01

    Multi-scale modeling of biological systems requires integration of various information about genes and proteins that are connected together in networks. Spatial, temporal and functional information is available; however, it is still a challenge to retrieve and explore this knowledge in an integrated, quick and user-friendly manner. We present GEMMER (GEnome-wide tool for Multi-scale Modelling data Extraction and Representation), a web-based data-integration tool that facilitates high quality visualization of physical, regulatory and genetic interactions between proteins/genes in Saccharomyces cerevisiae. GEMMER creates network visualizations that integrate information on function, temporal expression, localization and abundance from various existing databases. GEMMER supports modeling efforts by effortlessly gathering this information and providing convenient export options for images and their underlying data. GEMMER is freely available at http://gemmer.barberislab.com. Source code, written in Python, JavaScript library D3js, PHP and JSON, is freely available at https://github.com/barberislab/GEMMER. M.Barberis@uva.nl. Supplementary data are available at Bioinformatics online. © The Author(s) 2018. Published by Oxford University Press.

  14. Genome Sequences of Industrially Relevant Saccharomyces cerevisiae Strain M3707, Isolated from a Sample of Distillers Yeast and Four Haploid Derivatives

    Energy Technology Data Exchange (ETDEWEB)

    Brown, Steven D.; Klingeman, Dawn M.; Johnson, Courtney M.; Clum, Alicia; Aerts, Andrea; Salamov, Asaf; Sharma, Aditi; Zane, Matthew; Barry, Kerrie; Grigoriev, Igor V.; Davison, Brian H.; Lynd, Lee R.; Gilna, Paul; Hau, Heidi; Hogsett, David A.; Froehlich, Allan C.

    2013-04-19

    Saccharomyces cerevisiae strain M3707 was isolated from a sample of commercial distillers yeast, and its genome sequence together with the genome sequences for the four derived haploid strains M3836, M3837, M3838, and M3839 has been determined. Yeasts have potential for consolidated bioprocessing (CBP) for biofuel production, and access to these genome sequences will facilitate their development.

  15. The Ruby UCSC API: accessing the UCSC genome database using Ruby.

    Science.gov (United States)

    Mishima, Hiroyuki; Aerts, Jan; Katayama, Toshiaki; Bonnal, Raoul J P; Yoshiura, Koh-ichiro

    2012-09-21

    The University of California, Santa Cruz (UCSC) genome database is among the most used sources of genomic annotation in human and other organisms. The database offers an excellent web-based graphical user interface (the UCSC genome browser) and several means for programmatic queries. A simple application programming interface (API) in a scripting language aimed at the biologist was however not yet available. Here, we present the Ruby UCSC API, a library to access the UCSC genome database using Ruby. The API is designed as a BioRuby plug-in and built on the ActiveRecord 3 framework for the object-relational mapping, making writing SQL statements unnecessary. The current version of the API supports databases of all organisms in the UCSC genome database including human, mammals, vertebrates, deuterostomes, insects, nematodes, and yeast.The API uses the bin index-if available-when querying for genomic intervals. The API also supports genomic sequence queries using locally downloaded *.2bit files that are not stored in the official MySQL database. The API is implemented in pure Ruby and is therefore available in different environments and with different Ruby interpreters (including JRuby). Assisted by the straightforward object-oriented design of Ruby and ActiveRecord, the Ruby UCSC API will facilitate biologists to query the UCSC genome database programmatically. The API is available through the RubyGem system. Source code and documentation are available at https://github.com/misshie/bioruby-ucsc-api/ under the Ruby license. Feedback and help is provided via the website at http://rubyucscapi.userecho.com/.

  16. The Ruby UCSC API: accessing the UCSC genome database using Ruby

    Science.gov (United States)

    2012-01-01

    Background The University of California, Santa Cruz (UCSC) genome database is among the most used sources of genomic annotation in human and other organisms. The database offers an excellent web-based graphical user interface (the UCSC genome browser) and several means for programmatic queries. A simple application programming interface (API) in a scripting language aimed at the biologist was however not yet available. Here, we present the Ruby UCSC API, a library to access the UCSC genome database using Ruby. Results The API is designed as a BioRuby plug-in and built on the ActiveRecord 3 framework for the object-relational mapping, making writing SQL statements unnecessary. The current version of the API supports databases of all organisms in the UCSC genome database including human, mammals, vertebrates, deuterostomes, insects, nematodes, and yeast. The API uses the bin index—if available—when querying for genomic intervals. The API also supports genomic sequence queries using locally downloaded *.2bit files that are not stored in the official MySQL database. The API is implemented in pure Ruby and is therefore available in different environments and with different Ruby interpreters (including JRuby). Conclusions Assisted by the straightforward object-oriented design of Ruby and ActiveRecord, the Ruby UCSC API will facilitate biologists to query the UCSC genome database programmatically. The API is available through the RubyGem system. Source code and documentation are available at https://github.com/misshie/bioruby-ucsc-api/ under the Ruby license. Feedback and help is provided via the website at http://rubyucscapi.userecho.com/. PMID:22994508

  17. The Ruby UCSC API: accessing the UCSC genome database using Ruby

    Directory of Open Access Journals (Sweden)

    Mishima Hiroyuki

    2012-09-01

    Full Text Available Abstract Background The University of California, Santa Cruz (UCSC genome database is among the most used sources of genomic annotation in human and other organisms. The database offers an excellent web-based graphical user interface (the UCSC genome browser and several means for programmatic queries. A simple application programming interface (API in a scripting language aimed at the biologist was however not yet available. Here, we present the Ruby UCSC API, a library to access the UCSC genome database using Ruby. Results The API is designed as a BioRuby plug-in and built on the ActiveRecord 3 framework for the object-relational mapping, making writing SQL statements unnecessary. The current version of the API supports databases of all organisms in the UCSC genome database including human, mammals, vertebrates, deuterostomes, insects, nematodes, and yeast. The API uses the bin index—if available—when querying for genomic intervals. The API also supports genomic sequence queries using locally downloaded *.2bit files that are not stored in the official MySQL database. The API is implemented in pure Ruby and is therefore available in different environments and with different Ruby interpreters (including JRuby. Conclusions Assisted by the straightforward object-oriented design of Ruby and ActiveRecord, the Ruby UCSC API will facilitate biologists to query the UCSC genome database programmatically. The API is available through the RubyGem system. Source code and documentation are available at https://github.com/misshie/bioruby-ucsc-api/ under the Ruby license. Feedback and help is provided via the website at http://rubyucscapi.userecho.com/.

  18. KAIKObase: An integrated silkworm genome database and data mining tool

    Directory of Open Access Journals (Sweden)

    Nagaraju Javaregowda

    2009-10-01

    Full Text Available Abstract Background The silkworm, Bombyx mori, is one of the most economically important insects in many developing countries owing to its large-scale cultivation for silk production. With the development of genomic and biotechnological tools, B. mori has also become an important bioreactor for production of various recombinant proteins of biomedical interest. In 2004, two genome sequencing projects for B. mori were reported independently by Chinese and Japanese teams; however, the datasets were insufficient for building long genomic scaffolds which are essential for unambiguous annotation of the genome. Now, both the datasets have been merged and assembled through a joint collaboration between the two groups. Description Integration of the two data sets of silkworm whole-genome-shotgun sequencing by the Japanese and Chinese groups together with newly obtained fosmid- and BAC-end sequences produced the best continuity (~3.7 Mb in N50 scaffold size among the sequenced insect genomes and provided a high degree of nucleotide coverage (88% of all 28 chromosomes. In addition, a physical map of BAC contigs constructed by fingerprinting BAC clones and a SNP linkage map constructed using BAC-end sequences were available. In parallel, proteomic data from two-dimensional polyacrylamide gel electrophoresis in various tissues and developmental stages were compiled into a silkworm proteome database. Finally, a Bombyx trap database was constructed for documenting insertion positions and expression data of transposon insertion lines. Conclusion For efficient usage of genome information for functional studies, genomic sequences, physical and genetic map information and EST data were compiled into KAIKObase, an integrated silkworm genome database which consists of 4 map viewers, a gene viewer, and sequence, keyword and position search systems to display results and data at the level of nucleotide sequence, gene, scaffold and chromosome. Integration of the

  19. Genomic evolution of Saccharomyces cerevisiae under Chinese rice wine fermentation.

    Science.gov (United States)

    Li, Yudong; Zhang, Weiping; Zheng, Daoqiong; Zhou, Zhan; Yu, Wenwen; Zhang, Lei; Feng, Lifang; Liang, Xinle; Guan, Wenjun; Zhou, Jingwen; Chen, Jian; Lin, Zhenguo

    2014-09-10

    Rice wine fermentation represents a unique environment for the evolution of the budding yeast, Saccharomyces cerevisiae. To understand how the selection pressure shaped the yeast genome and gene regulation, we determined the genome sequence and transcriptome of a S. cerevisiae strain YHJ7 isolated from Chinese rice wine (Huangjiu), a popular traditional alcoholic beverage in China. By comparing the genome of YHJ7 to the lab strain S288c, a Japanese sake strain K7, and a Chinese industrial bioethanol strain YJSH1, we identified many genomic sequence and structural variations in YHJ7, which are mainly located in subtelomeric regions, suggesting that these regions play an important role in genomic evolution between strains. In addition, our comparative transcriptome analysis between YHJ7 and S288c revealed a set of differentially expressed genes, including those involved in glucose transport (e.g., HXT2, HXT7) and oxidoredutase activity (e.g., AAD10, ADH7). Interestingly, many of these genomic and transcriptional variations are directly or indirectly associated with the adaptation of YHJ7 strain to its specific niches. Our molecular evolution analysis suggested that Japanese sake strains (K7/UC5) were derived from Chinese rice wine strains (YHJ7) at least approximately 2,300 years ago, providing the first molecular evidence elucidating the origin of Japanese sake strains. Our results depict interesting insights regarding the evolution of yeast during rice wine fermentation, and provided a valuable resource for genetic engineering to improve industrial wine-making strains. © The Author(s) 2014. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

  20. Specialized microbial databases for inductive exploration of microbial genome sequences

    Directory of Open Access Journals (Sweden)

    Cabau Cédric

    2005-02-01

    Full Text Available Abstract Background The enormous amount of genome sequence data asks for user-oriented databases to manage sequences and annotations. Queries must include search tools permitting function identification through exploration of related objects. Methods The GenoList package for collecting and mining microbial genome databases has been rewritten using MySQL as the database management system. Functions that were not available in MySQL, such as nested subquery, have been implemented. Results Inductive reasoning in the study of genomes starts from "islands of knowledge", centered around genes with some known background. With this concept of "neighborhood" in mind, a modified version of the GenoList structure has been used for organizing sequence data from prokaryotic genomes of particular interest in China. GenoChore http://bioinfo.hku.hk/genochore.html, a set of 17 specialized end-user-oriented microbial databases (including one instance of Microsporidia, Encephalitozoon cuniculi, a member of Eukarya has been made publicly available. These databases allow the user to browse genome sequence and annotation data using standard queries. In addition they provide a weekly update of searches against the world-wide protein sequences data libraries, allowing one to monitor annotation updates on genes of interest. Finally, they allow users to search for patterns in DNA or protein sequences, taking into account a clustering of genes into formal operons, as well as providing extra facilities to query sequences using predefined sequence patterns. Conclusion This growing set of specialized microbial databases organize data created by the first Chinese bacterial genome programs (ThermaList, Thermoanaerobacter tencongensis, LeptoList, with two different genomes of Leptospira interrogans and SepiList, Staphylococcus epidermidis associated to related organisms for comparison.

  1. Genome Sequence of Saccharomyces cerevisiae Strain Kagoshima No. 2, Used for Brewing the Japanese Distilled Spirit Shōchū.

    Science.gov (United States)

    Mori, Kazuki; Kadooka, Chihiro; Masuda, Chika; Muto, Ai; Okutsu, Kayu; Yoshizaki, Yumiko; Takamine, Kazunori; Futagami, Taiki; Tamaki, Hisanori

    2017-10-12

    Here, we report a draft genome sequence of Saccharomyces cerevisiae strain Kagoshima no. 2, which is used for brewing shōchū, a traditional distilled spirit in Japan. The genome data will facilitate an understanding of the evolutional traits and genetic background related to the characteristic features of strain Kagoshima no. 2. Copyright © 2017 Mori et al.

  2. Enhanced annotations and features for comparing thousands of Pseudomonas genomes in the Pseudomonas genome database.

    Science.gov (United States)

    Winsor, Geoffrey L; Griffiths, Emma J; Lo, Raymond; Dhillon, Bhavjinder K; Shay, Julie A; Brinkman, Fiona S L

    2016-01-04

    The Pseudomonas Genome Database (http://www.pseudomonas.com) is well known for the application of community-based annotation approaches for producing a high-quality Pseudomonas aeruginosa PAO1 genome annotation, and facilitating whole-genome comparative analyses with other Pseudomonas strains. To aid analysis of potentially thousands of complete and draft genome assemblies, this database and analysis platform was upgraded to integrate curated genome annotations and isolate metadata with enhanced tools for larger scale comparative analysis and visualization. Manually curated gene annotations are supplemented with improved computational analyses that help identify putative drug targets and vaccine candidates or assist with evolutionary studies by identifying orthologs, pathogen-associated genes and genomic islands. The database schema has been updated to integrate isolate metadata that will facilitate more powerful analysis of genomes across datasets in the future. We continue to place an emphasis on providing high-quality updates to gene annotations through regular review of the scientific literature and using community-based approaches including a major new Pseudomonas community initiative for the assignment of high-quality gene ontology terms to genes. As we further expand from thousands of genomes, we plan to provide enhancements that will aid data visualization and analysis arising from whole-genome comparative studies including more pan-genome and population-based approaches. © The Author(s) 2015. Published by Oxford University Press on behalf of Nucleic Acids Research.

  3. MIPS PlantsDB: a database framework for comparative plant genome research.

    Science.gov (United States)

    Nussbaumer, Thomas; Martis, Mihaela M; Roessner, Stephan K; Pfeifer, Matthias; Bader, Kai C; Sharma, Sapna; Gundlach, Heidrun; Spannagl, Manuel

    2013-01-01

    The rapidly increasing amount of plant genome (sequence) data enables powerful comparative analyses and integrative approaches and also requires structured and comprehensive information resources. Databases are needed for both model and crop plant organisms and both intuitive search/browse views and comparative genomics tools should communicate the data to researchers and help them interpret it. MIPS PlantsDB (http://mips.helmholtz-muenchen.de/plant/genomes.jsp) was initially described in NAR in 2007 [Spannagl,M., Noubibou,O., Haase,D., Yang,L., Gundlach,H., Hindemitt, T., Klee,K., Haberer,G., Schoof,H. and Mayer,K.F. (2007) MIPSPlantsDB-plant database resource for integrative and comparative plant genome research. Nucleic Acids Res., 35, D834-D840] and was set up from the start to provide data and information resources for individual plant species as well as a framework for integrative and comparative plant genome research. PlantsDB comprises database instances for tomato, Medicago, Arabidopsis, Brachypodium, Sorghum, maize, rice, barley and wheat. Building up on that, state-of-the-art comparative genomics tools such as CrowsNest are integrated to visualize and investigate syntenic relationships between monocot genomes. Results from novel genome analysis strategies targeting the complex and repetitive genomes of triticeae species (wheat and barley) are provided and cross-linked with model species. The MIPS Repeat Element Database (mips-REdat) and Catalog (mips-REcat) as well as tight connections to other databases, e.g. via web services, are further important components of PlantsDB.

  4. Omics analysis of acetic acid tolerance in Saccharomyces cerevisiae.

    Science.gov (United States)

    Geng, Peng; Zhang, Liang; Shi, Gui Yang

    2017-05-01

    Acetic acid is an inhibitor in industrial processes such as wine making and bioethanol production from cellulosic hydrolysate. It causes energy depletion, inhibition of metabolic enzyme activity, growth arrest and ethanol productivity losses in Saccharomyces cerevisiae. Therefore, understanding the mechanisms of the yeast responses to acetic acid stress is essential for improving acetic acid tolerance and ethanol production. Although 329 genes associated with acetic acid tolerance have been identified in the Saccharomyces genome and included in the database ( http://www.yeastgenome.org/observable/resistance_to_acetic_acid/overview ), the cellular mechanistic responses to acetic acid remain unclear in this organism. Post-genomic approaches such as transcriptomics, proteomics, metabolomics and chemogenomics are being applied to yeast and are providing insight into the mechanisms and interactions of genes, proteins and other components that together determine complex quantitative phenotypic traits such as acetic acid tolerance. This review focuses on these omics approaches in the response to acetic acid in S. cerevisiae. Additionally, several novel strains with improved acetic acid tolerance have been engineered by modifying key genes, and the application of these strains and recently acquired knowledge to industrial processes is also discussed.

  5. Genomic diversity of Saccharomyces cerevisiae yeasts associated with alcoholic fermentation of bacanora produced by artisanal methods.

    Science.gov (United States)

    Álvarez-Ainza, M L; Zamora-Quiñonez, K A; Moreno-Ibarra, G M; Acedo-Félix, E

    2015-03-01

    Bacanora is a spirituous beverage elaborated with Agave angustifolia Haw in an artisanal process. Natural fermentation is mostly performed with native yeasts and bacteria. In this study, 228 strains of yeast like Saccharomyces were isolated from the natural alcoholic fermentation on the production of bacanora. Restriction analysis of the amplified region ITS1-5.8S-ITS2 of the ribosomal DNA genes (RFLPr) were used to confirm the genus, and 182 strains were identified as Saccharomyces cerevisiae. These strains displayed high genomic variability in their chromosomes profiles by karyotyping. Electrophoretic profiles of the strains evaluated showed a large number of chromosomes the size of which ranged between 225 and 2200 kpb approximately.

  6. The catfish genome database cBARBEL: an informatic platform for genome biology of ictalurid catfish.

    Science.gov (United States)

    Lu, Jianguo; Peatman, Eric; Yang, Qing; Wang, Shaolin; Hu, Zhiliang; Reecy, James; Kucuktas, Huseyin; Liu, Zhanjiang

    2011-01-01

    The catfish genome database, cBARBEL (abbreviated from catfish Breeder And Researcher Bioinformatics Entry Location) is an online open-access database for genome biology of ictalurid catfish (Ictalurus spp.). It serves as a comprehensive, integrative platform for all aspects of catfish genetics, genomics and related data resources. cBARBEL provides BLAST-based, fuzzy and specific search functions, visualization of catfish linkage, physical and integrated maps, a catfish EST contig viewer with SNP information overlay, and GBrowse-based organization of catfish genomic data based on sequence similarity with zebrafish chromosomes. Subsections of the database are tightly related, allowing a user with a sequence or search string of interest to navigate seamlessly from one area to another. As catfish genome sequencing proceeds and ongoing quantitative trait loci (QTL) projects bear fruit, cBARBEL will allow rapid data integration and dissemination within the catfish research community and to interested stakeholders. cBARBEL can be accessed at http://catfishgenome.org.

  7. Human Ageing Genomic Resources: new and updated databases

    Science.gov (United States)

    Tacutu, Robi; Thornton, Daniel; Johnson, Emily; Budovsky, Arie; Barardo, Diogo; Craig, Thomas; Diana, Eugene; Lehmann, Gilad; Toren, Dmitri; Wang, Jingwei; Fraifeld, Vadim E

    2018-01-01

    Abstract In spite of a growing body of research and data, human ageing remains a poorly understood process. Over 10 years ago we developed the Human Ageing Genomic Resources (HAGR), a collection of databases and tools for studying the biology and genetics of ageing. Here, we present HAGR’s main functionalities, highlighting new additions and improvements. HAGR consists of six core databases: (i) the GenAge database of ageing-related genes, in turn composed of a dataset of >300 human ageing-related genes and a dataset with >2000 genes associated with ageing or longevity in model organisms; (ii) the AnAge database of animal ageing and longevity, featuring >4000 species; (iii) the GenDR database with >200 genes associated with the life-extending effects of dietary restriction; (iv) the LongevityMap database of human genetic association studies of longevity with >500 entries; (v) the DrugAge database with >400 ageing or longevity-associated drugs or compounds; (vi) the CellAge database with >200 genes associated with cell senescence. All our databases are manually curated by experts and regularly updated to ensure a high quality data. Cross-links across our databases and to external resources help researchers locate and integrate relevant information. HAGR is freely available online (http://genomics.senescence.info/). PMID:29121237

  8. Kazusa Marker DataBase: a database for genomics, genetics, and molecular breeding in plants

    Science.gov (United States)

    Shirasawa, Kenta; Isobe, Sachiko; Tabata, Satoshi; Hirakawa, Hideki

    2014-01-01

    In order to provide useful genomic information for agronomical plants, we have established a database, the Kazusa Marker DataBase (http://marker.kazusa.or.jp). This database includes information on DNA markers, e.g., SSR and SNP markers, genetic linkage maps, and physical maps, that were developed at the Kazusa DNA Research Institute. Keyword searches for the markers, sequence data used for marker development, and experimental conditions are also available through this database. Currently, 10 plant species have been targeted: tomato (Solanum lycopersicum), pepper (Capsicum annuum), strawberry (Fragaria × ananassa), radish (Raphanus sativus), Lotus japonicus, soybean (Glycine max), peanut (Arachis hypogaea), red clover (Trifolium pratense), white clover (Trifolium repens), and eucalyptus (Eucalyptus camaldulensis). In addition, the number of plant species registered in this database will be increased as our research progresses. The Kazusa Marker DataBase will be a useful tool for both basic and applied sciences, such as genomics, genetics, and molecular breeding in crops. PMID:25320561

  9. Genomic structural variation contributes to phenotypic change of industrial bioethanol yeast Saccharomyces cerevisiae.

    Science.gov (United States)

    Zhang, Ke; Zhang, Li-Jie; Fang, Ya-Hong; Jin, Xin-Na; Qi, Lei; Wu, Xue-Chang; Zheng, Dao-Qiong

    2016-03-01

    Genomic structural variation (GSV) is a ubiquitous phenomenon observed in the genomes of Saccharomyces cerevisiae strains with different genetic backgrounds; however, the physiological and phenotypic effects of GSV are not well understood. Here, we first revealed the genetic characteristics of a widely used industrial S. cerevisiae strain, ZTW1, by whole genome sequencing. ZTW1 was identified as an aneuploidy strain and a large-scale GSV was observed in the ZTW1 genome compared with the genome of a diploid strain YJS329. These GSV events led to copy number variations (CNVs) in many chromosomal segments as well as one whole chromosome in the ZTW1 genome. Changes in the DNA dosage of certain functional genes directly affected their expression levels and the resultant ZTW1 phenotypes. Moreover, CNVs of large chromosomal regions triggered an aneuploidy stress in ZTW1. This stress decreased the proliferation ability and tolerance of ZTW1 to various stresses, while aneuploidy response stress may also provide some benefits to the fermentation performance of the yeast, including increased fermentation rates and decreased byproduct generation. This work reveals genomic characters of the bioethanol S. cerevisiae strain ZTW1 and suggests that GSV is an important kind of mutation that changes the traits of industrial S. cerevisiae strains. © FEMS 2016. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

  10. Unlimited Thirst for Genome Sequencing, Data Interpretation, and Database Usage in Genomic Era: The Road towards Fast-Track Crop Plant Improvement

    Directory of Open Access Journals (Sweden)

    Arun Prabhu Dhanapal

    2015-01-01

    Full Text Available The number of sequenced crop genomes and associated genomic resources is growing rapidly with the advent of inexpensive next generation sequencing methods. Databases have become an integral part of all aspects of science research, including basic and applied plant and animal sciences. The importance of databases keeps increasing as the volume of datasets from direct and indirect genomics, as well as other omics approaches, keeps expanding in recent years. The databases and associated web portals provide at a minimum a uniform set of tools and automated analysis across a wide range of crop plant genomes. This paper reviews some basic terms and considerations in dealing with crop plant databases utilization in advancing genomic era. The utilization of databases for variation analysis with other comparative genomics tools, and data interpretation platforms are well described. The major focus of this review is to provide knowledge on platforms and databases for genome-based investigations of agriculturally important crop plants. The utilization of these databases in applied crop improvement program is still being achieved widely; otherwise, the end for sequencing is not far away.

  11. GDR (Genome Database for Rosaceae: integrated web resources for Rosaceae genomics and genetics research

    Directory of Open Access Journals (Sweden)

    Ficklin Stephen

    2004-09-01

    Full Text Available Abstract Background Peach is being developed as a model organism for Rosaceae, an economically important family that includes fruits and ornamental plants such as apple, pear, strawberry, cherry, almond and rose. The genomics and genetics data of peach can play a significant role in the gene discovery and the genetic understanding of related species. The effective utilization of these peach resources, however, requires the development of an integrated and centralized database with associated analysis tools. Description The Genome Database for Rosaceae (GDR is a curated and integrated web-based relational database. GDR contains comprehensive data of the genetically anchored peach physical map, an annotated peach EST database, Rosaceae maps and markers and all publicly available Rosaceae sequences. Annotations of ESTs include contig assembly, putative function, simple sequence repeats, and anchored position to the peach physical map where applicable. Our integrated map viewer provides graphical interface to the genetic, transcriptome and physical mapping information. ESTs, BACs and markers can be queried by various categories and the search result sites are linked to the integrated map viewer or to the WebFPC physical map sites. In addition to browsing and querying the database, users can compare their sequences with the annotated GDR sequences via a dedicated sequence similarity server running either the BLAST or FASTA algorithm. To demonstrate the utility of the integrated and fully annotated database and analysis tools, we describe a case study where we anchored Rosaceae sequences to the peach physical and genetic map by sequence similarity. Conclusions The GDR has been initiated to meet the major deficiency in Rosaceae genomics and genetics research, namely a centralized web database and bioinformatics tools for data storage, analysis and exchange. GDR can be accessed at http://www.genome.clemson.edu/gdr/.

  12. GDR (Genome Database for Rosaceae): integrated web resources for Rosaceae genomics and genetics research.

    Science.gov (United States)

    Jung, Sook; Jesudurai, Christopher; Staton, Margaret; Du, Zhidian; Ficklin, Stephen; Cho, Ilhyung; Abbott, Albert; Tomkins, Jeffrey; Main, Dorrie

    2004-09-09

    Peach is being developed as a model organism for Rosaceae, an economically important family that includes fruits and ornamental plants such as apple, pear, strawberry, cherry, almond and rose. The genomics and genetics data of peach can play a significant role in the gene discovery and the genetic understanding of related species. The effective utilization of these peach resources, however, requires the development of an integrated and centralized database with associated analysis tools. The Genome Database for Rosaceae (GDR) is a curated and integrated web-based relational database. GDR contains comprehensive data of the genetically anchored peach physical map, an annotated peach EST database, Rosaceae maps and markers and all publicly available Rosaceae sequences. Annotations of ESTs include contig assembly, putative function, simple sequence repeats, and anchored position to the peach physical map where applicable. Our integrated map viewer provides graphical interface to the genetic, transcriptome and physical mapping information. ESTs, BACs and markers can be queried by various categories and the search result sites are linked to the integrated map viewer or to the WebFPC physical map sites. In addition to browsing and querying the database, users can compare their sequences with the annotated GDR sequences via a dedicated sequence similarity server running either the BLAST or FASTA algorithm. To demonstrate the utility of the integrated and fully annotated database and analysis tools, we describe a case study where we anchored Rosaceae sequences to the peach physical and genetic map by sequence similarity. The GDR has been initiated to meet the major deficiency in Rosaceae genomics and genetics research, namely a centralized web database and bioinformatics tools for data storage, analysis and exchange. GDR can be accessed at http://www.genome.clemson.edu/gdr/.

  13. Viral Genome DataBase: storing and analyzing genes and proteins from complete viral genomes.

    Science.gov (United States)

    Hiscock, D; Upton, C

    2000-05-01

    The Viral Genome DataBase (VGDB) contains detailed information of the genes and predicted protein sequences from 15 completely sequenced genomes of large (&100 kb) viruses (2847 genes). The data that is stored includes DNA sequence, protein sequence, GenBank and user-entered notes, molecular weight (MW), isoelectric point (pI), amino acid content, A + T%, nucleotide frequency, dinucleotide frequency and codon use. The VGDB is a mySQL database with a user-friendly JAVA GUI. Results of queries can be easily sorted by any of the individual parameters. The software and additional figures and information are available at http://athena.bioc.uvic.ca/genomes/index.html .

  14. Nencki Genomics Database--Ensembl funcgen enhanced with intersections, user data and genome-wide TFBS motifs.

    Science.gov (United States)

    Krystkowiak, Izabella; Lenart, Jakub; Debski, Konrad; Kuterba, Piotr; Petas, Michal; Kaminska, Bozena; Dabrowski, Michal

    2013-01-01

    We present the Nencki Genomics Database, which extends the functionality of Ensembl Regulatory Build (funcgen) for the three species: human, mouse and rat. The key enhancements over Ensembl funcgen include the following: (i) a user can add private data, analyze them alongside the public data and manage access rights; (ii) inside the database, we provide efficient procedures for computing intersections between regulatory features and for mapping them to the genes. To Ensembl funcgen-derived data, which include data from ENCODE, we add information on conserved non-coding (putative regulatory) sequences, and on genome-wide occurrence of transcription factor binding site motifs from the current versions of two major motif libraries, namely, Jaspar and Transfac. The intersections and mapping to the genes are pre-computed for the public data, and the result of any procedure run on the data added by the users is stored back into the database, thus incrementally increasing the body of pre-computed data. As the Ensembl funcgen schema for the rat is currently not populated, our database is the first database of regulatory features for this frequently used laboratory animal. The database is accessible without registration using the mysql client: mysql -h database.nencki-genomics.org -u public. Registration is required only to add or access private data. A WSDL webservice provides access to the database from any SOAP client, including the Taverna Workbench with a graphical user interface.

  15. Xylella fastidiosa comparative genomic database is an information resource to explore the annotation, genomic features, and biology of different strains

    Directory of Open Access Journals (Sweden)

    Alessandro M. Varani

    2012-01-01

    Full Text Available The Xylella fastidiosa comparative genomic database is a scientific resource with the aim to provide a user-friendly interface for accessing high-quality manually curated genomic annotation and comparative sequence analysis, as well as for identifying and mapping prophage-like elements, a marked feature of Xylella genomes. Here we describe a database and tools for exploring the biology of this important plant pathogen. The hallmarks of this database are the high quality genomic annotation, the functional and comparative genomic analysis and the identification and mapping of prophage-like elements. It is available from web site http://www.xylella.lncc.br.

  16. MBGD update 2015: microbial genome database for flexible ortholog analysis utilizing a diverse set of genomic data.

    Science.gov (United States)

    Uchiyama, Ikuo; Mihara, Motohiro; Nishide, Hiroyo; Chiba, Hirokazu

    2015-01-01

    The microbial genome database for comparative analysis (MBGD) (available at http://mbgd.genome.ad.jp/) is a comprehensive ortholog database for flexible comparative analysis of microbial genomes, where the users are allowed to create an ortholog table among any specified set of organisms. Because of the rapid increase in microbial genome data owing to the next-generation sequencing technology, it becomes increasingly challenging to maintain high-quality orthology relationships while allowing the users to incorporate the latest genomic data available into an analysis. Because many of the recently accumulating genomic data are draft genome sequences for which some complete genome sequences of the same or closely related species are available, MBGD now stores draft genome data and allows the users to incorporate them into a user-specific ortholog database using the MyMBGD functionality. In this function, draft genome data are incorporated into an existing ortholog table created only from the complete genome data in an incremental manner to prevent low-quality draft data from affecting clustering results. In addition, to provide high-quality orthology relationships, the standard ortholog table containing all the representative genomes, which is first created by the rapid classification program DomClust, is now refined using DomRefine, a recently developed program for improving domain-level clustering using multiple sequence alignment information. © The Author(s) 2014. Published by Oxford University Press on behalf of Nucleic Acids Research.

  17. GDR (Genome Database for Rosaceae): integrated web-database for Rosaceae genomics and genetics data.

    Science.gov (United States)

    Jung, Sook; Staton, Margaret; Lee, Taein; Blenda, Anna; Svancara, Randall; Abbott, Albert; Main, Dorrie

    2008-01-01

    The Genome Database for Rosaceae (GDR) is a central repository of curated and integrated genetics and genomics data of Rosaceae, an economically important family which includes apple, cherry, peach, pear, raspberry, rose and strawberry. GDR contains annotated databases of all publicly available Rosaceae ESTs, the genetically anchored peach physical map, Rosaceae genetic maps and comprehensively annotated markers and traits. The ESTs are assembled to produce unigene sets of each genus and the entire Rosaceae. Other annotations include putative function, microsatellites, open reading frames, single nucleotide polymorphisms, gene ontology terms and anchored map position where applicable. Most of the published Rosaceae genetic maps can be viewed and compared through CMap, the comparative map viewer. The peach physical map can be viewed using WebFPC/WebChrom, and also through our integrated GDR map viewer, which serves as a portal to the combined genetic, transcriptome and physical mapping information. ESTs, BACs, markers and traits can be queried by various categories and the search result sites are linked to the mapping visualization tools. GDR also provides online analysis tools such as a batch BLAST/FASTA server for the GDR datasets, a sequence assembly server and microsatellite and primer detection tools. GDR is available at http://www.rosaceae.org.

  18. MIPS: curated databases and comprehensive secondary data resources in 2010.

    Science.gov (United States)

    Mewes, H Werner; Ruepp, Andreas; Theis, Fabian; Rattei, Thomas; Walter, Mathias; Frishman, Dmitrij; Suhre, Karsten; Spannagl, Manuel; Mayer, Klaus F X; Stümpflen, Volker; Antonov, Alexey

    2011-01-01

    The Munich Information Center for Protein Sequences (MIPS at the Helmholtz Center for Environmental Health, Neuherberg, Germany) has many years of experience in providing annotated collections of biological data. Selected data sets of high relevance, such as model genomes, are subjected to careful manual curation, while the bulk of high-throughput data is annotated by automatic means. High-quality reference resources developed in the past and still actively maintained include Saccharomyces cerevisiae, Neurospora crassa and Arabidopsis thaliana genome databases as well as several protein interaction data sets (MPACT, MPPI and CORUM). More recent projects are PhenomiR, the database on microRNA-related phenotypes, and MIPS PlantsDB for integrative and comparative plant genome research. The interlinked resources SIMAP and PEDANT provide homology relationships as well as up-to-date and consistent annotation for 38,000,000 protein sequences. PPLIPS and CCancer are versatile tools for proteomics and functional genomics interfacing to a database of compilations from gene lists extracted from literature. A novel literature-mining tool, EXCERBT, gives access to structured information on classified relations between genes, proteins, phenotypes and diseases extracted from Medline abstracts by semantic analysis. All databases described here, as well as the detailed descriptions of our projects can be accessed through the MIPS WWW server (http://mips.helmholtz-muenchen.de).

  19. PairWise Neighbours database: overlaps and spacers among prokaryote genomes

    Directory of Open Access Journals (Sweden)

    Garcia-Vallvé Santiago

    2009-06-01

    Full Text Available Abstract Background Although prokaryotes live in a variety of habitats and possess different metabolic and genomic complexity, they have several genomic architectural features in common. The overlapping genes are a common feature of the prokaryote genomes. The overlapping lengths tend to be short because as the overlaps become longer they have more risk of deleterious mutations. The spacers between genes tend to be short too because of the tendency to reduce the non coding DNA among prokaryotes. However they must be long enough to maintain essential regulatory signals such as the Shine-Dalgarno (SD sequence, which is responsible of an efficient translation. Description PairWise Neighbours is an interactive and intuitive database used for retrieving information about the spacers and overlapping genes among bacterial and archaeal genomes. It contains 1,956,294 gene pairs from 678 fully sequenced prokaryote genomes and is freely available at the URL http://genomes.urv.cat/pwneigh. This database provides information about the overlaps and their conservation across species. Furthermore, it allows the wide analysis of the intergenic regions providing useful information such as the location and strength of the SD sequence. Conclusion There are experiments and bioinformatic analysis that rely on correct annotations of the initiation site. Therefore, a database that studies the overlaps and spacers among prokaryotes appears to be desirable. PairWise Neighbours database permits the reliability analysis of the overlapping structures and the study of the SD presence and location among the adjacent genes, which may help to check the annotation of the initiation sites.

  20. Gramene database: Navigating plant comparative genomics resources

    Directory of Open Access Journals (Sweden)

    Parul Gupta

    2016-11-01

    Full Text Available Gramene (http://www.gramene.org is an online, open source, curated resource for plant comparative genomics and pathway analysis designed to support researchers working in plant genomics, breeding, evolutionary biology, system biology, and metabolic engineering. It exploits phylogenetic relationships to enrich the annotation of genomic data and provides tools to perform powerful comparative analyses across a wide spectrum of plant species. It consists of an integrated portal for querying, visualizing and analyzing data for 44 plant reference genomes, genetic variation data sets for 12 species, expression data for 16 species, curated rice pathways and orthology-based pathway projections for 66 plant species including various crops. Here we briefly describe the functions and uses of the Gramene database.

  1. GenomeRNAi: a database for cell-based RNAi phenotypes.

    Science.gov (United States)

    Horn, Thomas; Arziman, Zeynep; Berger, Juerg; Boutros, Michael

    2007-01-01

    RNA interference (RNAi) has emerged as a powerful tool to generate loss-of-function phenotypes in a variety of organisms. Combined with the sequence information of almost completely annotated genomes, RNAi technologies have opened new avenues to conduct systematic genetic screens for every annotated gene in the genome. As increasing large datasets of RNAi-induced phenotypes become available, an important challenge remains the systematic integration and annotation of functional information. Genome-wide RNAi screens have been performed both in Caenorhabditis elegans and Drosophila for a variety of phenotypes and several RNAi libraries have become available to assess phenotypes for almost every gene in the genome. These screens were performed using different types of assays from visible phenotypes to focused transcriptional readouts and provide a rich data source for functional annotation across different species. The GenomeRNAi database provides access to published RNAi phenotypes obtained from cell-based screens and maps them to their genomic locus, including possible non-specific regions. The database also gives access to sequence information of RNAi probes used in various screens. It can be searched by phenotype, by gene, by RNAi probe or by sequence and is accessible at http://rnai.dkfz.de.

  2. Brassica database (BRAD) version 2.0: integrating and mining Brassicaceae species genomic resources.

    Science.gov (United States)

    Wang, Xiaobo; Wu, Jian; Liang, Jianli; Cheng, Feng; Wang, Xiaowu

    2015-01-01

    The Brassica database (BRAD) was built initially to assist users apply Brassica rapa and Arabidopsis thaliana genomic data efficiently to their research. However, many Brassicaceae genomes have been sequenced and released after its construction. These genomes are rich resources for comparative genomics, gene annotation and functional evolutionary studies of Brassica crops. Therefore, we have updated BRAD to version 2.0 (V2.0). In BRAD V2.0, 11 more Brassicaceae genomes have been integrated into the database, namely those of Arabidopsis lyrata, Aethionema arabicum, Brassica oleracea, Brassica napus, Camelina sativa, Capsella rubella, Leavenworthia alabamica, Sisymbrium irio and three extremophiles Schrenkiella parvula, Thellungiella halophila and Thellungiella salsuginea. BRAD V2.0 provides plots of syntenic genomic fragments between pairs of Brassicaceae species, from the level of chromosomes to genomic blocks. The Generic Synteny Browser (GBrowse_syn), a module of the Genome Browser (GBrowse), is used to show syntenic relationships between multiple genomes. Search functions for retrieving syntenic and non-syntenic orthologs, as well as their annotation and sequences are also provided. Furthermore, genome and annotation information have been imported into GBrowse so that all functional elements can be visualized in one frame. We plan to continually update BRAD by integrating more Brassicaceae genomes into the database. Database URL: http://brassicadb.org/brad/. © The Author(s) 2015. Published by Oxford University Press.

  3. An Open Access Database of Genome-wide Association Results

    Directory of Open Access Journals (Sweden)

    Johnson Andrew D

    2009-01-01

    Full Text Available Abstract Background The number of genome-wide association studies (GWAS is growing rapidly leading to the discovery and replication of many new disease loci. Combining results from multiple GWAS datasets may potentially strengthen previous conclusions and suggest new disease loci, pathways or pleiotropic genes. However, no database or centralized resource currently exists that contains anywhere near the full scope of GWAS results. Methods We collected available results from 118 GWAS articles into a database of 56,411 significant SNP-phenotype associations and accompanying information, making this database freely available here. In doing so, we met and describe here a number of challenges to creating an open access database of GWAS results. Through preliminary analyses and characterization of available GWAS, we demonstrate the potential to gain new insights by querying a database across GWAS. Results Using a genomic bin-based density analysis to search for highly associated regions of the genome, positive control loci (e.g., MHC loci were detected with high sensitivity. Likewise, an analysis of highly repeated SNPs across GWAS identified replicated loci (e.g., APOE, LPL. At the same time we identified novel, highly suggestive loci for a variety of traits that did not meet genome-wide significant thresholds in prior analyses, in some cases with strong support from the primary medical genetics literature (SLC16A7, CSMD1, OAS1, suggesting these genes merit further study. Additional adjustment for linkage disequilibrium within most regions with a high density of GWAS associations did not materially alter our findings. Having a centralized database with standardized gene annotation also allowed us to examine the representation of functional gene categories (gene ontologies containing one or more associations among top GWAS results. Genes relating to cell adhesion functions were highly over-represented among significant associations (p -14, a finding

  4. EuMicroSatdb: A database for microsatellites in the sequenced genomes of eukaryotes

    Directory of Open Access Journals (Sweden)

    Grover Atul

    2007-07-01

    Full Text Available Abstract Background Microsatellites have immense utility as molecular markers in different fields like genome characterization and mapping, phylogeny and evolutionary biology. Existing microsatellite databases are of limited utility for experimental and computational biologists with regard to their content and information output. EuMicroSatdb (Eukaryotic MicroSatellite database http://ipu.ac.in/usbt/EuMicroSatdb.htm is a web based relational database for easy and efficient positional mining of microsatellites from sequenced eukaryotic genomes. Description A user friendly web interface has been developed for microsatellite data retrieval using Active Server Pages (ASP. The backend database codes for data extraction and assembly have been written using Perl based scripts and C++. Precise need based microsatellites data retrieval is possible using different input parameters like microsatellite type (simple perfect or compound perfect, repeat unit length (mono- to hexa-nucleotide, repeat number, microsatellite length and chromosomal location in the genome. Furthermore, information about clustering of different microsatellites in the genome can also be retrieved. Finally, to facilitate primer designing for PCR amplification of any desired microsatellite locus, 200 bp upstream and downstream sequences are provided. Conclusion The database allows easy systematic retrieval of comprehensive information about simple and compound microsatellites, microsatellite clusters and their locus coordinates in 31 sequenced eukaryotic genomes. The information content of the database is useful in different areas of research like gene tagging, genome mapping, population genetics, germplasm characterization and in understanding microsatellite dynamics in eukaryotic genomes.

  5. BBGD: an online database for blueberry genomic data

    Directory of Open Access Journals (Sweden)

    Matthews Benjamin F

    2007-01-01

    Full Text Available Abstract Background Blueberry is a member of the Ericaceae family, which also includes closely related cranberry and more distantly related rhododendron, azalea, and mountain laurel. Blueberry is a major berry crop in the United States, and one that has great nutritional and economical value. Extreme low temperatures, however, reduce crop yield and cause major losses to US farmers. A better understanding of the genes and biochemical pathways that are up- or down-regulated during cold acclimation is needed to produce blueberry cultivars with enhanced cold hardiness. To that end, the blueberry genomics database (BBDG was developed. Along with the analysis tools and web-based query interfaces, the database serves both the broader Ericaceae research community and the blueberry research community specifically by making available ESTs and gene expression data in searchable formats and in elucidating the underlying mechanisms of cold acclimation and freeze tolerance in blueberry. Description BBGD is the world's first database for blueberry genomics. BBGD is both a sequence and gene expression database. It stores both EST and microarray data and allows scientists to correlate expression profiles with gene function. BBGD is a public online database. Presently, the main focus of the database is the identification of genes in blueberry that are significantly induced or suppressed after low temperature exposure. Conclusion By using the database, researchers have developed EST-based markers for mapping and have identified a number of "candidate" cold tolerance genes that are highly expressed in blueberry flower buds after exposure to low temperatures.

  6. The Genome Sequence of Saccharomyces eubayanus and the Domestication of Lager-Brewing Yeasts

    Science.gov (United States)

    Baker, EmilyClare; Wang, Bing; Bellora, Nicolas; Peris, David; Hulfachor, Amanda Beth; Koshalek, Justin A.; Adams, Marie; Libkind, Diego; Hittinger, Chris Todd

    2015-01-01

    The dramatic phenotypic changes that occur in organisms during domestication leave indelible imprints on their genomes. Although many domesticated plants and animals have been systematically compared with their wild genetic stocks, the molecular and genomic processes underlying fungal domestication have received less attention. Here, we present a nearly complete genome assembly for the recently described yeast species Saccharomyces eubayanus and compare it to the genomes of multiple domesticated alloploid hybrids of S. eubayanus × S. cerevisiae (S. pastorianus syn. S. carlsbergensis), which are used to brew lager-style beers. We find that the S. eubayanus subgenomes of lager-brewing yeasts have experienced increased rates of evolution since hybridization, and that certain genes involved in metabolism may have been particularly affected. Interestingly, the S. eubayanus subgenome underwent an especially strong shift in selection regimes, consistent with more extensive domestication of the S. cerevisiae parent prior to hybridization. In contrast to recent proposals that lager-brewing yeasts were domesticated following a single hybridization event, the radically different neutral site divergences between the subgenomes of the two major lager yeast lineages strongly favor at least two independent origins for the S. cerevisiae × S. eubayanus hybrids that brew lager beers. Our findings demonstrate how this industrially important hybrid has been domesticated along similar evolutionary trajectories on multiple occasions. PMID:26269586

  7. A Ruby API to query the Ensembl database for genomic features.

    Science.gov (United States)

    Strozzi, Francesco; Aerts, Jan

    2011-04-01

    The Ensembl database makes genomic features available via its Genome Browser. It is also possible to access the underlying data through a Perl API for advanced querying. We have developed a full-featured Ruby API to the Ensembl databases, providing the same functionality as the Perl interface with additional features. A single Ruby API is used to access different releases of the Ensembl databases and is also able to query multi-species databases. Most functionality of the API is provided using the ActiveRecord pattern. The library depends on introspection to make it release independent. The API is available through the Rubygem system and can be installed with the command gem install ruby-ensembl-api.

  8. PGSB/MIPS PlantsDB Database Framework for the Integration and Analysis of Plant Genome Data.

    Science.gov (United States)

    Spannagl, Manuel; Nussbaumer, Thomas; Bader, Kai; Gundlach, Heidrun; Mayer, Klaus F X

    2017-01-01

    Plant Genome and Systems Biology (PGSB), formerly Munich Institute for Protein Sequences (MIPS) PlantsDB, is a database framework for the integration and analysis of plant genome data, developed and maintained for more than a decade now. Major components of that framework are genome databases and analysis resources focusing on individual (reference) genomes providing flexible and intuitive access to data. Another main focus is the integration of genomes from both model and crop plants to form a scaffold for comparative genomics, assisted by specialized tools such as the CrowsNest viewer to explore conserved gene order (synteny). Data exchange and integrated search functionality with/over many plant genome databases is provided within the transPLANT project.

  9. Uniform standards for genome databases in forest and fruit trees

    Science.gov (United States)

    TreeGenes and tfGDR serve the international forestry and fruit tree genomics research communities, respectively. These databases hold similar sequence data and provide resources for the submission and recovery of this information in order to enable comparative genomics research. Large-scale genotype...

  10. MIPS: a database for protein sequences, homology data and yeast genome information.

    Science.gov (United States)

    Mewes, H W; Albermann, K; Heumann, K; Liebl, S; Pfeiffer, F

    1997-01-01

    The MIPS group (Martinsried Institute for Protein Sequences) at the Max-Planck-Institute for Biochemistry, Martinsried near Munich, Germany, collects, processes and distributes protein sequence data within the framework of the tripartite association of the PIR-International Protein Sequence Database (,). MIPS contributes nearly 50% of the data input to the PIR-International Protein Sequence Database. The database is distributed on CD-ROM together with PATCHX, an exhaustive supplement of unique, unverified protein sequences from external sources compiled by MIPS. Through its WWW server (http://www.mips.biochem.mpg.de/ ) MIPS permits internet access to sequence databases, homology data and to yeast genome information. (i) Sequence similarity results from the FASTA program () are stored in the FASTA database for all proteins from PIR-International and PATCHX. The database is dynamically maintained and permits instant access to FASTA results. (ii) Starting with FASTA database queries, proteins have been classified into families and superfamilies (PROT-FAM). (iii) The HPT (hashed position tree) data structure () developed at MIPS is a new approach for rapid sequence and pattern searching. (iv) MIPS provides access to the sequence and annotation of the complete yeast genome (), the functional classification of yeast genes (FunCat) and its graphical display, the 'Genome Browser' (). A CD-ROM based on the JAVA programming language providing dynamic interactive access to the yeast genome and the related protein sequences has been compiled and is available on request. PMID:9016498

  11. Requirements and standards for organelle genome databases

    Energy Technology Data Exchange (ETDEWEB)

    Boore, Jeffrey L.

    2006-01-09

    Mitochondria and plastids (collectively called organelles)descended from prokaryotes that adopted an intracellular, endosymbioticlifestyle within early eukaryotes. Comparisons of their remnant genomesaddress a wide variety of biological questions, especially when includingthe genomes of their prokaryotic relatives and the many genes transferredto the eukaryotic nucleus during the transitions from endosymbiont toorganelle. The pace of producing complete organellar genome sequences nowmakes it unfeasible to do broad comparisons using the primary literatureand, even if it were feasible, it is now becoming uncommon for journalsto accept detailed descriptions of genome-level features. Unfortunatelyno database is currently useful for this task, since they have littlestandardization and are riddled with error. Here I outline what iscurrently wrong and what must be done to make this data useful to thescientific community.

  12. License - TMBETA-GENOME | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available List Contact us TMBETA-GENOME License License to Use This Database Last updated : 2015/03/09 You may use this database... the license terms regarding the use of this database and the requirements you must follow in using this database.... The license for this database is specified in the Creative Commons Attribu...tion-Share Alike 2.1 Japan . If you use data from this database, please be sure attribute this database as f....1 Japan . The summary of the Creative Commons Attribution-Share Alike 2.1 Japan is found here . With regard to this database

  13. BRAD, the genetics and genomics database for Brassica plants

    Directory of Open Access Journals (Sweden)

    Li Pingxia

    2011-10-01

    Full Text Available Abstract Background Brassica species include both vegetable and oilseed crops, which are very important to the daily life of common human beings. Meanwhile, the Brassica species represent an excellent system for studying numerous aspects of plant biology, specifically for the analysis of genome evolution following polyploidy, so it is also very important for scientific research. Now, the genome of Brassica rapa has already been assembled, it is the time to do deep mining of the genome data. Description BRAD, the Brassica database, is a web-based resource focusing on genome scale genetic and genomic data for important Brassica crops. BRAD was built based on the first whole genome sequence and on further data analysis of the Brassica A genome species, Brassica rapa (Chiifu-401-42. It provides datasets, such as the complete genome sequence of B. rapa, which was de novo assembled from Illumina GA II short reads and from BAC clone sequences, predicted genes and associated annotations, non coding RNAs, transposable elements (TE, B. rapa genes' orthologous to those in A. thaliana, as well as genetic markers and linkage maps. BRAD offers useful searching and data mining tools, including search across annotation datasets, search for syntenic or non-syntenic orthologs, and to search the flanking regions of a certain target, as well as the tools of BLAST and Gbrowse. BRAD allows users to enter almost any kind of information, such as a B. rapa or A. thaliana gene ID, physical position or genetic marker. Conclusion BRAD, a new database which focuses on the genetics and genomics of the Brassica plants has been developed, it aims at helping scientists and breeders to fully and efficiently use the information of genome data of Brassica plants. BRAD will be continuously updated and can be accessed through http://brassicadb.org.

  14. Properties of promoters cloned randomly from the Saccharomyces cerevisiae genome.

    Science.gov (United States)

    Santangelo, G M; Tornow, J; McLaughlin, C S; Moldave, K

    1988-01-01

    Promoters were isolated at random from the genome of Saccharomyces cerevisiae by using a plasmid that contains a divergently arrayed pair of promoterless reporter genes. A comprehensive library was constructed by inserting random (DNase I-generated) fragments into the intergenic region upstream from the reporter genes. Simple in vivo assays for either reporter gene product (alcohol dehydrogenase or beta-galactosidase) allowed the rapid identification of promoters from among these random fragments. Poly(dA-dT) homopolymer tracts were present in three of five randomly cloned promoters. With two exceptions, each RNA start site detected was 40 to 100 base pairs downstream from a TATA element. All of the randomly cloned promoters were capable of activating reporter gene transcription bidirectionally. Interestingly, one of the promoter fragments originated in a region of the S. cerevisiae rDNA spacer; regulated divergent transcription (presumably by RNA polymerase II) initiated in the same region. Images PMID:2847031

  15. Building a genome database using an object-oriented approach.

    Science.gov (United States)

    Barbasiewicz, Anna; Liu, Lin; Lang, B Franz; Burger, Gertraud

    2002-01-01

    GOBASE is a relational database that integrates data associated with mitochondria and chloroplasts. The most important data in GOBASE, i. e., molecular sequences and taxonomic information, are obtained from the public sequence data repository at the National Center for Biotechnology Information (NCBI), and are validated by our experts. Maintaining a curated genomic database comes with a towering labor cost, due to the shear volume of available genomic sequences and the plethora of annotation errors and omissions in records retrieved from public repositories. Here we describe our approach to increase automation of the database population process, thereby reducing manual intervention. As a first step, we used Unified Modeling Language (UML) to construct a list of potential errors. Each case was evaluated independently, and an expert solution was devised, and represented as a diagram. Subsequently, the UML diagrams were used as templates for writing object-oriented automation programs in the Java programming language.

  16. Evaluating the Cassandra NoSQL Database Approach for Genomic Data Persistency

    Directory of Open Access Journals (Sweden)

    Rodrigo Aniceto

    2015-01-01

    Full Text Available Rapid advances in high-throughput sequencing techniques have created interesting computational challenges in bioinformatics. One of them refers to management of massive amounts of data generated by automatic sequencers. We need to deal with the persistency of genomic data, particularly storing and analyzing these large-scale processed data. To find an alternative to the frequently considered relational database model becomes a compelling task. Other data models may be more effective when dealing with a very large amount of nonconventional data, especially for writing and retrieving operations. In this paper, we discuss the Cassandra NoSQL database approach for storing genomic data. We perform an analysis of persistency and I/O operations with real data, using the Cassandra database system. We also compare the results obtained with a classical relational database system and another NoSQL database approach, MongoDB.

  17. Evaluating the Cassandra NoSQL Database Approach for Genomic Data Persistency

    Science.gov (United States)

    Aniceto, Rodrigo; Xavier, Rene; Guimarães, Valeria; Hondo, Fernanda; Holanda, Maristela; Walter, Maria Emilia; Lifschitz, Sérgio

    2015-01-01

    Rapid advances in high-throughput sequencing techniques have created interesting computational challenges in bioinformatics. One of them refers to management of massive amounts of data generated by automatic sequencers. We need to deal with the persistency of genomic data, particularly storing and analyzing these large-scale processed data. To find an alternative to the frequently considered relational database model becomes a compelling task. Other data models may be more effective when dealing with a very large amount of nonconventional data, especially for writing and retrieving operations. In this paper, we discuss the Cassandra NoSQL database approach for storing genomic data. We perform an analysis of persistency and I/O operations with real data, using the Cassandra database system. We also compare the results obtained with a classical relational database system and another NoSQL database approach, MongoDB. PMID:26558254

  18. Evaluating the Cassandra NoSQL Database Approach for Genomic Data Persistency.

    Science.gov (United States)

    Aniceto, Rodrigo; Xavier, Rene; Guimarães, Valeria; Hondo, Fernanda; Holanda, Maristela; Walter, Maria Emilia; Lifschitz, Sérgio

    2015-01-01

    Rapid advances in high-throughput sequencing techniques have created interesting computational challenges in bioinformatics. One of them refers to management of massive amounts of data generated by automatic sequencers. We need to deal with the persistency of genomic data, particularly storing and analyzing these large-scale processed data. To find an alternative to the frequently considered relational database model becomes a compelling task. Other data models may be more effective when dealing with a very large amount of nonconventional data, especially for writing and retrieving operations. In this paper, we discuss the Cassandra NoSQL database approach for storing genomic data. We perform an analysis of persistency and I/O operations with real data, using the Cassandra database system. We also compare the results obtained with a classical relational database system and another NoSQL database approach, MongoDB.

  19. Considerations for creating and annotating the budding yeast Genome Map at SGD: a progress report.

    Science.gov (United States)

    Chan, Esther T; Cherry, J Michael

    2012-01-01

    The Saccharomyces Genome Database (SGD) is compiling and annotating a comprehensive catalogue of functional sequence elements identified in the budding yeast genome. Recent advances in deep sequencing technologies have enabled for example, global analyses of transcription profiling and assembly of maps of transcription factor occupancy and higher order chromatin organization, at nucleotide level resolution. With this growing influx of published genome-scale data, come new challenges for their storage, display, analysis and integration. Here, we describe SGD's progress in the creation of a consolidated resource for genome sequence elements in the budding yeast, the considerations taken in its design and the lessons learned thus far. The data within this collection can be accessed at http://browse.yeastgenome.org and downloaded from http://downloads.yeastgenome.org. DATABASE URL: http://www.yeastgenome.org.

  20. The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.

    OpenAIRE

    Sven Heinicke; Michael S Livstone; Charles Lu; Rose Oughtred; Fan Kang; Samuel V Angiuoli; Owen White; David Botstein; Kara Dolinski

    2007-01-01

    Many biological databases that provide comparative genomics information and tools are now available on the internet. While certainly quite useful, to our knowledge none of the existing databases combine results from multiple comparative genomics methods with manually curated information from the literature. Here we describe the Princeton Protein Orthology Database (P-POD, http://ortholog.princeton.edu), a user-friendly database system that allows users to find and visualize the phylogenetic r...

  1. Genome-wide high-resolution mapping of UV-induced mitotic recombination events in Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Yi Yin

    2013-10-01

    Full Text Available In the yeast Saccharomyces cerevisiae and most other eukaryotes, mitotic recombination is important for the repair of double-stranded DNA breaks (DSBs. Mitotic recombination between homologous chromosomes can result in loss of heterozygosity (LOH. In this study, LOH events induced by ultraviolet (UV light are mapped throughout the genome to a resolution of about 1 kb using single-nucleotide polymorphism (SNP microarrays. UV doses that have little effect on the viability of diploid cells stimulate crossovers more than 1000-fold in wild-type cells. In addition, UV stimulates recombination in G1-synchronized cells about 10-fold more efficiently than in G2-synchronized cells. Importantly, at high doses of UV, most conversion events reflect the repair of two sister chromatids that are broken at approximately the same position whereas at low doses, most conversion events reflect the repair of a single broken chromatid. Genome-wide mapping of about 380 unselected crossovers, break-induced replication (BIR events, and gene conversions shows that UV-induced recombination events occur throughout the genome without pronounced hotspots, although the ribosomal RNA gene cluster has a significantly lower frequency of crossovers.

  2. Supervised Learning for Detection of Duplicates in Genomic Sequence Databases.

    Directory of Open Access Journals (Sweden)

    Qingyu Chen

    Full Text Available First identified as an issue in 1996, duplication in biological databases introduces redundancy and even leads to inconsistency when contradictory information appears. The amount of data makes purely manual de-duplication impractical, and existing automatic systems cannot detect duplicates as precisely as can experts. Supervised learning has the potential to address such problems by building automatic systems that learn from expert curation to detect duplicates precisely and efficiently. While machine learning is a mature approach in other duplicate detection contexts, it has seen only preliminary application in genomic sequence databases.We developed and evaluated a supervised duplicate detection method based on an expert curated dataset of duplicates, containing over one million pairs across five organisms derived from genomic sequence databases. We selected 22 features to represent distinct attributes of the database records, and developed a binary model and a multi-class model. Both models achieve promising performance; under cross-validation, the binary model had over 90% accuracy in each of the five organisms, while the multi-class model maintains high accuracy and is more robust in generalisation. We performed an ablation study to quantify the impact of different sequence record features, finding that features derived from meta-data, sequence identity, and alignment quality impact performance most strongly. The study demonstrates machine learning can be an effective additional tool for de-duplication of genomic sequence databases. All Data are available as described in the supplementary material.

  3. Supervised Learning for Detection of Duplicates in Genomic Sequence Databases.

    Science.gov (United States)

    Chen, Qingyu; Zobel, Justin; Zhang, Xiuzhen; Verspoor, Karin

    2016-01-01

    First identified as an issue in 1996, duplication in biological databases introduces redundancy and even leads to inconsistency when contradictory information appears. The amount of data makes purely manual de-duplication impractical, and existing automatic systems cannot detect duplicates as precisely as can experts. Supervised learning has the potential to address such problems by building automatic systems that learn from expert curation to detect duplicates precisely and efficiently. While machine learning is a mature approach in other duplicate detection contexts, it has seen only preliminary application in genomic sequence databases. We developed and evaluated a supervised duplicate detection method based on an expert curated dataset of duplicates, containing over one million pairs across five organisms derived from genomic sequence databases. We selected 22 features to represent distinct attributes of the database records, and developed a binary model and a multi-class model. Both models achieve promising performance; under cross-validation, the binary model had over 90% accuracy in each of the five organisms, while the multi-class model maintains high accuracy and is more robust in generalisation. We performed an ablation study to quantify the impact of different sequence record features, finding that features derived from meta-data, sequence identity, and alignment quality impact performance most strongly. The study demonstrates machine learning can be an effective additional tool for de-duplication of genomic sequence databases. All Data are available as described in the supplementary material.

  4. The Genome Sequence of Saccharomyces eubayanus and the Domestication of Lager-Brewing Yeasts.

    Science.gov (United States)

    Baker, EmilyClare; Wang, Bing; Bellora, Nicolas; Peris, David; Hulfachor, Amanda Beth; Koshalek, Justin A; Adams, Marie; Libkind, Diego; Hittinger, Chris Todd

    2015-11-01

    The dramatic phenotypic changes that occur in organisms during domestication leave indelible imprints on their genomes. Although many domesticated plants and animals have been systematically compared with their wild genetic stocks, the molecular and genomic processes underlying fungal domestication have received less attention. Here, we present a nearly complete genome assembly for the recently described yeast species Saccharomyces eubayanus and compare it to the genomes of multiple domesticated alloploid hybrids of S. eubayanus × S. cerevisiae (S. pastorianus syn. S. carlsbergensis), which are used to brew lager-style beers. We find that the S. eubayanus subgenomes of lager-brewing yeasts have experienced increased rates of evolution since hybridization, and that certain genes involved in metabolism may have been particularly affected. Interestingly, the S. eubayanus subgenome underwent an especially strong shift in selection regimes, consistent with more extensive domestication of the S. cerevisiae parent prior to hybridization. In contrast to recent proposals that lager-brewing yeasts were domesticated following a single hybridization event, the radically different neutral site divergences between the subgenomes of the two major lager yeast lineages strongly favor at least two independent origins for the S. cerevisiae × S. eubayanus hybrids that brew lager beers. Our findings demonstrate how this industrially important hybrid has been domesticated along similar evolutionary trajectories on multiple occasions. © The Author 2015. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

  5. Use of Genomic Databases for Inquiry-Based Learning about Influenza

    Science.gov (United States)

    Ledley, Fred; Ndung'u, Eric

    2011-01-01

    The genome projects of the past decades have created extensive databases of biological information with applications in both research and education. We describe an inquiry-based exercise that uses one such database, the National Center for Biotechnology Information Influenza Virus Resource, to advance learning about influenza. This database…

  6. Genome-wide RNAi screen reveals the E3 SUMO-protein ligase gene SIZ1 as a novel determinant of furfural tolerance in Saccharomyces cerevisiae

    OpenAIRE

    Xiao, Han; Zhao, Huimin

    2014-01-01

    Background Furfural is a major growth inhibitor in lignocellulosic hydrolysates and improving furfural tolerance of microorganisms is critical for rapid and efficient fermentation of lignocellulosic biomass. In this study, we used the RNAi-Assisted Genome Evolution (RAGE) method to select for furfural resistant mutants of Saccharomyces cerevisiae, and identified a new determinant of furfural tolerance. Results By using a genome-wide RNAi (RNA-interference) screen in S. cerevisiae for genes in...

  7. Using relational databases for improved sequence similarity searching and large-scale genomic analyses.

    Science.gov (United States)

    Mackey, Aaron J; Pearson, William R

    2004-10-01

    Relational databases are designed to integrate diverse types of information and manage large sets of search results, greatly simplifying genome-scale analyses. Relational databases are essential for management and analysis of large-scale sequence analyses, and can also be used to improve the statistical significance of similarity searches by focusing on subsets of sequence libraries most likely to contain homologs. This unit describes using relational databases to improve the efficiency of sequence similarity searching and to demonstrate various large-scale genomic analyses of homology-related data. This unit describes the installation and use of a simple protein sequence database, seqdb_demo, which is used as a basis for the other protocols. These include basic use of the database to generate a novel sequence library subset, how to extend and use seqdb_demo for the storage of sequence similarity search results and making use of various kinds of stored search results to address aspects of comparative genomic analysis.

  8. Accessing the SEED genome databases via Web services API: tools for programmers.

    Science.gov (United States)

    Disz, Terry; Akhter, Sajia; Cuevas, Daniel; Olson, Robert; Overbeek, Ross; Vonstein, Veronika; Stevens, Rick; Edwards, Robert A

    2010-06-14

    The SEED integrates many publicly available genome sequences into a single resource. The database contains accurate and up-to-date annotations based on the subsystems concept that leverages clustering between genomes and other clues to accurately and efficiently annotate microbial genomes. The backend is used as the foundation for many genome annotation tools, such as the Rapid Annotation using Subsystems Technology (RAST) server for whole genome annotation, the metagenomics RAST server for random community genome annotations, and the annotation clearinghouse for exchanging annotations from different resources. In addition to a web user interface, the SEED also provides Web services based API for programmatic access to the data in the SEED, allowing the development of third-party tools and mash-ups. The currently exposed Web services encompass over forty different methods for accessing data related to microbial genome annotations. The Web services provide comprehensive access to the database back end, allowing any programmer access to the most consistent and accurate genome annotations available. The Web services are deployed using a platform independent service-oriented approach that allows the user to choose the most suitable programming platform for their application. Example code demonstrate that Web services can be used to access the SEED using common bioinformatics programming languages such as Perl, Python, and Java. We present a novel approach to access the SEED database. Using Web services, a robust API for access to genomics data is provided, without requiring large volume downloads all at once. The API ensures timely access to the most current datasets available, including the new genomes as soon as they come online.

  9. Investigation of mutations in the HBB gene using the 1,000 genomes database.

    Science.gov (United States)

    Carlice-Dos-Reis, Tânia; Viana, Jaime; Moreira, Fabiano Cordeiro; Cardoso, Greice de Lemos; Guerreiro, João; Santos, Sidney; Ribeiro-Dos-Santos, Ândrea

    2017-01-01

    Mutations in the HBB gene are responsible for several serious hemoglobinopathies, such as sickle cell anemia and β-thalassemia. Sickle cell anemia is one of the most common monogenic diseases worldwide. Due to its prevalence, diverse strategies have been developed for a better understanding of its molecular mechanisms. In silico analysis has been increasingly used to investigate the genotype-phenotype relationship of many diseases, and the sequences of healthy individuals deposited in the 1,000 Genomes database appear to be an excellent tool for such analysis. The objective of this study is to analyze the variations in the HBB gene in the 1,000 Genomes database, to describe the mutation frequencies in the different population groups, and to investigate the pattern of pathogenicity. The computational tool SNPEFF was used to align the data from 2,504 samples of the 1,000 Genomes database with the HG19 genome reference. The pathogenicity of each amino acid change was investigated using the databases CLINVAR, dbSNP and HbVar and five different predictors. Twenty different mutations were found in 209 healthy individuals. The African group had the highest number of individuals with mutations, and the European group had the lowest number. Thus, it is concluded that approximately 8.3% of phenotypically healthy individuals from the 1,000 Genomes database have some mutation in the HBB gene. The frequency of mutated genes was estimated at 0.042, so that the expected frequency of being homozygous or compound heterozygous for these variants in the next generation is approximately 0.002. In total, 193 subjects had a non-synonymous mutation, which 186 (7.4%) have a deleterious mutation. Considering that the 1,000 Genomes database is representative of the world's population, it can be estimated that fourteen out of every 10,000 individuals in the world will have a hemoglobinopathy in the next generation.

  10. The Genomes OnLine Database (GOLD) v.5: a metadata management system based on a four level (meta)genome project classification

    Science.gov (United States)

    Reddy, T.B.K.; Thomas, Alex D.; Stamatis, Dimitri; Bertsch, Jon; Isbandi, Michelle; Jansson, Jakob; Mallajosyula, Jyothi; Pagani, Ioanna; Lobos, Elizabeth A.; Kyrpides, Nikos C.

    2015-01-01

    The Genomes OnLine Database (GOLD; http://www.genomesonline.org) is a comprehensive online resource to catalog and monitor genetic studies worldwide. GOLD provides up-to-date status on complete and ongoing sequencing projects along with a broad array of curated metadata. Here we report version 5 (v.5) of the database. The newly designed database schema and web user interface supports several new features including the implementation of a four level (meta)genome project classification system and a simplified intuitive web interface to access reports and launch search tools. The database currently hosts information for about 19 200 studies, 56 000 Biosamples, 56 000 sequencing projects and 39 400 analysis projects. More than just a catalog of worldwide genome projects, GOLD is a manually curated, quality-controlled metadata warehouse. The problems encountered in integrating disparate and varying quality data into GOLD are briefly highlighted. GOLD fully supports and follows the Genomic Standards Consortium (GSC) Minimum Information standards. PMID:25348402

  11. The Genomes OnLine Database (GOLD) v.5: a metadata management system based on a four level (meta)genome project classification

    Energy Technology Data Exchange (ETDEWEB)

    Reddy, Tatiparthi B. K. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Thomas, Alex D. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Stamatis, Dimitri [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Bertsch, Jon [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Isbandi, Michelle [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Jansson, Jakob [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Mallajosyula, Jyothi [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Pagani, Ioanna [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Lobos, Elizabeth A. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Kyrpides, Nikos C. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); King Abdulaziz Univ., Jeddah (Saudi Arabia)

    2014-10-27

    The Genomes OnLine Database (GOLD; http://www.genomesonline.org) is a comprehensive online resource to catalog and monitor genetic studies worldwide. GOLD provides up-to-date status on complete and ongoing sequencing projects along with a broad array of curated metadata. Within this paper, we report version 5 (v.5) of the database. The newly designed database schema and web user interface supports several new features including the implementation of a four level (meta)genome project classification system and a simplified intuitive web interface to access reports and launch search tools. The database currently hosts information for about 19 200 studies, 56 000 Biosamples, 56 000 sequencing projects and 39 400 analysis projects. More than just a catalog of worldwide genome projects, GOLD is a manually curated, quality-controlled metadata warehouse. The problems encountered in integrating disparate and varying quality data into GOLD are briefly highlighted. Lastly, GOLD fully supports and follows the Genomic Standards Consortium (GSC) Minimum Information standards.

  12. EasyCloneMulti: A Set of Vectors for Simultaneous and Multiple Genomic Integrations in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Maury, Jerome; Germann, Susanne Manuela; Jacobsen, Simo Abdessamad

    2016-01-01

    Saccharomyces cerevisiae is widely used in the biotechnology industry for production of ethanol, recombinant proteins, food ingredients and other chemicals. In order to generate highly producing and stable strains, genome integration of genes encoding metabolic pathway enzymes is the preferred...... of integrative vectors, EasyCloneMulti, that enables multiple and simultaneous integration of genes in S. cerevisiae. By creating vector backbones that combine consensus sequences that aim at targeting subsets of Ty sequences and a quickly degrading selective marker, integrations at multiple genomic loci...... and a range of expression levels were obtained, as assessed with the green fluorescent protein (GFP) reporter system. The EasyCloneMulti vector set was applied to balance the expression of the rate-controlling step in the β-alanine pathway for biosynthesis of 3-hydroxypropionic acid (3HP). The best 3HP...

  13. Database Description - DGBY | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available base Description General information of database Database name DGBY Alternative name Database...EL: +81-29-838-8066 E-mail: Database classification Microarray Data and other Gene Expression Databases Orga...nism Taxonomy Name: Saccharomyces cerevisiae Taxonomy ID: 4932 Database descripti...-called phenomics). We uploaded these data on this website which is designated DGBY(Database for Gene expres...ma J, Ando A, Takagi H. Journal: Yeast. 2008 Mar;25(3):179-90. External Links: Original website information Database

  14. SoyTEdb: a comprehensive database of transposable elements in the soybean genome

    Directory of Open Access Journals (Sweden)

    Zhu Liucun

    2010-02-01

    Full Text Available Abstract Background Transposable elements are the most abundant components of all characterized genomes of higher eukaryotes. It has been documented that these elements not only contribute to the shaping and reshaping of their host genomes, but also play significant roles in regulating gene expression, altering gene function, and creating new genes. Thus, complete identification of transposable elements in sequenced genomes and construction of comprehensive transposable element databases are essential for accurate annotation of genes and other genomic components, for investigation of potential functional interaction between transposable elements and genes, and for study of genome evolution. The recent availability of the soybean genome sequence has provided an unprecedented opportunity for discovery, and structural and functional characterization of transposable elements in this economically important legume crop. Description Using a combination of structure-based and homology-based approaches, a total of 32,552 retrotransposons (Class I and 6,029 DNA transposons (Class II with clear boundaries and insertion sites were structurally annotated and clearly categorized, and a soybean transposable element database, SoyTEdb, was established. These transposable elements have been anchored in and integrated with the soybean physical map and genetic map, and are browsable and visualizable at any scale along the 20 soybean chromosomes, along with predicted genes and other sequence annotations. BLAST search and other infrastracture tools were implemented to facilitate annotation of transposable elements or fragments from soybean and other related legume species. The majority (> 95% of these elements (particularly a few hundred low-copy-number families are first described in this study. Conclusion SoyTEdb provides resources and information related to transposable elements in the soybean genome, representing the most comprehensive and the largest manually

  15. H2DB: a heritability database across multiple species by annotating trait-associated genomic loci.

    Science.gov (United States)

    Kaminuma, Eli; Fujisawa, Takatomo; Tanizawa, Yasuhiro; Sakamoto, Naoko; Kurata, Nori; Shimizu, Tokurou; Nakamura, Yasukazu

    2013-01-01

    H2DB (http://tga.nig.ac.jp/h2db/), an annotation database of genetic heritability estimates for humans and other species, has been developed as a knowledge database to connect trait-associated genomic loci. Heritability estimates have been investigated for individual species, particularly in human twin studies and plant/animal breeding studies. However, there appears to be no comprehensive heritability database for both humans and other species. Here, we introduce an annotation database for genetic heritabilities of various species that was annotated by manually curating online public resources in PUBMED abstracts and journal contents. The proposed heritability database contains attribute information for trait descriptions, experimental conditions, trait-associated genomic loci and broad- and narrow-sense heritability specifications. Annotated trait-associated genomic loci, for which most are single-nucleotide polymorphisms derived from genome-wide association studies, may be valuable resources for experimental scientists. In addition, we assigned phenotype ontologies to the annotated traits for the purposes of discussing heritability distributions based on phenotypic classifications.

  16. Genome Dynamics of Hybrid Saccharomyces cerevisiae During Vegetative and Meiotic Divisions

    Directory of Open Access Journals (Sweden)

    Abhishek Dutta

    2017-11-01

    Full Text Available Mutation and recombination are the major sources of genetic diversity in all organisms. In the baker’s yeast, all mutation rate estimates are in homozygous background. We determined the extent of genetic change through mutation and loss of heterozygosity (LOH in a heterozygous Saccharomyces cerevisiae genome during successive vegetative and meiotic divisions. We measured genome-wide LOH and base mutation rates during vegetative and meiotic divisions in a hybrid (S288c/YJM789 S. cerevisiae strain. The S288c/YJM789 hybrid showed nearly complete reduction in heterozygosity within 31 generations of meioses and improved spore viability. LOH in the meiotic lines was driven primarily by the mating of spores within the tetrad. The S288c/YJM789 hybrid lines propagated vegetatively for the same duration as the meiotic lines, showed variable LOH (from 2 to 3% and up to 35%. Two of the vegetative lines with extensive LOH showed frequent and large internal LOH tracts that suggest a high frequency of recombination repair. These results suggest significant LOH can occur in the S288c/YJM789 hybrid during vegetative propagation presumably due to return to growth events. The average base substitution rates for the vegetative lines (1.82 × 10−10 per base per division and the meiotic lines (1.22 × 10−10 per base per division are the first genome-wide mutation rate estimates for a hybrid yeast. This study therefore provides a novel context for the analysis of mutation rates (especially in the context of detecting LOH during vegetative divisions, compared to previous mutation accumulation studies in yeast that used homozygous backgrounds.

  17. Genome scan for nonadditive heterotic trait loci reveals mainly underdominant effects in Saccharomyces cerevisiae.

    Science.gov (United States)

    Laiba, Efrat; Glikaite, Ilana; Levy, Yael; Pasternak, Zohar; Fridman, Eyal

    2016-04-01

    The overdominant model of heterosis explains the superior phenotype of hybrids by synergistic allelic interaction within heterozygous loci. To map such genetic variation in yeast, we used a population doubling time dataset of Saccharomyces cerevisiae 16 × 16 diallel and searched for major contributing heterotic trait loci (HTL). Heterosis was observed for the majority of hybrids, as they surpassed their best parent growth rate. However, most of the local heterozygous loci identified by genome scan were surprisingly underdominant, i.e., reduced growth. We speculated that in these loci adverse effects on growth resulted from incompatible allelic interactions. To test this assumption, we eliminated these allelic interactions by creating hybrids with local hemizygosity for the underdominant HTLs, as well as for control random loci. Growth of hybrids was indeed elevated for most hemizygous to HTL genes but not for control genes, hence validating the results of our genome scan. Assessing the consequences of local heterozygosity by reciprocal hemizygosity and allele replacement assays revealed the influence of genetic background on the underdominant effects of HTLs. Overall, this genome-wide study on a multi-parental hybrid population provides a strong argument against single gene overdominance as a major contributor to heterosis, and favors the dominance complementation model.

  18. RICD: A rice indica cDNA database resource for rice functional genomics

    Directory of Open Access Journals (Sweden)

    Zhang Qifa

    2008-11-01

    Full Text Available Abstract Background The Oryza sativa L. indica subspecies is the most widely cultivated rice. During the last few years, we have collected over 20,000 putative full-length cDNAs and over 40,000 ESTs isolated from various cDNA libraries of two indica varieties Guangluai 4 and Minghui 63. A database of the rice indica cDNAs was therefore built to provide a comprehensive web data source for searching and retrieving the indica cDNA clones. Results Rice Indica cDNA Database (RICD is an online MySQL-PHP driven database with a user-friendly web interface. It allows investigators to query the cDNA clones by keyword, genome position, nucleotide or protein sequence, and putative function. It also provides a series of information, including sequences, protein domain annotations, similarity search results, SNPs and InDels information, and hyperlinks to gene annotation in both The Rice Annotation Project Database (RAP-DB and The TIGR Rice Genome Annotation Resource, expression atlas in RiceGE and variation report in Gramene of each cDNA. Conclusion The online rice indica cDNA database provides cDNA resource with comprehensive information to researchers for functional analysis of indica subspecies and for comparative genomics. The RICD database is available through our website http://www.ncgr.ac.cn/ricd.

  19. CTDB: An Integrated Chickpea Transcriptome Database for Functional and Applied Genomics.

    Directory of Open Access Journals (Sweden)

    Mohit Verma

    Full Text Available Chickpea is an important grain legume used as a rich source of protein in human diet. The narrow genetic diversity and limited availability of genomic resources are the major constraints in implementing breeding strategies and biotechnological interventions for genetic enhancement of chickpea. We developed an integrated Chickpea Transcriptome Database (CTDB, which provides the comprehensive web interface for visualization and easy retrieval of transcriptome data in chickpea. The database features many tools for similarity search, functional annotation (putative function, PFAM domain and gene ontology search and comparative gene expression analysis. The current release of CTDB (v2.0 hosts transcriptome datasets with high quality functional annotation from cultivated (desi and kabuli types and wild chickpea. A catalog of transcription factor families and their expression profiles in chickpea are available in the database. The gene expression data have been integrated to study the expression profiles of chickpea transcripts in major tissues/organs and various stages of flower development. The utilities, such as similarity search, ortholog identification and comparative gene expression have also been implemented in the database to facilitate comparative genomic studies among different legumes and Arabidopsis. Furthermore, the CTDB represents a resource for the discovery of functional molecular markers (microsatellites and single nucleotide polymorphisms between different chickpea types. We anticipate that integrated information content of this database will accelerate the functional and applied genomic research for improvement of chickpea. The CTDB web service is freely available at http://nipgr.res.in/ctdb.html.

  20. CRISPR–Cas system enables fast and simple genome editing of industrial Saccharomyces cerevisiae strains

    Directory of Open Access Journals (Sweden)

    Vratislav Stovicek

    2015-12-01

    Full Text Available There is a demand to develop 3rd generation biorefineries that integrate energy production with the production of higher value chemicals from renewable feedstocks. Here, robust and stress-tolerant industrial strains of Saccharomyces cerevisiae will be suitable production organisms. However, their genetic manipulation is challenging, as they are usually diploid or polyploid. Therefore, there is a need to develop more efficient genetic engineering tools. We applied a CRISPR–Cas9 system for genome editing of different industrial strains, and show simultaneous disruption of two alleles of a gene in several unrelated strains with the efficiency ranging between 65% and 78%. We also achieved simultaneous disruption and knock-in of a reporter gene, and demonstrate the applicability of the method by designing lactic acid-producing strains in a single transformation event, where insertion of a heterologous gene and disruption of two endogenous genes occurred simultaneously. Our study provides a foundation for efficient engineering of industrial yeast cell factories. Keywords: CRISPR–Cas9, Genome editing, Industrial yeast, Biorefineries, Chemical production

  1. Genomic reconstruction to improve bioethanol and ergosterol production of industrial yeast Saccharomyces cerevisiae.

    Science.gov (United States)

    Zhang, Ke; Tong, Mengmeng; Gao, Kehui; Di, Yanan; Wang, Pinmei; Zhang, Chunfang; Wu, Xuechang; Zheng, Daoqiong

    2015-02-01

    Baker's yeast (Saccharomyces cerevisiae) is the common yeast used in the fields of bread making, brewing, and bioethanol production. Growth rate, stress tolerance, ethanol titer, and byproducts yields are some of the most important agronomic traits of S. cerevisiae for industrial applications. Here, we developed a novel method of constructing S. cerevisiae strains for co-producing bioethanol and ergosterol. The genome of an industrial S. cerevisiae strain, ZTW1, was first reconstructed through treatment with an antimitotic drug followed by sporulation and hybridization. A total of 140 mutants were selected for ethanol fermentation testing, and a significant positive correlation between ergosterol content and ethanol production was observed. The highest performing mutant, ZG27, produced 7.9 % more ethanol and 43.2 % more ergosterol than ZTW1 at the end of fermentation. Chromosomal karyotyping and proteome analysis of ZG27 and ZTW1 suggested that this breeding strategy caused large-scale genome structural variations and global gene expression diversities in the mutants. Genetic manipulation further demonstrated that the altered expression activity of some genes (such as ERG1, ERG9, and ERG11) involved in ergosterol synthesis partly explained the trait improvement in ZG27.

  2. DRDB: An Online Date Palm Genomic Resource Database

    Directory of Open Access Journals (Sweden)

    Zilong He

    2017-11-01

    Full Text Available Background: Date palm (Phoenix dactylifera L. is a cultivated woody plant with agricultural and economic importance in many countries around the world. With the advantages of next generation sequencing technologies, genome sequences for many date palm cultivars have been released recently. Short sequence repeat (SSR and single nucleotide polymorphism (SNP can be identified from these genomic data, and have been proven to be very useful biomarkers in plant genome analysis and breeding.Results: Here, we first improved the date palm genome assembly using 130X of HiSeq data generated in our lab. Then 246,445 SSRs (214,901 SSRs and 31,544 compound SSRs were annotated in this genome assembly; among the SSRs, mononucleotide SSRs (58.92% were the most abundant, followed by di- (29.92%, tri- (8.14%, tetra- (2.47%, penta- (0.36%, and hexa-nucleotide SSRs (0.19%. The high-quality PCR primer pairs were designed for most (174,497; 70.81% out of total SSRs. We also annotated 6,375,806 SNPs with raw read depth≥3 in 90% cultivars. To further reduce false positive SNPs, we only kept 5,572,650 (87.40% out of total SNPs with at least 20% cultivars support for downstream analyses. The high-quality PCR primer pairs were also obtained for 4,177,778 (65.53% SNPs. We reconstructed the phylogenetic relationships among the 62 cultivars using these variants and found that they can be divided into three clusters, namely North Africa, Egypt – Sudan, and Middle East – South Asian, with Egypt – Sudan being the admixture of North Africa and Middle East – South Asian cultivars; we further confirmed these clusters using principal component analysis. Moreover, 34,346 SSRs and 4,177,778 SNPs with PCR primers were assigned to shared cultivars for cultivar classification and diversity analysis. All these SSRs, SNPs and their classification are available in our database, and can be used for cultivar identification, comparison, and molecular breeding.Conclusion:DRDB is a

  3. PSSRdb: a relational database of polymorphic simple sequence repeats extracted from prokaryotic genomes.

    Science.gov (United States)

    Kumar, Pankaj; Chaitanya, Pasumarthy S; Nagarajaram, Hampapathalu A

    2011-01-01

    PSSRdb (Polymorphic Simple Sequence Repeats database) (http://www.cdfd.org.in/PSSRdb/) is a relational database of polymorphic simple sequence repeats (PSSRs) extracted from 85 different species of prokaryotes. Simple sequence repeats (SSRs) are the tandem repeats of nucleotide motifs of the sizes 1-6 bp and are highly polymorphic. SSR mutations in and around coding regions affect transcription and translation of genes. Such changes underpin phase variations and antigenic variations seen in some bacteria. Although SSR-mediated phase variation and antigenic variations have been well-studied in some bacteria there seems a lot of other species of prokaryotes yet to be investigated for SSR mediated adaptive and other evolutionary advantages. As a part of our on-going studies on SSR polymorphism in prokaryotes we compared the genome sequences of various strains and isolates available for 85 different species of prokaryotes and extracted a number of SSRs showing length variations and created a relational database called PSSRdb. This database gives useful information such as location of PSSRs in genomes, length variation across genomes, the regions harboring PSSRs, etc. The information provided in this database is very useful for further research and analysis of SSRs in prokaryotes.

  4. Rapid storage and retrieval of genomic intervals from a relational database system using nested containment lists.

    Science.gov (United States)

    Wiley, Laura K; Sivley, R Michael; Bush, William S

    2013-01-01

    Efficient storage and retrieval of genomic annotations based on range intervals is necessary, given the amount of data produced by next-generation sequencing studies. The indexing strategies of relational database systems (such as MySQL) greatly inhibit their use in genomic annotation tasks. This has led to the development of stand-alone applications that are dependent on flat-file libraries. In this work, we introduce MyNCList, an implementation of the NCList data structure within a MySQL database. MyNCList enables the storage, update and rapid retrieval of genomic annotations from the convenience of a relational database system. Range-based annotations of 1 million variants are retrieved in under a minute, making this approach feasible for whole-genome annotation tasks. Database URL: https://github.com/bushlab/mynclist.

  5. Saccharomyces eubayanus and Saccharomyces arboricola reside in North Island native New Zealand forests.

    Science.gov (United States)

    Gayevskiy, Velimir; Goddard, Matthew R

    2016-04-01

    Saccharomyces is one of the best-studied microbial genera, but our understanding of the global distributions and evolutionary histories of its members is relatively poor. Recent studies have altered our view of Saccharomyces' origin, but a lack of sampling from the vast majority of the world precludes a holistic perspective. We evaluate alternate Gondwanan and Far East Asian hypotheses concerning the origin of these yeasts. Being part of Gondwana, and only colonized by humans in the last ∼1000 years, New Zealand represents a unique environment for testing these ideas. Genotyping and ribosomal sequencing of samples from North Island native forest parks identified a widespread population of Saccharomyces. Whole genome sequencing identified the presence of S. arboricola and S. eubayanus in New Zealand, which is the first report of S. arboricola outside Far East Asia, and also expands S. eubayanus' known distribution to include the Oceanic region. Phylogenomic approaches place the S. arboricola population as significantly diverged from the only other sequenced Chinese isolate but indicate that S. eubayanus might be a recent migrant from South America. These data tend to support the Far East Asian origin of the Saccharomyces, but the history of this group is still far from clear. © 2015 Society for Applied Microbiology and John Wiley & Sons Ltd.

  6. Genome Sequence of the Lager-Brewing Yeast Saccharomyces sp. Strain M14, Used in the High-Gravity Brewing Industry in China.

    Science.gov (United States)

    Liu, Chunfeng; Li, Qi; Niu, Chengtuo; Zheng, Feiyun; Li, Yongxian; Zhao, Yun; Yin, Xiangsheng

    2017-10-26

    Lager-brewing yeasts are mainly used for the production of lager beers. Illumina and PacBio-based sequence analyses revealed an approximate genome size of 22.8 Mb, with a GC content of 38.98%, for the Chinese lager-brewing yeast Saccharomyces sp. strain M14. Based on ab initio prediction, 9,970 coding genes were annotated. Copyright © 2017 Liu et al.

  7. Genome-wide map of Apn1 binding sites under oxidative stress in Saccharomyces cerevisiae.

    Science.gov (United States)

    Morris, Lydia P; Conley, Andrew B; Degtyareva, Natalya; Jordan, I King; Doetsch, Paul W

    2017-11-01

    The DNA is cells is continuously exposed to reactive oxygen species resulting in toxic and mutagenic DNA damage. Although the repair of oxidative DNA damage occurs primarily through the base excision repair (BER) pathway, the nucleotide excision repair (NER) pathway processes some of the same lesions. In addition, damage tolerance mechanisms, such as recombination and translesion synthesis, enable cells to tolerate oxidative DNA damage, especially when BER and NER capacities are exceeded. Thus, disruption of BER alone or disruption of BER and NER in Saccharomyces cerevisiae leads to increased mutations as well as large-scale genomic rearrangements. Previous studies demonstrated that a particular region of chromosome II is susceptible to chronic oxidative stress-induced chromosomal rearrangements, suggesting the existence of DNA damage and/or DNA repair hotspots. Here we investigated the relationship between oxidative damage and genomic instability utilizing chromatin immunoprecipitation combined with DNA microarray technology to profile DNA repair sites along yeast chromosomes under different oxidative stress conditions. We targeted the major yeast AP endonuclease Apn1 as a representative BER protein. Our results indicate that Apn1 target sequences are enriched for cytosine and guanine nucleotides. We predict that BER protects these sites in the genome because guanines and cytosines are thought to be especially susceptible to oxidative attack, thereby preventing large-scale genome destabilization from chronic accumulation of DNA damage. Information from our studies should provide insight into how regional deployment of oxidative DNA damage management systems along chromosomes protects against large-scale rearrangements. Copyright © 2017 John Wiley & Sons, Ltd. Copyright © 2017 John Wiley & Sons, Ltd.

  8. Efficient engineering of marker-free synthetic allotetraploids of Saccharomyces.

    Science.gov (United States)

    Alexander, William G; Peris, David; Pfannenstiel, Brandon T; Opulente, Dana A; Kuang, Meihua; Hittinger, Chris Todd

    2016-04-01

    Saccharomyces interspecies hybrids are critical biocatalysts in the fermented beverage industry, including in the production of lager beers, Belgian ales, ciders, and cold-fermented wines. Current methods for making synthetic interspecies hybrids are cumbersome and/or require genome modifications. We have developed a simple, robust, and efficient method for generating allotetraploid strains of prototrophic Saccharomyces without sporulation or nuclear genome manipulation. S. cerevisiae×S. eubayanus, S. cerevisiae×S. kudriavzevii, and S. cerevisiae×S. uvarum designer hybrid strains were created as synthetic lager, Belgian, and cider strains, respectively. The ploidy and hybrid nature of the strains were confirmed using flow cytometry and PCR-RFLP analysis, respectively. This method provides an efficient means for producing novel synthetic hybrids for beverage and biofuel production, as well as for constructing tetraploids to be used for basic research in evolutionary genetics and genome stability. Copyright © 2015 Elsevier Inc. All rights reserved.

  9. Industrial Relevance of Chromosomal Copy Number Variation in Saccharomyces Yeasts

    Science.gov (United States)

    Gorter de Vries, Arthur R.; Pronk, Jack T.

    2017-01-01

    ABSTRACT Chromosomal copy number variation (CCNV) plays a key role in evolution and health of eukaryotes. The unicellular yeast Saccharomyces cerevisiae is an important model for studying the generation, physiological impact, and evolutionary significance of CCNV. Fundamental studies of this yeast have contributed to an extensive set of methods for analyzing and introducing CCNV. Moreover, these studies provided insight into the balance between negative and positive impacts of CCNV in evolutionary contexts. A growing body of evidence indicates that CCNV not only frequently occurs in industrial strains of Saccharomyces yeasts but also is a key contributor to the diversity of industrially relevant traits. This notion is further supported by the frequent involvement of CCNV in industrially relevant traits acquired during evolutionary engineering. This review describes recent developments in genome sequencing and genome editing techniques and discusses how these offer opportunities to unravel contributions of CCNV in industrial Saccharomyces strains as well as to rationally engineer yeast chromosomal copy numbers and karyotypes. PMID:28341679

  10. A SNP-centric database for the investigation of the human genome

    Directory of Open Access Journals (Sweden)

    Kohane Isaac S

    2004-03-01

    Full Text Available Abstract Background Single Nucleotide Polymorphisms (SNPs are an increasingly important tool for genetic and biomedical research. Although current genomic databases contain information on several million SNPs and are growing at a very fast rate, the true value of a SNP in this context is a function of the quality of the annotations that characterize it. Retrieving and analyzing such data for a large number of SNPs often represents a major bottleneck in the design of large-scale association studies. Description SNPper is a web-based application designed to facilitate the retrieval and use of human SNPs for high-throughput research purposes. It provides a rich local database generated by combining SNP data with the Human Genome sequence and with several other data sources, and offers the user a variety of querying, visualization and data export tools. In this paper we describe the structure and organization of the SNPper database, we review the available data export and visualization options, and we describe how the architecture of SNPper and its specialized data structures support high-volume SNP analysis. Conclusions The rich annotation database and the powerful data manipulation and presentation facilities it offers make SNPper a very useful online resource for SNP research. Its success proves the great need for integrated and interoperable resources in the field of computational biology, and shows how such systems may play a critical role in supporting the large-scale computational analysis of our genome.

  11. BioQ: tracing experimental origins in public genomic databases using a novel data provenance model.

    Science.gov (United States)

    Saccone, Scott F; Quan, Jiaxi; Jones, Peter L

    2012-04-15

    Public genomic databases, which are often used to guide genetic studies of human disease, are now being applied to genomic medicine through in silico integrative genomics. These databases, however, often lack tools for systematically determining the experimental origins of the data. We introduce a new data provenance model that we have implemented in a public web application, BioQ, for assessing the reliability of the data by systematically tracing its experimental origins to the original subjects and biologics. BioQ allows investigators to both visualize data provenance as well as explore individual elements of experimental process flow using precise tools for detailed data exploration and documentation. It includes a number of human genetic variation databases such as the HapMap and 1000 Genomes projects. BioQ is freely available to the public at http://bioq.saclab.net.

  12. SNUGB: a versatile genome browser supporting comparative and functional fungal genomics

    Directory of Open Access Journals (Sweden)

    Kim Seungill

    2008-12-01

    Full Text Available Abstract Background Since the full genome sequences of Saccharomyces cerevisiae were released in 1996, genome sequences of over 90 fungal species have become publicly available. The heterogeneous formats of genome sequences archived in different sequencing centers hampered the integration of the data for efficient and comprehensive comparative analyses. The Comparative Fungal Genomics Platform (CFGP was developed to archive these data via a single standardized format that can support multifaceted and integrated analyses of the data. To facilitate efficient data visualization and utilization within and across species based on the architecture of CFGP and associated databases, a new genome browser was needed. Results The Seoul National University Genome Browser (SNUGB integrates various types of genomic information derived from 98 fungal/oomycete (137 datasets and 34 plant and animal (38 datasets species, graphically presents germane features and properties of each genome, and supports comparison between genomes. The SNUGB provides three different forms of the data presentation interface, including diagram, table, and text, and six different display options to support visualization and utilization of the stored information. Information for individual species can be quickly accessed via a new tool named the taxonomy browser. In addition, SNUGB offers four useful data annotation/analysis functions, including 'BLAST annotation.' The modular design of SNUGB makes its adoption to support other comparative genomic platforms easy and facilitates continuous expansion. Conclusion The SNUGB serves as a powerful platform supporting comparative and functional genomics within the fungal kingdom and also across other kingdoms. All data and functions are available at the web site http://genomebrowser.snu.ac.kr/.

  13. Additions, losses, and rearrangements on the evolutionary route from a reconstructed ancestor to the modern Saccharomyces cerevisiae genome.

    Directory of Open Access Journals (Sweden)

    Jonathan L Gordon

    2009-05-01

    Full Text Available Comparative genomics can be used to infer the history of genomic rearrangements that occurred during the evolution of a species. We used the principle of parsimony, applied to aligned synteny blocks from 11 yeast species, to infer the gene content and gene order that existed in the genome of an extinct ancestral yeast about 100 Mya, immediately before it underwent whole-genome duplication (WGD. The reconstructed ancestral genome contains 4,703 ordered loci on eight chromosomes. The reconstruction is complete except for the subtelomeric regions. We then inferred the series of rearrangement steps that led from this ancestor to the current Saccharomyces cerevisiae genome; relative to the ancestral genome we observe 73 inversions, 66 reciprocal translocations, and five translocations involving telomeres. Some fragile chromosomal sites were reused as evolutionary breakpoints multiple times. We identified 124 genes that have been gained by S. cerevisiae in the time since the WGD, including one that is derived from a hAT family transposon, and 88 ancestral loci at which S. cerevisiae did not retain either of the gene copies that were formed by WGD. Sites of gene gain and evolutionary breakpoints both tend to be associated with tRNA genes and, to a lesser extent, with origins of replication. Many of the gained genes in S. cerevisiae have functions associated with ethanol production, growth in hypoxic environments, or the uptake of alternative nutrient sources.

  14. EchoBASE: an integrated post-genomic database for Escherichia coli.

    Science.gov (United States)

    Misra, Raju V; Horler, Richard S P; Reindl, Wolfgang; Goryanin, Igor I; Thomas, Gavin H

    2005-01-01

    EchoBASE (http://www.ecoli-york.org) is a relational database designed to contain and manipulate information from post-genomic experiments using the model bacterium Escherichia coli K-12. Its aim is to collate information from a wide range of sources to provide clues to the functions of the approximately 1500 gene products that have no confirmed cellular function. The database is built on an enhanced annotation of the updated genome sequence of strain MG1655 and the association of experimental data with the E.coli genes and their products. Experiments that can be held within EchoBASE include proteomics studies, microarray data, protein-protein interaction data, structural data and bioinformatics studies. EchoBASE also contains annotated information on 'orphan' enzyme activities from this microbe to aid characterization of the proteins that catalyse these elusive biochemical reactions.

  15. The need for high-quality whole-genome sequence databases in microbial forensics.

    Science.gov (United States)

    Sjödin, Andreas; Broman, Tina; Melefors, Öjar; Andersson, Gunnar; Rasmusson, Birgitta; Knutsson, Rickard; Forsman, Mats

    2013-09-01

    Microbial forensics is an important part of a strengthened capability to respond to biocrime and bioterrorism incidents to aid in the complex task of distinguishing between natural outbreaks and deliberate acts. The goal of a microbial forensic investigation is to identify and criminally prosecute those responsible for a biological attack, and it involves a detailed analysis of the weapon--that is, the pathogen. The recent development of next-generation sequencing (NGS) technologies has greatly increased the resolution that can be achieved in microbial forensic analyses. It is now possible to identify, quickly and in an unbiased manner, previously undetectable genome differences between closely related isolates. This development is particularly relevant for the most deadly bacterial diseases that are caused by bacterial lineages with extremely low levels of genetic diversity. Whole-genome analysis of pathogens is envisaged to be increasingly essential for this purpose. In a microbial forensic context, whole-genome sequence analysis is the ultimate method for strain comparisons as it is informative during identification, characterization, and attribution--all 3 major stages of the investigation--and at all levels of microbial strain identity resolution (ie, it resolves the full spectrum from family to isolate). Given these capabilities, one bottleneck in microbial forensics investigations is the availability of high-quality reference databases of bacterial whole-genome sequences. To be of high quality, databases need to be curated and accurate in terms of sequences, metadata, and genetic diversity coverage. The development of whole-genome sequence databases will be instrumental in successfully tracing pathogens in the future.

  16. Genome-scale modeling enables metabolic engineering of Saccharomyces cerevisiae for succinic acid production.

    Science.gov (United States)

    Agren, Rasmus; Otero, José Manuel; Nielsen, Jens

    2013-07-01

    In this work, we describe the application of a genome-scale metabolic model and flux balance analysis for the prediction of succinic acid overproduction strategies in Saccharomyces cerevisiae. The top three single gene deletion strategies, Δmdh1, Δoac1, and Δdic1, were tested using knock-out strains cultivated anaerobically on glucose, coupled with physiological and DNA microarray characterization. While Δmdh1 and Δoac1 strains failed to produce succinate, Δdic1 produced 0.02 C-mol/C-mol glucose, in close agreement with model predictions (0.03 C-mol/C-mol glucose). Transcriptional profiling suggests that succinate formation is coupled to mitochondrial redox balancing, and more specifically, reductive TCA cycle activity. While far from industrial titers, this proof-of-concept suggests that in silico predictions coupled with experimental validation can be used to identify novel and non-intuitive metabolic engineering strategies.

  17. VaProS: a database-integration approach for protein/genome information retrieval

    KAUST Repository

    Gojobori, Takashi; Ikeo, Kazuho; Katayama, Yukie; Kawabata, Takeshi; Kinjo, Akira R.; Kinoshita, Kengo; Kwon, Yeondae; Migita, Ohsuke; Mizutani, Hisashi; Muraoka, Masafumi; Nagata, Koji; Omori, Satoshi; Sugawara, Hideaki; Yamada, Daichi; Yura, Kei

    2016-01-01

    Life science research now heavily relies on all sorts of databases for genome sequences, transcription, protein three-dimensional (3D) structures, protein–protein interactions, phenotypes and so forth. The knowledge accumulated by all the omics research is so vast that a computer-aided search of data is now a prerequisite for starting a new study. In addition, a combinatory search throughout these databases has a chance to extract new ideas and new hypotheses that can be examined by wet-lab experiments. By virtually integrating the related databases on the Internet, we have built a new web application that facilitates life science researchers for retrieving experts’ knowledge stored in the databases and for building a new hypothesis of the research target. This web application, named VaProS, puts stress on the interconnection between the functional information of genome sequences and protein 3D structures, such as structural effect of the gene mutation. In this manuscript, we present the notion of VaProS, the databases and tools that can be accessed without any knowledge of database locations and data formats, and the power of search exemplified in quest of the molecular mechanisms of lysosomal storage disease. VaProS can be freely accessed at http://p4d-info.nig.ac.jp/vapros/.

  18. VaProS: a database-integration approach for protein/genome information retrieval

    KAUST Repository

    Gojobori, Takashi

    2016-12-24

    Life science research now heavily relies on all sorts of databases for genome sequences, transcription, protein three-dimensional (3D) structures, protein–protein interactions, phenotypes and so forth. The knowledge accumulated by all the omics research is so vast that a computer-aided search of data is now a prerequisite for starting a new study. In addition, a combinatory search throughout these databases has a chance to extract new ideas and new hypotheses that can be examined by wet-lab experiments. By virtually integrating the related databases on the Internet, we have built a new web application that facilitates life science researchers for retrieving experts’ knowledge stored in the databases and for building a new hypothesis of the research target. This web application, named VaProS, puts stress on the interconnection between the functional information of genome sequences and protein 3D structures, such as structural effect of the gene mutation. In this manuscript, we present the notion of VaProS, the databases and tools that can be accessed without any knowledge of database locations and data formats, and the power of search exemplified in quest of the molecular mechanisms of lysosomal storage disease. VaProS can be freely accessed at http://p4d-info.nig.ac.jp/vapros/.

  19. A Utility Maximizing and Privacy Preserving Approach for Protecting Kinship in Genomic Databases.

    Science.gov (United States)

    Kale, Gulce; Ayday, Erman; Tastan, Oznur

    2017-09-12

    Rapid and low cost sequencing of genomes enabled widespread use of genomic data in research studies and personalized customer applications, where genomic data is shared in public databases. Although the identities of the participants are anonymized in these databases, sensitive information about individuals can still be inferred. One such information is kinship. We define two routes kinship privacy can leak and propose a technique to protect kinship privacy against these risks while maximizing the utility of shared data. The method involves systematic identification of minimal portions of genomic data to mask as new participants are added to the database. Choosing the proper positions to hide is cast as an optimization problem in which the number of positions to mask is minimized subject to privacy constraints that ensure the familial relationships are not revealed.We evaluate the proposed technique on real genomic data. Results indicate that concurrent sharing of data pertaining to a parent and an offspring results in high risks of kinship privacy, whereas the sharing data from further relatives together is often safer. We also show arrival order of family members have a high impact on the level of privacy risks and on the utility of sharing data. Available at: https://github.com/tastanlab/Kinship-Privacy. erman@cs.bilkent.edu.tr or oznur.tastan@cs.bilkent.edu.tr. Supplementary data are available at Bioinformatics online. © The Author (2017). Published by Oxford University Press. All rights reserved. For Permissions, please email: journals.permissions@oup.com

  20. Combinatorial Cis-regulation in Saccharomyces Species

    Directory of Open Access Journals (Sweden)

    Aaron T. Spivak

    2016-03-01

    Full Text Available Transcriptional control of gene expression requires interactions between the cis-regulatory elements (CREs controlling gene promoters. We developed a sensitive computational method to identify CRE combinations with conserved spacing that does not require genome alignments. When applied to seven sensu stricto and sensu lato Saccharomyces species, 80% of the predicted interactions displayed some evidence of combinatorial transcriptional behavior in several existing datasets including: (1 chromatin immunoprecipitation data for colocalization of transcription factors, (2 gene expression data for coexpression of predicted regulatory targets, and (3 gene ontology databases for common pathway membership of predicted regulatory targets. We tested several predicted CRE interactions with chromatin immunoprecipitation experiments in a wild-type strain and strains in which a predicted cofactor was deleted. Our experiments confirmed that transcription factor (TF occupancy at the promoters of the CRE combination target genes depends on the predicted cofactor while occupancy of other promoters is independent of the predicted cofactor. Our method has the additional advantage of identifying regulatory differences between species. By analyzing the S. cerevisiae and S. bayanus genomes, we identified differences in combinatorial cis-regulation between the species and showed that the predicted changes in gene regulation explain several of the species-specific differences seen in gene expression datasets. In some instances, the same CRE combinations appear to regulate genes involved in distinct biological processes in the two different species. The results of this research demonstrate that (1 combinatorial cis-regulation can be inferred by multi-genome analysis and (2 combinatorial cis-regulation can explain differences in gene expression between species.

  1. CTDB: An Integrated Chickpea Transcriptome Database for Functional and Applied Genomics

    OpenAIRE

    Verma, Mohit; Kumar, Vinay; Patel, Ravi K.; Garg, Rohini; Jain, Mukesh

    2015-01-01

    Chickpea is an important grain legume used as a rich source of protein in human diet. The narrow genetic diversity and limited availability of genomic resources are the major constraints in implementing breeding strategies and biotechnological interventions for genetic enhancement of chickpea. We developed an integrated Chickpea Transcriptome Database (CTDB), which provides the comprehensive web interface for visualization and easy retrieval of transcriptome data in chickpea. The database fea...

  2. Industrial Relevance of Chromosomal Copy Number Variation in Saccharomyces Yeasts.

    Science.gov (United States)

    Gorter de Vries, Arthur R; Pronk, Jack T; Daran, Jean-Marc G

    2017-06-01

    Chromosomal copy number variation (CCNV) plays a key role in evolution and health of eukaryotes. The unicellular yeast Saccharomyces cerevisiae is an important model for studying the generation, physiological impact, and evolutionary significance of CCNV. Fundamental studies of this yeast have contributed to an extensive set of methods for analyzing and introducing CCNV. Moreover, these studies provided insight into the balance between negative and positive impacts of CCNV in evolutionary contexts. A growing body of evidence indicates that CCNV not only frequently occurs in industrial strains of Saccharomyces yeasts but also is a key contributor to the diversity of industrially relevant traits. This notion is further supported by the frequent involvement of CCNV in industrially relevant traits acquired during evolutionary engineering. This review describes recent developments in genome sequencing and genome editing techniques and discusses how these offer opportunities to unravel contributions of CCNV in industrial Saccharomyce s strains as well as to rationally engineer yeast chromosomal copy numbers and karyotypes. Copyright © 2017 Gorter de Vries et al.

  3. Molecular genetic diversity of the Saccharomyces yeasts in Taiwan: Saccharomyces arboricola, Saccharomyces cerevisiae and Saccharomyces kudriavzevii.

    Science.gov (United States)

    Naumov, Gennadi I; Lee, Ching-Fu; Naumova, Elena S

    2013-01-01

    Genetic hybridization, sequence and karyotypic analyses of natural Saccharomyces yeasts isolated in different regions of Taiwan revealed three biological species: Saccharomyces arboricola, Saccharomyces cerevisiae and Saccharomyces kudriavzevii. Intraspecies variability of the D1/D2 and ITS1 rDNA sequences was detected among S. cerevisiae and S. kudriavzevii isolates. According to molecular and genetic analyses, the cosmopolitan species S. cerevisiae and S. kudriavzevii contain local divergent populations in Taiwan, Malaysia and Japan. Six of the seven known Saccharomyces species are documented in East Asia: S. arboricola, S. bayanus, S. cerevisiae, S. kudriavzevii, S. mikatae, and S. paradoxus.

  4. GenoMycDB: a database for comparative analysis of mycobacterial genes and genomes.

    Science.gov (United States)

    Catanho, Marcos; Mascarenhas, Daniel; Degrave, Wim; Miranda, Antonio Basílio de

    2006-03-31

    Several databases and computational tools have been created with the aim of organizing, integrating and analyzing the wealth of information generated by large-scale sequencing projects of mycobacterial genomes and those of other organisms. However, with very few exceptions, these databases and tools do not allow for massive and/or dynamic comparison of these data. GenoMycDB (http://www.dbbm.fiocruz.br/GenoMycDB) is a relational database built for large-scale comparative analyses of completely sequenced mycobacterial genomes, based on their predicted protein content. Its central structure is composed of the results obtained after pair-wise sequence alignments among all the predicted proteins coded by the genomes of six mycobacteria: Mycobacterium tuberculosis (strains H37Rv and CDC1551), M. bovis AF2122/97, M. avium subsp. paratuberculosis K10, M. leprae TN, and M. smegmatis MC2 155. The database stores the computed similarity parameters of every aligned pair, providing for each protein sequence the predicted subcellular localization, the assigned cluster of orthologous groups, the features of the corresponding gene, and links to several important databases. Tables containing pairs or groups of potential homologs between selected species/strains can be produced dynamically by user-defined criteria, based on one or multiple sequence similarity parameters. In addition, searches can be restricted according to the predicted subcellular localization of the protein, the DNA strand of the corresponding gene and/or the description of the protein. Massive data search and/or retrieval are available, and different ways of exporting the result are offered. GenoMycDB provides an on-line resource for the functional classification of mycobacterial proteins as well as for the analysis of genome structure, organization, and evolution.

  5. Genome wide transcriptional response of Saccharomyces cerevisiae to stress-induced perturbations

    Directory of Open Access Journals (Sweden)

    Hilal eTaymaz-Nikerel

    2016-02-01

    Full Text Available Cells respond to environmental and/or genetic perturbations in order to survive and proliferate. Characterization of the changes after various stimuli at different -omics levels is crucial to comprehend the adaptation of cells to changing conditions. Genome wide quantification and analysis of transcript levels, the genes affected by perturbations, extends our understanding of cellular metabolism by pointing out the mechanisms that play role in sensing the stress caused by those perturbations and related signaling pathways, and in this way guides us to achieve endeavors such as rational engineering of cells or interpretation of disease mechanisms. Saccharomyces cerevisiae as a model system has been studied in response to different perturbations and corresponding transcriptional profiles were followed either statically or/and dynamically, short- and long- term. This review focuses on response of yeast cells to diverse stress inducing perturbations including nutritional changes, ionic stress, salt stress, oxidative stress, osmotic shock, as well as to genetic interventions such as deletion and over-expression of genes. It is aimed to conclude on common regulatory phenomena that allow yeast to organize its transcriptomic response after any perturbation under different external conditions.

  6. Deciphering the hybridisation history leading to the Lager lineage based on the mosaic genomes of Saccharomyces bayanus strains NBRC1948 and CBS380.

    Directory of Open Access Journals (Sweden)

    Huu-Vang Nguyen

    Full Text Available Saccharomyces bayanus is a yeast species described as one of the two parents of the hybrid brewing yeast S. pastorianus. Strains CBS380(T and NBRC1948 have been retained successively as pure-line representatives of S. bayanus. In the present study, sequence analyses confirmed and upgraded our previous finding: S. bayanus type strain CBS380(T harbours a mosaic genome. The genome of strain NBRC1948 was also revealed to be mosaic. Both genomes were characterized by amplification and sequencing of different markers, including genes involved in maltotriose utilization or genes detected by array-CGH mapping. Sequence comparisons with public Saccharomyces spp. nucleotide sequences revealed that the CBS380(T and NBRC1948 genomes are composed of: a predominant non-cerevisiae genetic background belonging to S. uvarum, a second unidentified species provisionally named S. lagerae, and several introgressed S. cerevisiae fragments. The largest cerevisiae-introgressed DNA common to both genomes totals 70kb in length and is distributed in three contigs, cA, cB and cC. These vary in terms of length and presence of MAL31 or MTY1 (maltotriose-transporter gene. In NBRC1948, two additional cerevisiae-contigs, cD and cE, totaling 12kb in length, as well as several smaller cerevisiae fragments were identified. All of these contigs were partially detected in the genomes of S. pastorianus lager strains CBS1503 (S. monacensis and CBS1513 (S. carlsbergensis explaining the noticeable common ability of S. bayanus and S. pastorianus to metabolize maltotriose. NBRC1948 was shown to be inter-fertile with S. uvarum CBS7001. The cross involving these two strains produced F1 segregants resembling the strains CBS380(T or NRRLY-1551. This demonstrates that these S. bayanus strains were the offspring of a cross between S. uvarum and a strain similar to NBRC1948. Phylogenies established with selected cerevisiae and non-cerevisiae genes allowed us to decipher the complex hybridisation

  7. GiSAO.db: a database for ageing research

    Directory of Open Access Journals (Sweden)

    Grillari Johannes

    2011-05-01

    Full Text Available Abstract Background Age-related gene expression patterns of Homo sapiens as well as of model organisms such as Mus musculus, Saccharomyces cerevisiae, Caenorhabditis elegans and Drosophila melanogaster are a basis for understanding the genetic mechanisms of ageing. For an effective analysis and interpretation of expression profiles it is necessary to store and manage huge amounts of data in an organized way, so that these data can be accessed and processed easily. Description GiSAO.db (Genes involved in senescence, apoptosis and oxidative stress database is a web-based database system for storing and retrieving ageing-related experimental data. Expression data of genes and miRNAs, annotation data like gene identifiers and GO terms, orthologs data and data of follow-up experiments are stored in the database. A user-friendly web application provides access to the stored data. KEGG pathways were incorporated and links to external databases augment the information in GiSAO.db. Search functions facilitate retrieval of data which can also be exported for further processing. Conclusions We have developed a centralized database that is very well suited for the management of data for ageing research. The database can be accessed at https://gisao.genome.tugraz.at and all the stored data can be viewed with a guest account.

  8. Importance of databases of nucleic acids for bioinformatic analysis focused to genomics

    Science.gov (United States)

    Jimenez-Gutierrez, L. R.; Barrios-Hernández, C. J.; Pedraza-Ferreira, G. R.; Vera-Cala, L.; Martinez-Perez, F.

    2016-08-01

    Recently, bioinformatics has become a new field of science, indispensable in the analysis of millions of nucleic acids sequences, which are currently deposited in international databases (public or private); these databases contain information of genes, RNA, ORF, proteins, intergenic regions, including entire genomes from some species. The analysis of this information requires computer programs; which were renewed in the use of new mathematical methods, and the introduction of the use of artificial intelligence. In addition to the constant creation of supercomputing units trained to withstand the heavy workload of sequence analysis. However, it is still necessary the innovation on platforms that allow genomic analyses, faster and more effectively, with a technological understanding of all biological processes.

  9. GEAR: A database of Genomic Elements Associated with drug Resistance

    Science.gov (United States)

    Wang, Yin-Ying; Chen, Wei-Hua; Xiao, Pei-Pei; Xie, Wen-Bin; Luo, Qibin; Bork, Peer; Zhao, Xing-Ming

    2017-01-01

    Drug resistance is becoming a serious problem that leads to the failure of standard treatments, which is generally developed because of genetic mutations of certain molecules. Here, we present GEAR (A database of Genomic Elements Associated with drug Resistance) that aims to provide comprehensive information about genomic elements (including genes, single-nucleotide polymorphisms and microRNAs) that are responsible for drug resistance. Right now, GEAR contains 1631 associations between 201 human drugs and 758 genes, 106 associations between 29 human drugs and 66 miRNAs, and 44 associations between 17 human drugs and 22 SNPs. These relationships are firstly extracted from primary literature with text mining and then manually curated. The drug resistome deposited in GEAR provides insights into the genetic factors underlying drug resistance. In addition, new indications and potential drug combinations can be identified based on the resistome. The GEAR database can be freely accessed through http://gear.comp-sysbio.org. PMID:28294141

  10. ATGC: a database of orthologous genes from closely related prokaryotic genomes and a research platform for microevolution of prokaryotes

    Energy Technology Data Exchange (ETDEWEB)

    Novichkov, Pavel S.; Ratnere, Igor; Wolf, Yuri I.; Koonin, Eugene V.; Dubchak, Inna

    2009-07-23

    The database of Alignable Tight Genomic Clusters (ATGCs) consists of closely related genomes of archaea and bacteria, and is a resource for research into prokaryotic microevolution. Construction of a data set with appropriate characteristics is a major hurdle for this type of studies. With the current rate of genome sequencing, it is difficult to follow the progress of the field and to determine which of the available genome sets meet the requirements of a given research project, in particular, with respect to the minimum and maximum levels of similarity between the included genomes. Additionally, extraction of specific content, such as genomic alignments or families of orthologs, from a selected set of genomes is a complicated and time-consuming process. The database addresses these problems by providing an intuitive and efficient web interface to browse precomputed ATGCs, select appropriate ones and access ATGC-derived data such as multiple alignments of orthologous proteins, matrices of pairwise intergenomic distances based on genome-wide analysis of synonymous and nonsynonymous substitution rates and others. The ATGC database will be regularly updated following new releases of the NCBI RefSeq. The database is hosted by the Genomics Division at Lawrence Berkeley National laboratory and is publicly available at http://atgc.lbl.gov.

  11. Construction of an integrated database to support genomic sequence analysis

    Energy Technology Data Exchange (ETDEWEB)

    Gilbert, W.; Overbeek, R.

    1994-11-01

    The central goal of this project is to develop an integrated database to support comparative analysis of genomes including DNA sequence data, protein sequence data, gene expression data and metabolism data. In developing the logic-based system GenoBase, a broader integration of available data was achieved due to assistance from collaborators. Current goals are to easily include new forms of data as they become available and to easily navigate through the ensemble of objects described within the database. This report comments on progress made in these areas.

  12. Updates to the Cool Season Food Legume Genome Database: Resources for pea, lentil, faba bean and chickpea genetics, genomics and breeding

    Science.gov (United States)

    The Cool Season Food Legume Genome database (CSFL, www.coolseasonfoodlegume.org) is an online resource for genomics, genetics, and breeding research for chickpea, lentil,pea, and faba bean. The user-friendly and curated website allows for all publicly available map,marker,trait, gene,transcript, ger...

  13. The MetaCyc database of metabolic pathways and enzymes and the BioCyc collection of pathway/genome databases

    Science.gov (United States)

    Caspi, Ron; Altman, Tomer; Dale, Joseph M.; Dreher, Kate; Fulcher, Carol A.; Gilham, Fred; Kaipa, Pallavi; Karthikeyan, Athikkattuvalasu S.; Kothari, Anamika; Krummenacker, Markus; Latendresse, Mario; Mueller, Lukas A.; Paley, Suzanne; Popescu, Liviu; Pujar, Anuradha; Shearer, Alexander G.; Zhang, Peifen; Karp, Peter D.

    2010-01-01

    The MetaCyc database (MetaCyc.org) is a comprehensive and freely accessible resource for metabolic pathways and enzymes from all domains of life. The pathways in MetaCyc are experimentally determined, small-molecule metabolic pathways and are curated from the primary scientific literature. With more than 1400 pathways, MetaCyc is the largest collection of metabolic pathways currently available. Pathways reactions are linked to one or more well-characterized enzymes, and both pathways and enzymes are annotated with reviews, evidence codes, and literature citations. BioCyc (BioCyc.org) is a collection of more than 500 organism-specific Pathway/Genome Databases (PGDBs). Each BioCyc PGDB contains the full genome and predicted metabolic network of one organism. The network, which is predicted by the Pathway Tools software using MetaCyc as a reference, consists of metabolites, enzymes, reactions and metabolic pathways. BioCyc PGDBs also contain additional features, such as predicted operons, transport systems, and pathway hole-fillers. The BioCyc Web site offers several tools for the analysis of the PGDBs, including Omics Viewers that enable visualization of omics datasets on two different genome-scale diagrams and tools for comparative analysis. The BioCyc PGDBs generated by SRI are offered for adoption by any party interested in curation of metabolic, regulatory, and genome-related information about an organism. PMID:19850718

  14. dbEM: A database of epigenetic modifiers curated from cancerous and normal genomes

    Science.gov (United States)

    Singh Nanda, Jagpreet; Kumar, Rahul; Raghava, Gajendra P. S.

    2016-01-01

    We have developed a database called dbEM (database of Epigenetic Modifiers) to maintain the genomic information of about 167 epigenetic modifiers/proteins, which are considered as potential cancer targets. In dbEM, modifiers are classified on functional basis and comprise of 48 histone methyl transferases, 33 chromatin remodelers and 31 histone demethylases. dbEM maintains the genomic information like mutations, copy number variation and gene expression in thousands of tumor samples, cancer cell lines and healthy samples. This information is obtained from public resources viz. COSMIC, CCLE and 1000-genome project. Gene essentiality data retrieved from COLT database further highlights the importance of various epigenetic proteins for cancer survival. We have also reported the sequence profiles, tertiary structures and post-translational modifications of these epigenetic proteins in cancer. It also contains information of 54 drug molecules against different epigenetic proteins. A wide range of tools have been integrated in dbEM e.g. Search, BLAST, Alignment and Profile based prediction. In our analysis, we found that epigenetic proteins DNMT3A, HDAC2, KDM6A, and TET2 are highly mutated in variety of cancers. We are confident that dbEM will be very useful in cancer research particularly in the field of epigenetic proteins based cancer therapeutics. This database is available for public at URL: http://crdd.osdd.net/raghava/dbem.

  15. Saccharomyces Boulardii

    Science.gov (United States)

    Saccharomyces boulardii is a yeast, which is a type of fungus. Saccharomyces boulardii was previously identified as a unique species of ... be a strain of Saccharomyces cerevisiae (baker's yeast). Saccharomyces boulardii is used as medicine. Saccharomyces boulardii is most ...

  16. CpGislandEVO: A Database and Genome Browser for Comparative Evolutionary Genomics of CpG Islands

    Directory of Open Access Journals (Sweden)

    Guillermo Barturen

    2013-01-01

    Full Text Available Hypomethylated, CpG-rich DNA segments (CpG islands, CGIs are epigenome markers involved in key biological processes. Aberrant methylation is implicated in the appearance of several disorders as cancer, immunodeficiency, or centromere instability. Furthermore, methylation differences at promoter regions between human and chimpanzee strongly associate with genes involved in neurological/psychological disorders and cancers. Therefore, the evolutionary comparative analyses of CGIs can provide insights on the functional role of these epigenome markers in both health and disease. Given the lack of specific tools, we developed CpGislandEVO. Briefly, we first compile a database of statistically significant CGIs for the best assembled mammalian genome sequences available to date. Second, by means of a coupled browser front-end, we focus on the CGIs overlapping orthologous genes extracted from OrthoDB, thus ensuring the comparison between CGIs located on truly homologous genome segments. This allows comparing the main compositional features between homologous CGIs. Finally, to facilitate nucleotide comparisons, we lifted genome coordinates between assemblies from different species, which enables the analysis of sequence divergence by direct count of nucleotide substitutions and indels occurring between homologous CGIs. The resulting CpGislandEVO database, linking together CGIs and single-cytosine DNA methylation data from several mammalian species, is freely available at our website.

  17. PIPEMicroDB: microsatellite database and primer generation tool for pigeonpea genome.

    Science.gov (United States)

    Sarika; Arora, Vasu; Iquebal, M A; Rai, Anil; Kumar, Dinesh

    2013-01-01

    Molecular markers play a significant role for crop improvement in desirable characteristics, such as high yield, resistance to disease and others that will benefit the crop in long term. Pigeonpea (Cajanus cajan L.) is the recently sequenced legume by global consortium led by ICRISAT (Hyderabad, India) and been analysed for gene prediction, synteny maps, markers, etc. We present PIgeonPEa Microsatellite DataBase (PIPEMicroDB) with an automated primer designing tool for pigeonpea genome, based on chromosome wise as well as location wise search of primers. Total of 123 387 Short Tandem Repeats (STRs) were extracted from pigeonpea genome, available in public domain using MIcroSAtellite tool (MISA). The database is an online relational database based on 'three-tier architecture' that catalogues information of microsatellites in MySQL and user-friendly interface is developed using PHP. Search for STRs may be customized by limiting their location on chromosome as well as number of markers in that range. This is a novel approach and is not been implemented in any of the existing marker database. This database has been further appended with Primer3 for primer designing of selected markers with left and right flankings of size up to 500 bp. This will enable researchers to select markers of choice at desired interval over the chromosome. Furthermore, one can use individual STRs of a targeted region over chromosome to narrow down location of gene of interest or linked Quantitative Trait Loci (QTLs). Although it is an in silico approach, markers' search based on characteristics and location of STRs is expected to be beneficial for researchers. Database URL: http://cabindb.iasri.res.in/pigeonpea/

  18. Core Data of Yeast Interacting Proteins Database (Original Version) - Yeast Interacting Proteins Database | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available y are in the reverse direction. *1 A comprehensive two-hybrid analysis to explore the yeast protein interact...s. 2000 Jan 1;28(1):73-6. *2 The yeast proteome database (YPD) and Caenorhabditis elegans proteome database (WormPD): comprehensive...000 Jan 1;28(1):73-6. *3 A comprehensive analysis of protein-protein interactions in Saccharomyces cerevisia

  19. Developing genomic knowledge bases and databases to support clinical management: current perspectives.

    Science.gov (United States)

    Huser, Vojtech; Sincan, Murat; Cimino, James J

    2014-01-01

    Personalized medicine, the ability to tailor diagnostic and treatment decisions for individual patients, is seen as the evolution of modern medicine. We characterize here the informatics resources available today or envisioned in the near future that can support clinical interpretation of genomic test results. We assume a clinical sequencing scenario (germline whole-exome sequencing) in which a clinical specialist, such as an endocrinologist, needs to tailor patient management decisions within his or her specialty (targeted findings) but relies on a genetic counselor to interpret off-target incidental findings. We characterize the genomic input data and list various types of knowledge bases that provide genomic knowledge for generating clinical decision support. We highlight the need for patient-level databases with detailed lifelong phenotype content in addition to genotype data and provide a list of recommendations for personalized medicine knowledge bases and databases. We conclude that no single knowledge base can currently support all aspects of personalized recommendations and that consolidation of several current resources into larger, more dynamic and collaborative knowledge bases may offer a future path forward.

  20. KGCAK: a K-mer based database for genome-wide phylogeny and complexity evaluation.

    Science.gov (United States)

    Wang, Dapeng; Xu, Jiayue; Yu, Jun

    2015-09-16

    The K-mer approach, treating genomic sequences as simple characters and counting the relative abundance of each string upon a fixed K, has been extensively applied to phylogeny inference for genome assembly, annotation, and comparison. To meet increasing demands for comparing large genome sequences and to promote the use of the K-mer approach, we develop a versatile database, KGCAK ( http://kgcak.big.ac.cn/KGCAK/ ), containing ~8,000 genomes that include genome sequences of diverse life forms (viruses, prokaryotes, protists, animals, and plants) and cellular organelles of eukaryotic lineages. It builds phylogeny based on genomic elements in an alignment-free fashion and provides in-depth data processing enabling users to compare the complexity of genome sequences based on K-mer distribution. We hope that KGCAK becomes a powerful tool for exploring relationship within and among groups of species in a tree of life based on genomic data.

  1. How did Saccharomyces evolve to become a good brewer?

    Science.gov (United States)

    Piskur, Jure; Rozpedowska, Elzbieta; Polakova, Silvia; Merico, Annamaria; Compagno, Concetta

    2006-04-01

    Brewing and wine production are among the oldest technologies and their products are almost indispensable in our lives. The central biological agents of beer and wine fermentation are yeasts belonging to the genus Saccharomyces, which can accumulate ethanol. Recent advances in comparative genomics and bioinformatics have made it possible to elucidate when and why yeasts produce ethanol in high concentrations, and how this remarkable trait originated and developed during their evolutionary history. Two research groups have shed light on the origin of the genes encoding alcohol dehydrogenase and the process of ethanol accumulation in Saccharomyces cerevisiae.

  2. Multiplex metabolic pathway engineering using CRISPR/Cas9 in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Jakociunas, Tadas; Bonde, Ida; Herrgard, Markus

    2015-01-01

    CRISPR/Cas9 is a simple and efficient tool for targeted and marker-free genome engineering. Here, we report the development and successful application of a multiplex CRISPR/Cas9 system for genome engineering of up to 5 different genomic loci in one transformation step in baker's yeast Saccharomyces...... cerevisiae. To assess the specificity of the tool we employed genome re-sequencing to screen for off-target sites in all single knock-out strains targeted by different gRNAs. This extensive analysis identified no more genome variants in CRISPR/Cas9 engineered strains compared to wild-type reference strains...

  3. G-quadruplex DNA sequences are evolutionarily conserved and associated with distinct genomic features in Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    John A Capra

    2010-07-01

    Full Text Available G-quadruplex DNA is a four-stranded DNA structure formed by non-Watson-Crick base pairing between stacked sets of four guanines. Many possible functions have been proposed for this structure, but its in vivo role in the cell is still largely unresolved. We carried out a genome-wide survey of the evolutionary conservation of regions with the potential to form G-quadruplex DNA structures (G4 DNA motifs across seven yeast species. We found that G4 DNA motifs were significantly more conserved than expected by chance, and the nucleotide-level conservation patterns suggested that the motif conservation was the result of the formation of G4 DNA structures. We characterized the association of conserved and non-conserved G4 DNA motifs in Saccharomyces cerevisiae with more than 40 known genome features and gene classes. Our comprehensive, integrated evolutionary and functional analysis confirmed the previously observed associations of G4 DNA motifs with promoter regions and the rDNA, and it identified several previously unrecognized associations of G4 DNA motifs with genomic features, such as mitotic and meiotic double-strand break sites (DSBs. Conserved G4 DNA motifs maintained strong associations with promoters and the rDNA, but not with DSBs. We also performed the first analysis of G4 DNA motifs in the mitochondria, and surprisingly found a tenfold higher concentration of the motifs in the AT-rich yeast mitochondrial DNA than in nuclear DNA. The evolutionary conservation of the G4 DNA motif and its association with specific genome features supports the hypothesis that G4 DNA has in vivo functions that are under evolutionary constraint.

  4. MetReS, an Efficient Database for Genomic Applications.

    Science.gov (United States)

    Vilaplana, Jordi; Alves, Rui; Solsona, Francesc; Mateo, Jordi; Teixidó, Ivan; Pifarré, Marc

    2018-02-01

    MetReS (Metabolic Reconstruction Server) is a genomic database that is shared between two software applications that address important biological problems. Biblio-MetReS is a data-mining tool that enables the reconstruction of molecular networks based on automated text-mining analysis of published scientific literature. Homol-MetReS allows functional (re)annotation of proteomes, to properly identify both the individual proteins involved in the processes of interest and their function. The main goal of this work was to identify the areas where the performance of the MetReS database performance could be improved and to test whether this improvement would scale to larger datasets and more complex types of analysis. The study was started with a relational database, MySQL, which is the current database server used by the applications. We also tested the performance of an alternative data-handling framework, Apache Hadoop. Hadoop is currently used for large-scale data processing. We found that this data handling framework is likely to greatly improve the efficiency of the MetReS applications as the dataset and the processing needs increase by several orders of magnitude, as expected to happen in the near future.

  5. MaizeGDB: The Maize Genetics and Genomics Database.

    Science.gov (United States)

    Harper, Lisa; Gardiner, Jack; Andorf, Carson; Lawrence, Carolyn J

    2016-01-01

    MaizeGDB is the community database for biological information about the crop plant Zea mays. Genomic, genetic, sequence, gene product, functional characterization, literature reference, and person/organization contact information are among the datatypes stored at MaizeGDB. At the project's website ( http://www.maizegdb.org ) are custom interfaces enabling researchers to browse data and to seek out specific information matching explicit search criteria. In addition, pre-compiled reports are made available for particular types of data and bulletin boards are provided to facilitate communication and coordination among members of the community of maize geneticists.

  6. PReMod: a database of genome-wide mammalian cis-regulatory module predictions.

    Science.gov (United States)

    Ferretti, Vincent; Poitras, Christian; Bergeron, Dominique; Coulombe, Benoit; Robert, François; Blanchette, Mathieu

    2007-01-01

    We describe PReMod, a new database of genome-wide cis-regulatory module (CRM) predictions for both the human and the mouse genomes. The prediction algorithm, described previously in Blanchette et al. (2006) Genome Res., 16, 656-668, exploits the fact that many known CRMs are made of clusters of phylogenetically conserved and repeated transcription factors (TF) binding sites. Contrary to other existing databases, PReMod is not restricted to modules located proximal to genes, but in fact mostly contains distal predicted CRMs (pCRMs). Through its web interface, PReMod allows users to (i) identify pCRMs around a gene of interest; (ii) identify pCRMs that have binding sites for a given TF (or a set of TFs) or (iii) download the entire dataset for local analyses. Queries can also be refined by filtering for specific chromosomal regions, for specific regions relative to genes or for the presence of CpG islands. The output includes information about the binding sites predicted within the selected pCRMs, and a graphical display of their distribution within the pCRMs. It also provides a visual depiction of the chromosomal context of the selected pCRMs in terms of neighboring pCRMs and genes, all of which are linked to the UCSC Genome Browser and the NCBI. PReMod: http://genomequebec.mcgill.ca/PReMod.

  7. A Trichosporonales genome tree based on 27 haploid and three evolutionarily conserved 'natural' hybrid genomes.

    Science.gov (United States)

    Takashima, Masako; Sriswasdi, Sira; Manabe, Ri-Ichiroh; Ohkuma, Moriya; Sugita, Takashi; Iwasaki, Wataru

    2018-01-01

    To construct a backbone tree consisting of basidiomycetous yeasts, draft genome sequences from 25 species of Trichosporonales (Tremellomycetes, Basidiomycota) were generated. In addition to the hybrid genomes of Trichosporon coremiiforme and Trichosporon ovoides that we described previously, we identified an interspecies hybrid genome in Cutaneotrichosporon mucoides (formerly Trichosporon mucoides). This hybrid genome had a gene retention rate of ~55%, and its closest haploid relative was Cutaneotrichosporon dermatis. After constructing the C. mucoides subgenomes, we generated a phylogenetic tree using genome data from the 27 haploid species and the subgenome data from the three hybrid genome species. It was a high-quality tree with 100% bootstrap support for all of the branches. The genome-based tree provided superior resolution compared with previous multi-gene analyses. Although our backbone tree does not include all Trichosporonales genera (e.g. Cryptotrichosporon), it will be valuable for future analyses of genome data. Interest in interspecies hybrid fungal genomes has recently increased because they may provide a basis for new technologies. The three Trichosporonales hybrid genomes described in this study are different from well-characterized hybrid genomes (e.g. those of Saccharomyces pastorianus and Saccharomyces bayanus) because these hybridization events probably occurred in the distant evolutionary past. Hence, they will be useful for studying genome stability following hybridization and speciation events. Copyright © 2017 John Wiley & Sons, Ltd. Copyright © 2017 John Wiley & Sons, Ltd.

  8. Database Description - RMOS | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available base Description General information of database Database name RMOS Alternative nam...arch Unit Shoshi Kikuchi E-mail : Database classification Plant databases - Rice Microarray Data and other Gene Expression Database...s Organism Taxonomy Name: Oryza sativa Taxonomy ID: 4530 Database description The Ric...19&lang=en Whole data download - Referenced database Rice Expression Database (RED) Rice full-length cDNA Database... (KOME) Rice Genome Integrated Map Database (INE) Rice Mutant Panel Database (Tos17) Rice Genome Annotation Database

  9. Hybridization and adaptive evolution of diverse Saccharomyces species for cellulosic biofuel production.

    Science.gov (United States)

    Peris, David; Moriarty, Ryan V; Alexander, William G; Baker, EmilyClare; Sylvester, Kayla; Sardi, Maria; Langdon, Quinn K; Libkind, Diego; Wang, Qi-Ming; Bai, Feng-Yan; Leducq, Jean-Baptiste; Charron, Guillaume; Landry, Christian R; Sampaio, José Paulo; Gonçalves, Paula; Hyma, Katie E; Fay, Justin C; Sato, Trey K; Hittinger, Chris Todd

    2017-01-01

    Lignocellulosic biomass is a common resource across the globe, and its fermentation offers a promising option for generating renewable liquid transportation fuels. The deconstruction of lignocellulosic biomass releases sugars that can be fermented by microbes, but these processes also produce fermentation inhibitors, such as aromatic acids and aldehydes. Several research projects have investigated lignocellulosic biomass fermentation by the baker's yeast Saccharomyces cerevisiae . Most projects have taken synthetic biological approaches or have explored naturally occurring diversity in S. cerevisiae to enhance stress tolerance, xylose consumption, or ethanol production. Despite these efforts, improved strains with new properties are needed. In other industrial processes, such as wine and beer fermentation, interspecies hybrids have combined important traits from multiple species, suggesting that interspecies hybridization may also offer potential for biofuel research. To investigate the efficacy of this approach for traits relevant to lignocellulosic biofuel production, we generated synthetic hybrids by crossing engineered xylose-fermenting strains of S. cerevisiae with wild strains from various Saccharomyces species. These interspecies hybrids retained important parental traits, such as xylose consumption and stress tolerance, while displaying intermediate kinetic parameters and, in some cases, heterosis (hybrid vigor). Next, we exposed them to adaptive evolution in ammonia fiber expansion-pretreated corn stover hydrolysate and recovered strains with improved fermentative traits. Genome sequencing showed that the genomes of these evolved synthetic hybrids underwent rearrangements, duplications, and deletions. To determine whether the genus Saccharomyces contains additional untapped potential, we screened a genetically diverse collection of more than 500 wild, non-engineered Saccharomyces isolates and uncovered a wide range of capabilities for traits relevant to

  10. The phytophthora genome initiative database: informatics and analysis for distributed pathogenomic research.

    Science.gov (United States)

    Waugh, M; Hraber, P; Weller, J; Wu, Y; Chen, G; Inman, J; Kiphart, D; Sobral, B

    2000-01-01

    The Phytophthora Genome Initiative (PGI) is a distributed collaboration to study the genome and evolution of a particularly destructive group of plant pathogenic oomycete, with the goal of understanding the mechanisms of infection and resistance. NCGR provides informatics support for the collaboration as well as a centralized data repository. In the pilot phase of the project, several investigators prepared Phytophthora infestans and Phytophthora sojae EST and Phytophthora sojae BAC libraries and sent them to another laboratory for sequencing. Data from sequencing reactions were transferred to NCGR for analysis and curation. An analysis pipeline transforms raw data by performing simple analyses (i.e., vector removal and similarity searching) that are stored and can be retrieved by investigators using a web browser. Here we describe the database and access tools, provide an overview of the data therein and outline future plans. This resource has provided a unique opportunity for the distributed, collaborative study of a genus from which relatively little sequence data are available. Results may lead to insight into how better to control these pathogens. The homepage of PGI can be accessed at http:www.ncgr.org/pgi, with database access through the database access hyperlink.

  11. HpBase: A genome database of a sea urchin, Hemicentrotus pulcherrimus.

    Science.gov (United States)

    Kinjo, Sonoko; Kiyomoto, Masato; Yamamoto, Takashi; Ikeo, Kazuho; Yaguchi, Shunsuke

    2018-04-01

    To understand the mystery of life, it is important to accumulate genomic information for various organisms because the whole genome encodes the commands for all the genes. Since the genome of Strongylocentrotus purpratus was sequenced in 2006 as the first sequenced genome in echinoderms, the genomic resources of other North American sea urchins have gradually been accumulated, but no sea urchin genomes are available in other areas, where many scientists have used the local species and reported important results. In this manuscript, we report a draft genome of the sea urchin Hemincentrotus pulcherrimus because this species has a long history as the target of developmental and cell biology in East Asia. The genome of H. pulcherrimus was assembled into 16,251 scaffold sequences with an N50 length of 143 kbp, and approximately 25,000 genes were identified in the genome. The size of the genome and the sequencing coverage were estimated to be approximately 800 Mbp and 100×, respectively. To provide these data and information of annotation, we constructed a database, HpBase (http://cell-innovation.nig.ac.jp/Hpul/). In HpBase, gene searches, genome browsing, and blast searches are available. In addition, HpBase includes the "recipes" for experiments from each lab using H. pulcherrimus. These recipes will continue to be updated according to the circumstances of individual scientists and can be powerful tools for experimental biologists and for the community. HpBase is a suitable dataset for evolutionary, developmental, and cell biologists to compare H. pulcherrimus genomic information with that of other species and to isolate gene information. © 2018 Japanese Society of Developmental Biologists.

  12. Genome-wide transcription survey on flavour production in Saccharomyces cerevisiae

    NARCIS (Netherlands)

    Schoondermark-Stolk, Sung A.; Jansen, Michael; Verkleij, Arie J.; Verrips, C. Theo; Euverink, Gert-Jan W.; Dijkhuizen, Lubbert; Boonstra, Johannes

    2006-01-01

    The yeast Saccharomyces cerevisiae is widely used as aroma producer in the preparation of fermented foods and beverages. During food fermentations, secondary metabolites like 3-methyl-1-butanol, 4-methyl-2-oxopentanoate, 3-methyl-2-oxobutanoate and 3-methylbutyrate emerge. These four compounds have

  13. Context based computational analysis and characterization of ARS consensus sequences (ACS of Saccharomyces cerevisiae genome

    Directory of Open Access Journals (Sweden)

    Vinod Kumar Singh

    2016-09-01

    Full Text Available Genome-wide experimental studies in Saccharomyces cerevisiae reveal that autonomous replicating sequence (ARS requires an essential consensus sequence (ACS for replication activity. Computational studies identified thousands of ACS like patterns in the genome. However, only a few hundreds of these sites act as replicating sites and the rest are considered as dormant or evolving sites. In a bid to understand the sequence makeup of replication sites, a content and context-based analysis was performed on a set of replicating ACS sequences that binds to origin-recognition complex (ORC denoted as ORC-ACS and non-replicating ACS sequences (nrACS, that are not bound by ORC. In this study, DNA properties such as base composition, correlation, sequence dependent thermodynamic and DNA structural profiles, and their positions have been considered for characterizing ORC-ACS and nrACS. Analysis reveals that ORC-ACS depict marked differences in nucleotide composition and context features in its vicinity compared to nrACS. Interestingly, an A-rich motif was also discovered in ORC-ACS sequences within its nucleosome-free region. Profound changes in the conformational features, such as DNA helical twist, inclination angle and stacking energy between ORC-ACS and nrACS were observed. Distribution of ACS motifs in the non-coding segments points to the locations of ORC-ACS which are found far away from the adjacent gene start position compared to nrACS thereby enabling an accessible environment for ORC-proteins. Our attempt is novel in considering the contextual view of ACS and its flanking region along with nucleosome positioning in the S. cerevisiae genome and may be useful for any computational prediction scheme.

  14. Complete genomic and transcriptional landscape analysis using third-generation sequencing: a case study of Saccharomyces cerevisiae CEN.PK113-7D

    DEFF Research Database (Denmark)

    Jenjaroenpun, Piroon; Wongsurawat, Thidathip; Pereira, Rui

    2018-01-01

    Completion of eukaryal genomes can be difficult task with the highly repetitive sequences along the chromosomes and short read lengths of secondgeneration sequencing. Saccharomyces cerevisiae strain CEN. PK113-7D, widely used as a model organism and a cell factory, was selected for this study...... to demonstrate the superior capability of very long sequence reads for de novo genome assembly. We generated long reads using two common third-generation sequencing technologies (Oxford Nanopore Technology (ONT) and Pacific Biosciences (PacBio)) and used short reads obtained using Illumina sequencing for error...... correction. Assembly of the reads derived from all three technologies resulted in complete sequences for all 16 yeast chromosomes, as well as themitochondrial chromosome, in one step. Further, we identified three types of DNA methylation (5mC, 4mC and 6mA). Comparison between the reference strain S288C...

  15. Global mapping of DNA conformational flexibility on Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Giulia Menconi

    2015-04-01

    Full Text Available In this study we provide the first comprehensive map of DNA conformational flexibility in Saccharomyces cerevisiae complete genome. Flexibility plays a key role in DNA supercoiling and DNA/protein binding, regulating DNA transcription, replication or repair. Specific interest in flexibility analysis concerns its relationship with human genome instability. Enrichment in flexible sequences has been detected in unstable regions of human genome defined fragile sites, where genes map and carry frequent deletions and rearrangements in cancer. Flexible sequences have been suggested to be the determinants of fragile gene proneness to breakage; however, their actual role and properties remain elusive. Our in silico analysis carried out genome-wide via the StabFlex algorithm, shows the conserved presence of highly flexible regions in budding yeast genome as well as in genomes of other Saccharomyces sensu stricto species. Flexibile peaks in S. cerevisiae identify 175 ORFs mapping on their 3'UTR, a region affecting mRNA translation, localization and stability. (TAn repeats of different extension shape the central structure of peaks and co-localize with polyadenylation efficiency element (EE signals. ORFs with flexible peaks share common features. Transcripts are characterized by decreased half-life: this is considered peculiar of genes involved in regulatory systems with high turnover; consistently, their function affects biological processes such as cell cycle regulation or stress response. Our findings support the functional importance of flexibility peaks, suggesting that the flexible sequence may be derived by an expansion of canonical TAYRTA polyadenylation efficiency element. The flexible (TAn repeat amplification could be the outcome of an evolutionary neofunctionalization leading to a differential 3'-end processing and expression regulation in genes with peculiar function. Our study provides a new support to the functional role of flexibility in

  16. Global mapping of DNA conformational flexibility on Saccharomyces cerevisiae.

    Science.gov (United States)

    Menconi, Giulia; Bedini, Andrea; Barale, Roberto; Sbrana, Isabella

    2015-04-01

    In this study we provide the first comprehensive map of DNA conformational flexibility in Saccharomyces cerevisiae complete genome. Flexibility plays a key role in DNA supercoiling and DNA/protein binding, regulating DNA transcription, replication or repair. Specific interest in flexibility analysis concerns its relationship with human genome instability. Enrichment in flexible sequences has been detected in unstable regions of human genome defined fragile sites, where genes map and carry frequent deletions and rearrangements in cancer. Flexible sequences have been suggested to be the determinants of fragile gene proneness to breakage; however, their actual role and properties remain elusive. Our in silico analysis carried out genome-wide via the StabFlex algorithm, shows the conserved presence of highly flexible regions in budding yeast genome as well as in genomes of other Saccharomyces sensu stricto species. Flexibile peaks in S. cerevisiae identify 175 ORFs mapping on their 3'UTR, a region affecting mRNA translation, localization and stability. (TA)n repeats of different extension shape the central structure of peaks and co-localize with polyadenylation efficiency element (EE) signals. ORFs with flexible peaks share common features. Transcripts are characterized by decreased half-life: this is considered peculiar of genes involved in regulatory systems with high turnover; consistently, their function affects biological processes such as cell cycle regulation or stress response. Our findings support the functional importance of flexibility peaks, suggesting that the flexible sequence may be derived by an expansion of canonical TAYRTA polyadenylation efficiency element. The flexible (TA)n repeat amplification could be the outcome of an evolutionary neofunctionalization leading to a differential 3'-end processing and expression regulation in genes with peculiar function. Our study provides a new support to the functional role of flexibility in genomes and a

  17. Genome duplication and mutations in ACE2 cause multicellular, fast-sedimenting phenotypes in evolved Saccharomyces cerevisiae.

    Science.gov (United States)

    Oud, Bart; Guadalupe-Medina, Victor; Nijkamp, Jurgen F; de Ridder, Dick; Pronk, Jack T; van Maris, Antonius J A; Daran, Jean-Marc

    2013-11-05

    Laboratory evolution of the yeast Saccharomyces cerevisiae in bioreactor batch cultures yielded variants that grow as multicellular, fast-sedimenting clusters. Knowledge of the molecular basis of this phenomenon may contribute to the understanding of natural evolution of multicellularity and to manipulating cell sedimentation in laboratory and industrial applications of S. cerevisiae. Multicellular, fast-sedimenting lineages obtained from a haploid S. cerevisiae strain in two independent evolution experiments were analyzed by whole genome resequencing. The two evolved cell lines showed different frameshift mutations in a stretch of eight adenosines in ACE2, which encodes a transcriptional regulator involved in cell cycle control and mother-daughter cell separation. Introduction of the two ace2 mutant alleles into the haploid parental strain led to slow-sedimenting cell clusters that consisted of just a few cells, thus representing only a partial reconstruction of the evolved phenotype. In addition to single-nucleotide mutations, a whole-genome duplication event had occurred in both evolved multicellular strains. Construction of a diploid reference strain with two mutant ace2 alleles led to complete reconstruction of the multicellular-fast sedimenting phenotype. This study shows that whole-genome duplication and a frameshift mutation in ACE2 are sufficient to generate a fast-sedimenting, multicellular phenotype in S. cerevisiae. The nature of the ace2 mutations and their occurrence in two independent evolution experiments encompassing fewer than 500 generations of selective growth suggest that switching between unicellular and multicellular phenotypes may be relevant for competitiveness of S. cerevisiae in natural environments.

  18. Update History of This Database - PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available List Contact us PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods ...B link & Genome analysis methods English archive site is opened. 2012/08/08 PGDBj... Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods is opened. About This...ate History of This Database - PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods | LSDB Archive ...

  19. Origin of the duplicated regions in the yeast genomes

    DEFF Research Database (Denmark)

    Piskur, Jure

    2001-01-01

    The genome of Saccharomyces cerevisiae contains several duplicated regions. The recent sequencing results of several yeast species suggest that the duplicated regions found in the modern Saccharomyces species are probably the result of a single gross duplication, as well as a series of sporadic...

  20. TcruziDB, an Integrated Database, and the WWW Information Server for the Trypanosoma cruzi Genome Project

    Directory of Open Access Journals (Sweden)

    Degrave Wim

    1997-01-01

    Full Text Available Data analysis, presentation and distribution is of utmost importance to a genome project. A public domain software, ACeDB, has been chosen as the common basis for parasite genome databases, and a first release of TcruziDB, the Trypanosoma cruzi genome database, is available by ftp from ftp://iris.dbbm.fiocruz.br/pub/genomedb/TcruziDB as well as versions of the software for different operating systems (ftp://iris.dbbm.fiocruz.br/pub/unixsoft/. Moreover, data originated from the project are available from the WWW server at http://www.dbbm.fiocruz.br. It contains biological and parasitological data on CL Brener, its karyotype, all available T. cruzi sequences from Genbank, data on the EST-sequencing project and on available libraries, a T. cruzi codon table and a listing of activities and participating groups in the genome project, as well as meeting reports. T. cruzi discussion lists (tcruzi-l@iris.dbbm.fiocruz.br and tcgenics@iris.dbbm.fiocruz.br are being maintained for communication and to promote collaboration in the genome project

  1. CBS Genome Atlas Database: a dynamic storage for bioinformatic results and sequence data

    DEFF Research Database (Denmark)

    Hallin, Peter Fischer; Ussery, David

    2004-01-01

    , these results counts to more than 220 pieces of information. The backbone of this solution consists of a program package written in Perl, which enables administrators to synchronize and update the database content. The MySQL database has been connected to the CBS web-server via PHP4, to present a dynamic web...... and frequent addition of new models are factors that require a dynamic database layout. Using basic tools like the GNU Make system, csh, Perl and MySQL, we have created a flexible database environment for storing and maintaining such results for a collection of complete microbial genomes. Currently...... content for users outside the center. This solution is tightly fitted to existing server infrastructure and the solutions proposed here can perhaps serve as a template for other research groups to solve database issues....

  2. MBGD update 2013: the microbial genome database for exploring the diversity of microbial world.

    Science.gov (United States)

    Uchiyama, Ikuo; Mihara, Motohiro; Nishide, Hiroyo; Chiba, Hirokazu

    2013-01-01

    The microbial genome database for comparative analysis (MBGD, available at http://mbgd.genome.ad.jp/) is a platform for microbial genome comparison based on orthology analysis. As its unique feature, MBGD allows users to conduct orthology analysis among any specified set of organisms; this flexibility allows MBGD to adapt to a variety of microbial genomic study. Reflecting the huge diversity of microbial world, the number of microbial genome projects now becomes several thousands. To efficiently explore the diversity of the entire microbial genomic data, MBGD now provides summary pages for pre-calculated ortholog tables among various taxonomic groups. For some closely related taxa, MBGD also provides the conserved synteny information (core genome alignment) pre-calculated using the CoreAligner program. In addition, efficient incremental updating procedure can create extended ortholog table by adding additional genomes to the default ortholog table generated from the representative set of genomes. Combining with the functionalities of the dynamic orthology calculation of any specified set of organisms, MBGD is an efficient and flexible tool for exploring the microbial genome diversity.

  3. Genetic characterization of strains of Saccharomyces uvarum from New Zealand wineries.

    Science.gov (United States)

    Zhang, Hanyao; Richards, Keith D; Wilson, Sandra; Lee, Soon A; Sheehan, Hester; Roncoroni, Miguel; Gardner, Richard C

    2015-04-01

    We present a genetic characterization of 65 isolates of Saccharomyces uvarum isolated from wineries in New Zealand, along with the complete nucleotide sequence of a single sulfite-tolerant isolate. The genome of the New Zealand isolate averaged 99.85% nucleotide identity to CBS7001, the previously sequenced strain of S. uvarum. However, three genomic segments (37-87 kb) showed 10% nucleotide divergence from CBS7001 but 99% identity to Saccharomyces eubayanus. We conclude that these three segments appear to have been introgressed from that species. The nucleotide sequence of the internal transcribed spacer (ITS) region from other New Zealand isolates were also very similar to that of CBS7001, and hybrids showed complete genetic compatibility for some strains, with tetrads giving four viable progeny that showed 2:2 segregations of marker genes. Some strains showed high tolerance to sulfite, with genetic analysis indicating linkage of this trait to the transcription factor FZF1, but not to SSU1, the sulfite efflux pump that it regulates in order to confer sulfite tolerance in Saccharomyces cerevisiae. The fermentation characteristics of selected strains of S. uvarum showed exceptionally good cold fermentation characteristics, superior to the best commercially available strains of S. cerevisiae. Copyright © 2014 Elsevier Ltd. All rights reserved.

  4. LDSplitDB: a database for studies of meiotic recombination hotspots in MHC using human genomic data.

    Science.gov (United States)

    Guo, Jing; Chen, Hao; Yang, Peng; Lee, Yew Ti; Wu, Min; Przytycka, Teresa M; Kwoh, Chee Keong; Zheng, Jie

    2018-04-20

    Meiotic recombination happens during the process of meiosis when chromosomes inherited from two parents exchange genetic materials to generate chromosomes in the gamete cells. The recombination events tend to occur in narrow genomic regions called recombination hotspots. Its dysregulation could lead to serious human diseases such as birth defects. Although the regulatory mechanism of recombination events is still unclear, DNA sequence polymorphisms have been found to play crucial roles in the regulation of recombination hotspots. To facilitate the studies of the underlying mechanism, we developed a database named LDSplitDB which provides an integrative and interactive data mining and visualization platform for the genome-wide association studies of recombination hotspots. It contains the pre-computed association maps of the major histocompatibility complex (MHC) region in the 1000 Genomes Project and the HapMap Phase III datasets, and a genome-scale study of the European population from the HapMap Phase II dataset. Besides the recombination profiles, related data of genes, SNPs and different types of epigenetic modifications, which could be associated with meiotic recombination, are provided for comprehensive analysis. To meet the computational requirement of the rapidly increasing population genomics data, we prepared a lookup table of 400 haplotypes for recombination rate estimation using the well-known LDhat algorithm which includes all possible two-locus haplotype configurations. To the best of our knowledge, LDSplitDB is the first large-scale database for the association analysis of human recombination hotspots with DNA sequence polymorphisms. It provides valuable resources for the discovery of the mechanism of meiotic recombination hotspots. The information about MHC in this database could help understand the roles of recombination in human immune system. DATABASE URL: http://histone.scse.ntu.edu.sg/LDSplitDB.

  5. Genome sequence of the lager brewing yeast, an interspecies hybrid.

    Science.gov (United States)

    Nakao, Yoshihiro; Kanamori, Takeshi; Itoh, Takehiko; Kodama, Yukiko; Rainieri, Sandra; Nakamura, Norihisa; Shimonaga, Tomoko; Hattori, Masahira; Ashikari, Toshihiko

    2009-04-01

    This work presents the genome sequencing of the lager brewing yeast (Saccharomyces pastorianus) Weihenstephan 34/70, a strain widely used in lager beer brewing. The 25 Mb genome comprises two nuclear sub-genomes originating from Saccharomyces cerevisiae and Saccharomyces bayanus and one circular mitochondrial genome originating from S. bayanus. Thirty-six different types of chromosomes were found including eight chromosomes with translocations between the two sub-genomes, whose breakpoints are within the orthologous open reading frames. Several gene loci responsible for typical lager brewing yeast characteristics such as maltotriose uptake and sulfite production have been increased in number by chromosomal rearrangements. Despite an overall high degree of conservation of the synteny with S. cerevisiae and S. bayanus, the syntenies were not well conserved in the sub-telomeric regions that contain lager brewing yeast characteristic and specific genes. Deletion of larger chromosomal regions, a massive unilateral decrease of the ribosomal DNA cluster and bilateral truncations of over 60 genes reflect a post-hybridization evolution process. Truncations and deletions of less efficient maltose and maltotriose uptake genes may indicate the result of adaptation to brewing. The genome sequence of this interspecies hybrid yeast provides a new tool for better understanding of lager brewing yeast behavior in industrial beer production.

  6. Evidence for Divergent Evolution of Growth Temperature Preference in Sympatric Saccharomyces Species

    Science.gov (United States)

    Gonçalves, Paula; Valério, Elisabete; Correia, Cláudia; de Almeida, João M. G. C. F.; Sampaio, José Paulo

    2011-01-01

    The genus Saccharomyces currently includes eight species in addition to the model yeast Saccharomyces cerevisiae, most of which can be consistently isolated from tree bark and soil. We recently found sympatric pairs of Saccharomyces species, composed of one cryotolerant and one thermotolerant species in oak bark samples of various geographic origins. In order to contribute to explain the occurrence in sympatry of Saccharomyces species, we screened Saccharomyces genomic data for protein divergence that might be correlated to distinct growth temperature preferences of the species, using the dN/dS ratio as a measure of protein evolution rates and pair-wise species comparisons. In addition to proteins previously implicated in growth at suboptimal temperatures, we found that glycolytic enzymes were among the proteins exhibiting higher than expected divergence when one cryotolerant and one thermotolerant species are compared. By measuring glycolytic fluxes and glycolytic enzymatic activities in different species and at different temperatures, we subsequently show that the unusual divergence of glycolytic genes may be related to divergent evolution of the glycolytic pathway aligning its performance to the growth temperature profiles of the different species. In general, our results support the view that growth temperature preference is a trait that may have undergone divergent selection in the course of ecological speciation in Saccharomyces. PMID:21674061

  7. Evaluation of relational and NoSQL database architectures to manage genomic annotations.

    Science.gov (United States)

    Schulz, Wade L; Nelson, Brent G; Felker, Donn K; Durant, Thomas J S; Torres, Richard

    2016-12-01

    While the adoption of next generation sequencing has rapidly expanded, the informatics infrastructure used to manage the data generated by this technology has not kept pace. Historically, relational databases have provided much of the framework for data storage and retrieval. Newer technologies based on NoSQL architectures may provide significant advantages in storage and query efficiency, thereby reducing the cost of data management. But their relative advantage when applied to biomedical data sets, such as genetic data, has not been characterized. To this end, we compared the storage, indexing, and query efficiency of a common relational database (MySQL), a document-oriented NoSQL database (MongoDB), and a relational database with NoSQL support (PostgreSQL). When used to store genomic annotations from the dbSNP database, we found the NoSQL architectures to outperform traditional, relational models for speed of data storage, indexing, and query retrieval in nearly every operation. These findings strongly support the use of novel database technologies to improve the efficiency of data management within the biological sciences. Copyright © 2016 Elsevier Inc. All rights reserved.

  8. PGG.Population: a database for understanding the genomic diversity and genetic ancestry of human populations.

    Science.gov (United States)

    Zhang, Chao; Gao, Yang; Liu, Jiaojiao; Xue, Zhe; Lu, Yan; Deng, Lian; Tian, Lei; Feng, Qidi; Xu, Shuhua

    2018-01-04

    There are a growing number of studies focusing on delineating genetic variations that are associated with complex human traits and diseases due to recent advances in next-generation sequencing technologies. However, identifying and prioritizing disease-associated causal variants relies on understanding the distribution of genetic variations within and among populations. The PGG.Population database documents 7122 genomes representing 356 global populations from 107 countries and provides essential information for researchers to understand human genomic diversity and genetic ancestry. These data and information can facilitate the design of research studies and the interpretation of results of both evolutionary and medical studies involving human populations. The database is carefully maintained and constantly updated when new data are available. We included miscellaneous functions and a user-friendly graphical interface for visualization of genomic diversity, population relationships (genetic affinity), ancestral makeup, footprints of natural selection, and population history etc. Moreover, PGG.Population provides a useful feature for users to analyze data and visualize results in a dynamic style via online illustration. The long-term ambition of the PGG.Population, together with the joint efforts from other researchers who contribute their data to our database, is to create a comprehensive depository of geographic and ethnic variation of human genome, as well as a platform bringing influence on future practitioners of medicine and clinical investigators. PGG.Population is available at https://www.pggpopulation.org. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  9. Cpf1-Database: web-based genome-wide guide RNA library design for gene knockout screens using CRISPR-Cpf1.

    Science.gov (United States)

    Park, Jeongbin; Bae, Sangsu

    2018-03-15

    Following the type II CRISPR-Cas9 system, type V CRISPR-Cpf1 endonucleases have been found to be applicable for genome editing in various organisms in vivo. However, there are as yet no web-based tools capable of optimally selecting guide RNAs (gRNAs) among all possible genome-wide target sites. Here, we present Cpf1-Database, a genome-wide gRNA library design tool for LbCpf1 and AsCpf1, which have DNA recognition sequences of 5'-TTTN-3' at the 5' ends of target sites. Cpf1-Database provides a sophisticated but simple way to design gRNAs for AsCpf1 nucleases on the genome scale. One can easily access the data using a straightforward web interface, and using the powerful collections feature one can easily design gRNAs for thousands of genes in short time. Free access at http://www.rgenome.net/cpf1-database/. sangsubae@hanyang.ac.kr.

  10. Genome Mutational and Transcriptional Hotspots Are Traps for Duplicated Genes and Sources of Adaptations.

    Science.gov (United States)

    Fares, Mario A; Sabater-Muñoz, Beatriz; Toft, Christina

    2017-05-01

    Gene duplication generates new genetic material, which has been shown to lead to major innovations in unicellular and multicellular organisms. A whole-genome duplication occurred in the ancestor of Saccharomyces yeast species but 92% of duplicates returned to single-copy genes shortly after duplication. The persisting duplicated genes in Saccharomyces led to the origin of major metabolic innovations, which have been the source of the unique biotechnological capabilities in the Baker's yeast Saccharomyces cerevisiae. What factors have determined the fate of duplicated genes remains unknown. Here, we report the first demonstration that the local genome mutation and transcription rates determine the fate of duplicates. We show, for the first time, a preferential location of duplicated genes in the mutational and transcriptional hotspots of S. cerevisiae genome. The mechanism of duplication matters, with whole-genome duplicates exhibiting different preservation trends compared to small-scale duplicates. Genome mutational and transcriptional hotspots are rich in duplicates with large repetitive promoter elements. Saccharomyces cerevisiae shows more tolerance to deleterious mutations in duplicates with repetitive promoter elements, which in turn exhibit higher transcriptional plasticity against environmental perturbations. Our data demonstrate that the genome traps duplicates through the accelerated regulatory and functional divergence of their gene copies providing a source of novel adaptations in yeast. © The Author 2017. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

  11. The Genomes On Line Database (GOLD) in 2009: status of genomic and metagenomic projects and their associated metadata

    Science.gov (United States)

    Liolios, Konstantinos; Chen, I-Min A.; Mavromatis, Konstantinos; Tavernarakis, Nektarios; Hugenholtz, Philip; Markowitz, Victor M.; Kyrpides, Nikos C.

    2010-01-01

    The Genomes On Line Database (GOLD) is a comprehensive resource for centralized monitoring of genome and metagenome projects worldwide. Both complete and ongoing projects, along with their associated metadata, can be accessed in GOLD through precomputed tables and a search page. As of September 2009, GOLD contains information for more than 5800 sequencing projects, of which 1100 have been completed and their sequence data deposited in a public repository. GOLD continues to expand, moving toward the goal of providing the most comprehensive repository of metadata information related to the projects and their organisms/environments in accordance with the Minimum Information about a (Meta)Genome Sequence (MIGS/MIMS) specification. GOLD is available at: http://www.genomesonline.org and has a mirror site at the Institute of Molecular Biology and Biotechnology, Crete, Greece, at: http://gold.imbb.forth.gr/ PMID:19914934

  12. A database of PCR primers for the chloroplast genomes of higher plants

    Science.gov (United States)

    Heinze, Berthold

    2007-01-01

    Background Chloroplast genomes evolve slowly and many primers for PCR amplification and analysis of chloroplast sequences can be used across a wide array of genera. In some cases 'universal' primers have been designed for the purpose of working across species boundaries. However, the essential information on these primer sequences is scattered throughout the literature. Results A database is presented here which assembles published primer information for chloroplast DNA. Additional primers were designed to fill gaps where little or no primer information could be found. Amplicons are either the genes themselves (typically useful in studies of sequence variation in higher-order phylogeny) or they are spacers, introns, and intergenic regions (for studies of phylogeographic patterns within and among species). The current list of 'generic' primers consists of more than 700 sequences. Wherever possible, we give the locations of the primers in the thirteen fully sequenced chloroplast genomes (Nicotiana tabacum, Atropa belladonna, Spinacia oleracea, Arabidopsis thaliana, Populus trichocarpa, Oryza sativa, Pinus thunbergii, Marchantia polymorpha, Zea mays, Oenothera elata, Acorus calamus, Eucalyptus globulus, Medicago trunculata). Conclusion The database described here is designed to serve as a resource for researchers who are venturing into the study of poorly described chloroplast genomes, whether for large- or small-scale DNA sequencing projects, to study molecular variation or to investigate chloroplast evolution. PMID:17326828

  13. A database of PCR primers for the chloroplast genomes of higher plants

    Directory of Open Access Journals (Sweden)

    Heinze Berthold

    2007-02-01

    Full Text Available Abstract Background Chloroplast genomes evolve slowly and many primers for PCR amplification and analysis of chloroplast sequences can be used across a wide array of genera. In some cases 'universal' primers have been designed for the purpose of working across species boundaries. However, the essential information on these primer sequences is scattered throughout the literature. Results A database is presented here which assembles published primer information for chloroplast DNA. Additional primers were designed to fill gaps where little or no primer information could be found. Amplicons are either the genes themselves (typically useful in studies of sequence variation in higher-order phylogeny or they are spacers, introns, and intergenic regions (for studies of phylogeographic patterns within and among species. The current list of 'generic' primers consists of more than 700 sequences. Wherever possible, we give the locations of the primers in the thirteen fully sequenced chloroplast genomes (Nicotiana tabacum, Atropa belladonna, Spinacia oleracea, Arabidopsis thaliana, Populus trichocarpa, Oryza sativa, Pinus thunbergii, Marchantia polymorpha, Zea mays, Oenothera elata, Acorus calamus, Eucalyptus globulus, Medicago trunculata. Conclusion The database described here is designed to serve as a resource for researchers who are venturing into the study of poorly described chloroplast genomes, whether for large- or small-scale DNA sequencing projects, to study molecular variation or to investigate chloroplast evolution.

  14. Mitochondrial introgression suggests extensive ancestral hybridization events among Saccharomyces species.

    Science.gov (United States)

    Peris, David; Arias, Armando; Orlić, Sandi; Belloch, Carmela; Pérez-Través, Laura; Querol, Amparo; Barrio, Eladio

    2017-03-01

    Horizontal gene transfer (HGT) in eukaryotic plastids and mitochondrial genomes is common, and plays an important role in organism evolution. In yeasts, recent mitochondrial HGT has been suggested between S. cerevisiae and S. paradoxus. However, few strains have been explored given the lack of accurate mitochondrial genome annotations. Mitochondrial genome sequences are important to understand how frequent these introgressions occur, and their role in cytonuclear incompatibilities and fitness. Indeed, most of the Bateson-Dobzhansky-Muller genetic incompatibilities described in yeasts are driven by cytonuclear incompatibilities. We herein explored the mitochondrial inheritance of several worldwide distributed wild Saccharomyces species and their hybrids isolated from different sources and geographic origins. We demonstrated the existence of several recombination points in mitochondrial region COX2-ORF1, likely mediated by either the activity of the protein encoded by the ORF1 (F-SceIII) gene, a free-standing homing endonuclease, or mostly facilitated by A+T tandem repeats and regions of integration of GC clusters. These introgressions were shown to occur among strains of the same species and among strains of different species, which suggests a complex model of Saccharomyces evolution that involves several ancestral hybridization events in wild environments. Copyright © 2017 Elsevier Inc. All rights reserved.

  15. A database of phylogenetically atypical genes in archaeal and bacterial genomes, identified using the DarkHorse algorithm

    Directory of Open Access Journals (Sweden)

    Allen Eric E

    2008-10-01

    Full Text Available Abstract Background The process of horizontal gene transfer (HGT is believed to be widespread in Bacteria and Archaea, but little comparative data is available addressing its occurrence in complete microbial genomes. Collection of high-quality, automated HGT prediction data based on phylogenetic evidence has previously been impractical for large numbers of genomes at once, due to prohibitive computational demands. DarkHorse, a recently described statistical method for discovering phylogenetically atypical genes on a genome-wide basis, provides a means to solve this problem through lineage probability index (LPI ranking scores. LPI scores inversely reflect phylogenetic distance between a test amino acid sequence and its closest available database matches. Proteins with low LPI scores are good horizontal gene transfer candidates; those with high scores are not. Description The DarkHorse algorithm has been applied to 955 microbial genome sequences, and the results organized into a web-searchable relational database, called the DarkHorse HGT Candidate Resource http://darkhorse.ucsd.edu. Users can select individual genomes or groups of genomes to screen by LPI score, search for protein functions by descriptive annotation or amino acid sequence similarity, or select proteins with unusual G+C composition in their underlying coding sequences. The search engine reports LPI scores for match partners as well as query sequences, providing the opportunity to explore whether potential HGT donor sequences are phylogenetically typical or atypical within their own genomes. This information can be used to predict whether or not sufficient information is available to build a well-supported phylogenetic tree using the potential donor sequence. Conclusion The DarkHorse HGT Candidate database provides a powerful, flexible set of tools for identifying phylogenetically atypical proteins, allowing researchers to explore both individual HGT events in single genomes, and

  16. MicrobesOnline: an integrated portal for comparative and functional genomics

    Energy Technology Data Exchange (ETDEWEB)

    Dehal, Paramvir S.; Joachimiak, Marcin P.; Price, Morgan N.; Bates, John T.; Baumohl, Jason K.; Chivian, Dylan; Friedland, Greg D.; Huang, Katherine H.; Keller, Keith; Novichkov, Pavel S.; Dubchak, Inna L.; Alm, Eric J.; Arkin, Adam P.

    2009-09-17

    Since 2003, MicrobesOnline (http://www.microbesonline.org) has been providing a community resource for comparative and functional genome analysis. The portal includes over 1000 complete genomes of bacteria, archaea and fungi and thousands of expression microarrays from diverse organisms ranging from model organisms such as Escherichia coli and Saccharomyces cerevisiae to environmental microbes such as Desulfovibrio vulgaris and Shewanella oneidensis. To assist in annotating genes and in reconstructing their evolutionary history, MicrobesOnline includes a comparative genome browser based on phylogenetic trees for every gene family as well as a species tree. To identify co-regulated genes, MicrobesOnline can search for genes based on their expression profile, and provides tools for identifying regulatory motifs and seeing if they are conserved. MicrobesOnline also includes fast phylogenetic profile searches, comparative views of metabolic pathways, operon predictions, a workbench for sequence analysis and integration with RegTransBase and other microbial genome resources. The next update of MicrobesOnline will contain significant new functionality, including comparative analysis of metagenomic sequence data. Programmatic access to the database, along with source code and documentation, is available at http://microbesonline.org/programmers.html.

  17. MicrobesOnline: an integrated portal for comparative and functional genomics

    Energy Technology Data Exchange (ETDEWEB)

    Dehal, Paramvir; Joachimiak, Marcin; Price, Morgan; Bates, John; Baumohl, Jason; Chivian, Dylan; Friedland, Greg; Huang, Kathleen; Keller, Keith; Novichkov, Pavel; Dubchak, Inna; Alm, Eric; Arkin, Adam

    2011-07-14

    Since 2003, MicrobesOnline (http://www.microbesonline.org) has been providing a community resource for comparative and functional genome analysis. The portal includes over 1000 complete genomes of bacteria, archaea and fungi and thousands of expression microarrays from diverse organisms ranging from model organisms such as Escherichia coli and Saccharomyces cerevisiae to environmental microbes such as Desulfovibrio vulgaris and Shewanella oneidensis. To assist in annotating genes and in reconstructing their evolutionary history, MicrobesOnline includes a comparative genome browser based on phylogenetic trees for every gene family as well as a species tree. To identify co-regulated genes, MicrobesOnline can search for genes based on their expression profile, and provides tools for identifying regulatory motifs and seeing if they are conserved. MicrobesOnline also includes fast phylogenetic profile searches, comparative views of metabolic pathways, operon predictions, a workbench for sequence analysis and integration with RegTransBase and other microbial genome resources. The next update of MicrobesOnline will contain significant new functionality, including comparative analysis of metagenomic sequence data. Programmatic access to the database, along with source code and documentation, is available at http://microbesonline.org/programmers.html.

  18. SolCyc: a database hub at the Sol Genomics Network (SGN) for the manual curation of metabolic networks in Solanum and Nicotiana specific databases

    Science.gov (United States)

    Foerster, Hartmut; Bombarely, Aureliano; Battey, James N D; Sierro, Nicolas; Ivanov, Nikolai V; Mueller, Lukas A

    2018-01-01

    Abstract SolCyc is the entry portal to pathway/genome databases (PGDBs) for major species of the Solanaceae family hosted at the Sol Genomics Network. Currently, SolCyc comprises six organism-specific PGDBs for tomato, potato, pepper, petunia, tobacco and one Rubiaceae, coffee. The metabolic networks of those PGDBs have been computationally predicted by the pathologic component of the pathway tools software using the manually curated multi-domain database MetaCyc (http://www.metacyc.org/) as reference. SolCyc has been recently extended by taxon-specific databases, i.e. the family-specific SolanaCyc database, containing only curated data pertinent to species of the nightshade family, and NicotianaCyc, a genus-specific database that stores all relevant metabolic data of the Nicotiana genus. Through manual curation of the published literature, new metabolic pathways have been created in those databases, which are complemented by the continuously updated, relevant species-specific pathways from MetaCyc. At present, SolanaCyc comprises 199 pathways and 29 superpathways and NicotianaCyc accounts for 72 pathways and 13 superpathways. Curator-maintained, taxon-specific databases such as SolanaCyc and NicotianaCyc are characterized by an enrichment of data specific to these taxa and free of falsely predicted pathways. Both databases have been used to update recently created Nicotiana-specific databases for Nicotiana tabacum, Nicotiana benthamiana, Nicotiana sylvestris and Nicotiana tomentosiformis by propagating verifiable data into those PGDBs. In addition, in-depth curation of the pathways in N.tabacum has been carried out which resulted in the elimination of 156 pathways from the 569 pathways predicted by pathway tools. Together, in-depth curation of the predicted pathway network and the supplementation with curated data from taxon-specific databases has substantially improved the curation status of the species–specific N.tabacum PGDB. The implementation of this

  19. 2μ plasmid in Saccharomyces species and in Saccharomyces cerevisiae.

    Science.gov (United States)

    Strope, Pooja K; Kozmin, Stanislav G; Skelly, Daniel A; Magwene, Paul M; Dietrich, Fred S; McCusker, John H

    2015-12-01

    We determined that extrachromosomal 2μ plasmid was present in 67 of the Saccharomyces cerevisiae 100-genome strains; in addition to variation in the size and copy number of 2μ, we identified three distinct classes of 2μ. We identified 2μ presence/absence and class associations with populations, clinical origin and nuclear genotypes. We also screened genome sequences of S. paradoxus, S. kudriavzevii, S. uvarum, S. eubayanus, S. mikatae, S. arboricolus and S. bayanus strains for both integrated and extrachromosomal 2μ. Similar to S. cerevisiae, we found no integrated 2μ sequences in any S. paradoxus strains. However, we identified part of 2μ integrated into the genomes of some S. uvarum, S. kudriavzevii, S. mikatae and S. bayanus strains, which were distinct from each other and from all extrachromosomal 2μ. We identified extrachromosomal 2μ in one S. paradoxus, one S. eubayanus, two S. bayanus and 13 S. uvarum strains. The extrachromosomal 2μ in S. paradoxus, S. eubayanus and S. cerevisiae were distinct from each other. In contrast, the extrachromosomal 2μ in S. bayanus and S. uvarum strains were identical with each other and with one of the three classes of S. cerevisiae 2μ, consistent with interspecific transfer. © FEMS 2015. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

  20. OryzaGenome: Genome Diversity Database of Wild Oryza Species

    KAUST Repository

    Ohyanagi, Hajime

    2015-11-18

    The species in the genus Oryza, encompassing nine genome types and 23 species, are a rich genetic resource and may have applications in deeper genomic analyses aiming to understand the evolution of plant genomes. With the advancement of next-generation sequencing (NGS) technology, a flood of Oryza species reference genomes and genomic variation information has become available in recent years. This genomic information, combined with the comprehensive phenotypic information that we are accumulating in our Oryzabase, can serve as an excellent genotype-phenotype association resource for analyzing rice functional and structural evolution, and the associated diversity of the Oryza genus. Here we integrate our previous and future phenotypic/habitat information and newly determined genotype information into a united repository, named OryzaGenome, providing the variant information with hyperlinks to Oryzabase. The current version of OryzaGenome includes genotype information of 446 O. rufipogon accessions derived by imputation and of 17 accessions derived by imputation-free deep sequencing. Two variant viewers are implemented: SNP Viewer as a conventional genome browser interface and Variant Table as a textbased browser for precise inspection of each variant one by one. Portable VCF (variant call format) file or tabdelimited file download is also available. Following these SNP (single nucleotide polymorphism) data, reference pseudomolecules/ scaffolds/contigs and genome-wide variation information for almost all of the closely and distantly related wild Oryza species from the NIG Wild Rice Collection will be available in future releases. All of the resources can be accessed through http://viewer.shigen.info/oryzagenome/.

  1. Visualizing information across multidimensional post-genomic structured and textual databases.

    Science.gov (United States)

    Tao, Ying; Friedman, Carol; Lussier, Yves A

    2005-04-15

    Visualizing relationships among biological information to facilitate understanding is crucial to biological research during the post-genomic era. Although different systems have been developed to view gene-phenotype relationships for specific databases, very few have been designed specifically as a general flexible tool for visualizing multidimensional genotypic and phenotypic information together. Our goal is to develop a method for visualizing multidimensional genotypic and phenotypic information and a model that unifies different biological databases in order to present the integrated knowledge using a uniform interface. We developed a novel, flexible and generalizable visualization tool, called PhenoGenesviewer (PGviewer), which in this paper was used to display gene-phenotype relationships from a human-curated database (OMIM) and from an automatic method using a Natural Language Processing tool called BioMedLEE. Data obtained from multiple databases were first integrated into a uniform structure and then organized by PGviewer. PGviewer provides a flexible query interface that allows dynamic selection and ordering of any desired dimension in the databases. Based on users' queries, results can be visualized using hierarchical expandable trees that present views specified by users according to their research interests. We believe that this method, which allows users to dynamically organize and visualize multiple dimensions, is a potentially powerful and promising tool that should substantially facilitate biological research. PhenogenesViewer as well as its support and tutorial are available at http://www.dbmi.columbia.edu/pgviewer/ Lussier@dbmi.columbia.edu.

  2. The duplicated genes database: identification and functional annotation of co-localised duplicated genes across genomes.

    Directory of Open Access Journals (Sweden)

    Marion Ouedraogo

    Full Text Available BACKGROUND: There has been a surge in studies linking genome structure and gene expression, with special focus on duplicated genes. Although initially duplicated from the same sequence, duplicated genes can diverge strongly over evolution and take on different functions or regulated expression. However, information on the function and expression of duplicated genes remains sparse. Identifying groups of duplicated genes in different genomes and characterizing their expression and function would therefore be of great interest to the research community. The 'Duplicated Genes Database' (DGD was developed for this purpose. METHODOLOGY: Nine species were included in the DGD. For each species, BLAST analyses were conducted on peptide sequences corresponding to the genes mapped on a same chromosome. Groups of duplicated genes were defined based on these pairwise BLAST comparisons and the genomic location of the genes. For each group, Pearson correlations between gene expression data and semantic similarities between functional GO annotations were also computed when the relevant information was available. CONCLUSIONS: The Duplicated Gene Database provides a list of co-localised and duplicated genes for several species with the available gene co-expression level and semantic similarity value of functional annotation. Adding these data to the groups of duplicated genes provides biological information that can prove useful to gene expression analyses. The Duplicated Gene Database can be freely accessed through the DGD website at http://dgd.genouest.org.

  3. MitBASE : a comprehensive and integrated mitochondrial DNA database. The present status

    NARCIS (Netherlands)

    Attimonelli, M.; Altamura, N.; Benne, R.; Brennicke, A.; Cooper, J. M.; D'Elia, D.; Montalvo, A.; Pinto, B.; de Robertis, M.; Golik, P.; Knoop, V.; Lanave, C.; Lazowska, J.; Licciulli, F.; Malladi, B. S.; Memeo, F.; Monnerot, M.; Pasimeni, R.; Pilbout, S.; Schapira, A. H.; Sloof, P.; Saccone, C.

    2000-01-01

    MitBASE is an integrated and comprehensive database of mitochondrial DNA data which collects, under a single interface, databases for Plant, Vertebrate, Invertebrate, Human, Protist and Fungal mtDNA and a Pilot database on nuclear genes involved in mitochondrial biogenesis in Saccharomyces

  4. Enhancing sesquiterpene production in Saccharomyces cerevisiae through in silico driven metabolic engineering

    DEFF Research Database (Denmark)

    Asadollahi, Mohammadali; Maury, Jerome; Patil, Kiran Raosaheb

    2009-01-01

    A genome-scale metabolic model was used to identify new target genes for enhanced biosynthesis of sesquiterpenes in the yeast Saccharomyces cerevisiae. The effect of gene deletions on the flux distributions in the metabolic model of S. cerevisiae was assessed using OptGene as the modeling framework...

  5. An Integrated Molecular Database on Indian Insects.

    Science.gov (United States)

    Pratheepa, Maria; Venkatesan, Thiruvengadam; Gracy, Gandhi; Jalali, Sushil Kumar; Rangheswaran, Rajagopal; Antony, Jomin Cruz; Rai, Anil

    2018-01-01

    MOlecular Database on Indian Insects (MODII) is an online database linking several databases like Insect Pest Info, Insect Barcode Information System (IBIn), Insect Whole Genome sequence, Other Genomic Resources of National Bureau of Agricultural Insect Resources (NBAIR), Whole Genome sequencing of Honey bee viruses, Insecticide resistance gene database and Genomic tools. This database was developed with a holistic approach for collecting information about phenomic and genomic information of agriculturally important insects. This insect resource database is available online for free at http://cib.res.in. http://cib.res.in/.

  6. Metabolic Engineering of Probiotic Saccharomyces boulardii.

    Science.gov (United States)

    Liu, Jing-Jing; Kong, In Iok; Zhang, Guo-Chang; Jayakody, Lahiru N; Kim, Heejin; Xia, Peng-Fei; Kwak, Suryang; Sung, Bong Hyun; Sohn, Jung-Hoon; Walukiewicz, Hanna E; Rao, Christopher V; Jin, Yong-Su

    2016-04-01

    Saccharomyces boulardiiis a probiotic yeast that has been used for promoting gut health as well as preventing diarrheal diseases. This yeast not only exhibits beneficial phenotypes for gut health but also can stay longer in the gut than Saccharomyces cerevisiae Therefore, S. boulardiiis an attractive host for metabolic engineering to produce biomolecules of interest in the gut. However, the lack of auxotrophic strains with defined genetic backgrounds has hampered the use of this strain for metabolic engineering. Here, we report the development of well-defined auxotrophic mutants (leu2,ura3,his3, and trp1) through clustered regularly interspaced short palindromic repeat (CRISPR)-Cas9-based genome editing. The resulting auxotrophic mutants can be used as a host for introducing various genetic perturbations, such as overexpression or deletion of a target gene, using existing genetic tools forS. cerevisiae We demonstrated the overexpression of a heterologous gene (lacZ), the correct localization of a target protein (red fluorescent protein) into mitochondria by using a protein localization signal, and the introduction of a heterologous metabolic pathway (xylose-assimilating pathway) in the genome ofS. boulardii We further demonstrated that human lysozyme, which is beneficial for human gut health, could be secreted by S. boulardii Our results suggest that more sophisticated genetic perturbations to improveS. boulardii can be performed without using a drug resistance marker, which is a prerequisite for in vivo applications using engineeredS. boulardii. Copyright © 2016, American Society for Microbiology. All Rights Reserved.

  7. Adaptive Response and Tolerance to Acetic Acid in Saccharomyces cerevisiae and Zygosaccharomyces bailii: A Physiological Genomics Perspective.

    Science.gov (United States)

    Palma, Margarida; Guerreiro, Joana F; Sá-Correia, Isabel

    2018-01-01

    Acetic acid is an important microbial growth inhibitor in the food industry; it is used as a preservative in foods and beverages and is produced during normal yeast metabolism in biotechnological processes. Acetic acid is also a major inhibitory compound present in lignocellulosic hydrolysates affecting the use of this promising carbon source for sustainable bioprocesses. Although the molecular mechanisms underlying Saccharomyces cerevisiae response and adaptation to acetic acid have been studied for years, only recently they have been examined in more detail in Zygosaccharomyces bailii . However, due to its remarkable tolerance to acetic acid and other weak acids this yeast species is a major threat in the spoilage of acidic foods and beverages and considered as an interesting alternative cell factory in Biotechnology. This review paper emphasizes genome-wide strategies that are providing global insights into the molecular targets, signaling pathways and mechanisms behind S. cerevisiae and Z. bailii tolerance to acetic acid, and extends this information to other weak acids whenever relevant. Such comprehensive perspective and the knowledge gathered in these two yeast species allowed the identification of candidate molecular targets, either for the design of effective strategies to overcome yeast spoilage in acidic foods and beverages, or for the rational genome engineering to construct more robust industrial strains. Examples of successful applications are provided.

  8. Genome shuffling of Saccharomyces cerevisiae through recursive population mating to evolve tolerance to inhibitors of Spent Sulfite Liquor

    Energy Technology Data Exchange (ETDEWEB)

    Martin, V.J.J.; Pinel, D.J.; D' aoust, F. [Concordia Univ., Montreal, PQ (Canada). Dept. of Biological Sciences; Bajwa, P.K.; Trevors, J.T.; Lee, H. [Guelph Univ., ON (Canada). Dept. of Environmental Biology

    2009-07-01

    The biochemical steps in the conversion of cellulosics to biofuels include the pretreatment, hydrolysis and fermentation of substrates into a final product. Fermentation of lignocellulosic substrates derived from waste biomass requires metabolic engineering. A biochemical flow chart from the Tembec Biorefinery plant was presented in which Spent Sulfite Liquor (SSL) was used to add value to the pulp and paper industry. The sugars contained in this carbohydrate-rich effluent from sulfite pulping were used to produce ethanol. A robust, ethanologenic microorganism that can withstand the substrate toxicity was needed. Saccharomyces cerevisiae is currently used for the production of ethanol from SSL. This yeast will succumb to toxicity and inhibition, particularly in the most inhibitor rich forms of SSL such as hardwood SSL (HWSSL). A genome shuffling method was therefore developed to create a better SSL fermenting strain. This method was designed to improve polygenic traits by generating pools of mutants with improved phenotypes, followed by iterative recombination between their genomes. Through 5 rounds of recursive mating and screening, 3 strains that could survive and grow in undiluted HWSSL were obtained. The study demonstrated that the tolerance of these strains to SSL translates into an increased capacity to produce ethanol over time using this substrate, due to continued viability of the yeast population. Phenotypic analysis of the three strains revealed that the genome shuffling approach successfully co-evolved tolerance to acetic acid, NaCl (osmotic) and HMF. A systems biology analysis of strain R57 was initiated in order to establish the genetic basis for HWSSL tolerance. tabs., figs.

  9. Human Ageing Genomic Resources: Integrated databases and tools for the biology and genetics of ageing

    Science.gov (United States)

    Tacutu, Robi; Craig, Thomas; Budovsky, Arie; Wuttke, Daniel; Lehmann, Gilad; Taranukha, Dmitri; Costa, Joana; Fraifeld, Vadim E.; de Magalhães, João Pedro

    2013-01-01

    The Human Ageing Genomic Resources (HAGR, http://genomics.senescence.info) is a freely available online collection of research databases and tools for the biology and genetics of ageing. HAGR features now several databases with high-quality manually curated data: (i) GenAge, a database of genes associated with ageing in humans and model organisms; (ii) AnAge, an extensive collection of longevity records and complementary traits for >4000 vertebrate species; and (iii) GenDR, a newly incorporated database, containing both gene mutations that interfere with dietary restriction-mediated lifespan extension and consistent gene expression changes induced by dietary restriction. Since its creation about 10 years ago, major efforts have been undertaken to maintain the quality of data in HAGR, while further continuing to develop, improve and extend it. This article briefly describes the content of HAGR and details the major updates since its previous publications, in terms of both structure and content. The completely redesigned interface, more intuitive and more integrative of HAGR resources, is also presented. Altogether, we hope that through its improvements, the current version of HAGR will continue to provide users with the most comprehensive and accessible resources available today in the field of biogerontology. PMID:23193293

  10. A DATABASE FOR TRACKING TOXICOGENOMIC SAMPLES AND PROCEDURES WITH GENOMIC, PROTEOMIC AND METABONOMIC COMPONENTS

    Science.gov (United States)

    A Database for Tracking Toxicogenomic Samples and Procedures with Genomic, Proteomic and Metabonomic Components Wenjun Bao1, Jennifer Fostel2, Michael D. Waters2, B. Alex Merrick2, Drew Ekman3, Mitchell Kostich4, Judith Schmid1, David Dix1Office of Research and Developmen...

  11. PATtyFams: Protein families for the microbial genomes in the PATRIC database

    Directory of Open Access Journals (Sweden)

    James J Davis

    2016-02-01

    Full Text Available The ability to build accurate protein families is a fundamental operation in bioinformatics that influences comparative analyses, genome annotation and metabolic modeling. For several years we have been maintaining protein families for all microbial genomes in the PATRIC database (Pathosystems Resource Integration Center, patricbrc.org in order to drive many of the comparative analysis tools that are available through the PATRIC website. However, due to the burgeoning number of genomes, traditional approaches for generating protein families are becoming prohibitive. In this report, we describe a new approach for generating protein families, which we call PATtyFams. This method uses the k-mer-based function assignments available through RAST (Rapid Annotation using Subsystem Technology to rapidly guide family formation, and then differentiates the function-based groups into families using a Markov Cluster algorithm (MCL. This new approach for generating protein families is rapid, scalable and has properties that are consistent with alignment-based methods.

  12. Construction of an ortholog database using the semantic web technology for integrative analysis of genomic data.

    Science.gov (United States)

    Chiba, Hirokazu; Nishide, Hiroyo; Uchiyama, Ikuo

    2015-01-01

    Recently, various types of biological data, including genomic sequences, have been rapidly accumulating. To discover biological knowledge from such growing heterogeneous data, a flexible framework for data integration is necessary. Ortholog information is a central resource for interlinking corresponding genes among different organisms, and the Semantic Web provides a key technology for the flexible integration of heterogeneous data. We have constructed an ortholog database using the Semantic Web technology, aiming at the integration of numerous genomic data and various types of biological information. To formalize the structure of the ortholog information in the Semantic Web, we have constructed the Ortholog Ontology (OrthO). While the OrthO is a compact ontology for general use, it is designed to be extended to the description of database-specific concepts. On the basis of OrthO, we described the ortholog information from our Microbial Genome Database for Comparative Analysis (MBGD) in the form of Resource Description Framework (RDF) and made it available through the SPARQL endpoint, which accepts arbitrary queries specified by users. In this framework based on the OrthO, the biological data of different organisms can be integrated using the ortholog information as a hub. Besides, the ortholog information from different data sources can be compared with each other using the OrthO as a shared ontology. Here we show some examples demonstrating that the ortholog information described in RDF can be used to link various biological data such as taxonomy information and Gene Ontology. Thus, the ortholog database using the Semantic Web technology can contribute to biological knowledge discovery through integrative data analysis.

  13. Genome-scale consequences of cofactor balancing in engineered pentose utilization pathways in Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Amit Ghosh

    Full Text Available Biofuels derived from lignocellulosic biomass offer promising alternative renewable energy sources for transportation fuels. Significant effort has been made to engineer Saccharomyces cerevisiae to efficiently ferment pentose sugars such as D-xylose and L-arabinose into biofuels such as ethanol through heterologous expression of the fungal D-xylose and L-arabinose pathways. However, one of the major bottlenecks in these fungal pathways is that the cofactors are not balanced, which contributes to inefficient utilization of pentose sugars. We utilized a genome-scale model of S. cerevisiae to predict the maximal achievable growth rate for cofactor balanced and imbalanced D-xylose and L-arabinose utilization pathways. Dynamic flux balance analysis (DFBA was used to simulate batch fermentation of glucose, D-xylose, and L-arabinose. The dynamic models and experimental results are in good agreement for the wild type and for the engineered D-xylose utilization pathway. Cofactor balancing the engineered D-xylose and L-arabinose utilization pathways simulated an increase in ethanol batch production of 24.7% while simultaneously reducing the predicted substrate utilization time by 70%. Furthermore, the effects of cofactor balancing the engineered pentose utilization pathways were evaluated throughout the genome-scale metabolic network. This work not only provides new insights to the global network effects of cofactor balancing but also provides useful guidelines for engineering a recombinant yeast strain with cofactor balanced engineered pathways that efficiently co-utilizes pentose and hexose sugars for biofuels production. Experimental switching of cofactor usage in enzymes has been demonstrated, but is a time-consuming effort. Therefore, systems biology models that can predict the likely outcome of such strain engineering efforts are highly useful for motivating which efforts are likely to be worth the significant time investment.

  14. Cyclone: java-based querying and computing with Pathway/Genome databases.

    Science.gov (United States)

    Le Fèvre, François; Smidtas, Serge; Schächter, Vincent

    2007-05-15

    Cyclone aims at facilitating the use of BioCyc, a collection of Pathway/Genome Databases (PGDBs). Cyclone provides a fully extensible Java Object API to analyze and visualize these data. Cyclone can read and write PGDBs, and can write its own data in the CycloneML format. This format is automatically generated from the BioCyc ontology by Cyclone itself, ensuring continued compatibility. Cyclone objects can also be stored in a relational database CycloneDB. Queries can be written in SQL, and in an intuitive and concise object-oriented query language, Hibernate Query Language (HQL). In addition, Cyclone interfaces easily with Java software including the Eclipse IDE for HQL edition, the Jung API for graph algorithms or Cytoscape for graph visualization. Cyclone is freely available under an open source license at: http://sourceforge.net/projects/nemo-cyclone. For download and installation instructions, tutorials, use cases and examples, see http://nemo-cyclone.sourceforge.net.

  15. BarleyBase—an expression profiling database for plant genomics

    Science.gov (United States)

    Shen, Lishuang; Gong, Jian; Caldo, Rico A.; Nettleton, Dan; Cook, Dianne; Wise, Roger P.; Dickerson, Julie A.

    2005-01-01

    BarleyBase (BB) (www.barleybase.org) is an online database for plant microarrays with integrated tools for data visualization and statistical analysis. BB houses raw and normalized expression data from the two publicly available Affymetrix genome arrays, Barley1 and Arabidopsis ATH1 with plans to include the new Affymetrix 61K wheat, maize, soybean and rice arrays, as they become available. BB contains a broad set of query and display options at all data levels, ranging from experiments to individual hybridizations to probe sets down to individual probes. Users can perform cross-experiment queries on probe sets based on observed expression profiles and/or based on known biological information. Probe set queries are integrated with visualization and analysis tools such as the R statistical toolbox, data filters and a large variety of plot types. Controlled vocabularies for gene and plant ontologies, as well as interconnecting links to physical or genetic map and other genomic data in PlantGDB, Gramene and GrainGenes, allow users to perform EST alignments and gene function prediction using Barley1 exemplar sequences, thus, enhancing cross-species comparison. PMID:15608273

  16. Genome update: the 1000th genome - a cautionary tale

    DEFF Research Database (Denmark)

    Lagesen, Karin; Ussery, David; Wassenaar, Gertrude Maria

    2010-01-01

    conclusions for example about the largest bacterial genome sequenced. Biological diversity is far greater than many have thought. For example, analysis of multiple Escherichia coli genomes has led to an estimate of around 45 000 gene families more genes than are recognized in the human genome. Moreover......There are now more than 1000 sequenced prokaryotic genomes deposited in public databases and available for analysis. Currently, although the sequence databases GenBank, DNA Database of Japan and EMBL are synchronized continually, there are slight differences in content at the genomes level...... for a variety of logistical reasons, including differences in format and loading errors, such as those caused by file transfer protocol interruptions. This means that the 1000th genome will be different in the various databases. Some of the data on the highly accessed web pages are inaccurate, leading to false...

  17. Physiological impact and context dependency of transcriptional responses : A chemostat study in Saccharomyces cerevisiae

    NARCIS (Netherlands)

    Tai, S.L.

    2007-01-01

    This thesis is a compilation of a four-year PhD project on bakers' yeast (Saccharomyces cerevisiae). Since the entire S. cerevisiae genome sequence became available in 1996, DNA-microarray analysis has become a popular high-information-density tool for analyzing gene expression in this important

  18. ATGC database and ATGC-COGs: an updated resource for micro- and macro-evolutionary studies of prokaryotic genomes and protein family annotation.

    Science.gov (United States)

    Kristensen, David M; Wolf, Yuri I; Koonin, Eugene V

    2017-01-04

    The Alignable Tight Genomic Clusters (ATGCs) database is a collection of closely related bacterial and archaeal genomes that provides several tools to aid research into evolutionary processes in the microbial world. Each ATGC is a taxonomy-independent cluster of 2 or more completely sequenced genomes that meet the objective criteria of a high degree of local gene order (synteny) and a small number of synonymous substitutions in the protein-coding genes. As such, each ATGC is suited for analysis of microevolutionary variations within a cohesive group of organisms (e.g. species), whereas the entire collection of ATGCs is useful for macroevolutionary studies. The ATGC database includes many forms of pre-computed data, in particular ATGC-COGs (Clusters of Orthologous Genes), multiple sequence alignments, a set of 'index' orthologs representing the most well-conserved members of each ATGC-COG, the phylogenetic tree of the organisms within each ATGC, etc. Although the ATGC database contains several million proteins from thousands of genomes organized into hundreds of clusters (roughly a 4-fold increase since the last version of the ATGC database), it is now built with completely automated methods and will be regularly updated following new releases of the NCBI RefSeq database. The ATGC database is hosted jointly at the University of Iowa at dmk-brain.ecn.uiowa.edu/ATGC/ and the NCBI at ftp.ncbi.nlm.nih.gov/pub/kristensen/ATGC/atgc_home.html. Published by Oxford University Press on behalf of Nucleic Acids Research 2016. This work is written by (a) US Government employee(s) and is in the public domain in the US.

  19. RPO41-independent maintenance of [rho-] mitochondrial DNA in Saccharomyces cerevisiae.

    Science.gov (United States)

    Fangman, W L; Henly, J W; Brewer, B J

    1990-01-01

    A subset of promoters in the mitochondrial DNA (mtDNA) of the yeast Saccharomyces cerevisiae has been proposed to participate in replication initiation, giving rise to a primer through site-specific cleavage of an RNA transcript. To test whether transcription is essential for mtDNA maintenance, we examined two simple mtDNA deletion ([rho-]) genomes in yeast cells. One genome (HS3324) contains a consensus promoter (ATATAAGTA) for the mitochondrial RNA polymerase encoded by the nuclear gene RPO41, and the other genome (4a) does not. As anticipated, in RPO41 cells transcripts from the HS3324 genome were more abundant than were transcripts from the 4a genome. When the RPO41 gene was disrupted, both [rho-] genomes were efficiently maintained. The level of transcripts from HS3324 mtDNA was decreased greater than 400-fold in cells carrying the RPO41 disrupted gene; however, the low-level transcripts from 4a mtDNA were undiminished. These results indicate that replication of [rho-] genomes can be initiated in the absence of wild-type levels of the RPO41-encoded RNA polymerase.

  20. Cellular responses of Saccharomyces cerevisiae at near-zero growth rates : Transcriptome analysis of anaerobic retentostat cultures

    NARCIS (Netherlands)

    Boender, L.G.M.; Van Maris, A.J.A.; De Hulster, E.A.F.; Almering, M.J.H.; Van der Klei, I.J.; Veenhuis, M.; De Winde, J.H.; Pronk, J.T.; Daran-Lapujade, P.A.S.

    2011-01-01

    Extremely low specific growth rates (below 0.01 h?1) represent a largely unexplored area of microbial physiology. In this study, anaerobic, glucose-limited retentostats were used to analyse physiological and genome-wide transcriptional responses of Saccharomyces cerevisiae to cultivation at

  1. Loss of lager specific genes and subtelomeric regions define two different Saccharomyces cerevisiae lineages for Saccharomyces pastorianus Group I and II strains.

    Science.gov (United States)

    Monerawela, Chandre; James, Tharappel C; Wolfe, Kenneth H; Bond, Ursula

    2015-03-01

    Lager yeasts, Saccharomyces pastorianus, are interspecies hybrids between S. cerevisiae and S. eubayanus and are classified into Group I and Group II clades. The genome of the Group II strain, Weihenstephan 34/70, contains eight so-called 'lager-specific' genes that are located in subtelomeric regions. We evaluated the origins of these genes through bioinformatic and PCR analyses of Saccharomyces genomes. We determined that four are of cerevisiae origin while four originate from S. eubayanus. The Group I yeasts contain all four S. eubayanus genes but individual strains contain only a subset of the cerevisiae genes. We identified S. cerevisiae strains that contain all four cerevisiae 'lager-specific' genes, and distinct patterns of loss of these genes in other strains. Analysis of the subtelomeric regions uncovered patterns of loss in different S. cerevisiae strains. We identify two classes of S. cerevisiae strains: ale yeasts (Foster O) and stout yeasts with patterns of 'lager-specific' genes and subtelomeric regions identical to Group I and II S. pastorianus yeasts, respectively. These findings lead us to propose that Group I and II S. pastorianus strains originate from separate hybridization events involving different S. cerevisiae lineages. Using the combined bioinformatic and PCR data, we describe a potential classification map for industrial yeasts. © FEMS 2015. All rights reserved. For permissions, please e-mail: journals.permission@oup.com.

  2. Ethanol production from Jerusalem artichoke by strains of Saccharomyces cheresiensis and Saccharomyces beticus

    Energy Technology Data Exchange (ETDEWEB)

    Pourrat, H.; Barthomeuf, C.; Regerat, F.; Carnat, A.P.; Carnat, A.

    1983-03-01

    Ethanol production from Jerusalem artichoke which is the most interesting autochtonous material has been studied. Two selected and acclimatised strains of Saccharomyces: Saccharomyces cheresiensis and Saccharomyces beticus were retained. The fermentation conditions, exactly definited, makes it possible to obtain in 4 days a theoric yield.

  3. Profiling of Escherichia coli Chromosome database.

    Science.gov (United States)

    Yamazaki, Yukiko; Niki, Hironori; Kato, Jun-ichi

    2008-01-01

    The Profiling of Escherichia coli Chromosome (PEC) database (http://www.shigen.nig.ac.jp/ecoli/pec/) is designed to allow E. coli researchers to efficiently access information from functional genomics studies. The database contains two principal types of data: gene essentiality and a large collection of E. coli genetic research resources. The essentiality data are based on data compilation from published single-gene essentiality studies and on cell growth studies of large-deletion mutants. Using the circular and linear viewers for both whole genomes and the minimal genome, users can not only gain an overview of the genome structure but also retrieve information on contigs, gene products, mutants, deletions, and so forth. In particular, genome-wide exhaustive mutants are an essential resource for studying E. coli gene functions. Although the genomic database was constructed independently from the genetic resources database, users may seamlessly access both types of data. In addition to these data, the PEC database also provides a summary of homologous genes of other bacterial genomes and of protein structure information, with a comprehensive interface. The PEC is thus a convenient and useful platform for contemporary E. coli researchers.

  4. MIPS Arabidopsis thaliana Database (MAtDB): an integrated biological knowledge resource for plant genomics

    Science.gov (United States)

    Schoof, Heiko; Ernst, Rebecca; Nazarov, Vladimir; Pfeifer, Lukas; Mewes, Hans-Werner; Mayer, Klaus F. X.

    2004-01-01

    Arabidopsis thaliana is the most widely studied model plant. Functional genomics is intensively underway in many laboratories worldwide. Beyond the basic annotation of the primary sequence data, the annotated genetic elements of Arabidopsis must be linked to diverse biological data and higher order information such as metabolic or regulatory pathways. The MIPS Arabidopsis thaliana database MAtDB aims to provide a comprehensive resource for Arabidopsis as a genome model that serves as a primary reference for research in plants and is suitable for transfer of knowledge to other plants, especially crops. The genome sequence as a common backbone serves as a scaffold for the integration of data, while, in a complementary effort, these data are enhanced through the application of state-of-the-art bioinformatics tools. This information is visualized on a genome-wide and a gene-by-gene basis with access both for web users and applications. This report updates the information given in a previous report and provides an outlook on further developments. The MAtDB web interface can be accessed at http://mips.gsf.de/proj/thal/db. PMID:14681437

  5. The Genomes OnLine Database (GOLD) v.4: status of genomic and metagenomic projects and their associated metadata

    Science.gov (United States)

    Pagani, Ioanna; Liolios, Konstantinos; Jansson, Jakob; Chen, I-Min A.; Smirnova, Tatyana; Nosrat, Bahador; Markowitz, Victor M.; Kyrpides, Nikos C.

    2012-01-01

    The Genomes OnLine Database (GOLD, http://www.genomesonline.org/) is a comprehensive resource for centralized monitoring of genome and metagenome projects worldwide. Both complete and ongoing projects, along with their associated metadata, can be accessed in GOLD through precomputed tables and a search page. As of September 2011, GOLD, now on version 4.0, contains information for 11 472 sequencing projects, of which 2907 have been completed and their sequence data has been deposited in a public repository. Out of these complete projects, 1918 are finished and 989 are permanent drafts. Moreover, GOLD contains information for 340 metagenome studies associated with 1927 metagenome samples. GOLD continues to expand, moving toward the goal of providing the most comprehensive repository of metadata information related to the projects and their organisms/environments in accordance with the Minimum Information about any (x) Sequence specification and beyond. PMID:22135293

  6. Database Description - RED | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available ase Description General information of database Database name RED Alternative name Rice Expression Database...enome Research Unit Shoshi Kikuchi E-mail : Database classification Plant databases - Rice Database classifi...cation Microarray, Gene Expression Organism Taxonomy Name: Oryza sativa Taxonomy ID: 4530 Database descripti... Article title: Rice Expression Database: the gateway to rice functional genomics...nt Science (2002) Dec 7 (12):563-564 External Links: Original website information Database maintenance site

  7. Download - PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available List Contact us PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods ...t_db_link_en.zip (36.3 KB) - 6 Genome analysis methods pgdbj_dna_marker_linkage_map_genome_analysis_methods_... of This Database Site Policy | Contact Us Download - PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods | LSDB Archive ...

  8. Sequence modelling and an extensible data model for genomic database

    Energy Technology Data Exchange (ETDEWEB)

    Li, Peter Wei-Der [California Univ., San Francisco, CA (United States); Univ. of California, Berkeley, CA (United States)

    1992-01-01

    The Human Genome Project (HGP) plans to sequence the human genome by the beginning of the next century. It will generate DNA sequences of more than 10 billion bases and complex marker sequences (maps) of more than 100 million markers. All of these information will be stored in database management systems (DBMSs). However, existing data models do not have the abstraction mechanism for modelling sequences and existing DBMS`s do not have operations for complex sequences. This work addresses the problem of sequence modelling in the context of the HGP and the more general problem of an extensible object data model that can incorporate the sequence model as well as existing and future data constructs and operators. First, we proposed a general sequence model that is application and implementation independent. This model is used to capture the sequence information found in the HGP at the conceptual level. In addition, abstract and biological sequence operators are defined for manipulating the modelled sequences. Second, we combined many features of semantic and object oriented data models into an extensible framework, which we called the ``Extensible Object Model``, to address the need of a modelling framework for incorporating the sequence data model with other types of data constructs and operators. This framework is based on the conceptual separation between constructors and constraints. We then used this modelling framework to integrate the constructs for the conceptual sequence model. The Extensible Object Model is also defined with a graphical representation, which is useful as a tool for database designers. Finally, we defined a query language to support this model and implement the query processor to demonstrate the feasibility of the extensible framework and the usefulness of the conceptual sequence model.

  9. Sequence modelling and an extensible data model for genomic database

    Energy Technology Data Exchange (ETDEWEB)

    Li, Peter Wei-Der (California Univ., San Francisco, CA (United States) Lawrence Berkeley Lab., CA (United States))

    1992-01-01

    The Human Genome Project (HGP) plans to sequence the human genome by the beginning of the next century. It will generate DNA sequences of more than 10 billion bases and complex marker sequences (maps) of more than 100 million markers. All of these information will be stored in database management systems (DBMSs). However, existing data models do not have the abstraction mechanism for modelling sequences and existing DBMS's do not have operations for complex sequences. This work addresses the problem of sequence modelling in the context of the HGP and the more general problem of an extensible object data model that can incorporate the sequence model as well as existing and future data constructs and operators. First, we proposed a general sequence model that is application and implementation independent. This model is used to capture the sequence information found in the HGP at the conceptual level. In addition, abstract and biological sequence operators are defined for manipulating the modelled sequences. Second, we combined many features of semantic and object oriented data models into an extensible framework, which we called the Extensible Object Model'', to address the need of a modelling framework for incorporating the sequence data model with other types of data constructs and operators. This framework is based on the conceptual separation between constructors and constraints. We then used this modelling framework to integrate the constructs for the conceptual sequence model. The Extensible Object Model is also defined with a graphical representation, which is useful as a tool for database designers. Finally, we defined a query language to support this model and implement the query processor to demonstrate the feasibility of the extensible framework and the usefulness of the conceptual sequence model.

  10. Saccharomyces boulardii probiotic-associated fungemia: questioning the safety of this preventive probiotic's use.

    Science.gov (United States)

    Martin, Isabella W; Tonner, Rita; Trivedi, Julie; Miller, Heather; Lee, Richard; Liang, Xinglun; Rotello, Leo; Isenbergh, Elena; Anderson, Jennifer; Perl, Trish; Zhang, Sean X

    2017-03-01

    We report a case of fungemia in an immunocompetent patient after administration of probiotic containing Saccharomyces boulardii. We demonstrated the strain relatedness of the yeast from the probiotic capsule and the yeast causing fungal infection using genomic and proteomic typing methods. Our study questions the safety of this preventative biotherapy. Copyright © 2016 Elsevier Inc. All rights reserved.

  11. Mitochondrial genomic dysfunction causes dephosphorylation of Sch9 in the yeast Saccharomyces cerevisiae.

    Science.gov (United States)

    Kawai, Shigeyuki; Urban, Jörg; Piccolis, Manuele; Panchaud, Nicolas; De Virgilio, Claudio; Loewith, Robbie

    2011-10-01

    TORC1-dependent phosphorylation of Saccharomyces cerevisiae Sch9 was dramatically reduced upon exposure to a protonophore or in respiration-incompetent ρ(0) cells but not in respiration-incompetent pet mutants, providing important insight into the molecular mechanisms governing interorganellar signaling in general and retrograde signaling in particular.

  12. Database Description - RMG | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available ase Description General information of database Database name RMG Alternative name ...raki 305-8602, Japan National Institute of Agrobiological Sciences E-mail : Database... classification Nucleotide Sequence Databases Organism Taxonomy Name: Oryza sativa Japonica Group Taxonomy ID: 39947 Database...rnal: Mol Genet Genomics (2002) 268: 434–445 External Links: Original website information Database...available URL of Web services - Need for user registration Not available About This Database Database Descri

  13. Effects of an unusual poison identify a lifespan role for Topoisomerase 2 in Saccharomyces cerevisiae

    OpenAIRE

    Tombline, Gregory; Millen, Jonathan I.; Polevoda, Bogdan; Rapaport, Matan; Baxter, Bonnie; Van Meter, Michael; Gilbertson, Matthew; Madrey, Joe; Piazza, Gary A.; Rasmussen, Lynn; Wennerberg, Krister; White, E. Lucile; Nitiss, John L.; Goldfarb, David S.

    2017-01-01

    A progressive loss of genome maintenance has been implicated as both a cause and consequence of aging. Here we present evidence supporting the hypothesis that an age-associated decay in genome maintenance promotes aging in Saccharomyces cerevisiae (yeast) due to an inability to sense or repair DNA damage by topoisomerase 2 (yTop2). We describe the characterization of LS1, identified in a high throughput screen for small molecules that shorten the replicative lifespan of yeast. LS1 accelerates...

  14. Database Description - KOME | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available base Description General information of database Database name KOME Alternative nam... Sciences Plant Genome Research Unit Shoshi Kikuchi E-mail : Database classification Plant databases - Rice ...Organism Taxonomy Name: Oryza sativa Taxonomy ID: 4530 Database description Information about approximately ...Hayashizaki Y, Kikuchi S. Journal: PLoS One. 2007 Nov 28; 2(11):e1235. External Links: Original website information Database...OS) Rice mutant panel database (Tos17) A Database of Plant Cis-acting Regulatory

  15. Database Description - GETDB | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available abase Description General information of database Database name GETDB Alternative n...ame Gal4 Enhancer Trap Insertion Database DOI 10.18908/lsdba.nbdc00236-000 Creator Creator Name: Shigeo Haya... Chuo-ku, Kobe 650-0047 Tel: +81-78-306-3185 FAX: +81-78-306-3183 E-mail: Database classification Expression... Invertebrate genome database Organism Taxonomy Name: Drosophila melanogaster Taxonomy ID: 7227 Database des...riginal website information Database maintenance site Drosophila Genetic Resource

  16. SinEx DB: a database for single exon coding sequences in mammalian genomes.

    Science.gov (United States)

    Jorquera, Roddy; Ortiz, Rodrigo; Ossandon, F; Cárdenas, Juan Pablo; Sepúlveda, Rene; González, Carolina; Holmes, David S

    2016-01-01

    Eukaryotic genes are typically interrupted by intragenic, noncoding sequences termed introns. However, some genes lack introns in their coding sequence (CDS) and are generally known as 'single exon genes' (SEGs). In this work, a SEG is defined as a nuclear, protein-coding gene that lacks introns in its CDS. Whereas, many public databases of Eukaryotic multi-exon genes are available, there are only two specialized databases for SEGs. The present work addresses the need for a more extensive and diverse database by creating SinEx DB, a publicly available, searchable database of predicted SEGs from 10 completely sequenced mammalian genomes including human. SinEx DB houses the DNA and protein sequence information of these SEGs and includes their functional predictions (KOG) and the relative distribution of these functions within species. The information is stored in a relational database built with My SQL Server 5.1.33 and the complete dataset of SEG sequences and their functional predictions are available for downloading. SinEx DB can be interrogated by: (i) a browsable phylogenetic schema, (ii) carrying out BLAST searches to the in-house SinEx DB of SEGs and (iii) via an advanced search mode in which the database can be searched by key words and any combination of searches by species and predicted functions. SinEx DB provides a rich source of information for advancing our understanding of the evolution and function of SEGs.Database URL: www.sinex.cl. © The Author(s) 2016. Published by Oxford University Press.

  17. PineElm_SSRdb: a microsatellite marker database identified from genomic, chloroplast, mitochondrial and EST sequences of pineapple (Ananas comosus (L.) Merrill).

    Science.gov (United States)

    Chaudhary, Sakshi; Mishra, Bharat Kumar; Vivek, Thiruvettai; Magadum, Santoshkumar; Yasin, Jeshima Khan

    2016-01-01

    Simple Sequence Repeats or microsatellites are resourceful molecular genetic markers. There are only few reports of SSR identification and development in pineapple. Complete genome sequence of pineapple available in the public domain can be used to develop numerous novel SSRs. Therefore, an attempt was made to identify SSRs from genomic, chloroplast, mitochondrial and EST sequences of pineapple which will help in deciphering genetic makeup of its germplasm resources. A total of 359511 SSRs were identified in pineapple (356385 from genome sequence, 45 from chloroplast sequence, 249 in mitochondrial sequence and 2832 from EST sequences). The list of EST-SSR markers and their details are available in the database. PineElm_SSRdb is an open source database available for non-commercial academic purpose at http://app.bioelm.com/ with a mapping tool which can develop circular maps of selected marker set. This database will be of immense use to breeders, researchers and graduates working on Ananas spp. and to others working on cross-species transferability of markers, investigating diversity, mapping and DNA fingerprinting.

  18. Recent advances in the genome-wide study of DNA replication origins in yeast

    Directory of Open Access Journals (Sweden)

    Chong ePeng

    2015-02-01

    Full Text Available DNA replication, one of the central events in the cell cycle, is the basis of biological inheritance. In order to be duplicated, a DNA double helix must be opened at defined sites, which are called DNA replication origins (ORIs. Unlike in bacteria, where replication initiates from a single replication origin, multiple origins are utilized in the eukaryotic genome. Among them, the ORIs in budding yeast Saccharomyces cerevisiae and the fission yeast Schizosaccharomyces pombe have been best characterized. In recent years, advances in DNA microarray and next-generation sequencing technologies have increased the number of yeast species involved in ORIs research dramatically. The ORIs in some nonconventional yeast species such as Kluyveromyces lactis and Pichia pastoris have also been genome-widely identified. Relevant databases of replication origins in yeast were constructed, then the comparative genomic analysis can be carried out. Here, we review several experimental approaches that have been used to map replication origins in yeast and some of the available web resources related to yeast ORIs. We also discuss the sequence characteristics and chromosome structures of ORIs in the four yeast species, which can be utilized to improve the replication origins prediction.

  19. Recent advances in the genome-wide study of DNA replication origins in yeast

    Science.gov (United States)

    Peng, Chong; Luo, Hao; Zhang, Xi; Gao, Feng

    2015-01-01

    DNA replication, one of the central events in the cell cycle, is the basis of biological inheritance. In order to be duplicated, a DNA double helix must be opened at defined sites, which are called DNA replication origins (ORIs). Unlike in bacteria, where replication initiates from a single replication origin, multiple origins are utilized in the eukaryotic genomes. Among them, the ORIs in budding yeast Saccharomyces cerevisiae and the fission yeast Schizosaccharomyces pombe have been best characterized. In recent years, advances in DNA microarray and next-generation sequencing technologies have increased the number of yeast species involved in ORIs research dramatically. The ORIs in some non-conventional yeast species such as Kluyveromyces lactis and Pichia pastoris have also been genome-widely identified. Relevant databases of replication origins in yeast were constructed, then the comparative genomic analysis can be carried out. Here, we review several experimental approaches that have been used to map replication origins in yeast and some of the available web resources related to yeast ORIs. We also discuss the sequence characteristics and chromosome structures of ORIs in the four yeast species, which can be utilized to improve yeast replication origins prediction. PMID:25745419

  20. The MAR databases: development and implementation of databases specific for marine metagenomics.

    Science.gov (United States)

    Klemetsen, Terje; Raknes, Inge A; Fu, Juan; Agafonov, Alexander; Balasundaram, Sudhagar V; Tartari, Giacomo; Robertsen, Espen; Willassen, Nils P

    2018-01-04

    We introduce the marine databases; MarRef, MarDB and MarCat (https://mmp.sfb.uit.no/databases/), which are publicly available resources that promote marine research and innovation. These data resources, which have been implemented in the Marine Metagenomics Portal (MMP) (https://mmp.sfb.uit.no/), are collections of richly annotated and manually curated contextual (metadata) and sequence databases representing three tiers of accuracy. While MarRef is a database for completely sequenced marine prokaryotic genomes, which represent a marine prokaryote reference genome database, MarDB includes all incomplete sequenced prokaryotic genomes regardless level of completeness. The last database, MarCat, represents a gene (protein) catalog of uncultivable (and cultivable) marine genes and proteins derived from marine metagenomics samples. The first versions of MarRef and MarDB contain 612 and 3726 records, respectively. Each record is built up of 106 metadata fields including attributes for sampling, sequencing, assembly and annotation in addition to the organism and taxonomic information. Currently, MarCat contains 1227 records with 55 metadata fields. Ontologies and controlled vocabularies are used in the contextual databases to enhance consistency. The user-friendly web interface lets the visitors browse, filter and search in the contextual databases and perform BLAST searches against the corresponding sequence databases. All contextual and sequence databases are freely accessible and downloadable from https://s1.sfb.uit.no/public/mar/. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  1. Nuclear mitochondrial DNA activates replication in Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Laurent Chatre

    Full Text Available The nuclear genome of eukaryotes is colonized by DNA fragments of mitochondrial origin, called NUMTs. These insertions have been associated with a variety of germ-line diseases in humans. The significance of this uptake of potentially dangerous sequences into the nuclear genome is unclear. Here we provide functional evidence that sequences of mitochondrial origin promote nuclear DNA replication in Saccharomyces cerevisiae. We show that NUMTs are rich in key autonomously replicating sequence (ARS consensus motifs, whose mutation results in the reduction or loss of DNA replication activity. Furthermore, 2D-gel analysis of the mrc1 mutant exposed to hydroxyurea shows that several NUMTs function as late chromosomal origins. We also show that NUMTs located close to or within ARS provide key sequence elements for replication. Thus NUMTs can act as independent origins, when inserted in an appropriate genomic context or affect the efficiency of pre-existing origins. These findings show that migratory mitochondrial DNAs can impact on the replication of the nuclear region they are inserted in.

  2. Legume and Lotus japonicus Databases

    DEFF Research Database (Denmark)

    Hirakawa, Hideki; Mun, Terry; Sato, Shusei

    2014-01-01

    Since the genome sequence of Lotus japonicus, a model plant of family Fabaceae, was determined in 2008 (Sato et al. 2008), the genomes of other members of the Fabaceae family, soybean (Glycine max) (Schmutz et al. 2010) and Medicago truncatula (Young et al. 2011), have been sequenced. In this sec....... In this section, we introduce representative, publicly accessible online resources related to plant materials, integrated databases containing legume genome information, and databases for genome sequence and derived marker information of legume species including L. japonicus...

  3. Genome cluster database. A sequence family analysis platform for Arabidopsis and rice.

    Science.gov (United States)

    Horan, Kevin; Lauricha, Josh; Bailey-Serres, Julia; Raikhel, Natasha; Girke, Thomas

    2005-05-01

    The genome-wide protein sequences from Arabidopsis (Arabidopsis thaliana) and rice (Oryza sativa) spp. japonica were clustered into families using sequence similarity and domain-based clustering. The two fundamentally different methods resulted in separate cluster sets with complementary properties to compensate the limitations for accurate family analysis. Functional names for the identified families were assigned with an efficient computational approach that uses the description of the most common molecular function gene ontology node within each cluster. Subsequently, multiple alignments and phylogenetic trees were calculated for the assembled families. All clustering results and their underlying sequences were organized in the Web-accessible Genome Cluster Database (http://bioinfo.ucr.edu/projects/GCD) with rich interactive and user-friendly sequence family mining tools to facilitate the analysis of any given family of interest for the plant science community. An automated clustering pipeline ensures current information for future updates in the annotations of the two genomes and clustering improvements. The analysis allowed the first systematic identification of family and singlet proteins present in both organisms as well as those restricted to one of them. In addition, the established Web resources for mining these data provide a road map for future studies of the composition and structure of protein families between the two species.

  4. Genome-wide analysis reveals the vacuolar pH-stat of Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Christopher L Brett

    Full Text Available Protons, the smallest and most ubiquitous of ions, are central to physiological processes. Transmembrane proton gradients drive ATP synthesis, metabolite transport, receptor recycling and vesicle trafficking, while compartmental pH controls enzyme function. Despite this fundamental importance, the mechanisms underlying pH homeostasis are not entirely accounted for in any organelle or organism. We undertook a genome-wide survey of vacuole pH (pH(v in 4,606 single-gene deletion mutants of Saccharomyces cerevisiae under control, acid and alkali stress conditions to reveal the vacuolar pH-stat. Median pH(v (5.27±0.13 was resistant to acid stress (5.28±0.14 but shifted significantly in response to alkali stress (5.83±0.13. Of 107 mutants that displayed aberrant pH(v under more than one external pH condition, functional categories of transporters, membrane biogenesis and trafficking machinery were significantly enriched. Phospholipid flippases, encoded by the family of P4-type ATPases, emerged as pH regulators, as did the yeast ortholog of Niemann Pick Type C protein, implicated in sterol trafficking. An independent genetic screen revealed that correction of pH(v dysregulation in a neo1(ts mutant restored viability whereas cholesterol accumulation in human NPC1(-/- fibroblasts diminished upon treatment with a proton ionophore. Furthermore, while it is established that lumenal pH affects trafficking, this study revealed a reciprocal link with many mutants defective in anterograde pathways being hyperacidic and retrograde pathway mutants with alkaline vacuoles. In these and other examples, pH perturbations emerge as a hitherto unrecognized phenotype that may contribute to the cellular basis of disease and offer potential therapeutic intervention through pH modulation.

  5. KoVariome: Korean National Standard Reference Variome database of whole genomes with comprehensive SNV, indel, CNV, and SV analyses.

    Science.gov (United States)

    Kim, Jungeun; Weber, Jessica A; Jho, Sungwoong; Jang, Jinho; Jun, JeHoon; Cho, Yun Sung; Kim, Hak-Min; Kim, Hyunho; Kim, Yumi; Chung, OkSung; Kim, Chang Geun; Lee, HyeJin; Kim, Byung Chul; Han, Kyudong; Koh, InSong; Chae, Kyun Shik; Lee, Semin; Edwards, Jeremy S; Bhak, Jong

    2018-04-04

    High-coverage whole-genome sequencing data of a single ethnicity can provide a useful catalogue of population-specific genetic variations, and provides a critical resource that can be used to more accurately identify pathogenic genetic variants. We report a comprehensive analysis of the Korean population, and present the Korean National Standard Reference Variome (KoVariome). As a part of the Korean Personal Genome Project (KPGP), we constructed the KoVariome database using 5.5 terabases of whole genome sequence data from 50 healthy Korean individuals in order to characterize the benign ethnicity-relevant genetic variation present in the Korean population. In total, KoVariome includes 12.7M single-nucleotide variants (SNVs), 1.7M short insertions and deletions (indels), 4K structural variations (SVs), and 3.6K copy number variations (CNVs). Among them, 2.4M (19%) SNVs and 0.4M (24%) indels were identified as novel. We also discovered selective enrichment of 3.8M SNVs and 0.5M indels in Korean individuals, which were used to filter out 1,271 coding-SNVs not originally removed from the 1,000 Genomes Project when prioritizing disease-causing variants. KoVariome health records were used to identify novel disease-causing variants in the Korean population, demonstrating the value of high-quality ethnic variation databases for the accurate interpretation of individual genomes and the precise characterization of genetic variations.

  6. SNPpy--database management for SNP data from genome wide association studies.

    Directory of Open Access Journals (Sweden)

    Faheem Mitha

    Full Text Available BACKGROUND: We describe SNPpy, a hybrid script database system using the Python SQLAlchemy library coupled with the PostgreSQL database to manage genotype data from Genome-Wide Association Studies (GWAS. This system makes it possible to merge study data with HapMap data and merge across studies for meta-analyses, including data filtering based on the values of phenotype and Single-Nucleotide Polymorphism (SNP data. SNPpy and its dependencies are open source software. RESULTS: The current version of SNPpy offers utility functions to import genotype and annotation data from two commercial platforms. We use these to import data from two GWAS studies and the HapMap Project. We then export these individual datasets to standard data format files that can be imported into statistical software for downstream analyses. CONCLUSIONS: By leveraging the power of relational databases, SNPpy offers integrated management and manipulation of genotype and phenotype data from GWAS studies. The analysis of these studies requires merging across GWAS datasets as well as patient and marker selection. To this end, SNPpy enables the user to filter the data and output the results as standardized GWAS file formats. It does low level and flexible data validation, including validation of patient data. SNPpy is a practical and extensible solution for investigators who seek to deploy central management of their GWAS data.

  7. The roles of the Saccharomyces cerevisiae RecQ helicase SGS1 in meiotic genome surveillance.

    Directory of Open Access Journals (Sweden)

    Amit Dipak Amin

    2010-11-01

    Full Text Available The Saccharomyces cerevisiae RecQ helicase Sgs1 is essential for mitotic and meiotic genome stability. The stage at which Sgs1 acts during meiosis is subject to debate. Cytological experiments showed that a deletion of SGS1 leads to an increase in synapsis initiation complexes and axial associations leading to the proposal that it has an early role in unwinding surplus strand invasion events. Physical studies of recombination intermediates implicate it in the dissolution of double Holliday junctions between sister chromatids.In this work, we observed an increase in meiotic recombination between diverged sequences (homeologous recombination and an increase in unequal sister chromatid events when SGS1 is deleted. The first of these observations is most consistent with an early role of Sgs1 in unwinding inappropriate strand invasion events while the second is consistent with unwinding or dissolution of recombination intermediates in an Mlh1- and Top3-dependent manner. We also provide data that suggest that Sgs1 is involved in the rejection of 'second strand capture' when sequence divergence is present. Finally, we have identified a novel class of tetrads where non-sister spores (pairs of spores where each contains a centromere marker from a different parent are inviable. We propose a model for this unusual pattern of viability based on the inability of sgs1 mutants to untangle intertwined chromosomes. Our data suggest that this role of Sgs1 is not dependent on its interaction with Top3. We propose that in the absence of SGS1 chromosomes may sometimes remain entangled at the end of pre-meiotic replication. This, combined with reciprocal crossing over, could lead to physical destruction of the recombined and entangled chromosomes. We hypothesise that Sgs1, acting in concert with the topoisomerase Top2, resolves these structures.This work provides evidence that Sgs1 interacts with various partner proteins to maintain genome stability throughout

  8. Toward an interactive article: integrating journals and biological databases

    Directory of Open Access Journals (Sweden)

    Marygold Steven J

    2011-05-01

    Full Text Available Abstract Background Journal articles and databases are two major modes of communication in the biological sciences, and thus integrating these critical resources is of urgent importance to increase the pace of discovery. Projects focused on bridging the gap between journals and databases have been on the rise over the last five years and have resulted in the development of automated tools that can recognize entities within a document and link those entities to a relevant database. Unfortunately, automated tools cannot resolve ambiguities that arise from one term being used to signify entities that are quite distinct from one another. Instead, resolving these ambiguities requires some manual oversight. Finding the right balance between the speed and portability of automation and the accuracy and flexibility of manual effort is a crucial goal to making text markup a successful venture. Results We have established a journal article mark-up pipeline that links GENETICS journal articles and the model organism database (MOD WormBase. This pipeline uses a lexicon built with entities from the database as a first step. The entity markup pipeline results in links from over nine classes of objects including genes, proteins, alleles, phenotypes and anatomical terms. New entities and ambiguities are discovered and resolved by a database curator through a manual quality control (QC step, along with help from authors via a web form that is provided to them by the journal. New entities discovered through this pipeline are immediately sent to an appropriate curator at the database. Ambiguous entities that do not automatically resolve to one link are resolved by hand ensuring an accurate link. This pipeline has been extended to other databases, namely Saccharomyces Genome Database (SGD and FlyBase, and has been implemented in marking up a paper with links to multiple databases. Conclusions Our semi-automated pipeline hyperlinks articles published in GENETICS to

  9. Functional co-operation between the nuclei of Saccharomyces cerevisiae and mitochondria from other yeast species

    DEFF Research Database (Denmark)

    Spirek, M.; Horvath, A.; Piskur, Jure

    2000-01-01

    We elaborated a simple method that allows the transfer of mitochondria from collection yeasts to Saccharomyces cerevisiae. Protoplasts prepared from different yeasts were fused to the protoplasts of the ade2-1, ura3-52, kar1-1, rho (0) strain of S. cerevisiae and were selected for respiring cybrids....... italicus, S, oviformis, S. capensis and S. chevalieri) exhibited complete compatibility with S. cerevisiae nuclei. The closely related S. douglasii mitochondrial genome could also partially restore respiration-deficiency in rho (0) S. cerevisiae, whereas mitochondrial genomes from phylogenetically less...

  10. Freedom and Responsibility in Synthetic Genomics: The Synthetic Yeast Project

    OpenAIRE

    Sliva, Anna; Yang, Huanming; Boeke, Jef D.; Mathews, Debra J. H.

    2015-01-01

    First introduced in 2011, the Synthetic Yeast Genome (Sc2.0) Project is a large international synthetic genomics project that will culminate in the first eukaryotic cell (Saccharomyces cerevisiae) with a fully synthetic genome. With collaborators from across the globe and from a range of institutions spanning from do-it-yourself biology (DIYbio) to commercial enterprises, it is important that all scientists working on this project are cognizant of the ethical and policy issues associated with...

  11. Analysis of disease-associated objects at the Rat Genome Database

    Science.gov (United States)

    Wang, Shur-Jen; Laulederkind, Stanley J. F.; Hayman, G. T.; Smith, Jennifer R.; Petri, Victoria; Lowry, Timothy F.; Nigam, Rajni; Dwinell, Melinda R.; Worthey, Elizabeth A.; Munzenmaier, Diane H.; Shimoyama, Mary; Jacob, Howard J.

    2013-01-01

    The Rat Genome Database (RGD) is the premier resource for genetic, genomic and phenotype data for the laboratory rat, Rattus norvegicus. In addition to organizing biological data from rats, the RGD team focuses on manual curation of gene–disease associations for rat, human and mouse. In this work, we have analyzed disease-associated strains, quantitative trait loci (QTL) and genes from rats. These disease objects form the basis for seven disease portals. Among disease portals, the cardiovascular disease and obesity/metabolic syndrome portals have the highest number of rat strains and QTL. These two portals share 398 rat QTL, and these shared QTL are highly concentrated on rat chromosomes 1 and 2. For disease-associated genes, we performed gene ontology (GO) enrichment analysis across portals using RatMine enrichment widgets. Fifteen GO terms, five from each GO aspect, were selected to profile enrichment patterns of each portal. Of the selected biological process (BP) terms, ‘regulation of programmed cell death’ was the top enriched term across all disease portals except in the obesity/metabolic syndrome portal where ‘lipid metabolic process’ was the most enriched term. ‘Cytosol’ and ‘nucleus’ were common cellular component (CC) annotations for disease genes, but only the cancer portal genes were highly enriched with ‘nucleus’ annotations. Similar enrichment patterns were observed in a parallel analysis using the DAVID functional annotation tool. The relationship between the preselected 15 GO terms and disease terms was examined reciprocally by retrieving rat genes annotated with these preselected terms. The individual GO term–annotated gene list showed enrichment in physiologically related diseases. For example, the ‘regulation of blood pressure’ genes were enriched with cardiovascular disease annotations, and the ‘lipid metabolic process’ genes with obesity annotations. Furthermore, we were able to enhance enrichment of neurological

  12. Laboratory evolution of a biotin-requiring Saccharomyces cerevisiae strain for full biotin prototrophy and identification of causal mutations

    NARCIS (Netherlands)

    Bracher, J.M.; de Hulster, A.F.; van den Broek, M.A.; Daran, J.G.; van Maris, A.J.A.; Pronk, J.T.

    2017-01-01

    Biotin prototrophy is a rare, incompletely understood, and industrially relevant characteristic of Saccharomyces cerevisiae strains. The genome of the haploid laboratory strain CEN.PK113-7D contains a full complement of biotin biosynthesis genes, but its growth in biotin-free synthetic medium is

  13. Influence of genetic background of engineered xylose-fermenting industrial Saccharomyces cerevisiae strains for ethanol production from lignocellulosic hydrolysates

    Science.gov (United States)

    An industrial ethanol-producing Saccharomyces cerevisiae strain with genes needed for xylose-fermentation integrated into its genome was used to obtain haploids and diploid isogenic strains. The isogenic strains were more effective in metabolizing xylose than their parental strain (p < 0.05) and abl...

  14. Multiplexed CRISPR/Cas9 Genome Editing and Gene Regulation Using Csy4 in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Ferreira, Raphael; Skrekas, Christos; Nielsen, Jens

    2018-01-01

    Clustered regularly interspaced short palindromic repeats (CRISPR) technology has greatly accelerated the field of strain engineering. However, insufficient efforts have been made toward developing robust multiplexing tools in Saccharomyces cerevisiae. Here, we exploit the RNA processing capacity...

  15. Genomics Portals: integrative web-platform for mining genomics data.

    Science.gov (United States)

    Shinde, Kaustubh; Phatak, Mukta; Johannes, Freudenberg M; Chen, Jing; Li, Qian; Vineet, Joshi K; Hu, Zhen; Ghosh, Krishnendu; Meller, Jaroslaw; Medvedovic, Mario

    2010-01-13

    A large amount of experimental data generated by modern high-throughput technologies is available through various public repositories. Our knowledge about molecular interaction networks, functional biological pathways and transcriptional regulatory modules is rapidly expanding, and is being organized in lists of functionally related genes. Jointly, these two sources of information hold a tremendous potential for gaining new insights into functioning of living systems. Genomics Portals platform integrates access to an extensive knowledge base and a large database of human, mouse, and rat genomics data with basic analytical visualization tools. It provides the context for analyzing and interpreting new experimental data and the tool for effective mining of a large number of publicly available genomics datasets stored in the back-end databases. The uniqueness of this platform lies in the volume and the diversity of genomics data that can be accessed and analyzed (gene expression, ChIP-chip, ChIP-seq, epigenomics, computationally predicted binding sites, etc), and the integration with an extensive knowledge base that can be used in such analysis. The integrated access to primary genomics data, functional knowledge and analytical tools makes Genomics Portals platform a unique tool for interpreting results of new genomics experiments and for mining the vast amount of data stored in the Genomics Portals backend databases. Genomics Portals can be accessed and used freely at http://GenomicsPortals.org.

  16. Marker list - PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available List Contact us PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods ...Database Site Policy | Contact Us Marker list - PGDBj Registered plant list, Marker list, QTL list, Plant DB link & Genome analysis methods | LSDB Archive ...

  17. Measuring the Levels of Ribonucleotides Embedded in Genomic DNA.

    Science.gov (United States)

    Meroni, Alice; Nava, Giulia M; Sertic, Sarah; Plevani, Paolo; Muzi-Falconi, Marco; Lazzaro, Federico

    2018-01-01

    Ribonucleotides (rNTPs) are incorporated into genomic DNA at a relatively high frequency during replication. They have beneficial effects but, if not removed from the chromosomes, increase genomic instability. Here, we describe a fast method to easily estimate the amounts of embedded ribonucleotides into the genome. The protocol described is performed in Saccharomyces cerevisiae and allows us to quantify altered levels of rNMPs due to different mutations in the replicative polymerase ε. However, this protocol can be easily applied to cells derived from any organism.

  18. MIPS plant genome information resources.

    Science.gov (United States)

    Spannagl, Manuel; Haberer, Georg; Ernst, Rebecca; Schoof, Heiko; Mayer, Klaus F X

    2007-01-01

    The Munich Institute for Protein Sequences (MIPS) has been involved in maintaining plant genome databases since the Arabidopsis thaliana genome project. Genome databases and analysis resources have focused on individual genomes and aim to provide flexible and maintainable data sets for model plant genomes as a backbone against which experimental data, for example from high-throughput functional genomics, can be organized and evaluated. In addition, model genomes also form a scaffold for comparative genomics, and much can be learned from genome-wide evolutionary studies.

  19. Saccharomyces kudriavzevii and Saccharomyces uvarum differ from Saccharomyces cerevisiae during the production of aroma-active higher alcohols and acetate esters using their amino acidic precursors.

    Science.gov (United States)

    Stribny, Jiri; Gamero, Amparo; Pérez-Torrado, Roberto; Querol, Amparo

    2015-07-16

    Higher alcohols and acetate esters are important flavour and aroma components in the food industry. In alcoholic beverages these compounds are produced by yeast during fermentation. Although Saccharomyces cerevisiae is one of the most extensively used species, other species of the Saccharomyces genus have become common in fermentation processes. This study analyses and compares the production of higher alcohols and acetate esters from their amino acidic precursors in three Saccharomyces species: Saccharomyces kudriavzevii, Saccharomyces uvarum and S. cerevisiae. The global volatile compound analysis revealed that S. kudriavzevii produced large amounts of higher alcohols, whereas S. uvarum excelled in the production of acetate esters. Particularly from phenylalanine, S. uvarum produced the largest amounts of 2-phenylethyl acetate, while S. kudriavzevii obtained the greatest 2-phenylethanol formation from this precursor. The present data indicate differences in the amino acid metabolism and subsequent production of flavour-active higher alcohols and acetate esters among the closely related Saccharomyces species. This knowledge will prove useful for developing new enhanced processes in fragrance, flavour, and food industries. Copyright © 2015. Published by Elsevier B.V.

  20. YMDB 2.0: a significantly expanded version of the yeast metabolome database.

    Science.gov (United States)

    Ramirez-Gaona, Miguel; Marcu, Ana; Pon, Allison; Guo, An Chi; Sajed, Tanvir; Wishart, Noah A; Karu, Naama; Djoumbou Feunang, Yannick; Arndt, David; Wishart, David S

    2017-01-04

    YMDB or the Yeast Metabolome Database (http://www.ymdb.ca/) is a comprehensive database containing extensive information on the genome and metabolome of Saccharomyces cerevisiae Initially released in 2012, the YMDB has gone through a significant expansion and a number of improvements over the past 4 years. This manuscript describes the most recent version of YMDB (YMDB 2.0). More specifically, it provides an updated description of the database that was previously described in the 2012 NAR Database Issue and it details many of the additions and improvements made to the YMDB over that time. Some of the most important changes include a 7-fold increase in the number of compounds in the database (from 2007 to 16 042), a 430-fold increase in the number of metabolic and signaling pathway diagrams (from 66 to 28 734), a 16-fold increase in the number of compounds linked to pathways (from 742 to 12 733), a 17-fold increase in the numbers of compounds with nuclear magnetic resonance or MS spectra (from 783 to 13 173) and an increase in both the number of data fields and the number of links to external databases. In addition to these database expansions, a number of improvements to YMDB's web interface and its data visualization tools have been made. These additions and improvements should greatly improve the ease, the speed and the quantity of data that can be extracted, searched or viewed within YMDB. Overall, we believe these improvements should not only improve the understanding of the metabolism of S. cerevisiae, but also allow more in-depth exploration of its extensive metabolic networks, signaling pathways and biochemistry. © The Author(s) 2016. Published by Oxford University Press on behalf of Nucleic Acids Research.

  1. Comparing genomes: databases and computational tools for comparative analysis of prokaryotic genomes - DOI: 10.3395/reciis.v1i2.Sup.105en

    Directory of Open Access Journals (Sweden)

    Marcos Catanho

    2007-12-01

    Full Text Available Since the 1990's, the complete genetic code of more than 600 living organisms has been deciphered, such as bacteria, yeasts, protozoan parasites, invertebrates and vertebrates, including Homo sapiens, and plants. More than 2,000 other genome projects representing medical, commercial, environmental and industrial interests, or comprising model organisms, important for the development of the scientific research, are currently in progress. The achievement of complete genome sequences of numerous species combined with the tremendous progress in computation that occurred in the last few decades allowed the use of new holistic approaches in the study of genome structure, organization and evolution, as well as in the field of gene prediction and functional classification. Numerous public or proprietary databases and computational tools have been created attempting to optimize the access to this information through the web. In this review, we present the main resources available through the web for comparative analysis of prokaryotic genomes. We concentrated on the group of mycobacteria that contains important human and animal pathogens. The birth of Bioinformatics and Computational Biology and the contributions of these disciplines to the scientific development of this field are also discussed.

  2. MIPS Arabidopsis thaliana Database (MAtDB): an integrated biological knowledge resource based on the first complete plant genome

    Science.gov (United States)

    Schoof, Heiko; Zaccaria, Paolo; Gundlach, Heidrun; Lemcke, Kai; Rudd, Stephen; Kolesov, Grigory; Arnold, Roland; Mewes, H. W.; Mayer, Klaus F. X.

    2002-01-01

    Arabidopsis thaliana is the first plant for which the complete genome has been sequenced and published. Annotation of complex eukaryotic genomes requires more than the assignment of genetic elements to the sequence. Besides completing the list of genes, we need to discover their cellular roles, their regulation and their interactions in order to understand the workings of the whole plant. The MIPS Arabidopsis thaliana Database (MAtDB; http://mips.gsf.de/proj/thal/db) started out as a repository for genome sequence data in the European Scientists Sequencing Arabidopsis (ESSA) project and the Arabidopsis Genome Initiative. Our aim is to transform MAtDB into an integrated biological knowledge resource by integrating diverse data, tools, query and visualization capabilities and by creating a comprehensive resource for Arabidopsis as a reference model for other species, including crop plants. PMID:11752263

  3. Genome Sequence Databases (Overview): Sequencing and Assembly

    Energy Technology Data Exchange (ETDEWEB)

    Lapidus, Alla L.

    2009-01-01

    From the date its role in heredity was discovered, DNA has been generating interest among scientists from different fields of knowledge: physicists have studied the three dimensional structure of the DNA molecule, biologists tried to decode the secrets of life hidden within these long molecules, and technologists invent and improve methods of DNA analysis. The analysis of the nucleotide sequence of DNA occupies a special place among the methods developed. Thanks to the variety of sequencing technologies available, the process of decoding the sequence of genomic DNA (or whole genome sequencing) has become robust and inexpensive. Meanwhile the assembly of whole genome sequences remains a challenging task. In addition to the need to assemble millions of DNA fragments of different length (from 35 bp (Solexa) to 800 bp (Sanger)), great interest in analysis of microbial communities (metagenomes) of different complexities raises new problems and pushes some new requirements for sequence assembly tools to the forefront. The genome assembly process can be divided into two steps: draft assembly and assembly improvement (finishing). Despite the fact that automatically performed assembly (or draft assembly) is capable of covering up to 98% of the genome, in most cases, it still contains incorrectly assembled reads. The error rate of the consensus sequence produced at this stage is about 1/2000 bp. A finished genome represents the genome assembly of much higher accuracy (with no gaps or incorrectly assembled areas) and quality ({approx}1 error/10,000 bp), validated through a number of computer and laboratory experiments.

  4. The Importance of Biological Databases in Biological Discovery.

    Science.gov (United States)

    Baxevanis, Andreas D; Bateman, Alex

    2015-06-19

    Biological databases play a central role in bioinformatics. They offer scientists the opportunity to access a wide variety of biologically relevant data, including the genomic sequences of an increasingly broad range of organisms. This unit provides a brief overview of major sequence databases and portals, such as GenBank, the UCSC Genome Browser, and Ensembl. Model organism databases, including WormBase, The Arabidopsis Information Resource (TAIR), and those made available through the Mouse Genome Informatics (MGI) resource, are also covered. Non-sequence-centric databases, such as Online Mendelian Inheritance in Man (OMIM), the Protein Data Bank (PDB), MetaCyc, and the Kyoto Encyclopedia of Genes and Genomes (KEGG), are also discussed. Copyright © 2015 John Wiley & Sons, Inc.

  5. Genomics Portals: integrative web-platform for mining genomics data

    Directory of Open Access Journals (Sweden)

    Ghosh Krishnendu

    2010-01-01

    Full Text Available Abstract Background A large amount of experimental data generated by modern high-throughput technologies is available through various public repositories. Our knowledge about molecular interaction networks, functional biological pathways and transcriptional regulatory modules is rapidly expanding, and is being organized in lists of functionally related genes. Jointly, these two sources of information hold a tremendous potential for gaining new insights into functioning of living systems. Results Genomics Portals platform integrates access to an extensive knowledge base and a large database of human, mouse, and rat genomics data with basic analytical visualization tools. It provides the context for analyzing and interpreting new experimental data and the tool for effective mining of a large number of publicly available genomics datasets stored in the back-end databases. The uniqueness of this platform lies in the volume and the diversity of genomics data that can be accessed and analyzed (gene expression, ChIP-chip, ChIP-seq, epigenomics, computationally predicted binding sites, etc, and the integration with an extensive knowledge base that can be used in such analysis. Conclusion The integrated access to primary genomics data, functional knowledge and analytical tools makes Genomics Portals platform a unique tool for interpreting results of new genomics experiments and for mining the vast amount of data stored in the Genomics Portals backend databases. Genomics Portals can be accessed and used freely at http://GenomicsPortals.org.

  6. The Mouse Genome Database (MGD): facilitating mouse as a model for human biology and disease.

    Science.gov (United States)

    Eppig, Janan T; Blake, Judith A; Bult, Carol J; Kadin, James A; Richardson, Joel E

    2015-01-01

    The Mouse Genome Database (MGD, http://www.informatics.jax.org) serves the international biomedical research community as the central resource for integrated genomic, genetic and biological data on the laboratory mouse. To facilitate use of mouse as a model in translational studies, MGD maintains a core of high-quality curated data and integrates experimentally and computationally generated data sets. MGD maintains a unified catalog of genes and genome features, including functional RNAs, QTL and phenotypic loci. MGD curates and provides functional and phenotype annotations for mouse genes using the Gene Ontology and Mammalian Phenotype Ontology. MGD integrates phenotype data and associates mouse genotypes to human diseases, providing critical mouse-human relationships and access to repositories holding mouse models. MGD is the authoritative source of nomenclature for genes, genome features, alleles and strains following guidelines of the International Committee on Standardized Genetic Nomenclature for Mice. A new addition to MGD, the Human-Mouse: Disease Connection, allows users to explore gene-phenotype-disease relationships between human and mouse. MGD has also updated search paradigms for phenotypic allele attributes, incorporated incidental mutation data, added a module for display and exploration of genes and microRNA interactions and adopted the JBrowse genome browser. MGD resources are freely available to the scientific community. © The Author(s) 2014. Published by Oxford University Press on behalf of Nucleic Acids Research.

  7. dBBQs: dataBase of Bacterial Quality scores.

    Science.gov (United States)

    Wanchai, Visanu; Patumcharoenpol, Preecha; Nookaew, Intawat; Ussery, David

    2017-12-28

    It is well-known that genome sequencing technologies are becoming significantly cheaper and faster. As a result of this, the exponential growth in sequencing data in public databases allows us to explore ever growing large collections of genome sequences. However, it is less known that the majority of available sequenced genome sequences in public databases are not complete, drafts of varying qualities. We have calculated quality scores for around 100,000 bacterial genomes from all major genome repositories and put them in a fast and easy-to-use database. Prokaryotic genomic data from all sources were collected and combined to make a non-redundant set of bacterial genomes. The genome quality score for each was calculated by four different measurements: assembly quality, number of rRNA and tRNA genes, and the occurrence of conserved functional domains. The dataBase of Bacterial Quality scores (dBBQs) was designed to store and retrieve quality scores. It offers fast searching and download features which the result can be used for further analysis. In addition, the search results are shown in interactive JavaScript chart framework using DC.js. The analysis of quality scores across major public genome databases find that around 68% of the genomes are of acceptable quality for many uses. dBBQs (available at http://arc-gem.uams.edu/dbbqs ) provides genome quality scores for all available prokaryotic genome sequences with a user-friendly Web-interface. These scores can be used as cut-offs to get a high-quality set of genomes for testing bioinformatics tools or improving the analysis. Moreover, all data of the four measurements that were combined to make the quality score for each genome, which can potentially be used for further analysis. dBBQs will be updated regularly and is freely use for non-commercial purpose.

  8. Improved xylose and arabinose utilization by an industrial recombinant Saccharomyces cerevisiae strain using evolutionary engineering

    DEFF Research Database (Denmark)

    Sanchez, R.G.; Karhumaa, Kaisa; Fonseca, C.

    2010-01-01

    Background: Cost-effective fermentation of lignocellulosic hydrolysate to ethanol by Saccharomyces cerevisiae requires efficient mixed sugar utilization. Notably, the rate and yield of xylose and arabinose co-fermentation to ethanol must be enhanced. Results: Evolutionary engineering was used...... to improve the simultaneous conversion of xylose and arabinose to ethanol in a recombinant industrial Saccharomyces cerevisiae strain carrying the heterologous genes for xylose and arabinose utilization pathways integrated in the genome. The evolved strain TMB3130 displayed an increased consumption rate...... of our knowledge, this is the first report that characterizes the molecular mechanisms for improved mixed-pentose utilization obtained by evolutionary engineering of a recombinant S. cerevisiae strain. Increased transport of pentoses and increased activities of xylose converting enzymes contributed...

  9. Multiplexed precision genome editing with trackable genomic barcodes in yeast.

    Science.gov (United States)

    Roy, Kevin R; Smith, Justin D; Vonesch, Sibylle C; Lin, Gen; Tu, Chelsea Szu; Lederer, Alex R; Chu, Angela; Suresh, Sundari; Nguyen, Michelle; Horecka, Joe; Tripathi, Ashutosh; Burnett, Wallace T; Morgan, Maddison A; Schulz, Julia; Orsley, Kevin M; Wei, Wu; Aiyar, Raeka S; Davis, Ronald W; Bankaitis, Vytas A; Haber, James E; Salit, Marc L; St Onge, Robert P; Steinmetz, Lars M

    2018-07-01

    Our understanding of how genotype controls phenotype is limited by the scale at which we can precisely alter the genome and assess the phenotypic consequences of each perturbation. Here we describe a CRISPR-Cas9-based method for multiplexed accurate genome editing with short, trackable, integrated cellular barcodes (MAGESTIC) in Saccharomyces cerevisiae. MAGESTIC uses array-synthesized guide-donor oligos for plasmid-based high-throughput editing and features genomic barcode integration to prevent plasmid barcode loss and to enable robust phenotyping. We demonstrate that editing efficiency can be increased more than fivefold by recruiting donor DNA to the site of breaks using the LexA-Fkh1p fusion protein. We performed saturation editing of the essential gene SEC14 and identified amino acids critical for chemical inhibition of lipid signaling. We also constructed thousands of natural genetic variants, characterized guide mismatch tolerance at the genome scale, and ascertained that cryptic Pol III termination elements substantially reduce guide efficacy. MAGESTIC will be broadly useful to uncover the genetic basis of phenotypes in yeast.

  10. ANISEED 2017: extending the integrated ascidian database to the exploration and evolutionary comparison of genome-scale datasets.

    Science.gov (United States)

    Brozovic, Matija; Dantec, Christelle; Dardaillon, Justine; Dauga, Delphine; Faure, Emmanuel; Gineste, Mathieu; Louis, Alexandra; Naville, Magali; Nitta, Kazuhiro R; Piette, Jacques; Reeves, Wendy; Scornavacca, Céline; Simion, Paul; Vincentelli, Renaud; Bellec, Maelle; Aicha, Sameh Ben; Fagotto, Marie; Guéroult-Bellone, Marion; Haeussler, Maximilian; Jacox, Edwin; Lowe, Elijah K; Mendez, Mickael; Roberge, Alexis; Stolfi, Alberto; Yokomori, Rui; Brown, C Titus; Cambillau, Christian; Christiaen, Lionel; Delsuc, Frédéric; Douzery, Emmanuel; Dumollard, Rémi; Kusakabe, Takehiro; Nakai, Kenta; Nishida, Hiroki; Satou, Yutaka; Swalla, Billie; Veeman, Michael; Volff, Jean-Nicolas; Lemaire, Patrick

    2018-01-04

    ANISEED (www.aniseed.cnrs.fr) is the main model organism database for tunicates, the sister-group of vertebrates. This release gives access to annotated genomes, gene expression patterns, and anatomical descriptions for nine ascidian species. It provides increased integration with external molecular and taxonomy databases, better support for epigenomics datasets, in particular RNA-seq, ChIP-seq and SELEX-seq, and features novel interactive interfaces for existing and novel datatypes. In particular, the cross-species navigation and comparison is enhanced through a novel taxonomy section describing each represented species and through the implementation of interactive phylogenetic gene trees for 60% of tunicate genes. The gene expression section displays the results of RNA-seq experiments for the three major model species of solitary ascidians. Gene expression is controlled by the binding of transcription factors to cis-regulatory sequences. A high-resolution description of the DNA-binding specificity for 131 Ciona robusta (formerly C. intestinalis type A) transcription factors by SELEX-seq is provided and used to map candidate binding sites across the Ciona robusta and Phallusia mammillata genomes. Finally, use of a WashU Epigenome browser enhances genome navigation, while a Genomicus server was set up to explore microsynteny relationships within tunicates and with vertebrates, Amphioxus, echinoderms and hemichordates. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  11. KONAGAbase: a genomic and transcriptomic database for the diamondback moth, Plutella xylostella.

    Science.gov (United States)

    Jouraku, Akiya; Yamamoto, Kimiko; Kuwazaki, Seigo; Urio, Masahiro; Suetsugu, Yoshitaka; Narukawa, Junko; Miyamoto, Kazuhisa; Kurita, Kanako; Kanamori, Hiroyuki; Katayose, Yuichi; Matsumoto, Takashi; Noda, Hiroaki

    2013-07-09

    The diamondback moth (DBM), Plutella xylostella, is one of the most harmful insect pests for crucifer crops worldwide. DBM has rapidly evolved high resistance to most conventional insecticides such as pyrethroids, organophosphates, fipronil, spinosad, Bacillus thuringiensis, and diamides. Therefore, it is important to develop genomic and transcriptomic DBM resources for analysis of genes related to insecticide resistance, both to clarify the mechanism of resistance of DBM and to facilitate the development of insecticides with a novel mode of action for more effective and environmentally less harmful insecticide rotation. To contribute to this goal, we developed KONAGAbase, a genomic and transcriptomic database for DBM (KONAGA is the Japanese word for DBM). KONAGAbase provides (1) transcriptomic sequences of 37,340 ESTs/mRNAs and 147,370 RNA-seq contigs which were clustered and assembled into 84,570 unigenes (30,695 contigs, 50,548 pseudo singletons, and 3,327 singletons); and (2) genomic sequences of 88,530 WGS contigs with 246,244 degenerate contigs and 106,455 singletons from which 6,310 de novo identified repeat sequences and 34,890 predicted gene-coding sequences were extracted. The unigenes and predicted gene-coding sequences were clustered and 32,800 representative sequences were extracted as a comprehensive putative gene set. These sequences were annotated with BLAST descriptions, Gene Ontology (GO) terms, and Pfam descriptions, respectively. KONAGAbase contains rich graphical user interface (GUI)-based web interfaces for easy and efficient searching, browsing, and downloading sequences and annotation data. Five useful search interfaces consisting of BLAST search, keyword search, BLAST result-based search, GO tree-based search, and genome browser are provided. KONAGAbase is publicly available from our website (http://dbm.dna.affrc.go.jp/px/) through standard web browsers. KONAGAbase provides DBM comprehensive transcriptomic and draft genomic sequences with

  12. The Eukaryotic Pathogen Databases: a functional genomic resource integrating data from human and veterinary parasites.

    Science.gov (United States)

    Harb, Omar S; Roos, David S

    2015-01-01

    Over the past 20 years, advances in high-throughput biological techniques and the availability of computational resources including fast Internet access have resulted in an explosion of large genome-scale data sets "big data." While such data are readily available for download and personal use and analysis from a variety of repositories, often such analysis requires access to seldom-available computational skills. As a result a number of databases have emerged to provide scientists with online tools enabling the interrogation of data without the need for sophisticated computational skills beyond basic knowledge of Internet browser utility. This chapter focuses on the Eukaryotic Pathogen Databases (EuPathDB: http://eupathdb.org) Bioinformatic Resource Center (BRC) and illustrates some of the available tools and methods.

  13. Stabilization process in Saccharomyces intra and interspecific hybrids in fermentative conditions.

    Science.gov (United States)

    Pérez-Través, Laura; Lopes, Christian A; Barrio, Eladio; Querol, Amparo

    2014-12-01

    We evaluated the genetic stabilization of artificial intra- (Saccharomyces cerevisiae) and interspecific (S. cerevisiae × S. kudriavzevii) hybrids under wine fermentative conditions. Large-scale transitions in genome size and genome reorganizations were observed during this process. Interspecific hybrids seem to need fewer generations to reach genetic stability than intraspecific hybrids. The largest number of molecular patterns recovered among the derived clones was observed for intraspecific hybrids, particularly for those obtained by rare-mating. Molecular marker analyses revealed that unstable clones could change during the industrial process to obtain active dry yeast. When no changes in molecular markers and ploidy were observed after this process, no changes in genetic composition were confirmed by comparative genome hybridization, considering the clone as a stable hybrid. According to our results, under these conditions, fermentation steps 3 and 5 (30-50 generations) would suffice to obtain genetically stable interspecific and intraspecific hybrids, respectively. Copyright© by the Spanish Society for Microbiology and Institute for Catalan Studies.

  14. BISQUE: locus- and variant-specific conversion of genomic, transcriptomic and proteomic database identifiers.

    Science.gov (United States)

    Meyer, Michael J; Geske, Philip; Yu, Haiyuan

    2016-05-15

    Biological sequence databases are integral to efforts to characterize and understand biological molecules and share biological data. However, when analyzing these data, scientists are often left holding disparate biological currency-molecular identifiers from different databases. For downstream applications that require converting the identifiers themselves, there are many resources available, but analyzing associated loci and variants can be cumbersome if data is not given in a form amenable to particular analyses. Here we present BISQUE, a web server and customizable command-line tool for converting molecular identifiers and their contained loci and variants between different database conventions. BISQUE uses a graph traversal algorithm to generalize the conversion process for residues in the human genome, genes, transcripts and proteins, allowing for conversion across classes of molecules and in all directions through an intuitive web interface and a URL-based web service. BISQUE is freely available via the web using any major web browser (http://bisque.yulab.org/). Source code is available in a public GitHub repository (https://github.com/hyulab/BISQUE). haiyuan.yu@cornell.edu Supplementary data are available at Bioinformatics online. © The Author 2016. Published by Oxford University Press. All rights reserved. For Permissions, please e-mail: journals.permissions@oup.com.

  15. Studies of Saccharomyces cerevisiae and Non-Saccharomyces Yeasts during Alcoholic Fermentation

    DEFF Research Database (Denmark)

    Kemsawasd, Varongsiri

    The early death of non-Saccharomyces yeasts during mixed culture spontaneous wine fermentation has traditionally been attributed to the lower capacity of these yeast species to withstand high levels of ethanol, low pH, and other media properties that are a part of progressing fermentation. However......, other yeast-yeast interactions, such as cell-cell contact mediated growth arrest and/or toxininduced death may also be a significant factor in the relative fragility of these non-Saccharomyces yeasts in mixed culture fermentation. In the present work we evaluate the combined roles of cell-cell contact...... and/or antimicrobial peptides on the early death of Lachancea thermotolerans during mixed culture fermentations with Saccharomyces cerevisiae. Using a specially designed double compartment fermentation system, we established that both cell-to-cell contact and antimicrobial peptides contribute...

  16. A reference model systesm of industrial yeasts Saccharomyces cerevisiae is needed for development of the next-generation biocatalyst toward advanced biofuels production

    Science.gov (United States)

    Diploid industrial yeast Saccharomyces cerevisiae has demonstrated distinct characteristics that differ from haploid laboratory model strains. However, as a workhorse for a broad range of fermentation-based industrial applications, it was poorly characterized at the genome level. Observations on the...

  17. Improvement of Xylose Fermentation Ability under Heat and Acid Co-Stress in Saccharomyces cerevisiae Using Genome Shuffling Technique

    Directory of Open Access Journals (Sweden)

    Kentaro Inokuma

    2017-12-01

    Full Text Available Xylose-assimilating yeasts with tolerance to both fermentation inhibitors (such as weak organic acids and high temperature are required for cost-effective simultaneous saccharification and cofermentation (SSCF of lignocellulosic materials. Here, we demonstrate the construction of a novel xylose-utilizing Saccharomyces cerevisiae strain with improved fermentation ability under heat and acid co-stress using the drug resistance marker-aided genome shuffling technique. The mutagenized genome pools derived from xylose-utilizing diploid yeasts with thermotolerance or acid tolerance were shuffled by sporulation and mating. The shuffled strains were then subjected to screening under co-stress conditions of heat and acids, and the hybrid strain Hyb-8 was isolated. The hybrid strain displayed enhanced xylose fermentation ability in comparison to both parental strains under co-stress conditions of heat and acids. Hyb-8 consumed 33.1 ± 0.6 g/L xylose and produced 11.1 ± 0.4 g/L ethanol after 72 h of fermentation at 38°C with 20 mM acetic acid and 15 mM formic acid. We also performed transcriptomic analysis of the hybrid strain and its parental strains to screen for key genes for multiple stress tolerances. We found that 13 genes, including 5 associated with cellular transition metal ion homeostasis, were significantly upregulated in Hyb-8 compared to levels in both parental strains under co-stress conditions. The hybrid strain Hyb-8 has strong potential for cost-effective SSCF of lignocellulosic materials. Moreover, the transcriptome data gathered in this study will be useful for understanding the mechanisms of multiple tolerance to high temperature and acids in yeast and facilitate the development of robust yeast strains for SSCF.

  18. ProFITS of maize: a database of protein families involved in the transduction of signalling in the maize genome

    Directory of Open Access Journals (Sweden)

    Zhang Zhenhai

    2010-10-01

    Full Text Available Abstract Background Maize (Zea mays ssp. mays L. is an important model for plant basic and applied research. In 2009, the B73 maize genome sequencing made a great step forward, using clone by clone strategy; however, functional annotation and gene classification of the maize genome are still limited. Thus, a well-annotated datasets and informative database will be important for further research discoveries. Signal transduction is a fundamental biological process in living cells, and many protein families participate in this process in sensing, amplifying and responding to various extracellular or internal stimuli. Therefore, it is a good starting point to integrate information on the maize functional genes involved in signal transduction. Results Here we introduce a comprehensive database 'ProFITS' (Protein Families Involved in the Transduction of Signalling, which endeavours to identify and classify protein kinases/phosphatases, transcription factors and ubiquitin-proteasome-system related genes in the B73 maize genome. Users can explore gene models, corresponding transcripts and FLcDNAs using the three abovementioned protein hierarchical categories, and visualize them using an AJAX-based genome browser (JBrowse or Generic Genome Browser (GBrowse. Functional annotations such as GO annotation, protein signatures, protein best-hits in the Arabidopsis and rice genome are provided. In addition, pre-calculated transcription factor binding sites of each gene are generated and mutant information is incorporated into ProFITS. In short, ProFITS provides a user-friendly web interface for studies in signal transduction process in maize. Conclusion ProFITS, which utilizes both the B73 maize genome and full length cDNA (FLcDNA datasets, provides users a comprehensive platform of maize annotation with specific focus on the categorization of families involved in the signal transduction process. ProFITS is designed as a user-friendly web interface and it is

  19. iTRAQ-based proteome profiling of Saccharomyces cerevisiae and cryotolerant species Saccharomyces uvarum and Saccharomyces kudriavzevii during low-temperature wine fermentation.

    Science.gov (United States)

    García-Ríos, Estéfani; Querol, Amparo; Guillamón, José Manuel

    2016-09-02

    Temperature is one of the most important parameters to affect the duration and rate of alcoholic fermentation and final wine quality. Some species of the Saccharomyces genus have shown better adaptation at low temperature than Saccharomyces cerevisiae, which was the case of cryotolerant yeasts Saccharomyces uvarum and Saccharomyces kudriavzevii. In an attempt to detect inter-specific metabolic differences, we characterized the proteomic landscape of these cryotolerant species grown at 12°C and 28°C, which we compared with the proteome of S. cerevisiae (poorly adapted at low temperature). Our results showed that the main differences among the proteomic profiling of the three Saccharomyces strains grown at 12°C and 28°C lay in translation, glycolysis and amino acid metabolism. Our data corroborate previous transcriptomic results, which suggest that S. kudriavzevii is better adapted to grow at low temperature as a result of enhanced more efficient translation. Fitter amino acid biosynthetic pathways can also be mechanisms that better explain biomass yield in cryotolerant strains. Yet even at low temperature, S. cerevisiae is the most fermentative competitive species. A higher concentration of glycolytic and alcoholic fermentation enzymes in the S. cerevisiae strain might explain such greater fermentation activity. Temperature is one of the main relevant environmental variables that microorganisms have to cope with and it is also a key factor in some industrial processes that involve microorganisms. However, we are still far from understanding the molecular and physiological mechanisms of adaptation at low temperatures. The results obtained in this study provided a global atlas of the proteome changes triggered by temperature in three different species of the genus Saccharomyces with different degree of cryotolerance. These results would facilitate a better understanding of mechanisms for how yeast could adapt at the low temperature of growth. Copyright © 2016

  20. ODG: Omics database generator - a tool for generating, querying, and analyzing multi-omics comparative databases to facilitate biological understanding.

    Science.gov (United States)

    Guhlin, Joseph; Silverstein, Kevin A T; Zhou, Peng; Tiffin, Peter; Young, Nevin D

    2017-08-10

    Rapid generation of omics data in recent years have resulted in vast amounts of disconnected datasets without systemic integration and knowledge building, while individual groups have made customized, annotated datasets available on the web with few ways to link them to in-lab datasets. With so many research groups generating their own data, the ability to relate it to the larger genomic and comparative genomic context is becoming increasingly crucial to make full use of the data. The Omics Database Generator (ODG) allows users to create customized databases that utilize published genomics data integrated with experimental data which can be queried using a flexible graph database. When provided with omics and experimental data, ODG will create a comparative, multi-dimensional graph database. ODG can import definitions and annotations from other sources such as InterProScan, the Gene Ontology, ENZYME, UniPathway, and others. This annotation data can be especially useful for studying new or understudied species for which transcripts have only been predicted, and rapidly give additional layers of annotation to predicted genes. In better studied species, ODG can perform syntenic annotation translations or rapidly identify characteristics of a set of genes or nucleotide locations, such as hits from an association study. ODG provides a web-based user-interface for configuring the data import and for querying the database. Queries can also be run from the command-line and the database can be queried directly through programming language hooks available for most languages. ODG supports most common genomic formats as well as generic, easy to use tab-separated value format for user-provided annotations. ODG is a user-friendly database generation and query tool that adapts to the supplied data to produce a comparative genomic database or multi-layered annotation database. ODG provides rapid comparative genomic annotation and is therefore particularly useful for non-model or

  1. Identification and characterization of insect-specific proteins by genome data analysis

    DEFF Research Database (Denmark)

    Zhang, Guojie; Wang, Hongsheng; Shi, Junjie

    2007-01-01

    melanogaster, Anopheles gambiae, Bombyx mori, Tribolium castaneum, and Apis mellifera were compared to the complete genomes of three non-insect eukaryotes (opisthokonts) Homo sapiens, Caenorhabditis elegans and Saccharomyces cerevisiae. This operation yielded 154 groups of orthologous proteins in Drosophila...

  2. OryzaGenome: Genome Diversity Database of Wild Oryza Species

    KAUST Repository

    Ohyanagi, Hajime; Ebata, Toshinobu; Huang, Xuehui; Gong, Hao; Fujita, Masahiro; Mochizuki, Takako; Toyoda, Atsushi; Fujiyama, Asao; Kaminuma, Eli; Nakamura, Yasukazu; Feng, Qi; Wang, Zi Xuan; Han, Bin; Kurata, Nori

    2015-01-01

    . Portable VCF (variant call format) file or tabdelimited file download is also available. Following these SNP (single nucleotide polymorphism) data, reference pseudomolecules/ scaffolds/contigs and genome-wide variation information for almost all

  3. ATG18 and FAB1 are involved in dehydration stress tolerance in Saccharomyces cerevisiae.

    Science.gov (United States)

    López-Martínez, Gema; Margalef-Català, Mar; Salinas, Francisco; Liti, Gianni; Cordero-Otero, Ricardo

    2015-01-01

    Recently, different dehydration-based technologies have been evaluated for the purpose of cell and tissue preservation. Although some early results have been promising, they have not satisfied the requirements for large-scale applications. The long experience of using quantitative trait loci (QTLs) with the yeast Saccharomyces cerevisiae has proven to be a good model organism for studying the link between complex phenotypes and DNA variations. Here, we use QTL analysis as a tool for identifying the specific yeast traits involved in dehydration stress tolerance. Three hybrids obtained from stable haploids and sequenced in the Saccharomyces Genome Resequencing Project showed intermediate dehydration tolerance in most cases. The dehydration resistance trait of 96 segregants from each hybrid was quantified. A smooth, continuous distribution of the anhydrobiosis tolerance trait was found, suggesting that this trait is determined by multiple QTLs. Therefore, we carried out a QTL analysis to identify the determinants of this dehydration tolerance trait at the genomic level. Among the genes identified after reciprocal hemizygosity assays, RSM22, ATG18 and DBR1 had not been referenced in previous studies. We report new phenotypes for these genes using a previously validated test. Finally, our data illustrates the power of this approach in the investigation of the complex cell dehydration phenotype.

  4. ATG18 and FAB1 are involved in dehydration stress tolerance in Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Gema López-Martínez

    Full Text Available Recently, different dehydration-based technologies have been evaluated for the purpose of cell and tissue preservation. Although some early results have been promising, they have not satisfied the requirements for large-scale applications. The long experience of using quantitative trait loci (QTLs with the yeast Saccharomyces cerevisiae has proven to be a good model organism for studying the link between complex phenotypes and DNA variations. Here, we use QTL analysis as a tool for identifying the specific yeast traits involved in dehydration stress tolerance. Three hybrids obtained from stable haploids and sequenced in the Saccharomyces Genome Resequencing Project showed intermediate dehydration tolerance in most cases. The dehydration resistance trait of 96 segregants from each hybrid was quantified. A smooth, continuous distribution of the anhydrobiosis tolerance trait was found, suggesting that this trait is determined by multiple QTLs. Therefore, we carried out a QTL analysis to identify the determinants of this dehydration tolerance trait at the genomic level. Among the genes identified after reciprocal hemizygosity assays, RSM22, ATG18 and DBR1 had not been referenced in previous studies. We report new phenotypes for these genes using a previously validated test. Finally, our data illustrates the power of this approach in the investigation of the complex cell dehydration phenotype.

  5. LC-MS/MS-based proteome profiling in Daphnia pulex and Daphnia longicephala: the Daphnia pulex genome database as a key for high throughput proteomics in Daphnia

    Directory of Open Access Journals (Sweden)

    Mayr Tobias

    2009-04-01

    Full Text Available Abstract Background Daphniids, commonly known as waterfleas, serve as important model systems for ecology, evolution and the environmental sciences. The sequencing and annotation of the Daphnia pulex genome both open future avenues of research on this model organism. As proteomics is not only essential to our understanding of cell function, and is also a powerful validation tool for predicted genes in genome annotation projects, a first proteomic dataset is presented in this article. Results A comprehensive set of 701,274 peptide tandem-mass-spectra, derived from Daphnia pulex, was generated, which lead to the identification of 531 proteins. To measure the impact of the Daphnia pulex filtered models database for mass spectrometry based Daphnia protein identification, this result was compared with results obtained with the Swiss-Prot and the Drosophila melanogaster database. To further validate the utility of the Daphnia pulex database for research on other Daphnia species, additional 407,778 peptide tandem-mass-spectra, obtained from Daphnia longicephala, were generated and evaluated, leading to the identification of 317 proteins. Conclusion Peptides identified in our approach provide the first experimental evidence for the translation of a broad variety of predicted coding regions within the Daphnia genome. Furthermore it could be demonstrated that identification of Daphnia longicephala proteins using the Daphnia pulex protein database is feasible but shows a slightly reduced identification rate. Data provided in this article clearly demonstrates that the Daphnia genome database is the key for mass spectrometry based high throughput proteomics in Daphnia.

  6. Enzymatic activities produced by mixed Saccharomyces and non-Saccharomyces cultures: relationship with wine volatile composition.

    Science.gov (United States)

    Maturano, Yolanda Paola; Assof, Mariela; Fabani, María Paula; Nally, María Cristina; Jofré, Viviana; Rodríguez Assaf, Leticia Anahí; Toro, María Eugenia; Castellanos de Figueroa, Lucía Inés; Vazquez, Fabio

    2015-11-01

    During certain wine fermentation processes, yeasts, and mainly non-Saccharomyces strains, produce and secrete enzymes such as β-glucosidases, proteases, pectinases, xylanases and amylases. The effects of enzyme activity on the aromatic quality of wines during grape juice fermentation, using different co-inoculation strategies of non-Saccharomyces and Saccharomyces cerevisiae yeasts, were assessed in the current study. Three strains with appropriate enological performance and high enzymatic activities, BSc562 (S. cerevisiae), BDv566 (Debaryomyces vanrijiae) and BCs403 (Candida sake), were assayed in pure and mixed Saccharomyces/non-Saccharomyces cultures. β-Glucosidase, pectinase, protease, xylanase and amylase activities were quantified during fermentations. The aromatic profile of pure and mixed cultures was determined at the end of each fermentation. In mixed cultures, non-Saccharomyces species were detected until day 4-5 of the fermentation process, and highest populations were observed in MSD2 (10% S. cerevisiae/90% D. vanrijiae) and MSC1 (1% S. cerevisiae/99% C. sake). According to correlation and multivariate analysis, MSD2 presented the highest concentrations of terpenes and higher alcohols which were associated with pectinase, amylase and xylanase activities. On the other hand, MSC1 high levels of β-glucosidase, proteolytic and xylanolytic activities were correlated to esters and fatty acids. Our study contributes to a better understanding of the effect of enzymatic activities by yeasts on compound transformations that occur during wine fermentation.

  7. Detecting non-orthology in the COGs database and other approaches grouping orthologs using genome-specific best hits.

    Science.gov (United States)

    Dessimoz, Christophe; Boeckmann, Brigitte; Roth, Alexander C J; Gonnet, Gaston H

    2006-01-01

    Correct orthology assignment is a critical prerequisite of numerous comparative genomics procedures, such as function prediction, construction of phylogenetic species trees and genome rearrangement analysis. We present an algorithm for the detection of non-orthologs that arise by mistake in current orthology classification methods based on genome-specific best hits, such as the COGs database. The algorithm works with pairwise distance estimates, rather than computationally expensive and error-prone tree-building methods. The accuracy of the algorithm is evaluated through verification of the distribution of predicted cases, case-by-case phylogenetic analysis and comparisons with predictions from other projects using independent methods. Our results show that a very significant fraction of the COG groups include non-orthologs: using conservative parameters, the algorithm detects non-orthology in a third of all COG groups. Consequently, sequence analysis sensitive to correct orthology assignments will greatly benefit from these findings.

  8. Using FlyBase, a Database of Drosophila Genes and Genomes.

    Science.gov (United States)

    Marygold, Steven J; Crosby, Madeline A; Goodman, Joshua L

    2016-01-01

    For nearly 25 years, FlyBase (flybase.org) has provided a freely available online database of biological information about Drosophila species, focusing on the model organism D. melanogaster. The need for a centralized, integrated view of Drosophila research has never been greater as advances in genomic, proteomic, and high-throughput technologies add to the quantity and diversity of available data and resources.FlyBase has taken several approaches to respond to these changes in the research landscape. Novel report pages have been generated for new reagent types and physical interaction data; Drosophila models of human disease are now represented and showcased in dedicated Human Disease Model Reports; other integrated reports have been established that bring together related genes, datasets, or reagents; Gene Reports have been revised to improve access to new data types and to highlight functional data; links to external sites have been organized and expanded; and new tools have been developed to display and interrogate all these data, including improved batch processing and bulk file availability. In addition, several new community initiatives have served to enhance interactions between researchers and FlyBase, resulting in direct user contributions and improved feedback.This chapter provides an overview of the data content, organization, and available tools within FlyBase, focusing on recent improvements. We hope it serves as a guide for our diverse user base, enabling efficient and effective exploration of the database and thereby accelerating research discoveries.

  9. dBBQs: dataBase of Bacterial Quality scores

    OpenAIRE

    Wanchai, Visanu; Patumcharoenpol, Preecha; Nookaew, Intawat; Ussery, David

    2017-01-01

    Background: It is well-known that genome sequencing technologies are becoming significantly cheaper and faster. As a result of this, the exponential growth in sequencing data in public databases allows us to explore ever growing large collections of genome sequences. However, it is less known that the majority of available sequenced genome sequences in public databases are not complete, drafts of varying qualities. We have calculated quality scores for around 100,000 bacterial genomes from al...

  10. Genome-wide screen for universal individual identification SNPs based on the HapMap and 1000 Genomes databases.

    Science.gov (United States)

    Huang, Erwen; Liu, Changhui; Zheng, Jingjing; Han, Xiaolong; Du, Weian; Huang, Yuanjian; Li, Chengshi; Wang, Xiaoguang; Tong, Dayue; Ou, Xueling; Sun, Hongyu; Zeng, Zhaoshu; Liu, Chao

    2018-04-03

    Differences among SNP panels for individual identification in SNP-selecting and populations led to few common SNPs, compromising their universal applicability. To screen all universal SNPs, we performed a genome-wide SNP mining in multiple populations based on HapMap and 1000Genomes databases. SNPs with high minor allele frequencies (MAF) in 37 populations were selected. With MAF from ≥0.35 to ≥0.43, the number of selected SNPs decreased from 2769 to 0. A total of 117 SNPs with MAF ≥0.39 have no linkage disequilibrium with each other in every population. For 116 of the 117 SNPs, cumulative match probability (CMP) ranged from 2.01 × 10-48 to 1.93 × 10-50 and cumulative exclusion probability (CEP) ranged from 0.9999999996653 to 0.9999999999945. In 134 tested Han samples, 110 of the 117 SNPs remained within high MAF and conformed to Hardy-Weinberg equilibrium, with CMP = 4.70 × 10-47 and CEP = 0.999999999862. By analyzing the same number of autosomal SNPs as in the HID-Ion AmpliSeq Identity Panel, i.e. 90 randomized out of the 110 SNPs, our panel yielded preferable CMP and CEP. Taken together, the 110-SNPs panel is advantageous for forensic test, and this study provided plenty of highly informative SNPs for compiling final universal panels.

  11. MIPS: analysis and annotation of proteins from whole genomes.

    Science.gov (United States)

    Mewes, H W; Amid, C; Arnold, R; Frishman, D; Güldener, U; Mannhaupt, G; Münsterkötter, M; Pagel, P; Strack, N; Stümpflen, V; Warfsmann, J; Ruepp, A

    2004-01-01

    The Munich Information Center for Protein Sequences (MIPS-GSF), Neuherberg, Germany, provides protein sequence-related information based on whole-genome analysis. The main focus of the work is directed toward the systematic organization of sequence-related attributes as gathered by a variety of algorithms, primary information from experimental data together with information compiled from the scientific literature. MIPS maintains automatically generated and manually annotated genome-specific databases, develops systematic classification schemes for the functional annotation of protein sequences and provides tools for the comprehensive analysis of protein sequences. This report updates the information on the yeast genome (CYGD), the Neurospora crassa genome (MNCDB), the database of complete cDNAs (German Human Genome Project, NGFN), the database of mammalian protein-protein interactions (MPPI), the database of FASTA homologies (SIMAP), and the interface for the fast retrieval of protein-associated information (QUIPOS). The Arabidopsis thaliana database, the rice database, the plant EST databases (MATDB, MOsDB, SPUTNIK), as well as the databases for the comprehensive set of genomes (PEDANT genomes) are described elsewhere in the 2003 and 2004 NAR database issues, respectively. All databases described, and the detailed descriptions of our projects can be accessed through the MIPS web server (http://mips.gsf.de).

  12. Thoroughbred Horse Single Nucleotide Polymorphism and Expression Database: HSDB

    Directory of Open Access Journals (Sweden)

    Joon-Ho Lee

    2014-09-01

    Full Text Available Genetics is important for breeding and selection of horses but there is a lack of well-established horse-related browsers or databases. In order to better understand horses, more variants and other integrated information are needed. Thus, we construct a horse genomic variants database including expression and other information. Horse Single Nucleotide Polymorphism and Expression Database (HSDB (http://snugenome2.snu.ac.kr/HSDB provides the number of unexplored genomic variants still remaining to be identified in the horse genome including rare variants by using population genome sequences of eighteen horses and RNA-seq of four horses. The identified single nucleotide polymorphisms (SNPs were confirmed by comparing them with SNP chip data and variants of RNA-seq, which showed a concordance level of 99.02% and 96.6%, respectively. Moreover, the database provides the genomic variants with their corresponding transcriptional profiles from the same individuals to help understand the functional aspects of these variants. The database will contribute to genetic improvement and breeding strategies of Thoroughbreds.

  13. Effects of fermentation by Saccharomyces cerevisiae and ...

    African Journals Online (AJOL)

    yassine

    2013-02-13

    Feb 13, 2013 ... Effect of Saccharomyces cerevisiae fermentation on the ... beetroot, fermentation, Saccharomyces cerevisiae, betalain compounds. ... by Saccharomyces cerevisiae strains (González et al., .... Both red and yellow pigments were influenced during S. .... in beverages such as white wine, grape fruit, and green.

  14. Draft genome sequence of Sclerospora graminicola, the pearl millet downy mildew pathogen

    Directory of Open Access Journals (Sweden)

    Navajeet Chakravartty

    2017-12-01

    Full Text Available Sclerospora graminicola pathogen is the most important biotic production constraints of pearl millet in India, Africa and other parts of the world. We report a de novo whole genome assembly and analysis of pathotype 1, one of the most virulent pathotypes of S. graminicola from India. The whole genome sequencing was performed by sequencing of 7.38 Gb with 73,889,924 paired end reads from the paired-end library, and 1.15 Gb with 3,851,788 reads from the mate pair library generated from Illumina HiSeq 2500 and Illumina MiSeq, respectively. A total 597,293 filtered sub reads with average read length of 6.39 Kb was generated on PACBIO RSII with P6-C4 chemistry. Assembled draft genome sequence of S. graminicola pathotype 1 was 299,901,251 bp in length, N50 of 17,909 bp with a minimum of 1 Kb scaffold size. The GC content was 47.2 % consisting of 26,786 scaffolds with longest scaffold size of 238,843 bp. The overall coverage was 40X. The draft genome sequence was used for gene prediction using AUGUSTUS which resulted in 65,404 genes using Saccharomyces cerevisiae as a model. A total of 52,285 predicted genes found homology using BLASTX against nr database and 38,120 genes were observed with a significant BLASTX match with E-value cutoff of 1e-5 and 40% identity percentage. Out of 38,120 genes annotated a set of 11,873 genes had UniProt entries, while 7,248 were GO terms and 9,686 with KEGG IDs. Of the 7,248 GO terms, 2,724 were associated with the biological processes. The genome information of downy mildew pathogen is available in the NCBI GenBank database. The Sclerospora graminicola whole genome shotgun (WGS project has the project accession MIQA00000000. This version of the project (02 has the accession number MIQA02000000, and consists of sequences MIQA02000001-MIQA02026786, with BioProject ID PRJNA325098 and BioSample ID SAMN05219233. This study may help understand the evolutionary pattern of pathogen and aid elucidation of effector evolution for

  15. Expanded microbial genome coverage and improved protein family annotation in the COG database.

    Science.gov (United States)

    Galperin, Michael Y; Makarova, Kira S; Wolf, Yuri I; Koonin, Eugene V

    2015-01-01

    Microbial genome sequencing projects produce numerous sequences of deduced proteins, only a small fraction of which have been or will ever be studied experimentally. This leaves sequence analysis as the only feasible way to annotate these proteins and assign to them tentative functions. The Clusters of Orthologous Groups of proteins (COGs) database (http://www.ncbi.nlm.nih.gov/COG/), first created in 1997, has been a popular tool for functional annotation. Its success was largely based on (i) its reliance on complete microbial genomes, which allowed reliable assignment of orthologs and paralogs for most genes; (ii) orthology-based approach, which used the function(s) of the characterized member(s) of the protein family (COG) to assign function(s) to the entire set of carefully identified orthologs and describe the range of potential functions when there were more than one; and (iii) careful manual curation of the annotation of the COGs, aimed at detailed prediction of the biological function(s) for each COG while avoiding annotation errors and overprediction. Here we present an update of the COGs, the first since 2003, and a comprehensive revision of the COG annotations and expansion of the genome coverage to include representative complete genomes from all bacterial and archaeal lineages down to the genus level. This re-analysis of the COGs shows that the original COG assignments had an error rate below 0.5% and allows an assessment of the progress in functional genomics in the past 12 years. During this time, functions of many previously uncharacterized COGs have been elucidated and tentative functional assignments of many COGs have been validated, either by targeted experiments or through the use of high-throughput methods. A particularly important development is the assignment of functions to several widespread, conserved proteins many of which turned out to participate in translation, in particular rRNA maturation and tRNA modification. The new version of the

  16. Investigating genotype-phenotype relationships in Saccharomyces cerevisiae metabolic network through stoichiometric modeling

    DEFF Research Database (Denmark)

    Brochado, Ana Rita

    processes. Metabolism is an extensively studied and characterised subcellular system, for which several modeling approaches have been proposed over the last 20 years. Nowadays, stoichiometric modeling of metabolism is done at the genome scale and it has diverse applications, many of them for helping....... This chapter aims at providing the reader with relevant state-of-the-art information concerning Systems Biology, Genome-Scale Metabolic Modeling and Metabolic Engineering. Particular attention is given to the yeast Saccharomyces cerevisiae, the eukaryotic model organism used thought the thesis.......A holistic view of the cell is fundamental for gaining insights into genotype to phenotype relationships. Systems Biology is a discipline within Biology, which uses such holistic approach by focusing on the development and application of tools for studying the structure and dynamics of cellular...

  17. Oxidative DNA damage causes mitochondrial genomic instability in Saccharomyces cerevisiae.

    Science.gov (United States)

    Doudican, Nicole A; Song, Binwei; Shadel, Gerald S; Doetsch, Paul W

    2005-06-01

    Mitochondria contain their own genome, the integrity of which is required for normal cellular energy metabolism. Reactive oxygen species (ROS) produced by normal mitochondrial respiration can damage cellular macromolecules, including mitochondrial DNA (mtDNA), and have been implicated in degenerative diseases, cancer, and aging. We developed strategies to elevate mitochondrial oxidative stress by exposure to antimycin and H(2)O(2) or utilizing mutants lacking mitochondrial superoxide dismutase (sod2Delta). Experiments were conducted with strains compromised in mitochondrial base excision repair (ntg1Delta) and oxidative damage resistance (pif1Delta) in order to delineate the relationship between these pathways. We observed enhanced ROS production, resulting in a direct increase in oxidative mtDNA damage and mutagenesis. Repair-deficient mutants exposed to oxidative stress conditions exhibited profound genomic instability. Elimination of Ntg1p and Pif1p resulted in a synergistic corruption of respiratory competency upon exposure to antimycin and H(2)O(2). Mitochondrial genomic integrity was substantially compromised in ntg1Delta pif1Delta sod2Delta strains, since these cells exhibit a total loss of mtDNA. A stable respiration-defective strain, possessing a normal complement of mtDNA damage resistance pathways, exhibited a complete loss of mtDNA upon exposure to antimycin and H(2)O(2). This loss was preventable by Sod2p overexpression. These results provide direct evidence that oxidative mtDNA damage can be a major contributor to mitochondrial genomic instability and demonstrate cooperation of Ntg1p and Pif1p to resist the introduction of lesions into the mitochondrial genome.

  18. pico-PLAZA, a genome database of microbial photosynthetic eukaryotes.

    Science.gov (United States)

    Vandepoele, Klaas; Van Bel, Michiel; Richard, Guilhem; Van Landeghem, Sofie; Verhelst, Bram; Moreau, Hervé; Van de Peer, Yves; Grimsley, Nigel; Piganeau, Gwenael

    2013-08-01

    With the advent of next generation genome sequencing, the number of sequenced algal genomes and transcriptomes is rapidly growing. Although a few genome portals exist to browse individual genome sequences, exploring complete genome information from multiple species for the analysis of user-defined sequences or gene lists remains a major challenge. pico-PLAZA is a web-based resource (http://bioinformatics.psb.ugent.be/pico-plaza/) for algal genomics that combines different data types with intuitive tools to explore genomic diversity, perform integrative evolutionary sequence analysis and study gene functions. Apart from homologous gene families, multiple sequence alignments, phylogenetic trees, Gene Ontology, InterPro and text-mining functional annotations, different interactive viewers are available to study genome organization using gene collinearity and synteny information. Different search functions, documentation pages, export functions and an extensive glossary are available to guide non-expert scientists. To illustrate the versatility of the platform, different case studies are presented demonstrating how pico-PLAZA can be used to functionally characterize large-scale EST/RNA-Seq data sets and to perform environmental genomics. Functional enrichments analysis of 16 Phaeodactylum tricornutum transcriptome libraries offers a molecular view on diatom adaptation to different environments of ecological relevance. Furthermore, we show how complementary genomic data sources can easily be combined to identify marker genes to study the diversity and distribution of algal species, for example in metagenomes, or to quantify intraspecific diversity from environmental strains. © 2013 John Wiley & Sons Ltd and Society for Applied Microbiology.

  19. Genomic Prediction from Whole Genome Sequence in Livestock: The 1000 Bull Genomes Project

    DEFF Research Database (Denmark)

    Hayes, Benjamin J; MacLeod, Iona M; Daetwyler, Hans D

    Advantages of using whole genome sequence data to predict genomic estimated breeding values (GEBV) include better persistence of accuracy of GEBV across generations and more accurate GEBV across breeds. The 1000 Bull Genomes Project provides a database of whole genome sequenced key ancestor bulls....... In a dairy data set, predictions using BayesRC and imputed sequence data from 1000 Bull Genomes were 2% more accurate than with 800k data. We could demonstrate the method identified causal mutations in some cases. Further improvements will come from more accurate imputation of sequence variant genotypes...

  20. The development of large-scale de-identified biomedical databases in the age of genomics-principles and challenges.

    Science.gov (United States)

    Dankar, Fida K; Ptitsyn, Andrey; Dankar, Samar K

    2018-04-10

    Contemporary biomedical databases include a wide range of information types from various observational and instrumental sources. Among the most important features that unite biomedical databases across the field are high volume of information and high potential to cause damage through data corruption, loss of performance, and loss of patient privacy. Thus, issues of data governance and privacy protection are essential for the construction of data depositories for biomedical research and healthcare. In this paper, we discuss various challenges of data governance in the context of population genome projects. The various challenges along with best practices and current research efforts are discussed through the steps of data collection, storage, sharing, analysis, and knowledge dissemination.

  1. Molecular cloning and expression in Saccharomyces cerevisiae and Neurospora crassa of the invertase gene from Neurospora crassa.

    Science.gov (United States)

    Carú, M; Cifuentes, V; Pincheira, G; Jiménez, A

    1989-10-01

    A plasmid (named pCN2) carrying a 7.6 kb BamHI DNA insert was isolated from a Neurospora crassa genomic library raised in the yeast vector YRp7. Saccharomyces cerevisiae suco and N. crassa inv strains transformed with pNC2 were able to grow on sucrose-based media and expressed invertase activity. Saccharomyces cerevisiae suco (pNC2) expressed a product which immunoreacted with antibody raised against purified invertase from wild type N. crassa, although S. cerevisiae suc+ did not. The cloned DNA hybridized with a 7.6 kb DNA fragment from BamHI-restricted wild type N. crassa DNA. Plasmid pNC2 transformed N. crassa Inv- to Inv+ by integration either near to the endogenous inv locus (40% events) or at other genomic sites (60% events). It appears therefore that the cloned DNA piece encodes the N. crassa invertase enzyme. A 3.8 kb XhoI DNA fragment, derived from pNC2, inserted in YRp7, in both orientation, was able to express invertase activity in yeast, suggesting that it contains an intact invertase gene which is not expressed from a vector promoter.

  2. Replication dynamics of the yeast genome.

    Science.gov (United States)

    Raghuraman, M K; Winzeler, E A; Collingwood, D; Hunt, S; Wodicka, L; Conway, A; Lockhart, D J; Davis, R W; Brewer, B J; Fangman, W L

    2001-10-05

    Oligonucleotide microarrays were used to map the detailed topography of chromosome replication in the budding yeast Saccharomyces cerevisiae. The times of replication of thousands of sites across the genome were determined by hybridizing replicated and unreplicated DNAs, isolated at different times in S phase, to the microarrays. Origin activations take place continuously throughout S phase but with most firings near mid-S phase. Rates of replication fork movement vary greatly from region to region in the genome. The two ends of each of the 16 chromosomes are highly correlated in their times of replication. This microarray approach is readily applicable to other organisms, including humans.

  3. 'Yeast mail': a novel Saccharomyces application (NSA) to encrypt messages.

    Science.gov (United States)

    Rosemeyer, Helmut; Paululat, Achim; Heinisch, Jürgen J

    2014-09-01

    The universal genetic code is used by all life forms to encode biological information. It can also be used to encrypt semantic messages and convey them within organisms without anyone but the sender and recipient knowing, i.e., as a means of steganography. Several theoretical, but comparatively few experimental, approaches have been dedicated to this subject, so far. Here, we describe an experimental system to stably integrate encrypted messages within the yeast genome using a polymerase chain reaction (PCR)-based, one-step homologous recombination system. Thus, DNA sequences encoding alphabetical and/or numerical information will be inherited by yeast propagation and can be sent in the form of dried yeast. Moreover, due to the availability of triple shuttle vectors, Saccharomyces cerevisiae can also be used as an intermediate construction device for transfer of information to either Drosophila or mammalian cells as steganographic containers. Besides its classical use in alcoholic fermentation and its modern use for heterologous gene expression, we here show that baker's yeast can thus be employed in a novel Saccharomyces application (NSA) as a simple steganographic container to hide and convey messages. Copyright © 2014 Verlag Helvetica Chimica Acta AG, Zürich.

  4. Impact of oxygenation on the performance of three non-Saccharomyces yeasts in co-fermentation with Saccharomyces cerevisiae.

    Science.gov (United States)

    Shekhawat, Kirti; Bauer, Florian F; Setati, Mathabatha E

    2017-03-01

    The sequential or co-inoculation of grape must with non-Saccharomyces yeast species and Saccharomyces cerevisiae wine yeast strains has recently become a common practice in winemaking. The procedure intends to enhance unique aroma and flavor profiles of wine. The extent of the impact of non-Saccharomyces strains depends on their ability to produce biomass and to remain metabolically active for a sufficiently long period. However, mixed-culture wine fermentations tend to become rapidly dominated by S. cerevisiae, reducing or eliminating the non-Saccharomyces yeast contribution. For an efficient application of these yeasts, it is therefore essential to understand the environmental factors that modulate the population dynamics of such ecosystems. Several environmental parameters have been shown to influence population dynamics, but their specific effect remains largely uncharacterized. In this study, the population dynamics in co-fermentations of S. cerevisiae and three non-Saccharomyces yeast species: Torulaspora delbrueckii, Lachancea thermotolerans, and Metschnikowia pulcherrima, was investigated as a function of oxygen availability. In all cases, oxygen availability strongly influenced population dynamics, but clear species-dependent differences were observed. Our data show that L. thermotolerans required the least oxygen, followed by T. delbrueckii and M. pulcherrima. Distinct species-specific chemical volatile profiles correlated in all cases with increased persistence of non-Saccharomyces yeasts, in particular increases in some higher alcohols and medium chain fatty acids. The results highlight the role of oxygen in regulating the succession of yeasts during wine fermentations and suggests that more stringent aeration strategies would be necessary to support the persistence of non-Saccharomyces yeasts in real must fermentations.

  5. MAKER2: an annotation pipeline and genome-database management tool for second-generation genome projects.

    Science.gov (United States)

    Holt, Carson; Yandell, Mark

    2011-12-22

    Second-generation sequencing technologies are precipitating major shifts with regards to what kinds of genomes are being sequenced and how they are annotated. While the first generation of genome projects focused on well-studied model organisms, many of today's projects involve exotic organisms whose genomes are largely terra incognita. This complicates their annotation, because unlike first-generation projects, there are no pre-existing 'gold-standard' gene-models with which to train gene-finders. Improvements in genome assembly and the wide availability of mRNA-seq data are also creating opportunities to update and re-annotate previously published genome annotations. Today's genome projects are thus in need of new genome annotation tools that can meet the challenges and opportunities presented by second-generation sequencing technologies. We present MAKER2, a genome annotation and data management tool designed for second-generation genome projects. MAKER2 is a multi-threaded, parallelized application that can process second-generation datasets of virtually any size. We show that MAKER2 can produce accurate annotations for novel genomes where training-data are limited, of low quality or even non-existent. MAKER2 also provides an easy means to use mRNA-seq data to improve annotation quality; and it can use these data to update legacy annotations, significantly improving their quality. We also show that MAKER2 can evaluate the quality of genome annotations, and identify and prioritize problematic annotations for manual review. MAKER2 is the first annotation engine specifically designed for second-generation genome projects. MAKER2 scales to datasets of any size, requires little in the way of training data, and can use mRNA-seq data to improve annotation quality. It can also update and manage legacy genome annotation datasets.

  6. The genome-wide early temporal response of Saccharomyces cerevisiae to oxidative stress induced by cumene hydroperoxide.

    Directory of Open Access Journals (Sweden)

    Wei Sha

    Full Text Available Oxidative stress is a well-known biological process that occurs in all respiring cells and is involved in pathophysiological processes such as aging and apoptosis. Oxidative stress agents include peroxides such as hydrogen peroxide, cumene hydroperoxide, and linoleic acid hydroperoxide, the thiol oxidant diamide, and menadione, a generator of superoxide, amongst others. The present study analyzed the early temporal genome-wide transcriptional response of Saccharomyces cerevisiae to oxidative stress induced by the aromatic peroxide cumene hydroperoxide. The accurate dataset obtained, supported by the use of temporal controls, biological replicates and well controlled growth conditions, provided a detailed picture of the early dynamics of the process. We identified a set of genes previously not implicated in the oxidative stress response, including several transcriptional regulators showing a fast transient response, suggesting a coordinated process in the transcriptional reprogramming. We discuss the role of the glutathione, thioredoxin and reactive oxygen species-removing systems, the proteasome and the pentose phosphate pathway. A data-driven clustering of the expression patterns identified one specific cluster that mostly consisted of genes known to be regulated by the Yap1p and Skn7p transcription factors, emphasizing their mediator role in the transcriptional response to oxidants. Comparison of our results with data reported for hydrogen peroxide identified 664 genes that specifically respond to cumene hydroperoxide, suggesting distinct transcriptional responses to these two peroxides. Genes up-regulated only by cumene hydroperoxide are mainly related to the cell membrane and cell wall, and proteolysis process, while those down-regulated only by this aromatic peroxide are involved in mitochondrial function.

  7. Genome-Wide Screen for Saccharomyces cerevisiae Genes Contributing to Opportunistic Pathogenicity in an Invertebrate Model Host

    Directory of Open Access Journals (Sweden)

    Sujal S. Phadke

    2018-01-01

    Full Text Available Environmental opportunistic pathogens can exploit vulnerable hosts through expression of traits selected for in their natural environments. Pathogenicity is itself a complicated trait underpinned by multiple complex traits, such as thermotolerance, morphology, and stress response. The baker’s yeast, Saccharomyces cerevisiae, is a species with broad environmental tolerance that has been increasingly reported as an opportunistic pathogen of humans. Here we leveraged the genetic resources available in yeast and a model insect species, the greater waxmoth Galleria mellonella, to provide a genome-wide analysis of pathogenicity factors. Using serial passaging experiments of genetically marked wild-type strains, a hybrid strain was identified as the most fit genotype across all replicates. To dissect the genetic basis for pathogenicity in the hybrid isolate, bulk segregant analysis was performed which revealed eight quantitative trait loci significantly differing between the two bulks with alleles from both parents contributing to pathogenicity. A second passaging experiment with a library of deletion mutants for most yeast genes identified a large number of mutations whose relative fitness differed in vivo vs. in vitro, including mutations in genes controlling cell wall integrity, mitochondrial function, and tyrosine metabolism. Yeast is presumably subjected to a massive assault by the innate insect immune system that leads to melanization of the host and to a large bottleneck in yeast population size. Our data support that resistance to the innate immune response of the insect is key to survival in the host and identifies shared genetic mechanisms between S. cerevisiae and other opportunistic fungal pathogens.

  8. Multi-targeted priming for genome-wide gene expression assays

    Directory of Open Access Journals (Sweden)

    Adomas Aleksandra B

    2010-08-01

    Full Text Available Abstract Background Complementary approaches to assaying global gene expression are needed to assess gene expression in regions that are poorly assayed by current methodologies. A key component of nearly all gene expression assays is the reverse transcription of transcribed sequences that has traditionally been performed by priming the poly-A tails on many of the transcribed genes in eukaryotes with oligo-dT, or by priming RNA indiscriminately with random hexamers. We designed an algorithm to find common sequence motifs that were present within most protein-coding genes of Saccharomyces cerevisiae and of Neurospora crassa, but that were not present within their ribosomal RNA or transfer RNA genes. We then experimentally tested whether degenerately priming these motifs with multi-targeted primers improved the accuracy and completeness of transcriptomic assays. Results We discovered two multi-targeted primers that would prime a preponderance of genes in the genomes of Saccharomyces cerevisiae and Neurospora crassa while avoiding priming ribosomal RNA or transfer RNA. Examining the response of Saccharomyces cerevisiae to nitrogen deficiency and profiling Neurospora crassa early sexual development, we demonstrated that using multi-targeted primers in reverse transcription led to superior performance of microarray profiling and next-generation RNA tag sequencing. Priming with multi-targeted primers in addition to oligo-dT resulted in higher sensitivity, a larger number of well-measured genes and greater power to detect differences in gene expression. Conclusions Our results provide the most complete and detailed expression profiles of the yeast nitrogen starvation response and N. crassa early sexual development to date. Furthermore, our multi-targeting priming methodology for genome-wide gene expression assays provides selective targeting of multiple sequences and counter-selection against undesirable sequences, facilitating a more complete and

  9. Evaluation of different co-inoculation time of non-Saccharomyces/Saccharomyces yeasts in order to obtain reduced ethanol wines

    Directory of Open Access Journals (Sweden)

    Mestre María Victoria

    2016-01-01

    Full Text Available Decreasing ethanol content in wines has become one of the main objectives of winemakers in different areas of the world. The use of selected wine yeasts can be considered one of the most effective and simple tools. The aim of this study was to evaluate the effect of co-inoculation times of selected non-Saccharomyces/Saccharomyces yeasts on the reduction of ethanol levels in wines. Hanseniaspora uvarum BHu9, Starmerella bacillaris BSb55 and Candida membranaefasciens BCm71 were co-inoculate with Saccharomyces cerevisiae under fermentative conditions. Treatments assayed were: pure fermentations of S. cerevisiae BSc203 and non-Saccharomyces yeasts BHu9, BSb55 and BCm71; -co-fermentations: A-BHu9/BSc203; B-BSb55/BSc203 and C-BCm71/BSc203. These co-inoculations were carried out under mixed (simultaneous inoculation, and sequential conditions (non-Saccharomyces yeasts inoculated at initial time and S. cerevisiae at 48, 96 and 144 h. Lower fermentative efficiencies were registered when BHu9 and BSb55 remained pure more time. Conversely, the conversion efficiency was reduced in co-inocula of BCm71/BSc203, when both yeasts interact more time. Metabolites produced during all vinification processes were within acceptable concentration ranges according to the current legislations. Conclusion Time interaction during fermentation processes of non-Saccharomyces and Saccharomyces yeasts showed influence on ethanol production, and this effect would be dependent on the co-inoculated species.

  10. History of genome editing in yeast.

    Science.gov (United States)

    Fraczek, Marcin G; Naseeb, Samina; Delneri, Daniela

    2018-05-01

    For thousands of years humans have used the budding yeast Saccharomyces cerevisiae for the production of bread and alcohol; however, in the last 30-40 years our understanding of the yeast biology has dramatically increased, enabling us to modify its genome. Although S. cerevisiae has been the main focus of many research groups, other non-conventional yeasts have also been studied and exploited for biotechnological purposes. Our experiments and knowledge have evolved from recombination to high-throughput PCR-based transformations to highly accurate CRISPR methods in order to alter yeast traits for either research or industrial purposes. Since the release of the genome sequence of S. cerevisiae in 1996, the precise and targeted genome editing has increased significantly. In this 'Budding topic' we discuss the significant developments of genome editing in yeast, mainly focusing on Cre-loxP mediated recombination, delitto perfetto and CRISPR/Cas. © 2018 The Authors. Yeast published by John Wiley & Sons, Ltd.

  11. The master two-dimensional gel database of human AMA cell proteins: towards linking protein and genome sequence and mapping information (update 1991)

    DEFF Research Database (Denmark)

    Celis, J E; Leffers, H; Rasmussen, H H

    1991-01-01

    autoantigens" and "cDNAs". For convenience we have included an alphabetical list of all known proteins recorded in this database. In the long run, the main goal of this database is to link protein and DNA sequencing and mapping information (Human Genome Program) and to provide an integrated picture......The master two-dimensional gel database of human AMA cells currently lists 3801 cellular and secreted proteins, of which 371 cellular polypeptides (306 IEF; 65 NEPHGE) were added to the master images during the last 10 months. These include: (i) very basic and acidic proteins that do not focus...

  12. Genomic Testing

    Science.gov (United States)

    ... this database. Top of Page Evaluation of Genomic Applications in Practice and Prevention (EGAPP™) In 2004, the Centers for Disease Control and Prevention launched the EGAPP initiative to establish and test a ... and other applications of genomic technology that are in transition from ...

  13. Ribosomal DNA sequence heterogeneity reflects intraspecies phylogenies and predicts genome structure in two contrasting yeast species.

    Science.gov (United States)

    West, Claire; James, Stephen A; Davey, Robert P; Dicks, Jo; Roberts, Ian N

    2014-07-01

    The ribosomal RNA encapsulates a wealth of evolutionary information, including genetic variation that can be used to discriminate between organisms at a wide range of taxonomic levels. For example, the prokaryotic 16S rDNA sequence is very widely used both in phylogenetic studies and as a marker in metagenomic surveys and the internal transcribed spacer region, frequently used in plant phylogenetics, is now recognized as a fungal DNA barcode. However, this widespread use does not escape criticism, principally due to issues such as difficulties in classification of paralogous versus orthologous rDNA units and intragenomic variation, both of which may be significant barriers to accurate phylogenetic inference. We recently analyzed data sets from the Saccharomyces Genome Resequencing Project, characterizing rDNA sequence variation within multiple strains of the baker's yeast Saccharomyces cerevisiae and its nearest wild relative Saccharomyces paradoxus in unprecedented detail. Notably, both species possess single locus rDNA systems. Here, we use these new variation datasets to assess whether a more detailed characterization of the rDNA locus can alleviate the second of these phylogenetic issues, sequence heterogeneity, while controlling for the first. We demonstrate that a strong phylogenetic signal exists within both datasets and illustrate how they can be used, with existing methodology, to estimate intraspecies phylogenies of yeast strains consistent with those derived from whole-genome approaches. We also describe the use of partial Single Nucleotide Polymorphisms, a type of sequence variation found only in repetitive genomic regions, in identifying key evolutionary features such as genome hybridization events and show their consistency with whole-genome Structure analyses. We conclude that our approach can transform rDNA sequence heterogeneity from a problem to a useful source of evolutionary information, enabling the estimation of highly accurate phylogenies of

  14. Reference: 690 [Arabidopsis Phenome Database[Archive

    Lifescience Database Archive (English)

    Full Text Available of the cleavage reaction consisting of Spo11 covalently linked to the 5' termini of DNA. While Rad50 and Mre...th the DNA repair machinery, as the mammalian homologue of Com1/Sae2, with important implications for the mo...nit Top6A. In Saccharomyces cerevisiae, Rad50, Mre11 and Com1/Sae2 are essential to process an intermediate ...11 also confer genome stability to vegetative cells and are well conserved in evo...ore, DNA fragmentation in AtCom1 is suppressed by eliminating AtSPO11-1. In addition, AtCOM1 is specifically required

  15. A New Single Nucleotide Polymorphism Database for Rainbow Trout Generated Through Whole Genome Resequencing

    Directory of Open Access Journals (Sweden)

    Guangtu Gao

    2018-04-01

    heterozygosity within each population. We also provide functional annotation based on the genome position of each SNP and evaluate the use of clonal lines for filtering of PSVs and MSVs. These SNPs form a new database, which provides an important resource for a new high density SNP array design and for other SNP genotyping platforms used for genetic and genomics studies of this iconic salmonid fish species.

  16. Molecular Basis for Saccharomyces cerevisiae Biofilm Development

    DEFF Research Database (Denmark)

    Andersen, Kaj Scherz

    In this study, I sought to identify genes regulating the global molecular program for development of sessile multicellular communities, also known as biofilm, of the eukaryotic microorganism, Saccharomyces cerevisiae (yeast). Yeast biofilm has a clinical interest, as biofilms can cause chronic...... infections in humans. Biofilm is also interesting from an evolutionary standpoint, as an example of primitive multicellularity. By using a genome-wide screen of yeast deletion mutants, I show that 71 genes are essential for biofilm formation. Two-thirds of these genes are required for transcription of FLO11......, but only a small subset is previously described as regulators of FLO11. These results reveal that the regulation of biofilm formation and FLO11 is even more complex than what has previously been described. I find that the molecular program for biofilm formation shares many essential components with two...

  17. Combining magnetic sorting of mother cells and fluctuation tests to analyze genome instability during mitotic cell aging in Saccharomyces cerevisiae.

    Science.gov (United States)

    Patterson, Melissa N; Maxwell, Patrick H

    2014-10-16

    Saccharomyces cerevisiae has been an excellent model system for examining mechanisms and consequences of genome instability. Information gained from this yeast model is relevant to many organisms, including humans, since DNA repair and DNA damage response factors are well conserved across diverse species. However, S. cerevisiae has not yet been used to fully address whether the rate of accumulating mutations changes with increasing replicative (mitotic) age due to technical constraints. For instance, measurements of yeast replicative lifespan through micromanipulation involve very small populations of cells, which prohibit detection of rare mutations. Genetic methods to enrich for mother cells in populations by inducing death of daughter cells have been developed, but population sizes are still limited by the frequency with which random mutations that compromise the selection systems occur. The current protocol takes advantage of magnetic sorting of surface-labeled yeast mother cells to obtain large enough populations of aging mother cells to quantify rare mutations through phenotypic selections. Mutation rates, measured through fluctuation tests, and mutation frequencies are first established for young cells and used to predict the frequency of mutations in mother cells of various replicative ages. Mutation frequencies are then determined for sorted mother cells, and the age of the mother cells is determined using flow cytometry by staining with a fluorescent reagent that detects bud scars formed on their cell surfaces during cell division. Comparison of predicted mutation frequencies based on the number of cell divisions to the frequencies experimentally observed for mother cells of a given replicative age can then identify whether there are age-related changes in the rate of accumulating mutations. Variations of this basic protocol provide the means to investigate the influence of alterations in specific gene functions or specific environmental conditions on

  18. Visualization for genomics: the Microbial Genome Viewer.

    Science.gov (United States)

    Kerkhoven, Robert; van Enckevort, Frank H J; Boekhorst, Jos; Molenaar, Douwe; Siezen, Roland J

    2004-07-22

    A Web-based visualization tool, the Microbial Genome Viewer, is presented that allows the user to combine complex genomic data in a highly interactive way. This Web tool enables the interactive generation of chromosome wheels and linear genome maps from genome annotation data stored in a MySQL database. The generated images are in scalable vector graphics (SVG) format, which is suitable for creating high-quality scalable images and dynamic Web representations. Gene-related data such as transcriptome and time-course microarray experiments can be superimposed on the maps for visual inspection. The Microbial Genome Viewer 1.0 is freely available at http://www.cmbi.kun.nl/MGV

  19. G-InforBIO: integrated system for microbial genomics

    Directory of Open Access Journals (Sweden)

    Abe Takashi

    2006-08-01

    Full Text Available Abstract Background Genome databases contain diverse kinds of information, including gene annotations and nucleotide and amino acid sequences. It is not easy to integrate such information for genomic study. There are few tools for integrated analyses of genomic data, therefore, we developed software that enables users to handle, manipulate, and analyze genome data with a variety of sequence analysis programs. Results The G-InforBIO system is a novel tool for genome data management and sequence analysis. The system can import genome data encoded as eXtensible Markup Language documents as formatted text documents, including annotations and sequences, from DNA Data Bank of Japan and GenBank encoded as flat files. The genome database is constructed automatically after importing, and the database can be exported as documents formatted with eXtensible Markup Language or tab-deliminated text. Users can retrieve data from the database by keyword searches, edit annotation data of genes, and process data with G-InforBIO. In addition, information in the G-InforBIO database can be analyzed seamlessly with nine different software programs, including programs for clustering and homology analyses. Conclusion The G-InforBIO system simplifies genome analyses by integrating several available software programs to allow efficient handling and manipulation of genome data. G-InforBIO is freely available from the download site.

  20. Apoptosis - Triggering Effects: UVB-irradiation and Saccharomyces cerevisiae.

    Science.gov (United States)

    Behzadi, Payam; Behzadi, Elham

    2012-12-01

    The pathogenic disturbance of Saccharomyces cerevisiae is known as a rare but invasive nosocomial fungal infection. This survey is focused on the evaluation of apoptosis-triggering effects of UVB-irradiation in Saccharomyces cerevisiae. The well-growth colonies of Saccharomyces cerevisiae on Sabouraud Dextrose Agar (SDA) were irradiated within an interval of 10 minutes by UVB-light (302 nm). Subsequently, the harvested DNA molecules of control and UV-exposed yeast colonies were run through the 1% agarose gel electrophoresis comprising the luminescent dye of ethidium bromide. No unusual patterns including DNA laddering bands or smears were detected. The applied procedure for UV exposure was not effective for inducing apoptosis in Saccharomyces cerevisiae. So, it needs another UV-radiation protocol for inducing apoptosis phenomenon in Saccharomyces cerevisiae.

  1. Saccharomyces species in the Production of Beer

    Directory of Open Access Journals (Sweden)

    Graham G. Stewart

    2016-12-01

    Full Text Available The characteristic flavour and aroma of any beer is, in large part, determined by the yeast strain employed and the wort composition. In addition, properties such as flocculation, wort fermentation ability (including the uptake of wort sugars, amino acids, and peptides, ethanol and osmotic pressure tolerance together with oxygen requirements have a critical impact on fermentation performance. Yeast management between fermentations is also a critical brewing parameter. Brewer’s yeasts are mostly part of the genus Saccharomyces. Ale yeasts belong to the species Saccharomyces cerevisiae and lager yeasts to the species Saccharomyces pastorianus. The latter is an interspecies hybrid between S. cerevisiae and Saccharomyces eubayanus. Brewer’s yeast strains are facultative anaerobes—they are able to grow in the presence or absence of oxygen and this ability supports their property as an important industrial microorganism. This article covers important aspects of Saccharomyces molecular biology, physiology, and metabolism that is involved in wort fermentation and beer production.

  2. The ecology and evolution of non-domesticated Saccharomyces species.

    Science.gov (United States)

    Boynton, Primrose J; Greig, Duncan

    2014-12-01

    Yeast researchers need model systems for ecology and evolution, but the model yeast Saccharomyces cerevisiae is not ideal because its evolution has been affected by domestication. Instead, ecologists and evolutionary biologists are focusing on close relatives of S. cerevisiae, the seven species in the genus Saccharomyces. The best-studied Saccharomyces yeast, after S. cerevisiae, is S. paradoxus, an oak tree resident throughout the northern hemisphere. In addition, several more members of the genus Saccharomyces have recently been discovered. Some Saccharomyces species are only found in nature, while others include both wild and domesticated strains. Comparisons between domesticated and wild yeasts have pinpointed hybridization, introgression and high phenotypic diversity as signatures of domestication. But studies of wild Saccharomyces natural history, biogeography and ecology are only beginning. Much remains to be understood about wild yeasts' ecological interactions and life cycles in nature. We encourage researchers to continue to investigate Saccharomyces yeasts in nature, both to place S. cerevisiae biology into its ecological context and to develop the genus Saccharomyces as a model clade for ecology and evolution. © 2014 The Authors. Yeast published by John Wiley & Sons, Ltd.

  3. Rapid detection of structural variation in a human genome using nanochannel-based genome mapping technology

    DEFF Research Database (Denmark)

    Cao, Hongzhi; Hastie, Alex R.; Cao, Dandan

    2014-01-01

    mutations; however, none of the current detection methods are comprehensive, and currently available methodologies are incapable of providing sufficient resolution and unambiguous information across complex regions in the human genome. To address these challenges, we applied a high-throughput, cost......-effective genome mapping technology to comprehensively discover genome-wide SVs and characterize complex regions of the YH genome using long single molecules (>150 kb) in a global fashion. RESULTS: Utilizing nanochannel-based genome mapping technology, we obtained 708 insertions/deletions and 17 inversions larger...... fosmid data. Of the remaining 270 SVs, 260 are insertions and 213 overlap known SVs in the Database of Genomic Variants. Overall, 609 out of 666 (90%) variants were supported by experimental orthogonal methods or historical evidence in public databases. At the same time, genome mapping also provides...

  4. The Human Gene Mutation Database: building a comprehensive mutation repository for clinical and molecular genetics, diagnostic testing and personalized genomic medicine.

    Science.gov (United States)

    Stenson, Peter D; Mort, Matthew; Ball, Edward V; Shaw, Katy; Phillips, Andrew; Cooper, David N

    2014-01-01

    The Human Gene Mutation Database (HGMD®) is a comprehensive collection of germline mutations in nuclear genes that underlie, or are associated with, human inherited disease. By June 2013, the database contained over 141,000 different lesions detected in over 5,700 different genes, with new mutation entries currently accumulating at a rate exceeding 10,000 per annum. HGMD was originally established in 1996 for the scientific study of mutational mechanisms in human genes. However, it has since acquired a much broader utility as a central unified disease-oriented mutation repository utilized by human molecular geneticists, genome scientists, molecular biologists, clinicians and genetic counsellors as well as by those specializing in biopharmaceuticals, bioinformatics and personalized genomics. The public version of HGMD (http://www.hgmd.org) is freely available to registered users from academic institutions/non-profit organizations whilst the subscription version (HGMD Professional) is available to academic, clinical and commercial users under license via BIOBASE GmbH.

  5. Improving Microbial Genome Annotations in an Integrated Database Context

    Science.gov (United States)

    Chen, I-Min A.; Markowitz, Victor M.; Chu, Ken; Anderson, Iain; Mavromatis, Konstantinos; Kyrpides, Nikos C.; Ivanova, Natalia N.

    2013-01-01

    Effective comparative analysis of microbial genomes requires a consistent and complete view of biological data. Consistency regards the biological coherence of annotations, while completeness regards the extent and coverage of functional characterization for genomes. We have developed tools that allow scientists to assess and improve the consistency and completeness of microbial genome annotations in the context of the Integrated Microbial Genomes (IMG) family of systems. All publicly available microbial genomes are characterized in IMG using different functional annotation and pathway resources, thus providing a comprehensive framework for identifying and resolving annotation discrepancies. A rule based system for predicting phenotypes in IMG provides a powerful mechanism for validating functional annotations, whereby the phenotypic traits of an organism are inferred based on the presence of certain metabolic reactions and pathways and compared to experimentally observed phenotypes. The IMG family of systems are available at http://img.jgi.doe.gov/. PMID:23424620

  6. Improving microbial genome annotations in an integrated database context.

    Directory of Open Access Journals (Sweden)

    I-Min A Chen

    Full Text Available Effective comparative analysis of microbial genomes requires a consistent and complete view of biological data. Consistency regards the biological coherence of annotations, while completeness regards the extent and coverage of functional characterization for genomes. We have developed tools that allow scientists to assess and improve the consistency and completeness of microbial genome annotations in the context of the Integrated Microbial Genomes (IMG family of systems. All publicly available microbial genomes are characterized in IMG using different functional annotation and pathway resources, thus providing a comprehensive framework for identifying and resolving annotation discrepancies. A rule based system for predicting phenotypes in IMG provides a powerful mechanism for validating functional annotations, whereby the phenotypic traits of an organism are inferred based on the presence of certain metabolic reactions and pathways and compared to experimentally observed phenotypes. The IMG family of systems are available at http://img.jgi.doe.gov/.

  7. Mining for genotype-phenotype relations in Saccharomyces using partial least squares

    Directory of Open Access Journals (Sweden)

    Sæbø Solve

    2011-08-01

    Full Text Available Abstract Background Multivariate approaches are important due to their versatility and applications in many fields as it provides decisive advantages over univariate analysis in many ways. Genome wide association studies are rapidly emerging, but approaches in hand pay less attention to multivariate relation between genotype and phenotype. We introduce a methodology based on a BLAST approach for extracting information from genomic sequences and Soft- Thresholding Partial Least Squares (ST-PLS for mapping genotype-phenotype relations. Results Applying this methodology to an extensive data set for the model yeast Saccharomyces cerevisiae, we found that the relationship between genotype-phenotype involves surprisingly few genes in the sense that an overwhelmingly large fraction of the phenotypic variation can be explained by variation in less than 1% of the full gene reference set containing 5791 genes. These phenotype influencing genes were evolving 20% faster than non-influential genes and were unevenly distributed over cellular functions, with strong enrichments in functions such as cellular respiration and transposition. These genes were also enriched with known paralogs, stop codon variations and copy number variations, suggesting that such molecular adjustments have had a disproportionate influence on Saccharomyces yeasts recent adaptation to environmental changes in its ecological niche. Conclusions BLAST and PLS based multivariate approach derived results that adhere to the known yeast phylogeny and gene ontology and thus verify that the methodology extracts a set of fast evolving genes that capture the phylogeny of the yeast strains. The approach is worth pursuing, and future investigations should be made to improve the computations of genotype signals as well as variable selection procedure within the PLS framework.

  8. Database Description - RGP physicalmap | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available classification Plant databases - Rice Database classification Sequence Physical map Organism Taxonomy Name: ...inobe Journal: Nature Genetics (1994) 8: 365-372. External Links: Article title: Physical Mapping of Rice Ch...rnal: DNA Research (1997) 4(2): 133-140. External Links: Article title: Physical Mapping of Rice Chromosomes... T Sasaki Journal: Genome Research (1996) 6(10): 935-942. External Links: Article title: Physical mapping of

  9. IMG: the integrated microbial genomes database and comparative analysis system

    Science.gov (United States)

    Markowitz, Victor M.; Chen, I-Min A.; Palaniappan, Krishna; Chu, Ken; Szeto, Ernest; Grechkin, Yuri; Ratner, Anna; Jacob, Biju; Huang, Jinghua; Williams, Peter; Huntemann, Marcel; Anderson, Iain; Mavromatis, Konstantinos; Ivanova, Natalia N.; Kyrpides, Nikos C.

    2012-01-01

    The Integrated Microbial Genomes (IMG) system serves as a community resource for comparative analysis of publicly available genomes in a comprehensive integrated context. IMG integrates publicly available draft and complete genomes from all three domains of life with a large number of plasmids and viruses. IMG provides tools and viewers for analyzing and reviewing the annotations of genes and genomes in a comparative context. IMG's data content and analytical capabilities have been continuously extended through regular updates since its first release in March 2005. IMG is available at http://img.jgi.doe.gov. Companion IMG systems provide support for expert review of genome annotations (IMG/ER: http://img.jgi.doe.gov/er), teaching courses and training in microbial genome analysis (IMG/EDU: http://img.jgi.doe.gov/edu) and analysis of genomes related to the Human Microbiome Project (IMG/HMP: http://www.hmpdacc-resources.org/img_hmp). PMID:22194640

  10. A Gondwanan Imprint on Global Diversity and Domestication of Wine and Cider Yeast Saccharomyces uvarum

    Science.gov (United States)

    Almeida, Pedro; Gonçalves, Carla; Teixeira, Sara; Libkind, Diego; Bontrager, Martin; Masneuf-Pomarède, Isabelle; Albertin, Warren; Durrens, Pascal; Sherman, David; Marullo, Philippe; Hittinger, Chris Todd; Gonçalves, Paula; Sampaio, José Paulo

    2016-01-01

    In addition to Saccharomyces cerevisiae, the cryotolerant yeast species S. uvarum is also used for wine and cider fermentation but nothing is known about its natural history. Here we use a population genomics approach to investigate its global phylogeography and domestication fingerprints using a collection of isolates obtained from fermented beverages and from natural environments on five continents. South American isolates contain more genetic diversity than that found in the Northern Hemisphere. Moreover, coalescence analyses suggest that a Patagonian sub-population gave rise to the Holarctic population through a recent bottleneck. Holarctic strains display multiple introgressions from other Saccharomyces species, those from S. eubayanus being prevalent in European strains associated with human-driven fermentations. These introgressions are absent in the large majority of wild strains and gene ontology analyses indicate that several gene categories relevant for wine fermentation are overrepresented. Such findings constitute a first indication of domestication in S. uvarum. PMID:24887054

  11. Identification of mitochondrial carriers in Saccharomyces cerevisiae by transport assay of reconstituted recombinant proteins.

    Science.gov (United States)

    Palmieri, Ferdinando; Agrimi, Gennaro; Blanco, Emanuela; Castegna, Alessandra; Di Noia, Maria A; Iacobazzi, Vito; Lasorsa, Francesco M; Marobbio, Carlo M T; Palmieri, Luigi; Scarcia, Pasquale; Todisco, Simona; Vozza, Angelo; Walker, John

    2006-01-01

    The inner membranes of mitochondria contain a family of carrier proteins that are responsible for the transport in and out of the mitochondrial matrix of substrates, products, co-factors and biosynthetic precursors that are essential for the function and activities of the organelle. This family of proteins is characterized by containing three tandem homologous sequence repeats of approximately 100 amino acids, each folded into two transmembrane alpha-helices linked by an extensive polar loop. Each repeat contains a characteristic conserved sequence. These features have been used to determine the extent of the family in genome sequences. The genome of Saccharomyces cerevisiae contains 34 members of the family. The identity of five of them was known before the determination of the genome sequence, but the functions of the remaining family members were not. This review describes how the functions of 15 of these previously unknown transport proteins have been determined by a strategy that consists of expressing the genes in Escherichia coli or Saccharomyces cerevisiae, reconstituting the gene products into liposomes and establishing their functions by transport assay. Genetic and biochemical evidence as well as phylogenetic considerations have guided the choice of substrates that were tested in the transport assays. The physiological roles of these carriers have been verified by genetic experiments. Various pieces of evidence point to the functions of six additional members of the family, but these proposals await confirmation by transport assay. The sequences of many of the newly identified yeast carriers have been used to characterize orthologs in other species, and in man five diseases are presently known to be caused by defects in specific mitochondrial carrier genes. The roles of eight yeast mitochondrial carriers remain to be established.

  12. Genome-wide screen in Saccharomyces cerevisiae identifies vacuolar protein sorting, autophagy, biosynthetic, and tRNA methylation genes involved in life span regulation.

    Science.gov (United States)

    Fabrizio, Paola; Hoon, Shawn; Shamalnasab, Mehrnaz; Galbani, Abdulaye; Wei, Min; Giaever, Guri; Nislow, Corey; Longo, Valter D

    2010-07-15

    The study of the chronological life span of Saccharomyces cerevisiae, which measures the survival of populations of non-dividing yeast, has resulted in the identification of homologous genes and pathways that promote aging in organisms ranging from yeast to mammals. Using a competitive genome-wide approach, we performed a screen of a complete set of approximately 4,800 viable deletion mutants to identify genes that either increase or decrease chronological life span. Half of the putative short-/long-lived mutants retested from the primary screen were confirmed, demonstrating the utility of our approach. Deletion of genes involved in vacuolar protein sorting, autophagy, and mitochondrial function shortened life span, confirming that respiration and degradation processes are essential for long-term survival. Among the genes whose deletion significantly extended life span are ACB1, CKA2, and TRM9, implicated in fatty acid transport and biosynthesis, cell signaling, and tRNA methylation, respectively. Deletion of these genes conferred heat-shock resistance, supporting the link between life span extension and cellular protection observed in several model organisms. The high degree of conservation of these novel yeast longevity determinants in other species raises the possibility that their role in senescence might be conserved.

  13. phiGENOME: an integrative navigation throughout bacteriophage genomes.

    Science.gov (United States)

    Stano, Matej; Klucar, Lubos

    2011-11-01

    phiGENOME is a web-based genome browser generating dynamic and interactive graphical representation of phage genomes stored in the phiSITE, database of gene regulation in bacteriophages. phiGENOME is an integral part of the phiSITE web portal (http://www.phisite.org/phigenome) and it was optimised for visualisation of phage genomes with the emphasis on the gene regulatory elements. phiGENOME consists of three components: (i) genome map viewer built using Adobe Flash technology, providing dynamic and interactive graphical display of phage genomes; (ii) sequence browser based on precisely formatted HTML tags, providing detailed exploration of genome features on the sequence level and (iii) regulation illustrator, based on Scalable Vector Graphics (SVG) and designed for graphical representation of gene regulations. Bringing 542 complete genome sequences accompanied with their rich annotations and references, makes phiGENOME a unique information resource in the field of phage genomics. Copyright © 2011 Elsevier Inc. All rights reserved.

  14. Lightweight genome viewer: portable software for browsing genomics data in its chromosomal context.

    Science.gov (United States)

    Faith, Jeremiah J; Olson, Andrew J; Gardner, Timothy S; Sachidanandam, Ravi

    2007-09-18

    Lightweight genome viewer (lwgv) is a web-based tool for visualization of sequence annotations in their chromosomal context. It performs most of the functions of larger genome browsers, while relying on standard flat-file formats and bypassing the database needs of most visualization tools. Visualization as an aide to discovery requires display of novel data in conjunction with static annotations in their chromosomal context. With database-based systems, displaying dynamic results requires temporary tables that need to be tracked for removal. lwgv simplifies the visualization of user-generated results on a local computer. The dynamic results of these analyses are written to transient files, which can import static content from a more permanent file. lwgv is currently used in many different applications, from whole genome browsers to single-gene RNAi design visualization, demonstrating its applicability in a large variety of contexts and scales. lwgv provides a lightweight alternative to large genome browsers for visualizing biological annotations and dynamic analyses in their chromosomal context. It is particularly suited for applications ranging from short sequences to medium-sized genomes when the creation and maintenance of a large software and database infrastructure is not necessary or desired.

  15. Genomics and the making of yeast biodiversity.

    Science.gov (United States)

    Hittinger, Chris Todd; Rokas, Antonis; Bai, Feng-Yan; Boekhout, Teun; Gonçalves, Paula; Jeffries, Thomas W; Kominek, Jacek; Lachance, Marc-André; Libkind, Diego; Rosa, Carlos A; Sampaio, José Paulo; Kurtzman, Cletus P

    2015-12-01

    Yeasts are unicellular fungi that do not form fruiting bodies. Although the yeast lifestyle has evolved multiple times, most known species belong to the subphylum Saccharomycotina (syn. Hemiascomycota, hereafter yeasts). This diverse group includes the premier eukaryotic model system, Saccharomyces cerevisiae; the common human commensal and opportunistic pathogen, Candida albicans; and over 1000 other known species (with more continuing to be discovered). Yeasts are found in every biome and continent and are more genetically diverse than angiosperms or chordates. Ease of culture, simple life cycles, and small genomes (∼10-20Mbp) have made yeasts exceptional models for molecular genetics, biotechnology, and evolutionary genomics. Here we discuss recent developments in understanding the genomic underpinnings of the making of yeast biodiversity, comparing and contrasting natural and human-associated evolutionary processes. Only a tiny fraction of yeast biodiversity and metabolic capabilities has been tapped by industry and science. Expanding the taxonomic breadth of deep genomic investigations will further illuminate how genome function evolves to encode their diverse metabolisms and ecologies. Copyright © 2015 Elsevier Ltd. All rights reserved.

  16. Anti-Saccharomyces cerevisiae autoantibodies in autoimmune diseases: from bread baking to autoimmunity.

    Science.gov (United States)

    Rinaldi, Maurizio; Perricone, Roberto; Blank, Miri; Perricone, Carlo; Shoenfeld, Yehuda

    2013-10-01

    Saccharomyces cerevisiae is best known as the baker's and brewer's yeast, but its residual traces are also frequent excipients in some vaccines. Although anti-S. cerevisiae autoantibodies (ASCAs) are considered specific for Crohn's disease, a growing number of studies have detected high levels of ASCAs in patients affected with autoimmune diseases as compared with healthy controls, including antiphospholipid syndrome, systemic lupus erythematosus, type 1 diabetes mellitus, and rheumatoid arthritis. Commensal microorganisms such as Saccharomyces are required for nutrition, proper development of Peyer's aggregated lymphoid tissue, and tissue healing. However, even the commensal nonclassically pathogenic microbiota can trigger autoimmunity when fine regulation of immune tolerance does not work properly. For our purposes, the protein database of the National Center for Biotechnology Information (NCBI) was consulted, comparing Saccharomyces mannan to several molecules with a pathogenetic role in autoimmune diseases. Thanks to the NCBI bioinformation technology tool, several overlaps in molecular structures (50-100 %) were identified when yeast mannan, and the most common autoantigens were compared. The autoantigen U2 snRNP B″ was found to conserve a superfamily protein domain that shares 83 % of the S. cerevisiae mannan sequence. Furthermore, ASCAs may be present years before the diagnosis of some associated autoimmune diseases as they were retrospectively found in the preserved blood samples of soldiers who became affected by Crohn's disease years later. Our results strongly suggest that ASCAs' role in clinical practice should be better addressed in order to evaluate their predictive or prognostic relevance.

  17. PGSB/MIPS Plant Genome Information Resources and Concepts for the Analysis of Complex Grass Genomes.

    Science.gov (United States)

    Spannagl, Manuel; Bader, Kai; Pfeifer, Matthias; Nussbaumer, Thomas; Mayer, Klaus F X

    2016-01-01

    PGSB (Plant Genome and Systems Biology; formerly MIPS-Munich Institute for Protein Sequences) has been involved in developing, implementing and maintaining plant genome databases for more than a decade. Genome databases and analysis resources have focused on individual genomes and aim to provide flexible and maintainable datasets for model plant genomes as a backbone against which experimental data, e.g., from high-throughput functional genomics, can be organized and analyzed. In addition, genomes from both model and crop plants form a scaffold for comparative genomics, assisted by specialized tools such as the CrowsNest viewer to explore conserved gene order (synteny) between related species on macro- and micro-levels.The genomes of many economically important Triticeae plants such as wheat, barley, and rye present a great challenge for sequence assembly and bioinformatic analysis due to their enormous complexity and large genome size. Novel concepts and strategies have been developed to deal with these difficulties and have been applied to the genomes of wheat, barley, rye, and other cereals. This includes the GenomeZipper concept, reference-guided exome assembly, and "chromosome genomics" based on flow cytometry sorted chromosomes.

  18. Systematic Identification of Determinants for Single-Strand Annealing-Mediated Deletion Formation in Saccharomyces cerevisiae

    Directory of Open Access Journals (Sweden)

    Maia Segura-Wang

    2017-10-01

    Full Text Available To ensure genomic integrity, living organisms have evolved diverse molecular processes for sensing and repairing damaged DNA. If improperly repaired, DNA damage can give rise to different types of mutations, an important class of which are genomic structural variants (SVs. In spite of their importance for phenotypic variation and genome evolution, potential contributors to SV formation in Saccharomyces cerevisiae (budding yeast, a highly tractable model organism, are not fully recognized. Here, we developed and applied a genome-wide assay to identify yeast gene knockout mutants associated with de novo deletion formation, in particular single-strand annealing (SSA-mediated deletion formation, in a systematic manner. In addition to genes previously linked to genome instability, our approach implicates novel genes involved in chromatin remodeling and meiosis in affecting the rate of SSA-mediated deletion formation in the presence or absence of stress conditions induced by DNA-damaging agents. We closely examined two candidate genes, the chromatin remodeling gene IOC4 and the meiosis-related gene MSH4, which when knocked-out resulted in gene expression alterations affecting genes involved in cell division and chromosome organization, as well as DNA repair and recombination, respectively. Our high-throughput approach facilitates the systematic identification of processes linked to the formation of a major class of genetic variation.

  19. Genome-wide screening of Saccharomyces cerevisiae genes regulated by vanillin.

    Science.gov (United States)

    Park, Eun-Hee; Kim, Myoung-Dong

    2015-01-01

    During pretreatment of lignocellulosic biomass, a variety of fermentation inhibitors, including acetic acid and vanillin, are released. Using DNA microarray analysis, this study explored genes of the budding yeast Saccharomyces cerevisiae that respond to vanillin-induced stress. The expression of 273 genes was upregulated and that of 205 genes was downregulated under vanillin stress. Significantly induced genes included MCH2, SNG1, GPH1, and TMA10, whereas NOP2, UTP18, FUR1, and SPR1 were down regulated. Sequence analysis of the 5'-flanking region of upregulated genes suggested that vanillin might regulate gene expression in a stress response element (STRE)-dependent manner, in addition to a pathway that involved the transcription factor Yap1p. Retardation in the cell growth of mutant strains indicated that MCH2, SNG1, and GPH1 are intimately involved in vanillin stress response. Deletion of the genes whose expression levels were decreased under vanillin stress did not result in a notable change in S. cerevisiae growth under vanillin stress. This study will provide the basis for a better understanding of the stress response of the yeast S. cerevisiae to fermentation inhibitors.

  20. Using SQL Databases for Sequence Similarity Searching and Analysis.

    Science.gov (United States)

    Pearson, William R; Mackey, Aaron J

    2017-09-13

    Relational databases can integrate diverse types of information and manage large sets of similarity search results, greatly simplifying genome-scale analyses. By focusing on taxonomic subsets of sequences, relational databases can reduce the size and redundancy of sequence libraries and improve the statistical significance of homologs. In addition, by loading similarity search results into a relational database, it becomes possible to explore and summarize the relationships between all of the proteins in an organism and those in other biological kingdoms. This unit describes how to use relational databases to improve the efficiency of sequence similarity searching and demonstrates various large-scale genomic analyses of homology-related data. It also describes the installation and use of a simple protein sequence database, seqdb_demo, which is used as a basis for the other protocols. The unit also introduces search_demo, a database that stores sequence similarity search results. The search_demo database is then used to explore the evolutionary relationships between E. coli proteins and proteins in other organisms in a large-scale comparative genomic analysis. © 2017 by John Wiley & Sons, Inc. Copyright © 2017 John Wiley & Sons, Inc.

  1. Assembly and Multiplex Genome Integration of Metabolic Pathways in Yeast Using CasEMBLR

    DEFF Research Database (Denmark)

    Jakočiūnas, Tadas; Jensen, Emil D.; Jensen, Michael Krogh

    2018-01-01

    and marker-free integration of the carotenoid pathway from 15 exogenously supplied DNA parts into three targeted genomic loci. As a second proof-of-principle, a total of ten DNA parts were assembled and integrated in two genomic loci to construct a tyrosine production strain, and at the same time knocking......Genome integration is a vital step for implementing large biochemical pathways to build a stable microbial cell factory. Although traditional strain construction strategies are well established for the model organism Saccharomyces cerevisiae, recent advances in CRISPR/Cas9-mediated genome...... engineering allow much higher throughput and robustness in terms of strain construction. In this chapter, we describe CasEMBLR, a highly efficient and marker-free genome engineering method for one-step integration of in vivo assembled expression cassettes in multiple genomic sites simultaneously. Cas...

  2. Genomic characterization of large heterochromatic gaps in the human genome assembly.

    Directory of Open Access Journals (Sweden)

    Nicolas Altemose

    2014-05-01

    Full Text Available The largest gaps in the human genome assembly correspond to multi-megabase heterochromatic regions composed primarily of two related families of tandem repeats, Human Satellites 2 and 3 (HSat2,3. The abundance of repetitive DNA in these regions challenges standard mapping and assembly algorithms, and as a result, the sequence composition and potential biological functions of these regions remain largely unexplored. Furthermore, existing genomic tools designed to predict consensus-based descriptions of repeat families cannot be readily applied to complex satellite repeats such as HSat2,3, which lack a consistent repeat unit reference sequence. Here we present an alignment-free method to characterize complex satellites using whole-genome shotgun read datasets. Utilizing this approach, we classify HSat2,3 sequences into fourteen subfamilies and predict their chromosomal distributions, resulting in a comprehensive satellite reference database to further enable genomic studies of heterochromatic regions. We also identify 1.3 Mb of non-repetitive sequence interspersed with HSat2,3 across 17 unmapped assembly scaffolds, including eight annotated gene predictions. Finally, we apply our satellite reference database to high-throughput sequence data from 396 males to estimate array size variation of the predominant HSat3 array on the Y chromosome, confirming that satellite array sizes can vary between individuals over an order of magnitude (7 to 98 Mb and further demonstrating that array sizes are distributed differently within distinct Y haplogroups. In summary, we present a novel framework for generating initial reference databases for unassembled genomic regions enriched with complex satellite DNA, and we further demonstrate the utility of these reference databases for studying patterns of sequence variation within human populations.

  3. Transcription factor control of growth rate dependent genes in Saccharomyces cerevisiae: A three factor design

    DEFF Research Database (Denmark)

    Fazio, Alessandro; Jewett, Michael Christopher; Daran-Lapujade, Pascale

    2008-01-01

    , such as Ace2 and Swi6, and stress response regulators, such as Yap1, were also shown to have significantly enriched target sets. Conclusion: Our work, which is the first genome-wide gene expression study to investigate specific growth rate and consider the impact of oxygen availability, provides a more......Background: Characterization of cellular growth is central to understanding living systems. Here, we applied a three-factor design to study the relationship between specific growth rate and genome-wide gene expression in 36 steady-state chemostat cultures of Saccharomyces cerevisiae. The three...... factors we considered were specific growth rate, nutrient limitation, and oxygen availability. Results: We identified 268 growth rate dependent genes, independent of nutrient limitation and oxygen availability. The transcriptional response was used to identify key areas in metabolism around which m...

  4. FunCoup 3.0: database of genome-wide functional coupling networks.

    Science.gov (United States)

    Schmitt, Thomas; Ogris, Christoph; Sonnhammer, Erik L L

    2014-01-01

    We present an update of the FunCoup database (http://FunCoup.sbc.su.se) of functional couplings, or functional associations, between genes and gene products. Identifying these functional couplings is an important step in the understanding of higher level mechanisms performed by complex cellular processes. FunCoup distinguishes between four classes of couplings: participation in the same signaling cascade, participation in the same metabolic process, co-membership in a protein complex and physical interaction. For each of these four classes, several types of experimental and statistical evidence are combined by Bayesian integration to predict genome-wide functional coupling networks. The FunCoup framework has been completely re-implemented to allow for more frequent future updates. It contains many improvements, such as a regularization procedure to automatically downweight redundant evidences and a novel method to incorporate phylogenetic profile similarity. Several datasets have been updated and new data have been added in FunCoup 3.0. Furthermore, we have developed a new Web site, which provides powerful tools to explore the predicted networks and to retrieve detailed information about the data underlying each prediction.

  5. Effect of Saccharomyces, Non-Saccharomyces Yeasts and Malolactic Fermentation Strategies on Fermentation Kinetics and Flavor of Shiraz Wines

    Directory of Open Access Journals (Sweden)

    Heinrich du Plessis

    2017-12-01

    Full Text Available The use of non-Saccharomyces yeasts to improve complexity and diversify wine style is increasing; however, the interactions between non-Saccharomyces yeasts and lactic acid bacteria (LAB have not received much attention. This study investigated the interactions of seven non-Saccharomyces yeast strains of the genera Candida, Hanseniaspora, Lachancea, Metschnikowia and Torulaspora in combination with S. cerevisiae and three malolactic fermentation (MLF strategies in a Shiraz winemaking trial. Standard oenological parameters, volatile composition and sensory profiles of wines were investigated. Wines produced with non-Saccharomyces yeasts had lower alcohol and glycerol levels than wines produced with S. cerevisiae only. Malolactic fermentation also completed faster in these wines. Wines produced with non-Saccharomyces yeasts differed chemically and sensorially from wines produced with S. cerevisiae only. The Candida zemplinina and the one L. thermotolerans isolate slightly inhibited LAB growth in wines that underwent simultaneous MLF. Malolactic fermentation strategy had a greater impact on sensory profiles than yeast treatment. Both yeast selection and MLF strategy had a significant effect on berry aroma, but MLF strategy also had a significant effect on acid balance and astringency of wines. Winemakers should apply the optimal yeast combination and MLF strategy to ensure fast completion of MLF and improve wine complexity.

  6. ChickVD: a sequence variation database for the chicken genome

    DEFF Research Database (Denmark)

    Wang, Jing; He, Ximiao; Ruan, Jue

    2005-01-01

    Working in parallel with the efforts to sequence the chicken (Gallus gallus) genome, the Beijing Genomics Institute led an international team of scientists from China, USA, UK, Sweden, The Netherlands and Germany to map extensive DNA sequence variation throughout the chicken genome by sampling DN...... on quantitative trait loci using data from collaborating institutions and public resources. Our data can be queried by search engine and homology-based BLAST searches. ChickVD is publicly accessible at http://chicken.genomics.org.cn. Udgivelsesdato: 2005-Jan-1...

  7. Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Bojsen, Rasmus K; Andersen, Kaj Scherz; Regenberg, Birgitte

    2012-01-01

    Microbial biofilms can be defined as multi-cellular aggregates adhering to a surface and embedded in an extracellular matrix (ECM). The nonpathogenic yeast, Saccharomyces cerevisiae, follows the common traits of microbial biofilms with cell-cell and cell-surface adhesion. S. cerevisiae is shown t...

  8. Mechanisms and Regulation of Mitotic Recombination in Saccharomyces cerevisiae

    Science.gov (United States)

    Symington, Lorraine S.; Rothstein, Rodney; Lisby, Michael

    2014-01-01

    Homology-dependent exchange of genetic information between DNA molecules has a profound impact on the maintenance of genome integrity by facilitating error-free DNA repair, replication, and chromosome segregation during cell division as well as programmed cell developmental events. This chapter will focus on homologous mitotic recombination in budding yeast Saccharomyces cerevisiae. However, there is an important link between mitotic and meiotic recombination (covered in the forthcoming chapter by Hunter et al. 2015) and many of the functions are evolutionarily conserved. Here we will discuss several models that have been proposed to explain the mechanism of mitotic recombination, the genes and proteins involved in various pathways, the genetic and physical assays used to discover and study these genes, and the roles of many of these proteins inside the cell. PMID:25381364

  9. Lightweight genome viewer: portable software for browsing genomics data in its chromosomal context

    Directory of Open Access Journals (Sweden)

    Gardner Timothy S

    2007-09-01

    Full Text Available Abstract Background Lightweight genome viewer (lwgv is a web-based tool for visualization of sequence annotations in their chromosomal context. It performs most of the functions of larger genome browsers, while relying on standard flat-file formats and bypassing the database needs of most visualization tools. Visualization as an aide to discovery requires display of novel data in conjunction with static annotations in their chromosomal context. With database-based systems, displaying dynamic results requires temporary tables that need to be tracked for removal. Results lwgv simplifies the visualization of user-generated results on a local computer. The dynamic results of these analyses are written to transient files, which can import static content from a more permanent file. lwgv is currently used in many different applications, from whole genome browsers to single-gene RNAi design visualization, demonstrating its applicability in a large variety of contexts and scales. Conclusion lwgv provides a lightweight alternative to large genome browsers for visualizing biological annotations and dynamic analyses in their chromosomal context. It is particularly suited for applications ranging from short sequences to medium-sized genomes when the creation and maintenance of a large software and database infrastructure is not necessary or desired.

  10. A Bayesian method for identifying missing enzymes in predicted metabolic pathway databases

    Directory of Open Access Journals (Sweden)

    Karp Peter D

    2004-06-01

    Full Text Available Abstract Background The PathoLogic program constructs Pathway/Genome databases by using a genome's annotation to predict the set of metabolic pathways present in an organism. PathoLogic determines the set of reactions composing those pathways from the enzymes annotated in the organism's genome. Most annotation efforts fail to assign function to 40–60% of sequences. In addition, large numbers of sequences may have non-specific annotations (e.g., thiolase family protein. Pathway holes occur when a genome appears to lack the enzymes needed to catalyze reactions in a pathway. If a protein has not been assigned a specific function during the annotation process, any reaction catalyzed by that protein will appear as a missing enzyme or pathway hole in a Pathway/Genome database. Results We have developed a method that efficiently combines homology and pathway-based evidence to identify candidates for filling pathway holes in Pathway/Genome databases. Our program not only identifies potential candidate sequences for pathway holes, but combines data from multiple, heterogeneous sources to assess the likelihood that a candidate has the required function. Our algorithm emulates the manual sequence annotation process, considering not only evidence from homology searches, but also considering evidence from genomic context (i.e., is the gene part of an operon? and functional context (e.g., are there functionally-related genes nearby in the genome? to determine the posterior belief that a candidate has the required function. The method can be applied across an entire metabolic pathway network and is generally applicable to any pathway database. The program uses a set of sequences encoding the required activity in other genomes to identify candidate proteins in the genome of interest, and then evaluates each candidate by using a simple Bayes classifier to determine the probability that the candidate has the desired function. We achieved 71% precision at a

  11. Genome-wide identification of genes involved in growth and fermentation activity at low temperature in Saccharomyces cerevisiae.

    Science.gov (United States)

    Salvadó, Zoel; Ramos-Alonso, Lucía; Tronchoni, Jordi; Penacho, Vanessa; García-Ríos, Estéfani; Morales, Pilar; Gonzalez, Ramon; Guillamón, José Manuel

    2016-11-07

    Fermentation at low temperatures is one of the most popular current winemaking practices because of its reported positive impact on the aromatic profile of wines. However, low temperature is an additional hurdle to develop Saccharomyces cerevisiae wine yeasts, which are already stressed by high osmotic pressure, low pH and poor availability of nitrogen sources in grape must. Understanding the mechanisms of adaptation of S. cerevisiae to fermentation at low temperature would help to design strategies for process management, and to select and improve wine yeast strains specifically adapted to this winemaking practice. The problem has been addressed by several approaches in recent years, including transcriptomic and other high-throughput strategies. In this work we used a genome-wide screening of S. cerevisiae diploid mutant strain collections to identify genes that potentially contribute to adaptation to low temperature fermentation conditions. Candidate genes, impaired for growth at low temperatures (12°C and 18°C), but not at a permissive temperature (28°C), were deleted in an industrial homozygous genetic background, wine yeast strain FX10, in both heterozygosis and homozygosis. Some candidate genes were required for growth at low temperatures only in the laboratory yeast genetic background, but not in FX10 (namely the genes involved in aromatic amino acid biosynthesis). Other genes related to ribosome biosynthesis (SNU66 and PAP2) were required for low-temperature fermentation of synthetic must (SM) in the industrial genetic background. This result coincides with our previous findings about translation efficiency with the fitness of different wine yeast strains at low temperature. Copyright © 2016 Elsevier B.V. All rights reserved.

  12. Isolation and characterization of PEP3, a gene required for vacuolar biogenesis in Saccharomyces cerevisiae.

    OpenAIRE

    Preston, R A; Manolson, M F; Becherer, K; Weidenhammer, E; Kirkpatrick, D; Wright, R; Jones, E W

    1991-01-01

    The Saccharomyces cerevisiae PEP3 gene was cloned from a wild-type genomic library by complementation of the carboxypeptidase Y deficiency in a pep3-12 strain. Subclone complementation results localized the PEP3 gene to a 3.8-kb DNA fragment. The DNA sequence of the fragment was determined; a 2,754-bp open reading frame predicts that the PEP3 gene product is a hydrophilic, 107-kDa protein that has no significant similarity to any known protein. The PEP3 predicted protein has a zinc finger (CX...

  13. Construction of an Ostrea edulis database from genomic and expressed sequence tags (ESTs) obtained from Bonamia ostreae infected haemocytes: Development of an immune-enriched oligo-microarray.

    Science.gov (United States)

    Pardo, Belén G; Álvarez-Dios, José Antonio; Cao, Asunción; Ramilo, Andrea; Gómez-Tato, Antonio; Planas, Josep V; Villalba, Antonio; Martínez, Paulino

    2016-12-01

    The flat oyster, Ostrea edulis, is one of the main farmed oysters, not only in Europe but also in the United States and Canada. Bonamiosis due to the parasite Bonamia ostreae has been associated with high mortality episodes in this species. This parasite is an intracellular protozoan that infects haemocytes, the main cells involved in oyster defence. Due to the economical and ecological importance of flat oyster, genomic data are badly needed for genetic improvement of the species, but they are still very scarce. The objective of this study is to develop a sequence database, OedulisDB, with new genomic and transcriptomic resources, providing new data and convenient tools to improve our knowledge of the oyster's immune mechanisms. Transcriptomic and genomic sequences were obtained using 454 pyrosequencing and compiled into an O. edulis database, OedulisDB, consisting of two sets of 10,318 and 7159 unique sequences that represent the oyster's genome (WG) and de novo haemocyte transcriptome (HT), respectively. The flat oyster transcriptome was obtained from two strains (naïve and tolerant) challenged with B. ostreae, and from their corresponding non-challenged controls. Approximately 78.5% of 5619 HT unique sequences were successfully annotated by Blast search using public databases. A total of 984 sequences were identified as being related to immune response and several key immune genes were identified for the first time in flat oyster. Additionally, transcriptome information was used to design and validate the first oligo-microarray in flat oyster enriched with immune sequences from haemocytes. Our transcriptomic and genomic sequencing and subsequent annotation have largely increased the scarce resources available for this economically important species and have enabled us to develop an OedulisDB database and accompanying tools for gene expression analysis. This study represents the first attempt to characterize in depth the O. edulis haemocyte transcriptome in

  14. Phylogenetic distribution of large-scale genome patchiness

    Directory of Open Access Journals (Sweden)

    Hackenberg Michael

    2008-04-01

    Full Text Available Abstract Background The phylogenetic distribution of large-scale genome structure (i.e. mosaic compositional patchiness has been explored mainly by analytical ultracentrifugation of bulk DNA. However, with the availability of large, good-quality chromosome sequences, and the recently developed computational methods to directly analyze patchiness on the genome sequence, an evolutionary comparative analysis can be carried out at the sequence level. Results The local variations in the scaling exponent of the Detrended Fluctuation Analysis are used here to analyze large-scale genome structure and directly uncover the characteristic scales present in genome sequences. Furthermore, through shuffling experiments of selected genome regions, computationally-identified, isochore-like regions were identified as the biological source for the uncovered large-scale genome structure. The phylogenetic distribution of short- and large-scale patchiness was determined in the best-sequenced genome assemblies from eleven eukaryotic genomes: mammals (Homo sapiens, Pan troglodytes, Mus musculus, Rattus norvegicus, and Canis familiaris, birds (Gallus gallus, fishes (Danio rerio, invertebrates (Drosophila melanogaster and Caenorhabditis elegans, plants (Arabidopsis thaliana and yeasts (Saccharomyces cerevisiae. We found large-scale patchiness of genome structure, associated with in silico determined, isochore-like regions, throughout this wide phylogenetic range. Conclusion Large-scale genome structure is detected by directly analyzing DNA sequences in a wide range of eukaryotic chromosome sequences, from human to yeast. In all these genomes, large-scale patchiness can be associated with the isochore-like regions, as directly detected in silico at the sequence level.

  15. CyanoClust: comparative genome resources of cyanobacteria and plastids.

    Science.gov (United States)

    Sasaki, Naobumi V; Sato, Naoki

    2010-01-01

    Cyanobacteria, which perform oxygen-evolving photosynthesis as do chloroplasts of plants and algae, are one of the best-studied prokaryotic phyla and one from which many representative genomes have been sequenced. Lack of a suitable comparative genomic database has been a problem in cyanobacterial genomics because many proteins involved in physiological functions such as photosynthesis and nitrogen fixation are not catalogued in commonly used databases, such as Clusters of Orthologous Proteins (COG). CyanoClust is a database of homolog groups in cyanobacteria and plastids that are produced by the program Gclust. We have developed a web-server system for the protein homology database featuring cyanobacteria and plastids. Database URL: http://cyanoclust.c.u-tokyo.ac.jp/.

  16. Effects of fermentation by Saccharomyces cerevisiae and ...

    African Journals Online (AJOL)

    yassine

    2013-02-13

    Feb 13, 2013 ... Full Length Research Paper. Effect of Saccharomyces cerevisiae fermentation on the ... 2003). Besides, several alcoholic beverages such as wine or liqueurs are obtained from fruit juices fermented by Saccharomyces ..... (2003). Kinetics of pigment release from hairy root cultures of Beta vulgaris under the ...

  17. YMDB: the Yeast Metabolome Database

    Science.gov (United States)

    Jewison, Timothy; Knox, Craig; Neveu, Vanessa; Djoumbou, Yannick; Guo, An Chi; Lee, Jacqueline; Liu, Philip; Mandal, Rupasri; Krishnamurthy, Ram; Sinelnikov, Igor; Wilson, Michael; Wishart, David S.

    2012-01-01

    The Yeast Metabolome Database (YMDB, http://www.ymdb.ca) is a richly annotated ‘metabolomic’ database containing detailed information about the metabolome of Saccharomyces cerevisiae. Modeled closely after the Human Metabolome Database, the YMDB contains >2000 metabolites with links to 995 different genes/proteins, including enzymes and transporters. The information in YMDB has been gathered from hundreds of books, journal articles and electronic databases. In addition to its comprehensive literature-derived data, the YMDB also contains an extensive collection of experimental intracellular and extracellular metabolite concentration data compiled from detailed Mass Spectrometry (MS) and Nuclear Magnetic Resonance (NMR) metabolomic analyses performed in our lab. This is further supplemented with thousands of NMR and MS spectra collected on pure, reference yeast metabolites. Each metabolite entry in the YMDB contains an average of 80 separate data fields including comprehensive compound description, names and synonyms, structural information, physico-chemical data, reference NMR and MS spectra, intracellular/extracellular concentrations, growth conditions and substrates, pathway information, enzyme data, gene/protein sequence data, as well as numerous hyperlinks to images, references and other public databases. Extensive searching, relational querying and data browsing tools are also provided that support text, chemical structure, spectral, molecular weight and gene/protein sequence queries. Because of S. cervesiae's importance as a model organism for biologists and as a biofactory for industry, we believe this kind of database could have considerable appeal not only to metabolomics researchers, but also to yeast biologists, systems biologists, the industrial fermentation industry, as well as the beer, wine and spirit industry. PMID:22064855

  18. Comparative sequence analysis of Sordaria macrospora and Neurospora crassa as a means to improve genome annotation.

    Science.gov (United States)

    Nowrousian, Minou; Würtz, Christian; Pöggeler, Stefanie; Kück, Ulrich

    2004-03-01

    One of the most challenging parts of large scale sequencing projects is the identification of functional elements encoded in a genome. Recently, studies of genomes of up to six different Saccharomyces species have demonstrated that a comparative analysis of genome sequences from closely related species is a powerful approach to identify open reading frames and other functional regions within genomes [Science 301 (2003) 71, Nature 423 (2003) 241]. Here, we present a comparison of selected sequences from Sordaria macrospora to their corresponding Neurospora crassa orthologous regions. Our analysis indicates that due to the high degree of sequence similarity and conservation of overall genomic organization, S. macrospora sequence information can be used to simplify the annotation of the N. crassa genome.

  19. Experience with Saccharomyces boulardii Probiotic in Oncohaematological Patients.

    Science.gov (United States)

    Sulik-Tyszka, Beata; Snarski, Emilian; Niedźwiedzka, Magda; Augustyniak, Małgorzata; Myhre, Thorvald Nilsen; Kacprzyk, Anna; Swoboda-Kopeć, Ewa; Roszkowska, Marta; Dwilewicz-Trojaczek, Jadwiga; Jędrzejczak, Wiesław Wiktor; Wróblewska, Marta

    2018-06-01

    Very few reports have been published to date on the bloodstream infections caused by Saccharomyces spp. in oncohaematological patients, and there are no guidelines on the use of this probiotic microorganism in this population. We describe the use of probiotic preparation containing Saccharomyces boulardii in a large group of oncohaematological patients. We retrospectively analysed the data from 32,000 patient hospitalisations at the haematological centre during 2011-2013 (including 196 haematopoietic stem cell transplant recipients) in a tertiary care university-affiliated hospital. During the study period, 2270 doses of Saccharomyces boulardii probiotic were administered to the oncohaematological patients. In total, 2816 mycological cultures were performed, out of which 772 (27.4%) were positive, with 52 indicating digestive tract colonisation by Saccharomyces spp., mainly in patients with acute myeloid leukaemia (AML), myelodysplastic syndrome (MDS) or multiple myeloma (MM). While colonised, they were hospitalised for 1683 days and 416 microbiological cultures of their clinical samples were performed. In the studied group of patients, there were six blood cultures positive for fungi; however, they comprised Candida species: two C. glabrata, one C. albicans, one C. krusei, one C. tropicalis and one C. parapsilosis. There was no blood culture positive for Saccharomyces spp. Our study indicates that despite colonisation of many oncohaematological patients with Saccharomyces spp., there were no cases of fungal sepsis caused by this species.

  20. Saccharomyces eubayanus and Saccharomyces uvarum associated with the fermentation of Araucaria araucana seeds in Patagonia.

    Science.gov (United States)

    Rodríguez, M Eugenia; Pérez-Través, Laura; Sangorrín, Marcela P; Barrio, Eladio; Lopes, Christian A

    2014-09-01

    Mudai is a traditional fermented beverage, made from the seeds of the Araucaria araucana tree by Mapuche communities. The main goal of the present study was to identify and characterize the yeast microbiota responsible of Mudai fermentation as well as from A. araucana seeds and bark from different locations in Northern Patagonia. Only Hanseniaspora uvarum and a commercial bakery strain of Saccharomyces cerevisiae were isolated from Mudai and all Saccharomyces isolates recovered from A. araucana seed and bark samples belonged to the cryotolerant species Saccharomyces eubayanus and Saccharomyces uvarum. These two species were already reported in Nothofagus trees from Patagonia; however, this is the first time that they were isolated from A. araucana, which extends their ecological distribution. The presence of these species in A. araucana seeds and bark samples, led us to postulate a potential role for them as the original yeasts responsible for the elaboration of Mudai before the introduction of commercial S. cerevisiae cultures. The molecular and genetic characterization of the S. uvarum and S. eubayanus isolates and their comparison with European S. uvarum strains and S. eubayanus hybrids (S. bayanus and S. pastorianus), allowed their ecology and evolution us to be examined. © 2014 Federation of European Microbiological Societies. Published by John Wiley & Sons Ltd. All rights reserved.

  1. Truncation of Gal4p explains the inactivation of the GAL/MEL regulon in both Saccharomyces bayanus and some Saccharomyces cerevisiae wine strains.

    Science.gov (United States)

    Dulermo, Rémi; Legras, Jean-Luc; Brunel, François; Devillers, Hugo; Sarilar, Véronique; Neuvéglise, Cécile; Nguyen, Huu-Vang

    2016-09-01

    In the past, the galactose-negative (Gal(-)) phenotype was a key physiological character used to distinguish Saccharomyces bayanus from S. cerevisiae In this work, we investigated the inactivation of GAL gene networks in S. bayanus, which is an S. uvarum/S. eubayanus hybrid, and in S. cerevisiae wine strains erroneously labelled 'S. bayanus'. We made an inventory of their GAL genes using genomes that were either available publicly, re-sequenced by us, or assembled from public data and completed with targeted sequencing. In the S. eubayanus/S. uvarum CBS 380(T) hybrid, the GAL/MEL network is composed of genes from both parents: from S. uvarum, an otherwise complete set that lacks GAL4, and from S. eubayanus, a truncated version of GAL4 and an additional copy of GAL3 and GAL80 Similarly, two different truncated GAL4 alleles were found in S. cerevisiae wine strains EC1118 and LalvinQA23. The lack of GAL4 activity in these strains was corrected by introducing a full-length copy of S. cerevisiae GAL4 on a CEN4/ARS plasmid. Transformation with this plasmid restored galactose utilisation in Gal(-) strains, and melibiose fermentation in strain CBS 380(T) The melibiose fermentation phenotype, formerly regarded as characteristic of S. uvarum, turned out to be widespread among Saccharomyces species. © FEMS 2016. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

  2. NeisseriaBase: a specialised Neisseria genomic resource and analysis platform.

    Science.gov (United States)

    Zheng, Wenning; Mutha, Naresh V R; Heydari, Hamed; Dutta, Avirup; Siow, Cheuk Chuen; Jakubovics, Nicholas S; Wee, Wei Yee; Tan, Shi Yang; Ang, Mia Yang; Wong, Guat Jah; Choo, Siew Woh

    2016-01-01

    Background. The gram-negative Neisseria is associated with two of the most potent human epidemic diseases: meningococcal meningitis and gonorrhoea. In both cases, disease is caused by bacteria colonizing human mucosal membrane surfaces. Overall, the genus shows great diversity and genetic variation mainly due to its ability to acquire and incorporate genetic material from a diverse range of sources through horizontal gene transfer. Although a number of databases exist for the Neisseria genomes, they are mostly focused on the pathogenic species. In this present study we present the freely available NeisseriaBase, a database dedicated to the genus Neisseria encompassing the complete and draft genomes of 15 pathogenic and commensal Neisseria species. Methods. The genomic data were retrieved from National Center for Biotechnology Information (NCBI) and annotated using the RAST server which were then stored into the MySQL database. The protein-coding genes were further analyzed to obtain information such as calculation of GC content (%), predicted hydrophobicity and molecular weight (Da) using in-house Perl scripts. The web application was developed following the secure four-tier web application architecture: (1) client workstation, (2) web server, (3) application server, and (4) database server. The web interface was constructed using PHP, JavaScript, jQuery, AJAX and CSS, utilizing the model-view-controller (MVC) framework. The in-house developed bioinformatics tools implemented in NeisseraBase were developed using Python, Perl, BioPerl and R languages. Results. Currently, NeisseriaBase houses 603,500 Coding Sequences (CDSs), 16,071 RNAs and 13,119 tRNA genes from 227 Neisseria genomes. The database is equipped with interactive web interfaces. Incorporation of the JBrowse genome browser in the database enables fast and smooth browsing of Neisseria genomes. NeisseriaBase includes the standard BLAST program to facilitate homology searching, and for Virulence Factor

  3. NeisseriaBase: a specialised Neisseria genomic resource and analysis platform

    Directory of Open Access Journals (Sweden)

    Wenning Zheng

    2016-03-01

    Full Text Available Background. The gram-negative Neisseria is associated with two of the most potent human epidemic diseases: meningococcal meningitis and gonorrhoea. In both cases, disease is caused by bacteria colonizing human mucosal membrane surfaces. Overall, the genus shows great diversity and genetic variation mainly due to its ability to acquire and incorporate genetic material from a diverse range of sources through horizontal gene transfer. Although a number of databases exist for the Neisseria genomes, they are mostly focused on the pathogenic species. In this present study we present the freely available NeisseriaBase, a database dedicated to the genus Neisseria encompassing the complete and draft genomes of 15 pathogenic and commensal Neisseria species. Methods. The genomic data were retrieved from National Center for Biotechnology Information (NCBI and annotated using the RAST server which were then stored into the MySQL database. The protein-coding genes were further analyzed to obtain information such as calculation of GC content (%, predicted hydrophobicity and molecular weight (Da using in-house Perl scripts. The web application was developed following the secure four-tier web application architecture: (1 client workstation, (2 web server, (3 application server, and (4 database server. The web interface was constructed using PHP, JavaScript, jQuery, AJAX and CSS, utilizing the model-view-controller (MVC framework. The in-house developed bioinformatics tools implemented in NeisseraBase were developed using Python, Perl, BioPerl and R languages. Results. Currently, NeisseriaBase houses 603,500 Coding Sequences (CDSs, 16,071 RNAs and 13,119 tRNA genes from 227 Neisseria genomes. The database is equipped with interactive web interfaces. Incorporation of the JBrowse genome browser in the database enables fast and smooth browsing of Neisseria genomes. NeisseriaBase includes the standard BLAST program to facilitate homology searching, and for Virulence

  4. Outlining a future for non-Saccharomyces yeasts: selection of putative spoilage wine strains to be used in association with Saccharomyces cerevisiae for grape juice fermentation.

    Science.gov (United States)

    Domizio, Paola; Romani, Cristina; Lencioni, Livio; Comitini, Francesca; Gobbi, Mirko; Mannazzu, Ilaria; Ciani, Maurizio

    2011-06-30

    The use of non-Saccharomyces yeasts that are generally considered as spoilage yeasts, in association with Saccharomyces cerevisiae for grape must fermentation was here evaluated. Analysis of the main oenological characteristics of pure cultures of 55 yeasts belonging to the genera Hanseniaspora, Pichia, Saccharomycodes and Zygosaccharomyces revealed wide biodiversity within each genus. Moreover, many of these non-Saccharomyces strains had interesting oenological properties in terms of fermentation purity, and ethanol and secondary metabolite production. The use of four non-Saccharomyces yeasts (one per genus) in mixed cultures with a commercial S. cerevisiae strain at different S. cerevisiae/non-Saccharomyces inoculum ratios was investigated. This revealed that most of the compounds normally produced at high concentrations by pure cultures of non-Saccharomyces, and which are considered detrimental to wine quality, do not reach threshold taste levels in these mixed fermentations. On the other hand, the analytical profiles of the wines produced by these mixed cultures indicated that depending on the yeast species and the S. cerevisiae/non-Saccharomyces inoculum ratio, these non-Saccharomyces yeasts can be used to increase production of polysaccharides and to modulate the final concentrations of acetic acid and volatile compounds, such as ethyl acetate, phenyl-ethyl acetate, 2-phenyl ethanol, and 2-methyl 1-butanol. Copyright © 2011 Elsevier B.V. All rights reserved.

  5. gb4gv: a genome browser for geminivirus

    Directory of Open Access Journals (Sweden)

    Eric S. Ho

    2017-04-01

    Full Text Available Background Geminiviruses (family Geminiviridae are prevalent plant viruses that imperil agriculture globally, causing serious damage to the livelihood of farmers, particularly in developing countries. The virus evolves rapidly, attributing to its single-stranded genome propensity, resulting in worldwide circulation of diverse and viable genomes. Genomics is a prominent approach taken by researchers in elucidating the infectious mechanism of the virus. Currently, the NCBI Viral Genome website is a popular repository of viral genomes that conveniently provides researchers a centralized data source of genomic information. However, unlike the genome of living organisms, viral genomes most often maintain peculiar characteristics that fit into no single genome architecture. By imposing a unified annotation scheme on the myriad of viral genomes may downplay their hallmark features. For example, the viron of begomoviruses prevailing in America encapsulates two similar-sized circular DNA components and both are required for systemic infection of plants. However, the bipartite components are kept separately in NCBI as individual genomes with no explicit association in linking them. Thus, our goal is to build a comprehensive Geminivirus genomics database, namely gb4gv, that not only preserves genomic characteristics of the virus, but also supplements biologically relevant annotations that help to interrogate this virus, for example, the targeted host, putative iterons, siRNA targets, etc. Methods We have employed manual and automatic methods to curate 508 genomes from four major genera of Geminiviridae, and 161 associated satellites obtained from NCBI RefSeq and PubMed databases. Results These data are available for free access without registration from our website. Besides genomic content, our website provides visualization capability inherited from UCSC Genome Browser. Discussion With the genomic information readily accessible, we hope that our database

  6. Genome-wide screen in Saccharomyces cerevisiae identifies vacuolar protein sorting, autophagy, biosynthetic, and tRNA methylation genes involved in life span regulation.

    Directory of Open Access Journals (Sweden)

    Paola Fabrizio

    2010-07-01

    Full Text Available The study of the chronological life span of Saccharomyces cerevisiae, which measures the survival of populations of non-dividing yeast, has resulted in the identification of homologous genes and pathways that promote aging in organisms ranging from yeast to mammals. Using a competitive genome-wide approach, we performed a screen of a complete set of approximately 4,800 viable deletion mutants to identify genes that either increase or decrease chronological life span. Half of the putative short-/long-lived mutants retested from the primary screen were confirmed, demonstrating the utility of our approach. Deletion of genes involved in vacuolar protein sorting, autophagy, and mitochondrial function shortened life span, confirming that respiration and degradation processes are essential for long-term survival. Among the genes whose deletion significantly extended life span are ACB1, CKA2, and TRM9, implicated in fatty acid transport and biosynthesis, cell signaling, and tRNA methylation, respectively. Deletion of these genes conferred heat-shock resistance, supporting the link between life span extension and cellular protection observed in several model organisms. The high degree of conservation of these novel yeast longevity determinants in other species raises the possibility that their role in senescence might be conserved.

  7. A large set of newly created interspecific Saccharomyces hybrids increases aromatic diversity in lager beers.

    Science.gov (United States)

    Mertens, Stijn; Steensels, Jan; Saels, Veerle; De Rouck, Gert; Aerts, Guido; Verstrepen, Kevin J

    2015-12-01

    Lager beer is the most consumed alcoholic beverage in the world. Its production process is marked by a fermentation conducted at low (8 to 15°C) temperatures and by the use of Saccharomyces pastorianus, an interspecific hybrid between Saccharomyces cerevisiae and the cold-tolerant Saccharomyces eubayanus. Recent whole-genome-sequencing efforts revealed that the currently available lager yeasts belong to one of only two archetypes, "Saaz" and "Frohberg." This limited genetic variation likely reflects that all lager yeasts descend from only two separate interspecific hybridization events, which may also explain the relatively limited aromatic diversity between the available lager beer yeasts compared to, for example, wine and ale beer yeasts. In this study, 31 novel interspecific yeast hybrids were developed, resulting from large-scale robot-assisted selection and breeding between carefully selected strains of S. cerevisiae (six strains) and S. eubayanus (two strains). Interestingly, many of the resulting hybrids showed a broader temperature tolerance than their parental strains and reference S. pastorianus yeasts. Moreover, they combined a high fermentation capacity with a desirable aroma profile in laboratory-scale lager beer fermentations, thereby successfully enriching the currently available lager yeast biodiversity. Pilot-scale trials further confirmed the industrial potential of these hybrids and identified one strain, hybrid H29, which combines a fast fermentation, high attenuation, and the production of a complex, desirable fruity aroma. Copyright © 2015, American Society for Microbiology. All Rights Reserved.

  8. Whole-genome sequencing of a laboratory-evolved yeast strain

    Directory of Open Access Journals (Sweden)

    Dunham Maitreya J

    2010-02-01

    Full Text Available Abstract Background Experimental evolution of microbial populations provides a unique opportunity to study evolutionary adaptation in response to controlled selective pressures. However, until recently it has been difficult to identify the precise genetic changes underlying adaptation at a genome-wide scale. New DNA sequencing technologies now allow the genome of parental and evolved strains of microorganisms to be rapidly determined. Results We sequenced >93.5% of the genome of a laboratory-evolved strain of the yeast Saccharomyces cerevisiae and its ancestor at >28× depth. Both single nucleotide polymorphisms and copy number amplifications were found, with specific gains over array-based methodologies previously used to analyze these genomes. Applying a segmentation algorithm to quantify structural changes, we determined the approximate genomic boundaries of a 5× gene amplification. These boundaries guided the recovery of breakpoint sequences, which provide insights into the nature of a complex genomic rearrangement. Conclusions This study suggests that whole-genome sequencing can provide a rapid approach to uncover the genetic basis of evolutionary adaptations, with further applications in the study of laboratory selections and mutagenesis screens. In addition, we show how single-end, short read sequencing data can provide detailed information about structural rearrangements, and generate predictions about the genomic features and processes that underlie genome plasticity.

  9. REDIdb: the RNA editing database.

    Science.gov (United States)

    Picardi, Ernesto; Regina, Teresa Maria Rosaria; Brennicke, Axel; Quagliariello, Carla

    2007-01-01

    The RNA Editing Database (REDIdb) is an interactive, web-based database created and designed with the aim to allocate RNA editing events such as substitutions, insertions and deletions occurring in a wide range of organisms. The database contains both fully and partially sequenced DNA molecules for which editing information is available either by experimental inspection (in vitro) or by computational detection (in silico). Each record of REDIdb is organized in a specific flat-file containing a description of the main characteristics of the entry, a feature table with the editing events and related details and a sequence zone with both the genomic sequence and the corresponding edited transcript. REDIdb is a relational database in which the browsing and identification of editing sites has been simplified by means of two facilities to either graphically display genomic or cDNA sequences or to show the corresponding alignment. In both cases, all editing sites are highlighted in colour and their relative positions are detailed by mousing over. New editing positions can be directly submitted to REDIdb after a user-specific registration to obtain authorized secure access. This first version of REDIdb database stores 9964 editing events and can be freely queried at http://biologia.unical.it/py_script/search.html.

  10. Saccharomyces cerevisiae var. boulardii fungemia following probiotic treatment

    Directory of Open Access Journals (Sweden)

    Marcelo C. Appel-da-Silva

    2017-12-01

    Full Text Available Probiotics are commonly prescribed as an adjuvant in the treatment of antibiotic-associated diarrhea caused by Clostridium difficile. We report the case of an immunocompromised 73-year-old patient on chemotherapy who developed Saccharomyces cerevisiae var. boulardii fungemia in a central venous catheter during treatment of antibiotic-associated pseudomembranous colitis with the probiotic Saccharomyces cerevisiae var. boulardii. Fungemia was resolved after interruption of probiotic administration without the need to replace the central venous line. Keywords: Saccharomyces, Probiotics, Fungemia, Critical illness, Clostridium difficile

  11. Genome-wide identification of Saccharomyces cerevisiae genes required for tolerance to acetic acid

    Directory of Open Access Journals (Sweden)

    Sá-Correia Isabel

    2010-10-01

    Full Text Available Abstract Background Acetic acid is a byproduct of Saccharomyces cerevisiae alcoholic fermentation. Together with high concentrations of ethanol and other toxic metabolites, acetic acid may contribute to fermentation arrest and reduced ethanol productivity. This weak acid is also a present in lignocellulosic hydrolysates, a highly interesting non-feedstock substrate in industrial biotechnology. Therefore, the better understanding of the molecular mechanisms underlying S. cerevisiae tolerance to acetic acid is essential for the rational selection of optimal fermentation conditions and the engineering of more robust industrial strains to be used in processes in which yeast is explored as cell factory. Results The yeast genes conferring protection against acetic acid were identified in this study at a genome-wide scale, based on the screening of the EUROSCARF haploid mutant collection for susceptibility phenotypes to this weak acid (concentrations in the range 70-110 mM, at pH 4.5. Approximately 650 determinants of tolerance to acetic acid were identified. Clustering of these acetic acid-resistance genes based on their biological function indicated an enrichment of genes involved in transcription, internal pH homeostasis, carbohydrate metabolism, cell wall assembly, biogenesis of mitochondria, ribosome and vacuole, and in the sensing, signalling and uptake of various nutrients in particular iron, potassium, glucose and amino acids. A correlation between increased resistance to acetic acid and the level of potassium in the growth medium was found. The activation of the Snf1p signalling pathway, involved in yeast response to glucose starvation, is demonstrated to occur in response to acetic acid stress but no evidence was obtained supporting the acetic acid-induced inhibition of glucose uptake. Conclusions Approximately 490 of the 650 determinants of tolerance to acetic acid identified in this work are implicated, for the first time, in tolerance to

  12. The Yeast Deletion Collection: A Decade of Functional Genomics

    Science.gov (United States)

    Giaever, Guri; Nislow, Corey

    2014-01-01

    The yeast deletion collections comprise >21,000 mutant strains that carry precise start-to-stop deletions of ∼6000 open reading frames. This collection includes heterozygous and homozygous diploids, and haploids of both MATa and MATα mating types. The yeast deletion collection, or yeast knockout (YKO) set, represents the first and only complete, systematically constructed deletion collection available for any organism. Conceived during the Saccharomyces cerevisiae sequencing project, work on the project began in 1998 and was completed in 2002. The YKO strains have been used in numerous laboratories in >1000 genome-wide screens. This landmark genome project has inspired development of numerous genome-wide technologies in organisms from yeast to man. Notable spinoff technologies include synthetic genetic array and HIPHOP chemogenomics. In this retrospective, we briefly describe the yeast deletion project and some of its most noteworthy biological contributions and the impact that these collections have had on the yeast research community and on genomics in general. PMID:24939991

  13. High-Resolution Replication Profiles Define the Stochastic Nature of Genome Replication Initiation and Termination

    Directory of Open Access Journals (Sweden)

    Michelle Hawkins

    2013-11-01

    Full Text Available Eukaryotic genome replication is stochastic, and each cell uses a different cohort of replication origins. We demonstrate that interpreting high-resolution Saccharomyces cerevisiae genome replication data with a mathematical model allows quantification of the stochastic nature of genome replication, including the efficiency of each origin and the distribution of termination events. Single-cell measurements support the inferred values for stochastic origin activation time. A strain, in which three origins were inactivated, confirmed that the distribution of termination events is primarily dictated by the stochastic activation time of origins. Cell-to-cell variability in origin activity ensures that termination events are widely distributed across virtually the whole genome. We propose that the heterogeneity in origin usage contributes to genome stability by limiting potentially deleterious events from accumulating at particular loci.

  14. Techno-politics of genomic nationalism: tracing genomics and its use in drug regulation in Japan and Taiwan.

    Science.gov (United States)

    Kuo, Wen-Hua

    2011-10-01

    This paper compares the development of genomics as a form of state project in Japan and Taiwan. Broadening the concepts of genomic sovereignty and bionationalism, I argue that the establishment and use of genomic databases vary according to techno-political context. While both Japan and Taiwan hold population-based databases to be necessary for scientific advance and competitiveness, they differ in how they have attempted to transform the information produced by databases into regulatory schemes for drug approval. The effectiveness of Taiwan's biobank is severely limited by the IRB reviewing process. By contrast, while updating its regulations for drug approval, Japan, is using pharmacogenomics to deal with matters relating to ethnic identity. By analysing genomic initiatives in the political context that nurtures them, this paper seeks to capture how global science and local societies interact and offers insight into the assessment of state-sponsored science in East Asia as they become transnational. Copyright © 2011 Elsevier Ltd. All rights reserved.

  15. A Guide to the PLAZA 3.0 Plant Comparative Genomic Database.

    Science.gov (United States)

    Vandepoele, Klaas

    2017-01-01

    PLAZA 3.0 is an online resource for comparative genomics and offers a versatile platform to study gene functions and gene families or to analyze genome organization and evolution in the green plant lineage. Starting from genome sequence information for over 35 plant species, precomputed comparative genomic data sets cover homologous gene families, multiple sequence alignments, phylogenetic trees, and genomic colinearity information within and between species. Complementary functional data sets, a Workbench, and interactive visualization tools are available through a user-friendly web interface, making PLAZA an excellent starting point to translate sequence or omics data sets into biological knowledge. PLAZA is available at http://bioinformatics.psb.ugent.be/plaza/ .

  16. A Mutation in PGM2 Causing Inefficient Galactose Metabolism in the Probiotic Yeast Saccharomyces boulardii.

    Science.gov (United States)

    Liu, Jing-Jing; Zhang, Guo-Chang; Kong, In Iok; Yun, Eun Ju; Zheng, Jia-Qi; Kweon, Dae-Hyuk; Jin, Yong-Su

    2018-05-15

    The probiotic yeast Saccharomyces boulardii has been extensively studied for the prevention and treatment of diarrheal diseases, and it is now commercially available in some countries. S. boulardii displays notable phenotypic characteristics, such as a high optimal growth temperature, high tolerance against acidic conditions, and the inability to form ascospores, which differentiate S. boulardii from Saccharomyces cerevisiae The majority of prior studies stated that S. boulardii exhibits sluggish or halted galactose utilization. Nonetheless, the molecular mechanisms underlying inefficient galactose uptake have yet to be elucidated. When the galactose utilization of a widely used S. boulardii strain, ATCC MYA-796, was examined under various culture conditions, the S. boulardii strain could consume galactose, but at a much lower rate than that of S. cerevisiae While all GAL genes were present in the S. boulardii genome, according to analysis of genomic sequencing data in a previous study, a point mutation (G1278A) in PGM2 , which codes for phosphoglucomutase, was identified in the genome of the S. boulardii strain. As the point mutation resulted in the truncation of the Pgm2 protein, which is known to play a pivotal role in galactose utilization, we hypothesized that the truncated Pgm2 might be associated with inefficient galactose metabolism. Indeed, complementation of S. cerevisiae PGM2 in S. boulardii restored galactose utilization. After reverting the point mutation to a full-length PGM2 in S. boulardii by Cas9-based genome editing, the growth rates of wild-type (with a truncated PGM2 gene) and mutant (with a full-length PGM2 ) strains with glucose or galactose as the carbon source were examined. As expected, the mutant (with a full-length PGM2 ) was able to ferment galactose faster than the wild-type strain. Interestingly, the mutant showed a lower growth rate than that of the wild-type strain on glucose at 37°C. Also, the wild-type strain was enriched in the

  17. SIGMA: A System for Integrative Genomic Microarray Analysis of Cancer Genomes

    Directory of Open Access Journals (Sweden)

    Davies Jonathan J

    2006-12-01

    Full Text Available Abstract Background The prevalence of high resolution profiling of genomes has created a need for the integrative analysis of information generated from multiple methodologies and platforms. Although the majority of data in the public domain are gene expression profiles, and expression analysis software are available, the increase of array CGH studies has enabled integration of high throughput genomic and gene expression datasets. However, tools for direct mining and analysis of array CGH data are limited. Hence, there is a great need for analytical and display software tailored to cross platform integrative analysis of cancer genomes. Results We have created a user-friendly java application to facilitate sophisticated visualization and analysis such as cross-tumor and cross-platform comparisons. To demonstrate the utility of this software, we assembled array CGH data representing Affymetrix SNP chip, Stanford cDNA arrays and whole genome tiling path array platforms for cross comparison. This cancer genome database contains 267 profiles from commonly used cancer cell lines representing 14 different tissue types. Conclusion In this study we have developed an application for the visualization and analysis of data from high resolution array CGH platforms that can be adapted for analysis of multiple types of high throughput genomic datasets. Furthermore, we invite researchers using array CGH technology to deposit both their raw and processed data, as this will be a continually expanding database of cancer genomes. This publicly available resource, the System for Integrative Genomic Microarray Analysis (SIGMA of cancer genomes, can be accessed at http://sigma.bccrc.ca.

  18. Glucose repression in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Kayikci, Omur; Nielsen, Jens

    2015-01-01

    Glucose is the primary source of energy for the budding yeast Saccharomyces cerevisiae. Although yeast cells can utilize a wide range of carbon sources, presence of glucose suppresses molecular activities involved in the use of alternate carbon sources as well as it represses respiration and gluc......Glucose is the primary source of energy for the budding yeast Saccharomyces cerevisiae. Although yeast cells can utilize a wide range of carbon sources, presence of glucose suppresses molecular activities involved in the use of alternate carbon sources as well as it represses respiration...

  19. Transcription activator-like effector nucleases mediated metabolic engineering for enhanced fatty acids production in Saccharomyces cerevisiae

    KAUST Repository

    Aouida, Mustapha; Li, Lixin; Mahjoub, Ali; Alshareef, Sahar; Ali, Zahir; Piatek, Agnieszka Anna; Mahfouz, Magdy M.

    2015-01-01

    Targeted engineering of microbial genomes holds much promise for diverse biotechnological applications. Transcription activator-like effector nucleases (TALENs) and clustered regularly interspaced short palindromic repeats/Cas9 systems are capable of efficiently editing microbial genomes, including that of Saccharomyces cerevisiae. Here, we demonstrate the use of TALENs to edit the genome of S.cerevisiae with the aim of inducing the overproduction of fatty acids. Heterodimeric TALENs were designed to simultaneously edit the FAA1 and FAA4 genes encoding acyl-CoA synthetases in S.cerevisiae. Functional yeast double knockouts generated using these TALENs over-produce large amounts of free fatty acids into the cell. This study demonstrates the use of TALENs for targeted engineering of yeast and demonstrates that this technology can be used to stimulate the enhanced production of free fatty acids, which are potential substrates for biofuel production. This proof-of-principle study extends the utility of TALENs as excellent genome editing tools and highlights their potential use for metabolic engineering of yeast and other organisms, such as microalgae and plants, for biofuel production. © 2015 The Society for Biotechnology, Japan.

  20. Transcription activator-like effector nucleases mediated metabolic engineering for enhanced fatty acids production in Saccharomyces cerevisiae

    KAUST Repository

    Aouida, Mustapha

    2015-04-01

    Targeted engineering of microbial genomes holds much promise for diverse biotechnological applications. Transcription activator-like effector nucleases (TALENs) and clustered regularly interspaced short palindromic repeats/Cas9 systems are capable of efficiently editing microbial genomes, including that of Saccharomyces cerevisiae. Here, we demonstrate the use of TALENs to edit the genome of S.cerevisiae with the aim of inducing the overproduction of fatty acids. Heterodimeric TALENs were designed to simultaneously edit the FAA1 and FAA4 genes encoding acyl-CoA synthetases in S.cerevisiae. Functional yeast double knockouts generated using these TALENs over-produce large amounts of free fatty acids into the cell. This study demonstrates the use of TALENs for targeted engineering of yeast and demonstrates that this technology can be used to stimulate the enhanced production of free fatty acids, which are potential substrates for biofuel production. This proof-of-principle study extends the utility of TALENs as excellent genome editing tools and highlights their potential use for metabolic engineering of yeast and other organisms, such as microalgae and plants, for biofuel production. © 2015 The Society for Biotechnology, Japan.

  1. Diversity and adaptive evolution of Saccharomyces wine yeast: a review

    Science.gov (United States)

    Marsit, Souhir; Dequin, Sylvie

    2015-01-01

    Saccharomyces cerevisiae and related species, the main workhorses of wine fermentation, have been exposed to stressful conditions for millennia, potentially resulting in adaptive differentiation. As a result, wine yeasts have recently attracted considerable interest for studying the evolutionary effects of domestication. The widespread use of whole-genome sequencing during the last decade has provided new insights into the biodiversity, population structure, phylogeography and evolutionary history of wine yeasts. Comparisons between S. cerevisiae isolates from various origins have indicated that a variety of mechanisms, including heterozygosity, nucleotide and structural variations, introgressions, horizontal gene transfer and hybridization, contribute to the genetic and phenotypic diversity of S. cerevisiae. This review will summarize the current knowledge on the diversity and evolutionary history of wine yeasts, focusing on the domestication fingerprints identified in these strains. PMID:26205244

  2. Evolutionarily conserved elements in vertebrate, insect, worm, and yeast genomes

    DEFF Research Database (Denmark)

    Siepel, Adam; Bejerano, Gill; Pedersen, Jakob Skou

    2005-01-01

    We have conducted a comprehensive search for conserved elements in vertebrate genomes, using genome-wide multiple alignments of five vertebrate species (human, mouse, rat, chicken, and Fugu rubripes). Parallel searches have been performed with multiple alignments of four insect species (three...... species of Drosophila and Anopheles gambiae), two species of Caenorhabditis, and seven species of Saccharomyces. Conserved elements were identified with a computer program called phastCons, which is based on a two-state phylogenetic hidden Markov model (phylo-HMM). PhastCons works by fitting a phylo......-HMM to the data by maximum likelihood, subject to constraints designed to calibrate the model across species groups, and then predicting conserved elements based on this model. The predicted elements cover roughly 3%-8% of the human genome (depending on the details of the calibration procedure) and substantially...

  3. Virus Database and Online Inquiry System Based on Natural Vectors.

    Science.gov (United States)

    Dong, Rui; Zheng, Hui; Tian, Kun; Yau, Shek-Chung; Mao, Weiguang; Yu, Wenping; Yin, Changchuan; Yu, Chenglong; He, Rong Lucy; Yang, Jie; Yau, Stephen St

    2017-01-01

    We construct a virus database called VirusDB (http://yaulab.math.tsinghua.edu.cn/VirusDB/) and an online inquiry system to serve people who are interested in viral classification and prediction. The database stores all viral genomes, their corresponding natural vectors, and the classification information of the single/multiple-segmented viral reference sequences downloaded from National Center for Biotechnology Information. The online inquiry system serves the purpose of computing natural vectors and their distances based on submitted genomes, providing an online interface for accessing and using the database for viral classification and prediction, and back-end processes for automatic and manual updating of database content to synchronize with GenBank. Submitted genomes data in FASTA format will be carried out and the prediction results with 5 closest neighbors and their classifications will be returned by email. Considering the one-to-one correspondence between sequence and natural vector, time efficiency, and high accuracy, natural vector is a significant advance compared with alignment methods, which makes VirusDB a useful database in further research.

  4. Genetic, genomic, and molecular tools for studying the protoploid yeast, L. waltii.

    Science.gov (United States)

    Di Rienzi, Sara C; Lindstrom, Kimberly C; Lancaster, Ragina; Rolczynski, Lisa; Raghuraman, M K; Brewer, Bonita J

    2011-02-01

    Sequencing of the yeast Kluyveromyces waltii (recently renamed Lachancea waltii) provided evidence of a whole genome duplication event in the lineage leading to the well-studied Saccharomyces cerevisiae. While comparative genomic analyses of these yeasts have proven to be extremely instructive in modeling the loss or maintenance of gene duplicates, experimental tests of the ramifications following such genome alterations remain difficult. To transform L. waltii from an organism of the computational comparative genomic literature into an organism of the functional comparative genomic literature, we have developed genetic, molecular and genomic tools for working with L. waltii. In particular, we have characterized basic properties of L. waltii (growth, ploidy, molecular karyotype, mating type and the sexual cycle), developed transformation, cell cycle arrest and synchronization protocols, and have created centromeric and non-centromeric vectors as well as a genome browser for L. waltii. We hope that these tools will be used by the community to follow up on the ideas generated by sequence data and lead to a greater understanding of eukaryotic biology and genome evolution. 2010 John Wiley & Sons, Ltd.

  5. Construction of a novel kind of expression plasmid by homologous recombination in Saccharomyces cerevisiae

    Institute of Scientific and Technical Information of China (English)

    CHEN; Xiangling

    2005-01-01

    (2): 91―96.[13]Hong, M., Sam, K., Peter, J. S. et al., Plasmid construction by homologous recombination in yeast, Gene, 1987, 58: 201―216.[14]Prado, F., Aguilera, A., New in-vivo cloning methods, methods by homologous recombination in yeast, Curr. Genet., 1994, 20: 180―183.[15]Jacques, D., DNA insertion system for complex yeast shuttle vectors, Curr. Genet., 1995, 27: 309―311.[16]Erik, D., Bruno, D., Mireille, D. et al., In vivo cloning by homologous recombination in yeast using a two-plasmid-based system, Yeast, 1995, 11: 629―640.[17]Kevin, R. O., Kham, T. V., Susan, M., Chris, P., Recombination-mediated PCR-directed plasmid construction in vivo in yeast, Nucleic Acids Res., 1997, 25(2): 451―452.[18]Falco, S. C., Li, Y. Y., James, R. B., David, B., Genetic properties of chromosomally integrated 2μ plasmid DNA in yeast, Cell, 1982, 29: 573―584.[19]Francesca, S. L., Kevtn, L., Michael, A. R., In vivo site-directed mutagenesis using oligonucleotides, Nature Biotechnology, 2001, 19: 773―776.[20]Chulman, J., Hyuck, K., Sangmee, A. J., In vivo site-directed mutagenesis of yeast plasmids using a three-fragment homologous recombination system, Biotechniques, 2002, 33(2): 288―294.[21]Wach, A., Brachat, A., Pohlmann, R., Philippsen, P., New heterologus modules for classical or PCR-based gene disruption in Saccharomyces cerevisiea, Yeast, 1994, 10: 1793―1808.[22]Lorenz, M. C., Muir, R. S., Lim, E. et al., Gene disruption with PCR products in Saccharomyces cerevisiae, Gene, 1995, 158: 113―117.[23]Bhargava, J., Direct cloning of genomic DNA by recombinogenic targeting method using a yeast-bacterial shuttle vector, pClasper, Genomics, 1999, 62: 285―288.[24]Sambrook, J., Fritsch, E. F., Maniatis, T., Molecular Cloning: A Laboratory Manual, New York: Cold Spring Harbor Laboratory Press, 1989.[25]Gietz, R. D., Schiestl, R. H., Williems, A. R. et al., Studies on the transformation of intact yeast cells by the LiAc/SS-DNA/PEG procedure, Yeast, 1995, 11(4): 355

  6. A geographically-diverse collection of 418 human gut microbiome pathway genome databases

    KAUST Repository

    Hahn, Aria S.

    2017-04-11

    Advances in high-throughput sequencing are reshaping how we perceive microbial communities inhabiting the human body, with implications for therapeutic interventions. Several large-scale datasets derived from hundreds of human microbiome samples sourced from multiple studies are now publicly available. However, idiosyncratic data processing methods between studies introduce systematic differences that confound comparative analyses. To overcome these challenges, we developed GutCyc, a compendium of environmental pathway genome databases (ePGDBs) constructed from 418 assembled human microbiome datasets using MetaPathways, enabling reproducible functional metagenomic annotation. We also generated metabolic network reconstructions for each metagenome using the Pathway Tools software, empowering researchers and clinicians interested in visualizing and interpreting metabolic pathways encoded by the human gut microbiome. For the first time, GutCyc provides consistent annotations and metabolic pathway predictions, making possible comparative community analyses between health and disease states in inflammatory bowel disease, Crohn’s disease, and type 2 diabetes. GutCyc data products are searchable online, or may be downloaded and explored locally using MetaPathways and Pathway Tools.

  7. Databases and web tools for cancer genomics study.

    Science.gov (United States)

    Yang, Yadong; Dong, Xunong; Xie, Bingbing; Ding, Nan; Chen, Juan; Li, Yongjun; Zhang, Qian; Qu, Hongzhu; Fang, Xiangdong

    2015-02-01

    Publicly-accessible resources have promoted the advance of scientific discovery. The era of genomics and big data has brought the need for collaboration and data sharing in order to make effective use of this new knowledge. Here, we describe the web resources for cancer genomics research and rate them on the basis of the diversity of cancer types, sample size, omics data comprehensiveness, and user experience. The resources reviewed include data repository and analysis tools; and we hope such introduction will promote the awareness and facilitate the usage of these resources in the cancer research community. Copyright © 2015 The Authors. Production and hosting by Elsevier Ltd.. All rights reserved.

  8. TheCellMap.org: A Web-Accessible Database for Visualizing and Mining the Global Yeast Genetic Interaction Network.

    Science.gov (United States)

    Usaj, Matej; Tan, Yizhao; Wang, Wen; VanderSluis, Benjamin; Zou, Albert; Myers, Chad L; Costanzo, Michael; Andrews, Brenda; Boone, Charles

    2017-05-05

    Providing access to quantitative genomic data is key to ensure large-scale data validation and promote new discoveries. TheCellMap.org serves as a central repository for storing and analyzing quantitative genetic interaction data produced by genome-scale Synthetic Genetic Array (SGA) experiments with the budding yeast Saccharomyces cerevisiae In particular, TheCellMap.org allows users to easily access, visualize, explore, and functionally annotate genetic interactions, or to extract and reorganize subnetworks, using data-driven network layouts in an intuitive and interactive manner. Copyright © 2017 Usaj et al.

  9. Global Metabolic Reconstruction and Metabolic Gene Evolution in the Cattle Genome

    Science.gov (United States)

    Kim, Woonsu; Park, Hyesun; Seo, Seongwon

    2016-01-01

    The sequence of cattle genome provided a valuable opportunity to systematically link genetic and metabolic traits of cattle. The objectives of this study were 1) to reconstruct genome-scale cattle-specific metabolic pathways based on the most recent and updated cattle genome build and 2) to identify duplicated metabolic genes in the cattle genome for better understanding of metabolic adaptations in cattle. A bioinformatic pipeline of an organism for amalgamating genomic annotations from multiple sources was updated. Using this, an amalgamated cattle genome database based on UMD_3.1, was created. The amalgamated cattle genome database is composed of a total of 33,292 genes: 19,123 consensus genes between NCBI and Ensembl databases, 8,410 and 5,493 genes only found in NCBI or Ensembl, respectively, and 266 genes from NCBI scaffolds. A metabolic reconstruction of the cattle genome and cattle pathway genome database (PGDB) was also developed using Pathway Tools, followed by an intensive manual curation. The manual curation filled or revised 68 pathway holes, deleted 36 metabolic pathways, and added 23 metabolic pathways. Consequently, the curated cattle PGDB contains 304 metabolic pathways, 2,460 reactions including 2,371 enzymatic reactions, and 4,012 enzymes. Furthermore, this study identified eight duplicated genes in 12 metabolic pathways in the cattle genome compared to human and mouse. Some of these duplicated genes are related with specific hormone biosynthesis and detoxifications. The updated genome-scale metabolic reconstruction is a useful tool for understanding biology and metabolic characteristics in cattle. There has been significant improvements in the quality of cattle genome annotations and the MetaCyc database. The duplicated metabolic genes in the cattle genome compared to human and mouse implies evolutionary changes in the cattle genome and provides a useful information for further research on understanding metabolic adaptations of cattle. PMID

  10. Heterologous expression of MlcE in Saccharomyces cerevisiae provides resistance to natural and semi-synthetic statins

    Directory of Open Access Journals (Sweden)

    Ana Ley

    2015-12-01

    Full Text Available Statins are inhibitors of 3-hydroxy-3-methylglutaryl coenzyme A reductase, the key enzyme in cholesterol biosynthesis. Their extensive use in treatment and prevention of cardiovascular diseases placed statins among the best selling drugs. Construction of Saccharomyces cerevisiae cell factory for the production of high concentrations of natural statins will require establishment of a non-destructive self-resistance mechanism to overcome the undesirable growth inhibition effects of statins. To establish active export of statins from yeast, and thereby detoxification, we integrated a putative efflux pump-encoding gene mlcE from the mevastatin-producing Penicillium citrinum into the S. cerevisiae genome. The resulting strain showed increased resistance to both natural statins (mevastatin and lovastatin and semi-synthetic statin (simvastatin when compared to the wild type strain. Expression of RFP-tagged mlcE showed that MlcE is localized to the yeast plasma and vacuolar membranes. We provide a possible engineering strategy for improvement of future yeast based production of natural and semi-synthetic statins. Keywords: Polyketide, Statins, Saccharomyces cerevisiae, Transport, Cell factory, Resistance

  11. Genetic relationship and biological status of the industrially important yeast Saccharomyces eubayanus Sampaio et al.

    Science.gov (United States)

    Naumov, G I

    2017-03-01

    The genomes of the recently discovered yeast Saccharomyces eubayanus and traditional S. cerevisiae are known to be found in the yeast S. pastorianus (syn. S. carlsbergensis), which are essential for brewing. The cryotolerant yeast S. bayanus var. uvarum is of great importance for production of some wines. Based on ascospore viability and meiotic recombination of the control parental markers in hybrids, we have shown that there is no complete interspecies post-zygotic isolation between the yeasts S. eubayanus, S. bayanus var. bayanus and S. bayanus var. uvarum. The genetic data presented indicate that all of the three taxa belong to the same species.

  12. A genome browser database for rice (Oryza sativa) and Chinese ...

    African Journals Online (AJOL)

    STORAGESEVER

    2009-10-19

    Oct 19, 2009 ... sativa) and Chinese cabbage (Brassica rapa) genomes. The genome ... tant staple food for a large part of the world's human population. .... some banding region for selection and the overview panel shows the location of ...

  13. BGDB: a database of bivalent genes.

    Science.gov (United States)

    Li, Qingyan; Lian, Shuabin; Dai, Zhiming; Xiang, Qian; Dai, Xianhua

    2013-01-01

    Bivalent gene is a gene marked with both H3K4me3 and H3K27me3 epigenetic modification in the same area, and is proposed to play a pivotal role related to pluripotency in embryonic stem (ES) cells. Identification of these bivalent genes and understanding their functions are important for further research of lineage specification and embryo development. So far, lots of genome-wide histone modification data were generated in mouse and human ES cells. These valuable data make it possible to identify bivalent genes, but no comprehensive data repositories or analysis tools are available for bivalent genes currently. In this work, we develop BGDB, the database of bivalent genes. The database contains 6897 bivalent genes in human and mouse ES cells, which are manually collected from scientific literature. Each entry contains curated information, including genomic context, sequences, gene ontology and other relevant information. The web services of BGDB database were implemented with PHP + MySQL + JavaScript, and provide diverse query functions. Database URL: http://dailab.sysu.edu.cn/bgdb/

  14. Biotechnology of non-Saccharomyces yeasts--the ascomycetes.

    Science.gov (United States)

    Johnson, Eric A

    2013-01-01

    Saccharomyces cerevisiae and several other yeast species are among the most important groups of biotechnological organisms. S. cerevisiae and closely related ascomycetous yeasts are the major producer of biotechnology products worldwide, exceeding other groups of industrial microorganisms in productivity and economic revenues. Traditional industrial attributes of the S. cerevisiae group include their primary roles in food fermentations such as beers, cider, wines, sake, distilled spirits, bakery products, cheese, sausages, and other fermented foods. Other long-standing industrial processes involving S. cerevisae yeasts are production of fuel ethanol, single-cell protein (SCP), feeds and fodder, industrial enzymes, and small molecular weight metabolites. More recently, non-Saccharomyces yeasts (non-conventional yeasts) have been utilized as industrial organisms for a variety of biotechnological roles. Non-Saccharomyces yeasts are increasingly being used as hosts for expression of proteins, biocatalysts and multi-enzyme pathways for the synthesis of fine chemicals and small molecular weight compounds of medicinal and nutritional importance. Non-Saccharomyces yeasts also have important roles in agriculture as agents of biocontrol, bioremediation, and as indicators of environmental quality. Several of these products and processes have reached commercial utility, while others are in advanced development. The objective of this mini-review is to describe processes currently used by industry and those in developmental stages and close to commercialization primarily from non-Saccharomyces yeasts with an emphasis on new opportunities. The utility of S. cerevisiae in heterologous production of selected products is also described.

  15. LCGbase: A Comprehensive Database for Lineage-Based Co-regulated Genes.

    Science.gov (United States)

    Wang, Dapeng; Zhang, Yubin; Fan, Zhonghua; Liu, Guiming; Yu, Jun

    2012-01-01

    Animal genes of different lineages, such as vertebrates and arthropods, are well-organized and blended into dynamic chromosomal structures that represent a primary regulatory mechanism for body development and cellular differentiation. The majority of genes in a genome are actually clustered, which are evolutionarily stable to different extents and biologically meaningful when evaluated among genomes within and across lineages. Until now, many questions concerning gene organization, such as what is the minimal number of genes in a cluster and what is the driving force leading to gene co-regulation, remain to be addressed. Here, we provide a user-friendly database-LCGbase (a comprehensive database for lineage-based co-regulated genes)-hosting information on evolutionary dynamics of gene clustering and ordering within animal kingdoms in two different lineages: vertebrates and arthropods. The database is constructed on a web-based Linux-Apache-MySQL-PHP framework and effective interactive user-inquiry service. Compared to other gene annotation databases with similar purposes, our database has three comprehensible advantages. First, our database is inclusive, including all high-quality genome assemblies of vertebrates and representative arthropod species. Second, it is human-centric since we map all gene clusters from other genomes in an order of lineage-ranks (such as primates, mammals, warm-blooded, and reptiles) onto human genome and start the database from well-defined gene pairs (a minimal cluster where the two adjacent genes are oriented as co-directional, convergent, and divergent pairs) to large gene clusters. Furthermore, users can search for any adjacent genes and their detailed annotations. Third, the database provides flexible parameter definitions, such as the distance of transcription start sites between two adjacent genes, which is extendable to genes that flanking the cluster across species. We also provide useful tools for sequence alignment, gene

  16. TabSQL: a MySQL tool to facilitate mapping user data to public databases.

    Science.gov (United States)

    Xia, Xiao-Qin; McClelland, Michael; Wang, Yipeng

    2010-06-23

    With advances in high-throughput genomics and proteomics, it is challenging for biologists to deal with large data files and to map their data to annotations in public databases. We developed TabSQL, a MySQL-based application tool, for viewing, filtering and querying data files with large numbers of rows. TabSQL provides functions for downloading and installing table files from public databases including the Gene Ontology database (GO), the Ensembl databases, and genome databases from the UCSC genome bioinformatics site. Any other database that provides tab-delimited flat files can also be imported. The downloaded gene annotation tables can be queried together with users' data in TabSQL using either a graphic interface or command line. TabSQL allows queries across the user's data and public databases without programming. It is a convenient tool for biologists to annotate and enrich their data.

  17. Improving ethanol fermentation performance of Saccharomyces cerevisiae in very high-gravity fermentation through chemical mutagenesis and meiotic recombination

    Energy Technology Data Exchange (ETDEWEB)

    Liu, Jing-Jing; Ding, Wen-Tao; Zhang, Guo-Chang; Wang, Jing-Yu [Tianjin Univ. (China). Dept. of Biochemical Engineering

    2011-08-15

    Genome shuffling is an efficient way to improve complex phenotypes under the control of multiple genes. For the improvement of strain's performance in very high-gravity (VHG) fermentation, we developed a new method of genome shuffling. A diploid ste2/ste2 strain was subjected to EMS (ethyl methanesulfonate) mutagenesis followed by meiotic recombination-mediated genome shuffling. The resulting haploid progenies were intrapopulation sterile and therefore haploid recombinant cells with improved phenotypes were directly selected under selection condition. In VHG fermentation, strain WS1D and WS5D obtained by this approach exhibited remarkably enhanced tolerance to ethanol and osmolarity, increased metabolic rate, and 15.12% and 15.59% increased ethanol yield compared to the starting strain W303D, respectively. These results verified the feasibility of the strain improvement strategy and suggested that it is a powerful and high throughput method for development of Saccharomyces cerevisiae strains with desired phenotypes that is complex and cannot be addressed with rational approaches. (orig.)

  18. MASiVEdb: the Sirevirus Plant Retrotransposon Database

    Directory of Open Access Journals (Sweden)

    Bousios Alexandros

    2012-04-01

    Full Text Available Abstract Background Sireviruses are an ancient genus of the Copia superfamily of LTR retrotransposons, and the only one that has exclusively proliferated within plant genomes. Based on experimental data and phylogenetic analyses, Sireviruses have successfully infiltrated many branches of the plant kingdom, extensively colonizing the genomes of grass species. Notably, it was recently shown that they have been a major force in the make-up and evolution of the maize genome, where they currently occupy ~21% of the nuclear content and ~90% of the Copia population. It is highly likely, therefore, that their life dynamics have been fundamental in the genome composition and organization of a plethora of plant hosts. To assist studies into their impact on plant genome evolution and also facilitate accurate identification and annotation of transposable elements in sequencing projects, we developed MASiVEdb (Mapping and Analysis of SireVirus Elements Database, a collective and systematic resource of Sireviruses in plants. Description Taking advantage of the increasing availability of plant genomic sequences, and using an updated version of MASiVE, an algorithm specifically designed to identify Sireviruses based on their highly conserved genome structure, we populated MASiVEdb (http://bat.infspire.org/databases/masivedb/ with data on 16,243 intact Sireviruses (total length >158Mb discovered in 11 fully-sequenced plant genomes. MASiVEdb is unlike any other transposable element database, providing a multitude of highly curated and detailed information on a specific genus across its hosts, such as complete set of coordinates, insertion age, and an analytical breakdown of the structure and gene complement of each element. All data are readily available through basic and advanced query interfaces, batch retrieval, and downloadable files. A purpose-built system is also offered for detecting and visualizing similarity between user sequences and Sireviruses, as

  19. Autism genetic database (AGD: a comprehensive database including autism susceptibility gene-CNVs integrated with known noncoding RNAs and fragile sites

    Directory of Open Access Journals (Sweden)

    Talebizadeh Zohreh

    2009-09-01

    Full Text Available Abstract Background Autism is a highly heritable complex neurodevelopmental disorder, therefore identifying its genetic basis has been challenging. To date, numerous susceptibility genes and chromosomal abnormalities have been reported in association with autism, but most discoveries either fail to be replicated or account for a small effect. Thus, in most cases the underlying causative genetic mechanisms are not fully understood. In the present work, the Autism Genetic Database (AGD was developed as a literature-driven, web-based, and easy to access database designed with the aim of creating a comprehensive repository for all the currently reported genes and genomic copy number variations (CNVs associated with autism in order to further facilitate the assessment of these autism susceptibility genetic factors. Description AGD is a relational database that organizes data resulting from exhaustive literature searches for reported susceptibility genes and CNVs associated with autism. Furthermore, genomic information about human fragile sites and noncoding RNAs was also downloaded and parsed from miRBase, snoRNA-LBME-db, piRNABank, and the MIT/ICBP siRNA database. A web client genome browser enables viewing of the features while a web client query tool provides access to more specific information for the features. When applicable, links to external databases including GenBank, PubMed, miRBase, snoRNA-LBME-db, piRNABank, and the MIT siRNA database are provided. Conclusion AGD comprises a comprehensive list of susceptibility genes and copy number variations reported to-date in association with autism, as well as all known human noncoding RNA genes and fragile sites. Such a unique and inclusive autism genetic database will facilitate the evaluation of autism susceptibility factors in relation to known human noncoding RNAs and fragile sites, impacting on human diseases. As a result, this new autism database offers a valuable tool for the research

  20. Saccharomyces cerevisiae and non-Saccharomyces yeasts in grape varieties of the São Francisco Valley

    Directory of Open Access Journals (Sweden)

    Camila M.P.B.S. de Ponzzes-Gomes

    2014-06-01

    Full Text Available The aims of this work was to characterise indigenous Saccharomyces cerevisiae strains in the naturally fermented juice of grape varieties Cabernet Sauvignon, Grenache, Tempranillo, Sauvignon Blanc and Verdejo used in the São Francisco River Valley, northeastern Brazil. In this study, 155 S. cerevisiae and 60 non-Saccharomyces yeasts were isolated and identified using physiological tests and sequencing of the D1/D2 domains of the large subunit of the rRNA gene. Among the non-Saccharomyces species, Rhodotorula mucilaginosa was the most common species, followed by Pichia kudriavzevii, Candida parapsilosis, Meyerozyma guilliermondii, Wickerhamomyces anomalus, Kloeckera apis, P. manshurica, C. orthopsilosis and C. zemplinina. The population counts of these yeasts ranged among 1.0 to 19 x 10(5 cfu/mL. A total of 155 isolates of S. cerevisiae were compared by mitochondrial DNA restriction analysis, and five molecular mitochondrial DNA restriction profiles were detected. Indigenous strains of S. cerevisiae isolated from grapes of the São Francisco Valley can be further tested as potential starters for wine production.

  1. Gains and Losses of Transcription Factor Binding Sites in Saccharomyces cerevisiae and Saccharomyces paradoxus

    Science.gov (United States)

    Schaefke, Bernhard; Wang, Tzi-Yuan; Wang, Chuen-Yi; Li, Wen-Hsiung

    2015-01-01

    Gene expression evolution occurs through changes in cis- or trans-regulatory elements or both. Interactions between transcription factors (TFs) and their binding sites (TFBSs) constitute one of the most important points where these two regulatory components intersect. In this study, we investigated the evolution of TFBSs in the promoter regions of different Saccharomyces strains and species. We divided the promoter of a gene into the proximal region and the distal region, which are defined, respectively, as the 200-bp region upstream of the transcription starting site and as the 200-bp region upstream of the proximal region. We found that the predicted TFBSs in the proximal promoter regions tend to be evolutionarily more conserved than those in the distal promoter regions. Additionally, Saccharomyces cerevisiae strains used in the fermentation of alcoholic drinks have experienced more TFBS losses than gains compared with strains from other environments (wild strains, laboratory strains, and clinical strains). We also showed that differences in TFBSs correlate with the cis component of gene expression evolution between species (comparing S. cerevisiae and its sister species Saccharomyces paradoxus) and within species (comparing two closely related S. cerevisiae strains). PMID:26220934

  2. HyCCAPP as a tool to characterize promoter DNA-protein interactions in Saccharomyces cerevisiae.

    Science.gov (United States)

    Guillen-Ahlers, Hector; Rao, Prahlad K; Levenstein, Mark E; Kennedy-Darling, Julia; Perumalla, Danu S; Jadhav, Avinash Y L; Glenn, Jeremy P; Ludwig-Kubinski, Amy; Drigalenko, Eugene; Montoya, Maria J; Göring, Harald H; Anderson, Corianna D; Scalf, Mark; Gildersleeve, Heidi I S; Cole, Regina; Greene, Alexandra M; Oduro, Akua K; Lazarova, Katarina; Cesnik, Anthony J; Barfknecht, Jared; Cirillo, Lisa A; Gasch, Audrey P; Shortreed, Michael R; Smith, Lloyd M; Olivier, Michael

    2016-06-01

    Currently available methods for interrogating DNA-protein interactions at individual genomic loci have significant limitations, and make it difficult to work with unmodified cells or examine single-copy regions without specific antibodies. In this study, we describe a physiological application of the Hybridization Capture of Chromatin-Associated Proteins for Proteomics (HyCCAPP) methodology we have developed. Both novel and known locus-specific DNA-protein interactions were identified at the ENO2 and GAL1 promoter regions of Saccharomyces cerevisiae, and revealed subgroups of proteins present in significantly different levels at the loci in cells grown on glucose versus galactose as the carbon source. Results were validated using chromatin immunoprecipitation. Overall, our analysis demonstrates that HyCCAPP is an effective and flexible technology that does not require specific antibodies nor prior knowledge of locally occurring DNA-protein interactions and can now be used to identify changes in protein interactions at target regions in the genome in response to physiological challenges. Copyright © 2016 Elsevier Inc. All rights reserved.

  3. Metabolic Engineering of Probiotic Saccharomyces boulardii

    OpenAIRE

    Liu, Jing-Jing; Kong, In Iok; Zhang, Guo-Chang; Jayakody, Lahiru N.; Kim, Heejin; Xia, Peng-Fei; Kwak, Suryang; Sung, Bong Hyun; Sohn, Jung-Hoon; Walukiewicz, Hanna E.; Rao, Christopher V.; Jin, Yong-Su

    2016-01-01

    Saccharomyces boulardii is a probiotic yeast that has been used for promoting gut health as well as preventing diarrheal diseases. This yeast not only exhibits beneficial phenotypes for gut health but also can stay longer in the gut than Saccharomyces cerevisiae. Therefore, S. boulardii is an attractive host for metabolic engineering to produce biomolecules of interest in the gut. However, the lack of auxotrophic strains with defined genetic backgrounds has hampered the use of this strain for...

  4. Effects of an unusual poison identify a lifespan role for Topoisomerase 2 in Saccharomyces cerevisiae.

    Science.gov (United States)

    Tombline, Gregory; Millen, Jonathan I; Polevoda, Bogdan; Rapaport, Matan; Baxter, Bonnie; Van Meter, Michael; Gilbertson, Matthew; Madrey, Joe; Piazza, Gary A; Rasmussen, Lynn; Wennerberg, Krister; White, E Lucile; Nitiss, John L; Goldfarb, David S

    2017-01-05

    A progressive loss of genome maintenance has been implicated as both a cause and consequence of aging. Here we present evidence supporting the hypothesis that an age-associated decay in genome maintenance promotes aging in Saccharomyces cerevisiae (yeast) due to an inability to sense or repair DNA damage by topoisomerase 2 (yTop2). We describe the characterization of LS1, identified in a high throughput screen for small molecules that shorten the replicative lifespan of yeast. LS1 accelerates aging without affecting proliferative growth or viability. Genetic and biochemical criteria reveal LS1 to be a weak Top2 poison. Top2 poisons induce the accumulation of covalent Top2-linked DNA double strand breaks that, if left unrepaired, lead to genome instability and death. LS1 is toxic to cells deficient in homologous recombination, suggesting that the damage it induces is normally mitigated by genome maintenance systems. The essential roles of yTop2 in proliferating cells may come with a fitness trade-off in older cells that are less able to sense or repair yTop2-mediated DNA damage. Consistent with this idea, cells live longer when yTop2 expression levels are reduced. These results identify intrinsic yTop2-mediated DNA damage as potentially manageable cause of aging.

  5. Toward genome-enabled mycology.

    Science.gov (United States)

    Hibbett, David S; Stajich, Jason E; Spatafora, Joseph W

    2013-01-01

    Genome-enabled mycology is a rapidly expanding field that is characterized by the pervasive use of genome-scale data and associated computational tools in all aspects of fungal biology. Genome-enabled mycology is integrative and often requires teams of researchers with diverse skills in organismal mycology, bioinformatics and molecular biology. This issue of Mycologia presents the first complete fungal genomes in the history of the journal, reflecting the ongoing transformation of mycology into a genome-enabled science. Here, we consider the prospects for genome-enabled mycology and the technical and social challenges that will need to be overcome to grow the database of complete fungal genomes and enable all fungal biologists to make use of the new data.

  6. SoyDB: a knowledge database of soybean transcription factors

    Directory of Open Access Journals (Sweden)

    Valliyodan Babu

    2010-01-01

    Full Text Available Abstract Background Transcription factors play the crucial rule of regulating gene expression and influence almost all biological processes. Systematically identifying and annotating transcription factors can greatly aid further understanding their functions and mechanisms. In this article, we present SoyDB, a user friendly database containing comprehensive knowledge of soybean transcription factors. Description The soybean genome was recently sequenced by the Department of Energy-Joint Genome Institute (DOE-JGI and is publicly available. Mining of this sequence identified 5,671 soybean genes as putative transcription factors. These genes were comprehensively annotated as an aid to the soybean research community. We developed SoyDB - a knowledge database for all the transcription factors in the soybean genome. The database contains protein sequences, predicted tertiary structures, putative DNA binding sites, domains, homologous templates in the Protein Data Bank (PDB, protein family classifications, multiple sequence alignments, consensus protein sequence motifs, web logo of each family, and web links to the soybean transcription factor database PlantTFDB, known EST sequences, and other general protein databases including Swiss-Prot, Gene Ontology, KEGG, EMBL, TAIR, InterPro, SMART, PROSITE, NCBI, and Pfam. The database can be accessed via an interactive and convenient web server, which supports full-text search, PSI-BLAST sequence search, database browsing by protein family, and automatic classification of a new protein sequence into one of 64 annotated transcription factor families by hidden Markov models. Conclusions A comprehensive soybean transcription factor database was constructed and made publicly accessible at http://casp.rnet.missouri.edu/soydb/.

  7. A database and API for variation, dense genotyping and resequencing data

    Directory of Open Access Journals (Sweden)

    Flicek Paul

    2010-05-01

    Full Text Available Abstract Background Advances in sequencing and genotyping technologies are leading to the widespread availability of multi-species variation data, dense genotype data and large-scale resequencing projects. The 1000 Genomes Project and similar efforts in other species are challenging the methods previously used for storage and manipulation of such data necessitating the redesign of existing genome-wide bioinformatics resources. Results Ensembl has created a database and software library to support data storage, analysis and access to the existing and emerging variation data from large mammalian and vertebrate genomes. These tools scale to thousands of individual genome sequences and are integrated into the Ensembl infrastructure for genome annotation and visualisation. The database and software system is easily expanded to integrate both public and non-public data sources in the context of an Ensembl software installation and is already being used outside of the Ensembl project in a number of database and application environments. Conclusions Ensembl's powerful, flexible and open source infrastructure for the management of variation, genotyping and resequencing data is freely available at http://www.ensembl.org.

  8. Comparative genomics of xylose-fermenting fungi for enhanced biofuel production

    Energy Technology Data Exchange (ETDEWEB)

    Wohlbach, Dana J.; Kuo, Alan; Sato, Trey K.; Potts, Katlyn M.; Salamov, Asaf A.; LaButti, Kurt M.; Sun, Hui; Clum, Alicia; Pangilinan, Jasmyn L.; Lindquist, Erika A.; Lucas, Susan; Lapidus, Alla; Jin, Mingjie; Gunawan, Christa; Balan, Venkatesh; Dale, Bruce E.; Jeffries, Thomas W.; Zinkel, Robert; Barry, Kerrie W.; Grigoriev, Igor V.; Gasch, Audrey P.

    2011-02-24

    Cellulosic biomass is an abundant and underused substrate for biofuel production. The inability of many microbes to metabolize the pentose sugars abundant within hemicellulose creates specific challenges for microbial biofuel production from cellulosic material. Although engineered strains of Saccharomyces cerevisiae can use the pentose xylose, the fermentative capacity pales in comparison with glucose, limiting the economic feasibility of industrial fermentations. To better understand xylose utilization for subsequent microbial engineering, we sequenced the genomes of two xylose-fermenting, beetle-associated fungi, Spathaspora passalidarum and Candida tenuis. To identify genes involved in xylose metabolism, we applied a comparative genomic approach across 14 Ascomycete genomes, mapping phenotypes and genotypes onto the fungal phylogeny, and measured genomic expression across five Hemiascomycete species with different xylose-consumption phenotypes. This approach implicated many genes and processes involved in xylose assimilation. Several of these genes significantly improved xylose utilization when engineered into S. cerevisiae, demonstrating the power of comparative methods in rapidly identifying genes for biomass conversion while reflecting on fungal ecology.

  9. MALDI-TOF MS typing enables the classification of brewing yeasts of the genus Saccharomyces to major beer styles.

    Science.gov (United States)

    Lauterbach, Alexander; Usbeck, Julia C; Behr, Jürgen; Vogel, Rudi F

    2017-01-01

    Brewing yeasts of the genus Saccharomyces are either available from yeast distributor centers or from breweries employing their own "in-house strains". During the last years, the classification and characterization of yeasts of the genus Saccharomyces was achieved by using biochemical and DNA-based methods. The current lack of fast, cost-effective and simple methods to classify brewing yeasts to a beer type, may be closed by Matrix Assisted Laser Desorption/Ionization-Time-Of-Flight Mass Spectrometry (MALDI-TOF MS) upon establishment of a database based on sub-proteome spectra from reference strains of brewing yeasts. In this study an extendable "brewing yeast" spectra database was established including 52 brewing yeast strains of the most important types of bottom- and top-fermenting strains as well as beer-spoiling S. cerevisiae var. diastaticus strains. 1560 single spectra, prepared with a standardized sample preparation method, were finally compared against the established database and investigated by bioinformatic analyses for similarities and distinctions. A 100% separation between bottom-, top-fermenting and S. cerevisiae var. diastaticus strains was achieved. Differentiation between Alt and Kölsch strains was not achieved because of the high similarity of their protein patterns. Whereas the Ale strains show a high degree of dissimilarity with regard to their sub-proteome. These results were supported by MDS and DAPC analysis of all recorded spectra. Within five clusters of beer types that were distinguished, and the wheat beer (WB) cluster has a clear separation from other groups. With the establishment of this MALDI-TOF MS spectra database proof of concept is provided of the discriminatory power of this technique to classify brewing yeasts into different major beer types in a rapid, easy way, and focus brewing trails accordingly. It can be extended to yeasts for specialty beer types and other applications including wine making or baking.

  10. MALDI-TOF MS typing enables the classification of brewing yeasts of the genus Saccharomyces to major beer styles

    Science.gov (United States)

    Lauterbach, Alexander; Usbeck, Julia C.; Behr, Jürgen

    2017-01-01

    Brewing yeasts of the genus Saccharomyces are either available from yeast distributor centers or from breweries employing their own “in-house strains”. During the last years, the classification and characterization of yeasts of the genus Saccharomyces was achieved by using biochemical and DNA-based methods. The current lack of fast, cost-effective and simple methods to classify brewing yeasts to a beer type, may be closed by Matrix Assisted Laser Desorption/Ionization–Time-Of-Flight Mass Spectrometry (MALDI-TOF MS) upon establishment of a database based on sub-proteome spectra from reference strains of brewing yeasts. In this study an extendable “brewing yeast” spectra database was established including 52 brewing yeast strains of the most important types of bottom- and top-fermenting strains as well as beer-spoiling S. cerevisiae var. diastaticus strains. 1560 single spectra, prepared with a standardized sample preparation method, were finally compared against the established database and investigated by bioinformatic analyses for similarities and distinctions. A 100% separation between bottom-, top-fermenting and S. cerevisiae var. diastaticus strains was achieved. Differentiation between Alt and Kölsch strains was not achieved because of the high similarity of their protein patterns. Whereas the Ale strains show a high degree of dissimilarity with regard to their sub-proteome. These results were supported by MDS and DAPC analysis of all recorded spectra. Within five clusters of beer types that were distinguished, and the wheat beer (WB) cluster has a clear separation from other groups. With the establishment of this MALDI-TOF MS spectra database proof of concept is provided of the discriminatory power of this technique to classify brewing yeasts into different major beer types in a rapid, easy way, and focus brewing trails accordingly. It can be extended to yeasts for specialty beer types and other applications including wine making or baking. PMID

  11. MALDI-TOF MS typing enables the classification of brewing yeasts of the genus Saccharomyces to major beer styles.

    Directory of Open Access Journals (Sweden)

    Alexander Lauterbach

    Full Text Available Brewing yeasts of the genus Saccharomyces are either available from yeast distributor centers or from breweries employing their own "in-house strains". During the last years, the classification and characterization of yeasts of the genus Saccharomyces was achieved by using biochemical and DNA-based methods. The current lack of fast, cost-effective and simple methods to classify brewing yeasts to a beer type, may be closed by Matrix Assisted Laser Desorption/Ionization-Time-Of-Flight Mass Spectrometry (MALDI-TOF MS upon establishment of a database based on sub-proteome spectra from reference strains of brewing yeasts. In this study an extendable "brewing yeast" spectra database was established including 52 brewing yeast strains of the most important types of bottom- and top-fermenting strains as well as beer-spoiling S. cerevisiae var. diastaticus strains. 1560 single spectra, prepared with a standardized sample preparation method, were finally compared against the established database and investigated by bioinformatic analyses for similarities and distinctions. A 100% separation between bottom-, top-fermenting and S. cerevisiae var. diastaticus strains was achieved. Differentiation between Alt and Kölsch strains was not achieved because of the high similarity of their protein patterns. Whereas the Ale strains show a high degree of dissimilarity with regard to their sub-proteome. These results were supported by MDS and DAPC analysis of all recorded spectra. Within five clusters of beer types that were distinguished, and the wheat beer (WB cluster has a clear separation from other groups. With the establishment of this MALDI-TOF MS spectra database proof of concept is provided of the discriminatory power of this technique to classify brewing yeasts into different major beer types in a rapid, easy way, and focus brewing trails accordingly. It can be extended to yeasts for specialty beer types and other applications including wine making or baking.

  12. Final Technical Report on the Genome Sequence DataBase (GSDB): DE-FG03 95 ER 62062 September 1997-September 1999

    Energy Technology Data Exchange (ETDEWEB)

    Harger, Carol A.

    1999-10-28

    Since September 1997 NCGR has produced two web-based tools for researchers to use to access and analyze data in the Genome Sequence DataBase (GSDB). These tools are: Sequence Viewer, a nucleotide sequence and annotation visualization tool, and MAR-Finder, a tool that predicts, base upon statistical inferences, the location of matrix attachment regions (MARS) within a nucleotide sequence. [The annual report for June 1996 to August 1997 is included as an attachment to this final report.

  13. CyanoClust: comparative genome resources of cyanobacteria and plastids

    OpenAIRE

    Sasaki, Naobumi V.; Sato, Naoki

    2010-01-01

    Cyanobacteria, which perform oxygen-evolving photosynthesis as do chloroplasts of plants and algae, are one of the best-studied prokaryotic phyla and one from which many representative genomes have been sequenced. Lack of a suitable comparative genomic database has been a problem in cyanobacterial genomics because many proteins involved in physiological functions such as photosynthesis and nitrogen fixation are not catalogued in commonly used databases, such as Clusters of Orthologous Protein...

  14. Induction of different types of mutations in yeast Saccharomyces serevisiae by γ-radiation

    International Nuclear Information System (INIS)

    Lyubimova, K.A.; Shvaneva, N.V.; Koltovaya, N.A.

    2005-01-01

    Several tester systems were used to study a wide spectrum of genetic changes induced by γ-radiation in the yeast Saccharomyces cerevisiae. The tester systems allow one to identify a loss of chromosomes, recombination (crossing over) and point mutations (frame shifts and base-pair substitutions.) Large genome changes were induced by γ-rays more efficiently than the point mutations. The dose dependence of the point mutations frequency was linear. Spontaneous and induced mutation rates per base pair corresponded with the known literature data for the same tester systems. Our finding shows that the used tester systems are not specific. They are useful for further study of mutations induced by ionizing radiation with various physical characteristics

  15. Genome-wide data-mining of candidate human splice translational efficiency polymorphisms (STEPs and an online database.

    Directory of Open Access Journals (Sweden)

    Christopher A Raistrick

    2010-10-01

    Full Text Available Variation in pre-mRNA splicing is common and in some cases caused by genetic variants in intronic splicing motifs. Recent studies into the insulin gene (INS discovered a polymorphism in a 5' non-coding intron that influences the likelihood of intron retention in the final mRNA, extending the 5' untranslated region and maintaining protein quality. Retention was also associated with increased insulin levels, suggesting that such variants--splice translational efficiency polymorphisms (STEPs--may relate to disease phenotypes through differential protein expression. We set out to explore the prevalence of STEPs in the human genome and validate this new category of protein quantitative trait loci (pQTL using publicly available data.Gene transcript and variant data were collected and mined for candidate STEPs in motif regions. Sequences from transcripts containing potential STEPs were analysed for evidence of splice site recognition and an effect in expressed sequence tags (ESTs. 16 publicly released genome-wide association data sets of common diseases were searched for association to candidate polymorphisms with HapMap frequency data. Our study found 3324 candidate STEPs lying in motif sequences of 5' non-coding introns and further mining revealed 170 with transcript evidence of intron retention. 21 potential STEPs had EST evidence of intron retention or exon extension, as well as population frequency data for comparison.Results suggest that the insulin STEP was not a unique example and that many STEPs may occur genome-wide with potentially causal effects in complex disease. An online database of STEPs is freely accessible at http://dbstep.genes.org.uk/.

  16. Engineering and Evolution of Saccharomyces cerevisiae to Produce Biofuels and Chemicals.

    Science.gov (United States)

    Turner, Timothy L; Kim, Heejin; Kong, In Iok; Liu, Jing-Jing; Zhang, Guo-Chang; Jin, Yong-Su

    To mitigate global climate change caused partly by the use of fossil fuels, the production of fuels and chemicals from renewable biomass has been attempted. The conversion of various sugars from renewable biomass into biofuels by engineered baker's yeast (Saccharomyces cerevisiae) is one major direction which has grown dramatically in recent years. As well as shifting away from fossil fuels, the production of commodity chemicals by engineered S. cerevisiae has also increased significantly. The traditional approaches of biochemical and metabolic engineering to develop economic bioconversion processes in laboratory and industrial settings have been accelerated by rapid advancements in the areas of yeast genomics, synthetic biology, and systems biology. Together, these innovations have resulted in rapid and efficient manipulation of S. cerevisiae to expand fermentable substrates and diversify value-added products. Here, we discuss recent and major advances in rational (relying on prior experimentally-derived knowledge) and combinatorial (relying on high-throughput screening and genomics) approaches to engineer S. cerevisiae for producing ethanol, butanol, 2,3-butanediol, fatty acid ethyl esters, isoprenoids, organic acids, rare sugars, antioxidants, and sugar alcohols from glucose, xylose, cellobiose, galactose, acetate, alginate, mannitol, arabinose, and lactose.

  17. Mathematical Analysis of Genomic Evolution

    Directory of Open Access Journals (Sweden)

    Cedric Green

    2011-01-01

    Full Text Available Changes in nucleotide sequences, or mutations, accumulate from generation to generation in the genomes of all living organisms. The mutations can be advantageous, deleterious, or neutral. The goal of this project is to determine the amount of advantageous mutations it takes to get human (Homo sapiens DNA from the DNA of genetically distinct organisms. We do this by collecting the genomic data of such organisms, and estimating the amount of mutations it takes to transform yeast (Saccharomyces cerevisiae DNA to the DNA of a human. We calculate the typical number of mutations occurring annually through the organism's average life span and the average mutation rate. This allows us to determine the total number of mutations as well as the probability of advantageous mutations. Not surprisingly, this probability proves to be fairly small. A more precise estimate can be determined by accounting for the differences in the chromosomal structure and phenomena like horizontal gene transfer.

  18. Highly variable rates of genome rearrangements between hemiascomycetous yeast lineages.

    Directory of Open Access Journals (Sweden)

    2006-03-01

    Full Text Available Hemiascomycete yeasts cover an evolutionary span comparable to that of the entire phylum of chordates. Since this group currently contains the largest number of complete genome sequences it presents unique opportunities to understand the evolution of genome organization in eukaryotes. We inferred rates of genome instability on all branches of a phylogenetic tree for 11 species and calculated species-specific rates of genome rearrangements. We characterized all inversion events that occurred within synteny blocks between six representatives of the different lineages. We show that the rates of macro- and microrearrangements of gene order are correlated within individual lineages but are highly variable across different lineages. The most unstable genomes correspond to the pathogenic yeasts Candida albicans and Candida glabrata. Chromosomal maps have been intensively shuffled by numerous interchromosomal rearrangements, even between species that have retained a very high physical fraction of their genomes within small synteny blocks. Despite this intensive reshuffling of gene positions, essential genes, which cluster in low recombination regions in the genome of Saccharomyces cerevisiae, tend to remain syntenic during evolution. This work reveals that the high plasticity of eukaryotic genomes results from rearrangement rates that vary between lineages but also at different evolutionary times of a given lineage.

  19. Mining biological databases for candidate disease genes

    Science.gov (United States)

    Braun, Terry A.; Scheetz, Todd; Webster, Gregg L.; Casavant, Thomas L.

    2001-07-01

    The publicly-funded effort to sequence the complete nucleotide sequence of the human genome, the Human Genome Project (HGP), has currently produced more than 93% of the 3 billion nucleotides of the human genome into a preliminary `draft' format. In addition, several valuable sources of information have been developed as direct and indirect results of the HGP. These include the sequencing of model organisms (rat, mouse, fly, and others), gene discovery projects (ESTs and full-length), and new technologies such as expression analysis and resources (micro-arrays or gene chips). These resources are invaluable for the researchers identifying the functional genes of the genome that transcribe and translate into the transcriptome and proteome, both of which potentially contain orders of magnitude more complexity than the genome itself. Preliminary analyses of this data identified approximately 30,000 - 40,000 human `genes.' However, the bulk of the effort still remains -- to identify the functional and structural elements contained within the transcriptome and proteome, and to associate function in the transcriptome and proteome to genes. A fortuitous consequence of the HGP is the existence of hundreds of databases containing biological information that may contain relevant data pertaining to the identification of disease-causing genes. The task of mining these databases for information on candidate genes is a commercial application of enormous potential. We are developing a system to acquire and mine data from specific databases to aid our efforts to identify disease genes. A high speed cluster of Linux of workstations is used to analyze sequence and perform distributed sequence alignments as part of our data mining and processing. This system has been used to mine GeneMap99 sequences within specific genomic intervals to identify potential candidate disease genes associated with Bardet-Biedle Syndrome (BBS).

  20. Systems Biology of Saccharomyces cerevisiae Physiology and its DNA Damage Response

    DEFF Research Database (Denmark)

    Fazio, Alessandro

    The yeast Saccharomyces cerevisiae is a model organism in biology, being widely used in fundamental research, the first eukaryotic organism to be fully sequenced and the platform for the development of many genomics techniques. Therefore, it is not surprising that S. cerevisiae has also been widely...... used in the field of systems biology during the last decade. This thesis investigates S. cerevisiae growth physiology and DNA damage response by using a systems biology approach. Elucidation of the relationship between growth rate and gene expression is important to understand the mechanisms regulating...... set of growth dependent genes by using a multi-factorial experimental design. Moreover, new insights into the metabolic response and transcriptional regulation of these genes have been provided by using systems biology tools (Chapter 3). One of the prerequisite of systems biology should...

  1. A Saccharomyces cerevisiae mitochondrial DNA fragment activates Reg1p-dependent glucose-repressible transcription in the nucleus.

    Science.gov (United States)

    Santangelo, G M; Tornow, J

    1997-12-01

    As part of an effort to identify random carbon-source-regulated promoters in the Saccharomyces cerevisiae genome, we discovered that a mitochondrial DNA fragment is capable of directing glucose-repressible expression of a reporter gene. This fragment (CR24) originated from the mitochondrial genome adjacent to a transcription initiation site. Mutational analyses identified a GC cluster within the fragment that is required for transcriptional induction. Repression of nuclear CR24-driven transcription required Reg1p, indicating that this mitochondrially derived promoter is a member of a large group of glucose-repressible nuclear promoters that are similarly regulated by Reg1p. In vivo and in vitro binding assays indicated the presence of factors, located within the nucleus and the mitochondria, that bind to the GC cluster. One or more of these factors may provide a regulatory link between the nucleus and mitochondria.

  2. Review of Saccharomyces boulardii as a treatment option in IBD

    DEFF Research Database (Denmark)

    Sivananthan, Kavitha; Petersen, Andreas Munk

    2018-01-01

    CONTEXT: Review of the yeast Saccharomyces boulardii as a treatment option for the inflammatory bowel diseases (IBD) ulcerative colitis and Crohn's disease. OBJECTIVE: IBD is caused by an inappropriate immune response to gut microbiota. Treatment options could therefore be prebiotics, probiotics......, antibiotics and/or fecal transplant. In this review, we have looked at the evidence for the yeast S. boulardii as a treatment option. MATERIAL AND METHODS: Searches in PubMed and the Cochrane Library with the MeSH words 'Saccharomyces boulardii AND IBD', 'Saccharomyces boulardii AND Inflammatory Bowel Disease....... Saccharomyces boulardii is, however, a plausible treatment option in the future, but more placebo-controlled clinical studies on both patients with ulcerative colitis and Crohn's disease are needed....

  3. Gains and Losses of Transcription Factor Binding Sites in Saccharomyces cerevisiae and Saccharomyces paradoxus.

    Science.gov (United States)

    Schaefke, Bernhard; Wang, Tzi-Yuan; Wang, Chuen-Yi; Li, Wen-Hsiung

    2015-07-27

    Gene expression evolution occurs through changes in cis- or trans-regulatory elements or both. Interactions between transcription factors (TFs) and their binding sites (TFBSs) constitute one of the most important points where these two regulatory components intersect. In this study, we investigated the evolution of TFBSs in the promoter regions of different Saccharomyces strains and species. We divided the promoter of a gene into the proximal region and the distal region, which are defined, respectively, as the 200-bp region upstream of the transcription starting site and as the 200-bp region upstream of the proximal region. We found that the predicted TFBSs in the proximal promoter regions tend to be evolutionarily more conserved than those in the distal promoter regions. Additionally, Saccharomyces cerevisiae strains used in the fermentation of alcoholic drinks have experienced more TFBS losses than gains compared with strains from other environments (wild strains, laboratory strains, and clinical strains). We also showed that differences in TFBSs correlate with the cis component of gene expression evolution between species (comparing S. cerevisiae and its sister species Saccharomyces paradoxus) and within species (comparing two closely related S. cerevisiae strains). © The Author(s) 2015. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

  4. Tolerance to winemaking stress conditions of Patagonian strains of Saccharomyces eubayanus and Saccharomyces uvarum.

    Science.gov (United States)

    Origone, A C; Del Mónaco, S M; Ávila, J R; González Flores, M; Rodríguez, M E; Lopes, C A

    2017-08-01

    Evaluating the winemaking stress tolerance of a set of both Saccharomyces eubayanus and Saccharomyces uvarum strains from diverse Patagonian habitats. Yeast strains growth was analysed under increasing ethanol concentrations; all of them were able to grow until 8% v/v ethanol. The effect of different temperature and pH conditions as well as at SO 2 and hexose concentrations was evaluated by means of a central composite experimental design. Only two S. uvarum strains (NPCC 1289 and 1321) were able to grow in most stress conditions. Kinetic parameters analysed (μ max and λ) were statistically affected by temperature, pH and SO 2 , but not influenced by sugar concentration. The obtained growth model was used for predicting optimal growth conditions for both strains: 20°C, 0% w/v SO 2 and pH 4·5. Strains from human-associated environments (chichas) presented the highest diversity in the response to different stress factors. Two S. uvarum strains from chichas demonstrated to be the most tolerant to winemaking conditions. This work evidenced the potential use of two S. uvarum yeast strains as starter cultures in wines fermented at low temperatures. Saccharomyces eubayanus was significantly affected by winemaking stress conditions, limiting its use in this industry. © 2017 The Society for Applied Microbiology.

  5. Specific transcripts are elevated in Saccharomyces cerevisiae in response to DNA damage

    International Nuclear Information System (INIS)

    McClanahan, T.; McEntee, K.

    1984-01-01

    Differential hybridization has been used to identify genes in Saccharomyces cerevisiae displaying increased transcript levels after treatment of cells with UV irradiation or with the mutagen/carcinogen 4-nitroquinoline-1-oxide (NQO). The authors describe the isolation and characterization of four DNA damage responsive genes obtained from screening ca. 9000 yeast genomic clones. Two of these clones, lambda 78A and pBR178C, contain repetitive elements in the yeast genome as shown by Southern hybridization analysis. Although the genomic hybridization pattern is distinct for each of these two clones, both of these sequences hybridize to large polyadenylated transcripts ca. 5 kilobases in length. Two other DNA damage responsive sequences, pBRA2 and pBR3016B, are single-copy genes and hybridize to 0.5- and 3.2-kilobase transcripts, respectively. Kinetic analysis of the 0.5-kilobase transcript homologous to pBRA2 indicates that the level of this RNA increases more than 15-fold within 20 min after exposure to 4-nitroquinoline-1-oxide. Moreover, the level of this transcript is significantly elevated in cells containing the rad52-1 mutation which are deficient in DNA strand break repair and gene conversion. These results provide some of the first evidence that DNA damage stimulates transcription of specific genes in eucaryotic cells

  6. High-throughput transformation of Saccharomyces cerevisiae using liquid handling robots.

    Directory of Open Access Journals (Sweden)

    Guangbo Liu

    Full Text Available Saccharomyces cerevisiae (budding yeast is a powerful eukaryotic model organism ideally suited to high-throughput genetic analyses, which time and again has yielded insights that further our understanding of cell biology processes conserved in humans. Lithium Acetate (LiAc transformation of yeast with DNA for the purposes of exogenous protein expression (e.g., plasmids or genome mutation (e.g., gene mutation, deletion, epitope tagging is a useful and long established method. However, a reliable and optimized high throughput transformation protocol that runs almost no risk of human error has not been described in the literature. Here, we describe such a method that is broadly transferable to most liquid handling high-throughput robotic platforms, which are now commonplace in academic and industry settings. Using our optimized method, we are able to comfortably transform approximately 1200 individual strains per day, allowing complete transformation of typical genomic yeast libraries within 6 days. In addition, use of our protocol for gene knockout purposes also provides a potentially quicker, easier and more cost-effective approach to generating collections of double mutants than the popular and elegant synthetic genetic array methodology. In summary, our methodology will be of significant use to anyone interested in high throughput molecular and/or genetic analysis of yeast.

  7. Genomic Enzymology: Web Tools for Leveraging Protein Family Sequence-Function Space and Genome Context to Discover Novel Functions.

    Science.gov (United States)

    Gerlt, John A

    2017-08-22

    The exponentially increasing number of protein and nucleic acid sequences provides opportunities to discover novel enzymes, metabolic pathways, and metabolites/natural products, thereby adding to our knowledge of biochemistry and biology. The challenge has evolved from generating sequence information to mining the databases to integrating and leveraging the available information, i.e., the availability of "genomic enzymology" web tools. Web tools that allow identification of biosynthetic gene clusters are widely used by the natural products/synthetic biology community, thereby facilitating the discovery of novel natural products and the enzymes responsible for their biosynthesis. However, many novel enzymes with interesting mechanisms participate in uncharacterized small-molecule metabolic pathways; their discovery and functional characterization also can be accomplished by leveraging information in protein and nucleic acid databases. This Perspective focuses on two genomic enzymology web tools that assist the discovery novel metabolic pathways: (1) Enzyme Function Initiative-Enzyme Similarity Tool (EFI-EST) for generating sequence similarity networks to visualize and analyze sequence-function space in protein families and (2) Enzyme Function Initiative-Genome Neighborhood Tool (EFI-GNT) for generating genome neighborhood networks to visualize and analyze the genome context in microbial and fungal genomes. Both tools have been adapted to other applications to facilitate target selection for enzyme discovery and functional characterization. As the natural products community has demonstrated, the enzymology community needs to embrace the essential role of web tools that allow the protein and genome sequence databases to be leveraged for novel insights into enzymological problems.

  8. Fragile genomic sites are associated with origins of replication.

    Science.gov (United States)

    Di Rienzi, Sara C; Collingwood, David; Raghuraman, M K; Brewer, Bonita J

    2009-09-09

    Genome rearrangements are mediators of evolution and disease. Such rearrangements are frequently bounded by transfer RNAs (tRNAs), transposable elements, and other repeated elements, suggesting a functional role for these elements in creating or repairing breakpoints. Though not well explored, there is evidence that origins of replication also colocalize with breakpoints. To investigate a potential correlation between breakpoints and origins, we analyzed evolutionary breakpoints defined between Saccharomyces cerevisiae and Kluyveromyces waltii and S. cerevisiae and a hypothetical ancestor of both yeasts, as well as breakpoints reported in the experimental literature. We find that origins correlate strongly with both evolutionary breakpoints and those described in the literature. Specifically, we find that origins firing earlier in S phase are more strongly correlated with breakpoints than are later-firing origins. Despite origins being located in genomic regions also bearing tRNAs and Ty elements, the correlation we observe between origins and breakpoints appears to be independent of these genomic features. This study lays the groundwork for understanding the mechanisms by which origins of replication may impact genome architecture and disease.

  9. Fungal genome resources at NCBI

    Science.gov (United States)

    Robbertse, B.; Tatusova, T.

    2011-01-01

    The National Center for Biotechnology Information (NCBI) is well known for the nucleotide sequence archive, GenBank and sequence analysis tool BLAST. However, NCBI integrates many types of biomolecular data from variety of sources and makes it available to the scientific community as interactive web resources as well as organized releases of bulk data. These tools are available to explore and compare fungal genomes. Searching all databases with Fungi [organism] at http://www.ncbi.nlm.nih.gov/ is the quickest way to find resources of interest with fungal entries. Some tools though are resources specific and can be indirectly accessed from a particular database in the Entrez system. These include graphical viewers and comparative analysis tools such as TaxPlot, TaxMap and UniGene DDD (found via UniGene Homepage). Gene and BioProject pages also serve as portals to external data such as community annotation websites, BioGrid and UniProt. There are many different ways of accessing genomic data at NCBI. Depending on the focus and goal of research projects or the level of interest, a user would select a particular route for accessing genomic databases and resources. This review article describes methods of accessing fungal genome data and provides examples that illustrate the use of analysis tools. PMID:22737589

  10. Conducting Wine Symphonics with the Aid of Yeast Genomics

    Directory of Open Access Journals (Sweden)

    Isak S. Pretorius

    2016-12-01

    Full Text Available A perfectly balanced wine can be said to create a symphony in the mouth. To achieve the sublime, both in wine and music, requires imagination and skilled orchestration of artistic craftmanship. For wine, inventiveness starts in the vineyard. Similar to a composer of music, the grapegrower produces grapes through a multitude of specifications to achieve a quality result. Different Vitis vinifera grape varieties allow the creation of wine of different genres. Akin to a conductor of music, the winemaker decides what genre to create and considers resources required to realise the grape’s potential. A primary consideration is the yeast: whether to inoculate the grape juice or leave it ‘wild’; whether to inoculate with a specific strain of Saccharomyces or a combination of Saccharomyces strains; or whether to proceed with a non-Saccharomyces species? Whilst the various Saccharomyces and non-Saccharomyces yeasts perform their role during fermentation, the performance is not over until the ‘fat lady’ (S. cerevisiae has sung (i.e., the grape sugar has been fermented to specified dryness and alcoholic fermentation is complete. Is the wine harmonious or discordant? Will the consumer demand an encore and make a repeat purchase? Understanding consumer needs lets winemakers orchestrate different symphonies (i.e., wine styles using single- or multi-species ferments. Some consumers will choose the sounds of a philharmonic orchestra comprising a great range of diverse instrumentalists (as is the case with wine created from spontaneous fermentation; some will prefer to listen to a smaller ensemble (analogous to wine produced by a selected group of non-Saccharomyces and Saccharomyces yeast; and others will favour the well-known and reliable superstar soprano (i.e., S. cerevisiae. But what if a digital music synthesizer—such as a synthetic yeast—becomes available that can produce any music genre with the purest of sounds by the touch of a few buttons

  11. GenomePeek—an online tool for prokaryotic genome and metagenome analysis

    Directory of Open Access Journals (Sweden)

    Katelyn McNair

    2015-06-01

    Full Text Available As more and more prokaryotic sequencing takes place, a method to quickly and accurately analyze this data is needed. Previous tools are mainly designed for metagenomic analysis and have limitations; such as long runtimes and significant false positive error rates. The online tool GenomePeek (edwards.sdsu.edu/GenomePeek was developed to analyze both single genome and metagenome sequencing files, quickly and with low error rates. GenomePeek uses a sequence assembly approach where reads to a set of conserved genes are extracted, assembled and then aligned against the highly specific reference database. GenomePeek was found to be faster than traditional approaches while still keeping error rates low, as well as offering unique data visualization options.

  12. MicroScope: a platform for microbial genome annotation and comparative genomics.

    Science.gov (United States)

    Vallenet, D; Engelen, S; Mornico, D; Cruveiller, S; Fleury, L; Lajus, A; Rouy, Z; Roche, D; Salvignol, G; Scarpelli, C; Médigue, C

    2009-01-01

    The initial outcome of genome sequencing is the creation of long text strings written in a four letter alphabet. The role of in silico sequence analysis is to assist biologists in the act of associating biological knowledge with these sequences, allowing investigators to make inferences and predictions that can be tested experimentally. A wide variety of software is available to the scientific community, and can be used to identify genomic objects, before predicting their biological functions. However, only a limited number of biologically interesting features can be revealed from an isolated sequence. Comparative genomics tools, on the other hand, by bringing together the information contained in numerous genomes simultaneously, allow annotators to make inferences based on the idea that evolution and natural selection are central to the definition of all biological processes. We have developed the MicroScope platform in order to offer a web-based framework for the systematic and efficient revision of microbial genome annotation and comparative analysis (http://www.genoscope.cns.fr/agc/microscope). Starting with the description of the flow chart of the annotation processes implemented in the MicroScope pipeline, and the development of traditional and novel microbial annotation and comparative analysis tools, this article emphasizes the essential role of expert annotation as a complement of automatic annotation. Several examples illustrate the use of implemented tools for the review and curation of annotations of both new and publicly available microbial genomes within MicroScope's rich integrated genome framework. The platform is used as a viewer in order to browse updated annotation information of available microbial genomes (more than 440 organisms to date), and in the context of new annotation projects (117 bacterial genomes). The human expertise gathered in the MicroScope database (about 280,000 independent annotations) contributes to improve the quality of

  13. Global repeat discovery and estimation of genomic copy number in a large, complex genome using a high-throughput 454 sequence survey

    Directory of Open Access Journals (Sweden)

    Varala Kranthi

    2007-05-01

    Full Text Available Abstract Background Extensive computational and database tools are available to mine genomic and genetic databases for model organisms, but little genomic data is available for many species of ecological or agricultural significance, especially those with large genomes. Genome surveys using conventional sequencing techniques are powerful, particularly for detecting sequences present in many copies per genome. However these methods are time-consuming and have potential drawbacks. High throughput 454 sequencing provides an alternative method by which much information can be gained quickly and cheaply from high-coverage surveys of genomic DNA. Results We sequenced 78 million base-pairs of randomly sheared soybean DNA which passed our quality criteria. Computational analysis of the survey sequences provided global information on the abundant repetitive sequences in soybean. The sequence was used to determine the copy number across regions of large genomic clones or contigs and discover higher-order structures within satellite repeats. We have created an annotated, online database of sequences present in multiple copies in the soybean genome. The low bias of pyrosequencing against repeat sequences is demonstrated by the overall composition of the survey data, which matches well with past estimates of repetitive DNA content obtained by DNA re-association kinetics (Cot analysis. Conclusion This approach provides a potential aid to conventional or shotgun genome assembly, by allowing rapid assessment of copy number in any clone or clone-end sequence. In addition, we show that partial sequencing can provide access to partial protein-coding sequences.

  14. Hydrogen peroxide induced loss of heterozygosity correlates with replicative lifespan and mitotic asymmetry in Saccharomyces cerevisiae

    Science.gov (United States)

    Jackson, Erin D.; Parker, Meighan C.; Gupta, Nilin; Rodrigues, Jenny

    2016-01-01

    Cellular aging in Saccharomyces cerevisiae can lead to genomic instability and impaired mitotic asymmetry. To investigate the role of oxidative stress in cellular aging, we examined the effect of exogenous hydrogen peroxide on genomic instability and mitotic asymmetry in a collection of yeast strains with diverse backgrounds. We treated yeast cells with hydrogen peroxide and monitored the changes of viability and the frequencies of loss of heterozygosity (LOH) in response to hydrogen peroxide doses. The mid-transition points of viability and LOH were quantified using sigmoid mathematical functions. We found that the increase of hydrogen peroxide dependent genomic instability often occurs before a drop in viability. We previously observed that elevation of genomic instability generally lags behind the drop in viability during chronological aging. Hence, onset of genomic instability induced by exogenous hydrogen peroxide treatment is opposite to that induced by endogenous oxidative stress during chronological aging, with regards to the midpoint of viability. This contrast argues that the effect of endogenous oxidative stress on genome integrity is well suppressed up to the dying-off phase during chronological aging. We found that the leadoff of exogenous hydrogen peroxide induced genomic instability to viability significantly correlated with replicative lifespan (RLS), indicating that yeast cells’ ability to counter oxidative stress contributes to their replicative longevity. Surprisingly, this leadoff is positively correlated with an inverse measure of endogenous mitotic asymmetry, indicating a trade-off between mitotic asymmetry and cell’s ability to fend off hydrogen peroxide induced oxidative stress. Overall, our results demonstrate strong associations of oxidative stress to genomic instability and mitotic asymmetry at the population level of budding yeast. PMID:27833823

  15. Zymogram profiling of superoxide dismutase and catalase activities allows Saccharomyces and non-Saccharomyces species differentiation and correlates to their fermentation performance.

    Science.gov (United States)

    Gamero-Sandemetrio, Esther; Gómez-Pastor, Rocío; Matallana, Emilia

    2013-05-01

    Aerobic organisms have devised several enzymatic and non-enzymatic antioxidant defenses to deal with reactive oxygen species (ROS) produced by cellular metabolism. To combat such stress, cells induce ROS scavenging enzymes such as catalase, peroxidase, superoxide dismutase (SOD) and glutathione reductase. In the present research, we have used a double staining technique of SOD and catalase enzymes in the same polyacrylamide gel to analyze the different antioxidant enzymatic activities and protein isoforms present in Saccharomyces and non-Saccharomyces yeast species. Moreover, we used a technique to differentially detect Sod1p and Sod2p on gel by immersion in NaCN, which specifically inhibits the Sod1p isoform. We observed unique SOD and catalase zymogram profiles for all the analyzed yeasts and we propose this technique as a new approach for Saccharomyces and non-Saccharomyces yeast strains differentiation. In addition, we observed functional correlations between SOD and catalase enzyme activities, accumulation of essential metabolites, such as glutathione and trehalose, and the fermentative performance of different yeasts strains with industrial relevance.

  16. SWITCH: a dynamic CRISPR tool for genome engineering and metabolic pathway control for cell factory construction in Saccharomyces cerevisiae.

    Science.gov (United States)

    Vanegas, Katherina García; Lehka, Beata Joanna; Mortensen, Uffe Hasbro

    2017-02-08

    The yeast Saccharomyces cerevisiae is increasingly used as a cell factory. However, cell factory construction time is a major obstacle towards using yeast for bio-production. Hence, tools to speed up cell factory construction are desirable. In this study, we have developed a new Cas9/dCas9 based system, SWITCH, which allows Saccharomyces cerevisiae strains to iteratively alternate between a genetic engineering state and a pathway control state. Since Cas9 induced recombination events are crucial for SWITCH efficiency, we first developed a technique TAPE, which we have successfully used to address protospacer efficiency. As proof of concept of the use of SWITCH in cell factory construction, we have exploited the genetic engineering state of a SWITCH strain to insert the five genes necessary for naringenin production. Next, the naringenin cell factory was switched to the pathway control state where production was optimized by downregulating an essential gene TSC13, hence, reducing formation of a byproduct. We have successfully integrated two CRISPR tools, one for genetic engineering and one for pathway control, into one system and successfully used it for cell factory construction.

  17. Final Technical Report on the Genome Sequence DataBase (GSDB): DE-FG03 95 ER 62062 September 1997-September 1999; FINAL

    International Nuclear Information System (INIS)

    Harger, Carol A.

    1999-01-01

    Since September 1997 NCGR has produced two web-based tools for researchers to use to access and analyze data in the Genome Sequence DataBase (GSDB). These tools are: Sequence Viewer, a nucleotide sequence and annotation visualization tool, and MAR-Finder, a tool that predicts, base upon statistical inferences, the location of matrix attachment regions (MARS) within a nucleotide sequence.[The annual report for June 1996 to August 1997 is included as an attachment to this final report.

  18. Database Resources of the BIG Data Center in 2018.

    Science.gov (United States)

    2018-01-04

    The BIG Data Center at Beijing Institute of Genomics (BIG) of the Chinese Academy of Sciences provides freely open access to a suite of database resources in support of worldwide research activities in both academia and industry. With the vast amounts of omics data generated at ever-greater scales and rates, the BIG Data Center is continually expanding, updating and enriching its core database resources through big-data integration and value-added curation, including BioCode (a repository archiving bioinformatics tool codes), BioProject (a biological project library), BioSample (a biological sample library), Genome Sequence Archive (GSA, a data repository for archiving raw sequence reads), Genome Warehouse (GWH, a centralized resource housing genome-scale data), Genome Variation Map (GVM, a public repository of genome variations), Gene Expression Nebulas (GEN, a database of gene expression profiles based on RNA-Seq data), Methylation Bank (MethBank, an integrated databank of DNA methylomes), and Science Wikis (a series of biological knowledge wikis for community annotations). In addition, three featured web services are provided, viz., BIG Search (search as a service; a scalable inter-domain text search engine), BIG SSO (single sign-on as a service; a user access control system to gain access to multiple independent systems with a single ID and password) and Gsub (submission as a service; a unified submission service for all relevant resources). All of these resources are publicly accessible through the home page of the BIG Data Center at http://bigd.big.ac.cn. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  19. HEpD: a database describing epigenetic differences between Thoroughbred and Jeju horses.

    Science.gov (United States)

    Gim, Jeong-An; Lee, Sugi; Kim, Dae-Soo; Jeong, Kwang-Seuk; Hong, Chang Pyo; Bae, Jin-Han; Moon, Jae-Woo; Choi, Yong-Seok; Cho, Byung-Wook; Cho, Hwan-Gue; Bhak, Jong; Kim, Heui-Soo

    2015-04-10

    With the advent of next-generation sequencing technology, genome-wide maps of DNA methylation are now available. The Thoroughbred horse is bred for racing, while the Jeju horse is a traditional Korean horse bred for racing or food. The methylation profiles of equine organs may provide genomic clues underlying their athletic traits. We have developed a database to elucidate genome-wide DNA methylation patterns of the cerebrum, lung, heart, and skeletal muscle from Thoroughbred and Jeju horses. Using MeDIP-Seq, our database provides information regarding significantly enriched methylated regions beyond a threshold, methylation density of a specific region, and differentially methylated regions (DMRs) for tissues from two equine breeds. It provided methylation patterns at 784 gene regions in the equine genome. This database can potentially help researchers identify DMRs in the tissues of these horse species and investigate the differences between the Thoroughbred and Jeju horse breeds. Copyright © 2015 Elsevier B.V. All rights reserved.

  20. The Candidate Cancer Gene Database: a database of cancer driver genes from forward genetic screens in mice.

    Science.gov (United States)

    Abbott, Kenneth L; Nyre, Erik T; Abrahante, Juan; Ho, Yen-Yi; Isaksson Vogel, Rachel; Starr, Timothy K

    2015-01-01

    Identification of cancer driver gene mutations is crucial for advancing cancer therapeutics. Due to the overwhelming number of passenger mutations in the human tumor genome, it is difficult to pinpoint causative driver genes. Using transposon mutagenesis in mice many laboratories have conducted forward genetic screens and identified thousands of candidate driver genes that are highly relevant to human cancer. Unfortunately, this information is difficult to access and utilize because it is scattered across multiple publications using different mouse genome builds and strength metrics. To improve access to these findings and facilitate meta-analyses, we developed the Candidate Cancer Gene Database (CCGD, http://ccgd-starrlab.oit.umn.edu/). The CCGD is a manually curated database containing a unified description of all identified candidate driver genes and the genomic location of transposon common insertion sites (CISs) from all currently published transposon-based screens. To demonstrate relevance to human cancer, we performed a modified gene set enrichment analysis using KEGG pathways and show that human cancer pathways are highly enriched in the database. We also used hierarchical clustering to identify pathways enriched in blood cancers compared to solid cancers. The CCGD is a novel resource available to scientists interested in the identification of genetic drivers of cancer. © The Author(s) 2014. Published by Oxford University Press on behalf of Nucleic Acids Research.

  1. Genome-derived vaccines.

    Science.gov (United States)

    De Groot, Anne S; Rappuoli, Rino

    2004-02-01

    Vaccine research entered a new era when the complete genome of a pathogenic bacterium was published in 1995. Since then, more than 97 bacterial pathogens have been sequenced and at least 110 additional projects are now in progress. Genome sequencing has also dramatically accelerated: high-throughput facilities can draft the sequence of an entire microbe (two to four megabases) in 1 to 2 days. Vaccine developers are using microarrays, immunoinformatics, proteomics and high-throughput immunology assays to reduce the truly unmanageable volume of information available in genome databases to a manageable size. Vaccines composed by novel antigens discovered from genome mining are already in clinical trials. Within 5 years we can expect to see a novel class of vaccines composed by genome-predicted, assembled and engineered T- and Bcell epitopes. This article addresses the convergence of three forces--microbial genome sequencing, computational immunology and new vaccine technologies--that are shifting genome mining for vaccines onto the forefront of immunology research.

  2. Medicago truncatula transporter database: a comprehensive database resource for M. truncatula transporters

    Directory of Open Access Journals (Sweden)

    Miao Zhenyan

    2012-02-01

    Full Text Available Abstract Background Medicago truncatula has been chosen as a model species for genomic studies. It is closely related to an important legume, alfalfa. Transporters are a large group of membrane-spanning proteins. They deliver essential nutrients, eject waste products, and assist the cell in sensing environmental conditions by forming a complex system of pumps and channels. Although studies have effectively characterized individual M. truncatula transporters in several databases, until now there has been no available systematic database that includes all transporters in M. truncatula. Description The M. truncatula transporter database (MTDB contains comprehensive information on the transporters in M. truncatula. Based on the TransportTP method, we have presented a novel prediction pipeline. A total of 3,665 putative transporters have been annotated based on International Medicago Genome Annotated Group (IMGAG V3.5 V3 and the M. truncatula Gene Index (MTGI V10.0 releases and assigned to 162 families according to the transporter classification system. These families were further classified into seven types according to their transport mode and energy coupling mechanism. Extensive annotations referring to each protein were generated, including basic protein function, expressed sequence tag (EST mapping, genome locus, three-dimensional template prediction, transmembrane segment, and domain annotation. A chromosome distribution map and text-based Basic Local Alignment Search Tools were also created. In addition, we have provided a way to explore the expression of putative M. truncatula transporter genes under stress treatments. Conclusions In summary, the MTDB enables the exploration and comparative analysis of putative transporters in M. truncatula. A user-friendly web interface and regular updates make MTDB valuable to researchers in related fields. The MTDB is freely available now to all users at http://bioinformatics.cau.edu.cn/MtTransporter/.

  3. Functional expression of amine oxidase from Aspergillus niger (AO-I) in Saccharomyces cerevisiae.

    Science.gov (United States)

    Kolaríková, Katerina; Galuszka, Petr; Sedlárová, Iva; Sebela, Marek; Frébort, Ivo

    2009-01-01

    The aim of this work was to prepare recombinant amine oxidase from Aspergillus niger after overexpressing in yeast. The yeast expression vector pDR197 that includes a constitutive PMA1 promoter was used for the expression in Saccharomyces cerevisiae. Recombinant amine oxidase was extracted from the growth medium of the yeast, purified to homogeneity and identified by activity assay and MALDI-TOF peptide mass fingerprinting. Similarity search in the newly published A. niger genome identified six genes coding for copper amine oxidase, two of them corresponding to the previously described enzymes AO-I a methylamine oxidase and three other genes coding for FAD amine oxidases. Thus, A. niger possesses an enormous metabolic gear to grow on amine compounds and thus support its saprophytic lifestyle.

  4. A DNA sequence element that advances replication origin activation time in Saccharomyces cerevisiae.

    Science.gov (United States)

    Pohl, Thomas J; Kolor, Katherine; Fangman, Walton L; Brewer, Bonita J; Raghuraman, M K

    2013-11-06

    Eukaryotic origins of DNA replication undergo activation at various times in S-phase, allowing the genome to be duplicated in a temporally staggered fashion. In the budding yeast Saccharomyces cerevisiae, the activation times of individual origins are not intrinsic to those origins but are instead governed by surrounding sequences. Currently, there are two examples of DNA sequences that are known to advance origin activation time, centromeres and forkhead transcription factor binding sites. By combining deletion and linker scanning mutational analysis with two-dimensional gel electrophoresis to measure fork direction in the context of a two-origin plasmid, we have identified and characterized a 19- to 23-bp and a larger 584-bp DNA sequence that are capable of advancing origin activation time.

  5. Investigation of autonomous cell cycle oscillation in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Hansen, Morten Skov

    2007-01-01

    Autonome Oscillationer i kontinuert kultivering af Saccharomyces cerevisiae Udgangspunktet for dette Ph.d. projekt var at søge at forstå, hvad der gør det muligt at opnå multiple statiske tilstande ved kontinuert kultivering af Saccharomyces cerevisiae med glukose som begrænsende substrat...

  6. Saccharomyces boulardii CNCM I-745 in different clinical conditions.

    Science.gov (United States)

    Dinleyici, Ener Cagri; Kara, Ates; Ozen, Metehan; Vandenplas, Yvan

    2014-11-01

    Saccharomyces boulardii is a well-known probiotic worldwide, and there are numerous studies including experimental and clinical trials in children and adults by the use of S. boulardii. The objective of the present report is to provide an update on the evidence for the efficacy of S. boulardii CNCM I-745 in different clinical conditions. Saccharomyces boulardii is one of the best-studied probiotics in acute gastroenteritis (AGE) and is shown to be safe and to reduce the duration of diarrhea and hospitalization by about 1 day. Saccharomyces boulardii is one of the recommended probiotics for AGE in children by European Society of Paediatric Infectious Diseases and European Society for Paediatric Gastroenterology, Hepatology and Nutrition (ESPGHAN). Saccharomyces boulardii is also a recommended probiotic for the prevention of antibiotic-associated diarrhea (AAD), and a recent study showed promising results for the treatment of AAD in children. There is insufficient evidence to recommend the long-term use of S. boulardii in patients with irritable bowel syndrome. Although some clinical studies showed positive effects of S. boulardii on inflammation, there is no clinical evidence that S. boulardii is useful in inflammatory bowel disease. Saccharomyces boulardii could be used in patients needing Helicobacter pylori eradication because the S. boulardii improves compliance, decreases the side effects and moderately increases the eradication rate. There are new promising results (improving feeding tolerance, shorten the course of hyperbilirubinemia), but we do still not recommend the routine use of S. boulardii in newborns. Saccharomyces boulardii CNCM I-745 is a good example for the statement that each probiotic needs to be taxonomically characterized and its efficacy and safety should be documented individually in different clinical settings.

  7. VerSeDa: vertebrate secretome database.

    Science.gov (United States)

    Cortazar, Ana R; Oguiza, José A; Aransay, Ana M; Lavín, José L

    2017-01-01

    Based on the current tools, de novo secretome (full set of proteins secreted by an organism) prediction is a time consuming bioinformatic task that requires a multifactorial analysis in order to obtain reliable in silico predictions. Hence, to accelerate this process and offer researchers a reliable repository where secretome information can be obtained for vertebrates and model organisms, we have developed VerSeDa (Vertebrate Secretome Database). This freely available database stores information about proteins that are predicted to be secreted through the classical and non-classical mechanisms, for the wide range of vertebrate species deposited at the NCBI, UCSC and ENSEMBL sites. To our knowledge, VerSeDa is the only state-of-the-art database designed to store secretome data from multiple vertebrate genomes, thus, saving an important amount of time spent in the prediction of protein features that can be retrieved from this repository directly. VerSeDa is freely available at http://genomics.cicbiogune.es/VerSeDa/index.php. © The Author(s) 2017. Published by Oxford University Press.

  8. Mutant power: using mutant allele collections for yeast functional genomics.

    Science.gov (United States)

    Norman, Kaitlyn L; Kumar, Anuj

    2016-03-01

    The budding yeast has long served as a model eukaryote for the functional genomic analysis of highly conserved signaling pathways, cellular processes and mechanisms underlying human disease. The collection of reagents available for genomics in yeast is extensive, encompassing a growing diversity of mutant collections beyond gene deletion sets in the standard wild-type S288C genetic background. We review here three main types of mutant allele collections: transposon mutagen collections, essential gene collections and overexpression libraries. Each collection provides unique and identifiable alleles that can be utilized in genome-wide, high-throughput studies. These genomic reagents are particularly informative in identifying synthetic phenotypes and functions associated with essential genes, including those modeled most effectively in complex genetic backgrounds. Several examples of genomic studies in filamentous/pseudohyphal backgrounds are provided here to illustrate this point. Additionally, the limitations of each approach are examined. Collectively, these mutant allele collections in Saccharomyces cerevisiae and the related pathogenic yeast Candida albicans promise insights toward an advanced understanding of eukaryotic molecular and cellular biology. © The Author 2015. Published by Oxford University Press. All rights reserved. For permissions, please email: journals.permissions@oup.com.

  9. Genome-wide analytical approaches for reverse metabolic engineering of industrially relevant phenotypes in yeast

    Science.gov (United States)

    Oud, Bart; Maris, Antonius J A; Daran, Jean-Marc; Pronk, Jack T

    2012-01-01

    Successful reverse engineering of mutants that have been obtained by nontargeted strain improvement has long presented a major challenge in yeast biotechnology. This paper reviews the use of genome-wide approaches for analysis of Saccharomyces cerevisiae strains originating from evolutionary engineering or random mutagenesis. On the basis of an evaluation of the strengths and weaknesses of different methods, we conclude that for the initial identification of relevant genetic changes, whole genome sequencing is superior to other analytical techniques, such as transcriptome, metabolome, proteome, or array-based genome analysis. Key advantages of this technique over gene expression analysis include the independency of genome sequences on experimental context and the possibility to directly and precisely reproduce the identified changes in naive strains. The predictive value of genome-wide analysis of strains with industrially relevant characteristics can be further improved by classical genetics or simultaneous analysis of strains derived from parallel, independent strain improvement lineages. PMID:22152095

  10. Genome-wide analytical approaches for reverse metabolic engineering of industrially relevant phenotypes in yeast.

    Science.gov (United States)

    Oud, Bart; van Maris, Antonius J A; Daran, Jean-Marc; Pronk, Jack T

    2012-03-01

    Successful reverse engineering of mutants that have been obtained by nontargeted strain improvement has long presented a major challenge in yeast biotechnology. This paper reviews the use of genome-wide approaches for analysis of Saccharomyces cerevisiae strains originating from evolutionary engineering or random mutagenesis. On the basis of an evaluation of the strengths and weaknesses of different methods, we conclude that for the initial identification of relevant genetic changes, whole genome sequencing is superior to other analytical techniques, such as transcriptome, metabolome, proteome, or array-based genome analysis. Key advantages of this technique over gene expression analysis include the independency of genome sequences on experimental context and the possibility to directly and precisely reproduce the identified changes in naive strains. The predictive value of genome-wide analysis of strains with industrially relevant characteristics can be further improved by classical genetics or simultaneous analysis of strains derived from parallel, independent strain improvement lineages. © 2011 Federation of European Microbiological Societies. Published by Blackwell Publishing Ltd. All rights reserved.

  11. DPTEdb, an integrative database of transposable elements in dioecious plants.

    Science.gov (United States)

    Li, Shu-Fen; Zhang, Guo-Jun; Zhang, Xue-Jin; Yuan, Jin-Hong; Deng, Chuan-Liang; Gu, Lian-Feng; Gao, Wu-Jun

    2016-01-01

    Dioecious plants usually harbor 'young' sex chromosomes, providing an opportunity to study the early stages of sex chromosome evolution. Transposable elements (TEs) are mobile DNA elements frequently found in plants and are suggested to play important roles in plant sex chromosome evolution. The genomes of several dioecious plants have been sequenced, offering an opportunity to annotate and mine the TE data. However, comprehensive and unified annotation of TEs in these dioecious plants is still lacking. In this study, we constructed a dioecious plant transposable element database (DPTEdb). DPTEdb is a specific, comprehensive and unified relational database and web interface. We used a combination of de novo, structure-based and homology-based approaches to identify TEs from the genome assemblies of previously published data, as well as our own. The database currently integrates eight dioecious plant species and a total of 31 340 TEs along with classification information. DPTEdb provides user-friendly web interfaces to browse, search and download the TE sequences in the database. Users can also use tools, including BLAST, GetORF, HMMER, Cut sequence and JBrowse, to analyze TE data. Given the role of TEs in plant sex chromosome evolution, the database will contribute to the investigation of TEs in structural, functional and evolutionary dynamics of the genome of dioecious plants. In addition, the database will supplement the research of sex diversification and sex chromosome evolution of dioecious plants.Database URL: http://genedenovoweb.ticp.net:81/DPTEdb/index.php. © The Author(s) 2016. Published by Oxford University Press.

  12. Saccharomyces cerevisiae var. boulardii fungemia following probiotic treatment

    OpenAIRE

    Appel-da-Silva, Marcelo C.; Narvaez, Gabriel A.; Perez, Leandro R.R.; Drehmer, Laura; Lewgoy, Jairo

    2017-01-01

    Probiotics are commonly prescribed as an adjuvant in the treatment of antibiotic-associated diarrhea caused by Clostridium difficile. We report the case of an immunocompromised 73-year-old patient on chemotherapy who developed Saccharomyces cerevisiae var. boulardii fungemia in a central venous catheter during treatment of antibiotic-associated pseudomembranous colitis with the probiotic Saccharomyces cerevisiae var. boulardii. Fungemia was resolved after interruption of probiotic administrat...

  13. The COG database: an updated version includes eukaryotes

    Directory of Open Access Journals (Sweden)

    Sverdlov Alexander V

    2003-09-01

    Full Text Available Abstract Background The availability of multiple, essentially complete genome sequences of prokaryotes and eukaryotes spurred both the demand and the opportunity for the construction of an evolutionary classification of genes from these genomes. Such a classification system based on orthologous relationships between genes appears to be a natural framework for comparative genomics and should facilitate both functional annotation of genomes and large-scale evolutionary studies. Results We describe here a major update of the previously developed system for delineation of Clusters of Orthologous Groups of proteins (COGs from the sequenced genomes of prokaryotes and unicellular eukaryotes and the construction of clusters of predicted orthologs for 7 eukaryotic genomes, which we named KOGs after eukaryotic orthologous groups. The COG collection currently consists of 138,458 proteins, which form 4873 COGs and comprise 75% of the 185,505 (predicted proteins encoded in 66 genomes of unicellular organisms. The eukaryotic orthologous groups (KOGs include proteins from 7 eukaryotic genomes: three animals (the nematode Caenorhabditis elegans, the fruit fly Drosophila melanogaster and Homo sapiens, one plant, Arabidopsis thaliana, two fungi (Saccharomyces cerevisiae and Schizosaccharomyces pombe, and the intracellular microsporidian parasite Encephalitozoon cuniculi. The current KOG set consists of 4852 clusters of orthologs, which include 59,838 proteins, or ~54% of the analyzed eukaryotic 110,655 gene products. Compared to the coverage of the prokaryotic genomes with COGs, a considerably smaller fraction of eukaryotic genes could be included into the KOGs; addition of new eukaryotic genomes is expected to result in substantial increase in the coverage of eukaryotic genomes with KOGs. Examination of the phyletic patterns of KOGs reveals a conserved core represented in all analyzed species and consisting of ~20% of the KOG set. This conserved portion of the

  14. Regulation of Small Mitochondrial DNA Replicative Advantage by Ribonucleotide Reductase in Saccharomyces cerevisiae

    Directory of Open Access Journals (Sweden)

    Elliot Bradshaw

    2017-09-01

    Full Text Available Small mitochondrial genomes can behave as selfish elements by displacing wild-type genomes regardless of their detriment to the host organism. In the budding yeast Saccharomyces cerevisiae, small hypersuppressive mtDNA transiently coexist with wild-type in a state of heteroplasmy, wherein the replicative advantage of the small mtDNA outcompetes wild-type and produces offspring without respiratory capacity in >95% of colonies. The cytosolic enzyme ribonucleotide reductase (RNR catalyzes the rate-limiting step in dNTP synthesis and its inhibition has been correlated with increased petite colony formation, reflecting loss of respiratory function. Here, we used heteroplasmic diploids containing wild-type (rho+ and suppressive (rho− or hypersuppressive (HS rho− mitochondrial genomes to explore the effects of RNR activity on mtDNA heteroplasmy in offspring. We found that the proportion of rho+ offspring was significantly increased by RNR overexpression or deletion of its inhibitor, SML1, while reducing RNR activity via SML1 overexpression produced the opposite effects. In addition, using Ex Taq and KOD Dash polymerases, we observed a replicative advantage for small over large template DNA in vitro, but only at low dNTP concentrations. These results suggest that dNTP insufficiency contributes to the replicative advantage of small mtDNA over wild-type and cytosolic dNTP synthesis by RNR is an important regulator of heteroplasmy involving small mtDNA molecules in yeast.

  15. SpirPep: an in silico digestion-based platform to assist bioactive peptides discovery from a genome-wide database.

    Science.gov (United States)

    Anekthanakul, Krittima; Hongsthong, Apiradee; Senachak, Jittisak; Ruengjitchatchawalya, Marasri

    2018-04-20

    Bioactive peptides, including biological sources-derived peptides with different biological activities, are protein fragments that influence the functions or conditions of organisms, in particular humans and animals. Conventional methods of identifying bioactive peptides are time-consuming and costly. To quicken the processes, several bioinformatics tools are recently used to facilitate screening of the potential peptides prior their activity assessment in vitro and/or in vivo. In this study, we developed an efficient computational method, SpirPep, which offers many advantages over the currently available tools. The SpirPep web application tool is a one-stop analysis and visualization facility to assist bioactive peptide discovery. The tool is equipped with 15 customized enzymes and 1-3 miscleavage options, which allows in silico digestion of protein sequences encoded by protein-coding genes from single, multiple, or genome-wide scaling, and then directly classifies the peptides by bioactivity using an in-house database that contains bioactive peptides collected from 13 public databases. With this tool, the resulting peptides are categorized by each selected enzyme, and shown in a tabular format where the peptide sequences can be tracked back to their original proteins. The developed tool and webpages are coded in PHP and HTML with CSS/JavaScript. Moreover, the tool allows protein-peptide alignment visualization by Generic Genome Browser (GBrowse) to display the region and details of the proteins and peptides within each parameter, while considering digestion design for the desirable bioactivity. SpirPep is efficient; it takes less than 20 min to digest 3000 proteins (751,860 amino acids) with 15 enzymes and three miscleavages for each enzyme, and only a few seconds for single enzyme digestion. Obviously, the tool identified more bioactive peptides than that of the benchmarked tool; an example of validated pentapeptide (FLPIL) from LC-MS/MS was demonstrated. The

  16. Ebolavirus Database: Gene and Protein Information Resource for Ebolaviruses

    Directory of Open Access Journals (Sweden)

    Rayapadi G. Swetha

    2016-01-01

    Full Text Available Ebola Virus Disease (EVD is a life-threatening haemorrhagic fever in humans. Even though there are many reports on EVD, the protein precursor functions and virulent factors of ebolaviruses remain poorly understood. Comparative analyses of Ebolavirus genomes will help in the identification of these important features. This prompted us to develop the Ebolavirus Database (EDB and we have provided links to various tools that will aid researchers to locate important regions in both the genomes and proteomes of Ebolavirus. The genomic analyses of ebolaviruses will provide important clues for locating the essential and core functional genes. The aim of EDB is to act as an integrated resource for ebolaviruses and we strongly believe that the database will be a useful tool for clinicians, microbiologists, health care workers, and bioscience researchers.

  17. Genomes to Proteomes

    Energy Technology Data Exchange (ETDEWEB)

    Panisko, Ellen A. [Pacific Northwest National Lab. (PNNL), Richland, WA (United States); Grigoriev, Igor [USDOE Joint Genome Inst., Walnut Creek, CA (United States); Daly, Don S. [Pacific Northwest National Lab. (PNNL), Richland, WA (United States); Webb-Robertson, Bobbie-Jo [Pacific Northwest National Lab. (PNNL), Richland, WA (United States); Baker, Scott E. [Pacific Northwest National Lab. (PNNL), Richland, WA (United States)

    2009-03-01

    Biologists are awash with genomic sequence data. In large part, this is due to the rapid acceleration in the generation of DNA sequence that occurred as public and private research institutes raced to sequence the human genome. In parallel with the large human genome effort, mostly smaller genomes of other important model organisms were sequenced. Projects following on these initial efforts have made use of technological advances and the DNA sequencing infrastructure that was built for the human and other organism genome projects. As a result, the genome sequences of many organisms are available in high quality draft form. While in many ways this is good news, there are limitations to the biological insights that can be gleaned from DNA sequences alone; genome sequences offer only a bird's eye view of the biological processes endemic to an organism or community. Fortunately, the genome sequences now being produced at such a high rate can serve as the foundation for other global experimental platforms such as proteomics. Proteomic methods offer a snapshot of the proteins present at a point in time for a given biological sample. Current global proteomics methods combine enzymatic digestion, separations, mass spectrometry and database searching for peptide identification. One key aspect of proteomics is the prediction of peptide sequences from mass spectrometry data. Global proteomic analysis uses computational matching of experimental mass spectra with predicted spectra based on databases of gene models that are often generated computationally. Thus, the quality of gene models predicted from a genome sequence is crucial in the generation of high quality peptide identifications. Once peptides are identified they can be assigned to their parent protein. Proteins identified as expressed in a given experiment are most useful when compared to other expressed proteins in a larger biological context or biochemical pathway. In this chapter we will discuss the automatic

  18. User Guidelines for the Brassica Database: BRAD.

    Science.gov (United States)

    Wang, Xiaobo; Cheng, Feng; Wang, Xiaowu

    2016-01-01

    The genome sequence of Brassica rapa was first released in 2011. Since then, further Brassica genomes have been sequenced or are undergoing sequencing. It is therefore necessary to develop tools that help users to mine information from genomic data efficiently. This will greatly aid scientific exploration and breeding application, especially for those with low levels of bioinformatic training. Therefore, the Brassica database (BRAD) was built to collect, integrate, illustrate, and visualize Brassica genomic datasets. BRAD provides useful searching and data mining tools, and facilitates the search of gene annotation datasets, syntenic or non-syntenic orthologs, and flanking regions of functional genomic elements. It also includes genome-analysis tools such as BLAST and GBrowse. One of the important aims of BRAD is to build a bridge between Brassica crop genomes with the genome of the model species Arabidopsis thaliana, thus transferring the bulk of A. thaliana gene study information for use with newly sequenced Brassica crops.

  19. An apoptotic cell cycle mutant in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Villadsen, Ingrid

    1996-01-01

    The simple eukaryote Saccharomyces cerevisiae has proved to be a useful organism for elucidating the mechanisms that govern cell cycle progression in eukaryotic cells. The excellent in vivo system permits a cell cycle study using temperature sensitive mutants. In addition, it is possible to study...... many genes and gene products from higher eukaryotes in Saccharomyces cerevisiae because many genes and biological processes are homologous or similar in lower and in higher eukaryotes. The highly developed methods of genetics and molecular biology greatly facilitates studies of higher eukaryotic...... processes.Programmmed cell death with apoptosis plays a major role in development and homeostatis in most, if not all, animal cells. Apoptosis is a morphologically distinct form of death, that requires the activation of a highly regulated suicide program. Saccharomyces cerevisiae provides a new system...

  20. The genome portal of the Department of Energy Joint Genome Institute: 2014 updates

    Energy Technology Data Exchange (ETDEWEB)

    Nordberg, Henrik [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Cantor, Michael [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Dusheyko, Serge [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Hua, Susan [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Poliakov, Alexander [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Shabalov, Igor [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Smirnova, Tatyana [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Grigoriev, Igor V. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Dubchak, Inna [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States)

    2013-11-12

    The U.S. Department of Energy (DOE) Joint Genome Institute (JGI), a national user facility, serves the diverse scientific community by providing integrated high-throughput sequencing and computational analysis to enable system-based scientific approaches in support of DOE missions related to clean energy generation and environmental characterization. The JGI Genome Portal (http://genome.jgi.doe.gov) provides unified access to all JGI genomic databases and analytical tools. The JGI maintains extensive data management systems and specialized analytical capabilities to manage and interpret complex genomic data. A user can search, download and explore multiple data sets available for all DOE JGI sequencing projects including their status, assemblies and annotations of sequenced genomes. In this paper, we describe major updates of the Genome Portal in the past 2 years with a specific emphasis on efficient handling of the rapidly growing amount of diverse genomic data accumulated in JGI.

  1. Guided genome halving: hardness, heuristics and the history of the Hemiascomycetes.

    Science.gov (United States)

    Zheng, Chunfang; Zhu, Qian; Adam, Zaky; Sankoff, David

    2008-07-01

    Some present day species have incurred a whole genome doubling event in their evolutionary history, and this is reflected today in patterns of duplicated segments scattered throughout their chromosomes. These duplications may be used as data to 'halve' the genome, i.e. to reconstruct the ancestral genome at the moment of doubling, but the solution is often highly nonunique. To resolve this problem, we take account of outgroups, external reference genomes, to guide and narrow down the search. We improve on a previous, computationally costly, 'brute force' method by adapting the genome halving algorithm of El-Mabrouk and Sankoff so that it rapidly and accurately constructs an ancestor close the outgroups, prior to a local optimization heuristic. We apply this to reconstruct the predoubling ancestor of Saccharomyces cerevisiae and Candida glabrata, guided by the genomes of three other yeasts that diverged before the genome doubling event. We analyze the results in terms (1) of the minimum evolution criterion, (2) how close the genome halving result is to the final (local) minimum and (3) how close the final result is to an ancestor manually constructed by an expert with access to additional information. We also visualize the set of reconstructed ancestors using classic multidimensional scaling to see what aspects of the two doubled and three unduplicated genomes influence the differences among the reconstructions. The experimental software is available on request.

  2. [Effects of non-saccharomyces albicans metabolic products on the proliferation of human umbilical vein endothelial cell ECV304].

    Science.gov (United States)

    Chen, Bin; Che, Tuanjie; Bai, Decheng; He, Xiangyi

    2013-04-01

    To evaluate the effects of non-Saccharomyces albicans metabolic products on the cell cycle distribution and proliferation of human umbilical vein endothelial cell ECV304 cells in vitro. The parallel dilution supernatant of Saccharomyces tropicalis, Saccharomyces krusei and Saccharomyces glabrata were prepared, and 1, 4, 16-fold(s) diluted concentration and control group were set up. The line of human umbilical vein endothelial cell ECV304 was cultured in vitro and treated by non-Saccharomyces albicans supernatant. The proliferous effect of ECV304 induced by non-Saccharomyces albicans supernatant after 24, 48, 72 h was detected by the methods of MTT, and the changes of cell density and cycle after 48 h were investigated by inverted microscope and flow cytometry. At the 24th hour, all of the higher concentration (1-fold) of non-Saccharomyces albicans supernatant and the 4-folds diluted Saccharomyces krusei could promote ECV304 proliferation(P Saccharomyces albicans supernatant at 48h and 72th hour, Saccharomyces krusei supernatant and Saccharomyces glabrata supernatant significantly increased proliferation rate of ECV304, while Saccharomyces tropicalis supernatant group showed no significant change no matter which concentration was tested. At 48th hour after adding the non-Saccharomyces albicans supernatant, the ECV304 cells density treated by Saccharomyces krusei supernatant and Saccharomyces glabrata supernatant were significantly higher under the inverted microscope. The G0/G1 population of ECV304 cells decreased while cell proliferation index (PI) increased after incubated with Saccharomyces krusei supernatant and Saccharomyces glabrata supernatant for 48 hours (P Saccharomyces tropicalis group showed no significant change (P > 0.05). The metabolic products of Sacharoymces krusei and Saccharomyces glabrata could induce proliferation of ECV304 cell, which suggests non-Saccharomyces albicans should be undergone more attention clinically in detection and treatment.

  3. Saccharomyces interspecies hybrids as model organisms for studying yeast adaptation to stressful environments.

    Science.gov (United States)

    Lopandic, Ksenija

    2018-01-01

    The strong development of molecular biology techniques and next-generation sequencing technologies in the last two decades has significantly improved our understanding of the evolutionary history of Saccharomyces yeasts. It has been shown that many strains isolated from man-made environments are not pure genetic lines, but contain genetic materials from different species that substantially increase their genome complexity. A number of strains have been described as interspecies hybrids, implying different yeast species that under specific circumstances exchange and recombine their genomes. Such fusing usually results in a wide variety of alterations at the genetic and chromosomal levels. The observed changes have suggested a high genome plasticity and a significant role of interspecies hybridization in the adaptation of yeasts to environmental stresses and industrial processes. There is a high probability that harsh wine and beer fermentation environments, from which the majority of interspecies hybrids have been isolated so far, influence their selection and stabilization as well as their genomic and phenotypic heterogeneity. The lessons we have learned about geno- and phenotype plasticity and the diversity of natural and commercial yeast hybrids have already had a strong impact on the development of artificial hybrids that can be successfully used in the fermentation-based food and beverage industry. The creation of artificial hybrids through the crossing of strains with desired attributes is a possibility to obtain a vast variety of new, but not genetically modified yeasts with a range of improved and beneficial traits. Copyright © 2017 John Wiley & Sons, Ltd. Copyright © 2017 John Wiley & Sons, Ltd.

  4. Non-Saccharomyces in Wine: Effect Upon Oenococcus oeni and Malolactic Fermentation

    Directory of Open Access Journals (Sweden)

    Aitor Balmaseda

    2018-03-01

    Full Text Available This work is a short review of the interactions between oenological yeasts and lactic acid bacteria (LAB, especially Oenococcus oeni, the main species carrying out the malolactic fermentation (MLF. The emphasis has been placed on non-Saccharomyces effects due to their recent increased interest in winemaking. Those interactions are variable, ranging from inhibitory, to neutral and stimulatory and are mediated by some known compounds, which will be discussed. One phenomena responsible of inhibitory interactions is the media exhaustion by yeasts, and particularly a decrease in L-malic acid by some non-Saccharomyces. Clearly ethanol is the main inhibitory compound of LAB produced by S. cerevisiae, but non-Saccharomyces can be used to decrease it. Sulfur dioxide and medium chain fatty acids (MCFAs produced by yeasts can exhibit inhibitory effect upon LAB or even result lethal. Interestingly mixed fermentations with non-Saccharomyces present less MCFA concentration. Among organic acids derived as result of yeast metabolism, succinic acid seems to be the most related with MLF inhibition. Several protein factors produced by S. cerevisiae inhibiting O. oeni have been described, but they have not been studied in non-Saccharomyces. According to the stimulatory effects, the use of non-Saccharomyces can increase the concentration of favorable mediators such as citric acid, pyruvic acid, or other compounds derived of yeast autolysis such as peptides, glucans, or mannoproteins. The emergence of non-Saccharomyces in winemaking present a new scenario in which MLF has to take place. For this reason, new tools and approaches should be explored to better understand this new winemaking context.

  5. Non-Saccharomyces in Wine: Effect Upon Oenococcus oeni and Malolactic Fermentation.

    Science.gov (United States)

    Balmaseda, Aitor; Bordons, Albert; Reguant, Cristina; Bautista-Gallego, Joaquín

    2018-01-01

    This work is a short review of the interactions between oenological yeasts and lactic acid bacteria (LAB), especially Oenococcus oeni , the main species carrying out the malolactic fermentation (MLF). The emphasis has been placed on non- Saccharomyces effects due to their recent increased interest in winemaking. Those interactions are variable, ranging from inhibitory, to neutral and stimulatory and are mediated by some known compounds, which will be discussed. One phenomena responsible of inhibitory interactions is the media exhaustion by yeasts, and particularly a decrease in L-malic acid by some non- Saccharomyces . Clearly ethanol is the main inhibitory compound of LAB produced by S. cerevisiae , but non- Saccharomyces can be used to decrease it. Sulfur dioxide and medium chain fatty acids (MCFAs) produced by yeasts can exhibit inhibitory effect upon LAB or even result lethal. Interestingly mixed fermentations with non- Saccharomyces present less MCFA concentration. Among organic acids derived as result of yeast metabolism, succinic acid seems to be the most related with MLF inhibition. Several protein factors produced by S. cerevisiae inhibiting O. oeni have been described, but they have not been studied in non- Saccharomyces . According to the stimulatory effects, the use of non- Saccharomyces can increase the concentration of favorable mediators such as citric acid, pyruvic acid, or other compounds derived of yeast autolysis such as peptides, glucans, or mannoproteins. The emergence of non- Saccharomyces in winemaking present a new scenario in which MLF has to take place. For this reason, new tools and approaches should be explored to better understand this new winemaking context.

  6. ProOpDB: Prokaryotic Operon DataBase.

    Science.gov (United States)

    Taboada, Blanca; Ciria, Ricardo; Martinez-Guerrero, Cristian E; Merino, Enrique

    2012-01-01

    The Prokaryotic Operon DataBase (ProOpDB, http://operons.ibt.unam.mx/OperonPredictor) constitutes one of the most precise and complete repositories of operon predictions now available. Using our novel and highly accurate operon identification algorithm, we have predicted the operon structures of more than 1200 prokaryotic genomes. ProOpDB offers diverse alternatives by which a set of operon predictions can be retrieved including: (i) organism name, (ii) metabolic pathways, as defined by the KEGG database, (iii) gene orthology, as defined by the COG database, (iv) conserved protein domains, as defined by the Pfam database, (v) reference gene and (vi) reference operon, among others. In order to limit the operon output to non-redundant organisms, ProOpDB offers an efficient method to select the most representative organisms based on a precompiled phylogenetic distances matrix. In addition, the ProOpDB operon predictions are used directly as the input data of our Gene Context Tool to visualize their genomic context and retrieve the sequence of their corresponding 5' regulatory regions, as well as the nucleotide or amino acid sequences of their genes.

  7. The bovine QTL viewer: a web accessible database of bovine Quantitative Trait Loci

    Directory of Open Access Journals (Sweden)

    Xavier Suresh R

    2006-06-01

    Full Text Available Abstract Background Many important agricultural traits such as weight gain, milk fat content and intramuscular fat (marbling in cattle are quantitative traits. Most of the information on these traits has not previously been integrated into a genomic context. Without such integration application of these data to agricultural enterprises will remain slow and inefficient. Our goal was to populate a genomic database with data mined from the bovine quantitative trait literature and to make these data available in a genomic context to researchers via a user friendly query interface. Description The QTL (Quantitative Trait Locus data and related information for bovine QTL are gathered from published work and from existing databases. An integrated database schema was designed and the database (MySQL populated with the gathered data. The bovine QTL Viewer was developed for the integration of QTL data available for cattle. The tool consists of an integrated database of bovine QTL and the QTL viewer to display QTL and their chromosomal position. Conclusion We present a web accessible, integrated database of bovine (dairy and beef cattle QTL for use by animal geneticists. The viewer and database are of general applicability to any livestock species for which there are public QTL data. The viewer can be accessed at http://bovineqtl.tamu.edu.

  8. Mapping data - KOME | LSDB Archive [Life Science Database Archive metadata

    Lifescience Database Archive (English)

    Full Text Available switchLanguage; BLAST Search Image Search Home About Archive Update History Data ...tional Rice Genome Sequencing Project (IRGSP) Data file File name: kome_mapping_data.zip File URL: ftp://ftp.biosciencedbc.jp/archiv...(Transcriptional Unit) About This Database Database Description Download License Update History of This Database Site Policy | Contact Us Mapping data - KOME | LSDB Archive ...

  9. Functional expression of rat VPAC1 receptor in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Hansen, M.K.; Tams, J.W.; Fahrenkrug, Jan

    1999-01-01

    G protein-coupled receptor; heterologous expression; membrane protein; Saccharomyces cerevisiae, vasoactive intestinal polypeptide; yeast mating factor-pre-pro *Ga-leader peptide......G protein-coupled receptor; heterologous expression; membrane protein; Saccharomyces cerevisiae, vasoactive intestinal polypeptide; yeast mating factor-pre-pro *Ga-leader peptide...

  10. MIPS: analysis and annotation of proteins from whole genomes in 2005.

    Science.gov (United States)

    Mewes, H W; Frishman, D; Mayer, K F X; Münsterkötter, M; Noubibou, O; Pagel, P; Rattei, T; Oesterheld, M; Ruepp, A; Stümpflen, V

    2006-01-01

    The Munich Information Center for Protein Sequences (MIPS at the GSF), Neuherberg, Germany, provides resources related to genome information. Manually curated databases for several reference organisms are maintained. Several of these databases are described elsewhere in this and other recent NAR database issues. In a complementary effort, a comprehensive set of >400 genomes automatically annotated with the PEDANT system are maintained. The main goal of our current work on creating and maintaining genome databases is to extend gene centered information to information on interactions within a generic comprehensive framework. We have concentrated our efforts along three lines (i) the development of suitable comprehensive data structures and database technology, communication and query tools to include a wide range of different types of information enabling the representation of complex information such as functional modules or networks Genome Research Environment System, (ii) the development of databases covering computable information such as the basic evolutionary relations among all genes, namely SIMAP, the sequence similarity matrix and the CABiNet network analysis framework and (iii) the compilation and manual annotation of information related to interactions such as protein-protein interactions or other types of relations (e.g. MPCDB, MPPI, CYGD). All databases described and the detailed descriptions of our projects can be accessed through the MIPS WWW server (http://mips.gsf.de).

  11. RatMap--rat genome tools and data.

    Science.gov (United States)

    Petersen, Greta; Johnson, Per; Andersson, Lars; Klinga-Levan, Karin; Gómez-Fabre, Pedro M; Ståhl, Fredrik

    2005-01-01

    The rat genome database RatMap (http://ratmap.org or http://ratmap.gen.gu.se) has been one of the main resources for rat genome information since 1994. The database is maintained by CMB-Genetics at Goteborg University in Sweden and provides information on rat genes, polymorphic rat DNA-markers and rat quantitative trait loci (QTLs), all curated at RatMap. The database is under the supervision of the Rat Gene and Nomenclature Committee (RGNC); thus much attention is paid to rat gene nomenclature. RatMap presents information on rat idiograms, karyotypes and provides a unified presentation of the rat genome sequence and integrated rat linkage maps. A set of tools is also available to facilitate the identification and characterization of rat QTLs, as well as the estimation of exon/intron number and sizes in individual rat genes. Furthermore, comparative gene maps of rat in regard to mouse and human are provided.

  12. RatMap—rat genome tools and data

    Science.gov (United States)

    Petersen, Greta; Johnson, Per; Andersson, Lars; Klinga-Levan, Karin; Gómez-Fabre, Pedro M.; Ståhl, Fredrik

    2005-01-01

    The rat genome database RatMap (http://ratmap.org or http://ratmap.gen.gu.se) has been one of the main resources for rat genome information since 1994. The database is maintained by CMB–Genetics at Göteborg University in Sweden and provides information on rat genes, polymorphic rat DNA-markers and rat quantitative trait loci (QTLs), all curated at RatMap. The database is under the supervision of the Rat Gene and Nomenclature Committee (RGNC); thus much attention is paid to rat gene nomenclature. RatMap presents information on rat idiograms, karyotypes and provides a unified presentation of the rat genome sequence and integrated rat linkage maps. A set of tools is also available to facilitate the identification and characterization of rat QTLs, as well as the estimation of exon/intron number and sizes in individual rat genes. Furthermore, comparative gene maps of rat in regard to mouse and human are provided. PMID:15608244

  13. Review of Saccharomyces boulardii as a treatment option in IBD.

    Science.gov (United States)

    Sivananthan, Kavitha; Petersen, Andreas Munk

    2018-05-17

    Review of the yeast Saccharomyces boulardii as a treatment option for the inflammatory bowel diseases (IBD) ulcerative colitis and Crohn's disease. IBD is caused by an inappropriate immune response to gut microbiota. Treatment options could therefore be prebiotics, probiotics, antibiotics and/or fecal transplant. In this review, we have looked at the evidence for the yeast S. boulardii as a treatment option. Searches in PubMed and the Cochrane Library with the MeSH words 'Saccharomyces boulardii AND IBD', 'Saccharomyces boulardii AND Inflammatory Bowel Disease', 'Saccharomyces boulardii AND ulcerative colitis' and 'Saccharomyces boulardii AND Crohn's disease' gave total a total of 80 articles. After exclusions because of irrelevance, articles in other languages and some articles that were not available, 16 articles were included in this review. Three of the clinical trials showed a positive effect of S. boulardii in IBD patients (two Crohn's disease, one ulcerative colitis), while there was one trial that didn't prove any effect (Crohn's disease). Included Animal trials and cell assays describes different anti-inflammatory mechanisms of S. boulardii supporting a possible effect when treating IBD patients. The number of studies of S. boulardii as treatment for IBD is limited. Furthermore, the existing trials have small populations and short duration. We do not have enough evidence to prove the effect of S. boulardii in IBD. Saccharomyces boulardii is, however, a plausible treatment option in the future, but more placebo-controlled clinical studies on both patients with ulcerative colitis and Crohn's disease are needed.

  14. Detection of genomic rearrangements in cucumber using genomecmp software

    Science.gov (United States)

    Kulawik, Maciej; Pawełkowicz, Magdalena Ewa; Wojcieszek, Michał; PlÄ der, Wojciech; Nowak, Robert M.

    2017-08-01

    Comparative genomic by increasing information about the genomes sequences available in the databases is a rapidly evolving science. A simple comparison of the general features of genomes such as genome size, number of genes, and chromosome number presents an entry point into comparative genomic analysis. Here we present the utility of the new tool genomecmp for finding rearrangements across the compared sequences and applications in plant comparative genomics.

  15. Distribution patterns of Saccharomyces species in cultural landscapes of Germany.

    Science.gov (United States)

    Brysch-Herzberg, Michael; Seidel, Martin

    2017-08-01

    The distribution patterns of the three Saccharomyces species, Saccharomyces paradoxus, S. uvarum and S. cerevisiae, were investigated by a culture-dependent approach in order to understand better how these species propagate in the cultural landscape of Germany. Saccharomyces paradoxus, the closest relative of S. cerevisiae, is shown to be a true woodland species. It was frequently found in the soil under conifers indicating that S. paradoxus is an autochthonous member of the microbial community in this habitat. Physiological characteristics of the species like the Crabtree effect and high tolerance against ethanol suggest that the species is adapted to regular supply with considerable amounts of sugars. Additionally, a high proportion of the S. paradoxus strains isolated in this study are shown to have the rare ability to ferment melezitose. For these reasons, it is hypothesized that S. paradoxus may be closely associated with the honeydew system in forests. Saccharomyces cerevisiae was rare in most habitats and only exceeded the frequency of S. paradoxus in habitats characterized by modern agricultural mass production of fruit. Both the landscape structure and the agricultural system heavily influence the frequencies of Saccharomyces species. © FEMS 2017. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

  16. Introducing a new breed of wine yeast: interspecific hybridisation between a commercial Saccharomyces cerevisiae wine yeast and Saccharomyces mikatae.

    Science.gov (United States)

    Bellon, Jennifer R; Schmid, Frank; Capone, Dimitra L; Dunn, Barbara L; Chambers, Paul J

    2013-01-01

    Interspecific hybrids are commonplace in agriculture and horticulture; bread wheat and grapefruit are but two examples. The benefits derived from interspecific hybridisation include the potential of generating advantageous transgressive phenotypes. This paper describes the generation of a new breed of wine yeast by interspecific hybridisation between a commercial Saccharomyces cerevisiae wine yeast strain and Saccharomyces mikatae, a species hitherto not associated with industrial fermentation environs. While commercially available wine yeast strains provide consistent and reliable fermentations, wines produced using single inocula are thought to lack the sensory complexity and rounded palate structure obtained from spontaneous fermentations. In contrast, interspecific yeast hybrids have the potential to deliver increased complexity to wine sensory properties and alternative wine styles through the formation of novel, and wider ranging, yeast volatile fermentation metabolite profiles, whilst maintaining the robustness of the wine yeast parent. Screening of newly generated hybrids from a cross between a S. cerevisiae wine yeast and S. mikatae (closely-related but ecologically distant members of the Saccharomyces sensu stricto clade), has identified progeny with robust fermentation properties and winemaking potential. Chemical analysis showed that, relative to the S. cerevisiae wine yeast parent, hybrids produced wines with different concentrations of volatile metabolites that are known to contribute to wine flavour and aroma, including flavour compounds associated with non-Saccharomyces species. The new S. cerevisiae x S. mikatae hybrids have the potential to produce complex wines akin to products of spontaneous fermentation while giving winemakers the safeguard of an inoculated ferment.

  17. Cloning-free genome alterations in Saccharomyces cerevisiae using adaptamer-mediated PCR

    DEFF Research Database (Denmark)

    Reid, Robert J D; Lisby, Michael; Rothstein, Rodney

    2002-01-01

    . Furthermore, many of the techniques described here rely on preexisting and commercially available adaptamer sets that can be obtained inexpensively rather than designing new primers for every experiment. Although a cost is incurred when performing multiple PCR amplifications, the increase in recombination...... efficiency is dramatic. Finally, the adaptamer-mediated PCR fusion methodology is versatile and can be applied to varied genome manipulations....

  18. Respiratory cancer database: An open access database of respiratory cancer gene and miRNA.

    Science.gov (United States)

    Choubey, Jyotsna; Choudhari, Jyoti Kant; Patel, Ashish; Verma, Mukesh Kumar

    2017-01-01

    Respiratory cancer database (RespCanDB) is a genomic and proteomic database of cancer of respiratory organ. It also includes the information of medicinal plants used for the treatment of various respiratory cancers with structure of its active constituents as well as pharmacological and chemical information of drug associated with various respiratory cancers. Data in RespCanDB has been manually collected from published research article and from other databases. Data has been integrated using MySQL an object-relational database management system. MySQL manages all data in the back-end and provides commands to retrieve and store the data into the database. The web interface of database has been built in ASP. RespCanDB is expected to contribute to the understanding of scientific community regarding respiratory cancer biology as well as developments of new way of diagnosing and treating respiratory cancer. Currently, the database consist the oncogenomic information of lung cancer, laryngeal cancer, and nasopharyngeal cancer. Data for other cancers, such as oral and tracheal cancers, will be added in the near future. The URL of RespCanDB is http://ridb.subdic-bioinformatics-nitrr.in/.

  19. RPAN: rice pan-genome browser for ∼3000 rice genomes.

    Science.gov (United States)

    Sun, Chen; Hu, Zhiqiang; Zheng, Tianqing; Lu, Kuangchen; Zhao, Yue; Wang, Wensheng; Shi, Jianxin; Wang, Chunchao; Lu, Jinyuan; Zhang, Dabing; Li, Zhikang; Wei, Chaochun

    2017-01-25

    A pan-genome is the union of the gene sets of all the individuals of a clade or a species and it provides a new dimension of genome complexity with the presence/absence variations (PAVs) of genes among these genomes. With the progress of sequencing technologies, pan-genome study is becoming affordable for eukaryotes with large-sized genomes. The Asian cultivated rice, Oryza sativa L., is one of the major food sources for the world and a model organism in plant biology. Recently, the 3000 Rice Genome Project (3K RGP) sequenced more than 3000 rice genomes with a mean sequencing depth of 14.3×, which provided a tremendous resource for rice research. In this paper, we present a genome browser, Rice Pan-genome Browser (RPAN), as a tool to search and visualize the rice pan-genome derived from 3K RGP. RPAN contains a database of the basic information of 3010 rice accessions, including genomic sequences, gene annotations, PAV information and gene expression data of the rice pan-genome. At least 12 000 novel genes absent in the reference genome were included. RPAN also provides multiple search and visualization functions. RPAN can be a rich resource for rice biology and rice breeding. It is available at http://cgm.sjtu.edu.cn/3kricedb/ or http://www.rmbreeding.cn/pan3k. © The Author(s) 2016. Published by Oxford University Press on behalf of Nucleic Acids Research.

  20. Detecting microsatellites within genomes: significant variation among algorithms

    Directory of Open Access Journals (Sweden)

    Rivals Eric

    2007-04-01

    Full Text Available Abstract Background Microsatellites are short, tandemly-repeated DNA sequences which are widely distributed among genomes. Their structure, role and evolution can be analyzed based on exhaustive extraction from sequenced genomes. Several dedicated algorithms have been developed for this purpose. Here, we compared the detection efficiency of five of them (TRF, Mreps, Sputnik, STAR, and RepeatMasker. Results Our analysis was first conducted on the human X chromosome, and microsatellite distributions were characterized by microsatellite number, length, and divergence from a pure motif. The algorithms work with user-defined parameters, and we demonstrate that the parameter values chosen can strongly influence microsatellite distributions. The five algorithms were then compared by fixing parameters settings, and the analysis was extended to three other genomes (Saccharomyces cerevisiae, Neurospora crassa and Drosophila melanogaster spanning a wide range of size and structure. Significant differences for all characteristics of microsatellites were observed among algorithms, but not among genomes, for both perfect and imperfect microsatellites. Striking differences were detected for short microsatellites (below 20 bp, regardless of motif. Conclusion Since the algorithm used strongly influences empirical distributions, studies analyzing microsatellite evolution based on a comparison between empirical and theoretical size distributions should therefore be considered with caution. We also discuss why a typological definition of microsatellites limits our capacity to capture their genomic distributions.

  1. Improved Production of a Heterologous Amylase in Saccharomyces cerevisiae by Inverse Metabolic Engineering

    DEFF Research Database (Denmark)

    Liu, Zihe; Liu, Lifang; Osterlund, Tobias

    2014-01-01

    this modification alone, the amylase secretion could be improved by 35%. As a complement to the identification of genomic variants, transcriptome analysis was also performed in order to understand on a global level the transcriptional changes associated with the improved amylase production caused by UV mutagenesis.......The increasing demand for industrial enzymes and biopharmaceutical proteins relies on robust production hosts with high protein yield and productivity. Being one of the best-studied model organisms and capable of performing posttranslational modifications, the yeast Saccharomyces cerevisiae...... is widely used as a cell factory for recombinant protein production. However, many recombinant proteins are produced at only 1% (or less) of the theoretical capacity due to the complexity of the secretory pathway, which has not been fully exploited. In this study, we applied the concept of inverse metabolic...

  2. Mouse Genome Informatics (MGI)

    Data.gov (United States)

    U.S. Department of Health & Human Services — MGI is the international database resource for the laboratory mouse, providing integrated genetic, genomic, and biological data to facilitate the study of human...

  3. A Web-Based Comparative Genomics Tutorial for Investigating Microbial Genomes

    Directory of Open Access Journals (Sweden)

    Michael Strong

    2009-12-01

    Full Text Available As the number of completely sequenced microbial genomes continues to rise at an impressive rate, it is important to prepare students with the skills necessary to investigate microorganisms at the genomic level. As a part of the core curriculum for first-year graduate students in the biological sciences, we have implemented a web-based tutorial to introduce students to the fields of comparative and functional genomics. The tutorial focuses on recent computational methods for identifying functionally linked genes and proteins on a genome-wide scale and was used to introduce students to the Rosetta Stone, Phylogenetic Profile, conserved Gene Neighbor, and Operon computational methods. Students learned to use a number of publicly available web servers and databases to identify functionally linked genes in the Escherichia coli genome, with emphasis on genome organization and operon structure. The overall effectiveness of the tutorial was assessed based on student evaluations and homework assignments. The tutorial is available to other educators at http://www.doe-mbi.ucla.edu/~strong/m253.php.

  4. Replicative age induces mitotic recombination in the ribosomal RNA gene cluster of Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    Derek L Lindstrom

    2011-03-01

    Full Text Available Somatic mutations contribute to the development of age-associated disease. In earlier work, we found that, at high frequency, aging Saccharomyces cerevisiae diploid cells produce daughters without mitochondrial DNA, leading to loss of respiration competence and increased loss of heterozygosity (LOH in the nuclear genome. Here we used the recently developed Mother Enrichment Program to ask whether aging cells that maintain the ability to produce respiration-competent daughters also experience increased genomic instability. We discovered that this population exhibits a distinct genomic instability phenotype that primarily affects the repeated ribosomal RNA gene array (rDNA array. As diploid cells passed their median replicative life span, recombination rates between rDNA arrays on homologous chromosomes progressively increased, resulting in mutational events that generated LOH at >300 contiguous open reading frames on the right arm of chromosome XII. We show that, while these recombination events were dependent on the replication fork block protein Fob1, the aging process that underlies this phenotype is Fob1-independent. Furthermore, we provide evidence that this aging process is not driven by mechanisms that modulate rDNA recombination in young cells, including loss of cohesion within the rDNA array or loss of Sir2 function. Instead, we suggest that the age-associated increase in rDNA recombination is a response to increasing DNA replication stress generated in aging cells.

  5. Genome sequence of Aspergillus luchuensis NBRC 4314

    Science.gov (United States)

    Yamada, Osamu; Machida, Masayuki; Hosoyama, Akira; Goto, Masatoshi; Takahashi, Toru; Futagami, Taiki; Yamagata, Youhei; Takeuchi, Michio; Kobayashi, Tetsuo; Koike, Hideaki; Abe, Keietsu; Asai, Kiyoshi; Arita, Masanori; Fujita, Nobuyuki; Fukuda, Kazuro; Higa, Ken-ichi; Horikawa, Hiroshi; Ishikawa, Takeaki; Jinno, Koji; Kato, Yumiko; Kirimura, Kohtaro; Mizutani, Osamu; Nakasone, Kaoru; Sano, Motoaki; Shiraishi, Yohei; Tsukahara, Masatoshi; Gomi, Katsuya

    2016-01-01

    Awamori is a traditional distilled beverage made from steamed Thai-Indica rice in Okinawa, Japan. For brewing the liquor, two microbes, local kuro (black) koji mold Aspergillus luchuensis and awamori yeast Saccharomyces cerevisiae are involved. In contrast, that yeasts are used for ethanol fermentation throughout the world, a characteristic of Japanese fermentation industries is the use of Aspergillus molds as a source of enzymes for the maceration and saccharification of raw materials. Here we report the draft genome of a kuro (black) koji mold, A. luchuensis NBRC 4314 (RIB 2604). The total length of nonredundant sequences was nearly 34.7 Mb, comprising approximately 2,300 contigs with 16 telomere-like sequences. In total, 11,691 genes were predicted to encode proteins. Most of the housekeeping genes, such as transcription factors and N-and O-glycosylation system, were conserved with respect to Aspergillus niger and Aspergillus oryzae. An alternative oxidase and acid-stable α-amylase regarding citric acid production and fermentation at a low pH as well as a unique glutamic peptidase were also found in the genome. Furthermore, key biosynthetic gene clusters of ochratoxin A and fumonisin B were absent when compared with A. niger genome, showing the safety of A. luchuensis for food and beverage production. This genome information will facilitate not only comparative genomics with industrial kuro-koji molds, but also molecular breeding of the molds in improvements of awamori fermentation. PMID:27651094

  6. Assembly and Multiplex Genome Integration of Metabolic Pathways in Yeast Using CasEMBLR.

    Science.gov (United States)

    Jakočiūnas, Tadas; Jensen, Emil D; Jensen, Michael K; Keasling, Jay D

    2018-01-01

    Genome integration is a vital step for implementing large biochemical pathways to build a stable microbial cell factory. Although traditional strain construction strategies are well established for the model organism Saccharomyces cerevisiae, recent advances in CRISPR/Cas9-mediated genome engineering allow much higher throughput and robustness in terms of strain construction. In this chapter, we describe CasEMBLR, a highly efficient and marker-free genome engineering method for one-step integration of in vivo assembled expression cassettes in multiple genomic sites simultaneously. CasEMBLR capitalizes on the CRISPR/Cas9 technology to generate double-strand breaks in genomic loci, thus prompting native homologous recombination (HR) machinery to integrate exogenously derived homology templates. As proof-of-principle for microbial cell factory development, CasEMBLR was used for one-step assembly and marker-free integration of the carotenoid pathway from 15 exogenously supplied DNA parts into three targeted genomic loci. As a second proof-of-principle, a total of ten DNA parts were assembled and integrated in two genomic loci to construct a tyrosine production strain, and at the same time knocking out two genes. This new method complements and improves the field of genome engineering in S. cerevisiae by providing a more flexible platform for rapid and precise strain building.

  7. Influence of organic acids and organochlorinated insecticides on metabolism of Saccharomyces cerevisiae

    Directory of Open Access Journals (Sweden)

    Pejin Dušanka J.

    2005-01-01

    Full Text Available Saccharomyces cerevisiae is exposed to different stress factors during the production: osmotic, temperature, oxidative. The response to these stresses is the adaptive mechanism of cells. The raw materials Saccharomyces cerevisiae is produced from, contain metabolism products of present microorganisms and protective agents used during the growth of sugar beet for example the influence of acetic and butyric acid and organochlorinated insecticides, lindan and heptachlor, on the metabolism of Saccharomyces cerevisiae was investigated and presented in this work. The mentioned compounds affect negatively the specific growth rate, yield, content of proteins, phosphorus, total ribonucleic acids. These compounds influence the increase of trechalose and glycogen content in the Saccharomyces cerevisiae cells.

  8. 1.15 - Structural Chemogenomics Databases to Navigate Protein–Ligand Interaction Space

    NARCIS (Netherlands)

    Kanev, G.K.; Kooistra, A.J.; de Esch, I.J.P.; de Graaf, C.

    2017-01-01

    Structural chemogenomics databases allow the integration and exploration of heterogeneous genomic, structural, chemical, and pharmacological data in order to extract useful information that is applicable for the discovery of new protein targets and biologically active molecules. Integrated databases

  9. SoyFN: a knowledge database of soybean functional networks.

    Science.gov (United States)

    Xu, Yungang; Guo, Maozu; Liu, Xiaoyan; Wang, Chunyu; Liu, Yang

    2014-01-01

    Many databases for soybean genomic analysis have been built and made publicly available, but few of them contain knowledge specifically targeting the omics-level gene-gene, gene-microRNA (miRNA) and miRNA-miRNA interactions. Here, we present SoyFN, a knowledge database of soybean functional gene networks and miRNA functional networks. SoyFN provides user-friendly interfaces to retrieve, visualize, analyze and download the functional networks of soybean genes and miRNAs. In addition, it incorporates much information about KEGG pathways, gene ontology annotations and 3'-UTR sequences as well as many useful tools including SoySearch, ID mapping, Genome Browser, eFP Browser and promoter motif scan. SoyFN is a schema-free database that can be accessed as a Web service from any modern programming language using a simple Hypertext Transfer Protocol call. The Web site is implemented in Java, JavaScript, PHP, HTML and Apache, with all major browsers supported. We anticipate that this database will be useful for members of research communities both in soybean experimental science and bioinformatics. Database URL: http://nclab.hit.edu.cn/SoyFN.

  10. From Genome Sequence to Taxonomy - A Skeptic’s View

    DEFF Research Database (Denmark)

    Özen, Asli Ismihan; Vesth, Tammi Camilla; Ussery, David

    2012-01-01

    The relative ease of sequencing bacterial genomes has resulted in thousands of sequenced bacterial genomes available in the public databases. This same technology now allows for using the entire genome sequence as an identifier for an organism. There are many methods available which attempt to us...

  11. Trichoderma virens β-glucosidase I (BGLI) gene; expression in Saccharomyces cerevisiae including docking and molecular dynamics studies.

    Science.gov (United States)

    Wickramasinghe, Gammadde Hewa Ishan Maduka; Rathnayake, Pilimathalawe Panditharathna Attanayake Mudiyanselage Samith Indika; Chandrasekharan, Naduviladath Vishvanath; Weerasinghe, Mahindagoda Siril Samantha; Wijesundera, Ravindra Lakshman Chundananda; Wijesundera, Wijepurage Sandhya Sulochana

    2017-06-21

    Cellulose, a linear polymer of β 1-4, linked glucose, is the most abundant renewable fraction of plant biomass (lignocellulose). It is synergistically converted to glucose by endoglucanase (EG) cellobiohydrolase (CBH) and β-glucosidase (BGL) of the cellulase complex. BGL plays a major role in the conversion of randomly cleaved cellooligosaccharides into glucose. As it is well known, Saccharomyces cerevisiae can efficiently convert glucose into ethanol under anaerobic conditions. Therefore, S.cerevisiae was genetically modified with the objective of heterologous extracellular expression of the BGLI gene of Trichoderma virens making it capable of utilizing cellobiose to produce ethanol. The cDNA and a genomic sequence of the BGLI gene of Trichoderma virens was cloned in the yeast expression vector pGAPZα and separately transformed to Saccharomyces cerevisiae. The size of the BGLI cDNA clone was 1363 bp and the genomic DNA clone contained an additional 76 bp single intron following the first exon. The gene was 90% similar to the DNA sequence and 99% similar to the deduced amino acid sequence of 1,4-β-D-glucosidase of T. atroviride (AC237343.1). The BGLI activity expressed by the recombinant genomic clone was 3.4 times greater (1.7 x 10 -3  IU ml -1 ) than that observed for the cDNA clone (5 x 10 -4  IU ml -1 ). Furthermore, the activity was similar to the activity of locally isolated Trichoderma virens (1.5 x 10 -3  IU ml -1 ). The estimated size of the protein was 52 kDA. In fermentation studies, the maximum ethanol production by the genomic and the cDNA clones were 0.36 g and 0.06 g /g of cellobiose respectively. Molecular docking results indicated that the bare protein and cellobiose-protein complex behave in a similar manner with considerable stability in aqueous medium. The deduced binding site and the binding affinity of the constructed homology model appeared to be reasonable. Moreover, it was identified that the five hydrogen bonds formed

  12. Construction of killer industrial yeast Saccharomyces cerevisiae HAU-1 and its fermentation performance

    Directory of Open Access Journals (Sweden)

    Bijender K. Bajaj

    2010-06-01

    Full Text Available Saccharomyces cerevisiae HAU-1, a time tested industrial yeast possesses most of the desirable fermentation characteristics like fast growth and fermentation rate, osmotolerance, high ethanol tolerance, ability to ferment molasses, and to ferment at elevated temperatures etc. However, this yeast was found to be sensitive against the killer strains of Saccharomyces cerevisiae. In the present study, killer trait was introduced into Saccharomyces cerevisiae HAU-1 by protoplast fusion with Saccharomyces cerevisiae MTCC 475, a killer strain. The resultant fusants were characterized for desirable fermentation characteristics. All the technologically important characteristics of distillery yeast Saccharomyces cerevisiae HAU-1 were retained in the fusants, and in addition the killer trait was also introduced into them. Further, the killer activity was found to be stably maintained during hostile conditions of ethanol fermentations in dextrose or molasses, and even during biomass recycling.

  13. Investigating core genetic-and-epigenetic cell cycle networks for stemness and carcinogenic mechanisms, and cancer drug design using big database mining and genome-wide next-generation sequencing data.

    Science.gov (United States)

    Li, Cheng-Wei; Chen, Bor-Sen

    2016-10-01

    Recent studies have demonstrated that cell cycle plays a central role in development and carcinogenesis. Thus, the use of big databases and genome-wide high-throughput data to unravel the genetic and epigenetic mechanisms underlying cell cycle progression in stem cells and cancer cells is a matter of considerable interest. Real genetic-and-epigenetic cell cycle networks (GECNs) of embryonic stem cells (ESCs) and HeLa cancer cells were constructed by applying system modeling, system identification, and big database mining to genome-wide next-generation sequencing data. Real GECNs were then reduced to core GECNs of HeLa cells and ESCs by applying principal genome-wide network projection. In this study, we investigated potential carcinogenic and stemness mechanisms for systems cancer drug design by identifying common core and specific GECNs between HeLa cells and ESCs. Integrating drug database information with the specific GECNs of HeLa cells could lead to identification of multiple drugs for cervical cancer treatment with minimal side-effects on the genes in the common core. We found that dysregulation of miR-29C, miR-34A, miR-98, and miR-215; and methylation of ANKRD1, ARID5B, CDCA2, PIF1, STAMBPL1, TROAP, ZNF165, and HIST1H2AJ in HeLa cells could result in cell proliferation and anti-apoptosis through NFκB, TGF-β, and PI3K pathways. We also identified 3 drugs, methotrexate, quercetin, and mimosine, which repressed the activated cell cycle genes, ARID5B, STK17B, and CCL2, in HeLa cells with minimal side-effects.

  14. Thermal resistance of Saccharomyces yeast ascospores in beers.

    Science.gov (United States)

    Milani, Elham A; Gardner, Richard C; Silva, Filipa V M

    2015-08-03

    The industrial production of beer ends with a process of thermal pasteurization. Saccharomyces cerevisiae and Saccharomyces pastorianus are yeasts used to produce top and bottom fermenting beers, respectively. In this research, first the sporulation rate of 12 Saccharomyces strains was studied. Then, the thermal resistance of ascospores of three S. cerevisiae strains (DSMZ 1848, DSMZ 70487, Ethanol Red(®)) and one strain of S. pastorianus (ATCC 9080) was determined in 4% (v/v) ethanol lager beer. D60 °C-values of 11.2, 7.5, 4.6, and 6.0 min and z-values of 11.7, 14.3, 12.4, and 12.7 °C were determined for DSMZ 1848, DSMZ 70487, ATCC 9080, and Ethanol Red(®), respectively. Lastly, experiments with 0 and 7% (v/v) beers were carried out to investigate the effect of ethanol content on the thermal resistance of S. cerevisiae (DSMZ 1848). D55 °C-values of 34.2 and 15.3 min were obtained for 0 and 7% beers, respectively, indicating lower thermal resistance in the more alcoholic beer. These results demonstrate similar spore thermal resistance for different Saccharomyces strains and will assist in the design of appropriate thermal pasteurization conditions for preserving beers with different alcohol contents. Copyright © 2015 Elsevier B.V. All rights reserved.

  15. Introducing a new breed of wine yeast: interspecific hybridisation between a commercial Saccharomyces cerevisiae wine yeast and Saccharomyces mikatae.

    Directory of Open Access Journals (Sweden)

    Jennifer R Bellon

    Full Text Available Interspecific hybrids are commonplace in agriculture and horticulture; bread wheat and grapefruit are but two examples. The benefits derived from interspecific hybridisation include the potential of generating advantageous transgressive phenotypes. This paper describes the generation of a new breed of wine yeast by interspecific hybridisation between a commercial Saccharomyces cerevisiae wine yeast strain and Saccharomyces mikatae, a species hitherto not associated with industrial fermentation environs. While commercially available wine yeast strains provide consistent and reliable fermentations, wines produced using single inocula are thought to lack the sensory complexity and rounded palate structure obtained from spontaneous fermentations. In contrast, interspecific yeast hybrids have the potential to deliver increased complexity to wine sensory properties and alternative wine styles through the formation of novel, and wider ranging, yeast volatile fermentation metabolite profiles, whilst maintaining the robustness of the wine yeast parent. Screening of newly generated hybrids from a cross between a S. cerevisiae wine yeast and S. mikatae (closely-related but ecologically distant members of the Saccharomyces sensu stricto clade, has identified progeny with robust fermentation properties and winemaking potential. Chemical analysis showed that, relative to the S. cerevisiae wine yeast parent, hybrids produced wines with different concentrations of volatile metabolites that are known to contribute to wine flavour and aroma, including flavour compounds associated with non-Saccharomyces species. The new S. cerevisiae x S. mikatae hybrids have the potential to produce complex wines akin to products of spontaneous fermentation while giving winemakers the safeguard of an inoculated ferment.

  16. Introducing a New Breed of Wine Yeast: Interspecific Hybridisation between a Commercial Saccharomyces cerevisiae Wine Yeast and Saccharomyces mikatae

    Science.gov (United States)

    Bellon, Jennifer R.; Schmid, Frank; Capone, Dimitra L.; Dunn, Barbara L.; Chambers, Paul J.

    2013-01-01

    Interspecific hybrids are commonplace in agriculture and horticulture; bread wheat and grapefruit are but two examples. The benefits derived from interspecific hybridisation include the potential of generating advantageous transgressive phenotypes. This paper describes the generation of a new breed of wine yeast by interspecific hybridisation between a commercial Saccharomyces cerevisiae wine yeast strain and Saccharomyces mikatae, a species hitherto not associated with industrial fermentation environs. While commercially available wine yeast strains provide consistent and reliable fermentations, wines produced using single inocula are thought to lack the sensory complexity and rounded palate structure obtained from spontaneous fermentations. In contrast, interspecific yeast hybrids have the potential to deliver increased complexity to wine sensory properties and alternative wine styles through the formation of novel, and wider ranging, yeast volatile fermentation metabolite profiles, whilst maintaining the robustness of the wine yeast parent. Screening of newly generated hybrids from a cross between a S. cerevisiae wine yeast and S. mikatae (closely-related but ecologically distant members of the Saccharomyces sensu stricto clade), has identified progeny with robust fermentation properties and winemaking potential. Chemical analysis showed that, relative to the S. cerevisiae wine yeast parent, hybrids produced wines with different concentrations of volatile metabolites that are known to contribute to wine flavour and aroma, including flavour compounds associated with non-Saccharomyces species. The new S. cerevisiae x S. mikatae hybrids have the potential to produce complex wines akin to products of spontaneous fermentation while giving winemakers the safeguard of an inoculated ferment. PMID:23614011

  17. The Vigna Genome Server, 'VigGS': A Genomic Knowledge Base of the Genus Vigna Based on High-Quality, Annotated Genome Sequence of the Azuki Bean, Vigna angularis (Willd.) Ohwi & Ohashi.

    Science.gov (United States)

    Sakai, Hiroaki; Naito, Ken; Takahashi, Yu; Sato, Toshiyuki; Yamamoto, Toshiya; Muto, Isamu; Itoh, Takeshi; Tomooka, Norihiko

    2016-01-01

    The genus Vigna includes legume crops such as cowpea, mungbean and azuki bean, as well as >100 wild species. A number of the wild species are highly tolerant to severe environmental conditions including high-salinity, acid or alkaline soil; drought; flooding; and pests and diseases. These features of the genus Vigna make it a good target for investigation of genetic diversity in adaptation to stressful environments; however, a lack of genomic information has hindered such research in this genus. Here, we present a genome database of the genus Vigna, Vigna Genome Server ('VigGS', http://viggs.dna.affrc.go.jp), based on the recently sequenced azuki bean genome, which incorporates annotated exon-intron structures, along with evidence for transcripts and proteins, visualized in GBrowse. VigGS also facilitates user construction of multiple alignments between azuki bean genes and those of six related dicot species. In addition, the database displays sequence polymorphisms between azuki bean and its wild relatives and enables users to design primer sequences targeting any variant site. VigGS offers a simple keyword search in addition to sequence similarity searches using BLAST and BLAT. To incorporate up to date genomic information, VigGS automatically receives newly deposited mRNA sequences of pre-set species from the public database once a week. Users can refer to not only gene structures mapped on the azuki bean genome on GBrowse but also relevant literature of the genes. VigGS will contribute to genomic research into plant biotic and abiotic stresses and to the future development of new stress-tolerant crops. © The Author 2015. Published by Oxford University Press on behalf of Japanese Society of Plant Physiologists. All rights reserved. For permissions, please email: journals.permissions@oup.com.

  18. Microbial Genome Analysis and Comparisons: Web-based Protocols and Resources

    Science.gov (United States)

    Fully annotated genome sequences of many microorganisms are publicly available as a resource. However, in-depth analysis of these genomes using specialized tools is required to derive meaningful information. We describe here the utility of three powerful publicly available genome databases and ana...

  19. Functional conservation of nucleosome formation selectively biases presumably neutral molecular variation in yeast genomes.

    Science.gov (United States)

    Babbitt, Gregory A; Cotter, C R

    2011-01-01

    One prominent pattern of mutational frequency, long appreciated in comparative genomics, is the bias of purine/pyrimidine conserving substitutions (transitions) over purine/pyrimidine altering substitutions (transversions). Traditionally, this transitional bias has been thought to be driven by the underlying rates of DNA mutation and/or repair. However, recent sequencing studies of mutation accumulation lines in model organisms demonstrate that substitutions generally do not accumulate at rates that would indicate a transitional bias. These observations have called into question a very basic assumption of molecular evolution; that naturally occurring patterns of molecular variation in noncoding regions accurately reflect the underlying processes of randomly accumulating neutral mutation in nuclear genomes. Here, in Saccharomyces yeasts, we report a very strong inverse association (r = -0.951, P < 0.004) between the genome-wide frequency of substitutions and their average energetic effect on nucleosome formation, as predicted by a structurally based energy model of DNA deformation around the nucleosome core. We find that transitions occurring at sites positioned nearest the nucleosome surface, which are believed to function most importantly in nucleosome formation, alter the deformation energy of DNA to the nucleosome core by only a fraction of the energy changes typical of most transversions. When we examined the same substitutions set against random background sequences as well as an existing study reporting substitutions arising in mutation accumulation lines of Saccharomyces cerevisiae, we failed to find a similar relationship. These results support the idea that natural selection acting to functionally conserve chromatin organization may contribute significantly to genome-wide transitional bias, even in noncoding regions. Because nucleosome core structure is highly conserved across eukaryotes, our observations may also help to further explain locally elevated

  20. Anaerobic organic acid metabolism of Candida zemplinina in comparison with Saccharomyces wine yeasts.

    Science.gov (United States)

    Magyar, Ildikó; Nyitrai-Sárdy, Diána; Leskó, Annamária; Pomázi, Andrea; Kállay, Miklós

    2014-05-16

    Organic acid production under oxygen-limited conditions has been thoroughly studied in the Saccharomyces species, but practically never investigated in Candida zemplinina, which seems to be an acidogenic species under oxidative laboratory conditions. In this study, several strains of C. zemplinina were tested for organic acid metabolism, in comparison with Saccharomyces cerevisiae, Saccharomyces uvarum and Candida stellata, under fermentative conditions. Only C. stellata produced significantly higher acidity in simple minimal media (SM) with low sugar content and two different nitrogen sources (ammonia or glutamic acid) at low level. However, the acid profile differed largely between the Saccharomyces and Candida species and showed inverse types of N-dependence in some cases. Succinic acid production was strongly enhanced on glutamic acid in Saccharomyces species, but not in Candida species. 2-oxoglutarate production was strongly supported on ammonium nitrogen in Candida species, but remained low in Saccharomyces. Candida species, C. stellata in particular, produced more pyruvic acid regardless of N-sources. From the results, we concluded that the anaerobic organic acid metabolisms of C. zemplinina and C. stellata are different from each other and also from that of the Saccharomyces species. In the formation of succinic acid, the oxidative pathway from glutamic acid seems to play little or no role in C. zemplinina. The reductive branch of the TCA cycle, however, produces acidic intermediates (malic, fumaric, and succinic acid) in a level comparable with the production of the Saccharomyces species. An unidentified organic acid, which was produced on glutamic acid only by the Candida species, needs further investigation. Copyright © 2014 Elsevier B.V. All rights reserved.