WorldWideScience

Sample records for genome conserves centromeric

  1. Genome-wide characterization of centromeric satellites from multiple mammalian genomes.

    Science.gov (United States)

    Alkan, Can; Cardone, Maria Francesca; Catacchio, Claudia Rita; Antonacci, Francesca; O'Brien, Stephen J; Ryder, Oliver A; Purgato, Stefania; Zoli, Monica; Della Valle, Giuliano; Eichler, Evan E; Ventura, Mario

    2011-01-01

    Despite its importance in cell biology and evolution, the centromere has remained the final frontier in genome assembly and annotation due to its complex repeat structure. However, isolation and characterization of the centromeric repeats from newly sequenced species are necessary for a complete understanding of genome evolution and function. In recent years, various genomes have been sequenced, but the characterization of the corresponding centromeric DNA has lagged behind. Here, we present a computational method (RepeatNet) to systematically identify higher-order repeat structures from unassembled whole-genome shotgun sequence and test whether these sequence elements correspond to functional centromeric sequences. We analyzed genome datasets from six species of mammals representing the diversity of the mammalian lineage, namely, horse, dog, elephant, armadillo, opossum, and platypus. We define candidate monomer satellite repeats and demonstrate centromeric localization for five of the six genomes. Our analysis revealed the greatest diversity of centromeric sequences in horse and dog in contrast to elephant and armadillo, which showed high-centromeric sequence homogeneity. We could not isolate centromeric sequences within the platypus genome, suggesting that centromeres in platypus are not enriched in satellite DNA. Our method can be applied to the characterization of thousands of other vertebrate genomes anticipated for sequencing in the near future, providing an important tool for annotation of centromeres.

  2. The Past, Present, and Future of Human Centromere Genomics

    Directory of Open Access Journals (Sweden)

    Megan E. Aldrup-MacDonald

    2014-01-01

    Full Text Available The centromere is the chromosomal locus essential for chromosome inheritance and genome stability. Human centromeres are located at repetitive alpha satellite DNA arrays that compose approximately 5% of the genome. Contiguous alpha satellite DNA sequence is absent from the assembled reference genome, limiting current understanding of centromere organization and function. Here, we review the progress in centromere genomics spanning the discovery of the sequence to its molecular characterization and the work done during the Human Genome Project era to elucidate alpha satellite structure and sequence variation. We discuss exciting recent advances in alpha satellite sequence assembly that have provided important insight into the abundance and complex organization of this sequence on human chromosomes. In light of these new findings, we offer perspectives for future studies of human centromere assembly and function.

  3. Centromere Locations in Brassica A and C Genomes Revealed Through Half-Tetrad Analysis.

    Science.gov (United States)

    Mason, Annaliese S; Rousseau-Gueutin, Mathieu; Morice, Jérôme; Bayer, Philipp E; Besharat, Naghmeh; Cousin, Anouska; Pradhan, Aneeta; Parkin, Isobel A P; Chèvre, Anne-Marie; Batley, Jacqueline; Nelson, Matthew N

    2016-02-01

    Locating centromeres on genome sequences can be challenging. The high density of repetitive elements in these regions makes sequence assembly problematic, especially when using short-read sequencing technologies. It can also be difficult to distinguish between active and recently extinct centromeres through sequence analysis. An effective solution is to identify genetically active centromeres (functional in meiosis) by half-tetrad analysis. This genetic approach involves detecting heterozygosity along chromosomes in segregating populations derived from gametes (half-tetrads). Unreduced gametes produced by first division restitution mechanisms comprise complete sets of nonsister chromatids. Along these chromatids, heterozygosity is maximal at the centromeres, and homologous recombination events result in homozygosity toward the telomeres. We genotyped populations of half-tetrad-derived individuals (from Brassica interspecific hybrids) using a high-density array of physically anchored SNP markers (Illumina Brassica 60K Infinium array). Mapping the distribution of heterozygosity in these half-tetrad individuals allowed the genetic mapping of all 19 centromeres of the Brassica A and C genomes to the reference Brassica napus genome. Gene and transposable element density across the B. napus genome were also assessed and corresponded well to previously reported genetic map positions. Known centromere-specific sequences were located in the reference genome, but mostly matched unanchored sequences, suggesting that the core centromeric regions may not yet be assembled into the pseudochromosomes of the reference genome. The increasing availability of genetic markers physically anchored to reference genomes greatly simplifies the genetic and physical mapping of centromeres using half-tetrad analysis. We discuss possible applications of this approach, including in species where half-tetrads are currently difficult to isolate. Copyright © 2016 by the Genetics Society of America.

  4. Genome-wide analysis reveals a cell cycle-dependent mechanism controlling centromere propagation.

    Science.gov (United States)

    Erhardt, Sylvia; Mellone, Barbara G; Betts, Craig M; Zhang, Weiguo; Karpen, Gary H; Straight, Aaron F

    2008-12-01

    Centromeres are the structural and functional foundation for kinetochore formation, spindle attachment, and chromosome segregation. In this study, we isolated factors required for centromere propagation using genome-wide RNA interference screening for defects in centromere protein A (CENP-A; centromere identifier [CID]) localization in Drosophila melanogaster. We identified the proteins CAL1 and CENP-C as essential factors for CID assembly at the centromere. CID, CAL1, and CENP-C coimmunoprecipitate and are mutually dependent for centromere localization and function. We also identified the mitotic cyclin A (CYCA) and the anaphase-promoting complex (APC) inhibitor RCA1/Emi1 as regulators of centromere propagation. We show that CYCA is centromere localized and that CYCA and RCA1/Emi1 couple centromere assembly to the cell cycle through regulation of the fizzy-related/CDH1 subunit of the APC. Our findings identify essential components of the epigenetic machinery that ensures proper specification and propagation of the centromere and suggest a mechanism for coordinating centromere inheritance with cell division.

  5. Genome-wide analysis reveals a cell cycle–dependent mechanism controlling centromere propagation

    Science.gov (United States)

    Erhardt, Sylvia; Mellone, Barbara G.; Betts, Craig M.; Zhang, Weiguo; Karpen, Gary H.; Straight, Aaron F.

    2008-01-01

    Centromeres are the structural and functional foundation for kinetochore formation, spindle attachment, and chromosome segregation. In this study, we isolated factors required for centromere propagation using genome-wide RNA interference screening for defects in centromere protein A (CENP-A; centromere identifier [CID]) localization in Drosophila melanogaster. We identified the proteins CAL1 and CENP-C as essential factors for CID assembly at the centromere. CID, CAL1, and CENP-C coimmunoprecipitate and are mutually dependent for centromere localization and function. We also identified the mitotic cyclin A (CYCA) and the anaphase-promoting complex (APC) inhibitor RCA1/Emi1 as regulators of centromere propagation. We show that CYCA is centromere localized and that CYCA and RCA1/Emi1 couple centromere assembly to the cell cycle through regulation of the fizzy-related/CDH1 subunit of the APC. Our findings identify essential components of the epigenetic machinery that ensures proper specification and propagation of the centromere and suggest a mechanism for coordinating centromere inheritance with cell division. PMID:19047461

  6. De Novo Centromere Formation and Centromeric Sequence Expansion in Wheat and its Wide Hybrids.

    Directory of Open Access Journals (Sweden)

    Xiang Guo

    2016-04-01

    Full Text Available Centromeres typically contain tandem repeat sequences, but centromere function does not necessarily depend on these sequences. We identified functional centromeres with significant quantitative changes in the centromeric retrotransposons of wheat (CRW contents in wheat aneuploids (Triticum aestivum and the offspring of wheat wide hybrids. The CRW signals were strongly reduced or essentially lost in some wheat ditelosomic lines and in the addition lines from the wide hybrids. The total loss of the CRW sequences but the presence of CENH3 in these lines suggests that the centromeres were formed de novo. In wheat and its wide hybrids, which carry large complex genomes or no sequenced genome, we performed CENH3-ChIP-dot-blot methods alone or in combination with CENH3-ChIP-seq and identified the ectopic genomic sequences present at the new centromeres. In adcdition, the transcription of the identified DNA sequences was remarkably increased at the new centromere, suggesting that the transcription of the corresponding sequences may be associated with de novo centromere formation. Stable alien chromosomes with two and three regions containing CRW sequences induced by centromere breakage were observed in the wheat-Th. elongatum hybrid derivatives, but only one was a functional centromere. In wheat-rye (Secale cereale hybrids, the rye centromere-specific sequences spread along the chromosome arms and may have caused centromere expansion. Frequent and significant quantitative alterations in the centromere sequence via chromosomal rearrangement have been systematically described in wheat wide hybridizations, which may affect the retention or loss of the alien chromosomes in the hybrids. Thus, the centromere behavior in wide crosses likely has an important impact on the generation of biodiversity, which ultimately has implications for speciation.

  7. Epigenetically-inherited centromere and neocentromere DNA replicates earliest in S-phase.

    Directory of Open Access Journals (Sweden)

    Amnon Koren

    2010-08-01

    Full Text Available Eukaryotic centromeres are maintained at specific chromosomal sites over many generations. In the budding yeast Saccharomyces cerevisiae, centromeres are genetic elements defined by a DNA sequence that is both necessary and sufficient for function; whereas, in most other eukaryotes, centromeres are maintained by poorly characterized epigenetic mechanisms in which DNA has a less definitive role. Here we use the pathogenic yeast Candida albicans as a model organism to study the DNA replication properties of centromeric DNA. By determining the genome-wide replication timing program of the C. albicans genome, we discovered that each centromere is associated with a replication origin that is the first to fire on its respective chromosome. Importantly, epigenetic formation of new ectopic centromeres (neocentromeres was accompanied by shifts in replication timing, such that a neocentromere became the first to replicate and became associated with origin recognition complex (ORC components. Furthermore, changing the level of the centromere-specific histone H3 isoform led to a concomitant change in levels of ORC association with centromere regions, further supporting the idea that centromere proteins determine origin activity. Finally, analysis of centromere-associated DNA revealed a replication-dependent sequence pattern characteristic of constitutively active replication origins. This strand-biased pattern is conserved, together with centromere position, among related strains and species, in a manner independent of primary DNA sequence. Thus, inheritance of centromere position is correlated with a constitutively active origin of replication that fires at a distinct early time. We suggest a model in which the distinct timing of DNA replication serves as an epigenetic mechanism for the inheritance of centromere position.

  8. SWI/SNF-like chromatin remodeling factor Fun30 supports point centromere function in S. cerevisiae.

    Directory of Open Access Journals (Sweden)

    Mickaël Durand-Dubief

    2012-09-01

    Full Text Available Budding yeast centromeres are sequence-defined point centromeres and are, unlike in many other organisms, not embedded in heterochromatin. Here we show that Fun30, a poorly understood SWI/SNF-like chromatin remodeling factor conserved in humans, promotes point centromere function through the formation of correct chromatin architecture at centromeres. Our determination of the genome-wide binding and nucleosome positioning properties of Fun30 shows that this enzyme is consistently enriched over centromeres and that a majority of CENs show Fun30-dependent changes in flanking nucleosome position and/or CEN core micrococcal nuclease accessibility. Fun30 deletion leads to defects in histone variant Htz1 occupancy genome-wide, including at and around most centromeres. FUN30 genetically interacts with CSE4, coding for the centromere-specific variant of histone H3, and counteracts the detrimental effect of transcription through centromeres on chromosome segregation and suppresses transcriptional noise over centromere CEN3. Previous work has shown a requirement for fission yeast and mammalian homologs of Fun30 in heterochromatin assembly. As centromeres in budding yeast are not embedded in heterochromatin, our findings indicate a direct role of Fun30 in centromere chromatin by promoting correct chromatin architecture.

  9. Organization and evolution of primate centromeric DNA from whole-genome shotgun sequence data.

    Directory of Open Access Journals (Sweden)

    Can Alkan

    2007-09-01

    Full Text Available The major DNA constituent of primate centromeres is alpha satellite DNA. As much as 2%-5% of sequence generated as part of primate genome sequencing projects consists of this material, which is fragmented or not assembled as part of published genome sequences due to its highly repetitive nature. Here, we develop computational methods to rapidly recover and categorize alpha-satellite sequences from previously uncharacterized whole-genome shotgun sequence data. We present an algorithm to computationally predict potential higher-order array structure based on paired-end sequence data and then experimentally validate its organization and distribution by experimental analyses. Using whole-genome shotgun data from the human, chimpanzee, and macaque genomes, we examine the phylogenetic relationship of these sequences and provide further support for a model for their evolution and mutation over the last 25 million years. Our results confirm fundamental differences in the dispersal and evolution of centromeric satellites in the Old World monkey and ape lineages of evolution.

  10. Organization and evolution of primate centromeric DNA from whole-genome shotgun sequence data.

    Science.gov (United States)

    Alkan, Can; Ventura, Mario; Archidiacono, Nicoletta; Rocchi, Mariano; Sahinalp, S Cenk; Eichler, Evan E

    2007-09-01

    The major DNA constituent of primate centromeres is alpha satellite DNA. As much as 2%-5% of sequence generated as part of primate genome sequencing projects consists of this material, which is fragmented or not assembled as part of published genome sequences due to its highly repetitive nature. Here, we develop computational methods to rapidly recover and categorize alpha-satellite sequences from previously uncharacterized whole-genome shotgun sequence data. We present an algorithm to computationally predict potential higher-order array structure based on paired-end sequence data and then experimentally validate its organization and distribution by experimental analyses. Using whole-genome shotgun data from the human, chimpanzee, and macaque genomes, we examine the phylogenetic relationship of these sequences and provide further support for a model for their evolution and mutation over the last 25 million years. Our results confirm fundamental differences in the dispersal and evolution of centromeric satellites in the Old World monkey and ape lineages of evolution.

  11. Characterization of the genomic organization of the region bordering the centromere of chromosome V of Podospora anserina by direct sequencing.

    Science.gov (United States)

    Silar, Philippe; Barreau, Christian; Debuchy, Robert; Kicka, Sébastien; Turcq, Béatrice; Sainsard-Chanet, Annie; Sellem, Carole H; Billault, Alain; Cattolico, Laurence; Duprat, Simone; Weissenbach, Jean

    2003-08-01

    A Podospora anserina BAC library of 4800 clones has been constructed in the vector pBHYG allowing direct selection in fungi. Screening of the BAC collection for centromeric sequences of chromosome V allowed the recovery of clones localized on either sides of the centromere, but no BAC clone was found to contain the centromere. Seven BAC clones containing 322,195 and 156,244bp from either sides of the centromeric region were sequenced and annotated. One 5S rRNA gene, 5 tRNA genes, and 163 putative coding sequences (CDS) were identified. Among these, only six CDS seem specific to P. anserina. The gene density in the centromeric region is approximately one gene every 2.8kb. Extrapolation of this gene density to the whole genome of P. anserina suggests that the genome contains about 11,000 genes. Synteny analyses between P. anserina and Neurospora crassa show that co-linearity extends at the most to a few genes, suggesting rapid genome rearrangements between these two species.

  12. Recent advances in plant centromere biology.

    Science.gov (United States)

    Feng, Chao; Liu, YaLin; Su, HanDong; Wang, HeFei; Birchler, James; Han, FangPu

    2015-03-01

    The centromere, which is one of the essential parts of a chromosome, controls kinetochore formation and chromosome segregation during mitosis and meiosis. While centromere function is conserved in eukaryotes, the centromeric DNA sequences evolve rapidly and have few similarities among species. The histone H3 variant CENH3 (CENP-A in human), which mostly exists in centromeric nucleosomes, is a universal active centromere mark in eukaryotes and plays an essential role in centromere identity determination. The relationship between centromeric DNA sequences and centromere identity determination is one of the intriguing questions in studying centromere formation. Due to the discoveries in the past decades, including "neocentromeres" and "centromere inactivation", it is now believed that the centromere identity is determined by epigenetic mechanisms. This review will present recent progress in plant centromere biology.

  13. Dynamic epigenetic states of maize centromeres

    Directory of Open Access Journals (Sweden)

    Yalin eLiu

    2015-10-01

    Full Text Available The centromere is a specialized chromosomal region identified as the major constriction, upon which the kinetochore complex is formed, ensuring accurate chromosome orientation and segregation during cell division. The rapid evolution of centromere DNA sequence and the conserved centromere function are two contradictory aspects of centromere biology. Indeed, the sole presence of genetic sequence is not sufficient for centromere formation. Various dicentric chromosomes with one inactive centromere have been recognized. It has also been found that de novo centromere formation is common on fragments in which centromeric DNA sequences are lost. Epigenetic factors play important roles in centromeric chromatin assembly and maintenance. Nondisjunction of the supernumerary B chromosome early prophase of meiosis I requires an active centromere. This review discusses recent studies in maize about genetic and epigenetic elements regulating formation and maintenance of centromere chromatin, as well as centromere behavior in meiosis.

  14. Centromeric heterochromatin: the primordial segregation machine.

    Science.gov (United States)

    Bloom, Kerry S

    2014-01-01

    Centromeres are specialized domains of heterochromatin that provide the foundation for the kinetochore. Centromeric heterochromatin is characterized by specific histone modifications, a centromere-specific histone H3 variant (CENP-A), and the enrichment of cohesin, condensin, and topoisomerase II. Centromere DNA varies orders of magnitude in size from 125 bp (budding yeast) to several megabases (human). In metaphase, sister kinetochores on the surface of replicated chromosomes face away from each other, where they establish microtubule attachment and bi-orientation. Despite the disparity in centromere size, the distance between separated sister kinetochores is remarkably conserved (approximately 1 μm) throughout phylogeny. The centromere functions as a molecular spring that resists microtubule-based extensional forces in mitosis. This review explores the physical properties of DNA in order to understand how the molecular spring is built and how it contributes to the fidelity of chromosome segregation.

  15. A regulatory effect of INMAP on centromere proteins: antisense INMAP induces CENP-B variation and centromeric halo.

    Directory of Open Access Journals (Sweden)

    Tan Tan

    Full Text Available CENP-B is a highly conserved protein that facilitates the assembly of specific centromere structures both in interphase nuclei and on mitotic chromosomes. INMAP is a conserved protein that localizes at nucleus in interphase cells and at mitotic apparatus in mitotic cells. Our previous results showed that INMAP over-expression leads to spindle defects, mitotic arrest and formation of polycentrosomal and multinuclear cells, indicating that INMAP may modulate the function of (a key protein(s in mitotic apparatus. In this study, we demonstrate that INMAP interacts with CENP-B and promotes cleavage of the N-terminal DNA binding domain from CENP-B. The cleaved CENP-B cannot associate with centromeres and thus lose its centromere-related functions. Consistent with these results, CENP-B in INMAP knockdown cells becomes more diffused around kinetochores. Although INMAP knockdown cells do not exhibit gross defects in mitotic spindle formation, these cells go through mitosis, especially prophase and metaphase, with different relative timing, indicating subtle abnormality. These results identify INMAP as a model regulator of CENP-B and support the notion that INMAP regulates mitosis through modulating CENP-B-mediated centromere organization.

  16. Centromere separation and association in the nuclei of an interspecific hybrid between Torenia fournieri and T. baillonii (Scrophulariaceae) during mitosis and meiosis.

    Science.gov (United States)

    Kikuchi, Shinji; Tanaka, Hiroyuki; Wako, Toshiyuki; Tsujimoto, Hisashi

    2007-10-01

    In the nuclei of some interspecific hybrid and allopolyploid plant species, each genome occupies a separate spatial domain. To analyze this phenomenon, we studied localization of the centromeres in the nuclei of a hybrid between Torenia fournieri and T. baillonii during mitosis and meiosis using three-dimensional fluorescence in situ hybridization (3D-FISH) probed with species-specific centromere repeats. Centromeres of each genome were located separately in undifferentiated cells but not differentiated cells, suggesting that cell division might be the possible force causing centromere separation. However, no remarkable difference of dividing distance was detected between chromatids with different centromeres in anaphase and telophase, indicating that tension of the spindle fiber attached to each chromatid is not the cause of centromere separation in Torenia. In differentiated cells, centromeres in both genomes were not often observed for the expected chromosome number, indicating centromere association. In addition, association of centromeres from the same genome was observed at a higher frequency than between different genomes. This finding suggests that centromeres within one genome are spatially separated from those within the other. This close position may increase possibility of association between centromeres of the same genome. In meiotic prophase, all centromeres irrespective of the genome were associated in a certain portion of the nucleus. Since centromere association in the interspecific hybrid and amphiploid was tighter than that in the diploid parents, it is possible that this phenomenon may be involved in sorting and pairing of homologous chromosomes.

  17. Structure, Function, and Evolution of Rice Centromeres

    Energy Technology Data Exchange (ETDEWEB)

    Jiang, Jiming

    2010-02-04

    The centromere is the most characteristic landmark of eukaryotic chromosomes. Centromeres function as the site for kinetochore assembly and spindle attachment, allowing for the faithful pairing and segregation of sister chromatids during cell division. Characterization of centromeric DNA is not only essential to understand the structure and organization of plant genomes, but it is also a critical step in the development of plant artificial chromosomes. The centromeres of most model eukaryotic species, consist predominantly of long arrays of satellite DNA. Determining the precise DNA boundary of a centromere has proven to be a difficult task in multicellular eukaryotes. We have successfully cloned and sequenced the centromere of rice chromosome 8 (Cen8), representing the first fully sequenced centromere from any multicellular eukaryotes. The functional core of Cen8 spans ~800 kb of DNA, which was determined by chromatin immunoprecipitation (ChIP) using an antibody against the rice centromere-specific H3 histone. We discovered 16 actively transcribed genes distributed throughout the Cen8 region. In addition to Cen8, we have characterized eight additional rice centromeres using the next generation sequencing technology. We discovered four subfamilies of the CRR retrotransposon that is highly enriched in rice centromeres. CRR elements are constitutively transcribed and different CRR subfamilies are differentially processed by RNAi. These results suggest that different CRR subfamilies may play different roles in the RNAi-mediated pathway for formation and maintenance of centromeric chromatin.

  18. Centromeric DNA characterization in the model grass Brachypodium distachyon provides insights on the evolution of the genus.

    Science.gov (United States)

    Li, Yinjia; Zuo, Sheng; Zhang, Zhiliang; Li, Zhanjie; Han, Jinlei; Chu, Zhaoqing; Hasterok, Robert; Wang, Kai

    2018-03-01

    Brachypodium distachyon is a well-established model monocot plant, and its small and compact genome has been used as an accurate reference for the much larger and often polyploid genomes of cereals such as Avena sativa (oats), Hordeum vulgare (barley) and Triticum aestivum (wheat). Centromeres are indispensable functional units of chromosomes and they play a core role in genome polyploidization events during evolution. As the Brachypodium genus contains about 20 species that differ significantly in terms of their basic chromosome numbers, genome size, ploidy levels and life strategies, studying their centromeres may provide important insight into the structure and evolution of the genome in this interesting and important genus. In this study, we isolated the centromeric DNA of the B. distachyon reference line Bd21 and characterized its composition via the chromatin immunoprecipitation of the nucleosomes that contain the centromere-specific histone CENH3. We revealed that the centromeres of Bd21 have the features of typical multicellular eukaryotic centromeres. Strikingly, these centromeres contain relatively few centromeric satellite DNAs; in particular, the centromere of chromosome 5 (Bd5) consists of only ~40 kb. Moreover, the centromeric retrotransposons in B. distachyon (CRBds) are evolutionarily young. These transposable elements are located both within and adjacent to the CENH3 binding domains, and have similar compositions. Moreover, based on the presence of CRBds in the centromeres, the species in this study can be grouped into two distinct lineages. This may provide new evidence regarding the phylogenetic relationships within the Brachypodium genus. © 2018 The Authors The Plant Journal © 2018 John Wiley & Sons Ltd.

  19. A segment of the apospory-specific genomic region is highly microsyntenic not only between the apomicts Pennisetum squamulatum and buffelgrass, but also with a rice chromosome 11 centromeric-proximal genomic region.

    Science.gov (United States)

    Gualtieri, Gustavo; Conner, Joann A; Morishige, Daryl T; Moore, L David; Mullet, John E; Ozias-Akins, Peggy

    2006-03-01

    Bacterial artificial chromosome (BAC) clones from apomicts Pennisetum squamulatum and buffelgrass (Cenchrus ciliaris), isolated with the apospory-specific genomic region (ASGR) marker ugt197, were assembled into contigs that were extended by chromosome walking. Gene-like sequences from contigs were identified by shotgun sequencing and BLAST searches, and used to isolate orthologous rice contigs. Additional gene-like sequences in the apomicts' contigs were identified by bioinformatics using fully sequenced BACs from orthologous rice contigs as templates, as well as by interspecies, whole-contig cross-hybridizations. Hierarchical contig orthology was rapidly assessed by constructing detailed long-range contig molecular maps showing the distribution of gene-like sequences and markers, and searching for microsyntenic patterns of sequence identity and spatial distribution within and across species contigs. We found microsynteny between P. squamulatum and buffelgrass contigs. Importantly, this approach also enabled us to isolate from within the rice (Oryza sativa) genome contig Rice A, which shows the highest microsynteny and is most orthologous to the ugt197-containing C1C buffelgrass contig. Contig Rice A belongs to the rice genome database contig 77 (according to the current September 12, 2003, rice fingerprint contig build) that maps proximal to the chromosome 11 centromere, a feature that interestingly correlates with the mapping of ASGR-linked BACs proximal to the centromere or centromere-like sequences. Thus, relatedness between these two orthologous contigs is supported both by their molecular microstructure and by their centromeric-proximal location. Our discoveries promote the use of a microsynteny-based positional-cloning approach using the rice genome as a template to aid in constructing the ASGR toward the isolation of genes underlying apospory.

  20. Identification of the centromeric repeat in the threespine stickleback fish (Gasterosteus aculeatus).

    Science.gov (United States)

    Cech, Jennifer N; Peichel, Catherine L

    2015-12-01

    Centromere sequences exist as gaps in many genome assemblies due to their repetitive nature. Here we take an unbiased approach utilizing centromere protein A (CENP-A) chomatin immunoprecipitation followed by high-throughput sequencing to identify the centromeric repeat sequence in the threespine stickleback fish (Gasterosteus aculeatus). A 186-bp, AT-rich repeat was validated as centromeric using both fluorescence in situ hybridization (FISH) and immunofluorescence combined with FISH (IF-FISH) on interphase nuclei and metaphase spreads. This repeat hybridizes strongly to the centromere on all chromosomes, with the exception of weak hybridization to the Y chromosome. Together, our work provides the first validated sequence information for the threespine stickleback centromere.

  1. A Segment of the Apospory-Specific Genomic Region Is Highly Microsyntenic Not Only between the Apomicts Pennisetum squamulatum and Buffelgrass, But Also with a Rice Chromosome 11 Centromeric-Proximal Genomic Region1[W

    Science.gov (United States)

    Gualtieri, Gustavo; Conner, Joann A.; Morishige, Daryl T.; Moore, L. David; Mullet, John E.; Ozias-Akins, Peggy

    2006-01-01

    Bacterial artificial chromosome (BAC) clones from apomicts Pennisetum squamulatum and buffelgrass (Cenchrus ciliaris), isolated with the apospory-specific genomic region (ASGR) marker ugt197, were assembled into contigs that were extended by chromosome walking. Gene-like sequences from contigs were identified by shotgun sequencing and BLAST searches, and used to isolate orthologous rice contigs. Additional gene-like sequences in the apomicts' contigs were identified by bioinformatics using fully sequenced BACs from orthologous rice contigs as templates, as well as by interspecies, whole-contig cross-hybridizations. Hierarchical contig orthology was rapidly assessed by constructing detailed long-range contig molecular maps showing the distribution of gene-like sequences and markers, and searching for microsyntenic patterns of sequence identity and spatial distribution within and across species contigs. We found microsynteny between P. squamulatum and buffelgrass contigs. Importantly, this approach also enabled us to isolate from within the rice (Oryza sativa) genome contig Rice A, which shows the highest microsynteny and is most orthologous to the ugt197-containing C1C buffelgrass contig. Contig Rice A belongs to the rice genome database contig 77 (according to the current September 12, 2003, rice fingerprint contig build) that maps proximal to the chromosome 11 centromere, a feature that interestingly correlates with the mapping of ASGR-linked BACs proximal to the centromere or centromere-like sequences. Thus, relatedness between these two orthologous contigs is supported both by their molecular microstructure and by their centromeric-proximal location. Our discoveries promote the use of a microsynteny-based positional-cloning approach using the rice genome as a template to aid in constructing the ASGR toward the isolation of genes underlying apospory. PMID:16415213

  2. High quality maize centromere 10 sequence reveals evidence of frequent recombination events

    Directory of Open Access Journals (Sweden)

    Thomas Kai Wolfgruber

    2016-03-01

    Full Text Available The ancestral centromeres of maize contain long stretches of the tandemly arranged CentC repeat. The abundance of tandem DNA repeats and centromeric retrotransposons (CR have presented a significant challenge to completely assembling centromeres using traditional sequencing methods. Here we report a nearly complete assembly of the 1.85 Mb maize centromere 10 from inbred B73 using PacBio technology and BACs from the reference genome project. The error rates estimated from overlapping BAC sequences are 7 x 10-6 and 5 x 10-5 for mismatches and indels, respectively. The number of gaps in the region covered by the reassembly was reduced from 140 in the reference genome to three. Three expressed genes are located between 92 and 477 kb of the inferred ancestral CentC cluster, which lies within the region of highest centromeric repeat density. The improved assembly increased the count of full-length centromeric retrotransposons from 5 to 55 and revealed a 22.7 kb segmental duplication that occurred approximately 121,000 years ago. Our analysis provides evidence of frequent recombination events in the form of partial retrotransposons, deletions within retrotransposons, chimeric retrotransposons, segmental duplications including higher order CentC repeats, a deleted CentC monomer, centromere-proximal inversions, and insertion of mitochondrial sequences. Double-strand DNA break (DSB repair is the most plausible mechanism for these events and may be the major driver of centromere repeat evolution and diversity. This repair appears to be mediated by microhomology, suggesting that tandem repeats may have evolved to facilitate the repair of frequent DSBs in centromeres.

  3. Centromere Destiny in Dicentric Chromosomes: New Insights from the Evolution of Human Chromosome 2 Ancestral Centromeric Region.

    Science.gov (United States)

    Chiatante, Giorgia; Giannuzzi, Giuliana; Calabrese, Francesco Maria; Eichler, Evan E; Ventura, Mario

    2017-07-01

    Dicentric chromosomes are products of genomic rearrangements that place two centromeres on the same chromosome. Due to the presence of two primary constrictions, they are inherently unstable and overcome their instability by epigenetically inactivating and/or deleting one of the two centromeres, thus resulting in functionally monocentric chromosomes that segregate normally during cell division. Our understanding to date of dicentric chromosome formation, behavior and fate has been largely inferred from observational studies in plants and humans as well as artificially produced de novo dicentrics in yeast and in human cells. We investigate the most recent product of a chromosome fusion event fixed in the human lineage, human chromosome 2, whose stability was acquired by the suppression of one centromere, resulting in a unique difference in chromosome number between humans (46 chromosomes) and our most closely related ape relatives (48 chromosomes). Using molecular cytogenetics, sequencing, and comparative sequence data, we deeply characterize the relicts of the chromosome 2q ancestral centromere and its flanking regions, gaining insight into the ancestral organization that can be easily broadened to all acrocentric chromosome centromeres. Moreover, our analyses offered the opportunity to trace the evolutionary history of rDNA and satellite III sequences among great apes, thus suggesting a new hypothesis for the preferential inactivation of some human centromeres, including IIq. Our results suggest two possible centromere inactivation models to explain the evolutionarily stabilization of human chromosome 2 over the last 5-6 million years. Our results strongly favor centromere excision through a one-step process. © The Author 2017. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

  4. Molecular characterization and chromosomal distribution of a species-specific transcribed centromeric satellite repeat from the olive fruit fly, Bactrocera oleae.

    Directory of Open Access Journals (Sweden)

    Konstantina T Tsoumani

    Full Text Available Satellite repetitive sequences that accumulate in the heterochromatin consist a large fraction of a genome and due to their properties are suggested to be implicated in centromere function. Current knowledge of heterochromatic regions of Bactrocera oleae genome, the major pest of the olive tree, is practically nonexistent. In our effort to explore the repetitive DNA portion of B. oleae genome, a novel satellite sequence designated BoR300 was isolated and cloned. The present study describes the genomic organization, abundance and chromosomal distribution of BoR300 which is organized in tandem, forming arrays of 298 bp-long monomers. Sequence analysis showed an AT content of 60.4%, a CENP-B like-motif and a high curvature value based on predictive models. Comparative analysis among randomly selected monomers demonstrated a high degree of sequence homogeneity (88%-97% of BoR300 repeats, which are present at approximately 3,000 copies per haploid genome accounting for about 0.28% of the total genomic DNA, based on two independent qPCR approaches. In addition, expression of the repeat was also confirmed through RT-PCR, by which BoR300 transcripts were detected in both sexes. Fluorescence in situ hybridization (FISH of BoR300 on mitotic metaphases and polytene chromosomes revealed signals to the centromeres of two out of the six chromosomes which indicated a chromosome-specific centromeric localization. Moreover, BoR300 is not conserved in the closely related Bactrocera species tested and it is also absent in other dipterans, but it's rather restricted to the B. oleae genome. This feature of species-specificity attributed to BoR300 satellite makes it a good candidate as an identification probe of the insect among its relatives at early development stages.

  5. RNA Pol II promotes transcription of centromeric satellite DNA in beetles.

    Directory of Open Access Journals (Sweden)

    Zeljka Pezer

    Full Text Available Transcripts of centromeric satellite DNAs are known to play a role in heterochromatin formation as well as in establishment of the kinetochore. However, little is known about basic mechanisms of satellite DNA expression within constitutive heterochromatin and its regulation. Here we present comprehensive analysis of transcription of abundant centromeric satellite DNA, PRAT from beetle Palorus ratzeburgii (Coleoptera. This satellite is characterized by preservation and extreme sequence conservation among evolutionarily distant insect species. PRAT is expressed in all three developmental stages: larvae, pupae and adults at similar level. Transcripts are abundant comprising 0.033% of total RNA and are heterogeneous in size ranging from 0.5 kb up to more than 5 kb. Transcription proceeds from both strands but with 10 fold different expression intensity and transcripts are not processed into siRNAs. Most of the transcripts (80% are not polyadenylated and remain in the nucleus while a small portion is exported to the cytoplasm. Multiple, irregularly distributed transcription initiation sites as well as termination sites have been mapped within the PRAT sequence using primer extension and RLM-RACE. The presence of cap structure as well as poly(A tails in a portion of the transcripts indicate RNA polymerase II-dependent transcription and a putative polymerase II promoter site overlaps the most conserved part of the PRAT sequence. The treatment of larvae with alpha-amanitin decreases the level of PRAT transcripts at concentrations that selectively inhibit pol II activity. In conclusion, stable, RNA polymerase II dependant transcripts of abundant centromeric satellite DNA, not regulated by RNAi, have been identified and characterized. This study offers a basic understanding of expression of highly abundant heterochromatic DNA which in beetle species constitutes up to 50% of the genome.

  6. The Role of Dicentric Chromosome Formation and Secondary Centromere Deletion in the Evolution of Myeloid Malignancy

    Science.gov (United States)

    MacKinnon, Ruth N.; Campbell, Lynda J.

    2011-01-01

    Dicentric chromosomes have been identified as instigators of the genome instability associated with cancer, but this instability is often resolved by one of a number of different secondary events. These include centromere inactivation, inversion, and intercentromeric deletion. Deletion or excision of one of the centromeres may be a significant occurrence in myeloid malignancy and other malignancies but has not previously been widely recognized, and our reports are the first describing centromere deletion in cancer cells. We review what is known about dicentric chromosomes and the mechanisms by which they can undergo stabilization in both constitutional and cancer genomes. The failure to identify centromere deletion in cancer cells until recently can be partly explained by the standard approaches to routine diagnostic cancer genome analysis, which do not identify centromeres in the context of chromosome organization. This hitherto hidden group of primary dicentric, secondary monocentric chromosomes, together with other unrecognized dicentric chromosomes, points to a greater role for dicentric chromosomes in cancer initiation and progression than is generally acknowledged. We present a model that predicts and explains a significant role for dicentric chromosomes in the formation of unbalanced translocations in malignancy. PMID:22567363

  7. The Role of Dicentric Chromosome Formation and Secondary Centromere Deletion in the Evolution of Myeloid Malignancy

    Directory of Open Access Journals (Sweden)

    Ruth N. MacKinnon

    2011-01-01

    Full Text Available Dicentric chromosomes have been identified as instigators of the genome instability associated with cancer, but this instability is often resolved by one of a number of different secondary events. These include centromere inactivation, inversion, and intercentromeric deletion. Deletion or excision of one of the centromeres may be a significant occurrence in myeloid malignancy and other malignancies but has not previously been widely recognized, and our reports are the first describing centromere deletion in cancer cells. We review what is known about dicentric chromosomes and the mechanisms by which they can undergo stabilization in both constitutional and cancer genomes. The failure to identify centromere deletion in cancer cells until recently can be partly explained by the standard approaches to routine diagnostic cancer genome analysis, which do not identify centromeres in the context of chromosome organization. This hitherto hidden group of primary dicentric, secondary monocentric chromosomes, together with other unrecognized dicentric chromosomes, points to a greater role for dicentric chromosomes in cancer initiation and progression than is generally acknowledged. We present a model that predicts and explains a significant role for dicentric chromosomes in the formation of unbalanced translocations in malignancy.

  8. Centromeric DNA replication reconstitution reveals DNA loops and ATR checkpoint suppression.

    Science.gov (United States)

    Aze, Antoine; Sannino, Vincenzo; Soffientini, Paolo; Bachi, Angela; Costanzo, Vincenzo

    2016-06-01

    Half of the human genome is made up of repetitive DNA. However, mechanisms underlying replication of chromosome regions containing repetitive DNA are poorly understood. We reconstituted replication of defined human chromosome segments using bacterial artificial chromosomes in Xenopus laevis egg extract. Using this approach we characterized the chromatin assembly and replication dynamics of centromeric alpha-satellite DNA. Proteomic analysis of centromeric chromatin revealed replication-dependent enrichment of a network of DNA repair factors including the MSH2-6 complex, which was required for efficient centromeric DNA replication. However, contrary to expectations, the ATR-dependent checkpoint monitoring DNA replication fork arrest could not be activated on highly repetitive DNA due to the inability of the single-stranded DNA binding protein RPA to accumulate on chromatin. Electron microscopy of centromeric DNA and supercoil mapping revealed the presence of topoisomerase I-dependent DNA loops embedded in a protein matrix enriched for SMC2-4 proteins. This arrangement suppressed ATR signalling by preventing RPA hyper-loading, facilitating replication of centromeric DNA. These findings have important implications for our understanding of repetitive DNA metabolism and centromere organization under normal and stressful conditions.

  9. Conservation of the Centromere/Kinetochore Protein ZW10

    OpenAIRE

    Starr, Daniel A.; Williams, Byron C.; Li, Zexiao; Etemad-Moghadam, Bijan; Dawe, R. Kelly; Goldberg, Michael L.

    1997-01-01

    Mutations in the essential Drosophila melanogaster gene zw10 disrupt chromosome segregation, producing chromosomes that lag at the metaphase plate during anaphase of mitosis and both meiotic divisions. Recent evidence suggests that the product of this gene, DmZW10, acts at the kinetochore as part of a tension-sensing checkpoint at anaphase onset. DmZW10 displays an intriguing cell cycle–dependent intracellular distribution, apparently moving from the centromere/kinetochore at prometaphase to ...

  10. Unique small RNA signatures uncovered in the tammar wallaby genome

    Directory of Open Access Journals (Sweden)

    Lindsay James

    2012-10-01

    Full Text Available Abstract Background Small RNAs have proven to be essential regulatory molecules encoded within eukaryotic genomes. These short RNAs participate in a diverse array of cellular processes including gene regulation, chromatin dynamics and genome defense. The tammar wallaby, a marsupial mammal, is a powerful comparative model for studying the evolution of regulatory networks. As part of the genome sequencing initiative for the tammar, we have explored the evolution of each of the major classes of mammalian small RNAs in an Australian marsupial for the first time, including the first genome-scale analysis of the newest class of small RNAs, centromere repeat associated short interacting RNAs (crasiRNAs. Results Using next generation sequencing, we have characterized the major classes of small RNAs, micro (mi RNAs, piwi interacting (pi RNAs, and the centromere repeat associated short interacting (crasi RNAs in the tammar. We examined each of these small RNA classes with respect to the newly assembled tammar wallaby genome for gene and repeat features, salient features that define their canonical sequences, and the constitution of both highly conserved and species-specific members. Using a combination of miRNA hairpin predictions and co-mapping with miRBase entries, we identified a highly conserved cluster of miRNA genes on the X chromosome in the tammar and a total of 94 other predicted miRNA producing genes. Mapping all miRNAs to the tammar genome and comparing target genes among tammar, mouse and human, we identified 163 conserved target genes. An additional nine genes were identified in tammar that do not have an orthologous miRNA target in human and likely represent novel miRNA-regulated genes in the tammar. A survey of the tammar gonadal piRNAs shows that these small RNAs are enriched in retroelements and carry members from both marsupial and tammar-specific repeat classes. Lastly, this study includes the first in-depth analyses of the newly

  11. Analysis of DNA restriction fragments greater than 5.7 Mb in size from the centromeric region of human chromosomes.

    Science.gov (United States)

    Arn, P H; Li, X; Smith, C; Hsu, M; Schwartz, D C; Jabs, E W

    1991-01-01

    Pulsed electrophoresis was used to study the organization of the human centromeric region. Genomic DNA was digested with rare-cutting enzymes. DNA fragments from 0.2 to greater than 5.7 Mb were separated by electrophoresis and hybridized with alphoid and simple DNA repeats. Rare-cutting enzymes (Mlu I, Nar I, Not I, Nru I, Sal I, Sfi I, Sst II) demonstrated fewer restriction sites at centromeric regions than elsewhere in the genome. The enzyme Not I had the fewest restriction sites at centromeric regions. As much as 70% of these sequences from the centromeric region are present in Not I DNA fragments greater than 5.7 and estimated to be as large as 10 Mb in size. Other repetitive sequences such as short interspersed repeated segments (SINEs), long interspersed repeated segments (LINEs), ribosomal DNA, and mini-satellite DNA that are not enriched at the centromeric region, are not enriched in Not I fragments of greater than 5.7 Mb in size.

  12. Telomere-Centromere-Driven Genomic Instability Contributes to Karyotype Evolution in a Mouse Model of Melanoma

    Directory of Open Access Journals (Sweden)

    Amanda Gonçalves dos Santos Silva

    2010-01-01

    Full Text Available Aneuploidy and chromosomal instability (CIN are hallmarks of most solid tumors. These alterations may result from inaccurate chromosomal segregation during mitosis, which can occur through several mechanisms including defective telomere metabolism, centrosome amplification, dysfunctional centromeres, and/or defective spindle checkpoint control. In this work, we used an in vitro murine melanoma model that uses a cellular adhesion blockade as a transforming factor to characterize telomeric and centromeric alterations that accompany melanocyte transformation. To study the timing of the occurrence of telomere shortening in this transformation model, we analyzed the profile of telomere length by quantitative fluorescent in situ hybridization and found that telomere length significantly decreased as additional rounds of cell adhesion blockages were performed. Together with it, an increase in telomere-free ends and complex karyotypic aberrations were also found, which include Robertsonian fusions in 100% of metaphases of the metastatic melanoma cells. These findings are in agreement with the idea that telomere length abnormalities seem to be one of the earliest genetic alterations acquired in the multistep process of malignant transformation and that telomere abnormalities result in telomere aggregation, breakage-bridge-fusion cycles, and CIN. Another remarkable feature of this model is the abundance of centromeric instability manifested as centromere fragments and centromeric fusions. Taken together, our results illustrate for this melanoma model CIN with a structural signature of centromere breakage and telomeric loss.

  13. Cis-Acting Determinants Affecting Centromere Function, Sister-Chromatid Cohesion and Reciprocal Recombination during Meiosis in Saccharomyces Cerevisiae

    OpenAIRE

    Sears, D. D.; Hegemann, J. H.; Shero, J. H.; Hieter, P.

    1995-01-01

    We have employed a system that utilizes homologous pairs of human DNA-derived yeast artificial chromosomes (YACs) as marker chromosomes to assess the specific role (s) of conserved centromere DNA elements (CDEI, CDEII and CDEIII) in meiotic chromosome disjunction fidelity. Thirteen different centromere (CEN) mutations were tested for their effects on meiotic centromere function. YACs containing a wild-type CEN DNA sequence segregate with high fidelity in meiosis I (99% normal segregation) and...

  14. Mislocalization of the Drosophila centromere-specific histone CIDpromotes formation of functional ectopic kinetochores

    Energy Technology Data Exchange (ETDEWEB)

    Heun, Patrick; Erhardt, Sylvia; Blower, Michael D.; Weiss,Samara; Skora, Andrew D.; Karpen, Gary H.

    2006-01-30

    The centromere-specific histone variant CENP-A (CID in Drosophila) is a structural and functional foundation for kinetochore formation and chromosome segregation. Here, we show that overexpressed CID is mislocalized into normally non-centromeric regions in Drosophila tissue culture cells and animals. Analysis of mitoses in living and fixed cells reveals that mitotic delays, anaphase bridges, chromosome fragmentation, and cell and organismal lethality are all direct consequences of CID mislocalization. In addition, proteins that are normally restricted to endogenous kinetochores assemble at a subset of ectopic CID incorporation regions. The presence of microtubule motors and binding proteins, spindle attachments, and aberrant chromosome morphologies demonstrate that these ectopic kinetochores are functional. We conclude that CID mislocalization promotes formation of ectopic centromeres and multicentric chromosomes, which causes chromosome missegregation, aneuploidy, and growth defects. Thus, CENP-A mislocalization is one possible mechanism for genome instability during cancer progression, as well as centromere plasticity during evolution.

  15. Analysis of ParB-centromere interactions by multiplex SPR imaging reveals specific patterns for binding ParB in six centromeres of Burkholderiales chromosomes and plasmids.

    Science.gov (United States)

    Pillet, Flavien; Passot, Fanny Marie; Pasta, Franck; Anton Leberre, Véronique; Bouet, Jean-Yves

    2017-01-01

    Bacterial centromeres-also called parS, are cis-acting DNA sequences which, together with the proteins ParA and ParB, are involved in the segregation of chromosomes and plasmids. The specific binding of ParB to parS nucleates the assembly of a large ParB/DNA complex from which ParA-the motor protein, segregates the sister replicons. Closely related families of partition systems, called Bsr, were identified on the chromosomes and large plasmids of the multi-chromosomal bacterium Burkholderia cenocepacia and other species from the order Burkholeriales. The centromeres of the Bsr partition families are 16 bp palindromes, displaying similar base compositions, notably a central CG dinucleotide. Despite centromeres bind the cognate ParB with a narrow specificity, weak ParB-parS non cognate interactions were nevertheless detected between few Bsr partition systems of replicons not belonging to the same genome. These observations suggested that Bsr partition systems could have a common ancestry but that evolution mostly erased the possibilities of cross-reactions between them, in particular to prevent replicon incompatibility. To detect novel similarities between Bsr partition systems, we have analyzed the binding of six Bsr parS sequences and a wide collection of modified derivatives, to their cognate ParB. The study was carried out by Surface Plasmon Resonance imaging (SPRi) mulitplex analysis enabling a systematic survey of each nucleotide position within the centromere. We found that in each parS some positions could be changed while maintaining binding to ParB. Each centromere displays its own pattern of changes, but some positions are shared more or less widely. In addition from these changes we could speculate evolutionary links between these centromeres.

  16. Neocentromeres Provide Chromosome Segregation Accuracy and Centromere Clustering to Multiple Loci along a Candida albicans Chromosome.

    Directory of Open Access Journals (Sweden)

    Laura S Burrack

    2016-09-01

    Full Text Available Assembly of kinetochore complexes, involving greater than one hundred proteins, is essential for chromosome segregation and genome stability. Neocentromeres, or new centromeres, occur when kinetochores assemble de novo, at DNA loci not previously associated with kinetochore proteins, and they restore chromosome segregation to chromosomes lacking a functional centromere. Neocentromeres have been observed in a number of diseases and may play an evolutionary role in adaptation or speciation. However, the consequences of neocentromere formation on chromosome missegregation rates, gene expression, and three-dimensional (3D nuclear structure are not well understood. Here, we used Candida albicans, an organism with small, epigenetically-inherited centromeres, as a model system to study the functions of twenty different neocentromere loci along a single chromosome, chromosome 5. Comparison of neocentromere properties relative to native centromere functions revealed that all twenty neocentromeres mediated chromosome segregation, albeit to different degrees. Some neocentromeres also caused reduced levels of transcription from genes found within the neocentromere region. Furthermore, like native centromeres, neocentromeres clustered in 3D with active/functional centromeres, indicating that formation of a new centromere mediates the reorganization of 3D nuclear architecture. This demonstrates that centromere clustering depends on epigenetically defined function and not on the primary DNA sequence, and that neocentromere function is independent of its distance from the native centromere position. Together, the results show that a neocentromere can form at many loci along a chromosome and can support the assembly of a functional kinetochore that exhibits native centromere functions including chromosome segregation accuracy and centromere clustering within the nucleus.

  17. Identification and Preliminary Analysis of Several Centromere-associated Bacterial Artificial Chromosome Clones from a Diploid Wheat Library

    Institute of Scientific and Technical Information of China (English)

    2006-01-01

    Although the centromeres of some plants have been investigated previously, our knowledge of the wheat centromere is still very limited. To understand the structure and function of the wheat centromere, we used two centromeric repeats (RCS1 and CCS1-5ab) to obtain some centromere-associated bacterial artificial chromosome (BAC) clones in 32 RCS1-related BAC clones that had been screened out from a diploid wheat (Triticum boeoticum Boiss.; 2n=2x=14) BAC library. Southern hybridization results indicated that, of the 32 candidates,there were 28 RCS1-positive clones. Based on gel blot patterns, the frequency of RCS1 was approximately one copy every 69.4 kb in these 28 RCS1-positive BAC clones. More bands were detected when the same filter was probed with CCS1-5ab. Furthermore, the CCS1 bands covered all the bands detected by RCS1, which suggests that some CCS1 repeats were distributed together with RCS1. The frequency of CCS1 families was once every 35.8 kb, nearly twice that of RCS1. Fluorescence in situ hybridization (FISH) analysis indicated that the five BAC clones containing RCS1 and CCS1 sequences all detected signals at the centromeric regions in hexaploid wheat, but the signal intensities on the A-genome chromosomes were stronger than those on the B- and/or D-genome chromosomes. The FISH analysis among nine Triticeae cereals indicated that there were A-genomespecific (or rich) sequences dispersing on chromosome arms in the BAC clone TbBAC5. In addition, at the interphase cells, the centromeres of diploid species usually clustered at one pole and formed a ring-like allocation in the period before metaphase.

  18. Analysis of ParB-centromere interactions by multiplex SPR imaging reveals specific patterns for binding ParB in six centromeres of Burkholderiales chromosomes and plasmids.

    Directory of Open Access Journals (Sweden)

    Flavien Pillet

    Full Text Available Bacterial centromeres-also called parS, are cis-acting DNA sequences which, together with the proteins ParA and ParB, are involved in the segregation of chromosomes and plasmids. The specific binding of ParB to parS nucleates the assembly of a large ParB/DNA complex from which ParA-the motor protein, segregates the sister replicons. Closely related families of partition systems, called Bsr, were identified on the chromosomes and large plasmids of the multi-chromosomal bacterium Burkholderia cenocepacia and other species from the order Burkholeriales. The centromeres of the Bsr partition families are 16 bp palindromes, displaying similar base compositions, notably a central CG dinucleotide. Despite centromeres bind the cognate ParB with a narrow specificity, weak ParB-parS non cognate interactions were nevertheless detected between few Bsr partition systems of replicons not belonging to the same genome. These observations suggested that Bsr partition systems could have a common ancestry but that evolution mostly erased the possibilities of cross-reactions between them, in particular to prevent replicon incompatibility. To detect novel similarities between Bsr partition systems, we have analyzed the binding of six Bsr parS sequences and a wide collection of modified derivatives, to their cognate ParB. The study was carried out by Surface Plasmon Resonance imaging (SPRi mulitplex analysis enabling a systematic survey of each nucleotide position within the centromere. We found that in each parS some positions could be changed while maintaining binding to ParB. Each centromere displays its own pattern of changes, but some positions are shared more or less widely. In addition from these changes we could speculate evolutionary links between these centromeres.

  19. Centromeres cluster de novo at the beginning of meiosis in Brachypodium distachyon.

    Directory of Open Access Journals (Sweden)

    Ruoyu Wen

    Full Text Available In most eukaryotes that have been studied, the telomeres cluster into a bouquet early in meiosis, and in wheat and its relatives and in Arabidopsis the centromeres pair at the same time. In Arabidopsis, the telomeres do not cluster as a typical telomere bouquet on the nuclear membrane but are associated with the nucleolus both somatically and at the onset of meiosis. We therefore assessed whether Brachypodium distachyon, a monocot species related to cereals and whose genome is approximately twice the size of Arabidopsis thaliana, also exhibited an atypical telomere bouquet and centromere pairing. In order to investigate the occurrence of a bouquet and centromere pairing in B distachyon, we first had to establish protocols for studying meiosis in this species. This enabled us to visualize chromosome behaviour in meiocytes derived from young B distachyon spikelets in three-dimensions by fluorescent in situ hybridization (FISH, and accurately to stage meiosis based on chromatin morphology in relation to spikelet size and the timing of sample collection. Surprisingly, this study revealed that the centromeres clustered as a single site at the same time as the telomeres also formed a bouquet or single cluster.

  20. Centromere and telomere sequence alterations reflect the rapid genome evolution within the carnivorous plant genus Genlisea

    Czech Academy of Sciences Publication Activity Database

    Tran, T.D.; Cao, H.X.; Jovtchev, G.; Neumann, Pavel; Novák, Petr; Fojtová, M.; Vu, G.T.H.; Macas, Jiří; Fajkus, Jiří; Schubert, I.; Fuchs, J.

    2015-01-01

    Roč. 84, č. 6 (2015), s. 1087-1099 ISSN 0960-7412 R&D Projects: GA ČR GBP501/12/G090; GA ČR GA13-06943S Institutional support: RVO:60077344 ; RVO:68081707 Keywords : Centromeric tandem repeat * centromeric retrotransposons * Genlisea nigrocaulis, hispidula Subject RIV: EB - Genetics ; Molecular Biology; BO - Biophysics (BFU-R) Impact factor: 5.468, year: 2015

  1. Genomics and the challenging translation into conservation practice

    Science.gov (United States)

    Aaron B. A. Shafer; Jochen B. W. Wolf; Paulo C. Alves; Linnea Bergstrom; Michael W. Bruford; Ioana Brannstrom; Guy Colling; Love Dalen; Luc De Meester; Robert Ekblom; Katie D. Fawcett; Simone Fior; Mehrdad Hajibabaei; Jason A. Hill; A. Rus Hoezel; Jacob Hoglund; Evelyn L. Jensen; Johannes Krause; Torsten N. Kristensen; Michael Krutzen; John K. McKay; Anita J. Norman; Rob Ogden; E. Martin Osterling; N. Joop Ouborg; John Piccolo; Danijela Popovic; Craig R. Primmer; Floyd A. Reed; Marie Roumet; Jordi Salmona; Tamara Schenekar; Michael K. Schwartz; Gernot Segelbacher; Helen Senn; Jens Thaulow; Mia Valtonen; Andrew Veale; Philippine Vergeer; Nagarjun Vijay; Carles Vila; Matthias Weissensteiner; Lovisa Wennerstrom; Christopher W. Wheat; Piotr Zielinski

    2015-01-01

    The global loss of biodiversity continues at an alarming rate. Genomic approaches have been suggested as a promising tool for conservation practice as scaling up to genome-wide data can improve traditional conservation genetic inferences and provide qualitatively novel insights. However, the generation of genomic data and subsequent analyses and interpretations remain...

  2. DNA topoisomerase III localizes to centromeres and affects centromeric CENP-A levels in fission yeast.

    Directory of Open Access Journals (Sweden)

    Ulrika Norman-Axelsson

    Full Text Available Centromeres are specialized chromatin regions marked by the presence of nucleosomes containing the centromere-specific histone H3 variant CENP-A, which is essential for chromosome segregation. Assembly and disassembly of nucleosomes is intimately linked to DNA topology, and DNA topoisomerases have previously been implicated in the dynamics of canonical H3 nucleosomes. Here we show that Schizosaccharomyces pombe Top3 and its partner Rqh1 are involved in controlling the levels of CENP-A(Cnp1 at centromeres. Both top3 and rqh1 mutants display defects in chromosome segregation. Using chromatin immunoprecipitation and tiling microarrays, we show that Top3, unlike Top1 and Top2, is highly enriched at centromeric central domains, demonstrating that Top3 is the major topoisomerase in this region. Moreover, centromeric Top3 occupancy positively correlates with CENP-A(Cnp1 occupancy. Intriguingly, both top3 and rqh1 mutants display increased relative enrichment of CENP-A(Cnp1 at centromeric central domains. Thus, Top3 and Rqh1 normally limit the levels of CENP-A(Cnp1 in this region. This new role is independent of the established function of Top3 and Rqh1 in homologous recombination downstream of Rad51. Therefore, we hypothesize that the Top3-Rqh1 complex has an important role in controlling centromere DNA topology, which in turn affects the dynamics of CENP-A(Cnp1 nucleosomes.

  3. Comparative analyses reveal high levels of conserved colinearity between the finger millet and rice genomes.

    Science.gov (United States)

    Srinivasachary; Dida, Mathews M; Gale, Mike D; Devos, Katrien M

    2007-08-01

    Finger millet is an allotetraploid (2n = 4x = 36) grass that belongs to the Chloridoideae subfamily. A comparative analysis has been carried out to determine the relationship of the finger millet genome with that of rice. Six of the nine finger millet homoeologous groups corresponded to a single rice chromosome each. Each of the remaining three finger millet groups were orthologous to two rice chromosomes, and in all the three cases one rice chromosome was inserted into the centromeric region of a second rice chromosome to give the finger millet chromosomal configuration. All observed rearrangements were, among the grasses, unique to finger millet and, possibly, the Chloridoideae subfamily. Gene orders between rice and finger millet were highly conserved, with rearrangements being limited largely to single marker transpositions and small putative inversions encompassing at most three markers. Only some 10% of markers mapped to non-syntenic positions in rice and finger millet and the majority of these were located in the distal 14% of chromosome arms, supporting a possible correlation between recombination and sequence evolution as has previously been observed in wheat. A comparison of the organization of finger millet, Panicoideae and Pooideae genomes relative to rice allowed us to infer putative ancestral chromosome configurations in the grasses.

  4. Centromere replication timing determines different forms of genomic instability in Saccharomyces cerevisiae checkpoint mutants during replication stress.

    Science.gov (United States)

    Feng, Wenyi; Bachant, Jeff; Collingwood, David; Raghuraman, M K; Brewer, Bonita J

    2009-12-01

    Yeast replication checkpoint mutants lose viability following transient exposure to hydroxyurea, a replication-impeding drug. In an effort to understand the basis for this lethality, we discovered that different events are responsible for inviability in checkpoint-deficient cells harboring mutations in the mec1 and rad53 genes. By monitoring genomewide replication dynamics of cells exposed to hydroxyurea, we show that cells with a checkpoint deficient allele of RAD53, rad53K227A, fail to duplicate centromeres. Following removal of the drug, however, rad53K227A cells recover substantial DNA replication, including replication through centromeres. Despite this recovery, the rad53K227A mutant fails to achieve biorientation of sister centromeres during recovery from hydroxyurea, leading to secondary activation of the spindle assembly checkpoint (SAC), aneuploidy, and lethal chromosome segregation errors. We demonstrate that cell lethality from this segregation defect could be partially remedied by reinforcing bipolar attachment. In contrast, cells with the mec1-1 sml1-1 mutations suffer from severely impaired replication resumption upon removal of hydroxyurea. mec1-1 sml1-1 cells can, however, duplicate at least some of their centromeres and achieve bipolar attachment, leading to abortive segregation and fragmentation of incompletely replicated chromosomes. Our results highlight the importance of replicating yeast centromeres early and reveal different mechanisms of cell death due to differences in replication fork progression.

  5. Molecular structures of centromeric heterochromatin and karyotypic evolution in the Siamese crocodile (Crocodylus siamensis) (Crocodylidae, Crocodylia).

    Science.gov (United States)

    Kawagoshi, Taiki; Nishida, Chizuko; Ota, Hidetoshi; Kumazawa, Yoshinori; Endo, Hideki; Matsuda, Yoichi

    2008-01-01

    Crocodilians have several unique karyotypic features, such as small diploid chromosome numbers (30-42) and the absence of dot-shaped microchromosomes. Of the extant crocodilian species, the Siamese crocodile (Crocodylus siamensis) has no more than 2n = 30, comprising mostly bi-armed chromosomes with large centromeric heterochromatin blocks. To investigate the molecular structures of C-heterochromatin and genomic compartmentalization in the karyotype, characterized by the disappearance of tiny microchromosomes and reduced chromosome number, we performed molecular cloning of centromeric repetitive sequences and chromosome mapping of the 18S-28S rDNA and telomeric (TTAGGG)( n ) sequences. The centromeric heterochromatin was composed mainly of two repetitive sequence families whose characteristics were quite different. Two types of GC-rich CSI-HindIII family sequences, the 305 bp CSI-HindIII-S (G+C content, 61.3%) and 424 bp CSI-HindIII-M (63.1%), were localized to the intensely PI-stained centric regions of all chromosomes, except for chromosome 2 with PI-negative heterochromatin. The 94 bp CSI-DraI (G+C content, 48.9%) was tandem-arrayed satellite DNA and localized to chromosome 2 and four pairs of small-sized chromosomes. The chromosomal size-dependent genomic compartmentalization that is supposedly unique to the Archosauromorpha was probably lost in the crocodilian lineage with the disappearance of microchromosomes followed by the homogenization of centromeric repetitive sequences between chromosomes, except for chromosome 2.

  6. Phylogeny of horse chromosome 5q in the genus Equus and centromere repositioning.

    Science.gov (United States)

    Piras, F M; Nergadze, S G; Poletto, V; Cerutti, F; Ryder, O A; Leeb, T; Raimondi, E; Giulotto, E

    2009-01-01

    Horses, asses and zebras belong to the genus Equus and are the only extant species of the family Equidae in the order Perissodactyla. In a previous work we demonstrated that a key factor in the rapid karyotypic evolution of this genus was evolutionary centromere repositioning, that is, the shift of the centromeric function to a new position without alteration of the order of markers along the chromosome. In search of previously undiscovered evolutionarily new centromeres, we traced the phylogeny of horse chromosome 5, analyzing the order of BAC markers, derived from a horse genomic library, in 7 Equus species (E. caballus, E. hemionus onager, E. kiang, E. asinus, E. grevyi, E. burchelli and E. zebra hartmannae). This analysis showed that repositioned centromeres are present in E. asinus (domestic donkey, EAS) chromosome 16 and in E. burchelli (Burchell's zebra, EBU) chromosome 17, confirming that centromere repositioning is a strikingly frequent phenomenon in this genus. The observation that the neocentromeres in EAS16 and EBU17 are in the same chromosomal position suggests that they may derive from the same event and therefore, E. asinus and E. burchelli may be more closely related than previously proposed; alternatively, 2 centromere repositioning events, involving the same chromosomal region, may have occurred independently in different lineages, pointing to the possible existence of hot spots for neocentromere formation. Our comparative analysis also showed that, while E. caballus chromosome 5 seems to represent the ancestral configuration, centric fission followed by independent fusion events gave rise to 3 different submetacentric chromosomes in other Equus lineages. (c) 2009 S. Karger AG, Basel.

  7. Epigenetic Regulation of Centromere Chromatin Stability by Dietary and Environmental Factors.

    Science.gov (United States)

    Hernández-Saavedra, Diego; Strakovsky, Rita S; Ostrosky-Wegman, Patricia; Pan, Yuan-Xiang

    2017-11-01

    The centromere is a genomic locus required for the segregation of the chromosomes during cell division. This chromosomal region together with pericentromeres has been found to be susceptible to damage, and thus the perturbation of the centromere could lead to the development of aneuploidic events. Metabolic abnormalities that underlie the generation of cancer include inflammation, oxidative stress, cell cycle deregulation, and numerous others. The micronucleus assay, an early clinical marker of cancer, has been shown to provide a reliable measure of genotoxic damage that may signal cancer initiation. In the current review, we will discuss the events that lead to micronucleus formation and centromeric and pericentromeric chromatin instability, as well transcripts emanating from these regions, which were previously thought to be inactive. Studies were selected in PubMed if they reported the effects of nutritional status (macro- and micronutrients) or environmental toxicant exposure on micronucleus frequency or any other chromosomal abnormality in humans, animals, or cell models. Mounting evidence from epidemiologic, environmental, and nutritional studies provides a novel perspective on the origination of aneuploidic events. Although substantial evidence exists describing the role that nutritional status and environmental toxicants have on the generation of micronuclei and other nuclear aberrations, limited information is available to describe the importance of macro- and micronutrients on centromeric and pericentromeric chromatin stability. Moving forward, studies that specifically address the direct link between nutritional status, excess, or deficiency and the epigenetic regulation of the centromere will provide much needed insight into the nutritional and environmental regulation of this chromosomal region and the initiation of aneuploidy. © 2017 American Society for Nutrition.

  8. Premitotic assembly of human CENPs -T and -W switches centromeric chromatin to a mitotic state.

    Directory of Open Access Journals (Sweden)

    Lisa Prendergast

    2011-06-01

    Full Text Available Centromeres are differentiated chromatin domains, present once per chromosome, that direct segregation of the genome in mitosis and meiosis by specifying assembly of the kinetochore. They are distinct genetic loci in that their identity in most organisms is determined not by the DNA sequences they are associated with, but through specific chromatin composition and context. The core nucleosomal protein CENP-A/cenH3 plays a primary role in centromere determination in all species and directs assembly of a large complex of associated proteins in vertebrates. While CENP-A itself is stably transmitted from one generation to the next, the nature of the template for centromere replication and its relationship to kinetochore function are as yet poorly understood. Here, we investigate the assembly and inheritance of a histone fold complex of the centromere, the CENP-T/W complex, which is integrated with centromeric chromatin in association with canonical histone H3 nucleosomes. We have investigated the cell cycle regulation, timing of assembly, generational persistence, and requirement for function of CENPs -T and -W in the cell cycle in human cells. The CENP-T/W complex assembles through a dynamic exchange mechanism in late S-phase and G2, is required for mitosis in each cell cycle and does not persist across cell generations, properties reciprocal to those measured for CENP-A. We propose that the CENP-A and H3-CENP-T/W nucleosome components of the centromere are specialized for centromeric and kinetochore activities, respectively. Segregation of the assembly mechanisms for the two allows the cell to switch between chromatin configurations that reciprocally support the replication of the centromere and its conversion to a mitotic state on postreplicative chromatin.

  9. Positioning of chromosomes in human spermatozoa is determined by ordered centromere arrangement.

    Directory of Open Access Journals (Sweden)

    Olga S Mudrak

    Full Text Available The intranuclear positioning of chromosomes (CHRs is a well-documented fact; however, mechanisms directing such ordering remain unclear. Unlike somatic cells, human spermatozoa contain distinct spatial markers and have asymmetric nuclei which make them a unique model for localizing CHR territories and matching peri-centromere domains. In this study, we established statistically preferential longitudinal and lateral positioning for eight CHRs. Both parameters demonstrated a correlation with the CHR gene densities but not with their sizes. Intranuclear non-random positioning of the CHRs was found to be driven by a specific linear order of centromeres physically interconnected in continuous arrays. In diploid spermatozoa, linear order of peri-centromeres was identical in two genome sets and essentially matched the arrangement established for haploid cells. We propose that the non-random longitudinal order of CHRs in human spermatozoa is generated during meiotic stages of spermatogenesis. The specific arrangement of sperm CHRs may serve as an epigenetic basis for differential transcription/replication and direct spatial CHR organization during early embryogenesis.

  10. Maintaining replication origins in the face of genomic change.

    Science.gov (United States)

    Di Rienzi, Sara C; Lindstrom, Kimberly C; Mann, Tobias; Noble, William S; Raghuraman, M K; Brewer, Bonita J

    2012-10-01

    Origins of replication present a paradox to evolutionary biologists. As a collection, they are absolutely essential genomic features, but individually are highly redundant and nonessential. It is therefore difficult to predict to what extent and in what regard origins are conserved over evolutionary time. Here, through a comparative genomic analysis of replication origins and chromosomal replication patterns in the budding yeasts Saccharomyces cerevisiae and Lachancea waltii, we assess to what extent replication origins survived genomic change produced from 150 million years of evolution. We find that L. waltii origins exhibit a core consensus sequence and nucleosome occupancy pattern highly similar to those of S. cerevisiae origins. We further observe that the overall progression of chromosomal replication is similar between L. waltii and S. cerevisiae. Nevertheless, few origins show evidence of being conserved in location between the two species. Among the conserved origins are those surrounding centromeres and adjacent to histone genes, suggesting that proximity to an origin may be important for their regulation. We conclude that, over evolutionary time, origins maintain sequence, structure, and regulation, but are continually being created and destroyed, with the result that their locations are generally not conserved.

  11. Phylogenetic and structural analysis of centromeric DNA and kinetochore proteins

    OpenAIRE

    Meraldi, Patrick; McAinsh, Andrew D; Rheinbay, Esther; Sorger, Peter K

    2006-01-01

    Background: Kinetochores are large multi-protein structures that assemble on centromeric DNA (CEN DNA) and mediate the binding of chromosomes to microtubules. Comprising 125 base-pairs of CEN DNA and 70 or more protein components, Saccharomyces cerevisiae kinetochores are among the best understood. In contrast, most fungal, plant and animal cells assemble kinetochores on CENs that are longer and more complex, raising the question of whether kinetochore architecture has been conserved through ...

  12. Characterization of centromeric histone H3 (CENH3 variants in cultivated and wild carrots (Daucus sp..

    Directory of Open Access Journals (Sweden)

    Frank Dunemann

    Full Text Available In eukaryotes, centromeres are the assembly sites for the kinetochore, a multi-protein complex to which spindle microtubules are attached at mitosis and meiosis, thereby ensuring segregation of chromosomes during cell division. They are specified by incorporation of CENH3, a centromere specific histone H3 variant which replaces canonical histone H3 in the nucleosomes of functional centromeres. To lay a first foundation of a putative alternative haploidization strategy based on centromere-mediated genome elimination in cultivated carrots, in the presented research we aimed at the identification and cloning of functional CENH3 genes in Daucus carota and three distantly related wild species of genus Daucus varying in basic chromosome numbers. Based on mining the carrot transcriptome followed by a subsequent PCR-based cloning, homologous coding sequences for CENH3s of the four Daucus species were identified. The ORFs of the CENH3 variants were very similar, and an amino acid sequence length of 146 aa was found in three out of the four species. Comparison of Daucus CENH3 amino acid sequences with those of other plant CENH3s as well as their phylogenetic arrangement among other dicot CENH3s suggest that the identified genes are authentic CENH3 homologs. To verify the location of the CENH3 protein in the kinetochore regions of the Daucus chromosomes, a polyclonal antibody based on a peptide corresponding to the N-terminus of DcCENH3 was developed and used for anti-CENH3 immunostaining of mitotic root cells. The chromosomal location of CENH3 proteins in the centromere regions of the chromosomes could be confirmed. For genetic localization of the CENH3 gene in the carrot genome, a previously constructed linkage map for carrot was used for mapping a CENH3-specific simple sequence repeat (SSR marker, and the CENH3 locus was mapped on the carrot chromosome 9.

  13. Fluorescence In Situ Hybridization (FISH-Based Karyotyping Reveals Rapid Evolution of Centromeric and Subtelomeric Repeats in Common Bean (Phaseolus vulgaris and Relatives

    Directory of Open Access Journals (Sweden)

    Aiko Iwata-Otsubo

    2016-04-01

    Full Text Available Fluorescence in situ hybridization (FISH-based karyotyping is a powerful cytogenetics tool to study chromosome organization, behavior, and chromosome evolution. Here, we developed a FISH-based karyotyping system using a probe mixture comprised of centromeric and subtelomeric satellite repeats, 5S rDNA, and chromosome-specific BAC clones in common bean, which enables one to unambiguously distinguish all 11 chromosome pairs. Furthermore, we applied the karyotyping system to several wild relatives and landraces of common bean from two distinct gene pools, as well as other related Phaseolus species, to investigate repeat evolution in the genus Phaseolus. Comparison of karyotype maps within common bean indicates that chromosomal distribution of the centromeric and subtelomeric satellite repeats is stable, whereas the copy number of the repeats was variable, indicating rapid amplification/reduction of the repeats in specific genomic regions. In Phaseolus species that diverged approximately 2–4 million yr ago, copy numbers of centromeric repeats were largely reduced or diverged, and chromosomal distributions have changed, suggesting rapid evolution of centromeric repeats. We also detected variation in the distribution pattern of subtelomeric repeats in Phaseolus species. The FISH-based karyotyping system revealed that satellite repeats are actively and rapidly evolving, forming genomic features unique to individual common bean accessions and Phaseolus species.

  14. Conservation genetics and genomics of amphibians and reptiles.

    Science.gov (United States)

    Shaffer, H Bradley; Gidiş, Müge; McCartney-Melstad, Evan; Neal, Kevin M; Oyamaguchi, Hilton M; Tellez, Marisa; Toffelmier, Erin M

    2015-01-01

    Amphibians and reptiles as a group are often secretive, reach their greatest diversity often in remote tropical regions, and contain some of the most endangered groups of organisms on earth. Particularly in the past decade, genetics and genomics have been instrumental in the conservation biology of these cryptic vertebrates, enabling work ranging from the identification of populations subject to trade and exploitation, to the identification of cryptic lineages harboring critical genetic variation, to the analysis of genes controlling key life history traits. In this review, we highlight some of the most important ways that genetic analyses have brought new insights to the conservation of amphibians and reptiles. Although genomics has only recently emerged as part of this conservation tool kit, several large-scale data sources, including full genomes, expressed sequence tags, and transcriptomes, are providing new opportunities to identify key genes, quantify landscape effects, and manage captive breeding stocks of at-risk species.

  15. Uncoupling of satellite DNA and centromeric function in the genus Equus.

    Science.gov (United States)

    Piras, Francesca M; Nergadze, Solomon G; Magnani, Elisa; Bertoni, Livia; Attolini, Carmen; Khoriauli, Lela; Raimondi, Elena; Giulotto, Elena

    2010-02-12

    In a previous study, we showed that centromere repositioning, that is the shift along the chromosome of the centromeric function without DNA sequence rearrangement, has occurred frequently during the evolution of the genus Equus. In this work, the analysis of the chromosomal distribution of satellite tandem repeats in Equus caballus, E. asinus, E. grevyi, and E. burchelli highlighted two atypical features: 1) several centromeres, including the previously described evolutionary new centromeres (ENCs), seem to be devoid of satellite DNA, and 2) satellite repeats are often present at non-centromeric termini, probably corresponding to relics of ancestral now inactive centromeres. Immuno-FISH experiments using satellite DNA and antibodies against the kinetochore protein CENP-A demonstrated that satellite-less primary constrictions are actually endowed with centromeric function. The phylogenetic reconstruction of centromere repositioning events demonstrates that the acquisition of satellite DNA occurs after the formation of the centromere during evolution and that centromeres can function over millions of years and many generations without detectable satellite DNA. The rapidly evolving Equus species gave us the opportunity to identify different intermediate steps along the full maturation of ENCs.

  16. Uncoupling of satellite DNA and centromeric function in the genus Equus.

    Directory of Open Access Journals (Sweden)

    Francesca M Piras

    2010-02-01

    Full Text Available In a previous study, we showed that centromere repositioning, that is the shift along the chromosome of the centromeric function without DNA sequence rearrangement, has occurred frequently during the evolution of the genus Equus. In this work, the analysis of the chromosomal distribution of satellite tandem repeats in Equus caballus, E. asinus, E. grevyi, and E. burchelli highlighted two atypical features: 1 several centromeres, including the previously described evolutionary new centromeres (ENCs, seem to be devoid of satellite DNA, and 2 satellite repeats are often present at non-centromeric termini, probably corresponding to relics of ancestral now inactive centromeres. Immuno-FISH experiments using satellite DNA and antibodies against the kinetochore protein CENP-A demonstrated that satellite-less primary constrictions are actually endowed with centromeric function. The phylogenetic reconstruction of centromere repositioning events demonstrates that the acquisition of satellite DNA occurs after the formation of the centromere during evolution and that centromeres can function over millions of years and many generations without detectable satellite DNA. The rapidly evolving Equus species gave us the opportunity to identify different intermediate steps along the full maturation of ENCs.

  17. Genomic Imprinting Was Evolutionarily Conserved during Wheat Polyploidization.

    Science.gov (United States)

    Yang, Guanghui; Liu, Zhenshan; Gao, Lulu; Yu, Kuohai; Feng, Man; Yao, Yingyin; Peng, Huiru; Hu, Zhaorong; Sun, Qixin; Ni, Zhongfu; Xin, Mingming

    2018-01-01

    Genomic imprinting is an epigenetic phenomenon that causes genes to be differentially expressed depending on their parent of origin. To evaluate the evolutionary conservation of genomic imprinting and the effects of ploidy on this process, we investigated parent-of-origin-specific gene expression patterns in the endosperm of diploid ( Aegilops spp), tetraploid, and hexaploid wheat ( Triticum spp) at various stages of development via high-throughput transcriptome sequencing. We identified 91, 135, and 146 maternally or paternally expressed genes (MEGs or PEGs, respectively) in diploid, tetraploid, and hexaploid wheat, respectively, 52.7% of which exhibited dynamic expression patterns at different developmental stages. Gene Ontology enrichment analysis suggested that MEGs and PEGs were involved in metabolic processes and DNA-dependent transcription, respectively. Nearly half of the imprinted genes exhibited conserved expression patterns during wheat hexaploidization. In addition, 40% of the homoeolog pairs originating from whole-genome duplication were consistently maternally or paternally biased in the different subgenomes of hexaploid wheat. Furthermore, imprinted expression was found for 41.2% and 50.0% of homolog pairs that evolved by tandem duplication after genome duplication in tetraploid and hexaploid wheat, respectively. These results suggest that genomic imprinting was evolutionarily conserved between closely related Triticum and Aegilops species and in the face of polyploid hybridization between species in these genera. © 2018 American Society of Plant Biologists. All rights reserved.

  18. Perspectives of genomics for genetic conservation of livestock

    NARCIS (Netherlands)

    Windig, J.J.; Engelsma, K.A.

    2010-01-01

    Genomics provides new opportunities for conservation genetics. Conservation genetics in livestock is based on estimating diversity by pedigree relatedness and managing diversity by choosing those animals that maximize genetic diversity. Animals can be chosen as parents for the next generation, as

  19. Genome-wide identification of coding and non-coding conserved sequence tags in human and mouse genomes

    Directory of Open Access Journals (Sweden)

    Maggi Giorgio P

    2008-06-01

    Full Text Available Abstract Background The accurate detection of genes and the identification of functional regions is still an open issue in the annotation of genomic sequences. This problem affects new genomes but also those of very well studied organisms such as human and mouse where, despite the great efforts, the inventory of genes and regulatory regions is far from complete. Comparative genomics is an effective approach to address this problem. Unfortunately it is limited by the computational requirements needed to perform genome-wide comparisons and by the problem of discriminating between conserved coding and non-coding sequences. This discrimination is often based (thus dependent on the availability of annotated proteins. Results In this paper we present the results of a comprehensive comparison of human and mouse genomes performed with a new high throughput grid-based system which allows the rapid detection of conserved sequences and accurate assessment of their coding potential. By detecting clusters of coding conserved sequences the system is also suitable to accurately identify potential gene loci. Following this analysis we created a collection of human-mouse conserved sequence tags and carefully compared our results to reliable annotations in order to benchmark the reliability of our classifications. Strikingly we were able to detect several potential gene loci supported by EST sequences but not corresponding to as yet annotated genes. Conclusion Here we present a new system which allows comprehensive comparison of genomes to detect conserved coding and non-coding sequences and the identification of potential gene loci. Our system does not require the availability of any annotated sequence thus is suitable for the analysis of new or poorly annotated genomes.

  20. Distinct retroelement classes define evolutionary breakpoints demarcating sites of evolutionary novelty

    Science.gov (United States)

    Longo, Mark S; Carone, Dawn M; Green, Eric D; O'Neill, Michael J; O'Neill, Rachel J

    2009-01-01

    Background Large-scale genome rearrangements brought about by chromosome breaks underlie numerous inherited diseases, initiate or promote many cancers and are also associated with karyotype diversification during species evolution. Recent research has shown that these breakpoints are nonrandomly distributed throughout the mammalian genome and many, termed "evolutionary breakpoints" (EB), are specific genomic locations that are "reused" during karyotypic evolution. When the phylogenetic trajectory of orthologous chromosome segments is considered, many of these EB are coincident with ancient centromere activity as well as new centromere formation. While EB have been characterized as repeat-rich regions, it has not been determined whether specific sequences have been retained during evolution that would indicate previous centromere activity or a propensity for new centromere formation. Likewise, the conservation of specific sequence motifs or classes at EBs among divergent mammalian taxa has not been determined. Results To define conserved sequence features of EBs associated with centromere evolution, we performed comparative sequence analysis of more than 4.8 Mb within the tammar wallaby, Macropus eugenii, derived from centromeric regions (CEN), euchromatic regions (EU), and an evolutionary breakpoint (EB) that has undergone convergent breakpoint reuse and past centromere activity in marsupials. We found a dramatic enrichment for long interspersed nucleotide elements (LINE1s) and endogenous retroviruses (ERVs) and a depletion of short interspersed nucleotide elements (SINEs) shared between CEN and EBs. We analyzed the orthologous human EB (14q32.33), known to be associated with translocations in many cancers including multiple myelomas and plasma cell leukemias, and found a conserved distribution of similar repetitive elements. Conclusion Our data indicate that EBs tracked within the class Mammalia harbor sequence features retained since the divergence of marsupials

  1. Distinct retroelement classes define evolutionary breakpoints demarcating sites of evolutionary novelty

    Directory of Open Access Journals (Sweden)

    Green Eric D

    2009-07-01

    Full Text Available Abstract Background Large-scale genome rearrangements brought about by chromosome breaks underlie numerous inherited diseases, initiate or promote many cancers and are also associated with karyotype diversification during species evolution. Recent research has shown that these breakpoints are nonrandomly distributed throughout the mammalian genome and many, termed "evolutionary breakpoints" (EB, are specific genomic locations that are "reused" during karyotypic evolution. When the phylogenetic trajectory of orthologous chromosome segments is considered, many of these EB are coincident with ancient centromere activity as well as new centromere formation. While EB have been characterized as repeat-rich regions, it has not been determined whether specific sequences have been retained during evolution that would indicate previous centromere activity or a propensity for new centromere formation. Likewise, the conservation of specific sequence motifs or classes at EBs among divergent mammalian taxa has not been determined. Results To define conserved sequence features of EBs associated with centromere evolution, we performed comparative sequence analysis of more than 4.8 Mb within the tammar wallaby, Macropus eugenii, derived from centromeric regions (CEN, euchromatic regions (EU, and an evolutionary breakpoint (EB that has undergone convergent breakpoint reuse and past centromere activity in marsupials. We found a dramatic enrichment for long interspersed nucleotide elements (LINE1s and endogenous retroviruses (ERVs and a depletion of short interspersed nucleotide elements (SINEs shared between CEN and EBs. We analyzed the orthologous human EB (14q32.33, known to be associated with translocations in many cancers including multiple myelomas and plasma cell leukemias, and found a conserved distribution of similar repetitive elements. Conclusion Our data indicate that EBs tracked within the class Mammalia harbor sequence features retained since the

  2. Genetic mapping of centromeres in the nine Citrus clementina chromosomes using half-tetrad analysis and recombination patterns in unreduced and haploid gametes.

    Science.gov (United States)

    Aleza, Pablo; Cuenca, José; Hernández, María; Juárez, José; Navarro, Luis; Ollitrault, Patrick

    2015-03-08

    Mapping centromere locations in plant species provides essential information for the analysis of genetic structures and population dynamics. The centromere's position affects the distribution of crossovers along a chromosome and the parental heterozygosity restitution by 2n gametes is a direct function of the genetic distance to the centromere. Sexual polyploidisation is relatively frequent in Citrus species and is widely used to develop new seedless triploid cultivars. The study's objectives were to (i) map the positions of the centromeres of the nine Citrus clementina chromosomes; (ii) analyse the crossover interference in unreduced gametes; and (iii) establish the pattern of genetic recombination in haploid clementine gametes along each chromosome and its relationship with the centromere location and distribution of genic sequences. Triploid progenies were derived from unreduced megagametophytes produced by second-division restitution. Centromere positions were mapped genetically for all linkage groups using half-tetrad analysis. Inference of the physical locations of centromeres revealed one acrocentric, four metacentric and four submetacentric chromosomes. Crossover interference was observed in unreduced gametes, with variation seen between chromosome arms. For haploid gametes, a strong decrease in the recombination rate occurred in centromeric and pericentromeric regions, which contained a low density of genic sequences. In chromosomes VIII and IX, these low recombination rates extended beyond the pericentromeric regions. The genomic region corresponding to a genetic distance recombination pattern along each chromosome. However, regions with low recombination rates extended beyond the pericentromeric regions of some chromosomes into areas richer in genic sequences. The persistence of strong linkage disequilibrium between large numbers of genes promotes the stability of epistatic interactions and multilocus-controlled traits over successive generations but

  3. The major horse satellite DNA family is associated with centromere competence.

    Science.gov (United States)

    Cerutti, Federico; Gamba, Riccardo; Mazzagatti, Alice; Piras, Francesca M; Cappelletti, Eleonora; Belloni, Elisa; Nergadze, Solomon G; Raimondi, Elena; Giulotto, Elena

    2016-01-01

    The centromere is the specialized locus required for correct chromosome segregation during cell division. The DNA of most eukaryotic centromeres is composed of extended arrays of tandem repeats (satellite DNA). In the horse, we previously showed that, although the centromere of chromosome 11 is completely devoid of tandem repeat arrays, all other centromeres are characterized by the presence of satellite DNA. We isolated three horse satellite DNA sequences (37cen, 2P1 and EC137) and described their chromosomal localization in four species of the genus Equus. In the work presented here, using the ChIP-seq methodology, we showed that, in the horse, the 37cen satellite binds CENP-A, the centromere-specific histone-H3 variant. The 37cen sequence bound by CENP-A is GC-rich with 221 bp units organized in a head-to-tail fashion. The physical interaction of CENP-A with 37cen was confirmed through slot blot experiments. Immuno-FISH on stretched chromosomes and chromatin fibres demonstrated that the extension of satellite DNA stretches is variable and is not related to the organization of CENP-A binding domains. Finally, we proved that the centromeric satellite 37cen is transcriptionally active. Our data offer new insights into the organization of horse centromeres. Although three different satellite DNA families are cytogenetically located at centromeres, only the 37cen family is associated to the centromeric function. Moreover, similarly to other species, CENP-A binding domains are variable in size. The transcriptional competence of the 37cen satellite that we observed adds new evidence to the hypothesis that centromeric transcripts may be required for centromere function.

  4. Comprehensive cytological characterization of the Gossypium hirsutum genome based on the development of a set of chromosome cytological markers

    Institute of Scientific and Technical Information of China (English)

    Wenbo; Shan; Yanqin; Jiang; Jinlei; Han; Kai; Wang

    2016-01-01

    Cotton is the world’s most important natural fiber crop. It is also a model system for studying polyploidization, genomic organization, and genome-size variation. Integrating the cytological characterization of cotton with its genetic map will be essential for understanding its genome structure and evolution, as well as for performing further genetic-map based mapping and cloning. In this study, we isolated a complete set of bacterial artificial chromosome clones anchored to each of the 52 chromosome arms of the tetraploid cotton Gossypium hirsutum. Combining these with telomere and centromere markers, we constructed a standard karyotype for the G. hirsutum inbred line TM-1. We dissected the chromosome arm localizations of the 45 S and 5S r DNA and suggest a centromere repositioning event in the homoeologous chromosomes AT09 and DT09. By integrating a systematic karyotype analysis with the genetic linkage map, we observed different genome sizes and chromosomal structures between the subgenomes of the tetraploid cotton and those of its diploid ancestors. Using evidence of conserved coding sequences, we suggest that the different evolutionary paths of non-coding retrotransposons account for most of the variation in size between the subgenomes of tetraploid cotton and its diploid ancestors. These results provide insights into the cotton genome and will facilitate further genome studies in G. hirsutum.

  5. Comprehensive cytological characterization of the Gossypium hirsutum genome based on the development of a set of chromosome cytological markers

    Directory of Open Access Journals (Sweden)

    Wenbo Shan

    2016-08-01

    Full Text Available Cotton is the world's most important natural fiber crop. It is also a model system for studying polyploidization, genomic organization, and genome-size variation. Integrating the cytological characterization of cotton with its genetic map will be essential for understanding its genome structure and evolution, as well as for performing further genetic-map based mapping and cloning. In this study, we isolated a complete set of bacterial artificial chromosome clones anchored to each of the 52 chromosome arms of the tetraploid cotton Gossypium hirsutum. Combining these with telomere and centromere markers, we constructed a standard karyotype for the G. hirsutum inbred line TM-1. We dissected the chromosome arm localizations of the 45S and 5S rDNA and suggest a centromere repositioning event in the homoeologous chromosomes AT09 and DT09. By integrating a systematic karyotype analysis with the genetic linkage map, we observed different genome sizes and chromosomal structures between the subgenomes of the tetraploid cotton and those of its diploid ancestors. Using evidence of conserved coding sequences, we suggest that the different evolutionary paths of non-coding retrotransposons account for most of the variation in size between the subgenomes of tetraploid cotton and its diploid ancestors. These results provide insights into the cotton genome and will facilitate further genome studies in G. hirsutum.

  6. Accelerated Evolution of Conserved Noncoding Sequences in theHuman Genome

    Energy Technology Data Exchange (ETDEWEB)

    Prambhakar, Shyam; Noonan, James P.; Paabo, Svante; Rubin, EdwardM.

    2006-07-06

    Genomic comparisons between human and distant, non-primatemammals are commonly used to identify cis-regulatory elements based onconstrained sequence evolution. However, these methods fail to detect"cryptic" functional elements, which are too weakly conserved amongmammals to distinguish from nonfunctional DNA. To address this problem,we explored the potential of deep intra-primate sequence comparisons. Wesequenced the orthologs of 558 kb of human genomic sequence, coveringmultiple loci involved in cholesterol homeostasis, in 6 nonhumanprimates. Our analysis identified 6 noncoding DNA elements displayingsignificant conservation among primates, but undetectable in more distantcomparisons. In vitro and in vivo tests revealed that at least three ofthese 6 elements have regulatory function. Notably, the mouse orthologsof these three functional human sequences had regulatory activity despitetheir lack of significant sequence conservation, indicating that they arecryptic ancestral cis-regulatory elements. These regulatory elementscould still be detected in a smaller set of three primate speciesincluding human, rhesus and marmoset. Since the human and rhesus genomesequences are already available, and the marmoset genome is activelybeing sequenced, the primate-specific conservation analysis describedhere can be applied in the near future on a whole-genome scale, tocomplement the annotation provided by more distant speciescomparisons.

  7. Evolutionarily conserved elements in vertebrate, insect, worm, and yeast genomes

    DEFF Research Database (Denmark)

    Siepel, Adam; Bejerano, Gill; Pedersen, Jakob Skou

    2005-01-01

    We have conducted a comprehensive search for conserved elements in vertebrate genomes, using genome-wide multiple alignments of five vertebrate species (human, mouse, rat, chicken, and Fugu rubripes). Parallel searches have been performed with multiple alignments of four insect species (three...... species of Drosophila and Anopheles gambiae), two species of Caenorhabditis, and seven species of Saccharomyces. Conserved elements were identified with a computer program called phastCons, which is based on a two-state phylogenetic hidden Markov model (phylo-HMM). PhastCons works by fitting a phylo......-HMM to the data by maximum likelihood, subject to constraints designed to calibrate the model across species groups, and then predicting conserved elements based on this model. The predicted elements cover roughly 3%-8% of the human genome (depending on the details of the calibration procedure) and substantially...

  8. Constraints on genes shape long-term conservation of macro-synteny in metazoan genomes

    Directory of Open Access Journals (Sweden)

    Putnam Nicholas H

    2011-10-01

    Full Text Available Abstract Background Many metazoan genomes conserve chromosome-scale gene linkage relationships (“macro-synteny” from the common ancestor of multicellular animal life 1234, but the biological explanation for this conservation is still unknown. Double cut and join (DCJ is a simple, well-studied model of neutral genome evolution amenable to both simulation and mathematical analysis 5, but as we show here, it is not sufficent to explain long-term macro-synteny conservation. Results We examine a family of simple (one-parameter extensions of DCJ to identify models and choices of parameters consistent with the levels of macro- and micro-synteny conservation observed among animal genomes. Our software implements a flexible strategy for incorporating genomic context into the DCJ model to incorporate various types of genomic context (“DCJ-[C]”, and is available as open source software from http://github.com/putnamlab/dcj-c. Conclusions A simple model of genome evolution, in which DCJ moves are allowed only if they maintain chromosomal linkage among a set of constrained genes, can simultaneously account for the level of macro-synteny conservation and for correlated conservation among multiple pairs of species. Simulations under this model indicate that a constraint on approximately 7% of metazoan genes is sufficient to constrain genome rearrangement to an average rate of 25 inversions and 1.7 translocations per million years.

  9. Linkage disequilibrium of evolutionarily conserved regions in the human genome

    Directory of Open Access Journals (Sweden)

    Johnson Todd A

    2006-12-01

    Full Text Available Abstract Background The strong linkage disequilibrium (LD recently found in genic or exonic regions of the human genome demonstrated that LD can be increased by evolutionary mechanisms that select for functionally important loci. This suggests that LD might be stronger in regions conserved among species than in non-conserved regions, since regions exposed to natural selection tend to be conserved. To assess this hypothesis, we used genome-wide polymorphism data from the HapMap project and investigated LD within DNA sequences conserved between the human and mouse genomes. Results Unexpectedly, we observed that LD was significantly weaker in conserved regions than in non-conserved regions. To investigate why, we examined sequence features that may distort the relationship between LD and conserved regions. We found that interspersed repeats, and not other sequence features, were associated with the weak LD tendency in conserved regions. To appropriately understand the relationship between LD and conserved regions, we removed the effect of repetitive elements and found that the high degree of sequence conservation was strongly associated with strong LD in coding regions but not with that in non-coding regions. Conclusion Our work demonstrates that the degree of sequence conservation does not simply increase LD as predicted by the hypothesis. Rather, it implies that purifying selection changes the polymorphic patterns of coding sequences but has little influence on the patterns of functional units such as regulatory elements present in non-coding regions, since the former are generally restricted by the constraint of maintaining a functional protein product across multiple exons while the latter may exist more as individually isolated units.

  10. HACking the centromere chromatin code: insights from human artificial chromosomes.

    Science.gov (United States)

    Bergmann, Jan H; Martins, Nuno M C; Larionov, Vladimir; Masumoto, Hiroshi; Earnshaw, William C

    2012-07-01

    The centromere is a specialized chromosomal region that serves as the assembly site of the kinetochore. At the centromere, CENP-A nucleosomes form part of a chromatin landscape termed centrochromatin. This chromatin environment conveys epigenetic marks regulating kinetochore formation. Recent work sheds light on the intricate relationship between centrochromatin state, the CENP-A assembly pathway and the maintenance of centromere function. Here, we review the emerging picture of how chromatin affects mammalian kinetochore formation. We place particular emphasis on data obtained from Human Artificial Chromosome (HAC) biology and the targeted engineering of centrochromatin using synthetic HACs. We discuss implications of these findings, which indicate that a delicate balance of histone modifications and chromatin state dictates both de novo centromere formation and the maintenance of centromere identity in dividing cell populations.

  11. Genome-Wide Identification of Polycomb Target Genes Reveals a Functional Association of Pho with Scm in Bombyx mori

    OpenAIRE

    Li, Zhiqing; Cheng, Daojun; Mon, Hiroaki; Tatsuke, Tsuneyuki; Zhu, Li; Xu, Jian; Lee, Jae Man; Xia, Qingyou; Kusakabe, Takahiro

    2012-01-01

    Polycomb group (PcG) proteins are evolutionarily conserved chromatin modifiers and act together in three multimeric complexes, Polycomb repressive complex 1 (PRC1), Polycomb repressive complex 2 (PRC2), and Pleiohomeotic repressive complex (PhoRC), to repress transcription of the target genes. Here, we identified Polycomb target genes in Bombyx mori with holocentric centromere using genome-wide expression screening based on the knockdown of BmSCE, BmESC, BmPHO, or BmSCM gene, which represent ...

  12. Meiotic Studies on Combinations of Chromosomes With Different Sized Centromeres in Maize

    Directory of Open Access Journals (Sweden)

    Fangpu Han

    2018-06-01

    Full Text Available Multiple centromere misdivision derivatives of a translocation between the supernumerary B chromosome and the short arm of chromosome 9 (TB-9Sb permit investigation of how centromeres of different sizes behave in meiosis in opposition or in competition with each other. In the first analysis, heterozygotes were produced between the normal TB-9Sb and derivatives of it that resulted from centromere misdivision that reduced the amounts of centromeric DNA. These heterozygotes could test whether these drastic differences would result in meiotic drive of the larger chromosome in female meiosis. Cytological determinations of the segregation of large and small centromeres among thousands of progeny of four combinations were made. The recovery of the larger centromere was at a few percent higher frequency in two of four combinations. However, examination of phosphorylated histone H2A-Thr133, a characteristic of active centromeres, showed a lack of correlation with the size of the centromeric DNA, suggesting an expansion of the basal protein features of the kinetochore in two of the three cases despite the reduction in the size of the underlying DNA. In the second analysis, plants containing different sizes of the B chromosome centromere were crossed to plants with TB-9Sb with a foldback duplication of 9S (TB-9Sb-Dp9. In the progeny, plants containing large and small versions of the B chromosome centromere were selected by FISH. A meiotic “tug of war” occurred in hybrid combinations by recombination between the normal 9S and the foldback duplication in those cases in which pairing occurred. Such pairing and recombination produce anaphase I bridges but in some cases the large and small centromeres progressed to the same pole. In one combination, new dicentric chromosomes were found in the progeny. Collectively, the results indicate that the size of the underlying DNA of a centromere does not dramatically affect its segregation properties or its ability

  13. Holocentromeres in Rhynchospora are associated with genome-wide centromere-specific repeat arrays interspersed among euchromatin

    Czech Academy of Sciences Publication Activity Database

    Marques, A.; Ribeiro, T.; Neumann, Pavel; Macas, Jiří; Novák, Petr; Schubert, V.; Pellino, M.; Fuchs, J.; Ma, W.; Kuhlmann, M.; Brandt, R.; Vanzela, A.L.L.; Beseda, Tomáš; Šimková, Hana; Pedrosa-Harand, A.; Houben, A.

    2015-01-01

    Roč. 112, č. 44 (2015), s. 13633-13638 ISSN 0027-8424 R&D Projects: GA ČR GBP501/12/G090; GA MŠk(CZ) LO1204 Institutional support: RVO:60077344 ; RVO:61389030 Keywords : Centromere * satellite DNA * holokinetic * chromosome Subject RIV: EB - Genetics ; Molecular Biology; EB - Genetics ; Molecular Biology (UEB-Q) Impact factor: 9.423, year: 2015

  14. Restructuring of Holocentric Centromeres During Meiosis in the Plant Rhynchospora pubera.

    Science.gov (United States)

    Marques, André; Schubert, Veit; Houben, Andreas; Pedrosa-Harand, Andrea

    2016-10-01

    Centromeres are responsible for the correct segregation of chromosomes during mitosis and meiosis. Holocentric chromosomes, characterized by multiple centromere units along each chromatid, have particular adaptations to ensure regular disjunction during meiosis. Here we show by detecting CENH3, CENP-C, tubulin, and centromeric repeats that holocentromeres may be organized differently in mitosis and meiosis of Rhynchospora pubera Contrasting to the mitotic linear holocentromere organization, meiotic centromeres show several clusters of centromere units (cluster-holocentromeres) during meiosis I. They accumulate along the poleward surface of bivalents where spindle fibers perpendicularly attach. During meiosis II, the cluster-holocentromeres are mostly present in the midregion of each chromatid. A linear holocentromere organization is restored after meiosis during pollen mitosis. Thus, a not yet described case of a cluster-holocentromere organization, showing a clear centromere restructuration between mitosis and meiosis, was identified in a holocentric organism. Copyright © 2016 by the Genetics Society of America.

  15. Single-copy genes define a conserved order between rice and wheat for understanding differences caused by duplication, deletion, and transposition of genes.

    Science.gov (United States)

    Singh, Nagendra K; Dalal, Vivek; Batra, Kamlesh; Singh, Binay K; Chitra, G; Singh, Archana; Ghazi, Irfan A; Yadav, Mahavir; Pandit, Awadhesh; Dixit, Rekha; Singh, Pradeep K; Singh, Harvinder; Koundal, Kirpa R; Gaikwad, Kishor; Mohapatra, Trilochan; Sharma, Tilak R

    2007-01-01

    The high-quality rice genome sequence is serving as a reference for comparative genome analysis in crop plants, especially cereals. However, early comparisons with bread wheat showed complex patterns of conserved synteny (gene content) and colinearity (gene order). Here, we show the presence of ancient duplicated segments in the progenitor of wheat, which were first identified in the rice genome. We also show that single-copy (SC) rice genes, those representing unique matches with wheat expressed sequence tag (EST) unigene contigs in the whole rice genome, show more than twice the proportion of genes mapping to syntenic wheat chromosome as compared to the multicopy (MC) or duplicated rice genes. While 58.7% of the 1,244 mapped SC rice genes were located in single syntenic wheat chromosome groups, the remaining 41.3% were distributed randomly to the other six non-syntenic wheat groups. This could only be explained by a background dispersal of genes in the genome through transposition or other unknown mechanism. The breakdown of rice-wheat synteny due to such transpositions was much greater near the wheat centromeres. Furthermore, the SC rice genes revealed a conserved primordial gene order that gives clues to the origin of rice and wheat chromosomes from a common ancestor through polyploidy, aneuploidy, centromeric fusions, and translocations. Apart from the bin-mapped wheat EST contigs, we also compared 56,298 predicted rice genes with 39,813 wheat EST contigs assembled from 409,765 EST sequences and identified 7,241 SC rice gene homologs of wheat. Based on the conserved colinearity of 1,063 mapped SC rice genes across the bins of individual wheat chromosomes, we predicted the wheat bin location of 6,178 unmapped SC rice gene homologs and validated the location of 213 of these in the telomeric bins of 21 wheat chromosomes with 35.4% initial success. This opens up the possibility of directed mapping of a large number of conserved SC rice gene homologs in wheat

  16. Perturbation of Incenp function impedes anaphase chromatid movements and chromosomal passenger protein flux at centromeres.

    Science.gov (United States)

    Ahonen, Leena J; Kukkonen, Anu M; Pouwels, Jeroen; Bolton, Margaret A; Jingle, Christopher D; Stukenberg, P Todd; Kallio, Marko J

    2009-02-01

    Incenp is an essential mitotic protein that, together with Aurora B, Survivin, and Borealin, forms the core of the chromosomal passenger protein complex (CPC). The CPC regulates various mitotic processes and functions to maintain genomic stability. The proper subcellular localization of the CPC and its full catalytic activity require the presence of each core subunit in the complex. We have investigated the mitotic tasks of the CPC using a function blocking antibody against Incenp microinjected into cells at different mitotic phases. This method allowed temporal analysis of CPC functions without perturbation of complex assembly or activity prior to injection. We have also studied the dynamic properties of Incenp and Aurora B using fusion protein photobleaching. We found that in early mitotic cells, Incenp and Aurora B exhibit dynamic turnover at centromeres, which is prevented by the anti-Incenp antibody. In these cells, the loss of centromeric CPC turnover is accompanied by forced mitotic exit without the execution of cytokinesis. Introduction of anti-Incenp antibody into early anaphase cells causes abnormalities in sister chromatid separation through defects in anaphase spindle functions. In summary, our data uncovers new mitotic roles for the CPC in anaphase and proposes that CPC turnover at centromeres modulates spindle assembly checkpoint signaling.

  17. Rad51-Rad52 mediated maintenance of centromeric chromatin in Candida albicans.

    Directory of Open Access Journals (Sweden)

    Sreyoshi Mitra

    2014-04-01

    Full Text Available Specification of the centromere location in most eukaryotes is not solely dependent on the DNA sequence. However, the non-genetic determinants of centromere identity are not clearly defined. While multiple mechanisms, individually or in concert, may specify centromeres epigenetically, most studies in this area are focused on a universal factor, a centromere-specific histone H3 variant CENP-A, often considered as the epigenetic determinant of centromere identity. In spite of variable timing of its loading at centromeres across species, a replication coupled early S phase deposition of CENP-A is found in most yeast centromeres. Centromeres are the earliest replicating chromosomal regions in a pathogenic budding yeast Candida albicans. Using a 2-dimensional agarose gel electrophoresis assay, we identify replication origins (ORI7-LI and ORI7-RI proximal to an early replicating centromere (CEN7 in C. albicans. We show that the replication forks stall at CEN7 in a kinetochore dependent manner and fork stalling is reduced in the absence of the homologous recombination (HR proteins Rad51 and Rad52. Deletion of ORI7-RI causes a significant reduction in the stalled fork signal and an increased loss rate of the altered chromosome 7. The HR proteins, Rad51 and Rad52, have been shown to play a role in fork restart. Confocal microscopy shows declustered kinetochores in rad51 and rad52 mutants, which are evidence of kinetochore disintegrity. CENP-ACaCse4 levels at centromeres, as determined by chromatin immunoprecipitation (ChIP experiments, are reduced in absence of Rad51/Rad52 resulting in disruption of the kinetochore structure. Moreover, western blot analysis reveals that delocalized CENP-A molecules in HR mutants degrade in a similar fashion as in other kinetochore mutants described before. Finally, co-immunoprecipitation assays indicate that Rad51 and Rad52 physically interact with CENP-ACaCse4 in vivo. Thus, the HR proteins Rad51 and Rad52

  18. Conservation and losses of non-coding RNAs in avian genomes.

    Directory of Open Access Journals (Sweden)

    Paul P Gardner

    Full Text Available Here we present the results of a large-scale bioinformatics annotation of non-coding RNA loci in 48 avian genomes. Our approach uses probabilistic models of hand-curated families from the Rfam database to infer conserved RNA families within each avian genome. We supplement these annotations with predictions from the tRNA annotation tool, tRNAscan-SE and microRNAs from miRBase. We identify 34 lncRNA-associated loci that are conserved between birds and mammals and validate 12 of these in chicken. We report several intriguing cases where a reported mammalian lncRNA, but not its function, is conserved. We also demonstrate extensive conservation of classical ncRNAs (e.g., tRNAs and more recently discovered ncRNAs (e.g., snoRNAs and miRNAs in birds. Furthermore, we describe numerous "losses" of several RNA families, and attribute these to either genuine loss, divergence or missing data. In particular, we show that many of these losses are due to the challenges associated with assembling avian microchromosomes. These combined results illustrate the utility of applying homology-based methods for annotating novel vertebrate genomes.

  19. A cytogenetic study of hospital workers occupationally exposed to radionuclides in Serbia. Premature centromere division as novel biomarker of exposure?

    Energy Technology Data Exchange (ETDEWEB)

    Pajic, Jelena; Rakic, Boban [Serbian Institute of Occupational Health ' ' Dr Dragomir Karajovic' ' , Belgrade (Serbia). Biodosimetry Dept.; Jovicic, Dubravka [Univ. ' ' Singidunum' ' , Belgrade (Serbia). Genotoxicology Dept.; Milovanovic, Aleksandar [Serbian Institute of Occupational Health ' ' Dr Dragomir Karajovic' ' , Belgrade (Serbia). Biodosimetry Dept.; Belgrade Univ. (Serbia). Occupational Health Dept.

    2016-04-15

    The health risk of chronic exposure to radionuclides includes changes in the genome (e.g., chromosomal aberrations and micronuclei) that increase chromosomal instability. There are also other phenomena, which seem to appear more frequently in metaphases of exposed persons (such as premature centromere division). The aim of this study was to discover whether or not there is correlation between incidence of named cytogenetic changes in persons occupationally exposed to radionuclides in comparison with unexposed control group, and if significant correlation is determined, can premature centromere division be consider as a biomarker of radiation exposure? The exposed group comprised 50 individuals occupationally exposed to radionuclides. The reference control group consisted of 40 unexposed individuals. Chromosomal aberrations, micronuclei and premature centromere division were analyzed according to a standard International Atomic Energy Agency protocol. Statistical analyses were performed using SPSS 17.0 statistics.The means for analyzed cytogenetic changes were significantly higher in the exposed group. Positive correlation between them was found in exposed group. Premature centromere division parameter PCD5-10 was selected as particularly suitable for separating groups (exposed/unexposed). Identification of other phenomena related to radionuclide exposure, beside well known, may clarify recent problems in radiobiology concerning the biological response to low doses of ionizing radiation and its consequences.

  20. A cytogenetic study of hospital workers occupationally exposed to radionuclides in Serbia. Premature centromere division as novel biomarker of exposure?

    International Nuclear Information System (INIS)

    Pajic, Jelena; Rakic, Boban; Jovicic, Dubravka; Milovanovic, Aleksandar; Belgrade Univ.

    2016-01-01

    The health risk of chronic exposure to radionuclides includes changes in the genome (e.g., chromosomal aberrations and micronuclei) that increase chromosomal instability. There are also other phenomena, which seem to appear more frequently in metaphases of exposed persons (such as premature centromere division). The aim of this study was to discover whether or not there is correlation between incidence of named cytogenetic changes in persons occupationally exposed to radionuclides in comparison with unexposed control group, and if significant correlation is determined, can premature centromere division be consider as a biomarker of radiation exposure? The exposed group comprised 50 individuals occupationally exposed to radionuclides. The reference control group consisted of 40 unexposed individuals. Chromosomal aberrations, micronuclei and premature centromere division were analyzed according to a standard International Atomic Energy Agency protocol. Statistical analyses were performed using SPSS 17.0 statistics.The means for analyzed cytogenetic changes were significantly higher in the exposed group. Positive correlation between them was found in exposed group. Premature centromere division parameter PCD5-10 was selected as particularly suitable for separating groups (exposed/unexposed). Identification of other phenomena related to radionuclide exposure, beside well known, may clarify recent problems in radiobiology concerning the biological response to low doses of ionizing radiation and its consequences.

  1. Overview on the Role of Advance Genomics in Conservation Biology of Endangered Species.

    Science.gov (United States)

    Khan, Suliman; Nabi, Ghulam; Ullah, Muhammad Wajid; Yousaf, Muhammad; Manan, Sehrish; Siddique, Rabeea; Hou, Hongwei

    2016-01-01

    In the recent era, due to tremendous advancement in industrialization, pollution and other anthropogenic activities have created a serious scenario for biota survival. It has been reported that present biota is entering a "sixth" mass extinction, because of chronic exposure to anthropogenic activities. Various ex situ and in situ measures have been adopted for conservation of threatened and endangered plants and animal species; however, these have been limited due to various discrepancies associated with them. Current advancement in molecular technologies, especially, genomics, is playing a very crucial role in biodiversity conservation. Advance genomics helps in identifying the segments of genome responsible for adaptation. It can also improve our understanding about microevolution through a better understanding of selection, mutation, assertive matting, and recombination. Advance genomics helps in identifying genes that are essential for fitness and ultimately for developing modern and fast monitoring tools for endangered biodiversity. This review article focuses on the applications of advanced genomics mainly demographic, adaptive genetic variations, inbreeding, hybridization and introgression, and disease susceptibilities, in the conservation of threatened biota. In short, it provides the fundamentals for novice readers and advancement in genomics for the experts working for the conservation of endangered plant and animal species.

  2. Centromeres Off the Hook: Massive Changes in Centromere Size and Structure Following Duplication of CenH3 Gene in Fabeae Species

    Czech Academy of Sciences Publication Activity Database

    Neumann, Pavel; Pavlíková, Zuzana; Koblížková, Andrea; Vrbová, Iva; Jedličková, Veronika; Novák, Petr; Macas, Jiří

    2015-01-01

    Roč. 32, č. 7 (2015), s. 1862-1879 ISSN 0737-4038 R&D Projects: GA ČR(CZ) GAP501/11/1843; GA MŠk(CZ) LH11058 Institutional support: RVO:60077344 Keywords : Centromere * CenH3 * centromere drive * chromosome Subject RIV: EB - Genetics ; Molecular Biology Impact factor: 13.649, year: 2015

  3. Genome-wide conserved consensus transcription factor binding motifs are hyper-methylated

    Directory of Open Access Journals (Sweden)

    Down Thomas A

    2010-09-01

    Full Text Available Abstract Background DNA methylation can regulate gene expression by modulating the interaction between DNA and proteins or protein complexes. Conserved consensus motifs exist across the human genome ("predicted transcription factor binding sites": "predicted TFBS" but the large majority of these are proven by chromatin immunoprecipitation and high throughput sequencing (ChIP-seq not to be biological transcription factor binding sites ("empirical TFBS". We hypothesize that DNA methylation at conserved consensus motifs prevents promiscuous or disorderly transcription factor binding. Results Using genome-wide methylation maps of the human heart and sperm, we found that all conserved consensus motifs as well as the subset of those that reside outside CpG islands have an aggregate profile of hyper-methylation. In contrast, empirical TFBS with conserved consensus motifs have a profile of hypo-methylation. 40% of empirical TFBS with conserved consensus motifs resided in CpG islands whereas only 7% of all conserved consensus motifs were in CpG islands. Finally we further identified a minority subset of TF whose profiles are either hypo-methylated or neutral at their respective conserved consensus motifs implicating that these TF may be responsible for establishing or maintaining an un-methylated DNA state, or whose binding is not regulated by DNA methylation. Conclusions Our analysis supports the hypothesis that at least for a subset of TF, empirical binding to conserved consensus motifs genome-wide may be controlled by DNA methylation.

  4. Comparative genomic in situ hybridization analysis on the ...

    African Journals Online (AJOL)

    The nucleolar organizing regions (NORs), a few telomeres, most centromeric regions and numerous interstitial sites were detected. The signals in small genomes were relatively sparse and unevenly distributed along chromosomes, whereas those in large genomes were dense and basically evenly distributed.

  5. Widespread Positive Selection Drives Differentiation of Centromeric Proteins in the Drosophila melanogaster subgroup.

    Science.gov (United States)

    Beck, Emily A; Llopart, Ana

    2015-11-25

    Rapid evolution of centromeric satellite repeats is thought to cause compensatory amino acid evolution in interacting centromere-associated kinetochore proteins. Cid, a protein that mediates kinetochore/centromere interactions, displays particularly high amino acid turnover. Rapid evolution of both Cid and centromeric satellite repeats led us to hypothesize that the apparent compensatory evolution may extend to interacting partners in the Condensin I complex (i.e., SMC2, SMC4, Cap-H, Cap-D2, and Cap-G) and HP1s. Missense mutations in these proteins often result in improper centromere formation and aberrant chromosome segregation, thus selection for maintained function and coevolution among proteins of the complex is likely strong. Here, we report evidence of rapid evolution and recurrent positive selection in seven centromere-associated proteins in species of the Drosophila melanogaster subgroup, and further postulate that positive selection on these proteins could be a result of centromere drive and compensatory changes, with kinetochore proteins competing for optimal spindle attachment.

  6. The rapidly evolving centromere-specific histone has stringent functional requirements in Arabidopsis thaliana.

    Science.gov (United States)

    Ravi, Maruthachalam; Kwong, Pak N; Menorca, Ron M G; Valencia, Joel T; Ramahi, Joseph S; Stewart, Jodi L; Tran, Robert K; Sundaresan, Venkatesan; Comai, Luca; Chan, Simon W-L

    2010-10-01

    Centromeres control chromosome inheritance in eukaryotes, yet their DNA structure and primary sequence are hypervariable. Most animals and plants have megabases of tandem repeats at their centromeres, unlike yeast with unique centromere sequences. Centromere function requires the centromere-specific histone CENH3 (CENP-A in human), which replaces histone H3 in centromeric nucleosomes. CENH3 evolves rapidly, particularly in its N-terminal tail domain. A portion of the CENH3 histone-fold domain, the CENP-A targeting domain (CATD), has been previously shown to confer kinetochore localization and centromere function when swapped into human H3. Furthermore, CENP-A in human cells can be functionally replaced by CENH3 from distantly related organisms including Saccharomyces cerevisiae. We have used cenh3-1 (a null mutant in Arabidopsis thaliana) to replace endogenous CENH3 with GFP-tagged variants. A H3.3 tail domain-CENH3 histone-fold domain chimera rescued viability of cenh3-1, but CENH3's lacking a tail domain were nonfunctional. In contrast to human results, H3 containing the A. thaliana CATD cannot complement cenh3-1. GFP-CENH3 from the sister species A. arenosa functionally replaces A. thaliana CENH3. GFP-CENH3 from the close relative Brassica rapa was targeted to centromeres, but did not complement cenh3-1, indicating that kinetochore localization and centromere function can be uncoupled. We conclude that CENH3 function in A. thaliana, an organism with large tandem repeat centromeres, has stringent requirements for functional complementation in mitosis.

  7. Overview on the Role of Advance Genomics in Conservation Biology of Endangered Species

    Directory of Open Access Journals (Sweden)

    Suliman Khan

    2016-01-01

    Full Text Available In the recent era, due to tremendous advancement in industrialization, pollution and other anthropogenic activities have created a serious scenario for biota survival. It has been reported that present biota is entering a “sixth” mass extinction, because of chronic exposure to anthropogenic activities. Various ex situ and in situ measures have been adopted for conservation of threatened and endangered plants and animal species; however, these have been limited due to various discrepancies associated with them. Current advancement in molecular technologies, especially, genomics, is playing a very crucial role in biodiversity conservation. Advance genomics helps in identifying the segments of genome responsible for adaptation. It can also improve our understanding about microevolution through a better understanding of selection, mutation, assertive matting, and recombination. Advance genomics helps in identifying genes that are essential for fitness and ultimately for developing modern and fast monitoring tools for endangered biodiversity. This review article focuses on the applications of advanced genomics mainly demographic, adaptive genetic variations, inbreeding, hybridization and introgression, and disease susceptibilities, in the conservation of threatened biota. In short, it provides the fundamentals for novice readers and advancement in genomics for the experts working for the conservation of endangered plant and animal species.

  8. Conservation and divergence of ADAM family proteins in the Xenopus genome

    Directory of Open Access Journals (Sweden)

    Shah Anoop

    2010-07-01

    Full Text Available Abstract Background Members of the disintegrin metalloproteinase (ADAM family play important roles in cellular and developmental processes through their functions as proteases and/or binding partners for other proteins. The amphibian Xenopus has long been used as a model for early vertebrate development, but genome-wide analyses for large gene families were not possible until the recent completion of the X. tropicalis genome sequence and the availability of large scale expression sequence tag (EST databases. In this study we carried out a systematic analysis of the X. tropicalis genome and uncovered several interesting features of ADAM genes in this species. Results Based on the X. tropicalis genome sequence and EST databases, we identified Xenopus orthologues of mammalian ADAMs and obtained full-length cDNA clones for these genes. The deduced protein sequences, synteny and exon-intron boundaries are conserved between most human and X. tropicalis orthologues. The alternative splicing patterns of certain Xenopus ADAM genes, such as adams 22 and 28, are similar to those of their mammalian orthologues. However, we were unable to identify an orthologue for ADAM7 or 8. The Xenopus orthologue of ADAM15, an active metalloproteinase in mammals, does not contain the conserved zinc-binding motif and is hence considered proteolytically inactive. We also found evidence for gain of ADAM genes in Xenopus as compared to other species. There is a homologue of ADAM10 in Xenopus that is missing in most mammals. Furthermore, a single scaffold of X. tropicalis genome contains four genes encoding ADAM28 homologues, suggesting genome duplication in this region. Conclusions Our genome-wide analysis of ADAM genes in X. tropicalis revealed both conservation and evolutionary divergence of these genes in this amphibian species. On the one hand, all ADAMs implicated in normal development and health in other species are conserved in X. tropicalis. On the other hand, some

  9. Prospects and Challenges for the Conservation of Farm Animal Genomic Resources, 2015-2025

    Directory of Open Access Journals (Sweden)

    Michael William Bruford

    2015-10-01

    Full Text Available Livestock conservation practice is changing rapidly in light of policy, climate change and market demands. The last decade saw a step change in technological and analytical approaches to define, manage and conserve Farm Animal Genomic Resources (FAnGR. These changes pose challenges for FAnGR conservation in terms of technological continuity, analytical capacity and the methodologies needed to exploit new, multidimensional data. The ESF Genomic Resources program final conference addressed these problems attempting to contribute to the development of the research and policy agenda for the next decade. We broadly identified four areas related to methodological and analytical challenges, data management and conservation. The overall conclusion is that there is a need for the use of current state-of-the-art tools to characterise the state of genomic resources in non-commercial and local breeds. The livestock genomic sector, which has been relatively well-organised in applying such methodologies so far, needs to make a concerted effort in the coming decade to enable to the democratisation of the powerful tools that are now at its disposal, and to ensure that they are applied in the context of breed conservation as well as development.

  10. Prospects and challenges for the conservation of farm animal genomic resources, 2015-2025

    Science.gov (United States)

    Bruford, Michael W.; Ginja, Catarina; Hoffmann, Irene; Joost, Stéphane; Orozco-terWengel, Pablo; Alberto, Florian J.; Amaral, Andreia J.; Barbato, Mario; Biscarini, Filippo; Colli, Licia; Costa, Mafalda; Curik, Ino; Duruz, Solange; Ferenčaković, Maja; Fischer, Daniel; Fitak, Robert; Groeneveld, Linn F.; Hall, Stephen J. G.; Hanotte, Olivier; Hassan, Faiz-ul; Helsen, Philippe; Iacolina, Laura; Kantanen, Juha; Leempoel, Kevin; Lenstra, Johannes A.; Ajmone-Marsan, Paolo; Masembe, Charles; Megens, Hendrik-Jan; Miele, Mara; Neuditschko, Markus; Nicolazzi, Ezequiel L.; Pompanon, François; Roosen, Jutta; Sevane, Natalia; Smetko, Anamarija; Štambuk, Anamaria; Streeter, Ian; Stucki, Sylvie; Supakorn, China; Telo Da Gama, Luis; Tixier-Boichard, Michèle; Wegmann, Daniel; Zhan, Xiangjiang

    2015-01-01

    Livestock conservation practice is changing rapidly in light of policy developments, climate change and diversifying market demands. The last decade has seen a step change in technology and analytical approaches available to define, manage and conserve Farm Animal Genomic Resources (FAnGR). However, these rapid changes pose challenges for FAnGR conservation in terms of technological continuity, analytical capacity and integrative methodologies needed to fully exploit new, multidimensional data. The final conference of the ESF Genomic Resources program aimed to address these interdisciplinary problems in an attempt to contribute to the agenda for research and policy development directions during the coming decade. By 2020, according to the Convention on Biodiversity's Aichi Target 13, signatories should ensure that “…the genetic diversity of …farmed and domesticated animals and of wild relatives …is maintained, and strategies have been developed and implemented for minimizing genetic erosion and safeguarding their genetic diversity.” However, the real extent of genetic erosion is very difficult to measure using current data. Therefore, this challenging target demands better coverage, understanding and utilization of genomic and environmental data, the development of optimized ways to integrate these data with social and other sciences and policy analysis to enable more flexible, evidence-based models to underpin FAnGR conservation. At the conference, we attempted to identify the most important problems for effective livestock genomic resource conservation during the next decade. Twenty priority questions were identified that could be broadly categorized into challenges related to methodology, analytical approaches, data management and conservation. It should be acknowledged here that while the focus of our meeting was predominantly around genetics, genomics and animal science, many of the practical challenges facing conservation of genomic resources are

  11. Prospects and challenges for the conservation of farm animal genomic resources, 2015-2025.

    Science.gov (United States)

    Bruford, Michael W; Ginja, Catarina; Hoffmann, Irene; Joost, Stéphane; Orozco-terWengel, Pablo; Alberto, Florian J; Amaral, Andreia J; Barbato, Mario; Biscarini, Filippo; Colli, Licia; Costa, Mafalda; Curik, Ino; Duruz, Solange; Ferenčaković, Maja; Fischer, Daniel; Fitak, Robert; Groeneveld, Linn F; Hall, Stephen J G; Hanotte, Olivier; Hassan, Faiz-Ul; Helsen, Philippe; Iacolina, Laura; Kantanen, Juha; Leempoel, Kevin; Lenstra, Johannes A; Ajmone-Marsan, Paolo; Masembe, Charles; Megens, Hendrik-Jan; Miele, Mara; Neuditschko, Markus; Nicolazzi, Ezequiel L; Pompanon, François; Roosen, Jutta; Sevane, Natalia; Smetko, Anamarija; Štambuk, Anamaria; Streeter, Ian; Stucki, Sylvie; Supakorn, China; Telo Da Gama, Luis; Tixier-Boichard, Michèle; Wegmann, Daniel; Zhan, Xiangjiang

    2015-01-01

    Livestock conservation practice is changing rapidly in light of policy developments, climate change and diversifying market demands. The last decade has seen a step change in technology and analytical approaches available to define, manage and conserve Farm Animal Genomic Resources (FAnGR). However, these rapid changes pose challenges for FAnGR conservation in terms of technological continuity, analytical capacity and integrative methodologies needed to fully exploit new, multidimensional data. The final conference of the ESF Genomic Resources program aimed to address these interdisciplinary problems in an attempt to contribute to the agenda for research and policy development directions during the coming decade. By 2020, according to the Convention on Biodiversity's Aichi Target 13, signatories should ensure that "…the genetic diversity of …farmed and domesticated animals and of wild relatives …is maintained, and strategies have been developed and implemented for minimizing genetic erosion and safeguarding their genetic diversity." However, the real extent of genetic erosion is very difficult to measure using current data. Therefore, this challenging target demands better coverage, understanding and utilization of genomic and environmental data, the development of optimized ways to integrate these data with social and other sciences and policy analysis to enable more flexible, evidence-based models to underpin FAnGR conservation. At the conference, we attempted to identify the most important problems for effective livestock genomic resource conservation during the next decade. Twenty priority questions were identified that could be broadly categorized into challenges related to methodology, analytical approaches, data management and conservation. It should be acknowledged here that while the focus of our meeting was predominantly around genetics, genomics and animal science, many of the practical challenges facing conservation of genomic resources are

  12. Statistical analyses of conserved features of genomic islands in bacteria.

    Science.gov (United States)

    Guo, F-B; Xia, Z-K; Wei, W; Zhao, H-L

    2014-03-17

    We performed statistical analyses of five conserved features of genomic islands of bacteria. Analyses were made based on 104 known genomic islands, which were identified by comparative methods. Four of these features include sequence size, abnormal G+C content, flanking tRNA gene, and embedded mobility gene, which are frequently investigated. One relatively new feature, G+C homogeneity, was also investigated. Among the 104 known genomic islands, 88.5% were found to fall in the typical length of 10-200 kb and 80.8% had G+C deviations with absolute values larger than 2%. For the 88 genomic islands whose hosts have been sequenced and annotated, 52.3% of them were found to have flanking tRNA genes and 64.7% had embedded mobility genes. For the homogeneity feature, 85% had an h homogeneity index less than 0.1, indicating that their G+C content is relatively uniform. Taking all the five features into account, 87.5% of 88 genomic islands had three of them. Only one genomic island had only one conserved feature and none of the genomic islands had zero features. These statistical results should help to understand the general structure of known genomic islands. We found that larger genomic islands tend to have relatively small G+C deviations relative to absolute values. For example, the absolute G+C deviations of 9 genomic islands longer than 100,000 bp were all less than 5%. This is a novel but reasonable result given that larger genomic islands should have greater restrictions in their G+C contents, in order to maintain the stable G+C content of the recipient genome.

  13. RNAi and heterochromatin repress centromeric meiotic recombination

    DEFF Research Database (Denmark)

    Ellermeier, Chad; Higuchi, Emily C; Phadnis, Naina

    2010-01-01

    During meiosis, the formation of viable haploid gametes from diploid precursors requires that each homologous chromosome pair be properly segregated to produce an exact haploid set of chromosomes. Genetic recombination, which provides a physical connection between homologous chromosomes, is essen......During meiosis, the formation of viable haploid gametes from diploid precursors requires that each homologous chromosome pair be properly segregated to produce an exact haploid set of chromosomes. Genetic recombination, which provides a physical connection between homologous chromosomes....... Surprisingly, one mutant derepressed for recombination in the heterochromatic mating-type region during meiosis and several mutants derepressed for centromeric gene expression during mitotic growth are not derepressed for centromeric recombination during meiosis. These results reveal a complex relation between...... types of repression by heterochromatin. Our results also reveal a previously undemonstrated role for RNAi and heterochromatin in the repression of meiotic centromeric recombination and, potentially, in the prevention of birth defects by maintenance of proper chromosome segregation during meiosis....

  14. Community standards for genomic resources, genetic conservation, and data integration

    Science.gov (United States)

    Jill Wegrzyn; Meg Staton; Emily Grau; Richard Cronn; C. Dana Nelson

    2017-01-01

    Genetics and genomics are increasingly important in forestry management and conservation. Next generation sequencing can increase analytical power, but still relies on building on the structure of previously acquired data. Data standards and data sharing allow the community to maximize the analytical power of high throughput genomics data. The landscape of incomplete...

  15. Replicating centromeric chromatin: Spatial and temporal control of CENP-A assembly

    International Nuclear Information System (INIS)

    Nechemia-Arbely, Yael; Fachinetti, Daniele; Cleveland, Don W.

    2012-01-01

    The centromere is the fundamental unit for insuring chromosome inheritance. This complex region has a distinct type of chromatin in which histone H3 is replaced by a structurally different homologue identified in humans as CENP-A. In metazoans, specific DNA sequences are neither required nor sufficient for centromere identity. Rather, an epigenetic mark comprised of CENP-A containing chromatin is thought to be the major determinant of centromere identity. In this view, CENP-A deposition and chromatin assembly are fundamental processes for the maintenance of centromeric identity across mitotic and meiotic divisions. Several lines of evidence support CENP-A deposition in metazoans occurring at only one time in the cell cycle. Such cell cycle-dependent loading of CENP-A is found in divergent species from human to fission yeast, albeit with differences in the cell cycle point at which CENP-A is assembled. Cell cycle dependent CENP-A deposition requires multiple assembly factors for its deposition and maintenance. This review discusses the regulation of new CENP-A deposition and its relevance to centromere identity and inheritance.

  16. Differential Regulation of Strand-Specific Transcripts from Arabidopsis Centromeric Satellite Repeats.

    Directory of Open Access Journals (Sweden)

    2005-12-01

    Full Text Available Centromeres interact with the spindle apparatus to enable chromosome disjunction and typically contain thousands of tandemly arranged satellite repeats interspersed with retrotransposons. While their role has been obscure, centromeric repeats are epigenetically modified and centromere specification has a strong epigenetic component. In the yeast Schizosaccharomyces pombe, long heterochromatic repeats are transcribed and contribute to centromere function via RNA interference (RNAi. In the higher plant Arabidopsis thaliana, as in mammalian cells, centromeric satellite repeats are short (180 base pairs, are found in thousands of tandem copies, and are methylated. We have found transcripts from both strands of canonical, bulk Arabidopsis repeats. At least one subfamily of 180-base pair repeats is transcribed from only one strand and regulated by RNAi and histone modification. A second subfamily of repeats is also silenced, but silencing is lost on both strands in mutants in the CpG DNA methyltransferase MET1, the histone deacetylase HDA6/SIL1, or the chromatin remodeling ATPase DDM1. This regulation is due to transcription from Athila2 retrotransposons, which integrate in both orientations relative to the repeats, and differs between strains of Arabidopsis. Silencing lost in met1 or hda6 is reestablished in backcrosses to wild-type, but silencing lost in RNAi mutants and ddm1 is not. Twenty-four-nucleotide small interfering RNAs from centromeric repeats are retained in met1 and hda6, but not in ddm1, and may have a role in this epigenetic inheritance. Histone H3 lysine-9 dimethylation is associated with both classes of repeats. We propose roles for transcribed repeats in the epigenetic inheritance and evolution of centromeres.

  17. Harnessing genomics for delineating conservation units.

    Science.gov (United States)

    Funk, W Chris; McKay, John K; Hohenlohe, Paul A; Allendorf, Fred W

    2012-09-01

    Genomic data have the potential to revolutionize the delineation of conservation units (CUs) by allowing the detection of adaptive genetic variation, which is otherwise difficult for rare, endangered species. In contrast to previous recommendations, we propose that the use of neutral versus adaptive markers should not be viewed as alternatives. Rather, neutral and adaptive markers provide different types of information that should be combined to make optimal management decisions. Genetic patterns at neutral markers reflect the interaction of gene flow and genetic drift that affects genome-wide variation within and among populations. This population genetic structure is what natural selection operates on to cause adaptive divergence. Here, we provide a new framework to integrate data on neutral and adaptive markers to protect biodiversity. Copyright © 2012 Elsevier Ltd. All rights reserved.

  18. A Molecular-Cytogenetic Method for Locating Genes to Pericentromeric Regions Facilitates a Genome-Wide Comparison of Syntency Between the Centrometric Regions of Wheat and Rice

    Science.gov (United States)

    Centromeres, because of their repeat structure and lack of sequence conservation, are difficult to assemble and compare across organisms. It was recently discovered that rice centromeres often contain genes. This suggested a method for studying centromere homologies between wheat and rice chromosome...

  19. Conservation genomics of the endangered Burmese roofed turtle.

    Science.gov (United States)

    Çilingir, F Gözde; Rheindt, Frank E; Garg, Kritika M; Platt, Kalyar; Platt, Steven G; Bickford, David P

    2017-12-01

    The Burmese roofed turtle (Batagur trivittata) is one of the world's most endangered turtles. Only one wild population remains in Myanmar. There are thought to be 12 breeding turtles in the wild. Conservation efforts for the species have raised >700 captive turtles since 2002, predominantly from eggs collected in the wild. We collected tissue samples from 445 individuals (approximately 40% of the turtles' remaining global population), applied double-digest restriction-site associated DNA sequencing (ddRAD-Seq), and obtained approximately 1500 unlinked genome-wide single nucleotide polymorphisms. Individuals fell into 5 distinct genetic clusters, 4 of which represented full-sib families. We inferred a low effective population size (≤10 individuals) but did not detect signs of severe inbreeding, possibly because the population bottleneck occurred recently. Two groups of 30 individuals from the captive pool that were the most genetically diverse were reintroduced to the wild, leading to an increase in the number of fertile eggs (n = 27) in the wild. Another 25 individuals, selected based on the same criteria, were transferred to the Singapore Zoo as an assurance colony. Our study demonstrates that the research-to-application gap in conservation can be bridged through application of cutting-edge genomic methods. © 2017 Society for Conservation Biology.

  20. Whole-genome sequencing approaches for conservation biology: Advantages, limitations and practical recommendations.

    Science.gov (United States)

    Fuentes-Pardo, Angela P; Ruzzante, Daniel E

    2017-10-01

    Whole-genome resequencing (WGR) is a powerful method for addressing fundamental evolutionary biology questions that have not been fully resolved using traditional methods. WGR includes four approaches: the sequencing of individuals to a high depth of coverage with either unresolved or resolved haplotypes, the sequencing of population genomes to a high depth by mixing equimolar amounts of unlabelled-individual DNA (Pool-seq) and the sequencing of multiple individuals from a population to a low depth (lcWGR). These techniques require the availability of a reference genome. This, along with the still high cost of shotgun sequencing and the large demand for computing resources and storage, has limited their implementation in nonmodel species with scarce genomic resources and in fields such as conservation biology. Our goal here is to describe the various WGR methods, their pros and cons and potential applications in conservation biology. WGR offers an unprecedented marker density and surveys a wide diversity of genetic variations not limited to single nucleotide polymorphisms (e.g., structural variants and mutations in regulatory elements), increasing their power for the detection of signatures of selection and local adaptation as well as for the identification of the genetic basis of phenotypic traits and diseases. Currently, though, no single WGR approach fulfils all requirements of conservation genetics, and each method has its own limitations and sources of potential bias. We discuss proposed ways to minimize such biases. We envision a not distant future where the analysis of whole genomes becomes a routine task in many nonmodel species and fields including conservation biology. © 2017 John Wiley & Sons Ltd.

  1. Late replication domains are evolutionary conserved in the Drosophila genome.

    Science.gov (United States)

    Andreyenkova, Natalya G; Kolesnikova, Tatyana D; Makunin, Igor V; Pokholkova, Galina V; Boldyreva, Lidiya V; Zykova, Tatyana Yu; Zhimulev, Igor F; Belyaeva, Elena S

    2013-01-01

    Drosophila chromosomes are organized into distinct domains differing in their predominant chromatin composition, replication timing and evolutionary conservation. We show on a genome-wide level that genes whose order has remained unaltered across 9 Drosophila species display late replication timing and frequently map to the regions of repressive chromatin. This observation is consistent with the existence of extensive domains of repressive chromatin that replicate extremely late and have conserved gene order in the Drosophila genome. We suggest that such repressive chromatin domains correspond to a handful of regions that complete replication at the very end of S phase. We further demonstrate that the order of genes in these regions is rarely altered in evolution. Substantial proportion of such regions significantly coincide with large synteny blocks. This indicates that there are evolutionary mechanisms maintaining the integrity of these late-replicating chromatin domains. The synteny blocks corresponding to the extremely late-replicating regions in the D. melanogaster genome consistently display two-fold lower gene density across different Drosophila species.

  2. Genetic, genomic, and molecular tools for studying the protoploid yeast, L. waltii.

    Science.gov (United States)

    Di Rienzi, Sara C; Lindstrom, Kimberly C; Lancaster, Ragina; Rolczynski, Lisa; Raghuraman, M K; Brewer, Bonita J

    2011-02-01

    Sequencing of the yeast Kluyveromyces waltii (recently renamed Lachancea waltii) provided evidence of a whole genome duplication event in the lineage leading to the well-studied Saccharomyces cerevisiae. While comparative genomic analyses of these yeasts have proven to be extremely instructive in modeling the loss or maintenance of gene duplicates, experimental tests of the ramifications following such genome alterations remain difficult. To transform L. waltii from an organism of the computational comparative genomic literature into an organism of the functional comparative genomic literature, we have developed genetic, molecular and genomic tools for working with L. waltii. In particular, we have characterized basic properties of L. waltii (growth, ploidy, molecular karyotype, mating type and the sexual cycle), developed transformation, cell cycle arrest and synchronization protocols, and have created centromeric and non-centromeric vectors as well as a genome browser for L. waltii. We hope that these tools will be used by the community to follow up on the ideas generated by sequence data and lead to a greater understanding of eukaryotic biology and genome evolution. 2010 John Wiley & Sons, Ltd.

  3. Plasmodium falciparum centromeres display a unique epigenetic makeup and cluster prior to and during schizogony.

    Science.gov (United States)

    Hoeijmakers, Wieteke A M; Flueck, Christian; Françoijs, Kees-Jan; Smits, Arne H; Wetzel, Johanna; Volz, Jennifer C; Cowman, Alan F; Voss, Till; Stunnenberg, Hendrik G; Bártfai, Richárd

    2012-09-01

    Centromeres are essential for the faithful transmission of chromosomes to the next generation, therefore being essential in all eukaryotic organisms. The centromeres of Plasmodium falciparum, the causative agent of the most severe form of malaria, have been broadly mapped on most chromosomes, but their epigenetic composition remained undefined. Here, we reveal that the centromeric histone variant PfCENH3 occupies a 4-4.5 kb region on each P. falciparum chromosome, which is devoid of pericentric heterochromatin but harbours another histone variant, PfH2A.Z. These CENH3 covered regions pinpoint the exact position of the centromere on all chromosomes and revealed that all centromeric regions have similar size and sequence composition. Immunofluorescence assay of PfCENH3 strongly suggests that P. falciparum centromeres cluster to a single nuclear location prior to and during mitosis and cytokinesis but dissociate soon after invasion. In summary, we reveal a dynamic association of Plasmodium centromeres, which bear a unique epigenetic signature and conform to a strict structure. These findings suggest that DNA-associated and epigenetic elements play an important role in centromere establishment in this important human pathogen. © 2012 Blackwell Publishing Ltd.

  4. Comparative genomics of 12 strains of Erwinia amylovora identifies a pan-genome with a large conserved core.

    Directory of Open Access Journals (Sweden)

    Rachel A Mann

    Full Text Available The plant pathogen Erwinia amylovora can be divided into two host-specific groupings; strains infecting a broad range of hosts within the Rosaceae subfamily Spiraeoideae (e.g., Malus, Pyrus, Crataegus, Sorbus and strains infecting Rubus (raspberries and blackberries. Comparative genomic analysis of 12 strains representing distinct populations (e.g., geographic, temporal, host origin of E. amylovora was used to describe the pan-genome of this major pathogen. The pan-genome contains 5751 coding sequences and is highly conserved relative to other phytopathogenic bacteria comprising on average 89% conserved, core genes. The chromosomes of Spiraeoideae-infecting strains were highly homogeneous, while greater genetic diversity was observed between Spiraeoideae- and Rubus-infecting strains (and among individual Rubus-infecting strains, the majority of which was attributed to variable genomic islands. Based on genomic distance scores and phylogenetic analysis, the Rubus-infecting strain ATCC BAA-2158 was genetically more closely related to the Spiraeoideae-infecting strains of E. amylovora than it was to the other Rubus-infecting strains. Analysis of the accessory genomes of Spiraeoideae- and Rubus-infecting strains has identified putative host-specific determinants including variation in the effector protein HopX1(Ea and a putative secondary metabolite pathway only present in Rubus-infecting strains.

  5. Repeatless and Repeat-Based Centromeres in Potato: Implications for Centromere Evolution[C][W

    Czech Academy of Sciences Publication Activity Database

    Gong, Z.; Wu, Y.; Koblížková, Andrea; Torres, G.A.; Wang, K.; Iovene, M.; Neumann, Pavel; Zhang, W.; Novák, Petr; Buell, C.R.; Macas, Jiří; Jiang, J.

    2012-01-01

    Roč. 24, č. 9 (2012), s. 3559-3574 ISSN 1040-4651 R&D Projects: GA MŠk(CZ) LH11058 Institutional research plan: CEZ:AV0Z50510513 Institutional support: RVO:60077344 Keywords : repetitive sequences * plant satellite repeats * Arabidopsis thaliana * rice centromere * wild potatoes Subject RIV: EB - Genetics ; Molecular Biology Impact factor: 9.251, year: 2012

  6. Cross-species genome-wide identification of evolutionary conserved microproteins

    DEFF Research Database (Denmark)

    Straub, Daniel; Wenkel, Stephan

    2017-01-01

    Protein concept beyond transcription factors to other protein families. Here, we reveal potential microProtein candidates in several plant and animal reference genomes. A large number of these microProteins are species-specific while others evolved early and are evolutionary highly conserved. Most known micro...... act in plant transcriptional regulation, signal transduction and anatomical structure development. MiPFinder is freely available to find microProteins in any genome and will aid in the identification of novel microProteins in plants and animals....

  7. Evolutionary movement of centromeres in horse, donkey, and zebra.

    Science.gov (United States)

    Carbone, Lucia; Nergadze, Solomon G; Magnani, Elisa; Misceo, Doriana; Francesca Cardone, Maria; Roberto, Roberta; Bertoni, Livia; Attolini, Carmen; Francesca Piras, Maria; de Jong, Pieter; Raudsepp, Terje; Chowdhary, Bhanu P; Guérin, Gérard; Archidiacono, Nicoletta; Rocchi, Mariano; Giulotto, Elena

    2006-06-01

    Centromere repositioning (CR) is a recently discovered biological phenomenon consisting of the emergence of a new centromere along a chromosome and the inactivation of the old one. After a CR, the primary constriction and the centromeric function are localized in a new position while the order of physical markers on the chromosome remains unchanged. These events profoundly affect chromosomal architecture. Since horses, asses, and zebras, whose evolutionary divergence is relatively recent, show remarkable morphological similarity and capacity to interbreed despite their chromosomes differing considerably, we investigated the role of CR in the karyotype evolution of the genus Equus. Using appropriate panels of BAC clones in FISH experiments, we compared the centromere position and marker order arrangement among orthologous chromosomes of Burchelli's zebra (Equus burchelli), donkey (Equus asinus), and horse (Equus caballus). Surprisingly, at least eight CRs took place during the evolution of this genus. Even more surprisingly, five cases of CR have occurred in the donkey after its divergence from zebra, that is, in a very short evolutionary time (approximately 1 million years). These findings suggest that in some species the CR phenomenon could have played an important role in karyotype shaping, with potential consequences on population dynamics and speciation.

  8. Using genomic information to conserve genetic diversity in livestock

    NARCIS (Netherlands)

    Eynard, Sonia E.

    2018-01-01

    Concern about the status of livestock breeds and their conservation has increased as selection and small population sizes caused loss of genetic diversity. Meanwhile, dense SNP chips and whole genome sequences (WGS) became available, providing opportunities to accurately quantify the impact of

  9. Conservation genomics of natural and managed populations: building a conceptual and practical framework.

    Science.gov (United States)

    Benestan, Laura Marilyn; Ferchaud, Anne-Laure; Hohenlohe, Paul A; Garner, Brittany A; Naylor, Gavin J P; Baums, Iliana Brigitta; Schwartz, Michael K; Kelley, Joanna L; Luikart, Gordon

    2016-07-01

    The boom of massive parallel sequencing (MPS) technology and its applications in conservation of natural and managed populations brings new opportunities and challenges to meet the scientific questions that can be addressed. Genomic conservation offers a wide range of approaches and analytical techniques, with their respective strengths and weaknesses that rely on several implicit assumptions. However, finding the most suitable approaches and analysis regarding our scientific question are often difficult and time-consuming. To address this gap, a recent workshop entitled 'ConGen 2015' was held at Montana University in order to bring together the knowledge accumulated in this field and to provide training in conceptual and practical aspects of data analysis applied to the field of conservation and evolutionary genomics. Here, we summarize the expertise yield by each instructor that has led us to consider the importance of keeping in mind the scientific question from sampling to management practices along with the selection of appropriate genomics tools and bioinformatics challenges. © 2016 John Wiley & Sons Ltd.

  10. Conserved genomic organisation of Group B Sox genes in insects.

    Directory of Open Access Journals (Sweden)

    Woerfel Gertrud

    2005-05-01

    Full Text Available Abstract Background Sox domain containing genes are important metazoan transcriptional regulators implicated in a wide rage of developmental processes. The vertebrate B subgroup contains the Sox1, Sox2 and Sox3 genes that have early functions in neural development. Previous studies show that Drosophila Group B genes have been functionally conserved since they play essential roles in early neural specification and mutations in the Drosophila Dichaete and SoxN genes can be rescued with mammalian Sox genes. Despite their importance, the extent and organisation of the Group B family in Drosophila has not been fully characterised, an important step in using Drosophila to examine conserved aspects of Group B Sox gene function. Results We have used the directed cDNA sequencing along with the output from the publicly-available genome sequencing projects to examine the structure of Group B Sox domain genes in Drosophila melanogaster, Drosophila pseudoobscura, Anopheles gambiae and Apis mellifora. All of the insect genomes contain four genes encoding Group B proteins, two of which are intronless, as is the case with vertebrate group B genes. As has been previously reported and unusually for Group B genes, two of the insect group B genes, Sox21a and Sox21b, contain introns within their DNA-binding domains. We find that the highly unusual multi-exon structure of the Sox21b gene is common to the insects. In addition, we find that three of the group B Sox genes are organised in a linked cluster in the insect genomes. By in situ hybridisation we show that the pattern of expression of each of the four group B genes during embryogenesis is conserved between D. melanogaster and D. pseudoobscura. Conclusion The DNA-binding domain sequences and genomic organisation of the group B genes have been conserved over 300 My of evolution since the last common ancestor of the Hymenoptera and the Diptera. Our analysis suggests insects have two Group B1 genes, SoxN and

  11. Essentiality, conservation, evolutionary pressure and codon bias in bacterial genomes.

    Science.gov (United States)

    Dilucca, Maddalena; Cimini, Giulio; Giansanti, Andrea

    2018-07-15

    Essential genes constitute the core of genes which cannot be mutated too much nor lost along the evolutionary history of a species. Natural selection is expected to be stricter on essential genes and on conserved (highly shared) genes, than on genes that are either nonessential or peculiar to a single or a few species. In order to further assess this expectation, we study here how essentiality of a gene is connected with its degree of conservation among several unrelated bacterial species, each one characterised by its own codon usage bias. Confirming previous results on E. coli, we show the existence of a universal exponential relation between gene essentiality and conservation in bacteria. Moreover, we show that, within each bacterial genome, there are at least two groups of functionally distinct genes, characterised by different levels of conservation and codon bias: i) a core of essential genes, mainly related to cellular information processing; ii) a set of less conserved nonessential genes with prevalent functions related to metabolism. In particular, the genes in the first group are more retained among species, are subject to a stronger purifying conservative selection and display a more limited repertoire of synonymous codons. The core of essential genes is close to the minimal bacterial genome, which is in the focus of recent studies in synthetic biology, though we confirm that orthologs of genes that are essential in one species are not necessarily essential in other species. We also list a set of highly shared genes which, reasonably, could constitute a reservoir of targets for new anti-microbial drugs. Copyright © 2018 Elsevier B.V. All rights reserved.

  12. Mps1 kinase-dependent Sgo2 centromere localisation mediates cohesin protection in mouse oocyte meiosis I.

    Science.gov (United States)

    El Yakoubi, Warif; Buffin, Eulalie; Cladière, Damien; Gryaznova, Yulia; Berenguer, Inés; Touati, Sandra A; Gómez, Rocío; Suja, José A; van Deursen, Jan M; Wassmann, Katja

    2017-09-25

    A key feature of meiosis is the step-wise removal of cohesin, the protein complex holding sister chromatids together, first from arms in meiosis I and then from the centromere region in meiosis II. Centromeric cohesin is protected by Sgo2 from Separase-mediated cleavage, in order to maintain sister chromatids together until their separation in meiosis II. Failures in step-wise cohesin removal result in aneuploid gametes, preventing the generation of healthy embryos. Here, we report that kinase activities of Bub1 and Mps1 are required for Sgo2 localisation to the centromere region. Mps1 inhibitor-treated oocytes are defective in centromeric cohesin protection, whereas oocytes devoid of Bub1 kinase activity, which cannot phosphorylate H2A at T121, are not perturbed in cohesin protection as long as Mps1 is functional. Mps1 and Bub1 kinase activities localise Sgo2 in meiosis I preferentially to the centromere and pericentromere respectively, indicating that Sgo2 at the centromere is required for protection.In meiosis I centromeric cohesin is protected by Sgo2 from Separase-mediated cleavage ensuring that sister chromatids are kept together until their separation in meiosis II. Here the authors demonstrate that Bub1 and Mps1 kinase activities are required for Sgo2 localisation to the centromere region.

  13. Prospects and challenges for the conservation of farm animal genomic resources, 2015-2025

    NARCIS (Netherlands)

    Bruford, M.W.; Ginja, Catarina; Hoffmann, Irene; Megens, Hendrik Jan

    2015-01-01

    Livestock conservation practice is changing rapidly in light of policy developments, climate change and diversifying market demands. The last decade has seen a step change in technology and analytical approaches available to define, manage and conserve Farm Animal Genomic Resources (FAnGR).

  14. Prospects and challenges for the conservation of farm animal genomic resources, 2015-2025

    NARCIS (Netherlands)

    Bruford, Michael W; Ginja, Catarina; Hoffmann, Irene; Joost, Stéphane; Orozco-terWengel, Pablo; Alberto, Florian J; Amaral, Andreia J; Barbato, Mario; Biscarini, Filippo; Colli, Licia; Costa, Mafalda; Curik, Ino; Duruz, Solange; Ferenčaković, Maja; Fischer, Daniel; Fitak, Robert; Groeneveld, Linn F; Hall, Stephen J G; Hanotte, Olivier; Hassan, Faiz-Ul; Helsen, Philippe; Iacolina, Laura; Kantanen, Juha; Leempoel, Kevin; Lenstra, Johannes A; Ajmone-Marsan, Paolo; Masembe, Charles; Megens, Hendrik-Jan; Miele, Mara; Neuditschko, Markus; Nicolazzi, Ezequiel L; Pompanon, François; Roosen, Jutta; Sevane, Natalia; Smetko, Anamarija; Štambuk, Anamaria; Streeter, Ian; Stucki, Sylvie; Supakorn, China; Telo Da Gama, Luis; Tixier-Boichard, Michèle; Wegmann, Daniel; Zhan, Xiangjiang

    2015-01-01

    Livestock conservation practice is changing rapidly in light of policy developments, climate change and diversifying market demands. The last decade has seen a step change in technology and analytical approaches available to define, manage and conserve Farm Animal Genomic Resources (FAnGR). However,

  15. Identification and classification of conserved RNA secondary structures in the human genome

    DEFF Research Database (Denmark)

    Pedersen, Jakob Skou; Bejerano, Gill; Siepel, Adam

    2006-01-01

    The discoveries of microRNAs and riboswitches, among others, have shown functional RNAs to be biologically more important and genomically more prevalent than previously anticipated. We have developed a general comparative genomics method based on phylogenetic stochastic context-free grammars...... for identifying functional RNAs encoded in the human genome and used it to survey an eight-way genome-wide alignment of the human, chimpanzee, mouse, rat, dog, chicken, zebra-fish, and puffer-fish genomes for deeply conserved functional RNAs. At a loose threshold for acceptance, this search resulted in a set......, the results nevertheless provide evidence for many new human functional RNAs and present specific predictions to facilitate their further characterization....

  16. Recurrent Gene Duplication Leads to Diverse Repertoires of Centromeric Histones in Drosophila Species.

    Science.gov (United States)

    Kursel, Lisa E; Malik, Harmit S

    2017-06-01

    Despite their essential role in the process of chromosome segregation in most eukaryotes, centromeric histones show remarkable evolutionary lability. Not only have they been lost in multiple insect lineages, but they have also undergone gene duplication in multiple plant lineages. Based on detailed study of a handful of model organisms including Drosophila melanogaster, centromeric histone duplication is considered to be rare in animals. Using a detailed phylogenomic study, we find that Cid, the centromeric histone gene, has undergone at least four independent gene duplications during Drosophila evolution. We find duplicate Cid genes in D. eugracilis (Cid2), in the montium species subgroup (Cid3, Cid4) and in the entire Drosophila subgenus (Cid5). We show that Cid3, Cid4, and Cid5 all localize to centromeres in their respective species. Some Cid duplicates are primarily expressed in the male germline. With rare exceptions, Cid duplicates have been strictly retained after birth, suggesting that they perform nonredundant centromeric functions, independent from the ancestral Cid. Indeed, each duplicate encodes a distinct N-terminal tail, which may provide the basis for distinct protein-protein interactions. Finally, we show some Cid duplicates evolve under positive selection whereas others do not. Taken together, our results support the hypothesis that Drosophila Cid duplicates have subfunctionalized. Thus, these gene duplications provide an unprecedented opportunity to dissect the multiple roles of centromeric histones. © The Author 2017. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

  17. Contribution of genetics and genomics to seagrass biology and conservation

    NARCIS (Netherlands)

    Procaccini, Gabriele; Olsen, Jeanine L.; Reusch, Thorsten B. H.

    2007-01-01

    Genetic diversity is one of three forms of biodiversity recognized by the IUCN as deserving conservation along with species and ecosystems. Seagrasses provide all three levels in one. This review addresses the latest advances in our understanding of seagrass population genetics and genomics within

  18. Mitochondrial genome sequences illuminate maternal lineages of conservation concern in a rare carnivore

    Science.gov (United States)

    Brian J. Knaus; Richard Cronn; Aaron Liston; Kristine Pilgrim; Michael K. Schwartz

    2011-01-01

    Science-based wildlife management relies on genetic information to infer population connectivity and identify conservation units. The most commonly used genetic marker for characterizing animal biodiversity and identifying maternal lineages is the mitochondrial genome. Mitochondrial genotyping figures prominently in conservation and management plans, with much of the...

  19. Integrating population genetics and conservation biology in the era of genomics.

    Science.gov (United States)

    Ouborg, N Joop

    2010-02-23

    As one of the final activities of the ESF-CONGEN Networking programme, a conference entitled 'Integrating Population Genetics and Conservation Biology' was held at Trondheim, Norway, from 23 to 26 May 2009. Conference speakers and poster presenters gave a display of the state-of-the-art developments in the field of conservation genetics. Over the five-year running period of the successful ESF-CONGEN Networking programme, much progress has been made in theoretical approaches, basic research on inbreeding depression and other genetic processes associated with habitat fragmentation and conservation issues, and with applying principles of conservation genetics in the conservation of many species. Future perspectives were also discussed in the conference, and it was concluded that conservation genetics is evolving into conservation genomics, while at the same time basic and applied research on threatened species and populations from a population genetic point of view continues to be emphasized.

  20. Somatic association of telocentric chromosomes carrying homologous centromeres in common wheat.

    Science.gov (United States)

    Mello-Sampayo, T

    1973-01-01

    Measurements of distances between telocentric chromosomes, either homologous or representing the opposite arms of a metacentric chromosome (complementary telocentrics), were made at metaphase in root tip cells of common wheat carrying two homologous pairs of complementary telocentrics of chromosome 1 B or 6 B (double ditelosomic 1 B or 6 B). The aim was to elucidate the relative locations of the telocentric chromosomes within the cell. The data obtained strongly suggest that all four telocentrics of chromosome 1 B or 6 B are spacially and simultaneously co-associated. In plants carrying two complementary (6 B (S) and 6 B (L)) and a non-related (5 B (L)) telocentric, only the complementary chromosomes were found to be somatically associated. It is thought, therefore, that the somatic association of chromosomes may involve more than two chromosomes in the same association and, since complementary telocentrics are as much associated as homologous, that the homology between centromeres (probably the only homologous region that exists between complementary telocentrics) is a very important condition for somatic association of chromosomes. The spacial arrangement of chromosomes was studied at anaphase and prophase and the polar orientation of chromosomes at prophase was found to resemble anaphase orientation. This was taken as good evidence for the maintenance of the chromosome arrangement - the Rabl orientation - and of the peripheral location of the centromere and its association with the nuclear membrane. Within this general arrangement homologous telocentric chromosomes were frequently seen to have their centromeres associated or directed towards each other. The role of the centromere in somatic association as a spindle fibre attachment and chromosome binder is discussed. It is suggested that for non-homologous chromosomes to become associated in root tips, the only requirement needed should be the homology of centromeres such as exists between complementary

  1. Molecular cloning and characterization of satellite DNA sequences from constitutive heterochromatin of the habu snake (Protobothrops flavoviridis, Viperidae) and the Burmese python (Python bivittatus, Pythonidae).

    Science.gov (United States)

    Matsubara, Kazumi; Uno, Yoshinobu; Srikulnath, Kornsorn; Seki, Risako; Nishida, Chizuko; Matsuda, Yoichi

    2015-12-01

    Highly repetitive DNA sequences of the centromeric heterochromatin provide valuable molecular cytogenetic markers for the investigation of genomic compartmentalization in the macrochromosomes and microchromosomes of sauropsids. Here, the relationship between centromeric heterochromatin and karyotype evolution was examined using cloned repetitive DNA sequences from two snake species, the habu snake (Protobothrops flavoviridis, Crotalinae, Viperidae) and Burmese python (Python bivittatus, Pythonidae). Three satellite DNA (stDNA) families were isolated from the heterochromatin of these snakes: 168-bp PFL-MspI from P. flavoviridis and 196-bp PBI-DdeI and 174-bp PBI-MspI from P. bivittatus. The PFL-MspI and PBI-DdeI sequences were localized to the centromeric regions of most chromosomes in the respective species, suggesting that the two sequences were the major components of the centromeric heterochromatin in these organisms. The PBI-MspI sequence was localized to the pericentromeric region of four chromosome pairs. The PFL-MspI and the PBI-DdeI sequences were conserved only in the genome of closely related species, Gloydius blomhoffii (Crotalinae) and Python molurus, respectively, although their locations on the chromosomes were slightly different. In contrast, the PBI-MspI sequence was also in the genomes of P. molurus and Boa constrictor (Boidae), and additionally localized to the centromeric regions of eight chromosome pairs in B. constrictor, suggesting that this sequence originated in the genome of a common ancestor of Pythonidae and Boidae, approximately 86 million years ago. The three stDNA sequences showed no genomic compartmentalization between the macrochromosomes and microchromosomes, suggesting that homogenization of the centromeric and/or pericentromeric stDNA sequences occurred in the macrochromosomes and microchromosomes of these snakes.

  2. G-quadruplex DNA sequences are evolutionarily conserved and associated with distinct genomic features in Saccharomyces cerevisiae.

    Directory of Open Access Journals (Sweden)

    John A Capra

    2010-07-01

    Full Text Available G-quadruplex DNA is a four-stranded DNA structure formed by non-Watson-Crick base pairing between stacked sets of four guanines. Many possible functions have been proposed for this structure, but its in vivo role in the cell is still largely unresolved. We carried out a genome-wide survey of the evolutionary conservation of regions with the potential to form G-quadruplex DNA structures (G4 DNA motifs across seven yeast species. We found that G4 DNA motifs were significantly more conserved than expected by chance, and the nucleotide-level conservation patterns suggested that the motif conservation was the result of the formation of G4 DNA structures. We characterized the association of conserved and non-conserved G4 DNA motifs in Saccharomyces cerevisiae with more than 40 known genome features and gene classes. Our comprehensive, integrated evolutionary and functional analysis confirmed the previously observed associations of G4 DNA motifs with promoter regions and the rDNA, and it identified several previously unrecognized associations of G4 DNA motifs with genomic features, such as mitotic and meiotic double-strand break sites (DSBs. Conserved G4 DNA motifs maintained strong associations with promoters and the rDNA, but not with DSBs. We also performed the first analysis of G4 DNA motifs in the mitochondria, and surprisingly found a tenfold higher concentration of the motifs in the AT-rich yeast mitochondrial DNA than in nuclear DNA. The evolutionary conservation of the G4 DNA motif and its association with specific genome features supports the hypothesis that G4 DNA has in vivo functions that are under evolutionary constraint.

  3. Transgenerational propagation and quantitative maintenance of paternal centromeres depends on Cid/Cenp-A presence in Drosophila sperm.

    Directory of Open Access Journals (Sweden)

    Nitika Raychaudhuri

    Full Text Available In Drosophila melanogaster, as in many animal and plant species, centromere identity is specified epigenetically. In proliferating cells, a centromere-specific histone H3 variant (CenH3, named Cid in Drosophila and Cenp-A in humans, is a crucial component of the epigenetic centromere mark. Hence, maintenance of the amount and chromosomal location of CenH3 during mitotic proliferation is important. Interestingly, CenH3 may have different roles during meiosis and the onset of embryogenesis. In gametes of Caenorhabditis elegans, and possibly in plants, centromere marking is independent of CenH3. Moreover, male gamete differentiation in animals often includes global nucleosome for protamine exchange that potentially could remove CenH3 nucleosomes. Here we demonstrate that the control of Cid loading during male meiosis is distinct from the regulation observed during the mitotic cycles of early embryogenesis. But Cid is present in mature sperm. After strong Cid depletion in sperm, paternal centromeres fail to integrate into the gonomeric spindle of the first mitosis, resulting in gynogenetic haploid embryos. Furthermore, after moderate depletion, paternal centromeres are unable to re-acquire normal Cid levels in the next generation. We conclude that Cid in sperm is an essential component of the epigenetic centromere mark on paternal chromosomes and it exerts quantitative control over centromeric Cid levels throughout development. Hence, the amount of Cid that is loaded during each cell cycle appears to be determined primarily by the preexisting centromeric Cid, with little flexibility for compensation of accidental losses.

  4. Boom-Bust Turnovers of Megabase-Sized Centromeric DNA in Solanum Species: Rapid Evolution of DNA Sequences Associated with Centromeres

    Czech Academy of Sciences Publication Activity Database

    Zhang, H.Q.; Koblížková, Andrea; Wang, K.; Gong, Z.Y.; Oliveira, L.; Torres, G.A.; Wu, Y.; Zhang, W.; Novák, Petr; Buell, C.R.; Macas, Jiří; Jiang, J.

    2014-01-01

    Roč. 26, č. 4 (2014), s. 1436-1447 ISSN 1040-4651 Institutional support: RVO:60077344 Keywords : Alpha-satellite DNA * repetitive sequences * rice centromeres Subject RIV: EB - Genetics ; Molecular Biology Impact factor: 9.338, year: 2014

  5. Evolutionary growth process of highly conserved sequences in vertebrate genomes.

    Science.gov (United States)

    Ishibashi, Minaka; Noda, Akiko Ogura; Sakate, Ryuichi; Imanishi, Tadashi

    2012-08-01

    Genome sequence comparison between evolutionarily distant species revealed ultraconserved elements (UCEs) among mammals under strong purifying selection. Most of them were also conserved among vertebrates. Because they tend to be located in the flanking regions of developmental genes, they would have fundamental roles in creating vertebrate body plans. However, the evolutionary origin and selection mechanism of these UCEs remain unclear. Here we report that UCEs arose in primitive vertebrates, and gradually grew in vertebrate evolution. We searched for UCEs in two teleost fishes, Tetraodon nigroviridis and Oryzias latipes, and found 554 UCEs with 100% identity over 100 bps. Comparison of teleost and mammalian UCEs revealed 43 pairs of common, jawed-vertebrate UCEs (jUCE) with high sequence identities, ranging from 83.1% to 99.2%. Ten of them retain lower similarities to the Petromyzon marinus genome, and the substitution rates of four non-exonic jUCEs were reduced after the teleost-mammal divergence, suggesting that robust conservation had been acquired in the jawed vertebrate lineage. Our results indicate that prototypical UCEs originated before the divergence of jawed and jawless vertebrates and have been frozen as perfect conserved sequences in the jawed vertebrate lineage. In addition, our comparative sequence analyses of UCEs and neighboring regions resulted in a discovery of lineage-specific conserved sequences. They were added progressively to prototypical UCEs, suggesting step-wise acquisition of novel regulatory roles. Our results indicate that conserved non-coding elements (CNEs) consist of blocks with distinct evolutionary history, each having been frozen since different evolutionary era along the vertebrate lineage. Copyright © 2012 Elsevier B.V. All rights reserved.

  6. Human Artificial Chromosomes with Alpha Satellite-Based De Novo Centromeres Show Increased Frequency of Nondisjunction and Anaphase Lag

    OpenAIRE

    Rudd, M. Katharine; Mays, Robert W.; Schwartz, Stuart; Willard, Huntington F.

    2003-01-01

    Human artificial chromosomes have been used to model requirements for human chromosome segregation and to explore the nature of sequences competent for centromere function. Normal human centromeres require specialized chromatin that consists of alpha satellite DNA complexed with epigenetically modified histones and centromere-specific proteins. While several types of alpha satellite DNA have been used to assemble de novo centromeres in artificial chromosome assays, the extent to which they fu...

  7. Most Uv-Induced Reciprocal Translocations in SORDARIA MACROSPORA Occur in or near Centromere Regions.

    Science.gov (United States)

    Leblon, G; Zickler, D; Lebilcot, S

    1986-02-01

    In fungi, translocations can be identified and classified by the patterns of ascospore abortion in asci from crosses of rearrangement x normal sequence. Previous studies of UV-induced rearrangements in Sordaria macrospora revealed that a major class (called type III) appeared to be reciprocal translocations that were anomalous in producing an unexpected class of asci with four aborted ascospores in bbbbaaaa linear sequence (b = black; a = abortive). The present study shows that the anomalous type III rearrangements are, in fact, reciprocal translocations having both breakpoints within or adjacent to centromeres and that bbbbaaaa asci result from 3:1 disjunction from the translocation quadrivalent.-Electron microscopic observations of synaptonemal complexes enable centromeres to be visualized. Lengths of synaptonemal complexes lateral elements in translocation quadrivalents accurately reflect chromosome arm lengths, enabling breakpoints to be located reliably in centromere regions. All genetic data are consistent with the behavior expected of translocations with breakpoints at centromeres.-Two-thirds of the UV-induced reciprocal translocations are of this type. Certain centromere regions are involved preferentially. Among 73 type-III translocations, there were but 13 of the 21 possible chromosome combinations and 20 of the 42 possible combinations of chromosome arms.

  8. Absence of positive selection on CenH3 in Luzula suggests that holokinetic chromosomes may suppress centromere drive.

    Science.gov (United States)

    Zedek, František; Bureš, Petr

    2016-12-01

    The centromere drive theory explains diversity of eukaryotic centromeres as a consequence of the recurrent conflict between centromeric repeats and centromeric histone H3 (CenH3), in which selfish centromeres exploit meiotic asymmetry and CenH3 evolves adaptively to counterbalance deleterious consequences of driving centromeres. Accordingly, adaptively evolving CenH3 has so far been observed only in eukaryotes with asymmetric meiosis. However, if such evolution is a consequence of centromere drive, it should depend not only on meiotic asymmetry but also on monocentric or holokinetic chromosomal structure. Selective pressures acting on CenH3 have never been investigated in organisms with holokinetic meiosis despite the fact that holokinetic chromosomes have been hypothesized to suppress centromere drive. Therefore, the present study evaluates selective pressures acting on the CenH3 gene in holokinetic organisms for the first time, specifically in the representatives of the plant genus Luzula (Juncaceae), in which the kinetochore formation is not co-localized with any type of centromeric repeat. PCR, cloning and sequencing, and database searches were used to obtain coding CenH3 sequences from Luzula species. Codon substitution models were employed to infer selective regimes acting on CenH3 in Luzula KEY RESULTS: In addition to the two previously published CenH3 sequences from L. nivea, 16 new CenH3 sequences have been isolated from 12 Luzula species. Two CenH3 isoforms in Luzula that originated by a duplication event prior to the divergence of analysed species were found. No signs of positive selection acting on CenH3 in Luzula were detected. Instead, evidence was found that selection on CenH3 of Luzula might have been relaxed. The results indicate that holokinetism itself may suppress centromere drive and, therefore, holokinetic chromosomes might have evolved as a defence against centromere drive. © The Author 2016. Published by Oxford University Press on behalf of

  9. The Arabidopsis lyrata genome sequence and the basis of rapid genome size change

    Energy Technology Data Exchange (ETDEWEB)

    Hu, Tina T.; Pattyn, Pedro; Bakker, Erica G.; Cao, Jun; Cheng, Jan-Fang; Clark, Richard M.; Fahlgren, Noah; Fawcett, Jeffrey A.; Grimwood, Jane; Gundlach, Heidrun; Haberer, Georg; Hollister, Jesse D.; Ossowski, Stephan; Ottilar, Robert P.; Salamov, Asaf A.; Schneeberger, Korbinian; Spannagl, Manuel; Wang, Xi; Yang, Liang; Nasrallah, Mikhail E.; Bergelson, Joy; Carrington, James C.; Gaut, Brandon S.; Schmutz, Jeremy; Mayer, Klaus F. X.; Van de Peer, Yves; Grigoriev, Igor V.; Nordborg, Magnus; Weigel, Detlef; Guo, Ya-Long

    2011-04-29

    In our manuscript, we present a high-quality genome sequence of the Arabidopsis thaliana relative, Arabidopsis lyrata, produced by dideoxy sequencing. We have performed the usual types of genome analysis (gene annotation, dN/dS studies etc. etc.), but this is relegated to the Supporting Information. Instead, we focus on what was a major motivation for sequencing this genome, namely to understand how A. thaliana lost half its genome in a few million years and lived to tell the tale. The rather surprising conclusion is that there is not a single genomic feature that accounts for the reduced genome, but that every aspect centromeres, intergenic regions, transposable elements, gene family number is affected through hundreds of thousands of cuts. This strongly suggests that overall genome size in itself is what has been under selection, a suggestion that is strongly supported by our demonstration (using population genetics data from A. thaliana) that new deletions seem to be driven to fixation.

  10. Localizing recent adaptive evolution in the human genome

    DEFF Research Database (Denmark)

    Williamson, Scott H; Hubisz, Melissa J; Clark, Andrew G

    2007-01-01

    , clusters of olfactory receptors, genes involved in nervous system development and function, immune system genes, and heat shock genes. We also observe consistent evidence of selective sweeps in centromeric regions. In general, we find that recent adaptation is strikingly pervasive in the human genome......-nucleotide polymorphism ascertainment, while also providing fine-scale estimates of the position of the selected site, we analyzed a genomic dataset of 1.2 million human single-nucleotide polymorphisms genotyped in African-American, European-American, and Chinese samples. We identify 101 regions of the human genome...

  11. Meiosis-Specific Loading of the Centromere-Specific Histone CENH3 in Arabidopsis thaliana

    Science.gov (United States)

    Ravi, Maruthachalam; Shibata, Fukashi; Ramahi, Joseph S.; Nagaki, Kiyotaka; Chen, Changbin; Murata, Minoru; Chan, Simon W. L.

    2011-01-01

    Centromere behavior is specialized in meiosis I, so that sister chromatids of homologous chromosomes are pulled toward the same side of the spindle (through kinetochore mono-orientation) and chromosome number is reduced. Factors required for mono-orientation have been identified in yeast. However, comparatively little is known about how meiotic centromere behavior is specialized in animals and plants that typically have large tandem repeat centromeres. Kinetochores are nucleated by the centromere-specific histone CENH3. Unlike conventional histone H3s, CENH3 is rapidly evolving, particularly in its N-terminal tail domain. Here we describe chimeric variants of CENH3 with alterations in the N-terminal tail that are specifically defective in meiosis. Arabidopsis thaliana cenh3 mutants expressing a GFP-tagged chimeric protein containing the H3 N-terminal tail and the CENH3 C-terminus (termed GFP-tailswap) are sterile because of random meiotic chromosome segregation. These defects result from the specific depletion of GFP-tailswap protein from meiotic kinetochores, which contrasts with its normal localization in mitotic cells. Loss of the GFP-tailswap CENH3 variant in meiosis affects recruitment of the essential kinetochore protein MIS12. Our findings suggest that CENH3 loading dynamics might be regulated differently in mitosis and meiosis. As further support for our hypothesis, we show that GFP-tailswap protein is recruited back to centromeres in a subset of pollen grains in GFP-tailswap once they resume haploid mitosis. Meiotic recruitment of the GFP-tailswap CENH3 variant is not restored by removal of the meiosis-specific cohesin subunit REC8. Our results reveal the existence of a specialized loading pathway for CENH3 during meiosis that is likely to involve the hypervariable N-terminal tail. Meiosis-specific CENH3 dynamics may play a role in modulating meiotic centromere behavior. PMID:21695238

  12. Toxoplasma gondii chromodomain protein 1 binds to heterochromatin and colocalises with centromeres and telomeres at the nuclear periphery.

    Directory of Open Access Journals (Sweden)

    Mathieu Gissot

    Full Text Available BACKGROUND: Apicomplexan parasites are responsible for some of the most deadly parasitic diseases afflicting humans, including malaria and toxoplasmosis. These obligate intracellular parasites exhibit a complex life cycle and a coordinated cell cycle-dependant expression program. Their cell division is a coordinated multistep process. How this complex mechanism is organised remains poorly understood. METHODS AND FINDINGS: In this study, we provide evidence for a link between heterochromatin, cell division and the compartmentalisation of the nucleus in Toxoplasma gondii. We characterised a T. gondii chromodomain containing protein (named TgChromo1 that specifically binds to heterochromatin. Using ChIP-on-chip on a genome-wide scale, we report TgChromo1 enrichment at the peri-centromeric chromatin. In addition, we demonstrate that TgChromo1 is cell-cycle regulated and co-localised with markers of the centrocone. Through the loci-specific FISH technique for T. gondii, we confirmed that TgChromo1 occupies the same nuclear localisation as the peri-centromeric sequences. CONCLUSION: We propose that TgChromo1 may play a role in the sequestration of chromosomes at the nuclear periphery and in the process of T. gondii cell division.

  13. Fungal genome and mating system transitions facilitated by chromosomal translocations involving intercentromeric recombination.

    Directory of Open Access Journals (Sweden)

    Sheng Sun

    2017-08-01

    Full Text Available Species within the human pathogenic Cryptococcus species complex are major threats to public health, causing approximately 1 million annual infections globally. Cryptococcus amylolentus is the most closely known related species of the pathogenic Cryptococcus species complex, and it is non-pathogenic. Additionally, while pathogenic Cryptococcus species have bipolar mating systems with a single large mating type (MAT locus that represents a derived state in Basidiomycetes, C. amylolentus has a tetrapolar mating system with 2 MAT loci (P/R and HD located on different chromosomes. Thus, studying C. amylolentus will shed light on the transition from tetrapolar to bipolar mating systems in the pathogenic Cryptococcus species, as well as its possible link with the origin and evolution of pathogenesis. In this study, we sequenced, assembled, and annotated the genomes of 2 C. amylolentus isolates, CBS6039 and CBS6273, which are sexual and interfertile. Genome comparison between the 2 C. amylolentus isolates identified the boundaries and the complete gene contents of the P/R and HD MAT loci. Bioinformatic and chromatin immunoprecipitation sequencing (ChIP-seq analyses revealed that, similar to those of the pathogenic Cryptococcus species, C. amylolentus has regional centromeres (CENs that are enriched with species-specific transposable and repetitive DNA elements. Additionally, we found that while neither the P/R nor the HD locus is physically closely linked to its centromere in C. amylolentus, and the regions between the MAT loci and their respective centromeres show overall synteny between the 2 genomes, both MAT loci exhibit genetic linkage to their respective centromere during meiosis, suggesting the presence of recombinational suppressors and/or epistatic gene interactions in the MAT-CEN intervening regions. Furthermore, genomic comparisons between C. amylolentus and related pathogenic Cryptococcus species provide evidence that multiple chromosomal

  14. A Three-Dimensional Model of the Yeast Genome

    Science.gov (United States)

    Noble, William; Duan, Zhi-Jun; Andronescu, Mirela; Schutz, Kevin; McIlwain, Sean; Kim, Yoo Jung; Lee, Choli; Shendure, Jay; Fields, Stanley; Blau, C. Anthony

    Layered on top of information conveyed by DNA sequence and chromatin are higher order structures that encompass portions of chromosomes, entire chromosomes, and even whole genomes. Interphase chromosomes are not positioned randomly within the nucleus, but instead adopt preferred conformations. Disparate DNA elements co-localize into functionally defined aggregates or factories for transcription and DNA replication. In budding yeast, Drosophila and many other eukaryotes, chromosomes adopt a Rabl configuration, with arms extending from centromeres adjacent to the spindle pole body to telomeres that abut the nuclear envelope. Nonetheless, the topologies and spatial relationships of chromosomes remain poorly understood. Here we developed a method to globally capture intra- and inter-chromosomal interactions, and applied it to generate a map at kilobase resolution of the haploid genome of Saccharomyces cerevisiae. The map recapitulates known features of genome organization, thereby validating the method, and identifies new features. Extensive regional and higher order folding of individual chromosomes is observed. Chromosome XII exhibits a striking conformation that implicates the nucleolus as a formidable barrier to interaction between DNA sequences at either end. Inter-chromosomal contacts are anchored by centromeres and include interactions among transfer RNA genes, among origins of early DNA replication and among sites where chromosomal breakpoints occur. Finally, we constructed a three-dimensional model of the yeast genome. Our findings provide a glimpse of the interface between the form and function of a eukaryotic genome.

  15. Endogenous pararetrovirus sequences associated with 24 nt small RNAs at the centromeres of Fritillaria imperialis L. (Liliaceae), a species with a giant genome

    Czech Academy of Sciences Publication Activity Database

    Becher, H.; Ma, L.; Kelly, L.J.; Kovařík, Aleš; Leitch, I. J.; Leitch, Andrew R.

    2014-01-01

    Roč. 80, č. 5 (2014), s. 823-833 ISSN 0960-7412 R&D Projects: GA ČR(CZ) GA13-10057S Institutional support: RVO:68081707 Keywords : pararetrovirus * Fritillaria imperialis * centromere Subject RIV: BO - Biophysics Impact factor: 5.972, year: 2014

  16. SOLO: a meiotic protein required for centromere cohesion, coorientation, and SMC1 localization in Drosophila melanogaster.

    Science.gov (United States)

    Yan, Rihui; Thomas, Sharon E; Tsai, Jui-He; Yamada, Yukihiro; McKee, Bruce D

    2010-02-08

    Sister chromatid cohesion is essential to maintain stable connections between homologues and sister chromatids during meiosis and to establish correct centromere orientation patterns on the meiosis I and II spindles. However, the meiotic cohesion apparatus in Drosophila melanogaster remains largely uncharacterized. We describe a novel protein, sisters on the loose (SOLO), which is essential for meiotic cohesion in Drosophila. In solo mutants, sister centromeres separate before prometaphase I, disrupting meiosis I centromere orientation and causing nondisjunction of both homologous and sister chromatids. Centromeric foci of the cohesin protein SMC1 are absent in solo mutants at all meiotic stages. SOLO and SMC1 colocalize to meiotic centromeres from early prophase I until anaphase II in wild-type males, but both proteins disappear prematurely at anaphase I in mutants for mei-S332, which encodes the Drosophila homologue of the cohesin protector protein shugoshin. The solo mutant phenotypes and the localization patterns of SOLO and SMC1 indicate that they function together to maintain sister chromatid cohesion in Drosophila meiosis.

  17. Human artificial chromosomes with alpha satellite-based de novo centromeres show increased frequency of nondisjunction and anaphase lag.

    Science.gov (United States)

    Rudd, M Katharine; Mays, Robert W; Schwartz, Stuart; Willard, Huntington F

    2003-11-01

    Human artificial chromosomes have been used to model requirements for human chromosome segregation and to explore the nature of sequences competent for centromere function. Normal human centromeres require specialized chromatin that consists of alpha satellite DNA complexed with epigenetically modified histones and centromere-specific proteins. While several types of alpha satellite DNA have been used to assemble de novo centromeres in artificial chromosome assays, the extent to which they fully recapitulate normal centromere function has not been explored. Here, we have used two kinds of alpha satellite DNA, DXZ1 (from the X chromosome) and D17Z1 (from chromosome 17), to generate human artificial chromosomes. Although artificial chromosomes are mitotically stable over many months in culture, when we examined their segregation in individual cell divisions using an anaphase assay, artificial chromosomes exhibited more segregation errors than natural human chromosomes (P artificial chromosomes missegregate over a fivefold range, the data suggest that variable centromeric DNA content and/or epigenetic assembly can influence the mitotic behavior of artificial chromosomes.

  18. APC/C-Cdc20 mediates deprotection of centromeric cohesin at meiosis II in yeast.

    Science.gov (United States)

    Jonak, Katarzyna; Zagoriy, Ievgeniia; Oz, Tugce; Graf, Peter; Rojas, Julie; Mengoli, Valentina; Zachariae, Wolfgang

    2017-06-18

    Cells undergoing meiosis produce haploid gametes through one round of DNA replication followed by 2 rounds of chromosome segregation. This requires that cohesin complexes, which establish sister chromatid cohesion during S phase, are removed in a stepwise manner. At meiosis I, the separase protease triggers the segregation of homologous chromosomes by cleaving cohesin's Rec8 subunit on chromosome arms. Cohesin persists at centromeres because the PP2A phosphatase, recruited by the shugoshin protein, dephosphorylates Rec8 and thereby protects it from cleavage. While chromatids disjoin upon cleavage of centromeric Rec8 at meiosis II, it was unclear how and when centromeric Rec8 is liberated from its protector PP2A. One proposal is that bipolar spindle forces separate PP2A from Rec8 as cells enter metaphase II. We show here that sister centromere biorientation is not sufficient to "deprotect" Rec8 at meiosis II in yeast. Instead, our data suggest that the ubiquitin-ligase APC/C Cdc20 removes PP2A from centromeres by targeting for degradation the shugoshin Sgo1 and the kinase Mps1. This implies that Rec8 remains protected until entry into anaphase II when it is phosphorylated concurrently with the activation of separase. Here, we provide further support for this model and speculate on its relevance to mammalian oocytes.

  19. Full-length RNA structure prediction of the HIV-1 genome reveals a conserved core domain

    DEFF Research Database (Denmark)

    Sükösd, Zsuzsanna; Andersen, Ebbe Sloth; Seemann, Ernst Stefan

    2015-01-01

    of the HIV-1 genome is highly variable in most regions, with a limited number of stable and conserved RNA secondary structures. Most interesting, a set of long distance interactions form a core organizing structure (COS) that organize the genome into three major structural domains. Despite overlapping...

  20. Esperanto for histones: CENP-A, not CenH3, is the centromeric histone H3 variant.

    Science.gov (United States)

    Earnshaw, W C; Allshire, R C; Black, B E; Bloom, K; Brinkley, B R; Brown, W; Cheeseman, I M; Choo, K H A; Copenhaver, G P; Deluca, J G; Desai, A; Diekmann, S; Erhardt, S; Fitzgerald-Hayes, M; Foltz, D; Fukagawa, T; Gassmann, R; Gerlich, D W; Glover, D M; Gorbsky, G J; Harrison, S C; Heun, P; Hirota, T; Jansen, L E T; Karpen, G; Kops, G J P L; Lampson, M A; Lens, S M; Losada, A; Luger, K; Maiato, H; Maddox, P S; Margolis, R L; Masumoto, H; McAinsh, A D; Mellone, B G; Meraldi, P; Musacchio, A; Oegema, K; O'Neill, R J; Salmon, E D; Scott, K C; Straight, A F; Stukenberg, P T; Sullivan, B A; Sullivan, K F; Sunkel, C E; Swedlow, J R; Walczak, C E; Warburton, P E; Westermann, S; Willard, H F; Wordeman, L; Yanagida, M; Yen, T J; Yoda, K; Cleveland, D W

    2013-04-01

    The first centromeric protein identified in any species was CENP-A, a divergent member of the histone H3 family that was recognised by autoantibodies from patients with scleroderma-spectrum disease. It has recently been suggested to rename this protein CenH3. Here, we argue that the original name should be maintained both because it is the basis of a long established nomenclature for centromere proteins and because it avoids confusion due to the presence of canonical histone H3 at centromeres.

  1. Comparative genomics of Mycoplasma: analysis of conserved essential genes and diversity of the pan-genome.

    Directory of Open Access Journals (Sweden)

    Wei Liu

    Full Text Available Mycoplasma, the smallest self-replicating organism with a minimal metabolism and little genomic redundancy, is expected to be a close approximation to the minimal set of genes needed to sustain bacterial life. This study employs comparative evolutionary analysis of twenty Mycoplasma genomes to gain an improved understanding of essential genes. By analyzing the core genome of mycoplasmas, we finally revealed the conserved essential genes set for mycoplasma survival. Further analysis showed that the core genome set has many characteristics in common with experimentally identified essential genes. Several key genes, which are related to DNA replication and repair and can be disrupted in transposon mutagenesis studies, may be critical for bacteria survival especially over long period natural selection. Phylogenomic reconstructions based on 3,355 homologous groups allowed robust estimation of phylogenetic relatedness among mycoplasma strains. To obtain deeper insight into the relative roles of molecular evolution in pathogen adaptation to their hosts, we also analyzed the positive selection pressures on particular sites and lineages. There appears to be an approximate correlation between the divergence of species and the level of positive selection detected in corresponding lineages.

  2. Identification of conserved regulatory elements by comparative genome analysis

    Directory of Open Access Journals (Sweden)

    Jareborg Niclas

    2003-05-01

    Full Text Available Abstract Background For genes that have been successfully delineated within the human genome sequence, most regulatory sequences remain to be elucidated. The annotation and interpretation process requires additional data resources and significant improvements in computational methods for the detection of regulatory regions. One approach of growing popularity is based on the preferential conservation of functional sequences over the course of evolution by selective pressure, termed 'phylogenetic footprinting'. Mutations are more likely to be disruptive if they appear in functional sites, resulting in a measurable difference in evolution rates between functional and non-functional genomic segments. Results We have devised a flexible suite of methods for the identification and visualization of conserved transcription-factor-binding sites. The system reports those putative transcription-factor-binding sites that are both situated in conserved regions and located as pairs of sites in equivalent positions in alignments between two orthologous sequences. An underlying collection of metazoan transcription-factor-binding profiles was assembled to facilitate the study. This approach results in a significant improvement in the detection of transcription-factor-binding sites because of an increased signal-to-noise ratio, as demonstrated with two sets of promoter sequences. The method is implemented as a graphical web application, ConSite, which is at the disposal of the scientific community at http://www.phylofoot.org/. Conclusions Phylogenetic footprinting dramatically improves the predictive selectivity of bioinformatic approaches to the analysis of promoter sequences. ConSite delivers unparalleled performance using a novel database of high-quality binding models for metazoan transcription factors. With a dynamic interface, this bioinformatics tool provides broad access to promoter analysis with phylogenetic footprinting.

  3. Conserved elements within the genome of foot-and mouth disease virus; their influence on virus replication

    DEFF Research Database (Denmark)

    Kjær, Jonas; Poulsen, Line D.; Vinther, Jeppe

    Objectives: Several conserved elements within the genome of foot-and-mouth disease virus (FMDV) have been identified, e.g. the IRES. Such elements can be crucial for the efficient replication of the genomic RNA. Previously, SHAPE analysis of the entire FMDV genome (Poulsen et al., 2016 submitted......) has identified a conserved RNA structure within the 3Dpol coding region (the RNA-dependent RNA polymerase) which might have an important role in virus replication. The FMDV 2A peptide, another conserved element, is responsible for the primary “cleavage” at its own C-terminus (2A/2B junction......). It is believed that this “cleavage” is achieved by ribosomal skipping, in which the 2A peptide prevents the ribosome from linking the next amino acid (aa) to the growing polypeptide. The nature of this “cleavage” has so far not been investigated in the context of the full-length FMDV RNA within cells. Through...

  4. The Genome of the Basidiomycetous Yeast and Human Pathogen Cryptococcus neoformans

    OpenAIRE

    Loftus, Brendan J.; Fung, Eula; Roncaglia, Paola; Rowley, Don; Amedeo, Paolo; Bruno, Dan; Vamathevan, Jessica; Miranda, Molly; Anderson, Iain J.; Fraser, James A.; Allen, Jonathan E.; Bosdet, Ian E.; Brent, Michael R.; Chiu, Readman; Doering, Tamara L.

    2005-01-01

    Cryptococcus neoformans is a basidiomycetous yeast ubiquitous in the environment, a model for fungal pathogenesis, and an opportunistic human pathogen of global importance. We have sequenced its ~20-megabase genome, which contains ~6500 intron-rich gene structures and encodes a transcriptome abundant in alternatively spliced and antisense messages. The genome is rich in transposons, many of which cluster at candidate centromeric regions. The presence of these transposons may drive karyotype i...

  5. Centromere pairing by a plasmid-encoded type I ParB protein

    DEFF Research Database (Denmark)

    Ringgaard, Simon; Löwe, Jan; Gerdes, Kenn

    2007-01-01

    The par2 locus of Escherichia coli plasmid pB171 encodes two trans-acting proteins, ParA and ParB, and two cis-acting sites, parC1 and parC2, to which ParB binds cooperatively. ParA is related to MinD and oscillates in helical structures and thereby positions ParB/parC-carrying plasmids regularly......, hence identifying the N terminus of ParB as a requirement for ParB-mediated centromere pairing. These observations suggest that centromere pairing is an important intermediate step in plasmid partitioning mediated by the common type I loci....

  6. Identifying all moiety conservation laws in genome-scale metabolic networks.

    Science.gov (United States)

    De Martino, Andrea; De Martino, Daniele; Mulet, Roberto; Pagnani, Andrea

    2014-01-01

    The stoichiometry of a metabolic network gives rise to a set of conservation laws for the aggregate level of specific pools of metabolites, which, on one hand, pose dynamical constraints that cross-link the variations of metabolite concentrations and, on the other, provide key insight into a cell's metabolic production capabilities. When the conserved quantity identifies with a chemical moiety, extracting all such conservation laws from the stoichiometry amounts to finding all non-negative integer solutions of a linear system, a programming problem known to be NP-hard. We present an efficient strategy to compute the complete set of integer conservation laws of a genome-scale stoichiometric matrix, also providing a certificate for correctness and maximality of the solution. Our method is deployed for the analysis of moiety conservation relationships in two large-scale reconstructions of the metabolism of the bacterium E. coli, in six tissue-specific human metabolic networks, and, finally, in the human reactome as a whole, revealing that bacterial metabolism could be evolutionarily designed to cover broader production spectra than human metabolism. Convergence to the full set of moiety conservation laws in each case is achieved in extremely reduced computing times. In addition, we uncover a scaling relation that links the size of the independent pool basis to the number of metabolites, for which we present an analytical explanation.

  7. Identifying all moiety conservation laws in genome-scale metabolic networks.

    Directory of Open Access Journals (Sweden)

    Andrea De Martino

    Full Text Available The stoichiometry of a metabolic network gives rise to a set of conservation laws for the aggregate level of specific pools of metabolites, which, on one hand, pose dynamical constraints that cross-link the variations of metabolite concentrations and, on the other, provide key insight into a cell's metabolic production capabilities. When the conserved quantity identifies with a chemical moiety, extracting all such conservation laws from the stoichiometry amounts to finding all non-negative integer solutions of a linear system, a programming problem known to be NP-hard. We present an efficient strategy to compute the complete set of integer conservation laws of a genome-scale stoichiometric matrix, also providing a certificate for correctness and maximality of the solution. Our method is deployed for the analysis of moiety conservation relationships in two large-scale reconstructions of the metabolism of the bacterium E. coli, in six tissue-specific human metabolic networks, and, finally, in the human reactome as a whole, revealing that bacterial metabolism could be evolutionarily designed to cover broader production spectra than human metabolism. Convergence to the full set of moiety conservation laws in each case is achieved in extremely reduced computing times. In addition, we uncover a scaling relation that links the size of the independent pool basis to the number of metabolites, for which we present an analytical explanation.

  8. Functional Identification of the Plasmodium Centromere and Generation of a Plasmodium Artificial Chromosome

    OpenAIRE

    Iwanaga, Shiroh; Khan, Shahid M.; Kaneko, Izumi; Christodoulou, Zoe; Newbold, Chris; Yuda, Masao; Janse, Chris J.; Waters, Andrew P.

    2010-01-01

    Summary The artificial chromosome represents a useful tool for gene transfer, both as cloning vectors and in chromosome biology research. To generate a Plasmodium artificial chromosome (PAC), we had to first functionally identify and characterize the parasite's centromere. A putative centromere (pbcen5) was cloned from chromosome 5 of the rodent parasite P. berghei based on a Plasmodium gene-synteny map. Plasmids containing pbcen5 were stably maintained in parasites during a blood-stage infec...

  9. Genome sequence analysis of the model grass Brachypodium distachyon: insights into grass genome evolution

    Energy Technology Data Exchange (ETDEWEB)

    Schulman, Al

    2009-08-09

    Three subfamilies of grasses, the Erhardtoideae (rice), the Panicoideae (maize, sorghum, sugar cane and millet), and the Pooideae (wheat, barley and cool season forage grasses) provide the basis of human nutrition and are poised to become major sources of renewable energy. Here we describe the complete genome sequence of the wild grass Brachypodium distachyon (Brachypodium), the first member of the Pooideae subfamily to be completely sequenced. Comparison of the Brachypodium, rice and sorghum genomes reveals a precise sequence- based history of genome evolution across a broad diversity of the grass family and identifies nested insertions of whole chromosomes into centromeric regions as a predominant mechanism driving chromosome evolution in the grasses. The relatively compact genome of Brachypodium is maintained by a balance of retroelement replication and loss. The complete genome sequence of Brachypodium, coupled to its exceptional promise as a model system for grass research, will support the development of new energy and food crops

  10. The First Myriapod Genome Sequence Reveals Conservative Arthropod Gene Content and Genome Organisation in the Centipede Strigamia maritima

    Science.gov (United States)

    Chipman, Ariel D.; Ferrier, David E. K.; Brena, Carlo; Qu, Jiaxin; Hughes, Daniel S. T.; Schröder, Reinhard; Torres-Oliva, Montserrat; Znassi, Nadia; Jiang, Huaiyang; Almeida, Francisca C.; Alonso, Claudio R.; Apostolou, Zivkos; Aqrawi, Peshtewani; Arthur, Wallace; Barna, Jennifer C. J.; Blankenburg, Kerstin P.; Brites, Daniela; Capella-Gutiérrez, Salvador; Coyle, Marcus; Dearden, Peter K.; Du Pasquier, Louis; Duncan, Elizabeth J.; Ebert, Dieter; Eibner, Cornelius; Erikson, Galina; Evans, Peter D.; Extavour, Cassandra G.; Francisco, Liezl; Gabaldón, Toni; Gillis, William J.; Goodwin-Horn, Elizabeth A.; Green, Jack E.; Griffiths-Jones, Sam; Grimmelikhuijzen, Cornelis J. P.; Gubbala, Sai; Guigó, Roderic; Han, Yi; Hauser, Frank; Havlak, Paul; Hayden, Luke; Helbing, Sophie; Holder, Michael; Hui, Jerome H. L.; Hunn, Julia P.; Hunnekuhl, Vera S.; Jackson, LaRonda; Javaid, Mehwish; Jhangiani, Shalini N.; Jiggins, Francis M.; Jones, Tamsin E.; Kaiser, Tobias S.; Kalra, Divya; Kenny, Nathan J.; Korchina, Viktoriya; Kovar, Christie L.; Kraus, F. Bernhard; Lapraz, François; Lee, Sandra L.; Lv, Jie; Mandapat, Christigale; Manning, Gerard; Mariotti, Marco; Mata, Robert; Mathew, Tittu; Neumann, Tobias; Newsham, Irene; Ngo, Dinh N.; Ninova, Maria; Okwuonu, Geoffrey; Ongeri, Fiona; Palmer, William J.; Patil, Shobha; Patraquim, Pedro; Pham, Christopher; Pu, Ling-Ling; Putman, Nicholas H.; Rabouille, Catherine; Ramos, Olivia Mendivil; Rhodes, Adelaide C.; Robertson, Helen E.; Robertson, Hugh M.; Ronshaugen, Matthew; Rozas, Julio; Saada, Nehad; Sánchez-Gracia, Alejandro; Scherer, Steven E.; Schurko, Andrew M.; Siggens, Kenneth W.; Simmons, DeNard; Stief, Anna; Stolle, Eckart; Telford, Maximilian J.; Tessmar-Raible, Kristin; Thornton, Rebecca; van der Zee, Maurijn; von Haeseler, Arndt; Williams, James M.; Willis, Judith H.; Wu, Yuanqing; Zou, Xiaoyan; Lawson, Daniel; Muzny, Donna M.; Worley, Kim C.; Gibbs, Richard A.; Akam, Michael; Richards, Stephen

    2014-01-01

    Myriapods (e.g., centipedes and millipedes) display a simple homonomous body plan relative to other arthropods. All members of the class are terrestrial, but they attained terrestriality independently of insects. Myriapoda is the only arthropod class not represented by a sequenced genome. We present an analysis of the genome of the centipede Strigamia maritima. It retains a compact genome that has undergone less gene loss and shuffling than previously sequenced arthropods, and many orthologues of genes conserved from the bilaterian ancestor that have been lost in insects. Our analysis locates many genes in conserved macro-synteny contexts, and many small-scale examples of gene clustering. We describe several examples where S. maritima shows different solutions from insects to similar problems. The insect olfactory receptor gene family is absent from S. maritima, and olfaction in air is likely effected by expansion of other receptor gene families. For some genes S. maritima has evolved paralogues to generate coding sequence diversity, where insects use alternate splicing. This is most striking for the Dscam gene, which in Drosophila generates more than 100,000 alternate splice forms, but in S. maritima is encoded by over 100 paralogues. We see an intriguing linkage between the absence of any known photosensory proteins in a blind organism and the additional absence of canonical circadian clock genes. The phylogenetic position of myriapods allows us to identify where in arthropod phylogeny several particular molecular mechanisms and traits emerged. For example, we conclude that juvenile hormone signalling evolved with the emergence of the exoskeleton in the arthropods and that RR-1 containing cuticle proteins evolved in the lineage leading to Mandibulata. We also identify when various gene expansions and losses occurred. The genome of S. maritima offers us a unique glimpse into the ancestral arthropod genome, while also displaying many adaptations to its specific

  11. Genomic diversity guides conservation strategies among rare terrestrial orchid species when taxonomy remains uncertain.

    Science.gov (United States)

    Ahrens, Collin W; Supple, Megan A; Aitken, Nicola C; Cantrill, David J; Borevitz, Justin O; James, Elizabeth A

    2017-06-01

    Species are often used as the unit for conservation, but may not be suitable for species complexes where taxa are difficult to distinguish. Under such circumstances, it may be more appropriate to consider species groups or populations as evolutionarily significant units (ESUs). A population genomic approach was employed to investigate the diversity within and among closely related species to create a more robust, lineage-specific conservation strategy for a nationally endangered terrestrial orchid and its relatives from south-eastern Australia. Four putative species were sampled from a total of 16 populations in the Victorian Volcanic Plain (VVP) bioregion and one population of a sub-alpine outgroup in south-eastern Australia. Morphological measurements were taken in situ along with leaf material for genotyping by sequencing (GBS) and microsatellite analyses. Species could not be differentiated using morphological measurements. Microsatellite and GBS markers confirmed the outgroup as distinct, but only GBS markers provided resolution of population genetic structure. The nationally endangered Diuris basaltica was indistinguishable from two related species ( D. chryseopsis and D. behrii ), while the state-protected D. gregaria showed genomic differentiation. Genomic diversity identified among the four Diuris species suggests that conservation of this taxonomically complex group will be best served by considering them as one ESU rather than separately aligned with species as currently recognized. This approach will maximize evolutionary potential among all species during increased isolation and environmental change. The methods used here can be applied generally to conserve evolutionary processes for groups where taxonomic uncertainty hinders the use of species as conservation units. © The Author 2017. Published by Oxford University Press on behalf of the Annals of Botany Company. All rights reserved. For Permissions, please email: journals.permissions@oup.com

  12. A New Approach to Dissect Nuclear Organization: TALE-Mediated Genome Visualization (TGV).

    Science.gov (United States)

    Miyanari, Yusuke

    2016-01-01

    Spatiotemporal organization of chromatin within the nucleus has so far remained elusive. Live visualization of nuclear remodeling could be a promising approach to understand its functional relevance in genome functions and mechanisms regulating genome architecture. Recent technological advances in live imaging of chromosomes begun to explore the biological roles of the movement of the chromatin within the nucleus. Here I describe a new technique, called TALE-mediated genome visualization (TGV), which allows us to visualize endogenous repetitive sequence including centromeric, pericentromeric, and telomeric repeats in living cells.

  13. Consequences for diversity when prioritizing animals for conservation with pedigree or genomic information

    NARCIS (Netherlands)

    Engelsma, K.A.; Veerkamp, R.F.; Calus, M.P.L.; Windig, J.J.

    2011-01-01

    Up to now, prioritization of animals for conservation has been mainly based on pedigree information; however, genomic information may improve prioritization. In this study, we used two Holstein populations to investigate the consequences for genetic diversity when animals are prioritized with

  14. Expressed Centromere Specific Histone 3 (CENH3 Variants in Cultivated Triploid and Wild Diploid Bananas (Musa spp.

    Directory of Open Access Journals (Sweden)

    Kariuki S. Muiruri

    2017-06-01

    Full Text Available Centromeres are specified by a centromere specific histone 3 (CENH3 protein, which exists in a complex environment, interacting with conserved proteins and rapidly evolving satellite DNA sequences. The interactions may become more challenging if multiple CENH3 versions are introduced into the zygote as this can affect post-zygotic mitosis and ultimately sexual reproduction. Here, we characterize CENH3 variant transcripts expressed in cultivated triploid and wild diploid progenitor bananas. We describe both splice- and allelic-[Single Nucleotide Polymorphisms (SNP] variants and their effects on the predicted secondary structures of protein. Expressed CENH3 transcripts from six banana genotypes were characterized and clustered into three groups (MusaCENH-1A, MusaCENH-1B, and MusaCENH-2 based on similarity. The CENH3 groups differed with SNPs as well as presence of indels resulting from retained and/or skipped exons. The CENH3 transcripts from different banana genotypes were spliced in either 7/6, 5/4 or 6/5 exons/introns. The 7/6 and the 5/4 exon/intron structures were found in both diploids and triploids, however, 7/6 was most predominant. The 6/5 exon/introns structure was a result of failure of the 7/6 to splice correctly. The various transcripts obtained were predicted to encode highly variable N-terminal tails and a relatively conserved C-terminal histone fold domain (HFD. The SNPs were predicted in some cases to affect the secondary structure of protein by lengthening or shorting the affected domains. Sequencing of banana CENH3 transcripts predicts SNP variations that affect amino acid sequences and alternatively spliced transcripts. Most of these changes affect the N-terminal tail of CENH3.

  15. Mps1 kinase-dependent Sgo2 centromere localisation mediates cohesin protection in mouse oocyte meiosis I

    NARCIS (Netherlands)

    Yakoubi, W. El; Buffin, E.; Cladiere, D.; Gryaznova, Y.; Berenguer, I.; Touati, S.A.; Gomez, R.; Suja, J.A.; Deursen, J.M.A. van; Wassmann, K.

    2017-01-01

    A key feature of meiosis is the step-wise removal of cohesin, the protein complex holding sister chromatids together, first from arms in meiosis I and then from the centromere region in meiosis II. Centromeric cohesin is protected by Sgo2 from Separase-mediated cleavage, in order to maintain sister

  16. The genome of the polar eukaryotic microalga Coccomyxa subellipsoidea reveals traits of cold adaptation

    Energy Technology Data Exchange (ETDEWEB)

    Blanc, Guillaume; Agarkova, Irina; Grimwood, Jane; Kuo, Alan; Brueggeman, Andrew; Dunigan, David D.; Gurnon, James; Ladunga, Istvan; Lindquist, Erika; Lucas, Susan; Pangilinan, Jasmyn; Proschold, Thomas; Salamov, Asaf; Schmutz, Jeremy; Weeks, Donald; Tamada, Takashi; Lomsadze, Alexandre; Borodovsky, Mark; Claverie, Jean-Michel; Grigoriev, Igor V.; Van Etten, James L.

    2012-02-13

    Background Little is known about the mechanisms of adaptation of life to the extreme environmental conditions encountered in polar regions. Here we present the genome sequence of a unicellular green alga from the division chlorophyta, Coccomyxa subellipsoidea C-169, which we will hereafter refer to as C-169. This is the first eukaryotic microorganism from a polar environment to have its genome sequenced. Results The 48.8 Mb genome contained in 20 chromosomes exhibits significant synteny conservation with the chromosomes of its relatives Chlorella variabilis and Chlamydomonas reinhardtii. The order of the genes is highly reshuffled within synteny blocks, suggesting that intra-chromosomal rearrangements were more prevalent than inter-chromosomal rearrangements. Remarkably, Zepp retrotransposons occur in clusters of nested elements with strictly one cluster per chromosome probably residing at the centromere. Several protein families overrepresented in C. subellipsoidae include proteins involved in lipid metabolism, transporters, cellulose synthases and short alcohol dehydrogenases. Conversely, C-169 lacks proteins that exist in all other sequenced chlorophytes, including components of the glycosyl phosphatidyl inositol anchoring system, pyruvate phosphate dikinase and the photosystem 1 reaction center subunit N (PsaN). Conclusions We suggest that some of these gene losses and gains could have contributed to adaptation to low temperatures. Comparison of these genomic features with the adaptive strategies of psychrophilic microbes suggests that prokaryotes and eukaryotes followed comparable evolutionary routes to adapt to cold environments.

  17. Identification of putative regulatory upstream ORFs in the yeast genome using heuristics and evolutionary conservation

    Directory of Open Access Journals (Sweden)

    Bilsland Elizabeth

    2007-08-01

    Full Text Available Abstract Background The translational efficiency of an mRNA can be modulated by upstream open reading frames (uORFs present in certain genes. A uORF can attenuate translation of the main ORF by interfering with translational reinitiation at the main start codon. uORFs also occur by chance in the genome, in which case they do not have a regulatory role. Since the sequence determinants for functional uORFs are not understood, it is difficult to discriminate functional from spurious uORFs by sequence analysis. Results We have used comparative genomics to identify novel uORFs in yeast with a high likelihood of having a translational regulatory role. We examined uORFs, previously shown to play a role in regulation of translation in Saccharomyces cerevisiae, for evolutionary conservation within seven Saccharomyces species. Inspection of the set of conserved uORFs yielded the following three characteristics useful for discrimination of functional from spurious uORFs: a length between 4 and 6 codons, a distance from the start of the main ORF between 50 and 150 nucleotides, and finally a lack of overlap with, and clear separation from, neighbouring uORFs. These derived rules are inherently associated with uORFs with properties similar to the GCN4 locus, and may not detect most uORFs of other types. uORFs with high scores based on these rules showed a much higher evolutionary conservation than randomly selected uORFs. In a genome-wide scan in S. cerevisiae, we found 34 conserved uORFs from 32 genes that we predict to be functional; subsequent analysis showed the majority of these to be located within transcripts. A total of 252 genes were found containing conserved uORFs with properties indicative of a functional role; all but 7 are novel. Functional content analysis of this set identified an overrepresentation of genes involved in transcriptional control and development. Conclusion Evolutionary conservation of uORFs in yeasts can be traced up to 100

  18. Genome Size Diversity in Lilium (Liliaceae Is Correlated with Karyotype and Environmental Traits

    Directory of Open Access Journals (Sweden)

    Yun-peng Du

    2017-07-01

    Full Text Available Genome size (GS diversity is of fundamental biological importance. The occurrence of giant genomes in angiosperms is restricted to just a few lineages in the analyzed genome size of plant species so far. It is still an open question whether GS diversity is shaped by neutral or natural selection. The genus Lilium, with giant genomes, is phylogenetically and horticulturally important and is distributed throughout the northern hemisphere. GS diversity in Lilium and the underlying evolutionary mechanisms are poorly understood. We performed a comprehensive study involving phylogenetically independent analysis on 71 species to explore the diversity and evolution of GS and its correlation with karyological and environmental traits within Lilium (including Nomocharis. The strong phylogenetic signal detected for GS in the genus provides evidence consistent with that the repetitive DNA may be the primary contributors to the GS diversity, while the significant positive relationships detected between GS and the haploid chromosome length (HCL provide insights into patterns of genome evolution. The relationships between GS and karyotypes indicate that ancestral karyotypes of Lilium are likely to have exhibited small genomes, low diversity in centromeric index (CVCI values and relatively high relative variation in chromosome length (CVCL values. Significant relationships identified between GS and annual temperature and between GS and annual precipitation suggest that adaptation to habitat strongly influences GS diversity. We conclude that GS in Lilium is shaped by both neutral (genetic drift and adaptive evolution. These findings will have important consequences for understanding the evolution of giant plant genomes, and exploring the role of repetitive DNA fraction and chromosome changes in a plant group with large genomes and conservation of chromosome number.

  19. Genomic dissection of conserved transcriptional regulation in intestinal epithelial cells.

    Directory of Open Access Journals (Sweden)

    Colin R Lickwar

    2017-08-01

    Full Text Available The intestinal epithelium serves critical physiologic functions that are shared among all vertebrates. However, it is unknown how the transcriptional regulatory mechanisms underlying these functions have changed over the course of vertebrate evolution. We generated genome-wide mRNA and accessible chromatin data from adult intestinal epithelial cells (IECs in zebrafish, stickleback, mouse, and human species to determine if conserved IEC functions are achieved through common transcriptional regulation. We found evidence for substantial common regulation and conservation of gene expression regionally along the length of the intestine from fish to mammals and identified a core set of genes comprising a vertebrate IEC signature. We also identified transcriptional start sites and other putative regulatory regions that are differentially accessible in IECs in all 4 species. Although these sites rarely showed sequence conservation from fish to mammals, surprisingly, they drove highly conserved IEC expression in a zebrafish reporter assay. Common putative transcription factor binding sites (TFBS found at these sites in multiple species indicate that sequence conservation alone is insufficient to identify much of the functionally conserved IEC regulatory information. Among the rare, highly sequence-conserved, IEC-specific regulatory regions, we discovered an ancient enhancer upstream from her6/HES1 that is active in a distinct population of Notch-positive cells in the intestinal epithelium. Together, these results show how combining accessible chromatin and mRNA datasets with TFBS prediction and in vivo reporter assays can reveal tissue-specific regulatory information conserved across 420 million years of vertebrate evolution. We define an IEC transcriptional regulatory network that is shared between fish and mammals and establish an experimental platform for studying how evolutionarily distilled regulatory information commonly controls IEC development

  20. Comparative Annotation of Viral Genomes with Non-Conserved Gene Structure

    DEFF Research Database (Denmark)

    de Groot, Saskia; Mailund, Thomas; Hein, Jotun

    2007-01-01

    Motivation: Detecting genes in viral genomes is a complex task. Due to the biological necessity of them being constrained in length, RNA viruses in particular tend to code in overlapping reading frames. Since one amino acid is encoded by a triplet of nucleic acids, up to three genes may be coded...... allows for coding in unidirectional nested and overlapping reading frames, to annotate two homologous aligned viral genomes. Our method does not insist on conserved gene structure between the two sequences, thus making it applicable for the pairwise comparison of more distantly related sequences. Results...... and HIV2, as well as of two different Hepatitis Viruses, attaining results of ~87% sensitivity and ~98.5% specificity. We subsequently incorporate prior knowledge by "knowing" the gene structure of one sequence and annotating the other conditional on it. Boosting accuracy close to perfect we demonstrate...

  1. Loss of maternal ATRX results in centromere instability and aneuploidy in the mammalian oocyte and pre-implantation embryo.

    Directory of Open Access Journals (Sweden)

    Claudia Baumann

    2010-09-01

    Full Text Available The α-thalassemia/mental retardation X-linked protein (ATRX is a chromatin-remodeling factor known to regulate DNA methylation at repetitive sequences of the human genome. We have previously demonstrated that ATRX binds to pericentric heterochromatin domains in mouse oocytes at the metaphase II stage where it is involved in mediating chromosome alignment at the meiotic spindle. However, the role of ATRX in the functional differentiation of chromatin structure during meiosis is not known. To test ATRX function in the germ line, we developed an oocyte-specific transgenic RNAi knockdown mouse model. Our results demonstrate that ATRX is required for heterochromatin formation and maintenance of chromosome stability during meiosis. During prophase I arrest, ATRX is necessary to recruit the transcriptional regulator DAXX (death domain associated protein to pericentric heterochromatin. At the metaphase II stage, transgenic ATRX-RNAi oocytes exhibit abnormal chromosome morphology associated with reduced phosphorylation of histone 3 at serine 10 as well as chromosome segregation defects leading to aneuploidy and severely reduced fertility. Notably, a large proportion of ATRX-depleted oocytes and 1-cell stage embryos exhibit chromosome fragments and centromeric DNA-containing micronuclei. Our results provide novel evidence indicating that ATRX is required for centromere stability and the epigenetic control of heterochromatin function during meiosis and the transition to the first mitosis.

  2. Functional conservation of nucleosome formation selectively biases presumably neutral molecular variation in yeast genomes.

    Science.gov (United States)

    Babbitt, Gregory A; Cotter, C R

    2011-01-01

    One prominent pattern of mutational frequency, long appreciated in comparative genomics, is the bias of purine/pyrimidine conserving substitutions (transitions) over purine/pyrimidine altering substitutions (transversions). Traditionally, this transitional bias has been thought to be driven by the underlying rates of DNA mutation and/or repair. However, recent sequencing studies of mutation accumulation lines in model organisms demonstrate that substitutions generally do not accumulate at rates that would indicate a transitional bias. These observations have called into question a very basic assumption of molecular evolution; that naturally occurring patterns of molecular variation in noncoding regions accurately reflect the underlying processes of randomly accumulating neutral mutation in nuclear genomes. Here, in Saccharomyces yeasts, we report a very strong inverse association (r = -0.951, P < 0.004) between the genome-wide frequency of substitutions and their average energetic effect on nucleosome formation, as predicted by a structurally based energy model of DNA deformation around the nucleosome core. We find that transitions occurring at sites positioned nearest the nucleosome surface, which are believed to function most importantly in nucleosome formation, alter the deformation energy of DNA to the nucleosome core by only a fraction of the energy changes typical of most transversions. When we examined the same substitutions set against random background sequences as well as an existing study reporting substitutions arising in mutation accumulation lines of Saccharomyces cerevisiae, we failed to find a similar relationship. These results support the idea that natural selection acting to functionally conserve chromatin organization may contribute significantly to genome-wide transitional bias, even in noncoding regions. Because nucleosome core structure is highly conserved across eukaryotes, our observations may also help to further explain locally elevated

  3. Separase Is Required for Homolog and Sister Disjunction during Drosophila melanogaster Male Meiosis, but Not for Biorientation of Sister Centromeres.

    Science.gov (United States)

    Blattner, Ariane C; Chaurasia, Soumya; McKee, Bruce D; Lehner, Christian F

    2016-04-01

    Spatially controlled release of sister chromatid cohesion during progression through the meiotic divisions is of paramount importance for error-free chromosome segregation during meiosis. Cohesion is mediated by the cohesin protein complex and cleavage of one of its subunits by the endoprotease separase removes cohesin first from chromosome arms during exit from meiosis I and later from the pericentromeric region during exit from meiosis II. At the onset of the meiotic divisions, cohesin has also been proposed to be present within the centromeric region for the unification of sister centromeres into a single functional entity, allowing bipolar orientation of paired homologs within the meiosis I spindle. Separase-mediated removal of centromeric cohesin during exit from meiosis I might explain sister centromere individualization which is essential for subsequent biorientation of sister centromeres during meiosis II. To characterize a potential involvement of separase in sister centromere individualization before meiosis II, we have studied meiosis in Drosophila melanogaster males where homologs are not paired in the canonical manner. Meiosis does not include meiotic recombination and synaptonemal complex formation in these males. Instead, an alternative homolog conjunction system keeps homologous chromosomes in pairs. Using independent strategies for spermatocyte-specific depletion of separase complex subunits in combination with time-lapse imaging, we demonstrate that separase is required for the inactivation of this alternative conjunction at anaphase I onset. Mutations that abolish alternative homolog conjunction therefore result in random segregation of univalents during meiosis I also after separase depletion. Interestingly, these univalents become bioriented during meiosis II, suggesting that sister centromere individualization before meiosis II does not require separase.

  4. Nuclear organization in human sperm: preliminary evidence for altered sex chromosome centromere position in infertile males.

    Science.gov (United States)

    Finch, K A; Fonseka, K G L; Abogrein, A; Ioannou, D; Handyside, A H; Thornhill, A R; Hickson, N; Griffin, D K

    2008-06-01

    Many genetic defects with a chromosomal basis affect male reproduction via a range of different mechanisms. Chromosome position is a well-known marker of nuclear organization, and alterations in standard patterns can lead to disease phenotypes such as cancer, laminopathies and epilepsy. It has been demonstrated that normal mammalian sperm adopt a pattern with the centromeres aligning towards the nuclear centre. The purpose of this study was to test the hypothesis that altered chromosome position in the sperm head is associated with male infertility. The average nuclear positions of fluorescence in-situ hybridization signals for three centromeric probes (for chromosomes X, Y and 18) were compared in normoozoospermic men and in men with compromised semen parameters. In controls, the centromeres of chromosomes X, Y and 18 all occupied a central nuclear location. In infertile men the sex chromosomes appeared more likely to be distributed in a pattern not distinguishable from a random model. Our findings cast doubt on the reliability of centromeric probes for aneuploidy screening. The analysis of chromosome position in sperm heads should be further investigated for the screening of infertile men.

  5. The TubR-centromere complex adopts a double-ring segrosome structure in Type III partition systems.

    Science.gov (United States)

    Martín-García, Bárbara; Martín-González, Alejandro; Carrasco, Carolina; Hernández-Arriaga, Ana M; Ruíz-Quero, Rubén; Díaz-Orejas, Ramón; Aicart-Ramos, Clara; Moreno-Herrero, Fernando; Oliva, María A

    2018-05-14

    In prokaryotes, the centromere is a specialized segment of DNA that promotes the assembly of the segrosome upon binding of the Centromere Binding Protein (CBP). The segrosome structure exposes a specific surface for the interaction of the CBP with the motor protein that mediates DNA movement during cell division. Additionally, the CBP usually controls the transcriptional regulation of the segregation system as a cell cycle checkpoint. Correct segrosome functioning is therefore indispensable for accurate DNA segregation. Here, we combine biochemical reconstruction and structural and biophysical analysis to bring light to the architecture of the segrosome complex in Type III partition systems. We present the particular features of the centromere site, tubC, of the model system encoded in Clostridium botulinum prophage c-st. We find that the split centromere site contains two different iterons involved in the binding and spreading of the CBP, TubR. The resulting nucleoprotein complex consists of a novel double-ring structure that covers part of the predicted promoter. Single molecule data provides a mechanism for the formation of the segrosome structure based on DNA bending and unwinding upon TubR binding.

  6. Comparative genome analysis reveals a conserved family of actin-like proteins in apicomplexan parasites

    Directory of Open Access Journals (Sweden)

    Sibley L David

    2005-12-01

    Full Text Available Abstract Background The phylum Apicomplexa is an early-branching eukaryotic lineage that contains a number of important human and animal pathogens. Their complex life cycles and unique cytoskeletal features distinguish them from other model eukaryotes. Apicomplexans rely on actin-based motility for cell invasion, yet the regulation of this system remains largely unknown. Consequently, we focused our efforts on identifying actin-related proteins in the recently completed genomes of Toxoplasma gondii, Plasmodium spp., Cryptosporidium spp., and Theileria spp. Results Comparative genomic and phylogenetic studies of apicomplexan genomes reveals that most contain only a single conventional actin and yet they each have 8–10 additional actin-related proteins. Among these are a highly conserved Arp1 protein (likely part of a conserved dynactin complex, and Arp4 and Arp6 homologues (subunits of the chromatin-remodeling machinery. In contrast, apicomplexans lack canonical Arp2 or Arp3 proteins, suggesting they lost the Arp2/3 actin polymerization complex on their evolutionary path towards intracellular parasitism. Seven of these actin-like proteins (ALPs are novel to apicomplexans. They show no phylogenetic associations to the known Arp groups and likely serve functions specific to this important group of intracellular parasites. Conclusion The large diversity of actin-like proteins in apicomplexans suggests that the actin protein family has diverged to fulfill various roles in the unique biology of intracellular parasites. Conserved Arps likely participate in vesicular transport and gene expression, while apicomplexan-specific ALPs may control unique biological traits such as actin-based gliding motility.

  7. The most conserved genome segments for life detection on Earth and other planets.

    Science.gov (United States)

    Isenbarger, Thomas A; Carr, Christopher E; Johnson, Sarah Stewart; Finney, Michael; Church, George M; Gilbert, Walter; Zuber, Maria T; Ruvkun, Gary

    2008-12-01

    On Earth, very simple but powerful methods to detect and classify broad taxa of life by the polymerase chain reaction (PCR) are now standard practice. Using DNA primers corresponding to the 16S ribosomal RNA gene, one can survey a sample from any environment for its microbial inhabitants. Due to massive meteoritic exchange between Earth and Mars (as well as other planets), a reasonable case can be made for life on Mars or other planets to be related to life on Earth. In this case, the supremely sensitive technologies used to study life on Earth, including in extreme environments, can be applied to the search for life on other planets. Though the 16S gene has become the standard for life detection on Earth, no genome comparisons have established that the ribosomal genes are, in fact, the most conserved DNA segments across the kingdoms of life. We present here a computational comparison of full genomes from 13 diverse organisms from the Archaea, Bacteria, and Eucarya to identify genetic sequences conserved across the widest divisions of life. Our results identify the 16S and 23S ribosomal RNA genes as well as other universally conserved nucleotide sequences in genes encoding particular classes of transfer RNAs and within the nucleotide binding domains of ABC transporters as the most conserved DNA sequence segments across phylogeny. This set of sequences defines a core set of DNA regions that have changed the least over billions of years of evolution and provides a means to identify and classify divergent life, including ancestrally related life on other planets.

  8. Recombination patterns reveal information about centromere location on linkage maps

    DEFF Research Database (Denmark)

    Limborg, Morten T.; McKinney, Garrett J.; Seeb, Lisa W.

    2016-01-01

    . mykiss) characterized by low and unevenly distributed recombination – a general feature of male meiosis in many species. Further, a high frequency of double crossovers along chromosome arms in barley reduced resolution for locating centromeric regions on most linkage groups. Despite these limitations...

  9. Chromosome segregation regulation in human zygotes : Altered mitotic histone phosphorylation dynamics underlying centromeric targeting of the chromosomal passenger complex

    NARCIS (Netherlands)

    Van De Werken, C.; Avo Santos, M.; Laven, J. S E; Eleveld, C.; Fauser, B. C J M; Lens, S. M A; Baart, E. B.

    2015-01-01

    STUDY QUESTION Are the kinase feedback loops that regulate activation and centromeric targeting of the chromosomal passenger complex (CPC), functional during mitosis in human embryos? SUMMARY ANSWER Investigation of the regulatory kinase pathways involved in centromeric CPC targeting revealed normal

  10. Subchromosomal karyotype evolution in Equidae.

    Science.gov (United States)

    Musilova, P; Kubickova, S; Vahala, J; Rubes, J

    2013-04-01

    Equidae is a small family which comprises horses, African and Asiatic asses, and zebras. Despite equids having diverged quite recently, their karyotypes underwent rapid evolution which resulted in extensive differences among chromosome complements in respective species. Comparative mapping using whole-chromosome painting probes delineated genome-wide chromosome homologies among extant equids, enabling us to trace chromosome rearrangements that occurred during evolution. In the present study, we performed subchromosomal comparative mapping among seven Equidae species, representing the whole family. Region-specific painting and bacterial artificial chromosome probes were used to determine the orientation of evolutionarily conserved segments with respect to centromere positions. This allowed assessment of the configuration of all fusions occurring during the evolution of Equidae, as well as revealing discrepancies in centromere location caused by centromere repositioning or inversions. Our results indicate that the prevailing type of fusion in Equidae is centric fusion. Tandem fusions of the type telomere-telomere occur almost exclusively in the karyotype of Hartmann's zebra and are characteristic of this species' evolution. We revealed inversions in segments homologous to horse chromosomes 3p/10p and 13 in zebras and confirmed inversions in segments 4/31 in African ass, 7 in horse and 8p/20 in zebras. Furthermore, our mapping results suggested that centromere repositioning events occurred in segments homologous to horse chromosomes 7, 8q, 10p and 19 in the African ass and an element homologous to horse chromosome 16 in Asiatic asses. Centromere repositioning in chromosome 1 resulted in three different chromosome types occurring in extant species. Heterozygosity of the centromere position of this chromosome was observed in the kiang. Other subtle changes in centromere position were described in several evolutionary conserved chromosomal segments, suggesting that tiny

  11. Abnormal centromere-chromatid apposition (ACCA) and Peters' anomaly.

    Science.gov (United States)

    Wertelecki, W; Dev, V G; Superneau, D W

    1985-08-01

    Abnormal centromere-chromatid apposition (ACCA) was noted in a patient with Peters' anomaly. Previous reports of ACCA emphasized its association with tetraphocomelia and other congenital malformations (Roberts, SC Phocomelia, Pseudothalidomide Syndromes). This report expands the array of congenital malformations associated with ACCA and emphasizes the diagnostic importance of ocular defects for the ascertainment of additional cases of ACCA and its possible relationship with abnormal cell division.

  12. Regulation of Centromere Localization of the Drosophila Shugoshin MEI-S332 and Sister-Chromatid Cohesion in Meiosis

    Science.gov (United States)

    Nogueira, Cristina; Kashevsky, Helena; Pinto, Belinda; Clarke, Astrid; Orr-Weaver, Terry L.

    2014-01-01

    The Shugoshin (Sgo) protein family helps to ensure proper chromosome segregation by protecting cohesion at the centromere by preventing cleavage of the cohesin complex. Some Sgo proteins also influence other aspects of kinetochore-microtubule attachments. Although many Sgo members require Aurora B kinase to localize to the centromere, factors controlling delocalization are poorly understood and diverse. Moreover, it is not clear how Sgo function is inactivated and whether this is distinct from delocalization. We investigated these questions in Drosophila melanogaster, an organism with superb chromosome cytology to monitor Sgo localization and quantitative assays to test its function in sister-chromatid segregation in meiosis. Previous research showed that in mitosis in cell culture, phosphorylation of the Drosophila Sgo, MEI-S332, by Aurora B promotes centromere localization, whereas Polo phosphorylation promotes delocalization. These studies also suggested that MEI-S332 can be inactivated independently of delocalization, a conclusion supported here by localization and function studies in meiosis. Phosphoresistant and phosphomimetic mutants for the Aurora B and Polo phosphorylation sites were examined for effects on MEI-S332 localization and chromosome segregation in meiosis. Strikingly, MEI-S332 with a phosphomimetic mutation in the Aurora B phosphorylation site prematurely dissociates from the centromeres in meiosis I. Despite the absence of MEI-S332 on meiosis II centromeres in male meiosis, sister chromatids segregate normally, demonstrating that detectable levels of this Sgo are not essential for chromosome congression, kinetochore biorientation, or spindle assembly. PMID:25081981

  13. The Genome of the Basidiomycetous Yeast and Human Pathogen Cryptococcus neoformans

    Science.gov (United States)

    Loftus, Brendan J.; Fung, Eula; Roncaglia, Paola; Rowley, Don; Amedeo, Paolo; Bruno, Dan; Vamathevan, Jessica; Miranda, Molly; Anderson, Iain J.; Fraser, James A.; Allen, Jonathan E.; Bosdet, Ian E.; Brent, Michael R.; Chiu, Readman; Doering, Tamara L.; Donlin, Maureen J.; D’Souza, Cletus A.; Fox, Deborah S.; Grinberg, Viktoriya; Fu, Jianmin; Fukushima, Marilyn; Haas, Brian J.; Huang, James C.; Janbon, Guilhem; Jones, Steven J. M.; Koo, Hean L.; Krzywinski, Martin I.; Kwon-Chung, June K.; Lengeler, Klaus B.; Maiti, Rama; Marra, Marco A.; Marra, Robert E.; Mathewson, Carrie A.; Mitchell, Thomas G.; Pertea, Mihaela; Riggs, Florenta R.; Salzberg, Steven L.; Schein, Jacqueline E.; Shvartsbeyn, Alla; Shin, Heesun; Shumway, Martin; Specht, Charles A.; Suh, Bernard B.; Tenney, Aaron; Utterback, Terry R.; Wickes, Brian L.; Wortman, Jennifer R.; Wye, Natasja H.; Kronstad, James W.; Lodge, Jennifer K.; Heitman, Joseph; Davis, Ronald W.; Fraser, Claire M.; Hyman, Richard W.

    2012-01-01

    Cryptococcus neoformans is a basidiomycetous yeast ubiquitous in the environment, a model for fungal pathogenesis, and an opportunistic human pathogen of global importance. We have sequenced its ~20-megabase genome, which contains ~6500 intron-rich gene structures and encodes a transcriptome abundant in alternatively spliced and antisense messages. The genome is rich in transposons, many of which cluster at candidate centromeric regions. The presence of these transposons may drive karyotype instability and phenotypic variation. C. neoformans encodes unique genes that may contribute to its unusual virulence properties, and comparison of two phenotypically distinct strains reveals variation in gene content in addition to sequence polymorphisms between the genomes. PMID:15653466

  14. Spatial organization of the budding yeast genome in the cell nucleus and identification of specific chromatin interactions from multi-chromosome constrained chromatin model.

    Science.gov (United States)

    Gürsoy, Gamze; Xu, Yun; Liang, Jie

    2017-07-01

    Nuclear landmarks and biochemical factors play important roles in the organization of the yeast genome. The interaction pattern of budding yeast as measured from genome-wide 3C studies are largely recapitulated by model polymer genomes subject to landmark constraints. However, the origin of inter-chromosomal interactions, specific roles of individual landmarks, and the roles of biochemical factors in yeast genome organization remain unclear. Here we describe a multi-chromosome constrained self-avoiding chromatin model (mC-SAC) to gain understanding of the budding yeast genome organization. With significantly improved sampling of genome structures, both intra- and inter-chromosomal interaction patterns from genome-wide 3C studies are accurately captured in our model at higher resolution than previous studies. We show that nuclear confinement is a key determinant of the intra-chromosomal interactions, and centromere tethering is responsible for the inter-chromosomal interactions. In addition, important genomic elements such as fragile sites and tRNA genes are found to be clustered spatially, largely due to centromere tethering. We uncovered previously unknown interactions that were not captured by genome-wide 3C studies, which are found to be enriched with tRNA genes, RNAPIII and TFIIS binding. Moreover, we identified specific high-frequency genome-wide 3C interactions that are unaccounted for by polymer effects under landmark constraints. These interactions are enriched with important genes and likely play biological roles.

  15. Molecular analysis and genomic organization of major DNA satellites in banana (Musa spp.).

    Science.gov (United States)

    Čížková, Jana; Hřibová, Eva; Humplíková, Lenka; Christelová, Pavla; Suchánková, Pavla; Doležel, Jaroslav

    2013-01-01

    Satellite DNA sequences consist of tandemly arranged repetitive units up to thousands nucleotides long in head-to-tail orientation. The evolutionary processes by which satellites arise and evolve include unequal crossing over, gene conversion, transposition and extra chromosomal circular DNA formation. Large blocks of satellite DNA are often observed in heterochromatic regions of chromosomes and are a typical component of centromeric and telomeric regions. Satellite-rich loci may show specific banding patterns and facilitate chromosome identification and analysis of structural chromosome changes. Unlike many other genomes, nuclear genomes of banana (Musa spp.) are poor in satellite DNA and the information on this class of DNA remains limited. The banana cultivars are seed sterile clones originating mostly from natural intra-specific crosses within M. acuminata (A genome) and inter-specific crosses between M. acuminata and M. balbisiana (B genome). Previous studies revealed the closely related nature of the A and B genomes, including similarities in repetitive DNA. In this study we focused on two main banana DNA satellites, which were previously identified in silico. Their genomic organization and molecular diversity was analyzed in a set of nineteen Musa accessions, including representatives of A, B and S (M. schizocarpa) genomes and their inter-specific hybrids. The two DNA satellites showed a high level of sequence conservation within, and a high homology between Musa species. FISH with probes for the satellite DNA sequences, rRNA genes and a single-copy BAC clone 2G17 resulted in characteristic chromosome banding patterns in M. acuminata and M. balbisiana which may aid in determining genomic constitution in interspecific hybrids. In addition to improving the knowledge on Musa satellite DNA, our study increases the number of cytogenetic markers and the number of individual chromosomes, which can be identified in Musa.

  16. Mapping cis-Regulatory Domains in the Human Genome UsingMulti-Species Conservation of Synteny

    Energy Technology Data Exchange (ETDEWEB)

    Ahituv, Nadav; Prabhakar, Shyam; Poulin, Francis; Rubin, EdwardM.; Couronne, Olivier

    2005-06-13

    Our inability to associate distant regulatory elements with the genes that they regulate has largely precluded their examination for sequence alterations contributing to human disease. One major obstacle is the large genomic space surrounding targeted genes in which such elements could potentially reside. In order to delineate gene regulatory boundaries we used whole-genome human-mouse-chicken (HMC) and human-mouse-frog (HMF) multiple alignments to compile conserved blocks of synteny (CBS), under the hypothesis that these blocks have been kept intact throughout evolution at least in part by the requirement of regulatory elements to stay linked to the genes that they regulate. A total of 2,116 and 1,942 CBS>200 kb were assembled for HMC and HMF respectively, encompassing 1.53 and 0.86 Gb of human sequence. To support the existence of complex long-range regulatory domains within these CBS we analyzed the prevalence and distribution of chromosomal aberrations leading to position effects (disruption of a genes regulatory environment), observing a clear bias not only for mapping onto CBS but also for longer CBS size. Our results provide a genome wide data set characterizing the regulatory domains of genes and the conserved regulatory elements within them.

  17. Consequences for diversity when animals are prioritized for conservation of the whole genome or of one specific allele

    NARCIS (Netherlands)

    Engelsma, K.A.; Veerkamp, R.F.; Calus, M.P.L.; Windig, J.J.

    2014-01-01

    When animals are selected for one specific allele, for example for inclusion in a gene bank, this may result in the loss of diversity in other parts of the genome. The aim of this study was to quantify the risk of losing diversity across the genome when targeting a single allele for conservation

  18. Interstitial telomere-like repeats in the Arabidopsis thaliana genome.

    Science.gov (United States)

    Uchida, Wakana; Matsunaga, Sachihiro; Sugiyama, Ryuji; Kawano, Shigeyuki

    2002-02-01

    Eukaryotic chromosomal ends are protected by telomeres, which are thought to play an important role in ensuring the complete replication of chromosomes. On the other hand, non-functional telomere-like repeats in the interchromosomal regions (interstitial telomeric repeats; ITRs) have been reported in several eukaryotes. In this study, we identified eight ITRs in the Arabidopsis thaliana genome, each consisting of complete and degenerate 300- to 1200-bp sequences. The ITRs were grouped into three classes (class IA-B, class II, and class IIIA-E) based on the degeneracy of the telomeric repeats in ITRs. The telomeric repeats of the two ITRs in class I were conserved for the most part, whereas the single ITR in class II, and the five ITRs in class III were relatively degenerated. In addition, degenerate ITRs were surrounded by common sequences that shared 70-100% homology to each other; these are named ITR-adjacent sequences (IAS). Although the genomic regions around ITRs in class I lacked IAS, those around ITRs in class II contained IAS (IASa), and those around five ITRs in class III had nine types of IAS (IASb, c, d, e, f, g, h, i, and j). Ten IAS types in classes II and III showed no significant homology to each other. The chromosomal locations of ITRs and IAS were not category-related, but most of them were adjacent to, or part of, a centromere. These results show that the A. thaliana genome has undergone chromosomal rearrangements, such as end-fusions and segmental duplications.

  19. St2-80: a new FISH marker for St genome and genome analysis in Triticeae.

    Science.gov (United States)

    Wang, Long; Shi, Qinghua; Su, Handong; Wang, Yi; Sha, Lina; Fan, Xing; Kang, Houyang; Zhang, Haiqin; Zhou, Yonghong

    2017-07-01

    The St genome is one of the most fundamental genomes in Triticeae. Repetitive sequences are widely used to distinguish different genomes or species. The primary objectives of this study were to (i) screen a new sequence that could easily distinguish the chromosome of the St genome from those of other genomes by fluorescence in situ hybridization (FISH) and (ii) investigate the genome constitution of some species that remain uncertain and controversial. We used degenerated oligonucleotide primer PCR (Dop-PCR), Dot-blot, and FISH to screen for a new marker of the St genome and to test the efficiency of this marker in the detection of the St chromosome at different ploidy levels. Signals produced by a new FISH marker (denoted St 2 -80) were present on the entire arm of chromosomes of the St genome, except in the centromeric region. On the contrary, St 2 -80 signals were present in the terminal region of chromosomes of the E, H, P, and Y genomes. No signal was detected in the A and B genomes, and only weak signals were detected in the terminal region of chromosomes of the D genome. St 2 -80 signals were obvious and stable in chromosomes of different genomes, whether diploid or polyploid. Therefore, St 2 -80 is a potential and useful FISH marker that can be used to distinguish the St genome from those of other genomes in Triticeae.

  20. Clinical spectrum of immunodeficiency, centromeric instability and facial dysmorphism (ICF syndrome).

    NARCIS (Netherlands)

    Hagleitner, M.M.; Lankester, A.; Maraschio, P.; Hulten, M.; Fryns, J.P.; Schuetz, C.; Gimelli, G.; Davies, E.G.; Gennery, A.R.; Belohradsky, B.H.; Groot, R. de; Gerritsen, E.J.; Mattina, T.; Howard, P.J.; Fasth, A.; Reisli, I.; Furthner, D.; Slatter, M.A.; Cant, A.J.; Cazzola, G.; Dijken, P.J. van; Deuren, M. van; Greef, J.C. de; Maarel, S.M. van der; Weemaes, C.M.R.

    2008-01-01

    BACKGROUND: Immunodeficiency, centromeric instability and facial dysmorphism (ICF syndrome) is a rare autosomal recessive disease characterised by facial dysmorphism, immunoglobulin deficiency and branching of chromosomes 1, 9 and 16 after PHA stimulation of lymphocytes. Hypomethylation of DNA of a

  1. Conservation priorities for endangered Indian tigers through a genomic lens.

    Science.gov (United States)

    Natesh, Meghana; Atla, Goutham; Nigam, Parag; Jhala, Yadvendradev V; Zachariah, Arun; Borthakur, Udayan; Ramakrishnan, Uma

    2017-08-29

    Tigers have lost 93% of their historical range worldwide. India plays a vital role in the conservation of tigers since nearly 60% of all wild tigers are currently found here. However, as protected areas are small (<300 km 2 on average), with only a few individuals in each, many of them may not be independently viable. It is thus important to identify and conserve genetically connected populations, as well as to maintain connectivity within them. We collected samples from wild tigers (Panthera tigris tigris) across India and used genome-wide SNPs to infer genetic connectivity. We genotyped 10,184 SNPs from 38 individuals across 17 protected areas and identified three genetically distinct clusters (corresponding to northwest, southern and central India). The northwest cluster was isolated with low variation and high relatedness. The geographically large central cluster included tigers from central, northeastern and northern India, and had the highest variation. Most genetic diversity (62%) was shared among clusters, while unique variation was highest in the central cluster (8.5%) and lowest in the northwestern one (2%). We did not detect signatures of differential selection or local adaptation. We highlight that the northwest population requires conservation attention to ensure persistence of these tigers.

  2. A Solution to the C-Value Paradox and the Function of Junk DNA: The Genome Balance Hypothesis.

    Science.gov (United States)

    Freeling, Michael; Xu, Jie; Woodhouse, Margaret; Lisch, Damon

    2015-06-01

    The Genome Balance Hypothesis originated from a recent study that provided a mechanism for the phenomenon of genome dominance in ancient polyploids: unique 24nt RNA coverage near genes is greater in genes on the recessive subgenome irrespective of differences in gene expression. 24nt RNAs target transposons. Transposon position effects are now hypothesized to balance the expression of networked genes and provide spring-like tension between pericentromeric heterochromatin and microtubules. The balance (coordination) of gene expression and centromere movement is under selection. Our hypothesis states that this balance can be maintained by many or few transposons about equally well. We explain known balanced distributions of junk DNA within genomes and between subgenomes in allopolyploids (and our hypothesis passes "the onion test" for any so-called solution to the C-value paradox). Importantly, when the allotetraploid maize chromosomes delete redundant genes, their nearby transposons are also lost; this result is explained if transposons near genes function. The Genome Balance Hypothesis is hypothetical because the position effect mechanisms implicated are not proved to apply to all junk DNA, and the continuous nature of the centromeric and gene position effects have not yet been studied as a single phenomenon. Copyright © 2015 The Author. Published by Elsevier Inc. All rights reserved.

  3. Complete genome sequence and comparative genomic analysis of Mycobacterium massiliense JCM 15300 in the Mycobacterium abscessus group reveal a conserved genomic island MmGI-1 related to putative lipid metabolism.

    Directory of Open Access Journals (Sweden)

    Tsuyoshi Sekizuka

    Full Text Available Mycobacterium abscessus group subsp., such as M. massiliense, M. abscessus sensu stricto and M. bolletii, are an environmental organism found in soil, water and other ecological niches, and have been isolated from respiratory tract infection, skin and soft tissue infection, postoperative infection of cosmetic surgery. To determine the unique genetic feature of M. massiliense, we sequenced the complete genome of M. massiliense type strain JCM 15300 (corresponding to CCUG 48898. Comparative genomic analysis was performed among Mycobacterium spp. and among M. abscessus group subspp., showing that additional ß-oxidation-related genes and, notably, the mammalian cell entry (mce operon were located on a genomic island, M. massiliense Genomic Island 1 (MmGI-1, in M. massiliense. In addition, putative anaerobic respiration system-related genes and additional mycolic acid cyclopropane synthetase-related genes were found uniquely in M. massiliense. Japanese isolates of M. massiliense also frequently possess the MmGI-1 (14/44, approximately 32% and three unique conserved regions (26/44; approximately 60%, 34/44; approximately 77% and 40/44; approximately 91%, as well as isolates of other countries (Malaysia, France, United Kingdom and United States. The well-conserved genomic island MmGI-1 may play an important role in high growth potential with additional lipid metabolism, extra factors for survival in the environment or synthesis of complex membrane-associated lipids. ORFs on MmGI-1 showed similarities to ORFs of phylogenetically distant M. avium complex (MAC, suggesting that horizontal gene transfer or genetic recombination events might have occurred within MmGI-1 among M. massiliense and MAC.

  4. The Cell Cycle Timing of Centromeric Chromatin Assembly in Drosophila Meiosis Is Distinct from Mitosis Yet Requires CAL1 and CENP-C

    Science.gov (United States)

    Gorgescu, Walter; Tang, Jonathan; Costes, Sylvain V.; Karpen, Gary H.

    2012-01-01

    CENP-A (CID in flies) is the histone H3 variant essential for centromere specification, kinetochore formation, and chromosome segregation during cell division. Recent studies have elucidated major cell cycle mechanisms and factors critical for CENP-A incorporation in mitosis, predominantly in cultured cells. However, we do not understand the roles, regulation, and cell cycle timing of CENP-A assembly in somatic tissues in multicellular organisms and in meiosis, the specialized cell division cycle that gives rise to haploid gametes. Here we investigate the timing and requirements for CID assembly in mitotic tissues and male and female meiosis in Drosophila melanogaster, using fixed and live imaging combined with genetic approaches. We find that CID assembly initiates at late telophase and continues during G1 phase in somatic tissues in the organism, later than the metaphase assembly observed in cultured cells. Furthermore, CID assembly occurs at two distinct cell cycle phases during male meiosis: prophase of meiosis I and after exit from meiosis II, in spermatids. CID assembly in prophase I is also conserved in female meiosis. Interestingly, we observe a novel decrease in CID levels after the end of meiosis I and before meiosis II, which correlates temporally with changes in kinetochore organization and orientation. We also demonstrate that CID is retained on mature sperm despite the gross chromatin remodeling that occurs during protamine exchange. Finally, we show that the centromere proteins CAL1 and CENP-C are both required for CID assembly in meiosis and normal progression through spermatogenesis. We conclude that the cell cycle timing of CID assembly in meiosis is different from mitosis and that the efficient propagation of CID through meiotic divisions and on sperm is likely to be important for centromere specification in the developing zygote. PMID:23300382

  5. Analysis of Primary Structural Determinants That Distinguish the Centromere-Specific Function of Histone Variant Cse4p from Histone H3

    OpenAIRE

    Keith, Kevin C.; Baker, Richard E.; Chen, Yinhuai; Harris, Kendra; Stoler, Sam; Fitzgerald-Hayes, Molly

    1999-01-01

    Cse4p is a variant of histone H3 that has an essential role in chromosome segregation and centromere chromatin structure in budding yeast. Cse4p has a unique 135-amino-acid N terminus and a C-terminal histone-fold domain that is more than 60% identical to histone H3 and the mammalian centromere protein CENP-A. Cse4p and CENP-A have biochemical properties similar to H3 and probably replace H3 in centromere-specific nucleosomes in yeasts and mammals, respectively. In order to identify regions o...

  6. B chromosomes are more frequent in mammals with acrocentric karyotypes: support for the theory of centromeric drive.

    Science.gov (United States)

    Palestis, Brian G; Burt, Austin; Jones, R Neil; Trivers, Robert

    2004-02-07

    The chromosomes of mammals tend to be either mostly acrocentric (having one long arm) or mostly bi-armed, with few species having intermediate karyotypes. The theory of centromeric drive suggests that this observation reflects a bias during female meiosis, favouring either more centromeres or fewer, and that the direction of this bias changes frequently over evolutionary time. B chromosomes are selfish genetic elements found in some individuals within some species. B chromosomes are often harmful, but persist because they drive (i.e. they are transmitted more frequently than expected). We predicted that species with mainly acrocentric chromosomes would be more likely to harbour B chromosomes than those with mainly bi-armed chromosomes, because female meiosis would favour more centromeres over fewer in species with one-armed chromosomes. Our results show that B chromosomes are indeed more common in species with acrocentric chromosomes, across all mammals, among rodents, among non-rodents and in a test of independent taxonomic contrasts. These results provide independent evidence supporting the theory of centromeric drive and also help to explain the distribution of selfish DNA across species. In addition, we demonstrate an association between the shape of the B chromosomes and the shape of the typical ('A') chromosomes.

  7. The cohesion protein SOLO associates with SMC1 and is required for synapsis, recombination, homolog bias and cohesion and pairing of centromeres in Drosophila Meiosis.

    Science.gov (United States)

    Yan, Rihui; McKee, Bruce D

    2013-01-01

    Cohesion between sister chromatids is mediated by cohesin and is essential for proper meiotic segregation of both sister chromatids and homologs. solo encodes a Drosophila meiosis-specific cohesion protein with no apparent sequence homology to cohesins that is required in male meiosis for centromere cohesion, proper orientation of sister centromeres and centromere enrichment of the cohesin subunit SMC1. In this study, we show that solo is involved in multiple aspects of meiosis in female Drosophila. Null mutations in solo caused the following phenotypes: 1) high frequencies of homolog and sister chromatid nondisjunction (NDJ) and sharply reduced frequencies of homolog exchange; 2) reduced transmission of a ring-X chromosome, an indicator of elevated frequencies of sister chromatid exchange (SCE); 3) premature loss of centromere pairing and cohesion during prophase I, as indicated by elevated foci counts of the centromere protein CID; 4) instability of the lateral elements (LE)s and central regions of synaptonemal complexes (SCs), as indicated by fragmented and spotty staining of the chromosome core/LE component SMC1 and the transverse filament protein C(3)G, respectively, at all stages of pachytene. SOLO and SMC1 are both enriched on centromeres throughout prophase I, co-align along the lateral elements of SCs and reciprocally co-immunoprecipitate from ovarian protein extracts. Our studies demonstrate that SOLO is closely associated with meiotic cohesin and required both for enrichment of cohesin on centromeres and stable assembly of cohesin into chromosome cores. These events underlie and are required for stable cohesion of centromeres, synapsis of homologous chromosomes, and a recombination mechanism that suppresses SCE to preferentially generate homolog crossovers (homolog bias). We propose that SOLO is a subunit of a specialized meiotic cohesin complex that mediates both centromeric and axial arm cohesion and promotes homolog bias as a component of chromosome

  8. Live visualization of genomic loci with BiFC-TALE.

    Science.gov (United States)

    Hu, Huan; Zhang, Hongmin; Wang, Sheng; Ding, Miao; An, Hui; Hou, Yingping; Yang, Xiaojing; Wei, Wensheng; Sun, Yujie; Tang, Chao

    2017-01-11

    Tracking the dynamics of genomic loci is important for understanding the mechanisms of fundamental intracellular processes. However, fluorescent labeling and imaging of such loci in live cells have been challenging. One of the major reasons is the low signal-to-background ratio (SBR) of images mainly caused by the background fluorescence from diffuse full-length fluorescent proteins (FPs) in the living nucleus, hampering the application of live cell genomic labeling methods. Here, combining bimolecular fluorescence complementation (BiFC) and transcription activator-like effector (TALE) technologies, we developed a novel method for labeling genomic loci (BiFC-TALE), which largely reduces the background fluorescence level. Using BiFC-TALE, we demonstrated a significantly improved SBR by imaging telomeres and centromeres in living cells in comparison with the methods using full-length FP.

  9. Comparative genomics reveals conservative evolution of the xylem transcriptome in vascular plants.

    Science.gov (United States)

    Li, Xinguo; Wu, Harry X; Southerton, Simon G

    2010-06-21

    Wood is a valuable natural resource and a major carbon sink. Wood formation is an important developmental process in vascular plants which played a crucial role in plant evolution. Although genes involved in xylem formation have been investigated, the molecular mechanisms of xylem evolution are not well understood. We use comparative genomics to examine evolution of the xylem transcriptome to gain insights into xylem evolution. The xylem transcriptome is highly conserved in conifers, but considerably divergent in angiosperms. The functional domains of genes in the xylem transcriptome are moderately to highly conserved in vascular plants, suggesting the existence of a common ancestral xylem transcriptome. Compared to the total transcriptome derived from a range of tissues, the xylem transcriptome is relatively conserved in vascular plants. Of the xylem transcriptome, cell wall genes, ancestral xylem genes, known proteins and transcription factors are relatively more conserved in vascular plants. A total of 527 putative xylem orthologs were identified, which are unevenly distributed across the Arabidopsis chromosomes with eight hot spots observed. Phylogenetic analysis revealed that evolution of the xylem transcriptome has paralleled plant evolution. We also identified 274 conifer-specific xylem unigenes, all of which are of unknown function. These xylem orthologs and conifer-specific unigenes are likely to have played a crucial role in xylem evolution. Conifers have highly conserved xylem transcriptomes, while angiosperm xylem transcriptomes are relatively diversified. Vascular plants share a common ancestral xylem transcriptome. The xylem transcriptomes of vascular plants are more conserved than the total transcriptomes. Evolution of the xylem transcriptome has largely followed the trend of plant evolution.

  10. Deleterious mutation accumulation in organelle genomes.

    Science.gov (United States)

    Lynch, M; Blanchard, J L

    1998-01-01

    It is well established on theoretical grounds that the accumulation of mildly deleterious mutations in nonrecombining genomes is a major extinction risk in obligately asexual populations. Sexual populations can also incur mutational deterioration in genomic regions that experience little or no recombination, i.e., autosomal regions near centromeres, Y chromosomes, and organelle genomes. Our results suggest, for a wide array of genes (transfer RNAs, ribosomal RNAs, and proteins) in a diverse collection of species (animals, plants, and fungi), an almost universal increase in the fixation probabilities of mildly deleterious mutations arising in mitochondrial and chloroplast genomes relative to those arising in the recombining nuclear genome. This enhanced width of the selective sieve in organelle genomes does not appear to be a consequence of relaxed selection, but can be explained by the decline in the efficiency of selection that results from the reduction of effective population size induced by uniparental inheritance. Because of the very low mutation rates of organelle genomes (on the order of 10(-4) per genome per year), the reduction in fitness resulting from mutation accumulation in such genomes is a very long-term process, not likely to imperil many species on time scales of less than a million years, but perhaps playing some role in phylogenetic lineage sorting on time scales of 10 to 100 million years.

  11. Decoding the principles underlying the frequency of association with nucleoli for RNA polymerase III–transcribed genes in budding yeast

    Science.gov (United States)

    Belagal, Praveen; Normand, Christophe; Shukla, Ashutosh; Wang, Renjie; Léger-Silvestre, Isabelle; Dez, Christophe; Bhargava, Purnima; Gadal, Olivier

    2016-01-01

    The association of RNA polymerase III (Pol III)–transcribed genes with nucleoli seems to be an evolutionarily conserved property of the spatial organization of eukaryotic genomes. However, recent studies of global chromosome architecture in budding yeast have challenged this view. We used live-cell imaging to determine the intranuclear positions of 13 Pol III–transcribed genes. The frequency of association with nucleolus and nuclear periphery depends on linear genomic distance from the tethering elements—centromeres or telomeres. Releasing the hold of the tethering elements by inactivating centromere attachment to the spindle pole body or changing the position of ribosomal DNA arrays resulted in the association of Pol III–transcribed genes with nucleoli. Conversely, ectopic insertion of a Pol III–transcribed gene in the vicinity of a centromere prevented its association with nucleolus. Pol III–dependent transcription was independent of the intranuclear position of the gene, but the nucleolar recruitment of Pol III–transcribed genes required active transcription. We conclude that the association of Pol III–transcribed genes with the nucleolus, when permitted by global chromosome architecture, provides nucleolar and/or nuclear peripheral anchoring points contributing locally to intranuclear chromosome organization. PMID:27559135

  12. Chromosome segregation regulation in human zygotes: altered mitotic histone phosphorylation dynamics underlying centromeric targeting of the chromosomal passenger complex.

    Science.gov (United States)

    van de Werken, C; Avo Santos, M; Laven, J S E; Eleveld, C; Fauser, B C J M; Lens, S M A; Baart, E B

    2015-10-01

    Are the kinase feedback loops that regulate activation and centromeric targeting of the chromosomal passenger complex (CPC), functional during mitosis in human embryos? Investigation of the regulatory kinase pathways involved in centromeric CPC targeting revealed normal phosphorylation dynamics of histone H2A at T120 (H2ApT120) by Bub1 kinase and subsequent recruitment of Shugoshin, but phosphorylation of histone H3 at threonine 3 (H3pT3) by Haspin failed to show the expected centromeric enrichment on metaphase chromosomes in the zygote. Human cleavage stage embryos show high levels of chromosomal instability. What causes this high error rate is unknown, as mechanisms used to ensure proper chromosome segregation in mammalian embryos are poorly described. In this study, we investigated the pathways regulating CPC targeting to the inner centromere in human embryos. We characterized the distribution of the CPC in relation to activity of its two main centromeric targeting pathways: the Bub1-H2ApT120-Sgo-CPC and Haspin-H3pT3-CPC pathways. The study was conducted between May 2012 and March 2014 on human surplus embryos resulting from in vitro fertilization treatment and donated for research. In zygotes, nuclear envelope breakdown was monitored by time-lapse imaging to allow timed incubations with specific inhibitors to arrest at prometaphase and metaphase, and to interfere with Haspin and Aurora B/C kinase activity. Functionality of the targeting pathways was assessed through characterization of histone phosphorylation dynamics by immunofluorescent analysis, combined with gene expression by RT-qPCR and immunofluorescent localization of key pathway proteins. Immunofluorescent analysis of the CPC subunit Inner Centromere Protein revealed the pool of stably bound CPC proteins was not strictly confined to the inner centromere of prometaphase chromosomes in human zygotes, as observed in later stages of preimplantation development and somatic cells. Investigation of the

  13. Mind the gap; seven reasons to close fragmented genome assemblies.

    Science.gov (United States)

    Thomma, Bart P H J; Seidl, Michael F; Shi-Kunne, Xiaoqian; Cook, David E; Bolton, Melvin D; van Kan, Jan A L; Faino, Luigi

    2016-05-01

    Like other domains of life, research into the biology of filamentous microbes has greatly benefited from the advent of whole-genome sequencing. Next-generation sequencing (NGS) technologies have revolutionized sequencing, making genomic sciences accessible to many academic laboratories including those that study non-model organisms. Thus, hundreds of fungal genomes have been sequenced and are publically available today, although these initiatives have typically yielded considerably fragmented genome assemblies that often lack large contiguous genomic regions. Many important genomic features are contained in intergenic DNA that is often missing in current genome assemblies, and recent studies underscore the significance of non-coding regions and repetitive elements for the life style, adaptability and evolution of many organisms. The study of particular types of genetic elements, such as telomeres, centromeres, repetitive elements, effectors, and clusters of co-regulated genes, but also of phenomena such as structural rearrangements, genome compartmentalization and epigenetics, greatly benefits from having a contiguous and high-quality, preferably even complete and gapless, genome assembly. Here we discuss a number of important reasons to produce gapless, finished, genome assemblies to help answer important biological questions. Copyright © 2015 Elsevier Inc. All rights reserved.

  14. Conserved elements within the genome of foot-and-mouth disease virus; their influence on viral replication

    DEFF Research Database (Denmark)

    Kjær, Jonas

    -and-mouth disease virus (FMDV) have been identified, e.g. the IRES. Such elements can be crucial for the efficient replication of the genomic RNA. A better understanding of the influence of these elements is required to identify currently unrecognized interactions within the viruses which may be important...... for the development of anti-viral agents. SHAPE analysis of the entire FMDV genome (Poulsen, 2015) has identified three conserved RNA structures within the coding regions for 2B, 3C and 3D (RNA-dependent RNA polymerase) which might have an important role in virus replication. The FMDV 2A peptide, another conserved...... polypeptide. The nature of this “cleavage” has so far not been investigated in the context of the full-length FMDV RNA within cells. The focus of this PhD thesis has been to characterize these elements and their influence on the FMDV replication. In order to fulfil the aims of this thesis a series of studies...

  15. Genomic features shaping the landscape of meiotic double-strand-break hotspots in maize.

    Science.gov (United States)

    He, Yan; Wang, Minghui; Dukowic-Schulze, Stefanie; Zhou, Adele; Tiang, Choon-Lin; Shilo, Shay; Sidhu, Gaganpreet K; Eichten, Steven; Bradbury, Peter; Springer, Nathan M; Buckler, Edward S; Levy, Avraham A; Sun, Qi; Pillardy, Jaroslaw; Kianian, Penny M A; Kianian, Shahryar F; Chen, Changbin; Pawlowski, Wojciech P

    2017-11-14

    Meiotic recombination is the most important source of genetic variation in higher eukaryotes. It is initiated by formation of double-strand breaks (DSBs) in chromosomal DNA in early meiotic prophase. The DSBs are subsequently repaired, resulting in crossovers (COs) and noncrossovers (NCOs). Recombination events are not distributed evenly along chromosomes but cluster at recombination hotspots. How specific sites become hotspots is poorly understood. Studies in yeast and mammals linked initiation of meiotic recombination to active chromatin features present upstream from genes, such as absence of nucleosomes and presence of trimethylation of lysine 4 in histone H3 (H3K4me3). Core recombination components are conserved among eukaryotes, but it is unclear whether this conservation results in universal characteristics of recombination landscapes shared by a wide range of species. To address this question, we mapped meiotic DSBs in maize, a higher eukaryote with a large genome that is rich in repetitive DNA. We found DSBs in maize to be frequent in all chromosome regions, including sites lacking COs, such as centromeres and pericentromeric regions. Furthermore, most DSBs are formed in repetitive DNA, predominantly Gypsy retrotransposons, and only one-quarter of DSB hotspots are near genes. Genic and nongenic hotspots differ in several characteristics, and only genic DSBs contribute to crossover formation. Maize hotspots overlap regions of low nucleosome occupancy but show only limited association with H3K4me3 sites. Overall, maize DSB hotspots exhibit distribution patterns and characteristics not reported previously in other species. Understanding recombination patterns in maize will shed light on mechanisms affecting dynamics of the plant genome.

  16. Single molecule localization imaging of telomeres and centromeres using fluorescence in situ hybridization and semiconductor quantum dots.

    Science.gov (United States)

    Wang, Le; Zong, Shenfei; Wang, Zhuyuan; Lu, Ju; Chen, Chen; Zhang, Ruohu; Cui, Yiping

    2018-07-13

    Single molecule localization microscopy (SMLM) is a powerful tool for imaging biological targets at the nanoscale. In this report, we present SMLM imaging of telomeres and centromeres using fluorescence in situ hybridization (FISH). The FISH probes were fabricated by decorating CdSSe/ZnS quantum dots (QDs) with telomere or centromere complementary DNA strands. SMLM imaging experiments using commercially available peptide nucleic acid (PNA) probes labeled with organic fluorophores were also conducted to demonstrate the advantages of using QDs FISH probes. Compared with the PNA probes, the QDs probes have the following merits. First, the fluorescence blinking of QDs can be realized in aqueous solution or PBS buffer without thiol, which is a key buffer component for organic fluorophores' blinking. Second, fluorescence blinking of the QDs probe needs only one excitation light (i.e. 405 nm). While fluorescence blinking of the organic fluorophores usually requires two illumination lights, that is, the activation light (i.e. 405 nm) and the imaging light. Third, the high quantum yield, multiple switching times and a good optical stability make the QDs more suitable for long-term imaging. The localization precision achieved in telomeres and centromeres imaging experiments is about 30 nm, which is far beyond the diffraction limit. SMLM has enabled new insights into telomeres or centromeres on the molecular level, and it is even possible to determine the length of telomere and become a potential technique for telomere-related investigation.

  17. Linking the potato genome to the conserved ortholog set (COS) markers

    Science.gov (United States)

    2013-01-01

    Background Conserved ortholog set (COS) markers are an important functional genomics resource that has greatly improved orthology detection in Asterid species. A comprehensive list of these markers is available at Sol Genomics Network (http://solgenomics.net/) and many of these have been placed on the genetic maps of a number of solanaceous species. Results We amplified over 300 COS markers from eight potato accessions involving two diploid landraces of Solanum tuberosum Andigenum group (formerly classified as S. goniocalyx, S. phureja), and a dihaploid clone derived from a modern tetraploid cultivar of S. tuberosum and the wild species S. berthaultii, S. chomatophilum, and S. paucissectum. By BLASTn (Basic Local Alignment Search Tool of the NCBI, National Center for Biotechnology Information) algorithm we mapped the DNA sequences of these markers into the potato genome sequence. Additionally, we mapped a subset of these markers genetically in potato and present a comparison between the physical and genetic locations of these markers in potato and in comparison with the genetic location in tomato. We found that most of the COS markers are single-copy in the reference genome of potato and that the genetic location in tomato and physical location in potato sequence are mostly in agreement. However, we did find some COS markers that are present in multiple copies and those that map in unexpected locations. Sequence comparisons between species show that some of these markers may be paralogs. Conclusions The sequence-based physical map becomes helpful in identification of markers for traits of interest thereby reducing the number of markers to be tested for applications like marker assisted selection, diversity, and phylogenetic studies. PMID:23758607

  18. A Simple Predictive Enhancer Syntax for Hindbrain Patterning Is Conserved in Vertebrate Genomes.

    Directory of Open Access Journals (Sweden)

    Joseph Grice

    Full Text Available Determining the function of regulatory elements is fundamental for our understanding of development, disease and evolution. However, the sequence features that mediate these functions are often unclear and the prediction of tissue-specific expression patterns from sequence alone is non-trivial. Previous functional studies have demonstrated a link between PBX-HOX and MEIS/PREP binding interactions and hindbrain enhancer activity, but the defining grammar of these sites, if any exists, has remained elusive.Here, we identify a shared sequence signature (syntax within a heterogeneous set of conserved vertebrate hindbrain enhancers composed of spatially co-occurring PBX-HOX and MEIS/PREP transcription factor binding motifs. We use this syntax to accurately predict hindbrain enhancers in 89% of cases (67/75 predicted elements from a set of conserved non-coding elements (CNEs. Furthermore, mutagenesis of the sites abolishes activity or generates ectopic expression, demonstrating their requirement for segmentally restricted enhancer activity in the hindbrain. We refine and use our syntax to predict over 3,000 hindbrain enhancers across the human genome. These sequences tend to be located near developmental transcription factors and are enriched in known hindbrain activating elements, demonstrating the predictive power of this simple model.Our findings support the theory that hundreds of CNEs, and perhaps thousands of regions across the human genome, function to coordinate gene expression in the developing hindbrain. We speculate that deeply conserved sequences of this kind contributed to the co-option of new genes into the hindbrain gene regulatory network during early vertebrate evolution by linking patterns of hox expression to downstream genes involved in segmentation and patterning, and evolutionarily newer instances may have continued to contribute to lineage-specific elaboration of the hindbrain.

  19. Different patterns of evolution in the centromeric and telomeric regions of group A and B haplotypes of the human killer cell Ig-like receptor locus.

    Directory of Open Access Journals (Sweden)

    Chul-Woo Pyo

    Full Text Available The fast evolving human KIR gene family encodes variable lymphocyte receptors specific for polymorphic HLA class I determinants. Nucleotide sequences for 24 representative human KIR haplotypes were determined. With three previously defined haplotypes, this gave a set of 12 group A and 15 group B haplotypes for assessment of KIR variation. The seven gene-content haplotypes are all combinations of four centromeric and two telomeric motifs. 2DL5, 2DS5 and 2DS3 can be present in centromeric and telomeric locations. With one exception, haplotypes having identical gene content differed in their combinations of KIR alleles. Sequence diversity varied between haplotype groups and between centromeric and telomeric halves of the KIR locus. The most variable A haplotype genes are in the telomeric half, whereas the most variable genes characterizing B haplotypes are in the centromeric half. Of the highly polymorphic genes, only the 3DL3 framework gene exhibits a similar diversity when carried by A and B haplotypes. Phylogenetic analysis and divergence time estimates, point to the centromeric gene-content motifs that distinguish A and B haplotypes having emerged ~6 million years ago, contemporaneously with the separation of human and chimpanzee ancestors. In contrast, the telomeric motifs that distinguish A and B haplotypes emerged more recently, ~1.7 million years ago, before the emergence of Homo sapiens. Thus the centromeric and telomeric motifs that typify A and B haplotypes have likely been present throughout human evolution. The results suggest the common ancestor of A and B haplotypes combined a B-like centromeric region with an A-like telomeric region.

  20. Genome-wide conserved non-coding microsatellite (CNMS) marker-based integrative genetical genomics for quantitative dissection of seed weight in chickpea.

    Science.gov (United States)

    Bajaj, Deepak; Saxena, Maneesha S; Kujur, Alice; Das, Shouvik; Badoni, Saurabh; Tripathi, Shailesh; Upadhyaya, Hari D; Gowda, C L L; Sharma, Shivali; Singh, Sube; Tyagi, Akhilesh K; Parida, Swarup K

    2015-03-01

    Phylogenetic footprinting identified 666 genome-wide paralogous and orthologous CNMS (conserved non-coding microsatellite) markers from 5'-untranslated and regulatory regions (URRs) of 603 protein-coding chickpea genes. The (CT)n and (GA)n CNMS carrying CTRMCAMV35S and GAGA8BKN3 regulatory elements, respectively, are abundant in the chickpea genome. The mapped genic CNMS markers with robust amplification efficiencies (94.7%) detected higher intraspecific polymorphic potential (37.6%) among genotypes, implying their immense utility in chickpea breeding and genetic analyses. Seventeen differentially expressed CNMS marker-associated genes showing strong preferential and seed tissue/developmental stage-specific expression in contrasting genotypes were selected to narrow down the gene targets underlying seed weight quantitative trait loci (QTLs)/eQTLs (expression QTLs) through integrative genetical genomics. The integration of transcript profiling with seed weight QTL/eQTL mapping, molecular haplotyping, and association analyses identified potential molecular tags (GAGA8BKN3 and RAV1AAT regulatory elements and alleles/haplotypes) in the LOB-domain-containing protein- and KANADI protein-encoding transcription factor genes controlling the cis-regulated expression for seed weight in the chickpea. This emphasizes the potential of CNMS marker-based integrative genetical genomics for the quantitative genetic dissection of complex seed weight in chickpea. © The Author 2014. Published by Oxford University Press on behalf of the Society for Experimental Biology.

  1. Mitotic control of human papillomavirus genome-containing cells is regulated by the function of the PDZ-binding motif of the E6 oncoprotein

    Science.gov (United States)

    Marsh, Elizabeth K.; Delury, Craig P.; Davies, Nicholas J.; Weston, Christopher J.; Miah, Mohammed A.L.; Banks, Lawrence; Parish, Joanna L.

    2017-01-01

    The function of a conserved PDS95/DLG1/ZO1 (PDZ) binding motif (E6 PBM) at the C-termini of E6 oncoproteins of high-risk human papillomavirus (HPV) types contributes to the development of HPV-associated malignancies. Here, using a primary human keratinocyte-based model of the high-risk HPV18 life cycle, we identify a novel link between the E6 PBM and mitotic stability. In cultures containing a mutant genome in which the E6 PBM was deleted there was an increase in the frequency of abnormal mitoses, including multinucleation, compared to cells harboring the wild type HPV18 genome. The loss of the E6 PBM was associated with a significant increase in the frequency of mitotic spindle defects associated with anaphase and telophase. Furthermore, cells carrying this mutant genome had increased chromosome segregation defects and they also exhibited greater levels of genomic instability, as shown by an elevated level of centromere-positive micronuclei. In wild type HPV18 genome-containing organotypic cultures, the majority of mitotic cells reside in the suprabasal layers, in keeping with the hyperplastic morphology of the structures. However, in mutant genome-containing structures a greater proportion of mitotic cells were retained in the basal layer, which were often of undefined polarity, thus correlating with their reduced thickness. We conclude that the ability of E6 to target cellular PDZ proteins plays a critical role in maintaining mitotic stability of HPV infected cells, ensuring stable episome persistence and vegetative amplification. PMID:28061478

  2. Mitotic control of human papillomavirus genome-containing cells is regulated by the function of the PDZ-binding motif of the E6 oncoprotein.

    Science.gov (United States)

    Marsh, Elizabeth K; Delury, Craig P; Davies, Nicholas J; Weston, Christopher J; Miah, Mohammed A L; Banks, Lawrence; Parish, Joanna L; Higgs, Martin R; Roberts, Sally

    2017-03-21

    The function of a conserved PDS95/DLG1/ZO1 (PDZ) binding motif (E6 PBM) at the C-termini of E6 oncoproteins of high-risk human papillomavirus (HPV) types contributes to the development of HPV-associated malignancies. Here, using a primary human keratinocyte-based model of the high-risk HPV18 life cycle, we identify a novel link between the E6 PBM and mitotic stability. In cultures containing a mutant genome in which the E6 PBM was deleted there was an increase in the frequency of abnormal mitoses, including multinucleation, compared to cells harboring the wild type HPV18 genome. The loss of the E6 PBM was associated with a significant increase in the frequency of mitotic spindle defects associated with anaphase and telophase. Furthermore, cells carrying this mutant genome had increased chromosome segregation defects and they also exhibited greater levels of genomic instability, as shown by an elevated level of centromere-positive micronuclei. In wild type HPV18 genome-containing organotypic cultures, the majority of mitotic cells reside in the suprabasal layers, in keeping with the hyperplastic morphology of the structures. However, in mutant genome-containing structures a greater proportion of mitotic cells were retained in the basal layer, which were often of undefined polarity, thus correlating with their reduced thickness. We conclude that the ability of E6 to target cellular PDZ proteins plays a critical role in maintaining mitotic stability of HPV infected cells, ensuring stable episome persistence and vegetative amplification.

  3. The Agassiz's desert tortoise genome provides a resource for the conservation of a threatened species.

    Directory of Open Access Journals (Sweden)

    Marc Tollis

    Full Text Available Agassiz's desert tortoise (Gopherus agassizii is a long-lived species native to the Mojave Desert and is listed as threatened under the US Endangered Species Act. To aid conservation efforts for preserving the genetic diversity of this species, we generated a whole genome reference sequence with an annotation based on deep transcriptome sequences of adult skeletal muscle, lung, brain, and blood. The draft genome assembly for G. agassizii has a scaffold N50 length of 252 kbp and a total length of 2.4 Gbp. Genome annotation reveals 20,172 protein-coding genes in the G. agassizii assembly, and that gene structure is more similar to chicken than other turtles. We provide a series of comparative analyses demonstrating (1 that turtles are among the slowest-evolving genome-enabled reptiles, (2 amino acid changes in genes controlling desert tortoise traits such as shell development, longevity and osmoregulation, and (3 fixed variants across the Gopherus species complex in genes related to desert adaptations, including circadian rhythm and innate immune response. This G. agassizii genome reference and annotation is the first such resource for any tortoise, and will serve as a foundation for future analysis of the genetic basis of adaptations to the desert environment, allow for investigation into genomic factors affecting tortoise health, disease and longevity, and serve as a valuable resource for additional studies in this species complex.

  4. Conservation and divergence of chemical defense system in the tunicate Oikopleura dioica revealed by genome wide response to two xenobiotics

    Directory of Open Access Journals (Sweden)

    Yadetie Fekadu

    2012-02-01

    Full Text Available Abstract Background Animals have developed extensive mechanisms of response to xenobiotic chemical attacks. Although recent genome surveys have suggested a broad conservation of the chemical defensome across metazoans, global gene expression responses to xenobiotics have not been well investigated in most invertebrates. Here, we performed genome survey for key defensome genes in Oikopleura dioica genome, and explored genome-wide gene expression using high density tiling arrays with over 2 million probes, in response to two model xenobiotic chemicals - the carcinogenic polycyclic aromatic hydrocarbon benzo[a]pyrene (BaP the pharmaceutical compound Clofibrate (Clo. Results Oikopleura genome surveys for key genes of the chemical defensome suggested a reduced repertoire. Not more than 23 cytochrome P450 (CYP genes could be identified, and neither CYP1 family genes nor their transcriptional activator AhR was detected. These two genes were present in deuterostome ancestors. As in vertebrates, the genotoxic compound BaP induced xenobiotic biotransformation and oxidative stress responsive genes. Notable exceptions were genes of the aryl hydrocarbon receptor (AhR signaling pathway. Clo also affected the expression of many biotransformation genes and markedly repressed genes involved in energy metabolism and muscle contraction pathways. Conclusions Oikopleura has the smallest number of CYP genes among sequenced animal genomes and lacks the AhR signaling pathway. However it appears to have basic xenobiotic inducible biotransformation genes such as a conserved genotoxic stress response gene set. Our genome survey and expression study does not support a role of AhR signaling pathway in the chemical defense of metazoans prior to the emergence of vertebrates.

  5. The B73 maize genome: complexity, diversity, and dynamics.

    Science.gov (United States)

    Schnable, Patrick S; Ware, Doreen; Fulton, Robert S; Stein, Joshua C; Wei, Fusheng; Pasternak, Shiran; Liang, Chengzhi; Zhang, Jianwei; Fulton, Lucinda; Graves, Tina A; Minx, Patrick; Reily, Amy Denise; Courtney, Laura; Kruchowski, Scott S; Tomlinson, Chad; Strong, Cindy; Delehaunty, Kim; Fronick, Catrina; Courtney, Bill; Rock, Susan M; Belter, Eddie; Du, Feiyu; Kim, Kyung; Abbott, Rachel M; Cotton, Marc; Levy, Andy; Marchetto, Pamela; Ochoa, Kerri; Jackson, Stephanie M; Gillam, Barbara; Chen, Weizu; Yan, Le; Higginbotham, Jamey; Cardenas, Marco; Waligorski, Jason; Applebaum, Elizabeth; Phelps, Lindsey; Falcone, Jason; Kanchi, Krishna; Thane, Thynn; Scimone, Adam; Thane, Nay; Henke, Jessica; Wang, Tom; Ruppert, Jessica; Shah, Neha; Rotter, Kelsi; Hodges, Jennifer; Ingenthron, Elizabeth; Cordes, Matt; Kohlberg, Sara; Sgro, Jennifer; Delgado, Brandon; Mead, Kelly; Chinwalla, Asif; Leonard, Shawn; Crouse, Kevin; Collura, Kristi; Kudrna, Dave; Currie, Jennifer; He, Ruifeng; Angelova, Angelina; Rajasekar, Shanmugam; Mueller, Teri; Lomeli, Rene; Scara, Gabriel; Ko, Ara; Delaney, Krista; Wissotski, Marina; Lopez, Georgina; Campos, David; Braidotti, Michele; Ashley, Elizabeth; Golser, Wolfgang; Kim, HyeRan; Lee, Seunghee; Lin, Jinke; Dujmic, Zeljko; Kim, Woojin; Talag, Jayson; Zuccolo, Andrea; Fan, Chuanzhu; Sebastian, Aswathy; Kramer, Melissa; Spiegel, Lori; Nascimento, Lidia; Zutavern, Theresa; Miller, Beth; Ambroise, Claude; Muller, Stephanie; Spooner, Will; Narechania, Apurva; Ren, Liya; Wei, Sharon; Kumari, Sunita; Faga, Ben; Levy, Michael J; McMahan, Linda; Van Buren, Peter; Vaughn, Matthew W; Ying, Kai; Yeh, Cheng-Ting; Emrich, Scott J; Jia, Yi; Kalyanaraman, Ananth; Hsia, An-Ping; Barbazuk, W Brad; Baucom, Regina S; Brutnell, Thomas P; Carpita, Nicholas C; Chaparro, Cristian; Chia, Jer-Ming; Deragon, Jean-Marc; Estill, James C; Fu, Yan; Jeddeloh, Jeffrey A; Han, Yujun; Lee, Hyeran; Li, Pinghua; Lisch, Damon R; Liu, Sanzhen; Liu, Zhijie; Nagel, Dawn Holligan; McCann, Maureen C; SanMiguel, Phillip; Myers, Alan M; Nettleton, Dan; Nguyen, John; Penning, Bryan W; Ponnala, Lalit; Schneider, Kevin L; Schwartz, David C; Sharma, Anupma; Soderlund, Carol; Springer, Nathan M; Sun, Qi; Wang, Hao; Waterman, Michael; Westerman, Richard; Wolfgruber, Thomas K; Yang, Lixing; Yu, Yeisoo; Zhang, Lifang; Zhou, Shiguo; Zhu, Qihui; Bennetzen, Jeffrey L; Dawe, R Kelly; Jiang, Jiming; Jiang, Ning; Presting, Gernot G; Wessler, Susan R; Aluru, Srinivas; Martienssen, Robert A; Clifton, Sandra W; McCombie, W Richard; Wing, Rod A; Wilson, Richard K

    2009-11-20

    We report an improved draft nucleotide sequence of the 2.3-gigabase genome of maize, an important crop plant and model for biological research. Over 32,000 genes were predicted, of which 99.8% were placed on reference chromosomes. Nearly 85% of the genome is composed of hundreds of families of transposable elements, dispersed nonuniformly across the genome. These were responsible for the capture and amplification of numerous gene fragments and affect the composition, sizes, and positions of centromeres. We also report on the correlation of methylation-poor regions with Mu transposon insertions and recombination, and copy number variants with insertions and/or deletions, as well as how uneven gene losses between duplicated regions were involved in returning an ancient allotetraploid to a genetically diploid state. These analyses inform and set the stage for further investigations to improve our understanding of the domestication and agricultural improvements of maize.

  6. Ecological genomics for coral and sea urchin conservation in times of climate change

    Science.gov (United States)

    Carpizo-Ituarte, E.; Hofmann, G.; Fangue, N.; Cupul-Magaña, A.; Rodríguez-Troncoso, A. P.; Díaz-Pérez, L.; Olivares Bañuelos, T.; Escobar Fernández, R.

    2010-03-01

    If atmospheric CO2 levels continue to increase, it is predicted that the average ocean sea surface temperature will also increase and ocean pH will decrease to levels not experienced by marine organisms for millions of years. Understanding the impact of these stressors will require the study of several marine organisms, and this knowledge will be fundamental to our ability to predict possible effects along large geographical regions and across phyla. Ecological genomics, defined as the use of molecular techniques to answer ecological questions, offers a set of tools that can help us better understand the responses of marine organisms to changes in their environment. In the present work we are using genomic tools to characterize the response of corals and sea urchins to environmental stress. On one side, coral species represent a useful model due to its functions as "environmental sentinels" in tropical ecosystems; on the other hand, species of sea urchins, with the recent sequence of the genome of the purple sea urchin S. purpuratus, offers important genomic resources. Recent results in corals and in sea urchins have shown that the response to stressful conditions can be detected using molecular genomic markers. Continued study of the mRNA expression patterns of several important gene families including calcification genes as well as genes involved in the cellular stress response such as heat shock proteins, will be valuable index of ecological stress in marine systems. These data can be integrated into better strategies of conservation and management of the oceans.

  7. HMMerThread: detecting remote, functional conserved domains in entire genomes by combining relaxed sequence-database searches with fold recognition.

    Directory of Open Access Journals (Sweden)

    Charles Richard Bradshaw

    Full Text Available Conserved domains in proteins are one of the major sources of functional information for experimental design and genome-level annotation. Though search tools for conserved domain databases such as Hidden Markov Models (HMMs are sensitive in detecting conserved domains in proteins when they share sufficient sequence similarity, they tend to miss more divergent family members, as they lack a reliable statistical framework for the detection of low sequence similarity. We have developed a greatly improved HMMerThread algorithm that can detect remotely conserved domains in highly divergent sequences. HMMerThread combines relaxed conserved domain searches with fold recognition to eliminate false positive, sequence-based identifications. With an accuracy of 90%, our software is able to automatically predict highly divergent members of conserved domain families with an associated 3-dimensional structure. We give additional confidence to our predictions by validation across species. We have run HMMerThread searches on eight proteomes including human and present a rich resource of remotely conserved domains, which adds significantly to the functional annotation of entire proteomes. We find ∼4500 cross-species validated, remotely conserved domain predictions in the human proteome alone. As an example, we find a DNA-binding domain in the C-terminal part of the A-kinase anchor protein 10 (AKAP10, a PKA adaptor that has been implicated in cardiac arrhythmias and premature cardiac death, which upon stress likely translocates from mitochondria to the nucleus/nucleolus. Based on our prediction, we propose that with this HLH-domain, AKAP10 is involved in the transcriptional control of stress response. Further remotely conserved domains we discuss are examples from areas such as sporulation, chromosome segregation and signalling during immune response. The HMMerThread algorithm is able to automatically detect the presence of remotely conserved domains in

  8. Epigenetic Histone Marks of Extended Meta-Polycentric Centromeres of Lathyrus and Pisum Chromosomes

    Czech Academy of Sciences Publication Activity Database

    Neumann, Pavel; Schubert, V.; Vrbová, Iva; Manning, Jasper Eugene; Houben, A.; Macas, Jiří

    2016-01-01

    Roč. 7, č. 234 (2016) ISSN 1664-462X R&D Projects: GA ČR(CZ) GAP501/11/1843 Institutional support: RVO:60077344 Keywords : Centromere structure * epigenetic modifications * histone phosphorylation * histone methylation Subject RIV: EB - Genetics ; Molecular Biology Impact factor: 4.298, year: 2016

  9. Analysis of 90 Mb of the potato genome reveals conservation of gene structures and order with tomato but divergence in repetitive sequence composition

    Directory of Open Access Journals (Sweden)

    O'Brien Kimberly

    2008-06-01

    Full Text Available Abstract Background The Solanaceae family contains a number of important crop species including potato (Solanum tuberosum which is grown for its underground storage organ known as a tuber. Albeit the 4th most important food crop in the world, other than a collection of ~220,000 Expressed Sequence Tags, limited genomic sequence information is currently available for potato and advances in potato yield and nutrition content would be greatly assisted through access to a complete genome sequence. While morphologically diverse, Solanaceae species such as potato, tomato, pepper, and eggplant share not only genes but also gene order thereby permitting highly informative comparative genomic analyses. Results In this study, we report on analysis 89.9 Mb of potato genomic sequence representing 10.2% of the genome generated through end sequencing of a potato bacterial artificial chromosome (BAC clone library (87 Mb and sequencing of 22 potato BAC clones (2.9 Mb. The GC content of potato is very similar to Solanum lycopersicon (tomato and other dicotyledonous species yet distinct from the monocotyledonous grass species, Oryza sativa. Parallel analyses of repetitive sequences in potato and tomato revealed substantial differences in their abundance, 34.2% in potato versus 46.3% in tomato, which is consistent with the increased genome size per haploid genome of these two Solanum species. Specific classes and types of repetitive sequences were also differentially represented between these two species including a telomeric-related repetitive sequence, ribosomal DNA, and a number of unclassified repetitive sequences. Comparative analyses between tomato and potato at the gene level revealed a high level of conservation of gene content, genic feature, and gene order although discordances in synteny were observed. Conclusion Genomic level analyses of potato and tomato confirm that gene sequence and gene order are conserved between these solanaceous species and that

  10. Genome-wide Control of Heterochromatin Replication by the Telomere Capping Protein TRF2.

    Science.gov (United States)

    Mendez-Bermudez, Aaron; Lototska, Liudmyla; Bauwens, Serge; Giraud-Panis, Marie-Josèphe; Croce, Olivier; Jamet, Karine; Irizar, Agurtzane; Mowinckel, Macarena; Koundrioukoff, Stephane; Nottet, Nicolas; Almouzni, Genevieve; Teulade-Fichou, Mare-Paule; Schertzer, Michael; Perderiset, Mylène; Londoño-Vallejo, Arturo; Debatisse, Michelle; Gilson, Eric; Ye, Jing

    2018-05-03

    Hard-to-replicate regions of chromosomes (e.g., pericentromeres, centromeres, and telomeres) impede replication fork progression, eventually leading, in the event of replication stress, to chromosome fragility, aging, and cancer. Our knowledge of the mechanisms controlling the stability of these regions is essentially limited to telomeres, where fragility is counteracted by the shelterin proteins. Here we show that the shelterin subunit TRF2 ensures progression of the replication fork through pericentromeric heterochromatin, but not centromeric chromatin. In a process involving its N-terminal basic domain, TRF2 binds to pericentromeric Satellite III sequences during S phase, allowing the recruitment of the G-quadruplex-resolving helicase RTEL1 to facilitate fork progression. We also show that TRF2 is required for the stability of other heterochromatic regions localized throughout the genome, paving the way for future research on heterochromatic replication and its relationship with aging and cancer. Copyright © 2018 Elsevier Inc. All rights reserved.

  11. Detection of alien chromatin introgression from Thinopyrum into wheat using S genomic DNA as a probe--a landmark approach for Thinopyrum genome research.

    Science.gov (United States)

    Chen, Q

    2005-01-01

    The introduction of alien genetic variation from the genus Thinopyrum through chromosome engineering into wheat is a valuable and proven technique for wheat improvement. A number of economically important traits have been transferred into wheat as single genes, chromosome arms or entire chromosomes. Successful transfers can be greatly assisted by the precise identification of alien chromatin in the recipient progenies. Chromosome identification and characterization are useful for genetic manipulation and transfer in wheat breeding following chromosome engineering. Genomic in situ hybridization (GISH) using an S genomic DNA probe from the diploid species Pseudoroegneria has proven to be a powerful diagnostic cytogenetic tool for monitoring the transfer of many promising agronomic traits from Thinopyrum. This specific S genomic probe not only allows the direct determination of the chromosome composition in wheat-Thinopyrum hybrids, but also can separate the Th. intermedium chromosomes into the J, J(S) and S genomes. The J(S) genome, which consists of a modified J genome chromosome distinguished by S genomic sequences of Pseudoroegneria near the centromere and telomere, carries many disease and mite resistance genes. Utilization of this S genomic probe leads to a better understanding of genomic affinities between Thinopyrum and wheat, and provides a molecular cytogenetic marker for monitoring the transfer of alien Thinopyrum agronomic traits into wheat recipient lines. Copyright 2005 S. Karger AG, Basel.

  12. Essential loci in centromeric heterochromatin of Drosophila melanogaster. I: the right arm of chromosome 2.

    Science.gov (United States)

    Coulthard, Alistair B; Alm, Christina; Cealiac, Iulia; Sinclair, Don A; Honda, Barry M; Rossi, Fabrizio; Dimitri, Patrizio; Hilliker, Arthur J

    2010-06-01

    With the most recent releases of the Drosophila melanogaster genome sequences, much of the previously absent heterochromatic sequences have now been annotated. We undertook an extensive genetic analysis of existing lethal mutations, as well as molecular mapping and sequence analysis (using a candidate gene approach) to identify as many essential genes as possible in the centromeric heterochromatin on the right arm of the second chromosome (2Rh) of D. melanogaster. We also utilized available RNA interference lines to knock down the expression of genes in 2Rh as another approach to identifying essential genes. In total, we verified the existence of eight novel essential loci in 2Rh: CG17665, CG17683, CG17684, CG17883, CG40127, CG41265, CG42595, and Atf6. Two of these essential loci, CG41265 and CG42595, are synonymous with the previously characterized loci l(2)41Ab and unextended, respectively. The genetic and molecular analysis of the previously reported locus, l(2)41Ae, revealed that this is not a single locus, but rather it is a large region of 2Rh that extends from unextended (CG42595) to CG17665 and includes four of the novel loci uncovered here.

  13. Correlation between centromere protein-F autoantibodies and cancer analyzed by enzyme-linked immunosorbent assay

    DEFF Research Database (Denmark)

    Welner, Simon; Trier, Nicole Hartwig; Morten Frisch, Morten

    2013-01-01

    Centromere protein-F (CENP-F) is a large nuclear protein of 367 kDa, which is involved in multiple mitosis-related events such as proper assembly of the kinetochores, stabilization of heterochromatin, chromosome alignment and mitotic checkpoint signaling. Several studies have shown a correlation...

  14. Conservation of Repeats at the Mammalian KCNQ1OT1-CDKN1C Region Suggests a Role in Genomic Imprinting

    Directory of Open Access Journals (Sweden)

    Marcos De Donato

    2017-06-01

    Full Text Available KCNQ1OT1 is located in the region with the highest number of genes showing genomic imprinting, but the mechanisms controlling the genes under its influence have not been fully elucidated. Therefore, we conducted a comparative analysis of the KCNQ1/KCNQ1OT1-CDKN1C region to study its conservation across the best assembled eutherian mammalian genomes sequenced to date and analyzed potential elements that may be implicated in the control of genomic imprinting in this region. The genomic features in these regions from human, mouse, cattle, and dog show a higher number of genes and CpG islands (detected using cpgplot from EMBOSS, but lower number of repetitive elements (including short interspersed nuclear elements and long interspersed nuclear elements, compared with their whole chromosomes (detected by RepeatMasker. The KCNQ1OT1-CDKN1C region contains the highest number of conserved noncoding sequences (CNS among mammals, where we found 16 regions containing about 38 different highly conserved repetitive elements (using mVista, such as LINE1 elements: L1M4, L1MB7, HAL1, L1M4a, L1Med, and an LTR element: MLT1H. From these elements, we found 74 CNS showing high sequence identity (>70% between human, cattle, and mouse, from which we identified 13 motifs (using Multiple Em for Motif Elicitation/Motif Alignment and Search Tool with a significant probability of occurrence, 3 of which were the most frequent and were used to find transcription factor–binding sites. We detected several transcription factors (using JASPAR suite from the families SOX, FOX, and GATA. A phylogenetic analysis of these CNS from human, marmoset, mouse, rat, cattle, dog, horse, and elephant shows branches with high levels of support and very similar phylogenetic relationships among these groups, confirming previous reports. Our results suggest that functional DNA elements identified by comparative genomics in a region densely populated with imprinted mammalian genes may be

  15. Modular architecture of the T4 phage superfamily: A conserved core genome and a plastic periphery

    International Nuclear Information System (INIS)

    Comeau, Andre M.; Bertrand, Claire; Letarov, Andrei; Tetart, Francoise; Krisch, H.M.

    2007-01-01

    Among the most numerous objects in the biosphere, phages show enormous diversity in morphology and genetic content. We have sequenced 7 T4-like phages and compared their genome architecture. All seven phages share a core genome with T4 that is interrupted by several hyperplastic regions (HPRs) where most of their divergence occurs. The core primarily includes homologues of essential T4 genes, such as the virion structure and DNA replication genes. In contrast, the HPRs contain mostly novel genes of unknown function and origin. A few of the HPR genes that can be assigned putative functions, such as a series of novel Internal Proteins, are implicated in phage adaptation to the host. Thus, the T4-like genome appears to be partitioned into discrete segments that fulfil different functions and behave differently in evolution. Such partitioning may be critical for these large and complex phages to maintain their flexibility, while simultaneously allowing them to conserve their highly successful virion design and mode of replication

  16. The perennial ryegrass GenomeZipper: targeted use of genome resources for comparative grass genomics.

    Science.gov (United States)

    Pfeifer, Matthias; Martis, Mihaela; Asp, Torben; Mayer, Klaus F X; Lübberstedt, Thomas; Byrne, Stephen; Frei, Ursula; Studer, Bruno

    2013-02-01

    Whole-genome sequences established for model and major crop species constitute a key resource for advanced genomic research. For outbreeding forage and turf grass species like ryegrasses (Lolium spp.), such resources have yet to be developed. Here, we present a model of the perennial ryegrass (Lolium perenne) genome on the basis of conserved synteny to barley (Hordeum vulgare) and the model grass genome Brachypodium (Brachypodium distachyon) as well as rice (Oryza sativa) and sorghum (Sorghum bicolor). A transcriptome-based genetic linkage map of perennial ryegrass served as a scaffold to establish the chromosomal arrangement of syntenic genes from model grass species. This scaffold revealed a high degree of synteny and macrocollinearity and was then utilized to anchor a collection of perennial ryegrass genes in silico to their predicted genome positions. This resulted in the unambiguous assignment of 3,315 out of 8,876 previously unmapped genes to the respective chromosomes. In total, the GenomeZipper incorporates 4,035 conserved grass gene loci, which were used for the first genome-wide sequence divergence analysis between perennial ryegrass, barley, Brachypodium, rice, and sorghum. The perennial ryegrass GenomeZipper is an ordered, information-rich genome scaffold, facilitating map-based cloning and genome assembly in perennial ryegrass and closely related Poaceae species. It also represents a milestone in describing synteny between perennial ryegrass and fully sequenced model grass genomes, thereby increasing our understanding of genome organization and evolution in the most important temperate forage and turf grass species.

  17. Genome-wide discovery of novel and conserved microRNAs in white shrimp (Litopenaeus vannamei).

    Science.gov (United States)

    Xi, Qian-Yun; Xiong, Yuan-Yan; Wang, Yuan-Mei; Cheng, Xiao; Qi, Qi-En; Shu, Gang; Wang, Song-Bo; Wang, Li-Na; Gao, Ping; Zhu, Xiao-Tong; Jiang, Qing-Yan; Zhang, Yong-Liang; Liu, Li

    2015-01-01

    Of late years, a large amount of conserved and species-specific microRNAs (miRNAs) have been performed on identification from species which are economically important but lack a full genome sequence. In this study, Solexa deep sequencing and cross-species miRNA microarray were used to detect miRNAs in white shrimp. We identified 239 conserved miRNAs, 14 miRNA* sequences and 20 novel miRNAs by bioinformatics analysis from 7,561,406 high-quality reads representing 325,370 distinct sequences. The all 20 novel miRNAs were species-specific in white shrimp and not homologous in other species. Using the conserved miRNAs from the miRBase database as a query set to search for homologs from shrimp expressed sequence tags (ESTs), 32 conserved computationally predicted miRNAs were discovered in shrimp. In addition, using microarray analysis in the shrimp fed with Panax ginseng polysaccharide complex, 151 conserved miRNAs were identified, 18 of which were significant up-expression, while 49 miRNAs were significant down-expression. In particular, qRT-PCR analysis was also performed for nine miRNAs in three shrimp tissues such as muscle, gill and hepatopancreas. Results showed that these miRNAs expression are tissue specific. Combining results of the three methods, we detected 20 novel and 394 conserved miRNAs. Verification with quantitative reverse transcription (qRT-PCR) and Northern blot showed a high confidentiality of data. The study provides the first comprehensive specific miRNA profile of white shrimp, which includes useful information for future investigations into the function of miRNAs in regulation of shrimp development and immunology.

  18. Isolation of BAC Clones Containing Conserved Genes from Libraries of Three Distantly Related Moths: A Useful Resource for Comparative Genomics of Lepidoptera

    Directory of Open Access Journals (Sweden)

    Yuji Yasukochi

    2011-01-01

    Full Text Available Lepidoptera, butterflies and moths, is the second largest animal order and includes numerous agricultural pests. To facilitate comparative genomics in Lepidoptera, we isolated BAC clones containing conserved and putative single-copy genes from libraries of three pests, Heliothis virescens, Ostrinia nubilalis, and Plutella xylostella, harboring the haploid chromosome number, =31, which are not closely related with each other or with the silkworm, Bombyx mori, (=28, the sequenced model lepidopteran. A total of 108–184 clones representing 101–182 conserved genes were isolated for each species. For 79 genes, clones were isolated from more than two species, which will be useful as common markers for analysis using fluorescence in situ hybridization (FISH, as well as for comparison of genome sequence among multiple species. The PCR-based clone isolation method presented here is applicable to species which lack a sequenced genome but have a significant collection of cDNA or EST sequences.

  19. Hybridization Capture Reveals Evolution and Conservation across the Entire Koala Retrovirus Genome

    Science.gov (United States)

    Ishida, Yasuko; Cui, Pin; Vielgrader, Hanna; Helgen, Kristofer M.; Roca, Alfred L.; Greenwood, Alex D.

    2014-01-01

    The koala retrovirus (KoRV) is the only retrovirus known to be in the midst of invading the germ line of its host species. Hybridization capture and next generation sequencing were used on modern and museum DNA samples of koala (Phascolarctos cinereus) to examine ca. 130 years of evolution across the full KoRV genome. Overall, the entire proviral genome appeared to be conserved across time in sequence, protein structure and transcriptional binding sites. A total of 138 polymorphisms were detected, of which 72 were found in more than one individual. At every polymorphic site in the museum koalas, one of the character states matched that of modern KoRV. Among non-synonymous polymorphisms, radical substitutions involving large physiochemical differences between amino acids were elevated in env, potentially reflecting anti-viral immune pressure or avoidance of receptor interference. Polymorphisms were not detected within two functional regions believed to affect infectivity. Host sequences flanking proviral integration sites were also captured; with few proviral loci shared among koalas. Recently described variants of KoRV, designated KoRV-B and KoRV-J, were not detected in museum samples, suggesting that these variants may be of recent origin. PMID:24752422

  20. Hybridization capture reveals evolution and conservation across the entire Koala retrovirus genome.

    Directory of Open Access Journals (Sweden)

    Kyriakos Tsangaras

    Full Text Available The koala retrovirus (KoRV is the only retrovirus known to be in the midst of invading the germ line of its host species. Hybridization capture and next generation sequencing were used on modern and museum DNA samples of koala (Phascolarctos cinereus to examine ca. 130 years of evolution across the full KoRV genome. Overall, the entire proviral genome appeared to be conserved across time in sequence, protein structure and transcriptional binding sites. A total of 138 polymorphisms were detected, of which 72 were found in more than one individual. At every polymorphic site in the museum koalas, one of the character states matched that of modern KoRV. Among non-synonymous polymorphisms, radical substitutions involving large physiochemical differences between amino acids were elevated in env, potentially reflecting anti-viral immune pressure or avoidance of receptor interference. Polymorphisms were not detected within two functional regions believed to affect infectivity. Host sequences flanking proviral integration sites were also captured; with few proviral loci shared among koalas. Recently described variants of KoRV, designated KoRV-B and KoRV-J, were not detected in museum samples, suggesting that these variants may be of recent origin.

  1. B chromosomes and Robertsonian fusions of Dichroplus pratensis (Acrididae): intraspecific support for the centromeric drive theory.

    Science.gov (United States)

    Bidau, C J; Martí, D A

    2004-01-01

    We tested the centromeric drive theory of karyotypic evolution in the grasshopper Dichroplus pratensis, which is simultaneously polymorphic for eight Robertsonian fusions and two classes of B chromosomes. A logistic regression analysis performed on 53 natural populations from Argentina revealed that B chromosomes are more probably found in populations with a higher proportion of acrocentric chromosomes, as the theory predicts. Furthermore, frequencies of B-carrying individuals are significantly negatively correlated with the mean frequency of different Robertsonian fusions per individual. No significant correlations between presence/absence or frequency of Bs, and latitude or altitude of the sampled populations, were found. We thus provide the first intraspecific evidence supporting the centromeric drive theory in relation to the establishment of B chromosomes in natural populations. Copyright 2004 S. Karger AG, Basel

  2. Visualizing conserved gene location across microbe genomes

    Science.gov (United States)

    Shaw, Chris D.

    2009-01-01

    This paper introduces an analysis-based zoomable visualization technique for displaying the location of genes across many related species of microbes. The purpose of this visualizatiuon is to enable a biologist to examine the layout of genes in the organism of interest with respect to the gene organization of related organisms. During the genomic annotation process, the ability to observe gene organization in common with previously annotated genomes can help a biologist better confirm the structure and function of newly analyzed microbe DNA sequences. We have developed a visualization and analysis tool that enables the biologist to observe and examine gene organization among genomes, in the context of the primary sequence of interest. This paper describes the visualization and analysis steps, and presents a case study using a number of Rickettsia genomes.

  3. ParABS Systems of the Four Replicons of Burkholderia cenocepacia: New Chromosome Centromeres Confer Partition Specificity†

    Science.gov (United States)

    Dubarry, Nelly; Pasta, Franck; Lane, David

    2006-01-01

    Most bacterial chromosomes carry an analogue of the parABS systems that govern plasmid partition, but their role in chromosome partition is ambiguous. parABS systems might be particularly important for orderly segregation of multipartite genomes, where their role may thus be easier to evaluate. We have characterized parABS systems in Burkholderia cenocepacia, whose genome comprises three chromosomes and one low-copy-number plasmid. A single parAB locus and a set of ParB-binding (parS) centromere sites are located near the origin of each replicon. ParA and ParB of the longest chromosome are phylogenetically similar to analogues in other multichromosome and monochromosome bacteria but are distinct from those of smaller chromosomes. The latter form subgroups that correspond to the taxa of their hosts, indicating evolution from plasmids. The parS sites on the smaller chromosomes and the plasmid are similar to the “universal” parS of the main chromosome but with a sequence specific to their replicon. In an Escherichia coli plasmid stabilization test, each parAB exhibits partition activity only with the parS of its own replicon. Hence, parABS function is based on the independent partition of individual chromosomes rather than on a single communal system or network of interacting systems. Stabilization by the smaller chromosome and plasmid systems was enhanced by mutation of parS sites and a promoter internal to their parAB operons, suggesting autoregulatory mechanisms. The small chromosome ParBs were found to silence transcription, a property relevant to autoregulation. PMID:16452432

  4. Genome-wide discovery and differential regulation of conserved and novel microRNAs in chickpea via deep sequencing.

    Science.gov (United States)

    Jain, Mukesh; Chevala, V V S Narayana; Garg, Rohini

    2014-11-01

    MicroRNAs (miRNAs) are essential components of complex gene regulatory networks that orchestrate plant development. Although several genomic resources have been developed for the legume crop chickpea, miRNAs have not been discovered until now. For genome-wide discovery of miRNAs in chickpea (Cicer arietinum), we sequenced the small RNA content from seven major tissues/organs employing Illumina technology. About 154 million reads were generated, which represented more than 20 million distinct small RNA sequences. We identified a total of 440 conserved miRNAs in chickpea based on sequence similarity with known miRNAs in other plants. In addition, 178 novel miRNAs were identified using a miRDeep pipeline with plant-specific scoring. Some of the conserved and novel miRNAs with significant sequence similarity were grouped into families. The chickpea miRNAs targeted a wide range of mRNAs involved in diverse cellular processes, including transcriptional regulation (transcription factors), protein modification and turnover, signal transduction, and metabolism. Our analysis revealed several miRNAs with differential spatial expression. Many of the chickpea miRNAs were expressed in a tissue-specific manner. The conserved and differential expression of members of the same miRNA family in different tissues was also observed. Some of the same family members were predicted to target different chickpea mRNAs, which suggested the specificity and complexity of miRNA-mediated developmental regulation. This study, for the first time, reveals a comprehensive set of conserved and novel miRNAs along with their expression patterns and putative targets in chickpea, and provides a framework for understanding regulation of developmental processes in legumes. © The Author 2014. Published by Oxford University Press on behalf of the Society for Experimental Biology.

  5. Characterization of the oncogenic function of centromere protein F in hepatocellular carcinoma

    Energy Technology Data Exchange (ETDEWEB)

    Dai, Yongdong; Liu, Lulu; Zeng, Tingting; Zhu, Ying-Hui [State Key Laboratory of Oncology in Southern China, Sun Yat-Sen University Cancer Center, Guangzhou (China); Li, Jiangchao [Vascular Biology Research Institute, Guangdong Pharmaceutical University, Guangzhou (China); Chen, Leilei [Department of Clinical Oncology, The University of Hong Kong, Pokfulam, Hong Kong (China); Li, Yan; Yuan, Yun-Fei [State Key Laboratory of Oncology in Southern China, Sun Yat-Sen University Cancer Center, Guangzhou (China); Ma, Stephanie, E-mail: stefma@hku.hk [Department of Clinical Oncology, The University of Hong Kong, Pokfulam, Hong Kong (China); State Key Laboratory for Liver Research, The University of Hong Kong, Pokfulam, Hong Kong (China); Guan, Xin-Yuan, E-mail: xyguan@hkucc.hku.hk [State Key Laboratory of Oncology in Southern China, Sun Yat-Sen University Cancer Center, Guangzhou (China); Department of Clinical Oncology, The University of Hong Kong, Pokfulam, Hong Kong (China); State Key Laboratory for Liver Research, The University of Hong Kong, Pokfulam, Hong Kong (China)

    2013-07-12

    Highlights: •Overexpression of CENPF is frequently detected in HCC. •Upregulation of CENPF serves as an independent prognosis factor in HCC patients. •CENPF functions as an oncogene in HCC by promoting cell G2/M transition. -- Abstract: Centromere protein F (CENPF) is an essential nuclear protein associated with the centromere-kinetochore complex and plays a critical role in chromosome segregation during mitosis. Up-regulation of CENPF expression has previously been detected in several solid tumors. In this study, we aim to study the expression and functional role of CENPF in hepatocellular carcinoma (HCC). We found CENPF was frequently overexpressed in HCC as compared with non-tumor tissue. Up-regulated CENPF expression in HCC was positively correlated with serum AFP, venous invasion, advanced differentiation stage and a shorter overall survival. Cox regression analysis found that overexpression of CENPF was an independent prognosis factor in HCC. Functional studies found that silencing CENPF could decrease the ability of the cells to proliferate, form colonies and induce tumor formation in nude mice. Silencing CENPF also resulted in the cell cycle arrest at G2/M checkpoint by down-regulating cell cycle proteins cdc2 and cyclin B1. Our data suggest that CENPF is frequently overexpressed in HCC and plays a critical role in driving HCC tumorigenesis.

  6. MHF1-2/CENP-S-X performs distinct roles in centromere metabolism and genetic recombination.

    Science.gov (United States)

    Bhattacharjee, Sonali; Osman, Fekret; Feeney, Laura; Lorenz, Alexander; Bryer, Claire; Whitby, Matthew C

    2013-09-11

    The histone-fold proteins Mhf1/CENP-S and Mhf2/CENP-X perform two important functions in vertebrate cells. First, they are components of the constitutive centromere-associated network, aiding kinetochore assembly and function. Second, they work with the FANCM DNA translocase to promote DNA repair. However, it has been unclear whether there is crosstalk between these roles. We show that Mhf1 and Mhf2 in fission yeast, as in vertebrates, serve a dual function, aiding DNA repair/recombination and localizing to centromeres to promote chromosome segregation. Importantly, these functions are distinct, with the former being dependent on their interaction with the FANCM orthologue Fml1 and the latter not. Together with Fml1, they play a second role in aiding chromosome segregation by processing sister chromatid junctions. However, a failure of this activity does not manifest dramatically increased levels of chromosome missegregation due to the Mus81-Eme1 endonuclease, which acts as a failsafe to resolve DNA junctions before the end of mitosis.

  7. Comparative cytogenetic analysis of chromosomal aberrations and premature centromere division in persons exposed to radionuclides

    International Nuclear Information System (INIS)

    Jovicic, D.; Rakic, B.; Vukov, T.; Pajic, J.; Milacic, S.; Kovacevic, R.; Stevanovic, M.; Drakulic, D.; Bukvic, N.

    2009-01-01

    The aim of the research was to determine the presence of correlation between the frequency of premature centromere division (PCD) and chromosomal aberrations (CA) in metaphases in persons professionally exposed to radionuclides. Biological dosimetry was performed by conventional cytogenetic technique. The presence of PCD was confirmed by Fluorescent in situ hybridization (FISH). The L1.84 probe (specific for centromeric region of chromosome 18) was used. The analysis included 50 subjects employed in the Clinical Center of Serbia (C) (the average age of 45.24 ± 1.18 and the average exposition time 17.96 ± 1.15) and 40 subjects in control group (K) (the average age of 44.40 ± 0.98 and the average years of employment 19.67 ± 0.98 years) which were not exposed to genotoxic agents in their workplaces. The results showed that frequencies of CA and PCD were statistically significantly higher in subjects exposed to radionuclides than in the control group (Mann-Whitney U test, P [sr

  8. Condensin HEAT subunits required for DNA repair, kinetochore/centromere function and ploidy maintenance in fission yeast.

    Directory of Open Access Journals (Sweden)

    Xingya Xu

    Full Text Available Condensin, a central player in eukaryotic chromosomal dynamics, contains five evolutionarily-conserved subunits. Two SMC (structural maintenance of chromosomes subunits contain ATPase, hinge, and coiled-coil domains. One non-SMC subunit is similar to bacterial kleisin, and two other non-SMC subunits contain HEAT (similar to armadillo repeats. Here we report isolation and characterization of 21 fission yeast (Schizosaccharomyces pombe mutants for three non-SMC subunits, created using error-prone mutagenesis that resulted in single-amino acid substitutions. Beside condensation, segregation, and DNA repair defects, similar to those observed in previously isolated SMC and cnd2 mutants, novel phenotypes were observed for mutants of HEAT-repeats containing Cnd1 and Cnd3 subunits. cnd3-L269P is hypersensitive to the microtubule poison, thiabendazole, revealing defects in kinetochore/centromere and spindle assembly checkpoints. Three cnd1 and three cnd3 mutants increased cell size and doubled DNA content, thereby eliminating the haploid state. Five of these mutations reside in helix B of HEAT repeats. Two non-SMC condensin subunits, Cnd1 and Cnd3, are thus implicated in ploidy maintenance.

  9. Correlation between sequence conservation and structural thermodynamics of microRNA precursors from human, mouse, and chicken genomes

    Directory of Open Access Journals (Sweden)

    Wang Shengqi

    2010-10-01

    Full Text Available Abstract Background Previous studies have shown that microRNA precursors (pre-miRNAs have considerably more stable secondary structures than other native RNAs (tRNA, rRNA, and mRNA and artificial RNA sequences. However, pre-miRNAs with ultra stable secondary structures have not been investigated. It is not known if there is a tendency in pre-miRNA sequences towards or against ultra stable structures? Furthermore, the relationship between the structural thermodynamic stability of pre-miRNA and their evolution remains unclear. Results We investigated the correlation between pre-miRNA sequence conservation and structural stability as measured by adjusted minimum folding free energies in pre-miRNAs isolated from human, mouse, and chicken. The analysis revealed that conserved and non-conserved pre-miRNA sequences had structures with similar average stabilities. However, the relatively ultra stable and unstable pre-miRNAs were more likely to be non-conserved than pre-miRNAs with moderate stability. Non-conserved pre-miRNAs had more G+C than A+U nucleotides, while conserved pre-miRNAs contained more A+U nucleotides. Notably, the U content of conserved pre-miRNAs was especially higher than that of non-conserved pre-miRNAs. Further investigations showed that conserved and non-conserved pre-miRNAs exhibited different structural element features, even though they had comparable levels of stability. Conclusions We proposed that there is a correlation between structural thermodynamic stability and sequence conservation for pre-miRNAs from human, mouse, and chicken genomes. Our analyses suggested that pre-miRNAs with relatively ultra stable or unstable structures were less favoured by natural selection than those with moderately stable structures. Comparison of nucleotide compositions between non-conserved and conserved pre-miRNAs indicated the importance of U nucleotides in the pre-miRNA evolutionary process. Several characteristic structural elements were

  10. CpGislandEVO: A Database and Genome Browser for Comparative Evolutionary Genomics of CpG Islands

    Directory of Open Access Journals (Sweden)

    Guillermo Barturen

    2013-01-01

    Full Text Available Hypomethylated, CpG-rich DNA segments (CpG islands, CGIs are epigenome markers involved in key biological processes. Aberrant methylation is implicated in the appearance of several disorders as cancer, immunodeficiency, or centromere instability. Furthermore, methylation differences at promoter regions between human and chimpanzee strongly associate with genes involved in neurological/psychological disorders and cancers. Therefore, the evolutionary comparative analyses of CGIs can provide insights on the functional role of these epigenome markers in both health and disease. Given the lack of specific tools, we developed CpGislandEVO. Briefly, we first compile a database of statistically significant CGIs for the best assembled mammalian genome sequences available to date. Second, by means of a coupled browser front-end, we focus on the CGIs overlapping orthologous genes extracted from OrthoDB, thus ensuring the comparison between CGIs located on truly homologous genome segments. This allows comparing the main compositional features between homologous CGIs. Finally, to facilitate nucleotide comparisons, we lifted genome coordinates between assemblies from different species, which enables the analysis of sequence divergence by direct count of nucleotide substitutions and indels occurring between homologous CGIs. The resulting CpGislandEVO database, linking together CGIs and single-cytosine DNA methylation data from several mammalian species, is freely available at our website.

  11. The Perennial Ryegrass GenomeZipper: Targeted Use of Genome Resources for Comparative Grass Genomics1[C][W

    Science.gov (United States)

    Pfeifer, Matthias; Martis, Mihaela; Asp, Torben; Mayer, Klaus F.X.; Lübberstedt, Thomas; Byrne, Stephen; Frei, Ursula; Studer, Bruno

    2013-01-01

    Whole-genome sequences established for model and major crop species constitute a key resource for advanced genomic research. For outbreeding forage and turf grass species like ryegrasses (Lolium spp.), such resources have yet to be developed. Here, we present a model of the perennial ryegrass (Lolium perenne) genome on the basis of conserved synteny to barley (Hordeum vulgare) and the model grass genome Brachypodium (Brachypodium distachyon) as well as rice (Oryza sativa) and sorghum (Sorghum bicolor). A transcriptome-based genetic linkage map of perennial ryegrass served as a scaffold to establish the chromosomal arrangement of syntenic genes from model grass species. This scaffold revealed a high degree of synteny and macrocollinearity and was then utilized to anchor a collection of perennial ryegrass genes in silico to their predicted genome positions. This resulted in the unambiguous assignment of 3,315 out of 8,876 previously unmapped genes to the respective chromosomes. In total, the GenomeZipper incorporates 4,035 conserved grass gene loci, which were used for the first genome-wide sequence divergence analysis between perennial ryegrass, barley, Brachypodium, rice, and sorghum. The perennial ryegrass GenomeZipper is an ordered, information-rich genome scaffold, facilitating map-based cloning and genome assembly in perennial ryegrass and closely related Poaceae species. It also represents a milestone in describing synteny between perennial ryegrass and fully sequenced model grass genomes, thereby increasing our understanding of genome organization and evolution in the most important temperate forage and turf grass species. PMID:23184232

  12. Dynamics of Rex3 in the genomes of endangered Iberian Leuciscinae (Teleostei, Cyprinidae and their natural hybrids

    Directory of Open Access Journals (Sweden)

    Carla Sofia A. Pereira

    2015-10-01

    Full Text Available Iberian Leuciscinae are greatly diverse comprising taxa of hybrid origin. With highly conservative karyotypes, Iberian Chondrostoma s.l. have recently demonstrated sub-chromosomal differentiation and rapid genome restructuring in natural hybrids, which was confirmed by ribosomal DNA (rDNA transposition and/or multiplication. To understand the role of repetitive DNAs in the differentiation of their genomes, a genetic and molecular cytogenetic survey was conducted in Achondrostoma oligolepis, Anaecypris hispanica, Iberochondrostoma lemmingii, I. lusitanicum, Pseudochondrostoma duriense, P. polylepis, Squalius pyrenaicus and hybrids between A. oligolepis x (P. duriense/P. polylepis, which represent 'alburnine', chondrostomine and Squalius lineages. The chromosomal distribution of Rex3 retroelement was found highly compartmentalized at centromeres and moderately at telomeres, co-localizing with 5S rDNA loci, and grossly correlating with heterochromatin and blocks of C0t-1 DNA. This accumulation was evident in at least 10 chromosome pairs, a pattern that seemed to be shared among the different species, likely predating their divergence. Nevertheless, species-specific clusters were detected in I. lusitanicum, P. duriense, P. polylepis and S. pyrenaicus demonstrating rapid and independent differentiation. Natural hybrids followed the same accumulation pattern and association with repetitive sequences but with increased number of Rex3 clusters and correlating with translocated 45S rDNA clusters. Rex3 sequence phylogeny didn't agree with its hosts' phylogeny but the observed distribution pattern is congruent with an evolutionary tendency to protect its activity, a robust regulatory system and/or events of horizontal transfer. This is the first report of retroelement physical mapping in Cyprinidae. It helped outlining conceivable ancestral homologies and recognizing retrotransposon activation in hybrids, being possibly associated with genome

  13. Genetic and physical mapping of two centromere-proximal regions of chromosome IV in Aspergillus nidulans

    DEFF Research Database (Denmark)

    Aleksenko, Alexei Y.; Nielsen, Michael Lynge; Clutterbuck, A.J.

    2001-01-01

    revision of the genetic map of the chromosome, including the position of the centromere, Comparison of physical and genetic maps indicates that meiotic recombination is low in subcentromeric DNA, its frequency being reduced from 1 crossover per 0.8 Mb to approximately 1 crossover per 5 Mb per meiosis...

  14. Genome-wide association scan identifies a risk locus for preeclampsia on 2q14, near the inhibin, beta B gene.

    Directory of Open Access Journals (Sweden)

    Matthew P Johnson

    Full Text Available Elucidating the genetic architecture of preeclampsia is a major goal in obstetric medicine. We have performed a genome-wide association study (GWAS for preeclampsia in unrelated Australian individuals of Caucasian ancestry using the Illumina OmniExpress-12 BeadChip to successfully genotype 648,175 SNPs in 538 preeclampsia cases and 540 normal pregnancy controls. Two SNP associations (rs7579169, p = 3.58×10(-7, OR = 1.57; rs12711941, p = 4.26×10(-7, OR = 1.56 satisfied our genome-wide significance threshold (modified Bonferroni p0.92. We attempted to provide evidence of a putative regulatory role for these SNPs using bioinformatic analyses and found that they all reside within regions of low sequence conservation and/or low complexity, suggesting functional importance is low. We also explored the mRNA expression in decidua of genes ±500 kb of INHBB and found a nominally significant correlation between a transcript encoded by the EPB41L5 gene, ∼250 kb centromeric to INHBB, and preeclampsia (p = 0.03. We were unable to replicate the associations shown by the significant GWAS SNPs in case-control cohorts from Norway and Finland, leading us to conclude that it is more likely that these SNPs are in LD with as yet unidentified causal variant(s.

  15. Comparison of C. elegans and C. briggsae genome sequences reveals extensive conservation of chromosome organization and synteny.

    Directory of Open Access Journals (Sweden)

    LaDeana W Hillier

    2007-07-01

    Full Text Available To determine whether the distinctive features of Caenorhabditis elegans chromosomal organization are shared with the C. briggsae genome, we constructed a single nucleotide polymorphism-based genetic map to order and orient the whole genome shotgun assembly along the six C. briggsae chromosomes. Although these species are of the same genus, their most recent common ancestor existed 80-110 million years ago, and thus they are more evolutionarily distant than, for example, human and mouse. We found that, like C. elegans chromosomes, C. briggsae chromosomes exhibit high levels of recombination on the arms along with higher repeat density, a higher fraction of intronic sequence, and a lower fraction of exonic sequence compared with chromosome centers. Despite extensive intrachromosomal rearrangements, 1:1 orthologs tend to remain in the same region of the chromosome, and colinear blocks of orthologs tend to be longer in chromosome centers compared with arms. More strikingly, the two species show an almost complete conservation of synteny, with 1:1 orthologs present on a single chromosome in one species also found on a single chromosome in the other. The conservation of both chromosomal organization and synteny between these two distantly related species suggests roles for chromosome organization in the fitness of an organism that are only poorly understood presently.

  16. A Trichosporonales genome tree based on 27 haploid and three evolutionarily conserved 'natural' hybrid genomes.

    Science.gov (United States)

    Takashima, Masako; Sriswasdi, Sira; Manabe, Ri-Ichiroh; Ohkuma, Moriya; Sugita, Takashi; Iwasaki, Wataru

    2018-01-01

    To construct a backbone tree consisting of basidiomycetous yeasts, draft genome sequences from 25 species of Trichosporonales (Tremellomycetes, Basidiomycota) were generated. In addition to the hybrid genomes of Trichosporon coremiiforme and Trichosporon ovoides that we described previously, we identified an interspecies hybrid genome in Cutaneotrichosporon mucoides (formerly Trichosporon mucoides). This hybrid genome had a gene retention rate of ~55%, and its closest haploid relative was Cutaneotrichosporon dermatis. After constructing the C. mucoides subgenomes, we generated a phylogenetic tree using genome data from the 27 haploid species and the subgenome data from the three hybrid genome species. It was a high-quality tree with 100% bootstrap support for all of the branches. The genome-based tree provided superior resolution compared with previous multi-gene analyses. Although our backbone tree does not include all Trichosporonales genera (e.g. Cryptotrichosporon), it will be valuable for future analyses of genome data. Interest in interspecies hybrid fungal genomes has recently increased because they may provide a basis for new technologies. The three Trichosporonales hybrid genomes described in this study are different from well-characterized hybrid genomes (e.g. those of Saccharomyces pastorianus and Saccharomyces bayanus) because these hybridization events probably occurred in the distant evolutionary past. Hence, they will be useful for studying genome stability following hybridization and speciation events. Copyright © 2017 John Wiley & Sons, Ltd. Copyright © 2017 John Wiley & Sons, Ltd.

  17. An online conserved SSR discovery through cross-species comparison

    Directory of Open Access Journals (Sweden)

    Tun-Wen Pai

    2009-02-01

    Full Text Available Tun-Wen Pai1, Chien-Ming Chen1, Meng-Chang Hsiao1, Ronshan Cheng2, Wen-Shyong Tzou3, Chin-Hua Hu31Department of Computer Science and Engineering; 2Department of Aquaculture, 3Institute of Bioscience and Biotechnology, National Taiwan Ocean University, Keelung, Taiwan, Republic of ChinaAbstract: Simple sequence repeats (SSRs play important roles in gene regulation and genome evolution. Although there exist several online resources for SSR mining, most of them only extract general SSR patterns without providing functional information. Here, an online search tool, CG-SSR (Comparative Genomics SSR discovery, has been developed for discovering potential functional SSRs from vertebrate genomes through cross-species comparison. In addition to revealing SSR candidates in conserved regions among various species, it also combines accurate coordinate and functional genomics information. CG-SSR is the first comprehensive and efficient online tool for conserved SSR discovery.Keywords: microsatellites, genome, comparative genomics, functional SSR, gene ontology, conserved region

  18. Highly conserved non-coding elements on either side of SOX9 associated with Pierre Robin sequence.

    Science.gov (United States)

    Benko, Sabina; Fantes, Judy A; Amiel, Jeanne; Kleinjan, Dirk-Jan; Thomas, Sophie; Ramsay, Jacqueline; Jamshidi, Negar; Essafi, Abdelkader; Heaney, Simon; Gordon, Christopher T; McBride, David; Golzio, Christelle; Fisher, Malcolm; Perry, Paul; Abadie, Véronique; Ayuso, Carmen; Holder-Espinasse, Muriel; Kilpatrick, Nicky; Lees, Melissa M; Picard, Arnaud; Temple, I Karen; Thomas, Paul; Vazquez, Marie-Paule; Vekemans, Michel; Roest Crollius, Hugues; Hastie, Nicholas D; Munnich, Arnold; Etchevers, Heather C; Pelet, Anna; Farlie, Peter G; Fitzpatrick, David R; Lyonnet, Stanislas

    2009-03-01

    Pierre Robin sequence (PRS) is an important subgroup of cleft palate. We report several lines of evidence for the existence of a 17q24 locus underlying PRS, including linkage analysis results, a clustering of translocation breakpoints 1.06-1.23 Mb upstream of SOX9, and microdeletions both approximately 1.5 Mb centromeric and approximately 1.5 Mb telomeric of SOX9. We have also identified a heterozygous point mutation in an evolutionarily conserved region of DNA with in vitro and in vivo features of a developmental enhancer. This enhancer is centromeric to the breakpoint cluster and maps within one of the microdeletion regions. The mutation abrogates the in vitro enhancer function and alters binding of the transcription factor MSX1 as compared to the wild-type sequence. In the developing mouse mandible, the 3-Mb region bounded by the microdeletions shows a regionally specific chromatin decompaction in cells expressing Sox9. Some cases of PRS may thus result from developmental misexpression of SOX9 due to disruption of very-long-range cis-regulatory elements.

  19. Novel Centromeric Loci of the Wine and Beer Yeast Dekkera bruxellensis CEN1 and CEN2

    DEFF Research Database (Denmark)

    Ishchuk, Olena P.; Vojvoda Zeljko, Tanja; Schifferdecker, Anna J.

    2016-01-01

    The wine and beer yeast Dekkera bruxellensis thrives in environments that are harsh and limiting, especially in concentrations with low oxygen and high ethanol. Its different strains' chromosomes greatly vary in number (karyotype). This study isolates two novel centromeric loci (CEN1 and CEN2...

  20. The complete mitochondrial genome of Gossypium hirsutum and evolutionary analysis of higher plant mitochondrial genomes.

    Science.gov (United States)

    Liu, Guozheng; Cao, Dandan; Li, Shuangshuang; Su, Aiguo; Geng, Jianing; Grover, Corrinne E; Hu, Songnian; Hua, Jinping

    2013-01-01

    Mitochondria are the main manufacturers of cellular ATP in eukaryotes. The plant mitochondrial genome contains large number of foreign DNA and repeated sequences undergone frequently intramolecular recombination. Upland Cotton (Gossypium hirsutum L.) is one of the main natural fiber crops and also an important oil-producing plant in the world. Sequencing of the cotton mitochondrial (mt) genome could be helpful for the evolution research of plant mt genomes. We utilized 454 technology for sequencing and combined with Fosmid library of the Gossypium hirsutum mt genome screening and positive clones sequencing and conducted a series of evolutionary analysis on Cycas taitungensis and 24 angiosperms mt genomes. After data assembling and contigs joining, the complete mitochondrial genome sequence of G. hirsutum was obtained. The completed G.hirsutum mt genome is 621,884 bp in length, and contained 68 genes, including 35 protein genes, four rRNA genes and 29 tRNA genes. Five gene clusters are found conserved in all plant mt genomes; one and four clusters are specifically conserved in monocots and dicots, respectively. Homologous sequences are distributed along the plant mt genomes and species closely related share the most homologous sequences. For species that have both mt and chloroplast genome sequences available, we checked the location of cp-like migration and found several fragments closely linked with mitochondrial genes. The G. hirsutum mt genome possesses most of the common characters of higher plant mt genomes. The existence of syntenic gene clusters, as well as the conservation of some intergenic sequences and genic content among the plant mt genomes suggest that evolution of mt genomes is consistent with plant taxonomy but independent among different species.

  1. Validation of rice genome sequence by optical mapping

    Directory of Open Access Journals (Sweden)

    Pape Louise

    2007-08-01

    Full Text Available Abstract Background Rice feeds much of the world, and possesses the simplest genome analyzed to date within the grass family, making it an economically relevant model system for other cereal crops. Although the rice genome is sequenced, validation and gap closing efforts require purely independent means for accurate finishing of sequence build data. Results To facilitate ongoing sequencing finishing and validation efforts, we have constructed a whole-genome SwaI optical restriction map of the rice genome. The physical map consists of 14 contigs, covering 12 chromosomes, with a total genome size of 382.17 Mb; this value is about 11% smaller than original estimates. 9 of the 14 optical map contigs are without gaps, covering chromosomes 1, 2, 3, 4, 5, 7, 8 10, and 12 in their entirety – including centromeres and telomeres. Alignments between optical and in silico restriction maps constructed from IRGSP (International Rice Genome Sequencing Project and TIGR (The Institute for Genomic Research genome sequence sources are comprehensive and informative, evidenced by map coverage across virtually all published gaps, discovery of new ones, and characterization of sequence misassemblies; all totalling ~14 Mb. Furthermore, since optical maps are ordered restriction maps, identified discordances are pinpointed on a reliable physical scaffold providing an independent resource for closure of gaps and rectification of misassemblies. Conclusion Analysis of sequence and optical mapping data effectively validates genome sequence assemblies constructed from large, repeat-rich genomes. Given this conclusion we envision new applications of such single molecule analysis that will merge advantages offered by high-resolution optical maps with inexpensive, but short sequence reads generated by emerging sequencing platforms. Lastly, map construction techniques presented here points the way to new types of comparative genome analysis that would focus on discernment of

  2. New Genome Similarity Measures based on Conserved Gene Adjacencies.

    Science.gov (United States)

    Doerr, Daniel; Kowada, Luis Antonio B; Araujo, Eloi; Deshpande, Shachi; Dantas, Simone; Moret, Bernard M E; Stoye, Jens

    2017-06-01

    Many important questions in molecular biology, evolution, and biomedicine can be addressed by comparative genomic approaches. One of the basic tasks when comparing genomes is the definition of measures of similarity (or dissimilarity) between two genomes, for example, to elucidate the phylogenetic relationships between species. The power of different genome comparison methods varies with the underlying formal model of a genome. The simplest models impose the strong restriction that each genome under study must contain the same genes, each in exactly one copy. More realistic models allow several copies of a gene in a genome. One speaks of gene families, and comparative genomic methods that allow this kind of input are called gene family-based. The most powerful-but also most complex-models avoid this preprocessing of the input data and instead integrate the family assignment within the comparative analysis. Such methods are called gene family-free. In this article, we study an intermediate approach between family-based and family-free genomic similarity measures. Introducing this simpler model, called gene connections, we focus on the combinatorial aspects of gene family-free genome comparison. While in most cases, the computational costs to the general family-free case are the same, we also find an instance where the gene connections model has lower complexity. Within the gene connections model, we define three variants of genomic similarity measures that have different expression powers. We give polynomial-time algorithms for two of them, while we show NP-hardness for the third, most powerful one. We also generalize the measures and algorithms to make them more robust against recent local disruptions in gene order. Our theoretical findings are supported by experimental results, proving the applicability and performance of our newly defined similarity measures.

  3. In situ optical sequencing and structure analysis of a trinucleotide repeat genome region by localization microscopy after specific COMBO-FISH nano-probing

    Science.gov (United States)

    Stuhlmüller, M.; Schwarz-Finsterle, J.; Fey, E.; Lux, J.; Bach, M.; Cremer, C.; Hinderhofer, K.; Hausmann, M.; Hildenbrand, G.

    2015-10-01

    Trinucleotide repeat expansions (like (CGG)n) of chromatin in the genome of cell nuclei can cause neurological disorders such as for example the Fragile-X syndrome. Until now the mechanisms are not clearly understood as to how these expansions develop during cell proliferation. Therefore in situ investigations of chromatin structures on the nanoscale are required to better understand supra-molecular mechanisms on the single cell level. By super-resolution localization microscopy (Spectral Position Determination Microscopy; SPDM) in combination with nano-probing using COMBO-FISH (COMBinatorial Oligonucleotide FISH), novel insights into the nano-architecture of the genome will become possible. The native spatial structure of trinucleotide repeat expansion genome regions was analysed and optical sequencing of repetitive units was performed within 3D-conserved nuclei using SPDM after COMBO-FISH. We analysed a (CGG)n-expansion region inside the 5' untranslated region of the FMR1 gene. The number of CGG repeats for a full mutation causing the Fragile-X syndrome was found and also verified by Southern blot. The FMR1 promotor region was similarly condensed like a centromeric region whereas the arrangement of the probes labelling the expansion region seemed to indicate a loop-like nano-structure. These results for the first time demonstrate that in situ chromatin structure measurements on the nanoscale are feasible. Due to further methodological progress it will become possible to estimate the state of trinucleotide repeat mutations in detail and to determine the associated chromatin strand structural changes on the single cell level. In general, the application of the described approach to any genome region will lead to new insights into genome nano-architecture and open new avenues for understanding mechanisms and their relevance in the development of heredity diseases.

  4. Comparisons of Copy Number, Genomic Structure, and Conserved Motifs for α-Amylase Genes from Barley, Rice, and Wheat

    Directory of Open Access Journals (Sweden)

    Qisen Zhang

    2017-10-01

    Full Text Available Barley is an important crop for the production of malt and beer. However, crops such as rice and wheat are rarely used for malting. α-amylase is the key enzyme that degrades starch during malting. In this study, we compared the genomic properties, gene copies, and conserved promoter motifs of α-amylase genes in barley, rice, and wheat. In all three crops, α-amylase consists of four subfamilies designated amy1, amy2, amy3, and amy4. In wheat and barley, members of amy1 and amy2 genes are localized on chromosomes 6 and 7, respectively. In rice, members of amy1 genes are found on chromosomes 1 and 2, and amy2 genes on chromosome 6. The barley genome has six amy1 members and three amy2 members. The wheat B genome contains four amy1 members and three amy2 members, while the rice genome has three amy1 members and one amy2 member. The B genome has mostly amy1 and amy2 members among the three wheat genomes. Amy1 promoters from all three crop genomes contain a GA-responsive complex consisting of a GA-responsive element (CAATAAA, pyrimidine box (CCTTTT and TATCCAT/C box. This study has shown that amy1 and amy2 from both wheat and barley have similar genomic properties, including exon/intron structures and GA-responsive elements on promoters, but these differ in rice. Like barley, wheat should have sufficient amy activity to degrade starch completely during malting. Other factors, such as high protein with haze issues and the lack of husk causing Lauting difficulty, may limit the use of wheat for brewing.

  5. An Emerging Tick-Borne Disease of Humans Is Caused by a Subset of Strains with Conserved Genome Structure

    Science.gov (United States)

    Barbet, Anthony F.; Al-Khedery, Basima; Stuen, Snorre; Granquist, Erik G.; Felsheim, Roderick F.; Munderloh, Ulrike G.

    2013-01-01

    The prevalence of tick-borne diseases is increasing worldwide. One such emerging disease is human anaplasmosis. The causative organism, Anaplasma phagocytophilum, is known to infect multiple animal species and cause human fatalities in the U.S., Europe and Asia. Although long known to infect ruminants, it is unclear why there are increasing numbers of human infections. We analyzed the genome sequences of strains infecting humans, animals and ticks from diverse geographic locations. Despite extensive variability amongst these strains, those infecting humans had conserved genome structure including the pfam01617 superfamily that encodes the major, neutralization-sensitive, surface antigen. These data provide potential targets to identify human-infective strains and have significance for understanding the selective pressures that lead to emergence of disease in new species. PMID:25437207

  6. Exploiting Genomic Resources for Efficient Conservation and Use of Chickpea, Groundnut, and Pigeonpea Collections for Crop Improvement

    Directory of Open Access Journals (Sweden)

    C. L. Laxmipathi Gowda

    2013-11-01

    Full Text Available Both chickpea ( L. and pigeonpea [ (L. Millsp.] are important dietary source of protein while groundnut ( L. is one of the major oil crops. Globally, approximately 1.1 million grain legume accessions are conserved in genebanks, of which the ICRISAT genebank holds 49,485 accessions of cultivated species and wild relatives of chickpea, pigeonpea, and groundnut from 133 countries. These genetic resources are reservoirs of many useful genes for present and future crop improvement programs. Representative subsets in the form of core and mini core collections have been used to identify trait-specific genetically diverse germplasm for use in breeding and genomic studies in these crops. Chickpea, groundnut, and pigeonpea have moved from “orphan” to “genomic resources rich crops.” The chickpea and pigeonpea genomes have been decoded, and the sequences of groundnut genome will soon be available. With the availability of these genomic resources, the germplasm curators, breeders, and molecular biologists will have abundant opportunities to enhance the efficiency of genebank operations, mine allelic variations in germplasm collection, identify genetically diverse germplasm with beneficial traits, broaden the cultigen’s genepool, and accelerate the cultivar development to address new challenges to production, particularly with respect to climate change and variability. Marker-assisted breeding approaches have already been initiated for some traits in chickpea and groundnut, which should lead to enhanced efficiency and efficacy of crop improvement. Resistance to some pests and diseases has been successfully transferred from wild relatives to cultivated species.

  7. Decoding the principles underlying the frequency of association with nucleoli for RNA polymerase III-transcribed genes in budding yeast.

    Science.gov (United States)

    Belagal, Praveen; Normand, Christophe; Shukla, Ashutosh; Wang, Renjie; Léger-Silvestre, Isabelle; Dez, Christophe; Bhargava, Purnima; Gadal, Olivier

    2016-10-15

    The association of RNA polymerase III (Pol III)-transcribed genes with nucleoli seems to be an evolutionarily conserved property of the spatial organization of eukaryotic genomes. However, recent studies of global chromosome architecture in budding yeast have challenged this view. We used live-cell imaging to determine the intranuclear positions of 13 Pol III-transcribed genes. The frequency of association with nucleolus and nuclear periphery depends on linear genomic distance from the tethering elements-centromeres or telomeres. Releasing the hold of the tethering elements by inactivating centromere attachment to the spindle pole body or changing the position of ribosomal DNA arrays resulted in the association of Pol III-transcribed genes with nucleoli. Conversely, ectopic insertion of a Pol III-transcribed gene in the vicinity of a centromere prevented its association with nucleolus. Pol III-dependent transcription was independent of the intranuclear position of the gene, but the nucleolar recruitment of Pol III-transcribed genes required active transcription. We conclude that the association of Pol III-transcribed genes with the nucleolus, when permitted by global chromosome architecture, provides nucleolar and/or nuclear peripheral anchoring points contributing locally to intranuclear chromosome organization. © 2016 Belagal et al. This article is distributed by The American Society for Cell Biology under license from the author(s). Two months after publication it is available to the public under an Attribution–Noncommercial–Share Alike 3.0 Unported Creative Commons License (http://creativecommons.org/licenses/by-nc-sa/3.0).

  8. PHYLUCE is a software package for the analysis of conserved genomic loci.

    Science.gov (United States)

    Faircloth, Brant C

    2016-03-01

    Targeted enrichment of conserved and ultraconserved genomic elements allows universal collection of phylogenomic data from hundreds of species at multiple time scales ( 300 Ma). Prior to downstream inference, data from these types of targeted enrichment studies must undergo preprocessing to assemble contigs from sequence data; identify targeted, enriched loci from the off-target background data; align enriched contigs representing conserved loci to one another; and prepare and manipulate these alignments for subsequent phylogenomic inference. PHYLUCE is an efficient and easy-to-install software package that accomplishes these tasks across hundreds of taxa and thousands of enriched loci. PHYLUCE is written for Python 2.7. PHYLUCE is supported on OSX and Linux (RedHat/CentOS) operating systems. PHYLUCE source code is distributed under a BSD-style license from https://www.github.com/faircloth-lab/phyluce/ PHYLUCE is also available as a package (https://binstar.org/faircloth-lab/phyluce) for the Anaconda Python distribution that installs all dependencies, and users can request a PHYLUCE instance on iPlant Atmosphere (tag: phyluce). The software manual and a tutorial are available from http://phyluce.readthedocs.org/en/latest/ and test data are available from doi: 10.6084/m9.figshare.1284521. brant@faircloth-lab.org Supplementary data are available at Bioinformatics online. © The Author 2015. Published by Oxford University Press. All rights reserved. For Permissions, please e-mail: journals.permissions@oup.com.

  9. DNA deformability changes of single base pair mutants within CDE binding sites in S. Cerevisiae centromere DNA correlate with measured chromosomal loss rates and CDE binding site symmetries

    Directory of Open Access Journals (Sweden)

    Marx Kenneth A

    2006-03-01

    Full Text Available Abstract Background The centromeres in yeast (S. cerevisiae are organized by short DNA sequences (125 bp on each chromosome consisting of 2 conserved elements: CDEI and CDEIII spaced by a CDEII region. CDEI and CDEIII are critical sequence specific protein binding sites necessary for correct centromere formation and following assembly with proteins, are positioned near each other on a specialized nucleosome. Hegemann et al. BioEssays 1993, 15: 451–460 reported single base DNA mutants within the critical CDEI and CDEIII binding sites on the centromere of chromosome 6 and quantitated centromere loss of function, which they measured as loss rates for the different chromosome 6 mutants during cell division. Olson et al. Proc Natl Acad Sci USA 1998, 95: 11163–11168 reported the use of protein-DNA crystallography data to produce a DNA dinucleotide protein deformability energetic scale (PD-scale that describes local DNA deformability by sequence specific binding proteins. We have used the PD-scale to investigate the DNA sequence dependence of the yeast chromosome 6 mutants' loss rate data. Each single base mutant changes 2 PD-scale values at that changed base position relative to the wild type. In this study, we have utilized these mutants to demonstrate a correlation between the change in DNA deformability of the CDEI and CDEIII core sites and the overall experimentally measured chromosome loss rates of the chromosome 6 mutants. Results In the CDE I and CDEIII core binding regions an increase in the magnitude of change in deformability of chromosome 6 single base mutants with respect to the wild type correlates to an increase in the measured chromosome loss rate. These correlations were found to be significant relative to 105 Monte Carlo randomizations of the dinucleotide PD-scale applied to the same calculation. A net loss of deformability also tends to increase the loss rate. Binding site position specific, 4 data-point correlations were also

  10. The Chthonomonas calidirosea Genome Is Highly Conserved across Geographic Locations and Distinct Chemical and Microbial Environments in New Zealand's Taupō Volcanic Zone.

    Science.gov (United States)

    Lee, Kevin C; Stott, Matthew B; Dunfield, Peter F; Huttenhower, Curtis; McDonald, Ian R; Morgan, Xochitl C

    2016-06-15

    Chthonomonas calidirosea T49(T) is a low-abundance, carbohydrate-scavenging, and thermophilic soil bacterium with a seemingly disorganized genome. We hypothesized that the C. calidirosea genome would be highly responsive to local selection pressure, resulting in the divergence of its genomic content, genome organization, and carbohydrate utilization phenotype across environments. We tested this hypothesis by sequencing the genomes of four C. calidirosea isolates obtained from four separate geothermal fields in the Taupō Volcanic Zone, New Zealand. For each isolation site, we measured physicochemical attributes and defined the associated microbial community by 16S rRNA gene sequencing. Despite their ecological and geographical isolation, the genome sequences showed low divergence (maximum, 1.17%). Isolate-specific variations included single-nucleotide polymorphisms (SNPs), restriction-modification systems, and mobile elements but few major deletions and no major rearrangements. The 50-fold variation in C. calidirosea relative abundance among the four sites correlated with site environmental characteristics but not with differences in genomic content. Conversely, the carbohydrate utilization profiles of the C. calidirosea isolates corresponded to the inferred isolate phylogenies, which only partially paralleled the geographical relationships among the sample sites. Genomic sequence conservation does not entirely parallel geographic distance, suggesting that stochastic dispersal and localized extinction, which allow for rapid population homogenization with little restriction by geographical barriers, are possible mechanisms of C. calidirosea distribution. This dispersal and extinction mechanism is likely not limited to C. calidirosea but may shape the populations and genomes of many other low-abundance free-living taxa. This study compares the genomic sequence variations and metabolisms of four strains of Chthonomonas calidirosea, a rare thermophilic bacterium from

  11. Genomic tools for evolution and conservation in the chimpanzee: Pan troglodytes ellioti is a genetically distinct population.

    Directory of Open Access Journals (Sweden)

    Rory Bowden

    Full Text Available In spite of its evolutionary significance and conservation importance, the population structure of the common chimpanzee, Pan troglodytes, is still poorly understood. An issue of particular controversy is whether the proposed fourth subspecies of chimpanzee, Pan troglodytes ellioti, from parts of Nigeria and Cameroon, is genetically distinct. Although modern high-throughput SNP genotyping has had a major impact on our understanding of human population structure and demographic history, its application to ecological, demographic, or conservation questions in non-human species has been extremely limited. Here we apply these tools to chimpanzee population structure, using ∼700 autosomal SNPs derived from chimpanzee genomic data and a further ∼100 SNPs from targeted re-sequencing. We demonstrate conclusively the existence of P. t. ellioti as a genetically distinct subgroup. We show that there is clear differentiation between the verus, troglodytes, and ellioti populations at the SNP and haplotype level, on a scale that is greater than that separating continental human populations. Further, we show that only a small set of SNPs (10-20 is needed to successfully assign individuals to these populations. Tellingly, use of only mitochondrial DNA variation to classify individuals is erroneous in 4 of 54 cases, reinforcing the dangers of basing demographic inference on a single locus and implying that the demographic history of the species is more complicated than that suggested analyses based solely on mtDNA. In this study we demonstrate the feasibility of developing economical and robust tests of individual chimpanzee origin as well as in-depth studies of population structure. These findings have important implications for conservation strategies and our understanding of the evolution of chimpanzees. They also act as a proof-of-principle for the use of cheap high-throughput genomic methods for ecological questions.

  12. Asymmetrical distribution of non-conserved regulatory sequences at PHOX2B is reflected at the ENCODE loci and illuminates a possible genome-wide trend

    Directory of Open Access Journals (Sweden)

    McCallion Andrew S

    2009-01-01

    Full Text Available Abstract Background Transcriptional regulatory elements are central to development and interspecific phenotypic variation. Current regulatory element prediction tools rely heavily upon conservation for prediction of putative elements. Recent in vitro observations from the ENCODE project combined with in vivo analyses at the zebrafish phox2b locus suggests that a significant fraction of regulatory elements may fall below commonly applied metrics of conservation. We propose to explore these observations in vivo at the human PHOX2B locus, and also evaluate the potential evidence for genome-wide applicability of these observations through a novel analysis of extant data. Results Transposon-based transgenic analysis utilizing a tiling path proximal to human PHOX2B in zebrafish recapitulates the observations at the zebrafish phox2b locus of both conserved and non-conserved regulatory elements. Analysis of human sequences conserved with previously identified zebrafish phox2b regulatory elements demonstrates that the orthologous sequences exhibit overlapping regulatory control. Additionally, analysis of non-conserved sequences scattered over 135 kb 5' to PHOX2B, provides evidence of non-conserved regulatory elements positively biased with close proximity to the gene. Furthermore, we provide a novel analysis of data from the ENCODE project, finding a non-uniform distribution of regulatory elements consistent with our in vivo observations at PHOX2B. These observations remain largely unchanged when one accounts for the sequence repeat content of the assayed intervals, when the intervals are sub-classified by biological role (developmental versus non-developmental, or by gene density (gene desert versus non-gene desert. Conclusion While regulatory elements frequently display evidence of evolutionary conservation, a fraction appears to be undetected by current metrics of conservation. In vivo observations at the PHOX2B locus, supported by our analyses of in

  13. Genes involved in complex adaptive processes tend to have highly conserved upstream regions in mammalian genomes

    Directory of Open Access Journals (Sweden)

    Kohane Isaac

    2005-11-01

    Full Text Available Abstract Background Recent advances in genome sequencing suggest a remarkable conservation in gene content of mammalian organisms. The similarity in gene repertoire present in different organisms has increased interest in studying regulatory mechanisms of gene expression aimed at elucidating the differences in phenotypes. In particular, a proximal promoter region contains a large number of regulatory elements that control the expression of its downstream gene. Although many studies have focused on identification of these elements, a broader picture on the complexity of transcriptional regulation of different biological processes has not been addressed in mammals. The regulatory complexity may strongly correlate with gene function, as different evolutionary forces must act on the regulatory systems under different biological conditions. We investigate this hypothesis by comparing the conservation of promoters upstream of genes classified in different functional categories. Results By conducting a rank correlation analysis between functional annotation and upstream sequence alignment scores obtained by human-mouse and human-dog comparison, we found a significantly greater conservation of the upstream sequence of genes involved in development, cell communication, neural functions and signaling processes than those involved in more basic processes shared with unicellular organisms such as metabolism and ribosomal function. This observation persists after controlling for G+C content. Considering conservation as a functional signature, we hypothesize a higher density of cis-regulatory elements upstream of genes participating in complex and adaptive processes. Conclusion We identified a class of functions that are associated with either high or low promoter conservation in mammals. We detected a significant tendency that points to complex and adaptive processes were associated with higher promoter conservation, despite the fact that they have emerged

  14. Generation of an ICF syndrome model by efficient genome editing of human induced pluripotent stem cells using the CRISPR system.

    Science.gov (United States)

    Horii, Takuro; Tamura, Daiki; Morita, Sumiyo; Kimura, Mika; Hatada, Izuho

    2013-09-30

    Genome manipulation of human induced pluripotent stem (iPS) cells is essential to achieve their full potential as tools for regenerative medicine. To date, however, gene targeting in human pluripotent stem cells (hPSCs) has proven to be extremely difficult. Recently, an efficient genome manipulation technology using the RNA-guided DNase Cas9, the clustered regularly interspaced short palindromic repeats (CRISPR) system, has been developed. Here we report the efficient generation of an iPS cell model for immunodeficiency, centromeric region instability, facial anomalies syndrome (ICF) syndrome using the CRISPR system. We obtained iPS cells with mutations in both alleles of DNA methyltransferase 3B (DNMT3B) in 63% of transfected clones. Our data suggest that the CRISPR system is highly efficient and useful for genome engineering of human iPS cells.

  15. Annotating non-coding regions of the genome.

    Science.gov (United States)

    Alexander, Roger P; Fang, Gang; Rozowsky, Joel; Snyder, Michael; Gerstein, Mark B

    2010-08-01

    Most of the human genome consists of non-protein-coding DNA. Recently, progress has been made in annotating these non-coding regions through the interpretation of functional genomics experiments and comparative sequence analysis. One can conceptualize functional genomics analysis as involving a sequence of steps: turning the output of an experiment into a 'signal' at each base pair of the genome; smoothing this signal and segmenting it into small blocks of initial annotation; and then clustering these small blocks into larger derived annotations and networks. Finally, one can relate functional genomics annotations to conserved units and measures of conservation derived from comparative sequence analysis.

  16. Autoantibodies directed to centromere protein F in a patient with BRCA1 gene mutation

    OpenAIRE

    Moghaddas, Fiona; Joshua, Fredrick; Taylor, Roberta; Fritzler, Marvin J.; Toh, Ban Hock

    2016-01-01

    Background Autoantibodies directed to centromere protein F were first reported in 1993 and their association with malignancy has been well documented. Case We present the case of a 48-year-old Caucasian female with a BRCA1 gene mutation associated with bilateral breast cancer. Antinuclear autoantibody immunofluorescence performed for workup of possible inflammatory arthropathy showed a high titre cell cycle related nuclear speckled pattern, with subsequent confirmation by addressable laser be...

  17. Annotation of the Clostridium Acetobutylicum Genome

    Energy Technology Data Exchange (ETDEWEB)

    Daly, M. J.

    2004-06-09

    The genome sequence of the solvent producing bacterium Clostridium acetobutylicum ATCC824, has been determined by the shotgun approach. The genome consists of a 3.94 Mb chromosome and a 192 kb megaplasmid that contains the majority of genes responsible for solvent production. Comparison of C. acetobutylicum to Bacillus subtilis reveals significant local conservation of gene order, which has not been seen in comparisons of other genomes with similar, or, in some cases, closer, phylogenetic proximity. This conservation allows the prediction of many previously undetected operons in both bacteria.

  18. Conservation of HIV-1 T cell epitopes across time and clades

    DEFF Research Database (Denmark)

    Levitz, Lauren; Koita, Ousmane A; Sangare, Kotou

    2012-01-01

    HIV genomic sequence variability has complicated efforts to generate an effective globally relevant vaccine. Regions of the viral genome conserved in sequence and across time may represent the "Achilles' heel" of HIV. In this study, highly conserved T-cell epitopes were selected using immunoinfor...

  19. Conserved microstructure of the Brassica B Genome of Brassica nigra in relation to homologous regions of Arabidopsis thaliana, B. rapa and B. oleracea

    Science.gov (United States)

    2013-01-01

    Background The Brassica B genome is known to carry several important traits, yet there has been limited analyses of its underlying genome structure, especially in comparison to the closely related A and C genomes. A bacterial artificial chromosome (BAC) library of Brassica nigra was developed and screened with 17 genes from a 222 kb region of A. thaliana that had been well characterised in both the Brassica A and C genomes. Results Fingerprinting of 483 apparently non-redundant clones defined physical contigs for the corresponding regions in B. nigra. The target region is duplicated in A. thaliana and six homologous contigs were found in B. nigra resulting from the whole genome triplication event shared by the Brassiceae tribe. BACs representative of each region were sequenced to elucidate the level of microscale rearrangements across the Brassica species divide. Conclusions Although the B genome species separated from the A/C lineage some 6 Mya, comparisons between the three paleopolyploid Brassica genomes revealed extensive conservation of gene content and sequence identity. The level of fractionation or gene loss varied across genomes and genomic regions; however, the greatest loss of genes was observed to be common to all three genomes. One large-scale chromosomal rearrangement differentiated the B genome suggesting such events could contribute to the lack of recombination observed between B genome species and those of the closely related A/C lineage. PMID:23586706

  20. Comparative genomics reveals conservation of filaggrin and loss of caspase-14 in dolphins.

    Science.gov (United States)

    Strasser, Bettina; Mlitz, Veronika; Fischer, Heinz; Tschachler, Erwin; Eckhart, Leopold

    2015-05-01

    The expression of filaggrin and its stepwise proteolytic degradation are critical events in the terminal differentiation of epidermal keratinocytes and in the formation of the skin barrier to the environment. Here, we investigated whether the evolutionary transition from a terrestrial to a fully aquatic lifestyle of cetaceans, that is dolphins and whales, has been associated with changes in genes encoding filaggrin and proteins involved in the processing of filaggrin. We used comparative genomics, PCRs and re-sequencing of gene segments to screen for the presence and integrity of genes coding for filaggrin and proteases implicated in the maturation of (pro)filaggrin. Filaggrin has been conserved in dolphins (bottlenose dolphin, orca and baiji) but has been lost in whales (sperm whale and minke whale). All other S100 fused-type genes have been lost in cetaceans. Among filaggrin-processing proteases, aspartic peptidase retroviral-like 1 (ASPRV1), also known as saspase, has been conserved, whereas caspase-14 has been lost in all cetaceans investigated. In conclusion, our results suggest that filaggrin is dispensable for the acquisition of fully aquatic lifestyles of whales, whereas it appears to confer an evolutionary advantage to dolphins. The discordant evolution of filaggrin, saspase and caspase-14 in cetaceans indicates that the biological roles of these proteins are not strictly interdependent. © 2015 The Authors. Experimental Dermatology Published by John Wiley & Sons Ltd.

  1. Analysis of the conservation of synteny between Fugu and human chromosome 12

    Directory of Open Access Journals (Sweden)

    Koop Ben F

    2003-07-01

    Full Text Available Abstract Background The pufferfish Fugu rubripes (Fugu with its compact genome is increasingly recognized as an important vertebrate model for comparative genomic studies. In particular, large regions of conserved synteny between human and Fugu genomes indicate its utility to identify disease-causing genes. The human chromosome 12p12 is frequently deleted in various hematological malignancies and solid tumors, but the actual tumor suppressor gene remains unidentified. Results We investigated approximately 200 kb of the genomic region surrounding the ETV6 locus in Fugu (fETV6 in order to find conserved functional features, such as genes or regulatory regions, that could give insight into the nature of the genes targeted by deletions in human cancer cells. Seven genes were identified near the fETV6 locus. We found that the synteny with human chromosome 12 was conserved, but extensive genomic rearrangements occurred between the Fugu and human ETV6 loci. Conclusion This comparative analysis led to the identification of previously uncharacterized genes in the human genome and some potentially important regulatory sequences as well. This is a good indication that the analysis of the compact Fugu genome will be valuable to identify functional features that have been conserved throughout the evolution of vertebrates.

  2. Generation of an ICF Syndrome Model by Efficient Genome Editing of Human Induced Pluripotent Stem Cells Using the CRISPR System

    Directory of Open Access Journals (Sweden)

    Izuho Hatada

    2013-09-01

    Full Text Available Genome manipulation of human induced pluripotent stem (iPS cells is essential to achieve their full potential as tools for regenerative medicine. To date, however, gene targeting in human pluripotent stem cells (hPSCs has proven to be extremely difficult. Recently, an efficient genome manipulation technology using the RNA-guided DNase Cas9, the clustered regularly interspaced short palindromic repeats (CRISPR system, has been developed. Here we report the efficient generation of an iPS cell model for immunodeficiency, centromeric region instability, facial anomalies syndrome (ICF syndrome using the CRISPR system. We obtained iPS cells with mutations in both alleles of DNA methyltransferase 3B (DNMT3B in 63% of transfected clones. Our data suggest that the CRISPR system is highly efficient and useful for genome engineering of human iPS cells.

  3. Genomic Variation in Natural Populations of Drosophila melanogaster

    Science.gov (United States)

    Langley, Charles H.; Stevens, Kristian; Cardeno, Charis; Lee, Yuh Chwen G.; Schrider, Daniel R.; Pool, John E.; Langley, Sasha A.; Suarez, Charlyn; Corbett-Detig, Russell B.; Kolaczkowski, Bryan; Fang, Shu; Nista, Phillip M.; Holloway, Alisha K.; Kern, Andrew D.; Dewey, Colin N.; Song, Yun S.; Hahn, Matthew W.; Begun, David J.

    2012-01-01

    This report of independent genome sequences of two natural populations of Drosophila melanogaster (37 from North America and 6 from Africa) provides unique insight into forces shaping genomic polymorphism and divergence. Evidence of interactions between natural selection and genetic linkage is abundant not only in centromere- and telomere-proximal regions, but also throughout the euchromatic arms. Linkage disequilibrium, which decays within 1 kbp, exhibits a strong bias toward coupling of the more frequent alleles and provides a high-resolution map of recombination rate. The juxtaposition of population genetics statistics in small genomic windows with gene structures and chromatin states yields a rich, high-resolution annotation, including the following: (1) 5′- and 3′-UTRs are enriched for regions of reduced polymorphism relative to lineage-specific divergence; (2) exons overlap with windows of excess relative polymorphism; (3) epigenetic marks associated with active transcription initiation sites overlap with regions of reduced relative polymorphism and relatively reduced estimates of the rate of recombination; (4) the rate of adaptive nonsynonymous fixation increases with the rate of crossing over per base pair; and (5) both duplications and deletions are enriched near origins of replication and their density correlates negatively with the rate of crossing over. Available demographic models of X and autosome descent cannot account for the increased divergence on the X and loss of diversity associated with the out-of-Africa migration. Comparison of the variation among these genomes to variation among genomes from D. simulans suggests that many targets of directional selection are shared between these species. PMID:22673804

  4. A Somatically Acquired Enhancer of the Androgen Receptor Is a Noncoding Driver in Advanced Prostate Cancer. | Office of Cancer Genomics

    Science.gov (United States)

    Increased androgen receptor (AR) activity drives therapeutic resistance in advanced prostate cancer. The most common resistance mechanism is amplification of this locus presumably targeting the AR gene. Here, we identify and characterize a somatically acquired AR enhancer located 650 kb centromeric to the AR. Systematic perturbation of this enhancer using genome editing decreased proliferation by suppressing AR levels. Insertion of an additional copy of this region sufficed to increase proliferation under low androgen conditions and to decrease sensitivity to enzalutamide.

  5. Genome re-assignment of Arachis trinitensis (Sect. Arachis, Leguminosae and its implications for the genetic origin of cultivated peanut

    Directory of Open Access Journals (Sweden)

    Germán Robledo

    2010-01-01

    Full Text Available The karyotype structure of Arachis trinitensis was studied by conventional Feulgen staining, CMA/DAPI banding and rDNA loci detection by fluorescence in situ hybridization (FISH in order to establish its genome status and test the hypothesis that this species is a genome donor of cultivated peanut. Conventional staining revealed that the karyotype lacked the small "A chromosomes" characteristic of the A genome. In agreement with this, chromosomal banding showed that none of the chromosomes had the large centromeric bands expected for A chromosomes. FISH revealed one pair each of 5S and 45S rDNA loci, located in different medium-sized metacentric chromosomes. Collectively, these results suggest that A. trinitensis should be removed from the A genome and be considered as a B or non-A genome species. The pattern of heterochromatic bands and rDNA loci of A. trinitensis differ markedly from any of the complements of A. hypogaea, suggesting that the former species is unlikely to be one of the wild diploid progenitors of the latter.

  6. Genome re-assignment of Arachis trinitensis (Sect. Arachis, Leguminosae) and its implications for the genetic origin of cultivated peanut

    Science.gov (United States)

    2010-01-01

    The karyotype structure of Arachis trinitensis was studied by conventional Feulgen staining, CMA/DAPI banding and rDNA loci detection by fluorescence in situ hybridization (FISH) in order to establish its genome status and test the hypothesis that this species is a genome donor of cultivated peanut. Conventional staining revealed that the karyotype lacked the small “A chromosomes” characteristic of the A genome. In agreement with this, chromosomal banding showed that none of the chromosomes had the large centromeric bands expected for A chromosomes. FISH revealed one pair each of 5S and 45S rDNA loci, located in different medium-sized metacentric chromosomes. Collectively, these results suggest that A. trinitensis should be removed from the A genome and be considered as a B or non-A genome species. The pattern of heterochromatic bands and rDNA loci of A. trinitensis differ markedly from any of the complements of A. hypogaea, suggesting that the former species is unlikely to be one of the wild diploid progenitors of the latter. PMID:21637581

  7. Conserved PCR primer set designing for closely-related species to complete mitochondrial genome sequencing using a sliding window-based PSO algorithm.

    Directory of Open Access Journals (Sweden)

    Cheng-Hong Yang

    Full Text Available BACKGROUND: Complete mitochondrial (mt genome sequencing is becoming increasingly common for phylogenetic reconstruction and as a model for genome evolution. For long template sequencing, i.e., like the entire mtDNA, it is essential to design primers for Polymerase Chain Reaction (PCR amplicons which are partly overlapping each other. The presented chromosome walking strategy provides the overlapping design to solve the problem for unreliable sequencing data at the 5' end and provides the effective sequencing. However, current algorithms and tools are mostly focused on the primer design for a local region in the genomic sequence. Accordingly, it is still challenging to provide the primer sets for the entire mtDNA. METHODOLOGY/PRINCIPAL FINDINGS: The purpose of this study is to develop an integrated primer design algorithm for entire mt genome in general, and for the common primer sets for closely-related species in particular. We introduce ClustalW to generate the multiple sequence alignment needed to find the conserved sequences in closely-related species. These conserved sequences are suitable for designing the common primers for the entire mtDNA. Using a heuristic algorithm particle swarm optimization (PSO, all the designed primers were computationally validated to fit the common primer design constraints, such as the melting temperature, primer length and GC content, PCR product length, secondary structure, specificity, and terminal limitation. The overlap requirement for PCR amplicons in the entire mtDNA is satisfied by defining the overlapping region with the sliding window technology. Finally, primer sets were designed within the overlapping region. The primer sets for the entire mtDNA sequences were successfully demonstrated in the example of two closely-related fish species. The pseudo code for the primer design algorithm is provided. CONCLUSIONS/SIGNIFICANCE: In conclusion, it can be said that our proposed sliding window-based PSO

  8. Insights into bilaterian evolution from three spiralian genomes

    Energy Technology Data Exchange (ETDEWEB)

    Simakov, Oleg; Marletaz, Ferdinand; Cho, Sung-Jin; Edsinger-Gonzales, Eric; Havlak, Paul; Hellsten, Uffe; Kuo, Dian-Han; Larsson, Tomas; Lv, Jie; Arendt, Detlev; Savage, Robert; Osoegawa, Kazutoyo; de Jong, Pieter; Grimwood, Jane; Chapman, Jarrod A.; Shapiro, Harris; Otillar, Robert P.; Terry, Astrid Y.; Boore, Jeffrey L.; Grigoriev, Igor V.; Lindberg, David R.; Seaver, Elaine C.; Weisblat, David A.; Putnam, Nicholas H.; Rokhsar, Daniel S.; Aerts, Andrea

    2012-01-07

    Current genomic perspectives on animal diversity neglect two prominent phyla, the molluscs and annelids, that together account for nearly one-third of known marine species and are important both ecologically and as experimental systems in classical embryology1, 2, 3. Here we describe the draft genomes of the owl limpet (Lottia gigantea), a marine polychaete (Capitella teleta) and a freshwater leech (Helobdella robusta), and compare them with other animal genomes to investigate the origin and diversification of bilaterians from a genomic perspective. We find that the genome organization, gene structure and functional content of these species are more similar to those of some invertebrate deuterostome genomes (for example, amphioxus and sea urchin) than those of other protostomes that have been sequenced to date (flies, nematodes and flatworms). The conservation of these genomic features enables us to expand the inventory of genes present in the last common bilaterian ancestor, establish the tripartite diversification of bilaterians using multiple genomic characteristics and identify ancient conserved long- and short-range genetic linkages across metazoans. Superimposed on this broadly conserved pan-bilaterian background we find examples of lineage-specific genome evolution, including varying rates of rearrangement, intron gain and loss, expansions and contractions of gene families, and the evolution of clade-specific genes that produce the unique content of each genome.

  9. Genome-wide recombination rate variation in a recombination map of cotton.

    Science.gov (United States)

    Shen, Chao; Li, Ximei; Zhang, Ruiting; Lin, Zhongxu

    2017-01-01

    Recombination is crucial for genetic evolution, which not only provides new allele combinations but also influences the biological evolution and efficacy of natural selection. However, recombination variation is not well understood outside of the complex species' genomes, and it is particularly unclear in Gossypium. Cotton is the most important natural fibre crop and the second largest oil-seed crop. Here, we found that the genetic and physical maps distances did not have a simple linear relationship. Recombination rates were unevenly distributed throughout the cotton genome, which showed marked changes along the chromosome lengths and recombination was completely suppressed in the centromeric regions. Recombination rates significantly varied between A-subgenome (At) (range = 1.60 to 3.26 centimorgan/megabase [cM/Mb]) and D-subgenome (Dt) (range = 2.17 to 4.97 cM/Mb), which explained why the genetic maps of At and Dt are similar but the physical map of Dt is only half that of At. The translocation regions between A02 and A03 and between A04 and A05, and the inversion regions on A10, D10, A07 and D07 indicated relatively high recombination rates in the distal regions of the chromosomes. Recombination rates were positively correlated with the densities of genes, markers and the distance from the centromere, and negatively correlated with transposable elements (TEs). The gene ontology (GO) categories showed that genes in high recombination regions may tend to response to environmental stimuli, and genes in low recombination regions are related to mitosis and meiosis, which suggested that they may provide the primary driving force in adaptive evolution and assure the stability of basic cell cycle in a rapidly changing environment. Global knowledge of recombination rates will facilitate genetics and breeding in cotton.

  10. Assessing Telomere Length Using Surface Enhanced Raman Scattering

    Science.gov (United States)

    Zong, Shenfei; Wang, Zhuyuan; Chen, Hui; Cui, Yiping

    2014-11-01

    Telomere length can provide valuable insight into telomeres and telomerase related diseases, including cancer. Here, we present a brand-new optical telomere length measurement protocol using surface enhanced Raman scattering (SERS). In this protocol, two single strand DNA are used as SERS probes. They are labeled with two different Raman molecules and can specifically hybridize with telomeres and centromere, respectively. First, genome DNA is extracted from cells. Then the telomere and centromere SERS probes are added into the genome DNA. After hybridization with genome DNA, excess SERS probes are removed by magnetic capturing nanoparticles. Finally, the genome DNA with SERS probes attached is dropped onto a SERS substrate and subjected to SERS measurement. Longer telomeres result in more attached telomere probes, thus a stronger SERS signal. Consequently, SERS signal can be used as an indicator of telomere length. Centromere is used as the inner control. By calibrating the SERS intensity of telomere probe with that of the centromere probe, SERS based telomere measurement is realized. This protocol does not require polymerase chain reaction (PCR) or electrophoresis procedures, which greatly simplifies the detection process. We anticipate that this easy-operation and cost-effective protocol is a fine alternative for the assessment of telomere length.

  11. The Perennial Ryegrass GenomeZipper – Targeted Use of Genome Resources for Comparative Grass Genomics

    DEFF Research Database (Denmark)

    Pfeiffer, Matthias; Martis, Mihaela; Asp, Torben

    2013-01-01

    (Lolium perenne) genome on the basis of conserved synteny to barley (Hordeum vulgare) and the model grass genome Brachypodium (Brachypodium distachyon) as well as rice (Oryza sativa) and sorghum (Sorghum bicolor). A transcriptome-based genetic linkage map of perennial ryegrass served as a scaffold......Whole-genome sequences established for model and major crop species constitute a key resource for advanced genomic research. For outbreeding forage and turf grass species like ryegrasses (Lolium spp.), such resources have yet to be developed. Here, we present a model of the perennial ryegrass...... to establish the chromosomal arrangement of syntenic genes from model grass species. This scaffold revealed a high degree of synteny and macrocollinearity and was then utilized to anchor a collection of perennial ryegrass genes in silico to their predicted genome positions. This resulted in the unambiguous...

  12. Genome-wide identification of conserved microRNA and their response to drought stress in Dongxiang wild rice (Oryza rufipogon Griff.).

    Science.gov (United States)

    Zhang, Fantao; Luo, Xiangdong; Zhou, Yi; Xie, Jiankun

    2016-04-01

    To identify drought stress-responsive conserved microRNA (miRNA) from Dongxiang wild rice (Oryza rufipogon Griff., DXWR) on a genome-wide scale, high-throughput sequencing technology was used to sequence libraries of DXWR samples, treated with and without drought stress. 505 conserved miRNAs corresponding to 215 families were identified. 17 were significantly down-regulated and 16 were up-regulated under drought stress. Stem-loop qRT-PCR revealed the same expression patterns as high-throughput sequencing, suggesting the accuracy of the sequencing result was high. Potential target genes of the drought-responsive miRNA were predicted to be involved in diverse biological processes. Furthermore, 16 miRNA families were first identified to be involved in drought stress response from plants. These results present a comprehensive view of the conserved miRNA and their expression patterns under drought stress for DXWR, which will provide valuable information and sequence resources for future basis studies.

  13. Whole-genome sequence of the Tibetan frog Nanorana parkeri and the comparative evolution of tetrapod genomes.

    Science.gov (United States)

    Sun, Yan-Bo; Xiong, Zi-Jun; Xiang, Xue-Yan; Liu, Shi-Ping; Zhou, Wei-Wei; Tu, Xiao-Long; Zhong, Li; Wang, Lu; Wu, Dong-Dong; Zhang, Bao-Lin; Zhu, Chun-Ling; Yang, Min-Min; Chen, Hong-Man; Li, Fang; Zhou, Long; Feng, Shao-Hong; Huang, Chao; Zhang, Guo-Jie; Irwin, David; Hillis, David M; Murphy, Robert W; Yang, Huan-Ming; Che, Jing; Wang, Jun; Zhang, Ya-Ping

    2015-03-17

    The development of efficient sequencing techniques has resulted in large numbers of genomes being available for evolutionary studies. However, only one genome is available for all amphibians, that of Xenopus tropicalis, which is distantly related from the majority of frogs. More than 96% of frogs belong to the Neobatrachia, and no genome exists for this group. This dearth of amphibian genomes greatly restricts genomic studies of amphibians and, more generally, our understanding of tetrapod genome evolution. To fill this gap, we provide the de novo genome of a Tibetan Plateau frog, Nanorana parkeri, and compare it to that of X. tropicalis and other vertebrates. This genome encodes more than 20,000 protein-coding genes, a number similar to that of Xenopus. Although the genome size of Nanorana is considerably larger than that of Xenopus (2.3 vs. 1.5 Gb), most of the difference is due to the respective number of transposable elements in the two genomes. The two frogs exhibit considerable conserved whole-genome synteny despite having diverged approximately 266 Ma, indicating a slow rate of DNA structural evolution in anurans. Multigenome synteny blocks further show that amphibians have fewer interchromosomal rearrangements than mammals but have a comparable rate of intrachromosomal rearrangements. Our analysis also identifies 11 Mb of anuran-specific highly conserved elements that will be useful for comparative genomic analyses of frogs. The Nanorana genome offers an improved understanding of evolution of tetrapod genomes and also provides a genomic reference for other evolutionary studies.

  14. Casein Kinase 1 Coordinates Cohesin Cleavage, Gametogenesis, and Exit from M Phase in Meiosis II.

    Science.gov (United States)

    Argüello-Miranda, Orlando; Zagoriy, Ievgeniia; Mengoli, Valentina; Rojas, Julie; Jonak, Katarzyna; Oz, Tugce; Graf, Peter; Zachariae, Wolfgang

    2017-01-09

    Meiosis consists of DNA replication followed by two consecutive nuclear divisions and gametogenesis or spore formation. While meiosis I has been studied extensively, less is known about the regulation of meiosis II. Here we show that Hrr25, the conserved casein kinase 1δ of budding yeast, links three mutually independent key processes of meiosis II. First, Hrr25 induces nuclear division by priming centromeric cohesin for cleavage by separase. Hrr25 simultaneously phosphorylates Rec8, the cleavable subunit of cohesin, and removes from centromeres the cohesin protector composed of shugoshin and the phosphatase PP2A. Second, Hrr25 initiates the sporulation program by inducing the synthesis of membranes that engulf the emerging nuclei at anaphase II. Third, Hrr25 mediates exit from meiosis II by activating pathways that trigger the destruction of M-phase-promoting kinases. Thus, Hrr25 synchronizes formation of the single-copy genome with gamete differentiation and termination of meiosis. Copyright © 2017 Elsevier Inc. All rights reserved.

  15. Expression and genomic analysis of midasin, a novel and highly conserved AAA protein distantly related to dynein

    Directory of Open Access Journals (Sweden)

    Gibbons I R

    2002-07-01

    Full Text Available Abstract Background The largest open reading frame in the Saccharomyces genome encodes midasin (MDN1p, YLR106p, an AAA ATPase of 560 kDa that is essential for cell viability. Orthologs of midasin have been identified in the genome projects for Drosophila, Arabidopsis, and Schizosaccharomyces pombe. Results Midasin is present as a single-copy gene encoding a well-conserved protein of ~600 kDa in all eukaryotes for which data are available. In humans, the gene maps to 6q15 and encodes a predicted protein of 5596 residues (632 kDa. Sequence alignments of midasin from humans, yeast, Giardia and Encephalitozoon indicate that its domain structure comprises an N-terminal domain (35 kDa, followed by an AAA domain containing six tandem AAA protomers (~30 kDa each, a linker domain (260 kDa, an acidic domain (~70 kDa containing 35–40% aspartate and glutamate, and a carboxy-terminal M-domain (30 kDa that possesses MIDAS sequence motifs and is homologous to the I-domain of integrins. Expression of hemagglutamin-tagged midasin in yeast demonstrates a polypeptide of the anticipated size that is localized principally in the nucleus. Conclusions The highly conserved structure of midasin in eukaryotes, taken in conjunction with its nuclear localization in yeast, suggests that midasin may function as a nuclear chaperone and be involved in the assembly/disassembly of macromolecular complexes in the nucleus. The AAA domain of midasin is evolutionarily related to that of dynein, but it appears to lack a microtubule-binding site.

  16. The identification and functional annotation of RNA structures conserved in vertebrates

    DEFF Research Database (Denmark)

    Seemann, Ernst Stefan; Mirza, Aashiq Hussain; Hansen, Claus

    2017-01-01

    Structured elements of RNA molecules are essential in, e.g., RNA stabilization, localization and protein interaction, and their conservation across species suggests a common functional role. We computationally screened vertebrate genomes for Conserved RNA Structures (CRSs), leveraging structure-b......-structured counterparts. Our findings of transcribed uncharacterized regulatory regions that contain CRSs support their RNA-mediated functionality.......Structured elements of RNA molecules are essential in, e.g., RNA stabilization, localization and protein interaction, and their conservation across species suggests a common functional role. We computationally screened vertebrate genomes for Conserved RNA Structures (CRSs), leveraging structure......-based, rather than sequence-based, alignments. After careful correction for sequence identity and GC content, we predict ~516k human genomic regions containing CRSs. We find that a substantial fraction of human-mouse CRS regions (i) co-localize consistently with binding sites of the same RNA binding proteins...

  17. Unique and conserved genome regions in Vibrio harveyi and related species in comparison with the shrimp pathogen Vibrio harveyi CAIM 1792

    DEFF Research Database (Denmark)

    Valles, Iliana Espinoza; Vora, Gary J; Lin, Baochuan

    2015-01-01

    Vibrio harveyi CAIM 1792 is a marine bacterial strain that causes mortality in farmed shrimp in north-west Mexico, and the identification of virulence genes in this strain is important for understanding its pathogenicity. The aim of this work was to compare the V. harveyi CAIM 1792 genome....... The proteome of CAIM 1792 had higher similarity to those of other V. harveyi strains (78 %) than to those of the other closely related species Vibrio owensii (67 %), Vibrio rotiferianus (63 %) and Vibrio campbellii (59 %). Pan-genome ORFans trees showed the best fit with the accepted phylogeny based on DNA......-DNA hybridization and multi-locus sequence analysis of 11 concatenated housekeeping genes. SNP analysis clustered 34/38 genomes within their accepted species. The pangenomic and SNP trees showed that V. harveyi is the most conserved of the four species studied and V. campbellii may be divided into at least three...

  18. Delineating slowly and rapidly evolving fractions of the Drosophila genome.

    Science.gov (United States)

    Keith, Jonathan M; Adams, Peter; Stephen, Stuart; Mattick, John S

    2008-05-01

    Evolutionary conservation is an important indicator of function and a major component of bioinformatic methods to identify non-protein-coding genes. We present a new Bayesian method for segmenting pairwise alignments of eukaryotic genomes while simultaneously classifying segments into slowly and rapidly evolving fractions. We also describe an information criterion similar to the Akaike Information Criterion (AIC) for determining the number of classes. Working with pairwise alignments enables detection of differences in conservation patterns among closely related species. We analyzed three whole-genome and three partial-genome pairwise alignments among eight Drosophila species. Three distinct classes of conservation level were detected. Sequences comprising the most slowly evolving component were consistent across a range of species pairs, and constituted approximately 62-66% of the D. melanogaster genome. Almost all (>90%) of the aligned protein-coding sequence is in this fraction, suggesting much of it (comprising the majority of the Drosophila genome, including approximately 56% of non-protein-coding sequences) is functional. The size and content of the most rapidly evolving component was species dependent, and varied from 1.6% to 4.8%. This fraction is also enriched for protein-coding sequence (while containing significant amounts of non-protein-coding sequence), suggesting it is under positive selection. We also classified segments according to conservation and GC content simultaneously. This analysis identified numerous sub-classes of those identified on the basis of conservation alone, but was nevertheless consistent with that classification. Software, data, and results available at www.maths.qut.edu.au/-keithj/. Genomic segments comprising the conservation classes available in BED format.

  19. Annotation Of Novel And Conserved MicroRNA Genes In The Build 10 Sus scrofa Reference Genome And Determination Of Their Expression Levels In Ten Different Tissues

    DEFF Research Database (Denmark)

    Thomsen, Bo; Nielsen, Mathilde; Hedegaard, Jakob

    The DNA template used in the pig genome sequencing project was provided by a Duroc pig named TJ Tabasco. In an effort to annotate microRNA (miRNA) genes in the reference genome we have conducted deep sequencing to determine the miRNA transcriptomes in ten different tissues isolated from Pinky......, a genetically identical clone of TJ Tabasco. The purpose was to generate miRNA sequences that are highly homologous to the reference genome sequence, which along with computational prediction will improve confidence in the genomic annotation of miRNA genes. Based on homology searches of the sequence data...... against miRBase, we identified more than 600 conserved known miRNA/miRNA*, which is a significant increase relative to the 211 porcine miRNA/miRNA* deposited in the current version of miRBase. Furthermore, the genome-wide transcript profiles provided important information on the relative abundance...

  20. Short and long-term genome stability analysis of prokaryotic genomes.

    Science.gov (United States)

    Brilli, Matteo; Liò, Pietro; Lacroix, Vincent; Sagot, Marie-France

    2013-05-08

    Gene organization dynamics is actively studied because it provides useful evolutionary information, makes functional annotation easier and often enables to characterize pathogens. There is therefore a strong interest in understanding the variability of this trait and the possible correlations with life-style. Two kinds of events affect genome organization: on one hand translocations and recombinations change the relative position of genes shared by two genomes (i.e. the backbone gene order); on the other, insertions and deletions leave the backbone gene order unchanged but they alter the gene neighborhoods by breaking the syntenic regions. A complete picture about genome organization evolution therefore requires to account for both kinds of events. We developed an approach where we model chromosomes as graphs on which we compute different stability estimators; we consider genome rearrangements as well as the effect of gene insertions and deletions. In a first part of the paper, we fit a measure of backbone gene order conservation (hereinafter called backbone stability) against phylogenetic distance for over 3000 genome comparisons, improving existing models for the divergence in time of backbone stability. Intra- and inter-specific comparisons were treated separately to focus on different time-scales. The use of multiple genomes of a same species allowed to identify genomes with diverging gene order with respect to their conspecific. The inter-species analysis indicates that pathogens are more often unstable with respect to non-pathogens. In a second part of the text, we show that in pathogens, gene content dynamics (insertions and deletions) have a much more dramatic effect on genome organization stability than backbone rearrangements. In this work, we studied genome organization divergence taking into account the contribution of both genome order rearrangements and genome content dynamics. By studying species with multiple sequenced genomes available, we were

  1. Packaging signals in two single-stranded RNA viruses imply a conserved assembly mechanism and geometry of the packaged genome.

    Science.gov (United States)

    Dykeman, Eric C; Stockley, Peter G; Twarock, Reidun

    2013-09-09

    The current paradigm for assembly of single-stranded RNA viruses is based on a mechanism involving non-sequence-specific packaging of genomic RNA driven by electrostatic interactions. Recent experiments, however, provide compelling evidence for sequence specificity in this process both in vitro and in vivo. The existence of multiple RNA packaging signals (PSs) within viral genomes has been proposed, which facilitates assembly by binding coat proteins in such a way that they promote the protein-protein contacts needed to build the capsid. The binding energy from these interactions enables the confinement or compaction of the genomic RNAs. Identifying the nature of such PSs is crucial for a full understanding of assembly, which is an as yet untapped potential drug target for this important class of pathogens. Here, for two related bacterial viruses, we determine the sequences and locations of their PSs using Hamiltonian paths, a concept from graph theory, in combination with bioinformatics and structural studies. Their PSs have a common secondary structure motif but distinct consensus sequences and positions within the respective genomes. Despite these differences, the distributions of PSs in both viruses imply defined conformations for the packaged RNA genomes in contact with the protein shell in the capsid, consistent with a recent asymmetric structure determination of the MS2 virion. The PS distributions identified moreover imply a preferred, evolutionarily conserved assembly pathway with respect to the RNA sequence with potentially profound implications for other single-stranded RNA viruses known to have RNA PSs, including many animal and human pathogens. Copyright © 2013 Elsevier Ltd. All rights reserved.

  2. Evolutionary conservation of regulatory elements in vertebrate HOX gene clusters

    Energy Technology Data Exchange (ETDEWEB)

    Santini, Simona; Boore, Jeffrey L.; Meyer, Axel

    2003-12-31

    Due to their high degree of conservation, comparisons of DNA sequences among evolutionarily distantly-related genomes permit to identify functional regions in noncoding DNA. Hox genes are optimal candidate sequences for comparative genome analyses, because they are extremely conserved in vertebrates and occur in clusters. We aligned (Pipmaker) the nucleotide sequences of HoxA clusters of tilapia, pufferfish, striped bass, zebrafish, horn shark, human and mouse (over 500 million years of evolutionary distance). We identified several highly conserved intergenic sequences, likely to be important in gene regulation. Only a few of these putative regulatory elements have been previously described as being involved in the regulation of Hox genes, while several others are new elements that might have regulatory functions. The majority of these newly identified putative regulatory elements contain short fragments that are almost completely conserved and are identical to known binding sites for regulatory proteins (Transfac). The conserved intergenic regions located between the most rostrally expressed genes in the developing embryo are longer and better retained through evolution. We document that presumed regulatory sequences are retained differentially in either A or A clusters resulting from a genome duplication in the fish lineage. This observation supports both the hypothesis that the conserved elements are involved in gene regulation and the Duplication-Deletion-Complementation model.

  3. Pan-Genome Analysis of Human Gastric Pathogen H. pylori: Comparative Genomics and Pathogenomics Approaches to Identify Regions Associated with Pathogenicity and Prediction of Potential Core Therapeutic Targets

    DEFF Research Database (Denmark)

    Ali, Amjad; Naz, Anam; Soares, Siomar C.

    2015-01-01

    -genome approach; the predicted conserved gene families (1,193) constitute similar to 77% of the average H. pylori genome and 45% of the global gene repertoire of the species. Reverse vaccinology strategies have been adopted to identify and narrow down the potential core-immunogenic candidates. Total of 28 nonhost....... Pan-genome analyses of the global representative H. pylori isolates consisting of 39 complete genomes are presented in this paper. Phylogenetic analyses have revealed close relationships among geographically diverse strains of H. pylori. The conservation among these genomes was further analyzed by pan...

  4. GenomePeek—an online tool for prokaryotic genome and metagenome analysis

    Directory of Open Access Journals (Sweden)

    Katelyn McNair

    2015-06-01

    Full Text Available As more and more prokaryotic sequencing takes place, a method to quickly and accurately analyze this data is needed. Previous tools are mainly designed for metagenomic analysis and have limitations; such as long runtimes and significant false positive error rates. The online tool GenomePeek (edwards.sdsu.edu/GenomePeek was developed to analyze both single genome and metagenome sequencing files, quickly and with low error rates. GenomePeek uses a sequence assembly approach where reads to a set of conserved genes are extracted, assembled and then aligned against the highly specific reference database. GenomePeek was found to be faster than traditional approaches while still keeping error rates low, as well as offering unique data visualization options.

  5. Genomic organization, annotation, and ligand-receptor inferences of chicken chemokines and chemokine receptor genes based on comparative genomics

    Directory of Open Access Journals (Sweden)

    Sze Sing-Hoi

    2005-03-01

    Full Text Available Abstract Background Chemokines and their receptors play important roles in host defense, organogenesis, hematopoiesis, and neuronal communication. Forty-two chemokines and 19 cognate receptors have been found in the human genome. Prior to this report, only 11 chicken chemokines and 7 receptors had been reported. The objectives of this study were to systematically identify chicken chemokines and their cognate receptor genes in the chicken genome and to annotate these genes and ligand-receptor binding by a comparative genomics approach. Results Twenty-three chemokine and 14 chemokine receptor genes were identified in the chicken genome. All of the chicken chemokines contained a conserved CC, CXC, CX3C, or XC motif, whereas all the chemokine receptors had seven conserved transmembrane helices, four extracellular domains with a conserved cysteine, and a conserved DRYLAIV sequence in the second intracellular domain. The number of coding exons in these genes and the syntenies are highly conserved between human, mouse, and chicken although the amino acid sequence homologies are generally low between mammalian and chicken chemokines. Chicken genes were named with the systematic nomenclature used in humans and mice based on phylogeny, synteny, and sequence homology. Conclusion The independent nomenclature of chicken chemokines and chemokine receptors suggests that the chicken may have ligand-receptor pairings similar to mammals. All identified chicken chemokines and their cognate receptors were identified in the chicken genome except CCR9, whose ligand was not identified in this study. The organization of these genes suggests that there were a substantial number of these genes present before divergence between aves and mammals and more gene duplications of CC, CXC, CCR, and CXCR subfamilies in mammals than in aves after the divergence.

  6. Detection of Weakly Conserved Ancestral Mammalian RegulatorySequences by Primate Comparisons

    Energy Technology Data Exchange (ETDEWEB)

    Wang, Qian-fei; Prabhakar, Shyam; Chanan, Sumita; Cheng,Jan-Fang; Rubin, Edward M.; Boffelli, Dario

    2006-06-01

    Genomic comparisons between human and distant, non-primatemammals are commonly used to identify cis-regulatory elements based onconstrained sequence evolution. However, these methods fail to detectcryptic functional elements, which are too weakly conserved among mammalsto distinguish from nonfunctional DNA. To address this problem, weexplored the potential of deep intra-primate sequence comparisons. Wesequenced the orthologs of 558 kb of human genomic sequence, coveringmultiple loci involved in cholesterol homeostasis, in 6 nonhumanprimates. Our analysis identified 6 noncoding DNA elements displayingsignificant conservation among primates, but undetectable in more distantcomparisons. In vitro and in vivo tests revealed that at least three ofthese 6 elements have regulatory function. Notably, the mouse orthologsof these three functional human sequences had regulatory activity despitetheir lack of significant sequence conservation, indicating that they arecryptic ancestral cis-regulatory elements. These regulatory elementscould still be detected in a smaller set of three primate speciesincluding human, rhesus and marmoset. Since the human and rhesus genomesequences are already available, and the marmoset genome is activelybeing sequenced, the primate-specific conservation analysis describedhere can be applied in the near future on a whole-genome scale, tocomplement the annotation provided by more distant speciescomparisons.

  7. PSAT: A web tool to compare genomic neighborhoods of multiple prokaryotic genomes

    Directory of Open Access Journals (Sweden)

    Wasnick Michael

    2008-03-01

    Full Text Available Abstract Background The conservation of gene order among prokaryotic genomes can provide valuable insight into gene function, protein interactions, or events by which genomes have evolved. Although some tools are available for visualizing and comparing the order of genes between genomes of study, few support an efficient and organized analysis between large numbers of genomes. The Prokaryotic Sequence homology Analysis Tool (PSAT is a web tool for comparing gene neighborhoods among multiple prokaryotic genomes. Results PSAT utilizes a database that is preloaded with gene annotation, BLAST hit results, and gene-clustering scores designed to help identify regions of conserved gene order. Researchers use the PSAT web interface to find a gene of interest in a reference genome and efficiently retrieve the sequence homologs found in other bacterial genomes. The tool generates a graphic of the genomic neighborhood surrounding the selected gene and the corresponding regions for its homologs in each comparison genome. Homologs in each region are color coded to assist users with analyzing gene order among various genomes. In contrast to common comparative analysis methods that filter sequence homolog data based on alignment score cutoffs, PSAT leverages gene context information for homologs, including those with weak alignment scores, enabling a more sensitive analysis. Features for constraining or ordering results are designed to help researchers browse results from large numbers of comparison genomes in an organized manner. PSAT has been demonstrated to be useful for helping to identify gene orthologs and potential functional gene clusters, and detecting genome modifications that may result in loss of function. Conclusion PSAT allows researchers to investigate the order of genes within local genomic neighborhoods of multiple genomes. A PSAT web server for public use is available for performing analyses on a growing set of reference genomes through any

  8. The spotted gar genome illuminates vertebrate evolution and facilitates human-teleost comparisons.

    Science.gov (United States)

    Braasch, Ingo; Gehrke, Andrew R; Smith, Jeramiah J; Kawasaki, Kazuhiko; Manousaki, Tereza; Pasquier, Jeremy; Amores, Angel; Desvignes, Thomas; Batzel, Peter; Catchen, Julian; Berlin, Aaron M; Campbell, Michael S; Barrell, Daniel; Martin, Kyle J; Mulley, John F; Ravi, Vydianathan; Lee, Alison P; Nakamura, Tetsuya; Chalopin, Domitille; Fan, Shaohua; Wcisel, Dustin; Cañestro, Cristian; Sydes, Jason; Beaudry, Felix E G; Sun, Yi; Hertel, Jana; Beam, Michael J; Fasold, Mario; Ishiyama, Mikio; Johnson, Jeremy; Kehr, Steffi; Lara, Marcia; Letaw, John H; Litman, Gary W; Litman, Ronda T; Mikami, Masato; Ota, Tatsuya; Saha, Nil Ratan; Williams, Louise; Stadler, Peter F; Wang, Han; Taylor, John S; Fontenot, Quenton; Ferrara, Allyse; Searle, Stephen M J; Aken, Bronwen; Yandell, Mark; Schneider, Igor; Yoder, Jeffrey A; Volff, Jean-Nicolas; Meyer, Axel; Amemiya, Chris T; Venkatesh, Byrappa; Holland, Peter W H; Guiguen, Yann; Bobe, Julien; Shubin, Neil H; Di Palma, Federica; Alföldi, Jessica; Lindblad-Toh, Kerstin; Postlethwait, John H

    2016-04-01

    To connect human biology to fish biomedical models, we sequenced the genome of spotted gar (Lepisosteus oculatus), whose lineage diverged from teleosts before teleost genome duplication (TGD). The slowly evolving gar genome has conserved in content and size many entire chromosomes from bony vertebrate ancestors. Gar bridges teleosts to tetrapods by illuminating the evolution of immunity, mineralization and development (mediated, for example, by Hox, ParaHox and microRNA genes). Numerous conserved noncoding elements (CNEs; often cis regulatory) undetectable in direct human-teleost comparisons become apparent using gar: functional studies uncovered conserved roles for such cryptic CNEs, facilitating annotation of sequences identified in human genome-wide association studies. Transcriptomic analyses showed that the sums of expression domains and expression levels for duplicated teleost genes often approximate the patterns and levels of expression for gar genes, consistent with subfunctionalization. The gar genome provides a resource for understanding evolution after genome duplication, the origin of vertebrate genomes and the function of human regulatory sequences.

  9. The Release 6 reference sequence of the Drosophila melanogaster genome.

    Science.gov (United States)

    Hoskins, Roger A; Carlson, Joseph W; Wan, Kenneth H; Park, Soo; Mendez, Ivonne; Galle, Samuel E; Booth, Benjamin W; Pfeiffer, Barret D; George, Reed A; Svirskas, Robert; Krzywinski, Martin; Schein, Jacqueline; Accardo, Maria Carmela; Damia, Elisabetta; Messina, Giovanni; Méndez-Lago, María; de Pablos, Beatriz; Demakova, Olga V; Andreyeva, Evgeniya N; Boldyreva, Lidiya V; Marra, Marco; Carvalho, A Bernardo; Dimitri, Patrizio; Villasante, Alfredo; Zhimulev, Igor F; Rubin, Gerald M; Karpen, Gary H; Celniker, Susan E

    2015-03-01

    Drosophila melanogaster plays an important role in molecular, genetic, and genomic studies of heredity, development, metabolism, behavior, and human disease. The initial reference genome sequence reported more than a decade ago had a profound impact on progress in Drosophila research, and improving the accuracy and completeness of this sequence continues to be important to further progress. We previously described improvement of the 117-Mb sequence in the euchromatic portion of the genome and 21 Mb in the heterochromatic portion, using a whole-genome shotgun assembly, BAC physical mapping, and clone-based finishing. Here, we report an improved reference sequence of the single-copy and middle-repetitive regions of the genome, produced using cytogenetic mapping to mitotic and polytene chromosomes, clone-based finishing and BAC fingerprint verification, ordering of scaffolds by alignment to cDNA sequences, incorporation of other map and sequence data, and validation by whole-genome optical restriction mapping. These data substantially improve the accuracy and completeness of the reference sequence and the order and orientation of sequence scaffolds into chromosome arm assemblies. Representation of the Y chromosome and other heterochromatic regions is particularly improved. The new 143.9-Mb reference sequence, designated Release 6, effectively exhausts clone-based technologies for mapping and sequencing. Highly repeat-rich regions, including large satellite blocks and functional elements such as the ribosomal RNA genes and the centromeres, are largely inaccessible to current sequencing and assembly methods and remain poorly represented. Further significant improvements will require sequencing technologies that do not depend on molecular cloning and that produce very long reads. © 2015 Hoskins et al.; Published by Cold Spring Harbor Laboratory Press.

  10. Genome Analysis of Conserved Dehydrin Motifs in Vascular Plants

    Directory of Open Access Journals (Sweden)

    Ahmad A. Malik

    2017-05-01

    Full Text Available Dehydrins, a large family of abiotic stress proteins, are defined by the presence of a mostly conserved motif known as the K-segment, and may also contain two other conserved motifs known as the Y-segment and S-segment. Using the dehydrin literature, we developed a sequence motif definition of the K-segment, which we used to create a large dataset of dehydrin sequences by searching the Pfam00257 dehydrin dataset and the Phytozome 10 sequences of vascular plants. A comprehensive analysis of these sequences reveals that lysine residues are highly conserved in the K-segment, while the amino acid type is often conserved at other positions. Despite the Y-segment name, the central tyrosine is somewhat conserved, but can be substituted with two other small aromatic amino acids (phenylalanine or histidine. The S-segment contains a series of serine residues, but in some proteins is also preceded by a conserved LHR sequence. In many dehydrins containing all three of these motifs the S-segment is linked to the K-segment by a GXGGRRKK motif (where X can be any amino acid, suggesting a functional linkage between these two motifs. An analysis of the sequences shows that the dehydrin architecture and several biochemical properties (isoelectric point, molecular mass, and hydrophobicity score are dependent on each other, and that some dehydrin architectures are overexpressed during certain abiotic stress, suggesting that they may be optimized for a specific abiotic stress while others are involved in all forms of dehydration stress (drought, cold, and salinity.

  11. Comparative genomics of chondrichthyan Hoxa clusters

    Directory of Open Access Journals (Sweden)

    Zhong Ying-Fu

    2009-09-01

    Full Text Available Abstract Background The chondrichthyan or cartilaginous fish (chimeras, sharks, skates and rays occupy an important phylogenetic position as the sister group to all other jawed vertebrates and as an early lineage to diverge from the vertebrate lineage following two whole genome duplication events in vertebrate evolution. There have been few comparative genomic analyses incorporating data from chondrichthyan fish and none comparing genomic information from within the group. We have sequenced the complete Hoxa cluster of the Little Skate (Leucoraja erinacea and compared to the published Hoxa cluster of the Horn Shark (Heterodontus francisci and to available data from the Elephant Shark (Callorhinchus milii genome project. Results A BAC clone containing the full Little Skate Hoxa cluster was fully sequenced and assembled. Analyses of coding sequences and conserved non-coding elements reveal a strikingly high level of conservation across the cartilaginous fish, with twenty ultraconserved elements (100%,100 bp found between Skate and Horn Shark, compared to three between human and marsupials. We have also identified novel potential non-coding RNAs in the Skate BAC clone, some of which are conserved to other species. Conclusion We find that the Little Skate Hoxa cluster is remarkably similar to the previously published Horn Shark Hoxa cluster with respect to sequence identity, gene size and intergenic distance despite over 180 million years of separation between the two lineages. We suggest that the genomes of cartilaginous fish are more highly conserved than those of tetrapods or teleost fish and so are more likely to have retained ancestral non-coding elements. While useful for isolating homologous DNA, this complicates bioinformatic approaches to identify chondrichthyan-specific non-coding DNA elements

  12. Strong trans-Pacific break and local conservation units in the Galapagos shark (Carcharhinus galapagensis) revealed by genome-wide cytonuclear markers.

    Science.gov (United States)

    Pazmiño, Diana A; Maes, Gregory E; Green, Madeline E; Simpfendorfer, Colin A; Hoyos-Padilla, E Mauricio; Duffy, Clinton J A; Meyer, Carl G; Kerwath, Sven E; Salinas-de-León, Pelayo; van Herwerden, Lynne

    2018-05-01

    The application of genome-wide cytonuclear molecular data to identify management and adaptive units at various spatio-temporal levels is particularly important for overharvested large predatory organisms, often characterized by smaller, localized populations. Despite being "near threatened", current understanding of habitat use and population structure of Carcharhinus galapagensis is limited to specific areas within its distribution. We evaluated population structure and connectivity across the Pacific Ocean using genome-wide single-nucleotide polymorphisms (~7200 SNPs) and mitochondrial control region sequences (945 bp) for 229 individuals. Neutral SNPs defined at least two genetically discrete geographic groups: an East Tropical Pacific (Mexico, east and west Galapagos Islands), and another central-west Pacific (Lord Howe Island, Middleton Reef, Norfolk Island, Elizabeth Reef, Kermadec, Hawaii and Southern Africa). More fine-grade population structure was suggested using outlier SNPs: west Pacific, Hawaii, Mexico, and Galapagos. Consistently, mtDNA pairwise Φ ST defined three regional stocks: east, central and west Pacific. Compared to neutral SNPs (F ST  = 0.023-0.035), mtDNA exhibited more divergence (Φ ST  = 0.258-0.539) and high overall genetic diversity (h = 0.794 ± 0.014; π = 0.004 ± 0.000), consistent with the longstanding eastern Pacific barrier between the east and central-west Pacific. Hawaiian and Southern African populations group within the west Pacific cluster. Effective population sizes were moderate/high for east/west populations (738 and 3421, respectively). Insights into the biology, connectivity, genetic diversity, and population demographics informs for improved conservation of this species, by delineating three to four conservation units across their Pacific distribution. Implementing such conservation management may be challenging, but is necessary to achieve long-term population resilience at basin and

  13. Comparative analysis of catfish BAC end sequences with the zebrafish genome

    Directory of Open Access Journals (Sweden)

    Abernathy Jason

    2009-12-01

    Full Text Available Abstract Background Comparative mapping is a powerful tool to transfer genomic information from sequenced genomes to closely related species for which whole genome sequence data are not yet available. However, such an approach is still very limited in catfish, the most important aquaculture species in the United States. This project was initiated to generate additional BAC end sequences and demonstrate their applications in comparative mapping in catfish. Results We reported the generation of 43,000 BAC end sequences and their applications for comparative genome analysis in catfish. Using these and the additional 20,000 existing BAC end sequences as a resource along with linkage mapping and existing physical map, conserved syntenic regions were identified between the catfish and zebrafish genomes. A total of 10,943 catfish BAC end sequences (17.3% had significant BLAST hits to the zebrafish genome (cutoff value ≤ e-5, of which 3,221 were unique gene hits, providing a platform for comparative mapping based on locations of these genes in catfish and zebrafish. Genetic linkage mapping of microsatellites associated with contigs allowed identification of large conserved genomic segments and construction of super scaffolds. Conclusion BAC end sequences and their associated polymorphic markers are great resources for comparative genome analysis in catfish. Highly conserved chromosomal regions were identified to exist between catfish and zebrafish. However, it appears that the level of conservation at local genomic regions are high while a high level of chromosomal shuffling and rearrangements exist between catfish and zebrafish genomes. Orthologous regions established through comparative analysis should facilitate both structural and functional genome analysis in catfish.

  14. Combining genetical genomics and bulked segregant analysis differential expression: an approach to gene localization

    NARCIS (Netherlands)

    Chen, Xinwei; Hedley, P.E.; Morris, J.; Liu, Hui; Niks, R.E.; Waugh, R.

    2011-01-01

    Positional gene isolation in unsequenced species generally requires either a reference genome sequence or an inference of gene content based on conservation of synteny with a genomic model. In the large unsequenced genomes of the Triticeae cereals the latter, i.e. conservation of synteny with the

  15. A Limousin specific myostatin allele affects longissimus muscle area and fatty acid profiles in a Wagyu-Limousin F*2* population

    Science.gov (United States)

    A microsatellite-based genome scan of a Wagyu x Limousin F2 cross population previously demonstrated QTL affecting longissimus muscle area (LMA) and fatty acid composition were present in regions near the centromere of BTA 2. In this study we used 70 SNP markers to examine the centromeric 20 megabas...

  16. Comparative genomics reveals insights into avian genome evolution and adaptation

    Science.gov (United States)

    Zhang, Guojie; Li, Cai; Li, Qiye; Li, Bo; Larkin, Denis M.; Lee, Chul; Storz, Jay F.; Antunes, Agostinho; Greenwold, Matthew J.; Meredith, Robert W.; Ödeen, Anders; Cui, Jie; Zhou, Qi; Xu, Luohao; Pan, Hailin; Wang, Zongji; Jin, Lijun; Zhang, Pei; Hu, Haofu; Yang, Wei; Hu, Jiang; Xiao, Jin; Yang, Zhikai; Liu, Yang; Xie, Qiaolin; Yu, Hao; Lian, Jinmin; Wen, Ping; Zhang, Fang; Li, Hui; Zeng, Yongli; Xiong, Zijun; Liu, Shiping; Zhou, Long; Huang, Zhiyong; An, Na; Wang, Jie; Zheng, Qiumei; Xiong, Yingqi; Wang, Guangbiao; Wang, Bo; Wang, Jingjing; Fan, Yu; da Fonseca, Rute R.; Alfaro-Núñez, Alonzo; Schubert, Mikkel; Orlando, Ludovic; Mourier, Tobias; Howard, Jason T.; Ganapathy, Ganeshkumar; Pfenning, Andreas; Whitney, Osceola; Rivas, Miriam V.; Hara, Erina; Smith, Julia; Farré, Marta; Narayan, Jitendra; Slavov, Gancho; Romanov, Michael N; Borges, Rui; Machado, João Paulo; Khan, Imran; Springer, Mark S.; Gatesy, John; Hoffmann, Federico G.; Opazo, Juan C.; Håstad, Olle; Sawyer, Roger H.; Kim, Heebal; Kim, Kyu-Won; Kim, Hyeon Jeong; Cho, Seoae; Li, Ning; Huang, Yinhua; Bruford, Michael W.; Zhan, Xiangjiang; Dixon, Andrew; Bertelsen, Mads F.; Derryberry, Elizabeth; Warren, Wesley; Wilson, Richard K; Li, Shengbin; Ray, David A.; Green, Richard E.; O’Brien, Stephen J.; Griffin, Darren; Johnson, Warren E.; Haussler, David; Ryder, Oliver A.; Willerslev, Eske; Graves, Gary R.; Alström, Per; Fjeldså, Jon; Mindell, David P.; Edwards, Scott V.; Braun, Edward L.; Rahbek, Carsten; Burt, David W.; Houde, Peter; Zhang, Yong; Yang, Huanming; Wang, Jian; Jarvis, Erich D.; Gilbert, M. Thomas P.; Wang, Jun

    2015-01-01

    Birds are the most species-rich class of tetrapod vertebrates and have wide relevance across many research fields. We explored bird macroevolution using full genomes from 48 avian species representing all major extant clades. The avian genome is principally characterized by its constrained size, which predominantly arose because of lineage-specific erosion of repetitive elements, large segmental deletions, and gene loss. Avian genomes furthermore show a remarkably high degree of evolutionary stasis at the levels of nucleotide sequence, gene synteny, and chromosomal structure. Despite this pattern of conservation, we detected many non-neutral evolutionary changes in protein-coding genes and noncoding regions. These analyses reveal that pan-avian genomic diversity covaries with adaptations to different lifestyles and convergent evolution of traits. PMID:25504712

  17. Conservation and Role of Electrostatics in Thymidylate Synthase.

    Science.gov (United States)

    Garg, Divita; Skouloubris, Stephane; Briffotaux, Julien; Myllykallio, Hannu; Wade, Rebecca C

    2015-11-27

    Conservation of function across families of orthologous enzymes is generally accompanied by conservation of their active site electrostatic potentials. To study the electrostatic conservation in the highly conserved essential enzyme, thymidylate synthase (TS), we conducted a systematic species-based comparison of the electrostatic potential in the vicinity of its active site. Whereas the electrostatics of the active site of TS are generally well conserved, the TSs from minimal organisms do not conform to the overall trend. Since the genomes of minimal organisms have a high thymidine content compared to other organisms, the observation of non-conserved electrostatics was surprising. Analysis of the symbiotic relationship between minimal organisms and their hosts, and the genetic completeness of the thymidine synthesis pathway suggested that TS from the minimal organism Wigglesworthia glossinidia (W.g.b.) must be active. Four residues in the vicinity of the active site of Escherichia coli TS were mutated individually and simultaneously to mimic the electrostatics of W.g.b TS. The measured activities of the E. coli TS mutants imply that conservation of electrostatics in the region of the active site is important for the activity of TS, and suggest that the W.g.b. TS has the minimal activity necessary to support replication of its reduced genome.

  18. Forces shaping the fastest evolving regions in the human genome

    DEFF Research Database (Denmark)

    Pollard, Katherine S; Salama, Sofie R; King, Bryan

    2006-01-01

    Comparative genomics allow us to search the human genome for segments that were extensively changed in the last approximately 5 million years since divergence from our common ancestor with chimpanzee, but are highly conserved in other species and thus are likely to be functional. We found 202...... genomic elements that are highly conserved in vertebrates but show evidence of significantly accelerated substitution rates in human. These are mostly in non-coding DNA, often near genes associated with transcription and DNA binding. Resequencing confirmed that the five most accelerated elements...... contributed to accelerated evolution of the fastest evolving elements in the human genome....

  19. Mutations in CENPE define a novel kinetochore-centromeric mechanism for microcephalic primordial dwarfism.

    Science.gov (United States)

    Mirzaa, Ghayda M; Vitre, Benjamin; Carpenter, Gillian; Abramowicz, Iga; Gleeson, Joseph G; Paciorkowski, Alex R; Cleveland, Don W; Dobyns, William B; O'Driscoll, Mark

    2014-08-01

    Defects in centrosome, centrosomal-associated and spindle-associated proteins are the most frequent cause of primary microcephaly (PM) and microcephalic primordial dwarfism (MPD) syndromes in humans. Mitotic progression and segregation defects, microtubule spindle abnormalities and impaired DNA damage-induced G2-M cell cycle checkpoint proficiency have been documented in cell lines from these patients. This suggests that impaired mitotic entry, progression and exit strongly contribute to PM and MPD. Considering the vast protein networks involved in coordinating this cell cycle stage, the list of potential target genes that could underlie novel developmental disorders is large. One such complex network, with a direct microtubule-mediated physical connection to the centrosome, is the kinetochore. This centromeric-associated structure nucleates microtubule attachments onto mitotic chromosomes. Here, we described novel compound heterozygous variants in CENPE in two siblings who exhibit a profound MPD associated with developmental delay, simplified gyri and other isolated abnormalities. CENPE encodes centromere-associated protein E (CENP-E), a core kinetochore component functioning to mediate chromosome congression initially of misaligned chromosomes and in subsequent spindle microtubule capture during mitosis. Firstly, we present a comprehensive clinical description of these patients. Then, using patient cells we document abnormalities in spindle microtubule organization, mitotic progression and segregation, before modeling the cellular pathogenicity of these variants in an independent cell system. Our cellular analysis shows that a pathogenic defect in CENP-E, a kinetochore-core protein, largely phenocopies PCNT-mutated microcephalic osteodysplastic primordial dwarfism-type II patient cells. PCNT encodes a centrosome-associated protein. These results highlight a common underlying pathomechanism. Our findings provide the first evidence for a kinetochore-based route to

  20. Highly conserved non-coding sequences are associated with vertebrate development.

    Directory of Open Access Journals (Sweden)

    Adam Woolfe

    2005-01-01

    Full Text Available In addition to protein coding sequence, the human genome contains a significant amount of regulatory DNA, the identification of which is proving somewhat recalcitrant to both in silico and functional methods. An approach that has been used with some success is comparative sequence analysis, whereby equivalent genomic regions from different organisms are compared in order to identify both similarities and differences. In general, similarities in sequence between highly divergent organisms imply functional constraint. We have used a whole-genome comparison between humans and the pufferfish, Fugu rubripes, to identify nearly 1,400 highly conserved non-coding sequences. Given the evolutionary divergence between these species, it is likely that these sequences are found in, and furthermore are essential to, all vertebrates. Most, and possibly all, of these sequences are located in and around genes that act as developmental regulators. Some of these sequences are over 90% identical across more than 500 bases, being more highly conserved than coding sequence between these two species. Despite this, we cannot find any similar sequences in invertebrate genomes. In order to begin to functionally test this set of sequences, we have used a rapid in vivo assay system using zebrafish embryos that allows tissue-specific enhancer activity to be identified. Functional data is presented for highly conserved non-coding sequences associated with four unrelated developmental regulators (SOX21, PAX6, HLXB9, and SHH, in order to demonstrate the suitability of this screen to a wide range of genes and expression patterns. Of 25 sequence elements tested around these four genes, 23 show significant enhancer activity in one or more tissues. We have identified a set of non-coding sequences that are highly conserved throughout vertebrates. They are found in clusters across the human genome, principally around genes that are implicated in the regulation of development

  1. The Drosophila wings apart gene anchors a novel, evolutionarily conserved pathway of neuromuscular development.

    Science.gov (United States)

    Morriss, Ginny R; Jaramillo, Carmelita T; Mikolajczak, Crystal M; Duong, Sandy; Jaramillo, Maryann S; Cripps, Richard M

    2013-11-01

    wings apart (wap) is a recessive, semilethal gene located on the X chromosome in Drosophila melanogaster, which is required for normal wing-vein patterning. We show that the wap mutation also results in loss of the adult jump muscle. We use complementation mapping and gene-specific RNA interference to localize the wap locus to the proximal X chromosome. We identify the annotated gene CG14614 as the gene affected by the wap mutation, since one wap allele contains a non-sense mutation in CG14614, and a genomic fragment containing only CG14614 rescues the jump-muscle phenotypes of two wap mutant alleles. The wap gene lies centromere-proximal to touch-insensitive larva B and centromere-distal to CG14619, which is tentatively assigned as the gene affected in introverted mutants. In mutant wap animals, founder cell precursors for the jump muscle are specified early in development, but are later lost. Through tissue-specific knockdowns, we demonstrate that wap function is required in both the musculature and the nervous system for normal jump-muscle formation. wap/CG14614 is homologous to vertebrate wdr68, DDB1 and CUL4 associated factor 7, which also are expressed in neuromuscular tissues. Thus, our findings provide insight into mechanisms of neuromuscular development in higher animals and facilitate the understanding of neuromuscular diseases that may result from mis-expression of muscle-specific or neuron-specific genes.

  2. The spotted gar genome illuminates vertebrate evolution and facilitates human-to-teleost comparisons

    Science.gov (United States)

    Braasch, Ingo; Gehrke, Andrew R.; Smith, Jeramiah J.; Kawasaki, Kazuhiko; Manousaki, Tereza; Pasquier, Jeremy; Amores, Angel; Desvignes, Thomas; Batzel, Peter; Catchen, Julian; Berlin, Aaron M.; Campbell, Michael S.; Barrell, Daniel; Martin, Kyle J.; Mulley, John F.; Ravi, Vydianathan; Lee, Alison P.; Nakamura, Tetsuya; Chalopin, Domitille; Fan, Shaohua; Wcisel, Dustin; Cañestro, Cristian; Sydes, Jason; Beaudry, Felix E. G.; Sun, Yi; Hertel, Jana; Beam, Michael J.; Fasold, Mario; Ishiyama, Mikio; Johnson, Jeremy; Kehr, Steffi; Lara, Marcia; Letaw, John H.; Litman, Gary W.; Litman, Ronda T.; Mikami, Masato; Ota, Tatsuya; Saha, Nil Ratan; Williams, Louise; Stadler, Peter F.; Wang, Han; Taylor, John S.; Fontenot, Quenton; Ferrara, Allyse; Searle, Stephen M. J.; Aken, Bronwen; Yandell, Mark; Schneider, Igor; Yoder, Jeffrey A.; Volff, Jean-Nicolas; Meyer, Axel; Amemiya, Chris T.; Venkatesh, Byrappa; Holland, Peter W. H.; Guiguen, Yann; Bobe, Julien; Shubin, Neil H.; Di Palma, Federica; Alföldi, Jessica; Lindblad-Toh, Kerstin; Postlethwait, John H.

    2016-01-01

    To connect human biology to fish biomedical models, we sequenced the genome of spotted gar (Lepisosteus oculatus), whose lineage diverged from teleosts before the teleost genome duplication (TGD). The slowly evolving gar genome conserved in content and size many entire chromosomes from bony vertebrate ancestors. Gar bridges teleosts to tetrapods by illuminating the evolution of immunity, mineralization, and development (e.g., Hox, ParaHox, and miRNA genes). Numerous conserved non-coding elements (CNEs, often cis-regulatory) undetectable in direct human-teleost comparisons become apparent using gar: functional studies uncovered conserved roles of such cryptic CNEs, facilitating annotation of sequences identified in human genome-wide association studies. Transcriptomic analyses revealed that the sum of expression domains and levels from duplicated teleost genes often approximate patterns and levels of gar genes, consistent with subfunctionalization. The gar genome provides a resource for understanding evolution after genome duplication, the origin of vertebrate genomes, and the function of human regulatory sequences. PMID:26950095

  3. Conservation of the Type IV secretion system throughout Wolbachia evolution

    DEFF Research Database (Denmark)

    Pichon, Samuel; Bouchon, Didier; Cordaux, Richard

    2009-01-01

    , encoding a T4SS were previously identified and characterized at two separate genomic loci. Using the largest data set of Wolbachia strains studied so far, we show that vir gene sequence and organization are strictly conserved among 37 Wolbachia strains inducing various phenotypes such as cytoplasmic...... incompatibility, feminization, or oogenesis in their arthropod hosts. In sharp contrast, extensive variation of genomic sequences flanking the virB8-D4 operon suggested its distinct location among Wolbachia genomes. Long term conservation of the T4SS may imply maintenance of a functional effector translocation...... system in Wolbachia, thereby suggesting the importance for the T4SS in Wolbachia biology and survival inside host cells....

  4. In Vivo Enhancer Analysis Chromosome 16 Conserved NoncodingSequences

    Energy Technology Data Exchange (ETDEWEB)

    Pennacchio, Len A.; Ahituv, Nadav; Moses, Alan M.; Nobrega,Marcelo; Prabhakar, Shyam; Shoukry, Malak; Minovitsky, Simon; Visel,Axel; Dubchak, Inna; Holt, Amy; Lewis, Keith D.; Plajzer-Frick, Ingrid; Akiyama, Jennifer; De Val, Sarah; Afzal, Veena; Black, Brian L.; Couronne, Olivier; Eisen, Michael B.; Rubin, Edward M.

    2006-02-01

    The identification of enhancers with predicted specificitiesin vertebrate genomes remains a significant challenge that is hampered bya lack of experimentally validated training sets. In this study, weleveraged extreme evolutionary sequence conservation as a filter toidentify putative gene regulatory elements and characterized the in vivoenhancer activity of human-fish conserved and ultraconserved1 noncodingelements on human chromosome 16 as well as such elements from elsewherein the genome. We initially tested 165 of these extremely conservedsequences in a transgenic mouse enhancer assay and observed that 48percent (79/165) functioned reproducibly as tissue-specific enhancers ofgene expression at embryonic day 11.5. While driving expression in abroad range of anatomical structures in the embryo, the majority of the79 enhancers drove expression in various regions of the developingnervous system. Studying a set of DNA elements that specifically droveforebrain expression, we identified DNA signatures specifically enrichedin these elements and used these parameters to rank all ~;3,400human-fugu conserved noncoding elements in the human genome. The testingof the top predictions in transgenic mice resulted in a three-foldenrichment for sequences with forebrain enhancer activity. These datadramatically expand the catalogue of in vivo-characterized human geneenhancers and illustrate the future utility of such training sets for avariety of iological applications including decoding the regulatoryvocabulary of the human genome.

  5. Large scale genomic reorganization of topological domains at the HoxD locus.

    Science.gov (United States)

    Fabre, Pierre J; Leleu, Marion; Mormann, Benjamin H; Lopez-Delisle, Lucille; Noordermeer, Daan; Beccari, Leonardo; Duboule, Denis

    2017-08-07

    The transcriptional activation of HoxD genes during mammalian limb development involves dynamic interactions with two topologically associating domains (TADs) flanking the HoxD cluster. In particular, the activation of the most posterior HoxD genes in developing digits is controlled by regulatory elements located in the centromeric TAD (C-DOM) through long-range contacts. To assess the structure-function relationships underlying such interactions, we measured compaction levels and TAD discreteness using a combination of chromosome conformation capture (4C-seq) and DNA FISH. We assessed the robustness of the TAD architecture by using a series of genomic deletions and inversions that impact the integrity of this chromatin domain and that remodel long-range contacts. We report multi-partite associations between HoxD genes and up to three enhancers. We find that the loss of native chromatin topology leads to the remodeling of TAD structure following distinct parameters. Our results reveal that the recomposition of TAD architectures after large genomic re-arrangements is dependent on a boundary-selection mechanism in which CTCF mediates the gating of long-range contacts in combination with genomic distance and sequence specificity. Accordingly, the building of a recomposed TAD at this locus depends on distinct functional and constitutive parameters.

  6. Transcription of tandemly repetitive DNA: functional roles.

    Science.gov (United States)

    Biscotti, Maria Assunta; Canapa, Adriana; Forconi, Mariko; Olmo, Ettore; Barucca, Marco

    2015-09-01

    A considerable fraction of the eukaryotic genome is made up of satellite DNA constituted of tandemly repeated sequences. These elements are mainly located at centromeres, pericentromeres, and telomeres and are major components of constitutive heterochromatin. Although originally satellite DNA was thought silent and inert, an increasing number of studies are providing evidence on its transcriptional activity supporting, on the contrary, an unexpected dynamicity. This review summarizes the multiple structural roles of satellite noncoding RNAs at chromosome level. Indeed, satellite noncoding RNAs play a role in the establishment of a heterochromatic state at centromere and telomere. These highly condensed structures are indispensable to preserve chromosome integrity and genome stability, preventing recombination events, and ensuring the correct chromosome pairing and segregation. Moreover, these RNA molecules seem to be involved also in maintaining centromere identity and in elongation, capping, and replication of telomere. Finally, the abnormal variation of centromeric and pericentromeric DNA transcription across major eukaryotic lineages in stress condition and disease has evidenced the critical role that these transcripts may play and the potentially dire consequences for the organism.

  7. Function of Junk: Pericentromeric Satellite DNA in Chromosome Maintenance.

    Science.gov (United States)

    Jagannathan, Madhav; Yamashita, Yukiko M

    2018-04-02

    Satellite DNAs are simple tandem repeats that exist at centromeric and pericentromeric regions on eukaryotic chromosomes. Unlike the centromeric satellite DNA that comprises the vast majority of natural centromeres, function(s) for the much more abundant pericentromeric satellite repeats are poorly understood. In fact, the lack of coding potential allied with rapid divergence of repeat sequences across eukaryotes has led to their dismissal as "junk DNA" or "selfish parasites." Although implicated in various biological processes, a conserved function for pericentromeric satellite DNA remains unidentified. We have addressed the role of satellite DNA through studying chromocenters, a cytological aggregation of pericentromeric satellite DNA from multiple chromosomes into DNA-dense nuclear foci. We have shown that multivalent satellite DNA-binding proteins cross-link pericentromeric satellite DNA on chromosomes into chromocenters. Disruption of chromocenters results in the formation of micronuclei, which arise by budding off the nucleus during interphase. We propose a model that satellite DNAs are critical chromosome elements that are recognized by satellite DNA-binding proteins and incorporated into chromocenters. We suggest that chromocenters function to preserve the entire chromosomal complement in a single nucleus, a fundamental and unquestioned feature of eukaryotic genomes. We speculate that the rapid divergence of satellite DNA sequences between closely related species results in discordant chromocenter function and may underlie speciation and hybrid incompatibility. © 2017 Jagannathan and Yamashita; Published by Cold Spring Harbor Laboratory Press.

  8. Localization of latency-associated nuclear antigen (LANA) on mitotic chromosomes

    Energy Technology Data Exchange (ETDEWEB)

    Rahayu, Retno; Ohsaki, Eriko [Division of Virology, Department of Microbiology and Immunology, Osaka University Graduate School of Medicine, 2-2 Yamada-oka, Suita, Osaka 565-0871 (Japan); Omori, Hiroko [Central Instrumentation Laboratory Research Institute for Microbial Diseases (BIKEN), Osaka University, Osaka 565-0871 (Japan); Ueda, Keiji, E-mail: kueda@virus.med.osaka-u.ac.jp [Division of Virology, Department of Microbiology and Immunology, Osaka University Graduate School of Medicine, 2-2 Yamada-oka, Suita, Osaka 565-0871 (Japan)

    2016-09-15

    In latent infection of Kaposi's sarcoma-associated herpesvirus (KSHV), viral gene expression is extremely limited and copy numbers of viral genomes remain constant. Latency-associated nuclear antigen (LANA) is known to have a role in maintaining viral genome copy numbers in growing cells. Several studies have shown that LANA is localized in particular regions on mitotic chromosomes, such as centromeres/pericentromeres. We independently examined the distinct localization of LANA on mitotic chromosomes during mitosis, using super-resolution laser confocal microscopy and correlative fluorescence microscopy–electron microscopy (FM-EM) analyses. We found that the majority of LANA were not localized at particular regions such as telomeres/peritelomeres, centromeres/pericentromeres, and cohesion sites, but at the bodies of condensed chromosomes. Thus, LANA may undergo various interactions with the host factors on the condensed chromosomes in order to tether the viral genome to mitotic chromosomes and realize faithful viral genome segregation during cell division. - Highlights: • This is the first report showing LANA dots on mitotic chromosomes by fluorescent microscopy followed by electron microscopy. • LANA dots localized randomly on condensed chromosomes other than centromere/pericentromere and telomere/peritelomre. • Cellular mitotic checkpoint should not be always involved in the segregation of KSHV genomes in the latency.

  9. Localization of latency-associated nuclear antigen (LANA) on mitotic chromosomes

    International Nuclear Information System (INIS)

    Rahayu, Retno; Ohsaki, Eriko; Omori, Hiroko; Ueda, Keiji

    2016-01-01

    In latent infection of Kaposi's sarcoma-associated herpesvirus (KSHV), viral gene expression is extremely limited and copy numbers of viral genomes remain constant. Latency-associated nuclear antigen (LANA) is known to have a role in maintaining viral genome copy numbers in growing cells. Several studies have shown that LANA is localized in particular regions on mitotic chromosomes, such as centromeres/pericentromeres. We independently examined the distinct localization of LANA on mitotic chromosomes during mitosis, using super-resolution laser confocal microscopy and correlative fluorescence microscopy–electron microscopy (FM-EM) analyses. We found that the majority of LANA were not localized at particular regions such as telomeres/peritelomeres, centromeres/pericentromeres, and cohesion sites, but at the bodies of condensed chromosomes. Thus, LANA may undergo various interactions with the host factors on the condensed chromosomes in order to tether the viral genome to mitotic chromosomes and realize faithful viral genome segregation during cell division. - Highlights: • This is the first report showing LANA dots on mitotic chromosomes by fluorescent microscopy followed by electron microscopy. • LANA dots localized randomly on condensed chromosomes other than centromere/pericentromere and telomere/peritelomre. • Cellular mitotic checkpoint should not be always involved in the segregation of KSHV genomes in the latency.

  10. Identification of a 450-bp region of human papillomavirus type 1 that promotes episomal replication in Saccharomyces cerevisiae

    International Nuclear Information System (INIS)

    Chattopadhyay, Anasuya; Schmidt, Martin C.; Khan, Saleem A.

    2005-01-01

    Human papillomaviruses (HPVs) replicate as nuclear plasmids in infected cells. Since the DNA replication machinery is generally conserved between humans and Saccharomyces cerevisiae, we studied whether HPV-1 DNA can replicate in yeast. Plasmids containing a selectable marker (with or without a yeast centromere) and either the full-length HPV-1 genome or various regions of the viral long control region (LCR) and the 3' end of the L1 gene were introduced into S. cerevisiae and their ability to replicate episomally was investigated. Our results show that HPV-1 sequences promote episomal replication of plasmids although the yeast centromere is required for plasmid retention. We have mapped the autonomously replicating sequence activity of HPV-1 DNA to a 450 base-pair sequence (HPV-1 nt 6783-7232) that includes 293 nucleotides from the 5' region of the viral LCR and 157 nucleotides from the 3' end of the L1 gene. The HPV-1 ARS does not include the binding sites for the viral E1 and E2 proteins, and these proteins are dispensable for replication in S. cerevisiae

  11. Genome analysis and comparative genomics of a Giardia intestinalis assemblage E isolate

    Directory of Open Access Journals (Sweden)

    Andersson Jan O

    2010-10-01

    Full Text Available Abstract Background Giardia intestinalis is a protozoan parasite that causes diarrhea in a wide range of mammalian species. To further understand the genetic diversity between the Giardia intestinalis species, we have performed genome sequencing and analysis of a wild-type Giardia intestinalis sample from the assemblage E group, isolated from a pig. Results We identified 5012 protein coding genes, the majority of which are conserved compared to the previously sequenced genomes of the WB and GS strains in terms of microsynteny and sequence identity. Despite this, there is an unexpectedly large number of chromosomal rearrangements and several smaller structural changes that are present in all chromosomes. Novel members of the VSP, NEK Kinase and HCMP gene families were identified, which may reveal possible mechanisms for host specificity and new avenues for antigenic variation. We used comparative genomics of the three diverse Giardia intestinalis isolates P15, GS and WB to define a core proteome for this species complex and to identify lineage-specific genes. Extensive analyses of polymorphisms in the core proteome of Giardia revealed differential rates of divergence among cellular processes. Conclusions Our results indicate that despite a well conserved core of genes there is significant genome variation between Giardia isolates, both in terms of gene content, gene polymorphisms, structural chromosomal variations and surface molecule repertoires. This study improves the annotation of the Giardia genomes and enables the identification of functionally important variation.

  12. The genome sequence of Caenorhabditis briggsae: a platform for comparative genomics.

    Directory of Open Access Journals (Sweden)

    Lincoln D Stein

    2003-11-01

    Full Text Available The soil nematodes Caenorhabditis briggsae and Caenorhabditis elegans diverged from a common ancestor roughly 100 million years ago and yet are almost indistinguishable by eye. They have the same chromosome number and genome sizes, and they occupy the same ecological niche. To explore the basis for this striking conservation of structure and function, we have sequenced the C. briggsae genome to a high-quality draft stage and compared it to the finished C. elegans sequence. We predict approximately 19,500 protein-coding genes in the C. briggsae genome, roughly the same as in C. elegans. Of these, 12,200 have clear C. elegans orthologs, a further 6,500 have one or more clearly detectable C. elegans homologs, and approximately 800 C. briggsae genes have no detectable matches in C. elegans. Almost all of the noncoding RNAs (ncRNAs known are shared between the two species. The two genomes exhibit extensive colinearity, and the rate of divergence appears to be higher in the chromosomal arms than in the centers. Operons, a distinctive feature of C. elegans, are highly conserved in C. briggsae, with the arrangement of genes being preserved in 96% of cases. The difference in size between the C. briggsae (estimated at approximately 104 Mbp and C. elegans (100.3 Mbp genomes is almost entirely due to repetitive sequence, which accounts for 22.4% of the C. briggsae genome in contrast to 16.5% of the C. elegans genome. Few, if any, repeat families are shared, suggesting that most were acquired after the two species diverged or are undergoing rapid evolution. Coclustering the C. elegans and C. briggsae proteins reveals 2,169 protein families of two or more members. Most of these are shared between the two species, but some appear to be expanding or contracting, and there seem to be as many as several hundred novel C. briggsae gene families. The C. briggsae draft sequence will greatly improve the annotation of the C. elegans genome. Based on similarity to C

  13. Trade-off between Transcriptome Plasticity and Genome Evolution in Cephalopods.

    Science.gov (United States)

    Liscovitch-Brauer, Noa; Alon, Shahar; Porath, Hagit T; Elstein, Boaz; Unger, Ron; Ziv, Tamar; Admon, Arie; Levanon, Erez Y; Rosenthal, Joshua J C; Eisenberg, Eli

    2017-04-06

    RNA editing, a post-transcriptional process, allows the diversification of proteomes beyond the genomic blueprint; however it is infrequently used among animals for this purpose. Recent reports suggesting increased levels of RNA editing in squids thus raise the question of the nature and effects of these events. We here show that RNA editing is particularly common in behaviorally sophisticated coleoid cephalopods, with tens of thousands of evolutionarily conserved sites. Editing is enriched in the nervous system, affecting molecules pertinent for excitability and neuronal morphology. The genomic sequence flanking editing sites is highly conserved, suggesting that the process confers a selective advantage. Due to the large number of sites, the surrounding conservation greatly reduces the number of mutations and genomic polymorphisms in protein-coding regions. This trade-off between genome evolution and transcriptome plasticity highlights the importance of RNA recoding as a strategy for diversifying proteins, particularly those associated with neural function. PAPERCLIP. Copyright © 2017 Elsevier Inc. All rights reserved.

  14. Sigma factors in a thousand E. coli genomes

    DEFF Research Database (Denmark)

    Cook, Helen Victoria; Ussery, David

    2013-01-01

    , 2013), only less than half (983) are of sufficient quality to use in comparative genomic work. Unfortunately, even some of the ‘complete’ E. coli genomes are in pieces, and a few ‘draft’ genomes are good quality. Six of the seven known sigma factors in E. coli strain K‐12 are extremely well conserved...

  15. A Web-Based Comparative Genomics Tutorial for Investigating Microbial Genomes

    Directory of Open Access Journals (Sweden)

    Michael Strong

    2009-12-01

    Full Text Available As the number of completely sequenced microbial genomes continues to rise at an impressive rate, it is important to prepare students with the skills necessary to investigate microorganisms at the genomic level. As a part of the core curriculum for first-year graduate students in the biological sciences, we have implemented a web-based tutorial to introduce students to the fields of comparative and functional genomics. The tutorial focuses on recent computational methods for identifying functionally linked genes and proteins on a genome-wide scale and was used to introduce students to the Rosetta Stone, Phylogenetic Profile, conserved Gene Neighbor, and Operon computational methods. Students learned to use a number of publicly available web servers and databases to identify functionally linked genes in the Escherichia coli genome, with emphasis on genome organization and operon structure. The overall effectiveness of the tutorial was assessed based on student evaluations and homework assignments. The tutorial is available to other educators at http://www.doe-mbi.ucla.edu/~strong/m253.php.

  16. Genomic profiling of rice sperm cell transcripts reveals conserved and distinct elements in the flowering plant male germ lineage.

    Science.gov (United States)

    Russell, Scott D; Gou, Xiaoping; Wong, Chui E; Wang, Xinkun; Yuan, Tong; Wei, Xiaoping; Bhalla, Prem L; Singh, Mohan B

    2012-08-01

    Genomic assay of sperm cell RNA provides insight into functional control, modes of regulation, and contributions of male gametes to double fertilization. Sperm cells of rice (Oryza sativa) were isolated from field-grown, disease-free plants and RNA was processed for use with the full-genome Affymetrix microarray. Comparison with Gene Expression Omnibus (GEO) reference arrays confirmed expressionally distinct gene profiles. A total of 10,732 distinct gene sequences were detected in sperm cells, of which 1668 were not expressed in pollen or seedlings. Pathways enriched in male germ cells included ubiquitin-mediated pathways, pathways involved in chromatin modeling including histones, histone modification and nonhistone epigenetic modification, and pathways related to RNAi and gene silencing. Genome-wide expression patterns in angiosperm sperm cells indicate common and divergent themes in the male germline that appear to be largely self-regulating through highly up-regulated chromatin modification pathways. A core of highly conserved genes appear common to all sperm cells, but evidence is still emerging that another class of genes have diverged in expression between monocots and dicots since their divergence. Sperm cell transcripts present at fusion may be transmitted through plasmogamy during double fertilization to effect immediate post-fertilization expression of early embryo and (or) endosperm development. © 2012 The Authors. New Phytologist © 2012 New Phytologist Trust.

  17. The complete mitochondrial genome of eastern lowland gorilla, Gorilla beringei graueri, and comparative mitochondrial genomics of Gorilla species.

    Science.gov (United States)

    Hu, Xiao-di; Gao, Li-zhi

    2016-01-01

    In this study, we determined the complete mitochondrial (mt) genome of eastern lowland gorilla, Gorilla beringei graueri for the first time. The total genome was 16,416 bp in length. It contained a total of 13 protein-coding genes, 22 transfer RNA genes, 2 ribosomal RNA genes and 1 control region (D-loop region). The base composition was A (30.88%), G (13.10%), C (30.89%) and T (25.13%), indicating that the percentage of A+T (56.01%) was higher than G+C (43.99%). Comparisons with the other publicly available Gorilla mitogenome showed the conservation of gene order and base compositions but a bunch of nucleotide diversity. This complete mitochondrial genome sequence will provide valuable genetic information for further studies on conservation genetics of eastern lowland gorilla.

  18. Bat biology, genomes, and the Bat1K project

    DEFF Research Database (Denmark)

    Teeling, Emma C; Vernes, Sonja C; Dávalos, Liliana M

    2018-01-01

    and endangered. Here we announce Bat1K, an initiative to sequence the genomes of all living bat species (n∼1,300) to chromosome-level assembly. The Bat1K genome consortium unites bat biologists (>148 members as of writing), computational scientists, conservation organizations, genome technologists, and any...

  19. Immunohistochemical Assessment of Expression of Centromere Protein—A (CENPA) in Human Invasive Breast Cancer

    International Nuclear Information System (INIS)

    Rajput, Ashish B.; Hu, Nianping; Varma, Sonal; Chen, Chien-Hung; Ding, Keyue; Park, Paul C.; Chapman, Judy-Anne W.; SenGupta, Sandip K.; Madarnas, Yolanda; Elliott, Bruce E.; Feilotter, Harriet E.

    2011-01-01

    Abnormal cell division leading to the gain or loss of entire chromosomes and consequent genetic instability is a hallmark of cancer. Centromere protein –A (CENPA) is a centromere-specific histone-H3-like variant gene involved in regulating chromosome segregation during cell division. CENPA is one of the genes included in some of the commercially available RNA based prognostic assays for breast cancer (BCa)—the 70 gene signature MammaPrint ® and the five gene Molecular Grade Index (MGI SM ). Our aim was to assess the immunohistochemical (IHC) expression of CENPA in normal and malignant breast tissue. Clinically annotated triplicate core tissue microarrays of 63 invasive BCa and 20 normal breast samples were stained with a monoclonal antibody against CENPA and scored for percentage of visibly stained nuclei. Survival analyses with Kaplan–Meier (KM) estimate and Cox proportional hazards regression models were applied to assess the associations between CENPA expression and disease free survival (DFS). Average percentage of nuclei visibly stained with CENPA antibody was significantly higher (p = 0.02) in BCa than normal tissue. The 3-year DFS in tumors over-expressing CENPA (>50% stained nuclei) was 79% compared to 85% in low expression tumors (<50% stained nuclei). On multivariate analysis, IHC expression of CENPA showed weak association with DFS (HR > 60.07; p = 0.06) within our small cohort. To the best of our knowledge, this is the first published report evaluating the implications of increased IHC expression of CENPA in paraffin embedded breast tissue samples. Our finding that increased CENPA expression may be associated with shorter DFS in BCa supports its exploration as a potential prognostic biomarker

  20. Immunohistochemical Assessment of Expression of Centromere Protein—A (CENPA) in Human Invasive Breast Cancer

    Energy Technology Data Exchange (ETDEWEB)

    Rajput, Ashish B. [Department of Pathology and Molecular Medicine, Queen' s University, Kingston, ON K7L 3N6 (Canada); Hu, Nianping [Cancer Research institute, Queen' s University, Kingston, ON K7L 3N6 (Canada); Varma, Sonal; Chen, Chien-Hung [Department of Pathology and Molecular Medicine, Queen' s University, Kingston, ON K7L 3N6 (Canada); Ding, Keyue [NCIC Clinical Trials Group, Queen' s University, Kingston, ON K7L 3N6 (Canada); Park, Paul C. [Department of Pathology and Molecular Medicine, Queen' s University, Kingston, ON K7L 3N6 (Canada); Chapman, Judy-Anne W. [NCIC Clinical Trials Group, Queen' s University, Kingston, ON K7L 3N6 (Canada); SenGupta, Sandip K. [Department of Pathology and Molecular Medicine, Queen' s University, Kingston, ON K7L 3N6 (Canada); Madarnas, Yolanda [Cancer Research institute, Queen' s University, Kingston, ON K7L 3N6 (Canada); Department of Oncology, Cancer Center of Southeastern Ontario, Kingston, ON K7L 2V7 (Canada); Elliott, Bruce E.; Feilotter, Harriet E., E-mail: feilotth@kgh.kari.net [Department of Pathology and Molecular Medicine, Queen' s University, Kingston, ON K7L 3N6 (Canada)

    2011-12-06

    Abnormal cell division leading to the gain or loss of entire chromosomes and consequent genetic instability is a hallmark of cancer. Centromere protein –A (CENPA) is a centromere-specific histone-H3-like variant gene involved in regulating chromosome segregation during cell division. CENPA is one of the genes included in some of the commercially available RNA based prognostic assays for breast cancer (BCa)—the 70 gene signature MammaPrint{sup ®} and the five gene Molecular Grade Index (MGI{sup SM}). Our aim was to assess the immunohistochemical (IHC) expression of CENPA in normal and malignant breast tissue. Clinically annotated triplicate core tissue microarrays of 63 invasive BCa and 20 normal breast samples were stained with a monoclonal antibody against CENPA and scored for percentage of visibly stained nuclei. Survival analyses with Kaplan–Meier (KM) estimate and Cox proportional hazards regression models were applied to assess the associations between CENPA expression and disease free survival (DFS). Average percentage of nuclei visibly stained with CENPA antibody was significantly higher (p = 0.02) in BCa than normal tissue. The 3-year DFS in tumors over-expressing CENPA (>50% stained nuclei) was 79% compared to 85% in low expression tumors (<50% stained nuclei). On multivariate analysis, IHC expression of CENPA showed weak association with DFS (HR > 60.07; p = 0.06) within our small cohort. To the best of our knowledge, this is the first published report evaluating the implications of increased IHC expression of CENPA in paraffin embedded breast tissue samples. Our finding that increased CENPA expression may be associated with shorter DFS in BCa supports its exploration as a potential prognostic biomarker.

  1. Defining functional DNA elements in the human genome

    Science.gov (United States)

    Kellis, Manolis; Wold, Barbara; Snyder, Michael P.; Bernstein, Bradley E.; Kundaje, Anshul; Marinov, Georgi K.; Ward, Lucas D.; Birney, Ewan; Crawford, Gregory E.; Dekker, Job; Dunham, Ian; Elnitski, Laura L.; Farnham, Peggy J.; Feingold, Elise A.; Gerstein, Mark; Giddings, Morgan C.; Gilbert, David M.; Gingeras, Thomas R.; Green, Eric D.; Guigo, Roderic; Hubbard, Tim; Kent, Jim; Lieb, Jason D.; Myers, Richard M.; Pazin, Michael J.; Ren, Bing; Stamatoyannopoulos, John A.; Weng, Zhiping; White, Kevin P.; Hardison, Ross C.

    2014-01-01

    With the completion of the human genome sequence, attention turned to identifying and annotating its functional DNA elements. As a complement to genetic and comparative genomics approaches, the Encyclopedia of DNA Elements Project was launched to contribute maps of RNA transcripts, transcriptional regulator binding sites, and chromatin states in many cell types. The resulting genome-wide data reveal sites of biochemical activity with high positional resolution and cell type specificity that facilitate studies of gene regulation and interpretation of noncoding variants associated with human disease. However, the biochemically active regions cover a much larger fraction of the genome than do evolutionarily conserved regions, raising the question of whether nonconserved but biochemically active regions are truly functional. Here, we review the strengths and limitations of biochemical, evolutionary, and genetic approaches for defining functional DNA segments, potential sources for the observed differences in estimated genomic coverage, and the biological implications of these discrepancies. We also analyze the relationship between signal intensity, genomic coverage, and evolutionary conservation. Our results reinforce the principle that each approach provides complementary information and that we need to use combinations of all three to elucidate genome function in human biology and disease. PMID:24753594

  2. A manually annotated Actinidia chinensis var. chinensis (kiwifruit) genome highlights the challenges associated with draft genomes and gene prediction in plants.

    Science.gov (United States)

    Pilkington, Sarah M; Crowhurst, Ross; Hilario, Elena; Nardozza, Simona; Fraser, Lena; Peng, Yongyan; Gunaseelan, Kularajathevan; Simpson, Robert; Tahir, Jibran; Deroles, Simon C; Templeton, Kerry; Luo, Zhiwei; Davy, Marcus; Cheng, Canhong; McNeilage, Mark; Scaglione, Davide; Liu, Yifei; Zhang, Qiong; Datson, Paul; De Silva, Nihal; Gardiner, Susan E; Bassett, Heather; Chagné, David; McCallum, John; Dzierzon, Helge; Deng, Cecilia; Wang, Yen-Yi; Barron, Lorna; Manako, Kelvina; Bowen, Judith; Foster, Toshi M; Erridge, Zoe A; Tiffin, Heather; Waite, Chethi N; Davies, Kevin M; Grierson, Ella P; Laing, William A; Kirk, Rebecca; Chen, Xiuyin; Wood, Marion; Montefiori, Mirco; Brummell, David A; Schwinn, Kathy E; Catanach, Andrew; Fullerton, Christina; Li, Dawei; Meiyalaghan, Sathiyamoorthy; Nieuwenhuizen, Niels; Read, Nicola; Prakash, Roneel; Hunter, Don; Zhang, Huaibi; McKenzie, Marian; Knäbel, Mareike; Harris, Alastair; Allan, Andrew C; Gleave, Andrew; Chen, Angela; Janssen, Bart J; Plunkett, Blue; Ampomah-Dwamena, Charles; Voogd, Charlotte; Leif, Davin; Lafferty, Declan; Souleyre, Edwige J F; Varkonyi-Gasic, Erika; Gambi, Francesco; Hanley, Jenny; Yao, Jia-Long; Cheung, Joey; David, Karine M; Warren, Ben; Marsh, Ken; Snowden, Kimberley C; Lin-Wang, Kui; Brian, Lara; Martinez-Sanchez, Marcela; Wang, Mindy; Ileperuma, Nadeesha; Macnee, Nikolai; Campin, Robert; McAtee, Peter; Drummond, Revel S M; Espley, Richard V; Ireland, Hilary S; Wu, Rongmei; Atkinson, Ross G; Karunairetnam, Sakuntala; Bulley, Sean; Chunkath, Shayhan; Hanley, Zac; Storey, Roy; Thrimawithana, Amali H; Thomson, Susan; David, Charles; Testolin, Raffaele; Huang, Hongwen; Hellens, Roger P; Schaffer, Robert J

    2018-04-16

    Most published genome sequences are drafts, and most are dominated by computational gene prediction. Draft genomes typically incorporate considerable sequence data that are not assigned to chromosomes, and predicted genes without quality confidence measures. The current Actinidia chinensis (kiwifruit) 'Hongyang' draft genome has 164 Mb of sequences unassigned to pseudo-chromosomes, and omissions have been identified in the gene models. A second genome of an A. chinensis (genotype Red5) was fully sequenced. This new sequence resulted in a 554.0 Mb assembly with all but 6 Mb assigned to pseudo-chromosomes. Pseudo-chromosomal comparisons showed a considerable number of translocation events have occurred following a whole genome duplication (WGD) event some consistent with centromeric Robertsonian-like translocations. RNA sequencing data from 12 tissues and ab initio analysis informed a genome-wide manual annotation, using the WebApollo tool. In total, 33,044 gene loci represented by 33,123 isoforms were identified, named and tagged for quality of evidential support. Of these 3114 (9.4%) were identical to a protein within 'Hongyang' The Kiwifruit Information Resource (KIR v2). Some proportion of the differences will be varietal polymorphisms. However, as most computationally predicted Red5 models required manual re-annotation this proportion is expected to be small. The quality of the new gene models was tested by fully sequencing 550 cloned 'Hort16A' cDNAs and comparing with the predicted protein models for Red5 and both the original 'Hongyang' assembly and the revised annotation from KIR v2. Only 48.9% and 63.5% of the cDNAs had a match with 90% identity or better to the original and revised 'Hongyang' annotation, respectively, compared with 90.9% to the Red5 models. Our study highlights the need to take a cautious approach to draft genomes and computationally predicted genes. Our use of the manual annotation tool WebApollo facilitated manual checking and

  3. Assembly of the Genome of the Disease Vector Aedes aegypti onto a Genetic Linkage Map Allows Mapping of Genes Affecting Disease Transmission

    KAUST Repository

    Juneja, Punita

    2014-01-30

    The mosquito Aedes aegypti transmits some of the most important human arboviruses, including dengue, yellow fever and chikungunya viruses. It has a large genome containing many repetitive sequences, which has resulted in the genome being poorly assembled - there are 4,758 scaffolds, few of which have been assigned to a chromosome. To allow the mapping of genes affecting disease transmission, we have improved the genome assembly by scoring a large number of SNPs in recombinant progeny from a cross between two strains of Ae. aegypti, and used these to generate a genetic map. This revealed a high rate of misassemblies in the current genome, where, for example, sequences from different chromosomes were found on the same scaffold. Once these were corrected, we were able to assign 60% of the genome sequence to chromosomes and approximately order the scaffolds along the chromosome. We found that there are very large regions of suppressed recombination around the centromeres, which can extend to as much as 47% of the chromosome. To illustrate the utility of this new genome assembly, we mapped a gene that makes Ae. aegypti resistant to the human parasite Brugia malayi, and generated a list of candidate genes that could be affecting the trait. © 2014 Juneja et al.

  4. The genomes and comparative genomics of Lactobacillus delbrueckii phages.

    Science.gov (United States)

    Riipinen, Katja-Anneli; Forsman, Päivi; Alatossava, Tapani

    2011-07-01

    Lactobacillus delbrueckii phages are a great source of genetic diversity. Here, the genome sequences of Lb. delbrueckii phages LL-Ku, c5 and JCL1032 were analyzed in detail, and the genetic diversity of Lb. delbrueckii phages belonging to different taxonomic groups was explored. The lytic isometric group b phages LL-Ku (31,080 bp) and c5 (31,841 bp) showed a minimum nucleotide sequence identity of 90% over about three-fourths of their genomes. The genomic locations of their lysis modules were unique, and the genomes featured several putative overlapping transcription units of genes. LL-Ku and c5 virions displayed peptidoglycan hydrolytic activity associated with a ~36-kDa protein similar in size to the endolysin. Unexpectedly, the 49,433-bp genome of the prolate phage JCL1032 (temperate, group c) revealed a conserved gene order within its structural genes. Lb. delbrueckii phages representing groups a (a phage LL-H), b and c possessed only limited protein sequence homology. Genomic comparison of LL-Ku and c5 suggested that diversification of Lb. delbrueckii phages is mainly due to insertions, deletions and recombination. For the first time, the complete genome sequences of group b and c Lb. delbrueckii phages are reported.

  5. Analysis of the grape MYB R2R3 subfamily reveals expanded wine quality-related clades and conserved gene structure organization across Vitis and Arabidopsis genomes

    Science.gov (United States)

    Matus, José Tomás; Aquea, Felipe; Arce-Johnson, Patricio

    2008-01-01

    Background The MYB superfamily constitutes the most abundant group of transcription factors described in plants. Members control processes such as epidermal cell differentiation, stomatal aperture, flavonoid synthesis, cold and drought tolerance and pathogen resistance. No genome-wide characterization of this family has been conducted in a woody species such as grapevine. In addition, previous analysis of the recently released grape genome sequence suggested expansion events of several gene families involved in wine quality. Results We describe and classify 108 members of the grape R2R3 MYB gene subfamily in terms of their genomic gene structures and similarity to their putative Arabidopsis thaliana orthologues. Seven gene models were derived and analyzed in terms of gene expression and their DNA binding domain structures. Despite low overall sequence homology in the C-terminus of all proteins, even in those with similar functions across Arabidopsis and Vitis, highly conserved motif sequences and exon lengths were found. The grape epidermal cell fate clade is expanded when compared with the Arabidopsis and rice MYB subfamilies. Two anthocyanin MYBA related clusters were identified in chromosomes 2 and 14, one of which includes the previously described grape colour locus. Tannin related loci were also detected with eight candidate homologues in chromosomes 4, 9 and 11. Conclusion This genome wide transcription factor analysis in Vitis suggests that clade-specific grape R2R3 MYB genes are expanded while other MYB genes could be well conserved compared to Arabidopsis. MYB gene abundance, homology and orientation within particular loci also suggests that expanded MYB clades conferring quality attributes of grapes and wines, such as colour and astringency, could possess redundant, overlapping and cooperative functions. PMID:18647406

  6. On-line sorting of human chromosomes by centromeric index, and identification of sorted populations by GTG-banding and fluorescent in situ hybridization

    NARCIS (Netherlands)

    Boschman, G. A.; Rens, W.; Manders, E.; van Oven, C.; Barendsen, G. W.; Aten, J. A.

    1990-01-01

    Using slit-scan flow cytometry, the shape of human metaphase chromosomes, as expressed in their centromeric index (CI), and the DNA content of the chromosomes have been used as parameters in bivariate flow karyotyping. The resolution of the DNA vs CI flow karyogram of the larger chromosomes up to

  7. Comparative Genomics of Carp Herpesviruses

    Science.gov (United States)

    Kurobe, Tomofumi; Gatherer, Derek; Cunningham, Charles; Korf, Ian; Fukuda, Hideo; Hedrick, Ronald P.; Waltzek, Thomas B.

    2013-01-01

    Three alloherpesviruses are known to cause disease in cyprinid fish: cyprinid herpesviruses 1 and 3 (CyHV1 and CyHV3) in common carp and koi and cyprinid herpesvirus 2 (CyHV2) in goldfish. We have determined the genome sequences of CyHV1 and CyHV2 and compared them with the published CyHV3 sequence. The CyHV1 and CyHV2 genomes are 291,144 and 290,304 bp, respectively, in size, and thus the CyHV3 genome, at 295,146 bp, remains the largest recorded among the herpesviruses. Each of the three genomes consists of a unique region flanked at each terminus by a sizeable direct repeat. The CyHV1, CyHV2, and CyHV3 genomes are predicted to contain 137, 150, and 155 unique, functional protein-coding genes, respectively, of which six, four, and eight, respectively, are duplicated in the terminal repeat. The three viruses share 120 orthologous genes in a largely colinear arrangement, of which up to 55 are also conserved in the other member of the genus Cyprinivirus, anguillid herpesvirus 1. Twelve genes are conserved convincingly in all sequenced alloherpesviruses, and two others are conserved marginally. The reference CyHV3 strain has been reported to contain five fragmented genes that are presumably nonfunctional. The CyHV2 strain has two fragmented genes, and the CyHV1 strain has none. CyHV1, CyHV2, and CyHV3 have five, six, and five families of paralogous genes, respectively. One family unique to CyHV1 is related to cellular JUNB, which encodes a transcription factor involved in oncogenesis. To our knowledge, this is the first time that JUNB-related sequences have been reported in a herpesvirus. PMID:23269803

  8. V-SINEs: A New Superfamily of Vertebrate SINEs That Are Widespread in Vertebrate Genomes and Retain a Strongly Conserved Segment within Each Repetitive Unit

    Science.gov (United States)

    Ogiwara, Ikuo; Miya, Masaki; Ohshima, Kazuhiko; Okada, Norihiro

    2002-01-01

    We have identified a new superfamily of vertebrate short interspersed repetitive elements (SINEs), designated V-SINEs, that are widespread in fishes and frogs. Each V-SINE includes a central conserved domain preceded by a 5′-end tRNA-related region and followed by a potentially recombinogenic (TG)n tract, with a 3′ tail derived from the 3′ untranslated region (UTR) of the corresponding partner long interspersed repetitive element (LINE) that encodes a functional reverse transcriptase. The central domain is strongly conserved and is even found in SINEs in the lamprey genome, suggesting that V-SINEs might be ∼550 Myr old or older in view of the timing of divergence of the lamprey lineage from the bony fish lineage. The central conserved domain might have been subject to some form of positive selection. Although the contemporary 3′ tails of V-SINEs differ from one another, it is possible that the original 3′ tail might have been replaced, via recombination, by the 3′ tails of more active partner LINEs, thereby retaining retropositional activity and the ability to survive for long periods on the evolutionary time scale. It seems plausible that V-SINEs may have some function(s) that have been maintained by the coevolution of SINEs and LINEs during the evolution of vertebrates. [The sequences reported in this paper have been deposited in the DDBJ/GenBank database under accession nos. AB072981–AB073004. Supplemental figures are available online at http://www.genome.org.] PMID:11827951

  9. Prospects for Genomic Research in Forestry

    Directory of Open Access Journals (Sweden)

    K. V. Krutovsky

    2014-08-01

    Full Text Available Conifers are keystone species of boreal forests. Their whole genome sequencing, assembly and annotation will allow us to understand the evolution of the complex ancient giant conifer genomes that are 4 times larger in larch and 7–9 times larger in pines than the human genome. Genomic studies will allow also to obtain important whole genome sequence data and develop highly polymorphic and informative genetic markers, such as microsatellites and single nucleotide polymorphisms (SNPs that can be efficiently used in timber origin identification, for genetic variation monitoring, to study local and climate change adaptation and in tree improvement and conservation programs.

  10. The utility of transcriptomics in fish conservation.

    Science.gov (United States)

    Connon, Richard E; Jeffries, Ken M; Komoroske, Lisa M; Todgham, Anne E; Fangue, Nann A

    2018-01-29

    There is growing recognition of the need to understand the mechanisms underlying organismal resilience (i.e. tolerance, acclimatization) to environmental change to support the conservation management of sensitive and economically important species. Here, we discuss how functional genomics can be used in conservation biology to provide a cellular-level understanding of organismal responses to environmental conditions. In particular, the integration of transcriptomics with physiological and ecological research is increasingly playing an important role in identifying functional physiological thresholds predictive of compensatory responses and detrimental outcomes, transforming the way we can study issues in conservation biology. Notably, with technological advances in RNA sequencing, transcriptome-wide approaches can now be applied to species where no prior genomic sequence information is available to develop species-specific tools and investigate sublethal impacts that can contribute to population declines over generations and undermine prospects for long-term conservation success. Here, we examine the use of transcriptomics as a means of determining organismal responses to environmental stressors and use key study examples of conservation concern in fishes to highlight the added value of transcriptome-wide data to the identification of functional response pathways. Finally, we discuss the gaps between the core science and policy frameworks and how thresholds identified through transcriptomic evaluations provide evidence that can be more readily used by resource managers. © 2018. Published by The Company of Biologists Ltd.

  11. Comparative Genome Viewer

    International Nuclear Information System (INIS)

    Molineris, I.; Sales, G.

    2009-01-01

    The amount of information about genomes, both in the form of complete sequences and annotations, has been exponentially increasing in the last few years. As a result there is the need for tools providing a graphical representation of such information that should be comprehensive and intuitive. Visual representation is especially important in the comparative genomics field since it should provide a combined view of data belonging to different genomes. We believe that existing tools are limited in this respect as they focus on a single genome at a time (conservation histograms) or compress alignment representation to a single dimension. We have therefore developed a web-based tool called Comparative Genome Viewer (Cgv): it integrates a bidimensional representation of alignments between two regions, both at small and big scales, with the richness of annotations present in other genome browsers. We give access to our system through a web-based interface that provides the user with an interactive representation that can be updated in real time using the mouse to move from region to region and to zoom in on interesting details.

  12. Comparative Genomics Reveals High Genomic Diversity in the Genus Photobacterium.

    Science.gov (United States)

    Machado, Henrique; Gram, Lone

    2017-01-01

    Vibrionaceae is a large marine bacterial family, which can constitute up to 50% of the prokaryotic population in marine waters. Photobacterium is the second largest genus in the family and we used comparative genomics on 35 strains representing 16 of the 28 species described so far, to understand the genomic diversity present in the Photobacterium genus. Such understanding is important for ecophysiology studies of the genus. We used whole genome sequences to evaluate phylogenetic relationships using several analyses (16S rRNA, MLSA, fur , amino-acid usage, ANI), which allowed us to identify two misidentified strains. Genome analyses also revealed occurrence of higher and lower GC content clades, correlating with phylogenetic clusters. Pan- and core-genome analysis revealed the conservation of 25% of the genome throughout the genus, with a large and open pan-genome. The major source of genomic diversity could be traced to the smaller chromosome and plasmids. Several of the physiological traits studied in the genus did not correlate with phylogenetic data. Since horizontal gene transfer (HGT) is often suggested as a source of genetic diversity and a potential driver of genomic evolution in bacterial species, we looked into evidence of such in Photobacterium genomes. Genomic islands were the source of genomic differences between strains of the same species. Also, we found transposase genes and CRISPR arrays that suggest multiple encounters with foreign DNA. Presence of genomic exchange traits was widespread and abundant in the genus, suggesting a role in genomic evolution. The high genetic variability and indications of genetic exchange make it difficult to elucidate genome evolutionary paths and raise the awareness of the roles of foreign DNA in the genomic evolution of environmental organisms.

  13. Divergence of RNA polymerase ? subunits in angiosperm plastid genomes is mediated by genomic rearrangement

    OpenAIRE

    Blazier, J. Chris; Ruhlman, Tracey A.; Weng, Mao-Lun; Rehman, Sumaiyah K.; Sabir, Jamal S. M.; Jansen, Robert K.

    2016-01-01

    Genes for the plastid-encoded RNA polymerase (PEP) persist in the plastid genomes of all photosynthetic angiosperms. However, three unrelated lineages (Annonaceae, Passifloraceae and Geraniaceae) have been identified with unusually divergent open reading frames (ORFs) in the conserved region of rpoA, the gene encoding the PEP ? subunit. We used sequence-based approaches to evaluate whether these genes retain function. Both gene sequences and complete plastid genome sequences were assembled an...

  14. Generation of an approximately 2.4 Mb human X centromere-based minichromosome by targeted telomere-associated chromosome fragmentation in DT40.

    Science.gov (United States)

    Mills, W; Critcher, R; Lee, C; Farr, C J

    1999-05-01

    A linear mammalian artificial chromosome (MAC) will require at least three types of functional element: a centromere, two telomeres and origins of replication. As yet, our understanding of these elements, as well as many other aspects of structure and organization which may be critical for a fully functional mammalian chromosome, remains poor. As a way of defining these various requirements, minichromosome reagents are being developed and analysed. Approaches for minichromosome generation fall into two broad categories: de novo assembly from candidate DNA sequences, or the fragmentation of an existing chromosome to reduce it to a minimal size. Here we describe the generation of a human minichromosome using the latter, top-down, approach. A human X chromosome, present in a DT40-human microcell hybrid, has been manipulated using homologous recombination and the targeted seeding of a de novo telomere. This strategy has generated a linear approximately 2.4 Mb human X centromere-based minichromosome capped by two artificially seeded telomeres: one immediately flanking the centromeric alpha-satellite DNA and the other targeted to the zinc finger gene ZXDA in Xp11.21. The chromosome retains an alpha-satellite domain of approximately 1. 8 Mb, a small array of gamma-satellite repeat ( approximately 40 kb) and approximately 400 kb of Xp proximal DNA sequence. The mitotic stability of this minichromosome has been examined, both in DT40 and following transfer into hamster and human cell lines. In all three backgrounds, the minichromosome is retained efficiently, but in the human and hamster microcell hybrids its copy number is poorly regulated. This approach of engineering well-defined chromosome reagents will allow key questions in MAC development (such as whether a lower size limit exists) to be addressed. In addition, the 2.4 Mb minichromosome described here has potential to be developed as a vector for gene delivery.

  15. Is It Time for Synthetic Biodiversity Conservation?

    Science.gov (United States)

    Piaggio, Antoinette J; Segelbacher, Gernot; Seddon, Philip J; Alphey, Luke; Bennett, Elizabeth L; Carlson, Robert H; Friedman, Robert M; Kanavy, Dona; Phelan, Ryan; Redford, Kent H; Rosales, Marina; Slobodian, Lydia; Wheeler, Keith

    2017-02-01

    Evidence indicates that, despite some critical successes, current conservation approaches are not slowing the overall rate of biodiversity loss. The field of synthetic biology, which is capable of altering natural genomes with extremely precise editing, might offer the potential to resolve some intractable conservation problems (e.g., invasive species or pathogens). However, it is our opinion that there has been insufficient engagement by the conservation community with practitioners of synthetic biology. We contend that rapid, large-scale engagement of these two communities is urgently needed to avoid unintended and deleterious ecological consequences. To this point we describe case studies where synthetic biology is currently being applied to conservation, and we highlight the benefits to conservation biologists from engaging with this emerging technology. Published by Elsevier Ltd.

  16. Molecular cytogenetic of the Amoy croaker, Argyrosomus amoyensis (Teleostei, Sciaenidae)

    Science.gov (United States)

    Liao, Mengxiang; Zheng, Jiao; Wang, Zhiyong; Wang, Yilei; Zhang, Jing; Cai, Mingyi

    2017-08-01

    The family Sciaenidae is remarkable for its species richness and economic importance. However, the cytogenetic data available in this fish group are still limited, especially those obtained using fluorescence in situ hybridization (FISH). In the present study, the chromosome characteristics of a sciaenid species, Argyrosomus amoyensis, were examined with several cytogenetic methods, including dual-FISH with 18S and 5S rDNA probes, and a self-genomic in situ hybridization procedure (Self-GISH). The karyotype of A. amoyensis comprised 2n=48 acrocentric chromosomes. A single pair of nucleolar organizer regions (NORs) was located at the proximal position of chromosome 1, which was positive for silver nitrate impregnation (AgNO3) staining and denaturation-propidium iodide (DPI) staining but negative for Giemsa staining and 4',6-diamidino-2-phenylindole (DAPI) staining, and was confirmed by FISH with 18S rDNA probes. The 5S rDNA sites were located at the centromeric region of chromosome 3. Telomeric FISH signals were detected at all chromosome ends with different intensities, but internal telomeric sequences (ITSs) were not found. Self-GISH resulted in strong signals distributed at the centromeric regions of all chromosomes. C-banding revealed not only centromeric heterochromatin, but also heterochromatin that located on NORs, in interstitial and distal telomeric regions of specific chromosomes. These results suggest that the karyotype of Amoy croaker was relatively conserved and primitive. By comparison with the reported cytogenetic data of other sciaenids, it can be deduced that although the karyotypic macrostructure and chromosomal localization of 18S rDNA are conserved, the distribution of 5S rDNA varies dynamically among sciaenid species. Thus, the 5S rDNA sites may have different evolutionary dynamics in relation to other chromosomal regions, and have the potential to be effective cytotaxonomic markers in Sciaenidae.

  17. Genomic diversity in two related plant species with and without sex chromosomes--Silene latifolia and S. vulgaris.

    Directory of Open Access Journals (Sweden)

    Radim Cegan

    Full Text Available Genome size evolution is a complex process influenced by polyploidization, satellite DNA accumulation, and expansion of retroelements. How this process could be affected by different reproductive strategies is still poorly understood.We analyzed differences in the number and distribution of major repetitive DNA elements in two closely related species, Silene latifolia and S. vulgaris. Both species are diploid and possess the same chromosome number (2n = 24, but differ in their genome size and mode of reproduction. The dioecious S. latifolia (1C = 2.70 pg DNA possesses sex chromosomes and its genome is 2.5× larger than that of the gynodioecious S. vulgaris (1C = 1.13 pg DNA, which does not possess sex chromosomes. We discovered that the genome of S. latifolia is larger mainly due to the expansion of Ogre retrotransposons. Surprisingly, the centromeric STAR-C and TR1 tandem repeats were found to be more abundant in S. vulgaris, the species with the smaller genome. We further examined the distribution of major repetitive sequences in related species in the Caryophyllaceae family. The results of FISH (fluorescence in situ hybridization on mitotic chromosomes with the Retand element indicate that large rearrangements occurred during the evolution of the Caryophyllaceae family.Our data demonstrate that the evolution of genome size in the genus Silene is accompanied by the expansion of different repetitive elements with specific patterns in the dioecious species possessing the sex chromosomes.

  18. Conserved domains and SINE diversity during animal evolution.

    Science.gov (United States)

    Luchetti, Andrea; Mantovani, Barbara

    2013-10-01

    Eukaryotic genomes harbour a number of mobile genetic elements (MGEs); moving from one genomic location to another, they are known to impact on the host genome. Short interspersed elements (SINEs) are well-represented, non-autonomous retroelements and they are likely the most diversified MGEs. In some instances, sequence domains conserved across unrelated SINEs have been identified; remarkably, one of these, called Nin, has been conserved since the Radiata-Bilateria splitting. Here we report on two new domains: Inv, derived from Nin, identified in insects and in deuterostomes, and Pln, restricted to polyneopteran insects. The identification of Inv and Pln sequences allowed us to retrieve new SINEs, two in insects and one in a hemichordate. The diverse structural combination of the different domains in different SINE families, during metazoan evolution, offers a clearer view of SINE diversity and their frequent de novo emergence through module exchange, possibly underlying the high evolutionary success of SINEs. © 2013 Elsevier Inc. All rights reserved.

  19. Are ribosomal DNA clusters rearrangement hotspots? A case study in the genus Mus (Rodentia, Muridae

    Directory of Open Access Journals (Sweden)

    Douzery Emmanuel JP

    2011-05-01

    Full Text Available Abstract Background Recent advances in comparative genomics have considerably improved our knowledge of the evolution of mammalian karyotype architecture. One of the breakthroughs was the preferential localization of evolutionary breakpoints in regions enriched in repetitive sequences (segmental duplications, telomeres and centromeres. In this context, we investigated the contribution of ribosomal genes to genome reshuffling since they are generally located in pericentromeric or subtelomeric regions, and form repeat clusters on different chromosomes. The target model was the genus Mus which exhibits a high rate of karyotypic change, a large fraction of which involves centromeres. Results The chromosomal distribution of rDNA clusters was determined by in situ hybridization of mouse probes in 19 species. Using a molecular-based reference tree, the phylogenetic distribution of clusters within the genus was reconstructed, and the temporal association between rDNA clusters, breakpoints and centromeres was tested by maximum likelihood analyses. Our results highlighted the following features of rDNA cluster dynamics in the genus Mus: i rDNA clusters showed extensive diversity in number between species and an almost exclusive pericentromeric location, ii a strong association between rDNA sites and centromeres was retrieved which may be related to their shared constraint of concerted evolution, iii 24% of the observed breakpoints mapped near an rDNA cluster, and iv a substantial rate of rDNA cluster change (insertion, deletion also occurred in the absence of chromosomal rearrangements. Conclusions This study on the dynamics of rDNA clusters within the genus Mus has revealed a strong evolutionary relationship between rDNA clusters and centromeres. Both of these genomic structures coincide with breakpoints in the genus Mus, suggesting that the accumulation of a large number of repeats in the centromeric region may contribute to the high level of chromosome

  20. Human Contamination in Public Genome Assemblies.

    Science.gov (United States)

    Kryukov, Kirill; Imanishi, Tadashi

    2016-01-01

    Contamination in genome assembly can lead to wrong or confusing results when using such genome as reference in sequence comparison. Although bacterial contamination is well known, the problem of human-originated contamination received little attention. In this study we surveyed 45,735 available genome assemblies for evidence of human contamination. We used lineage specificity to distinguish between contamination and conservation. We found that 154 genome assemblies contain fragments that with high confidence originate as contamination from human DNA. Majority of contaminating human sequences were present in the reference human genome assembly for over a decade. We recommend that existing contaminated genomes should be revised to remove contaminated sequence, and that new assemblies should be thoroughly checked for presence of human DNA before submitting them to public databases.

  1. Utilization of complete chloroplast genomes for phylogenetic studies

    NARCIS (Netherlands)

    Ramlee, Shairul Izan Binti

    2016-01-01

    Chloroplast DNA sequence polymorphisms are a primary source of data in many plant phylogenetic studies. The chloroplast genome is relatively conserved in its evolution making it an ideal molecule to retain phylogenetic signals. The chloroplast genome is also largely, but not completely, free from

  2. Novel ZBTB24 Mutation Associated with Immunodeficiency, Centromere Instability, and Facial Anomalies Type-2 Syndrome Identified in a Patient with Very Early Onset Inflammatory Bowel Disease.

    Science.gov (United States)

    Conrad, Máire A; Dawany, Noor; Sullivan, Kathleen E; Devoto, Marcella; Kelsen, Judith R

    2017-12-01

    Very early onset inflammatory bowel disease, diagnosed in children ≤5 years old, can be the initial presentation of some primary immunodeficiencies. In this study, we describe a 17-month-old boy with recurrent infections, growth failure, facial anomalies, and inflammatory bowel disease. Immune evaluation, whole-exome sequencing, karyotyping, and methylation array were performed to evaluate the child's constellation of symptoms and examination findings. Whole-exome sequencing revealed that the child was homozygous for a novel variant in ZBTB24, the gene associated with immunodeficiency, centromere instability, and facial anomalies type-2 syndrome. This describes the first case of inflammatory bowel disease associated with immunodeficiency, centromere instability, and facial anomalies type-2 syndrome in a child with a novel disease-causing mutation in ZBTB24 found on whole-exome sequencing.

  3. PGSB/MIPS Plant Genome Information Resources and Concepts for the Analysis of Complex Grass Genomes.

    Science.gov (United States)

    Spannagl, Manuel; Bader, Kai; Pfeifer, Matthias; Nussbaumer, Thomas; Mayer, Klaus F X

    2016-01-01

    PGSB (Plant Genome and Systems Biology; formerly MIPS-Munich Institute for Protein Sequences) has been involved in developing, implementing and maintaining plant genome databases for more than a decade. Genome databases and analysis resources have focused on individual genomes and aim to provide flexible and maintainable datasets for model plant genomes as a backbone against which experimental data, e.g., from high-throughput functional genomics, can be organized and analyzed. In addition, genomes from both model and crop plants form a scaffold for comparative genomics, assisted by specialized tools such as the CrowsNest viewer to explore conserved gene order (synteny) between related species on macro- and micro-levels.The genomes of many economically important Triticeae plants such as wheat, barley, and rye present a great challenge for sequence assembly and bioinformatic analysis due to their enormous complexity and large genome size. Novel concepts and strategies have been developed to deal with these difficulties and have been applied to the genomes of wheat, barley, rye, and other cereals. This includes the GenomeZipper concept, reference-guided exome assembly, and "chromosome genomics" based on flow cytometry sorted chromosomes.

  4. Tetrahedral gray code for visualization of genome information.

    Directory of Open Access Journals (Sweden)

    Natsuhiro Ichinose

    Full Text Available We propose a tetrahedral Gray code that facilitates visualization of genome information on the surfaces of a tetrahedron, where the relative abundance of each [Formula: see text]-mer in the genomic sequence is represented by a color of the corresponding cell of a triangular lattice. For biological significance, the code is designed such that the [Formula: see text]-mers corresponding to any adjacent pair of cells differ from each other by only one nucleotide. We present a simple procedure to draw such a pattern on the development surfaces of a tetrahedron. The thus constructed tetrahedral Gray code can demonstrate evolutionary conservation and variation of the genome information of many organisms at a glance. We also apply the tetrahedral Gray code to the honey bee (Apis mellifera genome to analyze its methylation structure. The results indicate that the honey bee genome exhibits CpG overrepresentation in spite of its methylation ability and that two conserved motifs, CTCGAG and CGCGCG, in the unmethylated regions are responsible for the overrepresentation of CpG.

  5. Sequencing and comparative genome analysis of two pathogenic Streptococcus gallolyticus subspecies: genome plasticity, adaptation and virulence.

    Directory of Open Access Journals (Sweden)

    I-Hsuan Lin

    Full Text Available Streptococcus gallolyticus infections in humans are often associated with bacteremia, infective endocarditis and colon cancers. The disease manifestations are different depending on the subspecies of S. gallolyticus causing the infection. Here, we present the complete genomes of S. gallolyticus ATCC 43143 (biotype I and S. pasteurianus ATCC 43144 (biotype II.2. The genomic differences between the two biotypes were characterized with comparative genomic analyses. The chromosome of ATCC 43143 and ATCC 43144 are 2,36 and 2,10 Mb in length and encode 2246 and 1869 CDS respectively. The organization and genomic contents of both genomes were most similar to the recently published S. gallolyticus UCN34, where 2073 (92% and 1607 (86% of the ATCC 43143 and ATCC 43144 CDS were conserved in UCN34 respectively. There are around 600 CDS conserved in all Streptococcus genomes, indicating the Streptococcus genus has a small core-genome (constitute around 30% of total CDS and substantial evolutionary plasticity. We identified eight and five regions of genome plasticity in ATCC 43143 and ATCC 43144 respectively. Within these regions, several proteins were recognized to contribute to the fitness and virulence of each of the two subspecies. We have also predicted putative cell-surface associated proteins that could play a role in adherence to host tissues, leading to persistent infections causing sub-acute and chronic diseases in humans. This study showed evidence that the S. gallolyticus still possesses genes making it suitable in a rumen environment, whereas the ability for S. pasteurianus to live in rumen is reduced. The genome heterogeneity and genetic diversity among the two biotypes, especially membrane and lipoproteins, most likely contribute to the differences in the pathogenesis of the two S. gallolyticus biotypes and the type of disease an infected patient eventually develops.

  6. Analysis of the grape MYB R2R3 subfamily reveals expanded wine quality-related clades and conserved gene structure organization across Vitis and Arabidopsis genomes

    Directory of Open Access Journals (Sweden)

    Arce-Johnson Patricio

    2008-07-01

    Full Text Available Abstract Background The MYB superfamily constitutes the most abundant group of transcription factors described in plants. Members control processes such as epidermal cell differentiation, stomatal aperture, flavonoid synthesis, cold and drought tolerance and pathogen resistance. No genome-wide characterization of this family has been conducted in a woody species such as grapevine. In addition, previous analysis of the recently released grape genome sequence suggested expansion events of several gene families involved in wine quality. Results We describe and classify 108 members of the grape R2R3 MYB gene subfamily in terms of their genomic gene structures and similarity to their putative Arabidopsis thaliana orthologues. Seven gene models were derived and analyzed in terms of gene expression and their DNA binding domain structures. Despite low overall sequence homology in the C-terminus of all proteins, even in those with similar functions across Arabidopsis and Vitis, highly conserved motif sequences and exon lengths were found. The grape epidermal cell fate clade is expanded when compared with the Arabidopsis and rice MYB subfamilies. Two anthocyanin MYBA related clusters were identified in chromosomes 2 and 14, one of which includes the previously described grape colour locus. Tannin related loci were also detected with eight candidate homologues in chromosomes 4, 9 and 11. Conclusion This genome wide transcription factor analysis in Vitis suggests that clade-specific grape R2R3 MYB genes are expanded while other MYB genes could be well conserved compared to Arabidopsis. MYB gene abundance, homology and orientation within particular loci also suggests that expanded MYB clades conferring quality attributes of grapes and wines, such as colour and astringency, could possess redundant, overlapping and cooperative functions.

  7. Functional genome analysis of Bifidobacterium breve UCC2003 reveals type IVb tight adherence (Tad) pili as an essential and conserved host-colonization factor

    Science.gov (United States)

    O'Connell Motherway, Mary; Zomer, Aldert; Leahy, Sinead C.; Reunanen, Justus; Bottacini, Francesca; Claesson, Marcus J.; O'Brien, Frances; Flynn, Kiera; Casey, Patrick G.; Moreno Munoz, Jose Antonio; Kearney, Breda; Houston, Aileen M.; O'Mahony, Caitlin; Higgins, Des G.; Shanahan, Fergus; Palva, Airi; de Vos, Willem M.; Fitzgerald, Gerald F.; Ventura, Marco; O'Toole, Paul W.; van Sinderen, Douwe

    2011-01-01

    Development of the human gut microbiota commences at birth, with bifidobacteria being among the first colonizers of the sterile newborn gastrointestinal tract. To date, the genetic basis of Bifidobacterium colonization and persistence remains poorly understood. Transcriptome analysis of the Bifidobacterium breve UCC2003 2.42-Mb genome in a murine colonization model revealed differential expression of a type IVb tight adherence (Tad) pilus-encoding gene cluster designated “tad2003.” Mutational analysis demonstrated that the tad2003 gene cluster is essential for efficient in vivo murine gut colonization, and immunogold transmission electron microscopy confirmed the presence of Tad pili at the poles of B. breve UCC2003 cells. Conservation of the Tad pilus-encoding locus among other B. breve strains and among sequenced Bifidobacterium genomes supports the notion of a ubiquitous pili-mediated host colonization and persistence mechanism for bifidobacteria. PMID:21690406

  8. Functional genome analysis of Bifidobacterium breve UCC2003 reveals type IVb tight adherence (Tad) pili as an essential and conserved host-colonization factor.

    Science.gov (United States)

    O'Connell Motherway, Mary; Zomer, Aldert; Leahy, Sinead C; Reunanen, Justus; Bottacini, Francesca; Claesson, Marcus J; O'Brien, Frances; Flynn, Kiera; Casey, Patrick G; Munoz, Jose Antonio Moreno; Kearney, Breda; Houston, Aileen M; O'Mahony, Caitlin; Higgins, Des G; Shanahan, Fergus; Palva, Airi; de Vos, Willem M; Fitzgerald, Gerald F; Ventura, Marco; O'Toole, Paul W; van Sinderen, Douwe

    2011-07-05

    Development of the human gut microbiota commences at birth, with bifidobacteria being among the first colonizers of the sterile newborn gastrointestinal tract. To date, the genetic basis of Bifidobacterium colonization and persistence remains poorly understood. Transcriptome analysis of the Bifidobacterium breve UCC2003 2.42-Mb genome in a murine colonization model revealed differential expression of a type IVb tight adherence (Tad) pilus-encoding gene cluster designated "tad(2003)." Mutational analysis demonstrated that the tad(2003) gene cluster is essential for efficient in vivo murine gut colonization, and immunogold transmission electron microscopy confirmed the presence of Tad pili at the poles of B. breve UCC2003 cells. Conservation of the Tad pilus-encoding locus among other B. breve strains and among sequenced Bifidobacterium genomes supports the notion of a ubiquitous pili-mediated host colonization and persistence mechanism for bifidobacteria.

  9. The tiger genome and comparative analysis with lion and snow leopard genomes.

    Science.gov (United States)

    Cho, Yun Sung; Hu, Li; Hou, Haolong; Lee, Hang; Xu, Jiaohui; Kwon, Soowhan; Oh, Sukhun; Kim, Hak-Min; Jho, Sungwoong; Kim, Sangsoo; Shin, Young-Ah; Kim, Byung Chul; Kim, Hyunmin; Kim, Chang-Uk; Luo, Shu-Jin; Johnson, Warren E; Koepfli, Klaus-Peter; Schmidt-Küntzel, Anne; Turner, Jason A; Marker, Laurie; Harper, Cindy; Miller, Susan M; Jacobs, Wilhelm; Bertola, Laura D; Kim, Tae Hyung; Lee, Sunghoon; Zhou, Qian; Jung, Hyun-Ju; Xu, Xiao; Gadhvi, Priyvrat; Xu, Pengwei; Xiong, Yingqi; Luo, Yadan; Pan, Shengkai; Gou, Caiyun; Chu, Xiuhui; Zhang, Jilin; Liu, Sanyang; He, Jing; Chen, Ying; Yang, Linfeng; Yang, Yulan; He, Jiaju; Liu, Sha; Wang, Junyi; Kim, Chul Hong; Kwak, Hwanjong; Kim, Jong-Soo; Hwang, Seungwoo; Ko, Junsu; Kim, Chang-Bae; Kim, Sangtae; Bayarlkhagva, Damdin; Paek, Woon Kee; Kim, Seong-Jin; O'Brien, Stephen J; Wang, Jun; Bhak, Jong

    2013-01-01

    Tigers and their close relatives (Panthera) are some of the world's most endangered species. Here we report the de novo assembly of an Amur tiger whole-genome sequence as well as the genomic sequences of a white Bengal tiger, African lion, white African lion and snow leopard. Through comparative genetic analyses of these genomes, we find genetic signatures that may reflect molecular adaptations consistent with the big cats' hypercarnivorous diet and muscle strength. We report a snow leopard-specific genetic determinant in EGLN1 (Met39>Lys39), which is likely to be associated with adaptation to high altitude. We also detect a TYR260G>A mutation likely responsible for the white lion coat colour. Tiger and cat genomes show similar repeat composition and an appreciably conserved synteny. Genomic data from the five big cats provide an invaluable resource for resolving easily identifiable phenotypes evident in very close, but distinct, species.

  10. The tiger genome and comparative analysis with lion and snow leopard genomes

    Science.gov (United States)

    Cho, Yun Sung; Hu, Li; Hou, Haolong; Lee, Hang; Xu, Jiaohui; Kwon, Soowhan; Oh, Sukhun; Kim, Hak-Min; Jho, Sungwoong; Kim, Sangsoo; Shin, Young-Ah; Kim, Byung Chul; Kim, Hyunmin; Kim, Chang-uk; Luo, Shu-Jin; Johnson, Warren E.; Koepfli, Klaus-Peter; Schmidt-Küntzel, Anne; Turner, Jason A.; Marker, Laurie; Harper, Cindy; Miller, Susan M.; Jacobs, Wilhelm; Bertola, Laura D.; Kim, Tae Hyung; Lee, Sunghoon; Zhou, Qian; Jung, Hyun-Ju; Xu, Xiao; Gadhvi, Priyvrat; Xu, Pengwei; Xiong, Yingqi; Luo, Yadan; Pan, Shengkai; Gou, Caiyun; Chu, Xiuhui; Zhang, Jilin; Liu, Sanyang; He, Jing; Chen, Ying; Yang, Linfeng; Yang, Yulan; He, Jiaju; Liu, Sha; Wang, Junyi; Kim, Chul Hong; Kwak, Hwanjong; Kim, Jong-Soo; Hwang, Seungwoo; Ko, Junsu; Kim, Chang-Bae; Kim, Sangtae; Bayarlkhagva, Damdin; Paek, Woon Kee; Kim, Seong-Jin; O’Brien, Stephen J.; Wang, Jun; Bhak, Jong

    2013-01-01

    Tigers and their close relatives (Panthera) are some of the world’s most endangered species. Here we report the de novo assembly of an Amur tiger whole-genome sequence as well as the genomic sequences of a white Bengal tiger, African lion, white African lion and snow leopard. Through comparative genetic analyses of these genomes, we find genetic signatures that may reflect molecular adaptations consistent with the big cats’ hypercarnivorous diet and muscle strength. We report a snow leopard-specific genetic determinant in EGLN1 (Met39>Lys39), which is likely to be associated with adaptation to high altitude. We also detect a TYR260G>A mutation likely responsible for the white lion coat colour. Tiger and cat genomes show similar repeat composition and an appreciably conserved synteny. Genomic data from the five big cats provide an invaluable resource for resolving easily identifiable phenotypes evident in very close, but distinct, species. PMID:24045858

  11. Elevated Rate of Genome Rearrangements in Radiation-Resistant Bacteria.

    Science.gov (United States)

    Repar, Jelena; Supek, Fran; Klanjscek, Tin; Warnecke, Tobias; Zahradka, Ksenija; Zahradka, Davor

    2017-04-01

    A number of bacterial, archaeal, and eukaryotic species are known for their resistance to ionizing radiation. One of the challenges these species face is a potent environmental source of DNA double-strand breaks, potential drivers of genome structure evolution. Efficient and accurate DNA double-strand break repair systems have been demonstrated in several unrelated radiation-resistant species and are putative adaptations to the DNA damaging environment. Such adaptations are expected to compensate for the genome-destabilizing effect of environmental DNA damage and may be expected to result in a more conserved gene order in radiation-resistant species. However, here we show that rates of genome rearrangements, measured as loss of gene order conservation with time, are higher in radiation-resistant species in multiple, phylogenetically independent groups of bacteria. Comparison of indicators of selection for genome organization between radiation-resistant and phylogenetically matched, nonresistant species argues against tolerance to disruption of genome structure as a strategy for radiation resistance. Interestingly, an important mechanism affecting genome rearrangements in prokaryotes, the symmetrical inversions around the origin of DNA replication, shapes genome structure of both radiation-resistant and nonresistant species. In conclusion, the opposing effects of environmental DNA damage and DNA repair result in elevated rates of genome rearrangements in radiation-resistant bacteria. Copyright © 2017 Repar et al.

  12. Insights from Human/Mouse genome comparisons

    Energy Technology Data Exchange (ETDEWEB)

    Pennacchio, Len A.

    2003-03-30

    Large-scale public genomic sequencing efforts have provided a wealth of vertebrate sequence data poised to provide insights into mammalian biology. These include deep genomic sequence coverage of human, mouse, rat, zebrafish, and two pufferfish (Fugu rubripes and Tetraodon nigroviridis) (Aparicio et al. 2002; Lander et al. 2001; Venter et al. 2001; Waterston et al. 2002). In addition, a high-priority has been placed on determining the genomic sequence of chimpanzee, dog, cow, frog, and chicken (Boguski 2002). While only recently available, whole genome sequence data have provided the unique opportunity to globally compare complete genome contents. Furthermore, the shared evolutionary ancestry of vertebrate species has allowed the development of comparative genomic approaches to identify ancient conserved sequences with functionality. Accordingly, this review focuses on the initial comparison of available mammalian genomes and describes various insights derived from such analysis.

  13. COMBO-FISH Enables High Precision Localization Microscopy as a Prerequisite for Nanostructure Analysis of Genome Loci

    Directory of Open Access Journals (Sweden)

    Rainer Kaufmann

    2010-10-01

    Full Text Available With the completeness of genome databases, it has become possible to develop a novel FISH (Fluorescence in Situ Hybridization technique called COMBO-FISH (COMBinatorial Oligo FISH. In contrast to other FISH techniques, COMBO-FISH makes use of a bioinformatics approach for probe set design. By means of computer genome database searching, several oligonucleotide stretches of typical lengths of 15–30 nucleotides are selected in such a way that all uniquely colocalize at the given genome target. The probes applied here were Peptide Nucleic Acids (PNAs—synthetic DNA analogues with a neutral backbone—which were synthesized under high purity conditions. For a probe repetitively highlighted in centromere 9, PNAs labeled with different dyes were tested, among which Alexa 488Ò showed reversible photobleaching (blinking between dark and bright state a prerequisite for the application of SPDM (Spectral Precision Distance/Position Determination Microscopy a novel technique of high resolution fluorescence localization microscopy. Although COMBO-FISH labeled cell nuclei under SPDM conditions sometimes revealed fluorescent background, the specific locus was clearly discriminated by the signal intensity and the resulting localization accuracy in the range of 10–20 nm for a detected oligonucleotide stretch. The results indicate that COMBO-FISH probes with blinking dyes are well suited for SPDM, which will open new perspectives on molecular nanostructural analysis of the genome.

  14. Discovery of cis-elements between sorghum and rice using co-expression and evolutionary conservation

    Directory of Open Access Journals (Sweden)

    Haberer Georg

    2009-06-01

    Full Text Available Abstract Background The spatiotemporal regulation of gene expression largely depends on the presence and absence of cis-regulatory sites in the promoter. In the economically highly important grass family, our knowledge of transcription factor binding sites and transcriptional networks is still very limited. With the completion of the sorghum genome and the available rice genome sequence, comparative promoter analyses now allow genome-scale detection of conserved cis-elements. Results In this study, we identified thousands of phylogenetic footprints conserved between orthologous rice and sorghum upstream regions that are supported by co-expression information derived from three different rice expression data sets. In a complementary approach, cis-motifs were discovered by their highly conserved co-occurrence in syntenic promoter pairs. Sequence conservation and matches to known plant motifs support our findings. Expression similarities of gene pairs positively correlate with the number of motifs that are shared by gene pairs and corroborate the importance of similar promoter architectures for concerted regulation. This strongly suggests that these motifs function in the regulation of transcript levels in rice and, presumably also in sorghum. Conclusion Our work provides the first large-scale collection of cis-elements for rice and sorghum and can serve as a paradigm for cis-element analysis through comparative genomics in grasses in general.

  15. The reduced genomes of Parcubacteria (OD1) contain signatures of a symbiotic lifestyle

    Energy Technology Data Exchange (ETDEWEB)

    Nelson, William C.; Stegen, James C.

    2015-07-21

    Candidate phylum OD1 bacteria (also referred to as Parcubacteria) have been identified in broad range of anoxic environments through community survey analysis. Although none of these species have been isolated in the laboratory, several genome sequences have been reconstructed from metagenomic sequence data and single-cell sequencing. The organisms have small (generally <1 Mb) genomes with severely reduced metabolic capabilities. We have reconstructed 8 partial to near-complete OD1 genomes from oxic groundwater samples, and compared them against existing genomic data. The conserved core gene set comprises 202 genes, or ~28% of the genomic complement. ‘Housekeeping’ genes and genes for biosynthesis of peptidoglycan and Type IV pilus production are conserved. Gene sets for biosynthesis of cofactors, amino acids, nucleotides and fatty acids are absent entirely or greatly reduced. The only aspects of energy metabolism conserved are the non-oxidative branch of the pentose-phosphate shunt and central glycolysis. These organisms also lack some activities conserved in almost all other known bacterial genomes, including signal recognition particle, pseudouridine synthase A, and FAD synthase. Pan-genome analysis indicates a broad genotypic diversity and perhaps a highly fluid gene complement, indicating historical adaptation to a wide range of growth environments and a high degree of specialization. The genomes were examined for signatures suggesting either a free-living, streamlined lifestyle or a symbiotic lifestyle. The lack of biosynthetic capabilities and DNA repair, along with the presence of potential attachment and adhesion proteins suggest the Parcubacteria are ectosymbionts or parasites of other organisms. The wide diversity of genes that potentially mediate cell-cell contact suggests a broad range of partner/prey organisms across the phylum.

  16. The complete mitochondrial genomes of two rice planthoppers, Nilaparvata lugens and Laodelphax striatellus: conserved genome rearrangement in Delphacidae and discovery of new characteristics of atp8 and tRNA genes.

    Science.gov (United States)

    Zhang, Kai-Jun; Zhu, Wen-Chao; Rong, Xia; Zhang, Yan-Kai; Ding, Xiu-Lei; Liu, Jing; Chen, Da-Song; Du, Yu; Hong, Xiao-Yue

    2013-06-22

    Nilaparvata lugens (the brown planthopper, BPH) and Laodelphax striatellus (the small brown planthopper, SBPH) are two of the most important pests of rice. Up to now, there was only one mitochondrial genome of rice planthopper has been sequenced and very few dependable information of mitochondria could be used for research on population genetics, phylogeographics and phylogenetic evolution of these pests. To get more valuable information from the mitochondria, we sequenced the complete mitochondrial genomes of BPH and SBPH. These two planthoppers were infected with two different functional Wolbachia (intracellular endosymbiont) strains (wLug and wStri). Since both mitochondria and Wolbachia are transmitted by cytoplasmic inheritance and it was difficult to separate them when purified the Wolbachia particles, concomitantly sequencing the genome of Wolbachia using next generation sequencing method, we also got nearly complete mitochondrial genome sequences of these two rice planthoppers. After gap closing, we present high quality and reliable complete mitochondrial genomes of these two planthoppers. The mitogenomes of N. lugens (BPH) and L. striatellus (SBPH) are 17, 619 bp and 16, 431 bp long with A + T contents of 76.95% and 77.17%, respectively. Both species have typical circular mitochondrial genomes that encode the complete set of 37 genes which are usually found in metazoans. However, the BPH mitogenome also possesses two additional copies of the trnC gene. In both mitochondrial genomes, the lengths of the atp8 gene were conspicuously shorter than that of all other known insect mitochondrial genomes (99 bp for BPH, 102 bp for SBPH). That two rearrangement regions (trnC-trnW and nad6-trnP-trnT) of mitochondrial genomes differing from other known insect were found in these two distantly related planthoppers revealed that the gene order of mitochondria might be conservative in Delphacidae. The large non-coding fragment (the A+T-rich region) putatively

  17. The Schizosaccharomyces pombe JmjC-protein, Msc1, prevents H2A.Z localization in centromeric and subtelomeric chromatin domains.

    Directory of Open Access Journals (Sweden)

    Luke Buchanan

    2009-11-01

    Full Text Available Eukaryotic genomes are repetitively packaged into chromatin by nucleosomes, however they are regulated by the differences between nucleosomes, which establish various chromatin states. Local chromatin cues direct the inheritance and propagation of chromatin status via self-reinforcing epigenetic mechanisms. Replication-independent histone exchange could potentially perturb chromatin status if histone exchange chaperones, such as Swr1C, loaded histone variants into wrong sites. Here we show that in Schizosaccharomyces pombe, like Saccharomyces cerevisiae, Swr1C is required for loading H2A.Z into specific sites, including the promoters of lowly expressed genes. However S. pombe Swr1C has an extra subunit, Msc1, which is a JumonjiC-domain protein of the Lid/Jarid1 family. Deletion of Msc1 did not disrupt the S. pombe Swr1C or its ability to bind and load H2A.Z into euchromatin, however H2A.Z was ectopically found in the inner centromere and in subtelomeric chromatin. Normally this subtelomeric region not only lacks H2A.Z but also shows uniformly lower levels of H3K4me2, H4K5, and K12 acetylation than euchromatin and disproportionately contains the most lowly expressed genes during vegetative growth, including many meiotic-specific genes. Genes within and adjacent to subtelomeric chromatin become overexpressed in the absence of either Msc1, Swr1, or paradoxically H2A.Z itself. We also show that H2A.Z is N-terminally acetylated before, and lysine acetylated after, loading into chromatin and that it physically associates with the Nap1 histone chaperone. However, we find a negative correlation between the genomic distributions of H2A.Z and Nap1/Hrp1/Hrp3, suggesting that the Nap1 chaperones remove H2A.Z from chromatin. These data describe H2A.Z action in S. pombe and identify a new mode of chromatin surveillance and maintenance based on negative regulation of histone variant misincorporation.

  18. The genome sequence of the commercially cultivated mushroom Agrocybe aegerita reveals a conserved repertoire of fruiting-related genes and a versatile suite of biopolymer-degrading enzymes.

    Science.gov (United States)

    Gupta, Deepak K; Rühl, Martin; Mishra, Bagdevi; Kleofas, Vanessa; Hofrichter, Martin; Herzog, Robert; Pecyna, Marek J; Sharma, Rahul; Kellner, Harald; Hennicke, Florian; Thines, Marco

    2018-01-15

    Agrocybe aegerita is an agaricomycete fungus with typical mushroom features, which is commercially cultivated for its culinary use. In nature, it is a saprotrophic or facultative pathogenic fungus causing a white-rot of hardwood in forests of warm and mild climate. The ease of cultivation and fructification on solidified media as well as its archetypal mushroom fruit body morphology render A. aegerita a well-suited model for investigating mushroom developmental biology. Here, the genome of the species is reported and analysed with respect to carbohydrate active genes and genes known to play a role during fruit body formation. In terms of fruit body development, our analyses revealed a conserved repertoire of fruiting-related genes, which corresponds well to the archetypal fruit body morphology of this mushroom. For some genes involved in fruit body formation, paralogisation was observed, but not all fruit body maturation-associated genes known from other agaricomycetes seem to be conserved in the genome sequence of A. aegerita. In terms of lytic enzymes, our analyses suggest a versatile arsenal of biopolymer-degrading enzymes that likely account for the flexible life style of this species. Regarding the amount of genes encoding CAZymes relevant for lignin degradation, A. aegerita shows more similarity to white-rot fungi than to litter decomposers, including 18 genes coding for unspecific peroxygenases and three dye-decolourising peroxidase genes expanding its lignocellulolytic machinery. The genome resource will be useful for developing strategies towards genetic manipulation of A. aegerita, which will subsequently allow functional genetics approaches to elucidate fundamentals of fruiting and vegetative growth including lignocellulolysis.

  19. The Arab genome: Health and wealth.

    Science.gov (United States)

    Zayed, Hatem

    2016-11-05

    The 22 Arab nations have a unique genetic structure, which reflects both conserved and diverse gene pools due to the prevalent endogamous and consanguineous marriage culture and the long history of admixture among different ethnic subcultures descended from the Asian, European, and African continents. Human genome sequencing has enabled large-scale genomic studies of different populations and has become a powerful tool for studying disease predictions and diagnosis. Despite the importance of the Arab genome for better understanding the dynamics of the human genome, discovering rare genetic variations, and studying early human migration out of Africa, it is poorly represented in human genome databases, such as HapMap and the 1000 Genomes Project. In this review, I demonstrate the significance of sequencing the Arab genome and setting an Arab genome reference(s) for better understanding the molecular pathogenesis of genetic diseases, discovering novel/rare variants, and identifying a meaningful genotype-phenotype correlation for complex diseases. Copyright © 2016. Published by Elsevier B.V.

  20. Prokaryotic Phylogenies Inferred from Whole-Genome Sequence and Annotation Data

    Directory of Open Access Journals (Sweden)

    Wei Du

    2013-01-01

    Full Text Available Phylogenetic trees are used to represent the evolutionary relationship among various groups of species. In this paper, a novel method for inferring prokaryotic phylogenies using multiple genomic information is proposed. The method is called CGCPhy and based on the distance matrix of orthologous gene clusters between whole-genome pairs. CGCPhy comprises four main steps. First, orthologous genes are determined by sequence similarity, genomic function, and genomic structure information. Second, genes involving potential HGT events are eliminated, since such genes are considered to be the highly conserved genes across different species and the genes located on fragments with abnormal genome barcode. Third, we calculate the distance of the orthologous gene clusters between each genome pair in terms of the number of orthologous genes in conserved clusters. Finally, the neighbor-joining method is employed to construct phylogenetic trees across different species. CGCPhy has been examined on different datasets from 617 complete single-chromosome prokaryotic genomes and achieved applicative accuracies on different species sets in agreement with Bergey's taxonomy in quartet topologies. Simulation results show that CGCPhy achieves high average accuracy and has a low standard deviation on different datasets, so it has an applicative potential for phylogenetic analysis.

  1. Telomere disruption results in non-random formation of de novo dicentric chromosomes involving acrocentric human chromosomes.

    Directory of Open Access Journals (Sweden)

    Kaitlin M Stimpson

    2010-08-01

    Full Text Available Genome rearrangement often produces chromosomes with two centromeres (dicentrics that are inherently unstable because of bridge formation and breakage during cell division. However, mammalian dicentrics, and particularly those in humans, can be quite stable, usually because one centromere is functionally silenced. Molecular mechanisms of centromere inactivation are poorly understood since there are few systems to experimentally create dicentric human chromosomes. Here, we describe a human cell culture model that enriches for de novo dicentrics. We demonstrate that transient disruption of human telomere structure non-randomly produces dicentric fusions involving acrocentric chromosomes. The induced dicentrics vary in structure near fusion breakpoints and like naturally-occurring dicentrics, exhibit various inter-centromeric distances. Many functional dicentrics persist for months after formation. Even those with distantly spaced centromeres remain functionally dicentric for 20 cell generations. Other dicentrics within the population reflect centromere inactivation. In some cases, centromere inactivation occurs by an apparently epigenetic mechanism. In other dicentrics, the size of the alpha-satellite DNA array associated with CENP-A is reduced compared to the same array before dicentric formation. Extra-chromosomal fragments that contained CENP-A often appear in the same cells as dicentrics. Some of these fragments are derived from the same alpha-satellite DNA array as inactivated centromeres. Our results indicate that dicentric human chromosomes undergo alternative fates after formation. Many retain two active centromeres and are stable through multiple cell divisions. Others undergo centromere inactivation. This event occurs within a broad temporal window and can involve deletion of chromatin that marks the locus as a site for CENP-A maintenance/replenishment.

  2. Genome-Scale Co-Expression Network Comparison across Escherichia coli and Salmonella enterica Serovar Typhimurium Reveals Significant Conservation at the Regulon Level of Local Regulators Despite Their Dissimilar Lifestyles

    Science.gov (United States)

    Zarrineh, Peyman; Sánchez-Rodríguez, Aminael; Hosseinkhan, Nazanin; Narimani, Zahra; Marchal, Kathleen; Masoudi-Nejad, Ali

    2014-01-01

    Availability of genome-wide gene expression datasets provides the opportunity to study gene expression across different organisms under a plethora of experimental conditions. In our previous work, we developed an algorithm called COMODO (COnserved MODules across Organisms) that identifies conserved expression modules between two species. In the present study, we expanded COMODO to detect the co-expression conservation across three organisms by adapting the statistics behind it. We applied COMODO to study expression conservation/divergence between Escherichia coli, Salmonella enterica, and Bacillus subtilis. We observed that some parts of the regulatory interaction networks were conserved between E. coli and S. enterica especially in the regulon of local regulators. However, such conservation was not observed between the regulatory interaction networks of B. subtilis and the two other species. We found co-expression conservation on a number of genes involved in quorum sensing, but almost no conservation for genes involved in pathogenicity across E. coli and S. enterica which could partially explain their different lifestyles. We concluded that despite their different lifestyles, no significant rewiring have occurred at the level of local regulons involved for instance, and notable conservation can be detected in signaling pathways and stress sensing in the phylogenetically close species S. enterica and E. coli. Moreover, conservation of local regulons seems to depend on the evolutionary time of divergence across species disappearing at larger distances as shown by the comparison with B. subtilis. Global regulons follow a different trend and show major rewiring even at the limited evolutionary distance that separates E. coli and S. enterica. PMID:25101984

  3. Systematic discovery of regulatory motifs in Fusarium graminearum by comparing four Fusarium genomes

    Directory of Open Access Journals (Sweden)

    Kistler Corby

    2010-03-01

    Full Text Available Abstract Background Fusarium graminearum (Fg, a major fungal pathogen of cultivated cereals, is responsible for billions of dollars in agriculture losses. There is a growing interest in understanding the transcriptional regulation of this organism, especially the regulation of genes underlying its pathogenicity. The generation of whole genome sequence assemblies for Fg and three closely related Fusarium species provides a unique opportunity for such a study. Results Applying comparative genomics approaches, we developed a computational pipeline to systematically discover evolutionarily conserved regulatory motifs in the promoter, downstream and the intronic regions of Fg genes, based on the multiple alignments of sequenced Fusarium genomes. Using this method, we discovered 73 candidate regulatory motifs in the promoter regions. Nearly 30% of these motifs are highly enriched in promoter regions of Fg genes that are associated with a specific functional category. Through comparison to Saccharomyces cerevisiae (Sc and Schizosaccharomyces pombe (Sp, we observed conservation of transcription factors (TFs, their binding sites and the target genes regulated by these TFs related to pathways known to respond to stress conditions or phosphate metabolism. In addition, this study revealed 69 and 39 conserved motifs in the downstream regions and the intronic regions, respectively, of Fg genes. The top intronic motif is the splice donor site. For the downstream regions, we noticed an intriguing absence of the mammalian and Sc poly-adenylation signals among the list of conserved motifs. Conclusion This study provides the first comprehensive list of candidate regulatory motifs in Fg, and underscores the power of comparative genomics in revealing functional elements among related genomes. The conservation of regulatory pathways among the Fusarium genomes and the two yeast species reveals their functional significance, and provides new insights in their

  4. Human social genomics.

    Directory of Open Access Journals (Sweden)

    Steven W Cole

    2014-08-01

    Full Text Available A growing literature in human social genomics has begun to analyze how everyday life circumstances influence human gene expression. Social-environmental conditions such as urbanity, low socioeconomic status, social isolation, social threat, and low or unstable social status have been found to associate with differential expression of hundreds of gene transcripts in leukocytes and diseased tissues such as metastatic cancers. In leukocytes, diverse types of social adversity evoke a common conserved transcriptional response to adversity (CTRA characterized by increased expression of proinflammatory genes and decreased expression of genes involved in innate antiviral responses and antibody synthesis. Mechanistic analyses have mapped the neural "social signal transduction" pathways that stimulate CTRA gene expression in response to social threat and may contribute to social gradients in health. Research has also begun to analyze the functional genomics of optimal health and thriving. Two emerging opportunities now stand to revolutionize our understanding of the everyday life of the human genome: network genomics analyses examining how systems-level capabilities emerge from groups of individual socially sensitive genomes and near-real-time transcriptional biofeedback to empirically optimize individual well-being in the context of the unique genetic, geographic, historical, developmental, and social contexts that jointly shape the transcriptional realization of our innate human genomic potential for thriving.

  5. Complete Genome Sequences of Mycobacteriophages Clautastrophe, Kingsolomon, Krypton555, and Nicholas

    OpenAIRE

    Chung, Hui-Min; D’Elia, Tom; Ross, Joseph F.; Alvarado, Samuel M.; Brantley, Molly-Catherine; Bricker, Lydia P.; Butler, Courtney R.; Crist, Carson; Dane, Julia M.; Farran, Brett W.; Hobbs, Sierra; Lapak, Michelle; Lovell, Conner; Ludergnani, Nicholas; McMullen, Allison

    2017-01-01

    ABSTRACT We report here the complete genome sequences of four subcluster L3 mycobacteriophages newly isolated from soil samples, using Mycobacterium smegmatis mc2155 as the host. Comparative genomic analyses with four previously described subcluster L3 phages reveal strong nucleotide similarity and gene conservation, with several large insertions/deletions near their right genome ends.

  6. Short Communication: An apospory-specific genomic region is conserved between Buffelgrass (Cenchrus ciliaris L.) and Pennisetum squamulatum Fresen.

    Science.gov (United States)

    Roche; Cong; Chen; Hanna; Gustine; Sherwood; Ozias-Akins

    1999-07-01

    Twelve molecular markers linked to pseudogamous apospory, a form of gametophytic apomixis, were previously isolated from Pennisetum squamulatum Fresen. No recombination between these markers was found in a segregating population of 397 individuals (Ozias-Akins et al. 1998, Proc. Natl Acad. Sci. USA, 95, 5127-5132). The objective of the present study was to test if these markers were also linked to the aposporous mode of reproduction in two small segregating populations of Cenchrus ciliaris (= Pennisetum ciliare (L.)Link), another apomictic grass species. Among 12 markers (sequence characterized amplified regions, SCARs), six were scored as dominant markers between aposporous and sexual C. ciliaris genotypes (presence/absence, respectively). Five were always linked to apospory and one showed a low level of recombination in 84 progenies. Restriction fragment length polymorphisms (RFLPs) were observed between sexual and apomictic phenotypes for three of the six remaining SCARs from P. squamulatum when used as probes. No recombination was observed in the F1 progenies. Preliminary data from megabase DNA analysis and sequencing in both species indicate that an apospory-specific genomic region (ASGR) is highly conserved between the two species. Although C. ciliaris has a smaller genome size to P. squamulatum, a higher copy number for markers linked to apospory found in the former may impair the progress of positional cloning of gene(s) for apomixis in this species.

  7. Arthropod phylogenetics in light of three novel millipede (myriapoda: diplopoda mitochondrial genomes with comments on the appropriateness of mitochondrial genome sequence data for inferring deep level relationships.

    Directory of Open Access Journals (Sweden)

    Michael S Brewer

    Full Text Available BACKGROUND: Arthropods are the most diverse group of eukaryotic organisms, but their phylogenetic relationships are poorly understood. Herein, we describe three mitochondrial genomes representing orders of millipedes for which complete genomes had not been characterized. Newly sequenced genomes are combined with existing data to characterize the protein coding regions of myriapods and to attempt to reconstruct the evolutionary relationships within the Myriapoda and Arthropoda. RESULTS: The newly sequenced genomes are similar to previously characterized millipede sequences in terms of synteny and length. Unique translocations occurred within the newly sequenced taxa, including one half of the Appalachioria falcifera genome, which is inverted with respect to other millipede genomes. Across myriapods, amino acid conservation levels are highly dependent on the gene region. Additionally, individual loci varied in the level of amino acid conservation. Overall, most gene regions showed low levels of conservation at many sites. Attempts to reconstruct the evolutionary relationships suffered from questionable relationships and low support values. Analyses of phylogenetic informativeness show the lack of signal deep in the trees (i.e., genes evolve too quickly. As a result, the myriapod tree resembles previously published results but lacks convincing support, and, within the arthropod tree, well established groups were recovered as polyphyletic. CONCLUSIONS: The novel genome sequences described herein provide useful genomic information concerning millipede groups that had not been investigated. Taken together with existing sequences, the variety of compositions and evolution of myriapod mitochondrial genomes are shown to be more complex than previously thought. Unfortunately, the use of mitochondrial protein-coding regions in deep arthropod phylogenetics appears problematic, a result consistent with previously published studies. Lack of phylogenetic

  8. Arthropod phylogenetics in light of three novel millipede (myriapoda: diplopoda) mitochondrial genomes with comments on the appropriateness of mitochondrial genome sequence data for inferring deep level relationships.

    Science.gov (United States)

    Brewer, Michael S; Swafford, Lynn; Spruill, Chad L; Bond, Jason E

    2013-01-01

    Arthropods are the most diverse group of eukaryotic organisms, but their phylogenetic relationships are poorly understood. Herein, we describe three mitochondrial genomes representing orders of millipedes for which complete genomes had not been characterized. Newly sequenced genomes are combined with existing data to characterize the protein coding regions of myriapods and to attempt to reconstruct the evolutionary relationships within the Myriapoda and Arthropoda. The newly sequenced genomes are similar to previously characterized millipede sequences in terms of synteny and length. Unique translocations occurred within the newly sequenced taxa, including one half of the Appalachioria falcifera genome, which is inverted with respect to other millipede genomes. Across myriapods, amino acid conservation levels are highly dependent on the gene region. Additionally, individual loci varied in the level of amino acid conservation. Overall, most gene regions showed low levels of conservation at many sites. Attempts to reconstruct the evolutionary relationships suffered from questionable relationships and low support values. Analyses of phylogenetic informativeness show the lack of signal deep in the trees (i.e., genes evolve too quickly). As a result, the myriapod tree resembles previously published results but lacks convincing support, and, within the arthropod tree, well established groups were recovered as polyphyletic. The novel genome sequences described herein provide useful genomic information concerning millipede groups that had not been investigated. Taken together with existing sequences, the variety of compositions and evolution of myriapod mitochondrial genomes are shown to be more complex than previously thought. Unfortunately, the use of mitochondrial protein-coding regions in deep arthropod phylogenetics appears problematic, a result consistent with previously published studies. Lack of phylogenetic signal renders the resulting tree topologies as suspect

  9. Genomic sequence around butterfly wing development genes: annotation and comparative analysis.

    Directory of Open Access Journals (Sweden)

    Inês C Conceição

    Full Text Available BACKGROUND: Analysis of genomic sequence allows characterization of genome content and organization, and access beyond gene-coding regions for identification of functional elements. BAC libraries, where relatively large genomic regions are made readily available, are especially useful for species without a fully sequenced genome and can increase genomic coverage of phylogenetic and biological diversity. For example, no butterfly genome is yet available despite the unique genetic and biological properties of this group, such as diversified wing color patterns. The evolution and development of these patterns is being studied in a few target species, including Bicyclus anynana, where a whole-genome BAC library allows targeted access to large genomic regions. METHODOLOGY/PRINCIPAL FINDINGS: We characterize ∼1.3 Mb of genomic sequence around 11 selected genes expressed in B. anynana developing wings. Extensive manual curation of in silico predictions, also making use of a large dataset of expressed genes for this species, identified repetitive elements and protein coding sequence, and highlighted an expansion of Alcohol dehydrogenase genes. Comparative analysis with orthologous regions of the lepidopteran reference genome allowed assessment of conservation of fine-scale synteny (with detection of new inversions and translocations and of DNA sequence (with detection of high levels of conservation of non-coding regions around some, but not all, developmental genes. CONCLUSIONS: The general properties and organization of the available B. anynana genomic sequence are similar to the lepidopteran reference, despite the more than 140 MY divergence. Our results lay the groundwork for further studies of new interesting findings in relation to both coding and non-coding sequence: 1 the Alcohol dehydrogenase expansion with higher similarity between the five tandemly-repeated B. anynana paralogs than with the corresponding B. mori orthologs, and 2 the high

  10. Mountain gorilla genomes reveal the impact of long-term population decline and inbreeding

    DEFF Research Database (Denmark)

    Xue, Yali; Prado-Martinez, Javier; Sudmant, Peter H

    2015-01-01

    Mountain gorillas are an endangered great ape subspecies and a prominent focus for conservation, yet we know little about their genomic diversity and evolutionary past. We sequenced whole genomes from multiple wild individuals and compared the genomes of all four Gorilla subspecies. We found that...

  11. Conservation Genetics of the Cheetah: Lessons Learned and New Opportunities.

    Science.gov (United States)

    O'Brien, Stephen J; Johnson, Warren E; Driscoll, Carlos A; Dobrynin, Pavel; Marker, Laurie

    2017-09-01

    The dwindling wildlife species of our planet have become a cause célèbre for conservation groups, governments, and concerned citizens throughout the world. The application of powerful new genetic technologies to surviving populations of threatened mammals has revolutionized our ability to recognize hidden perils that afflict them. We have learned new lessons of survival, adaptation, and evolution from viewing the natural history of genomes in hundreds of detailed studies. A single case history of one species, the African cheetah, Acinonyx jubatus, is here reviewed to reveal a long-term story of conservation challenges and action informed by genetic discoveries and insights. A synthesis of 3 decades of data, interpretation, and controversy, capped by whole genome sequence analysis of cheetahs, provides a compelling tale of conservation relevance and action to protect this species and other threatened wildlife. © The American Genetic Association 2017. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

  12. Complete Genome Sequences of Mycobacteriophages Clautastrophe, Kingsolomon, Krypton555, and Nicholas

    Science.gov (United States)

    Chung, Hui-Min; D’Elia, Tom; Ross, Joseph F.; Alvarado, Samuel M.; Brantley, Molly-Catherine; Bricker, Lydia P.; Butler, Courtney R.; Crist, Carson; Dane, Julia M.; Farran, Brett W.; Hobbs, Sierra; Lapak, Michelle; Lovell, Conner; McMullen, Allison; Mirza, Sohail A.; Thrift, Noah; Vaughan, Donald P.; Worley, Grace; Ejikemeuwa, Amara; Zaw, May; Albritton, Claude F.; Bertrand, Sarah C.; Chaudhry, Shanzay S.; Cheema, Vzair A.; Do, Camilla; Do, Michael L.; Duong, Huyen M.; El-Desoky, Dalia H.; Green, Kelsey M.; Lee, Rhea N.; Thornton, Lauren A.; Vu, James M.; Zahra, Mah Noor; Stoner, Ty H.; Garlena, Rebecca A.; Jacobs-Sera, Deborah; Russell, Daniel A.

    2017-01-01

    ABSTRACT We report here the complete genome sequences of four subcluster L3 mycobacteriophages newly isolated from soil samples, using Mycobacterium smegmatis mc2155 as the host. Comparative genomic analyses with four previously described subcluster L3 phages reveal strong nucleotide similarity and gene conservation, with several large insertions/deletions near their right genome ends. PMID:29122864

  13. Survey sequencing and comparative analysis of the elephant shark (Callorhinchus milii genome.

    Directory of Open Access Journals (Sweden)

    Byrappa Venkatesh

    2007-04-01

    Full Text Available Owing to their phylogenetic position, cartilaginous fishes (sharks, rays, skates, and chimaeras provide a critical reference for our understanding of vertebrate genome evolution. The relatively small genome of the elephant shark, Callorhinchus milii, a chimaera, makes it an attractive model cartilaginous fish genome for whole-genome sequencing and comparative analysis. Here, the authors describe survey sequencing (1.4x coverage and comparative analysis of the elephant shark genome, one of the first cartilaginous fish genomes to be sequenced to this depth. Repetitive sequences, represented mainly by a novel family of short interspersed element-like and long interspersed element-like sequences, account for about 28% of the elephant shark genome. Fragments of approximately 15,000 elephant shark genes reveal specific examples of genes that have been lost differentially during the evolution of tetrapod and teleost fish lineages. Interestingly, the degree of conserved synteny and conserved sequences between the human and elephant shark genomes are higher than that between human and teleost fish genomes. Elephant shark contains putative four Hox clusters indicating that, unlike teleost fish genomes, the elephant shark genome has not experienced an additional whole-genome duplication. These findings underscore the importance of the elephant shark as a critical reference vertebrate genome for comparative analysis of the human and other vertebrate genomes. This study also demonstrates that a survey-sequencing approach can be applied productively for comparative analysis of distantly related vertebrate genomes.

  14. A universal genomic coordinate translator for comparative genomics.

    Science.gov (United States)

    Zamani, Neda; Sundström, Görel; Meadows, Jennifer R S; Höppner, Marc P; Dainat, Jacques; Lantz, Henrik; Haas, Brian J; Grabherr, Manfred G

    2014-06-30

    Genomic duplications constitute major events in the evolution of species, allowing paralogous copies of genes to take on fine-tuned biological roles. Unambiguously identifying the orthology relationship between copies across multiple genomes can be resolved by synteny, i.e. the conserved order of genomic sequences. However, a comprehensive analysis of duplication events and their contributions to evolution would require all-to-all genome alignments, which increases at N2 with the number of available genomes, N. Here, we introduce Kraken, software that omits the all-to-all requirement by recursively traversing a graph of pairwise alignments and dynamically re-computing orthology. Kraken scales linearly with the number of targeted genomes, N, which allows for including large numbers of genomes in analyses. We first evaluated the method on the set of 12 Drosophila genomes, finding that orthologous correspondence computed indirectly through a graph of multiple synteny maps comes at minimal cost in terms of sensitivity, but reduces overall computational runtime by an order of magnitude. We then used the method on three well-annotated mammalian genomes, human, mouse, and rat, and show that up to 93% of protein coding transcripts have unambiguous pairwise orthologous relationships across the genomes. On a nucleotide level, 70 to 83% of exons match exactly at both splice junctions, and up to 97% on at least one junction. We last applied Kraken to an RNA-sequencing dataset from multiple vertebrates and diverse tissues, where we confirmed that brain-specific gene family members, i.e. one-to-many or many-to-many homologs, are more highly correlated across species than single-copy (i.e. one-to-one homologous) genes. Not limited to protein coding genes, Kraken also identifies thousands of newly identified transcribed loci, likely non-coding RNAs that are consistently transcribed in human, chimpanzee and gorilla, and maintain significant correlation of expression levels across

  15. Genome sequencing of chimpanzee malaria parasites reveals possible pathways of adaptation to human hosts

    KAUST Repository

    Otto, Thomas D.

    2014-09-09

    Plasmodium falciparum causes most human malaria deaths, having prehistorically evolved from parasites of African Great Apes. Here we explore the genomic basis of P. falciparum adaptation to human hosts by fully sequencing the genome of the closely related chimpanzee parasite species P. reichenowi, and obtaining partial sequence data from a more distantly related chimpanzee parasite (P. gaboni). The close relationship between P. reichenowi and P. falciparum is emphasized by almost complete conservation of genomic synteny, but against this strikingly conserved background we observe major differences at loci involved in erythrocyte invasion. The organization of most virulence-associated multigene families, including the hypervariable var genes, is broadly conserved, but P. falciparum has a smaller subset of rif and stevor genes whose products are expressed on the infected erythrocyte surface. Genome-wide analysis identifies other loci under recent positive selection, but a limited number of changes at the host–parasite interface may have mediated host switching.

  16. Gene expression in chicken reveals correlation with structural genomic features and conserved patterns of transcription in the terrestrial vertebrates.

    Directory of Open Access Journals (Sweden)

    Haisheng Nie

    Full Text Available BACKGROUND: The chicken is an important agricultural and avian-model species. A survey of gene expression in a range of different tissues will provide a benchmark for understanding expression levels under normal physiological conditions in birds. With expression data for birds being very scant, this benchmark is of particular interest for comparative expression analysis among various terrestrial vertebrates. METHODOLOGY/PRINCIPAL FINDINGS: We carried out a gene expression survey in eight major chicken tissues using whole genome microarrays. A global picture of gene expression is presented for the eight tissues, and tissue specific as well as common gene expression were identified. A Gene Ontology (GO term enrichment analysis showed that tissue-specific genes are enriched with GO terms reflecting the physiological functions of the specific tissue, and housekeeping genes are enriched with GO terms related to essential biological functions. Comparisons of structural genomic features between tissue-specific genes and housekeeping genes show that housekeeping genes are more compact. Specifically, coding sequence and particularly introns are shorter than genes that display more variation in expression between tissues, and in addition intergenic space was also shorter. Meanwhile, housekeeping genes are more likely to co-localize with other abundantly or highly expressed genes on the same chromosomal regions. Furthermore, comparisons of gene expression in a panel of five common tissues between birds, mammals and amphibians showed that the expression patterns across tissues are highly similar for orthologous genes compared to random gene pairs within each pair-wise comparison, indicating a high degree of functional conservation in gene expression among terrestrial vertebrates. CONCLUSIONS: The housekeeping genes identified in this study have shorter gene length, shorter coding sequence length, shorter introns, and shorter intergenic regions, there seems

  17. Investigation of genome sequences within the family Pasteurellaceae

    DEFF Research Database (Denmark)

    Angen, Øystein; Ussery, David

    Introduction The bacterial genome sequences are now available for an increasing number of strains within the family Pasteurellaceae. At present, 24 Pasteurellaceae genomes are publicly available through internet databases, and another 40 genomes are being sequenced. This investigation will describe...... the core genome for both the family Pasteurellaceae and for the species Haemophilus influenzae. Methods Twenty genome sequences from the following species were included: Haemophilus influenzae (11 strains), Haemophilus ducreyi (1 strain), Histophilus somni (2 strains), Haemophilus parasuis (1 strain......), Actinobacillus pleuropneumoniae (2 strains), Actinobacillus succinogenes (1 strain), Mannheimia succiniciproducens (1 strain), and Pasteurella multocida (1 strain). The predicted proteins for each genome were BLASTed against each other, and a set of conserved core gene families was determined as described...

  18. Development and bin mapping of a Rosaceae Conserved Ortholog Set (COS of markers

    Directory of Open Access Journals (Sweden)

    Kozik Alex

    2009-01-01

    Full Text Available Abstract Background Detailed comparative genome analyses within the economically important Rosaceae family have not been conducted. This is largely due to the lack of conserved gene-based molecular markers that are transferable among the important crop genera within the family [e.g. Malus (apple, Fragaria (strawberry, and Prunus (peach, cherry, apricot and almond]. The lack of molecular markers and comparative whole genome sequence analysis for this family severely hampers crop improvement efforts as well as QTL confirmation and validation studies. Results We identified a set of 3,818 rosaceaous unigenes comprised of two or more ESTs that correspond to single copy Arabidopsis genes. From this Rosaceae Conserved Orthologous Set (RosCOS, 1039 were selected from which 857 were used for the development of intron-flanking primers and allele amplification. This led to successful amplification and subsequent mapping of 613 RosCOS onto the Prunus TxE reference map resulting in a genome-wide coverage of 0.67 to 1.06 gene-based markers per cM per linkage group. Furthermore, the RosCOS primers showed amplification success rates from 23 to 100% across the family indicating that a substantial part of the RosCOS primers can be directly employed in other less studied rosaceaous crops. Comparisons of the genetic map positions of the RosCOS with the physical locations of the orthologs in the Populus trichocarpa genome identified regions of colinearity between the genomes of Prunus-Rosaceae and Populus-Salicaceae. Conclusion Conserved orthologous genes are extremely useful for the analysis of genome evolution among closely and distantly related species. The results presented in this study demonstrate the considerable potential of the mapped Prunus RosCOS for genome-wide marker employment and comparative whole genome studies within the Rosaceae family. Moreover, these markers will also function as useful anchor points for the genome sequencing efforts currently

  19. Development and bin mapping of a Rosaceae Conserved Ortholog Set (COS) of markers.

    Science.gov (United States)

    Cabrera, Antonio; Kozik, Alex; Howad, Werner; Arus, Pere; Iezzoni, Amy F; van der Knaap, Esther

    2009-11-29

    Detailed comparative genome analyses within the economically important Rosaceae family have not been conducted. This is largely due to the lack of conserved gene-based molecular markers that are transferable among the important crop genera within the family [e.g. Malus (apple), Fragaria (strawberry), and Prunus (peach, cherry, apricot and almond)]. The lack of molecular markers and comparative whole genome sequence analysis for this family severely hampers crop improvement efforts as well as QTL confirmation and validation studies. We identified a set of 3,818 rosaceaous unigenes comprised of two or more ESTs that correspond to single copy Arabidopsis genes. From this Rosaceae Conserved Orthologous Set (RosCOS), 1039 were selected from which 857 were used for the development of intron-flanking primers and allele amplification. This led to successful amplification and subsequent mapping of 613 RosCOS onto the Prunus TxE reference map resulting in a genome-wide coverage of 0.67 to 1.06 gene-based markers per cM per linkage group. Furthermore, the RosCOS primers showed amplification success rates from 23 to 100% across the family indicating that a substantial part of the RosCOS primers can be directly employed in other less studied rosaceaous crops. Comparisons of the genetic map positions of the RosCOS with the physical locations of the orthologs in the Populus trichocarpa genome identified regions of colinearity between the genomes of Prunus-Rosaceae and Populus-Salicaceae. Conserved orthologous genes are extremely useful for the analysis of genome evolution among closely and distantly related species. The results presented in this study demonstrate the considerable potential of the mapped Prunus RosCOS for genome-wide marker employment and comparative whole genome studies within the Rosaceae family. Moreover, these markers will also function as useful anchor points for the genome sequencing efforts currently ongoing in this family as well as for comparative QTL

  20. A web-based multi-genome synteny viewer for customized data

    Directory of Open Access Journals (Sweden)

    Revanna Kashi V

    2012-08-01

    Full Text Available Abstract Background Web-based synteny visualization tools are important for sharing data and revealing patterns of complicated genome conservation and rearrangements. Such tools should allow biologists to upload genomic data for their own analysis. This requirement is critical because individual biologists are generating large amounts of genomic sequences that quickly overwhelm any centralized web resources to collect and display all those data. Recently, we published a web-based synteny viewer, GSV, which was designed to satisfy the above requirement. However, GSV can only compare two genomes at a given time. Extending the functionality of GSV to visualize multiple genomes is important to meet the increasing demand of the research community. Results We have developed a multi-Genome Synteny Viewer (mGSV. Similar to GSV, mGSV is a web-based tool that allows users to upload their own genomic data files for visualization. Multiple genomes can be presented in a single integrated view with an enhanced user interface. Users can navigate through all the selected genomes in either pairwise or multiple viewing mode to examine conserved genomic regions as well as the accompanying genome annotations. Besides serving users who manually interact with the web server, mGSV also provides Web Services for machine-to-machine communication to accept data sent by other remote resources. The entire mGSV package can also be downloaded for easy local installation. Conclusions mGSV significantly enhances the original functionalities of GSV. A web server hosting mGSV is provided at http://cas-bioinfo.cas.unt.edu/mgsv.

  1. The Salmonella enterica Pan-genome

    DEFF Research Database (Denmark)

    Jacobsen, Annika; Hendriksen, Rene S.; Aarestrup, Frank Møller

    2011-01-01

    Salmonella enterica is divided into four subspecies containing a large number of different serovars, several of which are important zoonotic pathogens and some show a high degree of host specificity or host preference. We compare 45 sequenced S. enterica genomes that are publicly available (22......, and the core and pan-genome of Salmonella were estimated to be around 2,800 and 10,000 gene families, respectively. The constructed pan-genomic dendrograms suggest that gene content is often, but not uniformly correlated to serotype. Any given Salmonella strain has a large stable core, whilst...... there is an abundance of accessory genes, including the Salmonella pathogenicity islands (SPIs), transposable elements, phages, and plasmid DNA. We visualize conservation in the genomes in relation to chromosomal location and DNA structural features and find that variation in gene content is localized in a selection...

  2. Phylogenomic Analysis and Dynamic Evolution of Chloroplast Genomes in Salicaceae

    Directory of Open Access Journals (Sweden)

    Yuan Huang

    2017-06-01

    Full Text Available Chloroplast genomes of plants are highly conserved in both gene order and gene content. Analysis of the whole chloroplast genome is known to provide much more informative DNA sites and thus generates high resolution for plant phylogenies. Here, we report the complete chloroplast genomes of three Salix species in family Salicaceae. Phylogeny of Salicaceae inferred from complete chloroplast genomes is generally consistent with previous studies but resolved with higher statistical support. Incongruences of phylogeny, however, are observed in genus Populus, which most likely results from homoplasy. By comparing three Salix chloroplast genomes with the published chloroplast genomes of other Salicaceae species, we demonstrate that the synteny and length of chloroplast genomes in Salicaceae are highly conserved but experienced dynamic evolution among species. We identify seven positively selected chloroplast genes in Salicaceae, which might be related to the adaptive evolution of Salicaceae species. Comparative chloroplast genome analysis within the family also indicates that some chloroplast genes are lost or became pseudogenes, infer that the chloroplast genes horizontally transferred to the nucleus genome. Based on the complete nucleus genome sequences from two Salicaceae species, we remarkably identify that the entire chloroplast genome is indeed transferred and integrated to the nucleus genome in the individual of the reference genome of P. trichocarpa at least once. This observation, along with presence of the large nuclear plastid DNA (NUPTs and NUPTs-containing multiple chloroplast genes in their original order in the chloroplast genome, favors the DNA-mediated hypothesis of organelle to nucleus DNA transfer. Overall, the phylogenomic analysis using chloroplast complete genomes clearly elucidates the phylogeny of Salicaceae. The identification of positively selected chloroplast genes and dynamic chloroplast-to-nucleus gene transfers in

  3. Genome-wide identification of the regulatory targets of a transcription factor using biochemical characterization and computational genomic analysis

    Directory of Open Access Journals (Sweden)

    Jolly Emmitt R

    2005-11-01

    Full Text Available Abstract Background A major challenge in computational genomics is the development of methodologies that allow accurate genome-wide prediction of the regulatory targets of a transcription factor. We present a method for target identification that combines experimental characterization of binding requirements with computational genomic analysis. Results Our method identified potential target genes of the transcription factor Ndt80, a key transcriptional regulator involved in yeast sporulation, using the combined information of binding affinity, positional distribution, and conservation of the binding sites across multiple species. We have also developed a mathematical approach to compute the false positive rate and the total number of targets in the genome based on the multiple selection criteria. Conclusion We have shown that combining biochemical characterization and computational genomic analysis leads to accurate identification of the genome-wide targets of a transcription factor. The method can be extended to other transcription factors and can complement other genomic approaches to transcriptional regulation.

  4. Quality scores for 32,000 genomes

    DEFF Research Database (Denmark)

    Land, Miriam L.; Hyatt, Doug; Jun, Se-Ran

    2014-01-01

    Background More than 80% of the microbial genomes in GenBank are of ‘draft’ quality (12,553 draft vs. 2,679 finished, as of October, 2013). We have examined all the microbial DNA sequences available for complete, draft, and Sequence Read Archive genomes in GenBank as well as three other major...... public databases, and assigned quality scores for more than 30,000 prokaryotic genome sequences. Results Scores were assigned using four categories: the completeness of the assembly, the presence of full-length rRNA genes, tRNA composition and the presence of a set of 102 conserved genes in prokaryotes....... Most (~88%) of the genomes had quality scores of 0.8 or better and can be safely used for standard comparative genomics analysis. We compared genomes across factors that may influence the score. We found that although sequencing depth coverage of over 100x did not ensure a better score, sequencing read...

  5. Efficient oligonucleotide probe selection for pan-genomic tiling arrays

    Directory of Open Access Journals (Sweden)

    Zhang Wei

    2009-09-01

    Full Text Available Abstract Background Array comparative genomic hybridization is a fast and cost-effective method for detecting, genotyping, and comparing the genomic sequence of unknown bacterial isolates. This method, as with all microarray applications, requires adequate coverage of probes targeting the regions of interest. An unbiased tiling of probes across the entire length of the genome is the most flexible design approach. However, such a whole-genome tiling requires that the genome sequence is known in advance. For the accurate analysis of uncharacterized bacteria, an array must query a fully representative set of sequences from the species' pan-genome. Prior microarrays have included only a single strain per array or the conserved sequences of gene families. These arrays omit potentially important genes and sequence variants from the pan-genome. Results This paper presents a new probe selection algorithm (PanArray that can tile multiple whole genomes using a minimal number of probes. Unlike arrays built on clustered gene families, PanArray uses an unbiased, probe-centric approach that does not rely on annotations, gene clustering, or multi-alignments. Instead, probes are evenly tiled across all sequences of the pan-genome at a consistent level of coverage. To minimize the required number of probes, probes conserved across multiple strains in the pan-genome are selected first, and additional probes are used only where necessary to span polymorphic regions of the genome. The viability of the algorithm is demonstrated by array designs for seven different bacterial pan-genomes and, in particular, the design of a 385,000 probe array that fully tiles the genomes of 20 different Listeria monocytogenes strains with overlapping probes at greater than twofold coverage. Conclusion PanArray is an oligonucleotide probe selection algorithm for tiling multiple genome sequences using a minimal number of probes. It is capable of fully tiling all genomes of a species on

  6. CrusView: a Java-based visualization platform for comparative genomics analyses in Brassicaceae species.

    Science.gov (United States)

    Chen, Hao; Wang, Xiangfeng

    2013-09-01

    In plants and animals, chromosomal breakage and fusion events based on conserved syntenic genomic blocks lead to conserved patterns of karyotype evolution among species of the same family. However, karyotype information has not been well utilized in genomic comparison studies. We present CrusView, a Java-based bioinformatic application utilizing Standard Widget Toolkit/Swing graphics libraries and a SQLite database for performing visualized analyses of comparative genomics data in Brassicaceae (crucifer) plants. Compared with similar software and databases, one of the unique features of CrusView is its integration of karyotype information when comparing two genomes. This feature allows users to perform karyotype-based genome assembly and karyotype-assisted genome synteny analyses with preset karyotype patterns of the Brassicaceae genomes. Additionally, CrusView is a local program, which gives its users high flexibility when analyzing unpublished genomes and allows users to upload self-defined genomic information so that they can visually study the associations between genome structural variations and genetic elements, including chromosomal rearrangements, genomic macrosynteny, gene families, high-frequency recombination sites, and tandem and segmental duplications between related species. This tool will greatly facilitate karyotype, chromosome, and genome evolution studies using visualized comparative genomics approaches in Brassicaceae species. CrusView is freely available at http://www.cmbb.arizona.edu/CrusView/.

  7. Comparative genomic analysis of four representative plant growth-promoting rhizobacteria in Pseudomonas

    Science.gov (United States)

    2013-01-01

    Background Some Pseudomonas strains function as predominant plant growth-promoting rhizobacteria (PGPR). Within this group, Pseudomonas chlororaphis and Pseudomonas fluorescens are non-pathogenic biocontrol agents, and some Pseudomonas aeruginosa and Pseudomonas stutzeri strains are PGPR. P. chlororaphis GP72 is a plant growth-promoting rhizobacterium with a fully sequenced genome. We conducted a genomic analysis comparing GP72 with three other pseudomonad PGPR: P. fluorescens Pf-5, P. aeruginosa M18, and the nitrogen-fixing strain P. stutzeri A1501. Our aim was to identify the similarities and differences among these strains using a comparative genomic approach to clarify the mechanisms of plant growth-promoting activity. Results The genome sizes of GP72, Pf-5, M18, and A1501 ranged from 4.6 to 7.1 M, and the number of protein-coding genes varied among the four species. Clusters of Orthologous Groups (COGs) analysis assigned functions to predicted proteins. The COGs distributions were similar among the four species. However, the percentage of genes encoding transposases and their inactivated derivatives (COG L) was 1.33% of the total genes with COGs classifications in A1501, 0.21% in GP72, 0.02% in Pf-5, and 0.11% in M18. A phylogenetic analysis indicated that GP72 and Pf-5 were the most closely related strains, consistent with the genome alignment results. Comparisons of predicted coding sequences (CDSs) between GP72 and Pf-5 revealed 3544 conserved genes. There were fewer conserved genes when GP72 CDSs were compared with those of A1501 and M18. Comparisons among the four Pseudomonas species revealed 603 conserved genes in GP72, illustrating common plant growth-promoting traits shared among these PGPR. Conserved genes were related to catabolism, transport of plant-derived compounds, stress resistance, and rhizosphere colonization. Some strain-specific CDSs were related to different kinds of biocontrol activities or plant growth promotion. The GP72 genome

  8. Conservation, Divergence, and Genome-Wide Distribution of PAL and POX A Gene Families in Plants.

    Science.gov (United States)

    Rawal, H C; Singh, N K; Sharma, T R

    2013-01-01

    Genome-wide identification and phylogenetic and syntenic comparison were performed for the genes responsible for phenylalanine ammonia lyase (PAL) and peroxidase A (POX A) enzymes in nine plant species representing very diverse groups like legumes (Glycine max and Medicago truncatula), fruits (Vitis vinifera), cereals (Sorghum bicolor, Zea mays, and Oryza sativa), trees (Populus trichocarpa), and model dicot (Arabidopsis thaliana) and monocot (Brachypodium distachyon) species. A total of 87 and 1045 genes in PAL and POX A gene families, respectively, have been identified in these species. The phylogenetic and syntenic comparison along with motif distributions shows a high degree of conservation of PAL genes, suggesting that these genes may predate monocot/eudicot divergence. The POX A family genes, present in clusters at the subtelomeric regions of chromosomes, might be evolving and expanding with higher rate than the PAL gene family. Our analysis showed that during the expansion of POX A gene family, many groups and subgroups have evolved, resulting in a high level of functional divergence among monocots and dicots. These results will act as a first step toward the understanding of monocot/eudicot evolution and functional characterization of these gene families in the future.

  9. Conservation, Divergence, and Genome-Wide Distribution of PAL and POX A Gene Families in Plants

    Directory of Open Access Journals (Sweden)

    H. C. Rawal

    2013-01-01

    Full Text Available Genome-wide identification and phylogenetic and syntenic comparison were performed for the genes responsible for phenylalanine ammonia lyase (PAL and peroxidase A (POX A enzymes in nine plant species representing very diverse groups like legumes (Glycine max and Medicago truncatula, fruits (Vitis vinifera, cereals (Sorghum bicolor, Zea mays, and Oryza sativa, trees (Populus trichocarpa, and model dicot (Arabidopsis thaliana and monocot (Brachypodium distachyon species. A total of 87 and 1045 genes in PAL and POX A gene families, respectively, have been identified in these species. The phylogenetic and syntenic comparison along with motif distributions shows a high degree of conservation of PAL genes, suggesting that these genes may predate monocot/eudicot divergence. The POX A family genes, present in clusters at the subtelomeric regions of chromosomes, might be evolving and expanding with higher rate than the PAL gene family. Our analysis showed that during the expansion of POX A gene family, many groups and subgroups have evolved, resulting in a high level of functional divergence among monocots and dicots. These results will act as a first step toward the understanding of monocot/eudicot evolution and functional characterization of these gene families in the future.

  10. The Geobacillus pan-genome: implications for the evolution of the genus

    Directory of Open Access Journals (Sweden)

    Oliver Keoagile Ignatius Bezuidt

    2016-05-01

    Full Text Available The genus Geobacillus is comprised of a diverse group of spore-forming Gram-positive thermophilic bacterial species and is well known for both its ecological diversity and as a source of novel thermostable enzymes. Although the mechanisms underlying the thermophilicity of the organism and the thermostability of its macromolecules are reasonably well understood, relatively little is known of the evolutionary mechanisms, which underlie the structural and functional properties of members of this genus. In this study, we have compared 29 Geobacillus genomes, with a specific focus on the elements, which comprise the conserved core and flexible genomes. Based on comparisons of conserved core and flexible genomes, we present evidence of habitat delineation with specific Geobacillus genomes linked to specific niches. Interestingly, our analysis has shown that horizontal gene transfer is a major factor deriving the evolution of Geobacillus from Bacillus, with genetic contributions from other phylogenetically distant taxa.

  11. Evolutionary Genomics of an Ancient Prophage of the Order Sphingomonadales

    Science.gov (United States)

    Viswanathan, Vandana; Narjala, Anushree; Ravichandran, Aravind; Jayaprasad, Suvratha

    2017-01-01

    The order Sphingomonadales, containing the families Erythrobacteraceae and Sphingomonadaceae, is a relatively less well-studied phylogenetic branch within the class Alphaproteobacteria. Prophage elements are present in most bacterial genomes and are important determinants of adaptive evolution. An “intact” prophage was predicted within the genome of Sphingomonas hengshuiensis strain WHSC-8 and was designated Prophage IWHSC-8. Loci homologous to the region containing the first 22 open reading frames (ORFs) of Prophage IWHSC-8 were discovered among the genomes of numerous Sphingomonadales. In 17 genomes, the homologous loci were co-located with an ORF encoding a putative superoxide dismutase. Several other lines of molecular evidence implied that these homologous loci represent an ancient temperate bacteriophage integration, and this horizontal transfer event pre-dated niche-based speciation within the order Sphingomonadales. The “stabilization” of prophages in the genomes of their hosts is an indicator of “fitness” conferred by these elements and natural selection. Among the various ORFs predicted within the conserved prophages, an ORF encoding a putative proline-rich outer membrane protein A was consistently present among the genomes of many Sphingomonadales. Furthermore, the conserved prophages in six Sphingomonas sp. contained an ORF encoding a putative spermidine synthase. It is possible that one or more of these ORFs bestow selective fitness, and thus the prophages continue to be vertically transferred within the host strains. Although conserved prophages have been identified previously among closely related genera and species, this is the first systematic and detailed description of orthologous prophages at the level of an order that contains two diverse families and many pigmented species. PMID:28201618

  12. A linear mitochondrial genome of Cyclospora cayetanensis (Eimeriidae, Eucoccidiorida, Coccidiasina, Apicomplexa) suggests the ancestral start position within mitochondrial genomes of eimeriid coccidia.

    Science.gov (United States)

    Ogedengbe, Mosun E; Qvarnstrom, Yvonne; da Silva, Alexandre J; Arrowood, Michael J; Barta, John R

    2015-05-01

    The near complete mitochondrial genome for Cyclospora cayetanensis is 6184 bp in length with three protein-coding genes (Cox1, Cox3, CytB) and numerous lsrDNA and ssrDNA fragments. Gene arrangements were conserved with other coccidia in the Eimeriidae, but the C. cayetanensis mitochondrial genome is not circular-mapping. Terminal transferase tailing and nested PCR completed the 5'-terminus of the genome starting with a 21 bp A/T-only region that forms a potential stem-loop. Regions homologous to the C. cayetanensis mitochondrial genome 5'-terminus are found in all eimeriid mitochondrial genomes available and suggest this may be the ancestral start of eimeriid mitochondrial genomes. Copyright © 2015 Australian Society for Parasitology Inc. All rights reserved.

  13. The zebrafish genome: a review and msx gene case study.

    Science.gov (United States)

    Postlethwait, J H

    2006-01-01

    Zebrafish is one of several important teleost models for understanding principles of vertebrate developmental, molecular, organismal, genetic, evolutionary, and genomic biology. Efficient investigation of the molecular genetic basis of induced mutations depends on knowledge of the zebrafish genome. Principles of zebrafish genomic analysis, including gene mapping, ortholog identification, conservation of syntenies, genome duplication, and evolution of duplicate gene function are discussed here using as a case study the zebrafish msxa, msxb, msxc, msxd, and msxe genes, which together constitute zebrafish orthologs of tetrapod Msx1, Msx2, and Msx3. Genomic analysis suggests orthologs for this difficult to understand group of paralogs.

  14. Draft genome of the American Eel (Anguilla rostrata).

    Science.gov (United States)

    Pavey, Scott A; Laporte, Martin; Normandeau, Eric; Gaudin, Jérémy; Letourneau, Louis; Boisvert, Sébastien; Corbeil, Jacques; Audet, Céline; Bernatchez, Louis

    2017-07-01

    Freshwater eels (Anguilla sp.) have large economic, cultural, ecological and aesthetic importance worldwide, but they suffered more than 90% decline in global stocks over the past few decades. Proper genetic resources, such as sequenced, assembled and annotated genomes, are essential to help plan sustainable recoveries by identifying physiological, biochemical and genetic mechanisms that caused the declines or that may lead to recoveries. Here, we present the first sequenced genome of the American eel. This genome contained 305 043 contigs (N50 = 7397) and 79 209 scaffolds (N50 = 86 641) for a total size of 1.41 Gb, which is in the middle of the range of previous estimations for this species. In addition, protein-coding regions, including introns and flanking regions, are very well represented in the genome, as 95.2% of the 458 core eukaryotic genes and 98.8% of the 248 ultra-conserved subset were represented in the assembly and a total of 26 564 genes were annotated for future functional genomics studies. We performed a candidate gene analysis to compare three genes among all three freshwater eel species and, congruent with the phylogenetic relationships, Japanese eel (A. japanica) exhibited the most divergence. Overall, the sequenced genome presented in this study is a crucial addition to the presently available genetic tools to help guide future conservation efforts of freshwater eels. © 2016 John Wiley & Sons Ltd.

  15. Towards understanding the first genome sequence of a crenarchaeon by genome annotation using clusters of orthologous groups of proteins (COGs).

    Science.gov (United States)

    Natale, D A; Shankavaram, U T; Galperin, M Y; Wolf, Y I; Aravind, L; Koonin, E V

    2000-01-01

    Standard archival sequence databases have not been designed as tools for genome annotation and are far from being optimal for this purpose. We used the database of Clusters of Orthologous Groups of proteins (COGs) to reannotate the genomes of two archaea, Aeropyrum pernix, the first member of the Crenarchaea to be sequenced, and Pyrococcus abyssi. A. pernix and P. abyssi proteins were assigned to COGs using the COGNITOR program; the results were verified on a case-by-case basis and augmented by additional database searches using the PSI-BLAST and TBLASTN programs. Functions were predicted for over 300 proteins from A. pernix, which could not be assigned a function using conventional methods with a conservative sequence similarity threshold, an approximately 50% increase compared to the original annotation. A. pernix shares most of the conserved core of proteins that were previously identified in the Euryarchaeota. Cluster analysis or distance matrix tree construction based on the co-occurrence of genomes in COGs showed that A. pernix forms a distinct group within the archaea, although grouping with the two species of Pyrococci, indicative of similar repertoires of conserved genes, was observed. No indication of a specific relationship between Crenarchaeota and eukaryotes was obtained in these analyses. Several proteins that are conserved in Euryarchaeota and most bacteria are unexpectedly missing in A. pernix, including the entire set of de novo purine biosynthesis enzymes, the GTPase FtsZ (a key component of the bacterial and euryarchaeal cell-division machinery), and the tRNA-specific pseudouridine synthase, previously considered universal. A. pernix is represented in 48 COGs that do not contain any euryarchaeal members. Many of these proteins are TCA cycle and electron transport chain enzymes, reflecting the aerobic lifestyle of A. pernix. Special-purpose databases organized on the basis of phylogenetic analysis and carefully curated with respect to known and

  16. The Global Invertebrate Genomics Alliance (GIGA): Developing Community Resources to Study Diverse Invertebrate Genomes

    KAUST Repository

    Bracken-Grissom, Heather

    2013-12-12

    Over 95% of all metazoan (animal) species comprise the invertebrates, but very few genomes from these organisms have been sequenced. We have, therefore, formed a Global Invertebrate Genomics Alliance (GIGA). Our intent is to build a collaborative network of diverse scientists to tackle major challenges (e.g., species selection, sample collection and storage, sequence assembly, annotation, analytical tools) associated with genome/transcriptome sequencing across a large taxonomic spectrum. We aim to promote standards that will facilitate comparative approaches to invertebrate genomics and collaborations across the international scientific community. Candidate study taxa include species from Porifera, Ctenophora, Cnidaria, Placozoa, Mollusca, Arthropoda, Echinodermata, Annelida, Bryozoa, and Platyhelminthes, among others. GIGA will target 7000 noninsect/nonnematode species, with an emphasis on marine taxa because of the unrivaled phyletic diversity in the oceans. Priorities for selecting invertebrates for sequencing will include, but are not restricted to, their phylogenetic placement; relevance to organismal, ecological, and conservation research; and their importance to fisheries and human health. We highlight benefits of sequencing both whole genomes (DNA) and transcriptomes and also suggest policies for genomic-level data access and sharing based on transparency and inclusiveness. The GIGA Web site () has been launched to facilitate this collaborative venture.

  17. Divergence of RNA polymerase α subunits in angiosperm plastid genomes is mediated by genomic rearrangement.

    Science.gov (United States)

    Blazier, J Chris; Ruhlman, Tracey A; Weng, Mao-Lun; Rehman, Sumaiyah K; Sabir, Jamal S M; Jansen, Robert K

    2016-04-18

    Genes for the plastid-encoded RNA polymerase (PEP) persist in the plastid genomes of all photosynthetic angiosperms. However, three unrelated lineages (Annonaceae, Passifloraceae and Geraniaceae) have been identified with unusually divergent open reading frames (ORFs) in the conserved region of rpoA, the gene encoding the PEP α subunit. We used sequence-based approaches to evaluate whether these genes retain function. Both gene sequences and complete plastid genome sequences were assembled and analyzed from each of the three angiosperm families. Multiple lines of evidence indicated that the rpoA sequences are likely functional despite retaining as low as 30% nucleotide sequence identity with rpoA genes from outgroups in the same angiosperm order. The ratio of non-synonymous to synonymous substitutions indicated that these genes are under purifying selection, and bioinformatic prediction of conserved domains indicated that functional domains are preserved. One of the lineages (Pelargonium, Geraniaceae) contains species with multiple rpoA-like ORFs that show evidence of ongoing inter-paralog gene conversion. The plastid genomes containing these divergent rpoA genes have experienced extensive structural rearrangement, including large expansions of the inverted repeat. We propose that illegitimate recombination, not positive selection, has driven the divergence of rpoA.

  18. Genome sequence of the lager brewing yeast, an interspecies hybrid.

    Science.gov (United States)

    Nakao, Yoshihiro; Kanamori, Takeshi; Itoh, Takehiko; Kodama, Yukiko; Rainieri, Sandra; Nakamura, Norihisa; Shimonaga, Tomoko; Hattori, Masahira; Ashikari, Toshihiko

    2009-04-01

    This work presents the genome sequencing of the lager brewing yeast (Saccharomyces pastorianus) Weihenstephan 34/70, a strain widely used in lager beer brewing. The 25 Mb genome comprises two nuclear sub-genomes originating from Saccharomyces cerevisiae and Saccharomyces bayanus and one circular mitochondrial genome originating from S. bayanus. Thirty-six different types of chromosomes were found including eight chromosomes with translocations between the two sub-genomes, whose breakpoints are within the orthologous open reading frames. Several gene loci responsible for typical lager brewing yeast characteristics such as maltotriose uptake and sulfite production have been increased in number by chromosomal rearrangements. Despite an overall high degree of conservation of the synteny with S. cerevisiae and S. bayanus, the syntenies were not well conserved in the sub-telomeric regions that contain lager brewing yeast characteristic and specific genes. Deletion of larger chromosomal regions, a massive unilateral decrease of the ribosomal DNA cluster and bilateral truncations of over 60 genes reflect a post-hybridization evolution process. Truncations and deletions of less efficient maltose and maltotriose uptake genes may indicate the result of adaptation to brewing. The genome sequence of this interspecies hybrid yeast provides a new tool for better understanding of lager brewing yeast behavior in industrial beer production.

  19. Global MLST of Salmonella Typhi Revisited in Post-Genomic Era: Genetic conservation, Population Structure and Comparative genomics of rare sequence types

    Directory of Open Access Journals (Sweden)

    Kien-Pong eYap

    2016-03-01

    Full Text Available Typhoid fever, caused by Salmonella enterica serovar Typhi, remains an important public health burden in Southeast Asia and other endemic countries. Various genotyping methods have been applied to study the genetic variations of this human-restricted pathogen. Multilocus Sequence Typing (MLST is one of the widely accepted methods, and recently, there is a growing interest in the re-application of MLST in the post-genomic era. In this study, we provide the global MLST distribution of S. Typhi utilizing both publicly available 1,826 S. Typhi genome sequences in addition to performing conventional MLST on S. Typhi strains isolated from various endemic regions spanning over a century. Our global MLST analysis confirms the predominance of two sequence types (ST1 and ST2 co-existing in the endemic regions. Interestingly, S. Typhi strains with ST8 are currently confined within the African continent. Comparative genomic analyses of ST8 and other rare STs with genomes of ST1/ST2 revealed unique mutations in important virulence genes such as flhB, sipC and tviD that may explain the variations that differentiate between seemingly successful (widespread and unsuccessful (poor dissemination S. Typhi populations. Large scale whole-genome phylogeny demonstrated evidence of phylogeographical structuring and showed that ST8 may have diverged from the earlier ancestral population of ST1 and ST2, which later lost some of its fitness advantages, leading to poor worldwide dissemination. In response to the unprecedented increase in genomic data, this study demonstrates and highlights the utility of large-scale genome-based MLST as a quick and effective approach to narrow the scope of in-depth comparative genomic analysis and consequently provide new insights into the fine scale of pathogen evolution and population structure.

  20. Evolution of small prokaryotic genomes

    Directory of Open Access Journals (Sweden)

    David José Martínez-Cano

    2015-01-01

    Full Text Available As revealed by genome sequencing, the biology of prokaryotes with reduced genomes is strikingly diverse. These include free-living prokaryotes with ~800 genes as well as endosymbiotic bacteria with as few as ~140 genes. Comparative genomics is revealing the evolutionary mechanisms that led to these small genomes. In the case of free-living prokaryotes, natural selection directly favored genome reduction, while in the case of endosymbiotic prokaryotes neutral processes played a more prominent role. However, new experimental data suggest that selective processes may be at operation as well for endosymbiotic prokaryotes at least during the first stages of genome reduction. Endosymbiotic prokaryotes have evolved diverse strategies for living with reduced gene sets inside a host-defined medium. These include utilization of host-encoded functions (some of them coded by genes acquired by gene transfer from the endosymbiont and/or other bacteria; metabolic complementation between co-symbionts; and forming consortiums with other bacteria within the host. Recent genome sequencing projects of intracellular mutualistic bacteria showed that previously believed universal evolutionary trends like reduced G+C content and conservation of genome synteny are not always present in highly reduced genomes. Finally, the simplified molecular machinery of some of these organisms with small genomes may be used to aid in the design of artificial minimal cells. Here we review recent genomic discoveries of the biology of prokaryotes endowed with small gene sets and discuss the evolutionary mechanisms that have been proposed to explain their peculiar nature.

  1. Comparative Reannotation of 21 Aspergillus Genomes

    Energy Technology Data Exchange (ETDEWEB)

    Salamov, Asaf; Riley, Robert; Kuo, Alan; Grigoriev, Igor

    2013-03-08

    We used comparative gene modeling to reannotate 21 Aspergillus genomes. Initial automatic annotation of individual genomes may contain some errors of different nature, e.g. missing genes, incorrect exon-intron structures, 'chimeras', which fuse 2 or more real genes or alternatively splitting some real genes into 2 or more models. The main premise behind the comparative modeling approach is that for closely related genomes most orthologous families have the same conserved gene structure. The algorithm maps all gene models predicted in each individual Aspergillus genome to the other genomes and, for each locus, selects from potentially many competing models, the one which most closely resembles the orthologous genes from other genomes. This procedure is iterated until no further change in gene models is observed. For Aspergillus genomes we predicted in total 4503 new gene models ( ~;;2percent per genome), supported by comparative analysis, additionally correcting ~;;18percent of old gene models. This resulted in a total of 4065 more genes with annotated PFAM domains (~;;3percent increase per genome). Analysis of a few genomes with EST/transcriptomics data shows that the new annotation sets also have a higher number of EST-supported splice sites at exon-intron boundaries.

  2. The Complete Chloroplast and Mitochondrial Genome Sequences of Boea hygrometrica: Insights into the Evolution of Plant Organellar Genomes

    Science.gov (United States)

    Wang, Xumin; Deng, Xin; Zhang, Xiaowei; Hu, Songnian; Yu, Jun

    2012-01-01

    The complete nucleotide sequences of the chloroplast (cp) and mitochondrial (mt) genomes of resurrection plant Boea hygrometrica (Bh, Gesneriaceae) have been determined with the lengths of 153,493 bp and 510,519 bp, respectively. The smaller chloroplast genome contains more genes (147) with a 72% coding sequence, and the larger mitochondrial genome have less genes (65) with a coding faction of 12%. Similar to other seed plants, the Bh cp genome has a typical quadripartite organization with a conserved gene in each region. The Bh mt genome has three recombinant sequence repeats of 222 bp, 843 bp, and 1474 bp in length, which divide the genome into a single master circle (MC) and four isomeric molecules. Compared to other angiosperms, one remarkable feature of the Bh mt genome is the frequent transfer of genetic material from the cp genome during recent Bh evolution. We also analyzed organellar genome evolution in general regarding genome features as well as compositional dynamics of sequence and gene structure/organization, providing clues for the understanding of the evolution of organellar genomes in plants. The cp-derived sequences including tRNAs found in angiosperm mt genomes support the conclusion that frequent gene transfer events may have begun early in the land plant lineage. PMID:22291979

  3. The Methanosarcina barkeri genome: comparative analysis withMethanosarcina acetivorans and Methanosarcina mazei reveals extensiverearrangement within methanosarcinal genomes

    Energy Technology Data Exchange (ETDEWEB)

    Maeder, Dennis L.; Anderson, Iain; Brettin, Thomas S.; Bruce,David C.; Gilna, Paul; Han, Cliff S.; Lapidus, Alla; Metcalf, William W.; Saunders, Elizabeth; Tapia, Roxanne; Sowers, Kevin R.

    2006-05-19

    We report here a comparative analysis of the genome sequence of Methanosarcina barkeri with those of Methanosarcina acetivorans and Methanosarcina mazei. All three genomes share a conserved double origin of replication and many gene clusters. M. barkeri is distinguished by having an organization that is well conserved with respect to the other Methanosarcinae in the region proximal to the origin of replication with interspecies gene similarities as high as 95%. However it is disordered and marked by increased transposase frequency and decreased gene synteny and gene density in the proximal semi-genome. Of the 3680 open reading frames in M. barkeri, 678 had paralogs with better than 80% similarity to both M. acetivorans and M. mazei while 128 nonhypothetical orfs were unique (non-paralogous) amongst these species including a complete formate dehydrogenase operon, two genes required for N-acetylmuramic acid synthesis, a 14 gene gas vesicle cluster and a bacterial P450-specific ferredoxin reductase cluster not previously observed or characterized in this genus. A cryptic 36 kbp plasmid sequence was detected in M. barkeri that contains an orc1 gene flanked by a presumptive origin of replication consisting of 38 tandem repeats of a 143 nt motif. Three-way comparison of these genomes reveals differing mechanisms for the accrual of changes. Elongation of the large M. acetivorans is the result of multiple gene-scale insertions and duplications uniformly distributed in that genome, while M. barkeri is characterized by localized inversions associated with the loss of gene content. In contrast, the relatively short M. mazei most closely approximates the ancestral organizational state.

  4. Evidence for widespread degradation of gene control regions in hominid genomes.

    Directory of Open Access Journals (Sweden)

    Peter D Keightley

    2005-02-01

    Full Text Available Although sequences containing regulatory elements located close to protein-coding genes are often only weakly conserved during evolution, comparisons of rodent genomes have implied that these sequences are subject to some selective constraints. Evolutionary conservation is particularly apparent upstream of coding sequences and in first introns, regions that are enriched for regulatory elements. By comparing the human and chimpanzee genomes, we show here that there is almost no evidence for conservation in these regions in hominids. Furthermore, we show that gene expression is diverging more rapidly in hominids than in murids per unit of neutral sequence divergence. By combining data on polymorphism levels in human noncoding DNA and the corresponding human-chimpanzee divergence, we show that the proportion of adaptive substitutions in these regions in hominids is very low. It therefore seems likely that the lack of conservation and increased rate of gene expression divergence are caused by a reduction in the effectiveness of natural selection against deleterious mutations because of the low effective population sizes of hominids. This has resulted in the accumulation of a large number of deleterious mutations in sequences containing gene control elements and hence a widespread degradation of the genome during the evolution of humans and chimpanzees.

  5. Genes Important for Schizosaccharomyces pombe Meiosis Identified Through a Functional Genomics Screen

    Science.gov (United States)

    Blyth, Julie; Makrantoni, Vasso; Barton, Rachael E.; Spanos, Christos; Rappsilber, Juri; Marston, Adele L.

    2018-01-01

    Meiosis is a specialized cell division that generates gametes, such as eggs and sperm. Errors in meiosis result in miscarriages and are the leading cause of birth defects; however, the molecular origins of these defects remain unknown. Studies in model organisms are beginning to identify the genes and pathways important for meiosis, but the parts list is still poorly defined. Here we present a comprehensive catalog of genes important for meiosis in the fission yeast, Schizosaccharomyces pombe. Our genome-wide functional screen surveyed all nonessential genes for roles in chromosome segregation and spore formation. Novel genes important at distinct stages of the meiotic chromosome segregation and differentiation program were identified. Preliminary characterization implicated three of these genes in centrosome/spindle pole body, centromere, and cohesion function. Our findings represent a near-complete parts list of genes important for meiosis in fission yeast, providing a valuable resource to advance our molecular understanding of meiosis. PMID:29259000

  6. Azolla--a model organism for plant genomic studies.

    Science.gov (United States)

    Qiu, Yin-Long; Yu, Jun

    2003-02-01

    The aquatic ferns of the genus Azolla are nitrogen-fixing plants that have great potentials in agricultural production and environmental conservation. Azolla in many aspects is qualified to serve as a model organism for genomic studies because of its importance in agriculture, its unique position in plant evolution, its symbiotic relationship with the N2-fixing cyanobacterium, Anabaena azollae, and its moderate-sized genome. The goals of this genome project are not only to understand the biology of the Azolla genome to promote its applications in biological research and agriculture practice but also to gain critical insights about evolution of plant genomes. Together with the strategic and technical improvement as well as cost reduction of DNA sequencing, the deciphering of their genetic code is imminent.

  7. Does the evolutionary conservation of microsatellite loci imply function?

    Energy Technology Data Exchange (ETDEWEB)

    Shriver, M.D.; Deka, R.; Ferrell, R.E. [Univ. of Pittsburgh, PA (United States)] [and others

    1994-09-01

    Microsatellites are highly polymorphic tandem arrays of short (1-6 bp) sequence motifs which have been found widely distributed in the genomes of all eukaryotes. We have analyzed allele frequency data on 16 microsatellite loci typed in the great apes (human, chimp, orangutan, and gorilla). The majority of these loci (13) were isolated from human genomic libraries; three were cloned from chimpanzee genomic DNA. Most of these loci are not only present in all apes species, but are polymorphic with comparable levels of heterozygosity and have alleles which overlap in size. The extent of divergence of allele frequencies among these four species were studies using the stepwise-weighted genetic distance (Dsw), which was previously shown to conform to linearity with evolutionary time since divergence for loci where mutations exist in a stepwise fashion. The phylogenetic tree of the great apes constructed from this distance matrix was consistent with the expected topology, with a high bootstrap confidence (82%) for the human/chimp clade. However, the allele frequency distributions of these species are 10 times more similar to each other than expected when they were calibrated with a conservative estimate of the time since separation of humans and the apes. These results are in agreement with sequence-based surveys of microsatellites which have demonstrated that they are highly (90%) conserved over short periods of evolutionary time (< 10 million years) and moderately (30%) conserved over long periods of evolutionary time (> 60-80 million years). This evolutionary conservation has prompted some authors to speculate that there are functional constraints on microsatellite loci. In contrast, the presence of directional bias of mutations with constraints and/or selection against aberrant sized alleles can explain these results.

  8. RNA expression in a cartilaginous fish cell line reveals ancient 3′ noncoding regions highly conserved in vertebrates

    Science.gov (United States)

    Forest, David; Nishikawa, Ryuhei; Kobayashi, Hiroshi; Parton, Angela; Bayne, Christopher J.; Barnes, David W.

    2007-01-01

    We have established a cartilaginous fish cell line [Squalus acanthias embryo cell line (SAE)], a mesenchymal stem cell line derived from the embryo of an elasmobranch, the spiny dogfish shark S. acanthias. Elasmobranchs (sharks and rays) first appeared >400 million years ago, and existing species provide useful models for comparative vertebrate cell biology, physiology, and genomics. Comparative vertebrate genomics among evolutionarily distant organisms can provide sequence conservation information that facilitates identification of critical coding and noncoding regions. Although these genomic analyses are informative, experimental verification of functions of genomic sequences depends heavily on cell culture approaches. Using ESTs defining mRNAs derived from the SAE cell line, we identified lengthy and highly conserved gene-specific nucleotide sequences in the noncoding 3′ UTRs of eight genes involved in the regulation of cell growth and proliferation. Conserved noncoding 3′ mRNA regions detected by using the shark nucleotide sequences as a starting point were found in a range of other vertebrate orders, including bony fish, birds, amphibians, and mammals. Nucleotide identity of shark and human in these regions was remarkably well conserved. Our results indicate that highly conserved gene sequences dating from the appearance of jawed vertebrates and representing potential cis-regulatory elements can be identified through the use of cartilaginous fish as a baseline. Because the expression of genes in the SAE cell line was prerequisite for their identification, this cartilaginous fish culture system also provides a physiologically valid tool to test functional hypotheses on the role of these ancient conserved sequences in comparative cell biology. PMID:17227856

  9. CrusView: A Java-Based Visualization Platform for Comparative Genomics Analyses in Brassicaceae Species[OPEN

    Science.gov (United States)

    Chen, Hao; Wang, Xiangfeng

    2013-01-01

    In plants and animals, chromosomal breakage and fusion events based on conserved syntenic genomic blocks lead to conserved patterns of karyotype evolution among species of the same family. However, karyotype information has not been well utilized in genomic comparison studies. We present CrusView, a Java-based bioinformatic application utilizing Standard Widget Toolkit/Swing graphics libraries and a SQLite database for performing visualized analyses of comparative genomics data in Brassicaceae (crucifer) plants. Compared with similar software and databases, one of the unique features of CrusView is its integration of karyotype information when comparing two genomes. This feature allows users to perform karyotype-based genome assembly and karyotype-assisted genome synteny analyses with preset karyotype patterns of the Brassicaceae genomes. Additionally, CrusView is a local program, which gives its users high flexibility when analyzing unpublished genomes and allows users to upload self-defined genomic information so that they can visually study the associations between genome structural variations and genetic elements, including chromosomal rearrangements, genomic macrosynteny, gene families, high-frequency recombination sites, and tandem and segmental duplications between related species. This tool will greatly facilitate karyotype, chromosome, and genome evolution studies using visualized comparative genomics approaches in Brassicaceae species. CrusView is freely available at http://www.cmbb.arizona.edu/CrusView/. PMID:23898041

  10. Identify alternative splicing events based on position-specific evolutionary conservation.

    Directory of Open Access Journals (Sweden)

    Liang Chen

    Full Text Available The evolution of eukaryotes is accompanied by the increased complexity of alternative splicing which greatly expands genome information. One of the greatest challenges in the post-genome era is a complete revelation of human transcriptome with consideration of alternative splicing. Here, we introduce a comparative genomics approach to systemically identify alternative splicing events based on the differential evolutionary conservation between exons and introns and the high-quality annotation of the ENCODE regions. Specifically, we focus on exons that are included in some transcripts but are completely spliced out for others and we call them conditional exons. First, we characterize distinguishing features among conditional exons, constitutive exons and introns. One of the most important features is the position-specific conservation score. There are dramatic differences in conservation scores between conditional exons and constitutive exons. More importantly, the differences are position-specific. For flanking intronic regions, the differences between conditional exons and constitutive exons are also position-specific. Using the Random Forests algorithm, we can classify conditional exons with high specificities (97% for the identification of conditional exons from intron regions and 95% for the classification of known exons and fair sensitivities (64% and 32% respectively. We applied the method to the human genome and identified 39,640 introns that actually contain conditional exons and classified 8,813 conditional exons from the current RefSeq exon list. Among those, 31,673 introns containing conditional exons and 5,294 conditional exons classified from known exons cannot be inferred from RefSeq, UCSC or Ensembl annotations. Some of these de novo predictions were experimentally verified.

  11. Complete Genome Analysis of Thermus parvatiensis and Comparative Genomics of Thermus spp. Provide Insights into Genetic Variability and Evolution of Natural Competence as Strategic Survival Attributes

    Directory of Open Access Journals (Sweden)

    Charu Tripathi

    2017-07-01

    Full Text Available Thermophilic environments represent an interesting niche. Among thermophiles, the genus Thermus is among the most studied genera. In this study, we have sequenced the genome of Thermus parvatiensis strain RL, a thermophile isolated from Himalayan hot water springs (temperature >96°C using PacBio RSII SMRT technique. The small genome (2.01 Mbp comprises a chromosome (1.87 Mbp and a plasmid (143 Kbp, designated in this study as pTP143. Annotation revealed a high number of repair genes, a squeezed genome but containing highly plastic plasmid with transposases, integrases, mobile elements and hypothetical proteins (44%. We performed a comparative genomic study of the group Thermus with an aim of analysing the phylogenetic relatedness as well as niche specific attributes prevalent among the group. We compared the reference genome RL with 16 Thermus genomes to assess their phylogenetic relationships based on 16S rRNA gene sequences, average nucleotide identity (ANI, conserved marker genes (31 and 400, pan genome and tetranucleotide frequency. The core genome of the analyzed genomes contained 1,177 core genes and many singleton genes were detected in individual genomes, reflecting a conserved core but adaptive pan repertoire. We demonstrated the presence of metagenomic islands (chromosome:5, plasmid:5 by recruiting raw metagenomic data (from the same niche against the genomic replicons of T. parvatiensis. We also dissected the CRISPR loci wide all genomes and found widespread presence of this system across Thermus genomes. Additionally, we performed a comparative analysis of competence loci wide Thermus genomes and found evidence for recent horizontal acquisition of the locus and continued dispersal among members reflecting that natural competence is a beneficial survival trait among Thermus members and its acquisition depicts unending evolution in order to accomplish optimal fitness.

  12. Bat Biology, Genomes, and the Bat1K Project: To Generate Chromosome-Level Genomes for All Living Bat Species.

    Science.gov (United States)

    Teeling, Emma C; Vernes, Sonja C; Dávalos, Liliana M; Ray, David A; Gilbert, M Thomas P; Myers, Eugene

    2018-02-15

    Bats are unique among mammals, possessing some of the rarest mammalian adaptations, including true self-powered flight, laryngeal echolocation, exceptional longevity, unique immunity, contracted genomes, and vocal learning. They provide key ecosystem services, pollinating tropical plants, dispersing seeds, and controlling insect pest populations, thus driving healthy ecosystems. They account for more than 20% of all living mammalian diversity, and their crown-group evolutionary history dates back to the Eocene. Despite their great numbers and diversity, many species are threatened and endangered. Here we announce Bat1K, an initiative to sequence the genomes of all living bat species (n∼1,300) to chromosome-level assembly. The Bat1K genome consortium unites bat biologists (>148 members as of writing), computational scientists, conservation organizations, genome technologists, and any interested individuals committed to a better understanding of the genetic and evolutionary mechanisms that underlie the unique adaptations of bats. Our aim is to catalog the unique genetic diversity present in all living bats to better understand the molecular basis of their unique adaptations; uncover their evolutionary history; link genotype with phenotype; and ultimately better understand, promote, and conserve bats. Here we review the unique adaptations of bats and highlight how chromosome-level genome assemblies can uncover the molecular basis of these traits. We present a novel sequencing and assembly strategy and review the striking societal and scientific benefits that will result from the Bat1K initiative.

  13. Effects of using coding potential, sequence conservation and mRNA structure conservation for predicting pyrroly-sine containing genes

    DEFF Research Database (Denmark)

    Have, Christian Theil; Zambach, Sine; Christiansen, Henning

    2013-01-01

    for prediction of pyrrolysine incorporating genes in genomes of bacteria and archaea leading to insights about the factors driving pyrrolysine translation and identification of new gene candidates. The method predicts known conserved genes with high recall and predicts several other promising candidates...... for experimental verification. The method is implemented as a computational pipeline which is available on request....

  14. Evolutionary analysis of a large mtDNA translocation (numt) into the nuclear genome of the Panthera genus species.

    Science.gov (United States)

    Kim, Jae-Heup; Antunes, Agostinho; Luo, Shu-Jin; Menninger, Joan; Nash, William G; O'Brien, Stephen J; Johnson, Warren E

    2006-02-01

    Translocation of cymtDNA into the nuclear genome, also referred to as numt, has been reported in many species, including several closely related to the domestic cat (Felis catus). We describe the recent transposition of 12,536 bp of the 17 kb mitochondrial genome into the nucleus of the common ancestor of the five Panthera genus species: tiger, P. tigris; snow leopard, P. uncia; jaguar, P. onca; leopard, P. pardus; and lion, P. leo. This nuclear integration, representing 74% of the mitochondrial genome, is one of the largest to be reported in eukaryotes. The Panthera genus numt differs from the numt previously described in the Felis genus in: (1) chromosomal location (F2-telomeric region vs. D2-centromeric region), (2) gene make up (from the ND5 to the ATP8 vs. from the CR to the COII), (3) size (12.5 vs. 7.9 kb), and (4) structure (single monomer vs. tandemly repeated in Felis). These distinctions indicate that the origin of this large numt fragment in the nuclear genome of the Panthera species is an independent insertion from that of the domestic cat lineage, which has been further supported by phylogenetic analyses. The tiger cymtDNA shared around 90% sequence identity with the homologous numt sequence, suggesting an origin for the Panthera numt at around 3.5 million years ago, prior to the radiation of the five extant Panthera species.

  15. Dynamics of chromosome number and genome size variation in a cytogenetically variable sedge (Carex scoparia var. scoparia, Cyperaceae).

    Science.gov (United States)

    Chung, Kyong-Sook; Weber, Jaime A; Hipp, Andrew L

    2011-01-01

    High intraspecific cytogenetic variation in the sedge genus Carex (Cyperaceae) is hypothesized to be due to the "diffuse" or non-localized centromeres, which facilitate chromosome fission and fusion. If chromosome number changes are dominated by fission and fusion, then chromosome evolution will result primarily in changes in the potential for recombination among populations. Chromosome duplications, on the other hand, entail consequent opportunities for divergent evolution of paralogs. In this study, we evaluate whether genome size and chromosome number covary within species. We used flow cytometry to estimate genome sizes in Carex scoparia var. scoparia, sampling 99 plants (23 populations) in the Chicago region, and we used meiotic chromosome observations to document chromosome numbers and chromosome pairing relations. Chromosome numbers range from 2n = 62 to 2n = 68, and nuclear DNA 1C content from 0.342 to 0.361 pg DNA. Regressions of DNA content on chromosome number are nonsignificant for data analyzed by individual or population, and a regression model that excludes slope is favored over a model in which chromosome number predicts genome size. Chromosome rearrangements within cytogenetically variable Carex species are more likely a consequence of fission and fusion than of duplication and deletion. Moreover, neither genome size nor chromosome number is spatially autocorrelated, which suggests the potential for rapid chromosome evolution by fission and fusion at a relatively fine geographic scale (<350 km). These findings have important implications for ecological restoration and speciation within the largest angiosperm genus of the temperate zone.

  16. The Nuclear and Mitochondrial Genomes of the Facultatively Eusocial Orchid Bee Euglossa dilemma

    Directory of Open Access Journals (Sweden)

    Philipp Brand

    2017-09-01

    Full Text Available Bees provide indispensable pollination services to both agricultural crops and wild plant populations, and several species of bees have become important models for the study of learning and memory, plant–insect interactions, and social behavior. Orchid bees (Apidae: Euglossini are especially important to the fields of pollination ecology, evolution, and species conservation. Here we report the nuclear and mitochondrial genome sequences of the orchid bee Euglossa dilemma Bembé & Eltz. E. dilemma was selected because it is widely distributed, highly abundant, and it was recently naturalized in the southeastern United States. We provide a high-quality assembly of the 3.3 Gb genome, and an official gene set of 15,904 gene annotations. We find high conservation of gene synteny with the honey bee throughout 80 MY of divergence time. This genomic resource represents the first draft genome of the orchid bee genus Euglossa, and the first draft orchid bee mitochondrial genome, thus representing a valuable resource to the research community.

  17. The genome of Eucalyptus grandis

    Energy Technology Data Exchange (ETDEWEB)

    Myburg, Alexander A.; Grattapaglia, Dario; Tuskan, Gerald A.; Hellsten, Uffe; Hayes, Richard D.; Grimwood, Jane; Jenkins, Jerry; Lindquist, Erika; Tice, Hope; Bauer, Diane; Goodstein, David M.; Dubchak, Inna; Poliakov, Alexandre; Mizrachi, Eshchar; Kullan, Anand R. K.; Hussey, Steven G.; Pinard, Desre; van der Merwe, Karen; Singh, Pooja; van Jaarsveld, Ida; Silva-Junior, Orzenil B.; Togawa, Roberto C.; Pappas, Marilia R.; Faria, Danielle A.; Sansaloni, Carolina P.; Petroli, Cesar D.; Yang, Xiaohan; Ranjan, Priya; Tschaplinski, Timothy J.; Ye, Chu-Yu; Li, Ting; Sterck, Lieven; Vanneste, Kevin; Murat, Florent; Soler, Marçal; Clemente, Hélène San; Saidi, Naijib; Cassan-Wang, Hua; Dunand, Christophe; Hefer, Charles A.; Bornberg-Bauer, Erich; Kersting, Anna R.; Vining, Kelly; Amarasinghe, Vindhya; Ranik, Martin; Naithani, Sushma; Elser, Justin; Boyd, Alexander E.; Liston, Aaron; Spatafora, Joseph W.; Dharmwardhana, Palitha; Raja, Rajani; Sullivan, Christopher; Romanel, Elisson; Alves-Ferreira, Marcio; Külheim, Carsten; Foley, William; Carocha, Victor; Paiva, Jorge; Kudrna, David; Brommonschenkel, Sergio H.; Pasquali, Giancarlo; Byrne, Margaret; Rigault, Philippe; Tibbits, Josquin; Spokevicius, Antanas; Jones, Rebecca C.; Steane, Dorothy A.; Vaillancourt, René E.; Potts, Brad M.; Joubert, Fourie; Barry, Kerrie; Pappas, Georgios J.; Strauss, Steven H.; Jaiswal, Pankaj; Grima-Pettenati, Jacqueline; Salse, Jérôme; Van de Peer, Yves; Rokhsar, Daniel S.; Schmutz, Jeremy

    2014-06-11

    Eucalypts are the world s most widely planted hardwood trees. Their broad adaptability, rich species diversity, fast growth and superior multipurpose wood, have made them a global renewable resource of fiber and energy that mitigates human pressures on natural forests. We sequenced and assembled >94% of the 640 Mbp genome of Eucalyptus grandis into its 11 chromosomes. A set of 36,376 protein coding genes were predicted revealing that 34% occur in tandem duplications, the largest proportion found thus far in any plant genome. Eucalypts also show the highest diversity of genes for plant specialized metabolism that act as chemical defence against biotic agents and provide unique pharmaceutical oils. Resequencing of a set of inbred tree genomes revealed regions of strongly conserved heterozygosity, likely hotspots of inbreeding depression. The resequenced genome of the sister species E. globulus underscored the high inter-specific genome colinearity despite substantial genome size variation in the genus. The genome of E. grandis is the first reference for the early diverging Rosid order Myrtales and is placed here basal to the Eurosids. This resource expands knowledge on the unique biology of large woody perennials and provides a powerful tool to accelerate comparative biology, breeding and biotechnology.

  18. A Mitochondrial Genome of Rhyparochromidae (Hemiptera: Heteroptera) and a Comparative Analysis of Related Mitochondrial Genomes.

    Science.gov (United States)

    Li, Teng; Yang, Jie; Li, Yinwan; Cui, Ying; Xie, Qiang; Bu, Wenjun; Hillis, David M

    2016-10-19

    The Rhyparochromidae, the largest family of Lygaeoidea, encompasses more than 1,850 described species, but no mitochondrial genome has been sequenced to date. Here we describe the first mitochondrial genome for Rhyparochromidae: a complete mitochondrial genome of Panaorus albomaculatus (Scott, 1874). This mitochondrial genome is comprised of 16,345 bp, and contains the expected 37 genes and control region. The majority of the control region is made up of a large tandem-repeat region, which has a novel pattern not previously observed in other insects. The tandem-repeats region of P. albomaculatus consists of 53 tandem duplications (including one partial repeat), which is the largest number of tandem repeats among all the known insect mitochondrial genomes. Slipped-strand mispairing during replication is likely to have generated this novel pattern of tandem repeats. Comparative analysis of tRNA gene families in sequenced Pentatomomorpha and Lygaeoidea species shows that the pattern of nucleotide conservation is markedly higher on the J-strand. Phylogenetic reconstruction based on mitochondrial genomes suggests that Rhyparochromidae is not the sister group to all the remaining Lygaeoidea, and supports the monophyly of Lygaeoidea.

  19. Lack of independent prognostic and predictive value of centromere 17 copy number changes in breast cancer patients with known HER2 and TOP2A status

    DEFF Research Database (Denmark)

    Nielsen, Kirsten Vang; Ejlertsen, Bent; Møller, Susanne

    2011-01-01

    hybridization (FISH) with centromere 17 (CEN-17) and TOP2A was performed on 120 normal breast specimens. The diploid CEN-17 copy number was reduced from the expected two signals in whole nuclei to an average of 1.68 signals per nucleus in cut sections of normal breast. Ploidy levels determined in normal breast...

  20. Cinteny: flexible analysis and visualization of synteny and genome rearrangements in multiple organisms

    Directory of Open Access Journals (Sweden)

    Meller Jaroslaw

    2007-03-01

    Full Text Available Abstract Background Identifying syntenic regions, i.e., blocks of genes or other markers with evolutionary conserved order, and quantifying evolutionary relatedness between genomes in terms of chromosomal rearrangements is one of the central goals in comparative genomics. However, the analysis of synteny and the resulting assessment of genome rearrangements are sensitive to the choice of a number of arbitrary parameters that affect the detection of synteny blocks. In particular, the choice of a set of markers and the effect of different aggregation strategies, which enable coarse graining of synteny blocks and exclusion of micro-rearrangements, need to be assessed. Therefore, existing tools and resources that facilitate identification, visualization and analysis of synteny need to be further improved to provide a flexible platform for such analysis, especially in the context of multiple genomes. Results We present a new tool, Cinteny, for fast identification and analysis of synteny with different sets of markers and various levels of coarse graining of syntenic blocks. Using Hannenhalli-Pevzner approach and its extensions, Cinteny also enables interactive determination of evolutionary relationships between genomes in terms of the number of rearrangements (the reversal distance. In particular, Cinteny provides: i integration of synteny browsing with assessment of evolutionary distances for multiple genomes; ii flexibility to adjust the parameters and re-compute the results on-the-fly; iii ability to work with user provided data, such as orthologous genes, sequence tags or other conserved markers. In addition, Cinteny provides many annotated mammalian, invertebrate and fungal genomes that are pre-loaded and available for analysis at http://cinteny.cchmc.org. Conclusion Cinteny allows one to automatically compare multiple genomes and perform sensitivity analysis for synteny block detection and for the subsequent computation of reversal distances

  1. Enhancer Identification through Comparative Genomics

    Energy Technology Data Exchange (ETDEWEB)

    Visel, Axel; Bristow, James; Pennacchio, Len A.

    2006-10-01

    With the availability of genomic sequence from numerousvertebrates, a paradigm shift has occurred in the identification ofdistant-acting gene regulatory elements. In contrast to traditionalgene-centric studies in which investigators randomly scanned genomicfragments that flank genes of interest in functional assays, the modernapproach begins electronically with publicly available comparativesequence datasets that provide investigators with prioritized lists ofputative functional sequences based on their evolutionary conservation.However, although a large number of tools and resources are nowavailable, application of comparative genomic approaches remains far fromtrivial. In particular, it requires users to dynamically consider thespecies and methods for comparison depending on the specific biologicalquestion under investigation. While there is currently no single generalrule to this end, it is clear that when applied appropriately,comparative genomic approaches exponentially increase our power ingenerating biological hypotheses for subsequent experimentaltesting.

  2. Conservation of transcription factor binding events predicts gene expression across species

    Science.gov (United States)

    Hemberg, Martin; Kreiman, Gabriel

    2011-01-01

    Recent technological advances have made it possible to determine the genome-wide binding sites of transcription factors (TFs). Comparisons across species have suggested a relatively low degree of evolutionary conservation of experimentally defined TF binding events (TFBEs). Using binding data for six different TFs in hepatocytes and embryonic stem cells from human and mouse, we demonstrate that evolutionary conservation of TFBEs within orthologous proximal promoters is closely linked to function, defined as expression of the target genes. We show that (i) there is a significantly higher degree of conservation of TFBEs when the target gene is expressed in both species; (ii) there is increased conservation of binding events for groups of TFs compared to individual TFs; and (iii) conserved TFBEs have a greater impact on the expression of their target genes than non-conserved ones. These results link conservation of structural elements (TFBEs) to conservation of function (gene expression) and suggest a higher degree of functional conservation than implied by previous studies. PMID:21622661

  3. Unleashing the genome of Brassica rapa

    Directory of Open Access Journals (Sweden)

    Haibao eTang

    2012-07-01

    Full Text Available The completion and release of the Brassica rapa genome is of great benefit to researchers of the Brassicas, Arabidopsis, and genome evolution. While its lineage is closely related to the model organism Arabidopsis thaliana, the Brassicas experienced a whole genome triplication subsequent to their divergence. This event contemporaneously created three copies of its ancestral genome, which had diploidized through the process of homeologous gene loss known as fractionation. By the fractionation of homeologous gene content and genetic regulatory binding sites, Brassica’s genome is well placed to use comparative genomic techniques to identify syntenic regions, homeologous gene duplications, and putative regulatory sequences. Here, we use the comparative genomics platform CoGe to perform several different genomic analyses with which to study structural changes of its genome and dynamics of various genetic elements. Starting with whole genome comparisons, the Brassica paleohexaploidy is characterized, syntenic regions with Arabidopsis thaliana are identified, and the TOC1 gene in the circadian rhythm pathway from Arabidopsis thaliana is used to find duplicated orthologs in Brassica rapa. These TOC1 genes are further analyzed to identify conserved noncoding sequences that contain cis-acting regulatory elements and promoter sequences previously implicated in circadian rhythmicity. Each 'cookbook style' analysis includes a step-by-step walkthrough with links to CoGe to quickly reproduce each step of the analytical process.

  4. Comparative genomics of human and non-human Listeria monocytogenes sequence type 121 strains.

    Directory of Open Access Journals (Sweden)

    Kathrin Rychli

    Full Text Available The food-borne pathogen Listeria (L. monocytogenes is able to survive for months and even years in food production environments. Strains belonging to sequence type (ST121 are particularly found to be abundant and to persist in food and food production environments. To elucidate genetic determinants characteristic for L. monocytogenes ST121, we sequenced the genomes of 14 ST121 strains and compared them with currently available L. monocytogenes ST121 genomes. In total, we analyzed 70 ST121 genomes deriving from 16 different countries, different years of isolation, and different origins-including food, animal and human ST121 isolates. All ST121 genomes show a high degree of conservation sharing at least 99.7% average nucleotide identity. The main differences between the strains were found in prophage content and prophage conservation. We also detected distinct highly conserved subtypes of prophages inserted at the same genomic locus. While some of the prophages showed more than 99.9% similarity between strains from different sources and years, other prophages showed a higher level of diversity. 81.4% of the strains harbored virtually identical plasmids. 97.1% of the ST121 strains contain a truncated internalin A (inlA gene. Only one of the seven human ST121 isolates encodes a full-length inlA gene, illustrating the need of better understanding their survival and virulence mechanisms.

  5. The human noncoding genome defined by genetic diversity.

    Science.gov (United States)

    di Iulio, Julia; Bartha, Istvan; Wong, Emily H M; Yu, Hung-Chun; Lavrenko, Victor; Yang, Dongchan; Jung, Inkyung; Hicks, Michael A; Shah, Naisha; Kirkness, Ewen F; Fabani, Martin M; Biggs, William H; Ren, Bing; Venter, J Craig; Telenti, Amalio

    2018-03-01

    Understanding the significance of genetic variants in the noncoding genome is emerging as the next challenge in human genomics. We used the power of 11,257 whole-genome sequences and 16,384 heptamers (7-nt motifs) to build a map of sequence constraint for the human species. This build differed substantially from traditional maps of interspecies conservation and identified regulatory elements among the most constrained regions of the genome. Using new Hi-C experimental data, we describe a strong pattern of coordination over 2 Mb where the most constrained regulatory elements associate with the most essential genes. Constrained regions of the noncoding genome are up to 52-fold enriched for known pathogenic variants as compared to unconstrained regions (21-fold when compared to the genome average). This map of sequence constraint across thousands of individuals is an asset to help interpret noncoding elements in the human genome, prioritize variants and reconsider gene units at a larger scale.

  6. Comparative analysis of the full genome sequence of European bat lyssavirus type 1 and type 2 with other lyssaviruses and evidence for a conserved transcription termination and polyadenylation motif in the G-L 3' non-translated region.

    Science.gov (United States)

    Marston, D A; McElhinney, L M; Johnson, N; Müller, T; Conzelmann, K K; Tordo, N; Fooks, A R

    2007-04-01

    We report the first full-length genomic sequences for European bat lyssavirus type-1 (EBLV-1) and type-2 (EBLV-2). The EBLV-1 genomic sequence was derived from a virus isolated from a serotine bat in Hamburg, Germany, in 1968 and the EBLV-2 sequence was derived from a virus isolate from a human case of rabies that occurred in Scotland in 2002. A long-distance PCR strategy was used to amplify the open reading frames (ORFs), followed by standard and modified RACE (rapid amplification of cDNA ends) techniques to amplify the 3' and 5' ends. The lengths of each complete viral genome for EBLV-1 and EBLV-2 were 11 966 and 11 930 base pairs, respectively, and follow the standard rhabdovirus genome organization of five viral proteins. Comparison with other lyssavirus sequences demonstrates variation in degrees of homology, with the genomic termini showing a high degree of complementarity. The nucleoprotein was the most conserved, both intra- and intergenotypically, followed by the polymerase (L), matrix and glyco- proteins, with the phosphoprotein being the most variable. In addition, we have shown that the two EBLVs utilize a conserved transcription termination and polyadenylation (TTP) motif, approximately 50 nt upstream of the L gene start codon. All available lyssavirus sequences to date, with the exception of Pasteur virus (PV) and PV-derived isolates, use the second TTP site. This observation may explain differences in pathogenicity between lyssavirus strains, dependent on the length of the untranslated region, which might affect transcriptional activity and RNA stability.

  7. What It Takes to Be a Pseudomonas aeruginosa? The Core Genome of the Opportunistic Pathogen Updated.

    Directory of Open Access Journals (Sweden)

    Benoît Valot

    Full Text Available Pseudomonas aeruginosa is an opportunistic bacterial pathogen able to thrive in highly diverse ecological niches and to infect compromised patients. Its genome exhibits a mosaic structure composed of a core genome into which accessory genes are inserted en bloc at specific sites. The size and the content of the core genome are open for debate as their estimation depends on the set of genomes considered and the pipeline of gene detection and clustering. Here, we redefined the size and the content of the core genome of P. aeruginosa from fully re-analyzed genomes of 17 reference strains. After the optimization of gene detection and clustering parameters, the core genome was defined at 5,233 orthologs, which represented ~ 88% of the average genome. Extrapolation indicated that our panel was suitable to estimate the core genome that will remain constant even if new genomes are added. The core genome contained resistance determinants to the major antibiotic families as well as most metabolic, respiratory, and virulence genes. Although some virulence genes were accessory, they often related to conserved biological functions. Long-standing prophage elements were subjected to a genetic drift to eventually display a G+C content as higher as that of the core genome. This contrasts with the low G+C content of highly conserved ribosomal genes. The conservation of metabolic and respiratory genes could guarantee the ability of the species to thrive on a variety of carbon sources for energy in aerobiosis and anaerobiosis. Virtually all the strains, of environmental or clinical origin, have the complete toolkit to become resistant to the major antipseudomonal compounds and possess basic pathogenic mechanisms to infect humans. The knowledge of the genes shared by the majority of the P. aeruginosa isolates is a prerequisite for designing effective therapeutics to combat the wide variety of human infections.

  8. Comparative genomic sequence analysis of strawberry and other rosids reveals significant microsynteny

    Directory of Open Access Journals (Sweden)

    Abbott Albert

    2010-06-01

    Full Text Available Abstract Background Fragaria belongs to the Rosaceae, an economically important family that includes a number of important fruit producing genera such as Malus and Prunus. Using genomic sequences from 50 Fragaria fosmids, we have examined the microsynteny between Fragaria and other plant models. Results In more than half of the strawberry fosmids, we found syntenic regions that are conserved in Populus, Vitis, Medicago and/or Arabidopsis with Populus containing the greatest number of syntenic regions with Fragaria. The longest syntenic region was between LG VIII of the poplar genome and the strawberry fosmid 72E18, where seven out of twelve predicted genes were collinear. We also observed an unexpectedly high level of conserved synteny between Fragaria (rosid I and Vitis (basal rosid. One of the strawberry fosmids, 34E24, contained a cluster of R gene analogs (RGAs with NBS and LRR domains. We detected clusters of RGAs with high sequence similarity to those in 34E24 in all the genomes compared. In the phylogenetic tree we have generated, all the NBS-LRR genes grouped together with Arabidopsis CNL-A type NBS-LRR genes. The Fragaria RGA grouped together with those of Vitis and Populus in the phylogenetic tree. Conclusions Our analysis shows considerable microsynteny between Fragaria and other plant genomes such as Populus, Medicago, Vitis, and Arabidopsis to a lesser degree. We also detected a cluster of NBS-LRR type genes that are conserved in all the genomes compared.

  9. Conservation of transcription factor binding events predicts gene expression across species

    OpenAIRE

    Hemberg, Martin; Kreiman, Gabriel

    2011-01-01

    Recent technological advances have made it possible to determine the genome-wide binding sites of transcription factors (TFs). Comparisons across species have suggested a relatively low degree of evolutionary conservation of experimentally defined TF binding events (TFBEs). Using binding data for six different TFs in hepatocytes and embryonic stem cells from human and mouse, we demonstrate that evolutionary conservation of TFBEs within orthologous proximal promoters is closely linked to funct...

  10. Statistical properties of thermodynamically predicted RNA secondary structures in viral genomes

    Science.gov (United States)

    Spanò, M.; Lillo, F.; Miccichè, S.; Mantegna, R. N.

    2008-10-01

    By performing a comprehensive study on 1832 segments of 1212 complete genomes of viruses, we show that in viral genomes the hairpin structures of thermodynamically predicted RNA secondary structures are more abundant than expected under a simple random null hypothesis. The detected hairpin structures of RNA secondary structures are present both in coding and in noncoding regions for the four groups of viruses categorized as dsDNA, dsRNA, ssDNA and ssRNA. For all groups, hairpin structures of RNA secondary structures are detected more frequently than expected for a random null hypothesis in noncoding rather than in coding regions. However, potential RNA secondary structures are also present in coding regions of dsDNA group. In fact, we detect evolutionary conserved RNA secondary structures in conserved coding and noncoding regions of a large set of complete genomes of dsDNA herpesviruses.

  11. Sequence and comparative analysis of the chicken genome provide unique perspectives on vertebrate evolution.

    Science.gov (United States)

    2004-12-09

    We present here a draft genome sequence of the red jungle fowl, Gallus gallus. Because the chicken is a modern descendant of the dinosaurs and the first non-mammalian amniote to have its genome sequenced, the draft sequence of its genome--composed of approximately one billion base pairs of sequence and an estimated 20,000-23,000 genes--provides a new perspective on vertebrate genome evolution, while also improving the annotation of mammalian genomes. For example, the evolutionary distance between chicken and human provides high specificity in detecting functional elements, both non-coding and coding. Notably, many conserved non-coding sequences are far from genes and cannot be assigned to defined functional classes. In coding regions the evolutionary dynamics of protein domains and orthologous groups illustrate processes that distinguish the lineages leading to birds and mammals. The distinctive properties of avian microchromosomes, together with the inferred patterns of conserved synteny, provide additional insights into vertebrate chromosome architecture.

  12. Genome inventory and analysis of nuclear hormone receptors in ...

    Indian Academy of Sciences (India)

    Prakash

    2006-12-20

    Dec 20, 2006 ... progestins, as well as lipids, cholesterol metabolites, and. Genome ... Gene structure analysis shows strong conservation of exon structures among orthologoues. ..... earlier subfamily classification of NRs (Nuclear Receptors.

  13. The Genome of the Western Clawed Frog Xenopus tropicalis

    Energy Technology Data Exchange (ETDEWEB)

    Hellsten, Uffe; Harland, Richard M.; Gilchrist, Michael J.; Hendrix, David; Jurka, Jerzy; Kapitonov, Vladimir; Ovcharenko, Ivan; Putnam, Nicholas H.; Shu, Shengqiang; Taher, Leila; Blitz, Ira L.; Blumberg, Bruce; Dichmann, Darwin S.; Dubchak, Inna; Amaya, Enrique; Detter, John C.; Fletcher, Russell; Gerhard, Daniela S.; Goodstein, David; Graves, Tina; Grigoriev, Igor V.; Grimwood, Jane; Kawashima, Takeshi; Lindquist, Erika; Lucas, Susan M.; Mead, Paul E.; Mitros, Therese; Ogino, Hajime; Ohta, Yuko; Poliakov, Alexander V.; Pollet, Nicolas; Robert, Jacques; Salamov, Asaf; Sater, Amy K.; Schmutz, Jeremy; Terry, Astrid; Vize, Peter D.; Warren, Wesley C.; Wells, Dan; Wills, Andrea; Wilson, Richard K.; Zimmerman, Lyle B.; Zorn, Aaron M.; Grainger, Robert; Grammer, Timothy; Khokha, Mustafa K.; Richardson, Paul M.; Rokhsar, Daniel S.

    2009-10-01

    The western clawed frog Xenopus tropicalis is an important model for vertebrate development that combines experimental advantages of the African clawed frog Xenopus laevis with more tractable genetics. Here we present a draft genome sequence assembly of X. tropicalis. This genome encodes over 20,000 protein-coding genes, including orthologs of at least 1,700 human disease genes. Over a million expressed sequence tags validated the annotation. More than one-third of the genome consists of transposable elements, with unusually prevalent DNA transposons. Like other tetrapods, the genome contains gene deserts enriched for conserved non-coding elements. The genome exhibits remarkable shared synteny with human and chicken over major parts of large chromosomes, broken by lineage-specific chromosome fusions and fissions, mainly in the mammalian lineage.

  14. Allele frequencies of variants in ultra conserved elements identify selective pressure on transcription factor binding.

    Science.gov (United States)

    Silla, Toomas; Kepp, Katrin; Tai, E Shyong; Goh, Liang; Davila, Sonia; Catela Ivkovic, Tina; Calin, George A; Voorhoeve, P Mathijs

    2014-01-01

    Ultra-conserved genes or elements (UCGs/UCEs) in the human genome are extreme examples of conservation. We characterized natural variations in 2884 UCEs and UCGs in two distinct populations; Singaporean Chinese (n = 280) and Italian (n = 501) by using a pooled sample, targeted capture, sequencing approach. We identify, with high confidence, in these regions the abundance of rare SNVs (MAFpower for association studies. By combining our data with 1000 Genome Project data, we show in three independent datasets that prevalent UCE variants (MAF>5%) are more often found in relatively less-conserved nucleotides within UCEs, compared to rare variants. Moreover, prevalent variants are less likely to overlap transcription factor binding site. Using SNPfold we found no significant influence of RNA secondary structure on UCE conservation. All together, these results suggest UCEs are not under selective pressure as a stretch of DNA but are under differential evolutionary pressure on the single nucleotide level.

  15. The complete mitochondrial genome of the medicinal fungus Ganoderma applanatum (Polyporales, Basidiomycota).

    Science.gov (United States)

    Wang, Xin-Cun; Shao, Junjie; Liu, Chang

    2016-07-01

    We have determined the complete nucleotide sequence of the mitochondrial genome of the medicinal fungus Ganoderma applanatum (Pers.) Pat. using the next-generation sequencing technology. The circular molecule is 119,803 bp long with a GC content of 26.66%. Gene prediction revealed genes encoding 15 conserved proteins, 25 tRNAs, the large and small ribosomal RNAs, all genes are located on the same strand except trnW-CCA. Compared with previously sequenced genomes of G. lucidum, G. meredithiae and G. sinense, the order of the protein and rRNA genes is highly conserved; however, the types of tRNA genes are slightly different. The mitochondrial genome of G. applanatum will contribute to the understanding of the phylogeny and evolution of Ganoderma and Ganodermataceae, the group containing many species with high medicinal values.

  16. Dicentric chromosome aberration analysis using giemsa and centromere specific fluorescence in-situ hybridization for biological dosimetry: An inter- and intra-laboratory comparison in Indian laboratories

    International Nuclear Information System (INIS)

    Bhavani, M.; Tamizh Selvan, G.; Kaur, Harpreet; Adhikari, J.S.; Vijayalakshmi, J.; Venkatachalam, P.; Chaudhury, N.K.

    2014-01-01

    To facilitate efficient handling of large samples, an attempt towards networking of laboratories in India for biological dosimetry was carried out. Human peripheral blood samples were exposed to 60 Co γ-radiation for ten different doses (0–5 Gy) at a dose rate of 0.7 and 2 Gy/min. The chromosomal aberrations (CA) were scored in Giemsa-stained and fluorescence in-situ hybridization with centromere-specific probes. No significant difference (p>0.05) was observed in the CA yield for given doses except 4 and 5 Gy, between the laboratories, among the scorers and also staining methods adapted suggest the reliability and validates the inter-lab comparisons exercise for triage applications. - Highlights: • This is the first report from India on Networking for Biological Dosimetry preparedness using dicentric chromosomal (DC) aberration assay. • There is no significant difference in the in vitro dose response curve (Slope, Intercept, Curvature) constructed among the two labs. • No significant variation in the scoring of DC aberrations between the scorers irrespective of labs. • The DC results obtained by the labs from the Giemsa stained metaphase preparations were confirmed with centromere specific-FISH for further reliability and validity

  17. Relaxed selection against accidental binding of transcription factors with conserved chromatin contexts.

    Science.gov (United States)

    Babbitt, G A

    2010-10-15

    The spurious (or nonfunctional) binding of transcription factors (TF) to the wrong locations on DNA presents a formidable challenge to genomes given the relatively low ceiling for sequence complexity within the short lengths of most binding motifs. The high potential for the occurrence of random motifs and subsequent nonfunctional binding of many transcription factors should theoretically lead to natural selection against the occurrence of spurious motif throughout the genome. However, because of the active role that chromatin can influence over eukaryotic gene regulation, it may also be expected that many supposed spurious binding sites could escape purifying selection if (A) they simply occur in regions of high nucleosome occupancy or (B) their surrounding chromatin was dynamically involved in their identity and function. We compared nucleosome occupancy and the presence/absence of functionally conserved chromatin context to the strength of selection against spurious binding of various TF binding motifs in Saccharomyces yeast. While we find no direct relationship with nucleosome occupancy, we find strong evidence that transcription factors spatially associated with evolutionarily conserved chromatin states are under relaxed selection against accidental binding. Transcription factors (with/without) a conserved chromatin context were found to occur on average, (87.7%/49.3%) of their expected frequencies. Functional binding motifs with conserved chromatin contexts were also significantly shorter in length and more often clustered. These results indicate a role of chromatin context dependency in relaxing selection against spurious binding in nearly half of all TF binding motifs throughout the yeast genome. 2010 Elsevier B.V. All rights reserved.

  18. Elucidating the composition and conservation of the autophagy pathway in photosynthetic eukaryotes

    Science.gov (United States)

    Shemi, Adva; Ben-Dor, Shifra; Vardi, Assaf

    2015-01-01

    Aquatic photosynthetic eukaryotes represent highly diverse groups (green, red, and chromalveolate algae) derived from multiple endosymbiosis events, covering a wide spectrum of the tree of life. They are responsible for about 50% of the global photosynthesis and serve as the foundation for oceanic and fresh water food webs. Although the ecophysiology and molecular ecology of some algal species are extensively studied, some basic aspects of algal cell biology are still underexplored. The recent wealth of genomic resources from algae has opened new frontiers to decipher the role of cell signaling pathways and their function in an ecological and biotechnological context. Here, we took a bioinformatic approach to explore the distribution and conservation of TOR and autophagy-related (ATG) proteins (Atg in yeast) in diverse algal groups. Our genomic analysis demonstrates conservation of TOR and ATG proteins in green algae. In contrast, in all 5 available red algal genomes, we could not detect the sequences that encode for any of the 17 core ATG proteins examined, albeit TOR and its interacting proteins are conserved. This intriguing data suggests that the autophagy pathway is not conserved in red algae as it is in the entire eukaryote domain. In contrast, chromalveolates, despite being derived from the red-plastid lineage, retain and express ATG genes, which raises a fundamental question regarding the acquisition of ATG genes during algal evolution. Among chromalveolates, Emiliania huxleyi (Haptophyta), a bloom-forming coccolithophore, possesses the most complete set of ATG genes, and may serve as a model organism to study autophagy in marine protists with great ecological significance. PMID:25915714

  19. Whole Genome Epidemiological Typing of Salmonella

    DEFF Research Database (Denmark)

    Leekitcharoenphon, Pimlapas

    available Salmonella enterica genomes (accessed in April 2011). A consensus tree based on variation of the core genes gives better resolution than 16S rRNA and MLST that rarely provide separation between closely related strains. The performance of the pan-genome tree which is based on the presence....../absence of all genes across genomes, is similar to the consensus tree but with higher branching confidence value. The core genes can be divided into two categories: a few highly variable genes and a larger set of conserved core genes, with low variance. These core genes are useful for investigating molecular...... evolution and remain useful as candidate genes for bacterial genome typing-even if they cannot be expected to differentiate highly clonal isolates e.g. outbreak cases of Salmonella [I]. To achieve successful ‘real-time’ monitoring and identification of outbreaks, rapid and reliable sub-typing is essential...

  20. A novel program to design siRNAs simultaneously effective to highly variable virus genomes.

    Science.gov (United States)

    Lee, Hui Sun; Ahn, Jeonghyun; Jun, Eun Jung; Yang, Sanghwa; Joo, Chul Hyun; Kim, Yoo Kyum; Lee, Heuiran

    2009-07-10

    A major concern of antiviral therapy using small interfering RNAs (siRNAs) targeting RNA viral genome is high sequence diversity and mutation rate due to genetic instability. To overcome this problem, it is indispensable to design siRNAs targeting highly conserved regions. We thus designed CAPSID (Convenient Application Program for siRNA Design), a novel bioinformatics program to identify siRNAs targeting highly conserved regions within RNA viral genomes. From a set of input RNAs of diverse sequences, CAPSID rapidly searches conserved patterns and suggests highly potent siRNA candidates in a hierarchical manner. To validate the usefulness of this novel program, we investigated the antiviral potency of universal siRNA for various Human enterovirus B (HEB) serotypes. Assessment of antiviral efficacy using Hela cells, clearly demonstrates that HEB-specific siRNAs exhibit protective effects against all HEBs examined. These findings strongly indicate that CAPSID can be applied to select universal antiviral siRNAs against highly divergent viral genomes.

  1. Comparative analysis of rosaceous genomes and the reconstruction of a putative ancestral genome for the family.

    Science.gov (United States)

    Illa, Eudald; Sargent, Daniel J; Lopez Girona, Elena; Bushakra, Jill; Cestaro, Alessandro; Crowhurst, Ross; Pindo, Massimo; Cabrera, Antonio; van der Knaap, Esther; Iezzoni, Amy; Gardiner, Susan; Velasco, Riccardo; Arús, Pere; Chagné, David; Troggio, Michela

    2011-01-12

    Comparative genome mapping studies in Rosaceae have been conducted until now by aligning genetic maps within the same genus, or closely related genera and using a limited number of common markers. The growing body of genomics resources and sequence data for both Prunus and Fragaria permits detailed comparisons between these genera and the recently released Malus × domestica genome sequence. We generated a comparative analysis using 806 molecular markers that are anchored genetically to the Prunus and/or Fragaria reference maps, and physically to the Malus genome sequence. Markers in common for Malus and Prunus, and Malus and Fragaria, respectively were 784 and 148. The correspondence between marker positions was high and conserved syntenic blocks were identified among the three genera in the Rosaceae. We reconstructed a proposed ancestral genome for the Rosaceae. A genome containing nine chromosomes is the most likely candidate for the ancestral Rosaceae progenitor. The number of chromosomal translocations observed between the three genera investigated was low. However, the number of inversions identified among Malus and Prunus was much higher than any reported genome comparisons in plants, suggesting that small inversions have played an important role in the evolution of these two genera or of the Rosaceae.

  2. Comparative analysis of rosaceous genomes and the reconstruction of a putative ancestral genome for the family

    Directory of Open Access Journals (Sweden)

    Velasco Riccardo

    2011-01-01

    Full Text Available Abstract Background Comparative genome mapping studies in Rosaceae have been conducted until now by aligning genetic maps within the same genus, or closely related genera and using a limited number of common markers. The growing body of genomics resources and sequence data for both Prunus and Fragaria permits detailed comparisons between these genera and the recently released Malus × domestica genome sequence. Results We generated a comparative analysis using 806 molecular markers that are anchored genetically to the Prunus and/or Fragaria reference maps, and physically to the Malus genome sequence. Markers in common for Malus and Prunus, and Malus and Fragaria, respectively were 784 and 148. The correspondence between marker positions was high and conserved syntenic blocks were identified among the three genera in the Rosaceae. We reconstructed a proposed ancestral genome for the Rosaceae. Conclusions A genome containing nine chromosomes is the most likely candidate for the ancestral Rosaceae progenitor. The number of chromosomal translocations observed between the three genera investigated was low. However, the number of inversions identified among Malus and Prunus was much higher than any reported genome comparisons in plants, suggesting that small inversions have played an important role in the evolution of these two genera or of the Rosaceae.

  3. The human homolog of S. cerevisiae CDC27, CDC27 Hs, is encoded by a highly conserved intronless gene present in multiple copies in the human genome

    Energy Technology Data Exchange (ETDEWEB)

    Devor, E.J.; Dill-Devor, R.M. [Univ. of Iowa College of Medicine, Iowa City (United States)

    1994-09-01

    We have obtained a number of unique sequences via PCR amplification of human genomic DNA using degenerate primers under low stringency (42{degrees}C). One of these, an 853 bp product, has been identified as a partial genomic sequence of the human homolog of the S. cerevisiae CDC27 gene, CDC27Hs (GenBank No. U00001). This gene, reported by Turgendreich et al. is also designated EST00556 from Adams et al. We have undertaken a more detailed examination of our sequence, MCP34N, and have found that: 1. the genomic sequence is nearly identical to CDC27Hs over its entire 853 bp length; 2. an MCP34N-specific PCR assay of several non-human primate species reveals amplification products in chimpanzee and gorilla genomes having greater than 90% sequence identity with CDC27Hs; and 3. an MCP34N-specific PCR assay of the BIOS hybrid cell line panel gives a discordancy pattern suggesting multiple loci. Based upon these data, we present the following initial characterization: 1. the complete MCP34N sequence identity with CDC27Hs indicates that the latter is encoded by an intronless gene; 2. CDC27Hs is highly conserved among higher primates; and 3. CDC27Hs is present in multiple copies in the human genome. These characteristics, taken together with those initially reported for CDC27Hs, suggest that this is an old gene that carries out an important but, as yet, unknown function in the human brain.

  4. Identification of the conserved hypothetical protein BPSL0317 in Burkholderia pseudomallei K96243

    Science.gov (United States)

    Yusoff, Nur Syamimi; Damiri, Nadzirah; Firdaus-Raih, Mohd

    2014-09-01

    Burkholderia pseudomallei K96243 is the causative agent of melioidosis, a disease which is endemic in Northern Australia and Southeastern Asia. The genome encodes several essential proteins including those currently annotated as hypothetical proteins. We studied the conservation and the essentiality of expressed hypothetical proteins in normal and different stress conditions. Based on the comparative genomics, we identified a hypothetical protein, BPSL0317, a potential essential gene that is being expressed in all normal and stress conditions. BPSL0317 is also phylogenetically conserved in the Burkholderiales order suggesting that this protein is crucial for survival among the order's members. BPSL0317 therefore has a potential to be a candidate antimicrobial drug target for this group of bacteria.

  5. M-GCAT: interactively and efficiently constructing large-scale multiple genome comparison frameworks in closely related species

    Directory of Open Access Journals (Sweden)

    Messeguer Xavier

    2006-10-01

    Full Text Available Abstract Background Due to recent advances in whole genome shotgun sequencing and assembly technologies, the financial cost of decoding an organism's DNA has been drastically reduced, resulting in a recent explosion of genomic sequencing projects. This increase in related genomic data will allow for in depth studies of evolution in closely related species through multiple whole genome comparisons. Results To facilitate such comparisons, we present an interactive multiple genome comparison and alignment tool, M-GCAT, that can efficiently construct multiple genome comparison frameworks in closely related species. M-GCAT is able to compare and identify highly conserved regions in up to 20 closely related bacterial species in minutes on a standard computer, and as many as 90 (containing 75 cloned genomes from a set of 15 published enterobacterial genomes in an hour. M-GCAT also incorporates a novel comparative genomics data visualization interface allowing the user to globally and locally examine and inspect the conserved regions and gene annotations. Conclusion M-GCAT is an interactive comparative genomics tool well suited for quickly generating multiple genome comparisons frameworks and alignments among closely related species. M-GCAT is freely available for download for academic and non-commercial use at: http://alggen.lsi.upc.es/recerca/align/mgcat/intro-mgcat.html.

  6. Theory of microbial genome evolution

    Science.gov (United States)

    Koonin, Eugene

    Bacteria and archaea have small genomes tightly packed with protein-coding genes. This compactness is commonly perceived as evidence of adaptive genome streamlining caused by strong purifying selection in large microbial populations. In such populations, even the small cost incurred by nonfunctional DNA because of extra energy and time expenditure is thought to be sufficient for this extra genetic material to be eliminated by selection. However, contrary to the predictions of this model, there exists a consistent, positive correlation between the strength of selection at the protein sequence level, measured as the ratio of nonsynonymous to synonymous substitution rates, and microbial genome size. By fitting the genome size distributions in multiple groups of prokaryotes to predictions of mathematical models of population evolution, we show that only models in which acquisition of additional genes is, on average, slightly beneficial yield a good fit to genomic data. Thus, the number of genes in prokaryotic genomes seems to reflect the equilibrium between the benefit of additional genes that diminishes as the genome grows and deletion bias. New genes acquired by microbial genomes, on average, appear to be adaptive. Evolution of bacterial and archaeal genomes involves extensive horizontal gene transfer and gene loss. Many microbes have open pangenomes, where each newly sequenced genome contains more than 10% `ORFans', genes without detectable homologues in other species. A simple, steady-state evolutionary model reveals two sharply distinct classes of microbial genes, one of which (ORFans) is characterized by effectively instantaneous gene replacement, whereas the other consists of genes with finite, distributed replacement rates. These findings imply a conservative estimate of at least a billion distinct genes in the prokaryotic genomic universe.

  7. Broad genomic and transcriptional analysis reveals a highly derived genome in dinoflagellate mitochondria

    Directory of Open Access Journals (Sweden)

    Keeling Patrick J

    2007-09-01

    Full Text Available Abstract Background Dinoflagellates comprise an ecologically significant and diverse eukaryotic phylum that is sister to the phylum containing apicomplexan endoparasites. The mitochondrial genome of apicomplexans is uniquely reduced in gene content and size, encoding only three proteins and two ribosomal RNAs (rRNAs within a highly compacted 6 kb DNA. Dinoflagellate mitochondrial genomes have been comparatively poorly studied: limited available data suggest some similarities with apicomplexan mitochondrial genomes but an even more radical type of genomic organization. Here, we investigate structure, content and expression of dinoflagellate mitochondrial genomes. Results From two dinoflagellates, Crypthecodinium cohnii and Karlodinium micrum, we generated over 42 kb of mitochondrial genomic data that indicate a reduced gene content paralleling that of mitochondrial genomes in apicomplexans, i.e., only three protein-encoding genes and at least eight conserved components of the highly fragmented large and small subunit rRNAs. Unlike in apicomplexans, dinoflagellate mitochondrial genes occur in multiple copies, often as gene fragments, and in numerous genomic contexts. Analysis of cDNAs suggests several novel aspects of dinoflagellate mitochondrial gene expression. Polycistronic transcripts were found, standard start codons are absent, and oligoadenylation occurs upstream of stop codons, resulting in the absence of termination codons. Transcripts of at least one gene, cox3, are apparently trans-spliced to generate full-length mRNAs. RNA substitutional editing, a process previously identified for mRNAs in dinoflagellate mitochondria, is also implicated in rRNA expression. Conclusion The dinoflagellate mitochondrial genome shares the same gene complement and fragmentation of rRNA genes with its apicomplexan counterpart. However, it also exhibits several unique characteristics. Most notable are the expansion of gene copy numbers and their arrangements

  8. The Trichoplax Genome and the Nature of Placozoans

    Energy Technology Data Exchange (ETDEWEB)

    Srivastava, Mansi; Begovic, Emina; Chapman, Jarrod; Putnam, Nicholas H.; Hellsten, Uffe; Kawashima, Takeshi; Kuo, Alan; Mitros, Therese; Salamov, Asaf; Carpenter, Meredith L.; Signorovitch, Ana Y.; Moreno, Maria A.; Kamm, Kai; Grimwood, Jane; Schmutz, Jeremy; Shapiro, Harris; Grigoriev, Igor V.; Buss, Leo W.; Schierwater, Bernd; Dellaporta, Stephen L.; Rokhsar, Daniel S.

    2008-08-01

    Placozoans are arguably the simplest free-living animals, possibly evoking an early stage in metazoan evolution, yet their biology is poorly understood. Here we report the sequencing and analysis of the {approx}98 million base pair nuclear genome of the placozoan Trichoplax adhaerens. Whole genome phylogenetic analysis suggests that placozoans belong to a 'eumetazoan' clade that includes cnidarians and bilaterians, with sponges as the earliest diverging animals. The compact genome exhibits conserved gene content, gene structure, and synteny relative to the human and other complex eumetazoan genomes. Despite the apparent cellular and organismal simplicity of Trichoplax, its genome encodes a rich array of transcription factor and signaling pathway genes that are typically associated with diverse cell types and developmental processes in eumetazoans, motivating further searches for cryptic cellular complexity and/or as yet unobserved life history stages.

  9. Improved de novo genomic assembly for the domestic donkey

    DEFF Research Database (Denmark)

    Renaud, Gabriel; Petersen, Bent; Seguin-Orlando, Andaine

    2018-01-01

    Donkeys and horses share a common ancestor dating back to about 4 million years ago. Although a high-quality genome assembly at the chromosomal level is available for the horse, current assemblies available for the donkey are limited to moderately sized scaffolds. The absence of a better......-quality assembly for the donkey has hampered studies involving the characterization of patterns of genetic variation at the genome-wide scale. These range from the application of genomic tools to selective breeding and conservation to the more fundamental characterization of the genomic loci underlying speciation...... and domestication. We present a new high-quality donkey genome assembly obtained using the Chicago HiRise assembly technology, providing scaffolds of subchromosomal size. We make use of this new assembly to obtain more accurate measures of heterozygosity for equine species other than the horse, both genome...

  10. Conservation in Mammals of Genes Associated with Aggression-Related Behavioral Phenotypes in Honey Bees.

    Directory of Open Access Journals (Sweden)

    Hui Liu

    2016-06-01

    Full Text Available The emerging field of sociogenomics explores the relations between social behavior and genome structure and function. An important question is the extent to which associations between social behavior and gene expression are conserved among the Metazoa. Prior experimental work in an invertebrate model of social behavior, the honey bee, revealed distinct brain gene expression patterns in African and European honey bees, and within European honey bees with different behavioral phenotypes. The present work is a computational study of these previous findings in which we analyze, by orthology determination, the extent to which genes that are socially regulated in honey bees are conserved across the Metazoa. We found that the differentially expressed gene sets associated with alarm pheromone response, the difference between old and young bees, and the colony influence on soldier bees, are enriched in widely conserved genes, indicating that these differences have genomic bases shared with many other metazoans. By contrast, the sets of differentially expressed genes associated with the differences between African and European forager and guard bees are depleted in widely conserved genes, indicating that the genomic basis for this social behavior is relatively specific to honey bees. For the alarm pheromone response gene set, we found a particularly high degree of conservation with mammals, even though the alarm pheromone itself is bee-specific. Gene Ontology identification of human orthologs to the strongly conserved honey bee genes associated with the alarm pheromone response shows overrepresentation of protein metabolism, regulation of protein complex formation, and protein folding, perhaps associated with remodeling of critical neural circuits in response to alarm pheromone. We hypothesize that such remodeling may be an adaptation of social animals to process and respond appropriately to the complex patterns of conspecific communication essential for

  11. Conservation in Mammals of Genes Associated with Aggression-Related Behavioral Phenotypes in Honey Bees.

    Science.gov (United States)

    Liu, Hui; Robinson, Gene E; Jakobsson, Eric

    2016-06-01

    The emerging field of sociogenomics explores the relations between social behavior and genome structure and function. An important question is the extent to which associations between social behavior and gene expression are conserved among the Metazoa. Prior experimental work in an invertebrate model of social behavior, the honey bee, revealed distinct brain gene expression patterns in African and European honey bees, and within European honey bees with different behavioral phenotypes. The present work is a computational study of these previous findings in which we analyze, by orthology determination, the extent to which genes that are socially regulated in honey bees are conserved across the Metazoa. We found that the differentially expressed gene sets associated with alarm pheromone response, the difference between old and young bees, and the colony influence on soldier bees, are enriched in widely conserved genes, indicating that these differences have genomic bases shared with many other metazoans. By contrast, the sets of differentially expressed genes associated with the differences between African and European forager and guard bees are depleted in widely conserved genes, indicating that the genomic basis for this social behavior is relatively specific to honey bees. For the alarm pheromone response gene set, we found a particularly high degree of conservation with mammals, even though the alarm pheromone itself is bee-specific. Gene Ontology identification of human orthologs to the strongly conserved honey bee genes associated with the alarm pheromone response shows overrepresentation of protein metabolism, regulation of protein complex formation, and protein folding, perhaps associated with remodeling of critical neural circuits in response to alarm pheromone. We hypothesize that such remodeling may be an adaptation of social animals to process and respond appropriately to the complex patterns of conspecific communication essential for social organization.

  12. Whole Genome and Tandem Duplicate Retention facilitated Glucosinolate Pathway Diversification in the Mustard Family.

    NARCIS (Netherlands)

    Hofberger, J.A.; Lyons, E.; Edger, P.P.; Pires, J.C.; Schranz, M.E.

    2013-01-01

    Plants share a common history of successive whole genome duplication (WGD) events retaining genomic patterns of duplicate gene copies (ohnologs) organized in conserved syntenic blocks. Duplication was often proposed to affect the origin of novel traits during evolution. However, genetic evidence

  13. Dcode.org anthology of comparative genomic tools.

    Science.gov (United States)

    Loots, Gabriela G; Ovcharenko, Ivan

    2005-07-01

    Comparative genomics provides the means to demarcate functional regions in anonymous DNA sequences. The successful application of this method to identifying novel genes is currently shifting to deciphering the non-coding encryption of gene regulation across genomes. To facilitate the practical application of comparative sequence analysis to genetics and genomics, we have developed several analytical and visualization tools for the analysis of arbitrary sequences and whole genomes. These tools include two alignment tools, zPicture and Mulan; a phylogenetic shadowing tool, eShadow for identifying lineage- and species-specific functional elements; two evolutionary conserved transcription factor analysis tools, rVista and multiTF; a tool for extracting cis-regulatory modules governing the expression of co-regulated genes, Creme 2.0; and a dynamic portal to multiple vertebrate and invertebrate genome alignments, the ECR Browser. Here, we briefly describe each one of these tools and provide specific examples on their practical applications. All the tools are publicly available at the http://www.dcode.org/ website.

  14. Cytoplasmic ATR Activation Promotes Vaccinia Virus Genome Replication

    Directory of Open Access Journals (Sweden)

    Antonio Postigo

    2017-05-01

    Full Text Available In contrast to most DNA viruses, poxviruses replicate their genomes in the cytoplasm without host involvement. We find that vaccinia virus induces cytoplasmic activation of ATR early during infection, before genome uncoating, which is unexpected because ATR plays a fundamental nuclear role in maintaining host genome integrity. ATR, RPA, INTS7, and Chk1 are recruited to cytoplasmic DNA viral factories, suggesting canonical ATR pathway activation. Consistent with this, pharmacological and RNAi-mediated inhibition of canonical ATR signaling suppresses genome replication. RPA and the sliding clamp PCNA interact with the viral polymerase E9 and are required for DNA replication. Moreover, the ATR activator TOPBP1 promotes genome replication and associates with the viral replisome component H5. Our study suggests that, in contrast to long-held beliefs, vaccinia recruits conserved components of the eukaryote DNA replication and repair machinery to amplify its genome in the host cytoplasm.

  15. Flexibility and symmetry of prokaryotic genome rearrangement reveal lineage-associated core-gene-defined genome organizational frameworks.

    Science.gov (United States)

    Kang, Yu; Gu, Chaohao; Yuan, Lina; Wang, Yue; Zhu, Yanmin; Li, Xinna; Luo, Qibin; Xiao, Jingfa; Jiang, Daquan; Qian, Minping; Ahmed Khan, Aftab; Chen, Fei; Zhang, Zhang; Yu, Jun

    2014-11-25

    The prokaryotic pangenome partitions genes into core and dispensable genes. The order of core genes, albeit assumed to be stable under selection in general, is frequently interrupted by horizontal gene transfer and rearrangement, but how a core-gene-defined genome maintains its stability or flexibility remains to be investigated. Based on data from 30 species, including 425 genomes from six phyla, we grouped core genes into syntenic blocks in the context of a pangenome according to their stability across multiple isolates. A subset of the core genes, often species specific and lineage associated, formed a core-gene-defined genome organizational framework (cGOF). Such cGOFs are either single segmental (one-third of the species analyzed) or multisegmental (the rest). Multisegment cGOFs were further classified into symmetric or asymmetric according to segment orientations toward the origin-terminus axis. The cGOFs in Gram-positive species are exclusively symmetric and often reversible in orientation, as opposed to those of the Gram-negative bacteria, which are all asymmetric and irreversible. Meanwhile, all species showing strong strand-biased gene distribution contain symmetric cGOFs and often specific DnaE (α subunit of DNA polymerase III) isoforms. Furthermore, functional evaluations revealed that cGOF genes are hub associated with regard to cellular activities, and the stability of cGOF provides efficient indexes for scaffold orientation as demonstrated by assembling virtual and empirical genome drafts. cGOFs show species specificity, and the symmetry of multisegmental cGOFs is conserved among taxa and constrained by DNA polymerase-centric strand-biased gene distribution. The definition of species-specific cGOFs provides powerful guidance for genome assembly and other structure-based analysis. Prokaryotic genomes are frequently interrupted by horizontal gene transfer (HGT) and rearrangement. To know whether there is a set of genes not only conserved in position

  16. Isolation and Characterization of the Etheostoma tallapoosae (Teleostei: Percidae CENP-A Gene

    Directory of Open Access Journals (Sweden)

    Leos G. Kral

    2011-10-01

    Full Text Available Both centromeric alpha-satellite sequences as well as centromeric protein A (CENP-A are highly variable in eukaryotes. CENP-A, a histone H3 variant, is thought to act as the epigenetic “mark” for assembly of centromeric proteins. While most of the histone fold domain (HFD of the CENP-A is fairly well conserved, a portion of this HFD as well as the N-terminal tail show adaptive variation in both plants and animals. Such variation may establish reproductive barriers that may lead to speciation. The family Percidae contains over 200 species most of which are within the subfamily Etheostomatinae. This subfamily represents a species rich radiation of freshwater fishes in North America and these species exhibit both allopatric and sympatric distributions. In order to study the evolution of CENP-A in percid fish species, we have isolated and characterized the CENP-A gene from Etheostoma tallapoosae by PCR based gene walking. As a result of this study we have demonstrated that the Tallapoosa darter CENP-A gene HFD sequences can be isolated from genomic DNA by nested PCR in a manner that does not lead to the amplification of the highly sequence related histone H3 gene. We also demonstrated that PCR based walking can be subsequently used to isolate the rest of the CENP-A gene and adjacent gene sequences. These adjacent gene sequences provide us with a primer binding sites for PCR isolation of the CENP-A gene from other percid species of fishes. An initial comparison of three percid species shows that the N-terminal tail of the percid CENP-A gene shows adaptive evolution.

  17. Forces shaping the fastest evolving regions in the human genome.

    Directory of Open Access Journals (Sweden)

    Katherine S Pollard

    2006-10-01

    Full Text Available Comparative genomics allow us to search the human genome for segments that were extensively changed in the last approximately 5 million years since divergence from our common ancestor with chimpanzee, but are highly conserved in other species and thus are likely to be functional. We found 202 genomic elements that are highly conserved in vertebrates but show evidence of significantly accelerated substitution rates in human. These are mostly in non-coding DNA, often near genes associated with transcription and DNA binding. Resequencing confirmed that the five most accelerated elements are dramatically changed in human but not in other primates, with seven times more substitutions in human than in chimp. The accelerated elements, and in particular the top five, show a strong bias for adenine and thymine to guanine and cytosine nucleotide changes and are disproportionately located in high recombination and high guanine and cytosine content environments near telomeres, suggesting either biased gene conversion or isochore selection. In addition, there is some evidence of directional selection in the regions containing the two most accelerated regions. A combination of evolutionary forces has contributed to accelerated evolution of the fastest evolving elements in the human genome.

  18. Preferential transcription of conserved rif genes in two phenotypically distinct Plasmodium falciparum parasite lines

    DEFF Research Database (Denmark)

    Wang, Christian W; Magistrado, Pamela A; Nielsen, Morten A

    2009-01-01

    transcribed in the VAR2CSA-expressing parasite line. In addition, two rif genes were found transcribed at early and late intra-erythrocyte stages independently of var gene transcription. Rif genes are organised in groups and inter-genomic conserved gene families, suggesting that RIFIN sub-groups may have......Plasmodium falciparum variant surface antigens (VSA) are targets of protective immunity to malaria. Plasmodium falciparum erythrocyte membrane protein 1 (PfEMP1) and repetitive interspersed family (RIFIN) proteins are encoded by the two variable multigene families, var and rif genes, respectively...... novel rif gene groups, rifA1 and rifA2, containing inter-genomic conserved rif genes, were identified. All rifA1 genes were orientated head-to-head with a neighbouring Group A var gene whereas rifA2 was present in all parasite genomes as a single copy gene with a unique 5' untranslated region. Rif...

  19. Comparative Genome Analysis of Enterobacter cloacae

    Science.gov (United States)

    Liu, Wing-Yee; Wong, Chi-Fat; Chung, Karl Ming-Kar; Jiang, Jing-Wei; Leung, Frederick Chi-Ching

    2013-01-01

    The Enterobacter cloacae species includes an extremely diverse group of bacteria that are associated with plants, soil and humans. Publication of the complete genome sequence of the plant growth-promoting endophytic E. cloacae subsp. cloacae ENHKU01 provided an opportunity to perform the first comparative genome analysis between strains of this dynamic species. Examination of the pan-genome of E. cloacae showed that the conserved core genome retains the general physiological and survival genes of the species, while genomic factors in plasmids and variable regions determine the virulence of the human pathogenic E. cloacae strain; additionally, the diversity of fimbriae contributes to variation in colonization and host determination of different E. cloacae strains. Comparative genome analysis further illustrated that E. cloacae strains possess multiple mechanisms for antagonistic action against other microorganisms, which involve the production of siderophores and various antimicrobial compounds, such as bacteriocins, chitinases and antibiotic resistance proteins. The presence of Type VI secretion systems is expected to provide further fitness advantages for E. cloacae in microbial competition, thus allowing it to survive in different environments. Competition assays were performed to support our observations in genomic analysis, where E. cloacae subsp. cloacae ENHKU01 demonstrated antagonistic activities against a wide range of plant pathogenic fungal and bacterial species. PMID:24069314

  20. Genomic analyses of the microsporidian Nosema ceranae, an emergent pathogen of honey bees.

    Directory of Open Access Journals (Sweden)

    R Scott Cornman

    2009-06-01

    Full Text Available Recent steep declines in honey bee health have severely impacted the beekeeping industry, presenting new risks for agricultural commodities that depend on insect pollination. Honey bee declines could reflect increased pressures from parasites and pathogens. The incidence of the microsporidian pathogen Nosema ceranae has increased significantly in the past decade. Here we present a draft assembly (7.86 MB of the N. ceranae genome derived from pyrosequence data, including initial gene models and genomic comparisons with other members of this highly derived fungal lineage. N. ceranae has a strongly AT-biased genome (74% A+T and a diversity of repetitive elements, complicating the assembly. Of 2,614 predicted protein-coding sequences, we conservatively estimate that 1,366 have homologs in the microsporidian Encephalitozoon cuniculi, the most closely related published genome sequence. We identify genes conserved among microsporidia that lack clear homology outside this group, which are of special interest as potential virulence factors in this group of obligate parasites. A substantial fraction of the diminutive N. ceranae proteome consists of novel and transposable-element proteins. For a majority of well-supported gene models, a conserved sense-strand motif can be found within 15 bases upstream of the start codon; a previously uncharacterized version of this motif is also present in E. cuniculi. These comparisons provide insight into the architecture, regulation, and evolution of microsporidian genomes, and will drive investigations into honey bee-Nosema interactions.

  1. Puzzling sequences: studying microbial genomes from 'Ötzi'

    International Nuclear Information System (INIS)

    Rattei, T.

    2012-01-01

    Ancient remains, and mummies in particular, are of central value for archaeological research. The Tyrolean iceman “Ötzi” was conserved in a glacier of the Ötztal Alps about 5000 years ago. Aside from morphological and phenotypical classification, the determination of DNA sequences and the subsequent genome analyses have been first applied to mitochondrial DNA and then been extended to genomic DNA. Typically also ancient microbial DNA is sequenced. These sequences allow the identification of pathogens as well as studying the evolution of microorganisms. The talk will explain the metagenomic aspects of the “Ötzi” genome project and discuss the first results. (author)

  2. Lactobacillus paracasei comparative genomics: towards species pan-genome definition and exploitation of diversity.

    Directory of Open Access Journals (Sweden)

    Tamara Smokvina

    Full Text Available Lactobacillus paracasei is a member of the normal human and animal gut microbiota and is used extensively in the food industry in starter cultures for dairy products or as probiotics. With the development of low-cost, high-throughput sequencing techniques it has become feasible to sequence many different strains of one species and to determine its "pan-genome". We have sequenced the genomes of 34 different L. paracasei strains, and performed a comparative genomics analysis. We analysed genome synteny and content, focussing on the pan-genome, core genome and variable genome. Each genome was shown to contain around 2800-3100 protein-coding genes, and comparative analysis identified over 4200 ortholog groups that comprise the pan-genome of this species, of which about 1800 ortholog groups make up the conserved core. Several factors previously associated with host-microbe interactions such as pili, cell-envelope proteinase, hydrolases p40 and p75 or the capacity to produce short branched-chain fatty acids (bkd operon are part of the L. paracasei core genome present in all analysed strains. The variome consists mainly of hypothetical proteins, phages, plasmids, transposon/conjugative elements, and known functions such as sugar metabolism, cell-surface proteins, transporters, CRISPR-associated proteins, and EPS biosynthesis proteins. An enormous variety and variability of sugar utilization gene cassettes were identified, with each strain harbouring between 25-53 cassettes, reflecting the high adaptability of L. paracasei to different niches. A phylogenomic tree was constructed based on total genome contents, and together with an analysis of horizontal gene transfer events we conclude that evolution of these L. paracasei strains is complex and not always related to niche adaptation. The results of this genome content comparison was used, together with high-throughput growth experiments on various carbohydrates, to perform gene-trait matching analysis

  3. Mitochondrial genome sequencing helps show the evolutionary mechanism of mitochondrial genome formation in Brassica

    Science.gov (United States)

    2011-01-01

    Background Angiosperm mitochondrial genomes are more complex than those of other organisms. Analyses of the mitochondrial genome sequences of at least 11 angiosperm species have showed several common properties; these cannot easily explain, however, how the diverse mitotypes evolved within each genus or species. We analyzed the evolutionary relationships of Brassica mitotypes by sequencing. Results We sequenced the mitotypes of cam (Brassica rapa), ole (B. oleracea), jun (B. juncea), and car (B. carinata) and analyzed them together with two previously sequenced mitotypes of B. napus (pol and nap). The sizes of whole single circular genomes of cam, jun, ole, and car are 219,747 bp, 219,766 bp, 360,271 bp, and 232,241 bp, respectively. The mitochondrial genome of ole is largest as a resulting of the duplication of a 141.8 kb segment. The jun mitotype is the result of an inherited cam mitotype, and pol is also derived from the cam mitotype with evolutionary modifications. Genes with known functions are conserved in all mitotypes, but clear variation in open reading frames (ORFs) with unknown functions among the six mitotypes was observed. Sequence relationship analysis showed that there has been genome compaction and inheritance in the course of Brassica mitotype evolution. Conclusions We have sequenced four Brassica mitotypes, compared six Brassica mitotypes and suggested a mechanism for mitochondrial genome formation in Brassica, including evolutionary events such as inheritance, duplication, rearrangement, genome compaction, and mutation. PMID:21988783

  4. Human developmental enhancers conserved between deuterostomes and protostomes.

    Directory of Open Access Journals (Sweden)

    Shoa L Clarke

    Full Text Available The identification of homologies, whether morphological, molecular, or genetic, is fundamental to our understanding of common biological principles. Homologies bridging the great divide between deuterostomes and protostomes have served as the basis for current models of animal evolution and development. It is now appreciated that these two clades share a common developmental toolkit consisting of conserved transcription factors and signaling pathways. These patterning genes sometimes show common expression patterns and genetic interactions, suggesting the existence of similar or even conserved regulatory apparatus. However, previous studies have found no regulatory sequence conserved between deuterostomes and protostomes. Here we describe the first such enhancers, which we call bilaterian conserved regulatory elements (Bicores. Bicores show conservation of sequence and gene synteny. Sequence conservation of Bicores reflects conserved patterns of transcription factor binding sites. We predict that Bicores act as response elements to signaling pathways, and we show that Bicores are developmental enhancers that drive expression of transcriptional repressors in the vertebrate central nervous system. Although the small number of identified Bicores suggests extensive rewiring of cis-regulation between the protostome and deuterostome clades, additional Bicores may be revealed as our understanding of cis-regulatory logic and sample of bilaterian genomes continue to grow.

  5. Genome-wide identification of polycomb target genes reveals a functional association of Pho with Scm in Bombyx mori.

    Science.gov (United States)

    Li, Zhiqing; Cheng, Daojun; Mon, Hiroaki; Tatsuke, Tsuneyuki; Zhu, Li; Xu, Jian; Lee, Jae Man; Xia, Qingyou; Kusakabe, Takahiro

    2012-01-01

    Polycomb group (PcG) proteins are evolutionarily conserved chromatin modifiers and act together in three multimeric complexes, Polycomb repressive complex 1 (PRC1), Polycomb repressive complex 2 (PRC2), and Pleiohomeotic repressive complex (PhoRC), to repress transcription of the target genes. Here, we identified Polycomb target genes in Bombyx mori with holocentric centromere using genome-wide expression screening based on the knockdown of BmSCE, BmESC, BmPHO, or BmSCM gene, which represent the distinct complexes. As a result, the expressions of 29 genes were up-regulated after knocking down 4 PcG genes. Particularly, there is a significant overlap between targets of BmPho (331 out of 524) and BmScm (331 out of 532), and among these, 190 genes function as regulator factors playing important roles in development. We also found that BmPho, as well as BmScm, can interact with other Polycomb components examined in this study. Further detailed analysis revealed that the C-terminus of BmPho containing zinc finger domain is involved in the interaction between BmPho and BmScm. Moreover, the zinc finger domain in BmPho contributes to its inhibitory function and ectopic overexpression of BmScm is able to promote transcriptional repression by Gal4-Pho fusions including BmScm-interacting domain. Loss of BmPho expression causes relocalization of BmScm into the cytoplasm. Collectively, we provide evidence of a functional link between BmPho and BmScm, and propose two Polycomb-related repression mechanisms requiring only BmPho associated with BmScm or a whole set of PcG complexes.

  6. Genome-wide identification of polycomb target genes reveals a functional association of Pho with Scm in Bombyx mori.

    Directory of Open Access Journals (Sweden)

    Zhiqing Li

    Full Text Available Polycomb group (PcG proteins are evolutionarily conserved chromatin modifiers and act together in three multimeric complexes, Polycomb repressive complex 1 (PRC1, Polycomb repressive complex 2 (PRC2, and Pleiohomeotic repressive complex (PhoRC, to repress transcription of the target genes. Here, we identified Polycomb target genes in Bombyx mori with holocentric centromere using genome-wide expression screening based on the knockdown of BmSCE, BmESC, BmPHO, or BmSCM gene, which represent the distinct complexes. As a result, the expressions of 29 genes were up-regulated after knocking down 4 PcG genes. Particularly, there is a significant overlap between targets of BmPho (331 out of 524 and BmScm (331 out of 532, and among these, 190 genes function as regulator factors playing important roles in development. We also found that BmPho, as well as BmScm, can interact with other Polycomb components examined in this study. Further detailed analysis revealed that the C-terminus of BmPho containing zinc finger domain is involved in the interaction between BmPho and BmScm. Moreover, the zinc finger domain in BmPho contributes to its inhibitory function and ectopic overexpression of BmScm is able to promote transcriptional repression by Gal4-Pho fusions including BmScm-interacting domain. Loss of BmPho expression causes relocalization of BmScm into the cytoplasm. Collectively, we provide evidence of a functional link between BmPho and BmScm, and propose two Polycomb-related repression mechanisms requiring only BmPho associated with BmScm or a whole set of PcG complexes.

  7. Mitochondrial genome evolution in the Saccharomyces sensu stricto complex.

    Science.gov (United States)

    Ruan, Jiangxing; Cheng, Jian; Zhang, Tongcun; Jiang, Huifeng

    2017-01-01

    Exploring the evolutionary patterns of mitochondrial genomes is important for our understanding of the Saccharomyces sensu stricto (SSS) group, which is a model system for genomic evolution and ecological analysis. In this study, we first obtained the complete mitochondrial sequences of two important species, Saccharomyces mikatae and Saccharomyces kudriavzevii. We then compared the mitochondrial genomes in the SSS group with those of close relatives, and found that the non-coding regions evolved rapidly, including dramatic expansion of intergenic regions, fast evolution of introns and almost 20-fold higher rearrangement rates than those of the nuclear genomes. However, the coding regions, and especially the protein-coding genes, are more conserved than those in the nuclear genomes of the SSS group. The different evolutionary patterns of coding and non-coding regions in the mitochondrial and nuclear genomes may be related to the origin of the aerobic fermentation lifestyle in this group. Our analysis thus provides novel insights into the evolution of mitochondrial genomes.

  8. pTC Plasmids from Sulfolobus Species in the Geothermal Area of Tengchong, China: Genomic Conservation and Naturally-Occurring Variations as a Result of Transposition by Mobile Genetic Elements.

    Science.gov (United States)

    Xiang, Xiaoyu; Huang, Xiaoxing; Wang, Haina; Huang, Li

    2015-02-12

    Plasmids occur frequently in Archaea. A novel plasmid (denoted pTC1) containing typical conjugation functions has been isolated from Sulfolobus tengchongensis RT8-4, a strain obtained from a hot spring in Tengchong, China, and characterized. The plasmid is a circular double-stranded DNA molecule of 20,417 bp. Among a total of 26 predicted pTC1 ORFs, 23 have homologues in other known Sulfolobus conjugative plasmids (CPs). pTC1 resembles other Sulfolobus CPs in genome architecture, and is most highly conserved in the genomic region encoding conjugation functions. However, attempts to demonstrate experimentally the capacity of the plasmid for conjugational transfer were unsuccessful. A survey revealed that pTC1 and its closely related plasmid variants were widespread in the geothermal area of Tengchong. Variations of the plasmids at the target sites for transposition by an insertion sequence (IS) and a miniature inverted-repeat transposable element (MITE) were readily detected. The IS was efficiently inserted into the pTC1 genome, and the inserted sequence was inactivated and degraded more frequently in an imprecise manner than in a precise manner. These results suggest that the host organism has evolved a strategy to maintain a balance between the insertion and elimination of mobile genetic elements to permit genomic plasticity while inhibiting their fast spreading.

  9. pTC Plasmids from Sulfolobus Species in the Geothermal Area of Tengchong, China: Genomic Conservation and Naturally-Occurring Variations as a Result of Transposition by Mobile Genetic Elements

    Directory of Open Access Journals (Sweden)

    Xiaoyu Xiang

    2015-02-01

    Full Text Available Plasmids occur frequently in Archaea. A novel plasmid (denoted pTC1 containing typical conjugation functions has been isolated from Sulfolobus tengchongensis RT8-4, a strain obtained from a hot spring in Tengchong, China, and characterized. The plasmid is a circular double-stranded DNA molecule of 20,417 bp. Among a total of 26 predicted pTC1 ORFs, 23 have homologues in other known Sulfolobus conjugative plasmids (CPs. pTC1 resembles other Sulfolobus CPs in genome architecture, and is most highly conserved in the genomic region encoding conjugation functions. However, attempts to demonstrate experimentally the capacity of the plasmid for conjugational transfer were unsuccessful. A survey revealed that pTC1 and its closely related plasmid variants were widespread in the geothermal area of Tengchong. Variations of the plasmids at the target sites for transposition by an insertion sequence (IS and a miniature inverted-repeat transposable element (MITE were readily detected. The IS was efficiently inserted into the pTC1 genome, and the inserted sequence was inactivated and degraded more frequently in an imprecise manner than in a precise manner. These results suggest that the host organism has evolved a strategy to maintain a balance between the insertion and elimination of mobile genetic elements to permit genomic plasticity while inhibiting their fast spreading.

  10. Complete mitochondrial genome of the Freshwater Catfish Rita rita (Siluriformes, Bagridae).

    Science.gov (United States)

    Lashari, Punhal; Laghari, Muhammad Younis; Xu, Peng; Zhao, Zixia; Jiang, Li; Narejo, Naeem Tariq; Deng, Yulin; Sun, Xiaowen; Zhang, Yan

    2015-01-01

    The complete mitochondrial genome of Catfish, Rita rita, was isolated by LA PCR (TakaRa LAtaq, Dalian, China); and sequenced by Sanger's method to obtain the complete mitochondrial genome, which is listed Critically Endangered and Red Listed species. The complete mitogenome was 16,449 bp in length and contains 13 typical vertebrate protein-coding genes, 2 rRNA and 22 tRNA genes. The whole genome base composition was estimated to be 33.40% A, 27.43% C, 14.26% G and 24.89% T. The complete mitochondrial genome of catfish, Rita rita provides the basis for genetic breeding and conservation studies.

  11. Arabidopsis transcription factors: genome-wide comparative analysis among eukaryotes.

    Science.gov (United States)

    Riechmann, J L; Heard, J; Martin, G; Reuber, L; Jiang, C; Keddie, J; Adam, L; Pineda, O; Ratcliffe, O J; Samaha, R R; Creelman, R; Pilgrim, M; Broun, P; Zhang, J Z; Ghandehari, D; Sherman, B K; Yu, G

    2000-12-15

    The completion of the Arabidopsis thaliana genome sequence allows a comparative analysis of transcriptional regulators across the three eukaryotic kingdoms. Arabidopsis dedicates over 5% of its genome to code for more than 1500 transcription factors, about 45% of which are from families specific to plants. Arabidopsis transcription factors that belong to families common to all eukaryotes do not share significant similarity with those of the other kingdoms beyond the conserved DNA binding domains, many of which have been arranged in combinations specific to each lineage. The genome-wide comparison reveals the evolutionary generation of diversity in the regulation of transcription.

  12. Organization of plastid genomes in the freshwater red algal order Batrachospermales (Rhodophyta).

    Science.gov (United States)

    Paiano, Monica Orlandi; Del Cortona, Andrea; Costa, Joana F; Liu, Shao-Lun; Verbruggen, Heroen; De Clerck, Olivier; Necchi, Orlando

    2018-02-01

    Little is known about genome organization in members of the order Batrachospermales, and the infra-ordinal relationship remains unresolved. Plastid (cp) genomes of seven members of the freshwater red algal order Batrachospermales were sequenced, with the following aims: (i) to describe the characteristics of cp genomes and compare these with other red algal groups; (ii) to infer the phylogenetic relationships among these members to better understand the infra-ordinal classification. Cp genomes of Batrachospermales are large, with several cases of gene loss, they are gene-dense (high gene content for the genome size and short intergenic regions) and have highly conserved gene order. Phylogenetic analyses based on concatenated nucleotide genome data roughly supports the current taxonomic system for the order. Comparative analyses confirm data for members of the class Florideophyceae that cp genomes in Batrachospermales is highly conserved, with little variation in gene composition. However, relevant new features were revealed in our study: genome sizes in members of Batrachospermales are close to the lowest values reported for Florideophyceae; differences in cp genome size within the order are large in comparison with other orders (Ceramiales, Gelidiales, Gracilariales, Hildenbrandiales, and Nemaliales); and members of Batrachospermales have the lowest number of protein-coding genes among the Florideophyceae. In terms of gene loss, apcF, which encodes the allophycocyanin beta subunit, is absent in all sequenced taxa of Batrachospermales. We reinforce that the interordinal relationships between the freshwater orders Batrachospermales and Thoreales within the Nemaliophycidae is not well resolved due to limited taxon sampling. © 2017 Phycological Society of America.

  13. The impact of genome triplication on tandem gene evolution in Brassica rapa

    Directory of Open Access Journals (Sweden)

    Lu eFang

    2012-11-01

    Full Text Available Whole genome duplication (WGD and tandem duplication (TD are both important modes of gene expansion. However, how whole genome duplication influences tandemly duplicated genes is not well studied. We used Brassica rapa, which has undergone an additional genome triplication (WGT and shares a common ancestor with Arabidopsis thaliana, Arabidopsis lyrata and Thellungiella parvula, to investigate the impact of genome triplication on tandem gene evolution. We identified 2,137, 1,569, 1,751 and 1,135 tandem gene arrays in B. rapa, A. thaliana, A. lyrata and T. parvula respectively. Among them, 414 conserved tandem arrays are shared by the 3 species without WGT, which were also considered as existing in the diploid ancestor of B. rapa. Thus, after genome triplication, B. rapa should have 1,242 tandem arrays according to the 414 conserved tandems. Here, we found 400 out of the 414 tandems had at least one syntenic ortholog in the genome of B. rapa. Furthermore, 294 out of the 400 shared syntenic orthologs maintain tandem arrays (more than one gene for each syntenic hit in B. rapa. For the 294 tandem arrays, we obtained 426 copies of syntenic paralogous tandems in the triplicated genome of B. rapa. In this study, we demonstrated that tandem arrays in B. rapa were dramatically fractionated after WGT when compared either to non-tandem genes in the B. rapa genome or to the tandem arrays in closely related species that have not experienced a recent whole-genome polyploidization event.

  14. Conserved cis-regulatory regions in a large genomic landscape control SHH and BMP-regulated Gremlin1 expression in mouse limb buds

    Directory of Open Access Journals (Sweden)

    Zuniga Aimée

    2012-08-01

    Full Text Available Abstract Background Mouse limb bud is a prime model to study the regulatory interactions that control vertebrate organogenesis. Major aspects of limb bud development are controlled by feedback loops that define a self-regulatory signalling system. The SHH/GREM1/AER-FGF feedback loop forms the core of this signalling system that operates between the posterior mesenchymal organiser and the ectodermal signalling centre. The BMP antagonist Gremlin1 (GREM1 is a critical node in this system, whose dynamic expression is controlled by BMP, SHH, and FGF signalling and key to normal progression of limb bud development. Previous analysis identified a distant cis-regulatory landscape within the neighbouring Formin1 (Fmn1 locus that is required for Grem1 expression, reminiscent of the genomic landscapes controlling HoxD and Shh expression in limb buds. Results Three highly conserved regions (HMCO1-3 were identified within the previously defined critical genomic region and tested for their ability to regulate Grem1 expression in mouse limb buds. Using a combination of BAC and conventional transgenic approaches, a 9 kb region located ~70 kb downstream of the Grem1 transcription unit was identified. This region, termed Grem1 Regulatory Sequence 1 (GRS1, is able to recapitulate major aspects of Grem1 expression, as it drives expression of a LacZ reporter into the posterior and, to a lesser extent, in the distal-anterior mesenchyme. Crossing the GRS1 transgene into embryos with alterations in the SHH and BMP pathways established that GRS1 depends on SHH and is modulated by BMP signalling, i.e. integrates inputs from these pathways. Chromatin immunoprecipitation revealed interaction of endogenous GLI3 proteins with the core cis-regulatory elements in the GRS1 region. As GLI3 is a mediator of SHH signal transduction, these results indicated that SHH directly controls Grem1 expression through the GRS1 region. Finally, all cis-regulatory regions within the Grem1

  15. A decade of human genome project conclusion: Scientific diffusion about our genome knowledge.

    Science.gov (United States)

    Moraes, Fernanda; Góes, Andréa

    2016-05-06

    The Human Genome Project (HGP) was initiated in 1990 and completed in 2003. It aimed to sequence the whole human genome. Although it represented an advance in understanding the human genome and its complexity, many questions remained unanswered. Other projects were launched in order to unravel the mysteries of our genome, including the ENCyclopedia of DNA Elements (ENCODE). This review aims to analyze the evolution of scientific knowledge related to both the HGP and ENCODE projects. Data were retrieved from scientific articles published in 1990-2014, a period comprising the development and the 10 years following the HGP completion. The fact that only 20,000 genes are protein and RNA-coding is one of the most striking HGP results. A new concept about the organization of genome arose. The ENCODE project was initiated in 2003 and targeted to map the functional elements of the human genome. This project revealed that the human genome is pervasively transcribed. Therefore, it was determined that a large part of the non-protein coding regions are functional. Finally, a more sophisticated view of chromatin structure emerged. The mechanistic functioning of the genome has been redrafted, revealing a much more complex picture. Besides, a gene-centric conception of the organism has to be reviewed. A number of criticisms have emerged against the ENCODE project approaches, raising the question of whether non-conserved but biochemically active regions are truly functional. Thus, HGP and ENCODE projects accomplished a great map of the human genome, but the data generated still requires further in depth analysis. © 2016 by The International Union of Biochemistry and Molecular Biology, 44:215-223, 2016. © 2016 The International Union of Biochemistry and Molecular Biology.

  16. In the Beginning was the Genome: Genomics and the Bi-textuality of Human Existence.

    Science.gov (United States)

    Zwart, H A E Hub

    2018-04-01

    This paper addresses the cultural impact of genomics and the Human Genome Project (HGP) on human self-understanding. Notably, it addresses the claim made by Francis Collins (director of the HGP) that the genome is the language of God and the claim made by Max Delbrück (founding father of molecular life sciences research) that Aristotle must be credited with having predicted DNA as the soul that organises bio-matter. From a continental philosophical perspective I will argue that human existence results from a dialectical interaction between two types of texts: the language of molecular biology and the language of civilisation; the language of the genome and the language of our socio-cultural, symbolic ambiance. Whereas the former ultimately builds on the alphabets of genes and nucleotides, the latter is informed by primordial texts such as the Bible and the Quran. In applied bioethics deliberations on genomics, science is easily framed as liberating and progressive, religious world-views as conservative and restrictive (Zwart 1993). This paper focusses on the broader cultural ambiance of the debate to discern how the bi-textuality of human existence is currently undergoing a transition, as not only the physiological, but also the normative dimension is being reframed in biomolecular and terabyte terms.

  17. Transcriptional Response of Human Neurospheres to Helper-Dependent CAV-2 Vectors Involves the Modulation of DNA Damage Response, Microtubule and Centromere Gene Groups.

    Directory of Open Access Journals (Sweden)

    Stefania Piersanti

    Full Text Available Brain gene transfer using viral vectors will likely become a therapeutic option for several disorders. Helper-dependent (HD canine adenovirus type 2 vectors (CAV-2 are well suited for this goal. These vectors are poorly immunogenic, efficiently transduce neurons, are retrogradely transported to afferent structures in the brain and lead to long-term transgene expression. CAV-2 vectors are being exploited to unravel behavior, cognition, neural networks, axonal transport and therapy for orphan diseases. With the goal of better understanding and characterizing HD-CAV-2 for brain therapy, we analyzed the transcriptomic modulation induced by HD-CAV-2 in human differentiated neurospheres derived from midbrain progenitors. This 3D model system mimics several aspects of the dynamic nature of human brain. We found that differentiated neurospheres are readily transduced by HD-CAV-2 and that transduction generates two main transcriptional responses: a DNA damage response and alteration of centromeric and microtubule probes. Future investigations on the biochemistry of processes highlighted by probe modulations will help defining the implication of HD-CAV-2 and CAR receptor binding in enchaining these functional pathways. We suggest here that the modulation of DNA damage genes is related to viral DNA, while the alteration of centromeric and microtubule probes is possibly enchained by the interaction of the HD-CAV-2 fibre with CAR.

  18. Optimal packaging of FIV genomic RNA depends upon a conserved long-range interaction and a palindromic sequence within gag.

    Science.gov (United States)

    Rizvi, Tahir A; Kenyon, Julia C; Ali, Jahabar; Aktar, Suriya J; Phillip, Pretty S; Ghazawi, Akela; Mustafa, Farah; Lever, Andrew M L

    2010-10-15

    The feline immunodeficiency virus (FIV) is a lentivirus that is related to human immunodeficiency virus (HIV), causing a similar pathology in cats. It is a potential small animal model for AIDS and the FIV-based vectors are also being pursued for human gene therapy. Previous studies have mapped the FIV packaging signal (ψ) to two or more discontinuous regions within the 5' 511 nt of the genomic RNA and structural analyses have determined its secondary structure. The 5' and 3' sequences within ψ region interact through extensive long-range interactions (LRIs), including a conserved heptanucleotide interaction between R/U5 and gag. Other secondary structural elements identified include a conserved 150 nt stem-loop (SL2) and a small palindromic stem-loop within gag open reading frame that might act as a viral dimerization initiation site. We have performed extensive mutational analysis of these sequences and structures and ascertained their importance in FIV packaging using a trans-complementation assay. Disrupting the conserved heptanucleotide LRI to prevent base pairing between R/U5 and gag reduced packaging by 2.8-5.5 fold. Restoration of pairing using an alternative, non-wild type (wt) LRI sequence restored RNA packaging and propagation to wt levels, suggesting that it is the structure of the LRI, rather than its sequence, that is important for FIV packaging. Disrupting the palindrome within gag reduced packaging by 1.5-3-fold, but substitution with a different palindromic sequence did not restore packaging completely, suggesting that the sequence of this region as well as its palindromic nature is important. Mutation of individual regions of SL2 did not have a pronounced effect on FIV packaging, suggesting that either it is the structure of SL2 as a whole that is necessary for optimal packaging, or that there is redundancy within this structure. The mutational analysis presented here has further validated the previously predicted RNA secondary structure of FIV

  19. The mitochondrial genome of the stingless bee Melipona bicolor (Hymenoptera, Apidae, Meliponini: sequence, gene organization and a unique tRNA translocation event conserved across the tribe Meliponini

    Directory of Open Access Journals (Sweden)

    Daniela Silvestre

    2008-01-01

    Full Text Available At present a complete mtDNA sequence has been reported for only two hymenopterans, the Old World honey bee, Apis mellifera and the sawfly Perga condei. Among the bee group, the tribe Meliponini (stingless bees has some distinction due to its Pantropical distribution, great number of species and large importance as main pollinators in several ecosystems, including the Brazilian rain forest. However few molecular studies have been conducted on this group of bees and few sequence data from mitochondrial genomes have been described. In this project, we PCR amplified and sequenced 78% of the mitochondrial genome of the stingless bee Melipona bicolor (Apidae, Meliponini. The sequenced region contains all of the 13 mitochondrial protein-coding genes, 18 of 22 tRNA genes, and both rRNA genes (one of them was partially sequenced. We also report the genome organization (gene content and order, gene translation, genetic code, and other molecular features, such as base frequencies, codon usage, gene initiation and termination. We compare these characteristics of M. bicolor to those of the mitochondrial genome of A. mellifera and other insects. A highly biased A+T content is a typical characteristic of the A. mellifera mitochondrial genome and it was even more extreme in that of M. bicolor. Length and compositional differences between M. bicolor and A. mellifera genes were detected and the gene order was compared. Eleven tRNA gene translocations were observed between these two species. This latter finding was surprising, considering the taxonomic proximity of these two bee tribes. The tRNA Lys gene translocation was investigated within Meliponini and showed high conservation across the Pantropical range of the tribe.

  20. Genomic Approaches in Marine Biodiversity and Aquaculture

    Directory of Open Access Journals (Sweden)

    Jorge A Huete-Pérez

    2013-01-01

    Full Text Available Recent advances in genomic and post-genomic technologies have now established the new standard in medical and biotechnological research. The introduction of next-generation sequencing, NGS,has resulted in the generation of thousands of genomes from all domains of life, including the genomes of complex uncultured microbial communities revealed through metagenomics. Although the application of genomics to marine biodiversity remains poorly developed overall, some noteworthy progress has been made in recent years. The genomes of various model marine organisms have been published and a few more are underway. In addition, the recent large-scale analysis of marine microbes, along with transcriptomic and proteomic approaches to the study of teleost fishes, mollusks and crustaceans, to mention a few, has provided a better understanding of phenotypic variability and functional genomics. The past few years have also seen advances in applications relevant to marine aquaculture and fisheries. In this review we introduce several examples of recent discoveries and progress made towards engendering genomic resources aimed at enhancing our understanding of marine biodiversity and promoting the development of aquaculture. Finally, we discuss the need for auspicious science policies to address challenges confronting smaller nations in the appropriate oversight of this growing domain as they strive to guarantee food security and conservation of their natural resources.

  1. Comparison of carnivore, omnivore, and herbivore mammalian genomes with a new leopard assembly.

    Science.gov (United States)

    Kim, Soonok; Cho, Yun Sung; Kim, Hak-Min; Chung, Oksung; Kim, Hyunho; Jho, Sungwoong; Seomun, Hong; Kim, Jeongho; Bang, Woo Young; Kim, Changmu; An, Junghwa; Bae, Chang Hwan; Bhak, Youngjune; Jeon, Sungwon; Yoon, Hyejun; Kim, Yumi; Jun, JeHoon; Lee, HyeJin; Cho, Suan; Uphyrkina, Olga; Kostyria, Aleksey; Goodrich, John; Miquelle, Dale; Roelke, Melody; Lewis, John; Yurchenko, Andrey; Bankevich, Anton; Cho, Juok; Lee, Semin; Edwards, Jeremy S; Weber, Jessica A; Cook, Jo; Kim, Sangsoo; Lee, Hang; Manica, Andrea; Lee, Ilbeum; O'Brien, Stephen J; Bhak, Jong; Yeo, Joo-Hong

    2016-10-11

    There are three main dietary groups in mammals: carnivores, omnivores, and herbivores. Currently, there is limited comparative genomics insight into the evolution of dietary specializations in mammals. Due to recent advances in sequencing technologies, we were able to perform in-depth whole genome analyses of representatives of these three dietary groups. We investigated the evolution of carnivory by comparing 18 representative genomes from across Mammalia with carnivorous, omnivorous, and herbivorous dietary specializations, focusing on Felidae (domestic cat, tiger, lion, cheetah, and leopard), Hominidae, and Bovidae genomes. We generated a new high-quality leopard genome assembly, as well as two wild Amur leopard whole genomes. In addition to a clear contraction in gene families for starch and sucrose metabolism, the carnivore genomes showed evidence of shared evolutionary adaptations in genes associated with diet, muscle strength, agility, and other traits responsible for successful hunting and meat consumption. Additionally, an analysis of highly conserved regions at the family level revealed molecular signatures of dietary adaptation in each of Felidae, Hominidae, and Bovidae. However, unlike carnivores, omnivores and herbivores showed fewer shared adaptive signatures, indicating that carnivores are under strong selective pressure related to diet. Finally, felids showed recent reductions in genetic diversity associated with decreased population sizes, which may be due to the inflexible nature of their strict diet, highlighting their vulnerability and critical conservation status. Our study provides a large-scale family level comparative genomic analysis to address genomic changes associated with dietary specialization. Our genomic analyses also provide useful resources for diet-related genetic and health research.

  2. Comparative scaffolding and gap filling of ancient bacterial genomes applied to two ancient Yersinia pestis genomes

    Science.gov (United States)

    Doerr, Daniel; Chauve, Cedric

    2017-01-01

    Yersinia pestis is the causative agent of the bubonic plague, a disease responsible for several dramatic historical pandemics. Progress in ancient DNA (aDNA) sequencing rendered possible the sequencing of whole genomes of important human pathogens, including the ancient Y. pestis strains responsible for outbreaks of the bubonic plague in London in the 14th century and in Marseille in the 18th century, among others. However, aDNA sequencing data are still characterized by short reads and non-uniform coverage, so assembling ancient pathogen genomes remains challenging and often prevents a detailed study of genome rearrangements. It has recently been shown that comparative scaffolding approaches can improve the assembly of ancient Y. pestis genomes at a chromosome level. In the present work, we address the last step of genome assembly, the gap-filling stage. We describe an optimization-based method AGapEs (ancestral gap estimation) to fill in inter-contig gaps using a combination of a template obtained from related extant genomes and aDNA reads. We show how this approach can be used to refine comparative scaffolding by selecting contig adjacencies supported by a mix of unassembled aDNA reads and comparative signal. We applied our method to two Y. pestis data sets from the London and Marseilles outbreaks, for which we obtained highly improved genome assemblies for both genomes, comprised of, respectively, five and six scaffolds with 95 % of the assemblies supported by ancient reads. We analysed the genome evolution between both ancient genomes in terms of genome rearrangements, and observed a high level of synteny conservation between these strains. PMID:29114402

  3. Complete plastid genomes from Ophioglossum californicum, Psilotum nudum, and Equisetum hyemale reveal an ancestral land plant genome structure and resolve the position of Equisetales among monilophytes

    Directory of Open Access Journals (Sweden)

    Grewe Felix

    2013-01-01

    Full Text Available Abstract Background Plastid genome structure and content is remarkably conserved in land plants. This widespread conservation has facilitated taxon-rich phylogenetic analyses that have resolved organismal relationships among many land plant groups. However, the relationships among major fern lineages, especially the placement of Equisetales, remain enigmatic. Results In order to understand the evolution of plastid genomes and to establish phylogenetic relationships among ferns, we sequenced the plastid genomes from three early diverging species: Equisetum hyemale (Equisetales, Ophioglossum californicum (Ophioglossales, and Psilotum nudum (Psilotales. A comparison of fern plastid genomes showed that some lineages have retained inverted repeat (IR boundaries originating from the common ancestor of land plants, while other lineages have experienced multiple IR changes including expansions and inversions. Genome content has remained stable throughout ferns, except for a few lineage-specific losses of genes and introns. Notably, the losses of the rps16 gene and the rps12i346 intron are shared among Psilotales, Ophioglossales, and Equisetales, while the gain of a mitochondrial atp1 intron is shared between Marattiales and Polypodiopsida. These genomic structural changes support the placement of Equisetales as sister to Ophioglossales + Psilotales and Marattiales as sister to Polypodiopsida. This result is augmented by some molecular phylogenetic analyses that recover the same relationships, whereas others suggest a relationship between Equisetales and Polypodiopsida. Conclusions Although molecular analyses were inconsistent with respect to the position of Marattiales and Equisetales, several genomic structural changes have for the first time provided a clear placement of these lineages within the ferns. These results further demonstrate the power of using rare genomic structural changes in cases where molecular data fail to provide strong phylogenetic

  4. Biotechnological approaches for conservation and improvement of rare and endangered plants of Saudi Arabia.

    Science.gov (United States)

    Khan, Salim; Al-Qurainy, Fahad; Nadeem, Mohammad

    2012-01-01

    Genetic variation is believed to be a prerequisite for the short-and long-term survival of the plant species in their natural habitat. It depends on many environmental factors which determine the number of alleles on various loci in the genome. Therefore, it is important to understand the genetic composition and structure of the rare and endangered plant species from their natural habitat to develop successful management strategies for their conservation. However, rare and endangered plant species have low genetic diversity due to which their survival rate is decreasing in the wilds. The evaluation of genetic diversity of such species is very important for their conservation and gene manipulation. However, plant species can be conserved by in situ and in vitro methods and each has advantages and disadvantages. DNA banking can be considered as a means of complimentary method for the conservation of plant species by preserving their genomic DNA at low temperatures. Such approach of preservation of biological information provides opportunity for researchers to search novel genes and its products. Therefore, in this review we are describing some potential biotechnological approaches for the conservation and further manipulation of these rare and endangered plant species to enhance their yield and quality traits.

  5. Evolutionary Genomics and Conservation of the Endangered Przewalski’s Horse

    DEFF Research Database (Denmark)

    Der Sarkissian, Clio; Ermini, Luca; Schubert, Mikkel

    2015-01-01

    Przewalski’s horses (PHs, Equus ferus ssp. przewalskii) were discovered in the Asian steppes in the 1870s and represent the last remaining true wild horses. PHs became extinct in the wild in the 1960s but survived in captivity, thanks to major conservation efforts. The current population is still...

  6. Widespread of horizontal gene transfer in the human genome.

    Science.gov (United States)

    Huang, Wenze; Tsai, Lillian; Li, Yulong; Hua, Nan; Sun, Chen; Wei, Chaochun

    2017-04-04

    A fundamental concept in biology is that heritable material is passed from parents to offspring, a process called vertical gene transfer. An alternative mechanism of gene acquisition is through horizontal gene transfer (HGT), which involves movement of genetic materials between different species. Horizontal gene transfer has been found prevalent in prokaryotes but very rare in eukaryote. In this paper, we investigate horizontal gene transfer in the human genome. From the pair-wise alignments between human genome and 53 vertebrate genomes, 1,467 human genome regions (2.6 M bases) from all chromosomes were found to be more conserved with non-mammals than with most mammals. These human genome regions involve 642 known genes, which are enriched with ion binding. Compared to known horizontal gene transfer regions in the human genome, there were few overlapping regions, which indicated horizontal gene transfer is more common than we expected in the human genome. Horizontal gene transfer impacts hundreds of human genes and this study provided insight into potential mechanisms of HGT in the human genome.

  7. Insight into Energy Conservation via Alternative Carbon Monoxide Metabolism in Carboxydothermus pertinax Revealed by Comparative Genome Analysis.

    Science.gov (United States)

    Fukuyama, Yuto; Omae, Kimiho; Yoneda, Yasuko; Yoshida, Takashi; Sako, Yoshihiko

    2018-05-04

    Carboxydothermus species are some of the most studied thermophilic carboxydotrophs. Their varied carboxydotrophic growth properties suggest distinct strategies for energy conservation via CO metabolism. In this study, we used comparative genome analysis of the genus Carboxydothermus to show variations in the CO dehydrogenase/energy-converting hydrogenase gene cluster, which is responsible for CO metabolism with H 2 production (hydrogenogenic CO metabolism). Indeed, ability or inability to produce H 2 with CO oxidation is explained by the presence or absence of this gene cluster in C. hydrogenoformans , C. islandicus , and C. ferrireducens Interestingly, despite its hydrogenogenic CO metabolism, C. pertinax lacks the Ni-CO dehydrogenase catalytic subunit (CooS-I) and its transcriptional regulator encoding genes in this gene cluster probably due to inversion. Transcriptional analysis in C. pertinax showed that the Ni-CO dehydrogenase gene ( cooS-II ) and distantly encoded energy-converting hydrogenase related genes were remarkably upregulated under 100% CO. In addition, when thiosulfate was available as a terminal electron acceptor under 100% CO, C. pertinax maximum cell density and maximum specific growth rate were 3.1-fold and 1.5-fold higher, respectively, than when thiosulfate was absent. The amount of H 2 produced was only 63% of the consumed CO, less than expected according to hydrogenogenic CO oxidation: CO + H 2 O → CO 2 + H 2 Accordingly, C. pertinax would couple CO oxidation by Ni-CO dehydrogenase-II with simultaneous reduction of not only H 2 O but thiosulfate when grown under 100% CO. IMPORTANCE Anaerobic hydrogenogenic carboxydotrophs are thought to fill a vital niche with scavenging potentially toxic CO and producing H 2 as available energy source for thermophilic microbes. This hydrogenogenic carboxydotrophy relies on a Ni-CO dehydrogenase/energy-converting hydrogenase gene cluster. This feature is thought to be as common to these organisms. However

  8. Complete genome sequence and comparative genomics of the probiotic yeast Saccharomyces boulardii.

    Science.gov (United States)

    Khatri, Indu; Tomar, Rajul; Ganesan, K; Prasad, G S; Subramanian, Srikrishna

    2017-03-23

    The probiotic yeast, Saccharomyces boulardii (Sb) is known to be effective against many gastrointestinal disorders and antibiotic-associated diarrhea. To understand molecular basis of probiotic-properties ascribed to Sb we determined the complete genomes of two strains of Sb i.e. Biocodex and unique28 and the draft genomes for three other Sb strains that are marketed as probiotics in India. We compared these genomes with 145 strains of S. cerevisiae (Sc) to understand genome-level similarities and differences between these yeasts. A distinctive feature of Sb from other Sc is absence of Ty elements Ty1, Ty3, Ty4 and associated LTR. However, we could identify complete Ty2 and Ty5 elements in Sb. The genes for hexose transporters HXT11 and HXT9, and asparagine-utilization are absent in all Sb strains. We find differences in repeat periods and copy numbers of repeats in flocculin genes that are likely related to the differential adhesion of Sb as compared to Sc. Core-proteome based taxonomy places Sb strains along with wine strains of Sc. We find the introgression of five genes from Z. bailii into the chromosome IV of Sb and wine strains of Sc. Intriguingly, genes involved in conferring known probiotic properties to Sb are conserved in most Sc strains.

  9. The genome sequence of a widespread apex Predator, the golden eagle (Aquila chrysaetos)

    Science.gov (United States)

    Jacqueline M. Doyle; Todd E. Katzner; Peter H. Bloom; Yanzhu Ji; Bhagya K. Wijayawardena; J. Andrew DeWoody; Ludovic. Orlando

    2014-01-01

    Biologists routinely use molecular markers to identify conservation units, to quantify genetic connectivity, to estimate population sizes, and to identify targets of selection. Many imperiled eagle populations require such efforts and would benefit from enhanced genomic resources. We sequenced, assembled, and annotated the first eagle genome using DNA from a male...

  10. Specific patterns of gene space organisation revealed in wheat by using the combination of barley and wheat genomic resources

    Directory of Open Access Journals (Sweden)

    Waugh Robbie

    2010-12-01

    Full Text Available Abstract Background Because of its size, allohexaploid nature and high repeat content, the wheat genome has always been perceived as too complex for efficient molecular studies. We recently constructed the first physical map of a wheat chromosome (3B. However gene mapping is still laborious in wheat because of high redundancy between the three homoeologous genomes. In contrast, in the closely related diploid species, barley, numerous gene-based markers have been developed. This study aims at combining the unique genomic resources developed in wheat and barley to decipher the organisation of gene space on wheat chromosome 3B. Results Three dimensional pools of the minimal tiling path of wheat chromosome 3B physical map were hybridised to a barley Agilent 15K expression microarray. This led to the fine mapping of 738 barley orthologous genes on wheat chromosome 3B. In addition, comparative analyses revealed that 68% of the genes identified were syntenic between the wheat chromosome 3B and barley chromosome 3 H and 59% between wheat chromosome 3B and rice chromosome 1, together with some wheat-specific rearrangements. Finally, it indicated an increasing gradient of gene density from the centromere to the telomeres positively correlated with the number of genes clustered in islands on wheat chromosome 3B. Conclusion Our study shows that novel structural genomics resources now available in wheat and barley can be combined efficiently to overcome specific problems of genetic anchoring of physical contigs in wheat and to perform high-resolution comparative analyses with rice for deciphering the organisation of the wheat gene space.

  11. The chloroplast genome sequence of the green alga Leptosira terrestris: multiple losses of the inverted repeat and extensive genome rearrangements within the Trebouxiophyceae

    Directory of Open Access Journals (Sweden)

    Turmel Monique

    2007-07-01

    Full Text Available Abstract Background In the Chlorophyta – the green algal phylum comprising the classes Prasinophyceae, Ulvophyceae, Trebouxiophyceae and Chlorophyceae – the chloroplast genome displays a highly variable architecture. While chlorophycean chloroplast DNAs (cpDNAs deviate considerably from the ancestral pattern described for the prasinophyte Nephroselmis olivacea, the degree of remodelling sustained by the two ulvophyte cpDNAs completely sequenced to date is intermediate relative to those observed for chlorophycean and trebouxiophyte cpDNAs. Chlorella vulgaris (Chlorellales is currently the only photosynthetic trebouxiophyte whose complete cpDNA sequence has been reported. To gain insights into the evolutionary trends of the chloroplast genome in the Trebouxiophyceae, we sequenced cpDNA from the filamentous alga Leptosira terrestris (Ctenocladales. Results The 195,081-bp Leptosira chloroplast genome resembles the 150,613-bp Chlorella genome in lacking a large inverted repeat (IR but differs greatly in gene order. Six of the conserved genes present in Chlorella cpDNA are missing from the Leptosira gene repertoire. The 106 conserved genes, four introns and 11 free standing open reading frames (ORFs account for 48.3% of the genome sequence. This is the lowest gene density yet observed among chlorophyte cpDNAs. Contrary to the situation in Chlorella but similar to that in the chlorophycean Scenedesmus obliquus, the gene distribution is highly biased over the two DNA strands in Leptosira. Nine genes, compared to only three in Chlorella, have significantly expanded coding regions relative to their homologues in ancestral-type green algal cpDNAs. As observed in chlorophycean genomes, the rpoB gene is fragmented into two ORFs. Short repeats account for 5.1% of the Leptosira genome sequence and are present mainly in intergenic regions. Conclusion Our results highlight the great plasticity of the chloroplast genome in the Trebouxiophyceae and indicate

  12. Whole-Genome de novo Sequencing Of Quail And Grey Partridge

    DEFF Research Database (Denmark)

    Holm, Lars-Erik; Panitz, Frank; Burt, Dave

    2011-01-01

    The development in sequencing methods has made it possible to perform whole genome de novo sequencing of species without large commercial interests. Within the EU-financed QUANTOMICS project (KBBE-2A-222664), we have performed de novo sequencing of quail (Coturnix coturnix) and grey partridge...... (Perdix perdix) on a Genome Analyzer GAII (Illumina) using paired-end sequencing. The amount of generated sequences amounts to 8 to 9 Gb for each species. The analysis and assembly of the generated sequences is ongoing. Access to the whole genome sequence from these two species will enable enhanced...... comparative studies towards the chicken genome and will aid in identifying evolutionarily conserved sequences within the Galliformes. The obtained sequences from quail and partridge represent a beginning of generating the whole genome sequence for these species. The continuation of establishing the genome...

  13. Genome Wide Identification, Phylogeny, and Expression of Aquaporin Genes in Common Carp (Cyprinus carpio.

    Directory of Open Access Journals (Sweden)

    Chuanju Dong

    Full Text Available Aquaporins (Aqps are integral membrane proteins that facilitate the transport of water and small solutes across cell membranes. Among vertebrate species, Aqps are highly conserved in both gene structure and amino acid sequence. These proteins are vital for maintaining water homeostasis in living organisms, especially for aquatic animals such as teleost fish. Studies on teleost Aqps are mainly limited to several model species with diploid genomes. Common carp, which has a tetraploidized genome, is one of the most common aquaculture species being adapted to a wide range of aquatic environments. The complete common carp genome has recently been released, providing us the possibility for gene evolution of aqp gene family after whole genome duplication.In this study, we identified a total of 37 aqp genes from common carp genome. Phylogenetic analysis revealed that most of aqps are highly conserved. Comparative analysis was performed across five typical vertebrate genomes. We found that almost all of the aqp genes in common carp were duplicated in the evolution of the gene family. We postulated that the expansion of the aqp gene family in common carp was the result of an additional whole genome duplication event and that the aqp gene family in other teleosts has been lost in their evolution history with the reason that the functions of genes are redundant and conservation. Expression patterns were assessed in various tissues, including brain, heart, spleen, liver, intestine, gill, muscle, and skin, which demonstrated the comprehensive expression profiles of aqp genes in the tetraploidized genome. Significant gene expression divergences have been observed, revealing substantial expression divergences or functional divergences in those duplicated aqp genes post the latest WGD event.To some extent, the gene families are also considered as a unique source for evolutionary studies. Moreover, the whole set of common carp aqp gene family provides an

  14. Getting complete genomes from complex samples using nanopore sequencing

    DEFF Research Database (Denmark)

    Kirkegaard, Rasmus Hansen; Karst, Søren Michael; Albertsen, Mads

    Short read sequencing and metagenomic binning workflows have made it possible to extract bacterial genome bins from environmental microbial samples containing hundreds to thousands of different species. However, these genome bins often do not represent complete genomes, as they are mostly...... fragmented, incomplete and often contaminated with foreign DNA and with no robust strategies to validate the quality. The value of these `draft genomes` have limited, lasting value to the scientific community, as gene synteny is broken and the uncertainty of what is missing. The genetic material most often...... missed is important multi-copy and/or conserved marker genes such as the 16S rRNA gene, as sequence micro-heterogeneity prevents assembly of these genes in the de novo assembly. We demonstrate that using nanopore long reads it is now possible to overcome these issues and make complete genomes from...

  15. Getting complete genomes from complex samples using nanopore sequencing

    DEFF Research Database (Denmark)

    Kirkegaard, Rasmus Hansen; Karst, Søren Michael; Albertsen, Mads

    Background Short read DNA sequencing and metagenomic binning workflows have made it possible to extract bacterial genome bins from environmental microbial samples containing hundreds to thousands of different species. However, these genome bins often do not represent complete genomes......, as they are mostly fragmented, incomplete and often contaminated with foreign DNA. The value of these `draft genomes` have limited, lasting value to the scientific community, as gene synteny is broken and there is some uncertainty of what is missing1. The genetic material most often missed is important multi......-copy and/or conserved marker genes such as the 16S rRNA gene, as sequence micro-heterogeneity prevents assembly of these genes in the de novo assembly. However, long read sequencing technologies are emerging promising an end to fragmented genome assemblies2. Experimental design We extracted DNA from a full...

  16. Genome-wide association between DNA methylation and alternative splicing in an invertebrate

    Directory of Open Access Journals (Sweden)

    Flores Kevin

    2012-09-01

    Full Text Available Abstract Background Gene bodies are the most evolutionarily conserved targets of DNA methylation in eukaryotes. However, the regulatory functions of gene body DNA methylation remain largely unknown. DNA methylation in insects appears to be primarily confined to exons. Two recent studies in Apis mellifera (honeybee and Nasonia vitripennis (jewel wasp analyzed transcription and DNA methylation data for one gene in each species to demonstrate that exon-specific DNA methylation may be associated with alternative splicing events. In this study we investigated the relationship between DNA methylation, alternative splicing, and cross-species gene conservation on a genome-wide scale using genome-wide transcription and DNA methylation data. Results We generated RNA deep sequencing data (RNA-seq to measure genome-wide mRNA expression at the exon- and gene-level. We produced a de novo transcriptome from this RNA-seq data and computationally predicted splice variants for the honeybee genome. We found that exons that are included in transcription are higher methylated than exons that are skipped during transcription. We detected enrichment for alternative splicing among methylated genes compared to unmethylated genes using fisher’s exact test. We performed a statistical analysis to reveal that the presence of DNA methylation or alternative splicing are both factors associated with a longer gene length and a greater number of exons in genes. In concordance with this observation, a conservation analysis using BLAST revealed that each of these factors is also associated with higher cross-species gene conservation. Conclusions This study constitutes the first genome-wide analysis exhibiting a positive relationship between exon-level DNA methylation and mRNA expression in the honeybee. Our finding that methylated genes are enriched for alternative splicing suggests that, in invertebrates, exon-level DNA methylation may play a role in the construction of splice

  17. The complete chloroplast genome sequence of Dodonaea viscosa: comparative and phylogenetic analyses.

    Science.gov (United States)

    Saina, Josphat K; Gichira, Andrew W; Li, Zhi-Zhong; Hu, Guang-Wan; Wang, Qing-Feng; Liao, Kuo

    2018-02-01

    The plant chloroplast (cp) genome is a highly conserved structure which is beneficial for evolution and systematic research. Currently, numerous complete cp genome sequences have been reported due to high throughput sequencing technology. However, there is no complete chloroplast genome of genus Dodonaea that has been reported before. To better understand the molecular basis of Dodonaea viscosa chloroplast, we used Illumina sequencing technology to sequence its complete genome. The whole length of the cp genome is 159,375 base pairs (bp), with a pair of inverted repeats (IRs) of 27,099 bp separated by a large single copy (LSC) 87,204 bp, and small single copy (SSC) 17,972 bp. The annotation analysis revealed a total of 115 unique genes of which 81 were protein coding, 30 tRNA, and four ribosomal RNA genes. Comparative genome analysis with other closely related Sapindaceae members showed conserved gene order in the inverted and single copy regions. Phylogenetic analysis clustered D. viscosa with other species of Sapindaceae with strong bootstrap support. Finally, a total of 249 SSRs were detected. Moreover, a comparison of the synonymous (Ks) and nonsynonymous (Ka) substitution rates in D. viscosa showed very low values. The availability of cp genome reported here provides a valuable genetic resource for comprehensive further studies in genetic variation, taxonomy and phylogenetic evolution of Sapindaceae family. In addition, SSR markers detected will be used in further phylogeographic and population structure studies of the species in this genus.

  18. Conservation, duplication, and loss of the Tor signaling pathway in the fungal kingdom

    Directory of Open Access Journals (Sweden)

    Heitman Joseph

    2010-09-01

    Full Text Available Abstract Background The nutrient-sensing Tor pathway governs cell growth and is conserved in nearly all eukaryotic organisms from unicellular yeasts to multicellular organisms, including humans. Tor is the target of the immunosuppressive drug rapamycin, which in complex with the prolyl isomerase FKBP12 inhibits Tor functions. Rapamycin is a gold standard drug for organ transplant recipients that was approved by the FDA in 1999 and is finding additional clinical indications as a chemotherapeutic and antiproliferative agent. Capitalizing on the plethora of recently sequenced genomes we have conducted comparative genomic studies to annotate the Tor pathway throughout the fungal kingdom and related unicellular opisthokonts, including Monosiga brevicollis, Salpingoeca rosetta, and Capsaspora owczarzaki. Results Interestingly, the Tor signaling cascade is absent in three microsporidian species with available genome sequences, the only known instance of a eukaryotic group lacking this conserved pathway. The microsporidia are obligate intracellular pathogens with highly reduced genomes, and we hypothesize that they lost the Tor pathway as they adapted and streamlined their genomes for intracellular growth in a nutrient-rich environment. Two TOR paralogs are present in several fungal species as a result of either a whole genome duplication or independent gene/segmental duplication events. One such event was identified in the amphibian pathogen Batrachochytrium dendrobatidis, a chytrid responsible for worldwide global amphibian declines and extinctions. Conclusions The repeated independent duplications of the TOR gene in the fungal kingdom might reflect selective pressure acting upon this kinase that populates two proteinaceous complexes with different cellular roles. These comparative genomic analyses illustrate the evolutionary trajectory of a central nutrient-sensing cascade that enables diverse eukaryotic organisms to respond to their natural

  19. The Genome and Development-Dependent Transcriptomes of Pyronema confluens: A Window into Fungal Evolution

    Science.gov (United States)

    Traeger, Stefanie; Altegoer, Florian; Freitag, Michael; Gabaldon, Toni; Kempken, Frank; Kumar, Abhishek; Marcet-Houben, Marina; Pöggeler, Stefanie; Stajich, Jason E.; Nowrousian, Minou

    2013-01-01

    Fungi are a large group of eukaryotes found in nearly all ecosystems. More than 250 fungal genomes have already been sequenced, greatly improving our understanding of fungal evolution, physiology, and development. However, for the Pezizomycetes, an early-diverging lineage of filamentous ascomycetes, there is so far only one genome available, namely that of the black truffle, Tuber melanosporum, a mycorrhizal species with unusual subterranean fruiting bodies. To help close the sequence gap among basal filamentous ascomycetes, and to allow conclusions about the evolution of fungal development, we sequenced the genome and assayed transcriptomes during development of Pyronema confluens, a saprobic Pezizomycete with a typical apothecium as fruiting body. With a size of 50 Mb and ∼13,400 protein-coding genes, the genome is more characteristic of higher filamentous ascomycetes than the large, repeat-rich truffle genome; however, some typical features are different in the P. confluens lineage, e.g. the genomic environment of the mating type genes that is conserved in higher filamentous ascomycetes, but only partly conserved in P. confluens. On the other hand, P. confluens has a full complement of fungal photoreceptors, and expression studies indicate that light perception might be similar to distantly related ascomycetes and, thus, represent a basic feature of filamentous ascomycetes. Analysis of spliced RNA-seq sequence reads allowed the detection of natural antisense transcripts for 281 genes. The P. confluens genome contains an unusually high number of predicted orphan genes, many of which are upregulated during sexual development, consistent with the idea of rapid evolution of sex-associated genes. Comparative transcriptomics identified the transcription factor gene pro44 that is upregulated during development in P. confluens and the Sordariomycete Sordaria macrospora. The P. confluens pro44 gene (PCON_06721) was used to complement the S. macrospora pro44 deletion

  20. The genome and development-dependent transcriptomes of Pyronema confluens: a window into fungal evolution.

    Directory of Open Access Journals (Sweden)

    Stefanie Traeger

    Full Text Available Fungi are a large group of eukaryotes found in nearly all ecosystems. More than 250 fungal genomes have already been sequenced, greatly improving our understanding of fungal evolution, physiology, and development. However, for the Pezizomycetes, an early-diverging lineage of filamentous ascomycetes, there is so far only one genome available, namely that of the black truffle, Tuber melanosporum, a mycorrhizal species with unusual subterranean fruiting bodies. To help close the sequence gap among basal filamentous ascomycetes, and to allow conclusions about the evolution of fungal development, we sequenced the genome and assayed transcriptomes during development of Pyronema confluens, a saprobic Pezizomycete with a typical apothecium as fruiting body. With a size of 50 Mb and ~13,400 protein-coding genes, the genome is more characteristic of higher filamentous ascomycetes than the large, repeat-rich truffle genome; however, some typical features are different in the P. confluens lineage, e.g. the genomic environment of the mating type genes that is conserved in higher filamentous ascomycetes, but only partly conserved in P. confluens. On the other hand, P. confluens has a full complement of fungal photoreceptors, and expression studies indicate that light perception might be similar to distantly related ascomycetes and, thus, represent a basic feature of filamentous ascomycetes. Analysis of spliced RNA-seq sequence reads allowed the detection of natural antisense transcripts for 281 genes. The P. confluens genome contains an unusually high number of predicted orphan genes, many of which are upregulated during sexual development, consistent with the idea of rapid evolution of sex-associated genes. Comparative transcriptomics identified the transcription factor gene pro44 that is upregulated during development in P. confluens and the Sordariomycete Sordaria macrospora. The P. confluens pro44 gene (PCON_06721 was used to complement the S. macrospora