WorldWideScience

Sample records for populus genome integration

  1. Creation and genomic analysis of irradiation hybrids in Populus

    Science.gov (United States)

    Matthew S. Zinkgraf; K. Haiby; M.C. Lieberman; L. Comai; I.M. Henry; Andrew Groover

    2016-01-01

    Establishing efficient functional genomic systems for creating and characterizing genetic variation in forest trees is challenging. Here we describe protocols for creating novel gene-dosage variation in Populus through gamma-irradiation of pollen, followed by genomic analysis to identify chromosomal regions that have been deleted or inserted in...

  2. Diversity of Pseudomonas Genomes, Including Populus-Associated Isolates, as Revealed by Comparative Genome Analysis.

    Science.gov (United States)

    Jun, Se-Ran; Wassenaar, Trudy M; Nookaew, Intawat; Hauser, Loren; Wanchai, Visanu; Land, Miriam; Timm, Collin M; Lu, Tse-Yuan S; Schadt, Christopher W; Doktycz, Mitchel J; Pelletier, Dale A; Ussery, David W

    2016-01-01

    The Pseudomonas genus contains a metabolically versatile group of organisms that are known to occupy numerous ecological niches, including the rhizosphere and endosphere of many plants. Their diversity influences the phylogenetic diversity and heterogeneity of these communities. On the basis of average amino acid identity, comparative genome analysis of >1,000 Pseudomonas genomes, including 21 Pseudomonas strains isolated from the roots of native Populus deltoides (eastern cottonwood) trees resulted in consistent and robust genomic clusters with phylogenetic homogeneity. All Pseudomonas aeruginosa genomes clustered together, and these were clearly distinct from other Pseudomonas species groups on the basis of pangenome and core genome analyses. In contrast, the genomes of Pseudomonas fluorescens were organized into 20 distinct genomic clusters, representing enormous diversity and heterogeneity. Most of our 21 Populus-associated isolates formed three distinct subgroups within the major P. fluorescens group, supported by pathway profile analysis, while two isolates were more closely related to Pseudomonas chlororaphis and Pseudomonas putida. Genes specific to Populus-associated subgroups were identified. Genes specific to subgroup 1 include several sensory systems that act in two-component signal transduction, a TonB-dependent receptor, and a phosphorelay sensor. Genes specific to subgroup 2 contain hypothetical genes, and genes specific to subgroup 3 were annotated with hydrolase activity. This study justifies the need to sequence multiple isolates, especially from P. fluorescens, which displays the most genetic variation, in order to study functional capabilities from a pangenomic perspective. This information will prove useful when choosing Pseudomonas strains for use to promote growth and increase disease resistance in plants. Copyright © 2015 Jun et al.

  3. The Plant Genome Integrative Explorer Resource: PlantGenIE.org.

    Science.gov (United States)

    Sundell, David; Mannapperuma, Chanaka; Netotea, Sergiu; Delhomme, Nicolas; Lin, Yao-Cheng; Sjödin, Andreas; Van de Peer, Yves; Jansson, Stefan; Hvidsten, Torgeir R; Street, Nathaniel R

    2015-12-01

    Accessing and exploring large-scale genomics data sets remains a significant challenge to researchers without specialist bioinformatics training. We present the integrated PlantGenIE.org platform for exploration of Populus, conifer and Arabidopsis genomics data, which includes expression networks and associated visualization tools. Standard features of a model organism database are provided, including genome browsers, gene list annotation, Blast homology searches and gene information pages. Community annotation updating is supported via integration of WebApollo. We have produced an RNA-sequencing (RNA-Seq) expression atlas for Populus tremula and have integrated these data within the expression tools. An updated version of the ComPlEx resource for performing comparative plant expression analyses of gene coexpression network conservation between species has also been integrated. The PlantGenIE.org platform provides intuitive access to large-scale and genome-wide genomics data from model forest tree species, facilitating both community contributions to annotation improvement and tools supporting use of the included data resources to inform biological insight. © 2015 The Authors. New Phytologist © 2015 New Phytologist Trust.

  4. Twenty-one genome sequences from Pseudomonas species and 19 genome sequences from diverse bacteria isolated from the rhizosphere and endosphere of Populus deltoides.

    Science.gov (United States)

    Brown, Steven D; Utturkar, Sagar M; Klingeman, Dawn M; Johnson, Courtney M; Martin, Stanton L; Land, Miriam L; Lu, Tse-Yuan S; Schadt, Christopher W; Doktycz, Mitchel J; Pelletier, Dale A

    2012-11-01

    To aid in the investigation of the Populus deltoides microbiome, we generated draft genome sequences for 21 Pseudomonas strains and 19 other diverse bacteria isolated from Populus deltoides roots. Genome sequences for isolates similar to Acidovorax, Bradyrhizobium, Brevibacillus, Caulobacter, Chryseobacterium, Flavobacterium, Herbaspirillum, Novosphingobium, Pantoea, Phyllobacterium, Polaromonas, Rhizobium, Sphingobium, and Variovorax were generated.

  5. Twenty-One Genome Sequences from Pseudomonas Species and 19 Genome Sequences from Diverse Bacteria Isolated from the Rhizosphere and Endosphere of Populus deltoides

    Energy Technology Data Exchange (ETDEWEB)

    Brown, Steven D [ORNL; Utturkar, Sagar M [ORNL; Klingeman, Dawn Marie [ORNL; Johnson, Courtney M [ORNL; Martin, Stanton [ORNL; Land, Miriam L [ORNL; Lu, Tse-Yuan [ORNL; Schadt, Christopher Warren [ORNL; Doktycz, Mitchel John [ORNL; Pelletier, Dale A [ORNL

    2012-01-01

    To aid in the investigation of the Populus deltoides microbiome we generated draft genome sequences for twenty one Pseudomonas and twenty one other diverse bacteria isolated from Populus deltoides roots. Genome sequences for isolates similar to Acidovorax, Bradyrhizobium, Brevibacillus, Burkholderia, Caulobacter, Chryseobacterium, Flavobacterium, Herbaspirillum, Novosphingobium, Pantoea, Phyllobacterium, Polaromonas, Rhizobium, Sphingobium and Variovorax were generated.

  6. Comparative genomic analysis of the WRKY III gene family in populus, grape, arabidopsis and rice.

    Science.gov (United States)

    Wang, Yiyi; Feng, Lin; Zhu, Yuxin; Li, Yuan; Yan, Hanwei; Xiang, Yan

    2015-09-08

    WRKY III genes have significant functions in regulating plant development and resistance. In plant, WRKY gene family has been studied in many species, however, there still lack a comprehensive analysis of WRKY III genes in the woody plant species poplar, three representative lineages of flowering plant species are incorporated in most analyses: Arabidopsis (a model plant for annual herbaceous dicots), grape (one model plant for perennial dicots) and Oryza sativa (a model plant for monocots). In this study, we identified 10, 6, 13 and 28 WRKY III genes in the genomes of Populus trichocarpa, grape (Vitis vinifera), Arabidopsis thaliana and rice (Oryza sativa), respectively. Phylogenetic analysis revealed that the WRKY III proteins could be divided into four clades. By microsynteny analysis, we found that the duplicated regions were more conserved between poplar and grape than Arabidopsis or rice. We dated their duplications by Ks analysis of Populus WRKY III genes and demonstrated that all the blocks were formed after the divergence of monocots and dicots. Strong purifying selection has played a key role in the maintenance of WRKY III genes in Populus. Tissue expression analysis of the WRKY III genes in Populus revealed that five were most highly expressed in the xylem. We also performed quantitative real-time reverse transcription PCR analysis of WRKY III genes in Populus treated with salicylic acid, abscisic acid and polyethylene glycol to explore their stress-related expression patterns. This study highlighted the duplication and diversification of the WRKY III gene family in Populus and provided a comprehensive analysis of this gene family in the Populus genome. Our results indicated that the majority of WRKY III genes of Populus was expanded by large-scale gene duplication. The expression pattern of PtrWRKYIII gene identified that these genes play important roles in the xylem during poplar growth and development, and may play crucial role in defense to drought

  7. Evolutionary Quantitative Genomics of Populus trichocarpa.

    Directory of Open Access Journals (Sweden)

    Ilga Porth

    Full Text Available Forest trees generally show high levels of local adaptation and efforts focusing on understanding adaptation to climate will be crucial for species survival and management. Here, we address fundamental questions regarding the molecular basis of adaptation in undomesticated forest tree populations to past climatic environments by employing an integrative quantitative genetics and landscape genomics approach. Using this comprehensive approach, we studied the molecular basis of climate adaptation in 433 Populus trichocarpa (black cottonwood genotypes originating across western North America. Variation in 74 field-assessed traits (growth, ecophysiology, phenology, leaf stomata, wood, and disease resistance was investigated for signatures of selection (comparing QST-FST using clustering of individuals by climate of origin (temperature and precipitation. 29,354 SNPs were investigated employing three different outlier detection methods and marker-inferred relatedness was estimated to obtain the narrow-sense estimate of population differentiation in wild populations. In addition, we compared our results with previously assessed selection of candidate SNPs using the 25 topographical units (drainages across the P. trichocarpa sampling range as population groupings. Narrow-sense QST for 53% of distinct field traits was significantly divergent from expectations of neutrality (indicating adaptive trait variation; 2,855 SNPs showed signals of diversifying selection and of these, 118 SNPs (within 81 genes were associated with adaptive traits (based on significant QST. Many SNPs were putatively pleiotropic for functionally uncorrelated adaptive traits, such as autumn phenology, height, and disease resistance. Evolutionary quantitative genomics in P. trichocarpa provides an enhanced understanding regarding the molecular basis of climate-driven selection in forest trees and we highlight that important loci underlying adaptive trait variation also show

  8. Genotyping-by-sequencing for Populus population genomics: an assessment of genome sampling patterns and filtering approaches.

    Directory of Open Access Journals (Sweden)

    Martin P Schilling

    Full Text Available Continuing advances in nucleotide sequencing technology are inspiring a suite of genomic approaches in studies of natural populations. Researchers are faced with data management and analytical scales that are increasing by orders of magnitude. With such dramatic advances comes a need to understand biases and error rates, which can be propagated and magnified in large-scale data acquisition and processing. Here we assess genomic sampling biases and the effects of various population-level data filtering strategies in a genotyping-by-sequencing (GBS protocol. We focus on data from two species of Populus, because this genus has a relatively small genome and is emerging as a target for population genomic studies. We estimate the proportions and patterns of genomic sampling by examining the Populus trichocarpa genome (Nisqually-1, and demonstrate a pronounced bias towards coding regions when using the methylation-sensitive ApeKI restriction enzyme in this species. Using population-level data from a closely related species (P. tremuloides, we also investigate various approaches for filtering GBS data to retain high-depth, informative SNPs that can be used for population genetic analyses. We find a data filter that includes the designation of ambiguous alleles resulted in metrics of population structure and Hardy-Weinberg equilibrium that were most consistent with previous studies of the same populations based on other genetic markers. Analyses of the filtered data (27,910 SNPs also resulted in patterns of heterozygosity and population structure similar to a previous study using microsatellites. Our application demonstrates that technically and analytically simple approaches can readily be developed for population genomics of natural populations.

  9. Natural Selection and Recombination Rate Variation Shape Nucleotide Polymorphism Across the Genomes of Three Related Populus Species.

    Science.gov (United States)

    Wang, Jing; Street, Nathaniel R; Scofield, Douglas G; Ingvarsson, Pär K

    2016-03-01

    A central aim of evolutionary genomics is to identify the relative roles that various evolutionary forces have played in generating and shaping genetic variation within and among species. Here we use whole-genome resequencing data to characterize and compare genome-wide patterns of nucleotide polymorphism, site frequency spectrum, and population-scaled recombination rates in three species of Populus: Populus tremula, P. tremuloides, and P. trichocarpa. We find that P. tremuloides has the highest level of genome-wide variation, skewed allele frequencies, and population-scaled recombination rates, whereas P. trichocarpa harbors the lowest. Our findings highlight multiple lines of evidence suggesting that natural selection, due to both purifying and positive selection, has widely shaped patterns of nucleotide polymorphism at linked neutral sites in all three species. Differences in effective population sizes and rates of recombination largely explain the disparate magnitudes and signatures of linked selection that we observe among species. The present work provides the first phylogenetic comparative study on a genome-wide scale in forest trees. This information will also improve our ability to understand how various evolutionary forces have interacted to influence genome evolution among related species. Copyright © 2016 by the Genetics Society of America.

  10. A genomics investigation of partitioning into and among flavonoid-derived condensed tannins for carbon sequestration in Populus

    Energy Technology Data Exchange (ETDEWEB)

    Harding, Scott, A; Tsai, Chung-jui; Lindroth, Richard, L

    2013-03-24

    The project set out to use comparative (genotype and treatment) and transgenic approaches to investigate the determinants of condensed tannin (CT) accrual and chemical variability in Populus. CT type and amount are thought to effect the decomposition of plant detritus in the soil, and thereby the sequestering of carbon in the soil. The stated objectives were: 1. Genome-wide transcriptome profiling (microarrays) to analyze structural gene, transcription factor and metabolite control of CT partitioning; 2. Transcriptomic (microarray) and chemical analysis of ontogenetic effects on CT and PG partitioning; and 3. Transgenic manipulation of flavonoid biosynthetic pathway genes to modify the control of CT composition. Objective 1: A number of approaches for perturbing CT content and chemistry were tested in Objective 1, and those included nitrogen deficit, leaf wounding, drought, and salicylic acid spraying. Drought had little effect on CTs in the genotypes we used. Plants exhibited unpredictability in their response to salicylic acid spraying, leading us to abandon its use. Reduced plant nitrogen status and leaf wounding caused reproducible and magnitudinally striking increases in leaf CT content. Microarray submissions to NCBI from those experiments are the following: GSE ID 14515: Comparative transcriptomics analysis of Populus leaves under nitrogen limitation: clone 1979. Public on Jan 04, 2010; Contributor(s) Harding SA, Tsai C GSE ID 14893: Comparative transcriptomics analysis of Populus leaves under nitrogen limitation: clone 3200. Public on Feb 19, 2009; Contributor(s) Harding SA, Tsai C GSE ID 16783 Wound-induced gene expression changes in Populus: 1 week; clone RM5. Status Public on Dec 01, 2009; Contributor(s) Harding SA, Tsai C GSE ID 16785 Wound-induced gene expression changes in Populus: 90 hours; clone RM5 Status Public on Dec 01, 2009; Contributor(s) Harding SA, Tsai C Although CT amount changed in response to treatments, CT composition was essentially

  11. Efficient Agrobacterium-Mediated Transformation of Hybrid Poplar Populus davidiana Dode × Populus bollena Lauche

    Directory of Open Access Journals (Sweden)

    Xue Han

    2013-01-01

    Full Text Available Poplar is a model organism for high in vitro regeneration in woody plants. We have chosen a hybrid poplar Populus davidiana Dode × Populus bollena Lauche. By optimizing the Murashige and Skoog medium with (0.3 mg/L 6-benzylaminopurine and (0.08 mg/L naphthaleneacetic acid, we have achieved the highest frequency (90% for shoot regeneration from poplar leaves. It was also important to improve the transformation efficiency of poplar for genetic breeding and other applications. In this study, we found a significant improvement of the transformation frequency by controlling the leaf age. Transformation efficiency was enhanced by optimizing the Agrobacterium concentration (OD600 = 0.8–1.0 and an infection time (20–30 min. According to transmission electron microscopy observations, there were more Agrobacterium invasions in the 30-day-old leaf explants than in 60-day-old and 90-day-old explants. Using the green fluorescent protein (GFP marker, the expression of MD–GFP fusion proteins in the leaf, shoot, and root of hybrid poplar P. davidiana Dode × P. bollena Lauche was visualized for confirmation of transgene integration. Southern and Northern blot analysis also showed the integration of T-DNA into the genome and gene expression of transgenic plants. Our results suggest that younger leaves had higher transformation efficiency (~30% than older leaves (10%.

  12. Insertional mutagenesis in Populus: relevance and feasibility

    Science.gov (United States)

    Victor Busov; Matthias Fladung; Andrew Groover; Steven Strauss

    2005-01-01

    The recent sequencing of the first tree genome, that of the black cottonwood (Populus trichocarpa), opens a new chapter in tree functional genomics. While the completion of the genome is a milestone, mobilizing this significant resource for better understanding the growth and development of woody perennials will be an even greater undertaking in the...

  13. Genome-Wide Analysis of a TaLEA-Introduced Transgenic Populus simonii × Populus nigra Dwarf Mutant

    Directory of Open Access Journals (Sweden)

    Jing Jiang

    2012-03-01

    Full Text Available A dwarf mutant (dwf1 was obtained among 15 transgenic lines, when TaLEA (Tamarix androssowii late embryogenesis abundant gene was introduced into Populus simonii × Populus nigra by Agrobacterium tumefaciens-mediated transformation. Under the same growth conditions, dwf1 height was significantly reduced compared with the wild type and the other transgenic lines. Because only one transgenic line (dwf1 displayed the dwarf phenotype, we considered that T-DNA insertion sites may play a role in the mutant formation. The mechanisms underlying this effect were investigated using TAIL-PCR (thermal asymmetric interlaced PCR and microarrays methods. According to the TAIL-PCR results, two flanking sequences located on chromosome IV and VIII respectively, were cloned. The results indicated the integration of two independent T-DNA copies. We searched for the potential genes near to the T-DNA insertions. The nearest gene was a putative poplar AP2 transcription factor (GI: 224073210. Expression analysis showed that AP2 was up-regulated in dwf1 compared with the wild type and the other transgenic lines. According to the microarrays results, a total of 537 genes involved in hydrolase, kinase and transcription factor activities, as well as protein and nucleotide binding, showed significant alterations in gene expression. These genes were expressed in more than 60 metabolic pathways, including starch, sucrose, galactose and glycerolipid metabolism and phenylpropanoids and flavonoid biosyntheses. Our transcriptome and T-DNA insertion sites analyses might provide some useful insights into the dwarf mutant formation.

  14. Dense genetic linkage maps of three Populus species (Populus deltoides, P. nigra and P. trichocarpa) based on AFLP and microsatellite markers.

    Science.gov (United States)

    Cervera, M T; Storme, V; Ivens, B; Gusmão, J; Liu, B H; Hostyn, V; Van Slycken, J; Van Montagu, M; Boerjan, W

    2001-06-01

    Populus deltoides, P. nigra, and P. trichocarpa are the most important species for poplar breeding programs worldwide. In addition, Populus has become a model for fundamental research on trees. Linkage maps were constructed for these three species by analyzing progeny of two controlled crosses sharing the same female parent, Populus deltoides cv. S9-2 x P. nigra cv. Ghoy and P. deltoides cv. S9-2 x P. trichocarpa cv. V24. The two-way pseudotestcross mapping strategy was used to construct the maps. Amplified fragment length polymorphism (AFLP) markers that segregated 1:1 were used to form the four parental maps. Microsatellites and sequence-tagged sites were used to align homoeologous groups between the maps and to merge linkage groups within the individual maps. Linkage analysis and alignment of the homoeologous groups resulted in 566 markers distributed over 19 groups for P. deltoides covering 86% of the genome, 339 markers distributed over 19 groups for P. trichocarpa covering 73%, and 369 markers distributed over 28 groups for P. nigra covering 61%. Several tests for randomness showed that the AFLP markers were randomly distributed over the genome.

  15. A systems biology, whole-genome association analysis of the molecular regulation of biomass growth and composition in Populus deltoides

    Energy Technology Data Exchange (ETDEWEB)

    Kirst, Matias [Univ. of Florida, Gainesville, FL (United States)

    2014-04-14

    Poplars trees are well suited for biofuel production due to their fast growing habit, favorable wood composition and adaptation to a broad range of environments. The availability of a reference genome sequence, ease of vegetative propagation and availability of transformation methods also make poplar an ideal model for the study of wood formation and biomass growth in woody, perennial plants. The objective of this project was to conduct a genome-wide association genetics study to identify genes that regulate bioenergy traits in Populus deltoides (eastern cottonwood). Populus deltoides is a genetically diverse keystone forest species in North America and an important short rotation woody crop for the bioenergy industry. We searched for associations between eight growth and wood composition traits and common and low-frequency single-nucleotide polymorphisms (SNPs) detected by targeted resequencing of 18,153 genes in a population of 391 unrelated individuals. To increase power to detect associations with low-frequency variants, multiple-marker association tests were used in combination with single-marker association tests. Significant associations were discovered for all phenotypes and are indicative that low-frequency polymorphisms contribute to phenotypic variance of several bioenergy traits. These polymorphism are critical tools for the development of specialized plant feedstocks for bioenergy.

  16. A systems biology, whole-genome association analysis of the molecular regulation of biomass growth and composition in Populus deltoides

    Energy Technology Data Exchange (ETDEWEB)

    Kirst, Matias [Univ. of Florida, Gainesville, FL (United States)

    2015-04-15

    Poplars trees are well suited for biofuel production due to their fast growing habit, favorable wood composition and adaptation to a broad range of environments. The availability of a reference genome sequence, ease of vegetative propagation and availability of transformation methods also make poplar an ideal model for the study of wood formation and biomass growth in woody, perennial plants. The objective of this project was to conduct a genome-wide association genetics study to identify genes that regulate bioenergy traits in Populus deltoides (eastern cottonwood). Populus deltoides is a genetically diverse keystone forest species in North America and an important short rotation woody crop for the bioenergy industry. We searched for associations between eight growth and wood composition traits and common and low-frequency single-nucleotide polymorphisms (SNPs) detected by targeted resequencing of 18,153 genes in a population of 391 unrelated individuals. To increase power to detect associations with low-frequency variants, multiple-marker association tests were used in combination with single-marker association tests. Significant associations were discovered for all phenotypes and are indicative that low-frequency polymorphisms contribute to phenotypic variance of several bioenergy traits. These polymorphism are critical tools for the development of specialized plant feedstocks for bioenergy.

  17. Phytoremediation of landfill leachate using Populus

    Science.gov (United States)

    Jill A. Zalesny; Ronald S., Jr. Zalesny; Adam H. Wiese; Richard B. Hall; Bart Sexton

    2006-01-01

    Proper genotype selection is required for successful phytoremediation. We selected eight Populus clones (NC13460, NC14018, DM115, NC14104, NC14106, DN5, NM2, NM6) of four genomic groups after three cycles of phyto-recurrent selection for a field trial that began June 2005 at the Oneida County Landfill in Rhinelander, WI, USA.

  18. Hsf and Hsp gene families in Populus: genome-wide identification, organization and correlated expression during development and in stress responses.

    Science.gov (United States)

    Zhang, Jin; Liu, Bobin; Li, Jianbo; Zhang, Li; Wang, Yan; Zheng, Huanquan; Lu, Mengzhu; Chen, Jun

    2015-03-14

    Heat shock proteins (Hsps) are molecular chaperones that are involved in many normal cellular processes and stress responses, and heat shock factors (Hsfs) are the transcriptional activators of Hsps. Hsfs and Hsps are widely coordinated in various biological processes. Although the roles of Hsfs and Hsps in stress responses have been well characterized in Arabidopsis, their roles in perennial woody species undergoing various environmental stresses remain unclear. Here, a comprehensive identification and analysis of Hsf and Hsp families in poplars is presented. In Populus trichocarpa, we identified 42 paralogous pairs, 66.7% resulting from a whole genome duplication. The gene structure and motif composition are relatively conserved in each subfamily. Microarray and quantitative real-time RT-PCR analyses showed that most of the Populus Hsf and Hsp genes are differentially expressed upon exposure to various stresses. A coexpression network between Populus Hsf and Hsp genes was generated based on their expression. Coordinated relationships were validated by transient overexpression and subsequent qPCR analyses. The comprehensive analysis indicates that different sets of PtHsps are downstream of particular PtHsfs and provides a basis for functional studies aimed at revealing the roles of these families in poplar development and stress responses.

  19. Mitochondrial DNA variation and genetic relationships of Populus species.

    Science.gov (United States)

    Barrett, J W; Rajora, O P; Yeh, F C; Dancik, B P; Strobeck, C

    1993-02-01

    We examined variation in and around the region coding for the cytochrome c oxidase I (coxI) and ATPase 6 (atp6) genes in the mitochondrial genomes of four Populus species (P. nigra, P. deltoides, P. maximowiczii, and P. tremuloides) and the natural hybrid P. x canadensis (P. deltoides x P. nigra). Total cellular DNAs of these poplars were digested with 16 restriction endonucleases and probed with maize mtDNA-specific probes (CoxI and Atp6). The only variant observed for Atp6 was interspecific, with P. maximowiczii separated from the other species as revealed by EcoRI digestions. No intraspecific mtDNA variation was observed among individuals of P. nigra, P. maximowiczii, P. x canadensis, or P. tremuloides for the CoxI probe. However, two varieties of P. deltoides were distinct because of a single site change in the KpnI digestions, demonstrating that P. deltoides var. deltoides (eastern cottonwood) and var. occidentalis (plains cottonwood) have distinct mitochondrial genomes in the region of the coxI gene. Populus x canadensis shared the same restriction fragment patterns as its suspected maternal parent P. deltoides. Nucleotide substitutions per base in and around the coxI and atp6 genes among the Populus species and the hybrid ranged from 0.0017 to 0.0077. The interspecific estimates of nucleotide substitution per base suggested that P. tremuloides was furthest removed from P. deltoides and P. x canadensis and least diverged from P. nigra. Populus maximowiczii was placed between these two clusters.

  20. The Populus ARBORKNOX1 homeodomain transcription factor regulates woody growth through binding to evolutionarily conserved target genes of diverse function

    Science.gov (United States)

    Lijun Liu; Matthew S. Zinkgraf; H. Earl Petzold; Eric P. Beers; Vladimir Filkov; Andrew Groover

    2014-01-01

    The class I KNOX homeodomain transcription factor ARBORKNOX1 (ARK1) is a key regulator of vascular cambium maintenance and cell differentiation in Populus. Currently, basic information is lacking concerning the distribution, functional characteristics, and evolution of ARK1 binding in the Populus genome.

  1. Identification of Quantitative Trait Loci (QTL) and Candidate Genes for Cadmium Tolerance in Populus

    Energy Technology Data Exchange (ETDEWEB)

    Induri, Brahma R [West Virginia University; Ellis, Danielle R [West Virginia University; Slavov, Gancho [West Virginia University; Yin, Tongming [ORNL; Muchero, Wellington [ORNL; Tuskan, Gerald A [ORNL; DiFazio, Stephen P [West Virginia University

    2012-01-01

    Knowledge of genetic variation in response of Populus to heavy metals like cadmium (Cd) is an important step in understanding the underlying mechanisms of tolerance. In this study, a pseudo-backcross pedigree of Populus trichocarpa and Populus deltoides was characterized for Cd exposure. The pedigree showed significant variation for Cd tolerance thus enabling the identification of relatively tolerant and susceptible genotypes for intensive characterization. A total of 16 QTLs at logarithm of odds (LOD) ratio > 2.5, were found to be associated with total dry weight, its components, and root volume. Four major QTLs for total dry weight were mapped to different linkage groups in control (LG III) and Cd conditions (LG XVI) and had opposite allelic effects on Cd tolerance, suggesting that these genomic regions were differentially controlled. The phenotypic variation explained by Cd QTL for all traits under study varied from 5.9% to 11.6% and averaged 8.2% across all QTL. Leaf Cd contents also showed significant variation suggesting the phytoextraction potential of Populus genotypes, though heritability of this trait was low (0.22). A whole-genome microarray study was conducted by using two genotypes with extreme responses for Cd tolerance in the above study and differentially expressed genes were identified. Candidate genes including CAD2 (CADMIUM SENSITIVE 2), HMA5 (HEAVY METAL ATPase5), ATGTST1 (Arabidopsis thaliana Glutathione S-Transferase1), ATGPX6 (Glutathione peroxidase 6), and ATMRP 14 (Arabidopsis thaliana Multidrug Resistance associated Protein 14) were identified from QTL intervals and microarray study. Functional characterization of these candidate genes could enhance phytoremediation capabilities of Populus.

  2. Aldehyde Dehydrogenase Gene Superfamily in Populus: Organization and Expression Divergence between Paralogous Gene Pairs.

    Science.gov (United States)

    Tian, Feng-Xia; Zang, Jian-Lei; Wang, Tan; Xie, Yu-Li; Zhang, Jin; Hu, Jian-Jun

    2015-01-01

    Aldehyde dehydrogenases (ALDHs) constitute a superfamily of NAD(P)+-dependent enzymes that catalyze the irreversible oxidation of a wide range of reactive aldehydes to their corresponding nontoxic carboxylic acids. ALDHs have been studied in many organisms from bacteria to mammals; however, no systematic analyses incorporating genome organization, gene structure, expression profiles, and cis-acting elements have been conducted in the model tree species Populus trichocarpa thus far. In this study, a comprehensive analysis of the Populus ALDH gene superfamily was performed. A total of 26 Populus ALDH genes were found to be distributed across 12 chromosomes. Genomic organization analysis indicated that purifying selection may have played a pivotal role in the retention and maintenance of PtALDH gene families. The exon-intron organizations of PtALDHs were highly conserved within the same family, suggesting that the members of the same family also may have conserved functionalities. Microarray data and qRT-PCR analysis indicated that most PtALDHs had distinct tissue-specific expression patterns. The specificity of cis-acting elements in the promoter regions of the PtALDHs and the divergence of expression patterns between nine paralogous PtALDH gene pairs suggested that gene duplications may have freed the duplicate genes from the functional constraints. The expression levels of some ALDHs were up- or down-regulated by various abiotic stresses, implying that the products of these genes may be involved in the adaptation of Populus to abiotic stresses. Overall, the data obtained from our investigation contribute to a better understanding of the complexity of the Populus ALDH gene superfamily and provide insights into the function and evolution of ALDH gene families in vascular plants.

  3. Aldehyde Dehydrogenase Gene Superfamily in Populus: Organization and Expression Divergence between Paralogous Gene Pairs.

    Directory of Open Access Journals (Sweden)

    Feng-Xia Tian

    Full Text Available Aldehyde dehydrogenases (ALDHs constitute a superfamily of NAD(P+-dependent enzymes that catalyze the irreversible oxidation of a wide range of reactive aldehydes to their corresponding nontoxic carboxylic acids. ALDHs have been studied in many organisms from bacteria to mammals; however, no systematic analyses incorporating genome organization, gene structure, expression profiles, and cis-acting elements have been conducted in the model tree species Populus trichocarpa thus far. In this study, a comprehensive analysis of the Populus ALDH gene superfamily was performed. A total of 26 Populus ALDH genes were found to be distributed across 12 chromosomes. Genomic organization analysis indicated that purifying selection may have played a pivotal role in the retention and maintenance of PtALDH gene families. The exon-intron organizations of PtALDHs were highly conserved within the same family, suggesting that the members of the same family also may have conserved functionalities. Microarray data and qRT-PCR analysis indicated that most PtALDHs had distinct tissue-specific expression patterns. The specificity of cis-acting elements in the promoter regions of the PtALDHs and the divergence of expression patterns between nine paralogous PtALDH gene pairs suggested that gene duplications may have freed the duplicate genes from the functional constraints. The expression levels of some ALDHs were up- or down-regulated by various abiotic stresses, implying that the products of these genes may be involved in the adaptation of Populus to abiotic stresses. Overall, the data obtained from our investigation contribute to a better understanding of the complexity of the Populus ALDH gene superfamily and provide insights into the function and evolution of ALDH gene families in vascular plants.

  4. Genome-wide Identification of WRKY Genes in the Desert Poplar Populus euphratica and Adaptive Evolution of the Genes in Response to Salt Stress.

    Science.gov (United States)

    Ma, Jianchao; Lu, Jing; Xu, Jianmei; Duan, Bingbing; He, Xiaodong; Liu, Jianquan

    2015-01-01

    WRKY transcription factors play important roles in plant development and responses to various stresses in plants. However, little is known about the evolution of the WRKY genes in the desert poplar species Populus euphratica, which is highly tolerant of salt stress. In this study, we identified 107 PeWRKY genes from the P. euphratica genome and examined their evolutionary relationships with the WRKY genes of the salt-sensitive congener Populus trichocarpa. Ten PeWRKY genes are specific to P. euphratica, and five of these showed altered expression under salt stress. Furthermore, we found that two pairs of orthologs between the two species showed evidence of positive evolution, with dN/dS ratios>1 (nonsynonymous/synonymous substitutions), and both of them altered their expression in response to salinity stress. These findings suggested that both the development of new genes and positive evolution in some orthologs of the WRKY gene family may have played an important role in the acquisition of high salt tolerance by P. euphratica.

  5. A microarray-based genotyping and genetic mapping approach for highly heterozygous outcrossing species enables localization of a large fraction of the unassembled Populus trichocarpa genome sequence.

    Science.gov (United States)

    Drost, Derek R; Novaes, Evandro; Boaventura-Novaes, Carolina; Benedict, Catherine I; Brown, Ryan S; Yin, Tongming; Tuskan, Gerald A; Kirst, Matias

    2009-06-01

    Microarrays have demonstrated significant power for genome-wide analyses of gene expression, and recently have also revolutionized the genetic analysis of segregating populations by genotyping thousands of loci in a single assay. Although microarray-based genotyping approaches have been successfully applied in yeast and several inbred plant species, their power has not been proven in an outcrossing species with extensive genetic diversity. Here we have developed methods for high-throughput microarray-based genotyping in such species using a pseudo-backcross progeny of 154 individuals of Populus trichocarpa and P. deltoides analyzed with long-oligonucleotide in situ-synthesized microarray probes. Our analysis resulted in high-confidence genotypes for 719 single-feature polymorphism (SFP) and 1014 gene expression marker (GEM) candidates. Using these genotypes and an established microsatellite (SSR) framework map, we produced a high-density genetic map comprising over 600 SFPs, GEMs and SSRs. The abundance of gene-based markers allowed us to localize over 35 million base pairs of previously unplaced whole-genome shotgun (WGS) scaffold sequence to putative locations in the genome of P. trichocarpa. A high proportion of sampled scaffolds could be verified for their placement with independently mapped SSRs, demonstrating the previously un-utilized power that high-density genotyping can provide in the context of map-based WGS sequence reassembly. Our results provide a substantial contribution to the continued improvement of the Populus genome assembly, while demonstrating the feasibility of microarray-based genotyping in a highly heterozygous population. The strategies presented are applicable to genetic mapping efforts in all plant species with similarly high levels of genetic diversity.

  6. Genomic research in Eucalyptus.

    Science.gov (United States)

    Poke, Fiona S; Vaillancourt, René E; Potts, Brad M; Reid, James B

    2005-09-01

    Eucalyptus L'Hérit. is a genus comprised of more than 700 species that is of vital importance ecologically to Australia and to the forestry industry world-wide, being grown in plantations for the production of solid wood products as well as pulp for paper. With the sequencing of the genomes of Arabidopsis thaliana and Oryza sativa and the recent completion of the first tree genome sequence, Populus trichocarpa, attention has turned to the current status of genomic research in Eucalyptus. For several eucalypt species, large segregating families have been established, high-resolution genetic maps constructed and large EST databases generated. Collaborative efforts have been initiated for the integration of diverse genomic projects and will provide the framework for future research including exploiting the sequence of the entire eucalypt genome which is currently being sequenced. This review summarises the current position of genomic research in Eucalyptus and discusses the direction of future research.

  7. Diversification and expression of the PIN, AUX/LAX and ABCB families of putative auxin transporters in Populus

    Directory of Open Access Journals (Sweden)

    Nicola eCarraro

    2012-02-01

    Full Text Available Intercellular transport of the plant hormone auxin is mediated by three families of membrane-bound protein carriers, with the PIN and ABCB families coding primarily for efflux proteins and the AUX/LAX family coding for influx proteins. In the last decade our understanding of gene and protein function for these transporters in Arabidopsis has expanded rapidly but very little is known about their role in woody plant development. Here we present a comprehensive account of all three families in the model woody species Populus, including chromosome distribution, protein structure, quantitative gene expression, and evolutionary relationships. The PIN and AUX/LAX gene families in Populus comprise 16 and 8 members respectively, and show evidence for the retention of paralogs following a relatively recent whole genome duplication. There is also evidence for differential expression across tissues within many gene pairs. The ABCB family is previously undescribed in Populus and includes 20 members, showing a much deeper evolutionary history including both tandem and whole genome duplication as well as probable loss. A striking number of these transporters are expressed in developing Populus stems and we suggest that evolutionary and structural relationships with known auxin transporters in Arabidopsis can point toward candidate genes for further study in Populus. This is especially important for the ABCBs, which is a large family and includes members in Arabidopsis that are able to transport other substrates in addition to auxin. Protein modeling, sequence alignment and expression data all point to ABCB1.1 as a likely auxin transport protein in Populus. Given that basipetal auxin flow through the cambial zone shapes the development of woody stems, it is important that we identify the full complement of proteins involved in this process. This work should lay the foundation for studies targeting specific proteins for functional characterization and in situ

  8. Statistical Methods in Integrative Genomics

    Science.gov (United States)

    Richardson, Sylvia; Tseng, George C.; Sun, Wei

    2016-01-01

    Statistical methods in integrative genomics aim to answer important biology questions by jointly analyzing multiple types of genomic data (vertical integration) or aggregating the same type of data across multiple studies (horizontal integration). In this article, we introduce different types of genomic data and data resources, and then review statistical methods of integrative genomics, with emphasis on the motivation and rationale of these methods. We conclude with some summary points and future research directions. PMID:27482531

  9. Population Genomics of Infectious and Integrated Wolbachia pipientis Genomes in Drosophila ananassae

    Science.gov (United States)

    Choi, Jae Young; Bubnell, Jaclyn E.; Aquadro, Charles F.

    2015-01-01

    Coevolution between Drosophila and its endosymbiont Wolbachia pipientis has many intriguing aspects. For example, Drosophila ananassae hosts two forms of W. pipientis genomes: One being the infectious bacterial genome and the other integrated into the host nuclear genome. Here, we characterize the infectious and integrated genomes of W. pipientis infecting D. ananassae (wAna), by genome sequencing 15 strains of D. ananassae that have either the infectious or integrated wAna genomes. Results indicate evolutionarily stable maternal transmission for the infectious wAna genome suggesting a relatively long-term coevolution with its host. In contrast, the integrated wAna genome showed pseudogene-like characteristics accumulating many variants that are predicted to have deleterious effects if present in an infectious bacterial genome. Phylogenomic analysis of sequence variation together with genotyping by polymerase chain reaction of large structural variations indicated several wAna variants among the eight infectious wAna genomes. In contrast, only a single wAna variant was found among the seven integrated wAna genomes examined in lines from Africa, south Asia, and south Pacific islands suggesting that the integration occurred once from a single infectious wAna genome and then spread geographically. Further analysis revealed that for all D. ananassae we examined with the integrated wAna genomes, the majority of the integrated wAna genomic regions is represented in at least two copies suggesting a double integration or single integration followed by an integrated genome duplication. The possible evolutionary mechanism underlying the widespread geographical presence of the duplicate integration of the wAna genome is an intriguing question remaining to be answered. PMID:26254486

  10. Characterization of genomic sequence showing strong association with polyembryony among diverse Citrus species and cultivars, and its synteny with Vitis and Populus.

    Science.gov (United States)

    Nakano, Michiharu; Shimada, Takehiko; Endo, Tomoko; Fujii, Hiroshi; Nesumi, Hirohisa; Kita, Masayuki; Ebina, Masumi; Shimizu, Tokurou; Omura, Mitsuo

    2012-02-01

    Polyembryony, in which multiple somatic nucellar cell-derived embryos develop in addition to the zygotic embryo in a seed, is common in the genus Citrus. Previous genetic studies indicated polyembryony is mainly determined by a single locus, but the underlying molecular mechanism is still unclear. As a step towards identification and characterization of the gene or genes responsible for nucellar embryogenesis in Citrus, haplotype-specific physical maps around the polyembryony locus were constructed. By sequencing three BAC clones aligned on the polyembryony haplotype, a single contiguous draft sequence consisting of 380 kb containing 70 predicted open reading frames (ORFs) was reconstructed. Single nucleotide polymorphism genotypes detected in the sequenced genomic region showed strong association with embryo type in Citrus, indicating a common polyembryony locus is shared among widely diverse Citrus cultivars and species. The arrangement of the predicted ORFs in the characterized genomic region showed high collinearity to the genomic sequence of chromosome 4 of Vitis vinifera and linkage group VI of Populus trichocarpa, suggesting that the syntenic relationship among these species is conserved even though V. vinifera and P. trichocarpa are non-apomictic species. This is the first study to characterize in detail the genomic structure of an apomixis locus determining adventitious embryony. Copyright © 2011 Elsevier Ireland Ltd. All rights reserved.

  11. Epigenomics of Development in Populus

    Energy Technology Data Exchange (ETDEWEB)

    Strauss, Steve; Freitag, Michael; Mockler, Todd

    2013-01-10

    We conducted research to determine the role of epigenetic modifications during tree development using poplar (Populus trichocarpa), a model woody feedstock species. Using methylated DNA immunoprecipitation (MeDIP) or chromatin immunoprecipitation (ChIP), followed by high-throughput sequencing, we are analyzed DNA and histone methylation patterns in the P. trichocarpa genome in relation to four biological processes: bud dormancy and release, mature organ maintenance, in vitro organogenesis, and methylation suppression. Our project is now completed. We have 1) produced 22 transgenic events for a gene involved in DNA methylation suppression and studied its phenotypic consequences; 2) completed sequencing of methylated DNA from eleven target tissues in wildtype P. trichocarpa; 3) updated our customized poplar genome browser using the open-source software tools (2.13) and (V2.2) of the P. trichocarpa genome; 4) produced summary data for genome methylation in P. trichocarpa, including distribution of methylation across chromosomes and in and around genes; 5) employed bioinformatic and statistical methods to analyze differences in methylation patterns among tissue types; and 6) used bisulfite sequencing of selected target genes to confirm bioinformatics and sequencing results, and gain a higher-resolution view of methylation at selected genes 7) compared methylation patterns to expression using available microarray data. Our main findings of biological significance are the identification of extensive regions of the genome that display developmental variation in DNA methylation; highly distinctive gene-associated methylation profiles in reproductive tissues, particularly male catkins; a strong whole genome/all tissue inverse association of methylation at gene bodies and promoters with gene expression; a lack of evidence that tissue specificity of gene expression is associated with gene methylation; and evidence that genome methylation is a significant impediment to tissue

  12. Safeguarding genome integrity

    DEFF Research Database (Denmark)

    Sørensen, Claus Storgaard; Syljuåsen, Randi G

    2012-01-01

    Mechanisms that preserve genome integrity are highly important during the normal life cycle of human cells. Loss of genome protective mechanisms can lead to the development of diseases such as cancer. Checkpoint kinases function in the cellular surveillance pathways that help cells to cope with D...

  13. Wood property variation in Populus

    Science.gov (United States)

    Dean W. Einspahr; Miles K. Benson; John R. Peckham

    1968-01-01

    The use of bigtooth aspen (Populus grandidentata Michx.), quaking aspen (P. tremuloides Michx.), and cottonwood (P. deltoides Bartr.) by the pulp and paper industry has increased greatly during the past decade. This expanded use has stimulated research on the genetic improvement of Populus. For the past 12 years...

  14. Genomics Portals: integrative web-platform for mining genomics data.

    Science.gov (United States)

    Shinde, Kaustubh; Phatak, Mukta; Johannes, Freudenberg M; Chen, Jing; Li, Qian; Vineet, Joshi K; Hu, Zhen; Ghosh, Krishnendu; Meller, Jaroslaw; Medvedovic, Mario

    2010-01-13

    A large amount of experimental data generated by modern high-throughput technologies is available through various public repositories. Our knowledge about molecular interaction networks, functional biological pathways and transcriptional regulatory modules is rapidly expanding, and is being organized in lists of functionally related genes. Jointly, these two sources of information hold a tremendous potential for gaining new insights into functioning of living systems. Genomics Portals platform integrates access to an extensive knowledge base and a large database of human, mouse, and rat genomics data with basic analytical visualization tools. It provides the context for analyzing and interpreting new experimental data and the tool for effective mining of a large number of publicly available genomics datasets stored in the back-end databases. The uniqueness of this platform lies in the volume and the diversity of genomics data that can be accessed and analyzed (gene expression, ChIP-chip, ChIP-seq, epigenomics, computationally predicted binding sites, etc), and the integration with an extensive knowledge base that can be used in such analysis. The integrated access to primary genomics data, functional knowledge and analytical tools makes Genomics Portals platform a unique tool for interpreting results of new genomics experiments and for mining the vast amount of data stored in the Genomics Portals backend databases. Genomics Portals can be accessed and used freely at http://GenomicsPortals.org.

  15. Genome-wide patterns of differentiation and spatially varying selection between postglacial recolonization lineages of Populus alba (Salicaceae), a widespread forest tree.

    Science.gov (United States)

    Stölting, Kai N; Paris, Margot; Meier, Cécile; Heinze, Berthold; Castiglione, Stefano; Bartha, Denes; Lexer, Christian

    2015-08-01

    Studying the divergence continuum in plants is relevant to fundamental and applied biology because of the potential to reveal functionally important genetic variation. In this context, whole-genome sequencing (WGS) provides the necessary rigour for uncovering footprints of selection. We resequenced populations of two divergent phylogeographic lineages of Populus alba (n = 48), thoroughly characterized by microsatellites (n = 317), and scanned their genomes for regions of unusually high allelic differentiation and reduced diversity using > 1.7 million single nucleotide polymorphisms (SNPs) from WGS. Results were confirmed by Sanger sequencing. On average, 9134 high-differentiation (≥ 4 standard deviations) outlier SNPs were uncovered between populations, 848 of which were shared by ≥ three replicate comparisons. Annotation revealed that 545 of these were located in 437 predicted genes. Twelve percent of differentiation outlier genome regions exhibited significantly reduced genetic diversity. Gene ontology (GO) searches were successful for 327 high-differentiation genes, and these were enriched for 63 GO terms. Our results provide a snapshot of the roles of 'hard selective sweeps' vs divergent selection of standing genetic variation in distinct postglacial recolonization lineages of P. alba. Thus, this study adds to our understanding of the mechanisms responsible for the origin of functionally relevant variation in temperate trees. © 2015 The Authors. New Phytologist © 2015 New Phytologist Trust.

  16. Genomics Portals: integrative web-platform for mining genomics data

    Directory of Open Access Journals (Sweden)

    Ghosh Krishnendu

    2010-01-01

    Full Text Available Abstract Background A large amount of experimental data generated by modern high-throughput technologies is available through various public repositories. Our knowledge about molecular interaction networks, functional biological pathways and transcriptional regulatory modules is rapidly expanding, and is being organized in lists of functionally related genes. Jointly, these two sources of information hold a tremendous potential for gaining new insights into functioning of living systems. Results Genomics Portals platform integrates access to an extensive knowledge base and a large database of human, mouse, and rat genomics data with basic analytical visualization tools. It provides the context for analyzing and interpreting new experimental data and the tool for effective mining of a large number of publicly available genomics datasets stored in the back-end databases. The uniqueness of this platform lies in the volume and the diversity of genomics data that can be accessed and analyzed (gene expression, ChIP-chip, ChIP-seq, epigenomics, computationally predicted binding sites, etc, and the integration with an extensive knowledge base that can be used in such analysis. Conclusion The integrated access to primary genomics data, functional knowledge and analytical tools makes Genomics Portals platform a unique tool for interpreting results of new genomics experiments and for mining the vast amount of data stored in the Genomics Portals backend databases. Genomics Portals can be accessed and used freely at http://GenomicsPortals.org.

  17. Genetic analysis of post-mating reproductive barriers in hybridizing European Populus species.

    Science.gov (United States)

    Macaya-Sanz, D; Suter, L; Joseph, J; Barbará, T; Alba, N; González-Martínez, S C; Widmer, A; Lexer, C

    2011-10-01

    Molecular genetic analyses of experimental crosses provide important information on the strength and nature of post-mating barriers to gene exchange between divergent populations, which are topics of great interest to evolutionary geneticists and breeders. Although not a trivial task in long-lived organisms such as trees, experimental interspecific recombinants can sometimes be created through controlled crosses involving natural F(1)'s. Here, we used this approach to understand the genetics of post-mating isolation and barriers to introgression in Populus alba and Populus tremula, two ecologically divergent, hybridizing forest trees. We studied 86 interspecific backcross (BC(1)) progeny and >350 individuals from natural populations of these species for up to 98 nuclear genetic markers, including microsatellites, indels and single nucleotide polymorphisms, and inferred the origin of the cytoplasm of the cross with plastid DNA. Genetic analysis of the BC(1) revealed extensive segregation distortions on six chromosomes, and >90% of these (12 out of 13) favored P. tremula donor alleles in the heterospecific genomic background. Since selection was documented during early diploid stages of the progeny, this surprising result was attributed to epistasis, cyto-nuclear coadaptation, heterozygote advantage at nuclear loci experiencing introgression or a combination of these. Our results indicate that gene flow across 'porous' species barriers affects these poplars and aspens beyond neutral, Mendelian expectations and suggests the mechanisms responsible. Contrary to expectations, the Populus sex determination region is not protected from introgression. Understanding the population dynamics of the Populus sex determination region will require tests based on natural interspecific hybrid zones.

  18. Differentiation of Populus species using chloroplast single nucleotide polymorphism (SNP) markers--essential for comprehensible and reliable poplar breeding.

    Science.gov (United States)

    Schroeder, H; Hoeltken, A M; Fladung, M

    2012-03-01

    Within the genus Populus several species belonging to different sections are cross-compatible. Hence, high numbers of interspecies hybrids occur naturally and, additionally, have been artificially produced in huge breeding programmes during the last 100 years. Therefore, determination of a single poplar species, used for the production of 'multi-species hybrids' is often difficult, and represents a great challenge for the use of molecular markers in species identification. Within this study, over 20 chloroplast regions, both intergenic spacers and coding regions, have been tested for their ability to differentiate different poplar species using 23 already published barcoding primer combinations and 17 newly designed primer combinations. About half of the published barcoding primers yielded amplification products, whereas the new primers designed on the basis of the total sequenced cpDNA genome of Populus trichocarpa Torr. & Gray yielded much higher amplification success. Intergenic spacers were found to be more variable than coding regions within the genus Populus. The highest discrimination power of Populus species was found in the combination of two intergenic spacers (trnG-psbK, psbK-psbl) and the coding region rpoC. In barcoding projects, the coding regions matK and rbcL are often recommended, but within the genus Populus they only show moderate variability and are not efficient in species discrimination. © 2011 German Botanical Society and The Royal Botanical Society of the Netherlands.

  19. Growth of Populus and Salix Species under Compost Leachate Irrigation

    OpenAIRE

    Tooba Abedi; Shamim Moghaddami; Ebrahim Lashkar Bolouki

    2014-01-01

    According to the known broad variation in remediation capacity, three plant species were used in the experiment: two fast growing poplar’s clones - Populus deltoides, Populus euramericana, and willows Salix alba. Populus and Salix cuttings were collected from the nursery of the Populus Research Center of Safrabasteh in the eastern part of Guilan province at north of Iran. The Populus clones were chosen because of their high biomass production capacity and willow- because it is native in Iran....

  20. Comparative genomic sequence analysis of strawberry and other rosids reveals significant microsynteny

    Directory of Open Access Journals (Sweden)

    Abbott Albert

    2010-06-01

    Full Text Available Abstract Background Fragaria belongs to the Rosaceae, an economically important family that includes a number of important fruit producing genera such as Malus and Prunus. Using genomic sequences from 50 Fragaria fosmids, we have examined the microsynteny between Fragaria and other plant models. Results In more than half of the strawberry fosmids, we found syntenic regions that are conserved in Populus, Vitis, Medicago and/or Arabidopsis with Populus containing the greatest number of syntenic regions with Fragaria. The longest syntenic region was between LG VIII of the poplar genome and the strawberry fosmid 72E18, where seven out of twelve predicted genes were collinear. We also observed an unexpectedly high level of conserved synteny between Fragaria (rosid I and Vitis (basal rosid. One of the strawberry fosmids, 34E24, contained a cluster of R gene analogs (RGAs with NBS and LRR domains. We detected clusters of RGAs with high sequence similarity to those in 34E24 in all the genomes compared. In the phylogenetic tree we have generated, all the NBS-LRR genes grouped together with Arabidopsis CNL-A type NBS-LRR genes. The Fragaria RGA grouped together with those of Vitis and Populus in the phylogenetic tree. Conclusions Our analysis shows considerable microsynteny between Fragaria and other plant genomes such as Populus, Medicago, Vitis, and Arabidopsis to a lesser degree. We also detected a cluster of NBS-LRR type genes that are conserved in all the genomes compared.

  1. Production costs for SRIC Populus biomass

    International Nuclear Information System (INIS)

    Strauss, C.H.

    1991-01-01

    Production costs for short rotation, intensive culture (SRIC) Populus biomass were developed from commercial-sized plantations under investigation throughout the US. Populus hybrid planted on good quality agricultural sites at a density of 850 cuttings/acre was projected to yield an average of 7 ovendry (OD) tons/acre/year. Discounted cash-flow analysis of multiple rotations showed preharvest production costs of $14/ton (OD). Harvesting and transportation expenses would increase the delivered cost to $35/ton (OD). Although this total cost compared favorably with the regional market price for aspen (Populus tremuloides), future investments in SRIC systems will require the development of biomass energy markets

  2. phiGENOME: an integrative navigation throughout bacteriophage genomes.

    Science.gov (United States)

    Stano, Matej; Klucar, Lubos

    2011-11-01

    phiGENOME is a web-based genome browser generating dynamic and interactive graphical representation of phage genomes stored in the phiSITE, database of gene regulation in bacteriophages. phiGENOME is an integral part of the phiSITE web portal (http://www.phisite.org/phigenome) and it was optimised for visualisation of phage genomes with the emphasis on the gene regulatory elements. phiGENOME consists of three components: (i) genome map viewer built using Adobe Flash technology, providing dynamic and interactive graphical display of phage genomes; (ii) sequence browser based on precisely formatted HTML tags, providing detailed exploration of genome features on the sequence level and (iii) regulation illustrator, based on Scalable Vector Graphics (SVG) and designed for graphical representation of gene regulations. Bringing 542 complete genome sequences accompanied with their rich annotations and references, makes phiGENOME a unique information resource in the field of phage genomics. Copyright © 2011 Elsevier Inc. All rights reserved.

  3. MycoCosm, an Integrated Fungal Genomics Resource

    Energy Technology Data Exchange (ETDEWEB)

    Shabalov, Igor; Grigoriev, Igor

    2012-03-16

    MycoCosm is a web-based interactive fungal genomics resource, which was first released in March 2010, in response to an urgent call from the fungal community for integration of all fungal genomes and analytical tools in one place (Pan-fungal data resources meeting, Feb 21-22, 2010, Alexandria, VA). MycoCosm integrates genomics data and analysis tools to navigate through over 100 fungal genomes sequenced at JGI and elsewhere. This resource allows users to explore fungal genomes in the context of both genome-centric analysis and comparative genomics, and promotes user community participation in data submission, annotation and analysis. MycoCosm has over 4500 unique visitors/month or 35000+ visitors/year as well as hundreds of registered users contributing their data and expertise to this resource. Its scalable architecture allows significant expansion of the data expected from JGI Fungal Genomics Program, its users, and integration with external resources used by fungal community.

  4. SIGMA: A System for Integrative Genomic Microarray Analysis of Cancer Genomes

    Directory of Open Access Journals (Sweden)

    Davies Jonathan J

    2006-12-01

    Full Text Available Abstract Background The prevalence of high resolution profiling of genomes has created a need for the integrative analysis of information generated from multiple methodologies and platforms. Although the majority of data in the public domain are gene expression profiles, and expression analysis software are available, the increase of array CGH studies has enabled integration of high throughput genomic and gene expression datasets. However, tools for direct mining and analysis of array CGH data are limited. Hence, there is a great need for analytical and display software tailored to cross platform integrative analysis of cancer genomes. Results We have created a user-friendly java application to facilitate sophisticated visualization and analysis such as cross-tumor and cross-platform comparisons. To demonstrate the utility of this software, we assembled array CGH data representing Affymetrix SNP chip, Stanford cDNA arrays and whole genome tiling path array platforms for cross comparison. This cancer genome database contains 267 profiles from commonly used cancer cell lines representing 14 different tissue types. Conclusion In this study we have developed an application for the visualization and analysis of data from high resolution array CGH platforms that can be adapted for analysis of multiple types of high throughput genomic datasets. Furthermore, we invite researchers using array CGH technology to deposit both their raw and processed data, as this will be a continually expanding database of cancer genomes. This publicly available resource, the System for Integrative Genomic Microarray Analysis (SIGMA of cancer genomes, can be accessed at http://sigma.bccrc.ca.

  5. VEGETATIVNO RAZMNOŽEVANJE TOPOLOV (Populus spp.) S POTAKNJENCI

    OpenAIRE

    Sternad, Rebeka

    2010-01-01

    Raziskava je bila opravljena na treh vrstah topolov: črnem topolu (Populus nigra L.), belem topolu (Populus alba L.) in trepetliki (Populus tremula L.). Namen diplomskega dela je bil proučiti ukoreninjenje potaknjencev topolov glede na okolje koreninjenja, termin potikanja in vrsto uporabljenega substrata. Zeleni in pololeseneli potaknjenci so bili večinoma nabrani v okolici celjske regije in Žalca. Koreninjeni so bili v šestih terminih (od junija do septembra) v okolju meglenja in neposredne...

  6. Subcellular localization of cadmium in hyperaccumulator Populus ...

    African Journals Online (AJOL)

    In this study, subcellular localization of cadmium in hyperaccumulator grey poplar (Populus × canescens) was investigated by the transmission electron microscopy (TEM) method. Young Populus × canescens were grown and hydroponic experiments were conducted under four Cd2+ concentrations (10, 30, 50, and 70 μM) ...

  7. Sex-Specific Response to Stress in Populus

    Directory of Open Access Journals (Sweden)

    Nataliya V. Melnikova

    2017-10-01

    Full Text Available Populus is an effective model for genetic studies in trees. The genus Populus includes dioecious species, and the differences exhibited in males and females have been intensively studied. This review focused on the distinctions between male and female poplar and aspen plants under stress conditions, such as drought, salinity, heavy metals, and nutrient deficiency on morphological, physiological, proteome, and gene expression levels. In most studies, males of Populus species were more adaptive to the majority of the stress conditions and showed less damage, better growth, and higher photosynthetic capacity and antioxidant activity than that of the females. However, in two recent studies, no differences in non-reproductive traits were revealed for male and female trees. This discrepancy of the results could be associated with experimental design: different species and genotypes, stress conditions, types of plant materials, sampling sizes. Knowledge of sex-specific differences is crucial for basic and applied research in Populus species.

  8. Allelopathic potential of populus euphratica olivier

    International Nuclear Information System (INIS)

    Sher, Z.; Hussain, F.; Ahmad, B.; Wahab, M.

    2011-01-01

    Populus euphratica Olivier is frequently cultivated deciduous tree in Pakistan on agricultural land for its shade, fodder, timber and fuel wood. A relatively reduced under storey is often observed below it. Therefore the present study was conducted to assess the allelopathic potential of Populus euphratica against some crop species. Plant material of Populus euphratica were collected from the agriculture fields of Lahor, District Swabi in 2008 and were dried at room temperature (258 deg. C-308 deg. C). Allelopathic studies conducted by using aqueous extracts from various parts including young leaves, mature leaves, bark, litter and mulching in various experiments invariably retarded the germination, plumule, radical growth, fresh and dry weight of Sorghum vulgare Perse, Setaria italica (L.) P. Beauv and Triticum aestivum L., in laboratory experiments. The aqueous extracts obtained after 48 h were more inhibitory than 24 h. Leaves were more toxic than bark. Litter and mulching experiments also proved to be inhibitory. It is suggested that the various assayed parts of Populus euphratica have strong allelopathic potential at least against the tested species. Further investigation is required to see its allelopathic behavior under field condition against its associated species and to identify the toxic principles. (author)

  9. Integrative Genomics Viewer (IGV): high-performance genomics data visualization and exploration.

    Science.gov (United States)

    Thorvaldsdóttir, Helga; Robinson, James T; Mesirov, Jill P

    2013-03-01

    Data visualization is an essential component of genomic data analysis. However, the size and diversity of the data sets produced by today's sequencing and array-based profiling methods present major challenges to visualization tools. The Integrative Genomics Viewer (IGV) is a high-performance viewer that efficiently handles large heterogeneous data sets, while providing a smooth and intuitive user experience at all levels of genome resolution. A key characteristic of IGV is its focus on the integrative nature of genomic studies, with support for both array-based and next-generation sequencing data, and the integration of clinical and phenotypic data. Although IGV is often used to view genomic data from public sources, its primary emphasis is to support researchers who wish to visualize and explore their own data sets or those from colleagues. To that end, IGV supports flexible loading of local and remote data sets, and is optimized to provide high-performance data visualization and exploration on standard desktop systems. IGV is freely available for download from http://www.broadinstitute.org/igv, under a GNU LGPL open-source license.

  10. Increasing the productivity of short-rotation Populus plantations. Final report

    Energy Technology Data Exchange (ETDEWEB)

    DeBell, D.S.; Harrington, C.A.; Clendenen, G.W.; Radwan, M.A.; Zasada, J.C. [Forest Service, Olympia, WA (United States). Pacific Northwest Research Station

    1997-12-31

    This final report represents the culmination of eight years of biological research devoted to increasing the productivity of short rotation plantations of Populus trichocarpa and Populus hybrids in the Pacific Northwest. Studies provide an understanding of tree growth, stand development and biomass yield at various spacings, and how patterns differ by Populus clone in monoclonal and polyclonal plantings. Also included is some information about factors related to wind damage in Populus plantings, use of leaf size as a predictor of growth potential, and approaches for estimating tree and stand biomass and biomass growth. Seven research papers are included which provide detailed methods, results, and interpretations on these topics.

  11. Integrating cancer genomic data into electronic health records

    Directory of Open Access Journals (Sweden)

    Jeremy L. Warner

    2016-10-01

    Full Text Available Abstract The rise of genomically targeted therapies and immunotherapy has revolutionized the practice of oncology in the last 10–15 years. At the same time, new technologies and the electronic health record (EHR in particular have permeated the oncology clinic. Initially designed as billing and clinical documentation systems, EHR systems have not anticipated the complexity and variety of genomic information that needs to be reviewed, interpreted, and acted upon on a daily basis. Improved integration of cancer genomic data with EHR systems will help guide clinician decision making, support secondary uses, and ultimately improve patient care within oncology clinics. Some of the key factors relating to the challenge of integrating cancer genomic data into EHRs include: the bioinformatics pipelines that translate raw genomic data into meaningful, actionable results; the role of human curation in the interpretation of variant calls; and the need for consistent standards with regard to genomic and clinical data. Several emerging paradigms for integration are discussed in this review, including: non-standardized efforts between individual institutions and genomic testing laboratories; “middleware” products that portray genomic information, albeit outside of the clinical workflow; and application programming interfaces that have the potential to work within clinical workflow. The critical need for clinical-genomic knowledge bases, which can be independent or integrated into the aforementioned solutions, is also discussed.

  12. Identification and characterization of nuclear genes involved in photosynthesis in Populus

    Science.gov (United States)

    2014-01-01

    Background The gap between the real and potential photosynthetic rate under field conditions suggests that photosynthesis could potentially be improved. Nuclear genes provide possible targets for improving photosynthetic efficiency. Hence, genome-wide identification and characterization of the nuclear genes affecting photosynthetic traits in woody plants would provide key insights on genetic regulation of photosynthesis and identify candidate processes for improvement of photosynthesis. Results Using microarray and bulked segregant analysis strategies, we identified differentially expressed nuclear genes for photosynthesis traits in a segregating population of poplar. We identified 515 differentially expressed genes in this population (FC ≥ 2 or FC ≤ 0.5, P photosynthesis by the nuclear genome mainly involves transport, metabolism and response to stimulus functions. Conclusions This study provides new genome-scale strategies for the discovery of potential candidate genes affecting photosynthesis in Populus, and for identification of the functions of genes involved in regulation of photosynthesis. This work also suggests that improving photosynthetic efficiency under field conditions will require the consideration of multiple factors, such as stress responses. PMID:24673936

  13. Laboratory-scale measurements of N2O and CH4 emissions from hybrid poplars (Populus deltoides x Populus nigra).

    Science.gov (United States)

    McBain, M C; Warland, J S; McBride, R A; Wagner-Riddle, C

    2004-12-01

    The purpose of this study was to determine whether or not young hybrid poplar (Populus deltoides x Populus nigra) could transport landfill biogas internally from the root zone to the atmosphere, thereby acting as conduits for landfill gas release. Fluxes of methane (CH4) and nitrous oxide (N2O) from the seedlings to the atmosphere were measured under controlled conditions using dynamic flux chambers and a tunable diode laser trace gas analyser (TDLTGA). Nitrous oxide was emitted from the seedlings, but only when extremely high soil N2O concentrations were applied to the root zone. In contrast, no detectable emissions of CH4 were measured in a similar experimental trial. Visible plant morphological responses, characteristic of flood-tolerant trees attempting to cope with the negative effects of soil hypoxia, were observed during the CH4 experiments. Leaf chlorosis, leaf abscission and adventitious roots were all visible plant responses. In addition, seedling survival was observed to be highest in the biogas 'hot spot' areas of a local municipal solid waste landfill involved in this study. Based on the available literature, these observations suggest that CH4 can be transported internally by Populus deltoides x Populus nigra seedlings in trace amounts, although future research is required to fully test this hypothesis.

  14. Establishment and early management of Populus species in southern Sweden

    OpenAIRE

    Mc Carthy, Rebecka

    2016-01-01

    Populus species are among the most productive tree species in Sweden. Interest in growing them has increased during the 21st century due to political goals to increase the share of renewable energy and to increase the proportion of hardwood species in forests. Populus species have been shown to be potentially profitable, but currently they are mostly planted on abandoned agricultural land. There is a lack of knowledge about the establishment of Populus species on forest sites. There is also a...

  15. Barcoding poplars (Populus L. from western China.

    Directory of Open Access Journals (Sweden)

    Jianju Feng

    Full Text Available BACKGROUND: Populus is an ecologically and economically important genus of trees, but distinguishing between wild species is relatively difficult due to extensive interspecific hybridization and introgression, and the high level of intraspecific morphological variation. The DNA barcoding approach is a potential solution to this problem. METHODOLOGY/PRINCIPAL FINDINGS: Here, we tested the discrimination power of five chloroplast barcodes and one nuclear barcode (ITS among 95 trees that represent 21 Populus species from western China. Among all single barcode candidates, the discrimination power is highest for the nuclear ITS, progressively lower for chloroplast barcodes matK (M, trnG-psbK (G and psbK-psbI (P, and trnH-psbA (H and rbcL (R; the discrimination efficiency of the nuclear ITS (I is also higher than any two-, three-, or even the five-locus combination of chloroplast barcodes. Among the five combinations of a single chloroplast barcode plus the nuclear ITS, H+I and P+I differentiated the highest and lowest portion of species, respectively. The highest discrimination rate for the barcodes or barcode combinations examined here is 55.0% (H+I, and usually discrimination failures occurred among species from sympatric or parapatric areas. CONCLUSIONS/SIGNIFICANCE: In this case study, we showed that when discriminating Populus species from western China, the nuclear ITS region represents a more promising barcode than any maternally inherited chloroplast region or combination of chloroplast regions. Meanwhile, combining the ITS region with chloroplast regions may improve the barcoding success rate and assist in detecting recent interspecific hybridizations. Failure to discriminate among several groups of Populus species from sympatric or parapatric areas may have been the result of incomplete lineage sorting, frequent interspecific hybridizations and introgressions. We agree with a previous proposal for constructing a tiered barcoding system in

  16. G-InforBIO: integrated system for microbial genomics

    Directory of Open Access Journals (Sweden)

    Abe Takashi

    2006-08-01

    Full Text Available Abstract Background Genome databases contain diverse kinds of information, including gene annotations and nucleotide and amino acid sequences. It is not easy to integrate such information for genomic study. There are few tools for integrated analyses of genomic data, therefore, we developed software that enables users to handle, manipulate, and analyze genome data with a variety of sequence analysis programs. Results The G-InforBIO system is a novel tool for genome data management and sequence analysis. The system can import genome data encoded as eXtensible Markup Language documents as formatted text documents, including annotations and sequences, from DNA Data Bank of Japan and GenBank encoded as flat files. The genome database is constructed automatically after importing, and the database can be exported as documents formatted with eXtensible Markup Language or tab-deliminated text. Users can retrieve data from the database by keyword searches, edit annotation data of genes, and process data with G-InforBIO. In addition, information in the G-InforBIO database can be analyzed seamlessly with nine different software programs, including programs for clustering and homology analyses. Conclusion The G-InforBIO system simplifies genome analyses by integrating several available software programs to allow efficient handling and manipulation of genome data. G-InforBIO is freely available from the download site.

  17. The integrated microbial genome resource of analysis.

    Science.gov (United States)

    Checcucci, Alice; Mengoni, Alessio

    2015-01-01

    Integrated Microbial Genomes and Metagenomes (IMG) is a biocomputational system that allows to provide information and support for annotation and comparative analysis of microbial genomes and metagenomes. IMG has been developed by the US Department of Energy (DOE)-Joint Genome Institute (JGI). IMG platform contains both draft and complete genomes, sequenced by Joint Genome Institute and other public and available genomes. Genomes of strains belonging to Archaea, Bacteria, and Eukarya domains are present as well as those of viruses and plasmids. Here, we provide some essential features of IMG system and case study for pangenome analysis.

  18. Growth and biomass of Populus irrigated with landfill leachate

    Science.gov (United States)

    Jill A. Zalesny; Ronald S., Jr. Zalesny; David R. Coyle; Richard B. Hall

    2007-01-01

    Resource managers are challenged with waste disposal and leachate produced from its degradation. Poplar (Populus spp.) trees offer an opportunity for ecological leachate disposal as an irrigation source for managed tree systems. Our objective was to irrigate Populus trees with municipal solid waste landfill leachate or fertilized well water (control...

  19. Integrative Genomics Viewer (IGV) | Informatics Technology for Cancer Research (ITCR)

    Science.gov (United States)

    The Integrative Genomics Viewer (IGV) is a high-performance visualization tool for interactive exploration of large, integrated genomic datasets. It supports a wide variety of data types, including array-based and next-generation sequence data, and genomic annotations.

  20. The Genome-Scale Integrated Networks in Microorganisms

    Directory of Open Access Journals (Sweden)

    Tong Hao

    2018-02-01

    Full Text Available The genome-scale cellular network has become a necessary tool in the systematic analysis of microbes. In a cell, there are several layers (i.e., types of the molecular networks, for example, genome-scale metabolic network (GMN, transcriptional regulatory network (TRN, and signal transduction network (STN. It has been realized that the limitation and inaccuracy of the prediction exist just using only a single-layer network. Therefore, the integrated network constructed based on the networks of the three types attracts more interests. The function of a biological process in living cells is usually performed by the interaction of biological components. Therefore, it is necessary to integrate and analyze all the related components at the systems level for the comprehensively and correctly realizing the physiological function in living organisms. In this review, we discussed three representative genome-scale cellular networks: GMN, TRN, and STN, representing different levels (i.e., metabolism, gene regulation, and cellular signaling of a cell’s activities. Furthermore, we discussed the integration of the networks of the three types. With more understanding on the complexity of microbial cells, the development of integrated network has become an inevitable trend in analyzing genome-scale cellular networks of microorganisms.

  1. GENOME-ENABLED DISCOVERY OF CARBON SEQUESTRATION GENES IN POPLAR

    Energy Technology Data Exchange (ETDEWEB)

    DAVIS J M

    2007-10-11

    Plants utilize carbon by partitioning the reduced carbon obtained through photosynthesis into different compartments and into different chemistries within a cell and subsequently allocating such carbon to sink tissues throughout the plant. Since the phytohormones auxin and cytokinin are known to influence sink strength in tissues such as roots (Skoog & Miller 1957, Nordstrom et al. 2004), we hypothesized that altering the expression of genes that regulate auxin-mediated (e.g., AUX/IAA or ARF transcription factors) or cytokinin-mediated (e.g., RR transcription factors) control of root growth and development would impact carbon allocation and partitioning belowground (Fig. 1 - Renewal Proposal). Specifically, the ARF, AUX/IAA and RR transcription factor gene families mediate the effects of the growth regulators auxin and cytokinin on cell expansion, cell division and differentiation into root primordia. Invertases (IVR), whose transcript abundance is enhanced by both auxin and cytokinin, are critical components of carbon movement and therefore of carbon allocation. Thus, we initiated comparative genomic studies to identify the AUX/IAA, ARF, RR and IVR gene families in the Populus genome that could impact carbon allocation and partitioning. Bioinformatics searches using Arabidopsis gene sequences as queries identified regions with high degrees of sequence similarities in the Populus genome. These Populus sequences formed the basis of our transgenic experiments. Transgenic modification of gene expression involving members of these gene families was hypothesized to have profound effects on carbon allocation and partitioning.

  2. IMG: the integrated microbial genomes database and comparative analysis system

    Science.gov (United States)

    Markowitz, Victor M.; Chen, I-Min A.; Palaniappan, Krishna; Chu, Ken; Szeto, Ernest; Grechkin, Yuri; Ratner, Anna; Jacob, Biju; Huang, Jinghua; Williams, Peter; Huntemann, Marcel; Anderson, Iain; Mavromatis, Konstantinos; Ivanova, Natalia N.; Kyrpides, Nikos C.

    2012-01-01

    The Integrated Microbial Genomes (IMG) system serves as a community resource for comparative analysis of publicly available genomes in a comprehensive integrated context. IMG integrates publicly available draft and complete genomes from all three domains of life with a large number of plasmids and viruses. IMG provides tools and viewers for analyzing and reviewing the annotations of genes and genomes in a comparative context. IMG's data content and analytical capabilities have been continuously extended through regular updates since its first release in March 2005. IMG is available at http://img.jgi.doe.gov. Companion IMG systems provide support for expert review of genome annotations (IMG/ER: http://img.jgi.doe.gov/er), teaching courses and training in microbial genome analysis (IMG/EDU: http://img.jgi.doe.gov/edu) and analysis of genomes related to the Human Microbiome Project (IMG/HMP: http://www.hmpdacc-resources.org/img_hmp). PMID:22194640

  3. Repeated unidirectional introgression towards Populus balsamifera in contact zones of exotic and native poplars

    NARCIS (Netherlands)

    Thompson, S.L.; Lamothe, M.; Meirmans, P.G.; Périnet, P.; Isabel, N.

    2010-01-01

    As the evolutionary significance of hybridization is largely dictated by its extent beyond the first generation, we broadly surveyed patterns of introgression across a sympatric zone of two native poplars (Populus balsamifera, Populus deltoides) in Quebec, Canada within which European exotic Populus

  4. Exploration of plant genomes in the FLAGdb++ environment

    Directory of Open Access Journals (Sweden)

    Leplé Jean-Charles

    2011-03-01

    Full Text Available Abstract Background In the contexts of genomics, post-genomics and systems biology approaches, data integration presents a major concern. Databases provide crucial solutions: they store, organize and allow information to be queried, they enhance the visibility of newly produced data by comparing them with previously published results, and facilitate the exploration and development of both existing hypotheses and new ideas. Results The FLAGdb++ information system was developed with the aim of using whole plant genomes as physical references in order to gather and merge available genomic data from in silico or experimental approaches. Available through a JAVA application, original interfaces and tools assist the functional study of plant genes by considering them in their specific context: chromosome, gene family, orthology group, co-expression cluster and functional network. FLAGdb++ is mainly dedicated to the exploration of large gene groups in order to decipher functional connections, to highlight shared or specific structural or functional features, and to facilitate translational tasks between plant species (Arabidopsis thaliana, Oryza sativa, Populus trichocarpa and Vitis vinifera. Conclusion Combining original data with the output of experts and graphical displays that differ from classical plant genome browsers, FLAGdb++ presents a powerful complementary tool for exploring plant genomes and exploiting structural and functional resources, without the need for computer programming knowledge. First launched in 2002, a 15th version of FLAGdb++ is now available and comprises four model plant genomes and over eight million genomic features.

  5. VERSE: a novel approach to detect virus integration in host genomes through reference genome customization.

    Science.gov (United States)

    Wang, Qingguo; Jia, Peilin; Zhao, Zhongming

    2015-01-01

    Fueled by widespread applications of high-throughput next generation sequencing (NGS) technologies and urgent need to counter threats of pathogenic viruses, large-scale studies were conducted recently to investigate virus integration in host genomes (for example, human tumor genomes) that may cause carcinogenesis or other diseases. A limiting factor in these studies, however, is rapid virus evolution and resulting polymorphisms, which prevent reads from aligning readily to commonly used virus reference genomes, and, accordingly, make virus integration sites difficult to detect. Another confounding factor is host genomic instability as a result of virus insertions. To tackle these challenges and improve our capability to identify cryptic virus-host fusions, we present a new approach that detects Virus intEgration sites through iterative Reference SEquence customization (VERSE). To the best of our knowledge, VERSE is the first approach to improve detection through customizing reference genomes. Using 19 human tumors and cancer cell lines as test data, we demonstrated that VERSE substantially enhanced the sensitivity of virus integration site detection. VERSE is implemented in the open source package VirusFinder 2 that is available at http://bioinfo.mc.vanderbilt.edu/VirusFinder/.

  6. Integrating genomics into undergraduate nursing education.

    Science.gov (United States)

    Daack-Hirsch, Sandra; Dieter, Carla; Quinn Griffin, Mary T

    2011-09-01

    To prepare the next generation of nurses, faculty are now faced with the challenge of incorporating genomics into curricula. Here we discuss how to meet this challenge. Steps to initiate curricular changes to include genomics are presented along with a discussion on creating a genomic curriculum thread versus a standalone course. Ideas for use of print material and technology on genomic topics are also presented. Information is based on review of the literature and curriculum change efforts by the authors. In recognition of advances in genomics, the nursing profession is increasing an emphasis on the integration of genomics into professional practice and educational standards. Incorporating genomics into nurses' practices begins with changes in our undergraduate curricula. Information given in didactic courses should be reinforced in clinical practica, and Internet-based tools such as WebQuest, Second Life, and wikis offer attractive, up-to-date platforms to deliver this now crucial content. To provide information that may assist faculty to prepare the next generation of nurses to practice using genomics. © 2011 Sigma Theta Tau International.

  7. Growth of Populus and Salix Species under Compost Leachate Irrigation

    Directory of Open Access Journals (Sweden)

    Tooba Abedi

    2014-12-01

    Full Text Available According to the known broad variation in remediation capacity, three plant species were used in the experiment: two fast growing poplar’s clones - Populus deltoides, Populus euramericana, and willows Salix alba. Populus and Salix cuttings were collected from the nursery of the Populus Research Center of Safrabasteh in the eastern part of Guilan province at north of Iran. The Populus clones were chosen because of their high biomass production capacity and willow- because it is native in Iran. The highest diameter growth rate was exhibited for all three plant species by the 1:1 treatment with an average of 0.26, 0.22 and 0.16 cm in eight months period for P. euroamericana, P. deltoides and S. alba, respectively. Over a period of eight months a higher growth rate of height was observed in (P and (1:1 treatment for S. alba (33.70 and 15.77 cm, respectively and in (C treatment for P. deltoides (16.51 cm. P. deltoides and S. alba produced significantly (p<0.05 smaller aboveground biomass in (P treatment compared to all species. P. deltoides exhibited greater mean aboveground biomass in the (1:1 treatment compared to other species. There were significant differences (p<0.05 in the growth of roots between P. deltoides, P. euramericana and S. alba in all of the treatments.

  8. Characterization of cytokinin signaling and homeostasis gene families in two hardwood tree species: Populus trichocarpa and Prunus persica.

    Science.gov (United States)

    Immanen, Juha; Nieminen, Kaisa; Duchens Silva, Héctor; Rodríguez Rojas, Fernanda; Meisel, Lee A; Silva, Herman; Albert, Victor A; Hvidsten, Torgeir R; Helariutta, Ykä

    2013-12-16

    Through the diversity of cytokinin regulated processes, this phytohormone has a profound impact on plant growth and development. Cytokinin signaling is involved in the control of apical and lateral meristem activity, branching pattern of the shoot, and leaf senescence. These processes influence several traits, including the stem diameter, shoot architecture, and perennial life cycle, which define the development of woody plants. To facilitate research about the role of cytokinin in regulation of woody plant development, we have identified genes associated with cytokinin signaling and homeostasis pathways from two hardwood tree species. Taking advantage of the sequenced black cottonwood (Populus trichocarpa) and peach (Prunus persica) genomes, we have compiled a comprehensive list of genes involved in these pathways. We identified genes belonging to the six families of cytokinin oxidases (CKXs), isopentenyl transferases (IPTs), LONELY GUY genes (LOGs), two-component receptors, histidine containing phosphotransmitters (HPts), and response regulators (RRs). All together 85 Populus and 45 Prunus genes were identified, and compared to their Arabidopsis orthologs through phylogenetic analyses. In general, when compared to Arabidopsis, differences in gene family structure were often seen in only one of the two tree species. However, one class of genes associated with cytokinin signal transduction, the CKI1-like family of two-component histidine kinases, was larger in both Populus and Prunus than in Arabidopsis.

  9. Integrating genomics into evolutionary medicine.

    Science.gov (United States)

    Rodríguez, Juan Antonio; Marigorta, Urko M; Navarro, Arcadi

    2014-12-01

    The application of the principles of evolutionary biology into medicine was suggested long ago and is already providing insight into the ultimate causes of disease. However, a full systematic integration of medical genomics and evolutionary medicine is still missing. Here, we briefly review some cases where the combination of the two fields has proven profitable and highlight two of the main issues hindering the development of evolutionary genomic medicine as a mature field, namely the dissociation between fitness and health and the still considerable difficulties in predicting phenotypes from genotypes. We use publicly available data to illustrate both problems and conclude that new approaches are needed for evolutionary genomic medicine to overcome these obstacles. Copyright © 2014 Elsevier Ltd. All rights reserved.

  10. Characterization of Equine Infectious Anemia Virus Integration in the Horse Genome

    Directory of Open Access Journals (Sweden)

    Qiang Liu

    2015-06-01

    Full Text Available Human immunodeficiency virus (HIV-1 has a unique integration profile in the human genome relative to murine and avian retroviruses. Equine infectious anemia virus (EIAV is another well-studied lentivirus that can also be used as a promising retro-transfection vector, but its integration into its native host has not been characterized. In this study, we mapped 477 integration sites of the EIAV strain EIAVFDDV13 in fetal equine dermal (FED cells during in vitro infection. Published integration sites of EIAV and HIV-1 in the human genome were also analyzed as references. Our results demonstrated that EIAVFDDV13 tended to integrate into genes and AT-rich regions, and it avoided integrating into transcription start sites (TSS, which is consistent with EIAV and HIV-1 integration in the human genome. Notably, the integration of EIAVFDDV13 favored long interspersed elements (LINEs and DNA transposons in the horse genome, whereas the integration of HIV-1 favored short interspersed elements (SINEs in the human genome. The chromosomal environment near LINEs or DNA transposons potentially influences viral transcription and may be related to the unique EIAV latency states in equids. The data on EIAV integration in its natural host will facilitate studies on lentiviral infection and lentivirus-based therapeutic vectors.

  11. IMG 4 version of the integrated microbial genomes comparative analysis system

    Science.gov (United States)

    Markowitz, Victor M.; Chen, I-Min A.; Palaniappan, Krishna; Chu, Ken; Szeto, Ernest; Pillay, Manoj; Ratner, Anna; Huang, Jinghua; Woyke, Tanja; Huntemann, Marcel; Anderson, Iain; Billis, Konstantinos; Varghese, Neha; Mavromatis, Konstantinos; Pati, Amrita; Ivanova, Natalia N.; Kyrpides, Nikos C.

    2014-01-01

    The Integrated Microbial Genomes (IMG) data warehouse integrates genomes from all three domains of life, as well as plasmids, viruses and genome fragments. IMG provides tools for analyzing and reviewing the structural and functional annotations of genomes in a comparative context. IMG’s data content and analytical capabilities have increased continuously since its first version released in 2005. Since the last report published in the 2012 NAR Database Issue, IMG’s annotation and data integration pipelines have evolved while new tools have been added for recording and analyzing single cell genomes, RNA Seq and biosynthetic cluster data. Different IMG datamarts provide support for the analysis of publicly available genomes (IMG/W: http://img.jgi.doe.gov/w), expert review of genome annotations (IMG/ER: http://img.jgi.doe.gov/er) and teaching and training in the area of microbial genome analysis (IMG/EDU: http://img.jgi.doe.gov/edu). PMID:24165883

  12. IMG 4 version of the integrated microbial genomes comparative analysis system

    Energy Technology Data Exchange (ETDEWEB)

    Markowitz, Victor M. [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Chen, I-Min A. [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Palaniappan, Krishna [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Chu, Ken [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Szeto, Ernest [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Pillay, Manoj [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Ratner, Anna [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Huang, Jinghua [Lawrence Berkeley National Lab. (LBNL), Berkeley, CA (United States). Biological Data Management and Technology Center. Computational Research Division; Woyke, Tanja [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Huntemann, Marcel [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Anderson, Iain [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Billis, Konstantinos [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Varghese, Neha [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Mavromatis, Konstantinos [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Pati, Amrita [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Ivanova, Natalia N. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program; Kyrpides, Nikos C. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States). Microbial Genome and Metagenome Program

    2013-10-27

    The Integrated Microbial Genomes (IMG) data warehouse integrates genomes from all three domains of life, as well as plasmids, viruses and genome fragments. IMG provides tools for analyzing and reviewing the structural and functional annotations of genomes in a comparative context. IMG’s data content and analytical capabilities have increased continuously since its first version released in 2005. Since the last report published in the 2012 NAR Database Issue, IMG’s annotation and data integration pipelines have evolved while new tools have been added for recording and analyzing single cell genomes, RNA Seq and biosynthetic cluster data. Finally, different IMG datamarts provide support for the analysis of publicly available genomes (IMG/W: http://img.jgi.doe.gov/w), expert review of genome annotations (IMG/ER: http://img.jgi.doe.gov/er) and teaching and training in the area of microbial genome analysis (IMG/EDU: http://img.jgi.doe.gov/edu).

  13. The Integrated Microbial Genomes (IMG) System: An Expanding Comparative Analysis Resource

    Energy Technology Data Exchange (ETDEWEB)

    Markowitz, Victor M.; Chen, I-Min A.; Palaniappan, Krishna; Chu, Ken; Szeto, Ernest; Grechkin, Yuri; Ratner, Anna; Anderson, Iain; Lykidis, Athanasios; Mavromatis, Konstantinos; Ivanova, Natalia N.; Kyrpides, Nikos C.

    2009-09-13

    The integrated microbial genomes (IMG) system serves as a community resource for comparative analysis of publicly available genomes in a comprehensive integrated context. IMG contains both draft and complete microbial genomes integrated with other publicly available genomes from all three domains of life, together with a large number of plasmids and viruses. IMG provides tools and viewers for analyzing and reviewing the annotations of genes and genomes in a comparative context. Since its first release in 2005, IMG's data content and analytical capabilities have been constantly expanded through regular releases. Several companion IMG systems have been set up in order to serve domain specific needs, such as expert review of genome annotations. IMG is available at .

  14. Spatial Pattern of Populus euphratica Forest Change as Affected by Water Conveyance in the Lower Tarim River

    Directory of Open Access Journals (Sweden)

    Shuhong Peng

    2014-01-01

    Full Text Available To restore declining species, including Populus euphratica and other riparian communities, in the river ecosystem of the lower Tarim River, the ecological water conveyance project (EWCP, as a part of an integrated water resource management plan, was implemented in 2000. The EWCP aims to schedule and manage the water resources in the upper reaches and transfer water to the lower reaches by a series of intermittent water deliveries. The delivered water flows along a modified river channel and nourishes riparian communities by river overflow flooding. Since it began, it has caused a fierce debate over the response of riparian vegetation to the water conveyance scheme. This study focuses on the lower Tarim River, where Populus euphratica forests have undergone watering, due to the EWCP. Twelve Landsat sensor images and one IKONOS satellite imagery acquired between 1999 and 2009 were used to monitor the change in Populus euphratica forests. Bi-temporal change detection and temporal trajectory analysis were employed to represent the spatial pattern of the forest change. Field investigations were used to analyze the driving forces behind forest change from the perspectives of anthropogenic activities and natural forces. The results showed that Populus euphratica forest have been declining in area, which implies that ecological risks have been increased during the watering process. However, forests areas have increased in the regions where the water supply is abundant, and vice versa.

  15. Detection of a variable number of ribosomal DNA loci by fluorescent in situ hybridization in Populus species.

    Science.gov (United States)

    Prado, E A; Faivre-Rampant, P; Schneider, C; Darmency, M A

    1996-10-01

    Fluorescent in situ hybridization (FISH) was applied to related Populus species (2n = 19) in order to detect rDNA loci. An interspecific variability in the number of hybridization sites was revealed using as probe an homologous 25S clone from Populus deltoides. The application of image analysis methods to measure fluorescence intensity of the hybridization signals has enabled us to characterize major and minor loci in the 18S-5.8S-25S rDNA. We identified one pair of such rDNA clusters in Populus alba; two pairs, one major and one minor, in both Populus nigra and P. deltoides; and three pairs in Populus balsamifera, (two major and one minor) and Populus euroamericana (one major and two minor). FISH results are in agreement with those based on RFLP analysis. The pBG13 probe containing 5S sequence from flax detected two separate clusters corresponding to the two size classes of units that coexist within 5S rDNA of most Populus species. Key words : Populus spp., fluorescent in situ hybridization, FISH, rDNA variability, image analysis.

  16. The three-dimensional genome organization of Drosophila melanogaster through data integration.

    Science.gov (United States)

    Li, Qingjiao; Tjong, Harianto; Li, Xiao; Gong, Ke; Zhou, Xianghong Jasmine; Chiolo, Irene; Alber, Frank

    2017-07-31

    Genome structures are dynamic and non-randomly organized in the nucleus of higher eukaryotes. To maximize the accuracy and coverage of three-dimensional genome structural models, it is important to integrate all available sources of experimental information about a genome's organization. It remains a major challenge to integrate such data from various complementary experimental methods. Here, we present an approach for data integration to determine a population of complete three-dimensional genome structures that are statistically consistent with data from both genome-wide chromosome conformation capture (Hi-C) and lamina-DamID experiments. Our structures resolve the genome at the resolution of topological domains, and reproduce simultaneously both sets of experimental data. Importantly, this data deconvolution framework allows for structural heterogeneity between cells, and hence accounts for the expected plasticity of genome structures. As a case study we choose Drosophila melanogaster embryonic cells, for which both data types are available. Our three-dimensional genome structures have strong predictive power for structural features not directly visible in the initial data sets, and reproduce experimental hallmarks of the D. melanogaster genome organization from independent and our own imaging experiments. Also they reveal a number of new insights about genome organization and its functional relevance, including the preferred locations of heterochromatic satellites of different chromosomes, and observations about homologous pairing that cannot be directly observed in the original Hi-C or lamina-DamID data. Our approach allows systematic integration of Hi-C and lamina-DamID data for complete three-dimensional genome structure calculation, while also explicitly considering genome structural variability.

  17. Leaf, woody, and root biomass of Populus irrigated with landfill leachate

    Science.gov (United States)

    Jill A. Zalesny; Ronald S., Jr. Zalesny; D.R. Coyle; R.B. Hall

    2007-01-01

    Poplar (Populus spp.) trees can be utilized for ecological leachate disposal when applied as an irrigation source for managed tree systems. Our objective was to evaluate differences in tree height, diameter, volume, and biomass of leaf, stem, branch, and root tissues of Populus trees after two seasons of irrigation with municipal...

  18. Deserts in the Deluge: TerraPopulus and Big Human-Environment Data.

    Science.gov (United States)

    Manson, S M; Kugler, T A; Haynes, D

    2016-01-01

    Terra Populus, or TerraPop, is a cyberinfrastructure project that integrates, preserves, and disseminates massive data collections describing characteristics of the human population and environment over the last six decades. TerraPop has made a number of GIScience advances in the handling of big spatial data to make information interoperable between formats and across scientific communities. In this paper, we describe challenges of these data, or 'deserts in the deluge' of data, that are common to spatial big data more broadly, and explore computational solutions specific to microdata, raster, and vector data models.

  19. Characterization of MORE AXILLARY GROWTH genes in Populus.

    Directory of Open Access Journals (Sweden)

    Olaf Czarnecki

    Full Text Available Strigolactones are a new class of plant hormones that play a key role in regulating shoot branching. Studies of branching mutants in Arabidopsis, pea, rice and petunia have identified several key genes involved in strigolactone biosynthesis or signaling pathway. In the model plant Arabidopsis, MORE AXILLARY GROWTH1 (MAX1, MAX2, MAX3 and MAX4 are four founding members of strigolactone pathway genes. However, little is known about the strigolactone pathway genes in the woody perennial plants.Here we report the identification of MAX homologues in the woody model plant Populus trichocarpa. We identified the sequence homologues for each MAX protein in P. trichocarpa. Gene expression analysis revealed that Populus MAX paralogous genes are differentially expressed across various tissues and organs. Furthermore, we showed that Populus MAX genes could complement or partially complement the shoot branching phenotypes of the corresponding Arabidopsis max mutants.This study provides genetic evidence that strigolactone pathway genes are likely conserved in the woody perennial plants and lays a foundation for further characterization of strigolactone pathway and its functions in the woody perennial plants.

  20. Genome-wide analysis and heavy metal-induced expression profiling of the HMA gene family in Populus trichocarpa

    Directory of Open Access Journals (Sweden)

    Dandan eLi

    2015-12-01

    Full Text Available The heavy metal ATPase (HMA family plays an important role in transition metal transport in plants. However, this gene family has not been extensively studied in Populus trichocarpa. We identified 17 HMA genes in P. trichocarpa (PtHMAs, of which PtHMA1–PtHMA4 belonged to the zinc (Zn/cobalt (Co/cadmium (Cd/lead (Pb subgroup, and PtHMA5–PtHMA8 were members of the copper (Cu/silver (Ag subgroup. Most of the genes were localized to chromosomes I and III. Gene structure, gene chromosomal location, and synteny analyses of PtHMAs indicated that tandem and segmental duplications likely contributed to the expansion and evolution of the PtHMAs. Most of the HMA genes contained abiotic stress-related cis-elements. Tissue-specific expression of PtHMA genes showed that PtHMA1 and PtHMA4 had relatively high expression levels in the leaves, whereas Cu/Ag subgroup (PtHMA5.1- PtHMA8 genes were upregulated in the roots. High concentrations of Cu, Ag, Zn, Cd, Co, Pb and Mn differentially regulated the expression of PtHMAs in various tissues. The preliminary results of the present study generated basic information on the HMA family of Populus that may serve as foundation for future functional studies.

  1. Genome-wide profiling of HPV integration in cervical cancer identifies clustered genomic hot spots and a potential microhomology-mediated integration mechanism

    DEFF Research Database (Denmark)

    Hu, Zheng; Zhu, Da; Wang, Wei

    2015-01-01

    Human papillomavirus (HPV) integration is a key genetic event in cervical carcinogenesis1. By conducting whole-genome sequencing and high-throughput viral integration detection, we identified 3,667 HPV integration breakpoints in 26 cervical intraepithelial neoplasias, 104 cervical carcinomas and ...

  2. Identifying gene coexpression networks underlying the dynamic regulation of wood-forming tissues in Populus under diverse environmental conditions.

    Science.gov (United States)

    Zinkgraf, Matthew; Liu, Lijun; Groover, Andrew; Filkov, Vladimir

    2017-06-01

    Trees modify wood formation through integration of environmental and developmental signals in complex but poorly defined transcriptional networks, allowing trees to produce woody tissues appropriate to diverse environmental conditions. In order to identify relationships among genes expressed during wood formation, we integrated data from new and publically available datasets in Populus. These datasets were generated from woody tissue and include transcriptome profiling, transcription factor binding, DNA accessibility and genome-wide association mapping experiments. Coexpression modules were calculated, each of which contains genes showing similar expression patterns across experimental conditions, genotypes and treatments. Conserved gene coexpression modules (four modules totaling 8398 genes) were identified that were highly preserved across diverse environmental conditions and genetic backgrounds. Functional annotations as well as correlations with specific experimental treatments associated individual conserved modules with distinct biological processes underlying wood formation, such as cell-wall biosynthesis, meristem development and epigenetic pathways. Module genes were also enriched for DNase I hypersensitivity footprints and binding from four transcription factors associated with wood formation. The conserved modules are excellent candidates for modeling core developmental pathways common to wood formation in diverse environments and genotypes, and serve as testbeds for hypothesis generation and testing for future studies. No claim to original US government works. New Phytologist © 2017 New Phytologist Trust.

  3. GAPIT: genome association and prediction integrated tool.

    Science.gov (United States)

    Lipka, Alexander E; Tian, Feng; Wang, Qishan; Peiffer, Jason; Li, Meng; Bradbury, Peter J; Gore, Michael A; Buckler, Edward S; Zhang, Zhiwu

    2012-09-15

    Software programs that conduct genome-wide association studies and genomic prediction and selection need to use methodologies that maximize statistical power, provide high prediction accuracy and run in a computationally efficient manner. We developed an R package called Genome Association and Prediction Integrated Tool (GAPIT) that implements advanced statistical methods including the compressed mixed linear model (CMLM) and CMLM-based genomic prediction and selection. The GAPIT package can handle large datasets in excess of 10 000 individuals and 1 million single-nucleotide polymorphisms with minimal computational time, while providing user-friendly access and concise tables and graphs to interpret results. http://www.maizegenetics.net/GAPIT. zhiwu.zhang@cornell.edu Supplementary data are available at Bioinformatics online.

  4. Roles of gibberellin catabolism and signaling in growth and physiological response to drought and short-day photoperiods in Populus trees.

    Directory of Open Access Journals (Sweden)

    Christine Zawaski

    Full Text Available Survival and productivity of perennial plants in temperate zones are dependent on robust responses to prolonged and seasonal cycles of unfavorable conditions. Here we report whole-genome microarray, expression, physiological, and transgenic evidence in hybrid poplar (Populus tremula × Populus alba showing that gibberellin (GA catabolism and repressive signaling mediates shoot growth inhibition and physiological adaptation in response to drought and short-day (SD induced bud dormancy. Both water deprivation and SDs elicited activation of a suite of poplar GA2ox and DELLA encoding genes. Poplar transgenics with up-regulated GA 2-oxidase (GA2ox and DELLA domain proteins showed hypersensitive growth inhibition in response to both drought and SDs. In addition, the transgenic plants displayed greater drought resistance as evidenced by increased pigment concentrations (chlorophyll and carotenoid and reductions in electrolyte leakage (EL. Comparative transcriptome analysis using whole-genome microarray showed that the GA-deficiency and GA-insensitivity, SD-induced dormancy, and drought response in poplar share a common regulon of 684 differentially-expressed genes, which suggest GA metabolism and signaling plays a role in plant physiological adaptations in response to alterations in environmental factors. Our results demonstrate that GA catabolism and repressive signaling represents a major route for control of growth and physiological adaptation in response to immediate or imminent adverse conditions.

  5. Dynamic DNA cytosine methylation in the Populus trichocarpa genome: tissue-level variation and relationship to gene expression

    Directory of Open Access Journals (Sweden)

    Vining Kelly J

    2012-01-01

    Full Text Available Abstract Background DNA cytosine methylation is an epigenetic modification that has been implicated in many biological processes. However, large-scale epigenomic studies have been applied to very few plant species, and variability in methylation among specialized tissues and its relationship to gene expression is poorly understood. Results We surveyed DNA methylation from seven distinct tissue types (vegetative bud, male inflorescence [catkin], female catkin, leaf, root, xylem, phloem in the reference tree species black cottonwood (Populus trichocarpa. Using 5-methyl-cytosine DNA immunoprecipitation followed by Illumina sequencing (MeDIP-seq, we mapped a total of 129,360,151 36- or 32-mer reads to the P. trichocarpa reference genome. We validated MeDIP-seq results by bisulfite sequencing, and compared methylation and gene expression using published microarray data. Qualitative DNA methylation differences among tissues were obvious on a chromosome scale. Methylated genes had lower expression than unmethylated genes, but genes with methylation in transcribed regions ("gene body methylation" had even lower expression than genes with promoter methylation. Promoter methylation was more frequent than gene body methylation in all tissues except male catkins. Male catkins differed in demethylation of particular transposable element categories, in level of gene body methylation, and in expression range of genes with methylated transcribed regions. Tissue-specific gene expression patterns were correlated with both gene body and promoter methylation. Conclusions We found striking differences among tissues in methylation, which were apparent at the chromosomal scale and when genes and transposable elements were examined. In contrast to other studies in plants, gene body methylation had a more repressive effect on transcription than promoter methylation.

  6. Assembly and Multiplex Genome Integration of Metabolic Pathways in Yeast Using CasEMBLR

    DEFF Research Database (Denmark)

    Jakočiūnas, Tadas; Jensen, Emil D.; Jensen, Michael Krogh

    2018-01-01

    and marker-free integration of the carotenoid pathway from 15 exogenously supplied DNA parts into three targeted genomic loci. As a second proof-of-principle, a total of ten DNA parts were assembled and integrated in two genomic loci to construct a tyrosine production strain, and at the same time knocking......Genome integration is a vital step for implementing large biochemical pathways to build a stable microbial cell factory. Although traditional strain construction strategies are well established for the model organism Saccharomyces cerevisiae, recent advances in CRISPR/Cas9-mediated genome...... engineering allow much higher throughput and robustness in terms of strain construction. In this chapter, we describe CasEMBLR, a highly efficient and marker-free genome engineering method for one-step integration of in vivo assembled expression cassettes in multiple genomic sites simultaneously. Cas...

  7. Gene Structures, Classification, and Expression Models of the DREB Transcription Factor Subfamily in Populus trichocarpa

    Directory of Open Access Journals (Sweden)

    Yunlin Chen

    2013-01-01

    Full Text Available We identified 75 dehydration-responsive element-binding (DREB protein genes in Populus trichocarpa. We analyzed gene structures, phylogenies, domain duplications, genome localizations, and expression profiles. The phylogenic construction suggests that the PtrDREB gene subfamily can be classified broadly into six subtypes (DREB A-1 to A-6 in Populus. The chromosomal localizations of the PtrDREB genes indicated 18 segmental duplication events involving 36 genes and six redundant PtrDREB genes were involved in tandem duplication events. There were fewer introns in the PtrDREB subfamily. The motif composition of PtrDREB was highly conserved in the same subtype. We investigated expression profiles of this gene subfamily from different tissues and/or developmental stages. Sixteen genes present in the digital expression analysis had high levels of transcript accumulation. The microarray results suggest that 18 genes were upregulated. We further examined the stress responsiveness of 15 genes by qRT-PCR. A digital northern analysis showed that the PtrDREB17, 18, and 32 genes were highly induced in leaves under cold stress, and the same expression trends were shown by qRT-PCR. Taken together, these observations may lay the foundation for future functional analyses to unravel the biological roles of Populus’ DREB genes.

  8. Stock characterization and improvement: DNA fingerprinting and cold tolerance of Populus and Salix clones

    Energy Technology Data Exchange (ETDEWEB)

    Lin, Dolly; Hubbes, M.; Zsuffa, L. [Toronto Univ., ON (Canada). Faculty of Forestry; Tsarouhas, V.; Gullberg, U. [Swedish Univ. of Agricultural Sciences, Uppsala (Sweden). Dept. of Forest Genetics; Howe, G.; Hackett, W.; Gardner, G.; Furnier, G. [Minnesota Univ., St. Paul, MN (United States). Dept. of Forest Resources; Tuskan, G. [Oak Ridge National Lab., TN (United States)

    1998-12-31

    Molecular characterization of advanced-generation pedigrees and evaluation of cold tolerance are two aspects of Populus and Salix genetic improvement programmes worldwide that have traditionally received little emphasis. As such, chloroplast DNA markers were tested as a means of determining multi-generation parental contributions to hybrid progeny. Likewise, greenhouse, growth chamber and field studies were used to assess cold tolerance in Populus and Salix. Chloroplast DNA markers did not reveal size polymorphisms among four tested Populus species, but did produce sequence polymorphisms between P. maximowiczii and P. trichocarpa. Additional crosses between multiple genotypes from each species are being used to evaluate the utility of the detected polymorphism for ascertaining parentage within interspecific crosses. Short-day, cold tolerance greenhouse studies revealed that bud set date and frost damage are moderately heritable and genetically correlated in Populus. The relationship between greenhouse and field studies suggests that factors other than short days contribute to cold tolerance in Populus. In Salix, response to artificial fall conditioning varied among full-sibling families, with the fastest growing family displaying the greatest amount of cold tolerance 26 refs, 3 tabs

  9. Transcription as a Threat to Genome Integrity.

    Science.gov (United States)

    Gaillard, Hélène; Aguilera, Andrés

    2016-06-02

    Genomes undergo different types of sporadic alterations, including DNA damage, point mutations, and genome rearrangements, that constitute the basis for evolution. However, these changes may occur at high levels as a result of cell pathology and trigger genome instability, a hallmark of cancer and a number of genetic diseases. In the last two decades, evidence has accumulated that transcription constitutes an important natural source of DNA metabolic errors that can compromise the integrity of the genome. Transcription can create the conditions for high levels of mutations and recombination by its ability to open the DNA structure and remodel chromatin, making it more accessible to DNA insulting agents, and by its ability to become a barrier to DNA replication. Here we review the molecular basis of such events from a mechanistic perspective with particular emphasis on the role of transcription as a genome instability determinant.

  10. Human Papillomavirus Genome Integration and Head and Neck Cancer.

    Science.gov (United States)

    Pinatti, L M; Walline, H M; Carey, T E

    2018-06-01

    We conducted a critical review of human papillomavirus (HPV) integration into the host genome in oral/oropharyngeal cancer, reviewed the literature for HPV-induced cancers, and obtained current data for HPV-related oral and oropharyngeal cancers. In addition, we performed studies to identify HPV integration sites and the relationship of integration to viral-host fusion transcripts and whether integration is required for HPV-associated oncogenesis. Viral integration of HPV into the host genome is not required for the viral life cycle and might not be necessary for cellular transformation, yet HPV integration is frequently reported in cervical and head and neck cancer specimens. Studies of large numbers of early cervical lesions revealed frequent viral integration into gene-poor regions of the host genome with comparatively rare integration into cellular genes, suggesting that integration is a stochastic event and that site of integration may be largely a function of chance. However, more recent studies of head and neck squamous cell carcinomas (HNSCCs) suggest that integration may represent an additional oncogenic mechanism through direct effects on cancer-related gene expression and generation of hybrid viral-host fusion transcripts. In HNSCC cell lines as well as primary tumors, integration into cancer-related genes leading to gene disruption has been reported. The studies have shown that integration-induced altered gene expression may be associated with tumor recurrence. Evidence from several studies indicates that viral integration into genic regions is accompanied by local amplification, increased expression in some cases, interruption of gene expression, and likely additional oncogenic effects. Similarly, reported examples of viral integration near microRNAs suggest that altered expression of these regulatory molecules may also contribute to oncogenesis. Future work is indicated to identify the mechanisms of these events on cancer cell behavior.

  11. Perspectives of Integrative Cancer Genomics in Next Generation Sequencing Era

    Directory of Open Access Journals (Sweden)

    So Mee Kwon

    2012-06-01

    Full Text Available The explosive development of genomics technologies including microarrays and next generation sequencing (NGS has provided comprehensive maps of cancer genomes, including the expression of mRNAs and microRNAs, DNA copy numbers, sequence variations, and epigenetic changes. These genome-wide profiles of the genetic aberrations could reveal the candidates for diagnostic and/or prognostic biomarkers as well as mechanistic insights into tumor development and progression. Recent efforts to establish the huge cancer genome compendium and integrative omics analyses, so-called "integromics", have extended our understanding on the cancer genome, showing its daunting complexity and heterogeneity. However, the challenges of the structured integration, sharing, and interpretation of the big omics data still remain to be resolved. Here, we review several issues raised in cancer omics data analysis, including NGS, focusing particularly on the study design and analysis strategies. This might be helpful to understand the current trends and strategies of the rapidly evolving cancer genomics research.

  12. Assembly and Multiplex Genome Integration of Metabolic Pathways in Yeast Using CasEMBLR.

    Science.gov (United States)

    Jakočiūnas, Tadas; Jensen, Emil D; Jensen, Michael K; Keasling, Jay D

    2018-01-01

    Genome integration is a vital step for implementing large biochemical pathways to build a stable microbial cell factory. Although traditional strain construction strategies are well established for the model organism Saccharomyces cerevisiae, recent advances in CRISPR/Cas9-mediated genome engineering allow much higher throughput and robustness in terms of strain construction. In this chapter, we describe CasEMBLR, a highly efficient and marker-free genome engineering method for one-step integration of in vivo assembled expression cassettes in multiple genomic sites simultaneously. CasEMBLR capitalizes on the CRISPR/Cas9 technology to generate double-strand breaks in genomic loci, thus prompting native homologous recombination (HR) machinery to integrate exogenously derived homology templates. As proof-of-principle for microbial cell factory development, CasEMBLR was used for one-step assembly and marker-free integration of the carotenoid pathway from 15 exogenously supplied DNA parts into three targeted genomic loci. As a second proof-of-principle, a total of ten DNA parts were assembled and integrated in two genomic loci to construct a tyrosine production strain, and at the same time knocking out two genes. This new method complements and improves the field of genome engineering in S. cerevisiae by providing a more flexible platform for rapid and precise strain building.

  13. GDR (Genome Database for Rosaceae: integrated web resources for Rosaceae genomics and genetics research

    Directory of Open Access Journals (Sweden)

    Ficklin Stephen

    2004-09-01

    Full Text Available Abstract Background Peach is being developed as a model organism for Rosaceae, an economically important family that includes fruits and ornamental plants such as apple, pear, strawberry, cherry, almond and rose. The genomics and genetics data of peach can play a significant role in the gene discovery and the genetic understanding of related species. The effective utilization of these peach resources, however, requires the development of an integrated and centralized database with associated analysis tools. Description The Genome Database for Rosaceae (GDR is a curated and integrated web-based relational database. GDR contains comprehensive data of the genetically anchored peach physical map, an annotated peach EST database, Rosaceae maps and markers and all publicly available Rosaceae sequences. Annotations of ESTs include contig assembly, putative function, simple sequence repeats, and anchored position to the peach physical map where applicable. Our integrated map viewer provides graphical interface to the genetic, transcriptome and physical mapping information. ESTs, BACs and markers can be queried by various categories and the search result sites are linked to the integrated map viewer or to the WebFPC physical map sites. In addition to browsing and querying the database, users can compare their sequences with the annotated GDR sequences via a dedicated sequence similarity server running either the BLAST or FASTA algorithm. To demonstrate the utility of the integrated and fully annotated database and analysis tools, we describe a case study where we anchored Rosaceae sequences to the peach physical and genetic map by sequence similarity. Conclusions The GDR has been initiated to meet the major deficiency in Rosaceae genomics and genetics research, namely a centralized web database and bioinformatics tools for data storage, analysis and exchange. GDR can be accessed at http://www.genome.clemson.edu/gdr/.

  14. GDR (Genome Database for Rosaceae): integrated web resources for Rosaceae genomics and genetics research.

    Science.gov (United States)

    Jung, Sook; Jesudurai, Christopher; Staton, Margaret; Du, Zhidian; Ficklin, Stephen; Cho, Ilhyung; Abbott, Albert; Tomkins, Jeffrey; Main, Dorrie

    2004-09-09

    Peach is being developed as a model organism for Rosaceae, an economically important family that includes fruits and ornamental plants such as apple, pear, strawberry, cherry, almond and rose. The genomics and genetics data of peach can play a significant role in the gene discovery and the genetic understanding of related species. The effective utilization of these peach resources, however, requires the development of an integrated and centralized database with associated analysis tools. The Genome Database for Rosaceae (GDR) is a curated and integrated web-based relational database. GDR contains comprehensive data of the genetically anchored peach physical map, an annotated peach EST database, Rosaceae maps and markers and all publicly available Rosaceae sequences. Annotations of ESTs include contig assembly, putative function, simple sequence repeats, and anchored position to the peach physical map where applicable. Our integrated map viewer provides graphical interface to the genetic, transcriptome and physical mapping information. ESTs, BACs and markers can be queried by various categories and the search result sites are linked to the integrated map viewer or to the WebFPC physical map sites. In addition to browsing and querying the database, users can compare their sequences with the annotated GDR sequences via a dedicated sequence similarity server running either the BLAST or FASTA algorithm. To demonstrate the utility of the integrated and fully annotated database and analysis tools, we describe a case study where we anchored Rosaceae sequences to the peach physical and genetic map by sequence similarity. The GDR has been initiated to meet the major deficiency in Rosaceae genomics and genetics research, namely a centralized web database and bioinformatics tools for data storage, analysis and exchange. GDR can be accessed at http://www.genome.clemson.edu/gdr/.

  15. Phylogenomic Analysis and Dynamic Evolution of Chloroplast Genomes in Salicaceae

    Directory of Open Access Journals (Sweden)

    Yuan Huang

    2017-06-01

    Full Text Available Chloroplast genomes of plants are highly conserved in both gene order and gene content. Analysis of the whole chloroplast genome is known to provide much more informative DNA sites and thus generates high resolution for plant phylogenies. Here, we report the complete chloroplast genomes of three Salix species in family Salicaceae. Phylogeny of Salicaceae inferred from complete chloroplast genomes is generally consistent with previous studies but resolved with higher statistical support. Incongruences of phylogeny, however, are observed in genus Populus, which most likely results from homoplasy. By comparing three Salix chloroplast genomes with the published chloroplast genomes of other Salicaceae species, we demonstrate that the synteny and length of chloroplast genomes in Salicaceae are highly conserved but experienced dynamic evolution among species. We identify seven positively selected chloroplast genes in Salicaceae, which might be related to the adaptive evolution of Salicaceae species. Comparative chloroplast genome analysis within the family also indicates that some chloroplast genes are lost or became pseudogenes, infer that the chloroplast genes horizontally transferred to the nucleus genome. Based on the complete nucleus genome sequences from two Salicaceae species, we remarkably identify that the entire chloroplast genome is indeed transferred and integrated to the nucleus genome in the individual of the reference genome of P. trichocarpa at least once. This observation, along with presence of the large nuclear plastid DNA (NUPTs and NUPTs-containing multiple chloroplast genes in their original order in the chloroplast genome, favors the DNA-mediated hypothesis of organelle to nucleus DNA transfer. Overall, the phylogenomic analysis using chloroplast complete genomes clearly elucidates the phylogeny of Salicaceae. The identification of positively selected chloroplast genes and dynamic chloroplast-to-nucleus gene transfers in

  16. Ensembl Genomes: an integrative resource for genome-scale data from non-vertebrate species.

    Science.gov (United States)

    Kersey, Paul J; Staines, Daniel M; Lawson, Daniel; Kulesha, Eugene; Derwent, Paul; Humphrey, Jay C; Hughes, Daniel S T; Keenan, Stephan; Kerhornou, Arnaud; Koscielny, Gautier; Langridge, Nicholas; McDowall, Mark D; Megy, Karine; Maheswari, Uma; Nuhn, Michael; Paulini, Michael; Pedro, Helder; Toneva, Iliana; Wilson, Derek; Yates, Andrew; Birney, Ewan

    2012-01-01

    Ensembl Genomes (http://www.ensemblgenomes.org) is an integrative resource for genome-scale data from non-vertebrate species. The project exploits and extends technology (for genome annotation, analysis and dissemination) developed in the context of the (vertebrate-focused) Ensembl project and provides a complementary set of resources for non-vertebrate species through a consistent set of programmatic and interactive interfaces. These provide access to data including reference sequence, gene models, transcriptional data, polymorphisms and comparative analysis. Since its launch in 2009, Ensembl Genomes has undergone rapid expansion, with the goal of providing coverage of all major experimental organisms, and additionally including taxonomic reference points to provide the evolutionary context in which genes can be understood. Against the backdrop of a continuing increase in genome sequencing activities in all parts of the tree of life, we seek to work, wherever possible, with the communities actively generating and using data, and are participants in a growing range of collaborations involved in the annotation and analysis of genomes.

  17. Integrated Genome-Based Studies of Shewanella Ecophysiology

    Energy Technology Data Exchange (ETDEWEB)

    Zhou, Jizhong [Univ. of Oklahoma, Norman, OK (United States); He, Zhili [Univ. of Oklahoma, Norman, OK (United States)

    2014-04-08

    As a part of the Shewanella Federation project, we have used integrated genomic, proteomic and computational technologies to study various aspects of energy metabolism of two Shewanella strains from a systems-level perspective.

  18. Integrative Genome Comparison of Primary and Metastatic Melanomas

    Science.gov (United States)

    Feng, Bin; Nazarian, Rosalynn M.; Bosenberg, Marcus; Wu, Min; Scott, Kenneth L.; Kwong, Lawrence N.; Xiao, Yonghong; Cordon-Cardo, Carlos; Granter, Scott R.; Ramaswamy, Sridhar; Golub, Todd; Duncan, Lyn M.; Wagner, Stephan N.; Brennan, Cameron; Chin, Lynda

    2010-01-01

    A cardinal feature of malignant melanoma is its metastatic propensity. An incomplete view of the genetic events driving metastatic progression has been a major barrier to rational development of effective therapeutics and prognostic diagnostics for melanoma patients. In this study, we conducted global genomic characterization of primary and metastatic melanomas to examine the genomic landscape associated with metastatic progression. In addition to uncovering three genomic subclasses of metastastic melanomas, we delineated 39 focal and recurrent regions of amplification and deletions, many of which encompassed resident genes that have not been implicated in cancer or metastasis. To identify progression-associated metastasis gene candidates, we applied a statistical approach, Integrative Genome Comparison (IGC), to define 32 genomic regions of interest that were significantly altered in metastatic relative to primary melanomas, encompassing 30 resident genes with statistically significant expression deregulation. Functional assays on a subset of these candidates, including MET, ASPM, AKAP9, IMP3, PRKCA, RPA3, and SCAP2, validated their pro-invasion activities in human melanoma cells. Validity of the IGC approach was further reinforced by tissue microarray analysis of Survivin showing significant increased protein expression in thick versus thin primary cutaneous melanomas, and a progression correlation with lymph node metastases. Together, these functional validation results and correlative analysis of human tissues support the thesis that integrated genomic and pathological analyses of staged melanomas provide a productive entry point for discovery of melanoma metastases genes. PMID:20520718

  19. INE: a rice genome database with an integrated map view.

    Science.gov (United States)

    Sakata, K; Antonio, B A; Mukai, Y; Nagasaki, H; Sakai, Y; Makino, K; Sasaki, T

    2000-01-01

    The Rice Genome Research Program (RGP) launched a large-scale rice genome sequencing in 1998 aimed at decoding all genetic information in rice. A new genome database called INE (INtegrated rice genome Explorer) has been developed in order to integrate all the genomic information that has been accumulated so far and to correlate these data with the genome sequence. A web interface based on Java applet provides a rapid viewing capability in the database. The first operational version of the database has been completed which includes a genetic map, a physical map using YAC (Yeast Artificial Chromosome) clones and PAC (P1-derived Artificial Chromosome) contigs. These maps are displayed graphically so that the positional relationships among the mapped markers on each chromosome can be easily resolved. INE incorporates the sequences and annotations of the PAC contig. A site on low quality information ensures that all submitted sequence data comply with the standard for accuracy. As a repository of rice genome sequence, INE will also serve as a common database of all sequence data obtained by collaborating members of the International Rice Genome Sequencing Project (IRGSP). The database can be accessed at http://www. dna.affrc.go.jp:82/giot/INE. html or its mirror site at http://www.staff.or.jp/giot/INE.html

  20. Stakeholder engagement: a key component of integrating genomic information into electronic health records.

    Science.gov (United States)

    Hartzler, Andrea; McCarty, Catherine A; Rasmussen, Luke V; Williams, Marc S; Brilliant, Murray; Bowton, Erica A; Clayton, Ellen Wright; Faucett, William A; Ferryman, Kadija; Field, Julie R; Fullerton, Stephanie M; Horowitz, Carol R; Koenig, Barbara A; McCormick, Jennifer B; Ralston, James D; Sanderson, Saskia C; Smith, Maureen E; Trinidad, Susan Brown

    2013-10-01

    Integrating genomic information into clinical care and the electronic health record can facilitate personalized medicine through genetically guided clinical decision support. Stakeholder involvement is critical to the success of these implementation efforts. Prior work on implementation of clinical information systems provides broad guidance to inform effective engagement strategies. We add to this evidence-based recommendations that are specific to issues at the intersection of genomics and the electronic health record. We describe stakeholder engagement strategies employed by the Electronic Medical Records and Genomics Network, a national consortium of US research institutions funded by the National Human Genome Research Institute to develop, disseminate, and apply approaches that combine genomic and electronic health record data. Through select examples drawn from sites of the Electronic Medical Records and Genomics Network, we illustrate a continuum of engagement strategies to inform genomic integration into commercial and homegrown electronic health records across a range of health-care settings. We frame engagement as activities to consult, involve, and partner with key stakeholder groups throughout specific phases of health information technology implementation. Our aim is to provide insights into engagement strategies to guide genomic integration based on our unique network experiences and lessons learned within the broader context of implementation research in biomedical informatics. On the basis of our collective experience, we describe key stakeholder practices, challenges, and considerations for successful genomic integration to support personalized medicine.

  1. KAIKObase: An integrated silkworm genome database and data mining tool

    Directory of Open Access Journals (Sweden)

    Nagaraju Javaregowda

    2009-10-01

    Full Text Available Abstract Background The silkworm, Bombyx mori, is one of the most economically important insects in many developing countries owing to its large-scale cultivation for silk production. With the development of genomic and biotechnological tools, B. mori has also become an important bioreactor for production of various recombinant proteins of biomedical interest. In 2004, two genome sequencing projects for B. mori were reported independently by Chinese and Japanese teams; however, the datasets were insufficient for building long genomic scaffolds which are essential for unambiguous annotation of the genome. Now, both the datasets have been merged and assembled through a joint collaboration between the two groups. Description Integration of the two data sets of silkworm whole-genome-shotgun sequencing by the Japanese and Chinese groups together with newly obtained fosmid- and BAC-end sequences produced the best continuity (~3.7 Mb in N50 scaffold size among the sequenced insect genomes and provided a high degree of nucleotide coverage (88% of all 28 chromosomes. In addition, a physical map of BAC contigs constructed by fingerprinting BAC clones and a SNP linkage map constructed using BAC-end sequences were available. In parallel, proteomic data from two-dimensional polyacrylamide gel electrophoresis in various tissues and developmental stages were compiled into a silkworm proteome database. Finally, a Bombyx trap database was constructed for documenting insertion positions and expression data of transposon insertion lines. Conclusion For efficient usage of genome information for functional studies, genomic sequences, physical and genetic map information and EST data were compiled into KAIKObase, an integrated silkworm genome database which consists of 4 map viewers, a gene viewer, and sequence, keyword and position search systems to display results and data at the level of nucleotide sequence, gene, scaffold and chromosome. Integration of the

  2. Integrated analysis of whole genome and transcriptome sequencing reveals diverse transcriptomic aberrations driven by somatic genomic changes in liver cancers.

    Directory of Open Access Journals (Sweden)

    Yuichi Shiraishi

    Full Text Available Recent studies applying high-throughput sequencing technologies have identified several recurrently mutated genes and pathways in multiple cancer genomes. However, transcriptional consequences from these genomic alterations in cancer genome remain unclear. In this study, we performed integrated and comparative analyses of whole genomes and transcriptomes of 22 hepatitis B virus (HBV-related hepatocellular carcinomas (HCCs and their matched controls. Comparison of whole genome sequence (WGS and RNA-Seq revealed much evidence that various types of genomic mutations triggered diverse transcriptional changes. Not only splice-site mutations, but also silent mutations in coding regions, deep intronic mutations and structural changes caused splicing aberrations. HBV integrations generated diverse patterns of virus-human fusion transcripts depending on affected gene, such as TERT, CDK15, FN1 and MLL4. Structural variations could drive over-expression of genes such as WNT ligands, with/without creating gene fusions. Furthermore, by taking account of genomic mutations causing transcriptional aberrations, we could improve the sensitivity of deleterious mutation detection in known cancer driver genes (TP53, AXIN1, ARID2, RPS6KA3, and identified recurrent disruptions in putative cancer driver genes such as HNF4A, CPS1, TSC1 and THRAP3 in HCCs. These findings indicate genomic alterations in cancer genome have diverse transcriptomic effects, and integrated analysis of WGS and RNA-Seq can facilitate the interpretation of a large number of genomic alterations detected in cancer genome.

  3. Shifting dominance of riparian Populus and Tamarix along gradients of flow alteration in western North American rivers.

    Science.gov (United States)

    Merritt, David M; Poff, N LeRoy

    2010-01-01

    Tamarix ramosissima is a naturalized, nonnative plant species which has become widespread along riparian corridors throughout the western United States. We test the hypothesis that the distribution and success of Tamarix result from human modification of river-flow regimes. We conducted a natural experiment in eight ecoregions in arid and semiarid portions of the western United States, measuring Tamarix and native Populus recruitment and abundance at 64 sites along 13 perennial rivers spanning a range of altered flow regimes. We quantified biologically relevant attributes of flow alteration as an integrated measure (the index of flow modification, IFM), which was then used to explain between-site variation in abundance and recruitment of native and nonnative riparian plant species. We found the likelihood of successful recruitment of Tamarix to be highest along unregulated river reaches and to remain high across a gradient of regulated flows. Recruitment probability for Populus, in contrast, was highest under free-flowing conditions and declined abruptly under even slight flow modification (IFM > 0.1). Adult Tamarix was most abundant at intermediate levels of IFM. Populus abundance declined sharply with modest flow regulation (IFM > 0.2) and was not present at the most flow-regulated sites. Dominance of Tamarix was highest along rivers with the most altered flow regimes. At the 16 least regulated sites, Tamarix and Populus were equally abundant. Given observed patterns of Tamarix recruitment and abundance, we infer that Tamarix would likely have naturalized, spread, and established widely in riparian communities in the absence of dam construction, diversions, and flow regulation in western North America. However, Tamarix dominance over native species would likely be less extensive in the absence of human alteration of river-flow regimes. Restoration that combines active mechanical removal of established stands of Tamarix with a program of flow releases conducive to

  4. Biomass and genotype × environment interactions of Populus energy crops in the midwestern United States

    Science.gov (United States)

    Ronald S., Jr. Zalesny; Richard B. Hall; Jill A. Zalesny; Bernard G. McMahon; William E. Berguson; Glen R. Stanosz

    2009-01-01

    Using Populus feedstocks for biofuels, bioenergy, and bioproducts is becoming economically feasible as global fossil fuel prices increase. Maximizing Populus biomass production across regional landscapes largely depends on understanding genotype × environment interactions, given broad genetic variation at strategic (...

  5. Barrier to gene flow between two ecologically divergent Populus species, P. alba (white poplar) and P. tremula (European aspen): the role of ecology and life history in gene introgression.

    Science.gov (United States)

    Lexer, C; Fay, M F; Joseph, J A; Nica, M-S; Heinze, B

    2005-04-01

    The renewed interest in the use of hybrid zones for studying speciation calls for the identification and study of hybrid zones across a wide range of organisms, especially in long-lived taxa for which it is often difficult to generate interpopulation variation through controlled crosses. Here, we report on the extent and direction of introgression between two members of the "model tree" genus Populus: Populus alba (white poplar) and Populus tremula (European aspen), across a large zone of sympatry located in the Danube valley. We genotyped 93 hybrid morphotypes and samples from four parental reference populations from within and outside the zone of sympatry for a genome-wide set of 20 nuclear microsatellites and eight plastid DNA restriction site polymorphisms. Our results indicate that introgression occurs preferentially from P. tremula to P. alba via P. tremula pollen. This unidirectional pattern is facilitated by high levels of pollen vs. seed dispersal in P. tremula (pollen/seed flow = 23.9) and by great ecological opportunity in the lowland floodplain forest in proximity to P. alba seed parents, which maintains gene flow in the direction of P. alba despite smaller effective population sizes (N(e)) in this species (P. alba N(e)c. 500-550; P. tremula N(e)c. 550-700). Our results indicate that hybrid zones will be valuable tools for studying the genetic architecture of the barrier to gene flow between these two ecologically divergent Populus species.

  6. PGSB/MIPS PlantsDB Database Framework for the Integration and Analysis of Plant Genome Data.

    Science.gov (United States)

    Spannagl, Manuel; Nussbaumer, Thomas; Bader, Kai; Gundlach, Heidrun; Mayer, Klaus F X

    2017-01-01

    Plant Genome and Systems Biology (PGSB), formerly Munich Institute for Protein Sequences (MIPS) PlantsDB, is a database framework for the integration and analysis of plant genome data, developed and maintained for more than a decade now. Major components of that framework are genome databases and analysis resources focusing on individual (reference) genomes providing flexible and intuitive access to data. Another main focus is the integration of genomes from both model and crop plants to form a scaffold for comparative genomics, assisted by specialized tools such as the CrowsNest viewer to explore conserved gene order (synteny). Data exchange and integrated search functionality with/over many plant genome databases is provided within the transPLANT project.

  7. Patterns of genetic diversity and differentiation in resistance gene clusters of two hybridizing European Populus species

    OpenAIRE

    Casey, Céline; Stölting, Kai N.; Barbará, Thelma; González-Martínez, Santiago C.; Lexer, Christian

    2015-01-01

    Resistance genes (R-genes) are essential for long-lived organisms such as forest trees, which are exposed to diverse herbivores and pathogens. In short-lived model species, R-genes have been shown to be involved in species isolation. Here, we studied more than 400 trees from two natural hybrid zones of the European Populus species Populus alba and Populus tremula for microsatellite markers located in three R-gene clusters, including one cluster situated in the incipient sex chromosome region....

  8. Het geslacht Populus in verband met zijn beteekenis voor de houtteelt = The genus populus and its significance in silviculture

    NARCIS (Netherlands)

    Houtzagers, G.

    1937-01-01

    The genus Populus L. can be divided into 5 sections. This study deals with the classification and description of the species and varieties of the section Aigeiros Duby (black poplars), which contains almost all the important cultivated types in the Netherlands. The botanical information was

  9. Kinetic modeling of batch fermentation for Populus hydrolysate tolerant mutant and wild type strains of Clostridium thermocellum.

    Science.gov (United States)

    Linville, Jessica L; Rodriguez, Miguel; Mielenz, Jonathan R; Cox, Chris D

    2013-11-01

    The extent of inhibition of two strains of Clostridium thermocellum by a Populus hydrolysate was investigated. A Monod-based model of wild type (WT) and Populus hydrolysate tolerant mutant (PM) strains of the cellulolytic bacterium C. thermocellum was developed to quantify growth kinetics in standard media and the extent of inhibition to a Populus hydrolysate. The PM was characterized by a higher growth rate (μmax=1.223 vs. 0.571 h(-1)) and less inhibition (KI,gen=0.991 vs. 0.757) in 10% v/v Populus hydrolysate compared to the WT. In 17.5% v/v Populus hydrolysate inhibition of PM increased slightly (KI,gen=0.888), whereas the WT was strongly inhibited and did not grow in a reproducible manner. Of the individual inhibitors tested, 4-hydroxybenzoic acid was the most inhibitory, followed by galacturonic acid. The PM did not have a greater ability to detoxify the hydrolysate than the WT. Copyright © 2013 Elsevier Ltd. All rights reserved.

  10. Toward integration of genomic selection with crop modelling: the development of an integrated approach to predicting rice heading dates.

    Science.gov (United States)

    Onogi, Akio; Watanabe, Maya; Mochizuki, Toshihiro; Hayashi, Takeshi; Nakagawa, Hiroshi; Hasegawa, Toshihiro; Iwata, Hiroyoshi

    2016-04-01

    It is suggested that accuracy in predicting plant phenotypes can be improved by integrating genomic prediction with crop modelling in a single hierarchical model. Accurate prediction of phenotypes is important for plant breeding and management. Although genomic prediction/selection aims to predict phenotypes on the basis of whole-genome marker information, it is often difficult to predict phenotypes of complex traits in diverse environments, because plant phenotypes are often influenced by genotype-environment interaction. A possible remedy is to integrate genomic prediction with crop/ecophysiological modelling, which enables us to predict plant phenotypes using environmental and management information. To this end, in the present study, we developed a novel method for integrating genomic prediction with phenological modelling of Asian rice (Oryza sativa, L.), allowing the heading date of untested genotypes in untested environments to be predicted. The method simultaneously infers the phenological model parameters and whole-genome marker effects on the parameters in a Bayesian framework. By cultivating backcross inbred lines of Koshihikari × Kasalath in nine environments, we evaluated the potential of the proposed method in comparison with conventional genomic prediction, phenological modelling, and two-step methods that applied genomic prediction to phenological model parameters inferred from Nelder-Mead or Markov chain Monte Carlo algorithms. In predicting heading dates of untested lines in untested environments, the proposed and two-step methods tended to provide more accurate predictions than the conventional genomic prediction methods, particularly in environments where phenotypes from environments similar to the target environment were unavailable for training genomic prediction. The proposed method showed greater accuracy in prediction than the two-step methods in all cross-validation schemes tested, suggesting the potential of the integrated approach in

  11. Multiple factors affect pest and pathogen damage on 31 Populus clones in South Carolina

    Science.gov (United States)

    David R. Coyle; Mark D. Coleman; Jaclin A. Durant; Lee A. Newman

    2006-01-01

    Populus species and hybrids have many practical applications, but there is a paucity of data regarding selections that perform well in the southeastern US. We compared pest susceptibility of 31 Populus clones over 3 years in South Carolina, USA. Cuttings were planted in spring 2001 on two study sites. Clones planted in the...

  12. GIGGLE: a search engine for large-scale integrated genome analysis.

    Science.gov (United States)

    Layer, Ryan M; Pedersen, Brent S; DiSera, Tonya; Marth, Gabor T; Gertz, Jason; Quinlan, Aaron R

    2018-02-01

    GIGGLE is a genomics search engine that identifies and ranks the significance of genomic loci shared between query features and thousands of genome interval files. GIGGLE (https://github.com/ryanlayer/giggle) scales to billions of intervals and is over three orders of magnitude faster than existing methods. Its speed extends the accessibility and utility of resources such as ENCODE, Roadmap Epigenomics, and GTEx by facilitating data integration and hypothesis generation.

  13. Integrating genomic selection into dairy cattle breeding programmes: a review.

    Science.gov (United States)

    Bouquet, A; Juga, J

    2013-05-01

    Extensive genetic progress has been achieved in dairy cattle populations on many traits of economic importance because of efficient breeding programmes. Success of these programmes has relied on progeny testing of the best young males to accurately assess their genetic merit and hence their potential for breeding. Over the last few years, the integration of dense genomic information into statistical tools used to make selection decisions, commonly referred to as genomic selection, has enabled gains in predicting accuracy of breeding values for young animals without own performance. The possibility to select animals at an early stage allows defining new breeding strategies aimed at boosting genetic progress while reducing costs. The first objective of this article was to review methods used to model and optimize breeding schemes integrating genomic selection and to discuss their relative advantages and limitations. The second objective was to summarize the main results and perspectives on the use of genomic selection in practical breeding schemes, on the basis of the example of dairy cattle populations. Two main designs of breeding programmes integrating genomic selection were studied in dairy cattle. Genomic selection can be used either for pre-selecting males to be progeny tested or for selecting males to be used as active sires in the population. The first option produces moderate genetic gains without changing the structure of breeding programmes. The second option leads to large genetic gains, up to double those of conventional schemes because of a major reduction in the mean generation interval, but it requires greater changes in breeding programme structure. The literature suggests that genomic selection becomes more attractive when it is coupled with embryo transfer technologies to further increase selection intensity on the dam-to-sire pathway. The use of genomic information also offers new opportunities to improve preservation of genetic variation. However

  14. Integrated Genome-Based Studies of Shewanella Ecophysiology

    Energy Technology Data Exchange (ETDEWEB)

    Andrei L. Osterman, Ph.D.

    2012-12-17

    Integration of bioinformatics and experimental techniques was applied to mapping and characterization of the key components (pathways, enzymes, transporters, regulators) of the core metabolic machinery in Shewanella oneidensis and related species with main focus was on metabolic and regulatory pathways involved in utilization of various carbon and energy sources. Among the main accomplishments reflected in ten joint publications with other participants of Shewanella Federation are: (i) A systems-level reconstruction of carbohydrate utilization pathways in the genus of Shewanella (19 species). This analysis yielded reconstruction of 18 sugar utilization pathways including 10 novel pathway variants and prediction of > 60 novel protein families of enzymes, transporters and regulators involved in these pathways. Selected functional predictions were verified by focused biochemical and genetic experiments. Observed growth phenotypes were consistent with bioinformatic predictions providing strong validation of the technology and (ii) Global genomic reconstruction of transcriptional regulons in 16 Shewanella genomes. The inferred regulatory network includes 82 transcription factors, 8 riboswitches and 6 translational attenuators. Of those, 45 regulons were inferred directly from the genome context analysis, whereas others were propagated from previously characterized regulons in other species. Selected regulatory predictions were experimentally tested. Integration of this analysis with microarray data revealed overall consistency and provided additional layer of interactions between regulons. All the results were captured in the new database RegPrecise, which is a joint development with the LBNL team. A more detailed analysis of the individual subsystems, pathways and regulons in Shewanella spp included bioinfiormatics-based prediction and experimental characterization of: (i) N-Acetylglucosamine catabolic pathway; (ii)Lactate utilization machinery; (iii) Novel Nrt

  15. Populus (Salicaceae plantations

    Directory of Open Access Journals (Sweden)

    Gonzalo M. Romano

    2013-01-01

    Full Text Available Aunque los cultivos forestales son comunidades artificiales, modifican condiciones ambientales que pueden alterar la diversidad fúngica nativa. Se estudiaron los efectos del manejo forestal de una plantación de sauces (Salix y álamos (Populus sobre la biodiversidad de Agaromycetes durante un año en una isla del Delta del Paraná, Argentina. Se midieron el peso seco y el número de basidiomas. Se identificaron 28 especies pertenecientes a los Agaricomycetes: 26 especies de Agaricales, una de Polyporales y una de Russulales. Nuestros resultados sugieren que el manejo forestal de dicha plantación no afecta la abundancia ni la diversidad de basidiomas de Agaricomycetes.

  16. INDIGO - INtegrated data warehouse of microbial genomes with examples from the red sea extremophiles.

    KAUST Repository

    Alam, Intikhab

    2013-12-06

    The next generation sequencing technologies substantially increased the throughput of microbial genome sequencing. To functionally annotate newly sequenced microbial genomes, a variety of experimental and computational methods are used. Integration of information from different sources is a powerful approach to enhance such annotation. Functional analysis of microbial genomes, necessary for downstream experiments, crucially depends on this annotation but it is hampered by the current lack of suitable information integration and exploration systems for microbial genomes.

  17. INDIGO - INtegrated data warehouse of microbial genomes with examples from the red sea extremophiles.

    KAUST Repository

    Alam, Intikhab; Antunes, André ; Kamau, Allan; Ba Alawi, Wail; Kalkatawi, Manal M.; Stingl, Ulrich; Bajic, Vladimir B.

    2013-01-01

    The next generation sequencing technologies substantially increased the throughput of microbial genome sequencing. To functionally annotate newly sequenced microbial genomes, a variety of experimental and computational methods are used. Integration of information from different sources is a powerful approach to enhance such annotation. Functional analysis of microbial genomes, necessary for downstream experiments, crucially depends on this annotation but it is hampered by the current lack of suitable information integration and exploration systems for microbial genomes.

  18. GIGGLE: a search engine for large-scale integrated genome analysis

    Science.gov (United States)

    Layer, Ryan M; Pedersen, Brent S; DiSera, Tonya; Marth, Gabor T; Gertz, Jason; Quinlan, Aaron R

    2018-01-01

    GIGGLE is a genomics search engine that identifies and ranks the significance of genomic loci shared between query features and thousands of genome interval files. GIGGLE (https://github.com/ryanlayer/giggle) scales to billions of intervals and is over three orders of magnitude faster than existing methods. Its speed extends the accessibility and utility of resources such as ENCODE, Roadmap Epigenomics, and GTEx by facilitating data integration and hypothesis generation. PMID:29309061

  19. Identification and allelic dissection uncover roles of lncRNAs in secondary growth of Populus tomentosa.

    Science.gov (United States)

    Zhou, Daling; Du, Qingzhang; Chen, Jinhui; Wang, Qingshi; Zhang, Deqiang

    2017-10-01

    Long non-coding RNAs (lncRNAs) function in various biological processes. However, their roles in secondary growth of plants remain poorly understood. Here, 15,691 lncRNAs were identified from vascular cambium, developing xylem, and mature xylem of Populus tomentosa with high and low biomass using RNA-seq, including 1,994 lncRNAs that were differentially expressed (DE) among the six libraries. 3,569 cis-regulated and 3,297 trans-regulated protein-coding genes were predicted as potential target genes (PTGs) of the DE lncRNAs to participate in biological regulation. Then, 476 and 28 lncRNAs were identified as putative targets and endogenous target mimics (eTMs) of Populus known microRNAs (miRNAs), respectively. Genome re-sequencing of 435 individuals from a natural population of P. tomentosa found 34,015 single nucleotide polymorphisms (SNPs) within 178 lncRNA loci and 522 PTGs. Single-SNP associations analysis detected 2,993 associations with 10 growth and wood-property traits under additive and dominance model. Epistasis analysis identified 17,656 epistatic SNP pairs, providing evidence for potential regulatory interactions between lncRNAs and their PTGs. Furthermore, a reconstructed epistatic network, representing interactions of 8 lncRNAs and 15 PTGs, might enrich regulation roles of genes in the phenylpropanoid pathway. These findings may enhance our understanding of non-coding genes in plants. © The Author 2017. Published by Oxford University Press on behalf of Kazusa DNA Research Institute.

  20. Structured Matrix Completion with Applications to Genomic Data Integration.

    Science.gov (United States)

    Cai, Tianxi; Cai, T Tony; Zhang, Anru

    2016-01-01

    Matrix completion has attracted significant recent attention in many fields including statistics, applied mathematics and electrical engineering. Current literature on matrix completion focuses primarily on independent sampling models under which the individual observed entries are sampled independently. Motivated by applications in genomic data integration, we propose a new framework of structured matrix completion (SMC) to treat structured missingness by design. Specifically, our proposed method aims at efficient matrix recovery when a subset of the rows and columns of an approximately low-rank matrix are observed. We provide theoretical justification for the proposed SMC method and derive lower bound for the estimation errors, which together establish the optimal rate of recovery over certain classes of approximately low-rank matrices. Simulation studies show that the method performs well in finite sample under a variety of configurations. The method is applied to integrate several ovarian cancer genomic studies with different extent of genomic measurements, which enables us to construct more accurate prediction rules for ovarian cancer survival.

  1. Metabolic functions of Pseudomonas fluorescens strains from Populus deltoides depend on rhizosphere or endosphere isolation compartment

    Directory of Open Access Journals (Sweden)

    Collin M Timm

    2015-10-01

    Full Text Available The bacterial microbiota of plants is diverse, with 1,000s of operational taxonomic units (OTUs associated with any individual plant. In this work we investigate the differences between 19 sequenced Pseudomonas fluorescens strains, isolated from Populus deltoides rhizosphere and endosphere and which represent a single OTU, using phenotypic analysis, comparative genomics, and metabolic models. While no traits were exclusive to either endosphere or rhizosphere P. fluorescens isolates, multiple pathways relevant for plant-bacterial interactions are enriched in endosphere isolate genomes. Further, growth phenotypes such as phosphate solubilization, protease activity, denitrification and root growth promotion are biased towards endosphere isolates. Endosphere isolates have significantly more metabolic pathways for plant signaling compounds and an increased metabolic range that includes utilization of energy rich nucleotides and sugars, consistent with endosphere colonization. Rhizosphere P. fluorescens have fewer pathways representative of plant-bacterial interactions but show metabolic bias towards chemical substrates often found in root exudates. This work reveals the diverse functions that may contribute to colonization of the endosphere by bacteria and are enriched among closely related isolates.

  2. The Populus superoxide dismutase gene family and its responses to drought stress in transgenic poplar overexpressing a pine cytosolic glutamine synthetase (GS1a.

    Directory of Open Access Journals (Sweden)

    Juan Jesús Molina-Rueda

    Full Text Available BACKGROUND: Glutamine synthetase (GS plays a central role in plant nitrogen assimilation, a process intimately linked to soil water availability. We previously showed that hybrid poplar (Populus tremula X alba, INRA 717-1B4 expressing ectopically a pine cytosolic glutamine synthetase gene (GS1a display enhanced tolerance to drought. Preliminary transcriptome profiling revealed that during drought, members of the superoxide dismutase (SOD family were reciprocally regulated in GS poplar when compared with the wild-type control, in all tissues examined. SOD was the only gene family found to exhibit such patterns. RESULTS: In silico analysis of the Populus genome identified 12 SOD genes and two genes encoding copper chaperones for SOD (CCSs. The poplar SODs form three phylogenetic clusters in accordance with their distinct metal co-factor requirements and gene structure. Nearly all poplar SODs and CCSs are present in duplicate derived from whole genome duplication, in sharp contrast to their predominantly single-copy Arabidopsis orthologs. Drought stress triggered plant-wide down-regulation of the plastidic copper SODs (CSDs, with concomitant up-regulation of plastidic iron SODs (FSDs in GS poplar relative to the wild type; this was confirmed at the activity level. We also found evidence for coordinated down-regulation of other copper proteins, including plastidic CCSs and polyphenol oxidases, in GS poplar under drought conditions. CONCLUSIONS: Both gene duplication and expression divergence have contributed to the expansion and transcriptional diversity of the Populus SOD/CCS families. Coordinated down-regulation of major copper proteins in drought-tolerant GS poplars supports the copper cofactor economy model where copper supply is preferentially allocated for plastocyanins to sustain photosynthesis during drought. Our results also extend previous findings on the compensatory regulation between chloroplastic CSDs and FSDs, and suggest that this

  3. Integrated proteomic and genomic analysis of colorectal cancer

    Science.gov (United States)

    Investigators who analyzed 95 human colorectal tumor samples have determined how gene alterations identified in previous analyses of the same samples are expressed at the protein level. The integration of proteomic and genomic data, or proteogenomics, pro

  4. Integration of expression data in genome-scale metabolic network reconstructions

    Directory of Open Access Journals (Sweden)

    Anna S. Blazier

    2012-08-01

    Full Text Available With the advent of high-throughput technologies, the field of systems biology has amassed an abundance of omics data, quantifying thousands of cellular components across a variety of scales, ranging from mRNA transcript levels to metabolite quantities. Methods are needed to not only integrate this omics data but to also use this data to heighten the predictive capabilities of computational models. Several recent studies have successfully demonstrated how flux balance analysis (FBA, a constraint-based modeling approach, can be used to integrate transcriptomic data into genome-scale metabolic network reconstructions to generate predictive computational models. In this review, we summarize such FBA-based methods for integrating expression data into genome-scale metabolic network reconstructions, highlighting their advantages as well as their limitations.

  5. Micropropagation, genetic engineering, and molecular biology of Populus

    Science.gov (United States)

    N. B. Klopfenstein; Y. W. Chun; M. -S. Kim; M. A. Ahuja; M. C. Dillon; R. C. Carman; L. G. Eskew

    1997-01-01

    Thirty-four Populus biotechnology chapters, written by 85 authors, are comprised in 5 sections: 1) in vitro culture (micropropagation, somatic embryogenesis, protoplasts, somaclonal variation, and germplasm preservation); 2) transformation and foreign gene expression; 3) molecular biology (molecular/genetic characterization); 4) biotic and abiotic resistance (disease,...

  6. High throughput platforms for structural genomics of integral membrane proteins.

    Science.gov (United States)

    Mancia, Filippo; Love, James

    2011-08-01

    Structural genomics approaches on integral membrane proteins have been postulated for over a decade, yet specific efforts are lagging years behind their soluble counterparts. Indeed, high throughput methodologies for production and characterization of prokaryotic integral membrane proteins are only now emerging, while large-scale efforts for eukaryotic ones are still in their infancy. Presented here is a review of recent literature on actively ongoing structural genomics of membrane protein initiatives, with a focus on those aimed at implementing interesting techniques aimed at increasing our rate of success for this class of macromolecules. Copyright © 2011 Elsevier Ltd. All rights reserved.

  7. Improving Microbial Genome Annotations in an Integrated Database Context

    Science.gov (United States)

    Chen, I-Min A.; Markowitz, Victor M.; Chu, Ken; Anderson, Iain; Mavromatis, Konstantinos; Kyrpides, Nikos C.; Ivanova, Natalia N.

    2013-01-01

    Effective comparative analysis of microbial genomes requires a consistent and complete view of biological data. Consistency regards the biological coherence of annotations, while completeness regards the extent and coverage of functional characterization for genomes. We have developed tools that allow scientists to assess and improve the consistency and completeness of microbial genome annotations in the context of the Integrated Microbial Genomes (IMG) family of systems. All publicly available microbial genomes are characterized in IMG using different functional annotation and pathway resources, thus providing a comprehensive framework for identifying and resolving annotation discrepancies. A rule based system for predicting phenotypes in IMG provides a powerful mechanism for validating functional annotations, whereby the phenotypic traits of an organism are inferred based on the presence of certain metabolic reactions and pathways and compared to experimentally observed phenotypes. The IMG family of systems are available at http://img.jgi.doe.gov/. PMID:23424620

  8. Improving microbial genome annotations in an integrated database context.

    Directory of Open Access Journals (Sweden)

    I-Min A Chen

    Full Text Available Effective comparative analysis of microbial genomes requires a consistent and complete view of biological data. Consistency regards the biological coherence of annotations, while completeness regards the extent and coverage of functional characterization for genomes. We have developed tools that allow scientists to assess and improve the consistency and completeness of microbial genome annotations in the context of the Integrated Microbial Genomes (IMG family of systems. All publicly available microbial genomes are characterized in IMG using different functional annotation and pathway resources, thus providing a comprehensive framework for identifying and resolving annotation discrepancies. A rule based system for predicting phenotypes in IMG provides a powerful mechanism for validating functional annotations, whereby the phenotypic traits of an organism are inferred based on the presence of certain metabolic reactions and pathways and compared to experimentally observed phenotypes. The IMG family of systems are available at http://img.jgi.doe.gov/.

  9. INDIGO - INtegrated data warehouse of microbial genomes with examples from the red sea extremophiles.

    Science.gov (United States)

    Alam, Intikhab; Antunes, André; Kamau, Allan Anthony; Ba Alawi, Wail; Kalkatawi, Manal; Stingl, Ulrich; Bajic, Vladimir B

    2013-01-01

    The next generation sequencing technologies substantially increased the throughput of microbial genome sequencing. To functionally annotate newly sequenced microbial genomes, a variety of experimental and computational methods are used. Integration of information from different sources is a powerful approach to enhance such annotation. Functional analysis of microbial genomes, necessary for downstream experiments, crucially depends on this annotation but it is hampered by the current lack of suitable information integration and exploration systems for microbial genomes. We developed a data warehouse system (INDIGO) that enables the integration of annotations for exploration and analysis of newly sequenced microbial genomes. INDIGO offers an opportunity to construct complex queries and combine annotations from multiple sources starting from genomic sequence to protein domain, gene ontology and pathway levels. This data warehouse is aimed at being populated with information from genomes of pure cultures and uncultured single cells of Red Sea bacteria and Archaea. Currently, INDIGO contains information from Salinisphaera shabanensis, Haloplasma contractile, and Halorhabdus tiamatea - extremophiles isolated from deep-sea anoxic brine lakes of the Red Sea. We provide examples of utilizing the system to gain new insights into specific aspects on the unique lifestyle and adaptations of these organisms to extreme environments. We developed a data warehouse system, INDIGO, which enables comprehensive integration of information from various resources to be used for annotation, exploration and analysis of microbial genomes. It will be regularly updated and extended with new genomes. It is aimed to serve as a resource dedicated to the Red Sea microbes. In addition, through INDIGO, we provide our Automatic Annotation of Microbial Genomes (AAMG) pipeline. The INDIGO web server is freely available at http://www.cbrc.kaust.edu.sa/indigo.

  10. Evolution of endogenous non-retroviral genes integrated into plant genomes

    Directory of Open Access Journals (Sweden)

    Hyosub Chu

    2014-08-01

    Full Text Available Numerous comparative genome analyses have revealed the wide extent of horizontal gene transfer (HGT in living organisms, which contributes to their evolution and genetic diversity. Viruses play important roles in HGT. Endogenous viral elements (EVEs are defined as viral DNA sequences present within the genomes of non-viral organisms. In eukaryotic cells, the majority of EVEs are derived from RNA viruses using reverse transcription. In contrast, endogenous non-retroviral elements (ENREs are poorly studied. However, the increasing availability of genomic data and the rapid development of bioinformatics tools have enabled the identification of several ENREs in various eukaryotic organisms. To date, a small number of ENREs integrated into plant genomes have been identified. Of the known non-retroviruses, most identified ENREs are derived from double-strand (ds RNA viruses, followed by single-strand (ss DNA and ssRNA viruses. At least eight virus families have been identified. Of these, viruses in the family Partitiviridae are dominant, followed by viruses of the families Chrysoviridae and Geminiviridae. The identified ENREs have been primarily identified in eudicots, followed by monocots. In this review, we briefly discuss the current view on non-retroviral sequences integrated into plant genomes that are associated with plant-virus evolution and their possible roles in antiviral resistance.

  11. Genome-Wide Analysis of Transposon and Retroviral Insertions Reveals Preferential Integrations in Regions of DNA Flexibility.

    Science.gov (United States)

    Vrljicak, Pavle; Tao, Shijie; Varshney, Gaurav K; Quach, Helen Ngoc Bao; Joshi, Adita; LaFave, Matthew C; Burgess, Shawn M; Sampath, Karuna

    2016-04-07

    DNA transposons and retroviruses are important transgenic tools for genome engineering. An important consideration affecting the choice of transgenic vector is their insertion site preferences. Previous large-scale analyses of Ds transposon integration sites in plants were done on the basis of reporter gene expression or germ-line transmission, making it difficult to discern vertebrate integration preferences. Here, we compare over 1300 Ds transposon integration sites in zebrafish with Tol2 transposon and retroviral integration sites. Genome-wide analysis shows that Ds integration sites in the presence or absence of marker selection are remarkably similar and distributed throughout the genome. No strict motif was found, but a preference for structural features in the target DNA associated with DNA flexibility (Twist, Tilt, Rise, Roll, Shift, and Slide) was observed. Remarkably, this feature is also found in transposon and retroviral integrations in maize and mouse cells. Our findings show that structural features influence the integration of heterologous DNA in genomes, and have implications for targeted genome engineering. Copyright © 2016 Vrljicak et al.

  12. Performance of Salix viminalis and Populus nigra x Populus maximowiczii in short rotation intensive culture under high irrigation

    Energy Technology Data Exchange (ETDEWEB)

    Fillion, Maud; Brisson, Jacques [Departement de Sciences biologiques, Universite de Montreal, C.P. 6128, succ. Centre-ville, Montreal, Quebec (Canada); Institut de recherche en biologie vegetale, 4101 Sherbrooke East, Montreal, Quebec (Canada); Teodorescu, Traian I.; Labrecque, Michel [Institut de recherche en biologie vegetale, 4101 Sherbrooke East, Montreal, Quebec (Canada); Sauve, Sebastien [Departement de chimie, Universite de Montreal, C.P. 6128, succursale Centre-ville, Montreal, Quebec (Canada)

    2009-09-15

    On a plantation established in 2004 from stem cuttings at a density of 20,000 trees per hectare, we investigated growth and nutritional plant response to a high hydraulic regime for two species (Salix viminalis and Populus nigra x Populus maximowiczii), using a comparative approach with measurements from irrigated and control plots. The plantation was irrigated from June to September 2005 with about 140 mm per day. The equivalent of 120 Kg NO{sub 3}-N, 40 Kg P{sub 2}O{sub 5}-P and 85 Kg K{sub 2}O-K per hectare per year was applied by means of irrigation with wastewater. No mortality occurred and stem biomass production of both poplar and willow species were not statistically different on irrigated and control areas. However, S. viminalis revealed to be more tolerant to flooded conditions since these corresponded more closely to its nutritional requirements (foliar concentration of 20 mgN g{sup -1}). The capacity of S. viminalis to withstand waterlogged conditions could play an important role in the sustainability of a plantation for the filtration of effluent at low pollutant concentration. (author)

  13. Genome-wide survey of recurrent HBV integration in hepatocellular carcinoma

    DEFF Research Database (Denmark)

    Sung, Wing-Kin; Zheng, Hancheng; Li, Shuyu

    2012-01-01

    To survey hepatitis B virus (HBV) integration in liver cancer genomes, we conducted massively parallel sequencing of 81 HBV-positive and 7 HBV-negative hepatocellular carcinomas (HCCs) and adjacent normal tissues. We found that HBV integration is observed more frequently in the tumors (86.4%) than...

  14. Effects of Clone, Silvicultural, and Miticide Treatments on Cottonwood Leafcurl Mite (Acari: Eriophyidae) Damage in Plantation Populus

    Science.gov (United States)

    David R. Coyle

    2002-01-01

    Aculops lobuliferus (Keifer) is a little known pest of plantation Populus spp., which is capable of causing substantial damage. This is the First documented occurrence of A. lobuliferus in South Carolina. Previous anecdotal data indicated clonal variation in Populus susceptibility to A...

  15. INDIGO - INtegrated data warehouse of microbial genomes with examples from the red sea extremophiles.

    Directory of Open Access Journals (Sweden)

    Intikhab Alam

    Full Text Available The next generation sequencing technologies substantially increased the throughput of microbial genome sequencing. To functionally annotate newly sequenced microbial genomes, a variety of experimental and computational methods are used. Integration of information from different sources is a powerful approach to enhance such annotation. Functional analysis of microbial genomes, necessary for downstream experiments, crucially depends on this annotation but it is hampered by the current lack of suitable information integration and exploration systems for microbial genomes.We developed a data warehouse system (INDIGO that enables the integration of annotations for exploration and analysis of newly sequenced microbial genomes. INDIGO offers an opportunity to construct complex queries and combine annotations from multiple sources starting from genomic sequence to protein domain, gene ontology and pathway levels. This data warehouse is aimed at being populated with information from genomes of pure cultures and uncultured single cells of Red Sea bacteria and Archaea. Currently, INDIGO contains information from Salinisphaera shabanensis, Haloplasma contractile, and Halorhabdus tiamatea - extremophiles isolated from deep-sea anoxic brine lakes of the Red Sea. We provide examples of utilizing the system to gain new insights into specific aspects on the unique lifestyle and adaptations of these organisms to extreme environments.We developed a data warehouse system, INDIGO, which enables comprehensive integration of information from various resources to be used for annotation, exploration and analysis of microbial genomes. It will be regularly updated and extended with new genomes. It is aimed to serve as a resource dedicated to the Red Sea microbes. In addition, through INDIGO, we provide our Automatic Annotation of Microbial Genomes (AAMG pipeline. The INDIGO web server is freely available at http://www.cbrc.kaust.edu.sa/indigo.

  16. Selection against recombinant hybrids maintains reproductive isolation in hybridizing Populus species despite F1 fertility and recurrent gene flow.

    Science.gov (United States)

    Christe, Camille; Stölting, Kai N; Bresadola, Luisa; Fussi, Barbara; Heinze, Berthold; Wegmann, Daniel; Lexer, Christian

    2016-06-01

    Natural hybrid zones have proven to be precious tools for understanding the origin and maintenance of reproductive isolation (RI) and therefore species. Most available genomic studies of hybrid zones using whole- or partial-genome resequencing approaches have focused on comparisons of the parental source populations involved in genome admixture, rather than exploring fine-scale patterns of chromosomal ancestry across the full admixture gradient present between hybridizing species. We have studied three well-known European 'replicate' hybrid zones of Populus alba and P. tremula, two widespread, ecologically divergent forest trees, using up to 432 505 single-nucleotide polymorphisms (SNPs) from restriction site-associated DNA (RAD) sequencing. Estimates of fine-scale chromosomal ancestry, genomic divergence and differentiation across all 19 poplar chromosomes revealed strikingly contrasting results, including an unexpected preponderance of F1 hybrids in the centre of genomic clines on the one hand, and genomically localized, spatially variable shared variants consistent with ancient introgression between the parental species on the other. Genetic ancestry had a significant effect on survivorship of hybrid seedlings in a common garden trial, pointing to selection against early-generation recombinants. Our results indicate a role for selection against recombinant genotypes in maintaining RI in the face of apparent F1 fertility, consistent with the intragenomic 'coadaptation' model of barriers to introgression upon secondary contact. Whole-genome resequencing of hybridizing populations will clarify the roles of specific genetic pathways in RI between these model forest trees and may reveal which loci are affected most strongly by its cyclic breakdown. © 2016 John Wiley & Sons Ltd.

  17. Integrated Genome-Based Studies of Shewanella Echophysiology

    Energy Technology Data Exchange (ETDEWEB)

    Margrethe H. Serres

    2012-06-29

    Shewanella oneidensis MR-1 is a motile, facultative {gamma}-Proteobacterium with remarkable respiratory versatility; it can utilize a range of organic and inorganic compounds as terminal electronacceptors for anaerobic metabolism. The ability to effectively reduce nitrate, S0, polyvalent metals andradionuclides has established MR-1 as an important model dissimilatory metal-reducing microorganism for genome-based investigations of biogeochemical transformation of metals and radionuclides that are of concern to the U.S. Department of Energy (DOE) sites nationwide. Metal-reducing bacteria such as Shewanella also have a highly developed capacity for extracellular transfer of respiratory electrons to solid phase Fe and Mn oxides as well as directly to anode surfaces in microbial fuel cells. More broadly, Shewanellae are recognized free-living microorganisms and members of microbial communities involved in the decomposition of organic matter and the cycling of elements in aquatic and sedimentary systems. To function and compete in environments that are subject to spatial and temporal environmental change, Shewanella must be able to sense and respond to such changes and therefore require relatively robust sensing and regulation systems. The overall goal of this project is to apply the tools of genomics, leveraging the availability of genome sequence for 18 additional strains of Shewanella, to better understand the ecophysiology and speciation of respiratory-versatile members of this important genus. To understand these systems we propose to use genome-based approaches to investigate Shewanella as a system of integrated networks; first describing key cellular subsystems - those involved in signal transduction, regulation, and metabolism - then building towards understanding the function of whole cells and, eventually, cells within populations. As a general approach, this project will employ complimentary "top-down" - bioinformatics-based genome functional predictions, high

  18. Site-Specific Integration of Exogenous Genes Using Genome Editing Technologies in Zebrafish

    Directory of Open Access Journals (Sweden)

    Atsuo Kawahara

    2016-05-01

    Full Text Available The zebrafish (Danio rerio is an ideal vertebrate model to investigate the developmental molecular mechanism of organogenesis and regeneration. Recent innovation in genome editing technologies, such as zinc finger nucleases (ZFNs, transcription activator-like effector nucleases (TALENs and the clustered regularly interspaced short palindromic repeats (CRISPR/CRISPR associated protein 9 (Cas9 system, have allowed researchers to generate diverse genomic modifications in whole animals and in cultured cells. The CRISPR/Cas9 and TALEN techniques frequently induce DNA double-strand breaks (DSBs at the targeted gene, resulting in frameshift-mediated gene disruption. As a useful application of genome editing technology, several groups have recently reported efficient site-specific integration of exogenous genes into targeted genomic loci. In this review, we provide an overview of TALEN- and CRISPR/Cas9-mediated site-specific integration of exogenous genes in zebrafish.

  19. Genomic integrity and the ageing brain.

    Science.gov (United States)

    Chow, Hei-man; Herrup, Karl

    2015-11-01

    DNA damage is correlated with and may drive the ageing process. Neurons in the brain are postmitotic and are excluded from many forms of DNA repair; therefore, neurons are vulnerable to various neurodegenerative diseases. The challenges facing the field are to understand how and when neuronal DNA damage accumulates, how this loss of genomic integrity might serve as a 'time keeper' of nerve cell ageing and why this process manifests itself as different diseases in different individuals.

  20. Glycosylation-mediated phenylpropanoid partitioning in Populus tremuloides cell cultures

    Directory of Open Access Journals (Sweden)

    Babst Benjamin A

    2009-12-01

    Full Text Available Abstract Background Phenylpropanoid-derived phenolic glycosides (PGs and condensed tannins (CTs comprise large, multi-purpose non-structural carbon sinks in Populus. A negative correlation between PG and CT concentrations has been observed in several studies. However, the molecular mechanism underlying the relationship is not known. Results Populus cell cultures produce CTs but not PGs under normal conditions. Feeding salicyl alcohol resulted in accumulation of salicins, the simplest PG, in the cells, but not higher-order PGs. Salicin accrual reflected the stimulation of a glycosylation response which altered a number of metabolic activities. We utilized this suspension cell feeding system as a model for analyzing the possible role of glycosylation in regulating the metabolic competition between PG formation, CT synthesis and growth. Cells accumulated salicins in a dose-dependent manner following salicyl alcohol feeding. Higher feeding levels led to a decrease in cellular CT concentrations (at 5 or 10 mM, and a negative effect on cell growth (at 10 mM. The competition between salicin and CT formation was reciprocal, and depended on the metabolic status of the cells. We analyzed gene expression changes between controls and cells fed with 5 mM salicyl alcohol for 48 hr, a time point when salicin accumulation was near maximum and CT synthesis was reduced, with no effect on growth. Several stress-responsive genes were up-regulated, suggestive of a general stress response in the fed cells. Salicyl alcohol feeding also induced expression of genes associated with sucrose catabolism, glycolysis and the Krebs cycle. Transcript levels of phenylalanine ammonia lyase and most of the flavonoid pathway genes were reduced, consistent with down-regulated CT synthesis. Conclusions Exogenous salicyl alcohol was readily glycosylated in Populus cell cultures, a process that altered sugar utilization and phenolic partitioning in the cells. Using this system, we

  1. Brassica database (BRAD) version 2.0: integrating and mining Brassicaceae species genomic resources.

    Science.gov (United States)

    Wang, Xiaobo; Wu, Jian; Liang, Jianli; Cheng, Feng; Wang, Xiaowu

    2015-01-01

    The Brassica database (BRAD) was built initially to assist users apply Brassica rapa and Arabidopsis thaliana genomic data efficiently to their research. However, many Brassicaceae genomes have been sequenced and released after its construction. These genomes are rich resources for comparative genomics, gene annotation and functional evolutionary studies of Brassica crops. Therefore, we have updated BRAD to version 2.0 (V2.0). In BRAD V2.0, 11 more Brassicaceae genomes have been integrated into the database, namely those of Arabidopsis lyrata, Aethionema arabicum, Brassica oleracea, Brassica napus, Camelina sativa, Capsella rubella, Leavenworthia alabamica, Sisymbrium irio and three extremophiles Schrenkiella parvula, Thellungiella halophila and Thellungiella salsuginea. BRAD V2.0 provides plots of syntenic genomic fragments between pairs of Brassicaceae species, from the level of chromosomes to genomic blocks. The Generic Synteny Browser (GBrowse_syn), a module of the Genome Browser (GBrowse), is used to show syntenic relationships between multiple genomes. Search functions for retrieving syntenic and non-syntenic orthologs, as well as their annotation and sequences are also provided. Furthermore, genome and annotation information have been imported into GBrowse so that all functional elements can be visualized in one frame. We plan to continually update BRAD by integrating more Brassicaceae genomes into the database. Database URL: http://brassicadb.org/brad/. © The Author(s) 2015. Published by Oxford University Press.

  2. Improved bacteriophage genome data is necessary for integrating viral and bacterial ecology.

    Science.gov (United States)

    Bibby, Kyle

    2014-02-01

    The recent rise in "omics"-enabled approaches has lead to improved understanding in many areas of microbial ecology. However, despite the importance that viruses play in a broad microbial ecology context, viral ecology remains largely not integrated into high-throughput microbial ecology studies. A fundamental hindrance to the integration of viral ecology into omics-enabled microbial ecology studies is the lack of suitable reference bacteriophage genomes in reference databases-currently, only 0.001% of bacteriophage diversity is represented in genome sequence databases. This commentary serves to highlight this issue and to promote bacteriophage genome sequencing as a valuable scientific undertaking to both better understand bacteriophage diversity and move towards a more holistic view of microbial ecology.

  3. Modeling the integration of bacterial rRNA fragments into the human cancer genome.

    Science.gov (United States)

    Sieber, Karsten B; Gajer, Pawel; Dunning Hotopp, Julie C

    2016-03-21

    Cancer is a disease driven by the accumulation of genomic alterations, including the integration of exogenous DNA into the human somatic genome. We previously identified in silico evidence of DNA fragments from a Pseudomonas-like bacteria integrating into the 5'-UTR of four proto-oncogenes in stomach cancer sequencing data. The functional and biological consequences of these bacterial DNA integrations remain unknown. Modeling of these integrations suggests that the previously identified sequences cover most of the sequence flanking the junction between the bacterial and human DNA. Further examination of these reads reveals that these integrations are rich in guanine nucleotides and the integrated bacterial DNA may have complex transcript secondary structures. The models presented here lay the foundation for future experiments to test if bacterial DNA integrations alter the transcription of the human genes.

  4. An integrated semiconductor device enabling non-optical genome sequencing.

    Science.gov (United States)

    Rothberg, Jonathan M; Hinz, Wolfgang; Rearick, Todd M; Schultz, Jonathan; Mileski, William; Davey, Mel; Leamon, John H; Johnson, Kim; Milgrew, Mark J; Edwards, Matthew; Hoon, Jeremy; Simons, Jan F; Marran, David; Myers, Jason W; Davidson, John F; Branting, Annika; Nobile, John R; Puc, Bernard P; Light, David; Clark, Travis A; Huber, Martin; Branciforte, Jeffrey T; Stoner, Isaac B; Cawley, Simon E; Lyons, Michael; Fu, Yutao; Homer, Nils; Sedova, Marina; Miao, Xin; Reed, Brian; Sabina, Jeffrey; Feierstein, Erika; Schorn, Michelle; Alanjary, Mohammad; Dimalanta, Eileen; Dressman, Devin; Kasinskas, Rachel; Sokolsky, Tanya; Fidanza, Jacqueline A; Namsaraev, Eugeni; McKernan, Kevin J; Williams, Alan; Roth, G Thomas; Bustillo, James

    2011-07-20

    The seminal importance of DNA sequencing to the life sciences, biotechnology and medicine has driven the search for more scalable and lower-cost solutions. Here we describe a DNA sequencing technology in which scalable, low-cost semiconductor manufacturing techniques are used to make an integrated circuit able to directly perform non-optical DNA sequencing of genomes. Sequence data are obtained by directly sensing the ions produced by template-directed DNA polymerase synthesis using all-natural nucleotides on this massively parallel semiconductor-sensing device or ion chip. The ion chip contains ion-sensitive, field-effect transistor-based sensors in perfect register with 1.2 million wells, which provide confinement and allow parallel, simultaneous detection of independent sequencing reactions. Use of the most widely used technology for constructing integrated circuits, the complementary metal-oxide semiconductor (CMOS) process, allows for low-cost, large-scale production and scaling of the device to higher densities and larger array sizes. We show the performance of the system by sequencing three bacterial genomes, its robustness and scalability by producing ion chips with up to 10 times as many sensors and sequencing a human genome.

  5. Successful grafting in poplar species (Populus spp.) breeding

    Science.gov (United States)

    A. Assibi Mahama; Brian Sparks; Ronald S., Zalesny; Richard B. Hall

    2006-01-01

    Poor rooting of Populus deltoides Bartr. ex Marsh hardwood cuttings often has contributed to delays in breeding progress as a result of failures of scion wood before and/or after pollination. Seventeen clones were used, and the study was conducted in the greenhouse to test an "intervenous feeding" (IV) method, along with three different...

  6. An Integrative Bioinformatics Framework for Genome-scale Multiple Level Network Reconstruction of Rice

    Directory of Open Access Journals (Sweden)

    Liu Lili

    2013-06-01

    Full Text Available Understanding how metabolic reactions translate the genome of an organism into its phenotype is a grand challenge in biology. Genome-wide association studies (GWAS statistically connect genotypes to phenotypes, without any recourse to known molecular interactions, whereas a molecular mechanistic description ties gene function to phenotype through gene regulatory networks (GRNs, protein-protein interactions (PPIs and molecular pathways. Integration of different regulatory information levels of an organism is expected to provide a good way for mapping genotypes to phenotypes. However, the lack of curated metabolic model of rice is blocking the exploration of genome-scale multi-level network reconstruction. Here, we have merged GRNs, PPIs and genome-scale metabolic networks (GSMNs approaches into a single framework for rice via omics’ regulatory information reconstruction and integration. Firstly, we reconstructed a genome-scale metabolic model, containing 4,462 function genes, 2,986 metabolites involved in 3,316 reactions, and compartmentalized into ten subcellular locations. Furthermore, 90,358 pairs of protein-protein interactions, 662,936 pairs of gene regulations and 1,763 microRNA-target interactions were integrated into the metabolic model. Eventually, a database was developped for systematically storing and retrieving the genome-scale multi-level network of rice. This provides a reference for understanding genotype-phenotype relationship of rice, and for analysis of its molecular regulatory network.

  7. Multilocus analysis of nucleotide variation and speciation in three closely related Populus (Salicaceae) species.

    Science.gov (United States)

    Du, Shuhui; Wang, Zhaoshan; Ingvarsson, Pär K; Wang, Dongsheng; Wang, Junhui; Wu, Zhiqiang; Tembrock, Luke R; Zhang, Jianguo

    2015-10-01

    Historical tectonism and climate oscillations can isolate and contract the geographical distributions of many plant species, and they are even known to trigger species divergence and ultimately speciation. Here, we estimated the nucleotide variation and speciation in three closely related Populus species, Populus tremuloides, P. tremula and P. davidiana, distributed in North America and Eurasia. We analysed the sequence variation in six single-copy nuclear loci and three chloroplast (cpDNA) fragments in 497 individuals sampled from 33 populations of these three species across their geographic distributions. These three Populus species harboured relatively high levels of nucleotide diversity and showed high levels of nucleotide differentiation. Phylogenetic analysis revealed that P. tremuloides diverged earlier than the other two species. The cpDNA haplotype network result clearly illustrated the dispersal route from North America to eastern Asia and then into Europe. Molecular dating results confirmed that the divergence of these three species coincided with the sundering of the Bering land bridge in the late Miocene and a rapid uplift of the Qinghai-Tibetan Plateau around the Miocene/Pliocene boundary. Vicariance-driven successful allopatric speciation resulting from historical tectonism and climate oscillations most likely played roles in the formation of the disjunct distributions and divergence of these three Populus species. © 2015 John Wiley & Sons Ltd.

  8. PHYTOREMEDIATION OF CHLORPYRIFOS BY POPULUS AND SALIX

    OpenAIRE

    Young Lee, Keum; Strand, Stuart E.; Doty, Sharon L.

    2012-01-01

    Chlorpyrifos is one of the commonly used organophosphorus insecticides that are implicated in serious environmental and human health problems. To evaluate plant potential for uptake of chlorpyrifos, several plant species of poplar (Populus sp.) and willow (Salix sp.) were investigated. Chlorpyrifos was taken up from nutrient solution by all seven plant species. Significant amounts of chlorpyrifos accumulated in plant tissues, and roots accumulated higher concentrations of chlorpyrifos than di...

  9. Evaluation of Populus and Salix continuously irrigated with landfill leachate I. Genotype-specific elemental phytoremediation.

    Science.gov (United States)

    Zalesny, Ronald S; Bauer, Edmund O

    2007-01-01

    There is a need for the identification and selection of specific tree genotypes that can sequester elements from contaminated soils, with elevated rates of uptake. We irrigated Populus (DN17, DN182, DN34, NM2, NM6) and Salix (94003, 94012, S287, S566, SX61) genotypes planted in large soil-filled containers with landfill leachate or municipal water and tested for differences in inorganic element concentrations (P, K, Ca, Mg, S, Zn, B, Mn, Fe, Cu, Al, Na, and Cl) in the leaves, stems, and roots. Trees were irrigated with leachate or water during the final 12 wk of the 18-wk study. Genotype-specific uptake existed. For genera, tissue concentrations exhibited four responses. First, Populus had the greatest uptake of P, K, S, Cu, and Cl. Second, Salix exhibited the greatest uptake of Zn, B, Fe, and Al. Third, Salix had greater concentrations of Ca and Mg in leaves, while Populus had greater concentrations in stems and roots. Fourth, Populus had greater concentrations of Mn and Na in leaves and stems, while Salix had greater concentrations in roots. Populus deltoides x P. nigra clones exhibited better overall phytoremediation than the P. nigra x P. maximowiczii genotypes tested. Phytoremediation for S. purpurea clones 94003 and 94012 was generally less than for other Salix genotypes. Overall, concentrations of elements in the leaves, stems, and roots corroborated those in the plant-sciences literature. Uptake was dependent upon the specific genotype for most elements. Our results corroborated the need for further testing and selecting of specific clones for various phytoremediation needs, while providing a baseline for future researchers developing additional studies and resource managers conducting on-site remediation.

  10. Efficient genome-wide genotyping strategies and data integration in crop plants.

    Science.gov (United States)

    Torkamaneh, Davoud; Boyle, Brian; Belzile, François

    2018-03-01

    Next-generation sequencing (NGS) has revolutionized plant and animal research by providing powerful genotyping methods. This review describes and discusses the advantages, challenges and, most importantly, solutions to facilitate data processing, the handling of missing data, and cross-platform data integration. Next-generation sequencing technologies provide powerful and flexible genotyping methods to plant breeders and researchers. These methods offer a wide range of applications from genome-wide analysis to routine screening with a high level of accuracy and reproducibility. Furthermore, they provide a straightforward workflow to identify, validate, and screen genetic variants in a short time with a low cost. NGS-based genotyping methods include whole-genome re-sequencing, SNP arrays, and reduced representation sequencing, which are widely applied in crops. The main challenges facing breeders and geneticists today is how to choose an appropriate genotyping method and how to integrate genotyping data sets obtained from various sources. Here, we review and discuss the advantages and challenges of several NGS methods for genome-wide genetic marker development and genotyping in crop plants. We also discuss how imputation methods can be used to both fill in missing data in genotypic data sets and to integrate data sets obtained using different genotyping tools. It is our hope that this synthetic view of genotyping methods will help geneticists and breeders to integrate these NGS-based methods in crop plant breeding and research.

  11. Trinucleotide repeat microsatellite markers for Black Poplar (Populus nigra L.)

    NARCIS (Netherlands)

    Smulders, M.J.M.; Schoot, van der J.; Arens, P.; Vosman, B.

    2001-01-01

    Using an enrichment procedure, we have cloned microsatellite repeats from black poplar (Populus nigra L.) and developed primers for microsatellite marker analysis. Ten primer pairs, mostly for trinucleotide repeats, produced polymorphic fragments in P. nigra. Some of them also showed amplification

  12. Dicty_cDB: Contig-U13006-1 [Dicty_cDB

    Lifescience Database Archive (English)

    Full Text Available A-FP_182294 Lysiphlebus testaceipes adult whol... 34 3.0 2 ( BU896130 ) X036A07 Populus wood cDNA library Populus tremul...1.r GM_WBc Glycine max genomic clone ... 34 2.9 2 ( DN492755 ) X036A07.3pR Populus wood cDNA library Populus tre...... 34 2.9 2 ( DN502768 ) X036A07.5pR Populus wood cDNA library Populus tre... 34 2.9 2 ( EH014215 ) USD..... 34 2.4 2 ( EU795178 ) Uncultured bacterium HF0010_09O16 genomic sequence. 34 2.4 2 ( AC182691 ) Populus tric...itermes flavipes symbiont library ... 36 3.2 2 ( DQ927304 ) Tetrahymena paravorax strain RP m

  13. Variant Review with the Integrative Genomics Viewer.

    Science.gov (United States)

    Robinson, James T; Thorvaldsdóttir, Helga; Wenger, Aaron M; Zehir, Ahmet; Mesirov, Jill P

    2017-11-01

    Manual review of aligned reads for confirmation and interpretation of variant calls is an important step in many variant calling pipelines for next-generation sequencing (NGS) data. Visual inspection can greatly increase the confidence in calls, reduce the risk of false positives, and help characterize complex events. The Integrative Genomics Viewer (IGV) was one of the first tools to provide NGS data visualization, and it currently provides a rich set of tools for inspection, validation, and interpretation of NGS datasets, as well as other types of genomic data. Here, we present a short overview of IGV's variant review features for both single-nucleotide variants and structural variants, with examples from both cancer and germline datasets. IGV is freely available at https://www.igv.org Cancer Res; 77(21); e31-34. ©2017 AACR . ©2017 American Association for Cancer Research.

  14. Unexpected inheritance: multiple integrations of ancient bornavirus and ebolavirus/marburgvirus sequences in vertebrate genomes.

    Science.gov (United States)

    Belyi, Vladimir A; Levine, Arnold J; Skalka, Anna Marie

    2010-07-29

    Vertebrate genomes contain numerous copies of retroviral sequences, acquired over the course of evolution. Until recently they were thought to be the only type of RNA viruses to be so represented, because integration of a DNA copy of their genome is required for their replication. In this study, an extensive sequence comparison was conducted in which 5,666 viral genes from all known non-retroviral families with single-stranded RNA genomes were matched against the germline genomes of 48 vertebrate species, to determine if such viruses could also contribute to the vertebrate genetic heritage. In 19 of the tested vertebrate species, we discovered as many as 80 high-confidence examples of genomic DNA sequences that appear to be derived, as long ago as 40 million years, from ancestral members of 4 currently circulating virus families with single strand RNA genomes. Surprisingly, almost all of the sequences are related to only two families in the Order Mononegavirales: the Bornaviruses and the Filoviruses, which cause lethal neurological disease and hemorrhagic fevers, respectively. Based on signature landmarks some, and perhaps all, of the endogenous virus-like DNA sequences appear to be LINE element-facilitated integrations derived from viral mRNAs. The integrations represent genes that encode viral nucleocapsid, RNA-dependent-RNA-polymerase, matrix and, possibly, glycoproteins. Integrations are generally limited to one or very few copies of a related viral gene per species, suggesting that once the initial germline integration was obtained (or selected), later integrations failed or provided little advantage to the host. The conservation of relatively long open reading frames for several of the endogenous sequences, the virus-like protein regions represented, and a potential correlation between their presence and a species' resistance to the diseases caused by these pathogens, are consistent with the notion that their products provide some important biological

  15. Unexpected inheritance: multiple integrations of ancient bornavirus and ebolavirus/marburgvirus sequences in vertebrate genomes.

    Directory of Open Access Journals (Sweden)

    Vladimir A Belyi

    2010-07-01

    Full Text Available Vertebrate genomes contain numerous copies of retroviral sequences, acquired over the course of evolution. Until recently they were thought to be the only type of RNA viruses to be so represented, because integration of a DNA copy of their genome is required for their replication. In this study, an extensive sequence comparison was conducted in which 5,666 viral genes from all known non-retroviral families with single-stranded RNA genomes were matched against the germline genomes of 48 vertebrate species, to determine if such viruses could also contribute to the vertebrate genetic heritage. In 19 of the tested vertebrate species, we discovered as many as 80 high-confidence examples of genomic DNA sequences that appear to be derived, as long ago as 40 million years, from ancestral members of 4 currently circulating virus families with single strand RNA genomes. Surprisingly, almost all of the sequences are related to only two families in the Order Mononegavirales: the Bornaviruses and the Filoviruses, which cause lethal neurological disease and hemorrhagic fevers, respectively. Based on signature landmarks some, and perhaps all, of the endogenous virus-like DNA sequences appear to be LINE element-facilitated integrations derived from viral mRNAs. The integrations represent genes that encode viral nucleocapsid, RNA-dependent-RNA-polymerase, matrix and, possibly, glycoproteins. Integrations are generally limited to one or very few copies of a related viral gene per species, suggesting that once the initial germline integration was obtained (or selected, later integrations failed or provided little advantage to the host. The conservation of relatively long open reading frames for several of the endogenous sequences, the virus-like protein regions represented, and a potential correlation between their presence and a species' resistance to the diseases caused by these pathogens, are consistent with the notion that their products provide some important

  16. Integration of genomic information with biological networks using Cytoscape.

    Science.gov (United States)

    Bauer-Mehren, Anna

    2013-01-01

    Cytoscape is an open-source software for visualizing, analyzing, and modeling biological networks. This chapter explains how to use Cytoscape to analyze the functional effect of sequence variations in the context of biological networks such as protein-protein interaction networks and signaling pathways. The chapter is divided into five parts: (1) obtaining information about the functional effect of sequence variation in a Cytoscape readable format, (2) loading and displaying different types of biological networks in Cytoscape, (3) integrating the genomic information (SNPs and mutations) with the biological networks, and (4) analyzing the effect of the genomic perturbation onto the network structure using Cytoscape built-in functions. Finally, we briefly outline how the integrated data can help in building mathematical network models for analyzing the effect of the sequence variation onto the dynamics of the biological system. Each part is illustrated by step-by-step instructions on an example use case and visualized by many screenshots and figures.

  17. Variation in Linked Selection and Recombination Drive Genomic Divergence during Allopatric Speciation of European and American Aspens.

    Science.gov (United States)

    Wang, Jing; Street, Nathaniel R; Scofield, Douglas G; Ingvarsson, Pär K

    2016-07-01

    Despite the global economic and ecological importance of forest trees, the genomic basis of differential adaptation and speciation in tree species is still poorly understood. Populus tremula and Populus tremuloides are two of the most widespread tree species in the Northern Hemisphere. Using whole-genome re-sequencing data of 24 P. tremula and 22 P. tremuloides individuals, we find that the two species diverged ∼2.2-3.1 million years ago, coinciding with the severing of the Bering land bridge and the onset of dramatic climatic oscillations during the Pleistocene. Both species have experienced substantial population expansions following long-term declines after species divergence. We detect widespread and heterogeneous genomic differentiation between species, and in accordance with the expectation of allopatric speciation, coalescent simulations suggest that neutral evolutionary processes can account for most of the observed patterns of genetic differentiation. However, there is an excess of regions exhibiting extreme differentiation relative to those expected under demographic simulations, which is indicative of the action of natural selection. Overall genetic differentiation is negatively associated with recombination rate in both species, providing strong support for a role of linked selection in generating the heterogeneous genomic landscape of differentiation between species. Finally, we identify a number of candidate regions and genes that may have been subject to positive and/or balancing selection during the speciation process. © The Author 2016. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

  18. Integration sites of Epstein-Barr virus genome on chromosomes of human lymphoblastoid cell lines

    Energy Technology Data Exchange (ETDEWEB)

    Wuu, K.D.; Chen, Y.J.; Wang-Wuu, S. [Institute of Genetics, Taipei (Taiwan, Province of China)

    1994-09-01

    Epstein-Barr virus (EBV) is the pathogen of infectious mononucleosis. The viral genome is present in more than 95% of the African cases of Burkitt lymphoma and it is usually maintained in episomal form in the tumor cells. Viral integration has been described only for Nanalwa which is a Burkitt lymphoma cell line lacking episomes. In order to examine the role of EBV in the immortalization of human Blymphocytes, we investigated whether the EBV integration into the human genome is essential. If the integration does occur, we would like to know whether the integration is randomly distributed or whether the viral DNA integrates preferentially at certain sites. Fourteen in vitro immortalized human lymphoblastoid cell lines (LCLs) were examined by fluorescence in situ hybridization (FISH) with a biotinylated EBV BamHI w DNA fragment as probe. The episomal form of EBV DNA was found in all cells of these cell lines, while only about 65% of the cells have the integrated viral DNA. This might suggest that integration is not a pre-requisite for cell immortalization. Although all chromosomes, except Y, have been found with integrated viral genome, chromsomes 1 and 5 are the most frequent EBV DNA carrier (p<0.05). Nine chromosome bands, namely, 1p31, 1q31, 2q32, 3q13, 3q26, 5q14, 6q24, 7q31 and 12q21, are preferential targets for EBV integration (p<0.001). Eighty percent of the total 938 EBV hybridization signals were found to be at G-band-positive area. This suggests that the mechanism of EBV integration might be different from that of the retroviruses, which specifically integrate to G-band-negative areas. Thus, we conclude that the integration of EBV to host genome is non-random and it may have something to do with the structure of chromosome and DNA sequences.

  19. Preparation and Characterization of Lignocellulosic Oil Sorbent by Hydrothermal Treatment of Populus Fiber

    Directory of Open Access Journals (Sweden)

    Yue Zhang

    2014-09-01

    Full Text Available This study is aimed at achieving the optimum conditions of hydrothermal treatment and acetylation of Populus fiber to improve its oil sorption capacity (OSC in an oil-water mixture. The characteristics of the hydrolyzed and acetylated fibers were comparatively investigated by FT-IR, CP-MAS 13C-NMR, SEM and TGA. The optimum conditions of the hydrothermal treatment and acetylation were obtained at170 °C for 1 h and 120 °C for 2 h, respectively. The maximum OSC of the hydrolyzed fiber (16.78 g/g was slightly lower than that of the acetylated fiber (21.57 g/g, but they were both higher than the maximum OSC of the unmodified fiber (3.94 g/g. In addition, acetylation after hydrothermal treatment for the Populus fiber was unnecessary as the increment of the maximum OSC was only 3.53 g/g. The hydrolyzed and the acetylated Populus fibers both displayed a lumen orifice enabling a high oil entrapment. The thermal stability of the modified fibers was shown to be increased in comparison with that of the raw fiber. The hydrothermal treatment offers a new approach to prepare lignocellulosic oil sorbent.

  20. INDIGO – INtegrated Data Warehouse of MIcrobial GenOmes with Examples from the Red Sea Extremophiles

    Science.gov (United States)

    Alam, Intikhab; Antunes, André; Kamau, Allan Anthony; Ba alawi, Wail; Kalkatawi, Manal; Stingl, Ulrich; Bajic, Vladimir B.

    2013-01-01

    Background The next generation sequencing technologies substantially increased the throughput of microbial genome sequencing. To functionally annotate newly sequenced microbial genomes, a variety of experimental and computational methods are used. Integration of information from different sources is a powerful approach to enhance such annotation. Functional analysis of microbial genomes, necessary for downstream experiments, crucially depends on this annotation but it is hampered by the current lack of suitable information integration and exploration systems for microbial genomes. Results We developed a data warehouse system (INDIGO) that enables the integration of annotations for exploration and analysis of newly sequenced microbial genomes. INDIGO offers an opportunity to construct complex queries and combine annotations from multiple sources starting from genomic sequence to protein domain, gene ontology and pathway levels. This data warehouse is aimed at being populated with information from genomes of pure cultures and uncultured single cells of Red Sea bacteria and Archaea. Currently, INDIGO contains information from Salinisphaera shabanensis, Haloplasma contractile, and Halorhabdus tiamatea - extremophiles isolated from deep-sea anoxic brine lakes of the Red Sea. We provide examples of utilizing the system to gain new insights into specific aspects on the unique lifestyle and adaptations of these organisms to extreme environments. Conclusions We developed a data warehouse system, INDIGO, which enables comprehensive integration of information from various resources to be used for annotation, exploration and analysis of microbial genomes. It will be regularly updated and extended with new genomes. It is aimed to serve as a resource dedicated to the Red Sea microbes. In addition, through INDIGO, we provide our Automatic Annotation of Microbial Genomes (AAMG) pipeline. The INDIGO web server is freely available at http://www.cbrc.kaust.edu.sa/indigo. PMID

  1. Exploiting the transcriptome of Euphrates Poplar, Populus euphratica (Salicaceae to develop and characterize new EST-SSR markers and construct an EST-SSR database.

    Directory of Open Access Journals (Sweden)

    Fang K Du

    Full Text Available BACKGROUND: Microsatellite markers or Simple Sequence Repeats (SSRs are the most popular markers in population/conservation genetics. However, the development of novel microsatellite markers has been impeded by high costs, a lack of available sequence data and technical difficulties. New species-specific microsatellite markers were required to investigate the evolutionary history of the Euphratica tree, Populus euphratica, the only tree species found in the desert regions of Western China and adjacent Central Asian countries. METHODOLOGY/PRINCIPAL FINDINGS: A total of 94,090 non-redundant Expressed Sequence Tags (ESTs from P. euphratica comprising around 63 Mb of sequence data were searched for SSRs. 4,202 SSRs were found in 3,839 ESTs, with 311 ESTs containing multiple SSRs. The most common motif types were trinucleotides (37% and hexanucleotides (33% repeats. We developed primer pairs for all of the identified EST-SSRs (eSSRs and selected 673 of these pairs at random for further validation. 575 pairs (85% gave successful amplification, of which, 464 (80.7% were polymorphic in six to 24 individuals from natural populations across Northern China. We also tested the transferability of the polymorphic eSSRs to nine other Populus species. In addition, to facilitate the use of these new eSSR markers by other researchers, we mapped them onto Populus trichocarpa scaffolds in silico and compiled our data into a web-based database (http://202.205.131.253:8080/poplar/resources/static_page/index.html. CONCLUSIONS: The large set of validated eSSRs identified in this work will have many potential applications in studies on P. euphratica and other poplar species, in fields such as population genetics, comparative genomics, linkage mapping, QTL, and marker-assisted breeding. Their use will be facilitated by their incorporation into a user-friendly web-based database.

  2. Genome-wide conserved non-coding microsatellite (CNMS) marker-based integrative genetical genomics for quantitative dissection of seed weight in chickpea.

    Science.gov (United States)

    Bajaj, Deepak; Saxena, Maneesha S; Kujur, Alice; Das, Shouvik; Badoni, Saurabh; Tripathi, Shailesh; Upadhyaya, Hari D; Gowda, C L L; Sharma, Shivali; Singh, Sube; Tyagi, Akhilesh K; Parida, Swarup K

    2015-03-01

    Phylogenetic footprinting identified 666 genome-wide paralogous and orthologous CNMS (conserved non-coding microsatellite) markers from 5'-untranslated and regulatory regions (URRs) of 603 protein-coding chickpea genes. The (CT)n and (GA)n CNMS carrying CTRMCAMV35S and GAGA8BKN3 regulatory elements, respectively, are abundant in the chickpea genome. The mapped genic CNMS markers with robust amplification efficiencies (94.7%) detected higher intraspecific polymorphic potential (37.6%) among genotypes, implying their immense utility in chickpea breeding and genetic analyses. Seventeen differentially expressed CNMS marker-associated genes showing strong preferential and seed tissue/developmental stage-specific expression in contrasting genotypes were selected to narrow down the gene targets underlying seed weight quantitative trait loci (QTLs)/eQTLs (expression QTLs) through integrative genetical genomics. The integration of transcript profiling with seed weight QTL/eQTL mapping, molecular haplotyping, and association analyses identified potential molecular tags (GAGA8BKN3 and RAV1AAT regulatory elements and alleles/haplotypes) in the LOB-domain-containing protein- and KANADI protein-encoding transcription factor genes controlling the cis-regulated expression for seed weight in the chickpea. This emphasizes the potential of CNMS marker-based integrative genetical genomics for the quantitative genetic dissection of complex seed weight in chickpea. © The Author 2014. Published by Oxford University Press on behalf of the Society for Experimental Biology.

  3. Genome-wide analysis reveals the extent of EAV-HP integration in domestic chicken.

    Science.gov (United States)

    Wragg, David; Mason, Andrew S; Yu, Le; Kuo, Richard; Lawal, Raman A; Desta, Takele Taye; Mwacharo, Joram M; Cho, Chang-Yeon; Kemp, Steve; Burt, David W; Hanotte, Olivier

    2015-10-14

    EAV-HP is an ancient retrovirus pre-dating Gallus speciation, which continues to circulate in modern chicken populations, and led to the emergence of avian leukosis virus subgroup J causing significant economic losses to the poultry industry. We mapped EAV-HP integration sites in Ethiopian village chickens, a Silkie, Taiwan Country chicken, red junglefowl Gallus gallus and several inbred experimental lines using whole-genome sequence data. An average of 75.22 ± 9.52 integration sites per bird were identified, which collectively group into 279 intervals of which 5 % are common to 90 % of the genomes analysed and are suggestive of pre-domestication integration events. More than a third of intervals are specific to individual genomes, supporting active circulation of EAV-HP in modern chickens. Interval density is correlated with chromosome length (P < 2.31(-6)), and 27 % of intervals are located within 5 kb of a transcript. Functional annotation clustering of genes reveals enrichment for immune-related functions (P < 0.05). Our results illustrate a non-random distribution of EAV-HP in the genome, emphasising the importance it may have played in the adaptation of the species, and provide a platform from which to extend investigations on the co-evolutionary significance of endogenous retroviral genera with their hosts.

  4. LocusTrack: Integrated visualization of GWAS results and genomic annotation.

    Science.gov (United States)

    Cuellar-Partida, Gabriel; Renteria, Miguel E; MacGregor, Stuart

    2015-01-01

    Genome-wide association studies (GWAS) are an important tool for the mapping of complex traits and diseases. Visual inspection of genomic annotations may be used to generate insights into the biological mechanisms underlying GWAS-identified loci. We developed LocusTrack, a web-based application that annotates and creates plots of regional GWAS results and incorporates user-specified tracks that display annotations such as linkage disequilibrium (LD), phylogenetic conservation, chromatin state, and other genomic and regulatory elements. Currently, LocusTrack can integrate annotation tracks from the UCSC genome-browser as well as from any tracks provided by the user. LocusTrack is an easy-to-use application and can be accessed at the following URL: http://gump.qimr.edu.au/general/gabrieC/LocusTrack/. Users can upload and manage GWAS results and select from and/or provide annotation tracks using simple and intuitive menus. LocusTrack scripts and associated data can be downloaded from the website and run locally.

  5. Phylogeny reconstruction and hybrid analysis of populus (Salicaceae) based on nucleotide sequences of multiple single-copy nuclear genes and plastid fragments.

    Science.gov (United States)

    Wang, Zhaoshan; Du, Shuhui; Dayanandan, Selvadurai; Wang, Dongsheng; Zeng, Yanfei; Zhang, Jianguo

    2014-01-01

    Populus (Salicaceae) is one of the most economically and ecologically important genera of forest trees. The complex reticulate evolution and lack of highly variable orthologous single-copy DNA markers have posed difficulties in resolving the phylogeny of this genus. Based on a large data set of nuclear and plastid DNA sequences, we reconstructed robust phylogeny of Populus using parsimony, maximum likelihood and Bayesian inference methods. The resulting phylogenetic trees showed better resolution at both inter- and intra-sectional level than previous studies. The results revealed that (1) the plastid-based phylogenetic tree resulted in two main clades, suggesting an early divergence of the maternal progenitors of Populus; (2) three advanced sections (Populus, Aigeiros and Tacamahaca) are of hybrid origin; (3) species of the section Tacamahaca could be divided into two major groups based on plastid and nuclear DNA data, suggesting a polyphyletic nature of the section; and (4) many species proved to be of hybrid origin based on the incongruence between plastid and nuclear DNA trees. Reticulate evolution may have played a significant role in the evolution history of Populus by facilitating rapid adaptive radiations into different environments.

  6. Phylogeny reconstruction and hybrid analysis of populus (Salicaceae based on nucleotide sequences of multiple single-copy nuclear genes and plastid fragments.

    Directory of Open Access Journals (Sweden)

    Zhaoshan Wang

    Full Text Available Populus (Salicaceae is one of the most economically and ecologically important genera of forest trees. The complex reticulate evolution and lack of highly variable orthologous single-copy DNA markers have posed difficulties in resolving the phylogeny of this genus. Based on a large data set of nuclear and plastid DNA sequences, we reconstructed robust phylogeny of Populus using parsimony, maximum likelihood and Bayesian inference methods. The resulting phylogenetic trees showed better resolution at both inter- and intra-sectional level than previous studies. The results revealed that (1 the plastid-based phylogenetic tree resulted in two main clades, suggesting an early divergence of the maternal progenitors of Populus; (2 three advanced sections (Populus, Aigeiros and Tacamahaca are of hybrid origin; (3 species of the section Tacamahaca could be divided into two major groups based on plastid and nuclear DNA data, suggesting a polyphyletic nature of the section; and (4 many species proved to be of hybrid origin based on the incongruence between plastid and nuclear DNA trees. Reticulate evolution may have played a significant role in the evolution history of Populus by facilitating rapid adaptive radiations into different environments.

  7. Discovery and annotation of small proteins using genomics, proteomics and computational approaches

    Energy Technology Data Exchange (ETDEWEB)

    Yang, Xiaohan; Tschaplinski, Timothy J.; Hurst, Gregory B.; Jawdy, Sara; Abraham, Paul E.; Lankford, Patricia K.; Adams, Rachel M.; Shah, Manesh B.; Hettich, Robert L.; Lindquist, Erika; Kalluri, Udaya C.; Gunter, Lee E.; Pennacchio, Christa; Tuskan, Gerald A.

    2011-03-02

    Small proteins (10 200 amino acids aa in length) encoded by short open reading frames (sORF) play important regulatory roles in various biological processes, including tumor progression, stress response, flowering, and hormone signaling. However, ab initio discovery of small proteins has been relatively overlooked. Recent advances in deep transcriptome sequencing make it possible to efficiently identify sORFs at the genome level. In this study, we obtained 2.6 million expressed sequence tag (EST) reads from Populus deltoides leaf transcriptome and reconstructed full-length transcripts from the EST sequences. We identified an initial set of 12,852 sORFs encoding proteins of 10 200 aa in length. Three computational approaches were then used to enrich for bona fide protein-coding sORFs from the initial sORF set: (1) codingpotential prediction, (2) evolutionary conservation between P. deltoides and other plant species, and (3) gene family clustering within P. deltoides. As a result, a high-confidence sORF candidate set containing 1469 genes was obtained. Analysis of the protein domains, non-protein-coding RNA motifs, sequence length distribution, and protein mass spectrometry data supported this high-confidence sORF set. In the high-confidence sORF candidate set, known protein domains were identified in 1282 genes (higher-confidence sORF candidate set), out of which 611 genes, designated as highest-confidence candidate sORF set, were supported by proteomics data. Of the 611 highest-confidence candidate sORF genes, 56 were new to the current Populus genome annotation. This study not only demonstrates that there are potential sORF candidates to be annotated in sequenced genomes, but also presents an efficient strategy for discovery of sORFs in species with no genome annotation yet available.

  8. Genetic Modification of Lignin in Hybrid Poplar (Populus alba × Populus tremula) Does Not Substantially Alter Plant Defense or Arthropod Communities.

    Science.gov (United States)

    Buhl, Christine; Meilan, Richard; Lindroth, Richard L

    2017-05-01

    Lignin impedes access to cellulose during biofuel production and pulping but trees can be genetically modified to improve processing efficiency. Modification of lignin may have nontarget effects on mechanical and chemical resistance and subsequent arthropod community responses with respect to pest susceptibility and arthropod biodiversity. We quantified foliar mechanical and chemical resistance traits in lignin-modified and wild-type (WT) poplar (Populus alba × Populus tremula) grown in a plantation and censused arthropods present on these trees to determine total abundance, as well as species richness, diversity and community composition. Our results indicate that mechanical resistance was not affected by lignin modification and only one genetic construct resulted in a (modest) change in chemical resistance. Arthropod abundance and community composition were consistent across modified and WT trees, but transgenics produced using one construct exhibited higher species richness and diversity relative to the WT. Our findings indicate that modification of lignin in poplar does not negatively affect herbivore resistance traits or arthropod community response, and may even result in a source of increased genetic diversity in trees and arthropod communities. © The Authors 2017. Published by Oxford University Press on behalf of Entomological Society of America.

  9. MicroScope in 2017: an expanding and evolving integrated resource for community expertise of microbial genomes.

    Science.gov (United States)

    Vallenet, David; Calteau, Alexandra; Cruveiller, Stéphane; Gachet, Mathieu; Lajus, Aurélie; Josso, Adrien; Mercier, Jonathan; Renaux, Alexandre; Rollin, Johan; Rouy, Zoe; Roche, David; Scarpelli, Claude; Médigue, Claudine

    2017-01-04

    The annotation of genomes from NGS platforms needs to be automated and fully integrated. However, maintaining consistency and accuracy in genome annotation is a challenging problem because millions of protein database entries are not assigned reliable functions. This shortcoming limits the knowledge that can be extracted from genomes and metabolic models. Launched in 2005, the MicroScope platform (http://www.genoscope.cns.fr/agc/microscope) is an integrative resource that supports systematic and efficient revision of microbial genome annotation, data management and comparative analysis. Effective comparative analysis requires a consistent and complete view of biological data, and therefore, support for reviewing the quality of functional annotation is critical. MicroScope allows users to analyze microbial (meta)genomes together with post-genomic experiment results if any (i.e. transcriptomics, re-sequencing of evolved strains, mutant collections, phenotype data). It combines tools and graphical interfaces to analyze genomes and to perform the expert curation of gene functions in a comparative context. Starting with a short overview of the MicroScope system, this paper focuses on some major improvements of the Web interface, mainly for the submission of genomic data and on original tools and pipelines that have been developed and integrated in the platform: computation of pan-genomes and prediction of biosynthetic gene clusters. Today the resource contains data for more than 6000 microbial genomes, and among the 2700 personal accounts (65% of which are now from foreign countries), 14% of the users are performing expert annotations, on at least a weekly basis, contributing to improve the quality of microbial genome annotations. © The Author(s) 2016. Published by Oxford University Press on behalf of Nucleic Acids Research.

  10. Constitutively elevated salicylic acid levels alter photosynthesis and oxidative state but not growth in transgenic populus.

    Science.gov (United States)

    Xue, Liang-Jiao; Guo, Wenbing; Yuan, Yinan; Anino, Edward O; Nyamdari, Batbayar; Wilson, Mark C; Frost, Christopher J; Chen, Han-Yi; Babst, Benjamin A; Harding, Scott A; Tsai, Chung-Jui

    2013-07-01

    Salicylic acid (SA) has long been implicated in plant responses to oxidative stress. SA overproduction in Arabidopsis thaliana leads to dwarfism, making in planta assessment of SA effects difficult in this model system. We report that transgenic Populus tremula × alba expressing a bacterial SA synthase hyperaccumulated SA and SA conjugates without negative growth consequences. In the absence of stress, endogenously elevated SA elicited widespread metabolic and transcriptional changes that resembled those of wild-type plants exposed to oxidative stress-promoting heat treatments. Potential signaling and oxidative stress markers azelaic and gluconic acids as well as antioxidant chlorogenic acids were strongly coregulated with SA, while soluble sugars and other phenylpropanoids were inversely correlated. Photosynthetic responses to heat were attenuated in SA-overproducing plants. Network analysis identified potential drivers of SA-mediated transcriptome rewiring, including receptor-like kinases and WRKY transcription factors. Orthologs of Arabidopsis SA signaling components NON-EXPRESSOR OF PATHOGENESIS-RELATED GENES1 and thioredoxins were not represented. However, all members of the expanded Populus nucleoredoxin-1 family exhibited increased expression and increased network connectivity in SA-overproducing Populus, suggesting a previously undescribed role in SA-mediated redox regulation. The SA response in Populus involved a reprogramming of carbon uptake and partitioning during stress that is compatible with constitutive chemical defense and sustained growth, contrasting with the SA response in Arabidopsis, which is transient and compromises growth if sustained.

  11. The complete chloroplast genome sequence of Euonymus japonicus (Celastraceae).

    Science.gov (United States)

    Choi, Kyoung Su; Park, SeonJoo

    2016-09-01

    The complete chloroplast (cp) genome sequence of the Euonymus japonicus, the first sequenced of the genus Euonymus, was reported in this study. The total length was 157 637 bp, containing a pair of 26 678 bp inverted repeat region (IR), which were separated by small single copy (SSC) region and large single copy (LSC) region of 18 340 bp and 85 941 bp, respectively. This genome contains 107 unique genes, including 74 coding genes, four rRNA genes, and 29 tRNA genes. Seventeen genes contain intron of E. japonicus, of which three genes (clpP, ycf3, and rps12) include two introns. The maximum likelihood (ML) phylogenetic analysis revealed that E. japonicus was closely related to Manihot and Populus.

  12. Selecting and utilizing Populus and Salix for landfill covers: implications for leachate irrigation.

    Science.gov (United States)

    Zalesny, Ronald S; Bauer, Edmund O

    2007-01-01

    The success of using Populus and Salix for phytoremediation has prompted further use of leachate as a combination of irrigation and fertilization for the trees. A common protocol for such efforts has been to utilize a limited number of readily-available genotypes with decades of deployment in other applications, such as fiber or windbreaks. However, it may be possible to increase phytoremediation success with proper genotypic screening and selection, followed by the field establishment of clones that exhibited favorable potential for cleanup of specific contaminants. There is an overwhelming need for testing and subsequent deployment of diverse Populus and Salix genotypes, given current availability of clonal material and the inherent genetic variation among and within these genera. Therefore, we detail phyto-recurrent selection, a method that consists of revising and combining crop and tree improvement protocols to meet the objective of utilizing superior Populus and Salix clones for remediation applications. Although such information is lacking for environmental clean-up technologies, centuries of plant selection success in agronomy, horticulture, and forestry validate the need for similar approaches in phytoremediation. We bridge the gap between these disciplines by describing project development, clone selection, tree establishment, and evaluation of success metrics in the context of their importance to utilizing trees for phytoremediation.

  13. Human papillomavirus genome integration in squamous carcinogenesis: what have next-generation sequencing studies taught us?

    Science.gov (United States)

    Groves, Ian J; Coleman, Nicholas

    2018-05-01

    Human papillomavirus (HPV) infection is associated with ∼5% of all human cancers, including a range of squamous cell carcinomas. Persistent infection by high-risk HPVs (HRHPVs) is associated with the integration of virus genomes (which are usually stably maintained as extrachromosomal episomes) into host chromosomes. Although HRHPV integration rates differ across human sites of infection, this process appears to be an important event in HPV-associated neoplastic progression, leading to deregulation of virus oncogene expression, host gene expression modulation, and further genomic instability. However, the mechanisms by which HRHPV integration occur and by which the subsequent gene expression changes take place are incompletely understood. The advent of next-generation sequencing (NGS) of both RNA and DNA has allowed powerful interrogation of the association of HRHPVs with human disease, including precise determination of the sites of integration and the genomic rearrangements at integration loci. In turn, these data have indicated that integration occurs through two main mechanisms: looping integration and direct insertion. Improved understanding of integration sites is allowing further investigation of the factors that provide a competitive advantage to some integrants during disease progression. Furthermore, advanced approaches to the generation of genome-wide samples have given novel insights into the three-dimensional interactions within the nucleus, which could act as another layer of epigenetic control of both virus and host transcription. It is hoped that further advances in NGS techniques and analysis will not only allow the examination of further unanswered questions regarding HPV infection, but also direct new approaches to treating HPV-associated human disease. Copyright © 2018 Pathological Society of Great Britain and Ireland. Published by John Wiley & Sons, Ltd. Copyright © 2018 Pathological Society of Great Britain and Ireland. Published by John

  14. Partitioning of Multivariate Phenotypes using Regression Trees Reveals Complex Patterns of Adaptation to Climate across the Range of Black Cottonwood (Populus trichocarpa

    Directory of Open Access Journals (Sweden)

    Regis Wendpouire Oubida

    2015-03-01

    Full Text Available Local adaptation to climate in temperate forest trees involves the integration of multiple physiological, morphological, and phenological traits. Latitudinal clines are frequently observed for these traits, but environmental constraints also track longitude and altitude. We combined extensive phenotyping of 12 candidate adaptive traits, multivariate regression trees, quantitative genetics, and a genome-wide panel of SNP markers to better understand the interplay among geography, climate, and adaptation to abiotic factors in Populus trichocarpa. Heritabilities were low to moderate (0.13 to 0.32 and population differentiation for many traits exceeded the 99th percentile of the genome-wide distribution of FST, suggesting local adaptation. When climate variables were taken as predictors and the 12 traits as response variables in a multivariate regression tree analysis, evapotranspiration (Eref explained the most variation, with subsequent splits related to mean temperature of the warmest month, frost-free period (FFP, and mean annual precipitation (MAP. These grouping matched relatively well the splits using geographic variables as predictors: the northernmost groups (short FFP and low Eref had the lowest growth, and lowest cold injury index; the southern British Columbia group (low Eref and intermediate temperatures had average growth and cold injury index; the group from the coast of California and Oregon (high Eref and FFP had the highest growth performance and the highest cold injury index; and the southernmost, high-altitude group (with high Eref and low FFP performed poorly, had high cold injury index, and lower water use efficiency. Taken together, these results suggest variation in both temperature and water availability across the range shape multivariate adaptive traits in poplar.

  15. Climate, migration, and the local food security context: Introducing Terra Populus

    Science.gov (United States)

    Schlak, Allison M.; Kugler, Tracy A.

    2016-01-01

    Studies investigating the connection between environmental factors and migration are difficult to execute because they require the integration of microdata and spatial information. In this article, we introduce the novel, publically available data extraction system Terra Populus (TerraPop), which was designed to facilitate population-environment studies. We showcase the use of TerraPop by exploring variations in the climate-migration association in Burkina Faso and Senegal based on differences in the local food security context. Food security was approximated using anthropometric indicators of child stunting and wasting derived from Demographic and Health Surveys (DHS) and linked to the TerraPop extract of climate and migration information. We find that an increase in heat waves was associated with a decrease in international migration from Burkina Faso, while excessive precipitation increased international moves from Senegal. Significant interactions reveal that the adverse effects of heat waves and droughts are strongly amplified in highly food insecure Senegalese departments. PMID:27974863

  16. Cadmium accumulation and growth responses of a poplar (Populus deltoids x Populus nigra) in cadmium contaminated purple soil and alluvial soil

    Energy Technology Data Exchange (ETDEWEB)

    Wu Fuzhong [Faculty of Forestry, Sichuan Agricultural University, 625014, Ya' an (China); Yang Wanqin, E-mail: scyangwq@163.com [Faculty of Forestry, Sichuan Agricultural University, 625014, Ya' an (China); Zhang Jian; Zhou Liqiang [Faculty of Forestry, Sichuan Agricultural University, 625014, Ya' an (China)

    2010-05-15

    To characterize the phytoextraction efficiency of a hybrid poplar (Populus deltoids x Populus nigra) in cadmium contaminated purple soil and alluvial soil, a pot experiment in field was carried out in Sichuan basin, western China. After one growing period, the poplar accumulated the highest of 541.98 {+-} 19.22 and 576.75 {+-} 40.55 {mu}g cadmium per plant with 110.77 {+-} 12.68 and 202.54 {+-} 19.12 g dry mass in these contaminated purple soil and alluvial soil, respectively. Higher phytoextraction efficiency with higher cadmium concentration in tissues was observed in poplar growing in purple soil than that in alluvial soil at relative lower soil cadmium concentration. The poplar growing in alluvial soil had relative higher tolerance ability with lower reduction rates of morphological and growth characters than that in purple soil, suggesting that the poplar growing in alluvial soil might display the higher phytoextraction ability when cadmium contamination level increased. Even so, the poplars exhibited obvious cadmium transport from root to shoot in both soils regardless of cadmium contamination levels. It implies that this examined poplar can extract more cadmium than some hyperaccumulators. The results indicated that metal phytoextraction using the poplar can be applied to clean up soils moderately contaminated by cadmium in these purple soil and alluvial soil.

  17. An integrative and applicable phylogenetic footprinting framework for cis-regulatory motifs identification in prokaryotic genomes.

    Science.gov (United States)

    Liu, Bingqiang; Zhang, Hanyuan; Zhou, Chuan; Li, Guojun; Fennell, Anne; Wang, Guanghui; Kang, Yu; Liu, Qi; Ma, Qin

    2016-08-09

    Phylogenetic footprinting is an important computational technique for identifying cis-regulatory motifs in orthologous regulatory regions from multiple genomes, as motifs tend to evolve slower than their surrounding non-functional sequences. Its application, however, has several difficulties for optimizing the selection of orthologous data and reducing the false positives in motif prediction. Here we present an integrative phylogenetic footprinting framework for accurate motif predictions in prokaryotic genomes (MP(3)). The framework includes a new orthologous data preparation procedure, an additional promoter scoring and pruning method and an integration of six existing motif finding algorithms as basic motif search engines. Specifically, we collected orthologous genes from available prokaryotic genomes and built the orthologous regulatory regions based on sequence similarity of promoter regions. This procedure made full use of the large-scale genomic data and taxonomy information and filtered out the promoters with limited contribution to produce a high quality orthologous promoter set. The promoter scoring and pruning is implemented through motif voting by a set of complementary predicting tools that mine as many motif candidates as possible and simultaneously eliminate the effect of random noise. We have applied the framework to Escherichia coli k12 genome and evaluated the prediction performance through comparison with seven existing programs. This evaluation was systematically carried out at the nucleotide and binding site level, and the results showed that MP(3) consistently outperformed other popular motif finding tools. We have integrated MP(3) into our motif identification and analysis server DMINDA, allowing users to efficiently identify and analyze motifs in 2,072 completely sequenced prokaryotic genomes. The performance evaluation indicated that MP(3) is effective for predicting regulatory motifs in prokaryotic genomes. Its application may enhance

  18. Diversity of cuticular wax among Salix species and Populus species hybrids.

    Science.gov (United States)

    Cameron, Kimberly D; Teece, Mark A; Bevilacqua, Eddie; Smart, Lawrence B

    2002-08-01

    The leaf cuticular waxes of three Salix species and two Populus species hybrids, selected for their ability to produce high amounts of biomass, were characterized. Samples were extracted in CH(2)Cl(2) three times over the growing season. Low kV SEM was utilized to observe differences in the ultrastructure of leaf surfaces from each clone. Homologous series of wax components were classified into organic groups, and the variation in wax components due to clone, sample time, and their interaction was identified. All Salix species and Populus species hybrids showed differences in total wax load at each sampling period, whereas the pattern of wax deposition over time differed only between the Salix species. A strong positive relationship was identified between the entire homologous series of alcohols and total wax load in all clones. Similarly strong relationships were observed between fatty acids and total wax load as well as fatty acids and alcohols in two Salix species and one Populus species hybrid. One Salix species, S. dasyclados, also displayed a strong positive relationship between alcohols and alkanes. These data indicate that species grown under the same environmental conditions produce measurably different cuticular waxes and that regulation of wax production appears to be different in each species. The important roles cuticular waxes play in drought tolerance, pest, and pathogen resistance, as well as the ease of wax extraction and analysis, strongly suggest that the characteristics of the cuticular wax may prove to be useful selectable traits in a breeding program.

  19. Genome scale models of yeast: towards standardized evaluation and consistent omic integration

    DEFF Research Database (Denmark)

    Sanchez, Benjamin J.; Nielsen, Jens

    2015-01-01

    Genome scale models (GEMs) have enabled remarkable advances in systems biology, acting as functional databases of metabolism, and as scaffolds for the contextualization of high-throughput data. In the case of Saccharomyces cerevisiae (budding yeast), several GEMs have been published and are curre......Genome scale models (GEMs) have enabled remarkable advances in systems biology, acting as functional databases of metabolism, and as scaffolds for the contextualization of high-throughput data. In the case of Saccharomyces cerevisiae (budding yeast), several GEMs have been published...... in which all levels of omics data (from gene expression to flux) have been integrated in yeast GEMs. Relevant conclusions and current challenges for both GEM evaluation and omic integration are highlighted....

  20. Water use sources of desert riparian Populus euphratica forests.

    Science.gov (United States)

    Si, Jianhua; Feng, Qi; Cao, Shengkui; Yu, Tengfei; Zhao, Chunyan

    2014-09-01

    Desert riparian forests are the main body of natural oases in the lower reaches of inland rivers; its growth and distribution are closely related to water use sources. However, how does the desert riparian forest obtains a stable water source and which water sources it uses to effectively avoid or overcome water stress to survive? This paper describes an analysis of the water sources, using the stable oxygen isotope technique and the linear mixed model of the isotopic values and of desert riparian Populus euphratica forests growing at sites with different groundwater depths and conditions. The results showed that the main water source of Populus euphratica changes from water in a single soil layer or groundwater to deep subsoil water and groundwater as the depth of groundwater increases. This appears to be an adaptive selection to arid and water-deficient conditions and is a primary reason for the long-term survival of P. euphratica in the desert riparian forest of an extremely arid region. Water contributions from the various soil layers and from groundwater differed and the desert riparian P. euphratica forests in different habitats had dissimilar water use strategies.

  1. Decoding the genome with an integrative analysis tool: combinatorial CRM Decoder.

    Science.gov (United States)

    Kang, Keunsoo; Kim, Joomyeong; Chung, Jae Hoon; Lee, Daeyoup

    2011-09-01

    The identification of genome-wide cis-regulatory modules (CRMs) and characterization of their associated epigenetic features are fundamental steps toward the understanding of gene regulatory networks. Although integrative analysis of available genome-wide information can provide new biological insights, the lack of novel methodologies has become a major bottleneck. Here, we present a comprehensive analysis tool called combinatorial CRM decoder (CCD), which utilizes the publicly available information to identify and characterize genome-wide CRMs in a species of interest. CCD first defines a set of the epigenetic features which is significantly associated with a set of known CRMs as a code called 'trace code', and subsequently uses the trace code to pinpoint putative CRMs throughout the genome. Using 61 genome-wide data sets obtained from 17 independent mouse studies, CCD successfully catalogued ∼12 600 CRMs (five distinct classes) including polycomb repressive complex 2 target sites as well as imprinting control regions. Interestingly, we discovered that ∼4% of the identified CRMs belong to at least two different classes named 'multi-functional CRM', suggesting their functional importance for regulating spatiotemporal gene expression. From these examples, we show that CCD can be applied to any potential genome-wide datasets and therefore will shed light on unveiling genome-wide CRMs in various species.

  2. An integrated clinical and genomic information system for cancer precision medicine.

    Science.gov (United States)

    Jang, Yeongjun; Choi, Taekjin; Kim, Jongho; Park, Jisub; Seo, Jihae; Kim, Sangok; Kwon, Yeajee; Lee, Seungjae; Lee, Sanghyuk

    2018-04-20

    Increasing affordability of next-generation sequencing (NGS) has created an opportunity for realizing genomically-informed personalized cancer therapy as a path to precision oncology. However, the complex nature of genomic information presents a huge challenge for clinicians in interpreting the patient's genomic alterations and selecting the optimum approved or investigational therapy. An elaborate and practical information system is urgently needed to support clinical decision as well as to test clinical hypotheses quickly. Here, we present an integrated clinical and genomic information system (CGIS) based on NGS data analyses. Major components include modules for handling clinical data, NGS data processing, variant annotation and prioritization, drug-target-pathway analysis, and population cohort explorer. We built a comprehensive knowledgebase of genes, variants, drugs by collecting annotated information from public and in-house resources. Structured reports for molecular pathology are generated using standardized terminology in order to help clinicians interpret genomic variants and utilize them for targeted cancer therapy. We also implemented many features useful for testing hypotheses to develop prognostic markers from mutation and gene expression data. Our CGIS software is an attempt to provide useful information for both clinicians and scientists who want to explore genomic information for precision oncology.

  3. [A review of the genomic and gene cloning studies in trees].

    Science.gov (United States)

    Yin, Tong-Ming

    2010-07-01

    Supported by the Department of Energy (DOE) of U.S., the first tree genome, black cottonwood (Populus trichocarpa), has been completely sequenced and publicly release. This is the milestone that indicates the beginning of post-genome era for forest trees. Identification and cloning genes underlying important traits are one of the main tasks for the post-genome-era tree genomic studies. Recently, great achievements have been made in cloning genes coordinating important domestication traits in some crops, such as rice, tomato, maize and so on. Molecular breeding has been applied in the practical breeding programs for many crops. By contrast, molecular studies in trees are lagging behind. Trees possess some characteristics that make them as difficult organisms for studying on locating and cloning of genes. With the advances in techniques, given also the fast growth of tree genomic resources, great achievements are desirable in cloning unknown genes from trees, which will facilitate tree improvement programs by means of molecular breeding. In this paper, the author reviewed the progress in tree genomic and gene cloning studies, and prospected the future achievements in order to provide a useful reference for researchers working in this area.

  4. Uptake of macro- and micro-nutrients into leaf, woody, and root tissue of Populus after irrigation with landfill leachate

    Science.gov (United States)

    Jill A. Zalesny; Ronald S., Jr. Zalesny; Adam H. Wiese; Bart T. Sexton; Richard B. Hall

    2008-01-01

    Information about macro- and micro-nutrient uptake and distribution into tissues of Populus irrigated with landfill leachate helps to maximize biomass production and understand impacts of leachate chemistry on tree health. We irrigated eight Populus clones (NC 13460, NCI4O18, NC14104, NC14106, DM115, DN5, NM2, NM6) with fertilized (N, P, K) well...

  5. GDR (Genome Database for Rosaceae): integrated web-database for Rosaceae genomics and genetics data.

    Science.gov (United States)

    Jung, Sook; Staton, Margaret; Lee, Taein; Blenda, Anna; Svancara, Randall; Abbott, Albert; Main, Dorrie

    2008-01-01

    The Genome Database for Rosaceae (GDR) is a central repository of curated and integrated genetics and genomics data of Rosaceae, an economically important family which includes apple, cherry, peach, pear, raspberry, rose and strawberry. GDR contains annotated databases of all publicly available Rosaceae ESTs, the genetically anchored peach physical map, Rosaceae genetic maps and comprehensively annotated markers and traits. The ESTs are assembled to produce unigene sets of each genus and the entire Rosaceae. Other annotations include putative function, microsatellites, open reading frames, single nucleotide polymorphisms, gene ontology terms and anchored map position where applicable. Most of the published Rosaceae genetic maps can be viewed and compared through CMap, the comparative map viewer. The peach physical map can be viewed using WebFPC/WebChrom, and also through our integrated GDR map viewer, which serves as a portal to the combined genetic, transcriptome and physical mapping information. ESTs, BACs, markers and traits can be queried by various categories and the search result sites are linked to the mapping visualization tools. GDR also provides online analysis tools such as a batch BLAST/FASTA server for the GDR datasets, a sequence assembly server and microsatellite and primer detection tools. GDR is available at http://www.rosaceae.org.

  6. CTDB: An Integrated Chickpea Transcriptome Database for Functional and Applied Genomics.

    Directory of Open Access Journals (Sweden)

    Mohit Verma

    Full Text Available Chickpea is an important grain legume used as a rich source of protein in human diet. The narrow genetic diversity and limited availability of genomic resources are the major constraints in implementing breeding strategies and biotechnological interventions for genetic enhancement of chickpea. We developed an integrated Chickpea Transcriptome Database (CTDB, which provides the comprehensive web interface for visualization and easy retrieval of transcriptome data in chickpea. The database features many tools for similarity search, functional annotation (putative function, PFAM domain and gene ontology search and comparative gene expression analysis. The current release of CTDB (v2.0 hosts transcriptome datasets with high quality functional annotation from cultivated (desi and kabuli types and wild chickpea. A catalog of transcription factor families and their expression profiles in chickpea are available in the database. The gene expression data have been integrated to study the expression profiles of chickpea transcripts in major tissues/organs and various stages of flower development. The utilities, such as similarity search, ortholog identification and comparative gene expression have also been implemented in the database to facilitate comparative genomic studies among different legumes and Arabidopsis. Furthermore, the CTDB represents a resource for the discovery of functional molecular markers (microsatellites and single nucleotide polymorphisms between different chickpea types. We anticipate that integrated information content of this database will accelerate the functional and applied genomic research for improvement of chickpea. The CTDB web service is freely available at http://nipgr.res.in/ctdb.html.

  7. STINGRAY: system for integrated genomic resources and analysis.

    Science.gov (United States)

    Wagner, Glauber; Jardim, Rodrigo; Tschoeke, Diogo A; Loureiro, Daniel R; Ocaña, Kary A C S; Ribeiro, Antonio C B; Emmel, Vanessa E; Probst, Christian M; Pitaluga, André N; Grisard, Edmundo C; Cavalcanti, Maria C; Campos, Maria L M; Mattoso, Marta; Dávila, Alberto M R

    2014-03-07

    The STINGRAY system has been conceived to ease the tasks of integrating, analyzing, annotating and presenting genomic and expression data from Sanger and Next Generation Sequencing (NGS) platforms. STINGRAY includes: (a) a complete and integrated workflow (more than 20 bioinformatics tools) ranging from functional annotation to phylogeny; (b) a MySQL database schema, suitable for data integration and user access control; and (c) a user-friendly graphical web-based interface that makes the system intuitive, facilitating the tasks of data analysis and annotation. STINGRAY showed to be an easy to use and complete system for analyzing sequencing data. While both Sanger and NGS platforms are supported, the system could be faster using Sanger data, since the large NGS datasets could potentially slow down the MySQL database usage. STINGRAY is available at http://stingray.biowebdb.org and the open source code at http://sourceforge.net/projects/stingray-biowebdb/.

  8. MicroScope-an integrated resource for community expertise of gene functions and comparative analysis of microbial genomic and metabolic data.

    Science.gov (United States)

    Médigue, Claudine; Calteau, Alexandra; Cruveiller, Stéphane; Gachet, Mathieu; Gautreau, Guillaume; Josso, Adrien; Lajus, Aurélie; Langlois, Jordan; Pereira, Hugo; Planel, Rémi; Roche, David; Rollin, Johan; Rouy, Zoe; Vallenet, David

    2017-09-12

    The overwhelming list of new bacterial genomes becoming available on a daily basis makes accurate genome annotation an essential step that ultimately determines the relevance of thousands of genomes stored in public databanks. The MicroScope platform (http://www.genoscope.cns.fr/agc/microscope) is an integrative resource that supports systematic and efficient revision of microbial genome annotation, data management and comparative analysis. Starting from the results of our syntactic, functional and relational annotation pipelines, MicroScope provides an integrated environment for the expert annotation and comparative analysis of prokaryotic genomes. It combines tools and graphical interfaces to analyze genomes and to perform the manual curation of gene function in a comparative genomics and metabolic context. In this article, we describe the free-of-charge MicroScope services for the annotation and analysis of microbial (meta)genomes, transcriptomic and re-sequencing data. Then, the functionalities of the platform are presented in a way providing practical guidance and help to the nonspecialists in bioinformatics. Newly integrated analysis tools (i.e. prediction of virulence and resistance genes in bacterial genomes) and original method recently developed (the pan-genome graph representation) are also described. Integrated environments such as MicroScope clearly contribute, through the user community, to help maintaining accurate resources. © The Author 2017. Published by Oxford University Press.

  9. Integrated genomic and gene expression profiling identifies two major genomic circuits in urothelial carcinoma.

    Directory of Open Access Journals (Sweden)

    David Lindgren

    Full Text Available Similar to other malignancies, urothelial carcinoma (UC is characterized by specific recurrent chromosomal aberrations and gene mutations. However, the interconnection between specific genomic alterations, and how patterns of chromosomal alterations adhere to different molecular subgroups of UC, is less clear. We applied tiling resolution array CGH to 146 cases of UC and identified a number of regions harboring recurrent focal genomic amplifications and deletions. Several potential oncogenes were included in the amplified regions, including known oncogenes like E2F3, CCND1, and CCNE1, as well as new candidate genes, such as SETDB1 (1q21, and BCL2L1 (20q11. We next combined genome profiling with global gene expression, gene mutation, and protein expression data and identified two major genomic circuits operating in urothelial carcinoma. The first circuit was characterized by FGFR3 alterations, overexpression of CCND1, and 9q and CDKN2A deletions. The second circuit was defined by E3F3 amplifications and RB1 deletions, as well as gains of 5p, deletions at PTEN and 2q36, 16q, 20q, and elevated CDKN2A levels. TP53/MDM2 alterations were common for advanced tumors within the two circuits. Our data also suggest a possible RAS/RAF circuit. The tumors with worst prognosis showed a gene expression profile that indicated a keratinized phenotype. Taken together, our integrative approach revealed at least two separate networks of genomic alterations linked to the molecular diversity seen in UC, and that these circuits may reflect distinct pathways of tumor development.

  10. Genetic basis of aboveground productivity in two native Populus species and their hybrids.

    Science.gov (United States)

    Lojewski, Nathan R; Fischer, Dylan G; Bailey, Joseph K; Schweitzer, Jennifer A; Whitham, Thomas G; Hart, Stephen C

    2009-09-01

    Demonstration of genetic control over riparian tree productivity has major implications for responses of riparian systems to shifting environmental conditions and effects of genetics on ecosystems in general. We used field studies and common gardens, applying both molecular and quantitative techniques, to compare plot-level tree aboveground net primary productivity (ANPP(tree)) and individual tree growth rate constants in relation to plant genetic identity in two naturally occurring Populus tree species and their hybrids. In field comparisons of four cross types (Populus fremontii S. Wats., Populus angustifolia James, F(1) hybrids and backcross hybrids) across 11 natural stands, productivity was greatest for P. fremontii trees, followed by hybrids and lowest in P. angustifolia. A similar pattern was observed in four common gardens across a 290 m elevation and 100 km environmental gradient. Despite a doubling in productivity across the common gardens, the relative differences among the cross types remained constant. Using clonal replicates in a common garden, we found ANPP(tree) to be a heritable plant trait (i.e., broad-sense heritability), such that plant genetic factors explained between 38% and 82% of the variation in ANPP(tree). Furthermore, analysis of the genetic composition among individual tree genotypes using restriction fragment length polymorphism molecular markers showed that genetically similar trees also exhibited similar ANPP(tree). These findings indicate strong genetic contributions to natural variation in ANPP with important ecological implications.

  11. Transcriptional and Hormonal Regulation of Gravitropism of Woody Stems in Populus

    Science.gov (United States)

    Suzanne Gerttula; Matthew S. Zinkgraf; Gloria K. Muday; Daniel R. Lewis; Farid M. Ibatullin; Harry Brumer; Foster Hart; Shawn D. Mansfield; Vladimir Filkov; Andrew Groover

    2015-01-01

    Angiosperm trees reorient their woody stems by asymmetrically producing a specialized xylem tissue, tension wood, which exerts a strong contractile force resulting in negative gravitropism of the stem. Here, we show, in Populus trees, that initial gravity perception and response occurs in specialized cells through sedimentation of starch-filled...

  12. NGS-based approach to determine the presence of HPV and their sites of integration in human cancer genome.

    Science.gov (United States)

    Chandrani, P; Kulkarni, V; Iyer, P; Upadhyay, P; Chaubal, R; Das, P; Mulherkar, R; Singh, R; Dutt, A

    2015-06-09

    Human papilloma virus (HPV) accounts for the most common cause of all virus-associated human cancers. Here, we describe the first graphic user interface (GUI)-based automated tool 'HPVDetector', for non-computational biologists, exclusively for detection and annotation of the HPV genome based on next-generation sequencing data sets. We developed a custom-made reference genome that comprises of human chromosomes along with annotated genome of 143 HPV types as pseudochromosomes. The tool runs on a dual mode as defined by the user: a 'quick mode' to identify presence of HPV types and an 'integration mode' to determine genomic location for the site of integration. The input data can be a paired-end whole-exome, whole-genome or whole-transcriptome data set. The HPVDetector is available in public domain for download: http://www.actrec.gov.in/pi-webpages/AmitDutt/HPVdetector/HPVDetector.html. On the basis of our evaluation of 116 whole-exome, 23 whole-transcriptome and 2 whole-genome data, we were able to identify presence of HPV in 20 exomes and 4 transcriptomes of cervical and head and neck cancer tumour samples. Using the inbuilt annotation module of HPVDetector, we found predominant integration of viral gene E7, a known oncogene, at known 17q21, 3q27, 7q35, Xq28 and novel sites of integration in the human genome. Furthermore, co-infection with high-risk HPVs such as 16 and 31 were found to be mutually exclusive compared with low-risk HPV71. HPVDetector is a simple yet precise and robust tool for detecting HPV from tumour samples using variety of next-generation sequencing platforms including whole genome, whole exome and transcriptome. Two different modes (quick detection and integration mode) along with a GUI widen the usability of HPVDetector for biologists and clinicians with minimal computational knowledge.

  13. Prolonged Integration Site Selection of a Lentiviral Vector in the Genome of Human Keratinocytes.

    Science.gov (United States)

    Qian, Wei; Wang, Yong; Li, Rui-Fu; Zhou, Xin; Liu, Jing; Peng, Dai-Zhi

    2017-03-03

    BACKGROUND Lentiviral vectors have been successfully used for human skin cell gene transfer studies. Defining the selection of integration sites for retroviral vectors in the host genome is crucial in risk assessment analysis of gene therapy. However, genome-wide analyses of lentiviral integration sites in human keratinocytes, especially after prolonged growth, are poorly understood. MATERIAL AND METHODS In this study, 874 unique lentiviral vector integration sites in human HaCaT keratinocytes after long-term culture were identified and analyzed with the online tool GTSG-QuickMap and SPSS software. RESULTS The data indicated that lentiviral vectors showed integration site preferences for genes and gene-rich regions. CONCLUSIONS This study will likely assist in determining the relative risks of the lentiviral vector system and in the design of a safe lentiviral vector system in the gene therapy of skin diseases.

  14. Cadmium accumulation and growth responses of a poplar (Populus deltoids x Populus nigra) in cadmium contaminated purple soil and alluvial soil

    International Nuclear Information System (INIS)

    Wu Fuzhong; Yang Wanqin; Zhang Jian; Zhou Liqiang

    2010-01-01

    To characterize the phytoextraction efficiency of a hybrid poplar (Populus deltoids x Populus nigra) in cadmium contaminated purple soil and alluvial soil, a pot experiment in field was carried out in Sichuan basin, western China. After one growing period, the poplar accumulated the highest of 541.98 ± 19.22 and 576.75 ± 40.55 μg cadmium per plant with 110.77 ± 12.68 and 202.54 ± 19.12 g dry mass in these contaminated purple soil and alluvial soil, respectively. Higher phytoextraction efficiency with higher cadmium concentration in tissues was observed in poplar growing in purple soil than that in alluvial soil at relative lower soil cadmium concentration. The poplar growing in alluvial soil had relative higher tolerance ability with lower reduction rates of morphological and growth characters than that in purple soil, suggesting that the poplar growing in alluvial soil might display the higher phytoextraction ability when cadmium contamination level increased. Even so, the poplars exhibited obvious cadmium transport from root to shoot in both soils regardless of cadmium contamination levels. It implies that this examined poplar can extract more cadmium than some hyperaccumulators. The results indicated that metal phytoextraction using the poplar can be applied to clean up soils moderately contaminated by cadmium in these purple soil and alluvial soil.

  15. Molecular Assemblies, Genes and Genomics Integrated Efficiently (MAGGIE)

    Energy Technology Data Exchange (ETDEWEB)

    Baliga, Nitin S

    2011-05-26

    when applied to the manually curated training set. Applying this method to the data representing around a quarter of the fraction space for water soluble proteins in D. vulgaris, we obtained 854 reliable pair wise interactions. Further, we have developed algorithms to analyze and assign significance to protein interaction data from bait pull-down experiments and integrate these data with other systems biology data through associative biclustering in a parallel computing environment. We will 'fill-in' missing information in these interaction data using a 'Transitive Closure' algorithm and subsequently use 'Between Commonality Decomposition' algorithm to discover complexes within these large graphs of protein interactions. To characterize the metabolic activities of proteins and their complexes we are developing algorithms to deconvolute pure mass spectra, estimate chemical formula for m/z values, and fit isotopic fine structure to metabolomics data. We have discovered that in comparison to isotopic pattern fitting methods restricting the chemical formula by these two dimensions actually facilitates unique solutions for chemical formula generators. To understand how microbial functions are regulated we have developed complementary algorithms for reconstructing gene regulatory networks (GRNs). Whereas the network inference algorithms cMonkey and Inferelator developed enable de novo reconstruction of predictive models for GRNs from diverse systems biology data, the RegPrecise and RegPredict framework developed uses evolutionary comparisons of genomes from closely related organisms to reconstruct conserved regulons. We have integrated the two complementary algorithms to rapidly generate comprehensive models for gene regulation of understudied organisms. Our preliminary analyses of these reconstructed GRNs have revealed novel regulatory mechanisms and cis-regulatory motifs, as well asothers that are conserved across species. Finally, we are

  16. Construction of an integrated database to support genomic sequence analysis

    Energy Technology Data Exchange (ETDEWEB)

    Gilbert, W.; Overbeek, R.

    1994-11-01

    The central goal of this project is to develop an integrated database to support comparative analysis of genomes including DNA sequence data, protein sequence data, gene expression data and metabolism data. In developing the logic-based system GenoBase, a broader integration of available data was achieved due to assistance from collaborators. Current goals are to easily include new forms of data as they become available and to easily navigate through the ensemble of objects described within the database. This report comments on progress made in these areas.

  17. The RNAPII-CTD Maintains Genome Integrity through Inhibition of Retrotransposon Gene Expression and Transposition.

    Directory of Open Access Journals (Sweden)

    Maria J Aristizabal

    2015-10-01

    Full Text Available RNA polymerase II (RNAPII contains a unique C-terminal domain that is composed of heptapeptide repeats and which plays important regulatory roles during gene expression. RNAPII is responsible for the transcription of most protein-coding genes, a subset of non-coding genes, and retrotransposons. Retrotransposon transcription is the first step in their multiplication cycle, given that the RNA intermediate is required for the synthesis of cDNA, the material that is ultimately incorporated into a new genomic location. Retrotransposition can have grave consequences to genome integrity, as integration events can change the gene expression landscape or lead to alteration or loss of genetic information. Given that RNAPII transcribes retrotransposons, we sought to investigate if the RNAPII-CTD played a role in the regulation of retrotransposon gene expression. Importantly, we found that the RNAPII-CTD functioned to maintaining genome integrity through inhibition of retrotransposon gene expression, as reducing CTD length significantly increased expression and transposition rates of Ty1 elements. Mechanistically, the increased Ty1 mRNA levels in the rpb1-CTD11 mutant were partly due to Cdk8-dependent alterations to the RNAPII-CTD phosphorylation status. In addition, Cdk8 alone contributed to Ty1 gene expression regulation by altering the occupancy of the gene-specific transcription factor Ste12. Loss of STE12 and TEC1 suppressed growth phenotypes of the RNAPII-CTD truncation mutant. Collectively, our results implicate Ste12 and Tec1 as general and important contributors to the Cdk8, RNAPII-CTD regulatory circuitry as it relates to the maintenance of genome integrity.

  18. BiologicalNetworks 2.0 - an integrative view of genome biology data

    Directory of Open Access Journals (Sweden)

    Ponomarenko Julia

    2010-12-01

    Full Text Available Abstract Background A significant problem in the study of mechanisms of an organism's development is the elucidation of interrelated factors which are making an impact on the different levels of the organism, such as genes, biological molecules, cells, and cell systems. Numerous sources of heterogeneous data which exist for these subsystems are still not integrated sufficiently enough to give researchers a straightforward opportunity to analyze them together in the same frame of study. Systematic application of data integration methods is also hampered by a multitude of such factors as the orthogonal nature of the integrated data and naming problems. Results Here we report on a new version of BiologicalNetworks, a research environment for the integral visualization and analysis of heterogeneous biological data. BiologicalNetworks can be queried for properties of thousands of different types of biological entities (genes/proteins, promoters, COGs, pathways, binding sites, and other and their relations (interactions, co-expression, co-citations, and other. The system includes the build-pathways infrastructure for molecular interactions/relations and module discovery in high-throughput experiments. Also implemented in BiologicalNetworks are the Integrated Genome Viewer and Comparative Genomics Browser applications, which allow for the search and analysis of gene regulatory regions and their conservation in multiple species in conjunction with molecular pathways/networks, experimental data and functional annotations. Conclusions The new release of BiologicalNetworks together with its back-end database introduces extensive functionality for a more efficient integrated multi-level analysis of microarray, sequence, regulatory, and other data. BiologicalNetworks is freely available at http://www.biologicalnetworks.org.

  19. Gibberellins in shoots and developing capsules of Populus species.

    Science.gov (United States)

    Pearce, David W; Hutt, Oliver E; Rood, Stewart B; Mander, Lewis N

    2002-03-01

    Extracts of stems of growing shoots of Populus deltoides and P. trichocarpa, and developing capsules of P. deltoides were analysed for gibberellins (GAs) by gas chromatography-mass spectrometry. The following known GAs were identified by comparison of their Kovats retention indices (KRIs) and mass spectra with those of standards: GA1, GA8, GA9, GA19, GA20, 16 beta,17-dihydro-17-hydroxy GA20, GA23, GA28, GA29, GA34, GA44, and GA97. Several of these have not been previously reported from Populus. In addition, two new GAs were identified as 12 beta-hydroxy GA53 (GA127) and 16 beta,17-dihydro-17-hydroxy GA53 and their structures were confirmed by partial synthesis. Evidence was found of 16,17-dihydro-16,17-dihydroxy GA9, 16,17-dihydro-16,17-dihydroxy GA12, 12-hydroxy GA14, and GA34-catabolite by comparison of mass spectra and KRIs with published data. Several putative GAs (hydroxy- and dihydroxy-GA12-like) were also found. The catabolites of active GAs or of key precursors, hydroxylated at C-2 in stems and either C-2, C-12, C-17, or C-16,17 in capsules, were the major proportion of the GAs.

  20. MicrobesOnline: an integrated portal for comparative and functional genomics

    Energy Technology Data Exchange (ETDEWEB)

    Dehal, Paramvir S.; Joachimiak, Marcin P.; Price, Morgan N.; Bates, John T.; Baumohl, Jason K.; Chivian, Dylan; Friedland, Greg D.; Huang, Katherine H.; Keller, Keith; Novichkov, Pavel S.; Dubchak, Inna L.; Alm, Eric J.; Arkin, Adam P.

    2009-09-17

    Since 2003, MicrobesOnline (http://www.microbesonline.org) has been providing a community resource for comparative and functional genome analysis. The portal includes over 1000 complete genomes of bacteria, archaea and fungi and thousands of expression microarrays from diverse organisms ranging from model organisms such as Escherichia coli and Saccharomyces cerevisiae to environmental microbes such as Desulfovibrio vulgaris and Shewanella oneidensis. To assist in annotating genes and in reconstructing their evolutionary history, MicrobesOnline includes a comparative genome browser based on phylogenetic trees for every gene family as well as a species tree. To identify co-regulated genes, MicrobesOnline can search for genes based on their expression profile, and provides tools for identifying regulatory motifs and seeing if they are conserved. MicrobesOnline also includes fast phylogenetic profile searches, comparative views of metabolic pathways, operon predictions, a workbench for sequence analysis and integration with RegTransBase and other microbial genome resources. The next update of MicrobesOnline will contain significant new functionality, including comparative analysis of metagenomic sequence data. Programmatic access to the database, along with source code and documentation, is available at http://microbesonline.org/programmers.html.

  1. MicrobesOnline: an integrated portal for comparative and functional genomics

    Energy Technology Data Exchange (ETDEWEB)

    Dehal, Paramvir; Joachimiak, Marcin; Price, Morgan; Bates, John; Baumohl, Jason; Chivian, Dylan; Friedland, Greg; Huang, Kathleen; Keller, Keith; Novichkov, Pavel; Dubchak, Inna; Alm, Eric; Arkin, Adam

    2011-07-14

    Since 2003, MicrobesOnline (http://www.microbesonline.org) has been providing a community resource for comparative and functional genome analysis. The portal includes over 1000 complete genomes of bacteria, archaea and fungi and thousands of expression microarrays from diverse organisms ranging from model organisms such as Escherichia coli and Saccharomyces cerevisiae to environmental microbes such as Desulfovibrio vulgaris and Shewanella oneidensis. To assist in annotating genes and in reconstructing their evolutionary history, MicrobesOnline includes a comparative genome browser based on phylogenetic trees for every gene family as well as a species tree. To identify co-regulated genes, MicrobesOnline can search for genes based on their expression profile, and provides tools for identifying regulatory motifs and seeing if they are conserved. MicrobesOnline also includes fast phylogenetic profile searches, comparative views of metabolic pathways, operon predictions, a workbench for sequence analysis and integration with RegTransBase and other microbial genome resources. The next update of MicrobesOnline will contain significant new functionality, including comparative analysis of metagenomic sequence data. Programmatic access to the database, along with source code and documentation, is available at http://microbesonline.org/programmers.html.

  2. Integrative Genomic Analysis of Cholangiocarcinoma Identifies Distinct IDH-Mutant Molecular Profiles

    DEFF Research Database (Denmark)

    Farshidfar, Farshad; Zheng, Siyuan; Gingras, Marie-Claude

    2017-01-01

    Cholangiocarcinoma (CCA) is an aggressive malignancy of the bile ducts, with poor prognosis and limited treatment options. Here, we describe the integrated analysis of somatic mutations, RNA expression, copy number, and DNA methylation by The Cancer Genome Atlas of a set of predominantly intrahep...

  3. Litter Quality of Populus Species as Affected by Free-Air CO2

    NARCIS (Netherlands)

    Vermue, E.; Buurman, P.; Hoosbeek, M.R.

    2009-01-01

    The effect of elevated CO2 and nitrogen fertilization on the molecular chemistry of litter of three Populus species and associated soil organic matter (SOM) was investigated by pyrolysis-gas chromatography/mass spectrometry. The results are based on 147 quantified organic compounds in 24 litter

  4. Light regimes in Populus plantations using the Voxel-based Light Interception Model

    NARCIS (Netherlands)

    Van der Zande, D.; Dieussart, K.; Stuckens, J.; Verstraeten, W.W.; Coppin, P.

    2011-01-01

    Three-dimensional light interception by three uniform Populus canopies was studied using the Voxel-based Light Interception Model (VLIM) in combination with ground-based Light Detection and Ranging (LiDAR) measurements. As the VLIM was developed and validated in a virtual environment to ensure

  5. Data integration to prioritize drugs using genomics and curated data.

    Science.gov (United States)

    Louhimo, Riku; Laakso, Marko; Belitskin, Denis; Klefström, Juha; Lehtonen, Rainer; Hautaniemi, Sampsa

    2016-01-01

    Genomic alterations affecting drug target proteins occur in several tumor types and are prime candidates for patient-specific tailored treatments. Increasingly, patients likely to benefit from targeted cancer therapy are selected based on molecular alterations. The selection of a precision therapy benefiting most patients is challenging but can be enhanced with integration of multiple types of molecular data. Data integration approaches for drug prioritization have successfully integrated diverse molecular data but do not take full advantage of existing data and literature. We have built a knowledge-base which connects data from public databases with molecular results from over 2200 tumors, signaling pathways and drug-target databases. Moreover, we have developed a data mining algorithm to effectively utilize this heterogeneous knowledge-base. Our algorithm is designed to facilitate retargeting of existing drugs by stratifying samples and prioritizing drug targets. We analyzed 797 primary tumors from The Cancer Genome Atlas breast and ovarian cancer cohorts using our framework. FGFR, CDK and HER2 inhibitors were prioritized in breast and ovarian data sets. Estrogen receptor positive breast tumors showed potential sensitivity to targeted inhibitors of FGFR due to activation of FGFR3. Our results suggest that computational sample stratification selects potentially sensitive samples for targeted therapies and can aid in precision medicine drug repositioning. Source code is available from http://csblcanges.fimm.fi/GOPredict/.

  6. Molecular genetic analysis of black poplar (Populus nigra L.) along Dutch rivers

    NARCIS (Netherlands)

    Arens, P.; Coops, H.; Jansen, J.; Vosman, B.

    1998-01-01

    The genetic structure of remaining black poplar (Populus nigra) trees on the banks of the Dutch Rhine branches was investigated using the AFLP technique. In total, 143 trees, including one P. deltoides and some P. x euramericana, were analysed using six AFLP primer combinations which generated 319

  7. Widespread triploidy in Western North American aspen (Populus tremuloides.

    Directory of Open Access Journals (Sweden)

    Karen E Mock

    Full Text Available We document high rates of triploidy in aspen (Populus tremuloides across the western USA (up to 69% of genets, and ask whether the incidence of triploidy across the species range corresponds with latitude, glacial history (as has been documented in other species, climate, or regional variance in clone size. Using a combination of microsatellite genotyping, flow cytometry, and cytology, we demonstrate that triploidy is highest in unglaciated, drought-prone regions of North America, where the largest clone sizes have been reported for this species. While we cannot completely rule out a low incidence of undetected aneuploidy, tetraploidy or duplicated loci, our evidence suggests that these phenomena are unlikely to be significant contributors to our observed patterns. We suggest that the distribution of triploid aspen is due to a positive synergy between triploidy and ecological factors driving clonality. Although triploids are expected to have low fertility, they are hypothesized to be an evolutionary link to sexual tetraploidy. Thus, interactions between clonality and polyploidy may be a broadly important component of geographic speciation patterns in perennial plants. Further, cytotypes are expected to show physiological and structural differences which may influence susceptibility to ecological factors such as drought, and we suggest that cytotype may be a significant and previously overlooked factor in recent patterns of high aspen mortality in the southwestern portion of the species range. Finally, triploidy should be carefully considered as a source of variance in genomic and ecological studies of aspen, particularly in western U.S. landscapes.

  8. Unexpected ancestry of Populus seedlings from a hybrid zone implies a large role for postzygotic selection in the maintenance of species.

    Science.gov (United States)

    Lindtke, Dorothea; Gompert, Zachariah; Lexer, Christian; Buerkle, C Alex

    2014-09-01

    In the context of potential interspecific gene flow, the integrity of species will be maintained by reproductive barriers that reduce genetic exchange, including traits associated with prezygotic isolation or poor performance of hybrids. Hybrid zones can be used to study the importance of different reproductive barriers, particularly when both parental species and hybrids occur in close spatial proximity. We investigated the importance of barriers to gene flow that act early vs. late in the life cycle of European Populus by quantifying the prevalence of homospecific and hybrid matings within a mosaic hybrid zone. We obtained genotypic data for 11 976 loci from progeny and their maternal parents and constructed a Bayesian model to estimate individual admixture proportions and hybrid classes for sampled trees and for the unsampled pollen parent. Matings that included one or two hybrid parents were common, resulting in admixture proportions of progeny that spanned the whole range of potential ancestries between the two parental species. This result contrasts strongly with the distribution of admixture proportions in adult trees, where intermediate hybrids and each of the parental species are separated into three discrete ancestry clusters. The existence of the full range of hybrids in seedlings is consistent with weak reproductive isolation early in the life cycle of Populus. Instead, a considerable amount of selection must take place between the seedling stage and maturity to remove many hybrid seedlings. Our results highlight that high hybridization rates and appreciable hybrid fitness do not necessarily conflict with the maintenance of species integrity. © 2014 John Wiley & Sons Ltd.

  9. Sodium and chloride accumulation in leaf, woody, and root tissue of Populus after irrigation with landfill leachate

    International Nuclear Information System (INIS)

    Zalesny, Jill A.; Zalesny, Ronald S.; Wiese, Adam H.; Sexton, Bart; Hall, Richard B.

    2008-01-01

    The response of Populus to irrigation sources containing elevated levels of sodium (Na + ) and chloride (Cl - ) is poorly understood. We irrigated eight Populus clones with fertilized well water (control) (N, P, K) or municipal solid waste landfill leachate weekly during 2005 and 2006 in Rhinelander, Wisconsin, USA (45.6 deg. N, 89.4 deg. W). During August 2006, we tested for differences in total Na + and Cl - concentration in preplanting and harvest soils, and in leaf, woody (stems + branches), and root tissue. The leachate-irrigated soils at harvest had the greatest Na + and Cl - levels. Genotypes exhibited elevated total tree Cl - concentration and increased biomass (clones NC14104, NM2, NM6), elevated Cl - and decreased biomass (NC14018, NC14106, DM115), or mid levels of Cl - and biomass (NC13460, DN5). Leachate tissue concentrations were 17 (Na + ) and four (Cl - ) times greater than water. Sodium and Cl - levels were greatest in roots and leaves, respectively. - Sodium and chloride supplied via landfill leachate irrigation is accumulated at high concentrations in tissues of Populus

  10. A role for stomata in regulating water use efficiency in Populus x ...

    African Journals Online (AJOL)

    A role for stomata in regulating water use efficiency in Populus x euramericana and characterization of a related gene, PdERECTA. P Guo, X Xia, WL Yin. Abstract. The physiological mechanism of water use efficiency (WUE) remains elucidated, especially in poplar. We studied WUEi (instantaneous leaf transpiration ...

  11. [Seasonal variation of Tamarix ramosissima and Populus euphratica water potentials in southern fringe of Taklamakan Desert].

    Science.gov (United States)

    Zeng, Fanjiang; Zhang, Ximing; Li, Xiangyi; Foetzki, Andrea; Runge, Michael

    2005-08-01

    The measurement of the seasonal and diurnal variations of Tamarix ramosissima and Populus euphratica water potentials in the southern fringe of Taklamakan Desert indicated that there was no apparent water stress for the two species during their growth period, with little change of predawn water potential and some extent decrease of midday water potential. Irrigation once or thinning had no significant effects on the water status of the plants, while groundwater appeared to be a prerequisite for the survival and growth of these species. It is very important to ensure a stable groundwater table for the restoration of Tamarix ramosissima and Populus euphratica in this area.

  12. Elevated atmospheric CO2 concentration leads to increased whole-plant isoprene emission in hybrid aspen (Populus tremula × Populus tremuloides).

    Science.gov (United States)

    Sun, Zhihong; Niinemets, Ülo; Hüve, Katja; Rasulov, Bahtijor; Noe, Steffen M

    2013-05-01

    Effects of elevated atmospheric [CO2] on plant isoprene emissions are controversial. Relying on leaf-scale measurements, most models simulating isoprene emissions in future higher [CO2] atmospheres suggest reduced emission fluxes. However, combined effects of elevated [CO2] on leaf area growth, net assimilation and isoprene emission rates have rarely been studied on the canopy scale, but stimulation of leaf area growth may largely compensate for possible [CO2] inhibition reported at the leaf scale. This study tests the hypothesis that stimulated leaf area growth leads to increased canopy isoprene emission rates. We studied the dynamics of canopy growth, and net assimilation and isoprene emission rates in hybrid aspen (Populus tremula × Populus tremuloides) grown under 380 and 780 μmol mol(-1) [CO2]. A theoretical framework based on the Chapman-Richards function to model canopy growth and numerically compare the growth dynamics among ambient and elevated atmospheric [CO2]-grown plants was developed. Plants grown under elevated [CO2] had higher C : N ratio, and greater total leaf area, and canopy net assimilation and isoprene emission rates. During ontogeny, these key canopy characteristics developed faster and stabilized earlier under elevated [CO2]. However, on a leaf area basis, foliage physiological traits remained in a transient state over the whole experiment. These results demonstrate that canopy-scale dynamics importantly complements the leaf-scale processes, and that isoprene emissions may actually increase under higher [CO2] as a result of enhanced leaf area production. © 2013 The Authors. New Phytologist © 2013 New Phytologist Trust.

  13. Visualization of RNA structure models within the Integrative Genomics Viewer.

    Science.gov (United States)

    Busan, Steven; Weeks, Kevin M

    2017-07-01

    Analyses of the interrelationships between RNA structure and function are increasingly important components of genomic studies. The SHAPE-MaP strategy enables accurate RNA structure probing and realistic structure modeling of kilobase-length noncoding RNAs and mRNAs. Existing tools for visualizing RNA structure models are not suitable for efficient analysis of long, structurally heterogeneous RNAs. In addition, structure models are often advantageously interpreted in the context of other experimental data and gene annotation information, for which few tools currently exist. We have developed a module within the widely used and well supported open-source Integrative Genomics Viewer (IGV) that allows visualization of SHAPE and other chemical probing data, including raw reactivities, data-driven structural entropies, and data-constrained base-pair secondary structure models, in context with linear genomic data tracks. We illustrate the usefulness of visualizing RNA structure in the IGV by exploring structure models for a large viral RNA genome, comparing bacterial mRNA structure in cells with its structure under cell- and protein-free conditions, and comparing a noncoding RNA structure modeled using SHAPE data with a base-pairing model inferred through sequence covariation analysis. © 2017 Busan and Weeks; Published by Cold Spring Harbor Laboratory Press for the RNA Society.

  14. Genome Enabled Discovery of Carbon Sequestration Genes in Poplar

    Energy Technology Data Exchange (ETDEWEB)

    Filichkin, Sergei; Etherington, Elizabeth; Ma, Caiping; Strauss, Steve

    2007-02-22

    The goals of the S.H. Strauss laboratory portion of 'Genome-enabled discovery of carbon sequestration genes in poplar' are (1) to explore the functions of candidate genes using Populus transformation by inserting genes provided by Oakridge National Laboratory (ORNL) and the University of Florida (UF) into poplar; (2) to expand the poplar transformation toolkit by developing transformation methods for important genotypes; and (3) to allow induced expression, and efficient gene suppression, in roots and other tissues. As part of the transformation improvement effort, OSU developed transformation protocols for Populus trichocarpa 'Nisqually-1' clone and an early flowering P. alba clone, 6K10. Complete descriptions of the transformation systems were published (Ma et. al. 2004, Meilan et. al 2004). Twenty-one 'Nisqually-1' and 622 6K10 transgenic plants were generated. To identify root predominant promoters, a set of three promoters were tested for their tissue-specific expression patterns in poplar and in Arabidopsis as a model system. A novel gene, ET304, was identified by analyzing a collection of poplar enhancer trap lines generated at OSU (Filichkin et. al 2006a, 2006b). Other promoters include the pGgMT1 root-predominant promoter from Casuarina glauca and the pAtPIN2 promoter from Arabidopsis root specific PIN2 gene. OSU tested two induction systems, alcohol- and estrogen-inducible, in multiple poplar transgenics. Ethanol proved to be the more efficient when tested in tissue culture and greenhouse conditions. Two estrogen-inducible systems were evaluated in transgenic Populus, neither of which functioned reliably in tissue culture conditions. GATEWAY-compatible plant binary vectors were designed to compare the silencing efficiency of homologous (direct) RNAi vs. heterologous (transitive) RNAi inverted repeats. A set of genes was targeted for post transcriptional silencing in the model Arabidopsis system; these include the floral

  15. Constitutively Elevated Salicylic Acid Levels Alter Photosynthesis and Oxidative State but Not Growth in Transgenic Populus[C][W

    Science.gov (United States)

    Xue, Liang-Jiao; Guo, Wenbing; Yuan, Yinan; Anino, Edward O.; Nyamdari, Batbayar; Wilson, Mark C.; Frost, Christopher J.; Chen, Han-Yi; Babst, Benjamin A.; Harding, Scott A.; Tsai, Chung-Jui

    2013-01-01

    Salicylic acid (SA) has long been implicated in plant responses to oxidative stress. SA overproduction in Arabidopsis thaliana leads to dwarfism, making in planta assessment of SA effects difficult in this model system. We report that transgenic Populus tremula × alba expressing a bacterial SA synthase hyperaccumulated SA and SA conjugates without negative growth consequences. In the absence of stress, endogenously elevated SA elicited widespread metabolic and transcriptional changes that resembled those of wild-type plants exposed to oxidative stress-promoting heat treatments. Potential signaling and oxidative stress markers azelaic and gluconic acids as well as antioxidant chlorogenic acids were strongly coregulated with SA, while soluble sugars and other phenylpropanoids were inversely correlated. Photosynthetic responses to heat were attenuated in SA-overproducing plants. Network analysis identified potential drivers of SA-mediated transcriptome rewiring, including receptor-like kinases and WRKY transcription factors. Orthologs of Arabidopsis SA signaling components NON-EXPRESSOR OF PATHOGENESIS-RELATED GENES1 and thioredoxins were not represented. However, all members of the expanded Populus nucleoredoxin-1 family exhibited increased expression and increased network connectivity in SA-overproducing Populus, suggesting a previously undescribed role in SA-mediated redox regulation. The SA response in Populus involved a reprogramming of carbon uptake and partitioning during stress that is compatible with constitutive chemical defense and sustained growth, contrasting with the SA response in Arabidopsis, which is transient and compromises growth if sustained. PMID:23903318

  16. International regulatory landscape and integration of corrective genome editing into in vitro fertilization.

    Science.gov (United States)

    Araki, Motoko; Ishii, Tetsuya

    2014-11-24

    Genome editing technology, including zinc finger nucleases (ZFNs), transcription activator-like effector nucleases (TALENs), and clustered regularly interspaced short palindromic repeat (CRISPR)/Cas, has enabled far more efficient genetic engineering even in non-human primates. This biotechnology is more likely to develop into medicine for preventing a genetic disease if corrective genome editing is integrated into assisted reproductive technology, represented by in vitro fertilization. Although rapid advances in genome editing are expected to make germline gene correction feasible in a clinical setting, there are many issues that still need to be addressed before this could occur. We herein examine current status of genome editing in mammalian embryonic stem cells and zygotes and discuss potential issues in the international regulatory landscape regarding human germline gene modification. Moreover, we address some ethical and social issues that would be raised when each country considers whether genome editing-mediated germline gene correction for preventive medicine should be permitted.

  17. Comparative physiology of allopatric Populus species: geographic clines in photosynthesis, height growth, and carbon isotope discrimination in common gardens.

    Science.gov (United States)

    Soolanayakanahally, Raju Y; Guy, Robert D; Street, Nathaniel R; Robinson, Kathryn M; Silim, Salim N; Albrectsen, Benedicte R; Jansson, Stefan

    2015-01-01

    Populus species with wide geographic ranges display strong adaptation to local environments. We studied the clinal patterns in phenology and ecophysiology in allopatric Populus species adapted to similar environments on different continents under common garden settings. As a result of climatic adaptation, both Populus tremula L. and Populus balsamifera L. display latitudinal clines in photosynthetic rates (A), whereby high-latitude trees of P. tremula had higher A compared to low-latitude trees and nearly so in P. balsamifera (p = 0.06). Stomatal conductance (g s) and chlorophyll content index (CCI) follow similar latitudinal trends. However, foliar nitrogen was positively correlated with latitude in P. balsamifera and negatively correlated in P. tremula. No significant trends in carbon isotope composition of the leaf tissue (δ(13)C) were observed for both species; but, intrinsic water-use efficiency (WUEi) was negatively correlated with the latitude of origin in P. balsamifera. In spite of intrinsically higher A, high-latitude trees in both common gardens accomplished less height gain as a result of early bud set. Thus, shoot biomass was determined by height elongation duration (HED), which was well approximated by the number of days available for free growth between bud flush and bud set. We highlight the shortcoming of unreplicated outdoor common gardens for tree improvement and the crucial role of photoperiod in limiting height growth, further complicating interpretation of other secondary effects.

  18. An integrated CRISPR Bombyx mori genome editing system with improved efficiency and expanded target sites.

    Science.gov (United States)

    Ma, Sanyuan; Liu, Yue; Liu, Yuanyuan; Chang, Jiasong; Zhang, Tong; Wang, Xiaogang; Shi, Run; Lu, Wei; Xia, Xiaojuan; Zhao, Ping; Xia, Qingyou

    2017-04-01

    Genome editing enabled unprecedented new opportunities for targeted genomic engineering of a wide variety of organisms ranging from microbes, plants, animals and even human embryos. The serial establishing and rapid applications of genome editing tools significantly accelerated Bombyx mori (B. mori) research during the past years. However, the only CRISPR system in B. mori was the commonly used SpCas9, which only recognize target sites containing NGG PAM sequence. In the present study, we first improve the efficiency of our previous established SpCas9 system by 3.5 folds. The improved high efficiency was also observed at several loci in both BmNs cells and B. mori embryos. Then to expand the target sites, we showed that two newly discovered CRISPR system, SaCas9 and AsCpf1, could also induce highly efficient site-specific genome editing in BmNs cells, and constructed an integrated CRISPR system. Genome-wide analysis of targetable sites was further conducted and showed that the integrated system cover 69,144,399 sites in B. mori genome, and one site could be found in every 6.5 bp. The efficiency and resolution of this CRISPR platform will probably accelerate both fundamental researches and applicable studies in B. mori, and perhaps other insects. Copyright © 2017 Elsevier Ltd. All rights reserved.

  19. Increasing the productivity of biomass plantations of Populus species and hybrids in the Pacific Northwest. Final report, September 14, 1981--December 31, 1996

    Energy Technology Data Exchange (ETDEWEB)

    DeBell, D.S.; Harrington, C.A.; Clendenen, G.W. [USDA Forest Service, Olympia, WA (United States)] [and others

    1997-08-01

    This final report represents the culmination of eight years of biological research devoted to increasing the productivity of short rotation plantations of Populus trichocarpa and Populus hybrids in the Pacific Northwest. Studies described herein provide an understanding of tree growth, stand development and biomass yield at various spacings, and how patterns thereof differ by Populus clone in monoclonal and polyclonal plantings. Also included is some information about factors related to wind damage in Populus plantings, use of leaf size as a predictor of growth potential, and approaches for estimating tree and stand biomass and biomass growth. The work was accomplished in three research plantations, all established cooperatively with the Washington State Department of Natural Resources (DNR) and located at the DNR Tree Improvement Center near Olympia. The first plantation was established in Spring 1986 to evaluate the highly touted {open_quotes}woodgrass{close_quotes} concept and compare it with more conventional short-rotation management regimes, using two Populus hybrid clones planted at five spacings. Besides providing scientific data to resolve the politicized {open_quotes}wood-grass{close_quotes} dispute, this plantation has furnished excellent data on stand dynamics and woody biomass yield. A second plantation was established at the same time; groups of trees therein received two levels of irrigation and different amounts of four fertilizer amendments, resulting in microsites with diverse moisture and nutrient conditions.

  20. The Conjugative Relaxase TrwC Promotes Integration of Foreign DNA in the Human Genome.

    Science.gov (United States)

    González-Prieto, Coral; Gabriel, Richard; Dehio, Christoph; Schmidt, Manfred; Llosa, Matxalen

    2017-06-15

    Bacterial conjugation is a mechanism of horizontal DNA transfer. The relaxase TrwC of the conjugative plasmid R388 cleaves one strand of the transferred DNA at the oriT gene, covalently attaches to it, and leads the single-stranded DNA (ssDNA) into the recipient cell. In addition, TrwC catalyzes site-specific integration of the transferred DNA into its target sequence present in the genome of the recipient bacterium. Here, we report the analysis of the efficiency and specificity of the integrase activity of TrwC in human cells, using the type IV secretion system of the human pathogen Bartonella henselae to introduce relaxase-DNA complexes. Compared to Mob relaxase from plasmid pBGR1, we found that TrwC mediated a 10-fold increase in the rate of plasmid DNA transfer to human cells and a 100-fold increase in the rate of chromosomal integration of the transferred DNA. We used linear amplification-mediated PCR and plasmid rescue to characterize the integration pattern in the human genome. DNA sequence analysis revealed mostly reconstituted oriT sequences, indicating that TrwC is active and recircularizes transferred DNA in human cells. One TrwC-mediated site-specific integration event was detected, proving that TrwC is capable of mediating site-specific integration in the human genome, albeit with very low efficiency compared to the rate of random integration. Our results suggest that TrwC may stabilize the plasmid DNA molecules in the nucleus of the human cell, probably by recircularization of the transferred DNA strand. This stabilization would increase the opportunities for integration of the DNA by the host machinery. IMPORTANCE Different biotechnological applications, including gene therapy strategies, require permanent modification of target cells. Long-term expression is achieved either by extrachromosomal persistence or by integration of the introduced DNA. Here, we studied the utility of conjugative relaxase TrwC, a bacterial protein with site

  1. Nucleotide excision repair : a multi-step mechanism required to maintain genome integrity

    NARCIS (Netherlands)

    Moser, Jill

    2010-01-01

    DNA is continuously exposed to exogenous and genotoxic insults including ionizing and ultraviolet radiation as well as chemical agents. DNA damage can compromise the integrity of the genome and have potentially deleterious effects. Ultraviolet light (UV) can induce the formation of helix distorting

  2. Chromosomally Integrated Human Herpesvirus 6: Models of Viral Genome Release from the Telomere and Impacts on Human Health.

    Science.gov (United States)

    Wood, Michael L; Royle, Nicola J

    2017-07-12

    Human herpesvirus 6A and 6B, alongside some other herpesviruses, have the striking capacity to integrate into telomeres, the terminal repeated regions of chromosomes. The chromosomally integrated forms, ciHHV-6A and ciHHV-6B, are proposed to be a state of latency and it has been shown that they can both be inherited if integration occurs in the germ line. The first step in full viral reactivation must be the release of the integrated viral genome from the telomere and here we propose various models of this release involving transcription of the viral genome, replication fork collapse, and t-circle mediated release. In this review, we also discuss the relationship between ciHHV-6 and the telomere carrying the insertion, particularly how the presence and subsequent partial or complete release of the ciHHV-6 genome may affect telomere dynamics and the risk of disease.

  3. Using low energy x-ray radiography to evaluate root initiation and growth of Populus

    Science.gov (United States)

    Ronald S., Jr. Zalesny; A. L. Friend; B. Kodrzycki; D.W. McDonald; R. Michaels; A.H. Wiese; J.W. Powers

    2007-01-01

    Populus roots have been studied less than aboveground tissues. However, there is an overwhelming need to evaluate root initiation and growth in order to understand the genetics and physiology of rooting, along with genotype x environment interactions.

  4. SPTEdb: a database for transposable elements in salicaceous plants

    Science.gov (United States)

    Jia, Zirui; Xiao, Yao; Ma, Wenjun; Wang, Junhui

    2018-01-01

    Abstract Although transposable elements (TEs) play significant roles in structural, functional and evolutionary dynamics of the salicaceous plants genome and the accurate identification, definition and classification of TEs are still inadequate. In this study, we identified 18 393 TEs from Populus trichocarpa, Populus euphratica and Salix suchowensis using a combination of signature-based, similarity-based and De novo method, and annotated them into 1621 families. A comprehensive and user-friendly web-based database, SPTEdb, was constructed and served for researchers. SPTEdb enables users to browse, retrieve and download the TEs sequences from the database. Meanwhile, several analysis tools, including BLAST, HMMER, GetORF and Cut sequence, were also integrated into SPTEdb to help users to mine the TEs data easily and effectively. In summary, SPTEdb will facilitate the study of TEs biology and functional genomics in salicaceous plants. Database URL: http://genedenovoweb.ticp.net:81/SPTEdb/index.php PMID:29688371

  5. Determination of As in tree-rings of poplar (Populus alba L.) by U-shaped DC arc.

    Science.gov (United States)

    Marković, D M; Novović, I; Vilotić, D; Ignjatović, Lj

    2009-04-01

    An argon-stabilized U-shaped DC arc with a system for aerosol introduction was used for determination of As in poplar (Populus alba L.) tree-rings. After optimization of the operating parameters and selection of the most appropriate signal integration time (30 s), the limit of detection for As was reduced to 15.0 ng/mL. This detection limit obtained with the optimal integration time was compared with those for other methods: inductively coupled plasma-atomic emission spectrometry (ICP-AES), direct coupled plasma-atomic emission spectrometry (DCP-AES), microwave induced plasma-atomic emission spectrometry (MIP-AES) and improved thermospray flame furnace atomic absorption spectrometry (TS-FF-AAS). Arsenic is toxic trace element which can adversely affect plant, animal and human health. As an indicator of environment pollution we collected poplar tree-rings from two locations. The first area was close to the "Nikola Tesla" (TENT-A) power plant, Obrenovac, while the other was in the urban area of Novi Sad. In all cases elevated average concentrations of As were registered in poplar tree-rings from the Obrenovac location.

  6. GenomeCAT: a versatile tool for the analysis and integrative visualization of DNA copy number variants.

    Science.gov (United States)

    Tebel, Katrin; Boldt, Vivien; Steininger, Anne; Port, Matthias; Ebert, Grit; Ullmann, Reinhard

    2017-01-06

    The analysis of DNA copy number variants (CNV) has increasing impact in the field of genetic diagnostics and research. However, the interpretation of CNV data derived from high resolution array CGH or NGS platforms is complicated by the considerable variability of the human genome. Therefore, tools for multidimensional data analysis and comparison of patient cohorts are needed to assist in the discrimination of clinically relevant CNVs from others. We developed GenomeCAT, a standalone Java application for the analysis and integrative visualization of CNVs. GenomeCAT is composed of three modules dedicated to the inspection of single cases, comparative analysis of multidimensional data and group comparisons aiming at the identification of recurrent aberrations in patients sharing the same phenotype, respectively. Its flexible import options ease the comparative analysis of own results derived from microarray or NGS platforms with data from literature or public depositories. Multidimensional data obtained from different experiment types can be merged into a common data matrix to enable common visualization and analysis. All results are stored in the integrated MySQL database, but can also be exported as tab delimited files for further statistical calculations in external programs. GenomeCAT offers a broad spectrum of visualization and analysis tools that assist in the evaluation of CNVs in the context of other experiment data and annotations. The use of GenomeCAT does not require any specialized computer skills. The various R packages implemented for data analysis are fully integrated into GenomeCATs graphical user interface and the installation process is supported by a wizard. The flexibility in terms of data import and export in combination with the ability to create a common data matrix makes the program also well suited as an interface between genomic data from heterogeneous sources and external software tools. Due to the modular architecture the functionality of

  7. The CanOE strategy: integrating genomic and metabolic contexts across multiple prokaryote genomes to find candidate genes for orphan enzymes.

    Directory of Open Access Journals (Sweden)

    Adam Alexander Thil Smith

    2012-05-01

    Full Text Available Of all biochemically characterized metabolic reactions formalized by the IUBMB, over one out of four have yet to be associated with a nucleic or protein sequence, i.e. are sequence-orphan enzymatic activities. Few bioinformatics annotation tools are able to propose candidate genes for such activities by exploiting context-dependent rather than sequence-dependent data, and none are readily accessible and propose result integration across multiple genomes. Here, we present CanOE (Candidate genes for Orphan Enzymes, a four-step bioinformatics strategy that proposes ranked candidate genes for sequence-orphan enzymatic activities (or orphan enzymes for short. The first step locates "genomic metabolons", i.e. groups of co-localized genes coding proteins catalyzing reactions linked by shared metabolites, in one genome at a time. These metabolons can be particularly helpful for aiding bioanalysts to visualize relevant metabolic data. In the second step, they are used to generate candidate associations between un-annotated genes and gene-less reactions. The third step integrates these gene-reaction associations over several genomes using gene families, and summarizes the strength of family-reaction associations by several scores. In the final step, these scores are used to rank members of gene families which are proposed for metabolic reactions. These associations are of particular interest when the metabolic reaction is a sequence-orphan enzymatic activity. Our strategy found over 60,000 genomic metabolons in more than 1,000 prokaryote organisms from the MicroScope platform, generating candidate genes for many metabolic reactions, of which more than 70 distinct orphan reactions. A computational validation of the approach is discussed. Finally, we present a case study on the anaerobic allantoin degradation pathway in Escherichia coli K-12.

  8. RGmatch: matching genomic regions to proximal genes in omics data integration

    Directory of Open Access Journals (Sweden)

    Pedro Furió-Tarí

    2016-11-01

    Full Text Available Abstract Background The integrative analysis of multiple genomics data often requires that genome coordinates-based signals have to be associated with proximal genes. The relative location of a genomic region with respect to the gene (gene area is important for functional data interpretation; hence algorithms that match regions to genes should be able to deliver insight into this information. Results In this work we review the tools that are publicly available for making region-to-gene associations. We also present a novel method, RGmatch, a flexible and easy-to-use Python tool that computes associations either at the gene, transcript, or exon level, applying a set of rules to annotate each region-gene association with the region location within the gene. RGmatch can be applied to any organism as long as genome annotation is available. Furthermore, we qualitatively and quantitatively compare RGmatch to other tools. Conclusions RGmatch simplifies the association of a genomic region with its closest gene. At the same time, it is a powerful tool because the rules used to annotate these associations are very easy to modify according to the researcher’s specific interests. Some important differences between RGmatch and other similar tools already in existence are RGmatch’s flexibility, its wide range of user options, compatibility with any annotatable organism, and its comprehensive and user-friendly output.

  9. Ensembl Genomes 2016: more genomes, more complexity.

    Science.gov (United States)

    Kersey, Paul Julian; Allen, James E; Armean, Irina; Boddu, Sanjay; Bolt, Bruce J; Carvalho-Silva, Denise; Christensen, Mikkel; Davis, Paul; Falin, Lee J; Grabmueller, Christoph; Humphrey, Jay; Kerhornou, Arnaud; Khobova, Julia; Aranganathan, Naveen K; Langridge, Nicholas; Lowy, Ernesto; McDowall, Mark D; Maheswari, Uma; Nuhn, Michael; Ong, Chuang Kee; Overduin, Bert; Paulini, Michael; Pedro, Helder; Perry, Emily; Spudich, Giulietta; Tapanari, Electra; Walts, Brandon; Williams, Gareth; Tello-Ruiz, Marcela; Stein, Joshua; Wei, Sharon; Ware, Doreen; Bolser, Daniel M; Howe, Kevin L; Kulesha, Eugene; Lawson, Daniel; Maslen, Gareth; Staines, Daniel M

    2016-01-04

    Ensembl Genomes (http://www.ensemblgenomes.org) is an integrating resource for genome-scale data from non-vertebrate species, complementing the resources for vertebrate genomics developed in the context of the Ensembl project (http://www.ensembl.org). Together, the two resources provide a consistent set of programmatic and interactive interfaces to a rich range of data including reference sequence, gene models, transcriptional data, genetic variation and comparative analysis. This paper provides an update to the previous publications about the resource, with a focus on recent developments. These include the development of new analyses and views to represent polyploid genomes (of which bread wheat is the primary exemplar); and the continued up-scaling of the resource, which now includes over 23 000 bacterial genomes, 400 fungal genomes and 100 protist genomes, in addition to 55 genomes from invertebrate metazoa and 39 genomes from plants. This dramatic increase in the number of included genomes is one part of a broader effort to automate the integration of archival data (genome sequence, but also associated RNA sequence data and variant calls) within the context of reference genomes and make it available through the Ensembl user interfaces. © The Author(s) 2015. Published by Oxford University Press on behalf of Nucleic Acids Research.

  10. In vitro analysis of integrated global high-resolution DNA methylation profiling with genomic imbalance and gene expression in osteosarcoma.

    Directory of Open Access Journals (Sweden)

    Bekim Sadikovic

    Full Text Available Genetic and epigenetic changes contribute to deregulation of gene expression and development of human cancer. Changes in DNA methylation are key epigenetic factors regulating gene expression and genomic stability. Recent progress in microarray technologies resulted in developments of high resolution platforms for profiling of genetic, epigenetic and gene expression changes. OS is a pediatric bone tumor with characteristically high level of numerical and structural chromosomal changes. Furthermore, little is known about DNA methylation changes in OS. Our objective was to develop an integrative approach for analysis of high-resolution epigenomic, genomic, and gene expression profiles in order to identify functional epi/genomic differences between OS cell lines and normal human osteoblasts. A combination of Affymetrix Promoter Tilling Arrays for DNA methylation, Agilent array-CGH platform for genomic imbalance and Affymetrix Gene 1.0 platform for gene expression analysis was used. As a result, an integrative high-resolution approach for interrogation of genome-wide tumour-specific changes in DNA methylation was developed. This approach was used to provide the first genomic DNA methylation maps, and to identify and validate genes with aberrant DNA methylation in OS cell lines. This first integrative analysis of global cancer-related changes in DNA methylation, genomic imbalance, and gene expression has provided comprehensive evidence of the cumulative roles of epigenetic and genetic mechanisms in deregulation of gene expression networks.

  11. TIGER: Toolbox for integrating genome-scale metabolic models, expression data, and transcriptional regulatory networks

    Directory of Open Access Journals (Sweden)

    Jensen Paul A

    2011-09-01

    Full Text Available Abstract Background Several methods have been developed for analyzing genome-scale models of metabolism and transcriptional regulation. Many of these methods, such as Flux Balance Analysis, use constrained optimization to predict relationships between metabolic flux and the genes that encode and regulate enzyme activity. Recently, mixed integer programming has been used to encode these gene-protein-reaction (GPR relationships into a single optimization problem, but these techniques are often of limited generality and lack a tool for automating the conversion of rules to a coupled regulatory/metabolic model. Results We present TIGER, a Toolbox for Integrating Genome-scale Metabolism, Expression, and Regulation. TIGER converts a series of generalized, Boolean or multilevel rules into a set of mixed integer inequalities. The package also includes implementations of existing algorithms to integrate high-throughput expression data with genome-scale models of metabolism and transcriptional regulation. We demonstrate how TIGER automates the coupling of a genome-scale metabolic model with GPR logic and models of transcriptional regulation, thereby serving as a platform for algorithm development and large-scale metabolic analysis. Additionally, we demonstrate how TIGER's algorithms can be used to identify inconsistencies and improve existing models of transcriptional regulation with examples from the reconstructed transcriptional regulatory network of Saccharomyces cerevisiae. Conclusion The TIGER package provides a consistent platform for algorithm development and extending existing genome-scale metabolic models with regulatory networks and high-throughput data.

  12. Transcriptome responses to aluminum stress in roots of aspen (Populus tremula

    Directory of Open Access Journals (Sweden)

    Grisel Nadine

    2010-08-01

    Full Text Available Abstract Background Ionic aluminum (mainly Al3+ is rhizotoxic and can be present in acid soils at concentrations high enough to inhibit root growth. Many forest tree species grow naturally in acid soils and often tolerate high concentrations of Al. Previously, we have shown that aspen (Populus tremula releases citrate and oxalate from roots in response to Al exposure. To obtain further insights into the root responses of aspen to Al, we investigated root gene expression at Al conditions that inhibit root growth. Results Treatment of the aspen roots with 500 μM Al induced a strong inhibition of root growth within 6 h of exposure time. The root growth subsequently recovered, reaching growth rates comparable to that of control plants. Changes in gene expression were determined after 6 h, 2 d, and 10 d of Al exposure. Replicated transcriptome analyses using the Affymetrix poplar genome array revealed a total of 175 significantly up-regulated and 69 down-regulated genes, of which 70% could be annotated based on Arabidopsis genome resources. Between 6 h and 2 d, the number of responsive genes strongly decreased from 202 to 26, and then the number of changes remained low. The responses after 6 h were characterized by genes involved in cell wall modification, ion transport, and oxidative stress. Two genes with prolonged induction were closely related to the Arabidopsis Al tolerance genes ALS3 (for Al sensitive 3 and MATE (for multidrug and toxin efflux protein, mediating citrate efflux. Patterns of expression in different plant organs and in response to Al indicated that the two aspen genes are homologs of the Arabidopsis ALS3 and MATE. Conclusion Exposure of aspen roots to Al results in a rapid inhibition of root growth and a large change in root gene expression. The subsequent root growth recovery and the concomitant reduction in the number of responsive genes presumably reflect the success of the roots in activating Al tolerance mechanisms. The

  13. EasyCloneMulti: A Set of Vectors for Simultaneous and Multiple Genomic Integrations in Saccharomyces cerevisiae

    DEFF Research Database (Denmark)

    Maury, Jerome; Germann, Susanne Manuela; Jacobsen, Simo Abdessamad

    2016-01-01

    Saccharomyces cerevisiae is widely used in the biotechnology industry for production of ethanol, recombinant proteins, food ingredients and other chemicals. In order to generate highly producing and stable strains, genome integration of genes encoding metabolic pathway enzymes is the preferred...... of integrative vectors, EasyCloneMulti, that enables multiple and simultaneous integration of genes in S. cerevisiae. By creating vector backbones that combine consensus sequences that aim at targeting subsets of Ty sequences and a quickly degrading selective marker, integrations at multiple genomic loci...... and a range of expression levels were obtained, as assessed with the green fluorescent protein (GFP) reporter system. The EasyCloneMulti vector set was applied to balance the expression of the rate-controlling step in the β-alanine pathway for biosynthesis of 3-hydroxypropionic acid (3HP). The best 3HP...

  14. Bud removal affects shoot, root, and callus development of hardwood Populus cuttings

    Science.gov (United States)

    A.H. Wiese; J.A. Zalesny; D.M. Donner; Ronald S., Jr. Zalesny

    2006-01-01

    The inadvertent removal and/or damage of buds during processing and planting of hardwood poplar (Populus spp.) cuttings are a concern because of their potential impact on shoot and root development during establishment. The objective of the current study was to test for differences in shoot dry mass, root dry mass, number of roots, length of the...

  15. Integrative Functional Genomics for Systems Genetics in GeneWeaver.org.

    Science.gov (United States)

    Bubier, Jason A; Langston, Michael A; Baker, Erich J; Chesler, Elissa J

    2017-01-01

    The abundance of existing functional genomics studies permits an integrative approach to interpreting and resolving the results of diverse systems genetics studies. However, a major challenge lies in assembling and harmonizing heterogeneous data sets across species for facile comparison to the positional candidate genes and coexpression networks that come from systems genetic studies. GeneWeaver is an online database and suite of tools at www.geneweaver.org that allows for fast aggregation and analysis of gene set-centric data. GeneWeaver contains curated experimental data together with resource-level data such as GO annotations, MP annotations, and KEGG pathways, along with persistent stores of user entered data sets. These can be entered directly into GeneWeaver or transferred from widely used resources such as GeneNetwork.org. Data are analyzed using statistical tools and advanced graph algorithms to discover new relations, prioritize candidate genes, and generate function hypotheses. Here we use GeneWeaver to find genes common to multiple gene sets, prioritize candidate genes from a quantitative trait locus, and characterize a set of differentially expressed genes. Coupling a large multispecies repository curated and empirical functional genomics data to fast computational tools allows for the rapid integrative analysis of heterogeneous data for interpreting and extrapolating systems genetics results.

  16. Shoot position affects root initiation and growth of dormant unrooted cuttings of Populus

    Science.gov (United States)

    R.S., Jr. Zalesny; R.B. Hall; E.O. Bauer; D.E. Riemenschneider

    2003-01-01

    Rooting of dormant unrooted cuttings is crucial to the commercial deployment of intensively cultured poplar (Populus spp.) plantations because it is the first biological prerequisite to stand establishment. Rooting can be genetically controlled and subject to selection. Thus, our objective was to test for differences in rooting ability among cuttings...

  17. Clonal variation in lateral and basal rooting of Populus irrigated with landfill leachate

    Science.gov (United States)

    R.S. Zalesny Jr.; J.A. Zalesny

    2011-01-01

    Successful establishment and productivity of Populus depends upon adventitious rooting from: 1) lateral roots that develop from either preformed or induced primordia and 2) basal roots that differentiate from callus at the base of the cutting in response to wounding. Information is needed for phytotechnologies about the degree to which ...

  18. A DNMT3A2-HDAC2 Complex Is Essential for Genomic Imprinting and Genome Integrity in Mouse Oocytes

    Directory of Open Access Journals (Sweden)

    Pengpeng Ma

    2015-11-01

    Full Text Available Maternal genomic imprints are established during oogenesis. Histone deacetylases (HDACs 1 and 2 are required for oocyte development in mouse, but their role in genomic imprinting is unknown. We find that Hdac1:Hdac2−/− double-mutant growing oocytes exhibit global DNA hypomethylation and fail to establish imprinting marks for Igf2r, Peg3, and Srnpn. Global hypomethylation correlates with increased retrotransposon expression and double-strand DNA breaks. Nuclear-associated DNMT3A2 is reduced in double-mutant oocytes, and injecting these oocytes with Hdac2 partially restores DNMT3A2 nuclear staining. DNMT3A2 co-immunoprecipitates with HDAC2 in mouse embryonic stem cells. Partial loss of nuclear DNMT3A2 and HDAC2 occurs in Sin3a−/− oocytes, which exhibit decreased DNA methylation of imprinting control regions for Igf2r and Srnpn, but not Peg3. These results suggest seminal roles of HDAC1/2 in establishing maternal genomic imprints and maintaining genomic integrity in oocytes mediated in part through a SIN3A complex that interacts with DNMT3A2.

  19. RNA sequencing of Populus x canadensis roots identifies key molecular mechanisms underlying physiological adaption to excess zinc.

    Directory of Open Access Journals (Sweden)

    Andrea Ariani

    Full Text Available Populus x canadensis clone I-214 exhibits a general indicator phenotype in response to excess Zn, and a higher metal uptake in roots than in shoots with a reduced translocation to aerial parts under hydroponic conditions. This physiological adaptation seems mainly regulated by roots, although the molecular mechanisms that underlie these processes are still poorly understood. Here, differential expression analysis using RNA-sequencing technology was used to identify the molecular mechanisms involved in the response to excess Zn in root. In order to maximize specificity of detection of differentially expressed (DE genes, we consider the intersection of genes identified by three distinct statistical approaches (61 up- and 19 down-regulated and validate them by RT-qPCR, yielding an agreement of 93% between the two experimental techniques. Gene Ontology (GO terms related to oxidation-reduction processes, transport and cellular iron ion homeostasis were enriched among DE genes, highlighting the importance of metal homeostasis in adaptation to excess Zn by P. x canadensis clone I-214. We identified the up-regulation of two Populus metal transporters (ZIP2 and NRAMP1 probably involved in metal uptake, and the down-regulation of a NAS4 gene involved in metal translocation. We identified also four Fe-homeostasis transcription factors (two bHLH38 genes, FIT and BTS that were differentially expressed, probably for reducing Zn-induced Fe-deficiency. In particular, we suggest that the down-regulation of FIT transcription factor could be a mechanism to cope with Zn-induced Fe-deficiency in Populus. These results provide insight into the molecular mechanisms involved in adaption to excess Zn in Populus spp., but could also constitute a starting point for the identification and characterization of molecular markers or biotechnological targets for possible improvement of phytoremediation performances of poplar trees.

  20. MIPS Arabidopsis thaliana Database (MAtDB): an integrated biological knowledge resource based on the first complete plant genome

    Science.gov (United States)

    Schoof, Heiko; Zaccaria, Paolo; Gundlach, Heidrun; Lemcke, Kai; Rudd, Stephen; Kolesov, Grigory; Arnold, Roland; Mewes, H. W.; Mayer, Klaus F. X.

    2002-01-01

    Arabidopsis thaliana is the first plant for which the complete genome has been sequenced and published. Annotation of complex eukaryotic genomes requires more than the assignment of genetic elements to the sequence. Besides completing the list of genes, we need to discover their cellular roles, their regulation and their interactions in order to understand the workings of the whole plant. The MIPS Arabidopsis thaliana Database (MAtDB; http://mips.gsf.de/proj/thal/db) started out as a repository for genome sequence data in the European Scientists Sequencing Arabidopsis (ESSA) project and the Arabidopsis Genome Initiative. Our aim is to transform MAtDB into an integrated biological knowledge resource by integrating diverse data, tools, query and visualization capabilities and by creating a comprehensive resource for Arabidopsis as a reference model for other species, including crop plants. PMID:11752263

  1. A genome-wide analysis of lentivector integration sites using targeted sequence capture and next generation sequencing technology.

    Science.gov (United States)

    Ustek, Duran; Sirma, Sema; Gumus, Ergun; Arikan, Muzaffer; Cakiris, Aris; Abaci, Neslihan; Mathew, Jaicy; Emrence, Zeliha; Azakli, Hulya; Cosan, Fulya; Cakar, Atilla; Parlak, Mahmut; Kursun, Olcay

    2012-10-01

    One application of next-generation sequencing (NGS) is the targeted resequencing of interested genes which has not been used in viral integration site analysis of gene therapy applications. Here, we combined targeted sequence capture array and next generation sequencing to address the whole genome profiling of viral integration sites. Human 293T and K562 cells were transduced with a HIV-1 derived vector. A custom made DNA probe sets targeted pLVTHM vector used to capture lentiviral vector/human genome junctions. The captured DNA was sequenced using GS FLX platform. Seven thousand four hundred and eighty four human genome sequences flanking the long terminal repeats (LTR) of pLVTHM fragment sequences matched with an identity of at least 98% and minimum 50 bp criteria in both cells. In total, 203 unique integration sites were identified. The integrations in both cell lines were totally distant from the CpG islands and from the transcription start sites and preferentially located in introns. A comparison between the two cell lines showed that the lentiviral-transduced DNA does not have the same preferred regions in the two different cell lines. Copyright © 2012 Elsevier B.V. All rights reserved.

  2. BiGG Models: A platform for integrating, standardizing and sharing genome-scale models

    DEFF Research Database (Denmark)

    King, Zachary A.; Lu, Justin; Dräger, Andreas

    2016-01-01

    Genome-scale metabolic models are mathematically-structured knowledge bases that can be used to predict metabolic pathway usage and growth phenotypes. Furthermore, they can generate and test hypotheses when integrated with experimental data. To maximize the value of these models, centralized repo...

  3. Isolation and characterization of cDNAs encoding leucoanthocyanidin reductase and anthocyanidin reductase from Populus trichocarpa.

    Directory of Open Access Journals (Sweden)

    Lijun Wang

    Full Text Available Proanthocyanidins (PAs contribute to poplar defense mechanisms against biotic and abiotic stresses. Transcripts of PA biosynthetic genes accumulated rapidly in response to infection by the fungus Marssonina brunnea f.sp. multigermtubi, treatments of salicylic acid (SA and wounding, resulting in PA accumulation in poplar leaves. Anthocyanidin reductase (ANR and leucoanthocyanidin reductase (LAR are two key enzymes of the PA biosynthesis that produce the main subunits: (+-catechin and (--epicatechin required for formation of PA polymers. In Populus, ANR and LAR are encoded by at least two and three highly related genes, respectively. In this study, we isolated and functionally characterized genes PtrANR1 and PtrLAR1 from P. trichocarpa. Phylogenetic analysis shows that Populus ANR1 and LAR1 occurr in two distinct phylogenetic lineages, but both genes have little difference in their tissue distribution, preferentially expressed in roots. Overexpression of PtrANR1 in poplar resulted in a significant increase in PA levels but no impact on catechin levels. Antisense down-regulation of PtrANR1 showed reduced PA accumulation in transgenic lines, but increased levels of anthocyanin content. Ectopic expression of PtrLAR1 in poplar positively regulated the biosynthesis of PAs, whereas the accumulation of anthocyanin and flavonol was significantly reduced (P<0.05 in all transgenic plants compared to the control plants. These results suggest that both PtrANR1 and PtrLAR1 contribute to PA biosynthesis in Populus.

  4. Remetabolism of transpired ethanol by Populus deltoides

    International Nuclear Information System (INIS)

    MacDonald, R.C.; Kimmerer, T.W.

    1990-01-01

    Ethanol is present in the transpiration stream of flooded and unflooded trees in concentrations up to 0.5mM. Transpired ethanol does not evaporate but is remetabolized by foliage and upper stems in Populus deltoides. 14 C-ethanol was supplied in the transpiration stream to excised leaves and shoots; more than 98% was incorporated. Less than 1% was respired as CO 2 . Organic and amino acids were labelled initially, with eventual accumulations in water- and chloroform-soluble fractions and into protein. Much of the label was incorporated into stem tissue, with little reaching the lamina. These experiments suggest that ethanol is not lost transpirationally through the leaves, but is efficiently recycled in a manner resembling lactate recycling in mammals

  5. Leaf chemical composition of twenty-one Populus hybrid clones grown under intensive culture

    Science.gov (United States)

    Richard E. Dickson; Philip R. Larson

    1976-01-01

    Leaf material from 21 nursery-grown Populus hybrid clones was analyzed for three nitrogen fractions (total N, soluble protein, and soluble amino acids) and three carbhydrate fractions (reducing sugars, total soluble sugars, and total nonstructural carbohydrates-TNC). In addition, nursery-grown green ash and silver maple, field-grown bigtooth and trembling aspen, and...

  6. Thidiazuron: A potent cytokinin for efficient plant regeneration in Himalayan poplar (Populus ciliata Wall. using leaf explants

    Directory of Open Access Journals (Sweden)

    Gaurav Aggarwal

    2012-11-01

    Full Text Available Populus species are important resource for certain branches of industry and have special roles for scientific study on biological and agricultural systems. The present investigation was undertaken with an objective of enhancing the frequency of plant regeneration in Himalayan poplar (Populus ciliata Wall.. The effect of Thiadizuron (TDZ alone and in combination with adenine and α-Naphthalene acetic acid (NAA were studied on the regeneration potential of leaf explants. A high efficiency of shoot regeneration was observed in leaf (80.00% explants on MS basal medium supplemented with 0.024 mg/l TDZ and 79.7 mg/l adenine. Elongation and multiplication of shoots were obtained on Murashige and Skoog (MS basal medium, containing 0.5 mg/l 6. Benzyl aminopurine (BAP + 0.2mg/l Indole 3-acetic acid (IAA + 0.3 mg/l Gibberellic acid (GA3. High frequency root regeneration from in vitro developed shoots was observed on MS basal medium supplemented with 0.10 mg/l Indole 3-butyric acid(IBA. Maximum of the in vitro rooted plantlets were well accomplished to the mixture of sand: soil (1:1 and exhibited similar morphology with the field plants. A high efficiency plant regeneration protocol has been developedfrom leaf explants in Himalayan poplar (Populus ciliata Wall..

  7. Comparative interrogation of the developing xylem transcriptomes of two wood-forming species: Populus trichocarpa and Eucalyptus grandis.

    Science.gov (United States)

    Hefer, Charles A; Mizrachi, Eshchar; Myburg, Alexander A; Douglas, Carl J; Mansfield, Shawn D

    2015-06-01

    Wood formation is a complex developmental process governed by genetic and environmental stimuli. Populus and Eucalyptus are fast-growing, high-yielding tree genera that represent ecologically and economically important species suitable for generating significant lignocellulosic biomass. Comparative analysis of the developing xylem and leaf transcriptomes of Populus trichocarpa and Eucalyptus grandis together with phylogenetic analyses identified clusters of homologous genes preferentially expressed during xylem formation in both species. A conserved set of 336 single gene pairs showed highly similar xylem preferential expression patterns, as well as evidence of high functional constraint. Individual members of multi-gene orthologous clusters known to be involved in secondary cell wall biosynthesis also showed conserved xylem expression profiles. However, species-specific expression as well as opposite (xylem versus leaf) expression patterns observed for a subset of genes suggest subtle differences in the transcriptional regulation important for xylem development in each species. Using sequence similarity and gene expression status, we identified functional homologs likely to be involved in xylem developmental and biosynthetic processes in Populus and Eucalyptus. Our study suggests that, while genes involved in secondary cell wall biosynthesis show high levels of gene expression conservation, differential regulation of some xylem development genes may give rise to unique xylem properties. © 2015 The Authors. New Phytologist © 2015 New Phytologist Trust.

  8. IMGMD: A platform for the integration and standardisation of In silico Microbial Genome-scale Metabolic Models.

    Science.gov (United States)

    Ye, Chao; Xu, Nan; Dong, Chuan; Ye, Yuannong; Zou, Xuan; Chen, Xiulai; Guo, Fengbiao; Liu, Liming

    2017-04-07

    Genome-scale metabolic models (GSMMs) constitute a platform that combines genome sequences and detailed biochemical information to quantify microbial physiology at the system level. To improve the unity, integrity, correctness, and format of data in published GSMMs, a consensus IMGMD database was built in the LAMP (Linux + Apache + MySQL + PHP) system by integrating and standardizing 328 GSMMs constructed for 139 microorganisms. The IMGMD database can help microbial researchers download manually curated GSMMs, rapidly reconstruct standard GSMMs, design pathways, and identify metabolic targets for strategies on strain improvement. Moreover, the IMGMD database facilitates the integration of wet-lab and in silico data to gain an additional insight into microbial physiology. The IMGMD database is freely available, without any registration requirements, at http://imgmd.jiangnan.edu.cn/database.

  9. Xylan hydrolysis in Populus trichocarpa × P. deltoides and model substrates during hydrothermal pretreatment.

    Science.gov (United States)

    Trajano, Heather L; Pattathil, Sivakumar; Tomkins, Bruce A; Tschaplinski, Timothy J; Hahn, Michael G; Van Berkel, Gary J; Wyman, Charles E

    2015-03-01

    Previous studies defined easy and difficult to hydrolyze fractions of hemicellulose that may result from bonds among cellulose, hemicellulose, and lignin. To understand how such bonds affect hydrolysis, Populus trichocarpa × Populus deltoides, holocellulose isolated from P. trichocarpa × P. deltoides and birchwood xylan were subjected to hydrothermal flow-through pretreatment. Samples were characterized by glycome profiling, HPLC, and UPLC-MS. Glycome profiling revealed steady fragmentation and removal of glycans from solids during hydrolysis. The extent of polysaccharide fragmentation, hydrolysis rate, and total xylose yield were lowest for P. trichocarpa × P. deltoides and greatest for birchwood xylan. Comparison of results from P. trichocarpa × P. deltoides and holocellulose suggested that lignin-carbohydrate complexes reduce hydrolysis rates and limit release of large xylooligomers. Smaller differences between results with holocellulose and birchwood xylan suggest xylan-cellulose hydrogen bonds limited hydrolysis, but to a lesser extent. These findings imply cell wall structure strongly influences hydrolysis. Copyright © 2014 Elsevier Ltd. All rights reserved.

  10. Drosophila Sld5 is essential for normal cell cycle progression and maintenance of genomic integrity

    Energy Technology Data Exchange (ETDEWEB)

    Gouge, Catherine A. [Department of Biology, East Carolina University East Carolina University, Greenville, NC 27858 (United States); Christensen, Tim W., E-mail: christensent@ecu.edu [Department of Biology, East Carolina University East Carolina University, Greenville, NC 27858 (United States)

    2010-09-10

    Research highlights: {yields} Drosophila Sld5 interacts with Psf1, PPsf2, and Mcm10. {yields} Haploinsufficiency of Sld5 leads to M-phase delay and genomic instability. {yields} Sld5 is also required for normal S phase progression. -- Abstract: Essential for the normal functioning of a cell is the maintenance of genomic integrity. Failure in this process is often catastrophic for the organism, leading to cell death or mis-proliferation. Central to genomic integrity is the faithful replication of DNA during S phase. The GINS complex has recently come to light as a critical player in DNA replication through stabilization of MCM2-7 and Cdc45 as a member of the CMG complex which is likely responsible for the processivity of helicase activity during S phase. The GINS complex is made up of 4 members in a 1:1:1:1 ratio: Psf1, Psf2, Psf3, And Sld5. Here we present the first analysis of the function of the Sld5 subunit in a multicellular organism. We show that Drosophila Sld5 interacts with Psf1, Psf2, and Mcm10 and that mutations in Sld5 lead to M and S phase delays with chromosomes exhibiting hallmarks of genomic instability.

  11. IVAG: An Integrative Visualization Application for Various Types of Genomic Data Based on R-Shiny and the Docker Platform.

    Science.gov (United States)

    Lee, Tae-Rim; Ahn, Jin Mo; Kim, Gyuhee; Kim, Sangsoo

    2017-12-01

    Next-generation sequencing (NGS) technology has become a trend in the genomics research area. There are many software programs and automated pipelines to analyze NGS data, which can ease the pain for traditional scientists who are not familiar with computer programming. However, downstream analyses, such as finding differentially expressed genes or visualizing linkage disequilibrium maps and genome-wide association study (GWAS) data, still remain a challenge. Here, we introduce a dockerized web application written in R using the Shiny platform to visualize pre-analyzed RNA sequencing and GWAS data. In addition, we have integrated a genome browser based on the JBrowse platform and an automated intermediate parsing process required for custom track construction, so that users can easily build and navigate their personal genome tracks with in-house datasets. This application will help scientists perform series of downstream analyses and obtain a more integrative understanding about various types of genomic data by interactively visualizing them with customizable options.

  12. The European Renal Genome Project: An Integrated Approach Towards Understanding the Genetics of Kidney Development and Disease

    OpenAIRE

    Willnow, TE; Antignac, C; Brändli, AW; Christensen, EI; Cox, RD; Davidson, D; Davies, JA; Devuyst, O; Eichele, G; Hastie, ND; Verroust, PJ; Schedl, A; Meij, IC

    2005-01-01

    Rapid progress in genome research creates a wealth of information on the functional annotation of mammalian genome sequences. However, as we accumulate large amounts of scientific information we are facing problems of how to integrate and relate the data produced by various genomic approaches. Here, we propose the novel concept of an organ atlas where diverse data from expression maps to histological findings to mutant phenotypes can be queried, compared and visualized in the context of a thr...

  13. Influence of Populus Genotype on Gene Expression by the Wood Decay Fungus Phanerochaete chrysosporium

    Science.gov (United States)

    Jill Gaskell; Amber Marty; Michael Mozuch; Philip J. Kersten; Sandra Splinter Bondurant; Grzegorz Sabat; Ali Azarpira; John Ralph; Oleksandr Skyba; Shawn D. Mansfield; Robert A. Blanchette; Dan Cullen

    2014-01-01

    We examined gene expression patterns in the lignin-degrading fungus Phanerochaete chrysosporium when it colonizes hybrid poplar (Populus alba tremula) and syringyl (S)-rich transgenic derivatives. Acombination ofmicroarrays and liquid chromatography- tandem mass spectrometry (LC-MS/MS) allowed detection of a total of 9,959 transcripts and 793...

  14. RegPredict: an integrated system for regulon inference in prokaryotes by comparative genomics approach

    Energy Technology Data Exchange (ETDEWEB)

    Novichkov, Pavel S.; Rodionov, Dmitry A.; Stavrovskaya, Elena D.; Novichkova, Elena S.; Kazakov, Alexey E.; Gelfand, Mikhail S.; Arkin, Adam P.; Mironov, Andrey A.; Dubchak, Inna

    2010-05-26

    RegPredict web server is designed to provide comparative genomics tools for reconstruction and analysis of microbial regulons using comparative genomics approach. The server allows the user to rapidly generate reference sets of regulons and regulatory motif profiles in a group of prokaryotic genomes. The new concept of a cluster of co-regulated orthologous operons allows the user to distribute the analysis of large regulons and to perform the comparative analysis of multiple clusters independently. Two major workflows currently implemented in RegPredict are: (i) regulon reconstruction for a known regulatory motif and (ii) ab initio inference of a novel regulon using several scenarios for the generation of starting gene sets. RegPredict provides a comprehensive collection of manually curated positional weight matrices of regulatory motifs. It is based on genomic sequences, ortholog and operon predictions from the MicrobesOnline. An interactive web interface of RegPredict integrates and presents diverse genomic and functional information about the candidate regulon members from several web resources. RegPredict is freely accessible at http://regpredict.lbl.gov.

  15. CTDB: An Integrated Chickpea Transcriptome Database for Functional and Applied Genomics

    OpenAIRE

    Verma, Mohit; Kumar, Vinay; Patel, Ravi K.; Garg, Rohini; Jain, Mukesh

    2015-01-01

    Chickpea is an important grain legume used as a rich source of protein in human diet. The narrow genetic diversity and limited availability of genomic resources are the major constraints in implementing breeding strategies and biotechnological interventions for genetic enhancement of chickpea. We developed an integrated Chickpea Transcriptome Database (CTDB), which provides the comprehensive web interface for visualization and easy retrieval of transcriptome data in chickpea. The database fea...

  16. Quantitative and qualitative proteome characteristics extracted from in-depth integrated genomics and proteomics analysis

    NARCIS (Netherlands)

    Low, T.Y.; van Heesch, S.; van den Toorn, H.; Giansanti, P.; Cristobal, A.; Toonen, P.; Schafer, S.; Hubner, N.; van Breukelen, B.; Mohammed, S.; Cuppen, E.; Heck, A.J.R.; Guryev, V.

    2013-01-01

    Quantitative and qualitative protein characteristics are regulated at genomic, transcriptomic, and posttranscriptional levels. Here, we integrated in-depth transcriptome and proteome analyses of liver tissues from two rat strains to unravel the interactions within and between these layers. We

  17. Integrative genome analyses identify key somatic driver mutations of small-cell lung cancer

    NARCIS (Netherlands)

    Peifer, Martin; Fernandez-Cuesta, Lynnette; Sos, Martin L.; George, Julie; Seidel, Danila; Kasper, Lawryn H.; Plenker, Dennis; Leenders, Frauke; Sun, Ruping; Zander, Thomas; Menon, Roopika; Koker, Mirjam; Dahmen, Ilona; Mueller, Christian; Di Cerbo, Vincenzo; Schildhaus, Hans-Ulrich; Altmueller, Janine; Baessmann, Ingelore; Becker, Christian; de Wilde, Bram; Vandesompele, Jo; Boehm, Diana; Ansen, Sascha; Gabler, Franziska; Wilkening, Ines; Heynck, Stefanie; Heuckmann, Johannes M.; Lu, Xin; Carter, Scott L.; Cibulskis, Kristian; Banerji, Shantanu; Getz, Gad; Park, Kwon-Sik; Rauh, Daniel; Gruetter, Christian; Fischer, Matthias; Pasqualucci, Laura; Wright, Gavin; Wainer, Zoe; Russell, Prudence; Petersen, Iver; Chen, Yuan; Stoelben, Erich; Ludwig, Corinna; Schnabel, Philipp; Hoffmann, Hans; Muley, Thomas; Brockmann, Michael; Engel-Riedel, Walburga; Muscarella, Lucia A.; Fazio, Vito M.; Groen, Harry; Timens, Wim; Sietsma, Hannie; Thunnissen, Erik; Smit, Egbert; Heideman, Danielle A. M.; Snijders, Peter J. F.; Cappuzzo, Federico; Ligorio, Claudia; Damiani, Stefania; Field, John; Solberg, Steinar; Brustugun, Odd Terje; Lund-Iversen, Marius; Saenger, Joerg; Clement, Joachim H.; Soltermann, Alex; Moch, Holger; Weder, Walter; Solomon, Benjamin; Soria, Jean-Charles; Validire, Pierre; Besse, Benjamin; Brambilla, Elisabeth; Brambilla, Christian; Lantuejoul, Sylvie; Lorimier, Philippe; Schneider, Peter M.; Hallek, Michael; Pao, William; Meyerson, Matthew; Sage, Julien; Shendure, Jay; Schneider, Robert; Buettner, Reinhard; Wolf, Juergen; Nuernberg, Peter; Perner, Sven; Heukamp, Lukas C.; Brindle, Paul K.; Haas, Stefan; Thomas, Roman K.

    2012-01-01

    Small-cell lung cancer (SCLC) is an aggressive lung tumor subtype with poor prognosis(1-3). We sequenced 29 SCLC exomes, 2 genomes and 15 transcriptomes and found an extremely high mutation rate of 7.4 +/- 1 protein-changing mutations per million base pairs. Therefore, we conducted integrated

  18. Construction of an ortholog database using the semantic web technology for integrative analysis of genomic data.

    Science.gov (United States)

    Chiba, Hirokazu; Nishide, Hiroyo; Uchiyama, Ikuo

    2015-01-01

    Recently, various types of biological data, including genomic sequences, have been rapidly accumulating. To discover biological knowledge from such growing heterogeneous data, a flexible framework for data integration is necessary. Ortholog information is a central resource for interlinking corresponding genes among different organisms, and the Semantic Web provides a key technology for the flexible integration of heterogeneous data. We have constructed an ortholog database using the Semantic Web technology, aiming at the integration of numerous genomic data and various types of biological information. To formalize the structure of the ortholog information in the Semantic Web, we have constructed the Ortholog Ontology (OrthO). While the OrthO is a compact ontology for general use, it is designed to be extended to the description of database-specific concepts. On the basis of OrthO, we described the ortholog information from our Microbial Genome Database for Comparative Analysis (MBGD) in the form of Resource Description Framework (RDF) and made it available through the SPARQL endpoint, which accepts arbitrary queries specified by users. In this framework based on the OrthO, the biological data of different organisms can be integrated using the ortholog information as a hub. Besides, the ortholog information from different data sources can be compared with each other using the OrthO as a shared ontology. Here we show some examples demonstrating that the ortholog information described in RDF can be used to link various biological data such as taxonomy information and Gene Ontology. Thus, the ortholog database using the Semantic Web technology can contribute to biological knowledge discovery through integrative data analysis.

  19. Quantitative and Qualitative Proteome Characteristics Extracted from In-Depth Integrated Genomics and Proteomics Analysis

    NARCIS (Netherlands)

    Low, Teck Yew; van Heesch, Sebastiaan; van den Toorn, Henk; Giansanti, Piero; Cristobal, Alba; Toonen, Pim; Schafer, Sebastian; Huebner, Norbert; van Breukelen, Bas; Mohammed, Shabaz; Cuppen, Edwin; Heck, Albert J. R.; Guryev, Victor

    2013-01-01

    Quantitative and qualitative protein characteristics are regulated at genomic, transcriptomic, and post-transcriptional levels. Here, we integrated in-depth transcriptome and proteome analyses of liver tissues from two rat strains to unravel the interactions within and between these layers. We

  20. Soil temperature and precipitation affect the rooting ability of dormant hardwood cuttings of Populus

    Science.gov (United States)

    R.S., Jr. Zalesny; R.B. Hall; E.O. Bauer; D.E. Riemenschneider

    2005-01-01

    In addition to genetic control, responses to environmental stimuli affect the success of rooting. Our objectives were to: 1) assess the variation in rooting ability among 21 Populus clones grown under varying soil temperatures and amounts of precipitation and 2) identify combinations of soil temperature and precipitation that promote rooting. The...

  1. MicroScope—an integrated microbial resource for the curation and comparative analysis of genomic and metabolic data

    Science.gov (United States)

    Vallenet, David; Belda, Eugeni; Calteau, Alexandra; Cruveiller, Stéphane; Engelen, Stefan; Lajus, Aurélie; Le Fèvre, François; Longin, Cyrille; Mornico, Damien; Roche, David; Rouy, Zoé; Salvignol, Gregory; Scarpelli, Claude; Thil Smith, Adam Alexander; Weiman, Marion; Médigue, Claudine

    2013-01-01

    MicroScope is an integrated platform dedicated to both the methodical updating of microbial genome annotation and to comparative analysis. The resource provides data from completed and ongoing genome projects (automatic and expert annotations), together with data sources from post-genomic experiments (i.e. transcriptomics, mutant collections) allowing users to perfect and improve the understanding of gene functions. MicroScope (http://www.genoscope.cns.fr/agc/microscope) combines tools and graphical interfaces to analyse genomes and to perform the manual curation of gene annotations in a comparative context. Since its first publication in January 2006, the system (previously named MaGe for Magnifying Genomes) has been continuously extended both in terms of data content and analysis tools. The last update of MicroScope was published in 2009 in the Database journal. Today, the resource contains data for >1600 microbial genomes, of which ∼300 are manually curated and maintained by biologists (1200 personal accounts today). Expert annotations are continuously gathered in the MicroScope database (∼50 000 a year), contributing to the improvement of the quality of microbial genomes annotations. Improved data browsing and searching tools have been added, original tools useful in the context of expert annotation have been developed and integrated and the website has been significantly redesigned to be more user-friendly. Furthermore, in the context of the European project Microme (Framework Program 7 Collaborative Project), MicroScope is becoming a resource providing for the curation and analysis of both genomic and metabolic data. An increasing number of projects are related to the study of environmental bacterial (meta)genomes that are able to metabolize a large variety of chemical compounds that may be of high industrial interest. PMID:23193269

  2. Integrating sequencing technologies in personal genomics: optimal low cost reconstruction of structural variants.

    Directory of Open Access Journals (Sweden)

    Jiang Du

    2009-07-01

    Full Text Available The goal of human genome re-sequencing is obtaining an accurate assembly of an individual's genome. Recently, there has been great excitement in the development of many technologies for this (e.g. medium and short read sequencing from companies such as 454 and SOLiD, and high-density oligo-arrays from Affymetrix and NimbelGen, with even more expected to appear. The costs and sensitivities of these technologies differ considerably from each other. As an important goal of personal genomics is to reduce the cost of re-sequencing to an affordable point, it is worthwhile to consider optimally integrating technologies. Here, we build a simulation toolbox that will help us optimally combine different technologies for genome re-sequencing, especially in reconstructing large structural variants (SVs. SV reconstruction is considered the most challenging step in human genome re-sequencing. (It is sometimes even harder than de novo assembly of small genomes because of the duplications and repetitive sequences in the human genome. To this end, we formulate canonical problems that are representative of issues in reconstruction and are of small enough scale to be computationally tractable and simulatable. Using semi-realistic simulations, we show how we can combine different technologies to optimally solve the assembly at low cost. With mapability maps, our simulations efficiently handle the inhomogeneous repeat-containing structure of the human genome and the computational complexity of practical assembly algorithms. They quantitatively show how combining different read lengths is more cost-effective than using one length, how an optimal mixed sequencing strategy for reconstructing large novel SVs usually also gives accurate detection of SNPs/indels, how paired-end reads can improve reconstruction efficiency, and how adding in arrays is more efficient than just sequencing for disentangling some complex SVs. Our strategy should facilitate the sequencing of

  3. Selective Gene Delivery for Integrating Exogenous DNA into Plastid and Mitochondrial Genomes Using Peptide-DNA Complexes.

    Science.gov (United States)

    Yoshizumi, Takeshi; Oikawa, Kazusato; Chuah, Jo-Ann; Kodama, Yutaka; Numata, Keiji

    2018-05-14

    Selective gene delivery into organellar genomes (mitochondrial and plastid genomes) has been limited because of a lack of appropriate platform technology, even though these organelles are essential for metabolite and energy production. Techniques for selective organellar modification are needed to functionally improve organelles and produce transplastomic/transmitochondrial plants. However, no method for mitochondrial genome modification has yet been established for multicellular organisms including plants. Likewise, modification of plastid genomes has been limited to a few plant species and algae. In the present study, we developed ionic complexes of fusion peptides containing organellar targeting signal and plasmid DNA for selective delivery of exogenous DNA into the plastid and mitochondrial genomes of intact plants. This is the first report of exogenous DNA being integrated into the mitochondrial genomes of not only plants, but also multicellular organisms in general. This fusion peptide-mediated gene delivery system is a breakthrough platform for both plant organellar biotechnology and gene therapy for mitochondrial diseases in animals.

  4. GMATA: An Integrated Software Package for Genome-Scale SSR Mining, Marker Development and Viewing.

    Science.gov (United States)

    Wang, Xuewen; Wang, Le

    2016-01-01

    Simple sequence repeats (SSRs), also referred to as microsatellites, are highly variable tandem DNAs that are widely used as genetic markers. The increasing availability of whole-genome and transcript sequences provides information resources for SSR marker development. However, efficient software is required to efficiently identify and display SSR information along with other gene features at a genome scale. We developed novel software package Genome-wide Microsatellite Analyzing Tool Package (GMATA) integrating SSR mining, statistical analysis and plotting, marker design, polymorphism screening and marker transferability, and enabled simultaneously display SSR markers with other genome features. GMATA applies novel strategies for SSR analysis and primer design in large genomes, which allows GMATA to perform faster calculation and provides more accurate results than existing tools. Our package is also capable of processing DNA sequences of any size on a standard computer. GMATA is user friendly, only requires mouse clicks or types inputs on the command line, and is executable in multiple computing platforms. We demonstrated the application of GMATA in plants genomes and reveal a novel distribution pattern of SSRs in 15 grass genomes. The most abundant motifs are dimer GA/TC, the A/T monomer and the GCG/CGC trimer, rather than the rich G/C content in DNA sequence. We also revealed that SSR count is a linear to the chromosome length in fully assembled grass genomes. GMATA represents a powerful application tool that facilitates genomic sequence analyses. GAMTA is freely available at http://sourceforge.net/projects/gmata/?source=navbar.

  5. Genome puzzle master (GPM): an integrated pipeline for building and editing pseudomolecules from fragmented sequences.

    Science.gov (United States)

    Zhang, Jianwei; Kudrna, Dave; Mu, Ting; Li, Weiming; Copetti, Dario; Yu, Yeisoo; Goicoechea, Jose Luis; Lei, Yang; Wing, Rod A

    2016-10-15

    Next generation sequencing technologies have revolutionized our ability to rapidly and affordably generate vast quantities of sequence data. Once generated, raw sequences are assembled into contigs or scaffolds. However, these assemblies are mostly fragmented and inaccurate at the whole genome scale, largely due to the inability to integrate additional informative datasets (e.g. physical, optical and genetic maps). To address this problem, we developed a semi-automated software tool-Genome Puzzle Master (GPM)-that enables the integration of additional genomic signposts to edit and build 'new-gen-assemblies' that result in high-quality 'annotation-ready' pseudomolecules. With GPM, loaded datasets can be connected to each other via their logical relationships which accomplishes tasks to 'group,' 'merge,' 'order and orient' sequences in a draft assembly. Manual editing can also be performed with a user-friendly graphical interface. Final pseudomolecules reflect a user's total data package and are available for long-term project management. GPM is a web-based pipeline and an important part of a Laboratory Information Management System (LIMS) which can be easily deployed on local servers for any genome research laboratory. The GPM (with LIMS) package is available at https://github.com/Jianwei-Zhang/LIMS CONTACTS: jzhang@mail.hzau.edu.cn or rwing@mail.arizona.eduSupplementary information: Supplementary data are available at Bioinformatics online. © The Author 2016. Published by Oxford University Press.

  6. Family genome browser: visualizing genomes with pedigree information.

    Science.gov (United States)

    Juan, Liran; Liu, Yongzhuang; Wang, Yongtian; Teng, Mingxiang; Zang, Tianyi; Wang, Yadong

    2015-07-15

    Families with inherited diseases are widely used in Mendelian/complex disease studies. Owing to the advances in high-throughput sequencing technologies, family genome sequencing becomes more and more prevalent. Visualizing family genomes can greatly facilitate human genetics studies and personalized medicine. However, due to the complex genetic relationships and high similarities among genomes of consanguineous family members, family genomes are difficult to be visualized in traditional genome visualization framework. How to visualize the family genome variants and their functions with integrated pedigree information remains a critical challenge. We developed the Family Genome Browser (FGB) to provide comprehensive analysis and visualization for family genomes. The FGB can visualize family genomes in both individual level and variant level effectively, through integrating genome data with pedigree information. Family genome analysis, including determination of parental origin of the variants, detection of de novo mutations, identification of potential recombination events and identical-by-decent segments, etc., can be performed flexibly. Diverse annotations for the family genome variants, such as dbSNP memberships, linkage disequilibriums, genes, variant effects, potential phenotypes, etc., are illustrated as well. Moreover, the FGB can automatically search de novo mutations and compound heterozygous variants for a selected individual, and guide investigators to find high-risk genes with flexible navigation options. These features enable users to investigate and understand family genomes intuitively and systematically. The FGB is available at http://mlg.hit.edu.cn/FGB/. © The Author 2015. Published by Oxford University Press. All rights reserved. For Permissions, please e-mail: journals.permissions@oup.com.

  7. Filling the knowledge gap: Integrating quantitative genetics and genomics in graduate education and outreach

    Science.gov (United States)

    The genomics revolution provides vital tools to address global food security. Yet to be incorporated into livestock breeding, molecular techniques need to be integrated into a quantitative genetics framework. Within the U.S., with shrinking faculty numbers with the requisite skills, the capacity to ...

  8. The Proteins API: accessing key integrated protein and genome information.

    Science.gov (United States)

    Nightingale, Andrew; Antunes, Ricardo; Alpi, Emanuele; Bursteinas, Borisas; Gonzales, Leonardo; Liu, Wudong; Luo, Jie; Qi, Guoying; Turner, Edd; Martin, Maria

    2017-07-03

    The Proteins API provides searching and programmatic access to protein and associated genomics data such as curated protein sequence positional annotations from UniProtKB, as well as mapped variation and proteomics data from large scale data sources (LSS). Using the coordinates service, researchers are able to retrieve the genomic sequence coordinates for proteins in UniProtKB. This, the LSS genomics and proteomics data for UniProt proteins is programmatically only available through this service. A Swagger UI has been implemented to provide documentation, an interface for users, with little or no programming experience, to 'talk' to the services to quickly and easily formulate queries with the services and obtain dynamically generated source code for popular programming languages, such as Java, Perl, Python and Ruby. Search results are returned as standard JSON, XML or GFF data objects. The Proteins API is a scalable, reliable, fast, easy to use RESTful services that provides a broad protein information resource for users to ask questions based upon their field of expertise and allowing them to gain an integrated overview of protein annotations available to aid their knowledge gain on proteins in biological processes. The Proteins API is available at (http://www.ebi.ac.uk/proteins/api/doc). © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  9. Arbuscular mycorrhizal fungi associated with Populus-Salix stands in a semiarid riparian ecosystem

    Science.gov (United States)

    Beauchamp, Vanessa B.; Stromberg, J.C.; Stutz, J.C.

    2006-01-01

    ??? This study examined the activity, species richness, and species composition of the arbuscular mycorrhizal fungal (AMF) community of Populus-Salix stands on the Verde River (Arizona, USA), quantified patterns of AMF richness and colonization along complex floodplain gradients, and identified environmental variables responsible for structuring the AMF community. ??? Samples from 61 Populus-Salix stands were analyzed for AMF and herbaceous composition, AMF colonization, gravimetric soil moisture, soil texture, per cent organic matter, pH, and concentrations of nitrate, bicarbonate phosphorus and exchangeable potassium. ??? AMF species richness declined with stand age and distance from and elevation above the channel and was positively related to perennial species cover and richness and gravimetric soil moisture. Distance from and elevation above the active channel, forest age, annual species cover, perennial species richness, and exchangeable potassium concentration all played a role in structuring the AMF community in this riparian area. ??? Most AMF species were found across a wide range of soil conditions, but a subset of species tended to occur more often in hydric areas. This group of riparian affiliate AMF species includes several not previously encountered in the surrounding Sonoran desert. ?? New Phytologist (2006).

  10. Chloride and sodium uptake potential over an entire rotation of Populus irrigated with landfill leachate

    Science.gov (United States)

    Jill A. Zalesny; Ronald S., Jr. Zalesny

    2009-01-01

    There is a need for information about the response of Populus genotypes to repeated application of high-salinity water and nutrient sources throughout an entire rotation. We have combined establishment biomass and uptake data with mid- and full-rotation growth data to project potential chloride (Cl−) and sodium (Na...

  11. SNUGB: a versatile genome browser supporting comparative and functional fungal genomics

    Directory of Open Access Journals (Sweden)

    Kim Seungill

    2008-12-01

    Full Text Available Abstract Background Since the full genome sequences of Saccharomyces cerevisiae were released in 1996, genome sequences of over 90 fungal species have become publicly available. The heterogeneous formats of genome sequences archived in different sequencing centers hampered the integration of the data for efficient and comprehensive comparative analyses. The Comparative Fungal Genomics Platform (CFGP was developed to archive these data via a single standardized format that can support multifaceted and integrated analyses of the data. To facilitate efficient data visualization and utilization within and across species based on the architecture of CFGP and associated databases, a new genome browser was needed. Results The Seoul National University Genome Browser (SNUGB integrates various types of genomic information derived from 98 fungal/oomycete (137 datasets and 34 plant and animal (38 datasets species, graphically presents germane features and properties of each genome, and supports comparison between genomes. The SNUGB provides three different forms of the data presentation interface, including diagram, table, and text, and six different display options to support visualization and utilization of the stored information. Information for individual species can be quickly accessed via a new tool named the taxonomy browser. In addition, SNUGB offers four useful data annotation/analysis functions, including 'BLAST annotation.' The modular design of SNUGB makes its adoption to support other comparative genomic platforms easy and facilitates continuous expansion. Conclusion The SNUGB serves as a powerful platform supporting comparative and functional genomics within the fungal kingdom and also across other kingdoms. All data and functions are available at the web site http://genomebrowser.snu.ac.kr/.

  12. An evolvable oestrogen receptor activity sensor: development of a modular system for integrating multiple genes into the yeast genome

    NARCIS (Netherlands)

    Fox, J.E.; Bridgham, J.T.; Bovee, T.F.H.; Thornton, J.W.

    2007-01-01

    To study a gene interaction network, we developed a gene-targeting strategy that allows efficient and stable genomic integration of multiple genetic constructs at distinct target loci in the yeast genome. This gene-targeting strategy uses a modular plasmid with a recyclable selectable marker and a

  13. The Next Generation Precision Medical Record - A Framework for Integrating Genomes and Wearable Sensors with Medical Records

    OpenAIRE

    Batra, Prag; Singh, Enakshi; Bog, Anja; Wright, Mark; Ashley, Euan; Waggott, Daryl

    2016-01-01

    Current medical records are rigid with regards to emerging big biomedical data. Examples of poorly integrated big data that already exist in clinical practice include whole genome sequencing and wearable sensors for real time monitoring. Genome sequencing enables conventional diagnostic interrogation and forms the fundamental baseline for precision health throughout a patients lifetime. Mobile sensors enable tailored monitoring regimes for both reducing risk through precision health intervent...

  14. Effects of summer and winter harvesting on element phytoextraction efficiency of Salix and Populus clones planted on contaminated soil.

    Science.gov (United States)

    Kubátová, Pavla; Száková, Jiřina; Břendová, Kateřina; Kroulíková-Vondráčková, Stanislava; Mercl, Filip; Tlustoš, Pavel

    2018-04-16

    The clones of fast-growing trees (FGTs) were investigated for phytoextraction of soil contaminated with risk elements (REs), especially Cd, Pb, and Zn. As a main experimental factor, the potential effect of biomass harvesting time was assessed. The field experiment with two Salix clones (S1 - (Salix schwerinii × Salix viminalis) × S. viminalis, S2 - S. × smithiana) and two Populus clones (P1 - Populus maximowiczii × Populus nigra, P2 - P. nigra) was established in April 2009. Shoots of all clones were first harvested in February 2012. After two further growing seasons, the first half of the trees was harvested in September 2013 before leaf fall (summer harvest) and the second half in February 2014 (winter harvest). Remediation factors (RFs) for all clones and all REs (except Pb for clone S1) were higher in the summer harvest. The highest annual RFs for Cd and for Zn (1.34 and 0.67%, respectively) were found for clone S2 and were significantly higher than other clones. Although no increased mortality of trees harvested in the summer was detected in the following season, the effect of summer harvesting on the phytoextraction potential of FGTs clones should be investigated in long-term studies.

  15. Assembly and comparative analysis of complete mitochondrial genome sequence of an economic plant Salix suchowensis

    Directory of Open Access Journals (Sweden)

    Ning Ye

    2017-03-01

    Full Text Available Willow is a widely used dioecious woody plant of Salicaceae family in China. Due to their high biomass yields, willows are promising sources for bioenergy crops. In this study, we assembled the complete mitochondrial (mt genome sequence of S. suchowensis with the length of 644,437 bp using Roche-454 GS FLX Titanium sequencing technologies. Base composition of the S. suchowensis mt genome is A (27.43%, T (27.59%, C (22.34%, and G (22.64%, which shows a prevalent GC content with that of other angiosperms. This long circular mt genome encodes 58 unique genes (32 protein-coding genes, 23 tRNA genes and 3 rRNA genes, and 9 of the 32 protein-coding genes contain 17 introns. Through the phylogenetic analysis of 35 species based on 23 protein-coding genes, it is supported that Salix as a sister to Populus. With the detailed phylogenetic information and the identification of phylogenetic position, some ribosomal protein genes and succinate dehydrogenase genes are found usually lost during evolution. As a native shrub willow species, this worthwhile research of S. suchowensis mt genome will provide more desirable information for better understanding the genomic breeding and missing pieces of sex determination evolution in the future.

  16. Field performance of Populus in short-rotation intensive culture plantations in the north-central U.S.

    Science.gov (United States)

    Edward A. Hansen; Michael E. Ostry; Wendell D. Johnson; David N. Tolsted; Daniel A. Netzer; William E. Berguson; Richard B. Hall

    1994-01-01

    Describes a network of short-rotation, Populus research and demonstration plantations that has been established across a 5-state region in the north-central U.S. to identify suitable hybrid poplar clones for large-scale biomass plantations in the region. Reports 6-year results.

  17. MIPS Arabidopsis thaliana Database (MAtDB): an integrated biological knowledge resource for plant genomics

    Science.gov (United States)

    Schoof, Heiko; Ernst, Rebecca; Nazarov, Vladimir; Pfeifer, Lukas; Mewes, Hans-Werner; Mayer, Klaus F. X.

    2004-01-01

    Arabidopsis thaliana is the most widely studied model plant. Functional genomics is intensively underway in many laboratories worldwide. Beyond the basic annotation of the primary sequence data, the annotated genetic elements of Arabidopsis must be linked to diverse biological data and higher order information such as metabolic or regulatory pathways. The MIPS Arabidopsis thaliana database MAtDB aims to provide a comprehensive resource for Arabidopsis as a genome model that serves as a primary reference for research in plants and is suitable for transfer of knowledge to other plants, especially crops. The genome sequence as a common backbone serves as a scaffold for the integration of data, while, in a complementary effort, these data are enhanced through the application of state-of-the-art bioinformatics tools. This information is visualized on a genome-wide and a gene-by-gene basis with access both for web users and applications. This report updates the information given in a previous report and provides an outlook on further developments. The MAtDB web interface can be accessed at http://mips.gsf.de/proj/thal/db. PMID:14681437

  18. Use of belowground growing degree days to predict rooting of dormant hardwood cuttings of Populus

    Science.gov (United States)

    R.S., Jr. Zalesny; E.O. Bauer; D.E. Riemenschneider

    2004-01-01

    Planting Populus cuttings based on calendar days neglects soil temperature extremes and does not promote rooting based on specific genotypes. Our objectives were to: 1) test the biological efficacy of a thermal index based on belowground growing degree days (GDD) across the growing period, 2) test for interactions between belowground GDD and clones,...

  19. Litter Quality of Populus Species as Affected by Free-Air CO2

    OpenAIRE

    Vermue, E.; Buurman, P.; Hoosbeek, M.R.

    2009-01-01

    The effect of elevated CO2 and nitrogen fertilization on the molecular chemistry of litter of three Populus species and associated soil organic matter (SOM) was investigated by pyrolysis-gas chromatography/mass spectrometry. The results are based on 147 quantified organic compounds in 24 litter samples. Litter of P. euramerica was clearly different from that of P. nigra and P. alba. The latter two had higher contents of proteins, polysaccharides, and cutin/cutan, while the former had higher c...

  20. CoryneCenter – An online resource for the integrated analysis of corynebacterial genome and transcriptome data

    Directory of Open Access Journals (Sweden)

    Hüser Andrea T

    2007-11-01

    Full Text Available Abstract Background The introduction of high-throughput genome sequencing and post-genome analysis technologies, e.g. DNA microarray approaches, has created the potential to unravel and scrutinize complex gene-regulatory networks on a large scale. The discovery of transcriptional regulatory interactions has become a major topic in modern functional genomics. Results To facilitate the analysis of gene-regulatory networks, we have developed CoryneCenter, a web-based resource for the systematic integration and analysis of genome, transcriptome, and gene regulatory information for prokaryotes, especially corynebacteria. For this purpose, we extended and combined the following systems into a common platform: (1 GenDB, an open source genome annotation system, (2 EMMA, a MAGE compliant application for high-throughput transcriptome data storage and analysis, and (3 CoryneRegNet, an ontology-based data warehouse designed to facilitate the reconstruction and analysis of gene regulatory interactions. We demonstrate the potential of CoryneCenter by means of an application example. Using microarray hybridization data, we compare the gene expression of Corynebacterium glutamicum under acetate and glucose feeding conditions: Known regulatory networks are confirmed, but moreover CoryneCenter points out additional regulatory interactions. Conclusion CoryneCenter provides more than the sum of its parts. Its novel analysis and visualization features significantly simplify the process of obtaining new biological insights into complex regulatory systems. Although the platform currently focusses on corynebacteria, the integrated tools are by no means restricted to these species, and the presented approach offers a general strategy for the analysis and verification of gene regulatory networks. CoryneCenter provides freely accessible projects with the underlying genome annotation, gene expression, and gene regulation data. The system is publicly available at http://www.CoryneCenter.de.

  1. DNA damage in Populus tremuloides clones exposed to elevated O3

    International Nuclear Information System (INIS)

    Tai, Helen H.; Percy, Kevin E.; Karnosky, David F.

    2010-01-01

    The effects of elevated concentrations of atmospheric tropospheric ozone (O 3 ) on DNA damage in five trembling aspen (Populus tremuloides Michx.) clones growing in a free-air enrichment experiment in the presence and absence of elevated concentrations of carbon dioxide (CO 2 ) were examined. Growing season mean hourly O 3 concentrations were 36.3 and 47.3 ppb for ambient and elevated O 3 plots, respectively. The 4th highest daily maximum 8-h ambient and elevated O 3 concentrations were 79 and 89 ppb, respectively. Elevated CO 2 averaged 524 ppm (+150 ppm) over the growing season. Exposure to O 3 and CO 2 in combination with O 3 increased DNA damage levels above background as measured by the comet assay. Ozone-tolerant clones 271 and 8L showed the highest levels of DNA damage under elevated O 3 compared with ambient air; whereas less tolerant clone 216 and sensitive clones 42E and 259 had comparably lower levels of DNA damage with no significant differences between elevated O 3 and ambient air. Clone 8L was demonstrated to have the highest level of excision DNA repair. In addition, clone 271 had the highest level of oxidative damage as measured by lipid peroxidation. The results suggest that variation in cellular responses to DNA damage between aspen clones may contribute to O 3 tolerance or sensitivity. - Ozone tolerant clones and sensitive Populus tremuloides clones show differences in DNA damage and repair.

  2. Biochemical basis of drought tolerance in hybrid Populus grown under field production conditions. CRADA final report

    Energy Technology Data Exchange (ETDEWEB)

    Tschaplinski, T.J.; Tuskan, G.A. [Oak Ridge National Lab., TN (United States); Wierman, C. [Boise Cascade Corp., Wallula, WA (United States)

    1997-04-01

    The purpose of this cooperative effort was to assess the use of osmotically active compounds as molecular selection criteria for drought tolerance in Populus in a large-scale field trial. It is known that some plant species, and individuals within a plant species, can tolerate increasing stress associated with reduced moisture availability by accumulating solutes. The biochemical matrix of such metabolites varies among species and among individuals. The ability of Populus clones to tolerate drought has equal value to other fiber producers, i.e., the wood products industry, where irrigation is used in combination with other cultural treatments to obtain high dry weight yields. The research initially involved an assessment of drought stress under field conditions and characterization of changes in osmotic constitution among the seven clones across the six moisture levels. The near-term goal was to provide a mechanistic basis for clonal differences in productivity under various irrigation treatments over time.

  3. Alternative splicing studies of the reactive oxygen species gene network in Populus reveal two isoforms of high-isoelectric-point superoxide dismutase.

    Science.gov (United States)

    Srivastava, Vaibhav; Srivastava, Manoj Kumar; Chibani, Kamel; Nilsson, Robert; Rouhier, Nicolas; Melzer, Michael; Wingsle, Gunnar

    2009-04-01

    Recent evidence has shown that alternative splicing (AS) is widely involved in the regulation of gene expression, substantially extending the diversity of numerous proteins. In this study, a subset of expressed sequence tags representing members of the reactive oxygen species gene network was selected from the PopulusDB database to investigate AS mechanisms in Populus. Examples of all known types of AS were detected, but intron retention was the most common. Interestingly, the closest Arabidopsis (Arabidopsis thaliana) homologs of half of the AS genes identified in Populus are not reportedly alternatively spliced. Two genes encoding the protein of most interest in our study (high-isoelectric-point superoxide dismutase [hipI-SOD]) have been found in black cottonwood (Populus trichocarpa), designated PthipI-SODC1 and PthipI-SODC2. Analysis of the expressed sequence tag libraries has indicated the presence of two transcripts of PthipI-SODC1 (hipI-SODC1b and hipI-SODC1s). Alignment of these sequences with the PthipI-SODC1 gene showed that hipI-SODC1b was 69 bp longer than hipI-SODC1s due to an AS event involving the use of an alternative donor splice site in the sixth intron. Transcript analysis showed that the splice variant hipI-SODC1b was differentially expressed, being clearly expressed in cambial and xylem, but not phloem, regions. In addition, immunolocalization and mass spectrometric data confirmed the presence of hipI-SOD proteins in vascular tissue. The functionalities of the spliced gene products were assessed by expressing recombinant hipI-SOD proteins and in vitro SOD activity assays.

  4. Vernalization and the Chilling Requirement to Exit Bud Dormancy: Shared or Separate Regulation?

    Directory of Open Access Journals (Sweden)

    Amy M Brunner

    2014-12-01

    Full Text Available Similarities have long been recognized between vernalization, the prolonged exposure to cold temperatures that promotes the floral transition in many plants, and the chilling requirement to release bud dormancy in woody plants of temperate climates. In both cases the extended chilling period occurring during winter is used to coordinate developmental events to the appropriate seasonal time. However, whether or not these processes share common regulatory components and molecular mechanisms remain largely unknown. Both gene function and association genetics studies in Populus are beginning to answer this question. In Populus, studies have revealed that orthologs of the antagonistic flowering time genes FT and CEN/TFL1 might have central roles in both processes. We review Populus seasonal shoot development related to dormancy release and the floral transition and evidence for FT/TFL1-mediated regulation of these processes to consider the question of regulatory overlap. In addition, we discuss the potential for and challenges to integrating functional and population genomics studies to uncover the regulatory mechanisms underpinning these processes in woody plant systems.

  5. Formation of wood secondary cell wall may involve two type cellulose synthase complexes in Populus.

    Science.gov (United States)

    Xi, Wang; Song, Dongliang; Sun, Jiayan; Shen, Junhui; Li, Laigeng

    2017-03-01

    Cellulose biosynthesis is mediated by cellulose synthases (CesAs), which constitute into rosette-like cellulose synthase complexe (CSC) on the plasma membrane. Two types of CSCs in Arabidopsis are believed to be involved in cellulose synthesis in the primary cell wall and secondary cell walls, respectively. In this work, we found that the two type CSCs participated cellulose biosynthesis in differentiating xylem cells undergoing secondary cell wall thickening in Populus. During the cell wall thickening process, expression of one type CSC genes increased while expression of the other type CSC genes decreased. Suppression of different type CSC genes both affected the wall-thickening and disrupted the multilaminar structure of the secondary cell walls. When CesA7A was suppressed, crystalline cellulose content was reduced, which, however, showed an increase when CesA3D was suppressed. The CesA suppression also affected cellulose digestibility of the wood cell walls. The results suggest that two type CSCs are involved in coordinating the cellulose biosynthesis in formation of the multilaminar structure in Populus wood secondary cell walls.

  6. Cryopreservation of Populus trichocarpa and Salix dormant buds with recovery by grafting or direct rooting.

    Science.gov (United States)

    Bonnart, Remi; Waddell, John; Haiby, Kathy; Widrlenchner, Mark P; Volk, Gayle M

    2014-01-01

    Methods are needed for the conservation of clonally maintained trees of Populus and Salix. In this work, Populus trichocarpa and Salix genetic resources were cryopreserved using dormant scions as the source explant. We quantified the recovery of cryopreserved materials that originated from diverse field environments by using either direct sprouting or grafting. Scions (either at their original moisture content of 48 to 60% or dried to 30%) were slowly cooled to -35 degree C, transferred to the vapor phase of liquid nitrogen (LNV, -160 degree C), and warmed before determining survival. Dormant buds from P. trichocarpa clones from Westport and Boardman, OR had regrowth levels between 42 and 100%. Direct rooting of cryopreserved P. trichocarpa was also possible. Ten of 11 cryopreserved Salix accessions, representing 10 different species, exhibited at least 40% bud growth and rooting after 6 weeks when a bottom-heated rooting system was implemented. We demonstrate that dormant buds of P. trichocarpa and Salix accessions can be cryopreserved and successfully regenerated without the use of tissue culture.

  7. Integration of least angle regression with empirical Bayes for multi-locus genome-wide association studies

    Science.gov (United States)

    Multi-locus genome-wide association studies has become the state-of-the-art procedure to identify quantitative trait loci (QTL) associated with traits simultaneously. However, implementation of multi-locus model is still difficult. In this study, we integrated least angle regression with empirical B...

  8. Biomass traits and candidate genes for bioenergy revealed through association genetics in coppiced European Populus nigra (L.)

    NARCIS (Netherlands)

    Allwright, Mike Robert; Payne, Adrienne; Emiliani, Giovanni; Milner, Suzanne; Viger, Maud; Rouse, Franchesca; Keurentjes, Joost J.B.; Bérard, Aurélie; Wildhagen, Henning; Faivre-Rampant, Patricia; Polle, Andrea; Morgante, Michele; Taylor, Gail

    2016-01-01

    Background: Second generation (2G) bioenergy from lignocellulosic feedstocks has the potential to develop as a sustainable source of renewable energy; however, significant hurdles still remain for large-scale commercialisation. Populus is considered as a promising 2G feedstock and understanding

  9. Integrating Genomic Data Sets for Knowledge Discovery: An Informed Approach to Management of Captive Endangered Species

    Directory of Open Access Journals (Sweden)

    Kristopher J. L. Irizarry

    2016-01-01

    Full Text Available Many endangered captive populations exhibit reduced genetic diversity resulting in health issues that impact reproductive fitness and quality of life. Numerous cost effective genomic sequencing and genotyping technologies provide unparalleled opportunity for incorporating genomics knowledge in management of endangered species. Genomic data, such as sequence data, transcriptome data, and genotyping data, provide critical information about a captive population that, when leveraged correctly, can be utilized to maximize population genetic variation while simultaneously reducing unintended introduction or propagation of undesirable phenotypes. Current approaches aimed at managing endangered captive populations utilize species survival plans (SSPs that rely upon mean kinship estimates to maximize genetic diversity while simultaneously avoiding artificial selection in the breeding program. However, as genomic resources increase for each endangered species, the potential knowledge available for management also increases. Unlike model organisms in which considerable scientific resources are used to experimentally validate genotype-phenotype relationships, endangered species typically lack the necessary sample sizes and economic resources required for such studies. Even so, in the absence of experimentally verified genetic discoveries, genomics data still provides value. In fact, bioinformatics and comparative genomics approaches offer mechanisms for translating these raw genomics data sets into integrated knowledge that enable an informed approach to endangered species management.

  10. Integrating Genomic Data Sets for Knowledge Discovery: An Informed Approach to Management of Captive Endangered Species.

    Science.gov (United States)

    Irizarry, Kristopher J L; Bryant, Doug; Kalish, Jordan; Eng, Curtis; Schmidt, Peggy L; Barrett, Gini; Barr, Margaret C

    2016-01-01

    Many endangered captive populations exhibit reduced genetic diversity resulting in health issues that impact reproductive fitness and quality of life. Numerous cost effective genomic sequencing and genotyping technologies provide unparalleled opportunity for incorporating genomics knowledge in management of endangered species. Genomic data, such as sequence data, transcriptome data, and genotyping data, provide critical information about a captive population that, when leveraged correctly, can be utilized to maximize population genetic variation while simultaneously reducing unintended introduction or propagation of undesirable phenotypes. Current approaches aimed at managing endangered captive populations utilize species survival plans (SSPs) that rely upon mean kinship estimates to maximize genetic diversity while simultaneously avoiding artificial selection in the breeding program. However, as genomic resources increase for each endangered species, the potential knowledge available for management also increases. Unlike model organisms in which considerable scientific resources are used to experimentally validate genotype-phenotype relationships, endangered species typically lack the necessary sample sizes and economic resources required for such studies. Even so, in the absence of experimentally verified genetic discoveries, genomics data still provides value. In fact, bioinformatics and comparative genomics approaches offer mechanisms for translating these raw genomics data sets into integrated knowledge that enable an informed approach to endangered species management.

  11. Perspectives on Clinical Informatics: Integrating Large-Scale Clinical, Genomic, and Health Information for Clinical Care

    Directory of Open Access Journals (Sweden)

    In Young Choi

    2013-12-01

    Full Text Available The advances in electronic medical records (EMRs and bioinformatics (BI represent two significant trends in healthcare. The widespread adoption of EMR systems and the completion of the Human Genome Project developed the technologies for data acquisition, analysis, and visualization in two different domains. The massive amount of data from both clinical and biology domains is expected to provide personalized, preventive, and predictive healthcare services in the near future. The integrated use of EMR and BI data needs to consider four key informatics areas: data modeling, analytics, standardization, and privacy. Bioclinical data warehouses integrating heterogeneous patient-related clinical or omics data should be considered. The representative standardization effort by the Clinical Bioinformatics Ontology (CBO aims to provide uniquely identified concepts to include molecular pathology terminologies. Since individual genome data are easily used to predict current and future health status, different safeguards to ensure confidentiality should be considered. In this paper, we focused on the informatics aspects of integrating the EMR community and BI community by identifying opportunities, challenges, and approaches to provide the best possible care service for our patients and the population.

  12. Improved genome recovery and integrated cell-size analyses of individual uncultured microbial cells and viral particles.

    Science.gov (United States)

    Stepanauskas, Ramunas; Fergusson, Elizabeth A; Brown, Joseph; Poulton, Nicole J; Tupper, Ben; Labonté, Jessica M; Becraft, Eric D; Brown, Julia M; Pachiadaki, Maria G; Povilaitis, Tadas; Thompson, Brian P; Mascena, Corianna J; Bellows, Wendy K; Lubys, Arvydas

    2017-07-20

    Microbial single-cell genomics can be used to provide insights into the metabolic potential, interactions, and evolution of uncultured microorganisms. Here we present WGA-X, a method based on multiple displacement amplification of DNA that utilizes a thermostable mutant of the phi29 polymerase. WGA-X enhances genome recovery from individual microbial cells and viral particles while maintaining ease of use and scalability. The greatest improvements are observed when amplifying high G+C content templates, such as those belonging to the predominant bacteria in agricultural soils. By integrating WGA-X with calibrated index-cell sorting and high-throughput genomic sequencing, we are able to analyze genomic sequences and cell sizes of hundreds of individual, uncultured bacteria, archaea, protists, and viral particles, obtained directly from marine and soil samples, in a single experiment. This approach may find diverse applications in microbiology and in biomedical and forensic studies of humans and other multicellular organisms.Single-cell genomics can be used to study uncultured microorganisms. Here, Stepanauskas et al. present a method combining improved multiple displacement amplification and FACS, to obtain genomic sequences and cell size information from uncultivated microbial cells and viral particles in environmental samples.

  13. Transport and use of CO2 in the xylem sap of Populus deltoides

    International Nuclear Information System (INIS)

    Stringer, J.W.; Kimmerer, T.W.

    1990-01-01

    Results of recent experiments indicate an internal cycling of respiratory CO 2 in woody plants. The CO 2 concentration of xylem sap expressed from the twigs of field grown Populus deltoides ranged from .14 to .50 mM. The pH of the xylem sap was 5.7 to 6.7, providing a significant bicarbonate concentration in many samples. Total dissolved inorganic carbon (DIC = CO 2 + H 2 CO 3 + HCO 3 - ) was 0.5 mM to 1.3 mM. Results from the analysis of xylem sap of 10 other species of woody plants were similar. To determine the fate of DIC delivered to the leaves of Populus deltoides, excised leaves were fed 1mM NaHCO 3 (2 μCi NaH 14 CO 3 ml -1 ). Less than 0.4% of the label escaped from the leaves, and ≥93% was fixed. Of the carbon fixed 56% of the 14 C was found in the petiole and midrib, and 14% was in the major veins, with the remaining 30% in the minor veins and lamina. Shading of the peptiole and midrib of leaves decreased the amount of fixed carbon in these tissues to 38% and increased the amount in the lamina to 55%

  14. Global assessment of genomic variation in cattle by genome resequencing and high-throughput genotyping

    DEFF Research Database (Denmark)

    Zhan, Bujie; Fadista, João; Thomsen, Bo

    2011-01-01

    Background Integration of genomic variation with phenotypic information is an effective approach for uncovering genotype-phenotype associations. This requires an accurate identification of the different types of variation in individual genomes. Results We report the integration of the whole genome...... of split-read and read-pair approaches proved to be complementary in finding different signatures. CNVs were identified on the basis of the depth of sequenced reads, and by using SNP and CGH arrays. Conclusions Our results provide high resolution mapping of diverse classes of genomic variation...

  15. Modeling and interoperability of heterogeneous genomic big data for integrative processing and querying.

    Science.gov (United States)

    Masseroli, Marco; Kaitoua, Abdulrahman; Pinoli, Pietro; Ceri, Stefano

    2016-12-01

    While a huge amount of (epi)genomic data of multiple types is becoming available by using Next Generation Sequencing (NGS) technologies, the most important emerging problem is the so-called tertiary analysis, concerned with sense making, e.g., discovering how different (epi)genomic regions and their products interact and cooperate with each other. We propose a paradigm shift in tertiary analysis, based on the use of the Genomic Data Model (GDM), a simple data model which links genomic feature data to their associated experimental, biological and clinical metadata. GDM encompasses all the data formats which have been produced for feature extraction from (epi)genomic datasets. We specifically describe the mapping to GDM of SAM (Sequence Alignment/Map), VCF (Variant Call Format), NARROWPEAK (for called peaks produced by NGS ChIP-seq or DNase-seq methods), and BED (Browser Extensible Data) formats, but GDM supports as well all the formats describing experimental datasets (e.g., including copy number variations, DNA somatic mutations, or gene expressions) and annotations (e.g., regarding transcription start sites, genes, enhancers or CpG islands). We downloaded and integrated samples of all the above-mentioned data types and formats from multiple sources. The GDM is able to homogeneously describe semantically heterogeneous data and makes the ground for providing data interoperability, e.g., achieved through the GenoMetric Query Language (GMQL), a high-level, declarative query language for genomic big data. The combined use of the data model and the query language allows comprehensive processing of multiple heterogeneous data, and supports the development of domain-specific data-driven computations and bio-molecular knowledge discovery. Copyright © 2016 Elsevier Inc. All rights reserved.

  16. Generalized allometric regression to estimate biomass of Populus in short-rotation coppice

    Energy Technology Data Exchange (ETDEWEB)

    Ben Brahim, Mohammed; Gavaland, Andre; Cabanettes, Alain [INRA Centre de Toulouse, Castanet-Tolosane Cedex (France). Unite Agroforesterie et Foret Paysanne

    2000-07-01

    Data from four different stands were combined to establish a single generalized allometric equation to estimate above-ground biomass of individual Populus trees grown on short-rotation coppice. The generalized model was performed using diameter at breast height, the mean diameter and the mean height of each site as dependent variables and then compared with the stand-specific regressions using F-test. Results showed that this single regression estimates tree biomass well at each stand and does not introduce bias with increasing diameter.

  17. Reconstruction of putative DNA virus from endogenous rice tungro bacilliform virus-like sequences in the rice genome: implications for integration and evolution

    Directory of Open Access Journals (Sweden)

    Kishima Yuji

    2004-10-01

    Full Text Available Abstract Background Plant genomes contain various kinds of repetitive sequences such as transposable elements, microsatellites, tandem repeats and virus-like sequences. Most of them, with the exception of virus-like sequences, do not allow us to trace their origins nor to follow the process of their integration into the host genome. Recent discoveries of virus-like sequences in plant genomes led us to set the objective of elucidating the origin of the repetitive sequences. Endogenous rice tungro bacilliform virus (RTBV-like sequences (ERTBVs have been found throughout the rice genome. Here, we reconstructed putative virus structures from RTBV-like sequences in the rice genome and characterized to understand evolutionary implication, integration manner and involvements of endogenous virus segments in the corresponding disease response. Results We have collected ERTBVs from the rice genomes. They contain rearranged structures and no intact ORFs. The identified ERTBV segments were shown to be phylogenetically divided into three clusters. For each phylogenetic cluster, we were able to make a consensus alignment for a circular virus-like structure carrying two complete ORFs. Comparisons of DNA and amino acid sequences suggested the closely relationship between ERTBV and RTBV. The Oryza AA-genome species vary in the ERTBV copy number. The species carrying low-copy-number of ERTBV segments have been reported to be extremely susceptible to RTBV. The DNA methylation state of the ERTBV sequences was correlated with their copy number in the genome. Conclusions These ERTBV segments are unlikely to have functional potential as a virus. However, these sequences facilitate to establish putative virus that provided information underlying virus integration and evolutionary relationship with existing virus. Comparison of ERTBV among the Oryza AA-genome species allowed us to speculate a possible role of endogenous virus segments against its related disease.

  18. Deciphering the genomes of 16 Acanthamoeba species does not provide evidence of integration of known giant virus-associated mobile genetic elements.

    Science.gov (United States)

    Chelkha, Nisrine; Colson, Philippe; Levasseur, Anthony; La Scola, Bernard

    2018-06-02

    Giant viruses infect protozoa, especially amoebae of the genus Acanthamoeba. These viruses possess genetic elements named Mobilome. So far, this mobilome comprises provirophages which are integrated into the genome of their hosts, transpovirons, and Maverick/Polintons. Virophages replicate inside virus factories within Acanthamoeba and can decrease the infectivity of giant viruses. The virophage infecting CroV was found to be integrated in the host of CroV, Cafeteria roenbergensis, thus protecting C. roenbergensis by reduction of CroV multiplication. Because of this unique property, assessment of the mechanisms of replication of virophages and their relationship with giant viruses is a key element of this investigation. This work aimed at evaluating the presence and the dynamic of these mobile elements in sixteen Acanthamoeba genomes. No significant traces of the integration of genomes or sequences from known virophages were identified in all the available Acanthamoeba genomes. These results brought us to hypothesize that the interactions between mimiviruses and their virophages might occur through different mechanisms, or at low frequency. An additional explanation could be that our knowledge of the diversity of virophages is still very limited. Copyright © 2018 Elsevier B.V. All rights reserved.

  19. Stable integration of recombinant adeno-associated virus vector genomes after transduction of murine hematopoietic stem cells.

    Science.gov (United States)

    Han, Zongchao; Zhong, Li; Maina, Njeri; Hu, Zhongbo; Li, Xiaomiao; Chouthai, Nitin S; Bischof, Daniela; Weigel-Van Aken, Kirsten A; Slayton, William B; Yoder, Mervin C; Srivastava, Arun

    2008-03-01

    We previously reported that among single-stranded adeno-associated virus (ssAAV) vectors, serotypes 1 through 5, ssAAV1 is the most efficient in transducing murine hematopoietic stem cells (HSCs), but viral second-strand DNA synthesis remains a rate-limiting step. Subsequently, using double-stranded, self-complementary AAV (scAAV) vectors, serotypes 7 through 10, we observed that scAAV7 vectors also transduce murine HSCs efficiently. In the present study, we used scAAV1 and scAAV7 shuttle vectors to transduce HSCs in a murine bone marrow serial transplant model in vivo, which allowed examination of the AAV proviral integration pattern in the mouse genome, as well as recovery and nucleotide sequence analyses of AAV-HSC DNA junction fragments. The proviral genomes were stably integrated, and integration sites were localized to different mouse chromosomes. None of the integration sites was found to be in a transcribed gene, or near a cellular oncogene. None of the animals, monitored for up to 1 year, exhibited pathological abnormalities. Thus, AAV proviral integration-induced risk of oncogenesis was not found in our study, which provides functional confirmation of stable transduction of self-renewing multipotential HSCs by scAAV vectors as well as promise for the use of these vectors in the potential treatment of disorders of the hematopoietic system.

  20. Cytogenetic Analysis of Populus trichocarpa - Ribosomal DNA, Telomere Repeat Sequence, and Marker-selected BACs

    Science.gov (United States)

    M.N. lslam-Faridi; C.D. Nelson; S.P. DiFazio; L.E. Gunter; G.A. Tuskan

    2009-01-01

    The 185-285 rDNA and 55 rDNA loci in Populus trichocarpa were localized using fluorescent in situ hybridization (FISH). Two 185-285 rDNA sites and one 55 rDNA site were identified and located at the ends of 3 different chromosomes. FISH signals from the Arabidopsis-type telomere repeat sequence were observed at the distal ends of each chromosome. Six BAC clones...

  1. Inheritance of compartmentalization of wounds in sweetgum (Liquidambar styraciflua L.) and eastern cottonwood (Populus deltoides Bartr.)

    Science.gov (United States)

    P. W. Garrett; W. K. Randall; A. L. Shigo; W. C. Shortle

    1979-01-01

    Studies of half-sib progeny tests of sweetgum (Liquidambar styraciflua) and clonal plantings of eastern cottonwood (Populus deltoides) in Mississippi indicate that rate of wound closure and size of discolored columns associated with the wounds are both heritable traits. Both are independent of stem diameter, which was used as a...

  2. Cryopreservation of Populus trichocarpa and Salix using dormant buds with recovery by grafting or direct rooting

    Science.gov (United States)

    Populus trichocarpa and Salix can be successfully cryopreserved by using dormant scions as the source explants. These scions (either at their original moisture content of 48 to 60% or dried to 30%) were slowly cooled to –35 degree Celsius, transferred to the vapor phase of liquid nitrogen (LNV,-160...

  3. Polymorphic integrations of an endogenous gammaretrovirus in the mule deer genome.

    Science.gov (United States)

    Elleder, Daniel; Kim, Oekyung; Padhi, Abinash; Bankert, Jason G; Simeonov, Ivan; Schuster, Stephan C; Wittekindt, Nicola E; Motameny, Susanne; Poss, Mary

    2012-03-01

    Endogenous retroviruses constitute a significant genomic fraction in all mammalian species. Typically they are evolutionarily old and fixed in the host species population. Here we report on a novel endogenous gammaretrovirus (CrERVγ; for cervid endogenous gammaretrovirus) in the mule deer (Odocoileus hemionus) that is insertionally polymorphic among individuals from the same geographical location, suggesting that it has a more recent evolutionary origin. Using PCR-based methods, we identified seven CrERVγ proviruses and demonstrated that they show various levels of insertional polymorphism in mule deer individuals. One CrERVγ provirus was detected in all mule deer sampled but was absent from white-tailed deer, indicating that this virus originally integrated after the split of the two species, which occurred approximately one million years ago. There are, on average, 100 CrERVγ copies in the mule deer genome based on quantitative PCR analysis. A CrERVγ provirus was sequenced and contained intact open reading frames (ORFs) for three virus genes. Transcripts were identified covering the entire provirus. CrERVγ forms a distinct branch of the gammaretrovirus phylogeny, with the closest relatives of CrERVγ being endogenous gammaretroviruses from sheep and pig. We demonstrated that white-tailed deer (Odocoileus virginianus) and elk (Cervus canadensis) DNA contain proviruses that are closely related to mule deer CrERVγ in a conserved region of pol; more distantly related sequences can be identified in the genome of another member of the Cervidae, the muntjac (Muntiacus muntjak). The discovery of a novel transcriptionally active and insertionally polymorphic retrovirus in mammals could provide a useful model system to study the dynamic interaction between the host genome and an invading retrovirus.

  4. Examination of correlation between histidine and nickel absorption by Morus L., Robinia pseudoacacia L. and Populus nigra L. using HPLC-MS and ICP-MS.

    Science.gov (United States)

    Ozen, Sukran Akkus; Yaman, Mehmet

    2016-08-02

    In this study, HPLC-MS and ICP-MS methods were used for the determination of histidine and nickel in Morus L., Robinia pseudoacacia L., and Populus nigra L. leaves taken from industrial areas including Gaziantep and Bursa cities. In the determination of histidine by HPLC-MS, all of the system parameters such as flow rate of mobile phase, fragmentor potential, injection volume and column temperature were optimized and found to be 0.2 mL min(-1), 70 V, 15 µL, and 20°C, respectively. Under the optimum conditions, histidine was extracted from plant sample by distilled water at 90°C for 30 min. Concentrations of histidine as mg kg(-1) were found to be between 2-9 for Morus L., 6-13 for Robinia pseudoacacia L., and 2-10 for Populus nigra L. Concentrations of nickel were in the ranges of 5-10 mg kg(-1) for Morus L., 3-10 mg kg(-1) for Robinia pseudoacacia L., and 0.6-4 mg kg(-1) for Populus nigra L. A significant linear correlation (r = 0.78) between histidine and Ni was observed for Populus nigra L., whereas insignificant linear correlation for Robinia pseudoacacia L. (r = 0.22) were seen. Limits of detection (LOD) and quantitation (LOQ) were found to be 0.025 mg Ni L(-1) and 0.075 mg Ni L(-1), respectively.

  5. Integration of Multiple Genomic and Phenotype Data to Infer Novel miRNA-Disease Associations.

    Science.gov (United States)

    Shi, Hongbo; Zhang, Guangde; Zhou, Meng; Cheng, Liang; Yang, Haixiu; Wang, Jing; Sun, Jie; Wang, Zhenzhen

    2016-01-01

    MicroRNAs (miRNAs) play an important role in the development and progression of human diseases. The identification of disease-associated miRNAs will be helpful for understanding the molecular mechanisms of diseases at the post-transcriptional level. Based on different types of genomic data sources, computational methods for miRNA-disease association prediction have been proposed. However, individual source of genomic data tends to be incomplete and noisy; therefore, the integration of various types of genomic data for inferring reliable miRNA-disease associations is urgently needed. In this study, we present a computational framework, CHNmiRD, for identifying miRNA-disease associations by integrating multiple genomic and phenotype data, including protein-protein interaction data, gene ontology data, experimentally verified miRNA-target relationships, disease phenotype information and known miRNA-disease connections. The performance of CHNmiRD was evaluated by experimentally verified miRNA-disease associations, which achieved an area under the ROC curve (AUC) of 0.834 for 5-fold cross-validation. In particular, CHNmiRD displayed excellent performance for diseases without any known related miRNAs. The results of case studies for three human diseases (glioblastoma, myocardial infarction and type 1 diabetes) showed that all of the top 10 ranked miRNAs having no known associations with these three diseases in existing miRNA-disease databases were directly or indirectly confirmed by our latest literature mining. All these results demonstrated the reliability and efficiency of CHNmiRD, and it is anticipated that CHNmiRD will serve as a powerful bioinformatics method for mining novel disease-related miRNAs and providing a new perspective into molecular mechanisms underlying human diseases at the post-transcriptional level. CHNmiRD is freely available at http://www.bio-bigdata.com/CHNmiRD.

  6. Integration of Multiple Genomic and Phenotype Data to Infer Novel miRNA-Disease Associations.

    Directory of Open Access Journals (Sweden)

    Hongbo Shi

    Full Text Available MicroRNAs (miRNAs play an important role in the development and progression of human diseases. The identification of disease-associated miRNAs will be helpful for understanding the molecular mechanisms of diseases at the post-transcriptional level. Based on different types of genomic data sources, computational methods for miRNA-disease association prediction have been proposed. However, individual source of genomic data tends to be incomplete and noisy; therefore, the integration of various types of genomic data for inferring reliable miRNA-disease associations is urgently needed. In this study, we present a computational framework, CHNmiRD, for identifying miRNA-disease associations by integrating multiple genomic and phenotype data, including protein-protein interaction data, gene ontology data, experimentally verified miRNA-target relationships, disease phenotype information and known miRNA-disease connections. The performance of CHNmiRD was evaluated by experimentally verified miRNA-disease associations, which achieved an area under the ROC curve (AUC of 0.834 for 5-fold cross-validation. In particular, CHNmiRD displayed excellent performance for diseases without any known related miRNAs. The results of case studies for three human diseases (glioblastoma, myocardial infarction and type 1 diabetes showed that all of the top 10 ranked miRNAs having no known associations with these three diseases in existing miRNA-disease databases were directly or indirectly confirmed by our latest literature mining. All these results demonstrated the reliability and efficiency of CHNmiRD, and it is anticipated that CHNmiRD will serve as a powerful bioinformatics method for mining novel disease-related miRNAs and providing a new perspective into molecular mechanisms underlying human diseases at the post-transcriptional level. CHNmiRD is freely available at http://www.bio-bigdata.com/CHNmiRD.

  7. Construction of an integrated genetic linkage map for the A genome of Brassica napus using SSR markers derived from sequenced BACs in B. rapa

    Directory of Open Access Journals (Sweden)

    King Graham J

    2010-10-01

    Full Text Available Abstract Background The Multinational Brassica rapa Genome Sequencing Project (BrGSP has developed valuable genomic resources, including BAC libraries, BAC-end sequences, genetic and physical maps, and seed BAC sequences for Brassica rapa. An integrated linkage map between the amphidiploid B. napus and diploid B. rapa will facilitate the rapid transfer of these valuable resources from B. rapa to B. napus (Oilseed rape, Canola. Results In this study, we identified over 23,000 simple sequence repeats (SSRs from 536 sequenced BACs. 890 SSR markers (designated as BrGMS were developed and used for the construction of an integrated linkage map for the A genome in B. rapa and B. napus. Two hundred and nineteen BrGMS markers were integrated to an existing B. napus linkage map (BnaNZDH. Among these mapped BrGMS markers, 168 were only distributed on the A genome linkage groups (LGs, 18 distrubuted both on the A and C genome LGs, and 33 only distributed on the C genome LGs. Most of the A genome LGs in B. napus were collinear with the homoeologous LGs in B. rapa, although minor inversions or rearrangements occurred on A2 and A9. The mapping of these BAC-specific SSR markers enabled assignment of 161 sequenced B. rapa BACs, as well as the associated BAC contigs to the A genome LGs of B. napus. Conclusion The genetic mapping of SSR markers derived from sequenced BACs in B. rapa enabled direct links to be established between the B. napus linkage map and a B. rapa physical map, and thus the assignment of B. rapa BACs and the associated BAC contigs to the B. napus linkage map. This integrated genetic linkage map will facilitate exploitation of the B. rapa annotated genomic resources for gene tagging and map-based cloning in B. napus, and for comparative analysis of the A genome within Brassica species.

  8. Complete Genome Sequence of Germline Chromosomally Integrated Human Herpesvirus 6A and Analyses Integration Sites Define a New Human Endogenous Virus with Potential to Reactivate as an Emerging Infection.

    Science.gov (United States)

    Tweedy, Joshua; Spyrou, Maria Alexandra; Pearson, Max; Lassner, Dirk; Kuhl, Uwe; Gompels, Ursula A

    2016-01-15

    Human herpesvirus-6A and B (HHV-6A, HHV-6B) have recently defined endogenous genomes, resulting from integration into the germline: chromosomally-integrated "CiHHV-6A/B". These affect approximately 1.0% of human populations, giving potential for virus gene expression in every cell. We previously showed that CiHHV-6A was more divergent than CiHHV-6B by examining four genes in 44 European CiHHV-6A/B cardiac/haematology patients. There was evidence for gene expression/reactivation, implying functional non-defective genomes. To further define the relationship between HHV-6A and CiHHV-6A we used next-generation sequencing to characterize genomes from three CiHHV-6A cardiac patients. Comparisons to known exogenous HHV-6A showed CiHHV-6A genomes formed a separate clade; including all 85 non-interrupted genes and necessary cis-acting signals for reactivation as infectious virus. Greater single nucleotide polymorphism (SNP) density was defined in 16 genes and the direct repeats (DR) terminal regions. Using these SNPs, deep sequencing analyses demonstrated superinfection with exogenous HHV-6A in two of the CiHHV-6A patients with recurrent cardiac disease. Characterisation of the integration sites in twelve patients identified the human chromosome 17p subtelomere as a prevalent site, which had specific repeat structures and phylogenetically related CiHHV-6A coding sequences indicating common ancestral origins. Overall CiHHV-6A genomes were similar, but distinct from known exogenous HHV-6A virus, and have the capacity to reactivate as emerging virus infections.

  9. VarB Plus: An Integrated Tool for Visualization of Genome Variation Datasets

    KAUST Repository

    Hidayah, Lailatul

    2012-07-01

    Research on genomic sequences has been improving significantly as more advanced technology for sequencing has been developed. This opens enormous opportunities for sequence analysis. Various analytical tools have been built for purposes such as sequence assembly, read alignments, genome browsing, comparative genomics, and visualization. From the visualization perspective, there is an increasing trend towards use of large-scale computation. However, more than power is required to produce an informative image. This is a challenge that we address by providing several ways of representing biological data in order to advance the inference endeavors of biologists. This thesis focuses on visualization of variations found in genomic sequences. We develop several visualization functions and embed them in an existing variation visualization tool as extensions. The tool we improved is named VarB, hence the nomenclature for our enhancement is VarB Plus. To the best of our knowledge, besides VarB, there is no tool that provides the capability of dynamic visualization of genome variation datasets as well as statistical analysis. Dynamic visualization allows users to toggle different parameters on and off and see the results on the fly. The statistical analysis includes Fixation Index, Relative Variant Density, and Tajima’s D. Hence we focused our efforts on this tool. The scope of our work includes plots of per-base genome coverage, Principal Coordinate Analysis (PCoA), integration with a read alignment viewer named LookSeq, and visualization of geo-biological data. In addition to description of embedded functionalities, significance, and limitations, future improvements are discussed. The result is four extensions embedded successfully in the original tool, which is built on the Qt framework in C++. Hence it is portable to numerous platforms. Our extensions have shown acceptable execution time in a beta testing with various high-volume published datasets, as well as positive

  10. Phytoextraction potential of poplar (Populus alba L. var. pyramidalis Bunge) from calcareous agricultural soils contaminated by cadmium.

    Science.gov (United States)

    Hu, Yahu; Nan, Zhongren; Jin, Cheng; Wang, Ning; Luo, Huanzhang

    2014-01-01

    To investigate the phytoextraction potential of Populus alba L. var. pyramidalis Bunge for cadmium (Cd) contaminated calcareous soils, a concentration gradient experiment and a field sampling experiment (involving poplars of different ages) were conducted. The translocation factors for all experiments and treatments were greater than 1. The bioconcentration factor decreased from 2.37 to 0.25 with increasing soil Cd concentration in the concentration gradient experiment and generally decreased with stand age under field conditions. The Cd concentrations in P. pyramidalis organs decreased in the order of leaves > stems > roots. The shoot biomass production in the concentration gradient experiment was not significantly reduced with soil Cd concentrations up to or slightly over 50 mg kg(-1). The results show that the phytoextraction efficiency of P. pyramidalis depends on both the soil Cd concentration and the tree age. Populus pyramidalis is most suitable for remediation of slightly Cd contaminated calcareous soils through the combined harvest of stems and leaves under actual field conditions.

  11. Evolution of plant virus movement proteins from the 30K superfamily and of their homologs integrated in plant genomes

    Energy Technology Data Exchange (ETDEWEB)

    Mushegian, Arcady R., E-mail: mushegian2@gmail.com [Division of Molecular and Cellular Biosciences, National Science Foundation, 4201 Wilson Boulevard, Arlington, VA 22230 (United States); Elena, Santiago F., E-mail: sfelena@ibmcp.upv.es [Instituto de Biología Molecular y Celular de Plantas, CSIC-UPV, 46022 València (Spain); The Santa Fe Institute, Santa Fe, NM 87501 (United States)

    2015-02-15

    Homologs of Tobacco mosaic virus 30K cell-to-cell movement protein are encoded by diverse plant viruses. Mechanisms of action and evolutionary origins of these proteins remain obscure. We expand the picture of conservation and evolution of the 30K proteins, producing sequence alignment of the 30K superfamily with the broadest phylogenetic coverage thus far and illuminating structural features of the core all-beta fold of these proteins. Integrated copies of pararetrovirus 30K movement genes are prevalent in euphyllophytes, with at least one copy intact in nearly every examined species, and mRNAs detected for most of them. Sequence analysis suggests repeated integrations, pseudogenizations, and positive selection in those provirus genes. An unannotated 30K-superfamily gene in Arabidopsis thaliana genome is likely expressed as a fusion with the At1g37113 transcript. This molecular background of endopararetrovirus gene products in plants may change our view of virus infection and pathogenesis, and perhaps of cellular homeostasis in the hosts. - Highlights: • Sequence region shared by plant virus “30K” movement proteins has an all-beta fold. • Most euphyllophyte genomes contain integrated copies of pararetroviruses. • These integrated virus genomes often include intact movement protein genes. • Molecular evidence suggests that these “30K” genes may be selected for function.

  12. Genomics Mechanisms of Carbon Allocation and Partitioning in Poplar

    Energy Technology Data Exchange (ETDEWEB)

    Kirst, Matias; Peter, Gary; Martin, Timothy

    2009-07-30

    The genetic control of carbon allocation and partitioning in woody perennial plants is poorly understood despite its importance for carbon sequestration. It is also unclear how environmental cues such as nitrogen availability impact the genes that regulate growth, and biomass allocation and wood composition in trees. To address these questions we phenotyped 396 clonally replicated genotypes of an interspecific pseudo-backcross pedigree of Populus for wood composition and biomass traits in above and below ground organs. The loci that regulate growth, carbon allocation and partitioning under two nitrogen conditions were identified, defining the contribution of environmental cues to their genetic control. Fifty-seven quantitative trait loci (QTL) were identified for twenty traits analyzed. The majority of QTL are specific to one of the two nitrogen treatments, demonstrating significant nitrogen-dependent genetic control. A highly significant genetic correlation was observed between plant growth and lignin/cellulose composition, and QTL co-localization identified the genomic position of potential pleiotropic regulators. Gene expression analysis of all poplar genes was also characterized in differentiating xylem, whole-roots and developing leaves of 192 of the segregating population. By integrating the QTL and gene expression information we identified genes that regulate carbon partitioning and several biomass growth related properties. The work developed in this project resulted in the publication of three book chapters, four scientific articles (three others currently in preparation), 17 presentations in international conferences and two provisional patent applications.

  13. Macro- and micro-nutrient concentration in leaf, woody, and root tissue of Populus irrigated with landfill leachate

    Science.gov (United States)

    Jill A. Zalesny; Ronald S., Jr. Zalesny; Adam H. Wiese; Bart T. Sexton; Richard B. Hall

    2007-01-01

    Landfill leachate offers an opportunity to supply water and plant nutritional benefits at a lower cost than traditional sources. Information about nutrient uptake and distribution into tissues of Populus irrigated with landfill leachate helps increase biomass production along with evaluating the impacts of leachate chemistry on tree health.

  14. Geochemical peculiarities of black poplar leaves (Populus nigra L.) in the sites with heavy metals intensive fallouts

    Science.gov (United States)

    Yalaltdinova, Albina; Baranovskaya, Natalya; Rikhvanov, Leonid; Matveenko, Irina

    2013-04-01

    The article deals with the content of 28 chemical elements in the leaves ash of black poplar (Populus nigra L.) growing in Ust-Kamenogorsk city area. It is the major industrial center of Kazakhstan Republic on the territory where the industrial giants of non-ferrous metallurgy and nuclear energy are situated. Comparative analysis with the similar data obtained from leaves ash of Populus nigra L. in Tomsk, Ekibastuz, and Pavlodar cities has revealed that in comparison with other urban areas, leaves ash of black poplar (Populus nigra L.) from Ust-Kamenogorsk city is characterized by elevated concentration rates of Ta, U, Zn, Ag, As, Sb, Br, Sr and Na. Within the city, the sites and areas with abnormal contents of typomorphic pollutants have been revealed. In the central part of the city, in the vicinity of lead-zinc plant and Ulba metallurgical plant, the highest concentrations of Ta, U, Zn, Ag, Au, As, Sb, Cr and Fe were marked. In the northeast, where the titanium-magnesium plant is located, elevated concentrations of Br and Sr were stated. Thus, the impact of major city enterprises which are the main sources of heavy metals is reflected in the element composition. Zn, As, Sb, Ag and Au comes from lead-zinc plant and its refinery plants, while Ulba metallurgical plant can be considered source of Ta and U in the environment, producing tantalum and fuel pellets for nuclear power plants. These companies, due to the current objective circumstances, are located in the central part of the city, have a significant negative effect on the environment and form the risk factors for human health.

  15. Comparative physiology of allopatric Populus species: Geographic clines in photosynthesis, height growth and carbon isotope discrimination in common gardens

    Directory of Open Access Journals (Sweden)

    Raju Yaranna Soolanayakanahally

    2015-07-01

    Full Text Available Populus species with wide geographic ranges display strong adaptation to local environments. We studied the clinal patterns in phenology and ecophysiology in allopatric Populus species adapted to similar environments on different continents under common garden settings. As a result of climatic adaptation, both P. tremula L. and Populus balsamifera L. display latitudinal clines in photosynthetic rates (A, whereby high-latitude trees of P. tremula had higher A compared to low-latitude trees and nearly so in P. balsamifera (p = 0.06. Stomatal conductance (gs and chlorophyll content index (CCI follow similar latitudinal trends. However, foliar nitrogen was positively correlated with latitude in P. balsamifera and negatively correlated in P. tremula. No significant trends in carbon isotope composition of the leaf tissue (δ13C were observed for both species; but, intrinsic water-use efficiency (WUEi was negatively correlated with the latitude of origin in P. balsamifera. In spite of intrinsically higher A, high-latitude trees in both common gardens accomplished less height gain as a result of early bud set. Thus, shoot biomass was determined by height elongation duration (HED, which was well approximated by the number of days available for free growth between bud flush and bud set. In doing so, we highlight the shortcoming of unreplicated outdoor common gardens for tree improvement and the crucial role of photoperiod in limiting height growth, further complicating interpretation of other secondary effects.

  16. Complete Genome Sequence of Germline Chromosomally Integrated Human Herpesvirus 6A and Analyses Integration Sites Define a New Human Endogenous Virus with Potential to Reactivate as an Emerging Infection

    Science.gov (United States)

    Tweedy, Joshua; Spyrou, Maria Alexandra; Pearson, Max; Lassner, Dirk; Kuhl, Uwe; Gompels, Ursula A.

    2016-01-01

    Human herpesvirus-6A and B (HHV-6A, HHV-6B) have recently defined endogenous genomes, resulting from integration into the germline: chromosomally-integrated “CiHHV-6A/B”. These affect approximately 1.0% of human populations, giving potential for virus gene expression in every cell. We previously showed that CiHHV-6A was more divergent than CiHHV-6B by examining four genes in 44 European CiHHV-6A/B cardiac/haematology patients. There was evidence for gene expression/reactivation, implying functional non-defective genomes. To further define the relationship between HHV-6A and CiHHV-6A we used next-generation sequencing to characterize genomes from three CiHHV-6A cardiac patients. Comparisons to known exogenous HHV-6A showed CiHHV-6A genomes formed a separate clade; including all 85 non-interrupted genes and necessary cis-acting signals for reactivation as infectious virus. Greater single nucleotide polymorphism (SNP) density was defined in 16 genes and the direct repeats (DR) terminal regions. Using these SNPs, deep sequencing analyses demonstrated superinfection with exogenous HHV-6A in two of the CiHHV-6A patients with recurrent cardiac disease. Characterisation of the integration sites in twelve patients identified the human chromosome 17p subtelomere as a prevalent site, which had specific repeat structures and phylogenetically related CiHHV-6A coding sequences indicating common ancestral origins. Overall CiHHV-6A genomes were similar, but distinct from known exogenous HHV-6A virus, and have the capacity to reactivate as emerging virus infections. PMID:26784220

  17. A robust network of double-strand break repair pathways governs genome integrity during C. elegans development.

    NARCIS (Netherlands)

    Pontier, D.B.; Tijsterman, M.

    2009-01-01

    To preserve genomic integrity, various mechanisms have evolved to repair DNA double-strand breaks (DSBs). Depending on cell type or cell cycle phase, DSBs can be repaired error-free, by homologous recombination, or with concomitant loss of sequence information, via nonhomologous end-joining (NHEJ)

  18. metabolicMine: an integrated genomics, genetics and proteomics data warehouse for common metabolic disease research.

    Science.gov (United States)

    Lyne, Mike; Smith, Richard N; Lyne, Rachel; Aleksic, Jelena; Hu, Fengyuan; Kalderimis, Alex; Stepan, Radek; Micklem, Gos

    2013-01-01

    Common metabolic and endocrine diseases such as diabetes affect millions of people worldwide and have a major health impact, frequently leading to complications and mortality. In a search for better prevention and treatment, there is ongoing research into the underlying molecular and genetic bases of these complex human diseases, as well as into the links with risk factors such as obesity. Although an increasing number of relevant genomic and proteomic data sets have become available, the quantity and diversity of the data make their efficient exploitation challenging. Here, we present metabolicMine, a data warehouse with a specific focus on the genomics, genetics and proteomics of common metabolic diseases. Developed in collaboration with leading UK metabolic disease groups, metabolicMine integrates data sets from a range of experiments and model organisms alongside tools for exploring them. The current version brings together information covering genes, proteins, orthologues, interactions, gene expression, pathways, ontologies, diseases, genome-wide association studies and single nucleotide polymorphisms. Although the emphasis is on human data, key data sets from mouse and rat are included. These are complemented by interoperation with the RatMine rat genomics database, with a corresponding mouse version under development by the Mouse Genome Informatics (MGI) group. The web interface contains a number of features including keyword search, a library of Search Forms, the QueryBuilder and list analysis tools. This provides researchers with many different ways to analyse, view and flexibly export data. Programming interfaces and automatic code generation in several languages are supported, and many of the features of the web interface are available through web services. The combination of diverse data sets integrated with analysis tools and a powerful query system makes metabolicMine a valuable research resource. The web interface makes it accessible to first

  19. Genome Modeling System: A Knowledge Management Platform for Genomics.

    Directory of Open Access Journals (Sweden)

    Malachi Griffith

    2015-07-01

    Full Text Available In this work, we present the Genome Modeling System (GMS, an analysis information management system capable of executing automated genome analysis pipelines at a massive scale. The GMS framework provides detailed tracking of samples and data coupled with reliable and repeatable analysis pipelines. The GMS also serves as a platform for bioinformatics development, allowing a large team to collaborate on data analysis, or an individual researcher to leverage the work of others effectively within its data management system. Rather than separating ad-hoc analysis from rigorous, reproducible pipelines, the GMS promotes systematic integration between the two. As a demonstration of the GMS, we performed an integrated analysis of whole genome, exome and transcriptome sequencing data from a breast cancer cell line (HCC1395 and matched lymphoblastoid line (HCC1395BL. These data are available for users to test the software, complete tutorials and develop novel GMS pipeline configurations. The GMS is available at https://github.com/genome/gms.

  20. Statistical Viewer: a tool to upload and integrate linkage and association data as plots displayed within the Ensembl genome browser

    Directory of Open Access Journals (Sweden)

    Hauser Elizabeth R

    2005-04-01

    Full Text Available Abstract Background To facilitate efficient selection and the prioritization of candidate complex disease susceptibility genes for association analysis, increasingly comprehensive annotation tools are essential to integrate, visualize and analyze vast quantities of disparate data generated by genomic screens, public human genome sequence annotation and ancillary biological databases. We have developed a plug-in package for Ensembl called "Statistical Viewer" that facilitates the analysis of genomic features and annotation in the regions of interest defined by linkage analysis. Results Statistical Viewer is an add-on package to the open-source Ensembl Genome Browser and Annotation System that displays disease study-specific linkage and/or association data as 2 dimensional plots in new panels in the context of Ensembl's Contig View and Cyto View pages. An enhanced upload server facilitates the upload of statistical data, as well as additional feature annotation to be displayed in DAS tracts, in the form of Excel Files. The Statistical View panel, drawn directly under the ideogram, illustrates lod score values for markers from a study of interest that are plotted against their position in base pairs. A module called "Get Map" easily converts the genetic locations of markers to genomic coordinates. The graph is placed under the corresponding ideogram features a synchronized vertical sliding selection box that is seamlessly integrated into Ensembl's Contig- and Cyto- View pages to choose the region to be displayed in Ensembl's "Overview" and "Detailed View" panels. To resolve Association and Fine mapping data plots, a "Detailed Statistic View" plot corresponding to the "Detailed View" may be displayed underneath. Conclusion Features mapping to regions of linkage are accentuated when Statistic View is used in conjunction with the Distributed Annotation System (DAS to display supplemental laboratory information such as differentially expressed disease

  1. Increased leaf area dominates carbon flux response to elevated CO2 in stands of Populus deltoides (Bartr.)

    Science.gov (United States)

    Ramesh Murthy; Greg Barron-Gafford; Philip M. Dougherty; Victor c. Engels; Katie Grieve; Linda Handley; Christie Klimas; Mark J. Postosnaks; Stanley J. Zarnoch; Jianwei Zhang

    2005-01-01

    We examined the effects of atmospheric vapor pressure deficit (VPD) and soil moisture stress (SMS) on leaf- and stand-level CO2 exchange in model 3-year-old coppiced cottonwood (Populus deltoides Bartr.) plantations using the large-scale, controlled environments of the Biosphere 2 Laboratory. A short-term experiment was imposed...

  2. Abnormal meiosis in an intersectional allotriploid of Populus L. and segregation of ploidy levels in 2x × 3x progeny.

    Directory of Open Access Journals (Sweden)

    Jun Wang

    Full Text Available Triploid plants are usually highly aborted owing to unbalanced meiotic chromosome segregation, but limited viable gametes can participate in the transition to different ploidy levels. In this study, numerous meiotic abnormalities were found with high frequency in an intersectional allotriploid poplar (Populus alba × P. berolinensis 'Yinzhong', including univalents, precocious chromosome migration, lagging chromosomes, chromosome bridges, micronuclei, and precocious cytokinesis, indicating high genetic imbalance in this allotriploid. Some micronuclei trigger mini-spindle formation in metaphase II and participate in cytokinesis to form polyads with microcytes. Unbalanced chromosome segregation and chromosome elimination resulted in the formation of microspores with aneuploid chromosome sets. Fusion of sister nuclei occurs in microsporocytes with precocious cytokinesis, which could form second meiotic division restitution (SDR-type gametes. However, SDR-type gametes likely contain incomplete chromosome sets due to unbalanced segregation of homologous chromosomes during the first meiotic division in triploids. Misorientation of spindles during the second meiotic division, such as fused and tripolar spindles with low frequency, could result in the formation of first meiotic division restitution (FDR-type unreduced gametes, which most likely contain three complete chromosome sets. Although 'Yinzhong' yields 88.7% stainable pollen grains with wide diameter variation from 23.9 to 61.3 μm, the pollen viability is poor (2.78% ± 0.38. A cross of 'Yinzhong' pollen with a diploid female clone produced progeny with extensive segregation of ploidy levels, including 29 diploids, 18 triploids, 4 tetraploids, and 48 aneuploids, suggesting the formation of viable aneuploidy and unreduced pollen in 'Yinzhong'. Individuals with different chromosome compositions are potential to analyze chromosomal function and to integrate the chromosomal dosage variation into

  3. Microenvironmental Heterogeneity Parallels Breast Cancer Progression: A Histology-Genomic Integration Analysis.

    Directory of Open Access Journals (Sweden)

    Rachael Natrajan

    2016-02-01

    Full Text Available The intra-tumor diversity of cancer cells is under intense investigation; however, little is known about the heterogeneity of the tumor microenvironment that is key to cancer progression and evolution. We aimed to assess the degree of microenvironmental heterogeneity in breast cancer and correlate this with genomic and clinical parameters.We developed a quantitative measure of microenvironmental heterogeneity along three spatial dimensions (3-D in solid tumors, termed the tumor ecosystem diversity index (EDI, using fully automated histology image analysis coupled with statistical measures commonly used in ecology. This measure was compared with disease-specific survival, key mutations, genome-wide copy number, and expression profiling data in a retrospective study of 510 breast cancer patients as a test set and 516 breast cancer patients as an independent validation set. In high-grade (grade 3 breast cancers, we uncovered a striking link between high microenvironmental heterogeneity measured by EDI and a poor prognosis that cannot be explained by tumor size, genomics, or any other data types. However, this association was not observed in low-grade (grade 1 and 2 breast cancers. The prognostic value of EDI was superior to known prognostic factors and was enhanced with the addition of TP53 mutation status (multivariate analysis test set, p = 9 × 10-4, hazard ratio = 1.47, 95% CI 1.17-1.84; validation set, p = 0.0011, hazard ratio = 1.78, 95% CI 1.26-2.52. Integration with genome-wide profiling data identified losses of specific genes on 4p14 and 5q13 that were enriched in grade 3 tumors with high microenvironmental diversity that also substratified patients into poor prognostic groups. Limitations of this study include the number of cell types included in the model, that EDI has prognostic value only in grade 3 tumors, and that our spatial heterogeneity measure was dependent on spatial scale and tumor size.To our knowledge, this is the first

  4. A statistical framework to predict functional non-coding regions in the human genome through integrated analysis of annotation data.

    Science.gov (United States)

    Lu, Qiongshi; Hu, Yiming; Sun, Jiehuan; Cheng, Yuwei; Cheung, Kei-Hoi; Zhao, Hongyu

    2015-05-27

    Identifying functional regions in the human genome is a major goal in human genetics. Great efforts have been made to functionally annotate the human genome either through computational predictions, such as genomic conservation, or high-throughput experiments, such as the ENCODE project. These efforts have resulted in a rich collection of functional annotation data of diverse types that need to be jointly analyzed for integrated interpretation and annotation. Here we present GenoCanyon, a whole-genome annotation method that performs unsupervised statistical learning using 22 computational and experimental annotations thereby inferring the functional potential of each position in the human genome. With GenoCanyon, we are able to predict many of the known functional regions. The ability of predicting functional regions as well as its generalizable statistical framework makes GenoCanyon a unique and powerful tool for whole-genome annotation. The GenoCanyon web server is available at http://genocanyon.med.yale.edu.

  5. The catfish genome database cBARBEL: an informatic platform for genome biology of ictalurid catfish.

    Science.gov (United States)

    Lu, Jianguo; Peatman, Eric; Yang, Qing; Wang, Shaolin; Hu, Zhiliang; Reecy, James; Kucuktas, Huseyin; Liu, Zhanjiang

    2011-01-01

    The catfish genome database, cBARBEL (abbreviated from catfish Breeder And Researcher Bioinformatics Entry Location) is an online open-access database for genome biology of ictalurid catfish (Ictalurus spp.). It serves as a comprehensive, integrative platform for all aspects of catfish genetics, genomics and related data resources. cBARBEL provides BLAST-based, fuzzy and specific search functions, visualization of catfish linkage, physical and integrated maps, a catfish EST contig viewer with SNP information overlay, and GBrowse-based organization of catfish genomic data based on sequence similarity with zebrafish chromosomes. Subsections of the database are tightly related, allowing a user with a sequence or search string of interest to navigate seamlessly from one area to another. As catfish genome sequencing proceeds and ongoing quantitative trait loci (QTL) projects bear fruit, cBARBEL will allow rapid data integration and dissemination within the catfish research community and to interested stakeholders. cBARBEL can be accessed at http://catfishgenome.org.

  6. The openness of pluripotent epigenome - Defining the genomic integrity of stemness for regenerative medicine

    Directory of Open Access Journals (Sweden)

    Xuejun H Parsons

    2014-02-01

    Full Text Available This article is an editorial, and it doesn't include an abstract. Full text of this article is available in HTML and PDF.Cite this article as: Parsons XH. The openness of pluripotent epigenome - Defining the genomic Integrity of stemness for regenerative medicine. Int J Cancer Ther Oncol 2014; 2(1:020114.DOI: http://dx.doi.org/10.14319/ijcto.0201.14

  7. FISH Oracle 2: a web server for integrative visualization of genomic data in cancer research.

    Science.gov (United States)

    Mader, Malte; Simon, Ronald; Kurtz, Stefan

    2014-03-31

    A comprehensive view on all relevant genomic data is instrumental for understanding the complex patterns of molecular alterations typically found in cancer cells. One of the most effective ways to rapidly obtain an overview of genomic alterations in large amounts of genomic data is the integrative visualization of genomic events. We developed FISH Oracle 2, a web server for the interactive visualization of different kinds of downstream processed genomics data typically available in cancer research. A powerful search interface and a fast visualization engine provide a highly interactive visualization for such data. High quality image export enables the life scientist to easily communicate their results. A comprehensive data administration allows to keep track of the available data sets. We applied FISH Oracle 2 to published data and found evidence that, in colorectal cancer cells, the gene TTC28 may be inactivated in two different ways, a fact that has not been published before. The interactive nature of FISH Oracle 2 and the possibility to store, select and visualize large amounts of downstream processed data support life scientists in generating hypotheses. The export of high quality images supports explanatory data visualization, simplifying the communication of new biological findings. A FISH Oracle 2 demo server and the software is available at http://www.zbh.uni-hamburg.de/fishoracle.

  8. Ethylene and jasmonic acid act as negative modulators during mutualistic symbiosis between Laccaria bicolor and Populus roots.

    Science.gov (United States)

    Plett, Jonathan M; Khachane, Amit; Ouassou, Malika; Sundberg, Björn; Kohler, Annegret; Martin, Francis

    2014-04-01

    The plant hormones ethylene, jasmonic acid and salicylic acid have interconnecting roles during the response of plant tissues to mutualistic and pathogenic symbionts. We used morphological studies of transgenic- or hormone-treated Populus roots as well as whole-genome oligoarrays to examine how these hormones affect root colonization by the mutualistic ectomycorrhizal fungus Laccaria bicolor S238N. We found that genes regulated by ethylene, jasmonic acid and salicylic acid were regulated in the late stages of the interaction between L. bicolor and poplar. Both ethylene and jasmonic acid treatments were found to impede fungal colonization of roots, and this effect was correlated to an increase in the expression of certain transcription factors (e.g. ETHYLENE RESPONSE FACTOR1) and a decrease in the expression of genes associated with microbial perception and cell wall modification. Further, we found that ethylene and jasmonic acid showed extensive transcriptional cross-talk, cross-talk that was opposed by salicylic acid signaling. We conclude that ethylene and jasmonic acid pathways are induced late in the colonization of root tissues in order to limit fungal growth within roots. This induction is probably an adaptive response by the plant such that its growth and vigor are not compromised by the fungus. © 2013 The Authors New Phytologist © 2013 New Phytologist Trust.

  9. Date of shoot collection, genotype, and original shoot position affect early rooting of dormant hardwood cuttings of Populus

    Science.gov (United States)

    R. S., Jr. Zalesny; A.H. Wiese

    2006-01-01

    Identifying superior combinations among date of dormant- season shoot collection, genotype, and original shoot position can increase the rooting potential of Populus cuttings. Thus, the objectives of our study were to: 1) evaluate variation among clones in early rooting from hardwood cuttings processed every three weeks from shoots collected...

  10. Toward allotetraploid cotton genome assembly: integration of a high-density molecular genetic linkage map with DNA sequence information

    Science.gov (United States)

    2012-01-01

    Background Cotton is the world’s most important natural textile fiber and a significant oilseed crop. Decoding cotton genomes will provide the ultimate reference and resource for research and utilization of the species. Integration of high-density genetic maps with genomic sequence information will largely accelerate the process of whole-genome assembly in cotton. Results In this paper, we update a high-density interspecific genetic linkage map of allotetraploid cultivated cotton. An additional 1,167 marker loci have been added to our previously published map of 2,247 loci. Three new marker types, InDel (insertion-deletion) and SNP (single nucleotide polymorphism) developed from gene information, and REMAP (retrotransposon-microsatellite amplified polymorphism), were used to increase map density. The updated map consists of 3,414 loci in 26 linkage groups covering 3,667.62 cM with an average inter-locus distance of 1.08 cM. Furthermore, genome-wide sequence analysis was finished using 3,324 informative sequence-based markers and publicly-available Gossypium DNA sequence information. A total of 413,113 EST and 195 BAC sequences were physically anchored and clustered by 3,324 sequence-based markers. Of these, 14,243 ESTs and 188 BACs from different species of Gossypium were clustered and specifically anchored to the high-density genetic map. A total of 2,748 candidate unigenes from 2,111 ESTs clusters and 63 BACs were mined for functional annotation and classification. The 337 ESTs/genes related to fiber quality traits were integrated with 132 previously reported cotton fiber quality quantitative trait loci, which demonstrated the important roles in fiber quality of these genes. Higher-level sequence conservation between different cotton species and between the A- and D-subgenomes in tetraploid cotton was found, indicating a common evolutionary origin for orthologous and paralogous loci in Gossypium. Conclusion This study will serve as a valuable genomic resource

  11. Alternative Splicing Studies of the Reactive Oxygen Species Gene Network in Populus Reveal Two Isoforms of High-Isoelectric-Point Superoxide Dismutase1[C][W

    Science.gov (United States)

    Srivastava, Vaibhav; Srivastava, Manoj Kumar; Chibani, Kamel; Nilsson, Robert; Rouhier, Nicolas; Melzer, Michael; Wingsle, Gunnar

    2009-01-01

    Recent evidence has shown that alternative splicing (AS) is widely involved in the regulation of gene expression, substantially extending the diversity of numerous proteins. In this study, a subset of expressed sequence tags representing members of the reactive oxygen species gene network was selected from the PopulusDB database to investigate AS mechanisms in Populus. Examples of all known types of AS were detected, but intron retention was the most common. Interestingly, the closest Arabidopsis (Arabidopsis thaliana) homologs of half of the AS genes identified in Populus are not reportedly alternatively spliced. Two genes encoding the protein of most interest in our study (high-isoelectric-point superoxide dismutase [hipI-SOD]) have been found in black cottonwood (Populus trichocarpa), designated PthipI-SODC1 and PthipI-SODC2. Analysis of the expressed sequence tag libraries has indicated the presence of two transcripts of PthipI-SODC1 (hipI-SODC1b and hipI-SODC1s). Alignment of these sequences with the PthipI-SODC1 gene showed that hipI-SODC1b was 69 bp longer than hipI-SODC1s due to an AS event involving the use of an alternative donor splice site in the sixth intron. Transcript analysis showed that the splice variant hipI-SODC1b was differentially expressed, being clearly expressed in cambial and xylem, but not phloem, regions. In addition, immunolocalization and mass spectrometric data confirmed the presence of hipI-SOD proteins in vascular tissue. The functionalities of the spliced gene products were assessed by expressing recombinant hipI-SOD proteins and in vitro SOD activity assays. PMID:19176719

  12. Sex-specific responses to winter flooding, spring waterlogging and post-flooding recovery in Populus deltoides

    OpenAIRE

    Ling-Feng Miao; Fan Yang; Chun-Yu Han; Yu-Jin Pu; Yang Ding; Li-Jia Zhang

    2017-01-01

    Winter flooding events are common in some rivers and streams due to dam constructions, and flooding and waterlogging inhibit the growth of trees in riparian zones. This study investigated sex-specific morphological, physiological and ultrastructural responses to various durations of winter flooding and spring waterlogging stresses, and post-flooding recovery characteristics in Populus deltoides. There were no significant differences in the morphological, ultrastructural and the majority of ph...

  13. Inter-replicon Gene Flow Contributes to Transcriptional Integration in the Sinorhizobium meliloti Multipartite Genome.

    Science.gov (United States)

    diCenzo, George C; Wellappili, Deelaka; Golding, G Brian; Finan, Turlough M

    2018-05-04

    Integration of newly acquired genes into existing regulatory networks is necessary for successful horizontal gene transfer (HGT). Ten percent of bacterial species contain at least two DNA replicons over 300 kilobases in size, with the secondary replicons derived predominately through HGT. The Sinorhizobium meliloti genome is split between a 3.7 Mb chromosome, a 1.7 Mb chromid consisting largely of genes acquired through ancient HGT, and a 1.4 Mb megaplasmid consisting primarily of recently acquired genes. Here, RNA-sequencing is used to examine the transcriptional consequences of massive, synthetic genome reduction produced through the removal of the megaplasmid and/or the chromid. Removal of the pSymA megaplasmid influenced the transcription of only six genes. In contrast, removal of the chromid influenced expression of ∼8% of chromosomal genes and ∼4% of megaplasmid genes. This was mediated in part by the loss of the ETR DNA region whose presence on pSymB is due to a translocation from the chromosome. No obvious functional bias among the up-regulated genes was detected, although genes with putative homologs on the chromid were enriched. Down-regulated genes were enriched in motility and sensory transduction pathways. Four transcripts were examined further, and in each case the transcriptional change could be traced to loss of specific pSymB regions. In particularly, a chromosomal transporter was induced due to deletion of bdhA likely mediated through 3-hydroxybutyrate accumulation. These data provide new insights into the evolution of the multipartite bacterial genome, and more generally into the integration of horizontally acquired genes into the transcriptome. Copyright © 2018 diCenzo, et al.

  14. Inter-replicon Gene Flow Contributes to Transcriptional Integration in the Sinorhizobium meliloti Multipartite Genome

    Directory of Open Access Journals (Sweden)

    George C. diCenzo

    2018-05-01

    Full Text Available Integration of newly acquired genes into existing regulatory networks is necessary for successful horizontal gene transfer (HGT. Ten percent of bacterial species contain at least two DNA replicons over 300 kilobases in size, with the secondary replicons derived predominately through HGT. The Sinorhizobium meliloti genome is split between a 3.7 Mb chromosome, a 1.7 Mb chromid consisting largely of genes acquired through ancient HGT, and a 1.4 Mb megaplasmid consisting primarily of recently acquired genes. Here, RNA-sequencing is used to examine the transcriptional consequences of massive, synthetic genome reduction produced through the removal of the megaplasmid and/or the chromid. Removal of the pSymA megaplasmid influenced the transcription of only six genes. In contrast, removal of the chromid influenced expression of ∼8% of chromosomal genes and ∼4% of megaplasmid genes. This was mediated in part by the loss of the ETR DNA region whose presence on pSymB is due to a translocation from the chromosome. No obvious functional bias among the up-regulated genes was detected, although genes with putative homologs on the chromid were enriched. Down-regulated genes were enriched in motility and sensory transduction pathways. Four transcripts were examined further, and in each case the transcriptional change could be traced to loss of specific pSymB regions. In particularly, a chromosomal transporter was induced due to deletion of bdhA likely mediated through 3-hydroxybutyrate accumulation. These data provide new insights into the evolution of the multipartite bacterial genome, and more generally into the integration of horizontally acquired genes into the transcriptome.

  15. Expression Patterns of ERF Genes Underlying Abiotic Stresses in Di-Haploid Populus simonii × P. nigra

    Directory of Open Access Journals (Sweden)

    Shengji Wang

    2014-01-01

    Full Text Available 176 ERF genes from Populus were identified by bioinformatics analysis, 13 of these in di-haploid Populus simonii × P. nigra were investigate by real-time RT-PCR, the results demonstrated that 13 ERF genes were highly responsive to salt stress, drought stress and ABA treatment, and all were expressed in root, stem, and leaf tissues, whereas their expression levels were markedly different in the various tissues. In roots, PthERF99, 110, 119, and 168 were primarily downregulated under drought and ABA treatment but were specifically upregulated under high salt condition. Interestingly, in poplar stems, all ERF genes showed the similar trends in expression in response to NaCl stress, drought stress, and ABA treatment, indicating that they may not play either specific or unique roles in stems in abiotic stress responses. In poplar leaves, PthERF168 was highly induced by ABA treatment, but was suppressed by high salinity and drought stresses, implying that PthERF168 participated in the ABA signaling pathway. The results of this study indicated that ERF genes could play essential but distinct roles in various plant tissues in response to different environment cues and hormonal treatment.

  16. Effects of light regime and IBA concentration on adventitious rooting of an eastern cottonwood (Populus deltoides) clone

    Science.gov (United States)

    Alexander P. Hoffman; Joshua P. Adams; Andrew Nelson

    2016-01-01

    Eastern cottonwood (Populus deltoides) has received a substantial amount of interest from invitro studies within the past decade. The ability to efficiently multiply the stock of established clones such as clone 110412 is a valuable asset for forest endeavors. However, a common problem encountered is initiating adventitious rooting in new micropropagation protocols....

  17. Integrated genomic and interfacility patient-transfer data reveal the transmission pathways of multidrug-resistant Klebsiella pneumoniae in a regional outbreak.

    Science.gov (United States)

    Snitkin, Evan S; Won, Sarah; Pirani, Ali; Lapp, Zena; Weinstein, Robert A; Lolans, Karen; Hayden, Mary K

    2017-11-22

    Development of effective strategies to limit the proliferation of multidrug-resistant organisms requires a thorough understanding of how such organisms spread among health care facilities. We sought to uncover the chains of transmission underlying a 2008 U.S. regional outbreak of carbapenem-resistant Klebsiella pneumoniae by performing an integrated analysis of genomic and interfacility patient-transfer data. Genomic analysis yielded a high-resolution transmission network that assigned directionality to regional transmission events and discriminated between intra- and interfacility transmission when epidemiologic data were ambiguous or misleading. Examining the genomic transmission network in the context of interfacility patient transfers (patient-sharing networks) supported the role of patient transfers in driving the outbreak, with genomic analysis revealing that a small subset of patient-transfer events was sufficient to explain regional spread. Further integration of the genomic and patient-sharing networks identified one nursing home as an important bridge facility early in the outbreak-a role that was not apparent from analysis of genomic or patient-transfer data alone. Last, we found that when simulating a real-time regional outbreak, our methodology was able to accurately infer the facility at which patients acquired their infections. This approach has the potential to identify facilities with high rates of intra- or interfacility transmission, data that will be useful for triggering targeted interventions to prevent further spread of multidrug-resistant organisms. Copyright © 2017 The Authors, some rights reserved; exclusive licensee American Association for the Advancement of Science. No claim to original U.S. Government Works.

  18. Ex situ growth and biomass of Populus bioenergy crops irrigated and fertilized with landfill leachate

    International Nuclear Information System (INIS)

    Zalesny, Ronald S.; Wiese, Adam H.; Bauer, Edmund O.; Riemenschneider, Donald E.

    2009-01-01

    Merging traditional intensive forestry with waste management offers dual goals of fiber and bioenergy production, along with environmental benefits such as soil/water remediation and carbon sequestration. As part of an ongoing effort to acquire data about initial genotypic performance, we evaluated: (1) the early aboveground growth of trees belonging to currently utilized Populus genotypes subjected to irrigation with municipal solid waste landfill leachate or non-fertilized well water (control), and (2) the above- and below-ground biomass of the trees after 70 days of growth. We determined height, diameter, and number of leaves at 28, 42, 56, and 70 days after planting (DAP), along with stem, leaf, and root dry mass by testing six Populus clones (DN34, DN5, I4551, NC14104, NM2, NM6) grown in a greenhouse in a split-split plot, repeated measures design with two blocks, two treatments (whole-plots), six clones (sub-plots), and four sampling dates (sub-sub-plots, repeated measure). Treatments (leachate, water) were applied every other day beginning 42 DAP. The leachate-treated trees exhibited greater height, diameter, and number of leaves at 56 and 70 DAP (P 0.05). Overall, genotypic responses to the leachate treatment were clone-specific for all traits

  19. Integrative computational approach for genome-based study of microbial lipid-degrading enzymes.

    Science.gov (United States)

    Vorapreeda, Tayvich; Thammarongtham, Chinae; Laoteng, Kobkul

    2016-07-01

    Lipid-degrading or lipolytic enzymes have gained enormous attention in academic and industrial sectors. Several efforts are underway to discover new lipase enzymes from a variety of microorganisms with particular catalytic properties to be used for extensive applications. In addition, various tools and strategies have been implemented to unravel the functional relevance of the versatile lipid-degrading enzymes for special purposes. This review highlights the study of microbial lipid-degrading enzymes through an integrative computational approach. The identification of putative lipase genes from microbial genomes and metagenomic libraries using homology-based mining is discussed, with an emphasis on sequence analysis of conserved motifs and enzyme topology. Molecular modelling of three-dimensional structure on the basis of sequence similarity is shown to be a potential approach for exploring the structural and functional relationships of candidate lipase enzymes. The perspectives on a discriminative framework of cutting-edge tools and technologies, including bioinformatics, computational biology, functional genomics and functional proteomics, intended to facilitate rapid progress in understanding lipolysis mechanism and to discover novel lipid-degrading enzymes of microorganisms are discussed.

  20. Genome Variation Map: a data repository of genome variations in BIG Data Center

    OpenAIRE

    Song, Shuhui; Tian, Dongmei; Li, Cuiping; Tang, Bixia; Dong, Lili; Xiao, Jingfa; Bao, Yiming; Zhao, Wenming; He, Hang; Zhang, Zhang

    2017-01-01

    Abstract The Genome Variation Map (GVM; http://bigd.big.ac.cn/gvm/) is a public data repository of genome variations. As a core resource in the BIG Data Center, Beijing Institute of Genomics, Chinese Academy of Sciences, GVM dedicates to collect, integrate and visualize genome variations for a wide range of species, accepts submissions of different types of genome variations from all over the world and provides free open access to all publicly available data in support of worldwide research a...

  1. DNA double-strand break response in stem cells: mechanisms to maintain genomic integrity.

    Science.gov (United States)

    Nagaria, Pratik; Robert, Carine; Rassool, Feyruz V

    2013-02-01

    Embryonic stem cells (ESCs) represent the point of origin of all cells in a given organism and must protect their genomes from both endogenous and exogenous genotoxic stress. DNA double-strand breaks (DSBs) are one of the most lethal forms of damage, and failure to adequately repair DSBs would not only compromise the ability of SCs to self-renew and differentiate, but will also lead to genomic instability and disease. Herein, we describe the mechanisms by which ESCs respond to DSB-inducing agents such as reactive oxygen species (ROS) and ionizing radiation, compared to somatic cells. We will also discuss whether the DSB response is fully reprogrammed in induced pluripotent stem cells (iPSCs) and the role of the DNA damage response (DDR) in the reprogramming of these cells. ESCs have distinct mechanisms to protect themselves against DSBs and oxidative stress compared to somatic cells. The response to damage and stress is crucial for the maintenance of self-renewal and differentiation capacity in SCs. iPSCs appear to reprogram some of the responses to genotoxic stress. However, it remains to be determined if iPSCs also retain some DDR characteristics of the somatic cells of origin. The mechanisms regulating the genomic integrity in ESCs and iPSCs are critical for its safe use in regenerative medicine and may shed light on the pathways and factors that maintain genomic stability, preventing diseases such as cancer. This article is part of a Special Issue entitled Biochemistry of Stem Cells. Copyright © 2012 Elsevier B.V. All rights reserved.

  2. Integrative Genomic and Proteomic Analysis of the Response of Lactobacillus casei Zhang to Glucose Restriction.

    Science.gov (United States)

    Yu, Jie; Hui, Wenyan; Cao, Chenxia; Pan, Lin; Zhang, Heping; Zhang, Wenyi

    2018-03-02

    Nutrient starvation is an important survival challenge for bacteria during industrial production of functional foods. As next-generation sequencing technology has greatly advanced, we performed proteomic and genomic analysis to investigate the response of Lactobacillus casei Zhang to a glucose-restricted environment. L. casei Zhang strains were permitted to evolve in glucose-restricted or normal medium from a common ancestor over a 3 year period, and they were sampled at 1000, 2000, 3000, 4000, 5000, 6000, 7000, and 8000 generations and subjected to proteomic and genomic analyses. Genomic resequencing data revealed different point mutations and other mutational events in each selected generation of L. casei Zhang under glucose restriction stress. The differentially expressed proteins induced by glucose restriction were mostly related to fructose and mannose metabolism, carbohydrate metabolic processes, lyase activity, and amino-acid-transporting ATPase activity. Integrative proteomic and genomic analysis revealed that the mutations protected L. casei Zhang against glucose starvation by regulating other cellular carbohydrate, fatty acid, and amino acid catabolism; phosphoenolpyruvate system pathway activation; glycogen synthesis; ATP consumption; pyruvate metabolism; and general stress-response protein expression. The results help reveal the mechanisms of adapting to glucose starvation and provide new strategies for enhancing the industrial utility of L. casei Zhang.

  3. Integrative Genomics Reveals Mechanisms of Copy Number Alterations Responsible for Transcriptional Deregulation in Colorectal Cancer

    Science.gov (United States)

    Camps, Jordi; Nguyen, Quang Tri; Padilla-Nash, Hesed M.; Knutsen, Turid; McNeil, Nicole E.; Wangsa, Danny; Hummon, Amanda B.; Grade, Marian; Ried, Thomas; Difilippantonio, Michael J.

    2016-01-01

    To evaluate the mechanisms and consequences of chromosomal aberrations in colorectal cancer (CRC), we used a combination of spectral karyotyping, array comparative genomic hybridization (aCGH), and array-based global gene expression profiling on 31 primary carcinomas and 15 established cell lines. Importantly, aCGH showed that the genomic profiles of primary tumors are recapitulated in the cell lines. We revealed a preponderance of chromosome breakpoints at sites of copy number variants (CNVs) in the CRC cell lines, a novel mechanism of DNA breakage in cancer. The integration of gene expression and aCGH led to the identification of 157 genes localized within high-level copy number changes whose transcriptional deregulation was significantly affected across all of the samples, thereby suggesting that these genes play a functional role in CRC. Genomic amplification at 8q24 was the most recurrent event and led to the overexpression of MYC and FAM84B. Copy number dependent gene expression resulted in deregulation of known cancer genes such as APC, FGFR2, and ERBB2. The identification of only 36 genes whose localization near a breakpoint could account for their observed deregulated expression demonstrates that the major mechanism for transcriptional deregulation in CRC is genomic copy number changes resulting from chromosomal aberrations. PMID:19691111

  4. Genomic organization of plant aminopropyl transferases.

    Science.gov (United States)

    Rodríguez-Kessler, Margarita; Delgado-Sánchez, Pablo; Rodríguez-Kessler, Gabriela Theresia; Moriguchi, Takaya; Jiménez-Bremont, Juan Francisco

    2010-07-01

    Aminopropyl transferases like spermidine synthase (SPDS; EC 2.5.1.16), spermine synthase and thermospermine synthase (SPMS, tSPMS; EC 2.5.1.22) belong to a class of widely distributed enzymes that use decarboxylated S-adenosylmethionine as an aminopropyl donor and putrescine or spermidine as an amino acceptor to form in that order spermidine, spermine or thermospermine. We describe the analysis of plant genomic sequences encoding SPDS, SPMS, tSPMS and PMT (putrescine N-methyltransferase; EC 2.1.1.53). Genome organization (including exon size, gain and loss, as well as intron number, size, loss, retention, placement and phase, and the presence of transposons) of plant aminopropyl transferase genes were compared between the genomic sequences of SPDS, SPMS and tSPMS from Zea mays, Oryza sativa, Malus x domestica, Populus trichocarpa, Arabidopsis thaliana and Physcomitrella patens. In addition, the genomic organization of plant PMT genes, proposed to be derived from SPDS during the evolution of alkaloid metabolism, is illustrated. Herein, a particular conservation and arrangement of exon and intron sequences between plant SPDS, SPMS and PMT genes that clearly differs with that of ACL5 genes, is shown. The possible acquisition of the plant SPMS exon II and, in particular exon XI in the monocot SPMS genes, is a remarkable feature that allows their differentiation from SPDS genes. In accordance with our in silico analysis, functional complementation experiments of the maize ZmSPMS1 enzyme (previously considered to be SPDS) in yeast demonstrated its spermine synthase activity. Another significant aspect is the conservation of intron sequences among SPDS and PMT paralogs. In addition the existence of microsynteny among some SPDS paralogs, especially in P. trichocarpa and A. thaliana, supports duplication events of plant SPDS genes. Based in our analysis, we hypothesize that SPMS genes appeared with the divergence of vascular plants by a processes of gene duplication and the

  5. The MedSeq Project: a randomized trial of integrating whole genome sequencing into clinical medicine.

    Science.gov (United States)

    Vassy, Jason L; Lautenbach, Denise M; McLaughlin, Heather M; Kong, Sek Won; Christensen, Kurt D; Krier, Joel; Kohane, Isaac S; Feuerman, Lindsay Z; Blumenthal-Barby, Jennifer; Roberts, J Scott; Lehmann, Lisa Soleymani; Ho, Carolyn Y; Ubel, Peter A; MacRae, Calum A; Seidman, Christine E; Murray, Michael F; McGuire, Amy L; Rehm, Heidi L; Green, Robert C

    2014-03-20

    illuminate the impact of integrating genomic medicine into the clinical care of patients but also inform the design of future studies. ClinicalTrials.gov identifier NCT01736566.

  6. Genomics in Public Health: Perspective from the Office of Public Health Genomics at the Centers for Disease Control and Prevention (CDC).

    Science.gov (United States)

    Green, Ridgely Fisk; Dotson, W David; Bowen, Scott; Kolor, Katherine; Khoury, Muin J

    2015-01-01

    The national effort to use genomic knowledge to save lives is gaining momentum, as illustrated by the inclusion of genomics in key public health initiatives, including Healthy People 2020, and the recent launch of the precision medicine initiative. The Office of Public Health Genomics (OPHG) at the Centers for Disease Control and Prevention (CDC) partners with state public health departments and others to advance the translation of genome-based discoveries into disease prevention and population health. To do this, OPHG has adopted an "identify, inform, and integrate" model: identify evidence-based genomic applications ready for implementation, inform stakeholders about these applications, and integrate these applications into public health at the local, state, and national level. This paper addresses current and future work at OPHG for integrating genomics into public health programs.

  7. A database of PCR primers for the chloroplast genomes of higher plants

    Science.gov (United States)

    Heinze, Berthold

    2007-01-01

    Background Chloroplast genomes evolve slowly and many primers for PCR amplification and analysis of chloroplast sequences can be used across a wide array of genera. In some cases 'universal' primers have been designed for the purpose of working across species boundaries. However, the essential information on these primer sequences is scattered throughout the literature. Results A database is presented here which assembles published primer information for chloroplast DNA. Additional primers were designed to fill gaps where little or no primer information could be found. Amplicons are either the genes themselves (typically useful in studies of sequence variation in higher-order phylogeny) or they are spacers, introns, and intergenic regions (for studies of phylogeographic patterns within and among species). The current list of 'generic' primers consists of more than 700 sequences. Wherever possible, we give the locations of the primers in the thirteen fully sequenced chloroplast genomes (Nicotiana tabacum, Atropa belladonna, Spinacia oleracea, Arabidopsis thaliana, Populus trichocarpa, Oryza sativa, Pinus thunbergii, Marchantia polymorpha, Zea mays, Oenothera elata, Acorus calamus, Eucalyptus globulus, Medicago trunculata). Conclusion The database described here is designed to serve as a resource for researchers who are venturing into the study of poorly described chloroplast genomes, whether for large- or small-scale DNA sequencing projects, to study molecular variation or to investigate chloroplast evolution. PMID:17326828

  8. A database of PCR primers for the chloroplast genomes of higher plants

    Directory of Open Access Journals (Sweden)

    Heinze Berthold

    2007-02-01

    Full Text Available Abstract Background Chloroplast genomes evolve slowly and many primers for PCR amplification and analysis of chloroplast sequences can be used across a wide array of genera. In some cases 'universal' primers have been designed for the purpose of working across species boundaries. However, the essential information on these primer sequences is scattered throughout the literature. Results A database is presented here which assembles published primer information for chloroplast DNA. Additional primers were designed to fill gaps where little or no primer information could be found. Amplicons are either the genes themselves (typically useful in studies of sequence variation in higher-order phylogeny or they are spacers, introns, and intergenic regions (for studies of phylogeographic patterns within and among species. The current list of 'generic' primers consists of more than 700 sequences. Wherever possible, we give the locations of the primers in the thirteen fully sequenced chloroplast genomes (Nicotiana tabacum, Atropa belladonna, Spinacia oleracea, Arabidopsis thaliana, Populus trichocarpa, Oryza sativa, Pinus thunbergii, Marchantia polymorpha, Zea mays, Oenothera elata, Acorus calamus, Eucalyptus globulus, Medicago trunculata. Conclusion The database described here is designed to serve as a resource for researchers who are venturing into the study of poorly described chloroplast genomes, whether for large- or small-scale DNA sequencing projects, to study molecular variation or to investigate chloroplast evolution.

  9. Genome Maps, a new generation genome browser.

    Science.gov (United States)

    Medina, Ignacio; Salavert, Francisco; Sanchez, Rubén; de Maria, Alejandro; Alonso, Roberto; Escobar, Pablo; Bleda, Marta; Dopazo, Joaquín

    2013-07-01

    Genome browsers have gained importance as more genomes and related genomic information become available. However, the increase of information brought about by new generation sequencing technologies is, at the same time, causing a subtle but continuous decrease in the efficiency of conventional genome browsers. Here, we present Genome Maps, a genome browser that implements an innovative model of data transfer and management. The program uses highly efficient technologies from the new HTML5 standard, such as scalable vector graphics, that optimize workloads at both server and client sides and ensure future scalability. Thus, data management and representation are entirely carried out by the browser, without the need of any Java Applet, Flash or other plug-in technology installation. Relevant biological data on genes, transcripts, exons, regulatory features, single-nucleotide polymorphisms, karyotype and so forth, are imported from web services and are available as tracks. In addition, several DAS servers are already included in Genome Maps. As a novelty, this web-based genome browser allows the local upload of huge genomic data files (e.g. VCF or BAM) that can be dynamically visualized in real time at the client side, thus facilitating the management of medical data affected by privacy restrictions. Finally, Genome Maps can easily be integrated in any web application by including only a few lines of code. Genome Maps is an open source collaborative initiative available in the GitHub repository (https://github.com/compbio-bigdata-viz/genome-maps). Genome Maps is available at: http://www.genomemaps.org.

  10. Development of an integrated genome informatics, data management and workflow infrastructure: A toolbox for the study of complex disease genetics

    Directory of Open Access Journals (Sweden)

    Burren Oliver S

    2004-01-01

    Full Text Available Abstract The genetic dissection of complex disease remains a significant challenge. Sample-tracking and the recording, processing and storage of high-throughput laboratory data with public domain data, require integration of databases, genome informatics and genetic analyses in an easily updated and scaleable format. To find genes involved in multifactorial diseases such as type 1 diabetes (T1D, chromosome regions are defined based on functional candidate gene content, linkage information from humans and animal model mapping information. For each region, genomic information is extracted from Ensembl, converted and loaded into ACeDB for manual gene annotation. Homology information is examined using ACeDB tools and the gene structure verified. Manually curated genes are extracted from ACeDB and read into the feature database, which holds relevant local genomic feature data and an audit trail of laboratory investigations. Public domain information, manually curated genes, polymorphisms, primers, linkage and association analyses, with links to our genotyping database, are shown in Gbrowse. This system scales to include genetic, statistical, quality control (QC and biological data such as expression analyses of RNA or protein, all linked from a genomics integrative display. Our system is applicable to any genetic study of complex disease, of either large or small scale.

  11. Multiplexed precision genome editing with trackable genomic barcodes in yeast.

    Science.gov (United States)

    Roy, Kevin R; Smith, Justin D; Vonesch, Sibylle C; Lin, Gen; Tu, Chelsea Szu; Lederer, Alex R; Chu, Angela; Suresh, Sundari; Nguyen, Michelle; Horecka, Joe; Tripathi, Ashutosh; Burnett, Wallace T; Morgan, Maddison A; Schulz, Julia; Orsley, Kevin M; Wei, Wu; Aiyar, Raeka S; Davis, Ronald W; Bankaitis, Vytas A; Haber, James E; Salit, Marc L; St Onge, Robert P; Steinmetz, Lars M

    2018-07-01

    Our understanding of how genotype controls phenotype is limited by the scale at which we can precisely alter the genome and assess the phenotypic consequences of each perturbation. Here we describe a CRISPR-Cas9-based method for multiplexed accurate genome editing with short, trackable, integrated cellular barcodes (MAGESTIC) in Saccharomyces cerevisiae. MAGESTIC uses array-synthesized guide-donor oligos for plasmid-based high-throughput editing and features genomic barcode integration to prevent plasmid barcode loss and to enable robust phenotyping. We demonstrate that editing efficiency can be increased more than fivefold by recruiting donor DNA to the site of breaks using the LexA-Fkh1p fusion protein. We performed saturation editing of the essential gene SEC14 and identified amino acids critical for chemical inhibition of lipid signaling. We also constructed thousands of natural genetic variants, characterized guide mismatch tolerance at the genome scale, and ascertained that cryptic Pol III termination elements substantially reduce guide efficacy. MAGESTIC will be broadly useful to uncover the genetic basis of phenotypes in yeast.

  12. Early Events in Populus Hybrid and Fagus sylvatica Leaves Exposed to Ozone

    Directory of Open Access Journals (Sweden)

    R. Desotgiu

    2010-01-01

    Full Text Available This paper aims to investigate early responses to ozone in leaves of Fagus sylvatica (beech and Populus maximowiczii x Populus berolinensis (poplar. The experimental setup consisted of four open-air (OA plots, four charcoal-filtered (CF open-top chambers (OTCs, and four nonfiltered (NF OTCs. Qualitative and quantitative analyses were carried out on nonsymptomatic (CF and symptomatic (NF and OA leaves of both species. Qualitative analyses were performed applying microscopic techniques: Evans blue staining for detection of cell viability, CeCl3 staining of transmission electron microscope (TEM samples to detect the accumulation of H2O2, and multispectral fluorescence microimaging and microspectrofluorometry to investigate the accumulation of fluorescent phenolic compounds in the walls of the damaged cells. Quantitative analyses consisted of the analysis of the chlorophyll a fluorescence transients (fast kinetics. The early responses to ozone were demonstrated by the Evans blue and CeCl3 staining techniques that provided evidence of plant responses in both species 1 month before foliar symptoms became visible. The fluorescence transients analysis, too, demonstrated the breakdown of the oxygen evolving system and the inactivation of the end receptors of electrons at a very early stage, both in poplar and in beech. The accumulation of phenolic compounds in the cell walls, on the other hand, was a species-specific response detected in poplar, but not in beech. Evans blue and CeCl3 staining, as well as the multispectral fluorescence microimaging and microspectrofluorometry, can be used to support the field diagnosis of ozone injury, whereas the fast kinetics of chlorophyll fluorescence provides evidence of early physiological responses.

  13. cisMEP: an integrated repository of genomic epigenetic profiles and cis-regulatory modules in Drosophila.

    Science.gov (United States)

    Yang, Tzu-Hsien; Wang, Chung-Ching; Hung, Po-Cheng; Wu, Wei-Sheng

    2014-01-01

    Cis-regulatory modules (CRMs), or the DNA sequences required for regulating gene expression, play the central role in biological researches on transcriptional regulation in metazoan species. Nowadays, the systematic understanding of CRMs still mainly resorts to computational methods due to the time-consuming and small-scale nature of experimental methods. But the accuracy and reliability of different CRM prediction tools are still unclear. Without comparative cross-analysis of the results and combinatorial consideration with extra experimental information, there is no easy way to assess the confidence of the predicted CRMs. This limits the genome-wide understanding of CRMs. It is known that transcription factor binding and epigenetic profiles tend to determine functions of CRMs in gene transcriptional regulation. Thus integration of the genome-wide epigenetic profiles with systematically predicted CRMs can greatly help researchers evaluate and decipher the prediction confidence and possible transcriptional regulatory functions of these potential CRMs. However, these data are still fragmentary in the literatures. Here we performed the computational genome-wide screening for potential CRMs using different prediction tools and constructed the pioneer database, cisMEP (cis-regulatory module epigenetic profile database), to integrate these computationally identified CRMs with genomic epigenetic profile data. cisMEP collects the literature-curated TFBS location data and nine genres of epigenetic data for assessing the confidence of these potential CRMs and deciphering the possible CRM functionality. cisMEP aims to provide a user-friendly interface for researchers to assess the confidence of different potential CRMs and to understand the functions of CRMs through experimentally-identified epigenetic profiles. The deposited potential CRMs and experimental epigenetic profiles for confidence assessment provide experimentally testable hypotheses for the molecular mechanisms

  14. Eighteen microsatellite loci in Salix arbutifolia (Salicaceae) and cross-species amplification in Salix and Populus species.

    Science.gov (United States)

    Hoshikawa, Takeshi; Kikuchi, Satoshi; Nagamitsu, Teruyoshi; Tomaru, Nobuhiro

    2009-07-01

    Salix arbutifolia is a riparian dioecious tree species that is of conservation concern in Japan because of its highly restricted distribution. Eighteen polymorphic loci of dinucleotide microsatellites were isolated and characterized. Among these, estimates of the expected heterozygosity ranged from 0.350 to 0.879. Cross-species amplification was successful at 9-13 loci among six Salix species and at three loci in one Populus species. © 2009 Blackwell Publishing Ltd.

  15. Genetic origin and composition of a natural hybrid poplar Populus???jrtyschensis from two distantly related species

    OpenAIRE

    Jiang, Dechun; Feng, Jianju; Dong, Miao; Wu, Guili; Mao, Kangshan; Liu, Jianquan

    2016-01-01

    Background The factors that contribute to and maintain hybrid zones between distinct species are highly variable, depending on hybrid origins, frequencies and fitness. In this study, we aimed to examine genetic origins, compositions and possible maintenance of Populus???jrtyschensis, an assumed natural hybrid between two distantly related species. This hybrid poplar occurs mainly on the floodplains along the river valleys between the overlapping distributions of the two putative parents. Resu...

  16. Construction of reference chromosome-scale pseudomolecules for potato: integrating the potato genome with genetic and physical maps.

    Science.gov (United States)

    Sharma, Sanjeev Kumar; Bolser, Daniel; de Boer, Jan; Sønderkær, Mads; Amoros, Walter; Carboni, Martin Federico; D'Ambrosio, Juan Martín; de la Cruz, German; Di Genova, Alex; Douches, David S; Eguiluz, Maria; Guo, Xiao; Guzman, Frank; Hackett, Christine A; Hamilton, John P; Li, Guangcun; Li, Ying; Lozano, Roberto; Maass, Alejandro; Marshall, David; Martinez, Diana; McLean, Karen; Mejía, Nilo; Milne, Linda; Munive, Susan; Nagy, Istvan; Ponce, Olga; Ramirez, Manuel; Simon, Reinhard; Thomson, Susan J; Torres, Yerisf; Waugh, Robbie; Zhang, Zhonghua; Huang, Sanwen; Visser, Richard G F; Bachem, Christian W B; Sagredo, Boris; Feingold, Sergio E; Orjeda, Gisella; Veilleux, Richard E; Bonierbale, Merideth; Jacobs, Jeanne M E; Milbourne, Dan; Martin, David Michael Alan; Bryan, Glenn J

    2013-11-06

    The genome of potato, a major global food crop, was recently sequenced. The work presented here details the integration of the potato reference genome (DM) with a new sequence-tagged site marker-based linkage map and other physical and genetic maps of potato and the closely related species tomato. Primary anchoring of the DM genome assembly was accomplished by the use of a diploid segregating population, which was genotyped with several types of molecular genetic markers to construct a new ~936 cM linkage map comprising 2469 marker loci. In silico anchoring approaches used genetic and physical maps from the diploid potato genotype RH89-039-16 (RH) and tomato. This combined approach has allowed 951 superscaffolds to be ordered into pseudomolecules corresponding to the 12 potato chromosomes. These pseudomolecules represent 674 Mb (~93%) of the 723 Mb genome assembly and 37,482 (~96%) of the 39,031 predicted genes. The superscaffold order and orientation within the pseudomolecules are closely collinear with independently constructed high density linkage maps. Comparisons between marker distribution and physical location reveal regions of greater and lesser recombination, as well as regions exhibiting significant segregation distortion. The work presented here has led to a greatly improved ordering of the potato reference genome superscaffolds into chromosomal "pseudomolecules".

  17. The Banana Genome Hub

    Science.gov (United States)

    Droc, Gaëtan; Larivière, Delphine; Guignon, Valentin; Yahiaoui, Nabila; This, Dominique; Garsmeur, Olivier; Dereeper, Alexis; Hamelin, Chantal; Argout, Xavier; Dufayard, Jean-François; Lengelle, Juliette; Baurens, Franc-Christophe; Cenci, Alberto; Pitollat, Bertrand; D’Hont, Angélique; Ruiz, Manuel; Rouard, Mathieu; Bocs, Stéphanie

    2013-01-01

    Banana is one of the world’s favorite fruits and one of the most important crops for developing countries. The banana reference genome sequence (Musa acuminata) was recently released. Given the taxonomic position of Musa, the completed genomic sequence has particular comparative value to provide fresh insights about the evolution of the monocotyledons. The study of the banana genome has been enhanced by a number of tools and resources that allows harnessing its sequence. First, we set up essential tools such as a Community Annotation System, phylogenomics resources and metabolic pathways. Then, to support post-genomic efforts, we improved banana existing systems (e.g. web front end, query builder), we integrated available Musa data into generic systems (e.g. markers and genetic maps, synteny blocks), we have made interoperable with the banana hub, other existing systems containing Musa data (e.g. transcriptomics, rice reference genome, workflow manager) and finally, we generated new results from sequence analyses (e.g. SNP and polymorphism analysis). Several uses cases illustrate how the Banana Genome Hub can be used to study gene families. Overall, with this collaborative effort, we discuss the importance of the interoperability toward data integration between existing information systems. Database URL: http://banana-genome.cirad.fr/ PMID:23707967

  18. Toward genome-enabled mycology.

    Science.gov (United States)

    Hibbett, David S; Stajich, Jason E; Spatafora, Joseph W

    2013-01-01

    Genome-enabled mycology is a rapidly expanding field that is characterized by the pervasive use of genome-scale data and associated computational tools in all aspects of fungal biology. Genome-enabled mycology is integrative and often requires teams of researchers with diverse skills in organismal mycology, bioinformatics and molecular biology. This issue of Mycologia presents the first complete fungal genomes in the history of the journal, reflecting the ongoing transformation of mycology into a genome-enabled science. Here, we consider the prospects for genome-enabled mycology and the technical and social challenges that will need to be overcome to grow the database of complete fungal genomes and enable all fungal biologists to make use of the new data.

  19. Integrated Genomic Analysis Identifies Clinically Relevant Subtypes of Glioblastoma Characterized by Abnormalities in PDGFRA, IDH1, EGFR, and NF1

    Energy Technology Data Exchange (ETDEWEB)

    Verhaak, Roel GW; Hoadley, Katherine A; Purdom, Elizabeth; Wang, Victoria; Qi, Yuan; Wilkerson, Matthew D; Miller, C Ryan; Ding, Li; Golub, Todd; Mesirov, Jill P; Alexe, Gabriele; Lawrence, Michael; O' Kelly, Michael; Tamayo, Pablo; Weir, Barbara A; Gabriel, Stacey; Winckler, Wendy; Gupta, Supriya; Jakkula, Lakshmi; Feiler, Heidi S; Hodgson, J Graeme; James, C David; Sarkaria, Jann N; Brennan, Cameron; Kahn, Ari; Spellman, Paul T; Wilson, Richard K; Speed, Terence P; Gray, Joe W; Meyerson, Matthew; Getz, Gad; Perou, Charles M; Hayes, D Neil; Network, The Cancer Genome Atlas Research

    2009-09-03

    The Cancer Genome Atlas Network recently cataloged recurrent genomic abnormalities in glioblastoma multiforme (GBM). We describe a robust gene expression-based molecular classification of GBM into Proneural, Neural, Classical, and Mesenchymal subtypes and integrate multidimensional genomic data to establish patterns of somatic mutations and DNA copy number. Aberrations and gene expression of EGFR, NF1, and PDGFRA/IDH1 each define the Classical, Mesenchymal, and Proneural subtypes, respectively. Gene signatures of normal brain cell types show a strong relationship between subtypes and different neural lineages. Additionally, response to aggressive therapy differs by subtype, with the greatest benefit in the Classical subtype and no benefit in the Proneural subtype. We provide a framework that unifies transcriptomic and genomic dimensions for GBM molecular stratification with important implications for future studies.

  20. PROCESS OPTIMIZATION OF TETRA ACETYL ETHYLENE DIAMINE ACTIVATED HYDROGEN PEROXIDE BLEACHING OF POPULUS NIGRA CTMP

    OpenAIRE

    Qiang Zhao; Junwen Pu; Shulei Mao; Guibo Qi

    2010-01-01

    To enhance the bleaching efficiency, the activator of tetra acetyl ethylene diamine (TAED) was used in conventional H2O2 bleaching. The H2O2/TAED bleaching system can accelerate the reaction rate and shorten bleaching time at relative low temperature, which can reduce the production cost. In this research, the process with hydrogen peroxide activated by TAED bleaching of Populus nigra chemi-thermo mechanical pulp was optimized. Suitable bleaching conditions were confirmed as follows: pulp con...

  1. Stem respiration of Populus species in the third year of free-air CO2 enrichment

    OpenAIRE

    GIELEN, Birgit; Scarascia-Mugnozza, G.; Ceulemans, R.

    2003-01-01

    Carbon cycling in ecosystems, and especially in forests, is intensively studied to predict the effects of global climate change, and the role which forests may play in 'changing climate change'. One of the questions is whether the carbon balance of forests will be affected by increasing atmospheric CO2 concentrations. Regarding this question, effects of elevated [CO2 ] on woody-tissue respiration have frequently been neglected. Stem respiration of three Populus species (P. alba L. (Clone 2AS-...

  2. Integration of HIV in the Human Genome: Which Sites Are Preferential? A Genetic and Statistical Assessment

    Science.gov (United States)

    Gonçalves, Juliana; Moreira, Elsa; Sequeira, Inês J.; Rodrigues, António S.; Rueff, José; Brás, Aldina

    2016-01-01

    Chromosomal fragile sites (FSs) are loci where gaps and breaks may occur and are preferential integration targets for some viruses, for example, Hepatitis B, Epstein-Barr virus, HPV16, HPV18, and MLV vectors. However, the integration of the human immunodeficiency virus (HIV) in Giemsa bands and in FSs is not yet completely clear. This study aimed to assess the integration preferences of HIV in FSs and in Giemsa bands using an in silico study. HIV integration positions from Jurkat cells were used and two nonparametric tests were applied to compare HIV integration in dark versus light bands and in FS versus non-FS (NFSs). The results show that light bands are preferential targets for integration of HIV-1 in Jurkat cells and also that it integrates with equal intensity in FSs and in NFSs. The data indicates that HIV displays different preferences for FSs compared to other viruses. The aim was to develop and apply an approach to predict the conditions and constraints of HIV insertion in the human genome which seems to adequately complement empirical data. PMID:27294106

  3. Integration of Genome-Wide TF Binding and Gene Expression Data to Characterize Gene Regulatory Networks in Plant Development.

    Science.gov (United States)

    Chen, Dijun; Kaufmann, Kerstin

    2017-01-01

    Key transcription factors (TFs) controlling the morphogenesis of flowers and leaves have been identified in the model plant Arabidopsis thaliana. Recent genome-wide approaches based on chromatin immunoprecipitation (ChIP) followed by high-throughput DNA sequencing (ChIP-seq) enable systematic identification of genome-wide TF binding sites (TFBSs) of these regulators. Here, we describe a computational pipeline for analyzing ChIP-seq data to identify TFBSs and to characterize gene regulatory networks (GRNs) with applications to the regulatory studies of flower development. In particular, we provide step-by-step instructions on how to download, analyze, visualize, and integrate genome-wide data in order to construct GRNs for beginners of bioinformatics. The practical guide presented here is ready to apply to other similar ChIP-seq datasets to characterize GRNs of interest.

  4. Influence of irrigation and fertilization on transpiration and hydraulic properties of Populus deltoides.

    Energy Technology Data Exchange (ETDEWEB)

    Samuelson, Lisa, J.; Stokes, Thomas, A.; Coleman, Mark, D.

    2007-02-01

    Summary Long-term hydraulic acclimation to resource availability was explored in 3-year-bld Populus deltoides Bartr. ex Marsh. clones by examining transpiration. leaf-specific hydraulic conductance (GL), canopy stomatal conductance (Gs) and leaf to sapwood area ratio (AL:Asi)n response to imgation (13 and 551 mm year in addition to ambient precipitation) and fertilization (0 and 120 kg N ha-' year-'). Sap flow was measured continuously over one growing season with thermal dissipation probes. Fertilization had a greater effect on growth and hydraulic properties than imgation, and fertilization effects were independent of irrigation treatment. Transpiration on a ground area basis (E) ranged between 0.3 and 1.8 mm day-', and increased 66% and 90% in response to imgation and fertilization, respectively. Increases in GL, Gs at a reference vapor pressure deficit of 1 kPa, and transpiration per unit leaf areain response to increases in resource availability were associated with reductions in AL:As and consequently a minimal change in the water potential gradient from soil to leaf. Imgation and fertilization increased leaf area index similarly, from an average 1.16 in control stands to 1.45, but sapwood area was increased from 4.0 to 6.3 m ha-' by irrigation and from 3.7 to 6.7 m2 ha-' by fertilization. The balance between leaf area and sapwood area was important in understanding long-term hydraulic acclimation to resource availability and mechanisms controlling maximum productivity in Populus deltoides.

  5. Genome Variation Map: a data repository of genome variations in BIG Data Center.

    Science.gov (United States)

    Song, Shuhui; Tian, Dongmei; Li, Cuiping; Tang, Bixia; Dong, Lili; Xiao, Jingfa; Bao, Yiming; Zhao, Wenming; He, Hang; Zhang, Zhang

    2018-01-04

    The Genome Variation Map (GVM; http://bigd.big.ac.cn/gvm/) is a public data repository of genome variations. As a core resource in the BIG Data Center, Beijing Institute of Genomics, Chinese Academy of Sciences, GVM dedicates to collect, integrate and visualize genome variations for a wide range of species, accepts submissions of different types of genome variations from all over the world and provides free open access to all publicly available data in support of worldwide research activities. Unlike existing related databases, GVM features integration of a large number of genome variations for a broad diversity of species including human, cultivated plants and domesticated animals. Specifically, the current implementation of GVM not only houses a total of ∼4.9 billion variants for 19 species including chicken, dog, goat, human, poplar, rice and tomato, but also incorporates 8669 individual genotypes and 13 262 manually curated high-quality genotype-to-phenotype associations for non-human species. In addition, GVM provides friendly intuitive web interfaces for data submission, browse, search and visualization. Collectively, GVM serves as an important resource for archiving genomic variation data, helpful for better understanding population genetic diversity and deciphering complex mechanisms associated with different phenotypes. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  6. Genome Variation Map: a data repository of genome variations in BIG Data Center

    Science.gov (United States)

    Tian, Dongmei; Li, Cuiping; Tang, Bixia; Dong, Lili; Xiao, Jingfa; Bao, Yiming; Zhao, Wenming; He, Hang

    2018-01-01

    Abstract The Genome Variation Map (GVM; http://bigd.big.ac.cn/gvm/) is a public data repository of genome variations. As a core resource in the BIG Data Center, Beijing Institute of Genomics, Chinese Academy of Sciences, GVM dedicates to collect, integrate and visualize genome variations for a wide range of species, accepts submissions of different types of genome variations from all over the world and provides free open access to all publicly available data in support of worldwide research activities. Unlike existing related databases, GVM features integration of a large number of genome variations for a broad diversity of species including human, cultivated plants and domesticated animals. Specifically, the current implementation of GVM not only houses a total of ∼4.9 billion variants for 19 species including chicken, dog, goat, human, poplar, rice and tomato, but also incorporates 8669 individual genotypes and 13 262 manually curated high-quality genotype-to-phenotype associations for non-human species. In addition, GVM provides friendly intuitive web interfaces for data submission, browse, search and visualization. Collectively, GVM serves as an important resource for archiving genomic variation data, helpful for better understanding population genetic diversity and deciphering complex mechanisms associated with different phenotypes. PMID:29069473

  7. Ginseng Genome Database: an open-access platform for genomics of Panax ginseng.

    Science.gov (United States)

    Jayakodi, Murukarthick; Choi, Beom-Soon; Lee, Sang-Choon; Kim, Nam-Hoon; Park, Jee Young; Jang, Woojong; Lakshmanan, Meiyappan; Mohan, Shobhana V G; Lee, Dong-Yup; Yang, Tae-Jin

    2018-04-12

    The ginseng (Panax ginseng C.A. Meyer) is a perennial herbaceous plant that has been used in traditional oriental medicine for thousands of years. Ginsenosides, which have significant pharmacological effects on human health, are the foremost bioactive constituents in this plant. Having realized the importance of this plant to humans, an integrated omics resource becomes indispensable to facilitate genomic research, molecular breeding and pharmacological study of this herb. The first draft genome sequences of P. ginseng cultivar "Chunpoong" were reported recently. Here, using the draft genome, transcriptome, and functional annotation datasets of P. ginseng, we have constructed the Ginseng Genome Database http://ginsengdb.snu.ac.kr /, the first open-access platform to provide comprehensive genomic resources of P. ginseng. The current version of this database provides the most up-to-date draft genome sequence (of approximately 3000 Mbp of scaffold sequences) along with the structural and functional annotations for 59,352 genes and digital expression of genes based on transcriptome data from different tissues, growth stages and treatments. In addition, tools for visualization and the genomic data from various analyses are provided. All data in the database were manually curated and integrated within a user-friendly query page. This database provides valuable resources for a range of research fields related to P. ginseng and other species belonging to the Apiales order as well as for plant research communities in general. Ginseng genome database can be accessed at http://ginsengdb.snu.ac.kr /.

  8. Computational genomics of hyperthermophiles

    NARCIS (Netherlands)

    Werken, van de H.J.G.

    2008-01-01

    With the ever increasing number of completely sequenced prokaryotic genomes and the subsequent use of functional genomics tools, e.g. DNA microarray and proteomics, computational data analysis and the integration of microbial and molecular data is inevitable. This thesis describes the computational

  9. Evolutionary time-scale of the begomoviruses: evidence from integrated sequences in the Nicotiana genome.

    Directory of Open Access Journals (Sweden)

    Pierre Lefeuvre

    Full Text Available Despite having single stranded DNA genomes that are replicated by host DNA polymerases, viruses in the family Geminiviridae are apparently evolving as rapidly as some RNA viruses. The observed substitution rates of geminiviruses in the genera Begomovirus and Mastrevirus are so high that the entire family could conceivably have originated less than a million years ago (MYA. However, the existence of geminivirus related DNA (GRD integrated within the genomes of various Nicotiana species suggests that the geminiviruses probably originated >10 MYA. Some have even suggested that a distinct New-World (NW lineage of begomoviruses may have arisen following the separation by continental drift of African and American proto-begomoviruses ∼110 MYA. We evaluate these various geminivirus origin hypotheses using Bayesian coalescent-based approaches to date firstly the Nicotiana GRD integration events, and then the divergence of the NW and Old-World (OW begomoviruses. Besides rejecting the possibility of a<2 MYA OW-NW begomovirus split, we could also discount that it may have occurred concomitantly with the breakup of Gondwanaland 110 MYA. Although we could only confidently narrow the date of the split down to between 2 and 80 MYA, the most plausible (and best supported date for the split is between 20 and 30 MYA--a time when global cooling ended the dispersal of temperate species between Asia and North America via the Beringian land bridge.

  10. The differential response of photosynthesis to high temperature for a boreal and temperate Populus species relates to differences in Rubisco activation and Rubisco activase properties.

    Science.gov (United States)

    Hozain, Moh'd I; Salvucci, Michael E; Fokar, Mohamed; Holaday, A Scott

    2010-01-01

    Significant inhibition of photosynthesis occurs at temperatures only a few degrees (Populus species adapted to contrasting thermal environments for determining the factors that constrain photosynthetic assimilation (A) under moderate heat stress in tree species. Consistent with its native range in temperate regions, Populus deltoides Bartr. ex Marsh. exhibited a significantly higher temperature optimum for A than did Populus balsamifera L., a boreal species. The higher A exhibited by P. deltoides at 33-40 degrees C compared to that for P. balsamifera was associated with a higher activation state of Rubisco and correlated with a higher ATPase activity of Rubisco activase. The temperature response of minimal chlorophyll a fluorescence for darkened leaves was similar for both species and was not consistent with a thylakoid lipid phase change contributing to the decline in A in the range of 30-40 degrees C. Taken together, these data support the idea that the differences in the temperature response of A for the two Populus species could be attributed to the differences in the response of Rubisco activation and ultimately to the thermal properties of Rubisco activase. That the primary sequence of Rubisco activase differed between the species, especially in regions associated with ATPase activity and Rubisco recognition, indicates that the genotypic differences in Rubisco activase might underlie the differences in the heat sensitivity of Rubisco activase and photosynthesis at moderately high temperatures.

  11. Chloride and sodium uptake potential over an entire rotation of Populus irrigated with landfill leachate.

    Science.gov (United States)

    Zalesny, Jill A; Zalesny, Ronald S

    2009-07-01

    There is a need for information about the response of Populus genotypes to repeated application of high-salinity water and nutrient sources throughout an entire rotation. We have combined establishment biomass and uptake data with mid- and full-rotation growth data to project potential chloride (Cl-) and sodium (Na+) uptake for 2- to 11-year-old Populus in the north central United States. Our objectives were to identify potential levels of uptake as the trees developed and stages of plantation development that are conducive to variable application rates of high-salinity irrigation. The projected cumulative uptake of Cl- and Na+ during mid-rotation plantation development was stable 2 to 3 years after planting but increased steadily from year 3 to 6. Year six cumulative uptake ranged from 22 to 175 kg Cl- ha(-1) and 8 to 74 kg Na+ ha(-1), while annual uptake ranged from 8 to 54 kg Cl- ha(-1) yr(-1) and 3 to 23 kg Na+ ha(-1) yr(-1). Full-rotation uptake was greatest from 4 to 9 years (Cl-) and 4 to 8 years (Na+), with maximum levels of Cl- (32 kg ha(-1) yr(-1)) and Na+ (13 kg ha(-1) yr(-1)) occurring in year six. The relative uptake potential of Cl- and Na+ at peak accumulation (year six) was 2.7 times greater than at the end of the rotation.

  12. Integrated Genomic Analysis of the Ubiquitin Pathway across Cancer Types

    Directory of Open Access Journals (Sweden)

    Zhongqi Ge

    2018-04-01

    Full Text Available Summary: Protein ubiquitination is a dynamic and reversible process of adding single ubiquitin molecules or various ubiquitin chains to target proteins. Here, using multidimensional omic data of 9,125 tumor samples across 33 cancer types from The Cancer Genome Atlas, we perform comprehensive molecular characterization of 929 ubiquitin-related genes and 95 deubiquitinase genes. Among them, we systematically identify top somatic driver candidates, including mutated FBXW7 with cancer-type-specific patterns and amplified MDM2 showing a mutually exclusive pattern with BRAF mutations. Ubiquitin pathway genes tend to be upregulated in cancer mediated by diverse mechanisms. By integrating pan-cancer multiomic data, we identify a group of tumor samples that exhibit worse prognosis. These samples are consistently associated with the upregulation of cell-cycle and DNA repair pathways, characterized by mutated TP53, MYC/TERT amplification, and APC/PTEN deletion. Our analysis highlights the importance of the ubiquitin pathway in cancer development and lays a foundation for developing relevant therapeutic strategies. : Ge et al. analyze a cohort of 9,125 TCGA samples across 33 cancer types to provide a comprehensive characterization of the ubiquitin pathway. They detect somatic driver candidates in the ubiquitin pathway and identify a cluster of patients with poor survival, highlighting the importance of this pathway in cancer development. Keywords: ubiquitin pathway, pan-cancer analysis, The Cancer Genome Atlas, tumor subtype, cancer prognosis, therapeutic targets, biomarker, FBXW7

  13. Earthworms, arthropods and plant litter decomposition in aspen (Populus tremuloides) and lodgepole pine(Pinus contorta) forests in Colorado, USA

    Science.gov (United States)

    Grizelle Gonzalez; Timothy R. Seastedt; Zugeily Donato

    2003-01-01

    We compared the abundance and community composition of earthworms, soil macroarthropods, and litter microarthropods to test faunal effects on plant litter decomposition rates in two forests in the subalpine in Colorado, USA. Litterbags containing recently senesced litter of Populus tremuloides (aspen) and Pinus contorta (lodgepole pine) were placed in aspen and pine...

  14. Preserving genome integrity: the DdrA protein of Deinococcus radiodurans R1.

    Science.gov (United States)

    Harris, Dennis R; Tanaka, Masashi; Saveliev, Sergei V; Jolivet, Edmond; Earl, Ashlee M; Cox, Michael M; Battista, John R

    2004-10-01

    The bacterium Deinococcus radiodurans can withstand extraordinary levels of ionizing radiation, reflecting an equally extraordinary capacity for DNA repair. The hypothetical gene product DR0423 has been implicated in the recovery of this organism from DNA damage, indicating that this protein is a novel component of the D. radiodurans DNA repair system. DR0423 is a homologue of the eukaryotic Rad52 protein. Following exposure to ionizing radiation, DR0423 expression is induced relative to an untreated control, and strains carrying a deletion of the DR0423 gene exhibit increased sensitivity to ionizing radiation. When recovering from ionizing-radiation-induced DNA damage in the absence of nutrients, wild-type D. radiodurans reassembles its genome while the mutant lacking DR0423 function does not. In vitro, the purified DR0423 protein binds to single-stranded DNA with an apparent affinity for 3' ends, and protects those ends from nuclease degradation. We propose that DR0423 is part of a DNA end-protection system that helps to preserve genome integrity following exposure to ionizing radiation. We designate the DR0423 protein as DNA damage response A protein.

  15. Preserving genome integrity: the DdrA protein of Deinococcus radiodurans R1.

    Directory of Open Access Journals (Sweden)

    Dennis R Harris

    2004-10-01

    Full Text Available The bacterium Deinococcus radiodurans can withstand extraordinary levels of ionizing radiation, reflecting an equally extraordinary capacity for DNA repair. The hypothetical gene product DR0423 has been implicated in the recovery of this organism from DNA damage, indicating that this protein is a novel component of the D. radiodurans DNA repair system. DR0423 is a homologue of the eukaryotic Rad52 protein. Following exposure to ionizing radiation, DR0423 expression is induced relative to an untreated control, and strains carrying a deletion of the DR0423 gene exhibit increased sensitivity to ionizing radiation. When recovering from ionizing-radiation-induced DNA damage in the absence of nutrients, wild-type D. radiodurans reassembles its genome while the mutant lacking DR0423 function does not. In vitro, the purified DR0423 protein binds to single-stranded DNA with an apparent affinity for 3' ends, and protects those ends from nuclease degradation. We propose that DR0423 is part of a DNA end-protection system that helps to preserve genome integrity following exposure to ionizing radiation. We designate the DR0423 protein as DNA damage response A protein.

  16. Cancer prevention, the need to preserve the integrity of the genome at all cost.

    Science.gov (United States)

    Okafor, M T; Nwagha, T U; Anusiem, C; Okoli, U A; Nubila, N I; Al-Alloosh, F; Udenyia, I J

    2018-05-01

    The entire genetic information carried by an organism makes up its genome. Genes have a diverse number of functions. They code different proteins for normal proliferation of cells. However, changes in the base sequence of genes affect their protein by-products which act as messengers for normal cellular functions such as proliferation and repairs. Salient processes for maintaining the integrity of the genome are hinged on intricate mechanisms put in place for the evolution to tackle genomic stresses. To discuss how cells sense and repair damage to their deoxyribonucleic acid (DNA) as well as to highlight how defects in the genes involved in DNA repair contribute to cancer development. Methodology: Online searches on the following databases such as Google Scholar, PubMed, Biomed Central, and SciELO were done. Attempt was made to review articles with keywords such as cancer, cell cycle, tumor suppressor genes, and DNA repair. The cell cycle, tumor suppression genes, DNA repair mechanism, as well as their contribution to cancer development, were discussed and reviewed. Knowledge on how cells detect and repair DNA damage through an array of mechanisms should allay our anxiety as regards cancer development. More studies on DNA damage detection and repair processes are important toward a holistic approach to cancer treatment.

  17. Enhanced annotations and features for comparing thousands of Pseudomonas genomes in the Pseudomonas genome database.

    Science.gov (United States)

    Winsor, Geoffrey L; Griffiths, Emma J; Lo, Raymond; Dhillon, Bhavjinder K; Shay, Julie A; Brinkman, Fiona S L

    2016-01-04

    The Pseudomonas Genome Database (http://www.pseudomonas.com) is well known for the application of community-based annotation approaches for producing a high-quality Pseudomonas aeruginosa PAO1 genome annotation, and facilitating whole-genome comparative analyses with other Pseudomonas strains. To aid analysis of potentially thousands of complete and draft genome assemblies, this database and analysis platform was upgraded to integrate curated genome annotations and isolate metadata with enhanced tools for larger scale comparative analysis and visualization. Manually curated gene annotations are supplemented with improved computational analyses that help identify putative drug targets and vaccine candidates or assist with evolutionary studies by identifying orthologs, pathogen-associated genes and genomic islands. The database schema has been updated to integrate isolate metadata that will facilitate more powerful analysis of genomes across datasets in the future. We continue to place an emphasis on providing high-quality updates to gene annotations through regular review of the scientific literature and using community-based approaches including a major new Pseudomonas community initiative for the assignment of high-quality gene ontology terms to genes. As we further expand from thousands of genomes, we plan to provide enhancements that will aid data visualization and analysis arising from whole-genome comparative studies including more pan-genome and population-based approaches. © The Author(s) 2015. Published by Oxford University Press on behalf of Nucleic Acids Research.

  18. Integrative Genomics: Quantifying significance of phenotype-genotype relationships from multiple sources of high-throughput data

    Directory of Open Access Journals (Sweden)

    Eric eGamazon

    2013-05-01

    Full Text Available Given recent advances in the generation of high-throughput data such as whole genome genetic variation and transcriptome expression, it is critical to come up with novel methods to integrate these heterogeneous datasets and to assess the significance of identified phenotype-genotype relationships. Recent studies show that genome-wide association findings are likely to fall in loci with gene regulatory effects such as expression quantitative trait loci (eQTLs, demonstrating the utility of such integrative approaches. When genotype and gene expression data are available on the same individuals, we developed methods wherein top phenotype-associated genetic variants are prioritized if they are associated, as eQTLs, with gene expression traits that are themselves associated with the phenotype. Yet there has been no method to determine an overall p-value for the findings that arise specifically from the integrative nature of the approach. We propose a computationally feasible permutation method that accounts for the assimilative nature of the method and the correlation structure among gene expression traits and among genotypes. We apply the method to data from a study of cellular sensitivity to etoposide, one of the most widely used chemotherapeutic drugs. To our knowledge, this study is the first statistically sound quantification of the significance of the genotype-phenotype relationships resulting from applying an integrative approach. This method can be easily extended to cases in which gene expression data are replaced by other molecular phenotypes of interest, e.g., microRNA or proteomic data. This study has important implications for studies seeking to expand on genetic association studies by the use of omics data. Finally, we provide an R code to compute the empirical FDR when p-values for the observed and simulated phenotypes are available.

  19. Brassica ASTRA: an integrated database for Brassica genomic research.

    Science.gov (United States)

    Love, Christopher G; Robinson, Andrew J; Lim, Geraldine A C; Hopkins, Clare J; Batley, Jacqueline; Barker, Gary; Spangenberg, German C; Edwards, David

    2005-01-01

    Brassica ASTRA is a public database for genomic information on Brassica species. The database incorporates expressed sequences with Swiss-Prot and GenBank comparative sequence annotation as well as secondary Gene Ontology (GO) annotation derived from the comparison with Arabidopsis TAIR GO annotations. Simple sequence repeat molecular markers are identified within resident sequences and mapped onto the closely related Arabidopsis genome sequence. Bacterial artificial chromosome (BAC) end sequences derived from the Multinational Brassica Genome Project are also mapped onto the Arabidopsis genome sequence enabling users to identify candidate Brassica BACs corresponding to syntenic regions of Arabidopsis. This information is maintained in a MySQL database with a web interface providing the primary means of interrogation. The database is accessible at http://hornbill.cspp.latrobe.edu.au.

  20. JGI Fungal Genomics Program

    Energy Technology Data Exchange (ETDEWEB)

    Grigoriev, Igor V.

    2011-03-14

    Genomes of energy and environment fungi are in focus of the Fungal Genomic Program at the US Department of Energy Joint Genome Institute (JGI). Its key project, the Genomics Encyclopedia of Fungi, targets fungi related to plant health (symbionts, pathogens, and biocontrol agents) and biorefinery processes (cellulose degradation, sugar fermentation, industrial hosts), and explores fungal diversity by means of genome sequencing and analysis. Over 50 fungal genomes have been sequenced by JGI to date and released through MycoCosm (www.jgi.doe.gov/fungi), a fungal web-portal, which integrates sequence and functional data with genome analysis tools for user community. Sequence analysis supported by functional genomics leads to developing parts list for complex systems ranging from ecosystems of biofuel crops to biorefineries. Recent examples of such 'parts' suggested by comparative genomics and functional analysis in these areas are presented here

  1. Genomic Encyclopedia of Fungi

    Energy Technology Data Exchange (ETDEWEB)

    Grigoriev, Igor

    2012-08-10

    Genomes of fungi relevant to energy and environment are in focus of the Fungal Genomic Program at the US Department of Energy Joint Genome Institute (JGI). Its key project, the Genomics Encyclopedia of Fungi, targets fungi related to plant health (symbionts, pathogens, and biocontrol agents) and biorefinery processes (cellulose degradation, sugar fermentation, industrial hosts), and explores fungal diversity by means of genome sequencing and analysis. Over 150 fungal genomes have been sequenced by JGI to date and released through MycoCosm (www.jgi.doe.gov/fungi), a fungal web-portal, which integrates sequence and functional data with genome analysis tools for user community. Sequence analysis supported by functional genomics leads to developing parts list for complex systems ranging from ecosystems of biofuel crops to biorefineries. Recent examples of such parts suggested by comparative genomics and functional analysis in these areas are presented here.

  2. Savant Genome Browser 2: visualization and analysis for population-scale genomics.

    Science.gov (United States)

    Fiume, Marc; Smith, Eric J M; Brook, Andrew; Strbenac, Dario; Turner, Brian; Mezlini, Aziz M; Robinson, Mark D; Wodak, Shoshana J; Brudno, Michael

    2012-07-01

    High-throughput sequencing (HTS) technologies are providing an unprecedented capacity for data generation, and there is a corresponding need for efficient data exploration and analysis capabilities. Although most existing tools for HTS data analysis are developed for either automated (e.g. genotyping) or visualization (e.g. genome browsing) purposes, such tools are most powerful when combined. For example, integration of visualization and computation allows users to iteratively refine their analyses by updating computational parameters within the visual framework in real-time. Here we introduce the second version of the Savant Genome Browser, a standalone program for visual and computational analysis of HTS data. Savant substantially improves upon its predecessor and existing tools by introducing innovative visualization modes and navigation interfaces for several genomic datatypes, and synergizing visual and automated analyses in a way that is powerful yet easy even for non-expert users. We also present a number of plugins that were developed by the Savant Community, which demonstrate the power of integrating visual and automated analyses using Savant. The Savant Genome Browser is freely available (open source) at www.savantbrowser.com.

  3. HGVA: the Human Genome Variation Archive

    OpenAIRE

    Lopez, Javier; Coll, Jacobo; Haimel, Matthias; Kandasamy, Swaathi; Tarraga, Joaquin; Furio-Tari, Pedro; Bari, Wasim; Bleda, Marta; Rueda, Antonio; Gr?f, Stefan; Rendon, Augusto; Dopazo, Joaquin; Medina, Ignacio

    2017-01-01

    Abstract High-profile genomic variation projects like the 1000 Genomes project or the Exome Aggregation Consortium, are generating a wealth of human genomic variation knowledge which can be used as an essential reference for identifying disease-causing genotypes. However, accessing these data, contrasting the various studies and integrating those data in downstream analyses remains cumbersome. The Human Genome Variation Archive (HGVA) tackles these challenges and facilitates access to genomic...

  4. CHESS (CgHExpreSS): a comprehensive analysis tool for the analysis of genomic alterations and their effects on the expression profile of the genome.

    Science.gov (United States)

    Lee, Mikyung; Kim, Yangseok

    2009-12-16

    Genomic alterations frequently occur in many cancer patients and play important mechanistic roles in the pathogenesis of cancer. Furthermore, they can modify the expression level of genes due to altered copy number in the corresponding region of the chromosome. An accumulating body of evidence supports the possibility that strong genome-wide correlation exists between DNA content and gene expression. Therefore, more comprehensive analysis is needed to quantify the relationship between genomic alteration and gene expression. A well-designed bioinformatics tool is essential to perform this kind of integrative analysis. A few programs have already been introduced for integrative analysis. However, there are many limitations in their performance of comprehensive integrated analysis using published software because of limitations in implemented algorithms and visualization modules. To address this issue, we have implemented the Java-based program CHESS to allow integrative analysis of two experimental data sets: genomic alteration and genome-wide expression profile. CHESS is composed of a genomic alteration analysis module and an integrative analysis module. The genomic alteration analysis module detects genomic alteration by applying a threshold based method or SW-ARRAY algorithm and investigates whether the detected alteration is phenotype specific or not. On the other hand, the integrative analysis module measures the genomic alteration's influence on gene expression. It is divided into two separate parts. The first part calculates overall correlation between comparative genomic hybridization ratio and gene expression level by applying following three statistical methods: simple linear regression, Spearman rank correlation and Pearson's correlation. In the second part, CHESS detects the genes that are differentially expressed according to the genomic alteration pattern with three alternative statistical approaches: Student's t-test, Fisher's exact test and Chi square

  5. Integrate genome-based assessment of safety for probiotic strains: Bacillus coagulans GBI-30, 6086 as a case study.

    Science.gov (United States)

    Salvetti, Elisa; Orrù, Luigi; Capozzi, Vittorio; Martina, Alessia; Lamontanara, Antonella; Keller, David; Cash, Howard; Felis, Giovanna E; Cattivelli, Luigi; Torriani, Sandra; Spano, Giuseppe

    2016-05-01

    Probiotics are microorganisms that confer beneficial effects on the host; nevertheless, before being allowed for human consumption, their safety must be verified with accurate protocols. In the genomic era, such procedures should take into account the genomic-based approaches. This study aims at assessing the safety traits of Bacillus coagulans GBI-30, 6086 integrating the most updated genomics-based procedures and conventional phenotypic assays. Special attention was paid to putative virulence factors (VF), antibiotic resistance (AR) genes and genes encoding enzymes responsible for harmful metabolites (i.e. biogenic amines, BAs). This probiotic strain was phenotypically resistant to streptomycin and kanamycin, although the genome analysis suggested that the AR-related genes were not easily transferrable to other bacteria, and no other genes with potential safety risks, such as those related to VF or BA production, were retrieved. Furthermore, no unstable elements that could potentially lead to genomic rearrangements were detected. Moreover, a workflow is proposed to allow the proper taxonomic identification of a microbial strain and the accurate evaluation of risk-related gene traits, combining whole genome sequencing analysis with updated bioinformatics tools and standard phenotypic assays. The workflow presented can be generalized as a guideline for the safety investigation of novel probiotic strains to help stakeholders (from scientists to manufacturers and consumers) to meet regulatory requirements and avoid misleading information.

  6. An integrative approach to predicting the functional effects of small indels in non-coding regions of the human genome.

    Science.gov (United States)

    Ferlaino, Michael; Rogers, Mark F; Shihab, Hashem A; Mort, Matthew; Cooper, David N; Gaunt, Tom R; Campbell, Colin

    2017-10-06

    Small insertions and deletions (indels) have a significant influence in human disease and, in terms of frequency, they are second only to single nucleotide variants as pathogenic mutations. As the majority of mutations associated with complex traits are located outside the exome, it is crucial to investigate the potential pathogenic impact of indels in non-coding regions of the human genome. We present FATHMM-indel, an integrative approach to predict the functional effect, pathogenic or neutral, of indels in non-coding regions of the human genome. Our method exploits various genomic annotations in addition to sequence data. When validated on benchmark data, FATHMM-indel significantly outperforms CADD and GAVIN, state of the art models in assessing the pathogenic impact of non-coding variants. FATHMM-indel is available via a web server at indels.biocompute.org.uk. FATHMM-indel can accurately predict the functional impact and prioritise small indels throughout the whole non-coding genome.

  7. Trade-offs between xylem hydraulic properties, wood anatomy and yield in Populus.

    Science.gov (United States)

    Hajek, Peter; Leuschner, Christoph; Hertel, Dietrich; Delzon, Sylvain; Schuldt, Bernhard

    2014-07-01

    Trees face the dilemma that achieving high plant productivity is accompanied by a risk of drought-induced hydraulic failure due to a trade-off in the trees' vascular system between hydraulic efficiency and safety. By investigating the xylem anatomy of branches and coarse roots, and measuring branch axial hydraulic conductivity and vulnerability to cavitation in 4-year-old field-grown aspen plants of five demes (Populus tremula L. and Populus tremuloides Michx.) differing in growth rate, we tested the hypotheses that (i) demes differ in wood anatomical and hydraulic properties, (ii) hydraulic efficiency and safety are related to xylem anatomical traits, and (iii) aboveground productivity and hydraulic efficiency are negatively correlated to cavitation resistance. Significant deme differences existed in seven of the nine investigated branch-related anatomical and hydraulic traits but only in one of the four coarse-root-related anatomical traits; this likely is a consequence of high intra-plant variation in root morphology and the occurrence of a few 'high-conductivity roots'. Growth rate was positively related to branch hydraulic efficiency (xylem-specific conductivity) but not to cavitation resistance; this indicates that no marked trade-off exists between cavitation resistance and growth. Both branch hydraulic safety and hydraulic efficiency significantly depended on vessel size and were related to the genetic distance between the demes, while the xylem pressure causing 88% loss of hydraulic conductivity (P88 value) was more closely related to hydraulic efficiency than the commonly used P50 value. Deme-specific variation in the pit membrane structure may explain why vessel size was not directly linked to growth rate. We conclude that branch hydraulic efficiency is an important growth-influencing trait in aspen, while the assumed trade-off between productivity and hydraulic safety is weak. © The Author 2014. Published by Oxford University Press. All rights reserved

  8. Integration of transcriptome and whole genomic resequencing data to identify key genes affecting swine fat deposition.

    Directory of Open Access Journals (Sweden)

    Kai Xing

    Full Text Available Fat deposition is highly correlated with the growth, meat quality, reproductive performance and immunity of pigs. Fatty acid synthesis takes place mainly in the adipose tissue of pigs; therefore, in this study, a high-throughput massively parallel sequencing approach was used to generate adipose tissue transcriptomes from two groups of Songliao black pigs that had opposite backfat thickness phenotypes. The total number of paired-end reads produced for each sample was in the range of 39.29-49.36 millions. Approximately 188 genes were differentially expressed in adipose tissue and were enriched for metabolic processes, such as fatty acid biosynthesis, lipid synthesis, metabolism of fatty acids, etinol, caffeine and arachidonic acid and immunity. Additionally, many genetic variations were detected between the two groups through pooled whole-genome resequencing. Integration of transcriptome and whole-genome resequencing data revealed important genomic variations among the differentially expressed genes for fat deposition, for example, the lipogenic genes. Further studies are required to investigate the roles of candidate genes in fat deposition to improve pig breeding programs.

  9. PopGenome: An Efficient Swiss Army Knife for Population Genomic Analyses in R

    OpenAIRE

    Pfeifer, Bastian; Wittelsbürger, Ulrich; Ramos-Onsins, Sebastian E.; Lercher, Martin J.

    2014-01-01

    Although many computer programs can perform population genetics calculations, they are typically limited in the analyses and data input formats they offer; few applications can process the large data sets produced by whole-genome resequencing projects. Furthermore, there is no coherent framework for the easy integration of new statistics into existing pipelines, hindering the development and application of new population genetics and genomics approaches. Here, we present PopGenome, a populati...

  10. The human vascular endothelial cell line HUV-EC-C harbors the integrated HHV-6B genome which remains stable in long term culture.

    Science.gov (United States)

    Shioda, Setsuko; Kasai, Fumio; Ozawa, Midori; Hirayama, Noriko; Satoh, Motonobu; Kameoka, Yousuke; Watanabe, Ken; Shimizu, Norio; Tang, Huamin; Mori, Yasuko; Kohara, Arihiro

    2018-02-01

    Human herpes virus 6 (HHV-6) is a common human pathogen that is most often detected in hematopoietic cells. Although human cells harboring chromosomally integrated HHV-6 can be generated in vitro, the availability of such cell lines originating from in vivo tissues is limited. In this study, chromosomally integrated HHV-6B has been identified in a human vascular endothelial cell line, HUV-EC-C (IFO50271), derived from normal umbilical cord tissue. Sequence analysis revealed that the viral genome was similar to the HHV-6B HST strain. FISH analysis using a HHV-6 DNA probe showed one signal in each cell, detected at the distal end of the long arm of chromosome 9. This was consistent with a digital PCR assay, validating one copy of the viral DNA. Because exposure of HUV-EC-C to chemicals did not cause viral reactivation, long term cell culture of HUV-EC-C was carried out to assess the stability of viral integration. The growth rate was altered depending on passage numbers, and morphology also changed during culture. SNP microarray profiles showed some differences between low and high passages, implying that the HUV-EC-C genome had changed during culture. However, no detectable change was observed in chromosome 9, where HHV-6B integration and the viral copy number remained unchanged. Our results suggest that integrated HHV-6B is stable in HUV-EC-C despite genome instability.

  11. HiView: an integrative genome browser to leverage Hi-C results for the interpretation of GWAS variants.

    Science.gov (United States)

    Xu, Zheng; Zhang, Guosheng; Duan, Qing; Chai, Shengjie; Zhang, Baqun; Wu, Cong; Jin, Fulai; Yue, Feng; Li, Yun; Hu, Ming

    2016-03-11

    Genome-wide association studies (GWAS) have identified thousands of genetic variants associated with complex traits and diseases. However, most of them are located in the non-protein coding regions, and therefore it is challenging to hypothesize the functions of these non-coding GWAS variants. Recent large efforts such as the ENCODE and Roadmap Epigenomics projects have predicted a large number of regulatory elements. However, the target genes of these regulatory elements remain largely unknown. Chromatin conformation capture based technologies such as Hi-C can directly measure the chromatin interactions and have generated an increasingly comprehensive catalog of the interactome between the distal regulatory elements and their potential target genes. Leveraging such information revealed by Hi-C holds the promise of elucidating the functions of genetic variants in human diseases. In this work, we present HiView, the first integrative genome browser to leverage Hi-C results for the interpretation of GWAS variants. HiView is able to display Hi-C data and statistical evidence for chromatin interactions in genomic regions surrounding any given GWAS variant, enabling straightforward visualization and interpretation. We believe that as the first GWAS variants-centered Hi-C genome browser, HiView is a useful tool guiding post-GWAS functional genomics studies. HiView is freely accessible at: http://www.unc.edu/~yunmli/HiView .

  12. An integrative genomic and transcriptomic analysis reveals potential targets associated with cell proliferation in uterine leiomyomas.

    Directory of Open Access Journals (Sweden)

    Priscila Daniele Ramos Cirilo

    Full Text Available Uterine Leiomyomas (ULs are the most common benign tumours affecting women of reproductive age. ULs represent a major problem in public health, as they are the main indication for hysterectomy. Approximately 40-50% of ULs have non-random cytogenetic abnormalities, and half of ULs may have copy number alterations (CNAs. Gene expression microarrays studies have demonstrated that cell proliferation genes act in response to growth factors and steroids. However, only a few genes mapping to CNAs regions were found to be associated with ULs.We applied an integrative analysis using genomic and transcriptomic data to identify the pathways and molecular markers associated with ULs. Fifty-one fresh frozen specimens were evaluated by array CGH (JISTIC and gene expression microarrays (SAM. The CONEXIC algorithm was applied to integrate the data.The integrated analysis identified the top 30 significant genes (P<0.01, which comprised genes associated with cancer, whereas the protein-protein interaction analysis indicated a strong association between FANCA and BRCA1. Functional in silico analysis revealed target molecules for drugs involved in cell proliferation, including FGFR1 and IGFBP5. Transcriptional and protein analyses showed that FGFR1 (P = 0.006 and P<0.01, respectively and IGFBP5 (P = 0.0002 and P = 0.006, respectively were up-regulated in the tumours when compared with the adjacent normal myometrium.The integrative genomic and transcriptomic approach indicated that FGFR1 and IGFBP5 amplification, as well as the consequent up-regulation of the protein products, plays an important role in the aetiology of ULs and thus provides data for potential drug therapies development to target genes associated with cellular proliferation in ULs.

  13. Causes of genome instability

    DEFF Research Database (Denmark)

    Langie, Sabine A S; Koppen, Gudrun; Desaulniers, Daniel

    2015-01-01

    function, chromosome segregation, telomere length). The purpose of this review is to describe the crucial aspects of genome instability, to outline the ways in which environmental chemicals can affect this cancer hallmark and to identify candidate chemicals for further study. The overall aim is to make......Genome instability is a prerequisite for the development of cancer. It occurs when genome maintenance systems fail to safeguard the genome's integrity, whether as a consequence of inherited defects or induced via exposure to environmental agents (chemicals, biological agents and radiation). Thus...

  14. IW-Scoring: an Integrative Weighted Scoring framework for annotating and prioritizing genetic variations in the noncoding genome.

    Science.gov (United States)

    Wang, Jun; Dayem Ullah, Abu Z; Chelala, Claude

    2018-01-30

    The vast majority of germline and somatic variations occur in the noncoding part of the genome, only a small fraction of which are believed to be functional. From the tens of thousands of noncoding variations detectable in each genome, identifying and prioritizing driver candidates with putative functional significance is challenging. To address this, we implemented IW-Scoring, a new Integrative Weighted Scoring model to annotate and prioritise functionally relevant noncoding variations. We evaluate 11 scoring methods, and apply an unsupervised spectral approach for subsequent selective integration into two linear weighted functional scoring schemas for known and novel variations. IW-Scoring produces stable high-quality performance as the best predictors for three independent data sets. We demonstrate the robustness of IW-Scoring in identifying recurrent functional mutations in the TERT promoter, as well as disease SNPs in proximity to consensus motifs and with gene regulatory effects. Using follicular lymphoma as a paradigmatic cancer model, we apply IW-Scoring to locate 11 recurrently mutated noncoding regions in 14 follicular lymphoma genomes, and validate 9 of these regions in an extension cohort, including the promoter and enhancer regions of PAX5. Overall, IW-Scoring demonstrates greater versatility in identifying trait- and disease-associated noncoding variants. Scores from IW-Scoring as well as other methods are freely available from http://www.snp-nexus.org/IW-Scoring/. © The Author(s) 2018. Published by Oxford University Press on behalf of Nucleic Acids Research.

  15. SINEs as driving forces in genome evolution.

    Science.gov (United States)

    Schmitz, J

    2012-01-01

    SINEs are short interspersed elements derived from cellular RNAs that repetitively retropose via RNA intermediates and integrate more or less randomly back into the genome. SINEs propagate almost entirely vertically within their host cells and, once established in the germline, are passed on from generation to generation. As non-autonomous elements, their reverse transcription (from RNA to cDNA) and genomic integration depends on the activity of the enzymatic machinery of autonomous retrotransposons, such as long interspersed elements (LINEs). SINEs are widely distributed in eukaryotes, but are especially effectively propagated in mammalian species. For example, more than a million Alu-SINE copies populate the human genome (approximately 13% of genomic space), and few master copies of them are still active. In the organisms where they occur, SINEs are a challenge to genomic integrity, but in the long term also can serve as beneficial building blocks for evolution, contributing to phenotypic heterogeneity and modifying gene regulatory networks. They substantially expand the genomic space and introduce structural variation to the genome. SINEs have the potential to mutate genes, to alter gene expression, and to generate new parts of genes. A balanced distribution and controlled activity of such properties is crucial to maintaining the organism's dynamic and thriving evolution. Copyright © 2012 S. Karger AG, Basel.

  16. Implementing genomics and pharmacogenomics in the clinic: The National Human Genome Research Institute's genomic medicine portfolio.

    Science.gov (United States)

    Manolio, Teri A

    2016-10-01

    Increasing knowledge about the influence of genetic variation on human health and growing availability of reliable, cost-effective genetic testing have spurred the implementation of genomic medicine in the clinic. As defined by the National Human Genome Research Institute (NHGRI), genomic medicine uses an individual's genetic information in his or her clinical care, and has begun to be applied effectively in areas such as cancer genomics, pharmacogenomics, and rare and undiagnosed diseases. In 2011 NHGRI published its strategic vision for the future of genomic research, including an ambitious research agenda to facilitate and promote the implementation of genomic medicine. To realize this agenda, NHGRI is consulting and facilitating collaborations with the external research community through a series of "Genomic Medicine Meetings," under the guidance and leadership of the National Advisory Council on Human Genome Research. These meetings have identified and begun to address significant obstacles to implementation, such as lack of evidence of efficacy, limited availability of genomics expertise and testing, lack of standards, and difficulties in integrating genomic results into electronic medical records. The six research and dissemination initiatives comprising NHGRI's genomic research portfolio are designed to speed the evaluation and incorporation, where appropriate, of genomic technologies and findings into routine clinical care. Actual adoption of successful approaches in clinical care will depend upon the willingness, interest, and energy of professional societies, practitioners, patients, and payers to promote their responsible use and share their experiences in doing so. Published by Elsevier Ireland Ltd.

  17. Graph-based semi-supervised learning with genomic data integration using condition-responsive genes applied to phenotype classification.

    Science.gov (United States)

    Doostparast Torshizi, Abolfazl; Petzold, Linda R

    2018-01-01

    Data integration methods that combine data from different molecular levels such as genome, epigenome, transcriptome, etc., have received a great deal of interest in the past few years. It has been demonstrated that the synergistic effects of different biological data types can boost learning capabilities and lead to a better understanding of the underlying interactions among molecular levels. In this paper we present a graph-based semi-supervised classification algorithm that incorporates latent biological knowledge in the form of biological pathways with gene expression and DNA methylation data. The process of graph construction from biological pathways is based on detecting condition-responsive genes, where 3 sets of genes are finally extracted: all condition responsive genes, high-frequency condition-responsive genes, and P-value-filtered genes. The proposed approach is applied to ovarian cancer data downloaded from the Human Genome Atlas. Extensive numerical experiments demonstrate superior performance of the proposed approach compared to other state-of-the-art algorithms, including the latest graph-based classification techniques. Simulation results demonstrate that integrating various data types enhances classification performance and leads to a better understanding of interrelations between diverse omics data types. The proposed approach outperforms many of the state-of-the-art data integration algorithms. © The Author 2017. Published by Oxford University Press on behalf of the American Medical Informatics Association. All rights reserved. For Permissions, please email: journals.permissions@oup.com

  18. Multiple-integrations of HPV16 genome and altered transcription of viral oncogenes and cellular genes are associated with the development of cervical cancer.

    Directory of Open Access Journals (Sweden)

    Xulian Lu

    Full Text Available The constitutive expression of the high-risk HPV E6 and E7 viral oncogenes is the major cause of cervical cancer. To comprehensively explore the composition of HPV16 early transcripts and their genomic annotation, cervical squamous epithelial tissues from 40 HPV16-infected patients were collected for analysis of papillomavirus oncogene transcripts (APOT. We observed different transcription patterns of HPV16 oncogenes in progression of cervical lesions to cervical cancer and identified one novel transcript. Multiple-integration events in the tissues of cervical carcinoma (CxCa are significantly more often than those of low-grade squamous intraepithelial lesions (LSIL and high-grade squamous intraepithelial lesions (HSIL. Moreover, most cellular genes within or near these integration sites are cancer-associated genes. Taken together, this study suggests that the multiple-integrations of HPV genome during persistent viral infection, which thereby alters the expression patterns of viral oncogenes and integration-related cellular genes, play a crucial role in progression of cervical lesions to cervix cancer.

  19. Patient-controlled encrypted genomic data: an approach to advance clinical genomics

    Directory of Open Access Journals (Sweden)

    Trakadis Yannis J

    2012-07-01

    Full Text Available Abstract Background The revolution in DNA sequencing technologies over the past decade has made it feasible to sequence an individual’s whole genome at a relatively low cost. The potential value of the information generated by genomic technologies for medicine and society is enormous. However, in order for exome sequencing, and eventually whole genome sequencing, to be implemented clinically, a number of major challenges need to be overcome. For instance, obtaining meaningful informed-consent, managing incidental findings and the great volume of data generated (including multiple findings with uncertain clinical significance, re-interpreting the genomic data and providing additional counselling to patients as genetic knowledge evolves are issues that need to be addressed. It appears that medical genetics is shifting from the present “phenotype-first” medical model to a “data-first” model which leads to multiple complexities. Discussion This manuscript discusses the different challenges associated with integrating genomic technologies into clinical practice and describes a “phenotype-first” approach, namely, “Individualized Mutation-weighed Phenotype Search”, and its benefits. The proposed approach allows for a more efficient prioritization of the genes to be tested in a clinical lab based on both the patient’s phenotype and his/her entire genomic data. It simplifies “informed-consent” for clinical use of genomic technologies and helps to protect the patient’s autonomy and privacy. Overall, this approach could potentially render widespread use of genomic technologies, in the immediate future, practical, ethical and clinically useful. Summary The “Individualized Mutation-weighed Phenotype Search” approach allows for an incremental integration of genomic technologies into clinical practice. It ensures that we do not over-medicalize genomic data but, rather, continue our current medical model which is based on serving

  20. Robust and rapid algorithms facilitate large-scale whole genome sequencing downstream analysis in an integrative framework.

    Science.gov (United States)

    Li, Miaoxin; Li, Jiang; Li, Mulin Jun; Pan, Zhicheng; Hsu, Jacob Shujui; Liu, Dajiang J; Zhan, Xiaowei; Wang, Junwen; Song, Youqiang; Sham, Pak Chung

    2017-05-19

    Whole genome sequencing (WGS) is a promising strategy to unravel variants or genes responsible for human diseases and traits. However, there is a lack of robust platforms for a comprehensive downstream analysis. In the present study, we first proposed three novel algorithms, sequence gap-filled gene feature annotation, bit-block encoded genotypes and sectional fast access to text lines to address three fundamental problems. The three algorithms then formed the infrastructure of a robust parallel computing framework, KGGSeq, for integrating downstream analysis functions for whole genome sequencing data. KGGSeq has been equipped with a comprehensive set of analysis functions for quality control, filtration, annotation, pathogenic prediction and statistical tests. In the tests with whole genome sequencing data from 1000 Genomes Project, KGGSeq annotated several thousand more reliable non-synonymous variants than other widely used tools (e.g. ANNOVAR and SNPEff). It took only around half an hour on a small server with 10 CPUs to access genotypes of ∼60 million variants of 2504 subjects, while a popular alternative tool required around one day. KGGSeq's bit-block genotype format used 1.5% or less space to flexibly represent phased or unphased genotypes with multiple alleles and achieved a speed of over 1000 times faster to calculate genotypic correlation. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  1. SilkDB: a knowledgebase for silkworm biology and genomics

    DEFF Research Database (Denmark)

    Wang, Jing; Xia, Qingyou; He, Ximiao

    2005-01-01

    The Silkworm Knowledgebase (SilkDB) is a web-based repository for the curation, integration and study of silkworm genetic and genomic data. With the recent accomplishment of a approximately 6X draft genome sequence of the domestic silkworm (Bombyx mori), SilkDB provides an integrated representati....... SilkDB is publicly accessible at http://silkworm.genomics.org.cn. Udgivelsesdato: 2005-Jan-1...

  2. Salt stress induces the formation of a novel type of 'pressure wood' in two Populus species.

    Science.gov (United States)

    Janz, Dennis; Lautner, Silke; Wildhagen, Henning; Behnke, Katja; Schnitzler, Jörg-Peter; Rennenberg, Heinz; Fromm, Jörg; Polle, Andrea

    2012-04-01

    • Salinity causes osmotic stress and limits biomass production of plants. The goal of this study was to investigate mechanisms underlying hydraulic adaptation to salinity. • Anatomical, ecophysiological and transcriptional responses to salinity were investigated in the xylem of a salt-sensitive (Populus × canescens) and a salt-tolerant species (Populus euphratica). • Moderate salt stress, which suppressed but did not abolish photosynthesis and radial growth in P. × canescens, resulted in hydraulic adaptation by increased vessel frequencies and decreased vessel lumina. Transcript abundances of a suite of genes (FLA, COB-like, BAM, XET, etc.) previously shown to be activated during tension wood formation, were collectively suppressed in developing xylem, whereas those for stress and defense-related genes increased. A subset of cell wall-related genes was also suppressed in salt-exposed P. euphratica, although this species largely excluded sodium and showed no anatomical alterations. Salt exposure influenced cell wall composition involving increases in the lignin : carbohydrate ratio in both species. • In conclusion, hydraulic stress adaptation involves cell wall modifications reciprocal to tension wood formation that result in the formation of a novel type of reaction wood in upright stems named 'pressure wood'. Our data suggest that transcriptional co-regulation of a core set of genes determines reaction wood composition. © 2011 The Authors. New Phytologist © 2011 New Phytologist Trust.

  3. VaProS: a database-integration approach for protein/genome information retrieval

    KAUST Repository

    Gojobori, Takashi; Ikeo, Kazuho; Katayama, Yukie; Kawabata, Takeshi; Kinjo, Akira R.; Kinoshita, Kengo; Kwon, Yeondae; Migita, Ohsuke; Mizutani, Hisashi; Muraoka, Masafumi; Nagata, Koji; Omori, Satoshi; Sugawara, Hideaki; Yamada, Daichi; Yura, Kei

    2016-01-01

    Life science research now heavily relies on all sorts of databases for genome sequences, transcription, protein three-dimensional (3D) structures, protein–protein interactions, phenotypes and so forth. The knowledge accumulated by all the omics research is so vast that a computer-aided search of data is now a prerequisite for starting a new study. In addition, a combinatory search throughout these databases has a chance to extract new ideas and new hypotheses that can be examined by wet-lab experiments. By virtually integrating the related databases on the Internet, we have built a new web application that facilitates life science researchers for retrieving experts’ knowledge stored in the databases and for building a new hypothesis of the research target. This web application, named VaProS, puts stress on the interconnection between the functional information of genome sequences and protein 3D structures, such as structural effect of the gene mutation. In this manuscript, we present the notion of VaProS, the databases and tools that can be accessed without any knowledge of database locations and data formats, and the power of search exemplified in quest of the molecular mechanisms of lysosomal storage disease. VaProS can be freely accessed at http://p4d-info.nig.ac.jp/vapros/.

  4. VaProS: a database-integration approach for protein/genome information retrieval

    KAUST Repository

    Gojobori, Takashi

    2016-12-24

    Life science research now heavily relies on all sorts of databases for genome sequences, transcription, protein three-dimensional (3D) structures, protein–protein interactions, phenotypes and so forth. The knowledge accumulated by all the omics research is so vast that a computer-aided search of data is now a prerequisite for starting a new study. In addition, a combinatory search throughout these databases has a chance to extract new ideas and new hypotheses that can be examined by wet-lab experiments. By virtually integrating the related databases on the Internet, we have built a new web application that facilitates life science researchers for retrieving experts’ knowledge stored in the databases and for building a new hypothesis of the research target. This web application, named VaProS, puts stress on the interconnection between the functional information of genome sequences and protein 3D structures, such as structural effect of the gene mutation. In this manuscript, we present the notion of VaProS, the databases and tools that can be accessed without any knowledge of database locations and data formats, and the power of search exemplified in quest of the molecular mechanisms of lysosomal storage disease. VaProS can be freely accessed at http://p4d-info.nig.ac.jp/vapros/.

  5. The genome portal of the Department of Energy Joint Genome Institute: 2014 updates

    Energy Technology Data Exchange (ETDEWEB)

    Nordberg, Henrik [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Cantor, Michael [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Dusheyko, Serge [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Hua, Susan [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Poliakov, Alexander [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Shabalov, Igor [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Smirnova, Tatyana [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Grigoriev, Igor V. [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States); Dubchak, Inna [USDOE Joint Genome Institute (JGI), Walnut Creek, CA (United States)

    2013-11-12

    The U.S. Department of Energy (DOE) Joint Genome Institute (JGI), a national user facility, serves the diverse scientific community by providing integrated high-throughput sequencing and computational analysis to enable system-based scientific approaches in support of DOE missions related to clean energy generation and environmental characterization. The JGI Genome Portal (http://genome.jgi.doe.gov) provides unified access to all JGI genomic databases and analytical tools. The JGI maintains extensive data management systems and specialized analytical capabilities to manage and interpret complex genomic data. A user can search, download and explore multiple data sets available for all DOE JGI sequencing projects including their status, assemblies and annotations of sequenced genomes. In this paper, we describe major updates of the Genome Portal in the past 2 years with a specific emphasis on efficient handling of the rapidly growing amount of diverse genomic data accumulated in JGI.

  6. Implementing genomics and pharmacogenomics in the clinic: The National Human Genome Research Institute’s genomic medicine portfolio

    Science.gov (United States)

    Manolio, Teri A.

    2016-01-01

    Increasing knowledge about the influence of genetic variation on human health and growing availability of reliable, cost-effective genetic testing have spurred the implementation of genomic medicine in the clinic. As defined by the National Human Genome Research Institute (NHGRI), genomic medicine uses an individual’s genetic information in his or her clinical care, and has begun to be applied effectively in areas such as cancer genomics, pharmacogenomics, and rare and undiagnosed diseases. In 2011 NHGRI published its strategic vision for the future of genomic research, including an ambitious research agenda to facilitate and promote the implementation of genomic medicine. To realize this agenda, NHGRI is consulting and facilitating collaborations with the external research community through a series of “Genomic Medicine Meetings,” under the guidance and leadership of the National Advisory Council on Human Genome Research. These meetings have identified and begun to address significant obstacles to implementation, such as lack of evidence of efficacy, limited availability of genomics expertise and testing, lack of standards, and diffficulties in integrating genomic results into electronic medical records. The six research and dissemination initiatives comprising NHGRI’s genomic research portfolio are designed to speed the evaluation and incorporation, where appropriate, of genomic technologies and findings into routine clinical care. Actual adoption of successful approaches in clinical care will depend upon the willingness, interest, and energy of professional societies, practitioners, patients, and payers to promote their responsible use and share their experiences in doing so. PMID:27612677

  7. Variable Nitrogen Fixation in Wild Populus.

    Directory of Open Access Journals (Sweden)

    Sharon L Doty

    Full Text Available The microbiome of plants is diverse, and like that of animals, is important for overall health and nutrient acquisition. In legumes and actinorhizal plants, a portion of essential nitrogen (N is obtained through symbiosis with nodule-inhabiting, N2-fixing microorganisms. However, a variety of non-nodulating plant species can also thrive in natural, low-N settings. Some of these species may rely on endophytes, microorganisms that live within plants, to fix N2 gas into usable forms. Here we report the first direct evidence of N2 fixation in the early successional wild tree, Populus trichocarpa, a non-leguminous tree, from its native riparian habitat. In order to measure N2 fixation, surface-sterilized cuttings of wild poplar were assayed using both 15N2 incorporation and the commonly used acetylene reduction assay. The 15N label was incorporated at high levels in a subset of cuttings, suggesting a high level of N-fixation. Similarly, acetylene was reduced to ethylene in some samples. The microbiota of the cuttings was highly variable, both in numbers of cultured bacteria and in genetic diversity. Our results indicated that associative N2-fixation occurred within wild poplar and that a non-uniformity in the distribution of endophytic bacteria may explain the variability in N-fixation activity. These results point to the need for molecular studies to decipher the required microbial consortia and conditions for effective endophytic N2-fixation in trees.

  8. Integration of heterologous DNA into the genome of Paracoccus denitrificans is mediated by a family of IS1248-related elements and a second type of integrative recombination event

    NARCIS (Netherlands)

    Van Spanning, R J; Reijnders, W N; Stouthamer, A.H.

    All members of the IS1248 family residing in the genome of Paracoccus denitrificans have been isolated by using a set of insertion sequence entrapment vectors. The family consists of five closely related members that integrate the entrapment vectors at distinct sites. One of these, IS1248b, was

  9. Sparse multivariate factor analysis regression models and its applications to integrative genomics analysis.

    Science.gov (United States)

    Zhou, Yan; Wang, Pei; Wang, Xianlong; Zhu, Ji; Song, Peter X-K

    2017-01-01

    The multivariate regression model is a useful tool to explore complex associations between two kinds of molecular markers, which enables the understanding of the biological pathways underlying disease etiology. For a set of correlated response variables, accounting for such dependency can increase statistical power. Motivated by integrative genomic data analyses, we propose a new methodology-sparse multivariate factor analysis regression model (smFARM), in which correlations of response variables are assumed to follow a factor analysis model with latent factors. This proposed method not only allows us to address the challenge that the number of association parameters is larger than the sample size, but also to adjust for unobserved genetic and/or nongenetic factors that potentially conceal the underlying response-predictor associations. The proposed smFARM is implemented by the EM algorithm and the blockwise coordinate descent algorithm. The proposed methodology is evaluated and compared to the existing methods through extensive simulation studies. Our results show that accounting for latent factors through the proposed smFARM can improve sensitivity of signal detection and accuracy of sparse association map estimation. We illustrate smFARM by two integrative genomics analysis examples, a breast cancer dataset, and an ovarian cancer dataset, to assess the relationship between DNA copy numbers and gene expression arrays to understand genetic regulatory patterns relevant to the disease. We identify two trans-hub regions: one in cytoband 17q12 whose amplification influences the RNA expression levels of important breast cancer genes, and the other in cytoband 9q21.32-33, which is associated with chemoresistance in ovarian cancer. © 2016 WILEY PERIODICALS, INC.

  10. DNA-PKcs, ATM, and ATR Interplay Maintains Genome Integrity during Neurogenesis.

    Science.gov (United States)

    Enriquez-Rios, Vanessa; Dumitrache, Lavinia C; Downing, Susanna M; Li, Yang; Brown, Eric J; Russell, Helen R; McKinnon, Peter J

    2017-01-25

    The DNA damage response (DDR) orchestrates a network of cellular processes that integrates cell-cycle control and DNA repair or apoptosis, which serves to maintain genome stability. DNA-PKcs (the catalytic subunit of the DNA-dependent kinase, encoded by PRKDC), ATM (ataxia telangiectasia, mutated), and ATR (ATM and Rad3-related) are related PI3K-like protein kinases and central regulators of the DDR. Defects in these kinases have been linked to neurodegenerative or neurodevelopmental syndromes. In all cases, the key neuroprotective function of these kinases is uncertain. It also remains unclear how interactions between the three DNA damage-responsive kinases coordinate genome stability, particularly in a physiological context. Here, we used a genetic approach to identify the neural function of DNA-PKcs and the interplay between ATM and ATR during neurogenesis. We found that DNA-PKcs loss in the mouse sensitized neuronal progenitors to apoptosis after ionizing radiation because of excessive DNA damage. DNA-PKcs was also required to prevent endogenous DNA damage accumulation throughout the adult brain. In contrast, ATR coordinated the DDR during neurogenesis to direct apoptosis in cycling neural progenitors, whereas ATM regulated apoptosis in both proliferative and noncycling cells. We also found that ATR controls a DNA damage-induced G 2 /M checkpoint in cortical progenitors, independent of ATM and DNA-PKcs. These nonoverlapping roles were further confirmed via sustained murine embryonic or cortical development after all three kinases were simultaneously inactivated. Thus, our results illustrate how DNA-PKcs, ATM, and ATR have unique and essential roles during the DDR, collectively ensuring comprehensive genome maintenance in the nervous system. The DNA damage response (DDR) is essential for prevention of a broad spectrum of different human neurologic diseases. However, a detailed understanding of the DDR at a physiological level is lacking. In contrast to many in

  11. Endogenous viral elements in animal genomes.

    Directory of Open Access Journals (Sweden)

    Aris Katzourakis

    2010-11-01

    Full Text Available Integration into the nuclear genome of germ line cells can lead to vertical inheritance of retroviral genes as host alleles. For other viruses, germ line integration has only rarely been documented. Nonetheless, we identified endogenous viral elements (EVEs derived from ten non-retroviral families by systematic in silico screening of animal genomes, including the first endogenous representatives of double-stranded RNA, reverse-transcribing DNA, and segmented RNA viruses, and the first endogenous DNA viruses in mammalian genomes. Phylogenetic and genomic analysis of EVEs across multiple host species revealed novel information about the origin and evolution of diverse virus groups. Furthermore, several of the elements identified here encode intact open reading frames or are expressed as mRNA. For one element in the primate lineage, we provide statistically robust evidence for exaptation. Our findings establish that genetic material derived from all known viral genome types and replication strategies can enter the animal germ line, greatly broadening the scope of paleovirological studies and indicating a more significant evolutionary role for gene flow from virus to animal genomes than has previously been recognized.

  12. Genomics in Public Health: Perspective from the Office of Public Health Genomics at the Centers for Disease Control and Prevention (CDC

    Directory of Open Access Journals (Sweden)

    Ridgely Fisk Green

    2015-09-01

    Full Text Available The national effort to use genomic knowledge to save lives is gaining momentum, as illustrated by the inclusion of genomics in key public health initiatives, including Healthy People 2020, and the recent launch of the precision medicine initiative. The Office of Public Health Genomics (OPHG at the Centers for Disease Control and Prevention (CDC partners with state public health departments and others to advance the translation of genome-based discoveries into disease prevention and population health. To do this, OPHG has adopted an “identify, inform, and integrate” model: identify evidence-based genomic applications ready for implementation, inform stakeholders about these applications, and integrate these applications into public health at the local, state, and national level. This paper addresses current and future work at OPHG for integrating genomics into public health programs.

  13. HGVA: the Human Genome Variation Archive.

    Science.gov (United States)

    Lopez, Javier; Coll, Jacobo; Haimel, Matthias; Kandasamy, Swaathi; Tarraga, Joaquin; Furio-Tari, Pedro; Bari, Wasim; Bleda, Marta; Rueda, Antonio; Gräf, Stefan; Rendon, Augusto; Dopazo, Joaquin; Medina, Ignacio

    2017-07-03

    High-profile genomic variation projects like the 1000 Genomes project or the Exome Aggregation Consortium, are generating a wealth of human genomic variation knowledge which can be used as an essential reference for identifying disease-causing genotypes. However, accessing these data, contrasting the various studies and integrating those data in downstream analyses remains cumbersome. The Human Genome Variation Archive (HGVA) tackles these challenges and facilitates access to genomic data for key reference projects in a clean, fast and integrated fashion. HGVA provides an efficient and intuitive web-interface for easy data mining, a comprehensive RESTful API and client libraries in Python, Java and JavaScript for fast programmatic access to its knowledge base. HGVA calculates population frequencies for these projects and enriches their data with variant annotation provided by CellBase, a rich and fast annotation solution. HGVA serves as a proof-of-concept of the genome analysis developments being carried out by the University of Cambridge together with UK's 100 000 genomes project and the National Institute for Health Research BioResource Rare-Diseases, in particular, deploying open-source for Computational Biology (OpenCB) software platform for storing and analyzing massive genomic datasets. © The Author(s) 2017. Published by Oxford University Press on behalf of Nucleic Acids Research.

  14. Individual capacity for DNA repair and maintenance of genomic integrity: a fertile ground for studies in the field of assisted reproduction

    Directory of Open Access Journals (Sweden)

    Radoslava Vazharova

    2016-05-01

    Full Text Available Many factors may affect the chances for successful pregnancy, especially at a later age. Fertility evaluations including genetic analysis are recommended to couples that have not achieved pregnancy within 6–12 months of unprotected intercourse. This review discusses some of the common polymorphisms in genes coding for proteins functioning in DNA damage identification and repair and maintenance of genomic integrity that may affect the chances of success in natural conception as well as in assisted reproduction (AR. Common polymorphisms in genes coding for proteins functioning in DNA damage identification and repair and maintenance of genomic integrity may affect the chances of success in assisted reproduction as well as in natural conception. The effects of carriership of different alleles of key genes of DNA repair may have differential effects in men and women and at different ages, suggesting complex interactions with the mechanisms controlling cell and tissue aging and programmed cell death. Future studies in the field are needed in order to elucidate the genotype–phenotype relationships and to translate the knowledge about individual repair capacity and maintenance of genomic integrity to potential clinical applications. Abbreviations: aCGH: microarray-based comparative genomic hybridization; AR: assisted reproduction; ATM: ataxia-telangiectasia mutated; ATP: adenosine triphosphate; BER: base excision repair; BFE: basic fertility evaluation; DMSO: dimethyl sulfoxide; FSH: follicle-stimulating hormone; GNRHR: gonadotropin-releasing hormone receptor; HMG: high-mobility group; ICSI: intracytoplasmic sperm injection; IUI: intrauterine insemination; IVF: in vitro fertilization; LH: luteinizing hormone; LIF: leukaemia inhibitory factor; MTR: methionine synthase; MTRR: methionine synthase reductase; NGS: next-generation sequencing; NER: nucleotide excision repair; NHEJ: non-homologous end joining; PAH: polycyclic aromatic hydrocarbons; PCOS

  15. The genome sequence of the North-European cucumber (Cucumis sativus L.) unravels evolutionary adaptation mechanisms in plants.

    Science.gov (United States)

    Wóycicki, Rafał; Witkowicz, Justyna; Gawroński, Piotr; Dąbrowska, Joanna; Lomsadze, Alexandre; Pawełkowicz, Magdalena; Siedlecka, Ewa; Yagi, Kohei; Pląder, Wojciech; Seroczyńska, Anna; Śmiech, Mieczysław; Gutman, Wojciech; Niemirowicz-Szczytt, Katarzyna; Bartoszewski, Grzegorz; Tagashira, Norikazu; Hoshi, Yoshikazu; Borodovsky, Mark; Karpiński, Stanisław; Malepszy, Stefan; Przybecki, Zbigniew

    2011-01-01

    Cucumber (Cucumis sativus L.), a widely cultivated crop, has originated from Eastern Himalayas and secondary domestication regions includes highly divergent climate conditions e.g. temperate and subtropical. We wanted to uncover adaptive genome differences between the cucumber cultivars and what sort of evolutionary molecular mechanisms regulate genetic adaptation of plants to different ecosystems and organism biodiversity. Here we present the draft genome sequence of the Cucumis sativus genome of the North-European Borszczagowski cultivar (line B10) and comparative genomics studies with the known genomes of: C. sativus (Chinese cultivar--Chinese Long (line 9930)), Arabidopsis thaliana, Populus trichocarpa and Oryza sativa. Cucumber genomes show extensive chromosomal rearrangements, distinct differences in quantity of the particular genes (e.g. involved in photosynthesis, respiration, sugar metabolism, chlorophyll degradation, regulation of gene expression, photooxidative stress tolerance, higher non-optimal temperatures tolerance and ammonium ion assimilation) as well as in distributions of abscisic acid-, dehydration- and ethylene-responsive cis-regulatory elements (CREs) in promoters of orthologous group of genes, which lead to the specific adaptation features. Abscisic acid treatment of non-acclimated Arabidopsis and C. sativus seedlings induced moderate freezing tolerance in Arabidopsis but not in C. sativus. This experiment together with analysis of abscisic acid-specific CRE distributions give a clue why C. sativus is much more susceptible to moderate freezing stresses than A. thaliana. Comparative analysis of all the five genomes showed that, each species and/or cultivars has a specific profile of CRE content in promoters of orthologous genes. Our results constitute the substantial and original resource for the basic and applied research on environmental adaptations of plants, which could facilitate creation of new crops with improved growth and yield in

  16. Genome-wide investigation reveals high evolutionary rates in annual model plants.

    Science.gov (United States)

    Yue, Jia-Xing; Li, Jinpeng; Wang, Dan; Araki, Hitoshi; Tian, Dacheng; Yang, Sihai

    2010-11-09

    Rates of molecular evolution vary widely among species. While significant deviations from molecular clock have been found in many taxa, effects of life histories on molecular evolution are not fully understood. In plants, annual/perennial life history traits have long been suspected to influence the evolutionary rates at the molecular level. To date, however, the number of genes investigated on this subject is limited and the conclusions are mixed. To evaluate the possible heterogeneity in evolutionary rates between annual and perennial plants at the genomic level, we investigated 85 nuclear housekeeping genes, 10 non-housekeeping families, and 34 chloroplast genes using the genomic data from model plants including Arabidopsis thaliana and Medicago truncatula for annuals and grape (Vitis vinifera) and popular (Populus trichocarpa) for perennials. According to the cross-comparisons among the four species, 74-82% of the nuclear genes and 71-97% of the chloroplast genes suggested higher rates of molecular evolution in the two annuals than those in the two perennials. The significant heterogeneity in evolutionary rate between annuals and perennials was consistently found both in nonsynonymous sites and synonymous sites. While a linear correlation of evolutionary rates in orthologous genes between species was observed in nonsynonymous sites, the correlation was weak or invisible in synonymous sites. This tendency was clearer in nuclear genes than in chloroplast genes, in which the overall evolutionary rate was small. The slope of the regression line was consistently lower than unity, further confirming the higher evolutionary rate in annuals at the genomic level. The higher evolutionary rate in annuals than in perennials appears to be a universal phenomenon both in nuclear and chloroplast genomes in the four dicot model plants we investigated. Therefore, such heterogeneity in evolutionary rate should result from factors that have genome-wide influence, most likely those

  17. Integrative genome-wide expression profiling identifies three distinct molecular subgroups of renal cell carcinoma with different patient outcome

    Directory of Open Access Journals (Sweden)

    Beleut Manfred

    2012-07-01

    Full Text Available Abstract Background Renal cell carcinoma (RCC is characterized by a number of diverse molecular aberrations that differ among individuals. Recent approaches to molecularly classify RCC were based on clinical, pathological as well as on single molecular parameters. As a consequence, gene expression patterns reflecting the sum of genetic aberrations in individual tumors may not have been recognized. In an attempt to uncover such molecular features in RCC, we used a novel, unbiased and integrative approach. Methods We integrated gene expression data from 97 primary RCC of different pathologic parameters, 15 RCC metastases as well as 34 cancer cell lines for two-way nonsupervised hierarchical clustering using gene groups suggested by the PANTHER Classification System. We depicted the genomic landscape of the resulted tumor groups by means of Single Nuclear Polymorphism (SNP technology. Finally, the achieved results were immunohistochemically analyzed using a tissue microarray (TMA composed of 254 RCC. Results We found robust, genome wide expression signatures, which split RCC into three distinct molecular subgroups. These groups remained stable even if randomly selected gene sets were clustered. Notably, the pattern obtained from RCC cell lines was clearly distinguishable from that of primary tumors. SNP array analysis demonstrated differing frequencies of chromosomal copy number alterations among RCC subgroups. TMA analysis with group-specific markers showed a prognostic significance of the different groups. Conclusion We propose the existence of characteristic and histologically independent genome-wide expression outputs in RCC with potential biological and clinical relevance.

  18. Distinct high resolution genome profiles of early onset and late onset colorectal cancer integrated with gene expression data identify candidate susceptibility loci

    Directory of Open Access Journals (Sweden)

    Merok Marianne A

    2010-05-01

    Full Text Available Abstract Background Estimates suggest that up to 30% of colorectal cancers (CRC may develop due to an increased genetic risk. The mean age at diagnosis for CRC is about 70 years. Time of disease onset 20 years younger than the mean age is assumed to be indicative of genetic susceptibility. We have compared high resolution tumor genome copy number variation (CNV (Roche NimbleGen, 385 000 oligo CGH array in microsatellite stable (MSS tumors from two age groups, including 23 young at onset patients without known hereditary syndromes and with a median age of 44 years (range: 28-53 and 17 elderly patients with median age 79 years (range: 69-87. Our aim was to identify differences in the tumor genomes between these groups and pinpoint potential susceptibility loci. Integration analysis of CNV and genome wide mRNA expression data, available for the same tumors, was performed to identify a restricted candidate gene list. Results The total fraction of the genome with aberrant copy number, the overall genomic profile and the TP53 mutation spectrum were similar between the two age groups. However, both the number of chromosomal aberrations and the number of breakpoints differed significantly between the groups. Gains of 2q35, 10q21.3-22.1, 10q22.3 and 19q13.2-13.31 and losses from 1p31.3, 1q21.1, 2q21.2, 4p16.1-q28.3, 10p11.1 and 19p12, positions that in total contain more than 500 genes, were found significantly more often in the early onset group as compared to the late onset group. Integration analysis revealed a covariation of DNA copy number at these sites and mRNA expression for 107 of the genes. Seven of these genes, CLC, EIF4E, LTBP4, PLA2G12A, PPAT, RG9MTD2, and ZNF574, had significantly different mRNA expression comparing median expression levels across the transcriptome between the two groups. Conclusions Ten genomic loci, containing more than 500 protein coding genes, are identified as more often altered in tumors from early onset versus late

  19. Network analysis of epidermal growth factor signaling using integrated genomic, proteomic and phosphorylation data.

    Directory of Open Access Journals (Sweden)

    Katrina M Waters

    Full Text Available To understand how integration of multiple data types can help decipher cellular responses at the systems level, we analyzed the mitogenic response of human mammary epithelial cells to epidermal growth factor (EGF using whole genome microarrays, mass spectrometry-based proteomics and large-scale western blots with over 1000 antibodies. A time course analysis revealed significant differences in the expression of 3172 genes and 596 proteins, including protein phosphorylation changes measured by western blot. Integration of these disparate data types showed that each contributed qualitatively different components to the observed cell response to EGF and that varying degrees of concordance in gene expression and protein abundance measurements could be linked to specific biological processes. Networks inferred from individual data types were relatively limited, whereas networks derived from the integrated data recapitulated the known major cellular responses to EGF and exhibited more highly connected signaling nodes than networks derived from any individual dataset. While cell cycle regulatory pathways were altered as anticipated, we found the most robust response to mitogenic concentrations of EGF was induction of matrix metalloprotease cascades, highlighting the importance of the EGFR system as a regulator of the extracellular environment. These results demonstrate the value of integrating multiple levels of biological information to more accurately reconstruct networks of cellular response.

  20. Network Analysis of Epidermal Growth Factor Signaling using Integrated Genomic, Proteomic and Phosphorylation Data

    Energy Technology Data Exchange (ETDEWEB)

    Waters, Katrina M.; Liu, Tao; Quesenberry, Ryan D.; Willse, Alan R.; Bandyopadhyay, Somnath; Kathmann, Loel E.; Weber, Thomas J.; Smith, Richard D.; Wiley, H. S.; Thrall, Brian D.

    2012-03-29

    To understand how integration of multiple data types can help decipher cellular responses at the systems level, we analyzed the mitogenic response of human mammary epithelial cells to epidermal growth factor (EGF) using whole genome microarrays, mass spectrometry-based proteomics and large-scale western blots with over 1000 antibodies. A time course analysis revealed significant differences in the expression of 3172 genes and 596 proteins, including protein phosphorylation changes measured by western blot. Integration of these disparate data types showed that each contributed qualitatively different components to the observed cell response to EGF and that varying degrees of concordance in gene expression and protein abundance measurements could be linked to specific biological processes. Networks inferred from individual data types were relatively limited, whereas networks derived from the integrated data recapitulated the known major cellular responses to EGF and exhibited more highly connected signaling nodes than networks derived from any individual dataset. While cell cycle regulatory pathways were altered as anticipated, we found the most robust response to mitogenic concentrations of EGF was induction of matrix metalloprotease cascades, highlighting the importance of the EGFR system as a regulator of the extracellular environment. These results demonstrate the value of integrating multiple levels of biological information to more accurately reconstruct networks of cellular response.

  1. Collaborative Genomics Study Advances Precision Oncology

    Science.gov (United States)

    A collaborative study conducted by two Office of Cancer Genomics (OCG) initiatives highlights the importance of integrating structural and functional genomics programs to improve cancer therapies, and more specifically, contribute to precision oncology treatments for children.

  2. Zinc accumulation potential and toxicity threshold determined for a metal-accumulating Populus canescens clone in a dose-response study

    Energy Technology Data Exchange (ETDEWEB)

    Langer, Ingrid [Institute of Soil Science, University of Natural Resources and Applied Life Sciences, Peter Jordan-Strasse 82, A-1190 Vienna (Austria); Krpata, Doris [Institute of Microbiology, Innsbruck University, Technikerstrasse 25, A-6020 Innsbruck (Austria); Fitz, Walter J.; Wenzel, Walter W. [Institute of Soil Science, University of Natural Resources and Applied Life Sciences, Peter Jordan-Strasse 82, A-1190 Vienna (Austria); Schweiger, Peter F., E-mail: peter.schweiger@boku.ac.a [Institute of Soil Science, University of Natural Resources and Applied Life Sciences, Peter Jordan-Strasse 82, A-1190 Vienna (Austria)

    2009-10-15

    The effect of increasing soil Zn concentrations on growth and Zn tissue concentrations of a metal-accumulating aspen clone was examined in a dose-response study. Plants were grown in a soil with a low native Zn content which was spiked with Zn salt solutions and subsequently aged. Plant growth was not affected by NH{sub 4}NO{sub 3}-extractable soil Zn concentrations up to 60 mug Zn g{sup -1} soil, but it was completely inhibited at extractable concentrations above 90 mug Zn g{sup -1} soil. From these data an effective concentration of 68.5 mug extractable Zn g{sup -1} soil was calculated at which plant growth was reduced by 50%. The obtained information on toxicity threshold concentrations, and the relation between plant Zn accumulation and extractable soil Zn concentrations may be used to assess the suitability of the investigated Populus canescens clone for various phytoremediation strategies. The potential risk of metal transfer into food webs associated with P. canescens stands on Zn-polluted sites may also be estimated. - Quantitative information about the concentration-dependent Zn accumulation of Populus canescens contributes to assess its suitability for phytoremediation.

  3. Modelling the growth of Populus species using Ecosystem Demography (ED) model

    Science.gov (United States)

    Wang, D.; Lebauer, D. S.; Feng, X.; Dietze, M. C.

    2010-12-01

    Hybrid poplar plantations are an important source being evaluated for biomass production. Effective management of such plantations requires adequate growth and yield models. The Ecosystem Demography model (ED) makes predictions about the large scales of interest in above- and belowground ecosystem structure and the fluxes of carbon and water from a description of the fine-scale physiological processes. In this study, we used a workflow management tool, the Predictive Ecophysiological Carbon flux Analyzer (PECAn), to integrate literature data, field measurement and the ED model to provide predictions of ecosystem functioning. Parameters for the ED ensemble runs were sampled from the posterior distribution of ecophysiological traits of Populus species compiled from the literature using a Bayesian meta-analysis approach. Sensitivity analysis was performed to identify the parameters which contribute the most to the uncertainties of the ED model output. Model emulation techniques were used to update parameter posterior distributions using field-observed data in northern Wisconsin hybrid poplar plantations. Model results were evaluated with 5-year field-observed data in a hybrid poplar plantation at New Franklin, MO. ED was then used to predict the spatial variability of poplar yield in the coterminous United States (United States minus Alaska and Hawaii). Sensitivity analysis showed that root respiration, dark respiration, growth respiration, stomatal slope and specific leaf area contribute the most to the uncertainty, which suggests that our field measurements and data collection should focus on these parameters. The ED model successfully captured the inter-annual and spatial variability of the yield of poplar. Analyses in progress with the ED model focus on evaluating the ecosystem services of short-rotation woody plantations, such as impacts on soil carbon storage, water use, and nutrient retention.

  4. A study of polymerization of aspen (Populus) wood lipophilic extractives by SEC and Py-GC/MS

    CSIR Research Space (South Africa)

    Sithole, Bruce

    2013-03-01

    Full Text Available ) Orig inal manuscript received 13 June 2012, revision accepted 31 October 2012 Vol 66 No 1 January - March 2013 1 PEER REVIEWED A study of polymerization of aspen (Populus) wood lipophilic extractives by SEC and Py-GC/MS BRUCE SITHOLE1*, LUC... of polymerized wood resin that will be difficult to remove if present in pulp and paper products. On the other hand, these problems may be minor compared to using unseasoned wood. KEYWORDS: Aspen, extractives, polymerization, size exclusion chromatography, Py...

  5. Inferring genetic architecture of complex traits using Bayesian integrative analysis of genome and transcriptiome data

    DEFF Research Database (Denmark)

    Ehsani, Alireza; Sørensen, Peter; Pomp, Daniel

    2012-01-01

    Background To understand the genetic architecture of complex traits and bridge the genotype-phenotype gap, it is useful to study intermediate -omics data, e.g. the transcriptome. The present study introduces a method for simultaneous quantification of the contributions from single nucleotide......-modal distribution of genomic values collapses, when gene expressions are added to the model Conclusions With increased availability of various -omics data, integrative approaches are promising tools for understanding the genetic architecture of complex traits. Partitioning of explained variances at the chromosome...

  6. DFAST and DAGA: web-based integrated genome annotation tools and resources.

    Science.gov (United States)

    Tanizawa, Yasuhiro; Fujisawa, Takatomo; Kaminuma, Eli; Nakamura, Yasukazu; Arita, Masanori

    2016-01-01

    Quality assurance and correct taxonomic affiliation of data submitted to public sequence databases have been an everlasting problem. The DDBJ Fast Annotation and Submission Tool (DFAST) is a newly developed genome annotation pipeline with quality and taxonomy assessment tools. To enable annotation of ready-to-submit quality, we also constructed curated reference protein databases tailored for lactic acid bacteria. DFAST was developed so that all the procedures required for DDBJ submission could be done seamlessly online. The online workspace would be especially useful for users not familiar with bioinformatics skills. In addition, we have developed a genome repository, DFAST Archive of Genome Annotation (DAGA), which currently includes 1,421 genomes covering 179 species and 18 subspecies of two genera, Lactobacillus and Pediococcus , obtained from both DDBJ/ENA/GenBank and Sequence Read Archive (SRA). All the genomes deposited in DAGA were annotated consistently and assessed using DFAST. To assess the taxonomic position based on genomic sequence information, we used the average nucleotide identity (ANI), which showed high discriminative power to determine whether two given genomes belong to the same species. We corrected mislabeled or misidentified genomes in the public database and deposited the curated information in DAGA. The repository will improve the accessibility and reusability of genome resources for lactic acid bacteria. By exploiting the data deposited in DAGA, we found intraspecific subgroups in Lactobacillus gasseri and Lactobacillus jensenii , whose variation between subgroups is larger than the well-accepted ANI threshold of 95% to differentiate species. DFAST and DAGA are freely accessible at https://dfast.nig.ac.jp.

  7. Chromatin dynamics in genome stability

    DEFF Research Database (Denmark)

    Nair, Nidhi; Shoaib, Muhammad; Sørensen, Claus Storgaard

    2017-01-01

    Genomic DNA is compacted into chromatin through packaging with histone and non-histone proteins. Importantly, DNA accessibility is dynamically regulated to ensure genome stability. This is exemplified in the response to DNA damage where chromatin relaxation near genomic lesions serves to promote...... access of relevant enzymes to specific DNA regions for signaling and repair. Furthermore, recent data highlight genome maintenance roles of chromatin through the regulation of endogenous DNA-templated processes including transcription and replication. Here, we review research that shows the importance...... of chromatin structure regulation in maintaining genome integrity by multiple mechanisms including facilitating DNA repair and directly suppressing endogenous DNA damage....

  8. OryzaGenome: Genome Diversity Database of Wild Oryza Species

    KAUST Repository

    Ohyanagi, Hajime

    2015-11-18

    The species in the genus Oryza, encompassing nine genome types and 23 species, are a rich genetic resource and may have applications in deeper genomic analyses aiming to understand the evolution of plant genomes. With the advancement of next-generation sequencing (NGS) technology, a flood of Oryza species reference genomes and genomic variation information has become available in recent years. This genomic information, combined with the comprehensive phenotypic information that we are accumulating in our Oryzabase, can serve as an excellent genotype-phenotype association resource for analyzing rice functional and structural evolution, and the associated diversity of the Oryza genus. Here we integrate our previous and future phenotypic/habitat information and newly determined genotype information into a united repository, named OryzaGenome, providing the variant information with hyperlinks to Oryzabase. The current version of OryzaGenome includes genotype information of 446 O. rufipogon accessions derived by imputation and of 17 accessions derived by imputation-free deep sequencing. Two variant viewers are implemented: SNP Viewer as a conventional genome browser interface and Variant Table as a textbased browser for precise inspection of each variant one by one. Portable VCF (variant call format) file or tabdelimited file download is also available. Following these SNP (single nucleotide polymorphism) data, reference pseudomolecules/ scaffolds/contigs and genome-wide variation information for almost all of the closely and distantly related wild Oryza species from the NIG Wild Rice Collection will be available in future releases. All of the resources can be accessed through http://viewer.shigen.info/oryzagenome/.

  9. Genetic counselors' views and experiences with the clinical integration of genome sequencing.

    Science.gov (United States)

    Machini, Kalotina; Douglas, Jessica; Braxton, Alicia; Tsipis, Judith; Kramer, Kate

    2014-08-01

    In recent years, new sequencing technologies known as next generation sequencing (NGS) have provided scientists the ability to rapidly sequence all known coding as well as non-coding sequences in the human genome. As the two emerging approaches, whole exome (WES) and whole genome (WGS) sequencing, have started to be integrated in the clinical arena, we sought to survey health care professionals who are likely to be involved in the implementation process now and/or in the future (e.g., genetic counselors, geneticists and nurse practitioners). Two hundred twenty-one genetic counselors- one third of whom currently offer WES/WGS-participated in an anonymous online survey. The aims of the survey were first, to identify barriers to the implementation of WES/WGS, as perceived by survey participants; second, to provide the first systematic report of current practices regarding the integration of WES/WGS in clinic and/or research across the US and Canada and to illuminate the roles and challenges of genetic counselors participating in this process; and third to evaluate the impact of WES/WGS on patient care. Our results showed that genetic counseling practices with respect to WES/WGS are consistent with the criteria set forth in the ACMG 2012 policy statement, which highlights indications for testing, reporting, and pre/post test considerations. Our respondents described challenges related to offering WES/WGS, which included billing issues, the duration and content of the consent process, result interpretation and disclosure of incidental findings and variants of unknown significance. In addition, respondents indicated that specialty area (i.e., prenatal and cancer), lack of clinical utility of WES/WGS and concerns about interpretation of test results were factors that prevented them from offering this technology to patients. Finally, study participants identified the aspects of their professional training which have been most beneficial in aiding with the integration of

  10. Complete Genome Sequence of Germline Chromosomally Integrated Human Herpesvirus 6A and Analyses Integration Sites Define a New Human Endogenous Virus with Potential to Reactivate as an Emerging Infection.

    OpenAIRE

    Tweedy, J; Spyrou, MA; Pearson, M; Lassner, D; Kuhl, U; Gompels, UA

    2016-01-01

    Human herpesvirus-6A and B (HHV-6A, HHV-6B) have recently defined endogenous genomes, resulting from integration into the germline: chromosomally-integrated "CiHHV-6A/B". These affect approximately 1.0% of human populations, giving potential for virus gene expression in every cell. We previously showed that CiHHV-6A was more divergent than CiHHV-6B by examining four genes in 44 European CiHHV-6A/B cardiac/haematology patients. There was evidence for gene expression/reactivation, imp...

  11. Genome-wide distribution and organization of microsatellites in plants: an insight into marker development in Brachypodium.

    Directory of Open Access Journals (Sweden)

    Humira Sonah

    Full Text Available Plant genomes are complex and contain large amounts of repetitive DNA including microsatellites that are distributed across entire genomes. Whole genome sequences of several monocot and dicot plants that are available in the public domain provide an opportunity to study the origin, distribution and evolution of microsatellites, and also facilitate the development of new molecular markers. In the present investigation, a genome-wide analysis of microsatellite distribution in monocots (Brachypodium, sorghum and rice and dicots (Arabidopsis, Medicago and Populus was performed. A total of 797,863 simple sequence repeats (SSRs were identified in the whole genome sequences of six plant species. Characterization of these SSRs revealed that mono-nucleotide repeats were the most abundant repeats, and that the frequency of repeats decreased with increase in motif length both in monocots and dicots. However, the frequency of SSRs was higher in dicots than in monocots both for nuclear and chloroplast genomes. Interestingly, GC-rich repeats were the dominant repeats only in monocots, with the majority of them being present in the coding region. These coding GC-rich repeats were found to be involved in different biological processes, predominantly binding activities. In addition, a set of 22,879 SSR markers that were validated by e-PCR were developed and mapped on different chromosomes in Brachypodium for the first time, with a frequency of 101 SSR markers per Mb. Experimental validation of 55 markers showed successful amplification of 80% SSR markers in 16 Brachypodium accessions. An online database 'BraMi' (Brachypodium microsatellite markers of these genome-wide SSR markers was developed and made available in the public domain. The observed differential patterns of SSR marker distribution would be useful for studying microsatellite evolution in a monocot-dicot system. SSR markers developed in this study would be helpful for genomic studies in Brachypodium

  12. Comparative Genome Viewer

    International Nuclear Information System (INIS)

    Molineris, I.; Sales, G.

    2009-01-01

    The amount of information about genomes, both in the form of complete sequences and annotations, has been exponentially increasing in the last few years. As a result there is the need for tools providing a graphical representation of such information that should be comprehensive and intuitive. Visual representation is especially important in the comparative genomics field since it should provide a combined view of data belonging to different genomes. We believe that existing tools are limited in this respect as they focus on a single genome at a time (conservation histograms) or compress alignment representation to a single dimension. We have therefore developed a web-based tool called Comparative Genome Viewer (Cgv): it integrates a bidimensional representation of alignments between two regions, both at small and big scales, with the richness of annotations present in other genome browsers. We give access to our system through a web-based interface that provides the user with an interactive representation that can be updated in real time using the mouse to move from region to region and to zoom in on interesting details.

  13. CpGAVAS, an integrated web server for the annotation, visualization, analysis, and GenBank submission of completely sequenced chloroplast genome sequences

    Science.gov (United States)

    2012-01-01

    Background The complete sequences of chloroplast genomes provide wealthy information regarding the evolutionary history of species. With the advance of next-generation sequencing technology, the number of completely sequenced chloroplast genomes is expected to increase exponentially, powerful computational tools annotating the genome sequences are in urgent need. Results We have developed a web server CPGAVAS. The server accepts a complete chloroplast genome sequence as input. First, it predicts protein-coding and rRNA genes based on the identification and mapping of the most similar, full-length protein, cDNA and rRNA sequences by integrating results from Blastx, Blastn, protein2genome and est2genome programs. Second, tRNA genes and inverted repeats (IR) are identified using tRNAscan, ARAGORN and vmatch respectively. Third, it calculates the summary statistics for the annotated genome. Fourth, it generates a circular map ready for publication. Fifth, it can create a Sequin file for GenBank submission. Last, it allows the extractions of protein and mRNA sequences for given list of genes and species. The annotation results in GFF3 format can be edited using any compatible annotation editing tools. The edited annotations can then be uploaded to CPGAVAS for update and re-analyses repeatedly. Using known chloroplast genome sequences as test set, we show that CPGAVAS performs comparably to another application DOGMA, while having several superior functionalities. Conclusions CPGAVAS allows the semi-automatic and complete annotation of a chloroplast genome sequence, and the visualization, editing and analysis of the annotation results. It will become an indispensible tool for researchers studying chloroplast genomes. The software is freely accessible from http://www.herbalgenomics.org/cpgavas. PMID:23256920

  14. CpGAVAS, an integrated web server for the annotation, visualization, analysis, and GenBank submission of completely sequenced chloroplast genome sequences

    Directory of Open Access Journals (Sweden)

    Liu Chang

    2012-12-01

    Full Text Available Abstract Background The complete sequences of chloroplast genomes provide wealthy information regarding the evolutionary history of species. With the advance of next-generation sequencing technology, the number of completely sequenced chloroplast genomes is expected to increase exponentially, powerful computational tools annotating the genome sequences are in urgent need. Results We have developed a web server CPGAVAS. The server accepts a complete chloroplast genome sequence as input. First, it predicts protein-coding and rRNA genes based on the identification and mapping of the most similar, full-length protein, cDNA and rRNA sequences by integrating results from Blastx, Blastn, protein2genome and est2genome programs. Second, tRNA genes and inverted repeats (IR are identified using tRNAscan, ARAGORN and vmatch respectively. Third, it calculates the summary statistics for the annotated genome. Fourth, it generates a circular map ready for publication. Fifth, it can create a Sequin file for GenBank submission. Last, it allows the extractions of protein and mRNA sequences for given list of genes and species. The annotation results in GFF3 format can be edited using any compatible annotation editing tools. The edited annotations can then be uploaded to CPGAVAS for update and re-analyses repeatedly. Using known chloroplast genome sequences as test set, we show that CPGAVAS performs comparably to another application DOGMA, while having several superior functionalities. Conclusions CPGAVAS allows the semi-automatic and complete annotation of a chloroplast genome sequence, and the visualization, editing and analysis of the annotation results. It will become an indispensible tool for researchers studying chloroplast genomes. The software is freely accessible from http://www.herbalgenomics.org/cpgavas.

  15. The use of the white poplar (Populus alba L.) biomass as fuel

    Institute of Scientific and Technical Information of China (English)

    Tatiana Griu; Aurel Lunguleasa

    2016-01-01

    We determined the calorific value of white poplar (Populus alba L.) woody biomass to use it as fire-wood. The value of 19.133 MJ kg-1 obtained experimen-tally shows that the white poplar can be quite successfully used as firewood. Being of a lower quality in comparison with usual beech firewood, the white poplar has similar calorific value. The white poplar has a calorific density of 30.7%lower than that of current firewood. That is why the price of this firewood from white poplar is lower accord-ingly. Also, the prognosis of calorific value on the basis of the main chemical elements, being very close to the experimental value (?2.6%), indicates an appropriate value can be achieved to be used for investigation with the chemical element analysis.

  16. Bioinformatics decoding the genome

    CERN Multimedia

    CERN. Geneva; Deutsch, Sam; Michielin, Olivier; Thomas, Arthur; Descombes, Patrick

    2006-01-01

    Extracting the fundamental genomic sequence from the DNA From Genome to Sequence : Biology in the early 21st century has been radically transformed by the availability of the full genome sequences of an ever increasing number of life forms, from bacteria to major crop plants and to humans. The lecture will concentrate on the computational challenges associated with the production, storage and analysis of genome sequence data, with an emphasis on mammalian genomes. The quality and usability of genome sequences is increasingly conditioned by the careful integration of strategies for data collection and computational analysis, from the construction of maps and libraries to the assembly of raw data into sequence contigs and chromosome-sized scaffolds. Once the sequence is assembled, a major challenge is the mapping of biologically relevant information onto this sequence: promoters, introns and exons of protein-encoding genes, regulatory elements, functional RNAs, pseudogenes, transposons, etc. The methodological ...

  17. HHV-6A/B Integration and the Pathogenesis Associated with the Reactivation of Chromosomally Integrated HHV-6A/B.

    Science.gov (United States)

    Collin, Vanessa; Flamand, Louis

    2017-06-26

    Unlike other human herpesviruses, human herpesvirus 6A and 6B (HHV-6A/B) infection can lead to integration of the viral genome in human chromosomes. When integration occurs in germinal cells, the integrated HHV-6A/B genome can be transmitted to 50% of descendants. Such individuals, carrying one copy of the HHV-6A/B genome in every cell, are referred to as having inherited chromosomally-integrated HHV-6A/B (iciHHV-6) and represent approximately 1% of the world's population. Interestingly, HHV-6A/B integrate their genomes in a specific region of the chromosomes known as telomeres. Telomeres are located at chromosomes' ends and play essential roles in chromosomal stability and the long-term proliferative potential of cells. Considering that the integrated HHV-6A/B genome is mostly intact without any gross rearrangements or deletions, integration is likely used for viral maintenance into host cells. Knowing the roles played by telomeres in cellular homeostasis, viral integration in such structure is not likely to be without consequences. At present, the mechanisms and factors involved in HHV-6A/B integration remain poorly defined. In this review, we detail the potential biological and medical impacts of HHV-6A/B integration as well as the possible chromosomal integration and viral excision processes.

  18. In vivo genome editing via CRISPR/Cas9 mediated homology-independent targeted integration

    KAUST Repository

    Suzuki, Keiichiro

    2016-11-15

    Targeted genome editing via engineered nucleases is an exciting area of biomedical research and holds potential for clinical applications. Despite rapid advances in the field, in vivo targeted transgene integration is still infeasible because current tools are inefficient1, especially for non-dividing cells, which compose most adult tissues. This poses a barrier for uncovering fundamental biological principles and developing treatments for a broad range of genetic disorders2. Based on clustered regularly interspaced short palindromic repeat/Cas9 (CRISPR/Cas9)3, 4 technology, here we devise a homology-independent targeted integration (HITI) strategy, which allows for robust DNA knock-in in both dividing and non-dividing cells in vitro and, more importantly, in vivo (for example, in neurons of postnatal mammals). As a proof of concept of its therapeutic potential, we demonstrate the efficacy of HITI in improving visual function using a rat model of the retinal degeneration condition retinitis pigmentosa. The HITI method presented here establishes new avenues for basic research and targeted gene therapies.

  19. Papillomavirus genomes in human cervical carcinoma: Analysis of their integration and transcriptional activity

    International Nuclear Information System (INIS)

    Matulic, M.; Soric, J.

    1994-01-01

    Eighty-four biopsies derived from cervical tissues were analyzed for the presence of human papillomavirus (HPV) DNA types 6, 16 and 18 using Southern blot hybridization. HPV 6 was found in none of the cervical biopsies, and HPV types 16 and 18 were found in 44% of them. The rate of HPV 16/18 positive samples increased proportionally to the severity of the lesion. In normal tissue there were no positive samples, in mild and moderate dysplasia HPV 16/18 was present in 20% and in severe dysplasia and invasive carcinomas in 37 and 50%, respectively. In biopsies from 13 cases with squamous cell carcinoma of the uterine cervix and CIN III lesions HPV 16 was integrated within the host genome. It was concluded that the virus could be integrated at variable, presumably randomly selected chromosomal loci and with different number of copies. Transcription of HPV 16 and 18 was detected in one cervical cancer in HeLa cells, respectively. These results imply that HPV types 16 and 18 play an etiological role in the carcinogenesis of human cervical epithelial cells. (author)

  20. The Integrated Genomic Architecture and Evolution of Dental Divergence in East African Cichlid Fishes (Haplochromis chilotes x H. nyererei

    Directory of Open Access Journals (Sweden)

    C. Darrin Hulsey

    2017-09-01

    Full Text Available The independent evolution of the two toothed jaws of cichlid fishes is thought to have promoted their unparalleled ecological divergence and species richness. However, dental divergence in cichlids could exhibit substantial genetic covariance and this could dictate how traits like tooth numbers evolve in different African Lakes and on their two jaws. To test this hypothesis, we used a hybrid mapping cross of two trophically divergent Lake Victoria species (Haplochromis chilotes × Haplochromis nyererei to examine genomic regions associated with cichlid tooth diversity. Surprisingly, a similar genomic region was found to be associated with oral jaw tooth numbers in cichlids from both Lake Malawi and Lake Victoria. Likewise, this same genomic location was associated with variation in pharyngeal jaw tooth numbers. Similar relationships between tooth numbers on the two jaws in both our Victoria hybrid population and across the phylogenetic diversity of Malawi cichlids additionally suggests that tooth numbers on the two jaws of haplochromine cichlids might generally coevolve owing to shared genetic underpinnings. Integrated, rather than independent, genomic architectures could be key to the incomparable evolutionary divergence and convergence in cichlid tooth numbers.

  1. Personalized medicine: new genomics, old lessons

    OpenAIRE

    Offit, Kenneth

    2011-01-01

    Personalized medicine uses traditional, as well as emerging concepts of the genetic and environmental basis of disease to individualize prevention, diagnosis and treatment. Personalized genomics plays a vital, but not exclusive role in this evolving model of personalized medicine. The distinctions between genetic and genomic medicine are more quantitative than qualitative. Personalized genomics builds on principles established by the integration of genetics into medical practice. Principles s...

  2. Integrative Genomic Analysis of Cholangiocarcinoma Identifies Distinct IDH-Mutant Molecular Profiles

    Directory of Open Access Journals (Sweden)

    Farshad Farshidfar

    2017-03-01

    Full Text Available Cholangiocarcinoma (CCA is an aggressive malignancy of the bile ducts, with poor prognosis and limited treatment options. Here, we describe the integrated analysis of somatic mutations, RNA expression, copy number, and DNA methylation by The Cancer Genome Atlas of a set of predominantly intrahepatic CCA cases and propose a molecular classification scheme. We identified an IDH mutant-enriched subtype with distinct molecular features including low expression of chromatin modifiers, elevated expression of mitochondrial genes, and increased mitochondrial DNA copy number. Leveraging the multi-platform data, we observed that ARID1A exhibited DNA hypermethylation and decreased expression in the IDH mutant subtype. More broadly, we found that IDH mutations are associated with an expanded histological spectrum of liver tumors with molecular features that stratify with CCA. Our studies reveal insights into the molecular pathogenesis and heterogeneity of cholangiocarcinoma and provide classification information of potential therapeutic significance.

  3. Integrating population genetics and conservation biology in the era of genomics.

    Science.gov (United States)

    Ouborg, N Joop

    2010-02-23

    As one of the final activities of the ESF-CONGEN Networking programme, a conference entitled 'Integrating Population Genetics and Conservation Biology' was held at Trondheim, Norway, from 23 to 26 May 2009. Conference speakers and poster presenters gave a display of the state-of-the-art developments in the field of conservation genetics. Over the five-year running period of the successful ESF-CONGEN Networking programme, much progress has been made in theoretical approaches, basic research on inbreeding depression and other genetic processes associated with habitat fragmentation and conservation issues, and with applying principles of conservation genetics in the conservation of many species. Future perspectives were also discussed in the conference, and it was concluded that conservation genetics is evolving into conservation genomics, while at the same time basic and applied research on threatened species and populations from a population genetic point of view continues to be emphasized.

  4. Babelomics: an integrative platform for the analysis of transcriptomics, proteomics and genomic data with advanced functional profiling

    Science.gov (United States)

    Medina, Ignacio; Carbonell, José; Pulido, Luis; Madeira, Sara C.; Goetz, Stefan; Conesa, Ana; Tárraga, Joaquín; Pascual-Montano, Alberto; Nogales-Cadenas, Ruben; Santoyo, Javier; García, Francisco; Marbà, Martina; Montaner, David; Dopazo, Joaquín

    2010-01-01

    Babelomics is a response to the growing necessity of integrating and analyzing different types of genomic data in an environment that allows an easy functional interpretation of the results. Babelomics includes a complete suite of methods for the analysis of gene expression data that include normalization (covering most commercial platforms), pre-processing, differential gene expression (case-controls, multiclass, survival or continuous values), predictors, clustering; large-scale genotyping assays (case controls and TDTs, and allows population stratification analysis and correction). All these genomic data analysis facilities are integrated and connected to multiple options for the functional interpretation of the experiments. Different methods of functional enrichment or gene set enrichment can be used to understand the functional basis of the experiment analyzed. Many sources of biological information, which include functional (GO, KEGG, Biocarta, Reactome, etc.), regulatory (Transfac, Jaspar, ORegAnno, miRNAs, etc.), text-mining or protein–protein interaction modules can be used for this purpose. Finally a tool for the de novo functional annotation of sequences has been included in the system. This provides support for the functional analysis of non-model species. Mirrors of Babelomics or command line execution of their individual components are now possible. Babelomics is available at http://www.babelomics.org. PMID:20478823

  5. A novel bHLH transcription factor PebHLH35 from Populus euphratica confers drought tolerance through regulating stomatal development, photosynthesis and growth in Arabidopsis

    Energy Technology Data Exchange (ETDEWEB)

    Dong, Yan [College of Biological Sciences and Technology, National Engineering Laboratory for Tree Breeding, Beijing Forestry University, Beijing 100083 (China); Liaoning Forestry Vocational-Technical College, Shenyang 110101 (China); Wang, Congpeng; Han, Xiao; Tang, Sha; Liu, Sha [College of Biological Sciences and Technology, National Engineering Laboratory for Tree Breeding, Beijing Forestry University, Beijing 100083 (China); Xia, Xinli, E-mail: xiaxl@bjfu.edu.cn [College of Biological Sciences and Technology, National Engineering Laboratory for Tree Breeding, Beijing Forestry University, Beijing 100083 (China); Yin, Weilun, E-mail: yinwl@bjfu.edu.cn [College of Biological Sciences and Technology, National Engineering Laboratory for Tree Breeding, Beijing Forestry University, Beijing 100083 (China)

    2014-07-18

    Highlights: • PebHLH35 is firstly cloned from Populus euphratica and characterized its functions. • PebHLH35 is important for earlier seedling establishment and vegetative growth. • PebHLH35 enhances tolerance to drought by regulating growth. • PebHLH35 enhances tolerance to drought by regulating stomatal development. • PebHLH35 enhances tolerance to drought by regulating photosynthesis and transpiration. - Abstract: Plant basic helix-loop-helix (bHLH) transcription factors (TFs) are involved in a variety of physiological processes including the regulation of plant responses to various abiotic stresses. However, few drought-responsive bHLH family members in Populus have been reported. In this study, a novel bHLH gene (PebHLH35) was cloned from Populus euphratica. Expression analysis in P. euphratica revealed that PebHLH35 was induced by drought and abscisic acid. Subcellular localization studies using a PebHLH35-GFP fusion showed that the protein was localized to the nucleus. Ectopic overexpression of PebHLH35 in Arabidopsis resulted in a longer primary root, more leaves, and a greater leaf area under well-watered conditions compared with vector control plants. Notably, PebHLH35 overexpression lines showed enhanced tolerance to water-deficit stress. This finding was supported by anatomical and physiological analyses, which revealed a reduced stomatal density, stomatal aperture, transpiration rate, and water loss, and a higher chlorophyll content and photosynthetic rate. Our results suggest that PebHLH35 functions as a positive regulator of drought stress responses by regulating stomatal density, stomatal aperture, photosynthesis and growth.

  6. A novel bHLH transcription factor PebHLH35 from Populus euphratica confers drought tolerance through regulating stomatal development, photosynthesis and growth in Arabidopsis

    International Nuclear Information System (INIS)

    Dong, Yan; Wang, Congpeng; Han, Xiao; Tang, Sha; Liu, Sha; Xia, Xinli; Yin, Weilun

    2014-01-01

    Highlights: • PebHLH35 is firstly cloned from Populus euphratica and characterized its functions. • PebHLH35 is important for earlier seedling establishment and vegetative growth. • PebHLH35 enhances tolerance to drought by regulating growth. • PebHLH35 enhances tolerance to drought by regulating stomatal development. • PebHLH35 enhances tolerance to drought by regulating photosynthesis and transpiration. - Abstract: Plant basic helix-loop-helix (bHLH) transcription factors (TFs) are involved in a variety of physiological processes including the regulation of plant responses to various abiotic stresses. However, few drought-responsive bHLH family members in Populus have been reported. In this study, a novel bHLH gene (PebHLH35) was cloned from Populus euphratica. Expression analysis in P. euphratica revealed that PebHLH35 was induced by drought and abscisic acid. Subcellular localization studies using a PebHLH35-GFP fusion showed that the protein was localized to the nucleus. Ectopic overexpression of PebHLH35 in Arabidopsis resulted in a longer primary root, more leaves, and a greater leaf area under well-watered conditions compared with vector control plants. Notably, PebHLH35 overexpression lines showed enhanced tolerance to water-deficit stress. This finding was supported by anatomical and physiological analyses, which revealed a reduced stomatal density, stomatal aperture, transpiration rate, and water loss, and a higher chlorophyll content and photosynthetic rate. Our results suggest that PebHLH35 functions as a positive regulator of drought stress responses by regulating stomatal density, stomatal aperture, photosynthesis and growth

  7. Public trust and 'ethics review' as a commodity: the case of Genomics England Limited and the UK's 100,000 genomes project.

    Science.gov (United States)

    Samuel, Gabrielle Natalie; Farsides, Bobbie

    2018-06-01

    The UK Chief Medical Officer's 2016 Annual Report, Generation Genome, focused on a vision to fully integrate genomics into all aspects of the UK's National Health Service (NHS). This process of integration, which has now already begun, raises a wide range of social and ethical concerns, many of which were discussed in the final Chapter of the report. This paper explores how the UK's 100,000 Genomes Project (100 kGP)-the catalyst for Generation Genome, and for bringing genomics into the NHS-is negotiating these ethical concerns. The UK's 100 kGP, promoted and delivered by Genomics England Limited (GEL), is an innovative venture aiming to sequence 100,000 genomes from NHS patients who have a rare disease, cancer, or an infectious disease. GEL has emphasised the importance of ethical governance and decision-making. However, some sociological critique argues that biomedical/technological organisations presenting themselves as 'ethical' entities do not necessarily reflect a space within which moral thinking occurs. Rather, the 'ethical work' conducted (and displayed) by organisations is more strategic, relating to the politics of the organisation and the need to build public confidence. We set out to explore whether GEL's ethical framework was reflective of this critique, and what this tells us more broadly about how genomics is being integrated into the NHS in response to the ethical and social concerns raised in Generation Genome. We do this by drawing on a series of 20 interviews with individuals associated with or working at GEL.

  8. Ensembl Genomes 2013: scaling up access to genome-wide data.

    Science.gov (United States)

    Kersey, Paul Julian; Allen, James E; Christensen, Mikkel; Davis, Paul; Falin, Lee J; Grabmueller, Christoph; Hughes, Daniel Seth Toney; Humphrey, Jay; Kerhornou, Arnaud; Khobova, Julia; Langridge, Nicholas; McDowall, Mark D; Maheswari, Uma; Maslen, Gareth; Nuhn, Michael; Ong, Chuang Kee; Paulini, Michael; Pedro, Helder; Toneva, Iliana; Tuli, Mary Ann; Walts, Brandon; Williams, Gareth; Wilson, Derek; Youens-Clark, Ken; Monaco, Marcela K; Stein, Joshua; Wei, Xuehong; Ware, Doreen; Bolser, Daniel M; Howe, Kevin Lee; Kulesha, Eugene; Lawson, Daniel; Staines, Daniel Michael

    2014-01-01

    Ensembl Genomes (http://www.ensemblgenomes.org) is an integrating resource for genome-scale data from non-vertebrate species. The project exploits and extends technologies for genome annotation, analysis and dissemination, developed in the context of the vertebrate-focused Ensembl project, and provides a complementary set of resources for non-vertebrate species through a consistent set of programmatic and interactive interfaces. These provide access to data including reference sequence, gene models, transcriptional data, polymorphisms and comparative analysis. This article provides an update to the previous publications about the resource, with a focus on recent developments. These include the addition of important new genomes (and related data sets) including crop plants, vectors of human disease and eukaryotic pathogens. In addition, the resource has scaled up its representation of bacterial genomes, and now includes the genomes of over 9000 bacteria. Specific extensions to the web and programmatic interfaces have been developed to support users in navigating these large data sets. Looking forward, analytic tools to allow targeted selection of data for visualization and download are likely to become increasingly important in future as the number of available genomes increases within all domains of life, and some of the challenges faced in representing bacterial data are likely to become commonplace for eukaryotes in future.

  9. Variation of leaf margin serration in Populus nigra of industrial dumps

    Directory of Open Access Journals (Sweden)

    Yu. A. Shtirs

    2017-07-01

    Full Text Available The variability of leaf margin serration of Populus nigra L. in conditions of industrial dumps (coal mines dumps and overburden dumps and city park is estimated. The value of this indicator is in the range from 1.25 to 1.76 and significantly increases along the gradient: coal mines dumps – overburden dumps – city park. From the number of selected gradations of P. nigra leaf blades, the gradation with values of 1.45-1.55 is most pronounced according to the analyzed index for industrial dumps, for the park – with the values of 1.55-1.65. The degree of serration of edge leaf blade is characterized by low values of variation – coefficient of variation is less than 10.0%. We registered the significant positive correlation between the average values of leaf margin serration and the length of P. nigra leaf blade.

  10. Natural hybridization between Populus nigra L. and P. x canadensis Moench. Hybrid offspring competes for niches along the Rhine river in the Netherlands

    NARCIS (Netherlands)

    Smulders, M.J.M.; Beringen, R.; Volosyanchuk, R.; Vanden Broeck, A.; Schoot, van der J.; Arens, P.F.P.; Vosman, B.

    2008-01-01

    Black poplar (Populus nigra L.) is a major species for European riparian forests but its abundance has decreased over the decades due to human influences. For restoration of floodplain woodlands, the remaining black poplar stands may act as source population. A potential problem is that P. nigra and

  11. Integrated high-resolution array CGH and SKY analysis of homozygous deletions and other genomic alterations present in malignant mesothelioma cell lines.

    Science.gov (United States)

    Klorin, Geula; Rozenblum, Ester; Glebov, Oleg; Walker, Robert L; Park, Yoonsoo; Meltzer, Paul S; Kirsch, Ilan R; Kaye, Frederic J; Roschke, Anna V

    2013-05-01

    High-resolution oligonucleotide array comparative genomic hybridization (aCGH) and spectral karyotyping (SKY) were applied to a panel of malignant mesothelioma (MMt) cell lines. SKY has not been applied to MMt before, and complete karyotypes are reported based on the integration of SKY and aCGH results. A whole genome search for homozygous deletions (HDs) produced the largest set of recurrent and non-recurrent HDs for MMt (52 recurrent HDs in 10 genomic regions; 36 non-recurrent HDs). For the first time, LINGO2, RBFOX1/A2BP1, RPL29, DUSP7, and CCSER1/FAM190A were found to be homozygously deleted in MMt, and some of these genes could be new tumor suppressor genes for MMt. Integration of SKY and aCGH data allowed reconstruction of chromosomal rearrangements that led to the formation of HDs. Our data imply that only with acquisition of structural and/or numerical karyotypic instability can MMt cells attain a complete loss of tumor suppressor genes located in 9p21.3, which is the most frequently homozygously deleted region. Tetraploidization is a late event in the karyotypic progression of MMt cells, after HDs in the 9p21.3 region have already been acquired. Published by Elsevier Inc.

  12. Germline transgenic pigs by Sleeping Beauty transposition in porcine zygotes and targeted integration in the pig genome.

    Directory of Open Access Journals (Sweden)

    Wiebke Garrels

    Full Text Available Genetic engineering can expand the utility of pigs for modeling human diseases, and for developing advanced therapeutic approaches. However, the inefficient production of transgenic pigs represents a technological bottleneck. Here, we assessed the hyperactive Sleeping Beauty (SB100X transposon system for enzyme-catalyzed transgene integration into the embryonic porcine genome. The components of the transposon vector system were microinjected as circular plasmids into the cytoplasm of porcine zygotes, resulting in high frequencies of transgenic fetuses and piglets. The transgenic animals showed normal development and persistent reporter gene expression for >12 months. Molecular hallmarks of transposition were confirmed by analysis of 25 genomic insertion sites. We demonstrate germ-line transmission, segregation of individual transposons, and continued, copy number-dependent transgene expression in F1-offspring. In addition, we demonstrate target-selected gene insertion into transposon-tagged genomic loci by Cre-loxP-based cassette exchange in somatic cells followed by nuclear transfer. Transposase-catalyzed transgenesis in a large mammalian species expands the arsenal of transgenic technologies for use in domestic animals and will facilitate the development of large animal models for human diseases.

  13. MIPS PlantsDB: a database framework for comparative plant genome research.

    Science.gov (United States)

    Nussbaumer, Thomas; Martis, Mihaela M; Roessner, Stephan K; Pfeifer, Matthias; Bader, Kai C; Sharma, Sapna; Gundlach, Heidrun; Spannagl, Manuel

    2013-01-01

    The rapidly increasing amount of plant genome (sequence) data enables powerful comparative analyses and integrative approaches and also requires structured and comprehensive information resources. Databases are needed for both model and crop plant organisms and both intuitive search/browse views and comparative genomics tools should communicate the data to researchers and help them interpret it. MIPS PlantsDB (http://mips.helmholtz-muenchen.de/plant/genomes.jsp) was initially described in NAR in 2007 [Spannagl,M., Noubibou,O., Haase,D., Yang,L., Gundlach,H., Hindemitt, T., Klee,K., Haberer,G., Schoof,H. and Mayer,K.F. (2007) MIPSPlantsDB-plant database resource for integrative and comparative plant genome research. Nucleic Acids Res., 35, D834-D840] and was set up from the start to provide data and information resources for individual plant species as well as a framework for integrative and comparative plant genome research. PlantsDB comprises database instances for tomato, Medicago, Arabidopsis, Brachypodium, Sorghum, maize, rice, barley and wheat. Building up on that, state-of-the-art comparative genomics tools such as CrowsNest are integrated to visualize and investigate syntenic relationships between monocot genomes. Results from novel genome analysis strategies targeting the complex and repetitive genomes of triticeae species (wheat and barley) are provided and cross-linked with model species. The MIPS Repeat Element Database (mips-REdat) and Catalog (mips-REcat) as well as tight connections to other databases, e.g. via web services, are further important components of PlantsDB.

  14. Genome Sequence of the Plant Growth Promoting Endophytic Bacterium Enterobacter sp. 638

    Science.gov (United States)

    Taghavi, Safiyh; van der Lelie, Daniel; Hoffman, Adam; Zhang, Yian-Biao; Walla, Michael D.; Vangronsveld, Jaco; Newman, Lee; Monchy, Sébastien

    2010-01-01

    Enterobacter sp. 638 is an endophytic plant growth promoting gamma-proteobacterium that was isolated from the stem of poplar (Populus trichocarpa×deltoides cv. H11-11), a potentially important biofuel feed stock plant. The Enterobacter sp. 638 genome sequence reveals the presence of a 4,518,712 bp chromosome and a 157,749 bp plasmid (pENT638-1). Genome annotation and comparative genomics allowed the identification of an extended set of genes specific to the plant niche adaptation of this bacterium. This includes genes that code for putative proteins involved in survival in the rhizosphere (to cope with oxidative stress or uptake of nutrients released by plant roots), root adhesion (pili, adhesion, hemagglutinin, cellulose biosynthesis), colonization/establishment inside the plant (chemiotaxis, flagella, cellobiose phosphorylase), plant protection against fungal and bacterial infections (siderophore production and synthesis of the antimicrobial compounds 4-hydroxybenzoate and 2-phenylethanol), and improved poplar growth and development through the production of the phytohormones indole acetic acid, acetoin, and 2,3-butanediol. Metabolite analysis confirmed by quantitative RT–PCR showed that, the production of acetoin and 2,3-butanediol is induced by the presence of sucrose in the growth medium. Interestingly, both the genetic determinants required for sucrose metabolism and the synthesis of acetoin and 2,3-butanediol are clustered on a genomic island. These findings point to a close interaction between Enterobacter sp. 638 and its poplar host, where the availability of sucrose, a major plant sugar, affects the synthesis of plant growth promoting phytohormones by the endophytic bacterium. The availability of the genome sequence, combined with metabolome and transcriptome analysis, will provide a better understanding of the synergistic interactions between poplar and its growth promoting endophyte Enterobacter sp. 638. This information can be further exploited to

  15. New approaches to assessing the effects of mutagenic agents on the integrity of the human genome

    International Nuclear Information System (INIS)

    Elespuru, R.K.; Sankaranarayanan, K.

    2007-01-01

    Heritable genetic alterations, although individually rare, have a substantial collective health impact. Approximately 20% of these are new mutations of unknown cause. Assessment of the effect of exposures to DNA damaging agents, i.e. mutagenic chemicals and radiations, on the integrity of the human genome and on the occurrence of genetic disease remains a daunting challenge. Recent insights may explain why previous examination of human exposures to ionizing radiation, as in Hiroshima and Nagasaki, failed to reveal heritable genetic effects. New opportunities to assess the heritable genetic damaging effects of environmental mutagens are afforded by: (1) integration of knowledge on the molecular nature of genetic disorders and the molecular effects of mutagens; (2) the development of more practical assays for germline mutagenesis; (3) the likely use of population-based genetic screening in personalized medicine

  16. Developmental and environmental regulation of Aquaporin gene expression across Populus species: divergence or redundancy?

    Science.gov (United States)

    Cohen, David; Bogeat-Triboulot, Marie-Béatrice; Vialet-Chabrand, Silvère; Merret, Rémy; Courty, Pierre-Emmanuel; Moretti, Sébastien; Bizet, François; Guilliot, Agnès; Hummel, Irène

    2013-01-01

    Aquaporins (AQPs) are membrane channels belonging to the major intrinsic proteins family and are known for their ability to facilitate water movement. While in Populus trichocarpa, AQP proteins form a large family encompassing fifty-five genes, most of the experimental work focused on a few genes or subfamilies. The current work was undertaken to develop a comprehensive picture of the whole AQP gene family in Populus species by delineating gene expression domain and distinguishing responsiveness to developmental and environmental cues. Since duplication events amplified the poplar AQP family, we addressed the question of expression redundancy between gene duplicates. On these purposes, we carried a meta-analysis of all publicly available Affymetrix experiments. Our in-silico strategy controlled for previously identified biases in cross-species transcriptomics, a necessary step for any comparative transcriptomics based on multispecies design chips. Three poplar AQPs were not supported by any expression data, even in a large collection of situations (abiotic and biotic constraints, temporal oscillations and mutants). The expression of 11 AQPs was never or poorly regulated whatever the wideness of their expression domain and their expression level. Our work highlighted that PtTIP1;4 was the most responsive gene of the AQP family. A high functional divergence between gene duplicates was detected across species and in response to tested cues, except for the root-expressed PtTIP2;3/PtTIP2;4 pair exhibiting 80% convergent responses. Our meta-analysis assessed key features of aquaporin expression which had remained hidden in single experiments, such as expression wideness, response specificity and genotype and environment interactions. By consolidating expression profiles using independent experimental series, we showed that the large expansion of AQP family in poplar was accompanied with a strong divergence of gene expression, even if some cases of functional redundancy

  17. Developmental and environmental regulation of Aquaporin gene expression across Populus species: divergence or redundancy?

    Directory of Open Access Journals (Sweden)

    David Cohen

    Full Text Available Aquaporins (AQPs are membrane channels belonging to the major intrinsic proteins family and are known for their ability to facilitate water movement. While in Populus trichocarpa, AQP proteins form a large family encompassing fifty-five genes, most of the experimental work focused on a few genes or subfamilies. The current work was undertaken to develop a comprehensive picture of the whole AQP gene family in Populus species by delineating gene expression domain and distinguishing responsiveness to developmental and environmental cues. Since duplication events amplified the poplar AQP family, we addressed the question of expression redundancy between gene duplicates. On these purposes, we carried a meta-analysis of all publicly available Affymetrix experiments. Our in-silico strategy controlled for previously identified biases in cross-species transcriptomics, a necessary step for any comparative transcriptomics based on multispecies design chips. Three poplar AQPs were not supported by any expression data, even in a large collection of situations (abiotic and biotic constraints, temporal oscillations and mutants. The expression of 11 AQPs was never or poorly regulated whatever the wideness of their expression domain and their expression level. Our work highlighted that PtTIP1;4 was the most responsive gene of the AQP family. A high functional divergence between gene duplicates was detected across species and in response to tested cues, except for the root-expressed PtTIP2;3/PtTIP2;4 pair exhibiting 80% convergent responses. Our meta-analysis assessed key features of aquaporin expression which had remained hidden in single experiments, such as expression wideness, response specificity and genotype and environment interactions. By consolidating expression profiles using independent experimental series, we showed that the large expansion of AQP family in poplar was accompanied with a strong divergence of gene expression, even if some cases of

  18. A scored human protein-protein interaction network to catalyze genomic interpretation

    DEFF Research Database (Denmark)

    Li, Taibo; Wernersson, Rasmus; Hansen, Rasmus B

    2017-01-01

    Genome-scale human protein-protein interaction networks are critical to understanding cell biology and interpreting genomic data, but challenging to produce experimentally. Through data integration and quality control, we provide a scored human protein-protein interaction network (InWeb_InBioMap,......Genome-scale human protein-protein interaction networks are critical to understanding cell biology and interpreting genomic data, but challenging to produce experimentally. Through data integration and quality control, we provide a scored human protein-protein interaction network (In...

  19. Stem wood properties of Populus tremuloides, Betula papyrifera and Acer saccharum saplings after three years of treatments to elevated carbon dioxide and ozone

    Science.gov (United States)

    Seija Kaakinen; Katri Kostiainen; Fredrik Ek; Pekka Saranpaa; Mark E. Kubiske; Jaak Sober; David F. Karnosky; Elina Vapaavuori

    2004-01-01

    The aim of this study was to examine the effects of elevated carbon dioxide [CO2] and ozone [O3] and their interaction on wood chemistry and anatomy of five clones of 3-year-old trembling aspen (Populus tremuloides Michx.). Wood chemistry was studied also on paper birch (Betula papyrifera...

  20. Big Data Analysis of Human Genome Variations

    KAUST Repository

    Gojobori, Takashi

    2016-01-25

    Since the human genome draft sequence was in public for the first time in 2000, genomic analyses have been intensively extended to the population level. The following three international projects are good examples for large-scale studies of human genome variations: 1) HapMap Data (1,417 individuals) (http://hapmap.ncbi.nlm.nih.gov/downloads/genotypes/2010-08_phaseII+III/forward/), 2) HGDP (Human Genome Diversity Project) Data (940 individuals) (http://www.hagsc.org/hgdp/files.html), 3) 1000 genomes Data (2,504 individuals) http://ftp.1000genomes.ebi.ac.uk/vol1/ftp/release/20130502/ If we can integrate all three data into a single volume of data, we should be able to conduct a more detailed analysis of human genome variations for a total number of 4,861 individuals (= 1,417+940+2,504 individuals). In fact, we successfully integrated these three data sets by use of information on the reference human genome sequence, and we conducted the big data analysis. In particular, we constructed a phylogenetic tree of about 5,000 human individuals at the genome level. As a result, we were able to identify clusters of ethnic groups, with detectable admixture, that were not possible by an analysis of each of the three data sets. Here, we report the outcome of this kind of big data analyses and discuss evolutionary significance of human genomic variations. Note that the present study was conducted in collaboration with Katsuhiko Mineta and Kosuke Goto at KAUST.

  1. GI-POP: a combinational annotation and genomic island prediction pipeline for ongoing microbial genome projects.

    Science.gov (United States)

    Lee, Chi-Ching; Chen, Yi-Ping Phoebe; Yao, Tzu-Jung; Ma, Cheng-Yu; Lo, Wei-Cheng; Lyu, Ping-Chiang; Tang, Chuan Yi

    2013-04-10

    Sequencing of microbial genomes is important because of microbial-carrying antibiotic and pathogenetic activities. However, even with the help of new assembling software, finishing a whole genome is a time-consuming task. In most bacteria, pathogenetic or antibiotic genes are carried in genomic islands. Therefore, a quick genomic island (GI) prediction method is useful for ongoing sequencing genomes. In this work, we built a Web server called GI-POP (http://gipop.life.nthu.edu.tw) which integrates a sequence assembling tool, a functional annotation pipeline, and a high-performance GI predicting module, in a support vector machine (SVM)-based method called genomic island genomic profile scanning (GI-GPS). The draft genomes of the ongoing genome projects in contigs or scaffolds can be submitted to our Web server, and it provides the functional annotation and highly probable GI-predicting results. GI-POP is a comprehensive annotation Web server designed for ongoing genome project analysis. Researchers can perform annotation and obtain pre-analytic information include possible GIs, coding/non-coding sequences and functional analysis from their draft genomes. This pre-analytic system can provide useful information for finishing a genome sequencing project. Copyright © 2012 Elsevier B.V. All rights reserved.

  2. Convergent functional genomics of psychiatric disorders.

    Science.gov (United States)

    Niculescu, Alexander B

    2013-10-01

    Genetic and gene expression studies, in humans and animal models of psychiatric and other medical disorders, are becoming increasingly integrated. Particularly for genomics, the convergence and integration of data across species, experimental modalities and technical platforms is providing a fit-to-disease way of extracting reproducible and biologically important signal, in contrast to the fit-to-cohort effect and limited reproducibility of human genetic analyses alone. With the advent of whole-genome sequencing and the realization that a major portion of the non-coding genome may contain regulatory variants, Convergent Functional Genomics (CFG) approaches are going to be essential to identify disease-relevant signal from the tremendous polymorphic variation present in the general population. Such work in psychiatry can provide an example of how to address other genetically complex disorders, and in turn will benefit by incorporating concepts from other areas, such as cancer, cardiovascular diseases, and diabetes. © 2013 Wiley Periodicals, Inc.

  3. Determination of Fe, Hg, Mn, and Pb in three-rings of poplar (Populus alba L.) by U-shaped DC arc

    Science.gov (United States)

    Marković, D. M.; Novović, I.; Vilotić, D.; Ignjatović, Lj.

    2007-09-01

    The U-shaped DC arc with aerosol supply was applied for the determination of Fe, Hg, Mn, and Pb in poplar (Populus alba L.) tree-rings. By optimization of the operating parameters and by selection of the most appropriate signal integration time (20 s for Fe, Mn, and Pb and 30 s for Hg), the obtained limits of detection for Fe, Hg, Mn, and Pb are 5.8, 2.6, 1.6, and 2.0 ng/ml, respectively. The detection limits achieved by this method for Fe, Hg, Mn, and Pb are comparable with the detection limits obtained for these elements by such methods as inductively coupled plasma-atomic emission spectrometry (ICP-AES), direct coupled plasmatomic emission spectrometry (DCP-AES), and microwave-induced plasma-atomic emission spectrometry (MIP-AES). We used the tree-rings of poplar from two different locations. The first one is in the area close to the power plant “Nikola Tesla” TENT A, Obrenovac, while the other one is in the urban area of Novi Sad. In almost all cases, samples from the location at Obrenovac registered elevated average concentrations of Fe, Hg, Mn, and Pb in the tree-rings of poplar.

  4. Integration of genome-wide association studies with biological knowledge identifies six novel genes related to kidney function.

    Science.gov (United States)

    Chasman, Daniel I; Fuchsberger, Christian; Pattaro, Cristian; Teumer, Alexander; Böger, Carsten A; Endlich, Karlhans; Olden, Matthias; Chen, Ming-Huei; Tin, Adrienne; Taliun, Daniel; Li, Man; Gao, Xiaoyi; Gorski, Mathias; Yang, Qiong; Hundertmark, Claudia; Foster, Meredith C; O'Seaghdha, Conall M; Glazer, Nicole; Isaacs, Aaron; Liu, Ching-Ti; Smith, Albert V; O'Connell, Jeffrey R; Struchalin, Maksim; Tanaka, Toshiko; Li, Guo; Johnson, Andrew D; Gierman, Hinco J; Feitosa, Mary F; Hwang, Shih-Jen; Atkinson, Elizabeth J; Lohman, Kurt; Cornelis, Marilyn C; Johansson, Asa; Tönjes, Anke; Dehghan, Abbas; Lambert, Jean-Charles; Holliday, Elizabeth G; Sorice, Rossella; Kutalik, Zoltan; Lehtimäki, Terho; Esko, Tõnu; Deshmukh, Harshal; Ulivi, Sheila; Chu, Audrey Y; Murgia, Federico; Trompet, Stella; Imboden, Medea; Coassin, Stefan; Pistis, Giorgio; Harris, Tamara B; Launer, Lenore J; Aspelund, Thor; Eiriksdottir, Gudny; Mitchell, Braxton D; Boerwinkle, Eric; Schmidt, Helena; Cavalieri, Margherita; Rao, Madhumathi; Hu, Frank; Demirkan, Ayse; Oostra, Ben A; de Andrade, Mariza; Turner, Stephen T; Ding, Jingzhong; Andrews, Jeanette S; Freedman, Barry I; Giulianini, Franco; Koenig, Wolfgang; Illig, Thomas; Meisinger, Christa; Gieger, Christian; Zgaga, Lina; Zemunik, Tatijana; Boban, Mladen; Minelli, Cosetta; Wheeler, Heather E; Igl, Wilmar; Zaboli, Ghazal; Wild, Sarah H; Wright, Alan F; Campbell, Harry; Ellinghaus, David; Nöthlings, Ute; Jacobs, Gunnar; Biffar, Reiner; Ernst, Florian; Homuth, Georg; Kroemer, Heyo K; Nauck, Matthias; Stracke, Sylvia; Völker, Uwe; Völzke, Henry; Kovacs, Peter; Stumvoll, Michael; Mägi, Reedik; Hofman, Albert; Uitterlinden, Andre G; Rivadeneira, Fernando; Aulchenko, Yurii S; Polasek, Ozren; Hastie, Nick; Vitart, Veronique; Helmer, Catherine; Wang, Jie Jin; Stengel, Bénédicte; Ruggiero, Daniela; Bergmann, Sven; Kähönen, Mika; Viikari, Jorma; Nikopensius, Tiit; Province, Michael; Ketkar, Shamika; Colhoun, Helen; Doney, Alex; Robino, Antonietta; Krämer, Bernhard K; Portas, Laura; Ford, Ian; Buckley, Brendan M; Adam, Martin; Thun, Gian-Andri; Paulweber, Bernhard; Haun, Margot; Sala, Cinzia; Mitchell, Paul; Ciullo, Marina; Kim, Stuart K; Vollenweider, Peter; Raitakari, Olli; Metspalu, Andres; Palmer, Colin; Gasparini, Paolo; Pirastu, Mario; Jukema, J Wouter; Probst-Hensch, Nicole M; Kronenberg, Florian; Toniolo, Daniela; Gudnason, Vilmundur; Shuldiner, Alan R; Coresh, Josef; Schmidt, Reinhold; Ferrucci, Luigi; Siscovick, David S; van Duijn, Cornelia M; Borecki, Ingrid B; Kardia, Sharon L R; Liu, Yongmei; Curhan, Gary C; Rudan, Igor; Gyllensten, Ulf; Wilson, James F; Franke, Andre; Pramstaller, Peter P; Rettig, Rainer; Prokopenko, Inga; Witteman, Jacqueline; Hayward, Caroline; Ridker, Paul M; Parsa, Afshin; Bochud, Murielle; Heid, Iris M; Kao, W H Linda; Fox, Caroline S; Köttgen, Anna

    2012-12-15

    In conducting genome-wide association studies (GWAS), analytical approaches leveraging biological information may further understanding of the pathophysiology of clinical traits. To discover novel associations with estimated glomerular filtration rate (eGFR), a measure of kidney function, we developed a strategy for integrating prior biological knowledge into the existing GWAS data for eGFR from the CKDGen Consortium. Our strategy focuses on single nucleotide polymorphism (SNPs) in genes that are connected by functional evidence, determined by literature mining and gene ontology (GO) hierarchies, to genes near previously validated eGFR associations. It then requires association thresholds consistent with multiple testing, and finally evaluates novel candidates by independent replication. Among the samples of European ancestry, we identified a genome-wide significant SNP in FBXL20 (P = 5.6 × 10(-9)) in meta-analysis of all available data, and additional SNPs at the INHBC, LRP2, PLEKHA1, SLC3A2 and SLC7A6 genes meeting multiple-testing corrected significance for replication and overall P-values of 4.5 × 10(-4)-2.2 × 10(-7). Neither the novel PLEKHA1 nor FBXL20 associations, both further supported by association with eGFR among African Americans and with transcript abundance, would have been implicated by eGFR candidate gene approaches. LRP2, encoding the megalin receptor, was identified through connection with the previously known eGFR gene DAB2 and extends understanding of the megalin system in kidney function. These findings highlight integration of existing genome-wide association data with independent biological knowledge to uncover novel candidate eGFR associations, including candidates lacking known connections to kidney-specific pathways. The strategy may also be applicable to other clinical phenotypes, although more testing will be needed to assess its potential for discovery in general.

  5. Structure and diversity of ground mesofauna inUlmus and Populus consortia in the industrial areas of mining and smelting complex of krivyi rig basin

    Directory of Open Access Journals (Sweden)

    V. V. Kachinskaya

    2010-05-01

    Full Text Available The structure and biological diversity of ground mesofauna on a consortium level of organisation of ecosystems are considered. Indicators of structural organisation and biodiversity of ground mesofauna were analised in Ulmus and Populus consortia in the conditions of industrial territories of mining and smelting complex of Krivyi Rig Basin. It is established that taxonomical structure of ground mesofauna is characterised by insignificant number and quantity of taxonomical groups. Prevalence of hortobionts and herpetobionts in morpho-ecological structure of the community testifies to their attachment to consortium’s determinants and influence of steppe climate on its structure. Dominance of phytophages and polyphages in trophic structure is caused by a combination of consortium determinants specificity and «a zone source» of the fauna formations. The structural organisation of ground mesofauna in consortia of Ulmus and Populus in the conditions of industrial sites is characterised by simplified taxonomical structure with low biodiversity at all levels.

  6. Integrated Genomic Characterization of Papillary Thyroid Carcinoma

    Science.gov (United States)

    Agrawal, Nishant; Akbani, Rehan; Aksoy, B. Arman; Ally, Adrian; Arachchi, Harindra; Asa, Sylvia L.; Auman, J. Todd; Balasundaram, Miruna; Balu, Saianand; Baylin, Stephen B.; Behera, Madhusmita; Bernard, Brady; Beroukhim, Rameen; Bishop, Justin A.; Black, Aaron D.; Bodenheimer, Tom; Boice, Lori; Bootwalla, Moiz S.; Bowen, Jay; Bowlby, Reanne; Bristow, Christopher A.; Brookens, Robin; Brooks, Denise; Bryant, Robert; Buda, Elizabeth; Butterfield, Yaron S.N.; Carling, Tobias; Carlsen, Rebecca; Carter, Scott L.; Carty, Sally E.; Chan, Timothy A.; Chen, Amy Y.; Cherniack, Andrew D.; Cheung, Dorothy; Chin, Lynda; Cho, Juok; Chu, Andy; Chuah, Eric; Cibulskis, Kristian; Ciriello, Giovanni; Clarke, Amanda; Clayman, Gary L.; Cope, Leslie; Copland, John; Covington, Kyle; Danilova, Ludmila; Davidsen, Tanja; Demchok, John A.; DiCara, Daniel; Dhalla, Noreen; Dhir, Rajiv; Dookran, Sheliann S.; Dresdner, Gideon; Eldridge, Jonathan; Eley, Greg; El-Naggar, Adel K.; Eng, Stephanie; Fagin, James A.; Fennell, Timothy; Ferris, Robert L.; Fisher, Sheila; Frazer, Scott; Frick, Jessica; Gabriel, Stacey B.; Ganly, Ian; Gao, Jianjiong; Garraway, Levi A.; Gastier-Foster, Julie M.; Getz, Gad; Gehlenborg, Nils; Ghossein, Ronald; Gibbs, Richard A.; Giordano, Thomas J.; Gomez-Hernandez, Karen; Grimsby, Jonna; Gross, Benjamin; Guin, Ranabir; Hadjipanayis, Angela; Harper, Hollie A.; Hayes, D. Neil; Heiman, David I.; Herman, James G.; Hoadley, Katherine A.; Hofree, Matan; Holt, Robert A.; Hoyle, Alan P.; Huang, Franklin W.; Huang, Mei; Hutter, Carolyn M.; Ideker, Trey; Iype, Lisa; Jacobsen, Anders; Jefferys, Stuart R.; Jones, Corbin D.; Jones, Steven J.M.; Kasaian, Katayoon; Kebebew, Electron; Khuri, Fadlo R.; Kim, Jaegil; Kramer, Roger; Kreisberg, Richard; Kucherlapati, Raju; Kwiatkowski, David J.; Ladanyi, Marc; Lai, Phillip H.; Laird, Peter W.; Lander, Eric; Lawrence, Michael S.; Lee, Darlene; Lee, Eunjung; Lee, Semin; Lee, William; Leraas, Kristen M.; Lichtenberg, Tara M.; Lichtenstein, Lee; Lin, Pei; Ling, Shiyun; Liu, Jinze; Liu, Wenbin; Liu, Yingchun; LiVolsi, Virginia A.; Lu, Yiling; Ma, Yussanne; Mahadeshwar, Harshad S.; Marra, Marco A.; Mayo, Michael; McFadden, David G.; Meng, Shaowu; Meyerson, Matthew; Mieczkowski, Piotr A.; Miller, Michael; Mills, Gordon; Moore, Richard A.; Mose, Lisle E.; Mungall, Andrew J.; Murray, Bradley A.; Nikiforov, Yuri E.; Noble, Michael S.; Ojesina, Akinyemi I.; Owonikoko, Taofeek K.; Ozenberger, Bradley A.; Pantazi, Angeliki; Parfenov, Michael; Park, Peter J.; Parker, Joel S.; Paull, Evan O.; Pedamallu, Chandra Sekhar; Perou, Charles M.; Prins, Jan F.; Protopopov, Alexei; Ramalingam, Suresh S.; Ramirez, Nilsa C.; Ramirez, Ricardo; Raphael, Benjamin J.; Rathmell, W. Kimryn; Ren, Xiaojia; Reynolds, Sheila M.; Rheinbay, Esther; Ringel, Matthew D.; Rivera, Michael; Roach, Jeffrey; Robertson, A. Gordon; Rosenberg, Mara W.; Rosenthall, Matthew; Sadeghi, Sara; Saksena, Gordon; Sander, Chris; Santoso, Netty; Schein, Jacqueline E.; Schultz, Nikolaus; Schumacher, Steven E.; Seethala, Raja R.; Seidman, Jonathan; Senbabaoglu, Yasin; Seth, Sahil; Sharpe, Samantha; Mills Shaw, Kenna R.; Shen, John P.; Shen, Ronglai; Sherman, Steven; Sheth, Margi; Shi, Yan; Shmulevich, Ilya; Sica, Gabriel L.; Simons, Janae V.; Sipahimalani, Payal; Smallridge, Robert C.; Sofia, Heidi J.; Soloway, Matthew G.; Song, Xingzhi; Sougnez, Carrie; Stewart, Chip; Stojanov, Petar; Stuart, Joshua M.; Tabak, Barbara; Tam, Angela; Tan, Donghui; Tang, Jiabin; Tarnuzzer, Roy; Taylor, Barry S.; Thiessen, Nina; Thorne, Leigh; Thorsson, Vésteinn; Tuttle, R. Michael; Umbricht, Christopher B.; Van Den Berg, David J.; Vandin, Fabio; Veluvolu, Umadevi; Verhaak, Roel G.W.; Vinco, Michelle; Voet, Doug; Walter, Vonn; Wang, Zhining; Waring, Scot; Weinberger, Paul M.; Weinstein, John N.; Weisenberger, Daniel J.; Wheeler, David; Wilkerson, Matthew D.; Wilson, Jocelyn; Williams, Michelle; Winer, Daniel A.; Wise, Lisa; Wu, Junyuan; Xi, Liu; Xu, Andrew W.; Yang, Liming; Yang, Lixing; Zack, Travis I.; Zeiger, Martha A.; Zeng, Dong; Zenklusen, Jean Claude; Zhao, Ni; Zhang, Hailei; Zhang, Jianhua; Zhang, Jiashan (Julia); Zhang, Wei; Zmuda, Erik; Zou., Lihua

    2014-01-01

    Summary Papillary thyroid carcinoma (PTC) is the most common type of thyroid cancer. Here, we describe the genomic landscape of 496 PTCs. We observed a low frequency of somatic alterations (relative to other carcinomas) and extended the set of known PTC driver alterations to include EIF1AX, PPM1D and CHEK2 and diverse gene fusions. These discoveries reduced the fraction of PTC cases with unknown oncogenic driver from 25% to 3.5%. Combined analyses of genomic variants, gene expression, and methylation demonstrated that different driver groups lead to different pathologies with distinct signaling and differentiation characteristics. Similarly, we identified distinct molecular subgroups of BRAF-mutant tumors and multidimensional analyses highlighted a potential involvement of oncomiRs in less-differentiated subgroups. Our results propose a reclassification of thyroid cancers into molecular subtypes that better reflect their underlying signaling and differentiation properties, which has the potential to improve their pathological classification and better inform the management of the disease. PMID:25417114

  7. Integrated, multi-scale, spatial-temporal cell biology--A next step in the post genomic era.

    Science.gov (United States)

    Horwitz, Rick

    2016-03-01

    New microscopic approaches, high-throughput imaging, and gene editing promise major new insights into cellular behaviors. When coupled with genomic and other 'omic information and "mined" for correlations and associations, a new breed of powerful and useful cellular models should emerge. These top down, coarse-grained, and statistical models, in turn, can be used to form hypotheses merging with fine-grained, bottom up mechanistic studies and models that are the back bone of cell biology. The goal of the Allen Institute for Cell Science is to develop the top down approach by developing a high throughput microscopy pipeline that is integrated with modeling, using gene edited hiPS cell lines in various physiological and pathological contexts. The output of these experiments and models will be an "animated" cell, capable of integrating and analyzing image data generated from experiments and models. Copyright © 2015 Elsevier Inc. All rights reserved.

  8. Childhood Acute Lymphoblastic Leukemia: Integrating Genomics into Therapy

    Science.gov (United States)

    Tasian, Sarah K; Loh, Mignon L; Hunger, Stephen P

    2015-01-01

    Acute lymphoblastic leukemia (ALL), the most common malignancy of childhood, is a genetically complex entity that remains a major cause of childhood cancer-related mortality. Major advances in genomic and epigenomic profiling during the past decade have appreciably enhanced knowledge of the biology of de novo and relapsed ALL and have facilitated more precise risk stratification of patients. These achievements have also provided critical insights regarding potentially targetable lesions for development of new therapeutic approaches in the era of precision medicine. This review delineates the current genetic landscape of childhood ALL with emphasis upon patient outcomes with contemporary treatment regimens, as well as therapeutic implications of newly identified genomic alterations in specific subsets of ALL. PMID:26194091

  9. Big data or bust: realizing the microbial genomics revolution.

    Science.gov (United States)

    Raza, Sobia; Luheshi, Leila

    2016-02-01

    Pathogen genomics has the potential to transform the clinical and public health management of infectious diseases through improved diagnosis, detection and tracking of antimicrobial resistance and outbreak control. However, the wide-ranging benefits of this technology can only fully be realized through the timely collation, integration and sharing of genomic and clinical/epidemiological metadata by all those involved in the delivery of genomic-informed services. As part of our review on bringing pathogen genomics into 'health-service' practice, we undertook extensive stakeholder consultation to examine the factors integral to achieving effective data sharing and integration. Infrastructure tailored to the needs of clinical users, as well as practical support and policies to facilitate the timely and responsible sharing of data with relevant health authorities and beyond, are all essential. We propose a tiered data sharing and integration model to maximize the immediate and longer term utility of microbial genomics in healthcare. Realizing this model at the scale and sophistication necessary to support national and international infection management services is not uncomplicated. Yet the establishment of a clear data strategy is paramount if failures in containing disease spread due to inadequate knowledge sharing are to be averted, and substantial progress made in tackling the dangers posed by infectious diseases.

  10. Significance of stigma receptivity in intergeneric cross-pollination of Salix × Populus

    Directory of Open Access Journals (Sweden)

    Elżbieta Zenkteler

    2016-09-01

    Full Text Available The pollen–stigma interaction plays an important role in reproductive process and has been continuously studied in many interspecific and intergeneric crossing experiments. The aim of this study was to investigate stigma receptivity (SR of willow in order to determine the most suitable period for its pollination with poplar pollen and improve the effectiveness of Salix × Populus crosses. Tissue samples were examined histologically using light, epifluorescent, scanning, and transmission electron microscopy. Willow SR was determined by stigma morphological traits, test of pollen germination rate, Peroxtesmo test of peroxidase and esterase activity on stigma surface as well as papilla ultrastructure at anthesis. We have ascertained that the SR duration in willow is short, lasting from 1 to 2 DA. The poplar pollen germination rate on willow stigmas on 1 DA ranged from 26.3 to 11.2%.

  11. Community standards for genomic resources, genetic conservation, and data integration

    Science.gov (United States)

    Jill Wegrzyn; Meg Staton; Emily Grau; Richard Cronn; C. Dana Nelson

    2017-01-01

    Genetics and genomics are increasingly important in forestry management and conservation. Next generation sequencing can increase analytical power, but still relies on building on the structure of previously acquired data. Data standards and data sharing allow the community to maximize the analytical power of high throughput genomics data. The landscape of incomplete...

  12. Three Infectious Viral Species Lying in Wait in the Banana Genome

    Science.gov (United States)

    Chabannes, Matthieu; Baurens, Franc-Christophe; Duroy, Pierre-Olivier; Bocs, Stéphanie; Vernerey, Marie-Stéphanie; Rodier-Goud, Marguerite; Barbe, Valérie; Gayral, Philippe

    2013-01-01

    Plant pararetroviruses integrate serendipitously into their host genomes. The banana genome harbors integrated copies of banana streak virus (BSV) named endogenous BSV (eBSV) that are able to release infectious pararetrovirus. In this investigation, we characterized integrants of three BSV species—Goldfinger (eBSGFV), Imove (eBSImV), and Obino l'Ewai (eBSOLV)—in the seedy Musa balbisiana Pisang klutuk wulung (PKW) by studying their molecular structure, genomic organization, genomic landscape, and infectious capacity. All eBSVs exhibit extensive viral genome duplications and rearrangements. eBSV segregation analysis on an F1 population of PKW combined with fluorescent in situ hybridization analysis showed that eBSImV, eBSOLV, and eBSGFV are each present at a single locus. eBSOLV and eBSGFV contain two distinct alleles, whereas eBSImV has two structurally identical alleles. Genotyping of both eBSV and viral particles expressed in the progeny demonstrated that only one allele for each species is infectious. The infectious allele of eBSImV could not be identified since the two alleles are identical. Finally, we demonstrate that eBSGFV and eBSOLV are located on chromosome 1 and eBSImV is located on chromosome 2 of the reference Musa genome published recently. The structure and evolution of eBSVs suggest sequential integration into the plant genome, and haplotype divergence analysis confirms that the three loci display differential evolution. Based on our data, we propose a model for BSV integration and eBSV evolution in the Musa balbisiana genome. The mutual benefits of this unique host-pathogen association are also discussed. PMID:23720724

  13. SCREEN FOR DOMINANT BEHAVIORAL MUTATIONS CAUSED BY GENOMIC INSERTION OF P-ELEMENT TRANSPOSONS IN DROSOPHILA: AN EXAMINATION OF THE INTEGRATION OF VIRAL VECTOR SEQUENCES

    OpenAIRE

    FOX, LYLE E.; GREEN, DAVID; YAN, ZIYING; ENGELHARDT, JOHN F.; WU, CHUN-FANG

    2007-01-01

    Here we report the development of a high-throughput screen to assess dominant mutation rates caused by P-element transposition within the Drosophila genome that is suitable for assessing the undesirable effects of integrating foreign regulatory sequences (viral cargo) into a host genome. Three different behavioral paradigms were used: sensitivity to mechanical stress, response to heat stress, and ability to fly. The results, from our screen of 35,000 flies, indicate that mutations caused by t...

  14. The pig genome project has plenty to squeal about.

    Science.gov (United States)

    Fan, B; Gorbach, D M; Rothschild, M F

    2011-01-01

    Significant progress on pig genetics and genomics research has been witnessed in recent years due to the integration of advanced molecular biology techniques, bioinformatics and computational biology, and the collaborative efforts of researchers in the swine genomics community. Progress on expanding the linkage map has slowed down, but the efforts have created a higher-resolution physical map integrating the clone map and BAC end sequence. The number of QTL mapped is still growing and most of the updated QTL mapping results are available through PigQTLdb. Additionally, expression studies using high-throughput microarrays and other gene expression techniques have made significant advancements. The number of identified non-coding RNAs is rapidly increasing and their exact regulatory functions are being explored. A publishable draft (build 10) of the swine genome sequence was available for the pig genomics community by the end of December 2010. Build 9 of the porcine genome is currently available with Ensembl annotation; manual annotation is ongoing. These drafts provide useful tools for such endeavors as comparative genomics and SNP scans for fine QTL mapping. A recent community-wide effort to create a 60K porcine SNP chip has greatly facilitated whole-genome association analyses, haplotype block construction and linkage disequilibrium mapping, which can contribute to whole-genome selection. The future 'systems biology' that integrates and optimizes the information from all research levels can enhance the pig community's understanding of the full complexity of the porcine genome. These recent technological advances and where they may lead are reviewed. Copyright © 2011 S. Karger AG, Basel.

  15. Are electronic health records ready for genomic medicine?

    Science.gov (United States)

    Scheuner, Maren T; de Vries, Han; Kim, Benjamin; Meili, Robin C; Olmstead, Sarah H; Teleki, Stephanie

    2009-07-01

    The goal of this project was to assess genetic/genomic content in electronic health records. Semistructured interviews were conducted with key informants. Questions addressed documentation, organization, display, decision support and security of family history and genetic test information, and challenges and opportunities relating to integrating genetic/genomics content in electronic health records. There were 56 participants: 10 electronic health record specialists, 18 primary care clinicians, 16 medical geneticists, and 12 genetic counselors. Few clinicians felt their electronic record met their current genetic/genomic medicine needs. Barriers to integration were mostly related to problems with family history data collection, documentation, and organization. Lack of demand for genetics content and privacy concerns were also mentioned as challenges. Data elements and functionality requirements that clinicians see include: pedigree drawing; clinical decision support for familial risk assessment and genetic testing indications; a patient portal for patient-entered data; and standards for data elements, terminology, structure, interoperability, and clinical decision support rules. Although most said that there is little impact of genetics/genomics on electronic records today, many stated genetics/genomics would be a driver of content in the next 5-10 years. Electronic health records have the potential to enable clinical integration of genetic/genomic medicine and improve delivery of personalized health care; however, structured and standardized data elements and functionality requirements are needed.

  16. Using web services for linking genomic data to medical information systems.

    Science.gov (United States)

    Maojo, V; Crespo, J; de la Calle, G; Barreiro, J; Garcia-Remesal, M

    2007-01-01

    To develop a new perspective for biomedical information systems, regarding the introduction of ideas, methods and tools related to the new scenario of genomic medicine. Technological aspects related to the analysis and integration of heterogeneous clinical and genomic data include mapping clinical and genetic concepts, potential future standards or the development of integrated biomedical ontologies. In this clinicomics scenario, we describe the use of Web services technologies to improve access to and integrate different information sources. We give a concrete example of the use of Web services technologies: the OntoFusion project. Web services provide new biomedical informatics (BMI) approaches related to genomic medicine. Customized workflows will aid research tasks by linking heterogeneous Web services. Two significant examples of these European Commission-funded efforts are the INFOBIOMED Network of Excellence and the Advancing Clinico-Genomic Trials on Cancer (ACGT) integrated project. Supplying medical researchers and practitioners with omics data and biologists with clinical datasets can help to develop genomic medicine. BMI is contributing by providing the informatics methods and technological infrastructure needed for these collaborative efforts.

  17. A novel Sulfolobus non-conjugative extrachromosomal genetic element capable of integration into the host genome and spreading in the presence of a fusellovirus

    DEFF Research Database (Denmark)

    Wang, Ying; Duan, Zhenhong; Zhu, Haojun

    2007-01-01

    An integrative non-conjugative extrachromosomal genetic element, denoted as pSSVi, has been isolated from a Sulfolobus solfataricus P2 strain and was characterized. This genetic element is a double-stranded DNA of 5740 bp in size and contains eight open reading frames (ORFs). It resembles members....... Interestingly, pSSVi encodes an SSV-type integrase which probably catalyzes the integration of its genome into a specific site (a tRNA(Arg) gene) in the S. solfataricus P2 genome. Like pSSVx, pSSVi can be packaged into a spindle-like viral particle and spread with the help of SSV1 or SSV2. In addition, both SSV......1 and SSV2 appeared to replicate more efficiently in the presence of pSSVi. Given the versatile genetic abilities, pSSVi appears to be well suited for a role in horizontal gene transfer....

  18. DEFINING THE CHEMICAL SPACE OF PUBLIC GENOMIC ...

    Science.gov (United States)

    The current project aims to chemically index the genomics content of public genomic databases to make these data accessible in relation to other publicly available, chemically-indexed toxicological information. By defining the chemical space of public genomic data, it is possible to identify classes of chemicals on which to develop methodologies for the integration of chemogenomic data into predictive toxicology. The chemical space of public genomic data will be presented as well as the methodologies and tools developed to identify this chemical space.

  19. An Integrative Genomic Island Affects the Adaptations of Piezophilic Hyperthermophilic Archaeon Pyrococcus yayanosii to High Temperature and High Hydrostatic Pressure

    Directory of Open Access Journals (Sweden)

    Zhen Li

    2016-11-01

    Full Text Available Deep-sea hydrothermal vent environments are characterized by high hydrostatic pressure and sharp temperature and chemical gradients. Horizontal gene transfer is thought to play an important role in the microbial adaptation to such an extreme environment. In this study, a 21.4-kb DNA fragment was identified as a genomic island, designated PYG1, in the genomic sequence of the piezophilic hyperthermophile Pyrococcus yayanosii. According to the sequence alignment and functional annotation, the genes in PYG1 could tentatively be divided into five modules, with functions related to mobility, DNA repair, metabolic processes and the toxin-antitoxin system. Integrase can mediate the site-specific integration and excision of PYG1 in the chromosome of P. yayanosii A1. Gene replacement of PYG1 with a SimR cassette was successful. The growth of the mutant strain ∆PYG1 was compared with its parent strain P. yayanosii A2 under various stress conditions, including different pH, salinity, temperature and hydrostatic pressure. The ∆PYG1 mutant strain showed reduced growth when grown at 100 °C, while the biomass of ∆PYG1 increased significantly when cultured at 80 MPa. Differential expression of the genes in module Ⅲ of PYG1 was observed under different temperature and pressure conditions. This study demonstrates the first example of an archaeal integrative genomic island that could affect the adaptation of the hyperthermophilic piezophile P. yayanosii to high temperature and high hydrostatic pressure.

  20. The Planteome database: an integrated resource for reference ontologies, plant genomics and phenomics

    Science.gov (United States)

    Cooper, Laurel; Meier, Austin; Laporte, Marie-Angélique; Elser, Justin L; Mungall, Chris; Sinn, Brandon T; Cavaliere, Dario; Carbon, Seth; Dunn, Nathan A; Smith, Barry; Qu, Botong; Preece, Justin; Zhang, Eugene; Todorovic, Sinisa; Gkoutos, Georgios; Doonan, John H; Stevenson, Dennis W; Arnaud, Elizabeth

    2018-01-01

    Abstract The Planteome project (http://www.planteome.org) provides a suite of reference and species-specific ontologies for plants and annotations to genes and phenotypes. Ontologies serve as common standards for semantic integration of a large and growing corpus of plant genomics, phenomics and genetics data. The reference ontologies include the Plant Ontology, Plant Trait Ontology and the Plant Experimental Conditions Ontology developed by the Planteome project, along with the Gene Ontology, Chemical Entities of Biological Interest, Phenotype and Attribute Ontology, and others. The project also provides access to species-specific Crop Ontologies developed by various plant breeding and research communities from around the world. We provide integrated data on plant traits, phenotypes, and gene function and expression from 95 plant taxa, annotated with reference ontology terms. The Planteome project is developing a plant gene annotation platform; Planteome Noctua, to facilitate community engagement. All the Planteome ontologies are publicly available and are maintained at the Planteome GitHub site (https://github.com/Planteome) for sharing, tracking revisions and new requests. The annotated data are freely accessible from the ontology browser (http://browser.planteome.org/amigo) and our data repository. PMID:29186578

  1. Group sparse canonical correlation analysis for genomic data integration.

    Science.gov (United States)

    Lin, Dongdong; Zhang, Jigang; Li, Jingyao; Calhoun, Vince D; Deng, Hong-Wen; Wang, Yu-Ping

    2013-08-12

    The emergence of high-throughput genomic datasets from different sources and platforms (e.g., gene expression, single nucleotide polymorphisms (SNP), and copy number variation (CNV)) has greatly enhanced our understandings of the interplay of these genomic factors as well as their influences on the complex diseases. It is challenging to explore the relationship between these different types of genomic data sets. In this paper, we focus on a multivariate statistical method, canonical correlation analysis (CCA) method for this problem. Conventional CCA method does not work effectively if the number of data samples is significantly less than that of biomarkers, which is a typical case for genomic data (e.g., SNPs). Sparse CCA (sCCA) methods were introduced to overcome such difficulty, mostly using penalizations with l-1 norm (CCA-l1) or the combination of l-1and l-2 norm (CCA-elastic net). However, they overlook the structural or group effect within genomic data in the analysis, which often exist and are important (e.g., SNPs spanning a gene interact and work together as a group). We propose a new group sparse CCA method (CCA-sparse group) along with an effective numerical algorithm to study the mutual relationship between two different types of genomic data (i.e., SNP and gene expression). We then extend the model to a more general formulation that can include the existing sCCA models. We apply the model to feature/variable selection from two data sets and compare our group sparse CCA method with existing sCCA methods on both simulation and two real datasets (human gliomas data and NCI60 data). We use a graphical representation of the samples with a pair of canonical variates to demonstrate the discriminating characteristic of the selected features. Pathway analysis is further performed for biological interpretation of those features. The CCA-sparse group method incorporates group effects of features into the correlation analysis while performs individual feature

  2. Genome-wide survey and characterization of the WRKY gene family in Populus trichocarpa.

    Science.gov (United States)

    He, Hongsheng; Dong, Qing; Shao, Yuanhua; Jiang, Haiyang; Zhu, Suwen; Cheng, Beijiu; Xiang, Yan

    2012-07-01

    WRKY transcription factors participate in diverse physiological and developmental processes in plants. They have highly conserved WRKYGQK amino acid sequences in their N-termini, followed by the novel zinc-finger-like motifs, Cys₂His₂ or Cys₂HisCys. To date, numerous WRKY genes have been identified and characterized in a number of herbaceous species. Survey and characterization of WRKY genes in a ligneous species would facilitate a better understanding of the evolutionary processes and functions of this gene family. In this study, 104 poplar WRKY genes (PtWRKY) were identified in the latest poplar genome sequence. According to their structural features, the predicted members were divided into the previously defined groups I-III, as described in rice. In addition, chromosomal localization of the genes demonstrated that there might be WRKY gene hot spots in 2.3 Mb regions on chromosome 14. Furthermore, approximately 83% (86 out of 104) WRKY genes participated in gene duplication events, including 69% (29 out of 42) gene pairs which exhibited segmental duplication. Using semi-quantitative RT-PCR, the expression patterns of subgroup III genes were investigated under different stresses [cold, drought, salinity and salicylic acid (SA)]. The data revealed that these genes presented different expression levels in response to various stress conditions. Expression analysis exhibited PtWRKY76 gene induced markedly in 0.1 mM SA or 25% PEG-6000 treatment. The results presented here provide a fundamental clue for cloning specific function genes in further studies and applications. This study identified 104 poplar WRKY genes and demonstrated WRKY gene hot spots on chromosome 14. Furthermore, semi-quantitative RT-PCR showed variable stress responses in subgroup III.

  3. Analysis of a Farquhar-von Caemmerer-Berry leaf-level photosynthetic rate model for Populus tremuloides in the context of modeling and measurement limitations

    Science.gov (United States)

    K.E. Lenz; G.E. Host; K. Roskoski; A. Noormets; A. Sober; D.F. Karnosky

    2010-01-01

    The balance of mechanistic detail with mathematical simplicity contributes to the broad use of the Farquhar, von Caemmerer and Berry (FvCB) photosynthetic rate model. Here the FvCB model was coupled with a stomatal conductance model to form an [A,gs] model, and parameterized for mature Populus tremuloides leaves under varying CO2...

  4. Roles of CcrA and CcrB in Excision and Integration of Staphylococcal Cassette Chromosome mec, a Staphylococcus aureus Genomic Island▿

    OpenAIRE

    Wang, Lei; Archer, Gordon L.

    2010-01-01

    The gene encoding resistance to methicillin and other β-lactam antibiotics in staphylococci, mecA, is carried on a genomic island, SCCmec (for staphylococcal cassette chromosome mec). The chromosomal excision and integration of types I to IV SCCmec are catalyzed by the site-specific recombinases CcrA and CcrB, the genes for which are encoded on each element. We sought to identify the relative contributions of CcrA and CcrB in the excision and integration of SCCmec. Purified CcrB but not CcrA ...

  5. Analysis of a Farquhar-von Caemmerer-Berry leaf-level photosynthetic rate model for Populus tremuloides in the context of modeling and measurement limitations

    International Nuclear Information System (INIS)

    Lenz, Kathryn E.; Host, George E.; Roskoski, Kyle; Noormets, Asko; Sober, Anu; Karnosky, David F.

    2010-01-01

    The balance of mechanistic detail with mathematical simplicity contributes to the broad use of the Farquhar, von Caemmerer and Berry (FvCB) photosynthetic rate model. Here the FvCB model was coupled with a stomatal conductance model to form an [A,g s ] model, and parameterized for mature Populus tremuloides leaves under varying CO 2 and temperature levels. Data were selected to be within typical forest light, CO 2 and temperature ranges, reducing artifacts associated with data collected at extreme values. The error between model-predicted photosynthetic rate (A) and A data was measured in three ways and found to be up to three times greater for each of two independent data sets than for a base-line evaluation using parameterization data. The evaluation methods used here apply to comparisons of model validation results among data sets varying in number and distribution of data, as well as to performance comparisons of [A,g s ] models differing in internal-process components. - A photosynthetic rate model is parameterized for Populus tremuloides and evaluated based on its ability to predict dependent as well as independent data.

  6. Mouse Genome Informatics (MGI)

    Data.gov (United States)

    U.S. Department of Health & Human Services — MGI is the international database resource for the laboratory mouse, providing integrated genetic, genomic, and biological data to facilitate the study of human...

  7. Playing with heart and soul…and genomes: sports implications and applications of personal genomics.

    Science.gov (United States)

    Wagner, Jennifer K

    2013-01-01

    Whether the integration of genetic/omic technologies in sports contexts will facilitate player success, promote player safety, or spur genetic discrimination depends largely upon the game rules established by those currently designing genomic sports medicine programs. The integration has already begun, but there is not yet a playbook for best practices. Thus far discussions have focused largely on whether the integration would occur and how to prevent the integration from occurring, rather than how it could occur in such a way that maximizes benefits, minimizes risks, and avoids the exacerbation of racial disparities. Previous empirical research has identified members of the personal genomics industry offering sports-related DNA tests, and previous legal research has explored the impact of collective bargaining in professional sports as it relates to the employment protections of the Genetic Information Nondiscrimination Act (GINA). Building upon that research and upon participant observations with specific sports-related DNA tests purchased from four direct-to-consumer companies in 2011 and broader personal genomics (PGx) services, this anthropological, legal, and ethical (ALE) discussion highlights fundamental issues that must be addressed by those developing personal genomic sports medicine programs, either independently or through collaborations with commercial providers. For example, the vulnerability of student-athletes creates a number of issues that require careful, deliberate consideration. More broadly, however, this ALE discussion highlights potential sports-related implications (that ultimately might mitigate or, conversely, exacerbate racial disparities among athletes) of whole exome/genome sequencing conducted by biomedical researchers and clinicians for non-sports purposes. For example, the possibility that exome/genome sequencing of individuals who are considered to be non-patients, asymptomatic, normal, etc. will reveal the presence of variants of

  8. Software for computing and annotating genomic ranges.

    Science.gov (United States)

    Lawrence, Michael; Huber, Wolfgang; Pagès, Hervé; Aboyoun, Patrick; Carlson, Marc; Gentleman, Robert; Morgan, Martin T; Carey, Vincent J

    2013-01-01

    We describe Bioconductor infrastructure for representing and computing on annotated genomic ranges and integrating genomic data with the statistical computing features of R and its extensions. At the core of the infrastructure are three packages: IRanges, GenomicRanges, and GenomicFeatures. These packages provide scalable data structures for representing annotated ranges on the genome, with special support for transcript structures, read alignments and coverage vectors. Computational facilities include efficient algorithms for overlap and nearest neighbor detection, coverage calculation and other range operations. This infrastructure directly supports more than 80 other Bioconductor packages, including those for sequence analysis, differential expression analysis and visualization.

  9. Ensembl 2002: accommodating comparative genomics.

    Science.gov (United States)

    Clamp, M; Andrews, D; Barker, D; Bevan, P; Cameron, G; Chen, Y; Clark, L; Cox, T; Cuff, J; Curwen, V; Down, T; Durbin, R; Eyras, E; Gilbert, J; Hammond, M; Hubbard, T; Kasprzyk, A; Keefe, D; Lehvaslaiho, H; Iyer, V; Melsopp, C; Mongin, E; Pettett, R; Potter, S; Rust, A; Schmidt, E; Searle, S; Slater, G; Smith, J; Spooner, W; Stabenau, A; Stalker, J; Stupka, E; Ureta-Vidal, A; Vastrik, I; Birney, E

    2003-01-01

    The Ensembl (http://www.ensembl.org/) database project provides a bioinformatics framework to organise biology around the sequences of large genomes. It is a comprehensive source of stable automatic annotation of human, mouse and other genome sequences, available as either an interactive web site or as flat files. Ensembl also integrates manually annotated gene structures from external sources where available. As well as being one of the leading sources of genome annotation, Ensembl is an open source software engineering project to develop a portable system able to handle very large genomes and associated requirements. These range from sequence analysis to data storage and visualisation and installations exist around the world in both companies and at academic sites. With both human and mouse genome sequences available and more vertebrate sequences to follow, many of the recent developments in Ensembl have focusing on developing automatic comparative genome analysis and visualisation.

  10. The Ensembl genome database project.

    Science.gov (United States)

    Hubbard, T; Barker, D; Birney, E; Cameron, G; Chen, Y; Clark, L; Cox, T; Cuff, J; Curwen, V; Down, T; Durbin, R; Eyras, E; Gilbert, J; Hammond, M; Huminiecki, L; Kasprzyk, A; Lehvaslaiho, H; Lijnzaad, P; Melsopp, C; Mongin, E; Pettett, R; Pocock, M; Potter, S; Rust, A; Schmidt, E; Searle, S; Slater, G; Smith, J; Spooner, W; Stabenau, A; Stalker, J; Stupka, E; Ureta-Vidal, A; Vastrik, I; Clamp, M

    2002-01-01

    The Ensembl (http://www.ensembl.org/) database project provides a bioinformatics framework to organise biology around the sequences of large genomes. It is a comprehensive source of stable automatic annotation of the human genome sequence, with confirmed gene predictions that have been integrated with external data sources, and is available as either an interactive web site or as flat files. It is also an open source software engineering project to develop a portable system able to handle very large genomes and associated requirements from sequence analysis to data storage and visualisation. The Ensembl site is one of the leading sources of human genome sequence annotation and provided much of the analysis for publication by the international human genome project of the draft genome. The Ensembl system is being installed around the world in both companies and academic sites on machines ranging from supercomputers to laptops.

  11. Discovery of Cellular Proteins Required for the Early Steps of HCV Infection Using Integrative Genomics

    Science.gov (United States)

    Yang, Jae-Seong; Kwon, Oh Sung; Kim, Sanguk; Jang, Sung Key

    2013-01-01

    Successful viral infection requires intimate communication between virus and host cell, a process that absolutely requires various host proteins. However, current efforts to discover novel host proteins as therapeutic targets for viral infection are difficult. Here, we developed an integrative-genomics approach to predict human genes involved in the early steps of hepatitis C virus (HCV) infection. By integrating HCV and human protein associations, co-expression data, and tight junction-tetraspanin web specific networks, we identified host proteins required for the early steps in HCV infection. Moreover, we validated the roles of newly identified proteins in HCV infection by knocking down their expression using small interfering RNAs. Specifically, a novel host factor CD63 was shown to directly interact with HCV E2 protein. We further demonstrated that an antibody against CD63 blocked HCV infection, indicating that CD63 may serve as a new therapeutic target for HCV-related diseases. The candidate gene list provides a source for identification of new therapeutic targets. PMID:23593195

  12. On the irrigation requirements of cottonwood (Populus fremontii and Populus deltoides var. wislizenii) and willow (Salix gooddingii) grown in a desert environment

    Science.gov (United States)

    Hartwell, S.; Morino, K.; Nagler, P.L.; Glenn, E.P.

    2010-01-01

    Native tree plots have been established in river irrigation districts in the western U.S. to provide habitat for threatened and endangered birds. Information is needed on the effective irrigation requirements of the target species. Cottonwood (Populus spp.) and willow (Salix gooddingii) trees were grown for seven years in an outdoor plot in a desert environment in Tucson, Arizona. Plants were allowed to achieve a nearly complete canopy cover over the first four years, then were subjected to three daily summer irrigation schedules of 6.20??mm??d-1; 8.26??mm??d-1 and 15.7??mm??d-1. The lowest irrigation rate was sufficient to maintain growth and high leaf area index for cottonwoods over three years, while willows suffered considerable die-back on this rate in years six and seven. These irrigation rates were applied April 15-September 15, but only 0.88??mm??d-1 was applied during the dormant period of the year. Expressed as a fraction of reference crop evapotranspiration (ETo), recommended annual water applications plus precipitation (and including some deep drainage) were 0.83 ETo for cottonwood and 1.01 ETo for willow. Current practices tend to over-irrigate restoration plots, and this study can provide guidelines for more efficient water use. ?? 2010 Elsevier Ltd.

  13. Penalized differential pathway analysis of integrative oncogenomics studies

    NARCIS (Netherlands)

    van Wieringen, W.N.; van de Wiel, M.A.

    2014-01-01

    Through integration of genomic data from multiple sources, we may obtain a more accurate and complete picture of the molecular mechanisms underlying tumorigenesis. We discuss the integration of DNA copy number and mRNA gene expression data from an observational integrative genomics study involving

  14. D3GB: An Interactive Genome Browser for R, Python, and WordPress.

    Science.gov (United States)

    Barrios, David; Prieto, Carlos

    2017-05-01

    Genome browsers are useful not only for showing final results but also for improving analysis protocols, testing data quality, and generating result drafts. Its integration in analysis pipelines allows the optimization of parameters, which leads to better results. New developments that facilitate the creation and utilization of genome browsers could contribute to improving analysis results and supporting the quick visualization of genomic data. D3 Genome Browser is an interactive genome browser that can be easily integrated in analysis protocols and shared on the Web. It is distributed as an R package, a Python module, and a WordPress plugin to facilitate its integration in pipelines and the utilization of platform capabilities. It is compatible with popular data formats such as GenBank, GFF, BED, FASTA, and VCF, and enables the exploration of genomic data with a Web browser.

  15. Populus species from diverse habitats maintain high night-time conductance under drought.

    Science.gov (United States)

    Cirelli, Damián; Equiza, María Alejandra; Lieffers, Victor James; Tyree, Melvin Thomas

    2016-02-01

    We investigated the interspecific variability in nocturnal whole-plant stomatal conductance under well-watered and drought conditions in seedlings of four species of Populus from habitats characterized by abundant water supply (mesic and riparian) or from drier upland sites. The study was carried out to determine whether (i) nocturnal conductance varies across different species of Populus according to their natural habitat, (ii) nocturnal conductance is affected by water stress similarly to daytime conductance based on species habitat and (iii) differences in conductance among species could be explained partly by differences in stomatal traits. We measured whole-plant transpiration and conductance (G) of greenhouse-grown seedlings using an automated high-resolution gravimetric technique. No relationship was found between habitat preference and daytime G (GD), but night-time G (GN) was on average 1.5 times higher in riparian and mesic species (P. deltoides Bartr. ex Marsh. and P. trichocarpa Torr. & Gray) than in those from drier environments (P. tremuloides Michx. and P. × petrowskyana Schr.). GN was not significantly reduced under drought in riparian species. Upland species restricted GN significantly in response to drought, but it was still at least one order of magnitude greater that the cuticular conductance until leaf death was imminent. Under both well-watered and drought conditions, GN declined with increasing vapour pressure deficit (D). Also, a small increase in GN towards the end of the night period was observed in P. deltoides and P. × petrowskyana, suggesting the involvement of endogenous regulation. The anatomical analyses indicated a positive correlation between G and variable stomatal pore index among species and revealed that stomata are not likely to be leaky but instead seem capable of complete occlusion, which raises the question of the possible physiological role of the significant GN observed under drought. Further comparisons among

  16. Reframed Genome-Scale Metabolic Model to Facilitate Genetic Design and Integration with Expression Data.

    Science.gov (United States)

    Gu, Deqing; Jian, Xingxing; Zhang, Cheng; Hua, Qiang

    2017-01-01

    Genome-scale metabolic network models (GEMs) have played important roles in the design of genetically engineered strains and helped biologists to decipher metabolism. However, due to the complex gene-reaction relationships that exist in model systems, most algorithms have limited capabilities with respect to directly predicting accurate genetic design for metabolic engineering. In particular, methods that predict reaction knockout strategies leading to overproduction are often impractical in terms of gene manipulations. Recently, we proposed a method named logical transformation of model (LTM) to simplify the gene-reaction associations by introducing intermediate pseudo reactions, which makes it possible to generate genetic design. Here, we propose an alternative method to relieve researchers from deciphering complex gene-reactions by adding pseudo gene controlling reactions. In comparison to LTM, this new method introduces fewer pseudo reactions and generates a much smaller model system named as gModel. We showed that gModel allows two seldom reported applications: identification of minimal genomes and design of minimal cell factories within a modified OptKnock framework. In addition, gModel could be used to integrate expression data directly and improve the performance of the E-Fmin method for predicting fluxes. In conclusion, the model transformation procedure will facilitate genetic research based on GEMs, extending their applications.

  17. BGI-RIS: an integrated information resource and comparative analysis workbench for rice genomics

    DEFF Research Database (Denmark)

    Zhao, Wenming; Wang, Jing; He, Ximiao

    2004-01-01

    Rice is a major food staple for the world's population and serves as a model species in cereal genome research. The Beijing Genomics Institute (BGI) has long been devoting itself to sequencing, information analysis and biological research of the rice and other crop genomes. In order to facilitate....... Designed as a basic platform, BGI-RIS presents the sequenced genomes and related information in systematic and graphical ways for the convenience of in-depth comparative studies (http://rise.genomics.org.cn/). Udgivelsesdato: 2004-Jan-1...

  18. Fueling the Future with Fungal Genomes

    Energy Technology Data Exchange (ETDEWEB)

    Grigoriev, Igor V.

    2014-10-27

    Genomes of fungi relevant to energy and environment are in focus of the JGI Fungal Genomic Program. One of its projects, the Genomics Encyclopedia of Fungi, targets fungi related to plant health (symbionts and pathogens) and biorefinery processes (cellulose degradation and sugar fermentation) by means of genome sequencing and analysis. New chapters of the Encyclopedia can be opened with user proposals to the JGI Community Science Program (CSP). Another JGI project, the 1000 fungal genomes, explores fungal diversity on genome level at scale and is open for users to nominate new species for sequencing. Over 400 fungal genomes have been sequenced by JGI to date and released through MycoCosm (www.jgi.doe.gov/fungi), a fungal web-portal, which integrates sequence and functional data with genome analysis tools for user community. Sequence analysis supported by functional genomics will lead to developing parts list for complex systems ranging from ecosystems of biofuel crops to biorefineries. Recent examples of such ‘parts’ suggested by comparative genomics and functional analysis in these areas are presented here.

  19. Ten years of maintaining and expanding a microbial genome and metagenome analysis system.

    Science.gov (United States)

    Markowitz, Victor M; Chen, I-Min A; Chu, Ken; Pati, Amrita; Ivanova, Natalia N; Kyrpides, Nikos C

    2015-11-01

    Launched in March 2005, the Integrated Microbial Genomes (IMG) system is a comprehensive data management system that supports multidimensional comparative analysis of genomic data. At the core of the IMG system is a data warehouse that contains genome and metagenome datasets sequenced at the Joint Genome Institute or provided by scientific users, as well as public genome datasets available at the National Center for Biotechnology Information Genbank sequence data archive. Genomes and metagenome datasets are processed using IMG's microbial genome and metagenome sequence data processing pipelines and are integrated into the data warehouse using IMG's data integration toolkits. Microbial genome and metagenome application specific data marts and user interfaces provide access to different subsets of IMG's data and analysis toolkits. This review article revisits IMG's original aims, highlights key milestones reached by the system during the past 10 years, and discusses the main challenges faced by a rapidly expanding system, in particular the complexity of maintaining such a system in an academic setting with limited budgets and computing and data management infrastructure. Copyright © 2015 Elsevier Ltd. All rights reserved.

  20. Integrating the genomic architecture of human nucleolar organizer regions with the biophysical properties of nucleoli.

    Science.gov (United States)

    Mangan, Hazel; Gailín, Michael Ó; McStay, Brian

    2017-12-01

    Nucleoli are the sites of ribosome biogenesis and the largest membraneless subnuclear structures. They are intimately linked with growth and proliferation control and function as sensors of cellular stress. Nucleoli form around arrays of ribosomal gene (rDNA) repeats also called nucleolar organizer regions (NORs). In humans, NORs are located on the short arms of all five human acrocentric chromosomes. Multiple NORs contribute to the formation of large heterochromatin-surrounded nucleoli observed in most human cells. Here we will review recent findings about their genomic architecture. The dynamic nature of nucleoli began to be appreciated with the advent of photodynamic experiments using fluorescent protein fusions. We review more recent data on nucleoli in Xenopus germinal vesicles (GVs) which has revealed a liquid droplet-like behavior that facilitates nucleolar fusion. Further analysis in both XenopusGVs and Drosophila embryos indicates that the internal organization of nucleoli is generated by a combination of liquid-liquid phase separation and active processes involving rDNA. We will attempt to integrate these recent findings with the genomic architecture of human NORs to advance our understanding of how nucleoli form and respond to stress in human cells. © 2017 Federation of European Biochemical Societies.