ALBERT

All Library Books, journals and Electronic Records Telegrafenberg

Your email was sent successfully. Check your inbox.

An error occurred while sending the email. Please try again.

Proceed reservation?

Export
Filter
  • Computational Methods, Genomics  (140)
  • Computational Methods  (62)
  • Physiology & Biochemistry  (57)
  • Oxford University Press  (259)
  • Public Library of Science
Collection
Publisher
Years
  • 1
    Publication Date: 2015-09-19
    Description: Sequence alignment is a long standing problem in bioinformatics. The Basic Local Alignment Search Tool (BLAST) is one of the most popular and fundamental alignment tools. The explosive growth of biological sequences calls for speedup of sequence alignment tools such as BLAST. To this end, we develop high speed BLASTN (HS-BLASTN), a parallel and fast nucleotide database search tool that accelerates MegaBLAST—the default module of NCBI-BLASTN. HS-BLASTN builds a new lookup table using the FMD-index of the database and employs an accurate and effective seeding method to find short stretches of identities (called seeds) between the query and the database. HS-BLASTN produces the same alignment results as MegaBLAST and its computational speed is much faster than MegaBLAST. Specifically, our experiments conducted on a 12-core server show that HS-BLASTN can be 22 times faster than MegaBLAST and exhibits better parallel performance than MegaBLAST. HS-BLASTN is written in C++ and the related source code is available at https://github.com/chenying2016/queries under the GPLv3 license.
    Keywords: Computational Methods
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 2
    Publication Date: 2015-09-19
    Description: Recent releases of genome three-dimensional (3D) structures have the potential to transform our understanding of genomes. Nonetheless, the storage technology and visualization tools need to evolve to offer to the scientific community fast and convenient access to these data. We introduce simultaneously a database system to store and query 3D genomic data ( 3DBG ), and a 3D genome browser to visualize and explore 3D genome structures ( 3DGB ). We benchmark 3DBG against state-of-the-art systems and demonstrate that it is faster than previous solutions, and importantly gracefully scales with the size of data. We also illustrate the usefulness of our 3D genome Web browser to explore human genome structures. The 3D genome browser is available at http://3dgb.cs.mcgill.ca/ .
    Keywords: Computational Methods, Genomics
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 3
    Publication Date: 2015-05-29
    Description: Model evaluation is a necessary step for better prediction and design of 3D RNA structures. For proteins, this has been widely studied and the knowledge-based statistical potential has been proved to be one of effective ways to solve this problem. Currently, a few knowledge-based statistical potentials have also been proposed to evaluate predicted models of RNA tertiary structures. The benchmark tests showed that they can identify the native structures effectively but further improvements are needed to identify near-native structures and those with non-canonical base pairs. Here, we present a novel knowledge-based potential, 3dRNAscore, which combines distance-dependent and dihedral-dependent energies. The benchmarks on different testing datasets all show that 3dRNAscore are more efficient than existing evaluation methods in recognizing native state from a pool of near-native states of RNAs as well as in ranking near-native states of RNA models.
    Keywords: Computational Methods
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 4
    Publication Date: 2015-05-29
    Description: Identification of transcription units (TUs) encoded in a bacterial genome is essential to elucidation of transcriptional regulation of the organism. To gain a detailed understanding of the dynamically composed TU structures, we have used four strand-specific RNA-seq (ssRNA-seq) datasets collected under two experimental conditions to derive the genomic TU organization of Clostridium thermocellum using a machine-learning approach. Our method accurately predicted the genomic boundaries of individual TUs based on two sets of parameters measuring the RNA-seq expression patterns across the genome: expression-level continuity and variance. A total of 2590 distinct TUs are predicted based on the four RNA-seq datasets. Among the predicted TUs, 44% have multiple genes. We assessed our prediction method on an independent set of RNA-seq data with longer reads. The evaluation confirmed the high quality of the predicted TUs. Functional enrichment analyses on a selected subset of the predicted TUs revealed interesting biology. To demonstrate the generality of the prediction method, we have also applied the method to RNA-seq data collected on Escherichia coli and achieved high prediction accuracies. The TU prediction program named SeqTU is publicly available at https://code.google.com/p/seqtu/ . We expect that the predicted TUs can serve as the baseline information for studying transcriptional and post-transcriptional regulation in C. thermocellum and other bacteria.
    Keywords: Computational Methods, Genomics
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 5
    Publication Date: 2015-05-29
    Description: Detecting genetic variation is one of the main applications of high-throughput sequencing, but is still challenging wherever aligning short reads poses ambiguities. Current state-of-the-art variant calling approaches avoid such regions, arguing that it is necessary to sacrifice detection sensitivity to limit false discovery. We developed a method that links candidate variant positions within repetitive genomic regions into clusters. The technique relies on a resource, a thesaurus of genetic variation, that enumerates genomic regions with similar sequence. The resource is computationally intensive to generate, but once compiled can be applied efficiently to annotate and prioritize variants in repetitive regions. We show that thesaurus annotation can reduce the rate of false variant calls due to mappability by up to three orders of magnitude. We apply the technique to whole genome datasets and establish that called variants in low mappability regions annotated using the thesaurus can be experimentally validated. We then extend the analysis to a large panel of exomes to show that the annotation technique opens possibilities to study variation in hereto hidden and under-studied parts of the genome.
    Keywords: Computational Methods, Genomics
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 6
    Publication Date: 2016-07-31
    Description: The 16S rRNA gene (16S rDNA) codes for RNA that plays a fundamental role during translation in the ribosome and is used extensively as a marker gene to establish relationships among bacteria. However, the complementary non-coding 16S rDNA (nc16S rDNA) has been ignored. An idea emerged in the course of analyzing bacterial 16S rDNA sequences in search for nucleotide composition and substitution patterns: Does the nc16S rDNA code? If so, what does it code for? More importantly: Does 16S rDNA evolution reflect its own evolution or the evolution of its counterpart nc16S rDNA? The objective of this minireview is to discuss these thoughts. nc strands often encode small RNAs (sRNAs), ancient components of gene regulation. nc16S rDNA sequences from different bacterial groups were used to search for possible matches in the Bacterial Small Regulatory RNA Database. Intriguingly, the sequence of one published sRNA obtained from Legionella pneumophila (GenBank: AE017354.1) showed high non-random similarity with nc16S rDNA corresponding in part to the V5 region especially from Legionella and relatives. While the target(s) of this sRNA is unclear at the moment, its mere existence might open up a new chapter in the use of the 16S rDNA to study relationships among bacteria.
    Keywords: Physiology & Biochemistry
    Print ISSN: 0378-1097
    Electronic ISSN: 1574-6968
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 7
    Publication Date: 2016-08-05
    Description: Thermotolerance of the fungus Fomes sp. EUM1 was evaluated in solid state fermentation (SSF). This thermotolerant strain improved both hyphal invasiveness (38%) and length (17%) in adverse thermal conditions exceeding 30°C and to a maximum of 40°C. In contrast, hyphal branching decreased by 46% at 45°C. The production of cellulases over corn stover increased 1.6-fold in 30°C culture conditions, xylanases increased 2.8-fold at 40°C, while laccase production improved 2.7-fold at 35°C. Maximum production of lignocellulolytic enzymes was obtained at elevated temperatures in shorter fermentation times (8–6 days), although the proteases appeared as a thermal stress response associated with a drop in lignocellulolytic activities. Novel and multiple isoenzymes of xylanase (four bands) and cellulase (six bands) were secreted in the range of 20–150 kDa during growth in adverse temperature conditions. However, only a single laccase isoenzyme (46 kDa) was detected. This is the first report describing the advantages of a thermotolerant white-rot fungus in SSF. These results have important implications for large-scale SSF, where effects of metabolic heat are detrimental to growth and enzyme production, which are severely affected by the formation of high temperature gradients.
    Keywords: Physiology & Biochemistry
    Print ISSN: 0378-1097
    Electronic ISSN: 1574-6968
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 8
    Publication Date: 2016-06-21
    Description: Assigning cancer patients to the most effective treatments requires an understanding of the molecular basis of their disease. While DNA-based molecular profiling approaches have flourished over the past several years to transform our understanding of driver pathways across a broad range of tumors, a systematic characterization of key driver pathways based on RNA data has not been undertaken. Here we introduce a new approach for predicting the status of driver cancer pathways based on signature functions derived from RNA sequencing data. To identify the driver cancer pathways of interest, we mined DNA variant data from TCGA and nominated driver alterations in seven major cancer pathways in breast, ovarian and colon cancer tumors. The activation status of these driver pathways were then characterized using RNA sequencing data by constructing classification signature functions in training datasets and then testing the accuracy of the signatures in test datasets. The signature functions differentiate well tumors with nominated pathway activation from tumors with no signs of activation: average AUC equals to 0.83. Our results confirm that driver genomic alterations are distinctively displayed at the transcriptional level and that the transcriptional signatures can generally provide an alternative to DNA sequencing methods in detecting specific driver pathways.
    Keywords: Computational Methods, Genomics
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 9
    Publication Date: 2016-06-21
    Description: The goal of pathway analysis is to identify the pathways that are significantly impacted when a biological system is perturbed, e.g. by a disease or drug. Current methods treat pathways as independent entities. However, many signals are constantly sent from one pathway to another, essentially linking all pathways into a global, system-wide complex. In this work, we propose a set of three pathway analysis methods based on the impact analysis, that performs a system-level analysis by considering all signals between pathways, as well as their overlaps. Briefly, the global system is modeled in two ways: (i) considering the inter-pathway interaction exchange for each individual pathways, and (ii) combining all individual pathways to form a global, system-wide graph. The third analysis method is a hybrid of these two models. The new methods were compared with DAVID, GSEA, GSA, PathNet, Crosstalk and SPIA on 23 GEO data sets involving 19 tissues investigated in 12 conditions. The results show that both the ranking and the P -values of the target pathways are substantially improved when the analysis considers the system-wide dependencies and interactions between pathways.
    Keywords: Computational Methods
    Print ISSN: 0305-1048
    Electronic ISSN: 1362-4962
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
  • 10
    Publication Date: 2016-06-23
    Description: The rhizobacterium Serratia plymuthica 4Rx13 emits the novel and unique volatile sodorifen (C 16 H 26 ), which has a polymethylated bicyclic structure. Transcriptome analysis revealed that gene SOD_c20750 (annotated as terpene cyclase) is involved in the biosynthesis of sodorifen. Here we show that this gene is located in a small cluster of four genes ( SOD_c20750 – SOD_c20780 ), and the analysis of the knockout mutants demonstrated that SOD_c20760 (annotated as methyltransferase) and SOD_c20780 (annotated as isopentenyl pyrophosphate (IPP) isomerase) are needed for the biosynthesis of sodorifen, while a sodorifen-negative phenotype was not achieved with the SOD_c20770 (annotated as deoxy-xylulose-5-phosphate (DOXP) synthase) mutant. Altogether, the function of this new gene cluster was assigned to the biosynthesis of this structurally unusual volatile compound sodorifen.
    Keywords: Physiology & Biochemistry
    Print ISSN: 0378-1097
    Electronic ISSN: 1574-6968
    Topics: Biology
    Location Call Number Expected Availability
    BibTip Others were also interested in ...
Close ⊗
This website uses cookies and the analysis tool Matomo. More information can be found here...