Applied Computational Genomics
The Applied Computational Genomics group focuses on theoretical and computational aspects of modelling the process of genome evolution and adaptive change. With growing size and complexity of molecular data, we strive to keep pace providing accurate, scalable and practical computational solutions that enable a wide range of scientists to analyse patterns of evolution and natural selection in large genomic and omics data.
Our goal is to bring new bioinformatics methods to real applications ranging from biotechnology to biomedical research, ecology and agriculture.
Evolutionary analyses of selective pressures in genomic data have high potential for applications, since natural selection is a leading force in function conservation, in adaptation to emerging pathogens, new environments, and plays key role in immune and resistance systems.
We develop phylogenetic methods for protein-coding sequences that enable to evaluate selective pressure and detect adaptive instances based on genomic signatures. Our computational methods serve to generate new biological hypotheses and predictions for further experimental validation, with the ultimate goal to develop practical applications.
We embrace the interdisciplinary approach by integrating different data sources and combining methods from biological, mathematical, and computer science disciplines.
- Fast alignment and phylogeny in frequentist framework
Evolutionary thinking helps to disentangle underlying biological mechanisms shaping molecular data. Genomic sequences of common origin are routinely used to infer phylogenies, which provide test-base for biological hypotheses or support downstream analyses. Based on fast approximation algorithms, we aim to include the alignment uncertainty during phylogeny estimation in the frequentist setting. This will allow for more accurate phylogenetic inferences from vast high-throughput data.
- Biosoda - Data Integration in BioSoda, NRP75 Bigdata
This project aims at enabling sophisticated semantic queries across large, decentralized and heterogeneous databases via an intuitive interface. The system will enable scientists, without prior training, to perform powerful joint queries across resources in ways that cannot be anticipated and therefore goes far and above the query functionality of specialized knowledge bases. The project represents an interdisciplinary collaboration between information systems and bioinformatics.
- Evolution and function of genomic tandem repeats
We develop statistical phylogenetic methods for analysing tandem repeats in genomic sequences. For example, leucine rich repeats (LRRs) in plant resistance genes provide a source for adaptation to emerging pathogens, so detecting selection on LRRs can bring ideas how to improve crop resistance (Shaper and Anisimova 2014, New Phytol).
- Stochastic models for protein-coding genes
We develop methods to study effects of selection on amino acid and codon mutation patterns. These methods can help to identify drug targets and study somatic processes. Our recent antibody model captures the sui generis mechanism specific to somatic hypermutation in maturating antibodies (Mirsky et al 2015, Mol Biol Evol). This provides basis for new bioinformatics methods for antibody analysis necessary for antibody selection and synthesis in the commercial context.
- Lighthouse project (Singeria SnSF) - Trans-omic approach to colorectal cancer: an integrative computational and clinical perspective
Frontiers in Bioinformatics.
Available from: https://doi.org/10.3389/fbinf.2021.691865
Frontiers in Bioinformatics.
Available from: https://doi.org/10.3389/fbinf.2021.685844
Pecerska, Julija; Kühnert, Denise; Meehan, Conor J.; Coscollá, Mireia; de Jong, Bouke C.; Gagneux, Sebastien; Stadler, Tanja,
Available from: https://doi.org/10.1016/j.epidem.2021.100471
Frick, Thomas; Glüge, Stefan; Rahimi, Abbas; Benini, Luca; Brunschwiler, Thomas,
Wireless Mobile Communication and Healthcare.
International Conference on Wireless Mobile Communication and Healthcare (MobiHealth), Online, 18 December 2020.
Lecture Notes of the Institute for Computer Sciences, Social Informatics and Telecommunications Engineering ; 362.
Available from: https://doi.org/10.1007/978-3-030-70569-5_15
Maloca, Peter M.; Müller, Philipp L.; Lee, Aaron Y.; Tufail, Adnan; Balaskas, Konstantinos; Niklaus, Stephanie; Kaiser, Pascal; Suter, Susanne; Zarranz-Ventura, Javier; Egan, Catherine; Scholl, Hendrik P. N.; Schnitzer, Tobias K.; Singer, Thomas; Hasler, Pascal W.; Denk, Nora,
4(1), pp. 170.
Available from: https://doi.org/10.1038/s42003-021-01697-y
Digital Tools for Codon Optimization
This project proposes to develop, study and apply mathematical models to optimize protein production of a gene that stems from one organism in a different organism. The focus will lie specifically on genes involved in the biosynthesis of suberin. Suberin is a carbon-rich decay-resistant biopolyester, found majorly ...
Trans-omic approach to colorectal cancer: an integrative computational and clinical perspective
Colorectal Cancer (CRC) is an important cause of cancer-related mortality world-wide. The Consensus Molecular Subtypes represent the first comprehensive molecular classification with clinical implications, but many aspects are still missing. We use a transomic approach to improve the stratification, prognosis, and ...