fr

Roman Cheplyaka

the bioinformatics chat

A podcast about computational biology, bioinformatics, and next generation sequencing.

Auteur

Roman Cheplyaka

Catégorie

Science

Site du podcast

bioinformatics.chat

Dernier épisode

21 déc. 2023

Où écouter ?

Les podcasts dans l'appli Replaio Radio Bientôt disponible

Les podcasts arrivent très bientôt dans l'appli. Installe-la dès maintenant et découvre en avant-première une toute nouvelle façon de vivre les podcasts

Télécharger sur Google Play Installe-la gratuitement Android 5 M+ de téléchargements · note de 4,8 iOS bientôt

Épisodes

#70 Prioritizing drug target genes with Marie Sadler 21.12.2023

In this episode, Marie Sadler talks about her recent Cell Genomics paper, Multi-layered genetic approaches to identify approved drug targets . Previous studies have found that the drugs that target a gene linked to the disease are more likely to be approved. Yet there are many ways to define what it means for a gene to be linked to the disease. Perhaps the most straightforward approach is to rely...

#69 Suffix arrays in optimal compressed space and δ-SA with Tomasz Kociumaka and Dominik Kempa 29.09.2023

Today on the podcast we have Tomasz Kociumaka and Dominik Kempa , the authors of the preprint Collapsing the Hierarchy of Compressed Data Structures: Suffix Arrays in Optimal Compressed Space . The suffix array is one of the foundational data structures in bioinformatics, serving as an index that allows fast substring searches in a large text. However, in its raw form, the suffix array occupies th...

#68 Phylogenetic inference from raw reads and Read2Tree with David Dylus 28.08.2023

In this episode, David Dylus talks about Read2Tree , a tool that builds alignment matrices and phylogenetic trees from raw sequencing reads. By leveraging the database of orthologous genes called OMA , Read2Tree bypasses traditional, time-consuming steps such as genome assembly, annotation and all-versus-all sequence comparisons. Links: Inference of phylogenetic trees directly from raw sequencing...

#67 AlphaFold and variant effect prediction with Amelie Stein 29.07.2023

This is the third and final episode in the AlphaFold series, originally recorded on February 23, 2022, with Amelie Stein , now an associate professor at the University of Copenhagen. In the episode, Amelie explains what 𝛥𝛥G is, how it informs us whether a particular protein mutation affects its stability, and how AlphaFold 2 helps in this analysis. A note from Amelie: Something that has happened i...

#66 AlphaFold and shape-mers with Janani Durairaj 10.07.2023

This is the second episode in the AlphaFold series, originally recorded on February 14, 2022, with Janani Durairaj , a postdoctoral researcher at the University of Basel. Janani talks about how she used shape-mers and topic modelling to discover classes of proteins assembled by AlphaFold 2 that were absent from the Protein Data Bank (PDB). The bioinformatics discussion starts at 03:35. Links: A st...

#65 AlphaFold and protein interactions with Pedro Beltrao 21.06.2023

In this episode, originally recorded on February 9, 2022, Roman talks to Pedro Beltrao about AlphaFold, the software developed by DeepMind that predicts a protein’s 3D structure from its amino acid sequence. Pedro is an associate professor at ETH Zurich and the coordinator of the structural biology community assessment of AlphaFold2 applications project, which involved over 30 scientists from diff...

#64 Enformer: predicting gene expression from sequence with Žiga Avsec 09.11.2021

In this episode, Jacob Schreiber interviews Žiga Avsec about a recently released model, Enformer . Their discussion begins with life differences between academia and industry, specifically about how research is conducted in the two settings. Then, they discuss the Enformer model, how it builds on previous work, and the potential that models like it have for genomics research in the future. Finally...

#63 Bioinformatics Contest 2021 with Maksym Kovalchuk and James Matthew Holt 27.09.2021

The Bioinformatics Contest is back this year, and we are back to discuss it! This year’s contest winners Maksym Kovalchuk (1st prize) and Matt Holt (2nd prize) talk about how they approach participating in the contest and what strategies have earned them the top scores. Timestamps and links for the individual problems: 00:10:36 Genotype Imputation 00:21:26 Causative Mutation 00:30:27 Superspreader...

#62 Steady states of metabolic networks and Dingo with Apostolos Chalkis 28.07.2021

In this episode, Apostolos Chalkis presents sampling steady states of metabolic networks as an alternative to the widely used flux balance analysis (FBA). We also discuss dingo , a Python package written by Apostolos that employs geometric random walks to sample steady states. You can see dingo in action here . Links: Dingo on GitHub Searching for COVID-19 treatments using metabolic networks Tweag...

#61 3D genome organization and GRiNCH with Da-Inn Erika Lee 23.06.2021

In this episode, Jacob Schreiber interviews Da-Inn Erika Lee about data and computational methods for making sense of 3D genome structure. They begin their discussion by talking about 3D genome structure at a high level and the challenges in working with such data. Then, they discuss a method recently developed by Erika, named GRiNCH , that mines this data to identify spans of the genome that clus...

#60 Differential gene expression and DESeq2 with Michael Love 12.05.2021

In this episode, Michael Love joins us to talk about the differential gene expression analysis from bulk RNA-Seq data. We talk about the history of Mike’s own differential expression package, DESeq2 , as well as other packages in this space, like edgeR and limma , and the theory they are based upon. Mike also shares his experience of being the author and maintainer of a popular bioninformatics pac...

#59 Proteomics calibration with Lindsay Pino 21.04.2021

In this episode, Lindsay Pino discusses the challenges of making quantitative measurements in the field of proteomics. Specifically, she discusses the difficulties of comparing measurements across different samples, potentially acquired in different labs, as well as a method she has developed recently for calibrating these measurements without the need for expensive reagents. The discussion then t...

#58 B cell maturation and class switching with Hamish King 31.03.2021

In this episode, we learn about B cell maturation and class switching from Hamish King . Hamish recently published a paper on this subject in Science Immunology, where he and his coauthors analyzed gene expression and antibody repertoire data from human tonsils. In the episode Hamish talks about some of the interesting B cell states he uncovered and shares his thoughts on questions such as «When d...

#57 Enhancers with Molly Gasperini 10.03.2021

In this episode, Jacob Schreiber interviews Molly Gasperini about enhancer elements. They begin their discussion by talking about Octant Bio, and then dive into the surprisingly difficult task of defining enhancers and determining the mechanisms that enable them to regulate gene expression. Links: Octant Bio Towards a comprehensive catalogue of validated and target-linked human enhancers (Molly Ga...

#56 Polygenic risk scores in admixed populations with Bárbara Bitarello 17.02.2021

Polygenic risk scores (PRS) rely on the genome-wide association studies (GWAS) to predict the phenotype based on the genotype. However, the prediction accuracy suffers when GWAS from one population are used to calculate PRS within a different population, which is a problem because the majority of the GWAS are done on cohorts of European ancestry. In this episode, Bárbara Bitarello helps us underst...

#55 Phylogenetics and the likelihood gradient with Xiang Ji 13.01.2021

In this episode, we chat about phylogenetics with Xiang Ji . We start with a general introduction to the field and then go deeper into the likelihood-based methods (maximum likelihood and Bayesian inference). In particular, we talk about the different ways to calculate the likelihood gradient, including a linear-time exact gradient algorithm recently published by Xiang and his colleagues. Links: G...

#54 Seeding methods for read alignment with Markus Schmidt 16.12.2020

In this episode, Markus Schmidt explains how seeding in read alignment works. We define and compare k-mers, minimizers, MEMs, SMEMs, and maximal spanning seeds. Markus also presents his recent work on computing variable-sized seeds (MEMs, SMEMs, and maximal spanning seeds) from fixed-sized seeds (k-mers and minimizers) and his Modular Aligner . Links: A performant bridge between fixed-size and var...

#53 Real-time quantitative proteomics with Devin Schweppe 18.11.2020

In this episode, Jacob Schreiber interviews Devin Schweppe about the analysis of mass spectrometry data in the field of proteomics. They begin by delving into the different types of mass spectrometry methods, including MS1, MS2, and, MS3, and the reasons for using each. They then discuss a recent paper from Devin, Full-Featured, Real-Time Database Searching Platform Enables Fast and Accurate Multi...

#52 How 23andMe finds identical-by-descent segments with William Freyman 27.10.2020

In this episode, Will Freyman talks about identity-by-descent (IBD): how it’s used at 23andMe , and how the templated positional Burrows-Wheeler transform can find IBD segments in the presence of genotyping and phasing errors. Links: Fast and robust identity-by-descent inference with the templated positional Burrows-Wheeler transform (William A. Freyman, Kimberly F. McManus, Suyash S. Shringarpure...

#51 Basset and Basenji with David Kelley 07.10.2020

In this episode, Jacob Schreiber interviews David Kelley about machine learning models that can yield insight into the consequences of mutations on the genome. They begin their discussion by talking about Calico Labs, and then delve into a series of papers that David has written about using models, named Basset and Basenji, that connect genome sequence to functional activity and so can be used to...

#50 ENCODE3 with Jill Moore 10.09.2020

In this episode, Jacob Schreiber interviews Jill Moore about recent research from the ENCODE Project . They begin their discussion with an overview and goals of the ENCODE Project, and then discuss a bundle of papers that were recently published in various Nature journals and the flagship paper, Expanded encyclopaedias of DNA elements in the human and mouse genomes . They conclude their discussion...

#49 Most Permissive Boolean Networks with Loïc Paulevé 19.08.2020

In systems biology, Boolean networks are a way to model interactions such as gene regulation or cell signaling. The standard interpretations of Boolean networks are the synchronous, asynchronous, and fully asynchronous semantics. In this episode, Loïc Paulevé explains how the same Boolean networks can be interpreted in a new, “most permissive” way. Loïc proved mathematically that his semantics can...

#48 Machine learning for drug development with Marinka Zitnik 29.07.2020

In this episode, Jacob Schreiber interviews Marinka Zitnik about applications of machine learning to drug development. They begin their discussion with an overview of open research questions in the field, including limiting the search space of high-throughput testing methods, designing drugs entirely from scratch, predicting ways that existing drugs can be repurposed, and identifying likely side-e...

#47 Reproducible pipelines and NGLess with Luis Pedro Coelho 24.06.2020

NGLess is a programming language specifically targeted at next generation sequencing (NGS) data processing. In this episode we chat with its main developer, Luis Pedro Coelho , about the benefits of domain-specific languages, pros and cons of Haskell in bioinformatics, reproducibility, and of course NGLess itself. Links: NGLess on GitHub NG-meta-profiler: fast processing of metagenomes using NGLes...

#46 HiFi reads and HiCanu with Sergey Nurk and Sergey Koren 27.05.2020

In this episode, I continue to talk (but mostly listen) to Sergey Koren and Sergey Nurk . If you missed the previous episode , you should probably start there. Otherwise, join us to learn about HiFi reads, the tradeoff between read length and quality, and what tricks HiCanu employs to resolve highly similar repeats. Links: HiCanu: accurate assembly of segmental duplications, satellites, and alleli...

Écoute le podcast the bioinformatics chat sur Replaio

La radio et les podcasts dans une seule appli - gratuite, sans inscription. Installe-la dès aujourd'hui et ne rate pas le lancement

Télécharger sur Google Play

Replaio n'est pas éditeur de podcasts ; les noms des émissions, les visuels et l'audio appartiennent à leurs auteurs et sont diffusés via des flux RSS publics