Itzik Ben-Shabat

Talking Papers Podcast

Talking Papers Podcast: deep dives into research papers in computer vision, 3D, machine learning, and AI, with the authors who wrote them. Where research meets conversation. By researchers, for researchers. Each episode is structured like the paper itself: a TL;DR / abstract to set the stage, then related work, approach, results, conclusions, and future work. We close with a bonus segment called "What did Reviewer 2 say?", where the authors share the candid peer-review story behind the publication. Hosted by Itzik Ben-Shabat. Guests are PhD students, postdocs, and faculty from leading labs acr...

Author

Itzik Ben-Shabat

Category

Technology

Latest episode

Feb 17, 2025

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android almost 10M downloads · 4.8 rating iOS soon

Episodes

Dejan Azinović - Neural RGBD Surface Reconstruction 06.05.2022

 In this episode of the Talking Papers Podcast, I hosted Dejan Azinović to chat about his paper "Neural RGB-D Surface Reconstruction”, published in CVPR 2022. In this paper, they take on the task of RGBD surface reconstruction by using novel view synthesis.  They incorporate depth measurements into the radiance field formulation by learning a neural network that stores a truncated signed dist...

Yuliang Xiu - ICON 19.04.2022

In this episode of the Talking Papers Podcast, I hosted Yuliang Xiu to chat about his paper "ICON: Implicit Clothed humans Obtained from Normals”, published in CVPR 2022. SMPL(-X) body model to infer clothed humans (conditioned on the normals).  Additionally, they propose an inference-time feedback loop that alternates between refining the body's normals and the shape.  PAPER TITLE  &quo...

Itai Lang - SampleNet 28.03.2022

In this episode of the Talking Papers Podcast, I hosted Itai Lang to chat about his paper "SampleNet: Differentiable Point Cloud Sampling”, published in CVPR 2020. In this paper, they propose a point soft-projection to allow differentiating through the sampling operation and enable learning task-specific point sampling. Combined with their regularization and task-specific losses, they can red...

Manuel Dahnert - Panoptic 3D Scene Reconstruction 07.03.2022

In this episode of the Talking Papers Podcast, I hosted Manuel Dahnert to chat about his paper “Panoptic 3D Scene Reconstruction From a Single RGB Image”, published in NeurIPS 2021.  In this paper, they unify the task of reconstruction, semantic segmentation and instance segmentation in 3D from a single RGB image. They propose a holistic approach to lift the 2D features into a 3D grid.  Manuel is...

Songyou Peng - Shape As Points 24.02.2022

In this episode of the Talking Papers Podcast, I hosted Songyou Peng to chat about his paper “Shape As Points: A Differentiable Poisson Solver”, published in NeurIPS 2021. In this paper, they take on the task of surface reconstruction and propose a hybrid representation that unifies explicit and implicit representation in addition to a differentiable solver for the classic Poisson surface reconstr...

Yicong Hong - VLN BERT 17.02.2022

PAPER TITLE: " VLN BERT:  A Recurrent Vision-and-Language BERT for Navigation " AUTHORS:  Yicong Hong, Qi Wu, Yuankai Qi, Cristian Rodriguez-Opazo, Stephen Gould ABSTRACT: Accuracy of many visiolinguistic tasks has benefited significantly from the application of vision-and-language (V&L) BERT. However, its application for the task of vision and-language navigation (VLN) remains limit...

Despoina Paschalidou - Neural Parts 10.02.2022

PAPER TITLE  Neural Parts: Learning Expressive 3D Shape Abstractions with Invertible Neural Networks AUTHORS Despoina Paschalidou , Angelos Katharopoulos , Andreas Geiger , Sanja Fidler ABSTRACT Impressive progress in 3D shape extraction led to representations that can capture object geometries with high fidelity. In parallel, primitive-based methods seek to represent objects as semantically consi...

Guy Gafni - NerFACE 03.02.2022

PAPER TITLE: Dynamic Neural Radiance Fields for Monocular 4D Facial Avatar Reconstruction AUTHORS:  Guy Gafni      Justus Thies      Michael Zollhöfer     Matthias Nießner    Project page: https://gafniguy.github.io/4D-Facial-Avatars/ CODE: 💻 https://github.com/gafniguy/4D-Facial-Avatars ABSTRACT: We present dynamic neural radiance fields for modeling the appearance and dynamics of a human face....

Jing Zhang - UC-Net 20.01.2022

PAPER TITLE: " UC-Net: Uncertainty Inspired RGB-D Saliency Detection via Conditional Variational Autoencoders " AUTHORS: Jing Zhang, Deng-Ping Fan, Yuchao Dai, Saeed Anwar, Fatemeh Sadat Saleh, Tong Zhang, Nick Barnes ABSTRACT: In this paper, we propose the first framework (UCNet) to employ uncertainty for RGB-D saliency detection by learning from the data labeling process. Existing RGB-...

Dylan Campbell - Deep Declarative Networks 13.01.2022

PAPER TITLE: " Deep Declarative Networks: a new hope " AUTHORS: Stephen Gould, Richard Hartley, Dylan Campbell ABSTRACT: We explore a new class of end-to-end learnable models wherein data processing nodes (or network layers) are defined in terms of desired behaviour rather than an explicit forward function. Specifically, the forward function is implicitly defined as the solution to a mat...

Cristian Rodriguez-Opazo - DORi 05.01.2022

Paper title: " DORi: Discovering Object Relationships for Moment Localization of a Natural Language Query in a Video " Authors:  Cristian Rodriguez-Opazo, Edison Marrese-Taylor, Basura Fernando, Hongdong Li, Stephen Gould Abstract: This paper studies the task of temporal moment localization in a long untrimmed video using natural language query. Given a query sentence, the goal is to det...

Listen to the Talking Papers Podcast podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.