Yun Wu

Learning GenAI via SOTA Papers

This podcast is focusing on sharing the papers on GenAI related topic, especially the SOTA (State of the Art) papers that are the foundations of GenAI work. It shows how these researches paved the way to the GenAI tools that we are using every day such as ChatGPT, Gemini, Claude Code etc.

Husk å besøke podkastens nettsted og støtte skaperen: podcasters.spotify.com

Forfatter

Yun Wu

Kategori

Technology

Podkastens nettsted

podcasters.spotify.com

Siste episode

9. okt 2026

Hvor kan du lytte?

Podkaster i appen Replaio Radio Kommer snart

Podkaster kommer snart til appen. Installer nå, og bli den første som ser en helt ny tilnærming til podkaster

Last ned på Google Play Installer gratis Android nesten 10 mill. nedlastinger · 4,8 i vurdering iOS snart

Episoder

EP005: How BERT Mastered Language by Hiding Words 24.02.2026

The paper " BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding " introduces a new language representation model called BERT , which stands for Bidirectional Encoder Representations from Transformers . Unlike previous language models that were restricted to unidirectional (left-to-right) architectures, BERT is designed to pre-train deep bidirectional representations fr...

EP004: How 7000 Unpublished Books Birthed GPT 23.02.2026

The paper " Improving Language Understanding by Generative Pre-Training " by Alec Radford and colleagues at OpenAI introduces a semi-supervised framework to address the challenge of limited labeled data for diverse natural language understanding (NLU) tasks. The authors propose a two-stage training procedure : • Unsupervised Pre-training: A high-capacity 12-layer Transformer decoder is first train...

EP003: How ELMo Made Word Vectors Dynamic 23.02.2026

The paper " Deep contextualized word representations " introduces a novel type of word representation called ELMo (Embeddings from Language Models). Unlike traditional word embeddings that provide a single, context-independent vector for each word, ELMo representations are deep contextualized vectors derived from all internal layers of a deep bidirectional language model (biLM) pre-trained on a la...

EP002: ULMFiT Was the ImageNet Moment for Text 23.02.2026

The paper " Universal Language Model Fine-tuning for Text Classification " by Jeremy Howard and Sebastian Ruder introduces ULMFiT , an effective transfer learning method for Natural Language Processing (NLP). While transfer learning has long revolutionized computer vision, NLP models previously required significant task-specific modifications or training from scratch. ULMFiT enables ImageNet-like...

EP001: How Transformers Smashed the Sequential Bottleneck 22.02.2026

Attention is All You Need • The Shift to Transformers: The overview discusses the move away from complex recurrent and convolutional neural networks toward the Transformer architecture , which relies entirely on attention mechanisms to draw global dependencies between inputs and outputs,. • Self-Attention & Multi-Head Attention: It explains how the model uses self-attention to relate different...

Hør på podkasten Learning GenAI via SOTA Papers i Replaio

Radio og podkaster i én app - gratis og uten registrering. Installer i dag, og ikke gå glipp av lanseringen

Last ned på Google Play

Replaio er ikke podkastutgiver; programnavn, omslag og lyd tilhører opphavspersonene og distribueres via offentlige RSS-feeder.