Julio Alonzo

The ML Digest

Trending AI/ML papers explained

Koniecznie odwiedź stronę podcastu i wesprzyj twórcę: podcasters.spotify.com

Autor

Julio Alonzo

Kategoria

Technology

Strona podcastu

podcasters.spotify.com

Ostatni odcinek

9 wrz 2025

Gdzie słuchać?

Podcasty w aplikacji Replaio Radio Już wkrótce

Podcasty trafią do aplikacji już wkrótce. Zainstaluj teraz i jako pierwszy zobacz nowe podejście do podcastów

Pobierz z Google Play Zainstaluj za darmo Android prawie 10 mln pobrań · ocena 4,8 iOS niedługo

Odcinki

Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025

This episode of  The ML Digest  covers the paper  “Towards a Unified View of Large Language Model Post-Training”  from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...

Are Small Language Models the Future of Agentic AI? 08.09.2025

In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153

Słuchaj podcastu The ML Digest w Replaio

Radio i podcasty w jednej aplikacji - za darmo, bez zakładania konta. Zainstaluj już dziś i nie przegap premiery

Pobierz z Google Play

Replaio nie jest wydawcą podcastów; nazwy audycji, okładki i audio należą do ich autorów i są rozpowszechniane przez publiczne kanały RSS