Julio Alonzo
The ML Digest
Trending AI/ML papers explained
Koniecznie odwiedź stronę podcastu i wesprzyj twórcę: podcasters.spotify.com
Autor
Julio Alonzo
Kategoria
Strona podcastu
Ostatni odcinek
9 wrz 2025
Gdzie słuchać?
Podcasty w aplikacji Replaio Radio Już wkrótcePodcasty trafią do aplikacji już wkrótce. Zainstaluj teraz i jako pierwszy zobacz nowe podejście do podcastów
Odcinki
Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025 25:31
This episode of The ML Digest covers the paper “Towards a Unified View of Large Language Model Post-Training” from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...
Are Small Language Models the Future of Agentic AI? 08.09.2025 25:36
In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153
Podobne podcasty
Replaio nie jest wydawcą podcastów; nazwy audycji, okładki i audio należą do ich autorów i są rozpowszechniane przez publiczne kanały RSS