Julio Alonzo
The ML Digest
Trending AI/ML papers explained
Não deixe de visitar o site do podcast e apoiar quem o produz: podcasters.spotify.com
Autor
Julio Alonzo
Categoria
Site do podcast
Último episódio
9 de set de 2025
Onde ouvir?
Podcasts no app Replaio Radio Em breveOs podcasts estão chegando ao app. Instale agora e seja o primeiro a descobrir um jeito totalmente novo de curtir podcasts
Episódios
Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025 25:31
This episode of The ML Digest covers the paper “Towards a Unified View of Large Language Model Post-Training” from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...
Are Small Language Models the Future of Agentic AI? 08.09.2025 25:36
In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153
Podcasts parecidos
O Replaio não é o publicador dos podcasts; os nomes dos programas, as capas e o áudio pertencem aos seus autores e são distribuídos por feeds RSS públicos