Julio Alonzo
The ML Digest
Trending AI/ML papers explained
N'hésitez pas à visiter le site du podcast et à soutenir son créateur : podcasters.spotify.com
Auteur
Julio Alonzo
Catégorie
Site du podcast
Dernier épisode
9 sept. 2025
Où écouter ?
Les podcasts dans l'appli Replaio Radio Bientôt disponibleLes podcasts arrivent très bientôt dans l'appli. Installe-la dès maintenant et découvre en avant-première une toute nouvelle façon de vivre les podcasts
Épisodes
Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025 25:31
This episode of The ML Digest covers the paper “Towards a Unified View of Large Language Model Post-Training” from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...
Are Small Language Models the Future of Agentic AI? 08.09.2025 25:36
In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153
Podcasts similaires
Replaio n'est pas éditeur de podcasts ; les noms des émissions, les visuels et l'audio appartiennent à leurs auteurs et sont diffusés via des flux RSS publics