Julio Alonzo
The ML Digest
Trending AI/ML papers explained
No dejes de visitar la web del podcast y apoyar a su creador: podcasters.spotify.com
Autor
Julio Alonzo
Categoría
Web del podcast
Último episodio
9 de sep. de 2025
¿Dónde escuchar?
Podcasts en la app Replaio Radio Muy prontoLos podcasts llegarán muy pronto a la app. Instálala ahora y sé el primero en descubrir una forma totalmente nueva de vivir los podcasts
Episodios
Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025 25:31
This episode of The ML Digest covers the paper “Towards a Unified View of Large Language Model Post-Training” from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...
Are Small Language Models the Future of Agentic AI? 08.09.2025 25:36
In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153
Podcasts similares
Replaio no es editor de podcasts; los nombres de los programas, las portadas y el audio pertenecen a sus autores y se distribuyen a través de canales RSS públicos