Julio Alonzo

The ML Digest

Trending AI/ML papers explained

No dejes de visitar la web del podcast y apoyar a su creador: podcasters.spotify.com

Autor

Julio Alonzo

Categoría

Technology

Web del podcast

podcasters.spotify.com

Último episodio

9 de sep. de 2025

¿Dónde escuchar?

Podcasts en la app Replaio Radio Muy pronto

Los podcasts llegarán muy pronto a la app. Instálala ahora y sé el primero en descubrir una forma totalmente nueva de vivir los podcasts

Descárgala en Google Play Instálala gratis Android casi 10 M de descargas · valoración de 4,8 iOS muy pronto

Episodios

Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025

This episode of  The ML Digest  covers the paper  “Towards a Unified View of Large Language Model Post-Training”  from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...

Are Small Language Models the Future of Agentic AI? 08.09.2025

In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153

Escucha el podcast The ML Digest en Replaio

Radio y podcasts en una sola app - gratis y sin registro. Instálala hoy y no te pierdas el estreno

Descárgala en Google Play

Replaio no es editor de podcasts; los nombres de los programas, las portadas y el audio pertenecen a sus autores y se distribuyen a través de canales RSS públicos