Julio Alonzo

The ML Digest

Trending AI/ML papers explained

Não deixe de visitar o site do podcast e apoiar quem o produz: podcasters.spotify.com

Autor

Julio Alonzo

Categoria

Technology

Site do podcast

podcasters.spotify.com

Último episódio

9 de set de 2025

Onde ouvir?

Podcasts no app Replaio Radio Em breve

Os podcasts estão chegando ao app. Instale agora e seja o primeiro a descobrir um jeito totalmente novo de curtir podcasts

Baixe no Google Play Instale grátis Android 5 mi+ downloads · nota 4,8 iOS em breve

Episódios

Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025

This episode of  The ML Digest  covers the paper  “Towards a Unified View of Large Language Model Post-Training”  from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...

Are Small Language Models the Future of Agentic AI? 08.09.2025

In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153

Ouça o podcast The ML Digest no Replaio

Rádio e podcasts em um só app - grátis e sem cadastro. Instale hoje e não perca o lançamento

Baixe no Google Play

O Replaio não é o publicador dos podcasts; os nomes dos programas, as capas e o áudio pertencem aos seus autores e são distribuídos por feeds RSS públicos