Julio Alonzo
The ML Digest
Trending AI/ML papers explained
Besuch unbedingt die Website des Podcasts und unterstütze die Macher: podcasters.spotify.com
Autor
Julio Alonzo
Kategorie
Podcast-Website
Neueste Folge
9. Sep 2025
Wo hören?
Podcasts in der App Replaio Radio Bald verfügbarPodcasts kommen bald in die App. Installiere sie jetzt und erlebe als Erster einen ganz neuen Blick auf Podcasts
Folgen
Unifying LLM Post-Training: From SFT and RL to Hybrid Approaches 09.09.2025 25:31
This episode of The ML Digest covers the paper “Towards a Unified View of Large Language Model Post-Training” from researchers at Tsinghua University, Shanghai AI Lab, and WeChat AI. The authors argue that seemingly distinct approaches—Supervised Fine-Tuning (SFT) with offline demonstrations and Reinforcement Learning (RL) with online rollouts—are in fact instances of a single optimization pro...
Are Small Language Models the Future of Agentic AI? 08.09.2025 25:36
In this episode we go over the recent NVIDIA paper titled "Small Language Models are the Future of Agentic AI." Link to the original paper: https://arxiv.org/pdf/2506.02153
Ähnliche Podcasts
Replaio ist kein Herausgeber von Podcasts; die Namen der Sendungen, Cover und Audioinhalte gehören ihren Autoren und werden über öffentliche RSS-Feeds verbreitet