Next in AI
Next in AI: Your Daily News Podcast
Stay ahead of artificial intelligence daily. AI Daily Brief brings you the latest AI news, research, tools, and industry trends — explained clearly and quickly. This daily AI podcast helps founders, developers, and curious minds cut through the noise and understand what’s next in technology.
Author
Next in AI
Category
Podcast website
Latest episode
May 16, 2026
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
Qwen3-Next: Decoupling LLM Knowledge from Compute for Sustainable AI Performance 13.09.2025 21:50
The podcast introduces Qwen3-Next, a new generation of large language models developed by Alibaba, emphasizing its innovative hybrid architecture designed for efficiency and long-context processing. This model significantly advances the Mixture-of-Experts (MoE) paradigm by activating only a small fraction of its total parameters (around 3 billion out of 80 billion) during inference, drastica...
From LLMs to LRMs: Reinforcement Learning's Quest for Truly Reasoning AI 12.09.2025 23:50
This podcast explores the integration of Reinforcement Learning (RL) with Large Reasoning Models (LRMs) , highlighting its foundational components, current challenges, and diverse applications. It discusses various reward design strategies , including verifiable, generative, dense, and unsupervised rewards, along with reward shaping techniques to optimize learning. The text further categorizes...
ChatGPT Developer Mode: Unleashing AI Power & Unpacking the "Lethal Trifecta" of Security Risks 11.09.2025 17:31
The podcast discusses the recent release of ChatGPT's "Developer Mode" , which grants full Model Context Protocol (MCP) client access, enabling the AI to interact with external tools and services. This feature, while powerful for developers seeking to automate tasks and integrate with various APIs, raises significant security concerns due to prompt injection vulnerabilities . Many commentators...
Trillion-Parameter Titans: Alibaba's Qwen3-Max-Preview vs. Kimi K2's Agentic AI Showdown 10.09.2025 16:36
Unpack the latest breakthroughs in AI with our podcast. We delve into trillion-parameter language models like Alibaba's Qwen3-Max-Preview , which marks a significant advancement for Chinese AI technology in ultra-large-scale models, and the open-source Kimi K2 , as well as Qwen3 models. Discover their cutting-edge agentic capabilities , with Kimi K2 specifically designed for agentic intelligence a...
Meta REFRAG: 30x Faster and Smarter Knowledge Access 09.09.2025 20:21
Tune into "REFRAG: Rethinking RAG Decoding" to discover a cutting-edge framework revolutionizing Retrieval-Augmented Generation (RAG) in Large Language Models (LLMs). Learn how REFRAG tackles the challenges of long-context inputs , which typically cause high latency and memory demands. This podcast explores REFRAG's innovative "compress, sense, and expand context" approach,...
OpenAI: Why LLM Hallucinates and How Our Tests Make It Worse 07.09.2025 16:04
Why do AI chatbots confidently make up facts ? This podcast explores the surprising reasons language models 'hallucinate' . We'll uncover how these plausible falsehoods originate from statistical errors during pretraining and persist because evaluations reward guessing over acknowledging uncertainty . Learn why models are optimized to be good test-takers, much like students guessing on...
Beyond Chatbots: Building Robust LLM Agents with LangGraph 06.09.2025 19:39
Dive into LangGraph , the production-ready agent runtime designed to give you control and durability over your AI agents. Discover how LangGraph addresses the unique challenges of slow, flaky, and open-ended LLMs with features like parallelization, streaming, checkpointing, and human-in-the-loop . Whether you're building simple routers , dynamic tool-calling agents (like ReAct), or custom agen...
The Gemmaverse Unleashed: Private, Powerful AI in Your Pocket 05.09.2025 13:53
Welcome to the "Gemmaverse Unlocked" podcast! Dive into the world of Google's Gemma family of open models , where State-of-the-Art AI meets On-Device and Offline capabilities . Join us as we explore: EmbeddingGemma : The best-in-class, mobile-first embedding model designed for private, efficient semantic search and RAG pipelines directly on your hardware, even without internet connec...
Unpacking Implicit Reasoning: The Silent, Speedy Revolution in LLM Thinking 05.09.2025 20:29
Decoding the Silent Mind: Implicit Reasoning in LLMs Discover Implicit Reasoning , the cutting-edge method where Large Language Models (LLMs) solve complex, multi-step problems silently, using internal latent structures, without generating intermediate textual steps .Move beyond verbose "Chain-of-Thought" (CoT) prompting! Implicit reasoning offers significant benefits: Lower generation c...
LLMs Unleashed: How GLM-4.5, vLLM, and Cognitive Load Shape the Future of AI Software 04.09.2025 19:15
Explore the future of AI software development with a look into advanced LLMs, high-performance inference systems, and the human element of cognitive load. This episode covers: GLM-4.5: The AI Frontier : Discover GLM-4.5 , Z.ai's flagship LLM series, featuring 355 billion total parameters and an innovative MoE architecture with an MTP layer for speculative decoding . We'll delve into its un...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.