Next in AI

Next in AI: Your Daily News Podcast

Stay ahead of artificial intelligence daily. AI Daily Brief brings you the latest AI news, research, tools, and industry trends — explained clearly and quickly. This daily AI podcast helps founders, developers, and curious minds cut through the noise and understand what’s next in technology.

Author

Next in AI

Category

Technology

Podcast website

podcasters.spotify.com

Latest episode

May 16, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

Perplexity MoE Deployment Deep Dive: The Custom Kernels and Network Secrets That Make Massive AI Models Run 5X Faster 06.11.2025

The podcast describes the development of high-performance, portable communication kernels specifically designed to handle the challenging sparse expert parallelism (EP) communication requirements (Dispatch and Combine) of large-scale Mixture-of-Experts (MoE) models such as DeepSeek R1 and Kimi-K2. An initial open-source NVSHMEM-based library achieved performance up to 10x faster than standard All-...

Stop Vibe Coding! Cognition's Windsurf Codemaps Battles the "Comprehension Tax" to Turn Engineers' Brains On 05.11.2025

The provided podcast introduces and discuss es Windsurf Codemaps , a new AI-powered feature developed by Cognition.ai for code comprehension, designed to create  AI-annotated structured maps  of a codebase. The feature aims to shift AI developer tooling beyond simple code generation by addressing the complex, high-value problem of understanding large, intricate codebases for tasks like debugging a...

OpenAI's $38 Billion AWS Deal: How a Sovereign AI Power Built a $700 Billion Multi-Cloud Empire and the Financial Bubble That Could Pop It All 04.11.2025

The podcast provides an extensive analysis of  OpenAI's infrastructure strategy , highlighted by a new  multi-year, $38 billion partnership with Amazon Web Services (AWS)  for computing power. The AWS deal, which grants OpenAI access to  Amazon EC2 UltraServers  featuring advanced NVIDIA GPUs, is presented as part of a much larger,  multi-cloud portfolio  that includes massive contracts with M...

Karpathy's AI Divide: Why We're Summoning "Ghosts," Agents Will Take a Decade, and the Brutal "March of Nines" 18.10.2025

The podcast provides an extensive interview transcript with  Andrej Karpathy , discussing his views on the future of  Large Language Models (LLMs)  and  AI agents . Karpathy argues that the full realization of competent AI agents will take a  decade , primarily due to current models'  cognitive deficits , lack of continual learning, and insufficient multimodality. He contrasts the current appr...

30 Gigawatts and the AI Race: Inside OpenAI's Custom Chip Alliance with Broadcom to Build Compute Abundance 14.10.2025

The podcast provides excerpts from an OpenAI podcast episode announcing a major partnership between  OpenAI and Broadcom  to develop custom artificial intelligence infrastructure. This collaboration, which has been ongoing for approximately  18 months , focuses on designing a new  custom chip  and a complete  vertical system  to support advanced AI workloads. Speakers from both companies, includin...

AI's Tectonic Shift: The State of AI 2025—Superintelligence Race, Open Source Tsunami, and the Looming Cybersecurity Crisis 11.10.2025

The podcast provides an extensive overview of the  State of AI for 2025 , presented by Nathan Benaich, General Partner of Air Street Capital. This material, which is drawn from a long-form video presentation and associated report, meticulously analyzes recent developments across  AI research, industry, politics, and safety . Key research narratives include the rapid progress of  OpenAI and the nar...

Gemini 2.5 Computer Use Model: How Google's New AI Agent Is Learning to 'Live' Inside Your Browser and Conquer the Messy Web 09.10.2025

The podcast discusses the launch and implications of  Google's Gemini 2.5 Computer Use model , a specialized AI built on Gemini 2.5 Pro designed to  interact directly with user interfaces (UIs) , such as filling forms and navigating websites. The official announcement highlights the model's superior performance in web and mobile control benchmarks with  low latency , achieved through an it...

ChatGPT’s New Apps SDK: The Universal UI Dream vs. The Developer's Walled Garden 07.10.2025

The podcast provides an extensive overview of guidelines for developers building applications that integrate with ChatGPT, which are referred to as "Apps" and leverage the  Model Context Protocol (MCP) , allowing for dynamic user interfaces like inline cards, carousels, and fullscreen experiences within the chat environment. The  App developer guidelines  establish minimum standards cent...

End AI Amnesia: Anthropic's Context Editing and Memory Tool Solve LLM Forgetfulness and Token Limits 06.10.2025

The podcast discusses new features on the  Claude Developer Platform  to enhance agents' ability to manage long-running tasks by addressing context window limitations. Specifically, Anthropic introduces  context editing , which automatically removes stale information like old tool results to preserve conversation flow and extend operational time. Additionally, the  memory tool  allows agents t...

OpenAI's Money Furnace: How $13.5 Billion in Losses Fuels the AI Arms Race and the Inevitable Ad Strategy 04.10.2025

The podcast focuses heavily on the financial health and long-term viability of  OpenAI , particularly given its substantial revenue of  $4.3 billion  contrasted with a  $13.5 billion  net loss in the first half of 2025, which includes massive spending on R&D and employee stock compensation. A central debate revolves around whether the company can successfully monetize its product, ChatGPT, wit...

OpenAI Sora 2: Video Generation Advancements and Deployment 01.10.2025

The podcast discusses the launch of  Sora 2 , the company’s advanced video and audio generation model, highlighting its improved capabilities in  realism, physics modeling, and controllability . The documents emphasize a strong commitment to  responsible deployment , outlining comprehensive safety measures integrated into the new Sora iOS app and its web platform. Key safeguards include visible an...

Claude Sonnet 4.5: Best AI Coder or Vibe Coder? Deep Diving Anthropic's Agent Autonomy, Price Wars, and the 30-Hour Task Breakthrough 30.09.2025

The podcast discusses announcement from Anthropic introducing  Claude Sonnet 4.5 , which is presented as the world's best model for coding and building complex agents, showing substantial gains in reasoning and math capabilities. The text highlights major product upgrades, including  checkpoints in Claude Code  and a  native VS Code extension , alongside a new  Claude Agent SDK  to allow devel...

The Synergy Secret: How Gemini Robotics' Dual-Model Agent (GR 1.5 & GR-ER 1.5) Solves the General-Purpose Robot Problem 27.09.2025

The podcast introduces and explain the capabilities of the  Gemini Robotics 1.5 model family  from Google DeepMind, focusing on the  Vision-Language-Action (VLA) model (GR 1.5)  and the  Embodied Reasoning (ER) model (GR-ER 1.5) . These models are designed to enable general-purpose robots to perceive, reason, and execute complex, multi-step tasks in the physical world, leveraging innovations like...

OpenAI: Why the GDPval Benchmark Reveals Near-Human Parity and Catastrophic Failure Rates 26.09.2025

The podcast introduces  GDPval , a new benchmark created by OpenAI to evaluate AI models on  real-world economically valuable tasks  across major sectors contributing to U.S. GDP. This benchmark covers 44 occupations and is built using tasks sourced from  industry professionals  with extensive experience, focusing on digital knowledge work. The research finds that  frontier models  are improving l...

Alibaba's $53 Billion AI War: Unpacking the Qwen3 'Yunqi Declaration' and the New Global Race for ASI 24.09.2025

The podcast provides an extensive analysis of  Alibaba's Qwen3 AI strategy , describing it as a meticulous, multi-front assault on the global AI landscape, backed by a capital commitment exceeding $53 billion. Alibaba is executing a sophisticated  "pincer movement"  strategy: on one side, it offers the proprietary, trillion-parameter  Qwen3-Max  model to compete for high-value enterp...

The Great AI Coding Paradox: Mastering Context Engineering to Beat 'Slop' on 500k-Line Codebases 23.09.2025

The podcast discusses a GitHub repository titled  "advanced-context-engineering-for-coding-agents"  under the  "humanlayer"  profile, which is a  public resource  evidenced by the notification, fork, and star counts. The content focuses on the  navigation and feature set of the GitHub platform , highlighting numerous tools and services for developers. Key offerings include  AI-...

OpenAI's 10 Gigawatt Gamble: The $100 Billion NVIDIA AI Deal, Energy Crisis, and the "Round Tripping" Debate 23.09.2025

The podcast centers on a significant  NVIDIA-OpenAI partnership  to deploy at least ten gigawatts (10GW) of AI data centers, which is raising serious concerns about the  massive electricity demand  and its resulting economic and environmental impact. Many view this metric as a problematic way to measure success, highlighting that such large-scale consumption is already contributing to  skyrocketin...

When AI Breaks: Anthropic's Postmortem Reveals the Three Infrastructure Bugs That Tanked Claude's Quality 22.09.2025

The podcast discusses a  technical postmortem from Anthropic  detailing three infrastructure bugs that intermittently degraded the quality of Claude's responses between August and September 2025, and a  collection of commentary  discussing the implications of these issues. Anthropic explains the three overlapping bugs—a  context window routing error , an  output corruption misconfiguration  on...

98% Cost Revolution: How xAI's Grok 4 Fast Rewrites the Economics of Frontier AI 21.09.2025

The podcast discusses the launch of  Grok 4 Fast , a new model from xAI designed for maximum cost-efficiency and intelligence density. This model achieves performance comparable to the larger Grok 4 while utilizing  40% fewer thinking tokens , resulting in a  98% reduction in price  for similar results on key benchmarks. Grok 4 Fast features a  unified architecture  that integrates both reasoning...

NVIDIA's $5 Billion Intel Bet: How the Arc-Rival NVLink Fusion Rewires PCs and AI with Uniform Memory Access 20.09.2025

The podcast discusses a major strategic partnership between  NVIDIA and Intel , highlighted by  NVIDIA’s $5 billion equity investment  in Intel. This collaboration centers on the co-development of new processor types, including  "Intel x86 RTX SoCs"  for the PC market that integrate an Intel x86 CPU chiplet with an NVIDIA RTX GPU chiplet. A technically significant feature of these new ch...

AI vs. VC: How LLMs Surpassed Human Experts in Spotting Unicorn Startups 19.09.2025

The podcast introduces  VCBench , the first standardized, anonymized benchmark designed to evaluate  Large Language Models (LLMs)  in the challenging domain of  venture capital (VC) founder-success prediction . Built from 9,000 founder profiles, the benchmark utilizes a multi-stage pipeline of standardization and adversarial testing to ensure  data privacy  by reducing re-identification risk by ov...

AI Outsmarts World's Best Programmers: The ICPC Revolution and the Future of Human-AI Collaboration 18.09.2025

The podcast discusses significant achievement of AI models from DeepMind and OpenAI in the 2025 International Collegiate Programming Contest (ICPC) World Finals, where they attained gold-medal equivalent performances, with OpenAI's system even achieving a perfect score. These texts highlight the  AI's advanced abstract reasoning capabilities , noting its success on a problem that stumped a...

GPT-5 Codex Unveiled: Your AI Co-Worker Revolutionizing Software Development 16.09.2025

This podcast features discussing the evolution and future of AI in coding, particularly focusing on  OpenAI's Codex and GPT-5 models . It explains how early observations of language models completing code led to the development of powerful AI coding assistants, emphasizing the critical role of a "harness"—the tools and infrastructure that allow the AI to interact with its environment...

Stop Overthinking: How AI is Learning to Think Smarter, Not Just Longer 15.09.2025

This podcast provides a comprehensive overview of efficient reasoning in  Large Language Models (LLMs) , identifying the "overthinking phenomenon" where models generate excessively lengthy and redundant reasoning steps. It explores various methodologies to optimize reasoning length while preserving performance, categorizing them into  model-based ,  reasoning output-based , and  input pr...

Seedream 4.0: The AI Image Game Changer for Creative Pros 14.09.2025

The podcast introduces  Seedream 4.0 , a new AI model from ByteDance released in September 2025, which is presented as the  definitive leader in AI image editing and generation . It highlights Seedream 4.0's  revolutionary unified architecture , featuring a Mixture-of-Experts (MoE) framework for unprecedented speed and efficiency, enabling  commercial-grade 4K resolution images  with near-real...

Listen to the Next in AI: Your Daily News Podcast podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.