Nathan Lambert
Interconnects
Audio essays about the latest developments in AI and interviews with leading scientists in the field. Breaking the hype, understanding what's under the hood, and telling stories. www.interconnects.ai
Author
Nathan Lambert
Category
Podcast website
Latest episode
Jun 22, 2026
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
Use multiple models 11.01.2026 7:12
I’ll start by explaining my current AI stack and how it’s changed in recent months. For chat, I’m using a mix of: * GPT 5.2 Thinking / Pro : My most frequent AI use is getting information. This is often a detail about a paper I’m remembering, a method I’m verifying for my RLHF Book , or some other niche fact. I know GPT 5.2 can find it if it exists, and I use Thinking for queries that I think are...
Claude Code Hits Different 09.01.2026 4:58
There is an incredible amount of hype for Claude Code with Opus 4.5 across the web right now, which I for better or worse entirely agree with. Having used coding agents extensively for the past 6-9 months, where it felt like sometimes OpenAI’s Codex was the best and sometimes Claude, there was some meaningful jump over the last few weeks. The jump is well captured by this post, which called it the...
Open models: Hot or Not with Nathan Lambert & Florian Brand 18.12.2025 37:36
Nathan sits down with Florian, our open model analyst to get spicy into debates of which labs won and lost momentum in open models of 2025. Reflection 70B, Huawei repackaging someone else's model as their own, the fall of Llama — no drama is left unturned. We also dig into the nuances that we didn't get to in our post, predict GPT-OSS 2, the American v. China balance at the end of 2026, and many o...
New Talk: Building Olmo 3 Think 10.12.2025 1:02:22
It’s finally here! The public (and most complete) version of my talk covering every stage of the process to build Olmo 3 Think ( slides are available). I’ve been giving this, improving it, and getting great feedback at other venues such as The Conference on Language Modeling (COLM) & The PyTorch Conference. This involves changes and new considerations of every angle of the stack, from pretraining,...
Olmo 3: America’s truly open reasoning models 20.11.2025 10:57
We present Olmo 3, our next family of fully open, leading language models. This family of 7B and 32B models represents: * The best 32B base model. * The best 7B Western-origin thinking & instruct models. * The first 32B (or larger) fully open reasoning model. This is a big milestone for Ai2 and the Olmo project. These aren’t huge models (more on that later), but it’s crucial for the viability of f...
Why AI writing is mid 17.11.2025 8:28
First, on the topic of writing, the polished, and more importantly printed , version of my RLHF Book is available for pre-order. It’s 50% off for a limited time, you can pre-order it here ! Like a lot of writing, I’ve been sitting on this piece for many months thinking it’s not contributing enough, but the topic keeps coming up — most recently via Jasmine Sun — and people seem to like it, so I hop...
Interview: Ant Group's open model ambitions 12.11.2025 1:17:49
This is the first of a handful of interviews I’m doing with teams building the best open language models of the world. In 2025, the open model ecosystem has changed incredibly . It’s more populated, far more dominated by Chinese companies, and growing. DeepSeek R1 shocked the world and now there are a handful of teams in China training exceptional models. The Ling models, from InclusionAI — Ant Gr...
5 Thoughts on Kimi K2 Thinking 06.11.2025 7:37
First, congrats to the Moonshot AI team, one of the 6 “ AI Tigers ” in China, on the awesome release of Kimi K2 Thinking . One of the overlooked and inspiring things for me these days is just how many people are learning very quickly to train excellent AI models. The ability to train leading AI models and distribute them internationally is going to be pervasive globally. As people use AI more, tho...
Burning out 25.10.2025 10:09
One of the obvious topics of the Valley today is how hard everyone works . We’re inundated with comments on “The Great Lock In”, 996 , 997, and now even a snarky 002 (midnight to midnight with a 2 hour break). Plenty of this is performative flexing on social media, but enough of it is real and reflecting how trends are unfolding in the LLM space. I’m affected. My friends are affected. All of this...
How to scale RL 20.10.2025 13:01
Two quick housekeeping items before I get to the post.1. I’ll be in SF this week for the PyTorch conference (22-23), AI Infra Summit (21st), and other local events. Come say hi.2. I launched a new Substack AI bundle with 8 of my favorite publications packaged together for teams of 20+. Learn more at readsail.com .Onto the post! “Scaling reinforcement learning (RL)” is the zeitgeisty way to capture...
The State of Open Models 16.10.2025 47:04
This talk covers everything that’s happened this year in the open model landscape — DeepSeek kickstarting the Chinese open model norms, Llama’s fade, Qwen’s dominance, GPT-OSS — and what comes next. It is my attempt to share what people need to know about where open models are heading, building on all of my research here at Interconnects and in my day job of training these models, in order to help...
Thoughts on The Curve 07.10.2025 11:58
I spent the weekend debating AI timelines, among other things, at The Curve conference. This translates as spending the weekend thinking about the trajectory of AI progress with a mix of DC and SF types. This is a worthwhile event that served as a great, high-bandwidth way to check in on timelines and expectations of the AI industry. Updating timelines My most striking takeaway is that the AI 2027...
ChatGPT: The Agentic App 30.09.2025 9:24
Ever since ChatGPT exploded in popularity, there has been a looming “how” to its monetization plans. Much has been said about shopping and advertising as the likely paths, especially with Fidji Simo joining as CEO of Applications under Sam Altman. Advertising as a business model for AI is logical but difficult to personalize and specialize. We know tons of people spend a lot of time using AI model...
Thinking, Searching, and Acting 22.09.2025 9:22
The weaknesses of today’s best models are far from those of the original ChatGPT — we see they lack speed, we fear superhuman persuasion, and we aspire for our models to be more autonomous. These models are all reasoning models that have long surpassed the original weaknesses of ChatGPT-era language models, hallucinations, total lack of recent information, complete capitulations, and other hiccups...
Coding as the epicenter of AI progress and the path to general agents 18.09.2025 16:18
Coding, due to its breadth of use-cases, is arguably the last tractable, general domain of continued progress for frontier models that most people can interface with. This is a bold claim, so let’s consider some of the other crucial capabilities covered in the discourse of frontier models: * Chat and the quality of prose written by models has leveled off, other than finetuning to user measures suc...
On China's open source AI trajectory 09.09.2025 13:37
Hello everyone! I’m coming back online after two weeks of vacation. Thankfully it coincided with some of the slowest weeks of the year in the AI space. I’m excited to get back to writing and (soon) share projects that’ll wrap up in the last months of the year. It seemed like a good time to remind people of the full set of housekeeping for Interconnects. * Many people love the audio version of the...
Ranking the Chinese Open Model Builders 17.08.2025 12:41
The Chinese AI ecosystem has taken the AI world by storm this summer with an unrelenting pace of stellar open model releases. The flagship releases that got the most Western media coverage are the likes of Qwen 3 , Kimi K2 , or Zhipu GLM 4.5 , but there is a long-tail of providers close behind in both quality and cadence of releases. In this post we rank the top 19 Chinese labs by the quality and...
Contra Dwarkesh on Continual Learning 15.08.2025 10:04
Dwarkesh Patel ’s now well-read post on why he is extending his AI timelines focuses on the idea of continual learning. If you ask me, what we have already is AGI , so the core question is: Is continual learning a bottleneck on AI progress? In this post, I argue that continual learning as he describes it actually doesn’t matter for the trajectory of AI progress that we are on. Continual learning w...
GPT-5 and the arc of progress 07.08.2025 10:41
If you want a video version of this, check out the last 20 minutes of the livestream reaction (edit, fixed link) I did with Will Brown of Prime Intellect and Swyx of Smol AI & Latent Space. GPT-5 was set up to fail on some of the narratives it was expected to satisfy. The two central themes it had to decide between were the AGI (or superintelligence) narrative that Sam Altman & co. have been using...
gpt-oss: OpenAI validates the open ecosystem (finally) 05.08.2025 13:36
OpenAI released two open-weight, text-only reasoning models today, both mixture of experts (MoE) sized to run efficiently on a range of hardware from consumer GPUs to the cloud. These models have the Apache 2.0 license, so they’re available for distillation into other reasoning models, deployment into commercial products, and are free of downstream restrictions. These two models, the smaller gpt-o...
Towards American Truly Open Models: The ATOM Project 04.08.2025 22:12
I’m very excited to share a substantial project on invigorating investment in open language models and AI research in the U.S. The ATOM (American Truly Open Models) Project is the mature evolution of my original “ American DeepSeek Project ” and I hope it can help be a turning point in the current trajectory of losing open model relevance vis-a-vis China, and even the rest of the world. I’ve inclu...
Interviewing Ross Taylor on the state of AI: Chinese open models, scaling reasoning, useful tools, and what comes next 29.07.2025 1:14:40
I’m excited to welcome Ross Taylor back on the podcast (and sorry for the lack of episodes in general – I have a lot going on!). The first time Ross came on we focused on reasoning – before inference-time scaling and that sort of RL was popular, agents, Galactica, and more from his Llama days. Since then, and especially after DeepSeek R1, Ross and I have talked asynchronously about the happenings...
The White House's plan for open models & AI research in the U.S. 23.07.2025 13:10
Today, the White House released its AI Action Plan , the document we’ve been waiting for to understand how the new administration plans to achieve “global dominance in artificial intelligence (AI).” There’s a lot to unpack in this document, which you’ll be hearing a lot about from the entire AI ecosystem. This post covers one narrow piece of the puzzle — its limited comments on open models and AI...
Kimi K2 and when "DeepSeek Moments" become normal 14.07.2025 6:44
https://www.interconnects.ai/p/kimi-k2-and-when-deepseek-moments The DeepSeek R1 release earlier this year was more of a prequel than a one-off fluke in the trajectory of AI. Last week, a Chinese startup named Moonshot AI dropped Kimi K2 , an open model that is permissively licensed and competitive with leading frontier models in the U.S. If you're interested in the geopolitics of AI and the rapid...
The American DeepSeek Project 04.07.2025 10:36
https://www.interconnects.ai/p/the-american-deepseek-project While America has the best AI models in Gemini, Claude, o3, etc. and the best infrastructure with Nvidia it’s rapidly losing its influence over the future directions of AI that unfold in the open-source and academic communities. Chinese organizations are releasing the most notable open models and datasets across all modalities, from text...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.