nextbig.dev
The Rundown: Daily AI & Compute
The daily audio edition of nextbig.dev, read by the desk that wrote it. The biggest story in AI and compute, the wire, the desk's market tape, and The Call: one falsifiable position, settled in public. Six to nine minutes, every morning at 06:00 UTC.
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
OpenAI's strongest model went public today after a 12-day government gate, and Grok matched its tier by afternoon at a quarter of Anthropic's price. Frontier capability is no longer the scarce thing; the clearance to ship it and the power to run it are 09.07.2026 5:04
Two frontier models reached the public on the same Thursday, and how they got there matters more than what they can do. GPT-5.6 went generally available in three priced trims (Sol, Terra, Luna) after twelve days locked to government-approved partners under a "voluntary" White House review that functioned as preclearance; xAI's Grok 4.5 landed the same day claiming Opus-class work at $2/$6 against...
A hidden line in a public GitHub issue turned the platform's own AI agent into a private-repo leak, and the industry spent the same week handing agents more access, not less 09.07.2026 5:25
The value in AI keeps sliding off the model into the systems around it, and this week the risk followed. Researchers at Noma Security dubbed it GitLost: a plain-English command hidden in a public GitHub issue turned the platform's own AI agent, which held read access to private repos, into a data-exfiltration tool that posted private code as a public comment. The bypass was the word "Additionally....
Meta's expensive GPUs were sitting idle, waiting on storage. This week two of AI's biggest builders showed the same thing: the model and the chip are the cheap part now 06.07.2026 5:37
Two of AI's biggest builders spent this week showing the same thing from opposite ends of the stack: the model and the GPU are the cheap part now, and the value moved into the systems around them. Meta rebuilt its storage layer to stop GPUs sitting idle, cutting dataset loads from about 150 minutes to 10; OpenAI's Codex team teased GPT-5.6 Sol Ultra, whose benchmark gain comes from cooperating sub...
Capable AI got cheaper all week. This weekend, builders caught the newest models fumbling the simplest job: using the tools you give them 05.07.2026 5:10
A holiday-quiet Sunday, and the story that matters is a builder's complaint: this weekend Armin Ronacher and an OpenAI Codex thread both caught the newest AI models, Opus 4.8, Sonnet 5, the latest Codex, getting worse at using third-party tools, even as they get better at everything else. The week made capable AI cheap; the weekend showed the difficulty moved into the harness around the model. Plu...
On Independence Day, the story is a benchmark: frontier AI running cheap on hardware that isn't Nvidia's 04.07.2026 5:00
On the Fourth of July, the most on-theme story on a half-size wire is a benchmark: GLM 5.2 running competitively on AMD, with performance per dollar still falling. A reflective close to a week that made compute independence, freedom from any single model or chip vendor, concrete for builders.
"Right to Local Intelligence" got a manifesto this week, and the models to make it real 04.07.2026 5:13
A "Right to Local Intelligence" manifesto trended this week beside working guides for running state-of-the-art models on your own hardware, and beside Virginia banning geolocation-data sales and warnings of an American privacy emergency. Why cheap open weights plus a privacy mandate is turning on-prem inference from a hobby into a purchase order.
A Chinese open model showed up inside GitHub Copilot this week, one click from OpenAI 04.07.2026 5:00
GitHub Copilot now serves Kimi K2.7 Code, an open model from China's Moonshot, generally available to millions of developers, while z.ai ships its own GLM-5.2 agent harness and even OpenAI ships a Claude Code plugin. Why the coding tool is going model-neutral, and what a commodity shelf means for the closed labs.
Anthropic's cheap new model runs agents like the flagship did, and the price of capable AI fell again 04.07.2026 4:54
Anthropic shipped Claude Sonnet 5, flagship-grade agentic coding and tool use at mid-tier pricing with a 1M context window, the same week Etched hit a $5B valuation on $1B of booked inference orders. Why the price of capable AI is collapsing from both ends, software and silicon.
Fable 5 is back at full capacity, and Anthropic's answer to a brutal week is a model you can actually call 01.07.2026 4:35
Fable 5 is back. Anthropic returned its flagship to full general availability at full rate limits, reclaiming the frontier from GLM 5.2 and gated Sol. Why the fix that matters is availability, not the benchmark.
A Chinese open model caught Claude this week, and the moat that's left isn't the model 01.07.2026 4:33
China's open models caught the Western frontier this week: Semgrep's cyber eval put Zhipu's GLM 5.2 level with Claude, at a fraction of the cost, with the weights open to download. Why the closed labs' moat is no longer the model.
Wall Street goes hunting for the next Nvidia and lands on the people who make memory 28.06.2026 6:05
Wall Street is pricing Micron like the next Nvidia and Lenovo says the memory shortage is permanent. Plus China retakes the supercomputer crown, AI coding agents tricked into installing malware, and the open-model field widens.
Anthropic put a remembering Claude inside Slack, and Salesforce owns the channel it learns on 27.06.2026 4:31
Anthropic’s Claude Tag makes Claude a shared, remembering coworker inside Slack, a channel Salesforce owns and competes in. Why distribution decides the AI-coworker market more than capability.
OpenAI launched its strongest model to about twenty government-approved partners, and no one else yet 27.06.2026 4:13
OpenAI previewed GPT-5.6 (Sol, Terra, Luna) and shipped its flagship first to about twenty US-government-approved partners. Why the access list, more than the benchmark, is the story.
Claude burned through 2h37m of degraded models in a single day, and it is the third straight week 24.06.2026 9:06
Claude logged 2h37m of degraded models in one day, the third straight week of capacity-bound outages. What it means for any team running Claude Code in CI.
SpaceX becomes a compute landlord, and Reflection signs a $6.3B lease 23.06.2026 8:43
SpaceX turns its Colossus cluster into a compute landlord as Reflection commits up to $6.3B through 2029, while GB200 serving costs fall 2.5x in 70 days.
GLM-5.2 turns leaving Claude into a five-minute config change 22.06.2026 8:23
GLM-5.2 wires into Claude Code in five minutes as the White House forces Anthropic to cut SK Telecom, plus HBM tightening and SMPTE freeing 800+ standards.
GLM-5.2 puts top-tier coding within four points of Claude for a sixth the cost 21.06.2026 8:53
GLM-5.2 ships open MIT weights with coding scores four points behind Claude Opus and a sixth the cost, 48 hours after US export controls pulled a frontier model.
Builder.io ships an MIT-licensed framework that makes the agent a first-class user of your app 20.06.2026 10:00
Builder.io ships an MIT-licensed agent-native framework; ClickHouse hits $250M ARR; SemiAnalysis says 99% of custom AI ASICs fail before silicon.
Z.ai shipped GLM-5 under MIT and is already two releases past it 19.06.2026 9:54
Z.ai shipped GLM-5 under MIT and already pushed to a 1M-context GLM-5.2, reopening the default-model decision in every coding tool.
GLM-5.2 takes the open-weights crown at 51, and pays for it in tokens 18.06.2026 10:19
GLM-5.2 leads open weights at 51, but burns 43k tokens a task. Plus Epic open-sources Lore VCS and a code-graph MCP cuts agent reads 10x.
SpaceX buys Cursor for $60B and pulls developer tools inside the rocket 17.06.2026 8:40
SpaceX buys Cursor maker Anysphere for $60B, double its November valuation, pulling the top AI coding IDE inside its own compute stack.
Hetzner raises prices a third time in 2026, and AI is the reason 16.06.2026 8:07
Hetzner raises prices a third time in 2026, blaming AI demand for memory and GPUs, after April hikes of 30-53%. Plus Iroh 1.0 and OpenRouter Fusion.
Fabro turns agent orchestration into a version-controlled graph with per-node model routing 15.06.2026 7:21
Fabro ships an open source agent orchestrator with per-node model routing: Haiku on cheap steps, Sonnet on hard ones, fallback chains, Git commits per stage.
Washington Orders Anthropic to Pull Fable 5 and Mythos 5, So It Pulled Them for Everyone 14.06.2026 9:58
A US export order pulled Anthropic's Fable 5 and Mythos 5 worldwide over an investor's jailbreak claim. Plus GLM 5.2, 21 FFmpeg zero-days, TensorZero winds down.
Washington pulls two Anthropic models offline by federal order 13.06.2026 9:55
US Commerce orders Anthropic to disable Fable 5 and Mythos 5 for all customers by federal letter, the first frontier model pulled offline by directive.
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.