duan

HuggingFace 每日AI论文速递

每天10分钟,带您快速了解当日HuggingFace热门AI论文内容。每个工作日更新,欢迎订阅。📢播客节目在小宇宙、Apple Podcast平台搜索【HuggingFace 每日AI论文速递】🖼另外还有图文版,可在小红书搜索并关注【AI速递】

Author

duan

Category

Technology

Podcast website

www.xiaoyuzhoufm.com

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

2026.02.18 | GLM-5智能体工程登顶50分;SAE可解释性遭随机基线打脸 18.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:31] 🤖 GLM-5: from Vibe Coding to Agentic Engineering(GLM-5:从氛围编码到智能体工程) [01:11] 🔍 Sanity Checks for Sparse Autoencoders: Do SAEs Beat Random Baselines?(稀疏自编码器的合理性检验:SAE是否优于随机基线?) [01:57]...

2026.02.17 | 查询锚定用户画像;量子原生数据库 17.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:29] 🧠 Query as Anchor: Scenario-Adaptive User Representation via Large Language Model(查询作为锚点:基于大型语言模型的场景自适应用户表征) [01:14] ⚛ Qute: Towards Quantum-Native Database(Qute:迈向量子原生数据库) [01:59] 🧠...

2026.02.16 | 特征激活补数据;区域蒸馏藏放大 16.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:30] 🧠 Less is Enough: Synthesizing Diverse Data in Feature Space of LLMs(少即是够:在大型语言模型特征空间中合成多样化数据) [01:19] 🔍 Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Percepti...

【周末特辑】2月第3周最火AI论文 | OPUS精准选数据;弱模型反向助攻强模型 14.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 5 篇论文如下: [00:52] TOP1(🔥305) | 🚀 OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration(OPUS:迈向大规模语言模型预训练中高效且原理化的逐轮数据选择) [02:42] TOP2(🔥250) | 📈 Weak-Drive...

2026.02.13 | 自演化AI难守安全;音频大模型统一token 13.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:31] ⚠ The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies(魔书背后的魔鬼:在自我进化的AI社会中,人类安全价值总是趋于消失) [01:24] 🎵 MOSS-Audio-Tokenizer: Scaling Audio Tokenizers for...

2026.02.12 | 稀疏MoE比肩GPT-5;GENIUS测流体智能 12.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:28] ⚡ Step 3.5 Flash: Open Frontier-Level Intelligence with 11B Active Parameters(Step 3.5 Flash:拥有110亿活跃参数的前沿级智能模型) [01:06] 🧠 GENIUS: Generative Fluid Intelligence Evaluation Suite(GENIUS:生成式流体智能评...

2026.02.11 | OPUS对齐更新选数据;Code2World代码预演GUI 11.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:33] 🚀 OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration(OPUS:迈向大规模语言模型预训练中高效且原理化的逐轮数据选择) [01:17] 💻 Code2World: A GUI World Model via Ren...

2026.02.10 | ReAlign零训弥合图文隙;MOVA同步生成视音频 10.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:34] 🔀 Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models(面向多模态大语言模型的模态间隙驱动的子空间对齐训练范式) [01:23] 🎬 MOVA: Towards Scalable and Synchronized Video-Audio Gener...

2026.02.09 | AI问诊如住院医;互动悟规则才是真智能 09.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:32] 🩺 Baichuan-M3: Modeling Clinical Inquiry for Reliable Medical Decision-Making(Baichuan-M3:建模临床问询以实现可靠的医疗决策) [01:17] 🧭 OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Induct...

【周末特辑】2月第2周最火AI论文 | 分阶段统一动作空间;ERNIE 5.0大一统多模态 08.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 5 篇论文如下: [00:48] TOP1(🔥235) | 🤖 Green-VLA: Staged Vision-Language-Action Model for Generalist Robots(Green-VLA:面向通用机器人的分阶段视觉-语言-动作模型) [02:54] TOP2(🔥235) | 🧠 ERNIE 5.0 Technical Report(ERNIE 5.0 技术报告) [05:14] T...

2026.02.06 | RLVR去长度偏见;长镜头不换记忆 06.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:29] 📊 Length-Unbiased Sequence Policy Optimization: Revealing and Controlling Response Length Variation in RLVR(长度无偏序列策略优化:揭示与控制RLVR中的响应长度变化) [01:20] 🎬 Context Forcing: Consistent Autoregressive Vide...

2026.02.05 | ERNIE 5.0统一模态;FASA稀疏注意力省内存 05.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:29] 🧠 ERNIE 5.0 Technical Report(ERNIE 5.0 技术报告) [01:11] ⚡ FASA: Frequency-aware Sparse Attention(FASA:基于频率感知的稀疏注意力机制) [02:01] 📊 Training Data Efficiency in Multimodal Process Reward Models(多模态过程...

2026.02.04 | 看图写代码省token;临时组队降成本 04.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:32] 👁 CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding(CodeOCR:视觉语言模型在代码理解中的有效性研究) [01:18] 🤖 AOrchestra: Automating Sub-Agent Creation for Agentic Orchestration(AOrchestra:面...

2026.02.03 | 分阶段训练统一动作空间;MoE+视觉编码器并行智能体 04.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:32] 🤖 Green-VLA: Staged Vision-Language-Action Model for Generalist Robots(Green-VLA:面向通用机器人的分阶段视觉-语言-动作模型) [01:24] 🤖 Kimi K2.5: Visual Agentic Intelligence(Kimi K2.5:视觉智能体) [02:09] 🔍 Vision-Dee...

2026.02.02 | ASTRA合成轨迹炼工具;THINKSAFE自对齐保安全 02.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:33] 🤖 ASTRA: Automated Synthesis of agentic Trajectories and Reinforcement Arenas(ASTRA:基于自动化轨迹合成与强化学习竞技场的智能体训练框架) [01:22] 🛡 THINKSAFE: Self-Generated Safety Alignment for Reasoning Models(THINKSAF...

【月末特辑】1月最火AI论文 | mHC稳梯度;GDPO解多奖励 02.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 10 篇论文如下: [00:42] TOP1(🔥292) | 🧠 mHC: Manifold-Constrained Hyper-Connections(mHC:流形约束的超连接) [03:06] TOP2(🔥212) | 📈 GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization(GDPO:面向多奖...

【周末特辑】2月第1周最火AI论文 | LLM当管家,数据变净菜;LongCat训特工,上网打副本 01.02.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 5 篇论文如下: [00:39] TOP1(🔥181) | 🧹 Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs(大语言模型能否清理你的数据?基于LLM的应用就绪数据准备综述) [02:50] TOP2(🔥169) | 🧠 LongCat-Flash-Thinking-2601 Technic...

2026.01.30 | 空间智能基准测不准;Idea2Story一键成文 30.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:29] 🧭 Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models(万物归位:文本到图像模型空间智能基准测试) [01:21] 🧠 Idea2Story: An Automated Pipeline for Transforming Research Concepts into Complete...

2026.01.29 | 难题优先补数学推理;LingBot生成交互平行世界 29.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 13 篇论文如下: [00:33] 🧠 Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation(越难越好:通过难度感知GRPO与多角度问题重构提升数学推理能力) [01:21] 🌍 Advancing Open-source World Mod...

2026.01.28 | AgentDoG筑护栏诊断风险根源;AdaReasoner排工具小模型逆袭 28.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 14 篇论文如下: [00:30] 🛡 AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security(AgentDoG:面向AI智能体安全与安全的诊断性护栏框架) [01:21] 🧩 AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning(AdaReasone...

2026.01.27 | Agent原生训练刷新SWE-Bench;LLM重塑数据清洗 pipeline 27.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:33] 🤖 daVinci-Dev: Agent-native Mid-training for Software Engineering(daVinci-Dev:面向软件工程的智能体原生中期训练) [01:21] 🧹 Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs(大语言模...

2026.01.26 | LongCat练5600亿MoE代理满分;SWE-Pruner剪五成Token更快 26.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:32] 🧠 LongCat-Flash-Thinking-2601 Technical Report(LongCat-Flash-Thinking-2601 技术报告) [01:13] ✂ SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents(SWE-Pruner:面向编码代理的自适应上下文剪枝框架) [02:08] 🧠 Twin...

【周末特辑】1月第4周最火AI论文 | Agentic LLM进化成行动派;群体RL纠偏难度歧视 24.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 5 篇论文如下: [00:44] TOP1(🔥159) | 🤖 Agentic Reasoning for Large Language Models(大语言模型的智能体推理) [03:02] TOP2(🔥138) | ⚖ Your Group-Relative Advantage Is Biased(你的组相对优势存在偏差) [05:37] TOP3(🔥71) | 🤖 Being-H0.5: Scaling Hum...

2026.01.23 | BayesianVLA逼模型“读心”;扩散模型“按顺序”更聪明 23.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:32] 🤖 BayesianVLA: Bayesian Decomposition of Vision Language Action Models via Latent Action Queries(BayesianVLA:通过潜在动作查询对视觉语言动作模型进行贝叶斯分解) [01:22] ⚠ The Flexibility Trap: Why Arbitrary Order Limits R...

2026.01.22 | LLM变数字特工;视频模型先考后练 22.01.2026

【赞助商】 通勤路上就听AI每周谈。AI每周谈,每周带你回顾上周AI大事 传送门 🔗https://www.xiaoyuzhoufm.com/podcast/688a34636f5a275f1cba40fd 【目录】 本期的 15 篇论文如下: [00:30] 🤖 Agentic Reasoning for Large Language Models(大语言模型的智能体推理) [01:05] 🤖 Rethinking Video Generation Model for the Embodied World(为具身世界重新思考视频生成模型) [01:43] 🤖 Paper2Rebuttal: A Multi-Agent Framewo...

Listen to the HuggingFace 每日AI论文速递 podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.